Back to approved models
OpenAIactiverelay/request.mjsready to route CONTEXT400K contextprompt window MAX OUTPUT128Kper response INPUT$1.25 / Mtoken price OUTPUT$10.00 / Mtoken price
Approved model profile_
GPT-5.1
GPT-5.1-Codex-Max is OpenAI’s latest agentic coding model, designed for long-running, high-context software development tasks. It is based on an updated version of the 5.1 reasoning stack and trained on agentic workflows spanning software engineering, mathematics, and research. GPT-5.1-Codex-Max...
gpt-5.1-2025-11-13const completion = await client.chat.completions.create({
model: "gpt-5.1-2025-11-13",
messages: [
{ role: "user",
content: "Plan a multi-step task" }
]
});Route configuration_
Choose a supported call format.
Select a format below to view its documentation. Use the displayed converter route when calling through PipeLLM.
https://api.pipellm.aiOne base URL for OpenAI Chat Completions, Responses, Anthropic, and Gemini.SUPPORTED CALL FORMATS
OpenAI/openai/v1/chat/completionsDocs Responses/responses/v1/responsesDocs Anthropic/anthropic/v1/messagesDocs Gemini/gemini/v1beta/models/{model}:generateContentDocs Choose a call format to open its request and response documentation.gpt-5.1-2025-11-13Use this in every requestPROVIDER AVAILABILITY2 routes
| Provider | Context | Max output | Discount | Input / M | Output / M |
|---|---|---|---|---|---|
| OpenAI | 400K context | 128K | — | $1.25 / M | $10.00 / M |
| Azure | 400K context | 128K | — | $1.25 / M | $10.00 / M |
SUPPORTED SURFACESactive
input: textinput: imageoutput: texttool usecache control
Relay governs model access with the same routing policy that Lens keeps visible and auditable.
