Back to approved models
OpenAIactiverelay/request.mjsready to route CONTEXT1M contextprompt window MAX OUTPUT128Kper response INPUT$0.16 / Mtoken price OUTPUT$0.53 / Mtoken price
Approved model profile_
GLM 5.3 Flash
GLM-5.3-Flash is Z.ai's multimodal model for efficient software development and extended agent tasks. It combines text generation with an understanding of images and video, allowing visual information to inform its responses. Its long-context design supports workflows that carry substantial source...
glm-5.3-flashconst completion = await client.chat.completions.create({
model: "glm-5.3-flash",
messages: [
{ role: "user",
content: "Plan a multi-step task" }
]
});Route configuration_
Use the model through the contract you already have.
PipeLLM resolves provider access, policy, and supported model capabilities without changing the interface your application calls.
https://api.pipellm.aiOne base URL for OpenAI, Anthropic, and Gemini.FORMAT ROUTES
OPENAI
/openaiANTHROPIC
/anthropicGEMINI
Use a format route to normalize requests and responses./geminiglm-5.3-flashUse this in every requestPROVIDER AVAILABILITY1 route
| Provider | Context | Max output | Discount | Input / M | Output / M |
|---|---|---|---|---|---|
| Tokenhub | 1M context | 128K | — | $0.16 / M | $0.53 / M |
SUPPORTED SURFACESactive
input: textoutput: texttool use
Relay governs model access with the same routing policy that Lens keeps visible and auditable.
