Back to approved models
Z.AIactiverelay/request.mjsready to route CONTEXT200K contextprompt window MAX OUTPUT128Kper response INPUT$0.30 / Mtoken price OUTPUT$0.90 / Mtoken price
Approved model profile_
GLM 4.6V
GLM-4.6V is a fast multimodal model designed for high-fidelity visual understanding and long-context reasoning across images, documents, and mixed media, optimized for low-latency performance.
glm-4.6vconst completion = await client.chat.completions.create({
model: "glm-4.6v",
messages: [
{ role: "user",
content: "Plan a multi-step task" }
]
});Route configuration_
Use the model through the contract you already have.
PipeLLM resolves provider access, policy, and supported model capabilities without changing the interface your application calls.
https://api.pipellm.ai/openaiOpenAI-compatible routeglm-4.6vUse this in every requestPROVIDER AVAILABILITY1 route
| Provider | Context | Max output | Input / M | Output / M |
|---|---|---|---|---|
| Z.AI | 200K context | 128K | $0.30 / M | $0.90 / M |
SUPPORTED SURFACESactive
input: textinput: imageoutput: texttool use
Relay governs model access with the same routing policy that Lens keeps visible and auditable.
