Back to approved models
Z.AIactiverelay/request.mjsready to route CONTEXT200K contextprompt window MAX OUTPUT128Kper response INPUT$0.00 / Mtoken price OUTPUT$0.00 / Mtoken price
Approved model profile_
GLM 4.6v Flash
GLM-4.6V is a large multimodal model designed for high-fidelity visual understanding and long-context reasoning across images, documents, and mixed media. It supports up to 128K tokens, processes complex page layouts and charts directly as visual inputs, and integrates native multimodal function...
glm-4.6v-flashPOST https://api.pipellm.ai/chat/completions
Authorization: Bearer $PIPELLM_API_KEY
Content-Type: application/json
Model: glm-4.6v-flash
See the provider documentation for request parameters.Route configuration_
Call this model through its native API.
Use the documented HTTP endpoint and request body for this model.
https://api.pipellm.aiUse the native HTTP request format for this model.NATIVE API · NO CONVERSION
Native HTTP/chat/completionsDocs Use this native API directly. Protocol conversion is not supported.glm-4.6v-flashUse this in every requestPROVIDER AVAILABILITY1 route
| Provider | Context | Max output | Discount | Input / M | Output / M |
|---|---|---|---|---|---|
| Z.AI | 200K context | 128K | — | $0.00 / M | $0.00 / M |
SUPPORTED SURFACESactive
input: textinput: imageoutput: texttool usecache control
Relay governs model access with the same routing policy that Lens keeps visible and auditable.
