Back to approved models

Approved model profile_

Z.AIactive

GLM 4.6v Flash

GLM-4.6V is a large multimodal model designed for high-fidelity visual understanding and long-context reasoning across images, documents, and mixed media. It supports up to 128K tokens, processes complex page layouts and charts directly as visual inputs, and integrates native multimodal function...

glm-4.6v-flash
relay/request.mjsready to route
POST https://api.pipellm.ai/chat/completions
Authorization: Bearer $PIPELLM_API_KEY
Content-Type: application/json

Model: glm-4.6v-flash
See the provider documentation for request parameters.
policy checked provider selected
CONTEXT200K contextprompt window
MAX OUTPUT128Kper response
INPUT$0.00 / Mtoken price
OUTPUT$0.00 / Mtoken price

Route configuration_

Call this model through its native API.

Use the documented HTTP endpoint and request body for this model.

ENDPOINThttps://api.pipellm.aiUse the native HTTP request format for this model.

NATIVE API · NO CONVERSION

Native HTTP/chat/completionsDocs Use this native API directly. Protocol conversion is not supported.
MODEL IDglm-4.6v-flashUse this in every request
GOVERNANCEPolicy controlledApprovals and tool rules apply
PROVIDER AVAILABILITY1 route
ProviderContextMax outputDiscountInput / MOutput / M
Z.AI200K context128K—$0.00 / M$0.00 / M
Displayed pricing is per one million tokens and may vary by provider.
SUPPORTED SURFACESactive
input: textinput: imageoutput: texttool usecache control

Relay governs model access with the same routing policy that Lens keeps visible and auditable.