Back to approved models

Approved model profile_

OpenAIactive

GLM 5.3 Flash

GLM-5.3-Flash is Z.ai's multimodal model for efficient software development and extended agent tasks. It combines text generation with an understanding of images and video, allowing visual information to inform its responses. Its long-context design supports workflows that carry substantial source...

glm-5.3-flash
relay/request.mjsready to route
const completion = await client.chat.completions.create({
  model: "glm-5.3-flash",
  messages: [
    { role: "user",
      content: "Plan a multi-step task" }
  ]
});
policy checked provider selected
CONTEXT1M contextprompt window
MAX OUTPUT128Kper response
INPUT$0.16 / Mtoken price
OUTPUT$0.53 / Mtoken price

Route configuration_

Use the model through the contract you already have.

PipeLLM resolves provider access, policy, and supported model capabilities without changing the interface your application calls.

ENDPOINThttps://api.pipellm.aiOne base URL for OpenAI, Anthropic, and Gemini.

FORMAT ROUTES

OPENAI/openai
ANTHROPIC/anthropic
GEMINI/gemini
Use a format route to normalize requests and responses.
MODEL IDglm-5.3-flashUse this in every request
GOVERNANCEPolicy controlledApprovals and tool rules apply
PROVIDER AVAILABILITY1 route
ProviderContextMax outputDiscountInput / MOutput / M
Tokenhub1M context128K$0.16 / M$0.53 / M
Displayed pricing is per one million tokens and may vary by provider.
SUPPORTED SURFACESactive
input: textoutput: texttool use

Relay governs model access with the same routing policy that Lens keeps visible and auditable.