Back to approved models
Googleactiverelay/request.mjsready to route CONTEXTN/Aprompt window MAX OUTPUTN/Aper response INPUT$0.10 / Mtoken price OUTPUT$0.40 / Mtoken price
Approved model profile_
Gemini 2.5 Flash-Lite
Gemini 2.5 Flash-Lite is a lightweight reasoning model in the Gemini 2.5 family, optimized for ultra-low latency and cost efficiency. It offers improved throughput, faster token generation, and better performance...
gemini-2.5-flash-liteconst completion = await client.chat.completions.create({
model: "gemini-2.5-flash-lite",
messages: [
{ role: "user",
content: "Plan a multi-step task" }
]
});Route configuration_
Use the model through the contract you already have.
PipeLLM resolves provider access, policy, and supported model capabilities without changing the interface your application calls.
https://api.pipellm.aiOne base URL for OpenAI, Anthropic, and Gemini.FORMAT ROUTES
OPENAI
/openaiANTHROPIC
/anthropicGEMINI
Use a format route to normalize requests and responses./geminigemini-2.5-flash-liteUse this in every requestPROVIDER AVAILABILITY1 route
| Provider | Context | Max output | Discount | Input / M | Output / M |
|---|---|---|---|---|---|
| Vertex AI | N/A | N/A | — | $0.10 / M | $0.40 / M |
SUPPORTED SURFACESactive
input: textoutput: text
Relay governs model access with the same routing policy that Lens keeps visible and auditable.
