Back to approved models
Googleactiverelay/request.mjsready to route CONTEXT1M contextprompt window MAX OUTPUT65Kper response INPUT$0.05 / Mtoken price OUTPUT$3.00 / Mtoken price
Approved model profile_
Gemini 3 Flash Preview
Gemini 3 Flash Preview is a high speed, high value thinking model designed for agentic workflows, multi turn chat, and coding assistance. It delivers near Pro level reasoning and tool use performance with substantially lower latency than larger Gemini variants, making it well suited for interactive...
gemini-3-flash-previewconst completion = await client.chat.completions.create({
model: "gemini-3-flash-preview",
messages: [
{ role: "user",
content: "Plan a multi-step task" }
]
});Route configuration_
Choose a supported call format.
Select a format below to view its documentation. Use the displayed converter route when calling through PipeLLM.
https://api.pipellm.aiOne base URL for OpenAI Chat Completions, Responses, Anthropic, and Gemini.SUPPORTED CALL FORMATS
OpenAI/openai/v1/chat/completionsDocs Responses/responses/v1/responsesDocs Anthropic/anthropic/v1/messagesDocs Gemini/gemini/v1beta/models/{model}:generateContentDocs Choose a call format to open its request and response documentation.gemini-3-flash-previewUse this in every requestPROVIDER AVAILABILITY2 routes
| Provider | Context | Max output | Discount | Input / M | Output / M |
|---|---|---|---|---|---|
| GCP Vertex | 1M context | 65K | — | $0.05 / M | $3.00 / M |
| Vertex AI | N/A | N/A | — | $0.05 / M | $3.00 / M |
SUPPORTED SURFACESactive
input: textinput: imageoutput: texttool usecache controlcomputer use
Relay governs model access with the same routing policy that Lens keeps visible and auditable.
