Back to approved models
ByteDanceactiverelay/request.mjsready to route CONTEXT1M contextprompt window MAX OUTPUT384Kper response INPUT$0.105 / M30% OFF · list $0.15 / M OUTPUT$0.42 / M30% OFF · list $0.60 / M
Approved model profile_
Deepseek V4 1 Flash
Deepseek V4 1 Flash is available through the PipeLLM Model Router.
deepseek-v4-1-flashconst completion = await client.chat.completions.create({
model: "deepseek-v4-1-flash",
messages: [
{ role: "user",
content: "Plan a multi-step task" }
]
});Route configuration_
Choose a supported call format.
Select a format below to view its documentation. Use the displayed converter route when calling through PipeLLM.
https://api.pipellm.aiOne base URL for OpenAI Chat Completions, Responses, Anthropic, and Gemini.SUPPORTED CALL FORMATS
OpenAI/openai/v1/chat/completionsDocs Responses/responses/v1/responsesDocs Anthropic/anthropic/v1/messagesDocs Gemini/gemini/v1beta/models/{model}:generateContentDocs Choose a call format to open its request and response documentation.deepseek-v4-1-flashUse this in every requestPROVIDER AVAILABILITY1 route
| Provider | Context | Max output | Discount | Input / M | Output / M |
|---|---|---|---|---|---|
| Volcano Engine | 1M context | 384K | 30% OFF | $0.105 / M | $0.42 / M |
SUPPORTED SURFACESactive
input: textoutput: texttool usecache control
Relay governs model access with the same routing policy that Lens keeps visible and auditable.
