Approved catalogue_

Choose the model. Keep the contract.

Provider pricing stays visible. Routing and governance stay behind one API.

ROUTABLE MODELS107 of 110 available
TypeSafe

Jev 1.13.0

jev-1.13.0

Jev 1.13.0 provides typed AI judgments through PipeLLM. Use model ID jev-1.13.0 with /v1/systemone.

text
CONTEXT
64K context
INPUT
$0.042/M
OUTPUT
$0/M
OpenAI

GPT 6 Sol

gpt-6-sol

Developed by OpenAI as part of its GPT-6 family, GPT-6 Sol is positioned as a cost-effective, high-tier language model. It occupies an intermediate place in the lineup, designed to sit directly beneath the flagship GPT-6 Astra tier while providing more advanced capability than the rapid-performance GPT-6 Luna variant.Engineered for demanding professional applications, GPT-6 Sol targets workloads that require strong performance and precision alongside balanced operating expenses. Its positioning makes it an effective choice for complex business tasks that benefit from near-flagship intelligence with greater resource efficiency.

text
CONTEXT
1M context
INPUT
$2/M
OUTPUT
$10/M
OpenAI

GPT-6 Astra

gpt-6-astra

GPT-6 Astra provides text generation through PipeLLM. Use model ID gpt-6-astra with /v1/chat/completions.

texttoolscache
CONTEXT
1M context
INPUT
$10/M
OUTPUT
$50/M
Google

Gemini 3.8 Flash

gemini-3.8-flash

Gemini 3.8 Flash extends Google's Flash family with a focus on software engineering, agent tasks, and complex reasoning. It accepts text, visual material, audio, and documents, bringing these inputs together in text-based responses.The model is suited to coding assistants, multimodal analysis, and workflows that require repeated reasoning and tool use. Its long-context support helps applications work with substantial documents and extended conversations.

texttools
CONTEXT
1M context
INPUT
$0.75/M
OUTPUT
$3.75/M
Anthropic

Claude Opus 5 5

claude-opus-5-5

Claude Opus 5.5 is an artificial intelligence model developed by Anthropic, made available to developers through the Claude Platform. It forms part of the publisher's expanding suite of models designed to be deployed and integrated directly into software applications.Within the Claude ecosystem, the model aligns with products such as Claude Code and Claude Code Enterprise. This positioning provides developers with access to Anthropic's capabilities across enterprise workflows and code-focused development environments.

text
CONTEXT
1M context
INPUT
$4/M
OUTPUT
$20/M
Z.AI

GLM 5.3

glm-5.3

GLM-5.3 is Z.ai's reasoning model for demanding software engineering and extended agent workflows. It processes text and is designed to work through complex problems while retaining substantial contextual information.The model suits code analysis, implementation tasks, and multi-step technical work. Its focus on reasoning and token efficiency makes it a candidate for applications that need sustained problem solving rather than brief, isolated responses.

texttoolscache
CONTEXT
1M context
INPUT
$1.4/M
OUTPUT
$4.4/M
Anthropic

Claude Opus 5

claude-opus-5

Claude Opus 5 is Anthropic’s flagship model for demanding reasoning, coding, and long-horizon agentic work. It is particularly strong at end-to-end software tasks, code review and bug finding, visual analysis...

texttoolscache
CONTEXT
1M context
INPUT
$5/M
OUTPUT
$25/M
Kimi AI

Kimi K3

kimi-k3

Kimi K3 is a 2.8T parameter open-weight multimodal reasoning model from Moonshot AI. It is suited for complex coding, knowledge work, and long-horizon agentic workflows, and is particularly strong at...

texttoolscache
CONTEXT
1M context
INPUT
$3/M
OUTPUT
$15/M
DISCOUNT30% OFF
ByteDance

Deepseek V4 1 Flash

deepseek-v4-1-flash

Approved for governed routing through Relay.

texttoolscache
CONTEXT
1M context
INPUT
$0.105/M$0.15/M
OUTPUT
$0.42/M$0.6/M
DISCOUNT30% OFF
DeepSeek

Deepseek V4 Pro

deepseek-v4-pro

DeepSeek V4 Pro is a large-scale Mixture-of-Experts model aimed at advanced reasoning, coding, long-context analysis, and long-horizon agent workflows. It is suited for full-codebase analysis, multi-step automation, and large-scale information synthesis where capability and efficiency both matter.

texttoolscache
CONTEXT
1M context
INPUT
$1.26/M$1.8/M
OUTPUT
$2.52/M$3.6/M
OpenAI

GPT 6.1 Sol

gpt-6.1-sol

Approved for governed routing through Relay.

text
CONTEXT
1M context
INPUT
$2/M
OUTPUT
$10/M
OpenAI

GPT 6 Luna

gpt-6-luna

Approved for governed routing through Relay.

text
CONTEXT
1M context
INPUT
$0.1/M
OUTPUT
$0.5/M
ByteDance

Seedream 5.0 Lite

doubao-seedream-5-0-260128

Seedream 5.0 Lite provides image generation through PipeLLM. Use model ID doubao-seedream-5-0-260128 with /v1/images/generations.

imagetext
PRICE · USD
$0.029118 / image
ByteDance

Seedance 2.0 Fast

doubao-seedance-2-0-fast-260128

Seedance 2.0 Fast provides video generation through PipeLLM. Use model ID doubao-seedance-2-0-fast-260128 with /v2/videos.

videotext
PRICE · USD
480p: $4.080882 / million video tokens; 720p: $4.080882 / million video tokens
ByteDance

Seedance 2.5

doubao-seedance-2-5-260628

Seedance 2.5 provides video generation through PipeLLM. Use model ID doubao-seedance-2-5-260628 with /v2/videos.

videotext
PRICE · USD
480p: $10.294118 / million video tokens; 720p: $10.294118 / million video tokens; 1080p: $11.323529 / million video tokens
ByteDance

Seedance 2.0 Mini

doubao-seedance-2-0-mini-260615

Seedance 2.0 Mini provides video generation through PipeLLM. Use model ID doubao-seedance-2-0-mini-260615 with /v2/videos.

videotext
PRICE · USD
480p: $0.06 / output second; 720p: $0.12 / output second; 480p video reference: $0.0375 / input + output second; 720p video reference: $0.075 / input + output second
OpenAI

GPT-5.4 Nano

gpt-5.4-nano

GPT-5.4 Nano provides text generation through PipeLLM. Use model ID gpt-5.4-nano with /v1/chat/completions.

texttoolscache
CONTEXT
N/A
INPUT
$0.2/M
OUTPUT
$1.25/M
Google

Gemini 3.5 Flash Lite

gemini-3.5-flash-lite

Gemini 3.5 Flash Lite provides text generation through PipeLLM. Use model ID gemini-3.5-flash-lite with /v1/chat/completions.

texttoolscache
CONTEXT
1M context
INPUT
$0.3/M
OUTPUT
$2.5/M
Google

Gemini 3.1 Flash Image

gemini-3.1-flash-image

Gemini 3.1 Flash Image supports image generation and text through PipeLLM. Use model ID gemini-3.1-flash-image with the configured compatible endpoint.

textimage
CONTEXT
N/A
INPUT
$0.5/M
OUTPUT
$3/M
Cohere

Rerank v4.0 Pro

rerank-v4.0-pro

Rerank v4.0 Pro provides document reranking through PipeLLM. Use model ID rerank-v4.0-pro with /v2/rerank.

text
PRICE · USD
$0.0025 / request
Cohere

Rerank v4.0 Fast

rerank-v4.0-fast

Rerank v4.0 Fast provides document reranking through PipeLLM. Use model ID rerank-v4.0-fast with /v2/rerank.

text
PRICE · USD
$0.002 / request
ByteDance

Seedance 2.0

seedance-2.0

Seedance 2.0 generates video through PipeLLM. Use model seedance-2.0 with POST /v2/videos.

videotext
PRICE · USD
$0.48 / output second
ByteDance

Doubao Seed 2.0 Mini

doubao-seed-2-0-mini-260428

Doubao Seed 2.0 Mini provides text generation through PipeLLM. Use model ID doubao-seed-2-0-mini-260428 with /v3/chat/completions.

texttools
CONTEXT
25.6K context
INPUT
$0.8/M
OUTPUT
$8/M
OpenAI

GPT-5.2

gpt-5.2

GPT-5.2 provides text generation through PipeLLM. Use model ID gpt-5.2 with /v1/chat/completions.

texttoolscache
CONTEXT
N/A
INPUT
$1.75/M
OUTPUT
$14/M