Model catalogue
426 models from OpenRouter. Any of them can back a token; only backed ones can be talked to.
GPT-6 Astra
OpenAI
GPT-6 Astra is OpenAI's flagship model for demanding end-to-end work. It is suited for advanced analysis, software engineering, deep research, scientific work, and document creation, with particular strengths in long-hor...
Context
1.1M
In
$10.00/M
Out
$50.00/M
UnbackedLaunch for this →Claude Fable 5.1
Anthropic
Claude Fable 5.1 improves on Claude Fable 5 across the board, with the biggest gains in agentic coding, long-running agentic workflows, and knowledge work: long code refactors, front-end and visual...
Context
1.0M
In
$10.00/M
Out
$50.00/M
UnbackedLaunch for this →Gemma 4 26B A4B
Google
Gemma 4 26B A4B IT is an instruction-tuned Mixture-of-Experts (MoE) model from Google DeepMind. Despite 25.2B total parameters, only 3.8B activate per token during inference — delivering near-31B quality at...
Context
262K
In
$0.09/M
Out
$0.30/M
UnbackedLaunch for this →Grok 4.6
xAI
Grok 4.6 is SpaceXAI's smartest model with frontier performance on coding, knowledge work, and STEM.
Context
500K
In
$2.00/M
Out
$6.00/M
UnbackedLaunch for this →DeepSeek V4.1 Flash
DeepSeek
DeepSeek V4.1 Flash is a sparse mixture-of-experts model from DeepSeek, and the first built on the company's Causal Encoder-Decoder (CED) architecture. It activates 8B parameters on input and 16B on...
Context
1.0M
In
$0.30/M
Out
$1.20/M
Backed by TEST2Launch for this →Llama Guard 4 12B
Meta
Llama Guard 4 is a Llama 4 Scout-derived multimodal pretrained model, fine-tuned for content safety classification. Similar to previous versions, it can be used to classify content in both LLM...
Context
164K
In
$0.18/M
Out
$0.18/M
UnbackedLaunch for this →Claude Sonnet 5
Anthropic
Sonnet 5 is Anthropic's most capable Sonnet-class model, with frontier performance across coding, agents, and professional work. It supports adaptive thinking with selectable reasoning effort levels (low, medium, high, m...
Context
1.0M
In
$2.00/M
Out
$10.00/M
Backed by TESTLaunch for this →Claude Opus 5
Anthropic
Claude Opus 5 is Anthropic’s flagship model for demanding reasoning, coding, and long-horizon agentic work. It is particularly strong at end-to-end software tasks, code review and bug finding, visual analysis...
Context
1.0M
In
$5.00/M
Out
$25.00/M
UnbackedLaunch for this →Gemini 3.5 Flash
Google
Gemini 3.5 Flash is Google's high-efficiency multimodal model, bringing near-Pro level coding and reasoning at Flash-tier cost and speed. It is highly optimized for coding proficiency and parallel agentic execution...
Context
1.0M
In
$1.50/M
Out
$9.00/M
UnbackedLaunch for this →Qwen3.8 Flash
Qwen
Qwen3.8 Flash is a multimodal reasoning model from Alibaba. It is suited for coding assistance, agentic workflows, visual understanding, document and codebase analysis, desktop interaction, chart analysis, and long-video...
Context
1.0M
In
$0.15/M
Out
$0.47/M
UnbackedLaunch for this →Kimi K3
Moonshot
Kimi K3 is a 2.8T parameter open-weight multimodal reasoning model from Moonshot AI. It is suited for complex coding, knowledge work, and long-horizon agentic workflows, and is particularly strong at...
Context
1.0M
In
$2.65/M
Out
$13.28/M
UnbackedLaunch for this →GLM 5.3 Flash
Z.ai
GLM-5.3-Flash is a native multimodal model from Z.ai. It is suited for efficient coding and long-horizon agent tasks. Its hybrid sparse and linear attention architecture maintains accurate long-context behavior while...
Context
1.3M
In
$0.15/M
Out
$0.50/M
UnbackedLaunch for this →MiniMax M3
MiniMax
MiniMax-M3 is a multimodal foundation model from MiniMax. It supports text, image, and video inputs with text output, a 1M-token context window, and is suited for long-horizon agentic work, coding,...
Context
1.0M
In
$0.30/M
Out
$1.20/M
UnbackedLaunch for this →- SC
Schematron V2 Turbo
Inference Net
Schematron V2 Turbo is a 3B-parameter HTML-to-JSON extraction model from Inference.net. It prioritizes throughput for high-volume extraction workloads. Extraction instructions must be supplied through a JSON schema in re...
Context
128K
In
$0.03/M
Out
$0.15/M
UnbackedLaunch for this → - SC
Schematron V2 Small
Inference Net
Schematron V2 Small is a 3B-parameter HTML-to-JSON extraction model from Inference.net. It prioritizes extraction quality for complex schemas and long pages. Extraction instructions must be supplied through a JSON schema...
Context
128K
In
$0.05/M
Out
$0.23/M
UnbackedLaunch for this → - GP
GPT Astra Latest
~Openai
This model always redirects to the latest model in the GPT Astra family.
Context
1.1M
In
$10.00/M
Out
$50.00/M
UnbackedLaunch for this → - GP
GPT Sol Latest
~Openai
This model always redirects to the latest model in the GPT Sol family.
Context
1.1M
In
$2.00/M
Out
$10.00/M
UnbackedLaunch for this → - GP
GPT Terra Latest
~Openai
This model always redirects to the latest model in the GPT Terra family.
Context
1.1M
In
$2.00/M
Out
$12.00/M
UnbackedLaunch for this → - GP
GPT Luna Latest
~Openai
This model always redirects to the latest model in the GPT Luna family.
Context
1.1M
In
$0.20/M
Out
$1.20/M
UnbackedLaunch for this → - FU
Fugu Ultra v2
Sakana
Fugu Ultra v2 is the higher-performance model in Sakana AI's Fugu family. Rather than a single monolithic model, Fugu is a learned multi-agent orchestration system: a language model trained to...
Context
1.0M
In
$5.00/M
Out
$30.00/M
UnbackedLaunch for this → - FU
Fugu Max
Sakana
Fugu Max is the cost-performance model in Sakana AI's Fugu family. Rather than a single monolithic model, Fugu is a learned multi-agent orchestration system: a language model trained to route...
Context
1.0M
In
$2.00/M
Out
$6.00/M
UnbackedLaunch for this → - LI
Ling 3.0 Flash VL
Inclusionai
Ling 3.0 Flash VL builds on Ling 3.0 Flash (124B total / 5.5B active MoE from InclusionAI), further strengthening its language capabilities while adding native visual perception and advanced visual...
Context
131K
In
$0.06/M
Out
$0.18/M
UnbackedLaunch for this → - ME
Mercury 2.5
Inception
Mercury 2.5 is the fastest reasoning LLM, and the latest diffusion LLM (dLLM) from Inception. Instead of generating tokens sequentially, Mercury 2.5 produces and refines multiple tokens in parallel, achieving...
Context
260K
In
$0.04/M
Out
$0.15/M
UnbackedLaunch for this → GPT-6 Astra (batch)
OpenAI
GPT-6 Astra is OpenAI's flagship model for demanding end-to-end work. It is suited for advanced analysis, software engineering, deep research, scientific work, and document creation, with particular strengths in long-hor...
Context
1.1M
In
$5.00/M
Out
$25.00/M
UnbackedLaunch for this →GPT-6 Astra Pro
OpenAI
GPT-6 Astra Pro is the same underlying model as [GPT-6 Astra](https://openrouter.ai/openai/gpt-6-astra), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks. Learn more in OpenAI's do...
Context
1.1M
In
$10.00/M
Out
$50.00/M
UnbackedLaunch for this →GPT-6 Astra Pro (batch)
OpenAI
GPT-6 Astra Pro is the same underlying model as [GPT-6 Astra](https://openrouter.ai/openai/gpt-6-astra), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks. Learn more in OpenAI's do...
Context
1.1M
In
$5.00/M
Out
$25.00/M
UnbackedLaunch for this →Qwen3.8 Max (0902)
Qwen
Qwen3.8 Max 0902 is an updated snapshot of Qwen3.8 Max from Alibaba's Qwen team. It is a 2.4-trillion-parameter mixture-of-experts model that accepts text, image, and video input and returns text,...
Context
1.0M
In
$2.00/M
Out
$6.00/M
UnbackedLaunch for this →Muse Spark 1.3 Contributor
Meta
Muse Spark 1.3 Contributor is the cost-efficient contributor tier of Meta’s multimodal reasoning model for experimentation, learning, and early-stage agentic, multi-agent, and coding workflows. It is designed to track in...
Context
1.0M
In
$0.10/M
Out
$0.20/M
UnbackedLaunch for this →Muse Spark 1.3
Meta
Muse Spark 1.3 is a multimodal reasoning model from Meta for long-running agentic, multi-agent, and coding workflows. It is designed to keep track of information across extended tasks, work through...
Context
1.0M
In
$1.25/M
Out
$4.25/M
UnbackedLaunch for this →Gemini 3.8 Flash
Google
Gemini 3.8 Flash is Google's most intelligent Flash model with significant gains from 3.7 Flash across software engineering, agentic tasks, and multi-step reasoning.
Context
1.0M
In
$0.75/M
Out
$3.75/M
UnbackedLaunch for this →Gemini 3.8 Flash (batch)
Google
Gemini 3.8 Flash is Google's most intelligent Flash model with significant gains from 3.7 Flash across software engineering, agentic tasks, and multi-step reasoning.
Context
1.0M
In
$0.38/M
Out
$1.88/M
UnbackedLaunch for this →Claude Fable 5.1 (batch)
Anthropic
Claude Fable 5.1 improves on Claude Fable 5 across the board, with the biggest gains in agentic coding, long-running agentic workflows, and knowledge work: long code refactors, front-end and visual...
Context
1.0M
In
$5.00/M
Out
$25.00/M
UnbackedLaunch for this →- GR
Granite 4.2 8B
Ibm Granite
Granite 4.2 8B is a dense reasoning model from IBM. It is suited for mathematics, code generation, multilingual dialogue, and agentic workflows that need multi-step reasoning. It supports full, low-effort,...
Context
131K
In
$0.06/M
Out
$0.25/M
UnbackedLaunch for this → - HY
Hy4 preview
Tencent
Tencent: Hy4 preview is a mixture-of-experts model from Tencent, with 49B active parameters out of 770B total. It is designed for coding agents, complex tool-use workflows, and productivity tasks that...
Context
1.0M
In
$0.83/M
Out
$2.50/M
UnbackedLaunch for this → - LI
Ling 3.0 Flash Fin
Inclusionai
Ling 3.0 Flash Fin is a finance-focused mixture-of-experts model from InclusionAI, built on Ling 3.0 Flash with 5.1B active parameters out of 124B total. It is designed for real-world investment...
Context
262K
In
$0.06/M
Out
$0.18/M
UnbackedLaunch for this → - GL
GLM Flash Latest
~Z Ai
This model always redirects to the latest model in the GLM Flash family.
Context
1.3M
In
$0.07/M
Out
$0.25/M
UnbackedLaunch for this → GLM 5.3 Flash (batch)
Z.ai
GLM-5.3-Flash is a native multimodal model from Z.ai. It is suited for efficient coding and long-horizon agent tasks. Its hybrid sparse and linear attention architecture maintains accurate long-context behavior while...
Context
1.0M
In
$0.07/M
Out
$0.25/M
UnbackedLaunch for this →Muse Spark 1.2 Contributor
Meta
Muse Spark 1.2 contributor tier is a reasoning model from Meta designed for developers who want to start building at an even lower cost. It’s meaningfully cheaper than Muse Spark...
Context
1.0M
In
$0.10/M
Out
$0.20/M
UnbackedLaunch for this →DeepSeek V4 Flash Vision Exp
DeepSeek
DeepSeek V4 Flash Vision Exp is an experimental vision-enabled version of [DeepSeek V4 Flash 0731](https://openrouter.ai/deepseek/deepseek-v4-flash-0731) from DeepSeek, adding image understanding while matching the base ...
Context
1.0M
In
$0.22/M
Out
$0.66/M
UnbackedLaunch for this →DeepSeek V4 Flash Vision Exp (batch)
DeepSeek
DeepSeek V4 Flash Vision Exp is an experimental vision-enabled version of [DeepSeek V4 Flash 0731](https://openrouter.ai/deepseek/deepseek-v4-flash-0731) from DeepSeek, adding image understanding while matching the base ...
Context
1.0M
In
$0.11/M
Out
$0.33/M
UnbackedLaunch for this →