GoogleGoogle 4.6TrustPilotTrustPilot 4.9

Every model, unlimited possibilities

Frontier and open models, pre-warmed and ready. Route to any of them from a single Claws endpoint, swap providers without rewriting a line. Here’s what’s powering agents this week.

445 models online59 providers312ms median P95
All models
Ranked by live usage across every hosted agent this week. Filter by capability, every model routes through the same single endpoint.
OpenAI: GPT Astra Latest
Community
NEW
gpt-astra-latest
1 Credit = 1,000 Tokens
Long ContextVision

This model always redirects to the latest model in the GPT Astra family.

OpenAI: GPT-6 Astra
OpenAI
NEW
gpt-6-astra
1 Credit = 1,000 Tokens
Long ContextVision

GPT-6 Astra is OpenAI's flagship model for demanding end-to-end work.

Anthropic: Claude Fable 5.1
Anthropic
NEW
claude-fable-5.1
1 Credit = 1,000 Tokens
Long ContextVision

Claude Fable 5.1 improves on Claude Fable 5 across the board, with the biggest gains in agentic coding, long-running agentic workflows, a…

Anthropic: Claude Fable Latest
Community
claude-fable-latest
1 Credit = 1,000 Tokens
Long ContextVision

This model always redirects to the latest model in the Claude Fable family.

Sakana: Fugu Ultra v2
Community
NEW
fugu-ultra-v2
1 Credit = 2,000 Tokens
Long ContextVision

Fugu Ultra v2 is the higher-performance model in Sakana AI's Fugu family.

Qwen: Qwen3.8 Max (0902)
Alibaba
NEW
qwen3.8-max-0902
1 Credit = 5,000 Tokens
Long ContextReasoning

Qwen3.8 Max 0902 is an updated snapshot of Qwen3.8 Max from Alibaba's Qwen team.

Meta: Muse Spark 1.3
Meta
NEW
muse-spark-1.3
1 Credit = 8,000 Tokens
Long ContextVision

Muse Spark 1.3 is a multimodal reasoning model from Meta for long-running agentic, multi-agent, and coding workflows.

SpaceXAI: Grok 4.6
xAI
grok-4.6
1 Credit = 5,000 Tokens
Long ContextReasoning

Grok 4.6 is SpaceXAI's smartest model with frontier performance on coding, knowledge work, and STEM.

Google: Gemini 3.8 Flash
Google
NEW
gemini-3.8-flash
1 Credit = 13,333 Tokens
Long ContextVision

Gemini 3.8 Flash is Google's most intelligent Flash model with significant gains from 3.7 Flash across software engineering, agentic task…

Z.ai: GLM 5.3
Z.ai
NEW
glm-5.3
1 Credit = 7,143 Tokens
Long Context

GLM-5.3 is a large-scale reasoning model from Z.ai, built for complex software engineering and long-horizon agent tasks.

MoonshotAI: Kimi K3 (batch)
Community
kimi-k3:batch
1 Credit = 3,333 Tokens
Long ContextVision

Kimi K3 is a 2.8T parameter open-weight multimodal reasoning model from Moonshot AI.

Tencent: Hy4 preview
Tencent
NEW
hy4-preview
1 Credit = 11,990 Tokens
Long Context

Tencent: Hy4 preview is a mixture-of-experts model from Tencent, with 49B active parameters out of 770B total.

DeepSeek: DeepSeek V4.1 Flash
DeepSeek
NEW
deepseek-v4.1-flash
1 Credit = 66,667 Tokens
Long ContextVision

DeepSeek V4.1 Flash is a sparse mixture-of-experts model from DeepSeek, and the first built on the company's Causal Encoder-Decoder (CED)…

Inference.net: Schematron V2 Small
Community
NEW
schematron-v2-small
1 Credit = 200,000 Tokens
Mid Context

Schematron V2 Small is a 3B-parameter HTML-to-JSON extraction model from Inference.net.

inclusionAI: Ling 3.0 Flash VL
Community
NEW
ling-3.0-flash-vl
1 Credit = 166,667 Tokens
Mid ContextVision

Ling 3.0 Flash VL builds on Ling 3.0 Flash (124B total / 5.5B active MoE from InclusionAI), further strengthening its language capabiliti…

Inception: Mercury 2.5
Community
NEW
mercury-2.5
1 Credit = 250,000 Tokens
Long Context

Mercury 2.5 is the fastest reasoning LLM, and the latest diffusion LLM (dLLM) from Inception.

Nex AGI: Nex-N2.5-Mini (free)
Community
NEW
nex-n2.5-mini:free
Pricing on request
Long ContextVision

Nex-N2.5 is an agentic model built to turn goals into working, verified outcomes.

Z.ai: GLM Latest
Community
NEW
glm-latest
1 Credit = 11,459 Tokens
Long Context

This model always redirects to the latest GLM model from Z.ai.

AionLabs: Aion-3.0
Community
aion-3.0
1 Credit = 3,333 Tokens
Mid Context

Aion-3.0 is a multi-model roleplaying and storytelling system from AionLabs, built on the GLM family of models.

IBM: Granite 4.2 8B
Community
NEW
granite-4.2-8b
1 Credit = 166,667 Tokens
Mid Context

Granite 4.2 8B is a dense reasoning model from IBM.

ByteDance Seed: Seed 2.1 Turbo
Community
seed-2-1-turbo
1 Credit = 20,000 Tokens
Long ContextVision

Seed 2.1 Turbo is a multimodal model from ByteDance Seed for coding and long-horizon agent workflows.

xAI: Grok Latest
Community
grok-latest
1 Credit = 5,000 Tokens
Long ContextVision

This model always redirects to the latest Grok model from xAI.

Dots Studio: Dots3-Note Preview (free)
Community
NEW
dots-3-note-preview:free
Pricing on request
Long ContextVision

Dots3-Note Preview is an open-weight mixture-of-experts model from Dots Studio, with 16B active parameters out of 280B total.

NVIDIA: Nemotron 3.5 Lightning
Community
nemotron-3.5-lightning
1 Credit = 125,000 Tokens
Long Context

NVIDIA Nemotron 3.5 Lightning is an open mixture-of-experts model from NVIDIA, with 3B active parameters out of 30B total.

Upstage: Solar Pro 4
Community
solar-pro4
1 Credit = 111,111 Tokens
Long ContextReasoning

Solar Pro 4 is Upstage's cost-efficient large language model, featuring a 524K context window.

LiquidAI: LFM2.5-2.6B (free)
Community
lfm-2.5-2.6b:free
Pricing on request
Mid Context

LFM2.5-2.6B is a compact reasoning model from Liquid AI.

Thinking Machines: Inkling Small (batch)
Community
inkling-small:batch
1 Credit = 20,000 Tokens
Long ContextReasoning

Inkling Small is an open-weight multimodal mixture-of-experts model from Thinking Machines Lab, with 12B active parameters out of 276B to…

DeepSeek: DeepSeek V4 Flash Latest
Community
deepseek-v4-flash-latest
1 Credit = 333,333 Tokens
Long Context

This model always redirects to the latest model in the DeepSeek V4 Flash family.

Meituan: LongCat 2.0
Community
longcat-2.0
1 Credit = 33,333 Tokens
Long Context

LongCat 2.0 is a sparse mixture-of-experts language model from Meituan, with 48B active parameters out of 1.6T total.

Kwaipilot: KAT-Coder-Pro V2.5
Community
kat-coder-pro-v2.5
1 Credit = 13,514 Tokens
Long ContextCode

KAT-Coder-Pro V2.5 is a flagship-level Agentic Coding model that can directly hand over an entire issue or an entire business workflow to…

Poolside: Laguna S 2.1
Community
laguna-s-2.1
1 Credit = 111,111 Tokens
Long Context

Laguna S 2.1 is the latest coding agent model from [Poolside](<https://poolside.ai/>).

Cohere: North Mini Code (free)
Cohere
north-mini-code:free
Pricing on request
Long ContextCode

North Mini Code is Cohere's first agentic coding model and the debut of its North family.

MoonshotAI: Kimi Latest
Community
kimi-latest
1 Credit = 4,211 Tokens
Long ContextVision

This model always redirects to the latest model in the Kimi family.

Google: Gemini Pro Latest
Community
gemini-pro-latest
1 Credit = 5,000 Tokens
Long ContextReasoning

This model always redirects to the latest model in the Gemini Pro family.

MiniMax: MiniMax M3
Community
minimax-m3
1 Credit = 33,333 Tokens
Long ContextVision

MiniMax-M3 is a multimodal foundation model from MiniMax.

StepFun: Step 3.7 Flash
Community
step-3.7-flash
1 Credit = 50,000 Tokens
Long ContextVision

Step 3.7 Flash is StepFun's latest high-efficiency multimodal Mixture-of-Experts model.

Mistral: Mistral Medium 3.5
Mistral
mistral-medium-3-5
1 Credit = 6,667 Tokens
Long ContextVision

Mistral Medium 3.5 is a dense 128B instruction-following model from Mistral AI.

Perceptron: Perceptron Mk1
Community
perceptron-mk1
1 Credit = 66,667 Tokens
Mid ContextVision

Perceptron Mk1 (Mark One) is Perceptron's highest-quality vision-language model for video and embodied reasoning.** It accepts image and…

Xiaomi: MiMo-V2.5-Pro
Xiaomi
mimo-v2.5-pro
1 Credit = 22,989 Tokens
Long Context

MiMo-V2.5-Pro is Xiaomi’s flagship model, delivering strong performance in general agentic capabilities, complex software engineering, an…

Arcee AI: Trinity Large Thinking
Community
trinity-large-thinking
1 Credit = 40,000 Tokens
Long ContextReasoning

Trinity Large Thinking is a powerful open source reasoning model from the team at Arcee AI.

Reka Edge
Community
reka-edge
1 Credit = 100,000 Tokens
Vision

Reka Edge is an extremely efficient 7B multimodal vision-language model that accepts image/video+text inputs and generates text outputs.

Writer: Palmyra X5
Community
palmyra-x5
1 Credit = 16,667 Tokens
Long Context

Palmyra X5 is Writer's most advanced model, purpose-built for building and scaling AI agents across the enterprise.

Free Models Router
Community
free
Pricing on request
Long ContextVision

The simplest way to get free inference.

Perplexity: Sonar Pro Search
Perplexity
sonar-pro-search
1 Credit = 3,333 Tokens
Long ContextVision

Exclusively available on the OpenRouter API, Sonar Pro's new Pro Search mode is Perplexity's most advanced agentic search system.

Relace: Relace Search
Community
relace-search
1 Credit = 10,000 Tokens
Long Context

The relace-search model uses 4-12 `view_file` and `grep` tools in parallel to explore a codebase and return relevant files to the user re…

Amazon: Nova Premier 1.0
Community
nova-premier-v1
1 Credit = 4,000 Tokens
Long ContextVision

Amazon Nova Premier is the most capable of Amazon’s multimodal models for complex reasoning tasks and for use as the best teacher for dis…

Magnum v4 72B
Community
magnum-v4-72b
1 Credit = 4,000 Tokens
Mid Context

This is a series of models designed to replicate the prose quality of the Claude 3 models, specifically Sonnet(https://openrouter.ai/anth…

TheDrummer: Cydonia 24B V4.1
Community
cydonia-24b-v4.1
1 Credit = 33,333 Tokens
Mid Context

Uncensored and creative writing model based on Mistral Small 3.2 24B with good recall, prompt adherence, and intelligence.

Nous: Hermes 4 405B
Community
hermes-4-405b
1 Credit = 10,000 Tokens
Mid Context

Hermes 4 is a large-scale reasoning model built on Meta-Llama-3.1-405B and released by Nous Research.

Morph: Morph V3 Large
Community
morph-v3-large
1 Credit = 11,111 Tokens
Long Context

Morph's high-accuracy apply model for complex code edits.

Sao10K: Llama 3.1 Euryale 70B v2.2
Community
l3.1-euryale-70b
1 Credit = 11,765 Tokens
Mid ContextReasoning

Euryale L3.1 70B v2.2 is a model focused on creative roleplay from [Sao10k](https://ko-fi.com/sao10k).

WizardLM-2 8x22B
Microsoft
wizardlm-2-8x22b
1 Credit = 16,129 Tokens
Mid Context

WizardLM-2 8x22B is Microsoft AI's most advanced Wizard model.

Baidu: ERNIE 4.5 VL 424B A47B
Community
ernie-4.5-vl-424b-a47b
1 Credit = 23,810 Tokens
Mid ContextVision

ERNIE-4.5-VL-424B-A47B is a multimodal Mixture-of-Experts (MoE) model from Baidu’s ERNIE 4.5 series, featuring 424B total parameters with…

Meta: Llama 3.1 70B Instruct
Meta
llama-3.1-70b-instruct
1 Credit = 25,000 Tokens
Mid Context

Meta's latest class of model (Llama 3.1) launched with a variety of sizes & flavors.

Mancer: Weaver (alpha)
Community
weaver
1 Credit = 25,000 Tokens
Mid Context

An attempt to recreate Claude-style verbosity, but don't expect the same level of coherence or memory.

ReMM SLERP 13B
Community
remm-slerp-l2-13b
1 Credit = 28,571 Tokens
Mid Context

A recreation trial of the original MythoMax-L2-B13 but with updated models.

Venice: Uncensored
Community
dolphin-mistral-24b-venice-edition
1 Credit = 50,000 Tokens
Mid Context

Venice Uncensored Dolphin Mistral 24B Venice Edition is a fine-tuned variant of Mistral-Small-24B-Instruct-2501, developed by dphn.ai in…

ByteDance: UI-TARS 7B
Community
ui-tars-1.5-7b
1 Credit = 100,000 Tokens
Mid ContextVision

UI-TARS-1.5 is a multimodal vision-language agent optimized for GUI-based environments, including desktop interfaces, web browsers, mobil…

MythoMax 13B
Community
mythomax-l2-13b
1 Credit = 166,667 Tokens
Mid Context

One of the highest performing and most popular fine-tunes of Llama 2 13B, with rich descriptions and roleplay.

OpenAI: GPT Sol Latest
Community
NEW
gpt-sol-latest
1 Credit = 5,000 Tokens
Long ContextVision

This model always redirects to the latest model in the GPT Sol family.

OpenAI: GPT-6 Astra Pro
OpenAI
NEW
gpt-6-astra-pro
1 Credit = 1,000 Tokens
Long ContextVision

GPT-6 Astra Pro is the same underlying model as [GPT-6 Astra](https://openrouter.ai/openai/gpt-6-astra), served with `reasoning.mode` set…

Anthropic: Claude Fable 5
Anthropic
claude-fable-5
1 Credit = 1,000 Tokens
Long ContextVision

Claude Fable 5 is a Mythos-class model from Anthropic, built for autonomous knowledge work and coding.

Anthropic: Claude Opus Latest
Community
claude-opus-latest
1 Credit = 2,000 Tokens
Long ContextReasoning

This model always redirects to the latest model in the Claude Opus family.

Sakana: Fugu Max
Community
NEW
fugu-max
1 Credit = 5,000 Tokens
Long ContextVision

Fugu Max is the cost-performance model in Sakana AI's Fugu family.

Qwen: Qwen3.8 2.4T A95B
Alibaba
qwen3.8-2.4t-a95b
1 Credit = 5,000 Tokens
Long Context

Qwen3.8 2.4T A95B is an open-weight sparse mixture-of-experts model from Qwen and the open-weight variant of [Qwen3.8 Max](/qwen/qwen3.8-…

Meta: Muse Spark 1.3 Contributor
Meta
NEW
muse-spark-1.3-contributor
1 Credit = 100,000 Tokens
Long ContextVision

Muse Spark 1.3 Contributor is the cost-efficient contributor tier of Meta’s multimodal reasoning model for experimentation, learning, and…

SpaceXAI: Grok 4.5
xAI
grok-4.5
1 Credit = 5,000 Tokens
Long ContextReasoning

Grok 4.5 is a model from SpaceXAI with frontier performance on coding, knowledge work, and STEM.

Google: Gemini 3.8 Flash (batch)
Google
NEW
gemini-3.8-flash:batch
1 Credit = 26,667 Tokens
Long ContextVision

Gemini 3.8 Flash is Google's most intelligent Flash model with significant gains from 3.7 Flash across software engineering, agentic task…

Z.ai: GLM 5.3 (batch)
Z.ai
NEW
glm-5.3:batch
1 Credit = 14,286 Tokens
Long Context

GLM-5.3 is a large-scale reasoning model from Z.ai, built for complex software engineering and long-horizon agent tasks.

MoonshotAI: Kimi K3
Community
kimi-k3
1 Credit = 3,776 Tokens
Long ContextVision

Kimi K3 is a 2.8T parameter open-weight multimodal reasoning model from Moonshot AI.

Tencent: Hy-MT2-30B-A3B
Tencent
NEW
hy-mt2-30b-a3b
1 Credit = 135,135 Tokens
Mid Context

Hy-MT2-30B-A3B is Tencent's flagship translation model in the Hy-MT2 family.

DeepSeek: DeepSeek V4 Flash Vision Exp
DeepSeek
NEW
deepseek-v4-flash-vision-exp
1 Credit = 45,455 Tokens
Long ContextVision

DeepSeek V4 Flash Vision Exp is an experimental vision-enabled version of [DeepSeek V4 Flash 0731](https://openrouter.ai/deepseek/deepsee…

Inference.net: Schematron V2 Turbo
Community
NEW
schematron-v2-turbo
1 Credit = 333,333 Tokens
Mid Context

Schematron V2 Turbo is a 3B-parameter HTML-to-JSON extraction model from Inference.net.

inclusionAI: Ling 3.0 Flash VL (free)
Community
NEW
ling-3.0-flash-vl:free
Pricing on request
Long ContextVision

Ling 3.0 Flash VL builds on Ling 3.0 Flash (124B total / 5.5B active MoE from InclusionAI), further strengthening its language capabiliti…

Inception: Mercury 2
Community
mercury-2
1 Credit = 40,000 Tokens
Mid Context

Mercury 2 is an extremely fast reasoning LLM, and the first reasoning diffusion LLM (dLLM).

Nex AGI: Nex-N2.5-Pro (free)
Community
NEW
nex-n2.5-pro:free
Pricing on request
Long ContextVision

Nex-N2.5 is an agentic model built to turn goals into working, verified outcomes.

Z.ai: GLM Flash Latest
Community
NEW
glm-flash-latest
1 Credit = 133,333 Tokens
Long ContextVision

This model always redirects to the latest model in the GLM Flash family.

AionLabs: Aion-3.0-Mini
Community
aion-3.0-mini
1 Credit = 14,286 Tokens
Mid Context

Aion-3.0 Mini is a multi-model roleplaying and storytelling system from AionLabs, built on the DeepSeek family of models.

IBM: Granite 4.0 Micro
Community
granite-4.0-h-micro
1 Credit = 588,235 Tokens
Mid Context

Granite-4.0-H-Micro is a 3B parameter from the Granite 4 family of models.

ByteDance Seed: Seed-2.0-Code
Community
seed-2.0-code
1 Credit = 20,000 Tokens
Long ContextVision

Seed 2.0 Code is a model from ByteDance Seed optimized for agentic coding.

NVIDIA: Nemotron 3.5 Lightning (free)
Community
nemotron-3.5-lightning:free
Pricing on request
Long Context

NVIDIA Nemotron 3.5 Lightning is an open mixture-of-experts model from NVIDIA, with 3B active parameters out of 30B total.

Upstage: Solar Pro 3
Community
solar-pro-3
1 Credit = 66,667 Tokens
Mid Context

Solar Pro 3 is Upstage's powerful Mixture-of-Experts (MoE) language model.

Thinking Machines: Inkling Small
Community
inkling-small
1 Credit = 22,222 Tokens
Long ContextReasoning

Inkling Small is an open-weight multimodal mixture-of-experts model from Thinking Machines Lab, with 12B active parameters out of 276B to…

Kwaipilot: KAT-Coder-Pro V2
Community
kat-coder-pro-v2
1 Credit = 33,333 Tokens
Long ContextCode

KAT-Coder-Pro V2 is the latest high-performance model in KwaiKAT’s KAT-Coder series, designed for complex enterprise-grade software engin…

Poolside: Laguna S 2.1 (free)
Community
laguna-s-2.1:free
Pricing on request
Long Context

Laguna S 2.1 is the latest coding agent model from [Poolside](<https://poolside.ai/>).

Cohere: Command A
Cohere
command-a
1 Credit = 4,000 Tokens
Long ContextCode

Command A is an open-weights 111B parameter model with a 256k context window focused on delivering great performance across agentic, mult…

Google: Gemini Flash Latest
Community
gemini-flash-latest
1 Credit = 13,333 Tokens
Long ContextVision

This model always redirects to the latest model in the Gemini Flash family.

MiniMax: MiniMax M3 (batch)
Community
minimax-m3:batch
1 Credit = 33,333 Tokens
Long ContextVision

MiniMax-M3 is a multimodal foundation model from MiniMax.

StepFun: Step 3.5 Flash
Community
step-3.5-flash
1 Credit = 100,000 Tokens
Long Context

Step 3.5 Flash is StepFun's most capable open-source foundation model.

Mistral: Mistral Medium 3.5 (batch)
Mistral
mistral-medium-3-5:batch
1 Credit = 13,333 Tokens
Long ContextVision

Mistral Medium 3.5 is a dense 128B instruction-following model from Mistral AI.

Xiaomi: MiMo-V2.5
Xiaomi
mimo-v2.5
1 Credit = 71,429 Tokens
Long ContextVision

MiMo-V2.5 is a native omnimodal model by Xiaomi.

Reka Flash 3
Community
reka-flash-3
1 Credit = 100,000 Tokens
Mid Context

Reka Flash 3 is a general-purpose, instruction-tuned large language model with 21 billion parameters, developed by Reka.

Auto Router (Beta)
Community
auto-beta
Pricing on request
Long ContextVision

Auto Router (Beta) is a task-aware router from OpenRouter.

Perplexity: Sonar Pro
Perplexity
sonar-pro
1 Credit = 3,333 Tokens
Long ContextVision

Note: Sonar Pro pricing includes Perplexity search pricing.

Relace: Relace Apply 3
Community
relace-apply-3
1 Credit = 11,765 Tokens
Long Context

Relace Apply 3 is a specialized code-patching LLM that merges AI-suggested edits straight into your source files.

Amazon: Nova 2 Lite
Community
nova-2-lite-v1
1 Credit = 33,333 Tokens
Long ContextVision

Nova 2 Lite is a fast, cost-effective reasoning model for everyday workloads that can process text, images, and videos to generate text.

TheDrummer: Skyfall 36B V2
Community
skyfall-36b-v2
1 Credit = 18,182 Tokens
Mid Context

Skyfall 36B v2 is an enhanced iteration of Mistral Small 2501, specifically fine-tuned for improved creativity, nuanced writing, role-pla…

Nous: Hermes 3 405B Instruct
Community
hermes-3-llama-3.1-405b
1 Credit = 10,000 Tokens
Mid Context

Hermes 3 is a generalist language model with many improvements over Hermes 2, including advanced agentic capabilities, much better rolepl…

Morph: Morph V3 Fast
Community
morph-v3-fast
1 Credit = 12,500 Tokens
Mid Context

Morph's fastest apply model for code edits.

Sao10K: Llama 3.3 Euryale 70B
Community
l3.3-euryale-70b
1 Credit = 15,385 Tokens
Mid ContextReasoning

Euryale L3.3 70B is a model focused on creative roleplay from [Sao10k](https://ko-fi.com/sao10k).

Microsoft: Phi 4
Microsoft
phi-4
1 Credit = 142,857 Tokens
Mid Context

[Microsoft Research](/microsoft) Phi-4 is designed to perform well in complex reasoning tasks and can operate efficiently in situations w…

Meta: Llama 4 Maverick
Meta
llama-4-maverick
1 Credit = 50,000 Tokens
Long ContextVision

Llama 4 Maverick 17B Instruct (128E) is a high-capacity multimodal language model from Meta, built on a mixture-of-experts (MoE) architec…

OpenAI: GPT Terra Latest
Community
NEW
gpt-terra-latest
1 Credit = 5,000 Tokens
Long ContextVision

This model always redirects to the latest model in the GPT Terra family.

OpenAI: GPT-5.5 Pro
OpenAI
gpt-5.5-pro
1 Credit = 333 Tokens
Long ContextVision

GPT-5.5 Pro is OpenAI’s high-capability model optimized for deep reasoning and accuracy on complex, high-stakes workloads.

Anthropic: Claude Fable 5.1 (batch)
Anthropic
NEW
claude-fable-5.1:batch
1 Credit = 2,000 Tokens
Long ContextVision

Claude Fable 5.1 improves on Claude Fable 5 across the board, with the biggest gains in agentic coding, long-running agentic workflows, a…

Anthropic: Claude Sonnet Latest
Community
claude-sonnet-latest
1 Credit = 5,000 Tokens
Long ContextReasoning

This model always redirects to the latest model in the Claude Sonnet family.

Sakana: Fugu Ultra
Community
fugu-ultra
1 Credit = 2,000 Tokens
Long ContextVision

Fugu Ultra is the higher-performance model in Sakana AI's Fugu family.

Qwen: Qwen3.8 2.4T A95B (batch)
Alibaba
qwen3.8-2.4t-a95b:batch
1 Credit = 5,000 Tokens
Long Context

Qwen3.8 2.4T A95B is an open-weight sparse mixture-of-experts model from Qwen and the open-weight variant of [Qwen3.8 Max](/qwen/qwen3.8-…

Meta: Muse Spark 1.2
Meta
muse-spark-1.2
1 Credit = 8,000 Tokens
Long ContextVision

Muse Spark 1.2 is a reasoning model from Meta, designed for complex agentic tasks.

SpaceXAI: Grok Build 0.1
xAI
grok-build-0.1
1 Credit = 10,000 Tokens
Long ContextVision

Grok Build 0.1 is SpaceXAI’s fast coding model trained specifically for agentic software engineering workflows.

Google: Gemini 3.7 Flash
Google
gemini-3.7-flash
1 Credit = 13,333 Tokens
Long ContextVision

Gemini 3.7 Flash is a multimodal model from Google for fast agentic workflows, coding, and complex multi-step reasoning.

Z.ai: GLM 5.3 Flash
Z.ai
NEW
glm-5.3-flash
1 Credit = 66,667 Tokens
Long ContextVision

GLM-5.3-Flash is a native multimodal model from Z.ai.

MoonshotAI: Kimi K2.7 Code
Community
kimi-k2.7-code
1 Credit = 14,085 Tokens
Long ContextVision

MoonshotAI: Kimi K2.7 Code is a coding-focused model in Moonshot AI's Kimi K2 family, built to complete end-to-end programming tasks reli…

Tencent: Hy-MT2-1.8B
Tencent
NEW
hy-mt2-1.8b
1 Credit = 227,273 Tokens
Mid Context

Hy-MT2-1.8B is a compact 1.8B-parameter translation model from Tencent.

DeepSeek: DeepSeek V4 Pro 0813 (batch)
DeepSeek
deepseek-v4-pro-0813:batch
1 Credit = 15,152 Tokens
Long Context

DeepSeek V4 Pro 0813 is a large-scale mixture-of-experts model from DeepSeek.

inclusionAI: Ling 3.0 Flash Sante (free)
Community
NEW
ling-3.0-flash-sante:free
Pricing on request
Long Context

Ling 3.0 Flash Sante is a health and medicine-focused mixture-of-experts model from InclusionAI, built on Ling 3.0 Flash with 5.1B active…

AionLabs: Aion-2.0
Community
aion-2.0
1 Credit = 12,500 Tokens
Mid Context

Aion-2.0 is a variant of DeepSeek V3.2 optimized for immersive roleplaying and storytelling.

ByteDance Seed: Seed-2.0-Lite
Community
seed-2.0-lite
1 Credit = 40,000 Tokens
Long ContextVision

Seed-2.0-Lite is a versatile, cost‑efficient enterprise workhorse that delivers strong multimodal and agent capabilities while offering n…

NVIDIA: Nemotron 3 Ultra
Community
nemotron-3-ultra-550b-a55b
1 Credit = 16,000 Tokens
Long Context

NVIDIA Nemotron 3 Ultra is an open frontier-reasoning and orchestration model from NVIDIA, with 55B active parameters out of 550B total (…

Thinking Machines: Inkling
Community
inkling
1 Credit = 10,000 Tokens
Long ContextReasoning

Inkling is an open-weight multimodal mixture-of-experts model from Thinking Machines Lab, with 41B active parameters out of 975B total.

Poolside: Laguna XS 2.1
Community
laguna-xs-2.1
1 Credit = 166,667 Tokens
Long Context

Laguna XS 2.1 is the latest coding agent model in the 33B-A3B category from [Poolside](https://poolside.ai/) and a step forward from thei…

Cohere: Command R+ (08-2024)
Cohere
command-r-plus-08-2024
1 Credit = 4,000 Tokens
Mid ContextCode

command-r-plus-08-2024 is an update of the [Command R+](/models/cohere/command-r-plus) with roughly 50% higher throughput and 25% lower l…

MiniMax: MiniMax M2.7
Community
minimax-m2.7
1 Credit = 33,333 Tokens
Long Context

MiniMax-M2.7 is a next-generation large language model designed for autonomous, real-world productivity and continuous improvement.

Mistral: Mistral Small 4
Mistral
mistral-small-2603
1 Credit = 66,667 Tokens
Long ContextVision

Mistral Small 4 is the next major release in the Mistral Small family, unifying the capabilities of several flagship Mistral models into…

OpenRouter: Fusion
Community
fusion
Pricing on request
Long Context

Fusion turns your prompt into a small multi-model deliberation.

Perplexity: Sonar Reasoning Pro
Perplexity
sonar-reasoning-pro
1 Credit = 5,000 Tokens
Mid ContextReasoning

Note: Sonar Pro pricing includes Perplexity search pricing.

Amazon: Nova Pro 1.0
Community
nova-pro-v1
1 Credit = 12,500 Tokens
Long ContextVision

Amazon Nova Pro 1.0 is a capable multimodal model from Amazon focused on providing a combination of accuracy, speed, and cost for a wide…

TheDrummer: UnslopNemo 12B
Community
unslopnemo-12b
1 Credit = 25,000 Tokens
Long Context

UnslopNemo v4.1 is the latest addition from the creator of Rocinante, designed for adventure writing and role-play scenarios.

Nous: Hermes 3 70B Instruct
Community
hermes-3-llama-3.1-70b
1 Credit = 14,286 Tokens
Mid Context

Hermes 3 is a generalist language model with many improvements over [Hermes 2](/models/nousresearch/nous-hermes-2-mistral-7b-dpo), includ…

Sao10K: Llama 3 8B Lunaris
Community
l3-lunaris-8b
1 Credit = 250,000 Tokens
Reasoning

Lunaris 8B is a versatile generalist and roleplaying model based on Llama 3.

Meta: Llama Guard 4 12B
Meta
llama-guard-4-12b
1 Credit = 55,556 Tokens
Mid ContextVision

Llama Guard 4 is a Llama 4 Scout-derived multimodal pretrained model, fine-tuned for content safety classification.

OpenAI: GPT Luna Latest
Community
NEW
gpt-luna-latest
1 Credit = 50,000 Tokens
Long ContextVision

This model always redirects to the latest model in the GPT Luna family.

OpenAI: GPT-5.5 Pro (batch)
OpenAI
gpt-5.5-pro:batch
1 Credit = 667 Tokens
Long ContextVision

GPT-5.5 Pro is OpenAI’s high-capability model optimized for deep reasoning and accuracy on complex, high-stakes workloads.

Claude Opus 5
Anthropic
claude-opus-5
1 Credit = 2,000 Tokens
Long ContextReasoning

Claude Opus 5 is Anthropic’s flagship model for demanding reasoning, coding, and long-horizon agentic work.

Anthropic: Claude Haiku Latest
Community
claude-haiku-latest
1 Credit = 10,000 Tokens
Long ContextVision

This model always redirects to the latest model in the Claude Haiku family.

Sakana: Sakana Namazu
Community
sakana-namazu
1 Credit = 10,526 Tokens
Long ContextVision

Sakana Namazu is a Japanese-specialized reasoning model from Sakana AI, based on Kimi K2.6 with additional training for Japanese language…

Qwen: Qwen3.8 Flash
Alibaba
NEW
qwen3.8-flash
1 Credit = 66,667 Tokens
Long ContextVision

Qwen3.8 Flash is a multimodal reasoning model from Alibaba.

Meta: Muse Spark 1.2 Contributor
Meta
NEW
muse-spark-1.2-contributor
1 Credit = 100,000 Tokens
Long ContextVision

Muse Spark 1.2 contributor tier is a reasoning model from Meta designed for developers who want to start building at an even lower cost.

SpaceXAI: Grok 4.3
xAI
grok-4.3
1 Credit = 8,000 Tokens
Long ContextReasoning

Grok 4.3 is a reasoning model from SpaceXAI.

Google: Gemini 3.7 Flash (batch)
Google
gemini-3.7-flash:batch
1 Credit = 26,667 Tokens
Long ContextVision

Gemini 3.7 Flash is a multimodal model from Google for fast agentic workflows, coding, and complex multi-step reasoning.

Z.ai: GLM 5.3 Flash (batch)
Z.ai
NEW
glm-5.3-flash:batch
1 Credit = 133,333 Tokens
Long ContextVision

GLM-5.3-Flash is a native multimodal model from Z.ai.

MoonshotAI: Kimi K2.6
Community
kimi-k2.6
1 Credit = 10,526 Tokens
Long ContextVision

Kimi K2.6 is Moonshot AI's next-generation multimodal model, designed for long-horizon coding, coding-driven UI/UX generation, and multi-…

Tencent: Hy-MT2-7B
Tencent
NEW
hy-mt2-7b
1 Credit = 135,135 Tokens
Mid Context

Hy-MT2-7B is a 7B-parameter translation model from Tencent.

DeepSeek: DeepSeek V4 Pro 0813
DeepSeek
deepseek-v4-pro-0813
1 Credit = 17,257 Tokens
Long Context

DeepSeek V4 Pro 0813 is a large-scale mixture-of-experts model from DeepSeek.

inclusionAI: Ling 3.0 Flash Fin
Community
NEW
ling-3.0-flash-fin
1 Credit = 166,667 Tokens
Long Context

Ling 3.0 Flash Fin is a finance-focused mixture-of-experts model from InclusionAI, built on Ling 3.0 Flash with 5.1B active parameters ou…

AionLabs: Aion-RP 1.0 (8B)
Community
aion-rp-llama-3.1-8b
1 Credit = 12,500 Tokens
Mid Context

Aion-RP-Llama-3.1-8B ranks the highest in the character evaluation portion of the RPBench-Auto benchmark, a roleplaying-specific variant…

ByteDance Seed: Seed-2.0-Mini
Community
seed-2.0-mini
1 Credit = 100,000 Tokens
Long ContextVision

Seed-2.0-mini targets latency-sensitive, high-concurrency, and cost-sensitive scenarios, emphasizing fast response and flexible inference…

NVIDIA: Nemotron 3.5 Content Safety
Community
nemotron-3.5-content-safety
1 Credit = 50,000 Tokens
Mid ContextVision

NVIDIA Nemotron 3.5 Content Safety is a compact 4B-parameter multimodal guardrail model from NVIDIA, fine-tuned from Google Gemma-3-4B.

Thinking Machines: Inkling (batch)
Community
inkling:batch
1 Credit = 10,000 Tokens
Long ContextReasoning

Inkling is an open-weight multimodal mixture-of-experts model from Thinking Machines Lab, with 41B active parameters out of 975B total.

Poolside: Laguna XS 2.1 (free)
Community
laguna-xs-2.1:free
Pricing on request
Long Context

Laguna XS 2.1 is the latest coding agent model in the 33B-A3B category from [Poolside](https://poolside.ai/) and a step forward from thei…

Cohere: Command R (08-2024)
Cohere
command-r-08-2024
1 Credit = 66,667 Tokens
Mid ContextCode

command-r-08-2024 is an update of the [Command R](/models/cohere/command-r) with improved performance for multilingual retrieval-augmente…

MiniMax: MiniMax M2.5
Community
minimax-m2.5
1 Credit = 37,037 Tokens
Long Context

MiniMax-M2.5 is a SOTA large language model designed for real-world productivity.

Mistral: Mistral Small 4 (batch)
Mistral
mistral-small-2603:batch
1 Credit = 133,333 Tokens
Long ContextVision

Mistral Small 4 is the next major release in the Mistral Small family, unifying the capabilities of several flagship Mistral models into…

Pareto Code Router
Community
pareto-code
Pricing on request
Long ContextCode

The Pareto Router maintains a tiered shortlist of strong coding models, ranked by [Artificial Analysis](https://artificialanalysis.ai/) c…

Perplexity: Sonar Deep Research
Perplexity
sonar-deep-research
1 Credit = 5,000 Tokens
Mid Context

Sonar Deep Research is a research-focused model designed for multi-step retrieval, synthesis, and reasoning across complex topics.

Amazon: Nova Lite 1.0
Community
nova-lite-v1
1 Credit = 166,667 Tokens
Long ContextVision

Amazon Nova Lite 1.0 is a very low-cost multimodal model from Amazon that focused on fast processing of image, video, and text inputs to…

Meta: Llama 4 Scout
Meta
llama-4-scout
1 Credit = 100,000 Tokens
Long ContextVision

Llama 4 Scout 17B Instruct (16E) is a mixture-of-experts (MoE) language model developed by Meta, activating 17 billion parameters out of…

OpenAI: GPT Mini Latest
Community
gpt-mini-latest
1 Credit = 13,333 Tokens
Long ContextVision

This model always redirects to the latest model in the GPT Mini family.

OpenAI: GPT-5.4 Pro
OpenAI
gpt-5.4-pro
1 Credit = 333 Tokens
Long ContextVision

GPT-5.4 Pro is OpenAI's most advanced model, building on GPT-5.4's unified architecture with enhanced reasoning capabilities for complex,…

Anthropic: Claude Fable 5 (batch)
Anthropic
claude-fable-5:batch
1 Credit = 2,000 Tokens
Long ContextVision

Claude Fable 5 is a Mythos-class model from Anthropic, built for autonomous knowledge work and coding.

Qwen: Qwen3.8 27B
Alibaba
NEW
qwen3.8-27b
1 Credit = 46,729 Tokens
Long ContextVision

Qwen3.8 27B is an open-weight dense vision-language model from Qwen.

Meta: Muse Glimmer 30B
Meta
muse-glimmer-30b
1 Credit = 33,333 Tokens
Mid ContextVision

Muse Glimmer 30B is a dense, open-weight multimodal model from Meta Superintelligence Labs, distilled from Muse Spark and optimized for a…

SpaceXAI: Grok 4.3 (batch)
xAI
grok-4.3:batch
1 Credit = 10,000 Tokens
Long ContextReasoning

Grok 4.3 is a reasoning model from SpaceXAI.

Google: Gemini 3.6 Flash
Google
gemini-3.6-flash
1 Credit = 13,333 Tokens
Long ContextVision

Gemini 3.6 Flash is a high-efficiency model from Google for coding, agentic workflows, and web and app development.

Z.ai: GLM 5.2 (batch)
Z.ai
glm-5.2:batch
1 Credit = 14,286 Tokens
Long Context

GLM 5.2 is a large-scale reasoning model from Z.ai.

MoonshotAI: Kimi K2.5
Community
kimi-k2.5
1 Credit = 22,222 Tokens
Long ContextVision

Kimi K2.5 is Moonshot AI's native multimodal model, delivering state-of-the-art visual coding capability and a self-directed agent swarm…

Tencent: Hy3
Tencent
hy3
1 Credit = 121,212 Tokens
Long Context

Hy3 is a 295B-parameter Mixture-of-Experts model from Tencent (21B active, 192 experts with top-8 routing) built for reasoning, agentic w…

DeepSeek: DeepSeek V4 Flash Vision Exp (batch)
DeepSeek
NEW
deepseek-v4-flash-vision-exp:batch
1 Credit = 90,909 Tokens
Long ContextVision

DeepSeek V4 Flash Vision Exp is an experimental vision-enabled version of [DeepSeek V4 Flash 0731](https://openrouter.ai/deepseek/deepsee…

inclusionAI: Ling 3.0 Flash Fin (free)
Community
NEW
ling-3.0-flash-fin:free
Pricing on request
Long Context

Ling 3.0 Flash Fin is a finance-focused mixture-of-experts model from InclusionAI, built on Ling 3.0 Flash with 5.1B active parameters ou…

ByteDance Seed: Seed 1.6
Community
seed-1.6
1 Credit = 40,000 Tokens
Long ContextVision

Seed 1.6 is a general-purpose model released by the ByteDance Seed team.

NVIDIA: Nemotron 3.5 Content Safety (free)
Community
nemotron-3.5-content-safety:free
Pricing on request
Mid ContextVision

NVIDIA Nemotron 3.5 Content Safety is a compact 4B-parameter multimodal guardrail model from NVIDIA, fine-tuned from Google Gemma-3-4B.

Thinking Machines: Inkling Small (free)
Community
inkling-small:free
Pricing on request
Long ContextReasoning

Inkling Small is an open-weight multimodal mixture-of-experts model from Thinking Machines Lab, with 12B active parameters out of 276B to…

Cohere: Command R7B (12-2024)
Cohere
command-r7b-12-2024
1 Credit = 266,667 Tokens
Mid ContextCode

Command R7B (12-2024) is a small, fast update of the Command R+ model, delivered in December 2024.

MiniMax: MiniMax M2-her
Community
minimax-m2-her
1 Credit = 33,333 Tokens
Mid Context

MiniMax M2-her is a dialogue-first large language model built for immersive roleplay, character-driven chat, and expressive multi-turn co…

Mistral: Devstral 2 2512
Mistral
devstral-2512
1 Credit = 25,000 Tokens
Long Context

Devstral 2 is a state-of-the-art open-source model by Mistral AI specializing in agentic coding.

Body Builder (beta)
Community
bodybuilder
Pricing on request
Mid Context

Transform your natural language requests into structured OpenRouter API request objects.

Perplexity: Sonar
Perplexity
sonar
1 Credit = 10,000 Tokens
Mid ContextVision

Sonar is lightweight, affordable, fast, and simple to use — now featuring citations and the ability to customize sources.

Amazon: Nova Micro 1.0
Community
nova-micro-v1
1 Credit = 285,714 Tokens
Mid Context

Amazon Nova Micro 1.0 is a text-only model that delivers the lowest latency responses in the Amazon Nova family of models at a very low c…

Meta: Llama 3.3 70B Instruct
Meta
llama-3.3-70b-instruct
1 Credit = 100,000 Tokens
Mid Context

The Meta Llama 3.3 multilingual large language model (LLM) is a pretrained and instruction tuned generative model in 70B (text in/text out).

OpenAI: GPT-5.4 Pro (batch)
OpenAI
gpt-5.4-pro:batch
1 Credit = 667 Tokens
Long ContextVision

GPT-5.4 Pro is OpenAI's most advanced model, building on GPT-5.4's unified architecture with enhanced reasoning capabilities for complex,…

Claude Opus 5 (batch)
Anthropic
claude-opus-5:batch
1 Credit = 4,000 Tokens
Long ContextReasoning

Claude Opus 5 is Anthropic’s flagship model for demanding reasoning, coding, and long-horizon agentic work.

Qwen: Qwen3.7 Flash
Alibaba
qwen3.7-flash
1 Credit = 333,333 Tokens
Long ContextVision

Qwen3.7 Flash is a vision-language reasoning model from Alibaba.

Meta: Muse Glimmer 30B (batch)
Meta
muse-glimmer-30b:batch
1 Credit = 57,143 Tokens
Mid ContextVision

Muse Glimmer 30B is a dense, open-weight multimodal model from Meta Superintelligence Labs, distilled from Muse Spark and optimized for a…

SpaceXAI: Grok 4.20 Multi-Agent
xAI
grok-4.20-multi-agent
1 Credit = 8,000 Tokens
Long ContextReasoning

Grok 4.20 Multi-Agent is a variant of SpaceXAI’s Grok 4.20 designed for collaborative, agent-based workflows.

Google: Gemini 3.6 Flash (batch)
Google
gemini-3.6-flash:batch
1 Credit = 26,667 Tokens
Long ContextVision

Gemini 3.6 Flash is a high-efficiency model from Google for coding, agentic workflows, and web and app development.

Z.ai: GLM 5.2
Z.ai
glm-5.2
1 Credit = 16,667 Tokens
Long Context

GLM 5.2 is a large-scale reasoning model from Z.ai.

MoonshotAI: Kimi K2 Thinking
Community
kimi-k2-thinking
1 Credit = 16,667 Tokens
Long ContextReasoning

Kimi K2 Thinking is Moonshot AI’s most advanced open reasoning model to date, extending the K2 series into agentic, long-horizon reasoning.

Tencent: Hy3 preview
Tencent
hy3-preview
1 Credit = 55,556 Tokens
Long Context

Hy3 preview is a high-efficiency Mixture-of-Experts model from Tencent designed for agentic workflows and production use.

DeepSeek: DeepSeek V4 Flash 0731 (batch)
DeepSeek
deepseek-v4-flash-0731:batch
1 Credit = 90,909 Tokens
Long Context

DeepSeek V4 Flash 0731 is a sparse mixture-of-experts model from DeepSeek, with 13B active parameters out of 284B total.

inclusionAI: Ling 3.0 Flash
Community
ling-3.0-flash
1 Credit = 476,190 Tokens
Long Context

*Ling-3.0-flash* is a *124B-parameter Mixture-of-Experts (MoE) model*, with approximately *5.1B parameters activated per token*.

ByteDance Seed: Seed 1.6 Flash
Community
seed-1.6-flash
1 Credit = 133,333 Tokens
Long ContextVision

Seed 1.6 Flash is an ultra-fast multimodal deep thinking model by ByteDance Seed, supporting both text and visual understanding.

NVIDIA: Nemotron 3 Ultra (free)
Community
nemotron-3-ultra-550b-a55b:free
Pricing on request
Long Context

NVIDIA Nemotron 3 Ultra is an open frontier-reasoning and orchestration model from NVIDIA, with 55B active parameters out of 550B total (…

Thinking Machines: Inkling (free)
Community
inkling:free
Pricing on request
Long ContextReasoning

Inkling is an open-weight multimodal mixture-of-experts model from Thinking Machines Lab, with 41B active parameters out of 975B total.

MiniMax: MiniMax M2.1
Community
minimax-m2.1
1 Credit = 33,333 Tokens
Long Context

MiniMax-M2.1 is a lightweight, state-of-the-art large language model optimized for coding, agentic workflows, and modern application deve…

Mistral: Mistral Large 3 2512
Mistral
mistral-large-2512
1 Credit = 20,000 Tokens
Long ContextVision

Mistral Large 3 2512 is Mistral’s most capable model to date, featuring a sparse mixture-of-experts architecture with 41B active paramete…

Auto Router
Community
auto
Pricing on request
Long ContextVision

The Auto Router automatically selects the best model for your prompt, powered by the wisdom of the market.

Meta: Llama 3.2 3B Instruct
Meta
llama-3.2-3b-instruct
1 Credit = 200,000 Tokens
Mid Context

Llama 3.2 3B is a 3-billion-parameter multilingual large language model, optimized for advanced natural language processing tasks like di…

OpenAI: GPT-6 Astra (batch)
OpenAI
NEW
gpt-6-astra:batch
1 Credit = 2,000 Tokens
Long ContextVision

GPT-6 Astra is OpenAI's flagship model for demanding end-to-end work.

Anthropic: Claude Opus 4.8
Anthropic
claude-opus-4.8
1 Credit = 2,000 Tokens
Long ContextReasoning

Claude Opus 4.8 is Anthropic's most capable generally available model in the Opus family.

Qwen: Qwen3.7 Max
Alibaba
qwen3.7-max
1 Credit = 6,780 Tokens
Long ContextReasoning

Qwen3.7-Max is the flagship model in Alibaba's Qwen3.7 series.

Meta: Muse Spark 1.1
Meta
muse-spark-1.1
1 Credit = 8,000 Tokens
Long ContextVision

Muse Spark 1.1 is a multimodal reasoning model from Meta, built for agentic tasks.

SpaceXAI: Grok 4.20
xAI
grok-4.20
1 Credit = 8,000 Tokens
Long ContextReasoning

Grok 4.20 is a reasoning model from SpaceXAI with industry-leading speed and agentic tool calling capabilities.

Google: Nano Banana Pro (Gemini 3 Pro Image)
Google
gemini-3-pro-image
1 Credit = 5,000 Tokens
Mid ContextReasoning

Nano Banana Pro is Google’s most advanced image-generation and editing model, built on Gemini 3 Pro.

Z.ai: GLM 5.1
Z.ai
glm-5.1
1 Credit = 10,352 Tokens
Long Context

GLM-5.1 delivers a major leap in coding capability, with particularly significant gains in handling long-horizon tasks.

MoonshotAI: Kimi K2 0905
Community
kimi-k2-0905
1 Credit = 16,667 Tokens
Long Context

Kimi K2 0905 is the September update of [Kimi K2 0711](moonshotai/kimi-k2).

Tencent: Hunyuan A13B Instruct
Tencent
hunyuan-a13b-instruct
1 Credit = 71,429 Tokens
Mid Context

Hunyuan-A13B is a 13B active parameter Mixture-of-Experts (MoE) language model developed by Tencent, with a total parameter count of 80B…

DeepSeek: DeepSeek V4 Flash 0731
DeepSeek
deepseek-v4-flash-0731
1 Credit = 250,000 Tokens
Long Context

DeepSeek V4 Flash 0731 is a sparse mixture-of-experts model from DeepSeek, with 13B active parameters out of 284B total.

NVIDIA: Nemotron 3 Nano Omni (free)
Community
nemotron-3-nano-omni-30b-a3b-reasoning:free
Pricing on request
Long ContextReasoning

NVIDIA Nemotron™ 3 Nano Omni is a 30B-A3B open multimodal model designed to function as a perception and context sub-agent in enterprise…

MiniMax: MiniMax M2
Community
minimax-m2
1 Credit = 39,216 Tokens
Long Context

MiniMax-M2 is a compact, high-efficiency large language model optimized for end-to-end coding and agentic workflows.

Mistral: Mistral Large 3 2512 (batch)
Mistral
mistral-large-2512:batch
1 Credit = 40,000 Tokens
Long ContextVision

Mistral Large 3 2512 is Mistral’s most capable model to date, featuring a sparse mixture-of-experts architecture with 41B active paramete…

Meta: Llama 3.1 8B Instruct
Meta
llama-3.1-8b-instruct
1 Credit = 200,000 Tokens
Mid Context

Meta's latest class of model (Llama 3.1) launched with a variety of sizes & flavors.

OpenAI: GPT-6 Astra Pro (batch)
OpenAI
NEW
gpt-6-astra-pro:batch
1 Credit = 2,000 Tokens
Long ContextVision

GPT-6 Astra Pro is the same underlying model as [GPT-6 Astra](https://openrouter.ai/openai/gpt-6-astra), served with `reasoning.mode` set…

Anthropic: Claude Sonnet 5
Anthropic
claude-sonnet-5
1 Credit = 5,000 Tokens
Long ContextReasoning

Sonnet 5 is Anthropic's most capable Sonnet-class model, with frontier performance across coding, agents, and professional work.

Qwen: Qwen3.7 Plus
Alibaba
qwen3.7-plus
1 Credit = 31,250 Tokens
Long ContextVision

Qwen3.7-Plus is a cost-effective model in Alibaba's Qwen3.7 series.

Google: Gemini 3.5 Flash Lite
Google
gemini-3.5-flash-lite
1 Credit = 33,333 Tokens
Long ContextVision

Gemini 3.5 Flash Lite is a high-efficiency model from Google with upgraded agentic capabilities.

Z.ai: GLM 5V Turbo
Z.ai
glm-5v-turbo
1 Credit = 8,333 Tokens
Long ContextVision

GLM-5V-Turbo is Z.ai’s first native multimodal agent foundation model, built for vision-based coding and agent-driven tasks.

MoonshotAI: Kimi K2 0711
Community
kimi-k2
1 Credit = 17,544 Tokens
Mid Context

Kimi K2 Instruct is a large-scale Mixture-of-Experts (MoE) language model developed by Moonshot AI, featuring 1 trillion total parameters…

DeepSeek: DeepSeek V4 Pro 0423
DeepSeek
deepseek-v4-pro
1 Credit = 6,250 Tokens
Long Context

DeepSeek V4 Pro is a large-scale Mixture-of-Experts model from DeepSeek with 1.6T total parameters and 49B activated parameters, supporti…

NVIDIA: Nemotron 3 Super
Community
nemotron-3-super-120b-a12b
1 Credit = 117,647 Tokens
Long Context

NVIDIA Nemotron 3 Super is a 120B-parameter open hybrid MoE model, activating just 12B parameters for maximum compute efficiency and accu…

MiniMax: MiniMax M1
Community
minimax-m1
1 Credit = 18,182 Tokens
Long Context

MiniMax-M1 is a large-scale, open-weight reasoning model designed for extended context and high-efficiency inference.

Mistral: Ministral 3 14B 2512
Mistral
ministral-14b-2512
1 Credit = 50,000 Tokens
Long ContextVision

The largest model in the Ministral 3 family, Ministral 3 14B offers frontier capabilities and performance comparable to its larger Mistra…

Meta: Llama 3.2 1B Instruct
Meta
llama-3.2-1b-instruct
1 Credit = 370,370 Tokens
Mid Context

Llama 3.2 1B is a 1-billion-parameter language model focused on efficiently performing natural language tasks, such as summarization, dia…

OpenAI: GPT-5.2 Pro
OpenAI
gpt-5.2-pro
1 Credit = 476 Tokens
Long ContextVision

GPT-5.2 Pro is OpenAI’s most advanced model, offering major improvements in agentic coding and long context performance over GPT-5 Pro.

Anthropic: Claude Opus 4.7
Anthropic
claude-opus-4.7
1 Credit = 2,000 Tokens
Long ContextReasoning

Opus 4.7 is the next generation of Anthropic's Opus family, built for long-running, asynchronous agents.

Qwen: Qwen3.6 Max Preview
Alibaba
qwen3.6-max-preview
1 Credit = 9,737 Tokens
Long ContextReasoning

Qwen3.6-Max-Preview is a proprietary frontier model from Alibaba Cloud built on a sparse mixture-of-experts architecture with approximate…

Google: Gemini 3.5 Flash Lite (batch)
Google
gemini-3.5-flash-lite:batch
1 Credit = 66,667 Tokens
Long ContextVision

Gemini 3.5 Flash Lite is a high-efficiency model from Google with upgraded agentic capabilities.

Z.ai: GLM 5 Turbo
Z.ai
glm-5-turbo
1 Credit = 8,333 Tokens
Long Context

GLM-5 Turbo is a new model from Z.ai designed for fast inference and strong performance in agent-driven environments such as OpenClaw sce…

DeepSeek: DeepSeek V4 Flash 0423
DeepSeek
deepseek-v4-flash
1 Credit = 152,300 Tokens
Long Context

DeepSeek V4 Flash is an efficiency-optimized Mixture-of-Experts model from DeepSeek with 284B total parameters and 13B activated paramete…

NVIDIA: Nemotron 3 Super (free)
Community
nemotron-3-super-120b-a12b:free
Pricing on request
Long Context

NVIDIA Nemotron 3 Super is a 120B-parameter open hybrid MoE model, activating just 12B parameters for maximum compute efficiency and accu…

MiniMax: MiniMax-01
Community
minimax-01
1 Credit = 50,000 Tokens
Long ContextVision

MiniMax-01 is a combines MiniMax-Text-01 for text generation and MiniMax-VL-01 for image understanding.

Mistral: Ministral 3 8B 2512
Mistral
ministral-8b-2512
1 Credit = 66,667 Tokens
Long ContextVision

A balanced model in the Ministral 3 family, Ministral 3 8B is a powerful, efficient tiny language model with vision capabilities.

OpenAI: GPT-5.4 Image 2
OpenAI
gpt-5.4-image-2
1 Credit = 1,250 Tokens
Long ContextVision

[GPT-5.4](https://openrouter.ai/openai/gpt-5.4) Image 2 combines OpenAI's GPT-5.4 model with state-of-the-art image generation capabiliti…

Anthropic: Claude Sonnet 5 (batch)
Anthropic
claude-sonnet-5:batch
1 Credit = 10,000 Tokens
Long ContextReasoning

Sonnet 5 is Anthropic's most capable Sonnet-class model, with frontier performance across coding, agents, and professional work.

Qwen: Qwen3.5 Plus 2026-04-20
Alibaba
qwen3.5-plus-20260420
1 Credit = 33,333 Tokens
Long ContextVision

Qwen3.5 Plus (April 2026) is a large-scale multimodal language model from Alibaba.

Google: Nano Banana 2 Lite (Gemini 3.1 Flash Lite Image)
Google
gemini-3.1-flash-lite-image
1 Credit = 40,000 Tokens
Mid ContextVision

Nano Banana 2 Lite (Gemini 3.1 Flash Lite Image) is Google's fastest, most cost-efficient Gemini image model, built for high-velocity dev…

Z.ai: GLM 5
Z.ai
glm-5
1 Credit = 16,667 Tokens
Long Context

GLM-5 is Z.ai’s flagship open-source foundation model engineered for complex systems design and long-horizon agent workflows.

DeepSeek: DeepSeek V3.2
DeepSeek
deepseek-v3.2
1 Credit = 37,175 Tokens
Mid Context

DeepSeek-V3.2 is a large language model designed to harmonize high computational efficiency with strong reasoning and agentic tool-use pe…

NVIDIA: Nemotron 3 Nano 30B A3B
Community
nemotron-3-nano-30b-a3b
1 Credit = 200,000 Tokens
Long Context

NVIDIA Nemotron 3 Nano 30B A3B is a small language MoE model with highest compute efficiency and accuracy for developers to build special…

Mistral: Ministral 3 3B 2512
Mistral
ministral-3b-2512
1 Credit = 100,000 Tokens
Mid ContextVision

The smallest model in the Ministral 3 family, Ministral 3 3B is a powerful, efficient tiny language model with vision capabilities.

OpenAI: GPT-5.6 Terra Pro
OpenAI
gpt-5.6-terra-pro
1 Credit = 5,000 Tokens
Long ContextVision

GPT-5.6 Terra Pro is the same underlying model as [GPT-5.6 Terra](https://openrouter.ai/openai/gpt-5.6-terra), served with `reasoning.mod…

Anthropic: Claude Opus 4.8 (batch)
Anthropic
claude-opus-4.8:batch
1 Credit = 4,000 Tokens
Long ContextReasoning

Claude Opus 4.8 is Anthropic's most capable generally available model in the Opus family.

Qwen: Qwen3.6 27B
Alibaba
qwen3.6-27b
1 Credit = 33,333 Tokens
Long ContextVision

Qwen3.6 27B is a dense 27-billion-parameter language model from the Qwen Team at Alibaba, released in April 2026.

Google: Nano Banana 2 (Gemini 3.1 Flash Image)
Google
gemini-3.1-flash-image
1 Credit = 20,000 Tokens
Mid ContextVision

Gemini 3.1 Flash Image, a.k.a.

Z.ai: GLM 4.7 Flash
Z.ai
glm-4.7-flash
1 Credit = 165,289 Tokens
Long Context

As a 30B-class SOTA model, GLM-4.7-Flash offers a new option that balances performance and efficiency.

DeepSeek: DeepSeek V3.2 Exp
DeepSeek
deepseek-v3.2-exp
1 Credit = 37,037 Tokens
Mid Context

DeepSeek-V3.2-Exp is an experimental large language model released by DeepSeek as an intermediate step between V3.1 and future architectu…

Mistral: Ministral 3 8B 2512 (batch)
Mistral
ministral-8b-2512:batch
1 Credit = 133,333 Tokens
Long ContextVision

A balanced model in the Ministral 3 family, Ministral 3 8B is a powerful, efficient tiny language model with vision capabilities.

OpenAI: GPT-5.6 Terra
OpenAI
gpt-5.6-terra
1 Credit = 5,000 Tokens
Long ContextVision

GPT-5.6 Terra is a balanced model in OpenAI's GPT-5.6 series, positioned between the flagship Sol tier and the cost-efficient Luna tier.

Anthropic: Claude Opus 4.1
Anthropic
claude-opus-4.1
1 Credit = 667 Tokens
Long ContextReasoning

Claude Opus 4.1 is an updated version of Anthropic’s flagship model, offering improved performance in coding, reasoning, and agentic tasks.

Qwen: Qwen3.6 Flash
Alibaba
qwen3.6-flash
1 Credit = 53,333 Tokens
Long ContextVision

Qwen3.6 Flash is a fast, efficient language model from Alibaba's Qwen 3.6 series.

Google: Gemini 3.5 Flash
Google
gemini-3.5-flash
1 Credit = 6,667 Tokens
Long ContextVision

Gemini 3.5 Flash is Google's high-efficiency multimodal model, bringing near-Pro level coding and reasoning at Flash-tier cost and speed.

Z.ai: GLM 4.7
Z.ai
glm-4.7
1 Credit = 25,000 Tokens
Long Context

GLM-4.7 is Z.ai’s latest flagship model, featuring upgrades in two key areas: enhanced programming capabilities and more stable multi-ste…

DeepSeek: R1 Distill Llama 70B
DeepSeek
deepseek-r1-distill-llama-70b
1 Credit = 12,500 Tokens
Reasoning

DeepSeek R1 Distill Llama 70B is a distilled large language model based on [Llama-3.3-70B-Instruct](/meta-llama/llama-3.3-70b-instruct),…

Mistral: Voxtral Small 24B 2507
Mistral
voxtral-small-24b-2507
1 Credit = 100,000 Tokens
Mid Context

Voxtral Small is an enhancement of Mistral Small 3, incorporating state-of-the-art audio input capabilities while retaining best-in-class…

OpenAI: GPT-5.6 Sol Pro
OpenAI
gpt-5.6-sol-pro
1 Credit = 5,000 Tokens
Long ContextVision

GPT-5.6 Sol Pro is the same underlying model as [GPT-5.6 Sol](https://openrouter.ai/openai/gpt-5.6-sol), served with `reasoning.mode` set…

Anthropic: Claude Opus 4
Anthropic
claude-opus-4
1 Credit = 667 Tokens
Long ContextReasoning

Claude Opus 4 is benchmarked as the world’s best coding model, at time of release, bringing sustained performance on complex, long-runnin…

Qwen: Qwen3.6 35B A3B
Alibaba
qwen3.6-35b-a3b
1 Credit = 100,000 Tokens
Long ContextVision

Qwen3.6-35B-A3B is an open-weight multimodal model from Alibaba Cloud with 35 billion total parameters and 3 billion active parameters pe…

Google: Gemini 3.5 Flash (batch)
Google
gemini-3.5-flash:batch
1 Credit = 13,333 Tokens
Long ContextVision

Gemini 3.5 Flash is Google's high-efficiency multimodal model, bringing near-Pro level coding and reasoning at Flash-tier cost and speed.

Z.ai: GLM 4.6V
Z.ai
glm-4.6v
1 Credit = 33,333 Tokens
Mid ContextVision

GLM-4.6V is a large multimodal model designed for high-fidelity visual understanding and long-context reasoning across images, documents,…

DeepSeek: DeepSeek V3.1 Terminus
DeepSeek
deepseek-v3.1-terminus
1 Credit = 37,037 Tokens
Mid Context

DeepSeek-V3.1 Terminus is an update to [DeepSeek V3.1](/deepseek/deepseek-chat-v3.1) that maintains the model's original capabilities whi…

Mistral Large 2407
Mistral
mistral-large-2407
1 Credit = 5,000 Tokens
Mid Context

This is Mistral AI's flagship model, Mistral Large 2 (version mistral-large-2407).

OpenAI: GPT-5.6 Sol
OpenAI
gpt-5.6-sol
1 Credit = 5,000 Tokens
Long ContextVision

GPT-5.6 Sol is the flagship model in OpenAI's GPT-5.6 series.

Anthropic: Claude Opus 4.7 (batch)
Anthropic
claude-opus-4.7:batch
1 Credit = 4,000 Tokens
Long ContextReasoning

Opus 4.7 is the next generation of Anthropic's Opus family, built for long-running, asynchronous agents.

Qwen: Qwen3.6 Plus
Alibaba
qwen3.6-plus
1 Credit = 30,769 Tokens
Long ContextVision

Qwen 3.6 Plus builds on a hybrid architecture that combines efficient linear attention with sparse mixture-of-experts routing, enabling s…

Google: Gemini 3.1 Flash Lite
Google
gemini-3.1-flash-lite
1 Credit = 40,000 Tokens
Long ContextVision

Gemini 3.1 Flash Lite is Google’s GA high-efficiency multimodal model optimized for low-latency, high-volume workloads.

Z.ai: GLM 4.6
Z.ai
glm-4.6
1 Credit = 23,256 Tokens
Long Context

Compared with GLM-4.5, this generation brings several key improvements: Longer context window: The context window has been expanded from…

DeepSeek: R1
DeepSeek
deepseek-r1
1 Credit = 14,286 Tokens
Mid ContextReasoning

DeepSeek R1 is here: Performance on par with [OpenAI o1](/openai/o1), but open-sourced and with fully open reasoning tokens.

Mistral: Mixtral 8x22B Instruct
Mistral
mixtral-8x22b-instruct
1 Credit = 5,000 Tokens
Mid Context

Mistral's official instruct fine-tuned version of [Mixtral 8x22B](/models/mistralai/mixtral-8x22b).

OpenAI: GPT Chat Latest
OpenAI
gpt-chat-latest
1 Credit = 2,000 Tokens
Long ContextVision

GPT Chat Latest points to OpenAI's stable API alias `chat-latest` that always resolves to the latest Instant chat model used in ChatGPT.

Anthropic: Claude Opus 4.6
Anthropic
claude-opus-4.6
1 Credit = 2,000 Tokens
Long ContextReasoning

Opus 4.6 is Anthropic’s strongest model for coding and long-running professional tasks.

Qwen: Qwen3.5-9B (batch)
Alibaba
qwen3.5-9b:batch
1 Credit = 58,824 Tokens
Long ContextVision

Qwen3.5-9B is a multimodal foundation model from the Qwen3.5 family, designed to deliver strong reasoning, coding, and visual understandi…

Google: Gemini 3.1 Flash Lite (batch)
Google
gemini-3.1-flash-lite:batch
1 Credit = 80,000 Tokens
Long ContextVision

Gemini 3.1 Flash Lite is Google’s GA high-efficiency multimodal model optimized for low-latency, high-volume workloads.

Z.ai: GLM 4.5V
Z.ai
glm-4.5v
1 Credit = 16,667 Tokens
Mid ContextVision

GLM-4.5V is a vision-language foundation model for multimodal agent applications.

DeepSeek: R1 0528
DeepSeek
deepseek-r1-0528
1 Credit = 20,000 Tokens
Mid ContextReasoning

May 28th update to the [original DeepSeek R1](/deepseek/deepseek-r1) Performance on par with [OpenAI o1](/openai/o1), but open-sourced an…

Mistral Large
Mistral
mistral-large
1 Credit = 5,000 Tokens
Mid Context

This is Mistral AI's flagship model, Mistral Large 2 (version `mistral-large-2407`).

OpenAI: GPT-5 Pro
OpenAI
gpt-5-pro
1 Credit = 667 Tokens
Long ContextVision

GPT-5 Pro is OpenAI’s most advanced model, offering major improvements in reasoning, code quality, and user experience.

Anthropic: Claude Sonnet 4.6
Anthropic
claude-sonnet-4.6
1 Credit = 3,333 Tokens
Long ContextReasoning

Sonnet 4.6 is Anthropic's most capable Sonnet-class model yet, with frontier performance across coding, agents, and professional work.

Qwen: Qwen3.5-9B
Alibaba
qwen3.5-9b
1 Credit = 100,000 Tokens
Long ContextVision

Qwen3.5-9B is a multimodal foundation model from the Qwen3.5 family, designed to deliver strong reasoning, coding, and visual understandi…

Google: Gemma 4 31B (batch)
Google
gemma-4-31b-it:batch
1 Credit = 25,641 Tokens
Long ContextVision

Gemma 4 31B Instruct is Google DeepMind's 30.7B dense multimodal model supporting text and image input with text output.

Z.ai: GLM 4.5
Z.ai
glm-4.5
1 Credit = 16,667 Tokens
Mid Context

GLM-4.5 is our latest flagship foundation model, purpose-built for agent-based applications.

DeepSeek: DeepSeek V3
DeepSeek
deepseek-chat
1 Credit = 38,850 Tokens
Mid Context

DeepSeek-V3 is the latest model from the DeepSeek team, building upon the instruction following and coding abilities of the previous vers…

Mistral: Mistral Medium 3.1
Mistral
mistral-medium-3.1
1 Credit = 25,000 Tokens
Mid ContextVision

Mistral Medium 3.1 is an updated version of Mistral Medium 3, which is a high-performance enterprise-grade language model designed to del…

OpenAI: GPT-5.5
OpenAI
gpt-5.5
1 Credit = 2,000 Tokens
Long ContextVision

GPT-5.5 is OpenAI’s frontier model designed for complex professional workloads, building on GPT-5.4 with stronger reasoning, higher relia…

Anthropic: Claude Opus 4.6 (batch)
Anthropic
claude-opus-4.6:batch
1 Credit = 4,000 Tokens
Long ContextReasoning

Opus 4.6 is Anthropic’s strongest model for coding and long-running professional tasks.

Qwen: Qwen3.5-35B-A3B
Alibaba
qwen3.5-35b-a3b
1 Credit = 32,000 Tokens
Long ContextVision

The Qwen3.5 Series 35B-A3B is a native vision-language model designed with a hybrid architecture that integrates linear attention mechani…

Google: Gemini 3.1 Pro Preview Custom Tools
Google
gemini-3.1-pro-preview-customtools
1 Credit = 5,000 Tokens
Long ContextReasoning

Gemini 3.1 Pro Preview Custom Tools is a variant of Gemini 3.1 Pro that improves tool selection behavior by preventing overuse of a gener…

Z.ai: GLM 4.5 Air
Z.ai
glm-4.5-air
1 Credit = 76,923 Tokens
Mid Context

GLM-4.5-Air is the lightweight variant of our latest flagship model family, also purpose-built for agent-centric applications.

DeepSeek: DeepSeek V3.1
DeepSeek
deepseek-chat-v3.1
1 Credit = 40,000 Tokens
Mid Context

DeepSeek-V3.1 is a large hybrid reasoning model (671B parameters, 37B active) that supports both thinking and non-thinking modes via prom…

Mistral: Mistral Medium 3
Mistral
mistral-medium-3
1 Credit = 25,000 Tokens
Mid ContextVision

Mistral Medium 3 is a high-performance enterprise-grade language model designed to deliver frontier-level capabilities at significantly r…

OpenAI: GPT-5.6 Terra Pro (batch)
OpenAI
gpt-5.6-terra-pro:batch
1 Credit = 10,000 Tokens
Long ContextVision

GPT-5.6 Terra Pro is the same underlying model as [GPT-5.6 Terra](https://openrouter.ai/openai/gpt-5.6-terra), served with `reasoning.mod…

Anthropic: Claude Sonnet 4.6 (batch)
Anthropic
claude-sonnet-4.6:batch
1 Credit = 6,667 Tokens
Long ContextReasoning

Sonnet 4.6 is Anthropic's most capable Sonnet-class model yet, with frontier performance across coding, agents, and professional work.

Qwen: Qwen3.5-122B-A10B
Alibaba
qwen3.5-122b-a10b
1 Credit = 38,462 Tokens
Long ContextVision

The Qwen3.5 122B-A10B native vision-language model is built on a hybrid architecture that integrates a linear attention mechanism with a…

Google: Gemma 4 31B
Google
gemma-4-31b-it
1 Credit = 111,111 Tokens
Long ContextVision

Gemma 4 31B Instruct is Google DeepMind's 30.7B dense multimodal model supporting text and image input with text output.

DeepSeek: DeepSeek V3 0324
DeepSeek
deepseek-chat-v3-0324
1 Credit = 40,000 Tokens
Mid Context

DeepSeek V3, a 685B-parameter, mixture-of-experts model, is the latest iteration of the flagship chat model family from the DeepSeek team.

Mistral: Mistral Small 3.1 24B
Mistral
mistral-small-3.1-24b-instruct
1 Credit = 28,490 Tokens
Mid ContextVision

Mistral Small 3.1 24B Instruct is an upgraded variant of Mistral Small 3 (2501), featuring 24 billion parameters with advanced multimodal…

OpenAI: GPT-5.6 Terra (batch)
OpenAI
gpt-5.6-terra:batch
1 Credit = 10,000 Tokens
Long ContextVision

GPT-5.6 Terra is a balanced model in OpenAI's GPT-5.6 series, positioned between the flagship Sol tier and the cost-efficient Luna tier.

Anthropic: Claude Opus 4.5
Anthropic
claude-opus-4.5
1 Credit = 2,000 Tokens
Long ContextReasoning

Claude Opus 4.5 is Anthropic’s frontier reasoning model optimized for complex software engineering, agentic workflows, and long-horizon c…

Qwen: Qwen3.5-27B
Alibaba
qwen3.5-27b
1 Credit = 51,282 Tokens
Long ContextVision

The Qwen3.5 27B native vision-language Dense model incorporates a linear attention mechanism, delivering fast response times while balanc…

Google: Gemma 4 26B A4B
Google
gemma-4-26b-a4b-it
1 Credit = 238,095 Tokens
Long ContextVision

Gemma 4 26B A4B IT is an instruction-tuned Mixture-of-Experts (MoE) model from Google DeepMind.

Mistral: Codestral 2508
Mistral
codestral-2508
1 Credit = 33,333 Tokens
Long ContextCode

Mistral's cutting-edge language model for coding released end of July 2025.

OpenAI: GPT-5.6 Sol Pro (batch)
OpenAI
gpt-5.6-sol-pro:batch
1 Credit = 10,000 Tokens
Long ContextVision

GPT-5.6 Sol Pro is the same underlying model as [GPT-5.6 Sol](https://openrouter.ai/openai/gpt-5.6-sol), served with `reasoning.mode` set…

Anthropic: Claude Opus 4.1 (batch)
Anthropic
claude-opus-4.1:batch
1 Credit = 1,333 Tokens
Long ContextReasoning

Claude Opus 4.1 is an updated version of Anthropic’s flagship model, offering improved performance in coding, reasoning, and agentic tasks.

Qwen: Qwen3.5 397B A17B
Alibaba
qwen3.5-397b-a17b
1 Credit = 18,182 Tokens
Long ContextVision

The Qwen3.5 series 397B-A17B native vision-language model is built on a hybrid architecture that integrates a linear attention mechanism…

Google: Gemma 4 26B A4B (free)
Google
gemma-4-26b-a4b-it:free
Pricing on request
Long ContextVision

Gemma 4 26B A4B IT is an instruction-tuned Mixture-of-Experts (MoE) model from Google DeepMind.

Mistral: Mistral Medium 3.1 (batch)
Mistral
mistral-medium-3.1:batch
1 Credit = 50,000 Tokens
Mid ContextVision

Mistral Medium 3.1 is an updated version of Mistral Medium 3, which is a high-performance enterprise-grade language model designed to del…

OpenAI: GPT-5.6 Sol (batch)
OpenAI
gpt-5.6-sol:batch
1 Credit = 10,000 Tokens
Long ContextVision

GPT-5.6 Sol is the flagship model in OpenAI's GPT-5.6 series.

Anthropic: Claude Opus 4.5 (batch)
Anthropic
claude-opus-4.5:batch
1 Credit = 4,000 Tokens
Long ContextReasoning

Claude Opus 4.5 is Anthropic’s frontier reasoning model optimized for complex software engineering, agentic workflows, and long-horizon c…

Qwen: Qwen3.5-Flash
Alibaba
qwen3.5-flash-02-23
1 Credit = 153,846 Tokens
Long ContextVision

The Qwen3.5 native vision-language Flash models are built on a hybrid architecture that integrates a linear attention mechanism with a sp…

Google: Gemma 4 31B (free)
Google
gemma-4-31b-it:free
Pricing on request
Long ContextVision

Gemma 4 31B Instruct is Google DeepMind's 30.7B dense multimodal model supporting text and image input with text output.

Mistral: Saba
Mistral
mistral-saba
1 Credit = 50,000 Tokens
Mid Context

Mistral Saba is a 24B-parameter language model specifically designed for the Middle East and South Asia, delivering accurate and contextu…

OpenAI: GPT-5.6 Luna Pro
OpenAI
gpt-5.6-luna-pro
1 Credit = 50,000 Tokens
Long ContextVision

GPT-5.6 Luna Pro is the same underlying model as [GPT-5.6 Luna](https://openrouter.ai/openai/gpt-5.6-luna), served with `reasoning.mode`…

Anthropic: Claude Sonnet 4.5
Anthropic
claude-sonnet-4.5
1 Credit = 3,333 Tokens
Long ContextReasoning

Claude Sonnet 4.5 is Anthropic’s most advanced Sonnet model to date, optimized for real-world agents and coding workflows.

Qwen: Qwen3 Max Thinking
Alibaba
qwen3-max-thinking
1 Credit = 12,821 Tokens
Long ContextReasoning

Qwen3-Max-Thinking is the flagship reasoning model in the Qwen3 series, designed for high-stakes cognitive tasks that require deep, multi…

Google: Gemini 3.1 Pro Preview
Google
gemini-3.1-pro-preview
1 Credit = 5,000 Tokens
Long ContextReasoning

Gemini 3.1 Pro Preview is Google’s frontier reasoning model, delivering enhanced software engineering performance, improved agentic relia…

Mistral: Codestral 2508 (batch)
Mistral
codestral-2508:batch
1 Credit = 66,667 Tokens
Long ContextCode

Mistral's cutting-edge language model for coding released end of July 2025.

OpenAI: GPT-5.6 Luna
OpenAI
gpt-5.6-luna
1 Credit = 50,000 Tokens
Long ContextVision

GPT-5.6 Luna is a fast, cost-efficient model in OpenAI's GPT-5.6 series.

Anthropic: Claude Sonnet 4
Anthropic
claude-sonnet-4
1 Credit = 3,333 Tokens
Long ContextReasoning

Claude Sonnet 4 significantly enhances the capabilities of its predecessor, Sonnet 3.7, excelling in both coding and reasoning tasks with…

Qwen: Qwen3.5 Plus 2026-02-15
Alibaba
qwen3.5-plus-02-15
1 Credit = 38,462 Tokens
Long ContextVision

The Qwen3.5 native vision-language series Plus models are built on a hybrid architecture that integrates linear attention mechanisms with…

Google: Lyria 3 Pro Preview
Google
lyria-3-pro-preview
Pricing on request
Long ContextVision

Full-length songs are priced at $0.08 per song.

Mistral: Mistral Small 3.2 24B
Mistral
mistral-small-3.2-24b-instruct
1 Credit = 133,333 Tokens
Long ContextVision

Mistral-Small-3.2-24B-Instruct-2506 is an updated 24B parameter model from Mistral optimized for instruction following, repetition reduct…

OpenAI: GPT-5.6 Luna Pro (batch)
OpenAI
gpt-5.6-luna-pro:batch
1 Credit = 100,000 Tokens
Long ContextVision

GPT-5.6 Luna Pro is the same underlying model as [GPT-5.6 Luna](https://openrouter.ai/openai/gpt-5.6-luna), served with `reasoning.mode`…

Anthropic: Claude Haiku 4.5
Anthropic
claude-haiku-4.5
1 Credit = 10,000 Tokens
Long ContextReasoning

Claude Haiku 4.5 is Anthropic’s fastest and most efficient model, delivering near-frontier intelligence at a fraction of the cost and lat…

Qwen: Qwen3 Coder Next
Alibaba
qwen3-coder-next
1 Credit = 83,333 Tokens
Long ContextCode

Qwen3-Coder-Next is an open-weight causal language model optimized for coding agents and local development workflows.

Google: Lyria 3 Clip Preview
Google
lyria-3-clip-preview
Pricing on request
Long ContextVision

30 second duration clips are priced at $0.04 per clip.

Mistral: Mistral Small 3
Mistral
mistral-small-24b-instruct-2501
1 Credit = 200,000 Tokens
Mid Context

Mistral Small 3 is a 24B-parameter language model optimized for low-latency performance across common AI tasks.

OpenAI: GPT-5.6 Luna (batch)
OpenAI
gpt-5.6-luna:batch
1 Credit = 100,000 Tokens
Long ContextVision

GPT-5.6 Luna is a fast, cost-efficient model in OpenAI's GPT-5.6 series.

Anthropic: Claude Sonnet 4.5 (batch)
Anthropic
claude-sonnet-4.5:batch
1 Credit = 6,667 Tokens
Long ContextReasoning

Claude Sonnet 4.5 is Anthropic’s most advanced Sonnet model to date, optimized for real-world agents and coding workflows.

Qwen: Qwen3 VL 32B Instruct
Alibaba
qwen3-vl-32b-instruct
1 Credit = 96,154 Tokens
Mid ContextVision

Qwen3-VL-32B-Instruct is a large-scale multimodal vision-language model designed for high-precision understanding and reasoning across te…

Google: Gemini 3.1 Pro Preview (batch)
Google
gemini-3.1-pro-preview:batch
1 Credit = 10,000 Tokens
Long ContextReasoning

Gemini 3.1 Pro Preview is Google’s frontier reasoning model, delivering enhanced software engineering performance, improved agentic relia…

Mistral: Mistral Nemo
Mistral
mistral-nemo
1 Credit = 526,316 Tokens
Mid Context

A 12B parameter model with a 128k token context length built by Mistral in collaboration with NVIDIA.

OpenAI: o3 Pro
OpenAI
o3-pro
1 Credit = 500 Tokens
Long ContextReasoning

The o-series of models are trained with reinforcement learning to think before they answer and perform complex reasoning.

Anthropic: Claude Haiku 4.5 (batch)
Anthropic
claude-haiku-4.5:batch
1 Credit = 20,000 Tokens
Long ContextReasoning

Claude Haiku 4.5 is Anthropic’s fastest and most efficient model, delivering near-frontier intelligence at a fraction of the cost and lat…

Qwen: Qwen3 VL 8B Thinking
Alibaba
qwen3-vl-8b-thinking
1 Credit = 55,556 Tokens
Mid ContextReasoning

Qwen3-VL-8B-Thinking is the reasoning-optimized variant of the Qwen3-VL-8B multimodal model, designed for advanced visual and textual rea…

Google: Nano Banana 2 (Gemini 3.1 Flash Image Preview)
Google
gemini-3.1-flash-image-preview
1 Credit = 20,000 Tokens
Mid ContextVision

Gemini 3.1 Flash Image Preview, a.k.a.

OpenAI: o1-pro
OpenAI
o1-pro
1 Credit = 67 Tokens
Long ContextReasoning

The o1 series of models are trained with reinforcement learning to think before they answer and perform complex reasoning.

Anthropic: Claude 3 Haiku
Anthropic
claude-3-haiku
1 Credit = 40,000 Tokens
Long ContextVision

Claude 3 Haiku is Anthropic's fastest and most compact model for near-instant responsiveness.

Qwen: Qwen3 VL 8B Instruct
Alibaba
qwen3-vl-8b-instruct
1 Credit = 85,470 Tokens
Long ContextVision

Qwen3-VL-8B-Instruct is a multimodal vision-language model from the Qwen3-VL series, built for high-fidelity understanding and reasoning…

Google: Gemini 3.1 Flash Lite Preview
Google
gemini-3.1-flash-lite-preview
1 Credit = 40,000 Tokens
Long ContextVision

Gemini 3.1 Flash Lite Preview is Google's high-efficiency model optimized for high-volume use cases.

OpenAI: o1
OpenAI
o1
1 Credit = 667 Tokens
Long ContextReasoning

The latest and strongest model family from OpenAI, o1 is designed to spend more time thinking before responding.

Qwen: Qwen3 VL 30B A3B Thinking
Alibaba
qwen3-vl-30b-a3b-thinking
1 Credit = 50,000 Tokens
Long ContextReasoning

Qwen3-VL-30B-A3B-Thinking is a multimodal model that unifies strong text generation with visual understanding for images and videos.

Google: Nano Banana Pro (Gemini 3 Pro Image Preview)
Google
gemini-3-pro-image-preview
1 Credit = 5,000 Tokens
Mid ContextReasoning

Nano Banana Pro is Google’s most advanced image-generation and editing model, built on Gemini 3 Pro.

OpenAI: GPT-4
OpenAI
gpt-4
1 Credit = 333 Tokens
Mid Context

OpenAI's flagship model, GPT-4 is a large-scale multimodal language model capable of solving difficult problems with greater accuracy tha…

Qwen: Qwen3 VL 30B A3B Instruct
Alibaba
qwen3-vl-30b-a3b-instruct
1 Credit = 66,667 Tokens
Long ContextVision

Qwen3-VL-30B-A3B-Instruct is a multimodal model that unifies strong text generation with visual understanding for images and videos.

Google: Gemini 3 Flash Preview
Google
gemini-3-flash-preview
1 Credit = 20,000 Tokens
Long ContextVision

Gemini 3 Flash Preview is a high speed, high value thinking model designed for agentic workflows, multi turn chat, and coding assistance.

OpenAI: GPT-5.2 Pro (batch)
OpenAI
gpt-5.2-pro:batch
1 Credit = 952 Tokens
Long ContextVision

GPT-5.2 Pro is OpenAI’s most advanced model, offering major improvements in agentic coding and long context performance over GPT-5 Pro.

Qwen: Qwen3 Max
Alibaba
qwen3-max
1 Credit = 12,821 Tokens
Long ContextReasoning

Qwen3-Max is an updated release built on the Qwen3 series, offering major improvements in reasoning, instruction following, multilingual…

Google: Gemini 3 Flash Preview (batch)
Google
gemini-3-flash-preview:batch
1 Credit = 40,000 Tokens
Long ContextVision

Gemini 3 Flash Preview is a high speed, high value thinking model designed for agentic workflows, multi turn chat, and coding assistance.

OpenAI: GPT-5.5 (batch)
OpenAI
gpt-5.5:batch
1 Credit = 4,000 Tokens
Long ContextVision

GPT-5.5 is OpenAI’s frontier model designed for complex professional workloads, building on GPT-5.4 with stronger reasoning, higher relia…

Qwen: Qwen3 Coder Plus
Alibaba
qwen3-coder-plus
1 Credit = 15,385 Tokens
Long ContextCode

Qwen3 Coder Plus is Alibaba's proprietary version of the Open Source Qwen3 Coder 480B A35B.

Google: Nano Banana (Gemini 2.5 Flash Image)
Google
gemini-2.5-flash-image
1 Credit = 33,333 Tokens
Mid ContextVision

Gemini 2.5 Flash Image, a.k.a.

OpenAI: GPT-5 Image
OpenAI
gpt-5-image
1 Credit = 1,000 Tokens
Long ContextVision

[GPT-5](https://openrouter.ai/openai/gpt-5) Image combines OpenAI's GPT-5 model with state-of-the-art image generation capabilities.

Qwen: Qwen3 VL 235B A22B Thinking
Alibaba
qwen3-vl-235b-a22b-thinking
1 Credit = 25,000 Tokens
Mid ContextReasoning

Qwen3-VL-235B-A22B Thinking is a multimodal model that unifies strong text generation with visual understanding across images and video.

Google: Gemini 2.5 Pro
Google
gemini-2.5-pro
1 Credit = 8,000 Tokens
Long ContextReasoning

Gemini 2.5 Pro is Google’s state-of-the-art AI model designed for advanced reasoning, coding, mathematics, and scientific tasks.

OpenAI: GPT-5.4
OpenAI
gpt-5.4
1 Credit = 4,000 Tokens
Long ContextVision

GPT-5.4 is OpenAI’s latest frontier model, unifying the Codex and GPT lines into a single system.

Qwen: Qwen2.5 VL 72B Instruct
Alibaba
qwen2.5-vl-72b-instruct
1 Credit = 12,500 Tokens
Mid ContextVision

Qwen2.5-VL is proficient in recognizing common objects such as flowers, birds, fish, and insects.

Google: Gemini 2.5 Pro Preview 06-05
Google
gemini-2.5-pro-preview
1 Credit = 8,000 Tokens
Long ContextReasoning

Gemini 2.5 Pro is Google’s state-of-the-art AI model designed for advanced reasoning, coding, mathematics, and scientific tasks.

OpenAI: GPT-5.4 Mini
OpenAI
gpt-5.4-mini
1 Credit = 13,333 Tokens
Long ContextVision

GPT-5.4 mini brings the core capabilities of GPT-5.4 to a faster, more efficient model optimized for high-throughput workloads.

Qwen: Qwen3 VL 235B A22B Instruct
Alibaba
qwen3-vl-235b-a22b-instruct
1 Credit = 47,619 Tokens
Long ContextVision

Qwen3-VL-235B-A22B Instruct is an open-weight multimodal model that unifies strong text generation with visual understanding across image…

Google: Gemini 2.5 Pro Preview 05-06
Google
gemini-2.5-pro-preview-05-06
1 Credit = 8,000 Tokens
Long ContextReasoning

Gemini 2.5 Pro is Google’s state-of-the-art AI model designed for advanced reasoning, coding, mathematics, and scientific tasks.

OpenAI: GPT-4 Turbo
OpenAI
gpt-4-turbo
1 Credit = 1,000 Tokens
Mid ContextVision

The latest GPT-4 Turbo model with vision capabilities.

Qwen2.5 Coder 32B Instruct
Alibaba
qwen-2.5-coder-32b-instruct
1 Credit = 15,152 Tokens
Mid ContextCode

Qwen2.5-Coder is the latest series of Code-Specific Qwen large language models (formerly known as CodeQwen).

Google: Gemma 2 27B
Google
gemma-2-27b-it
1 Credit = 15,385 Tokens
Mid Context

Gemma 2 27B by Google is an open model built from the same research and technology used to create the [Gemini models](/models?q=gemini).

OpenAI: GPT-4 Turbo Preview
OpenAI
gpt-4-turbo-preview
1 Credit = 1,000 Tokens
Mid Context

The preview GPT-4 model with improved instruction following, JSON mode, reproducible outputs, parallel function calling, and more.

Qwen: Qwen3 235B A22B
Alibaba
qwen3-235b-a22b
1 Credit = 21,978 Tokens
Mid Context

Qwen3-235B-A22B is a 235B parameter mixture-of-experts (MoE) model developed by Qwen, activating 22B parameters per forward pass.

Google: Gemini 2.5 Pro (batch)
Google
gemini-2.5-pro:batch
1 Credit = 16,000 Tokens
Long ContextReasoning

Gemini 2.5 Pro is Google’s state-of-the-art AI model designed for advanced reasoning, coding, mathematics, and scientific tasks.

OpenAI: GPT-5.3-Codex
OpenAI
gpt-5.3-codex
1 Credit = 5,714 Tokens
Long ContextVision

GPT-5.3-Codex is OpenAI’s most advanced agentic coding model, combining the frontier software engineering performance of GPT-5.2-Codex wi…

Qwen: Qwen3 Coder Flash
Alibaba
qwen3-coder-flash
1 Credit = 51,282 Tokens
Long ContextCode

Qwen3 Coder Flash is Alibaba's fast and cost efficient version of their proprietary Qwen3 Coder Plus.

Google: Gemini 2.5 Flash
Google
gemini-2.5-flash
1 Credit = 33,333 Tokens
Long ContextVision

Gemini 2.5 Flash is Google's state-of-the-art workhorse model, specifically designed for advanced reasoning, coding, mathematics, and sci…

OpenAI: GPT-5.4 (batch)
OpenAI
gpt-5.4:batch
1 Credit = 8,000 Tokens
Long ContextVision

GPT-5.4 is OpenAI’s latest frontier model, unifying the Codex and GPT lines into a single system.

Qwen2.5 72B Instruct
Alibaba
qwen-2.5-72b-instruct
1 Credit = 27,778 Tokens
Mid Context

Qwen2.5 72B is the latest series of Qwen large language models.

Google: Gemini 2.5 Flash (batch)
Google
gemini-2.5-flash:batch
1 Credit = 66,667 Tokens
Long ContextVision

Gemini 2.5 Flash is Google's state-of-the-art workhorse model, specifically designed for advanced reasoning, coding, mathematics, and sci…

OpenAI: GPT-5.4 Mini (batch)
OpenAI
gpt-5.4-mini:batch
1 Credit = 26,667 Tokens
Long ContextVision

GPT-5.4 mini brings the core capabilities of GPT-5.4 to a faster, more efficient model optimized for high-throughput workloads.

Qwen: Qwen3 Coder 480B A35B
Alibaba
qwen3-coder
1 Credit = 33,333 Tokens
Long ContextCode

Qwen3-Coder-480B-A35B-Instruct is a Mixture-of-Experts (MoE) code generation model developed by the Qwen team.

Google: Gemini 2.5 Flash Lite
Google
gemini-2.5-flash-lite
1 Credit = 100,000 Tokens
Long ContextVision

Gemini 2.5 Flash-Lite is a lightweight reasoning model in the Gemini 2.5 family, optimized for ultra-low latency and cost efficiency.

OpenAI: GPT-5.4 Nano
OpenAI
gpt-5.4-nano
1 Credit = 50,000 Tokens
Long ContextVision

GPT-5.4 nano is the most lightweight and cost-efficient variant of the GPT-5.4 family, optimized for speed-critical and high-volume tasks.

Qwen: Qwen Plus 0728
Alibaba
qwen-plus-2025-07-28
1 Credit = 38,462 Tokens
Long Context

Qwen Plus 0728, based on the Qwen3 foundation model, is a 1 million context hybrid reasoning model with a balanced performance, speed, an…

Google: Gemma 3 27B
Google
gemma-3-27b-it
1 Credit = 125,000 Tokens
Mid ContextVision

Gemma 3 introduces multimodality, supporting vision-language input and text outputs.

OpenAI: GPT-5.4 Nano (batch)
OpenAI
gpt-5.4-nano:batch
1 Credit = 100,000 Tokens
Long ContextVision

GPT-5.4 nano is the most lightweight and cost-efficient variant of the GPT-5.4 family, optimized for speed-critical and high-volume tasks.

Qwen: Qwen-Plus
Alibaba
qwen-plus
1 Credit = 38,462 Tokens
Long Context

Qwen-Plus, based on the Qwen2.5 foundation model, is a 131K context model with a balanced performance, speed, and cost combination.

Google: Gemini 2.5 Flash Lite (batch)
Google
gemini-2.5-flash-lite:batch
1 Credit = 200,000 Tokens
Long ContextVision

Gemini 2.5 Flash-Lite is a lightweight reasoning model in the Gemini 2.5 family, optimized for ultra-low latency and cost efficiency.

OpenAI: GPT Audio
OpenAI
gpt-audio
1 Credit = 4,000 Tokens
Mid Context

The gpt-audio model is OpenAI's first generally available audio model.

Qwen: Qwen3 235B A22B Thinking 2507
Alibaba
qwen3-235b-a22b-thinking-2507
1 Credit = 43,478 Tokens
Mid ContextReasoning

Qwen3-235B-A22B-Thinking-2507 is a high-performance, open-weight Mixture-of-Experts (MoE) language model optimized for complex reasoning…

Google: Gemma 3 4B
Google
gemma-3-4b-it
1 Credit = 200,000 Tokens
Mid ContextVision

Gemma 3 introduces multimodality, supporting vision-language input and text outputs.

OpenAI: GPT-5 Pro (batch)
OpenAI
gpt-5-pro:batch
1 Credit = 1,333 Tokens
Long ContextVision

GPT-5 Pro is OpenAI’s most advanced model, offering major improvements in reasoning, code quality, and user experience.

Qwen: Qwen3 14B
Alibaba
qwen3-14b
1 Credit = 43,956 Tokens
Mid Context

Qwen3-14B is a dense 14.8B parameter causal language model from the Qwen3 series, designed for both complex reasoning and efficient dialo…

Google: Gemma 3 12B
Google
gemma-3-12b-it
1 Credit = 200,000 Tokens
Mid ContextVision

Gemma 3 introduces multimodality, supporting vision-language input and text outputs.

OpenAI: GPT-5.2-Codex
OpenAI
gpt-5.2-codex
1 Credit = 5,714 Tokens
Long ContextVision

GPT-5.2-Codex is an upgraded version of GPT-5.1-Codex optimized for software engineering and coding workflows.

Qwen: Qwen3 30B A3B Thinking 2507
Alibaba
qwen3-30b-a3b-thinking-2507
1 Credit = 50,000 Tokens
Mid ContextReasoning

Qwen3-30B-A3B-Thinking-2507 is a 30B parameter Mixture-of-Experts reasoning model optimized for complex tasks requiring extended multi-st…

OpenAI: GPT Audio Mini
OpenAI
gpt-audio-mini
1 Credit = 16,667 Tokens
Mid Context

A cost-efficient version of GPT Audio.

Qwen: Qwen3 Next 80B A3B Thinking
Alibaba
qwen3-next-80b-a3b-thinking
1 Credit = 66,667 Tokens
Long ContextReasoning

Qwen3-Next-80B-A3B-Thinking is a reasoning-first chat model in the Qwen3-Next line that outputs structured “thinking” traces by default.

OpenAI: GPT-5.2 Chat
OpenAI
gpt-5.2-chat
1 Credit = 5,714 Tokens
Mid ContextVision

GPT-5.2 Chat (AKA Instant) is the fast, lightweight member of the 5.2 family, optimized for low-latency chat while retaining strong gener…

Qwen: Qwen3 30B A3B
Alibaba
qwen3-30b-a3b
1 Credit = 83,333 Tokens
Mid Context

Qwen3, the latest generation in the Qwen large language model series, features both dense and mixture-of-experts (MoE) architectures to e…

OpenAI: GPT-5.2
OpenAI
gpt-5.2
1 Credit = 5,714 Tokens
Long ContextVision

GPT-5.2 is the latest frontier-grade model in the GPT-5 series, offering stronger agentic and long context perfomance compared to GPT-5.1.

Qwen: Qwen3 8B
Alibaba
qwen3-8b
1 Credit = 85,470 Tokens
Mid Context

Qwen3-8B is a dense 8.2B parameter causal language model from the Qwen3 series, designed for both reasoning-heavy tasks and efficient dia…

OpenAI: GPT-5.1-Codex-Max
OpenAI
gpt-5.1-codex-max
1 Credit = 8,000 Tokens
Long ContextVision

GPT-5.1-Codex-Max is OpenAI’s latest agentic coding model, designed for long-running, high-context software development tasks.

Qwen: Qwen2.5 7B Instruct
Alibaba
qwen-2.5-7b-instruct
1 Credit = 100,000 Tokens
Mid Context

Qwen2.5 7B is the latest series of Qwen large language models.

OpenAI: GPT-5.2 (batch)
OpenAI
gpt-5.2:batch
1 Credit = 11,429 Tokens
Long ContextVision

GPT-5.2 is the latest frontier-grade model in the GPT-5 series, offering stronger agentic and long context perfomance compared to GPT-5.1.

Qwen: Qwen3 Next 80B A3B Instruct
Alibaba
qwen3-next-80b-a3b-instruct
1 Credit = 111,111 Tokens
Long Context

Qwen3-Next-80B-A3B-Instruct is an instruction-tuned chat model in the Qwen3-Next series optimized for fast, stable responses without “thi…

OpenAI: GPT-4o (2024-05-13)
OpenAI
gpt-4o-2024-05-13
1 Credit = 2,000 Tokens
Mid ContextVision

GPT-4o ("o" for "omni") is OpenAI's latest AI model, supporting both text and image inputs with text outputs.

Qwen: Qwen3 30B A3B Instruct 2507
Alibaba
qwen3-30b-a3b-instruct-2507
1 Credit = 111,111 Tokens
Long Context

Qwen3-30B-A3B-Instruct-2507 is a 30.5B-parameter mixture-of-experts language model from Qwen, with 3.3B active parameters per inference.

OpenAI: GPT-4 Turbo (batch)
OpenAI
gpt-4-turbo:batch
1 Credit = 2,000 Tokens
Mid ContextVision

The latest GPT-4 Turbo model with vision capabilities.

Qwen: Qwen3 235B A22B Instruct 2507
Alibaba
qwen3-235b-a22b-2507
1 Credit = 114,286 Tokens
Long Context

Qwen3-235B-A22B-Instruct-2507 is a multilingual, instruction-tuned mixture-of-experts language model based on the Qwen3-235B architecture…

OpenAI: GPT-5.1
OpenAI
gpt-5.1
1 Credit = 8,000 Tokens
Long ContextVision

GPT-5.1 is the latest frontier-grade model in the GPT-5 series, offering stronger general-purpose reasoning, improved instruction adheren…

Qwen: Qwen3 32B
Alibaba
qwen3-32b
1 Credit = 125,000 Tokens
Mid Context

Qwen3-32B is a dense 32.8B parameter causal language model from the Qwen3 series, optimized for both complex reasoning and efficient dial…

OpenAI: GPT-5.1-Codex
OpenAI
gpt-5.1-codex
1 Credit = 8,000 Tokens
Long ContextVision

GPT-5.1-Codex is a specialized version of GPT-5.1 optimized for software engineering and coding workflows.

Qwen: Qwen3 Coder 30B A3B Instruct
Alibaba
qwen3-coder-30b-a3b-instruct
1 Credit = 142,857 Tokens
Long ContextCode

Qwen3-Coder-30B-A3B-Instruct is a 30.5B parameter Mixture-of-Experts (MoE) model with 128 experts (8 active per forward pass), designed f…

OpenAI: GPT-5 Image Mini
OpenAI
gpt-5-image-mini
1 Credit = 4,000 Tokens
Long ContextVision

GPT-5 Image Mini combines OpenAI's advanced language capabilities, powered by [GPT-5 Mini](https://openrouter.ai/openai/gpt-5-mini), with…

OpenAI: GPT-5.1 (batch)
OpenAI
gpt-5.1:batch
1 Credit = 16,000 Tokens
Long ContextVision

GPT-5.1 is the latest frontier-grade model in the GPT-5 series, offering stronger general-purpose reasoning, improved instruction adheren…

OpenAI: GPT-5.1-Codex-Mini
OpenAI
gpt-5.1-codex-mini
1 Credit = 40,000 Tokens
Long ContextVision

GPT-5.1-Codex-Mini is a smaller and faster version of GPT-5.1-Codex

OpenAI: GPT-3.5 Turbo 16k
OpenAI
gpt-3.5-turbo-16k
1 Credit = 3,333 Tokens
Mid Context

This model offers four times the context length of gpt-3.5-turbo, allowing it to support approximately 20 pages of text in a single reque…

OpenAI: GPT-4o (2024-11-20)
OpenAI
gpt-4o-2024-11-20
1 Credit = 4,000 Tokens
Mid ContextVision

The 2024-11-20 version of GPT-4o offers a leveled-up creative writing ability with more natural, engaging, and tailored writing to improv…

OpenAI: GPT-4o (2024-08-06)
OpenAI
gpt-4o-2024-08-06
1 Credit = 4,000 Tokens
Mid ContextVision

The 2024-08-06 version of GPT-4o offers improved performance in structured outputs, with the ability to supply a JSON schema in the respo…

OpenAI: GPT-4o
OpenAI
gpt-4o
1 Credit = 4,000 Tokens
Mid ContextVision

GPT-4o ("o" for "omni") is OpenAI's latest AI model, supporting both text and image inputs with text outputs.

OpenAI: gpt-oss-safeguard-20b
OpenAI
gpt-oss-safeguard-20b
1 Credit = 133,333 Tokens
Mid Context

gpt-oss-safeguard-20b is a safety reasoning model from OpenAI built upon gpt-oss-20b.

OpenAI: o3
OpenAI
o3
1 Credit = 5,000 Tokens
Long ContextReasoning

o3 is a well-rounded and powerful model across domains.

OpenAI: GPT-4.1
OpenAI
gpt-4.1
1 Credit = 5,000 Tokens
Long ContextVision

GPT-4.1 is a flagship large language model optimized for advanced instruction following, real-world software engineering, and long-contex…

OpenAI: GPT-3.5 Turbo Instruct
OpenAI
gpt-3.5-turbo-instruct
1 Credit = 6,667 Tokens
Mid Context

This model is a variant of GPT-3.5 Turbo tuned for instructional prompts and omitting chat-related optimizations.

OpenAI: GPT-5
OpenAI
gpt-5
1 Credit = 8,000 Tokens
Long ContextVision

GPT-5 is OpenAI’s most advanced model, offering major improvements in reasoning, code quality, and user experience.

OpenAI: GPT-4o (batch)
OpenAI
gpt-4o:batch
1 Credit = 8,000 Tokens
Mid ContextVision

GPT-4o ("o" for "omni") is OpenAI's latest AI model, supporting both text and image inputs with text outputs.

OpenAI: o4 Mini High
OpenAI
o4-mini-high
1 Credit = 9,091 Tokens
Long ContextReasoning

OpenAI o4-mini-high is the same model as [o4-mini](/openai/o4-mini) with reasoning_effort set to high.

OpenAI: o4 Mini
OpenAI
o4-mini
1 Credit = 9,091 Tokens
Long ContextReasoning

OpenAI o4-mini is a compact reasoning model in the o-series, optimized for fast, cost-efficient performance while retaining strong multim…

OpenAI: o3 Mini High
OpenAI
o3-mini-high
1 Credit = 9,091 Tokens
Long ContextReasoning

OpenAI o3-mini-high is the same model as [o3-mini](/openai/o3-mini) with reasoning_effort set to high.

OpenAI: o3 Mini
OpenAI
o3-mini
1 Credit = 9,091 Tokens
Long ContextReasoning

OpenAI o3-mini is a cost-efficient language model optimized for STEM reasoning tasks, particularly excelling in science, mathematics, and…

OpenAI: o3 (batch)
OpenAI
o3:batch
1 Credit = 10,000 Tokens
Long ContextReasoning

o3 is a well-rounded and powerful model across domains.

OpenAI: GPT-4.1 (batch)
OpenAI
gpt-4.1:batch
1 Credit = 10,000 Tokens
Long ContextVision

GPT-4.1 is a flagship large language model optimized for advanced instruction following, real-world software engineering, and long-contex…

OpenAI: GPT-3.5 Turbo (older v0613)
OpenAI
gpt-3.5-turbo-0613
1 Credit = 10,000 Tokens
Mid Context

GPT-3.5 Turbo is OpenAI's fastest model.

OpenAI: GPT-5 (batch)
OpenAI
gpt-5:batch
1 Credit = 16,000 Tokens
Long ContextVision

GPT-5 is OpenAI’s most advanced model, offering major improvements in reasoning, code quality, and user experience.

OpenAI: o4 Mini (batch)
OpenAI
o4-mini:batch
1 Credit = 18,182 Tokens
Long ContextReasoning

OpenAI o4-mini is a compact reasoning model in the o-series, optimized for fast, cost-efficient performance while retaining strong multim…

OpenAI: o3 Mini (batch)
OpenAI
o3-mini:batch
1 Credit = 18,182 Tokens
Long ContextReasoning

OpenAI o3-mini is a cost-efficient language model optimized for STEM reasoning tasks, particularly excelling in science, mathematics, and…

OpenAI: GPT-3.5 Turbo
OpenAI
gpt-3.5-turbo
1 Credit = 20,000 Tokens
Mid Context

GPT-3.5 Turbo is OpenAI's fastest model.

OpenAI: GPT-4.1 Mini
OpenAI
gpt-4.1-mini
1 Credit = 25,000 Tokens
Long ContextVision

GPT-4.1 Mini is a mid-sized model delivering performance competitive with GPT-4o at substantially lower latency and cost.

OpenAI: GPT-5 Mini
OpenAI
gpt-5-mini
1 Credit = 40,000 Tokens
Long ContextVision

GPT-5 Mini is a compact version of GPT-5, designed to handle lighter-weight reasoning tasks.

OpenAI: GPT-3.5 Turbo (batch)
OpenAI
gpt-3.5-turbo:batch
1 Credit = 40,000 Tokens
Mid Context

GPT-3.5 Turbo is OpenAI's fastest model.

OpenAI: GPT-4.1 Mini (batch)
OpenAI
gpt-4.1-mini:batch
1 Credit = 50,000 Tokens
Long ContextVision

GPT-4.1 Mini is a mid-sized model delivering performance competitive with GPT-4o at substantially lower latency and cost.

OpenAI: gpt-oss-120b (batch)
OpenAI
gpt-oss-120b:batch
1 Credit = 66,667 Tokens
Mid Context

gpt-oss-120b is an open-weight, 117B-parameter Mixture-of-Experts (MoE) language model from OpenAI designed for high-reasoning, agentic,…

OpenAI: GPT-4o-mini
OpenAI
gpt-4o-mini
1 Credit = 66,667 Tokens
Mid ContextVision

GPT-4o mini is OpenAI's newest model after [GPT-4 Omni](/models/openai/gpt-4o), supporting both text and image inputs with text outputs.

OpenAI: GPT-4o-mini (2024-07-18)
OpenAI
gpt-4o-mini-2024-07-18
1 Credit = 66,667 Tokens
Mid ContextVision

GPT-4o mini is OpenAI's newest model after [GPT-4 Omni](/models/openai/gpt-4o), supporting both text and image inputs with text outputs.

OpenAI: GPT-5 Mini (batch)
OpenAI
gpt-5-mini:batch
1 Credit = 80,000 Tokens
Long ContextVision

GPT-5 Mini is a compact version of GPT-5, designed to handle lighter-weight reasoning tasks.

OpenAI: GPT-4.1 Nano
OpenAI
gpt-4.1-nano
1 Credit = 100,000 Tokens
Long ContextVision

For tasks that demand low latency, GPT‑4.1 nano is the fastest and cheapest model in the GPT-4.1 series.

OpenAI: GPT-4o-mini (batch)
OpenAI
gpt-4o-mini:batch
1 Credit = 133,333 Tokens
Mid ContextVision

GPT-4o mini is OpenAI's newest model after [GPT-4 Omni](/models/openai/gpt-4o), supporting both text and image inputs with text outputs.

OpenAI: GPT-5 Nano
OpenAI
gpt-5-nano
1 Credit = 200,000 Tokens
Long ContextVision

GPT-5-Nano is the smallest and fastest variant in the GPT-5 system, optimized for developer tools, rapid interactions, and ultra-low late…

OpenAI: gpt-oss-20b (batch)
OpenAI
gpt-oss-20b:batch
1 Credit = 200,000 Tokens
Mid Context

gpt-oss-20b is an open-weight 21B parameter model released by OpenAI under the Apache 2.0 license.

OpenAI: GPT-4.1 Nano (batch)
OpenAI
gpt-4.1-nano:batch
1 Credit = 200,000 Tokens
Long ContextVision

For tasks that demand low latency, GPT‑4.1 nano is the fastest and cheapest model in the GPT-4.1 series.

OpenAI: gpt-oss-120b
OpenAI
gpt-oss-120b
1 Credit = 270,270 Tokens
Mid Context

gpt-oss-120b is an open-weight, 117B-parameter Mixture-of-Experts (MoE) language model from OpenAI designed for high-reasoning, agentic,…

OpenAI: gpt-oss-20b
OpenAI
gpt-oss-20b
1 Credit = 333,333 Tokens
Mid Context

gpt-oss-20b is an open-weight 21B parameter model released by OpenAI under the Apache 2.0 license.

OpenAI: GPT-5 Nano (batch)
OpenAI
gpt-5-nano:batch
1 Credit = 400,000 Tokens
Long ContextVision

GPT-5-Nano is the smallest and fastest variant in the GPT-5 system, optimized for developer tools, rapid interactions, and ultra-low late…

Pick a model and ship today.

Spin up any agent framework free, route to any of these models, and only pay for the tokens you use.