Curated directory of top-tier AI models, token pricing, and context limits.
Data updated at 2026-08-06 14:34 UTC · OpenRouter Models API
Model Name & ID
Provider
Max Context
Input Price (1M)
Output Price (1M)
Key Highlights
Link
GPT-5.6 Luna ProLatest
openai/gpt-5.6-luna-pro
OpenAI
1.05M
$0.10 / 1M
$0.60 / 1M
GPT-5.6 Luna Pro is the same underlying model as [GPT-5.6 Luna](https://openrouter.ai/openai/gpt-5.6-luna), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks.
Learn more in OpenAI's docs: https://developers.openai.com/api/docs/guides/reasoning#reasoning-modeReleased: 2026-07-09
GPT-5.6 Luna
openai/gpt-5.6-luna
OpenAI
1.05M
$0.10 / 1M
$0.60 / 1M
GPT-5.6 Luna is a fast, cost-efficient model in OpenAI's GPT-5.6 series. It is suited for high-volume, latency-sensitive tasks such as chat, classification, and lightweight agentic workflows, providing capable reasoning for...Released: 2026-07-09
GPT-5.6 Terra Pro
openai/gpt-5.6-terra-pro
OpenAI
1.05M
$1.00 / 1M
$6.00 / 1M
GPT-5.6 Terra Pro is the same underlying model as [GPT-5.6 Terra](https://openrouter.ai/openai/gpt-5.6-terra), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks.
Learn more in OpenAI's docs: https://developers.openai.com/api/docs/guides/reasoning#reasoning-modeReleased: 2026-07-09
GPT-5.6 Terra
openai/gpt-5.6-terra
OpenAI
1.05M
$1.00 / 1M
$6.00 / 1M
GPT-5.6 Terra is a balanced model in OpenAI's GPT-5.6 series, positioned between the flagship Sol tier and the cost-efficient Luna tier. It is suited for everyday coding, reasoning, and agentic...Released: 2026-07-09
GPT-5.6 Sol Pro
openai/gpt-5.6-sol-pro
OpenAI
1.05M
$5.00 / 1M
$30.00 / 1M
GPT-5.6 Sol Pro is the same underlying model as [GPT-5.6 Sol](https://openrouter.ai/openai/gpt-5.6-sol), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks.
Learn more in OpenAI's docs: https://developers.openai.com/api/docs/guides/reasoning#reasoning-modeReleased: 2026-07-09
Gemini 3.6 FlashLatest
google/gemini-3.6-flash
Google
1.05M
$1.50 / 1M
$7.50 / 1M
Gemini 3.6 Flash is a high-efficiency model from Google for coding, agentic workflows, and web and app development. It is designed to produce polished outputs with fewer unnecessary edits and...Released: 2026-07-21
Gemini 3.5 Flash Lite
google/gemini-3.5-flash-lite
Google
1.05M
$0.30 / 1M
$2.50 / 1M
Gemini 3.5 Flash Lite is a high-efficiency model from Google with upgraded agentic capabilities. It is suited for subagents that execute focused tasks within complex, multi-agent workflows.Released: 2026-07-21
Nano Banana 2 Lite (Gemini 3.1 Flash Lite Image)
google/gemini-3.1-flash-lite-image
Google
66K
$0.25 / 1M
$1.50 / 1M
Nano Banana 2 Lite (Gemini 3.1 Flash Lite Image) is Google's fastest, most cost-efficient Gemini image model, built for high-velocity developer pipelines and rapid-fire visual exploration. It delivers text-to-image generation...Released: 2026-06-30
Nano Banana 2 (Gemini 3.1 Flash Image)
google/gemini-3.1-flash-image
Google
131K
$0.50 / 1M
$3.00 / 1M
Gemini 3.1 Flash Image, a.k.a. "Nano Banana 2," is Google’s latest state of the art image generation and editing model, delivering Pro-level visual quality at Flash speed. It combines advanced...Released: 2026-06-18
Nano Banana Pro (Gemini 3 Pro Image)
google/gemini-3-pro-image
Google
131K
$2.00 / 1M
$12.00 / 1M
Nano Banana Pro is Google’s most advanced image-generation and editing model, built on Gemini 3 Pro. It extends the original Nano Banana with significantly improved multimodal reasoning, real-world grounding, and...Released: 2026-06-18
Claude Opus 5 (Fast)Latest
anthropic/claude-opus-5-fast
Anthropic
1M
$10.00 / 1M
$50.00 / 1M
Fast-mode variant of [Opus 5](/anthropic/claude-opus-5) - identical capabilities with higher output speed at 2x pricing relative to regular Opus 5.
Learn more in Anthropic's docs: https://platform.claude.com/docs/en/build-with-claude/fast-modeReleased: 2026-07-24
Claude Opus 5
anthropic/claude-opus-5
Anthropic
1M
$5.00 / 1M
$25.00 / 1M
Claude Opus 5 is Anthropic’s flagship model for demanding reasoning, coding, and long-horizon agentic work. It is particularly strong at end-to-end software tasks, code review and bug finding, visual analysis...Released: 2026-07-24
Claude Sonnet 5
anthropic/claude-sonnet-5
Anthropic
1M
$2.00 / 1M
$10.00 / 1M
Sonnet 5 is Anthropic's most capable Sonnet-class model, with frontier performance across coding, agents, and professional work. It supports adaptive thinking with selectable reasoning effort levels (low, medium, high, max,...Released: 2026-06-30
Claude Fable 5
anthropic/claude-fable-5
Anthropic
1M
$10.00 / 1M
$50.00 / 1M
Claude Fable 5 is a Mythos-class model from Anthropic, built for autonomous knowledge work and coding. It supports text, image, and file inputs with text output, with reasoning support and...Released: 2026-06-09
Claude Opus 4.8 (Fast)
anthropic/claude-opus-4.8-fast
Anthropic
1M
$10.00 / 1M
$50.00 / 1M
Fast-mode variant of [Opus 4.8](/anthropic/claude-opus-4.8) - identical capabilities with higher output speed at 2x pricing relative to regular Opus 4.8.
Learn more in Anthropic's docs: https://platform.claude.com/docs/en/build-with-claude/fast-modeReleased: 2026-05-27
GLM 5.2LatestOpen Source / Self-Hosted
z-ai/glm-5.2
BigModel
1.05M
$0.56 / 1M
$1.76 / 1M
GLM 5.2 is a large-scale reasoning model from Z.ai. It supports text input and output with a 1M-token context window, and is suited for long-horizon agent workflows, project-level software engineering,...Released: 2026-06-16
GLM 5.1Open Source / Self-Hosted
z-ai/glm-5.1
BigModel
205K
$0.95 / 1M
$2.99 / 1M
GLM-5.1 delivers a major leap in coding capability, with particularly significant gains in handling long-horizon tasks. Unlike previous models built around minute-level interactions, GLM-5.1 can work independently and continuously on...Released: 2026-04-07
GLM 5Open Source / Self-Hosted
z-ai/glm-5
BigModel
205K
$0.95 / 1M
$2.55 / 1M
GLM-5 is Z.ai’s flagship open-source foundation model engineered for complex systems design and long-horizon agent workflows. Built for expert developers, it delivers production-grade performance on large-scale programming tasks, rivaling leading...Released: 2026-02-11
GLM 4.7 FlashOpen Source / Self-Hosted
z-ai/glm-4.7-flash
BigModel
203K
$0.06 / 1M
$0.40 / 1M
As a 30B-class SOTA model, GLM-4.7-Flash offers a new option that balances performance and efficiency. It is further optimized for agentic coding use cases, strengthening coding capabilities, long-horizon task planning,...Released: 2026-01-19
GLM 4.7Open Source / Self-Hosted
z-ai/glm-4.7
BigModel
205K
$0.40 / 1M
$1.75 / 1M
GLM-4.7 is Z.ai’s latest flagship model, featuring upgrades in two key areas: enhanced programming capabilities and more stable multi-step reasoning/execution. It demonstrates significant improvements in executing complex agent tasks while...Released: 2025-12-22
DeepSeek V4 Flash 0731 is a sparse mixture-of-experts model from DeepSeek, with 13B active parameters out of 284B total. This re-post-trained revision is suited for coding, reasoning, and agent workflows.Released: 2026-07-31
DeepSeek V4 ProOpen Source / Self-Hosted
deepseek/deepseek-v4-pro
DeepSeek
1.05M
$0.44 / 1M
$0.87 / 1M
DeepSeek V4 Pro is a large-scale Mixture-of-Experts model from DeepSeek with 1.6T total parameters and 49B activated parameters, supporting a 1M-token context window. It is designed for advanced reasoning, coding,...Released: 2026-04-24
DeepSeek V4 Flash 0423Open Source / Self-Hosted
deepseek/deepseek-v4-flash
DeepSeek
1.05M
$0.09 / 1M
$0.18 / 1M
DeepSeek V4 Flash is an efficiency-optimized Mixture-of-Experts model from DeepSeek with 284B total parameters and 13B activated parameters, supporting a 1M-token context window. It is designed for fast inference and...Released: 2026-04-24
DeepSeek V3.2Open Source / Self-Hosted
deepseek/deepseek-v3.2
DeepSeek
164K
$0.27 / 1M
$0.40 / 1M
DeepSeek-V3.2 is a large language model designed to harmonize high computational efficiency with strong reasoning and agentic tool-use performance. It introduces DeepSeek Sparse Attention (DSA), a fine-grained sparse attention mechanism...Released: 2025-12-01
DeepSeek V3.2 ExpOpen Source / Self-Hosted
deepseek/deepseek-v3.2-exp
DeepSeek
164K
$0.27 / 1M
$0.41 / 1M
DeepSeek-V3.2-Exp is an experimental large language model released by DeepSeek as an intermediate step between V3.1 and future architectures. It introduces DeepSeek Sparse Attention (DSA), a fine-grained sparse attention mechanism...Released: 2025-09-29
Qwen3.8 MaxLatestOpen Source / Self-Hosted
qwen/qwen3.8-max
Qwen
1M
$2.00 / 1M
$6.00 / 1M
Qwen3.8 Max is the flagship model in Alibaba's Qwen3.8 series, the general-availability successor to the Qwen3.8 Max Preview. It is a multimodal reasoning model intended for complex reasoning, visual understanding,...Released: 2026-08-03
Qwen3.7 FlashOpen Source / Self-Hosted
qwen/qwen3.7-flash
Qwen
1M
$0.03 / 1M
$0.13 / 1M
Qwen3.7 Flash is a vision-language reasoning model from Alibaba. It is suited for multimodal agents, visual coding, search, and computer interaction, with strengths in object recognition, spatial understanding, and real-world...Released: 2026-07-27
Qwen3.7 PlusOpen Source / Self-Hosted
qwen/qwen3.7-plus
Qwen
1M
$0.32 / 1M
$1.28 / 1M
Qwen3.7-Plus is a cost-effective model in Alibaba's Qwen3.7 series. It supports text and image input with text output, building on the series' text capabilities with a comprehensive upgrade to its...Released: 2026-06-03
Qwen3.7 MaxOpen Source / Self-Hosted
qwen/qwen3.7-max
Qwen
1M
$1.48 / 1M
$4.43 / 1M
Qwen3.7-Max is the flagship model in Alibaba's Qwen3.7 series. It supports text input and output and is designed for agent-centric workloads, with particular strengths in coding, office and productivity tasks,...Released: 2026-05-21
Qwen3.5 Plus 2026-04-20Open Source / Self-Hosted
qwen/qwen3.5-plus-20260420
Qwen
1M
$0.30 / 1M
$1.80 / 1M
Qwen3.5 Plus (April 2026) is a large-scale multimodal language model from Alibaba. It accepts text, image, and video input and produces text output, with a 1M token context window. This...Released: 2026-04-27
Kimi K3LatestOpen Source / Self-Hosted
moonshotai/kimi-k3
Moonshot AI
1.05M
$3.00 / 1M
$15.00 / 1M
Kimi K3 is a 2.8T parameter open-weight multimodal reasoning model from Moonshot AI. It is suited for complex coding, knowledge work, and long-horizon agentic workflows, and is particularly strong at...Released: 2026-07-16
Kimi K2.7 CodeOpen Source / Self-Hosted
moonshotai/kimi-k2.7-code
Moonshot AI
262K
$0.70 / 1M
$3.50 / 1M
MoonshotAI: Kimi K2.7 Code is a coding-focused model in Moonshot AI's Kimi K2 family, built to complete end-to-end programming tasks reliably over long contexts. It uses a native multimodal mixture-of-experts...Released: 2026-06-12
Kimi K2.6Open Source / Self-Hosted
moonshotai/kimi-k2.6
Moonshot AI
262K
$0.57 / 1M
$2.40 / 1M
Kimi K2.6 is Moonshot AI's next-generation multimodal model, designed for long-horizon coding, coding-driven UI/UX generation, and multi-agent orchestration. It handles complex end-to-end coding tasks across Python, Rust, and Go, and...Released: 2026-04-20
Kimi K2.5Open Source / Self-Hosted
moonshotai/kimi-k2.5
Moonshot AI
262K
$0.57 / 1M
$2.85 / 1M
Kimi K2.5 is Moonshot AI's native multimodal model, delivering state-of-the-art visual coding capability and a self-directed agent swarm paradigm. Built on Kimi K2 with continued pretraining over approximately 15T mixed...Released: 2026-01-27
Kimi K2 ThinkingOpen Source / Self-Hosted
moonshotai/kimi-k2-thinking
Moonshot AI
262K
$0.60 / 1M
$2.50 / 1M
Kimi K2 Thinking is Moonshot AI’s most advanced open reasoning model to date, extending the K2 series into agentic, long-horizon reasoning. Built on the trillion-parameter Mixture-of-Experts (MoE) architecture introduced in...Released: 2025-11-06