Astrolabe Auto
astrolabe/auto
Managed routing that selects the lowest-cost model capable of satisfying each request.
Benchmarks
Context
—
Input / output
Varies / Varies
Coding task
Dynamic
estimated / action
Explore every billable text and embedding model available through Astrolabe. Compare benchmark quality, capabilities, context, token rates, and estimated cost for real workloads.
450 of 450 models
Astrolabe Auto
astrolabe/auto
Managed routing that selects the lowest-cost model capable of satisfying each request.
Benchmarks
Context
—
Input / output
Varies / Varies
Coding task
Dynamic
estimated / action
GPT-5.6 Sol
openai/gpt-5.6-sol
GPT-5.6 Sol is the flagship model in OpenAI's GPT-5.6 series. It is suited for complex reasoning, coding, and agentic workflows, and is particularly strong at command-line and multi-step coding tasks...
Benchmarks
Context
1.1M
Input / output
$2 / $10
Coding task
$0.050
estimated / action
Grok 4.5
x-ai/grok-4.5
Grok 4.5 is a model from SpaceXAI with frontier performance on coding, knowledge work, and STEM.
Benchmarks
Context
500K
Input / output
$2 / $6
Coding task
$0.038
estimated / action
GPT-5.6 Luna
openai/gpt-5.6-luna
GPT-5.6 Luna is a fast, cost-efficient model in OpenAI's GPT-5.6 series. It is suited for high-volume, latency-sensitive tasks such as chat, classification, and lightweight agentic workflows, providing capable reasoning for...
Benchmarks
Context
1.1M
Input / output
$0.200 / $1.2
Coding task
$0.0056
estimated / action
Gemini 3.5 Flash
google/gemini-3.5-flash
Gemini 3.5 Flash is Google's high-efficiency multimodal model, bringing near-Pro level coding and reasoning at Flash-tier cost and speed. It is highly optimized for coding proficiency and parallel agentic execution...
Benchmarks
Context
1.0M
Input / output
$1.5 / $9
Coding task
$0.042
estimated / action
DeepSeek V4 Pro 0423
deepseek/deepseek-v4-pro
DeepSeek V4 Pro is a large-scale Mixture-of-Experts model from DeepSeek with 1.6T total parameters and 49B activated parameters, supporting a 1M-token context window. It is designed for advanced reasoning, coding,...
Benchmarks
Context
1.0M
Input / output
$1.6 / $3.2
Coding task
$0.026
estimated / action
Claude Sonnet 4.6
anthropic/claude-sonnet-4.6
Sonnet 4.6 is Anthropic's most capable Sonnet-class model yet, with frontier performance across coding, agents, and professional work. It excels at iterative development, complex codebase navigation, end-to-end project management with...
Benchmarks
Context
1M
Input / output
$3 / $15
Coding task
$0.075
estimated / action
MiniMax M3
minimax/minimax-m3
MiniMax-M3 is a multimodal foundation model from MiniMax. It supports text, image, and video inputs with text output, a 1M-token context window, and is suited for long-horizon agentic work, coding,...
Benchmarks
Context
1.0M
Input / output
$0.300 / $1.2
Coding task
$0.0066
estimated / action
Kimi K2.6
moonshotai/kimi-k2.6
Kimi K2.6 is Moonshot AI's next-generation multimodal model, designed for long-horizon coding, coding-driven UI/UX generation, and multi-agent orchestration. It handles complex end-to-end coding tasks across Python, Rust, and Go, and...
Benchmarks
Context
262K
Input / output
$0.950 / $4
Coding task
$0.021
estimated / action
MiMo-V2.5-Pro
xiaomi/mimo-v2.5-pro
MiMo-V2.5-Pro is Xiaomi’s flagship model, delivering strong performance in general agentic capabilities, complex software engineering, and long-horizon tasks, with top rankings on benchmarks such as ClawEval, GDPVal, and SWE-bench Pro....
Benchmarks
Context
1.1M
Input / output
$0.435 / $0.870
Coding task
$0.0070
estimated / action
Qwen3.7 Plus
qwen/qwen3.7-plus
Qwen3.7-Plus is a cost-effective model in Alibaba's Qwen3.7 series. It supports text and image input with text output, building on the series' text capabilities with a comprehensive upgrade to its...
Benchmarks
Context
1M
Input / output
$0.320 / $1.28
Coding task
$0.0070
estimated / action
MiniMax M2.7
minimax/minimax-m2.7
MiniMax-M2.7 is a next-generation large language model designed for autonomous, real-world productivity and continuous improvement. Built to actively participate in its own evolution, M2.7 integrates advanced agentic capabilities through multi-agent...
Benchmarks
Context
205K
Input / output
$0.300 / $1.2
Coding task
$0.0066
estimated / action
Grok 4.3
x-ai/grok-4.3
Grok 4.3 is a reasoning model from SpaceXAI. It accepts text and image inputs with text output, and is suited for agentic workflows, instruction-following tasks, and applications requiring high factual...
Benchmarks
Context
1M
Input / output
$1.25 / $2.5
Coding task
$0.020
estimated / action
Gemma 4 31B
google/gemma-4-31b-it
Gemma 4 31B Instruct is Google DeepMind's 30.7B dense multimodal model supporting text and image input with text output. Features a 256K token context window, configurable thinking/reasoning mode, native function...
Benchmarks
Context
262K
Input / output
$0.090 / $0.340
Coding task
$0.0019
estimated / action
GPT-5.4
openai/gpt-5.4
GPT-5.4 is OpenAI’s latest frontier model, unifying the Codex and GPT lines into a single system. It features a 1M+ token context window (922K input, 128K output) with support for...
Benchmarks
Context
1.1M
Input / output
$2.5 / $15
Coding task
$0.070
estimated / action
Claude Fable 5
anthropic/claude-fable-5
Claude Fable 5 is a Mythos-class model from Anthropic, built for autonomous knowledge work and coding. It supports text, image, and file inputs with text output, with reasoning support and...
Benchmarks
Context
1M
Input / output
$10 / $50
Coding task
$0.250
estimated / action
Claude Opus 4.8
anthropic/claude-opus-4.8
Claude Opus 4.8 is Anthropic's most capable generally available model in the Opus family. It supports text, image, and file inputs with text output, with reasoning support and a 1M-token...
Benchmarks
Context
1M
Input / output
$5 / $25
Coding task
$0.125
estimated / action
GPT-5.5
openai/gpt-5.5
GPT-5.5 is OpenAI’s frontier model designed for complex professional workloads, building on GPT-5.4 with stronger reasoning, higher reliability, and improved token efficiency on hard tasks. It features a 1M+ token...
Benchmarks
Context
1.1M
Input / output
$5 / $30
Coding task
$0.140
estimated / action
GLM 5.2
z-ai/glm-5.2
GLM 5.2 is a large-scale reasoning model from Z.ai. It supports text input and output with a 1M-token context window, and is suited for long-horizon agent workflows, project-level software engineering,...
Benchmarks
Context
1.0M
Input / output
$1.4 / $4.4
Coding task
$0.027
estimated / action
GPT-5.6 Luna Pro
openai/gpt-5.6-luna-pro
GPT-5.6 Luna Pro is the same underlying model as [GPT-5.6 Luna](https://openrouter.ai/openai/gpt-5.6-luna), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks. Learn more in OpenAI's docs: https://developers.openai.com/api/docs/guides/reasoning#reasoning-mode
Benchmarks
Context
1.1M
Input / output
$0.200 / $1.2
Coding task
$0.0056
estimated / action
GPT-5.6 Sol Pro
openai/gpt-5.6-sol-pro
GPT-5.6 Sol Pro is the same underlying model as [GPT-5.6 Sol](https://openrouter.ai/openai/gpt-5.6-sol), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks. Learn more in OpenAI's docs: https://developers.openai.com/api/docs/guides/reasoning#reasoning-mode
Benchmarks
Context
1.1M
Input / output
$2 / $10
Coding task
$0.050
estimated / action
Kimi K2.7 Code
moonshotai/kimi-k2.7-code
MoonshotAI: Kimi K2.7 Code is a coding-focused model in Moonshot AI's Kimi K2 family, built to complete end-to-end programming tasks reliably over long contexts. It uses a native multimodal mixture-of-experts...
Benchmarks
Context
262K
Input / output
$0.706 / $3.21
Coding task
$0.017
estimated / action
GLM 5.1
z-ai/glm-5.1
GLM-5.1 delivers a major leap in coding capability, with particularly significant gains in handling long-horizon tasks. Unlike previous models built around minute-level interactions, GLM-5.1 can work independently and continuously on...
Benchmarks
Context
205K
Input / output
$0.966 / $3.04
Coding task
$0.019
estimated / action
GPT-5.4 Mini
openai/gpt-5.4-mini
GPT-5.4 mini brings the core capabilities of GPT-5.4 to a faster, more efficient model optimized for high-throughput workloads. It supports text and image inputs with strong performance across reasoning, coding,...
Benchmarks
Context
400K
Input / output
$0.750 / $4.5
Coding task
$0.021
estimated / action
DeepSeek V4 Flash 0423
deepseek/deepseek-v4-flash
DeepSeek V4 Flash is an efficiency-optimized Mixture-of-Experts model from DeepSeek with 284B total parameters and 13B activated parameters, supporting a 1M-token context window. It is designed for fast inference and...
Benchmarks
Context
1.0M
Input / output
$0.082 / $0.165
Coding task
$0.0013
estimated / action
GPT-5.4 Nano
openai/gpt-5.4-nano
GPT-5.4 nano is the most lightweight and cost-efficient variant of the GPT-5.4 family, optimized for speed-critical and high-volume tasks. It supports text and image inputs and is designed for low-latency...
Benchmarks
Context
400K
Input / output
$0.200 / $1.25
Coding task
$0.0057
estimated / action
Claude Fable 5.1 (batch)
anthropic/claude-fable-5.1:batch
Claude Fable 5.1 improves on Claude Fable 5 across the board, with the biggest gains in agentic coding, long-running agentic workflows, and knowledge work: long code refactors, front-end and visual...
Benchmarks
Context
1M
Input / output
$5 / $25
Coding task
$0.125
estimated / action
GPT-6 Astra (batch)
openai/gpt-6-astra:batch
GPT-6 Astra is OpenAI's flagship model for demanding end-to-end work. It is suited for advanced analysis, software engineering, deep research, scientific work, and document creation, with particular strengths in long-horizon...
Benchmarks
Context
1.1M
Input / output
$5 / $25
Coding task
$0.125
estimated / action
Claude Opus 4.7
anthropic/claude-opus-4.7
Opus 4.7 is the next generation of Anthropic's Opus family, built for long-running, asynchronous agents. Building on the coding and agentic strengths of Opus 4.6, it delivers stronger performance on...
Benchmarks
Context
1M
Input / output
$5 / $25
Coding task
$0.125
estimated / action
Claude Opus 5 (batch)
anthropic/claude-opus-5:batch
Claude Opus 5 is Anthropic’s flagship model for demanding reasoning, coding, and long-horizon agentic work. It is particularly strong at end-to-end software tasks, code review and bug finding, visual analysis...
Benchmarks
Context
1M
Input / output
$2.5 / $12.5
Coding task
$0.063
estimated / action
Claude Fable 5.1
anthropic/claude-fable-5.1
Claude Fable 5.1 improves on Claude Fable 5 across the board, with the biggest gains in agentic coding, long-running agentic workflows, and knowledge work: long code refactors, front-end and visual...
Benchmarks
Context
1M
Input / output
$10 / $50
Coding task
$0.250
estimated / action
Mistral Small 4
mistralai/mistral-small-2603
Mistral Small 4 is the next major release in the Mistral Small family, unifying the capabilities of several flagship Mistral models into a single system. It combines strong reasoning from...
Benchmarks
Context
262K
Input / output
$0.150 / $0.600
Coding task
$0.0033
estimated / action
Claude Opus 5
anthropic/claude-opus-5
Claude Opus 5 is Anthropic’s flagship model for demanding reasoning, coding, and long-horizon agentic work. It is particularly strong at end-to-end software tasks, code review and bug finding, visual analysis...
Benchmarks
Context
1M
Input / output
$5 / $25
Coding task
$0.125
estimated / action
GPT-6 Astra
openai/gpt-6-astra
GPT-6 Astra is OpenAI's flagship model for demanding end-to-end work. It is suited for advanced analysis, software engineering, deep research, scientific work, and document creation, with particular strengths in long-horizon...
Benchmarks
Context
1.1M
Input / output
$10 / $50
Coding task
$0.250
estimated / action
Claude Fable 5 (batch)
anthropic/claude-fable-5:batch
Claude Fable 5 is a Mythos-class model from Anthropic, built for autonomous knowledge work and coding. It supports text, image, and file inputs with text output, with reasoning support and...
Benchmarks
Context
1M
Input / output
$5 / $25
Coding task
$0.125
estimated / action
GPT-5.6 Sol (batch)
openai/gpt-5.6-sol:batch
GPT-5.6 Sol is the flagship model in OpenAI's GPT-5.6 series. It is suited for complex reasoning, coding, and agentic workflows, and is particularly strong at command-line and multi-step coding tasks...
Benchmarks
Context
1.1M
Input / output
$1 / $5
Coding task
$0.025
estimated / action
Qwen3.8 Max (0902)
qwen/qwen3.8-max-0902
Qwen3.8 Max 0902 is an updated snapshot of Qwen3.8 Max from Alibaba's Qwen team. It is a 2.4-trillion-parameter mixture-of-experts model that accepts text, image, and video input and returns text,...
Benchmarks
Context
1M
Input / output
$2 / $6
Coding task
$0.038
estimated / action
GLM 5.3 (batch)
z-ai/glm-5.3:batch
GLM-5.3 is a large-scale reasoning model from Z.ai, built for complex software engineering and long-horizon agent tasks. It supports text input and output with a 1M-token context window, and improves...
Benchmarks
Context
1.0M
Input / output
$0.700 / $2.2
Coding task
$0.014
estimated / action
GLM 5.3
z-ai/glm-5.3
GLM-5.3 is a large-scale reasoning model from Z.ai, built for complex software engineering and long-horizon agent tasks. It supports text input and output with a 1M-token context window, and improves...
Benchmarks
Context
1.3M
Input / output
$1.4 / $4.4
Coding task
$0.027
estimated / action
Grok 4.6
x-ai/grok-4.6
Grok 4.6 is SpaceXAI's smartest model with frontier performance on coding, knowledge work, and STEM.
Benchmarks
Context
500K
Input / output
$2 / $6
Coding task
$0.038
estimated / action
Kimi K3
moonshotai/kimi-k3
Kimi K3 is a 2.8T parameter open-weight multimodal reasoning model from Moonshot AI. It is suited for complex coding, knowledge work, and long-horizon agentic workflows, and is particularly strong at...
Benchmarks
Context
1.0M
Input / output
$3 / $15
Coding task
$0.075
estimated / action
Kimi K3 (batch)
moonshotai/kimi-k3:batch
Kimi K3 is a 2.8T parameter open-weight multimodal reasoning model from Moonshot AI. It is suited for complex coding, knowledge work, and long-horizon agentic workflows, and is particularly strong at...
Benchmarks
Context
1.0M
Input / output
$3 / $15
Coding task
$0.075
estimated / action
GPT-5.6 Terra (batch)
openai/gpt-5.6-terra:batch
GPT-5.6 Terra is a balanced model in OpenAI's GPT-5.6 series, positioned between the flagship Sol tier and the cost-efficient Luna tier. It is suited for everyday coding, reasoning, and agentic...
Benchmarks
Context
1.1M
Input / output
$1 / $6
Coding task
$0.028
estimated / action
GPT-5.6 Terra
openai/gpt-5.6-terra
GPT-5.6 Terra is a balanced model in OpenAI's GPT-5.6 series, positioned between the flagship Sol tier and the cost-efficient Luna tier. It is suited for everyday coding, reasoning, and agentic...
Benchmarks
Context
1.1M
Input / output
$2 / $12
Coding task
$0.056
estimated / action
Gemini 3.8 Flash (batch)
google/gemini-3.8-flash:batch
Gemini 3.8 Flash is Google's most intelligent Flash model with significant gains from 3.7 Flash across software engineering, agentic tasks, and multi-step reasoning.
Benchmarks
Context
1.0M
Input / output
$0.375 / $1.88
Coding task
$0.0094
estimated / action
Gemini 3.8 Flash
google/gemini-3.8-flash
Gemini 3.8 Flash is Google's most intelligent Flash model with significant gains from 3.7 Flash across software engineering, agentic tasks, and multi-step reasoning.
Benchmarks
Context
1.0M
Input / output
$0.750 / $3.75
Coding task
$0.019
estimated / action
GLM 5.3 Flash (batch)
z-ai/glm-5.3-flash:batch
GLM-5.3-Flash is a native multimodal model from Z.ai. It is suited for efficient coding and long-horizon agent tasks. Its hybrid sparse and linear attention architecture maintains accurate long-context behavior while...
Benchmarks
Context
1.0M
Input / output
$0.075 / $0.250
Coding task
$0.0015
estimated / action
GLM 5.3 Flash
z-ai/glm-5.3-flash
GLM-5.3-Flash is a native multimodal model from Z.ai. It is suited for efficient coding and long-horizon agent tasks. Its hybrid sparse and linear attention architecture maintains accurate long-context behavior while...
Benchmarks
Context
1.3M
Input / output
$0.090 / $0.300
Coding task
$0.0018
estimated / action
Claude Opus 4.8 (batch)
anthropic/claude-opus-4.8:batch
Claude Opus 4.8 is Anthropic's most capable generally available model in the Opus family. It supports text, image, and file inputs with text output, with reasoning support and a 1M-token...
Benchmarks
Context
1M
Input / output
$2.5 / $12.5
Coding task
$0.063
estimated / action
Gemini 3.7 Flash (batch)
google/gemini-3.7-flash:batch
Gemini 3.7 Flash is a multimodal model from Google for fast agentic workflows, coding, and complex multi-step reasoning. It is designed for tasks that require responsive performance and reliable multi-step...
Benchmarks
Context
1.0M
Input / output
$0.375 / $1.88
Coding task
$0.0094
estimated / action
One model ID
Use astrolabe/auto to route each request to the lowest-cost capable model in this catalog.