Astrolabe Auto
astrolabe/auto
Managed routing that selects the lowest-cost model capable of satisfying each request.
Benchmarks
Context
—
Input / output
Varies / Varies
Coding task
Dynamic
estimated / action
Explore every billable text and embedding model available through Astrolabe. Compare benchmark quality, capabilities, context, token rates, and estimated cost for real workloads.
345 of 345 models
Astrolabe Auto
astrolabe/auto
Managed routing that selects the lowest-cost model capable of satisfying each request.
Benchmarks
Context
—
Input / output
Varies / Varies
Coding task
Dynamic
estimated / action
GPT-5.6 Sol
openai/gpt-5.6-sol
GPT-5.6 Sol is the flagship model in OpenAI's GPT-5.6 series. It is suited for complex reasoning, coding, and agentic workflows, and is particularly strong at command-line and multi-step coding tasks...
Benchmarks
Context
1.1M
Input / output
$5 / $30
Coding task
$0.140
estimated / action
Grok 4.5
x-ai/grok-4.5
Grok 4.5 is SpaceXAI's smartest model with frontier performance on coding, knowledge work, and STEM.
Benchmarks
Context
500K
Input / output
$2 / $6
Coding task
$0.038
estimated / action
GPT-5.6 Luna
openai/gpt-5.6-luna
GPT-5.6 Luna is a fast, cost-efficient model in OpenAI's GPT-5.6 series. It is suited for high-volume, latency-sensitive tasks such as chat, classification, and lightweight agentic workflows, providing capable reasoning for...
Benchmarks
Context
1.1M
Input / output
$0.100 / $0.600
Coding task
$0.0028
estimated / action
Gemini 3.5 Flash
google/gemini-3.5-flash
Gemini 3.5 Flash is Google's high-efficiency multimodal model, bringing near-Pro level coding and reasoning at Flash-tier cost and speed. It is highly optimized for coding proficiency and parallel agentic execution...
Benchmarks
Context
1.0M
Input / output
$1.5 / $9
Coding task
$0.042
estimated / action
Claude Sonnet 4.6
anthropic/claude-sonnet-4.6
Sonnet 4.6 is Anthropic's most capable Sonnet-class model yet, with frontier performance across coding, agents, and professional work. It excels at iterative development, complex codebase navigation, end-to-end project management with...
Benchmarks
Context
1M
Input / output
$3 / $15
Coding task
$0.075
estimated / action
Kimi K2.6
moonshotai/kimi-k2.6
Kimi K2.6 is Moonshot AI's next-generation multimodal model, designed for long-horizon coding, coding-driven UI/UX generation, and multi-agent orchestration. It handles complex end-to-end coding tasks across Python, Rust, and Go, and...
Benchmarks
Context
262K
Input / output
$0.600 / $3.41
Coding task
$0.016
estimated / action
DeepSeek V4 Pro
deepseek/deepseek-v4-pro
DeepSeek V4 Pro is a large-scale Mixture-of-Experts model from DeepSeek with 1.6T total parameters and 49B activated parameters, supporting a 1M-token context window. It is designed for advanced reasoning, coding,...
Benchmarks
Context
1.0M
Input / output
$0.435 / $0.870
Coding task
$0.0070
estimated / action
MiniMax M3
minimax/minimax-m3
MiniMax-M3 is a multimodal foundation model from MiniMax. It supports text, image, and video inputs with text output, a 1M-token context window, and is suited for long-horizon agentic work, coding,...
Benchmarks
Context
1.0M
Input / output
$0.300 / $1.2
Coding task
$0.0066
estimated / action
MiMo-V2.5-Pro
xiaomi/mimo-v2.5-pro
MiMo-V2.5-Pro is Xiaomi’s flagship model, delivering strong performance in general agentic capabilities, complex software engineering, and long-horizon tasks, with top rankings on benchmarks such as ClawEval, GDPVal, and SWE-bench Pro....
Benchmarks
Context
1.1M
Input / output
$0.435 / $0.870
Coding task
$0.0070
estimated / action
Qwen3.7 Plus
qwen/qwen3.7-plus
Qwen3.7-Plus is a cost-effective model in Alibaba's Qwen3.7 series. It supports text and image input with text output, building on the series' text capabilities with a comprehensive upgrade to its...
Benchmarks
Context
1M
Input / output
$0.320 / $1.28
Coding task
$0.0070
estimated / action
MiniMax M2.7
minimax/minimax-m2.7
MiniMax-M2.7 is a next-generation large language model designed for autonomous, real-world productivity and continuous improvement. Built to actively participate in its own evolution, M2.7 integrates advanced agentic capabilities through multi-agent...
Benchmarks
Context
205K
Input / output
$0.250 / $1
Coding task
$0.0055
estimated / action
Grok 4.3
x-ai/grok-4.3
Grok 4.3 is a reasoning model from xAI. It accepts text and image inputs with text output, and is suited for agentic workflows, instruction-following tasks, and applications requiring high factual...
Benchmarks
Context
1M
Input / output
$1.25 / $2.5
Coding task
$0.020
estimated / action
Gemma 4 31B
google/gemma-4-31b-it
Gemma 4 31B Instruct is Google DeepMind's 30.7B dense multimodal model supporting text and image input with text output. Features a 256K token context window, configurable thinking/reasoning mode, native function...
Benchmarks
Context
262K
Input / output
$0.100 / $0.340
Coding task
$0.0020
estimated / action
Claude Fable 5
anthropic/claude-fable-5
Claude Fable 5 is a Mythos-class model from Anthropic, built for autonomous knowledge work and coding. It supports text, image, and file inputs with text output, with reasoning support and...
Benchmarks
Context
1M
Input / output
$10 / $50
Coding task
$0.250
estimated / action
Claude Opus 4.8
anthropic/claude-opus-4.8
Claude Opus 4.8 is Anthropic's most capable generally available model in the Opus family. It supports text, image, and file inputs with text output, with reasoning support and a 1M-token...
Benchmarks
Context
1M
Input / output
$5 / $25
Coding task
$0.125
estimated / action
GPT-5.5
openai/gpt-5.5
GPT-5.5 is OpenAI’s frontier model designed for complex professional workloads, building on GPT-5.4 with stronger reasoning, higher reliability, and improved token efficiency on hard tasks. It features a 1M+ token...
Benchmarks
Context
1.1M
Input / output
$5 / $30
Coding task
$0.140
estimated / action
GLM 5.2
z-ai/glm-5.2
GLM 5.2 is a large-scale reasoning model from Z.ai. It supports text input and output with a 1M-token context window, and is suited for long-horizon agent workflows, project-level software engineering,...
Benchmarks
Context
1.0M
Input / output
$0.284 / $0.893
Coding task
$0.0055
estimated / action
GPT-5.4
openai/gpt-5.4
GPT-5.4 is OpenAI’s latest frontier model, unifying the Codex and GPT lines into a single system. It features a 1M+ token context window (922K input, 128K output) with support for...
Benchmarks
Context
1.1M
Input / output
$2.5 / $15
Coding task
$0.070
estimated / action
Kimi K2.7 Code
moonshotai/kimi-k2.7-code
MoonshotAI: Kimi K2.7 Code is a coding-focused model in Moonshot AI's Kimi K2 family, built to complete end-to-end programming tasks reliably over long contexts. It uses a native multimodal mixture-of-experts...
Benchmarks
Context
262K
Input / output
$0.730 / $3.5
Coding task
$0.018
estimated / action
DeepSeek V4 Flash
deepseek/deepseek-v4-flash
DeepSeek V4 Flash is an efficiency-optimized Mixture-of-Experts model from DeepSeek with 284B total parameters and 13B activated parameters, supporting a 1M-token context window. It is designed for fast inference and...
Benchmarks
Context
1.0M
Input / output
$0.140 / $0.280
Coding task
$0.0022
estimated / action
GLM 5.1
z-ai/glm-5.1
GLM-5.1 delivers a major leap in coding capability, with particularly significant gains in handling long-horizon tasks. Unlike previous models built around minute-level interactions, GLM-5.1 can work independently and continuously on...
Benchmarks
Context
205K
Input / output
$0.966 / $3.04
Coding task
$0.019
estimated / action
GPT-5.4 Mini
openai/gpt-5.4-mini
GPT-5.4 mini brings the core capabilities of GPT-5.4 to a faster, more efficient model optimized for high-throughput workloads. It supports text and image inputs with strong performance across reasoning, coding,...
Benchmarks
Context
400K
Input / output
$0.750 / $4.5
Coding task
$0.021
estimated / action
GPT-5.4 Nano
openai/gpt-5.4-nano
GPT-5.4 nano is the most lightweight and cost-efficient variant of the GPT-5.4 family, optimized for speed-critical and high-volume tasks. It supports text and image inputs and is designed for low-latency...
Benchmarks
Context
400K
Input / output
$0.200 / $1.25
Coding task
$0.0057
estimated / action
GPT-5.6 Luna Pro
openai/gpt-5.6-luna-pro
GPT-5.6 Luna Pro is the same underlying model as [GPT-5.6 Luna](https://openrouter.ai/openai/gpt-5.6-luna), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks. Learn more in OpenAI's docs: https://developers.openai.com/api/docs/guides/reasoning#reasoning-mode
Benchmarks
Context
1.1M
Input / output
$0.100 / $0.600
Coding task
$0.0028
estimated / action
GPT-5.6 Sol Pro
openai/gpt-5.6-sol-pro
GPT-5.6 Sol Pro is the same underlying model as [GPT-5.6 Sol](https://openrouter.ai/openai/gpt-5.6-sol), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks. Learn more in OpenAI's docs: https://developers.openai.com/api/docs/guides/reasoning#reasoning-mode
Benchmarks
Context
1.1M
Input / output
$5 / $30
Coding task
$0.140
estimated / action
Claude Opus 5
anthropic/claude-opus-5
Claude Opus 5 is Anthropic’s flagship model for demanding reasoning, coding, and long-horizon agentic work. It is particularly strong at end-to-end software tasks, code review and bug finding, visual analysis...
Benchmarks
Context
1M
Input / output
$5 / $25
Coding task
$0.125
estimated / action
Mistral Small 4
mistralai/mistral-small-2603
Mistral Small 4 is the next major release in the Mistral Small family, unifying the capabilities of several flagship Mistral models into a single system. It combines strong reasoning from...
Benchmarks
Context
262K
Input / output
$0.150 / $0.600
Coding task
$0.0033
estimated / action
Kimi K3
moonshotai/kimi-k3
Kimi K3 is a 2.8T parameter open-weight multimodal reasoning model from Moonshot AI. It is suited for complex coding, knowledge work, and long-horizon agentic workflows, and is particularly strong at...
Benchmarks
Context
1.0M
Input / output
$3 / $15
Coding task
$0.075
estimated / action
GPT-5.6 Terra
openai/gpt-5.6-terra
GPT-5.6 Terra is a balanced model in OpenAI's GPT-5.6 series, positioned between the flagship Sol tier and the cost-efficient Luna tier. It is suited for everyday coding, reasoning, and agentic...
Benchmarks
Context
1.1M
Input / output
$1 / $6
Coding task
$0.028
estimated / action
Claude Sonnet 5
anthropic/claude-sonnet-5
Sonnet 5 is Anthropic's most capable Sonnet-class model, with frontier performance across coding, agents, and professional work. It supports adaptive thinking with selectable reasoning effort levels (low, medium, high, max,...
Benchmarks
Context
1M
Input / output
$2 / $10
Coding task
$0.050
estimated / action
Claude Opus 4.7
anthropic/claude-opus-4.7
Opus 4.7 is the next generation of Anthropic's Opus family, built for long-running, asynchronous agents. Building on the coding and agentic strengths of Opus 4.6, it delivers stronger performance on...
Benchmarks
Context
1M
Input / output
$5 / $25
Coding task
$0.125
estimated / action
Muse Spark 1.1
meta/muse-spark-1.1
Muse Spark 1.1 is a multimodal reasoning model from Meta, built for agentic tasks. It accepts text, images, video, audio, and PDF documents and returns text, with a 1M-token context...
Benchmarks
Context
1.0M
Input / output
$1.25 / $4.25
Coding task
$0.025
estimated / action
DeepSeek V4 Flash 0731
deepseek/deepseek-v4-flash-0731
DeepSeek V4 Flash 0731 is a sparse mixture-of-experts model from DeepSeek, with 13B active parameters out of 284B total. This re-post-trained revision is suited for coding, reasoning, and agent workflows.
Benchmarks
Context
1.0M
Input / output
$0.090 / $0.180
Coding task
$0.0014
estimated / action
Gemini 3.6 Flash
google/gemini-3.6-flash
Gemini 3.6 Flash is a high-efficiency model from Google for coding, agentic workflows, and web and app development. It is designed to produce polished outputs with fewer unnecessary edits and...
Benchmarks
Context
1.0M
Input / output
$1.5 / $7.5
Coding task
$0.037
estimated / action
Gemini 3.1 Pro Preview
google/gemini-3.1-pro-preview
Gemini 3.1 Pro Preview is Google’s frontier reasoning model, delivering enhanced software engineering performance, improved agentic reliability, and more efficient token usage across complex workflows. Building on the multimodal foundation...
Benchmarks
Context
1.0M
Input / output
$2 / $12
Coding task
$0.056
estimated / action
Qwen3.7 Max
qwen/qwen3.7-max
Qwen3.7-Max is the flagship model in Alibaba's Qwen3.7 series. It supports text input and output and is designed for agent-centric workloads, with particular strengths in coding, office and productivity tasks,...
Benchmarks
Context
1M
Input / output
$1.48 / $4.43
Coding task
$0.028
estimated / action
Hy3 preview
tencent/hy3-preview
Hy3 preview is a high-efficiency Mixture-of-Experts model from Tencent designed for agentic workflows and production use. It supports configurable reasoning levels across disabled, low, and high modes, allowing it to...
Benchmarks
Context
262K
Input / output
$0.063 / $0.210
Coding task
$0.0013
estimated / action
Nex-N2-Pro
nex-agi/nex-n2-pro
Nex-N2-Pro is an agentic mixture-of-experts model from Nex AGI, with 17B active parameters out of 397B total. Built on the Qwen3.5 architecture, it accepts text and image input and produces...
Benchmarks
Context
262K
Input / output
$0.250 / $1
Coding task
$0.0055
estimated / action
Inkling
thinkingmachines/inkling
Inkling is an open-weight multimodal mixture-of-experts model from Thinking Machines Lab, with 41B active parameters out of 975B total. It is designed for general-purpose reasoning, coding, agentic and tool-use systems,...
Benchmarks
Context
1.0M
Input / output
$1 / $4.05
Coding task
$0.022
estimated / action
Inkling Small
thinkingmachines/inkling-small
Inkling Small is an open-weight multimodal mixture-of-experts model from Thinking Machines Lab, with 12B active parameters out of 276B total. It is positioned as the smaller, more efficient member of...
Benchmarks
Context
524K
Input / output
$0.500 / $1.2
Coding task
$0.0086
estimated / action
Qwen3.6 Plus
qwen/qwen3.6-plus
Qwen 3.6 Plus builds on a hybrid architecture that combines efficient linear attention with sparse mixture-of-experts routing, enabling strong scalability and high-performance inference. Compared to the 3.5 series, it delivers...
Benchmarks
Context
1M
Input / output
$0.325 / $1.95
Coding task
$0.0091
estimated / action
Grok Build 0.1
x-ai/grok-build-0.1
Grok Build 0.1 is xAI’s fast coding model trained specifically for agentic software engineering workflows. It supports text and image inputs with text output, and is optimized for interactive coding...
Benchmarks
Context
256K
Input / output
$1 / $2
Coding task
$0.016
estimated / action
MiMo-V2.5
xiaomi/mimo-v2.5
MiMo-V2.5 is a native omnimodal model by Xiaomi. It delivers Pro-level agentic performance at roughly half the inference cost, while surpassing MiMo-V2-Omni in multimodal perception across image and video understanding...
Benchmarks
Context
1.1M
Input / output
$0.140 / $0.280
Coding task
$0.0022
estimated / action
Qwen3.6 27B
qwen/qwen3.6-27b
Qwen3.6 27B is a dense 27-billion-parameter language model from the Qwen Team at Alibaba, released in April 2026. It features hybrid multimodal capabilities — accepting text, image, and video inputs...
Benchmarks
Context
262K
Input / output
$0.300 / $2
Coding task
$0.0090
estimated / action
Nemotron 3 Ultra
nvidia/nemotron-3-ultra-550b-a55b
NVIDIA Nemotron 3 Ultra is an open frontier-reasoning and orchestration model from NVIDIA, with 55B active parameters out of 550B total (MoE). Built on a hybrid Transformer-Mamba mixture-of-experts architecture, it...
Benchmarks
Context
512K
Input / output
$0.600 / $3.6
Coding task
$0.017
estimated / action
Gemini 3.5 Flash Lite
google/gemini-3.5-flash-lite
Gemini 3.5 Flash Lite is a high-efficiency model from Google with upgraded agentic capabilities. It is suited for subagents that execute focused tasks within complex, multi-agent workflows.
Benchmarks
Context
1.0M
Input / output
$0.300 / $2.5
Coding task
$0.010
estimated / action
KAT-Coder-Pro V2
kwaipilot/kat-coder-pro-v2
KAT-Coder-Pro V2 is the latest high-performance model in KwaiKAT’s KAT-Coder series, designed for complex enterprise-grade software engineering and SaaS integration. It builds on the agentic coding strengths of earlier versions,...
Benchmarks
Context
262K
Input / output
$0.300 / $1.2
Coding task
$0.0066
estimated / action
GPT-5.1
openai/gpt-5.1
GPT-5.1 is the latest frontier-grade model in the GPT-5 series, offering stronger general-purpose reasoning, improved instruction adherence, and a more natural conversational style compared to GPT-5. It uses adaptive reasoning...
Benchmarks
Context
400K
Input / output
$1.25 / $10
Coding task
$0.042
estimated / action
Claude Sonnet 4.5
anthropic/claude-sonnet-4.5
Claude Sonnet 4.5 is Anthropic’s most advanced Sonnet model to date, optimized for real-world agents and coding workflows. It delivers state-of-the-art performance on coding benchmarks such as SWE-bench Verified, with...
Benchmarks
Context
1M
Input / output
$3 / $15
Coding task
$0.075
estimated / action
One model ID
Use astrolabe/auto to route each request to the lowest-cost capable model in this catalog.