Premium tier (66): Best quality, advanced reasoning
Fast tier (30): Quick responses, cost-effective
OpenAI:
•GPT-5.6 — Flagship GPT-5.6 alias — routes to GPT-5.6 Sol for complex reasoning and coding [premium, vision, reasoning]
•GPT-5.6 Sol — Frontier GPT-5.6 model for complex professional work, coding, and agentic tasks [premium, vision, reasoning]
•GPT-5.6 Terra — GPT-5.6 model balancing intelligence and cost — strong agentic execution [premium, vision, reasoning]
•GPT-5.6 Luna — Cost-efficient GPT-5.6 model for high-volume workloads and subagent loops [fast, vision, reasoning]
•GPT-5.5 — Prior frontier model — prefer GPT-5.6 Sol for new projects [premium, vision, reasoning]
•GPT-5.4 — Highly capable GPT model for coding and agentic tasks — prefer GPT-5.6 Terra at same price [premium, vision, reasoning]
•GPT-5.4 mini — Fast GPT-5.4 variant for coding, computer use, and high-volume subagent workloads [fast, vision, reasoning]
•GPT-5.4 Pro — Most powerful GPT model with maximum compute for complex reasoning [premium, vision, reasoning]
•GPT-5.2 — The best model for coding and agentic tasks across industries [premium, vision, reasoning]
•GPT-5.1 — Intelligent reasoning model with configurable reasoning effort [premium, vision, reasoning]
•GPT-5 — Previous intelligent reasoning model for coding and agentic tasks [premium, vision, reasoning]
•GPT-5 mini — A faster, cost-efficient version of GPT-5 [fast, vision]
•GPT-5 nano — Fastest, most cost-efficient version of GPT-5 [fast, vision]
•GPT-5.2 Pro — Smarter and more precise responses [premium, vision, reasoning]
•GPT-4.1 — Smartest non-reasoning model [premium, vision]
•GPT-4.1 mini — Smaller, faster version of GPT-4.1 [fast, vision]
•o3 — Reasoning model for complex tasks [premium, reasoning]
•o3 Pro — Version of o3 with more compute for better responses [premium, reasoning]
•o4 mini — Fast, cost-efficient reasoning model [fast, reasoning]
Anthropic:
•Claude Fable 5 — Most capable widely released Claude — currently unavailable due to US export controls. Use Claude Opus 4.8 instead. [premium, vision, reasoning]
•Claude Opus 4.8 — Most capable Opus — enhanced coding, agentic workflows, and long-horizon reasoning with 1M context [premium, vision, reasoning]
•Claude Opus 4.7 — Previous Opus — enhanced SWE, vision, and long-horizon agentic reasoning with 1M context [premium, vision, reasoning]
•Claude Opus 4.6 — Hybrid reasoning model with 1M context, top-tier coding and agentic performance [premium, vision, reasoning]
•Claude Opus 4.5 — Previous premium model with maximum intelligence [premium, vision, reasoning]
•Claude Sonnet 5 — Best combination of speed and intelligence — adaptive thinking with near-Opus quality [premium, vision, reasoning]
•Llama 3.3 70B — Meta's powerful open source model [premium]
•Llama 3.3 70B — Meta's powerful open source model [fast]
•Muse Spark 1.1 — Meta multimodal agent model — text, image, video, audio, and PDF input with 1M context, parallel tool calling, and built-in search. US availability on OpenRouter. Admin preview only. [premium, vision, reasoning]
Mistral:
•Mistral Medium 3.1 — Balanced Mistral model [premium]
•Mistral Large 2512 — Latest large Mistral model [premium]
•Mistral Small Creative — Creative writing focused model [fast]
Qwen:
•Qwen3 235B — Large Qwen model with 235B parameters [premium]
•Qwen3 VL 235B — Vision-language Qwen model [premium, vision]
•MiMo v2.5 — Native omnimodal model — Pro-level agentic performance at half the cost, with image & video understanding and 1M context [premium, vision, reasoning]
•MiMo v2.5 Pro — Xiaomi's 1T-parameter flagship — agentic workflows, tool calling, and advanced reasoning with 1M context [premium, reasoning]
•MiMo v2 Flash — Xiaomi's fast AI model [fast]
NVIDIA:
•Nemotron 3 Ultra — 550B MoE frontier reasoning model (55B active) — long-horizon agentic workflows, coding agents, and deep research with 1M context. Free on OpenRouter. Web search via Kunya smart search (native tool calling not supported). [premium, reasoning, free]
•Nemotron 3 Nano — Nvidia's compact free MoE model for fast chat and lightweight tasks [fast]
Moonshot:
•Kimi K3 — 2.8T MoE flagship for complex coding, knowledge work, and long-horizon agentic workflows with 1M context and native vision [premium, vision, reasoning]
•Kimi K2.7 Code — Coding-focused MoE model for long-horizon programming, agentic task decomposition, and multimodal reasoning [premium, vision, reasoning]
•Kimi K2.5 — State-of-the-art visual coding and agentic tool-calling with multimodal reasoning [premium, vision, reasoning]
StepFun:
•Step 3.5 Flash — 196B MoE reasoning model — activates 11B per token, extremely fast [fast, reasoning]
•Step 3.7 Flash — 196B MoE multimodal model — native image & video understanding, selectable reasoning depth [fast, vision, reasoning]
OpenRouter:
•Hunter Alpha — 1T parameter frontier model built for agentic multi-step reasoning [premium]
•Healer Alpha — Omni-modal frontier model with vision, hearing, reasoning, and action [premium, vision]
Nous Research:
•Hermes 4 405B — Flagship uncensored reasoning model from Nous Research — hybrid think/respond mode, low refusal rates, strong at math, code, and structured output [premium]
•Hermes 4 70B — Efficient uncensored reasoning model from Nous Research — hybrid think/respond mode, low refusal rates, strong at math, code, and structured output [fast]
Perplexity:
•Sonar Pro Search — Perplexity agentic search — multi-step research with citations (compare vs Brave / Gemini grounding) [premium, reasoning]
Tencent:
•Hy3 — 295B MoE reasoning model (21B active) — agentic workflows, 256K context, configurable chain-of-thought for coding, analysis, and tool use [premium, reasoning]
•Hy3 — 295B MoE reasoning model (21B active) — agentic workflows, 256K context, configurable chain-of-thought. Free on Kunya until July 21. [premium, reasoning, free]
Poolside:
•Laguna M.1 — Poolside flagship coding agent — tool calling, reasoning, 256K context, up to 32K output. Free on Kunya until July 28. [premium, reasoning, free]
Kunya:
•Kunya V1 — Intelligently routed model — Opus-level quality at budget cost. Routes to the best model for each request. [premium, vision, reasoning]
How to switch models: Use the model selector dropdown at the top of the Chat interface. You can change models at any time, even mid-conversation.