Model catalog
Every model available in AiFuse, with the price you pay. No subscriptions.
Claude Fable 5.1
Anthropic's latest Fable release for the hardest reasoning and long-horizon agentic work.
Claude Opus 5.5
Anthropic's Opus 5.5 — adaptive thinking always on, with effort controlling how deeply it reasons.
Claude Sonnet 5.5
Anthropic's Sonnet 5.5 — fast, capable everyday model, strong at coding.
Claude Fable 5
Anthropic's most advanced model for the hardest reasoning and long-horizon agentic work.
Claude Sonnet 5
Balanced, versatile workhorse for daily use and scaled production.
Claude Opus 5
Anthropic's latest flagship for deep reasoning and long-horizon agentic coding. 1M context.
Claude Opus 4.8
Highly autonomous flagship for complex reasoning, long-horizon agentic coding, and knowledge work.
Claude Haiku 4.5
Fastest and most affordable for quick answers
DeepSeek V4 Pro
Flagship 1.6T-parameter MoE reasoning for complex coding, multi-step analysis, and long-horizon agentic work, at frontier quality.
DeepSeek V4.1 Flash
DeepSeek's fast V4.1 release — low-cost general chat with aggressive prompt caching.
DeepSeek V4 Flash
Fast and affordable with strong coding and analysis. 128K context.
Gemini 3.1 Pro
Google's flagship reasoning model with strong multimodal and long-context performance.
Gemini 3.8 Flash
Google's headline Flash model. Fast, high quality, 1M context.
Gemini 3.5 Flash-Lite
Google's cheapest Flash tier. Fast and low cost, 1M context.
GLM 5.3
Z.ai's open model for long-context coding and agentic work, with a 1M-token context window.
Grok 4.7
xAI's latest Grok release for reasoning, coding and analysis.
Grok Build
Fast coding model for agentic workflows. 256K context.
Grok 4.6
Frontier reasoning model. 500K context.
Grok 4.5
Frontier reasoning model. 200K context.
Grok 4.3
xAI's most powerful model with extended reasoning capability.
Mistral Medium 3.5
Mistral's newest frontier model for agentic work and coding, with image input — newer and stronger than Large 3.
Mistral Large 3
Mistral's open-weight general model, multilingual with image input — older and cheaper than Medium 3.5, despite the name.
Codestral
Mistral's code specialist — fast, low-cost code generation.
Mistral Small 4
Mistral's lightweight model — fast and very low cost, with image input.
Kimi K3
Moonshot's reasoning model with a 1M-token context and native vision. Strong multi-step problem solving.
Kimi K2.7 Code
Moonshot's coding-specialist Kimi with a 1M-token context. Tuned for code generation, editing, and repo-scale reasoning.
GPT-6 Astra
OpenAI's GPT-6 flagship for the most complex reasoning and agentic work.
GPT-6.1 Sol
OpenAI's GPT-6.1 model for complex reasoning and agentic work.
GPT-6 Luna
OpenAI's fast, low-cost GPT-6 model for high-volume tasks.
GPT-5.6 Sol
OpenAI's top-tier GPT-5.6 model for the most complex reasoning and agentic tasks.
GPT-5.5 Pro
OpenAI's specialist GPT-5.5 Pro model for the hardest, highest-stakes reasoning.
GPT-5.6 Terra
OpenAI's balanced GPT-5.6 model for everyday complex tasks and reasoning.
GPT-5.6 Luna
OpenAI's fast, cost-efficient GPT-5.6 model for high-volume tasks.
GPT-5.5
Flagship reasoning model. 1.05M context, 128K output. Highest-end OpenAI tier.
GPT-5.4
OpenAI's flagship model for complex tasks and advanced reasoning.
GPT-5.4 Nano
Ultra-fast and ultra-cheap for simple and repetitive tasks.
Perplexity Sonar Pro
Advanced web-grounded search with deeper reasoning and more citations. 200K context.
Perplexity Sonar Reasoning Pro
Web-grounded reasoning with chain-of-thought and real-time citations. 128K context.
Perplexity Sonar
Web-grounded AI search with real-time citations. 127K context.