anthropic/claude-opus-5
Strongest Claude Opus model for coding, agents, and professional work
- Context
- 1M
- Input
- $5/M
- Output
- $25/M
Filter 494 source-linked models by creator, price, context, modality, and published benchmark coverage.
anthropic/claude-opus-5
Strongest Claude Opus model for coding, agents, and professional work
poolside/laguna-s-2.1
Agentic coding model from Poolside in the XS size class for local deployment
google/gemini-3.6-flash
Fast Gemini model balancing multimodal reasoning, tool use, and cost
google/gemini-3.5-flash-lite
Fast Gemini model balancing multimodal reasoning, tool use, and cost
alibaba/qwen3.8-max-preview
Preview Qwen flagship for million-token multimodal reasoning and long-horizon agentic workflows
moonshotai/kimi-k3
Multimodal Kimi model with 1M context and toggleable max-effort thinking for long-horizon agent work
thinkingmachines/inkling
Multimodal MoE reasoning model (975B total, 41B active) for text, image, and audio
openai/gpt-5.6-luna
Cost-efficient GPT-5.6 model for fast, high-volume workloads
openai/gpt-5.6-terra
Balanced GPT-5.6 model for capable, cost-efficient everyday work
openai/gpt-5.6-sol
Frontier GPT-5.6 model for complex professional work, coding, and agentic workflows
xai/grok-4.5
xAI's latest Grok for chat, coding, agentic tools, and lower hallucination risk
tencent/hy3
Tencent Hy reasoning model for coding, instruction following, and agent tasks
openai/gpt-realtime-2.1
Realtime speech-to-speech model with configurable reasoning, tool use, and robust voice-agent behavior
poolside/laguna-xs-2.1
Agentic coding model from Poolside in the XS size class for local deployment
anthropic/claude-sonnet-5
Everyday Claude agent model for coding, planning, browsing, and general work
google/gemini-omni-flash-preview
Video generation and editing model for fast, conversational text- and image-to-video workflows
meituan/longcat-2.0
Meituan LongCat-2.0, a reasoning model with tool calling and a 1M-token context window
deepreinforce/ornith-1.0-31b
Open coding-reasoning model for repository tasks and self-improving agents
deepreinforce/ornith-1.0-35b
Large coding-reasoning model for agentic software tasks and RL search
deepreinforce/ornith-1.0-9b
Open coding-reasoning model for repository tasks and self-improving agents
deepreinforce/ornith-1.0-397b
Large coding-reasoning model for agentic software tasks and RL search
sakana/fugu-ultra
Quality-first multi-agent model for hard research, analysis, and competitions
sakana/fugu
Multi-agent model for routing expert agents across complex analytical tasks
zhipuai/glm-5.2
Open flagship GLM for long-horizon coding agents and million-token context work
moonshotai/kimi-k2.7-code
Coding-focused Kimi model, stronger on long-horizon repo work with less overthinking
moonshotai/kimi-k2.7-code-highspeed
Lower-latency Kimi Code variant for interactive edits and coding-agent loops
anthropic/claude-fable-5
Claude model for creative writing, analysis, and controlled agent workflows
cohere/north-mini-code-1-0
Cohere coding model for practical software engineering and agentic edits
google/gemini-3.5-live-translate-preview
Low-latency audio-to-audio model for real-time speech translation across 70+ languages
xiaomi/mimo-v2.5-pro-ultraspeed
MiMo pro model for strong multimodal reasoning and agent execution
nvidia/nemotron-3-ultra-550b-a55b
Largest Nemotron 3 model for maximum open-weight reasoning and agent accuracy
nvidia/nemotron-3.5-content-safety
Safety model for policy screening, moderation, and risk-aware routing workflows
microsoft/mai-code-1-flash
Microsoft coding model built for fast, efficient assistance in everyday developer workflows
alibaba/qwen3.7-plus
Multimodal Qwen workhorse for long-context agents, visual inputs, and coding
minimax/MiniMax-M3
MiniMax multimodal model for long-context coding, perception, and agent planning
stepfun/step-3.7-flash
Newer StepFun flash model for faster agents, coding, and multimodal prompts
google/gemini-3.1-flash-image
Image model for prompt-driven generation, editing, and visual design workflows
anthropic/claude-opus-4-8
Top Claude Opus tier for the hardest reasoning, coding, and long-horizon agents
alibaba/qwen3.7-max
Qwen frontier model tuned for agent frameworks, coding assistants, and long tasks
cohere/command-a-plus-05-2026
Cohere's stronger command model for multilingual agents and enterprise workflows
google/gemini-3.5-flash
Fast Gemini model balancing multimodal reasoning, tool use, and cost
google/gemini-flash-latest
Fast Gemini model balancing multimodal reasoning, tool use, and cost
google/gemini-3.1-flash-lite
Low-latency Gemini model for high-volume multimodal and agent workloads
google/gemini-flash-lite-latest
Low-latency Gemini model for high-volume multimodal and agent workloads
openai/gpt-5.5-instant
Compact GPT model for low-latency assistance and high-volume workloads
mistral/mistral-medium-latest
Balanced Mistral model for enterprise assistants, multilingual work, and tools
mistral/mistral-medium-2604
Balanced Mistral model for enterprise assistants, multilingual work, and tools
poolside/laguna-m.1
Poolside's open-weight model for agentic coding and long-horizon work
poolside/laguna-xs.2
Agentic coding model from Poolside in the XS size class for local deployment
nvidia/nemotron-3-nano-omni-30b-a3b-reasoning
Open Nemotron omni model combining reasoning with text, vision, and audio
| Model | Creator | Input types | Context | Input / Output | Released | Compare |
|---|---|---|---|---|---|---|
| Claude Opus 5anthropic/claude-opus-5 | 1M | $5 / $25 | 2026-07-24 | |||
| Laguna S 2.1poolside/laguna-s-2.1 | 1.04858M | $0.1 / $0.2 | 2026-07-21 | |||
| Gemini 3.6 Flashgoogle/gemini-3.6-flash | 1.04858M | $1.5 / $7.5 | 2026-07-21 | |||
| Gemini 3.5 Flash Litegoogle/gemini-3.5-flash-lite | 1.04858M | $0.3 / $2.5 | 2026-07-21 | |||
| Qwen3.8 Max Previewalibaba/qwen3.8-max-preview | 1M | - / - | 2026-07-19 | |||
| Kimi K3moonshotai/kimi-k3 | 1.04858M | $3 / $15 | 2026-07-16 | |||
| Inklingthinkingmachines/inkling | 1.04858M | $1.87 / $4.68 | 2026-07-15 | |||
| GPT-5.6 Lunaopenai/gpt-5.6-luna | 1.05M | $1 / $6 | 2026-07-09 | |||
| GPT-5.6 Terraopenai/gpt-5.6-terra | 1.05M | $2.5 / $15 | 2026-07-09 | |||
| GPT-5.6 Solopenai/gpt-5.6-sol | 1.05M | $5 / $30 | 2026-07-09 | |||
| Grok 4.5xai/grok-4.5 | 500K | $2 / $6 | 2026-07-08 | |||
| Hy3tencent/hy3 | 256K | $0.132 / $0.528 | 2026-07-06 | |||
| GPT-Realtime-2.1openai/gpt-realtime-2.1 | 128K | $4 / $24 | 2026-07-06 | |||
| Laguna XS 2.1poolside/laguna-xs-2.1 | 262.144K | $0.06 / $0.12 | 2026-07-02 | |||
| Claude Sonnet 5anthropic/claude-sonnet-5 | 1M | $2 / $10 | 2026-06-30 | |||
| Gemini Omni Flash Previewgoogle/gemini-omni-flash-preview | 1.04858M | $1.5 / $17.5 | 2026-06-30 | |||
| LongCat-2.0meituan/longcat-2.0 | 1M | $0.3 / $1.2 | 2026-06-30 | |||
| Ornith 1.0 31Bdeepreinforce/ornith-1.0-31b | 262.144K | - / - | 2026-06-25 | |||
| Ornith 1.0 35Bdeepreinforce/ornith-1.0-35b | 262.144K | - / - | 2026-06-25 | |||
| Ornith 1.0 9Bdeepreinforce/ornith-1.0-9b | 262.144K | - / - | 2026-06-25 | |||
| Ornith 1.0 397Bdeepreinforce/ornith-1.0-397b | 262.144K | - / - | 2026-06-25 | |||
| Fugu Ultrasakana/fugu-ultra | 1M | $5 / $30 | 2026-06-15 | |||
| Fugusakana/fugu | 1M | - / - | 2026-06-15 | |||
| GLM-5.2zhipuai/glm-5.2 | 1M | $1.4 / $4.4 | 2026-06-13 | |||
| Kimi K2.7 Codemoonshotai/kimi-k2.7-code | 262.144K | $0.95 / $4 | 2026-06-12 | |||
| Kimi K2.7 Code Highspeedmoonshotai/kimi-k2.7-code-highspeed | 262.144K | $1.9 / $8 | 2026-06-12 | |||
| Claude Fable 5anthropic/claude-fable-5 | 1M | $10 / $50 | 2026-06-09 | |||
| North Mini Codecohere/north-mini-code-1-0 | 256K | - / - | 2026-06-09 | |||
| Gemini 3.5 Live Translate Previewgoogle/gemini-3.5-live-translate-preview | 131.072K | $3.5 / $21 | 2026-06-09 | |||
| MiMo-V2.5-Pro-UltraSpeedxiaomi/mimo-v2.5-pro-ultraspeed | 1.04858M | $1.305 / $2.61 | 2026-06-08 | |||
| Nemotron 3 Ultra 550B A55Bnvidia/nemotron-3-ultra-550b-a55b | 1M | $0.5 / $2.5 | 2026-06-04 | |||
| Nemotron 3.5 Content Safetynvidia/nemotron-3.5-content-safety | 128K | - / - | 2026-06-04 | |||
| MAI-Code-1-Flashmicrosoft/mai-code-1-flash | 256K | $0.75 / $4.5 | 2026-06-02 | |||
| Qwen3.7 Plusalibaba/qwen3.7-plus | 1M | $0.5 / $3 | 2026-06-02 | |||
| MiniMax-M3minimax/MiniMax-M3 | 512K | $0.3 / $1.2 | 2026-06-01 | |||
| Step 3.7 Flashstepfun/step-3.7-flash | 256K | $0.185 / $1.11 | 2026-05-29 | |||
| Nano Banana 2google/gemini-3.1-flash-image | 131.072K | $0.5 / $60 | 2026-05-28 | |||
| Claude Opus 4.8anthropic/claude-opus-4-8 | 1M | $5 / $25 | 2026-05-28 | |||
| Qwen3.7 Maxalibaba/qwen3.7-max | 1M | $2.5 / $7.5 | 2026-05-21 | |||
| Command A Pluscohere/command-a-plus-05-2026 | 128K | $2.5 / $10 | 2026-05-20 | |||
| Gemini 3.5 Flashgoogle/gemini-3.5-flash | 1.04858M | $1.5 / $9 | 2026-05-19 | |||
| Gemini Flash Latestgoogle/gemini-flash-latest | 1.04858M | $1.5 / $9 | 2026-05-19 | |||
| Gemini 3.1 Flash Litegoogle/gemini-3.1-flash-lite | 1.04858M | $0.25 / $1.5 | 2026-05-07 | |||
| Gemini Flash-Lite Latestgoogle/gemini-flash-lite-latest | 1.04858M | $0.25 / $1.5 | 2026-05-07 | |||
| GPT-5.5 Instantopenai/gpt-5.5-instant | 400K | $5 / $30 | 2026-05-05 | |||
| Mistral Medium (latest)mistral/mistral-medium-latest | 262.144K | $1.5 / $7.5 | 2026-04-29 | |||
| Mistral Medium 3.5mistral/mistral-medium-2604 | 262.144K | $1.5 / $7.5 | 2026-04-29 | |||
| Laguna M.1poolside/laguna-m.1 | 262.144K | $0.2 / $0.4 | 2026-04-28 | |||
| Laguna XS.2poolside/laguna-xs.2 | 262.144K | $0.2 / $0.4 | 2026-04-28 | |||
| Nemotron 3 Nano Omni 30B A3B Reasoningnvidia/nemotron-3-nano-omni-30b-a3b-reasoning | 256K | $0.105 / $0.42 | 2026-04-28 |