anthropic/claude-opus-5
Strongest Claude Opus model for coding, agents, and professional work
- Context
- 1M
- Input
- $5/M
- Output
- $25/M
Filter 384 source-linked models by creator, price, context, modality, and published benchmark coverage.
anthropic/claude-opus-5
Strongest Claude Opus model for coding, agents, and professional work
google/gemini-3.6-flash
Fast Gemini model balancing multimodal reasoning, tool use, and cost
google/gemini-3.5-flash-lite
Fast Gemini model balancing multimodal reasoning, tool use, and cost
alibaba/qwen3.8-max-preview
Preview Qwen flagship for million-token multimodal reasoning and long-horizon agentic workflows
moonshotai/kimi-k3
Multimodal Kimi model with 1M context and toggleable max-effort thinking for long-horizon agent work
thinkingmachines/inkling
Multimodal MoE reasoning model (975B total, 41B active) for text, image, and audio
openai/gpt-5.6-luna
Cost-efficient GPT-5.6 model for fast, high-volume workloads
openai/gpt-5.6-terra
Balanced GPT-5.6 model for capable, cost-efficient everyday work
openai/gpt-5.6-sol
Frontier GPT-5.6 model for complex professional work, coding, and agentic workflows
xai/grok-4.5
xAI's latest Grok for chat, coding, agentic tools, and lower hallucination risk
openai/gpt-realtime-2.1
Realtime speech-to-speech model with configurable reasoning, tool use, and robust voice-agent behavior
anthropic/claude-sonnet-5
Everyday Claude agent model for coding, planning, browsing, and general work
google/gemini-3.1-flash-lite-image
Fastest, most cost-efficient Gemini image model for high-volume 1K generation and editing
google/gemini-omni-flash-preview
Video generation and editing model for fast, conversational text- and image-to-video workflows
deepreinforce/ornith-1.0-31b
Open coding-reasoning model for repository tasks and self-improving agents
deepreinforce/ornith-1.0-35b
Large coding-reasoning model for agentic software tasks and RL search
deepreinforce/ornith-1.0-9b
Open coding-reasoning model for repository tasks and self-improving agents
deepreinforce/ornith-1.0-397b
Large coding-reasoning model for agentic software tasks and RL search
sakana/fugu-ultra
Quality-first multi-agent model for hard research, analysis, and competitions
sakana/fugu
Multi-agent model for routing expert agents across complex analytical tasks
moonshotai/kimi-k2.7-code
Coding-focused Kimi model, stronger on long-horizon repo work with less overthinking
moonshotai/kimi-k2.7-code-highspeed
Lower-latency Kimi Code variant for interactive edits and coding-agent loops
anthropic/claude-fable-5
Claude model for creative writing, analysis, and controlled agent workflows
nvidia/nemotron-3.5-content-safety
Safety model for policy screening, moderation, and risk-aware routing workflows
alibaba/qwen3.7-plus
Multimodal Qwen workhorse for long-context agents, visual inputs, and coding
minimax/MiniMax-M3
MiniMax multimodal model for long-context coding, perception, and agent planning
xai/grok-imagine-video-1.5
Video model for image-to-video generation, editing, and extension workflows
stepfun/step-3.7-flash
Newer StepFun flash model for faster agents, coding, and multimodal prompts
google/gemini-3.1-flash-image
Image model for prompt-driven generation, editing, and visual design workflows
google/gemini-3-pro-image
Nano Banana Pro for higher-fidelity image generation and design-heavy edits
anthropic/claude-opus-4-8
Top Claude Opus tier for the hardest reasoning, coding, and long-horizon agents
cohere/command-a-plus-05-2026
Cohere's stronger command model for multilingual agents and enterprise workflows
google/gemini-3.5-flash
Fast Gemini model balancing multimodal reasoning, tool use, and cost
google/gemini-flash-latest
Fast Gemini model balancing multimodal reasoning, tool use, and cost
google/gemini-3.1-flash-lite
Low-latency Gemini model for high-volume multimodal and agent workloads
google/gemini-flash-lite-latest
Low-latency Gemini model for high-volume multimodal and agent workloads
openai/gpt-5.5-instant
Compact GPT model for low-latency assistance and high-volume workloads
mistral/mistral-medium-latest
Balanced Mistral model for enterprise assistants, multilingual work, and tools
mistral/mistral-medium-2604
Balanced Mistral model for enterprise assistants, multilingual work, and tools
nvidia/nemotron-3-nano-omni-30b-a3b-reasoning
Open Nemotron omni model combining reasoning with text, vision, and audio
alibaba/qwen3.6-flash
Qwen vision-language model for visual reasoning, documents, and agent tasks
openai/gpt-5.5-pro
Highest-accuracy GPT-5.5 tier for slower, precision-heavy reasoning and coding
openai/gpt-5.5
Default frontier GPT for coding, computer use, research, and knowledge work
xiaomi/mimo-v2.5
Open MiMo model for multimodal coding agents and long-context automation
alibaba/qwen3.6-27b
Qwen vision-language model for visual reasoning, documents, and agent tasks
google/gemini-embedding-2
Multimodal embedding model mapping text, images, video, audio, and PDFs into a unified embedding space
moonshotai/kimi-k2.6
Multimodal Kimi workhorse for agent loops, coding tasks, and visual context
openai/gpt-image-2
Image model for prompt-driven generation, editing, and visual design workflows
google/deep-research-preview-04-2026
Agentic model for autonomous multi-step research, synthesis, and cited reports
google/deep-research-max-preview-04-2026
Maximum-comprehensiveness agentic researcher for multi-step investigation, synthesis, and cited reports
| Model | Creator | Input types | Context | Input / Output | Released | Compare |
|---|---|---|---|---|---|---|
| Claude Opus 5anthropic/claude-opus-5 | 1M | $5 / $25 | 2026-07-24 | |||
| Gemini 3.6 Flashgoogle/gemini-3.6-flash | 1.04858M | $1.5 / $7.5 | 2026-07-21 | |||
| Gemini 3.5 Flash Litegoogle/gemini-3.5-flash-lite | 1.04858M | $0.3 / $2.5 | 2026-07-21 | |||
| Qwen3.8 Max Previewalibaba/qwen3.8-max-preview | 1M | - / - | 2026-07-19 | |||
| Kimi K3moonshotai/kimi-k3 | 1.04858M | $3 / $15 | 2026-07-16 | |||
| Inklingthinkingmachines/inkling | 1.04858M | $1.87 / $4.68 | 2026-07-15 | |||
| GPT-5.6 Lunaopenai/gpt-5.6-luna | 1.05M | $1 / $6 | 2026-07-09 | |||
| GPT-5.6 Terraopenai/gpt-5.6-terra | 1.05M | $2.5 / $15 | 2026-07-09 | |||
| GPT-5.6 Solopenai/gpt-5.6-sol | 1.05M | $5 / $30 | 2026-07-09 | |||
| Grok 4.5xai/grok-4.5 | 500K | $2 / $6 | 2026-07-08 | |||
| GPT-Realtime-2.1openai/gpt-realtime-2.1 | 128K | $4 / $24 | 2026-07-06 | |||
| Claude Sonnet 5anthropic/claude-sonnet-5 | 1M | $2 / $10 | 2026-06-30 | |||
| Nano Banana 2 Litegoogle/gemini-3.1-flash-lite-image | 65.536K | $0.25 / $30 | 2026-06-30 | |||
| Gemini Omni Flash Previewgoogle/gemini-omni-flash-preview | 1.04858M | $1.5 / $17.5 | 2026-06-30 | |||
| Ornith 1.0 31Bdeepreinforce/ornith-1.0-31b | 262.144K | - / - | 2026-06-25 | |||
| Ornith 1.0 35Bdeepreinforce/ornith-1.0-35b | 262.144K | - / - | 2026-06-25 | |||
| Ornith 1.0 9Bdeepreinforce/ornith-1.0-9b | 262.144K | - / - | 2026-06-25 | |||
| Ornith 1.0 397Bdeepreinforce/ornith-1.0-397b | 262.144K | - / - | 2026-06-25 | |||
| Fugu Ultrasakana/fugu-ultra | 1M | $5 / $30 | 2026-06-15 | |||
| Fugusakana/fugu | 1M | - / - | 2026-06-15 | |||
| Kimi K2.7 Codemoonshotai/kimi-k2.7-code | 262.144K | $0.95 / $4 | 2026-06-12 | |||
| Kimi K2.7 Code Highspeedmoonshotai/kimi-k2.7-code-highspeed | 262.144K | $1.9 / $8 | 2026-06-12 | |||
| Claude Fable 5anthropic/claude-fable-5 | 1M | $10 / $50 | 2026-06-09 | |||
| Nemotron 3.5 Content Safetynvidia/nemotron-3.5-content-safety | 128K | - / - | 2026-06-04 | |||
| Qwen3.7 Plusalibaba/qwen3.7-plus | 1M | $0.5 / $3 | 2026-06-02 | |||
| MiniMax-M3minimax/MiniMax-M3 | 512K | $0.3 / $1.2 | 2026-06-01 | |||
| Grok Imagine Video 1.5xai/grok-imagine-video-1.5 | 1.024K | - / - | 2026-05-30 | |||
| Step 3.7 Flashstepfun/step-3.7-flash | 256K | $0.185 / $1.11 | 2026-05-29 | |||
| Nano Banana 2google/gemini-3.1-flash-image | 131.072K | $0.5 / $60 | 2026-05-28 | |||
| Nano Banana Progoogle/gemini-3-pro-image | 65.536K | $2 / $120 | 2026-05-28 | |||
| Claude Opus 4.8anthropic/claude-opus-4-8 | 1M | $5 / $25 | 2026-05-28 | |||
| Command A Pluscohere/command-a-plus-05-2026 | 128K | $2.5 / $10 | 2026-05-20 | |||
| Gemini 3.5 Flashgoogle/gemini-3.5-flash | 1.04858M | $1.5 / $9 | 2026-05-19 | |||
| Gemini Flash Latestgoogle/gemini-flash-latest | 1.04858M | $1.5 / $9 | 2026-05-19 | |||
| Gemini 3.1 Flash Litegoogle/gemini-3.1-flash-lite | 1.04858M | $0.25 / $1.5 | 2026-05-07 | |||
| Gemini Flash-Lite Latestgoogle/gemini-flash-lite-latest | 1.04858M | $0.25 / $1.5 | 2026-05-07 | |||
| GPT-5.5 Instantopenai/gpt-5.5-instant | 400K | $5 / $30 | 2026-05-05 | |||
| Mistral Medium (latest)mistral/mistral-medium-latest | 262.144K | $1.5 / $7.5 | 2026-04-29 | |||
| Mistral Medium 3.5mistral/mistral-medium-2604 | 262.144K | $1.5 / $7.5 | 2026-04-29 | |||
| Nemotron 3 Nano Omni 30B A3B Reasoningnvidia/nemotron-3-nano-omni-30b-a3b-reasoning | 256K | $0.105 / $0.42 | 2026-04-28 | |||
| Qwen3.6 Flashalibaba/qwen3.6-flash | 1M | $0.188 / $1.125 | 2026-04-27 | |||
| GPT-5.5 Proopenai/gpt-5.5-pro | 1.05M | $30 / $180 | 2026-04-23 | |||
| GPT-5.5openai/gpt-5.5 | 1.05M | $5 / $30 | 2026-04-23 | |||
| MiMo-V2.5xiaomi/mimo-v2.5 | 1.04858M | $0.14 / $0.28 | 2026-04-22 | |||
| Qwen3.6 27Balibaba/qwen3.6-27b | 262.144K | $0.6 / $3.6 | 2026-04-22 | |||
| Gemini Embedding 2google/gemini-embedding-2 | 8.192K | $0.2 / - | 2026-04-22 | |||
| Kimi K2.6moonshotai/kimi-k2.6 | 262.144K | $0.95 / $4 | 2026-04-21 | |||
| GPT-Image-2openai/gpt-image-2 | Not documented | $5 / $30 | 2026-04-21 | |||
| Gemini Deep Research Previewgoogle/deep-research-preview-04-2026 | 1.04858M | - / - | 2026-04-21 | |||
| Deep Research Max Previewgoogle/deep-research-max-preview-04-2026 | 1.04858M | - / - | 2026-04-21 |