meta/llama-4-scout-17b-instruct
Open Llama with long-context vision for efficient multimodal agents
- Context
- 3.5M
- Input
- $0.17/M
- Output
- $0.66/M
Filter 494 source-linked models by creator, price, context, modality, and published benchmark coverage.
meta/llama-4-scout-17b-instruct
Open Llama with long-context vision for efficient multimodal agents
meta/llama-4-maverick-17b-instruct
Open multimodal Llama for strong reasoning with efficient everyday serving
alibaba/qwen3-coder-30b-a3b-instruct
Smaller Qwen coder for efficient local agents and repo-level fixes
alibaba/qwen3-32b
Dense open Qwen model for self-hosted chat, reasoning, and coding
alibaba/qwen3-235b-a22b
Large open Qwen MoE for multilingual reasoning, coding, and tool use
alibaba/qwen3-coder-480b-a35b-instruct
Open Qwen coding heavyweight for repository reasoning and agentic engineering
openai/o1-pro
O-series reasoning model for hard analysis, math, coding, and planning
mistral/magistral-medium-latest
Mistral reasoning model for transparent analysis, math, and complex decisions
cohere/command-a-03-2025
Cohere command model for multilingual enterprise agents, tools, and chat
alibaba/qwq-plus
Qwen reasoning model for deliberate problem solving, math, and coding
cohere/command-r7b-arabic-02-2025
Open Command R model optimized for Arabic enterprise chat, RAG, and cultural knowledge
anthropic/claude-3-7-sonnet-20250219
Balanced Claude model for coding, analysis, agent workflows, and cost control
deepseek/deepseek-r1
Classic open reasoning model for transparent math, coding, and deliberate problem solving
openai/o3-mini
Smaller o-series reasoner for economical coding, math, and planning tasks
google/gemini-2.0-flash
Earlier Gemini Flash workhorse for responsive multimodal apps and tool use
google/gemini-2.0-flash-lite
Low-latency Gemini model for high-volume multimodal and agent workloads
meta/llama-3.3-70b-instruct
Popular open Llama workhorse for multilingual chat, coding, and self-hosting
openai/o1
O-series reasoning model for hard analysis, math, coding, and planning
cohere/command-r7b-12-2024
Cohere retrieval model for long-context chat and enterprise RAG workflows
openai/gpt-4o-2024-11-20
GPT model for general reasoning, writing, coding, and tool-assisted tasks
mistral/mistral-large-2411
Flagship Mistral model for advanced reasoning, coding, and multilingual work
mistral/mistral-large-latest
Flagship Mistral model for advanced reasoning, coding, and multilingual work
mistral/pixtral-large-latest
Mistral's larger vision model for document-heavy image understanding and chat
mistral/mistral-large-2512
Mistral's largest general model for enterprise agents, coding, and multilingual reasoning
alibaba/qwen-turbo
Efficient Qwen model for fast chat, extraction, and high-volume workloads
cohere/c4ai-aya-expanse-32b
Open multilingual model optimized for generation across 23 languages
anthropic/claude-3-5-haiku-20241022
Fast Claude model for responsive assistance, classification, and lightweight agents
anthropic/claude-3-5-sonnet-20241022
Balanced Claude model for coding, analysis, agent workflows, and cost control
mistral/pixtral-12b
Mistral vision-language model for image understanding and multimodal chat
alibaba/qwen2-5-vl-72b-instruct
Qwen vision-language model for visual reasoning, documents, and agent tasks
cohere/command-r-08-2024
Cohere retrieval model for long-context chat and enterprise RAG workflows
cohere/command-r-plus-08-2024
Cohere's RAG workhorse for long-context enterprise search and tool use
nvidia/nemotron-mini-4b-instruct
Compact Nemotron model for efficient reasoning and deployable AI agents
openai/gpt-4o-2024-08-06
GPT model for general reasoning, writing, coding, and tool-assisted tasks
openai/gpt-4o-mini
Small omni GPT for cheap multimodal assistance and production-scale traffic
mistral/mistral-nemo
Efficient Mistral-NVIDIA open model for multilingual chat and local deployment
openai/o3-deep-research
Research model for long-horizon investigation, synthesis, and analytical reports
openai/o4-mini-deep-research
Research model for long-horizon investigation, synthesis, and analytical reports
mistral/codestral-latest
Mistral code model for completions, refactors, and developer IDE workflows
openai/gpt-4o-2024-05-13
GPT model for general reasoning, writing, coding, and tool-assisted tasks
openai/gpt-4o
Omni-era GPT for multimodal chat, practical coding, and general assistants
alibaba/qwen-vl-max
Qwen vision-language model for visual reasoning, documents, and agent tasks
anthropic/claude-3-haiku-20240307
Legacy model retained for compatibility with older integrations
alibaba/qwen-plus
Qwen instruction model for multilingual chat, reasoning, and tool use
alibaba/qwen-vl-plus
Qwen vision-language model for visual reasoning, documents, and agent tasks
perplexity/sonar-reasoning-pro
Web-grounded Sonar for multi-step research questions that need cited reasoning
perplexity/sonar-pro
Deeper Sonar search model with broader retrieval and stronger synthesis
perplexity/sonar
Fast web-grounded Sonar for current answers, citations, and lightweight retrieval
openai/gpt-4-turbo
Compact GPT model for low-latency assistance and high-volume workloads
openai/gpt-5.6-luna-pro
GPT-5.6 Luna Pro is the same underlying model as [GPT-5.6 Luna](https://openrouter.ai/openai/gpt-5.6-luna), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks. Learn more in OpenAI's docs: https://developers.openai.com/api/docs/guides/reasoning#reasoning-mode
| Model | Creator | Input types | Context | Input / Output | Released | Compare |
|---|---|---|---|---|---|---|
| Llama 4 Scout 17B Instructmeta/llama-4-scout-17b-instruct | 3.5M | $0.17 / $0.66 | 2025-04-05 | |||
| Llama 4 Maverick 17B Instructmeta/llama-4-maverick-17b-instruct | 1M | $0.124 / $0.603 | 2025-04-05 | |||
| Qwen3-Coder 30B-A3B Instructalibaba/qwen3-coder-30b-a3b-instruct | 262.144K | $0.45 / $2.25 | 2025-04 | |||
| Qwen3 32Balibaba/qwen3-32b | 131.072K | $0.7 / $2.8 | 2025-04 | |||
| Qwen3 235B-A22Balibaba/qwen3-235b-a22b | 131.072K | $0.7 / $2.8 | 2025-04 | |||
| Qwen3-Coder 480B-A35B Instructalibaba/qwen3-coder-480b-a35b-instruct | 262.144K | $1.5 / $7.5 | 2025-04 | |||
| o1-proopenai/o1-pro | 200K | $150 / $600 | 2025-03-19 | |||
| Magistral Medium (latest)mistral/magistral-medium-latest | 128K | $2 / $5 | 2025-03-17 | |||
| Command Acohere/command-a-03-2025 | 256K | $2.5 / $10 | 2025-03-13 | |||
| QwQ Plusalibaba/qwq-plus | 131.072K | $0.8 / $2.4 | 2025-03-05 | |||
| Command R7B Arabiccohere/command-r7b-arabic-02-2025 | 128K | $0.037 / $0.15 | 2025-02-27 | |||
| Claude Sonnet 3.7anthropic/claude-3-7-sonnet-20250219 | 200K | $3 / $15 | 2025-02-19 | |||
| DeepSeek-R1deepseek/deepseek-r1 | 128K | $0.7 / $2.5 | 2025-01-20 | |||
| o3-miniopenai/o3-mini | 200K | $1.1 / $4.4 | 2024-12-20 | |||
| Gemini 2.0 Flashgoogle/gemini-2.0-flash | 1.04858M | $0.1 / $0.4 | 2024-12-11 | |||
| Gemini 2.0 Flash-Litegoogle/gemini-2.0-flash-lite | 1.04858M | $0.075 / $0.3 | 2024-12-11 | |||
| Llama-3.3-70B-Instructmeta/llama-3.3-70b-instruct | 128K | $0.13 / $0.4 | 2024-12-06 | |||
| o1openai/o1 | 200K | $15 / $60 | 2024-12-05 | |||
| Command R7Bcohere/command-r7b-12-2024 | 128K | $0.037 / $0.15 | 2024-12-02 | |||
| GPT-4o (2024-11-20)openai/gpt-4o-2024-11-20 | 128K | $2.5 / $10 | 2024-11-20 | |||
| Mistral Large 2.1mistral/mistral-large-2411 | 131.072K | $2 / $6 | 2024-11-18 | |||
| Mistral Large (latest)mistral/mistral-large-latest | 262.144K | $0.5 / $1.5 | 2024-11-01 | |||
| Pixtral Large (latest)mistral/pixtral-large-latest | 128K | $2 / $6 | 2024-11-01 | |||
| Mistral Large 3mistral/mistral-large-2512 | 262.144K | $0.5 / $1.5 | 2024-11-01 | |||
| Qwen Turboalibaba/qwen-turbo | 1M | $0.05 / $0.2 | 2024-11-01 | |||
| Aya Expanse 32Bcohere/c4ai-aya-expanse-32b | 128K | - / - | 2024-10-24 | |||
| Claude Haiku 3.5anthropic/claude-3-5-haiku-20241022 | 200K | $0.8 / $4 | 2024-10-22 | |||
| Claude Sonnet 3.5 v2anthropic/claude-3-5-sonnet-20241022 | 200K | $3 / $15 | 2024-10-22 | |||
| Pixtral 12Bmistral/pixtral-12b | 128K | $0.15 / $0.15 | 2024-09-01 | |||
| Qwen2.5-VL 72B Instructalibaba/qwen2-5-vl-72b-instruct | 131.072K | $2.8 / $8.4 | 2024-09 | |||
| Command Rcohere/command-r-08-2024 | 128K | $0.15 / $0.6 | 2024-08-30 | |||
| Command R+cohere/command-r-plus-08-2024 | 128K | $2.5 / $10 | 2024-08-30 | |||
| Nemotron Mini 4B Instructnvidia/nemotron-mini-4b-instruct | 128K | - / - | 2024-08-21 | |||
| GPT-4o (2024-08-06)openai/gpt-4o-2024-08-06 | 128K | $2.5 / $10 | 2024-08-06 | |||
| GPT-4o miniopenai/gpt-4o-mini | 128K | $0.15 / $0.6 | 2024-07-18 | |||
| Mistral Nemomistral/mistral-nemo | 128K | $0.15 / $0.15 | 2024-07-01 | |||
| o3-deep-researchopenai/o3-deep-research | 200K | $9 / $36 | 2024-06-26 | |||
| o4-mini-deep-researchopenai/o4-mini-deep-research | 200K | $1.8 / $7.2 | 2024-06-26 | |||
| Codestral (latest)mistral/codestral-latest | 256K | $0.3 / $0.9 | 2024-05-29 | |||
| GPT-4o (2024-05-13)openai/gpt-4o-2024-05-13 | 128K | $5 / $15 | 2024-05-13 | |||
| GPT-4oopenai/gpt-4o | 128K | $2.5 / $10 | 2024-05-13 | |||
| Qwen-VL Maxalibaba/qwen-vl-max | 131.072K | $0.8 / $3.2 | 2024-04-08 | |||
| Claude Haiku 3anthropic/claude-3-haiku-20240307 | 200K | $0.25 / $1.25 | 2024-03-13 | |||
| Qwen Plusalibaba/qwen-plus | 1M | $0.4 / $1.2 | 2024-01-25 | |||
| Qwen-VL Plusalibaba/qwen-vl-plus | 131.072K | $0.21 / $0.63 | 2024-01-25 | |||
| Sonar Reasoning Properplexity/sonar-reasoning-pro | 128K | $2 / $8 | 2024-01-01 | |||
| Sonar Properplexity/sonar-pro | 200K | $3 / $15 | 2024-01-01 | |||
| Sonarperplexity/sonar | 128K | $1 / $1 | 2024-01-01 | |||
| GPT-4 Turboopenai/gpt-4-turbo | 128K | $10 / $30 | 2023-11-06 | |||
| OpenAI: GPT-5.6 Luna Proopenai/gpt-5.6-luna-pro | 1.05M | $0.1 / $0.6 | Undated |