494 results Clear filters
2025-04-05
meta/llama-4-scout-17b-instruct

Open Llama with long-context vision for efficient multimodal agents

T Open weights
Context
3.5M
Input
$0.17/M
Output
$0.66/M
2025-04-05
meta/llama-4-maverick-17b-instruct

Open multimodal Llama for strong reasoning with efficient everyday serving

T Open weights
Context
1M
Input
$0.124/M
Output
$0.603/M
2025-04
alibaba/qwen3-coder-30b-a3b-instruct

Smaller Qwen coder for efficient local agents and repo-level fixes

T Tools Open weights
Context
262.144K
Input
$0.45/M
Output
$2.25/M
Qwen3 32Bby Alibaba Qwen
2025-04
alibaba/qwen3-32b

Dense open Qwen model for self-hosted chat, reasoning, and coding

T Reasoning Tools Open weights
Context
131.072K
Input
$0.7/M
Output
$2.8/M
Qwen3 235B-A22Bby Alibaba Qwen
2025-04
alibaba/qwen3-235b-a22b

Large open Qwen MoE for multilingual reasoning, coding, and tool use

T Reasoning Tools Open weights
Context
131.072K
Input
$0.7/M
Output
$2.8/M
2025-04
alibaba/qwen3-coder-480b-a35b-instruct

Open Qwen coding heavyweight for repository reasoning and agentic engineering

T Tools Open weights
Context
262.144K
Input
$1.5/M
Output
$7.5/M
o1-proby OpenAI
2025-03-19
openai/o1-pro

O-series reasoning model for hard analysis, math, coding, and planning

T Reasoning Tools Weight access not listed
Context
200K
Input
$150/M
Output
$600/M
2025-03-17
mistral/magistral-medium-latest

Mistral reasoning model for transparent analysis, math, and complex decisions

T Reasoning Tools Weight access not listed
Context
128K
Input
$2/M
Output
$5/M
Command Aby Cohere
2025-03-13
cohere/command-a-03-2025

Cohere command model for multilingual enterprise agents, tools, and chat

T Tools Open weights
Context
256K
Input
$2.5/M
Output
$10/M
QwQ Plusby Alibaba Qwen
2025-03-05
alibaba/qwq-plus

Qwen reasoning model for deliberate problem solving, math, and coding

T Reasoning Tools Weight access not listed
Context
131.072K
Input
$0.8/M
Output
$2.4/M
2025-02-27
cohere/command-r7b-arabic-02-2025

Open Command R model optimized for Arabic enterprise chat, RAG, and cultural knowledge

T Tools Open weights
Context
128K
Input
$0.037/M
Output
$0.15/M
Claude Sonnet 3.7by Anthropic
2025-02-19
anthropic/claude-3-7-sonnet-20250219

Balanced Claude model for coding, analysis, agent workflows, and cost control

T Reasoning Tools Weight access not listed
Context
200K
Input
$3/M
Output
$15/M
DeepSeek-R1by DeepSeek
2025-01-20
deepseek/deepseek-r1

Classic open reasoning model for transparent math, coding, and deliberate problem solving

T Reasoning Tools Open weights
Context
128K
Input
$0.7/M
Output
$2.5/M
o3-miniby OpenAI
2024-12-20
openai/o3-mini

Smaller o-series reasoner for economical coding, math, and planning tasks

T Reasoning Tools Structured output Weight access not listed
Context
200K
Input
$1.1/M
Output
$4.4/M
2024-12-11
google/gemini-2.0-flash

Earlier Gemini Flash workhorse for responsive multimodal apps and tool use

T Tools Weight access not listed
Context
1.04858M
Input
$0.1/M
Output
$0.4/M
2024-12-11
google/gemini-2.0-flash-lite

Low-latency Gemini model for high-volume multimodal and agent workloads

T Tools Weight access not listed
Context
1.04858M
Input
$0.075/M
Output
$0.3/M
2024-12-06
meta/llama-3.3-70b-instruct

Popular open Llama workhorse for multilingual chat, coding, and self-hosting

T Tools Open weights
Context
128K
Input
$0.13/M
Output
$0.4/M
o1by OpenAI
2024-12-05
openai/o1

O-series reasoning model for hard analysis, math, coding, and planning

T Reasoning Tools Weight access not listed
Context
200K
Input
$15/M
Output
$60/M
Command R7Bby Cohere
2024-12-02
cohere/command-r7b-12-2024

Cohere retrieval model for long-context chat and enterprise RAG workflows

T Tools Open weights
Context
128K
Input
$0.037/M
Output
$0.15/M
2024-11-20
openai/gpt-4o-2024-11-20

GPT model for general reasoning, writing, coding, and tool-assisted tasks

T Weight access not listed
Context
128K
Input
$2.5/M
Output
$10/M
Mistral Large 2.1by Mistral AI
2024-11-18
mistral/mistral-large-2411

Flagship Mistral model for advanced reasoning, coding, and multilingual work

T Tools Open weights
Context
131.072K
Input
$2/M
Output
$6/M
2024-11-01
mistral/mistral-large-latest

Flagship Mistral model for advanced reasoning, coding, and multilingual work

T Tools Open weights
Context
262.144K
Input
$0.5/M
Output
$1.5/M
2024-11-01
mistral/pixtral-large-latest

Mistral's larger vision model for document-heavy image understanding and chat

T Tools Open weights
Context
128K
Input
$2/M
Output
$6/M
Mistral Large 3by Mistral AI
2024-11-01
mistral/mistral-large-2512

Mistral's largest general model for enterprise agents, coding, and multilingual reasoning

T Tools Open weights
Context
262.144K
Input
$0.5/M
Output
$1.5/M
Qwen Turboby Alibaba Qwen
2024-11-01
alibaba/qwen-turbo

Efficient Qwen model for fast chat, extraction, and high-volume workloads

T Reasoning Tools Weight access not listed
Context
1M
Input
$0.05/M
Output
$0.2/M
2024-10-24
cohere/c4ai-aya-expanse-32b

Open multilingual model optimized for generation across 23 languages

T Open weights
Context
128K
Input
-
Output
-
Claude Haiku 3.5by Anthropic
2024-10-22
anthropic/claude-3-5-haiku-20241022

Fast Claude model for responsive assistance, classification, and lightweight agents

T Tools Weight access not listed
Context
200K
Input
$0.8/M
Output
$4/M
2024-10-22
anthropic/claude-3-5-sonnet-20241022

Balanced Claude model for coding, analysis, agent workflows, and cost control

T Tools Weight access not listed
Context
200K
Input
$3/M
Output
$15/M
Pixtral 12Bby Mistral AI
2024-09-01
mistral/pixtral-12b

Mistral vision-language model for image understanding and multimodal chat

T Tools Open weights
Context
128K
Input
$0.15/M
Output
$0.15/M
2024-09
alibaba/qwen2-5-vl-72b-instruct

Qwen vision-language model for visual reasoning, documents, and agent tasks

T Tools Open weights
Context
131.072K
Input
$2.8/M
Output
$8.4/M
Command Rby Cohere
2024-08-30
cohere/command-r-08-2024

Cohere retrieval model for long-context chat and enterprise RAG workflows

T Tools Open weights
Context
128K
Input
$0.15/M
Output
$0.6/M
Command R+by Cohere
2024-08-30
cohere/command-r-plus-08-2024

Cohere's RAG workhorse for long-context enterprise search and tool use

T Tools Open weights
Context
128K
Input
$2.5/M
Output
$10/M
2024-08-21
nvidia/nemotron-mini-4b-instruct

Compact Nemotron model for efficient reasoning and deployable AI agents

T Tools Open weights
Context
128K
Input
-
Output
-
2024-08-06
openai/gpt-4o-2024-08-06

GPT model for general reasoning, writing, coding, and tool-assisted tasks

T Weight access not listed
Context
128K
Input
$2.5/M
Output
$10/M
GPT-4o miniby OpenAI
2024-07-18
openai/gpt-4o-mini

Small omni GPT for cheap multimodal assistance and production-scale traffic

T Tools Structured output Weight access not listed
Context
128K
Input
$0.15/M
Output
$0.6/M
Mistral Nemoby Mistral AI
2024-07-01
mistral/mistral-nemo

Efficient Mistral-NVIDIA open model for multilingual chat and local deployment

T Open weights
Context
128K
Input
$0.15/M
Output
$0.15/M
2024-06-26
openai/o3-deep-research

Research model for long-horizon investigation, synthesis, and analytical reports

T Reasoning Tools Weight access not listed
Context
200K
Input
$9/M
Output
$36/M
2024-06-26
openai/o4-mini-deep-research

Research model for long-horizon investigation, synthesis, and analytical reports

T Reasoning Tools Weight access not listed
Context
200K
Input
$1.8/M
Output
$7.2/M
Codestral (latest)by Mistral AI
2024-05-29
mistral/codestral-latest

Mistral code model for completions, refactors, and developer IDE workflows

T Tools Open weights
Context
256K
Input
$0.3/M
Output
$0.9/M
2024-05-13
openai/gpt-4o-2024-05-13

GPT model for general reasoning, writing, coding, and tool-assisted tasks

T Tools Structured output Weight access not listed
Context
128K
Input
$5/M
Output
$15/M
GPT-4oby OpenAI
2024-05-13
openai/gpt-4o

Omni-era GPT for multimodal chat, practical coding, and general assistants

T Tools Weight access not listed
Context
128K
Input
$2.5/M
Output
$10/M
Qwen-VL Maxby Alibaba Qwen
2024-04-08
alibaba/qwen-vl-max

Qwen vision-language model for visual reasoning, documents, and agent tasks

T Tools Weight access not listed
Context
131.072K
Input
$0.8/M
Output
$3.2/M
Claude Haiku 3by Anthropic
2024-03-13
anthropic/claude-3-haiku-20240307

Legacy model retained for compatibility with older integrations

T Tools Weight access not listed
Context
200K
Input
$0.25/M
Output
$1.25/M
Qwen Plusby Alibaba Qwen
2024-01-25
alibaba/qwen-plus

Qwen instruction model for multilingual chat, reasoning, and tool use

T Reasoning Tools Weight access not listed
Context
1M
Input
$0.4/M
Output
$1.2/M
Qwen-VL Plusby Alibaba Qwen
2024-01-25
alibaba/qwen-vl-plus

Qwen vision-language model for visual reasoning, documents, and agent tasks

T Tools Weight access not listed
Context
131.072K
Input
$0.21/M
Output
$0.63/M
Sonar Reasoning Proby Perplexity
2024-01-01
perplexity/sonar-reasoning-pro

Web-grounded Sonar for multi-step research questions that need cited reasoning

T Reasoning Weight access not listed
Context
128K
Input
$2/M
Output
$8/M
Sonar Proby Perplexity
2024-01-01
perplexity/sonar-pro

Deeper Sonar search model with broader retrieval and stronger synthesis

T Weight access not listed
Context
200K
Input
$3/M
Output
$15/M
Sonarby Perplexity
2024-01-01
perplexity/sonar

Fast web-grounded Sonar for current answers, citations, and lightweight retrieval

T Weight access not listed
Context
128K
Input
$1/M
Output
$1/M
GPT-4 Turboby OpenAI
2023-11-06
openai/gpt-4-turbo

Compact GPT model for low-latency assistance and high-volume workloads

T Tools Weight access not listed
Context
128K
Input
$10/M
Output
$30/M
Undated
openai/gpt-5.6-luna-pro

GPT-5.6 Luna Pro is the same underlying model as [GPT-5.6 Luna](https://openrouter.ai/openai/gpt-5.6-luna), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks. Learn more in OpenAI's docs: https://developers.openai.com/api/docs/guides/reasoning#reasoning-mode

T Weight access not listed
Context
1.05M
Input
$0.1/M
Output
$0.6/M