google/gemini-2.0-flash
Earlier Gemini Flash workhorse for responsive multimodal apps and tool use
- Context
- 1.04858M
- Input
- $0.1/M
- Output
- $0.4/M
Filter 384 source-linked models by creator, price, context, modality, and published benchmark coverage.
google/gemini-2.0-flash
Earlier Gemini Flash workhorse for responsive multimodal apps and tool use
google/gemini-2.0-flash-lite
Low-latency Gemini model for high-volume multimodal and agent workloads
openai/o1
O-series reasoning model for hard analysis, math, coding, and planning
openai/gpt-4o-2024-11-20
GPT model for general reasoning, writing, coding, and tool-assisted tasks
mistral/mistral-large-latest
Flagship Mistral model for advanced reasoning, coding, and multilingual work
mistral/pixtral-large-latest
Mistral's larger vision model for document-heavy image understanding and chat
mistral/mistral-large-2512
Mistral's largest general model for enterprise agents, coding, and multilingual reasoning
anthropic/claude-3-5-haiku-20241022
Fast Claude model for responsive assistance, classification, and lightweight agents
anthropic/claude-3-5-sonnet-20241022
Balanced Claude model for coding, analysis, agent workflows, and cost control
mistral/pixtral-12b
Mistral vision-language model for image understanding and multimodal chat
alibaba/qwen2-5-vl-72b-instruct
Qwen vision-language model for visual reasoning, documents, and agent tasks
openai/gpt-4o-2024-08-06
GPT model for general reasoning, writing, coding, and tool-assisted tasks
openai/gpt-4o-mini
Small omni GPT for cheap multimodal assistance and production-scale traffic
openai/o3-deep-research
Research model for long-horizon investigation, synthesis, and analytical reports
openai/o4-mini-deep-research
Research model for long-horizon investigation, synthesis, and analytical reports
openai/gpt-4o-2024-05-13
GPT model for general reasoning, writing, coding, and tool-assisted tasks
openai/gpt-4o
Omni-era GPT for multimodal chat, practical coding, and general assistants
alibaba/qwen-vl-max
Qwen vision-language model for visual reasoning, documents, and agent tasks
anthropic/claude-3-haiku-20240307
Legacy model retained for compatibility with older integrations
alibaba/qwen-vl-plus
Qwen vision-language model for visual reasoning, documents, and agent tasks
perplexity/sonar-reasoning-pro
Web-grounded Sonar for multi-step research questions that need cited reasoning
perplexity/sonar-pro
Deeper Sonar search model with broader retrieval and stronger synthesis
openai/gpt-4-turbo
Compact GPT model for low-latency assistance and high-volume workloads
openai/gpt-5.6-luna-pro
GPT-5.6 Luna Pro is the same underlying model as [GPT-5.6 Luna](https://openrouter.ai/openai/gpt-5.6-luna), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks. Learn more in OpenAI's docs: https://developers.openai.com/api/docs/guides/reasoning#reasoning-mode
openai/gpt-5.6-terra-pro
GPT-5.6 Terra Pro is the same underlying model as [GPT-5.6 Terra](https://openrouter.ai/openai/gpt-5.6-terra), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks. Learn more in OpenAI's docs: https://developers.openai.com/api/docs/guides/reasoning#reasoning-mode
openai/gpt-5.6-sol-pro
GPT-5.6 Sol Pro is the same underlying model as [GPT-5.6 Sol](https://openrouter.ai/openai/gpt-5.6-sol), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks. Learn more in OpenAI's docs: https://developers.openai.com/api/docs/guides/reasoning#reasoning-mode
x-ai/grok-4.5
Grok 4.5 is SpaceXAI's smartest model with frontier performance on coding, knowledge work, and STEM.
~x-ai/grok-latest
This model always redirects to the latest Grok model from xAI.
nex-agi/nex-n2-mini
Nex-N2-Mini is an open-source agentic mixture-of-experts model from Nex AGI, the smaller sibling in the Nex-N2 series. It accepts text and image input and is built for coding, tool use,...
~anthropic/claude-fable-latest
This model always redirects to the latest model in the Claude Fable family.
nex-agi/nex-n2-pro
Nex-N2-Pro is an agentic mixture-of-experts model from Nex AGI, with 17B active parameters out of 397B total. Built on the Qwen3.5 architecture, it accepts text and image input and produces...
nvidia/nemotron-3.5-content-safety:free
NVIDIA Nemotron 3.5 Content Safety is a compact 4B-parameter multimodal guardrail model from NVIDIA, fine-tuned from Google Gemma-3-4B. It moderates both inputs to and responses from LLMs and VLMs, accepting...
qwen/qwen3.7-plus
Qwen3.7-Plus is a cost-effective model in Alibaba's Qwen3.7 series. It supports text and image input with text output, building on the series' text capabilities with a comprehensive upgrade to its...
minimax/minimax-m3
MiniMax-M3 is a multimodal foundation model from MiniMax. It supports text, image, and video inputs with text output, a 1M-token context window, and is suited for long-horizon agentic work, coding,...
anthropic/claude-opus-4.8-fast
Fast-mode variant of [Opus 4.8](/anthropic/claude-opus-4.8) - identical capabilities with higher output speed at 2x pricing relative to regular Opus 4.8. Learn more in Anthropic's docs: https://platform.claude.com/docs/en/build-with-claude/fast-mode
anthropic/claude-opus-4.8
Claude Opus 4.8 is Anthropic's most capable generally available model in the Opus family. It supports text, image, and file inputs with text output, with reasoning support and a 1M-token...
x-ai/grok-build-0.1
Grok Build 0.1 is xAI’s fast coding model trained specifically for agentic software engineering workflows. It supports text and image inputs with text output, and is optimized for interactive coding...
anthropic/claude-opus-4.7-fast
Fast-mode variant of [Opus 4.7](/anthropic/claude-opus-4.7) - identical capabilities with higher output speed at premium 6x pricing. Learn more in Anthropic's docs: https://platform.claude.com/docs/en/build-with-claude/fast-mode
perceptron/perceptron-mk1
Perceptron Mk1 (Mark One) is Perceptron's highest-quality vision-language model for video and embodied reasoning.** It accepts image and video inputs paired with natural language queries, and produces detailed visual understanding...
openai/gpt-chat-latest
GPT Chat Latest points to OpenAI's stable API alias `chat-latest` that always resolves to the latest Instant chat model used in ChatGPT. As OpenAI rolls out new Instant model updates...
x-ai/grok-4.3
Grok 4.3 is a reasoning model from xAI. It accepts text and image inputs with text output, and is suited for agentic workflows, instruction-following tasks, and applications requiring high factual...
mistralai/mistral-medium-3-5
Mistral Medium 3.5 is a dense 128B instruction-following model from Mistral AI. It supports text and image inputs with text output, and is designed for agentic workflows, coding, and complex...
nvidia/nemotron-3-nano-omni-30b-a3b-reasoning:free
NVIDIA Nemotron™ 3 Nano Omni is a 30B-A3B open multimodal model designed to function as a perception and context sub-agent in enterprise agent systems. It accepts text, image, video, and...
~anthropic/claude-haiku-latest
This model always redirects to the latest model in the Anthropic Claude Haiku family.
~openai/gpt-mini-latest
This model always redirects to the latest model in the OpenAI GPT Mini family.
~google/gemini-pro-latest
This model always redirects to the latest model in the Google Gemini Pro family.
~moonshotai/kimi-latest
This model always redirects to the latest model in the MoonshotAI Kimi family.
~google/gemini-flash-latest
This model always redirects to the latest model in the Google Gemini Flash family.
~anthropic/claude-sonnet-latest
This model always redirects to the latest model in the Anthropic Claude Sonnet family.
~openai/gpt-latest
This model always redirects to the latest model in the OpenAI GPT family.
| Model | Creator | Input types | Context | Input / Output | Released | Compare |
|---|---|---|---|---|---|---|
| Gemini 2.0 Flashgoogle/gemini-2.0-flash | 1.04858M | $0.1 / $0.4 | 2024-12-11 | |||
| Gemini 2.0 Flash-Litegoogle/gemini-2.0-flash-lite | 1.04858M | $0.075 / $0.3 | 2024-12-11 | |||
| o1openai/o1 | 200K | $15 / $60 | 2024-12-05 | |||
| GPT-4o (2024-11-20)openai/gpt-4o-2024-11-20 | 128K | $2.5 / $10 | 2024-11-20 | |||
| Mistral Large (latest)mistral/mistral-large-latest | 262.144K | $0.5 / $1.5 | 2024-11-01 | |||
| Pixtral Large (latest)mistral/pixtral-large-latest | 128K | $2 / $6 | 2024-11-01 | |||
| Mistral Large 3mistral/mistral-large-2512 | 262.144K | $0.5 / $1.5 | 2024-11-01 | |||
| Claude Haiku 3.5anthropic/claude-3-5-haiku-20241022 | 200K | $0.8 / $4 | 2024-10-22 | |||
| Claude Sonnet 3.5 v2anthropic/claude-3-5-sonnet-20241022 | 200K | $3 / $15 | 2024-10-22 | |||
| Pixtral 12Bmistral/pixtral-12b | 128K | $0.15 / $0.15 | 2024-09-01 | |||
| Qwen2.5-VL 72B Instructalibaba/qwen2-5-vl-72b-instruct | 131.072K | $2.8 / $8.4 | 2024-09 | |||
| GPT-4o (2024-08-06)openai/gpt-4o-2024-08-06 | 128K | $2.5 / $10 | 2024-08-06 | |||
| GPT-4o miniopenai/gpt-4o-mini | 128K | $0.15 / $0.6 | 2024-07-18 | |||
| o3-deep-researchopenai/o3-deep-research | 200K | $9 / $36 | 2024-06-26 | |||
| o4-mini-deep-researchopenai/o4-mini-deep-research | 200K | $1.8 / $7.2 | 2024-06-26 | |||
| GPT-4o (2024-05-13)openai/gpt-4o-2024-05-13 | 128K | $5 / $15 | 2024-05-13 | |||
| GPT-4oopenai/gpt-4o | 128K | $2.5 / $10 | 2024-05-13 | |||
| Qwen-VL Maxalibaba/qwen-vl-max | 131.072K | $0.8 / $3.2 | 2024-04-08 | |||
| Claude Haiku 3anthropic/claude-3-haiku-20240307 | 200K | $0.25 / $1.25 | 2024-03-13 | |||
| Qwen-VL Plusalibaba/qwen-vl-plus | 131.072K | $0.21 / $0.63 | 2024-01-25 | |||
| Sonar Reasoning Properplexity/sonar-reasoning-pro | 128K | $2 / $8 | 2024-01-01 | |||
| Sonar Properplexity/sonar-pro | 200K | $3 / $15 | 2024-01-01 | |||
| GPT-4 Turboopenai/gpt-4-turbo | 128K | $10 / $30 | 2023-11-06 | |||
| OpenAI: GPT-5.6 Luna Proopenai/gpt-5.6-luna-pro | 1.05M | $0.1 / $0.6 | Undated | |||
| OpenAI: GPT-5.6 Terra Proopenai/gpt-5.6-terra-pro | 1.05M | $1 / $6 | Undated | |||
| OpenAI: GPT-5.6 Sol Proopenai/gpt-5.6-sol-pro | 1.05M | $5 / $30 | Undated | |||
| xAI: Grok 4.5x-ai/grok-4.5 | 500K | $2 / $6 | Undated | |||
| xAI: Grok Latest~x-ai/grok-latest | 500K | $2 / $6 | Undated | |||
| Nex AGI: Nex-N2-Mininex-agi/nex-n2-mini | 262.144K | $0.025 / $0.1 | Undated | |||
| Anthropic: Claude Fable Latest~anthropic/claude-fable-latest | 1M | $10 / $50 | Undated | |||
| Nex AGI: Nex-N2-Pronex-agi/nex-n2-pro | 262.144K | $0.25 / $1 | Undated | |||
| NVIDIA: Nemotron 3.5 Content Safety (free)nvidia/nemotron-3.5-content-safety:free | 128K | Free / Free | Undated | |||
| Qwen: Qwen3.7 Plusqwen/qwen3.7-plus | 1M | $0.32 / $1.28 | Undated | |||
| MiniMax: MiniMax M3minimax/minimax-m3 | 524.288K | $0.3 / $1.2 | Undated | |||
| Anthropic: Claude Opus 4.8 (Fast)anthropic/claude-opus-4.8-fast | 1M | $10 / $50 | Undated | |||
| Anthropic: Claude Opus 4.8anthropic/claude-opus-4.8 | 1M | $5 / $25 | Undated | |||
| xAI: Grok Build 0.1x-ai/grok-build-0.1 | 256K | $1 / $2 | Undated | |||
| Anthropic: Claude Opus 4.7 (Fast)anthropic/claude-opus-4.7-fast | 1M | $30 / $150 | Undated | |||
| Perceptron: Perceptron Mk1perceptron/perceptron-mk1 | 32.768K | $0.15 / $1.5 | Undated | |||
| OpenAI: GPT Chat Latestopenai/gpt-chat-latest | 400K | $5 / $30 | Undated | |||
| xAI: Grok 4.3x-ai/grok-4.3 | 1M | $1.25 / $2.5 | Undated | |||
| Mistral: Mistral Medium 3.5mistralai/mistral-medium-3-5 | 262.144K | $1.5 / $7.5 | Undated | |||
| NVIDIA: Nemotron 3 Nano Omni (free)nvidia/nemotron-3-nano-omni-30b-a3b-reasoning:free | 256K | Free / Free | Undated | |||
| Anthropic Claude Haiku Latest~anthropic/claude-haiku-latest | 200K | $1 / $5 | Undated | |||
| OpenAI GPT Mini Latest~openai/gpt-mini-latest | 400K | $0.75 / $4.5 | Undated | |||
| Google Gemini Pro Latest~google/gemini-pro-latest | 1.04858M | $2 / $12 | Undated | |||
| MoonshotAI Kimi Latest~moonshotai/kimi-latest | 1.04858M | $2.9 / $14 | Undated | |||
| Google Gemini Flash Latest~google/gemini-flash-latest | 1.04858M | $1.5 / $7.5 | Undated | |||
| Anthropic Claude Sonnet Latest~anthropic/claude-sonnet-latest | 1M | $2 / $10 | Undated | |||
| OpenAI GPT Latest~openai/gpt-latest | 1.05M | $5 / $30 | Undated |