xai/grok-4.3
xAI's default Grok for chat, coding, agentic tools, and lower hallucination risk
- Context
- 1M
- Input
- $1.25/M
- Output
- $2.5/M
Filter 384 source-linked models by creator, price, context, modality, and published benchmark coverage.
xai/grok-4.3
xAI's default Grok for chat, coding, agentic tools, and lower hallucination risk
alibaba/qwen3.6-35b-a3b
Open multimodal Qwen MoE for local agents that need vision, audio, and code
xai/grok-build-0.1
Fast Grok coding model tuned for agentic engineering and iterative edits
anthropic/claude-opus-4-7
Stronger Opus tier for advanced software work and high-stakes reasoning
google/gemini-robotics-er-1.6-preview
Vision-language model for embodied reasoning: spatial understanding, task planning, and physical-world agentic robotics
meta/muse-spark-1.1
Muse Spark is a natively multimodal reasoning model with support for tool-use, visual chain of thought, and multi-agent orchestration.
google/gemma-4-26b-a4b-it
Open Gemma instruction model for efficient chat and self-hosted deployments
google/gemma-4-31b-it
Largest Gemma 4 instruction model for open, self-hosted chat and reasoning
google/gemma-4-E4B-it
Open Gemma instruction model for efficient chat and self-hosted deployments
google/gemma-4-E2B-it
Open Gemma instruction model for efficient chat and self-hosted deployments
alibaba/qwen3.6-plus
Earlier Qwen multimodal workhorse for million-token agent and document tasks
zhipuai/glm-5v-turbo
Fast GLM vision model for screenshots, documents, and multimodal agent tasks
nvidia/llama-nemotron-rerank-vl-1b-v2
Reranking model for improving retrieval quality in search and recommendation systems
google/veo-3.1-lite-generate-preview
Video model for prompt-guided generation, editing, and motion workflows
google/gemini-3.1-flash-live-preview
High-quality, low-latency Live API model for real-time dialogue and voice-first AI applications
google/lyria-3-pro-preview
Music generation model for full-length songs from text or images with vocals and structure
google/lyria-3-clip-preview
Music generation model for short 30-second clips, loops, and previews from text or image prompts
xiaomi/mimo-v2-omni
MiMo omni model for text, image, video, audio, and agents
openai/gpt-5.4-nano
Cheapest GPT-5.4 lane for simple routing, extraction, and bulk automation
openai/gpt-5.4-mini
Strong small GPT for coding subagents, quick tool use, and high-volume work
mistral/mistral-small-latest
Efficient Mistral model for fast chat, extraction, and production assistants
mistral/mistral-small-2603
Fast Mistral production model for chat, extraction, and cost-sensitive agents
xai/grok-4.20-0309-non-reasoning
Grok model for agentic tool use, reasoning, coding, and live assistance
xai/grok-4.20-0309-reasoning
Reasoning Grok for document-heavy analysis and long-horizon tool use
openai/gpt-5.4-pro
More exact GPT-5.4 tier for demanding professional reasoning and agent tasks
openai/gpt-5.4
Agent-ready GPT for coding and computer-use workflows at a lower cost
google/gemini-3.1-flash-lite-preview
Low-latency Gemini model for high-volume multimodal and agent workloads
openai/gpt-5.3-chat-latest
Chat-tuned GPT model for conversational assistance, writing, and tool workflows
google/gemini-3.1-flash-image-preview
Image model for prompt-driven generation, editing, and visual design workflows
alibaba/qwen3.5-27b
Qwen vision-language model for visual reasoning, documents, and agent tasks
alibaba/qwen3.5-35b-a3b
Qwen vision-language model for visual reasoning, documents, and agent tasks
alibaba/qwen3.5-122b-a10b
Qwen vision-language model for visual reasoning, documents, and agent tasks
google/gemini-3.1-pro-preview-customtools
Advanced Gemini model for complex reasoning, coding, and multimodal analysis
google/gemini-3.1-pro-preview
Reasoning-first Gemini preview for agentic coding and complex problem solving
anthropic/claude-sonnet-4-6
Claude workhorse for coding agents, careful analysis, and production cost control
alibaba/qwen3.5-plus
Qwen vision-language model for visual reasoning, documents, and agent tasks
alibaba/qwen3.5-397b-a17b
Large open Qwen multimodal MoE for visual agents and long technical tasks
nvidia/llama-nemotron-embed-vl-1b-v2
Embedding model for semantic search, retrieval, clustering, and ranking pipelines
openai/gpt-5.3-codex
Coding-optimized GPT model for repository edits, reviews, and agentic software work
anthropic/claude-opus-4-6
High-end Claude for difficult coding, planning, and slower expert reasoning
moonshotai/kimi-k2.5
Earlier Kimi frontier model for long-context agents, coding, and multimodal work
google/gemini-3-flash-preview
New Gemini flash lane bringing frontier-style multimodal reasoning to cheaper runs
openai/gpt-5.2-codex
Code-specialist GPT for repository edits, reviews, and long-running software agents
openai/gpt-5.2-pro
Higher-accuracy GPT-5.2 variant for tougher reasoning and review workflows
openai/gpt-5.2
Reliable GPT generation for broad coding, writing, and tool-assisted product work
openai/gpt-5.2-chat-latest
Chat-tuned GPT model for conversational assistance, writing, and tool workflows
zhipuai/glm-4.6v
GLM vision model for visual reasoning, documents, and multimodal agents
openai/gpt-image-1.5
Image model for prompt-driven generation, editing, and visual design workflows
anthropic/claude-opus-4-5
Flagship Claude model for deep reasoning, coding, and long-horizon agents
google/gemini-3-pro-image-preview
Nano Banana Pro for higher-fidelity image generation and design-heavy edits
| Model | Creator | Input types | Context | Input / Output | Released | Compare |
|---|---|---|---|---|---|---|
| Grok 4.3xai/grok-4.3 | 1M | $1.25 / $2.5 | 2026-04-17 | |||
| Qwen3.6 35B-A3Balibaba/qwen3.6-35b-a3b | 262.144K | $0.248 / $1.485 | 2026-04-17 | |||
| Grok Build 0.1xai/grok-build-0.1 | 256K | $1 / $2 | 2026-04-16 | |||
| Claude Opus 4.7anthropic/claude-opus-4-7 | 1M | $5 / $25 | 2026-04-16 | |||
| Gemini Robotics-ER 1.6 Previewgoogle/gemini-robotics-er-1.6-preview | 131.072K | $1 / $5 | 2026-04-14 | |||
| Muse Spark 1.1meta/muse-spark-1.1 | 1M | $1.25 / $4.25 | 2026-04-08 | |||
| Gemma 4 26B A4B ITgoogle/gemma-4-26b-a4b-it | 262.144K | $0.06 / $0.33 | 2026-04-02 | |||
| Gemma 4 31B ITgoogle/gemma-4-31b-it | 262.144K | $0.1 / $0.35 | 2026-04-02 | |||
| Gemma 4 E4B ITgoogle/gemma-4-E4B-it | 131.072K | $0.2 / $0.2 | 2026-04-02 | |||
| Gemma 4 E2B ITgoogle/gemma-4-E2B-it | 131.072K | $0.1 / $0.1 | 2026-04-02 | |||
| Qwen3.6 Plusalibaba/qwen3.6-plus | 1M | $0.5 / $3 | 2026-04-02 | |||
| GLM-5V-Turbozhipuai/glm-5v-turbo | 200K | $5 / $22 | 2026-04-01 | |||
| Llama Nemotron Rerank VL 1B v2nvidia/llama-nemotron-rerank-vl-1b-v2 | 128K | - / - | 2026-03-31 | |||
| Veo 3.1 Lite Previewgoogle/veo-3.1-lite-generate-preview | 1.024K | - / - | 2026-03-31 | |||
| Gemini 3.1 Flash Live Previewgoogle/gemini-3.1-flash-live-preview | 131.072K | $0.75 / $4.5 | 2026-03-26 | |||
| Lyria 3 Pro Previewgoogle/lyria-3-pro-preview | 131.072K | - / - | 2026-03-25 | |||
| Lyria 3 Clip Previewgoogle/lyria-3-clip-preview | 131.072K | - / - | 2026-03-25 | |||
| MiMo-V2-Omnixiaomi/mimo-v2-omni | 262.144K | $0.14 / $0.28 | 2026-03-18 | |||
| GPT-5.4 nanoopenai/gpt-5.4-nano | 400K | $0.2 / $1.25 | 2026-03-17 | |||
| GPT-5.4 miniopenai/gpt-5.4-mini | 400K | $0.75 / $4.5 | 2026-03-17 | |||
| Mistral Small (latest)mistral/mistral-small-latest | 256K | $0.15 / $0.6 | 2026-03-16 | |||
| Mistral Small 4mistral/mistral-small-2603 | 256K | $0.15 / $0.6 | 2026-03-16 | |||
| Grok 4.20 (Non-Reasoning)xai/grok-4.20-0309-non-reasoning | 1M | $1.25 / $2.5 | 2026-03-09 | |||
| Grok 4.20 (Reasoning)xai/grok-4.20-0309-reasoning | 1M | $1.25 / $2.5 | 2026-03-09 | |||
| GPT-5.4 Proopenai/gpt-5.4-pro | 1.05M | $30 / $180 | 2026-03-05 | |||
| GPT-5.4openai/gpt-5.4 | 1.05M | $2.5 / $15 | 2026-03-05 | |||
| Gemini 3.1 Flash Lite Previewgoogle/gemini-3.1-flash-lite-preview | 1.04858M | $0.25 / $1.5 | 2026-03-03 | |||
| GPT-5.3 Chat (latest)openai/gpt-5.3-chat-latest | 128K | $1.75 / $14 | 2026-03-03 | |||
| Nano Banana 2google/gemini-3.1-flash-image-preview | 65.536K | $0.5 / $3 | 2026-02-26 | |||
| Qwen3.5 27Balibaba/qwen3.5-27b | 262.144K | $0.3 / $2.4 | 2026-02-23 | |||
| Qwen3.5 35B-A3Balibaba/qwen3.5-35b-a3b | 262.144K | $0.25 / $2 | 2026-02-23 | |||
| Qwen3.5 122B-A10Balibaba/qwen3.5-122b-a10b | 262.144K | $0.4 / $3.2 | 2026-02-23 | |||
| Gemini 3.1 Pro Preview Custom Toolsgoogle/gemini-3.1-pro-preview-customtools | 1.04858M | $2 / $12 | 2026-02-19 | |||
| Gemini 3.1 Pro Previewgoogle/gemini-3.1-pro-preview | 1.04858M | $2 / $12 | 2026-02-19 | |||
| Claude Sonnet 4.6anthropic/claude-sonnet-4-6 | 1M | $3 / $15 | 2026-02-17 | |||
| Qwen3.5 Plusalibaba/qwen3.5-plus | 1M | $0.4 / $2.4 | 2026-02-16 | |||
| Qwen3.5 397B-A17Balibaba/qwen3.5-397b-a17b | 262.144K | $0.6 / $3.6 | 2026-02-15 | |||
| Llama Nemotron Embed VL 1B v2nvidia/llama-nemotron-embed-vl-1b-v2 | 32.768K | - / - | 2026-02-10 | |||
| GPT-5.3 Codexopenai/gpt-5.3-codex | 400K | $1.75 / $14 | 2026-02-05 | |||
| Claude Opus 4.6anthropic/claude-opus-4-6 | 1M | $5 / $25 | 2026-02-05 | |||
| Kimi K2.5moonshotai/kimi-k2.5 | 262.144K | $0.6 / $3 | 2026-01 | |||
| Gemini 3 Flash Previewgoogle/gemini-3-flash-preview | 1.04858M | $0.5 / $3 | 2025-12-17 | |||
| GPT-5.2 Codexopenai/gpt-5.2-codex | 400K | $0.14 / $1.14 | 2025-12-11 | |||
| GPT-5.2 Proopenai/gpt-5.2-pro | 400K | $21 / $168 | 2025-12-11 | |||
| GPT-5.2openai/gpt-5.2 | 400K | $1.75 / $14 | 2025-12-11 | |||
| GPT-5.2 Chatopenai/gpt-5.2-chat-latest | 128K | $1.75 / $14 | 2025-12-11 | |||
| GLM-4.6Vzhipuai/glm-4.6v | 128K | $0.3 / $0.9 | 2025-12-08 | |||
| GPT-Image-1.5openai/gpt-image-1.5 | Not documented | $5 / $32 | 2025-11-25 | |||
| Claude Opus 4.5 (latest)anthropic/claude-opus-4-5 | 200K | $5 / $25 | 2025-11-24 | |||
| Nano Banana Progoogle/gemini-3-pro-image-preview | 65.536K | $2 / $12 | 2025-11-20 |