nvidia/nemotron-voicechat
Nemotron multimodal model for visual reasoning and agentic AI workflows
- Context
- 128K
- Input
- -
- Output
- -
Filter 978 source-linked models by creator, price, context, modality, and published benchmark coverage.
nvidia/nemotron-voicechat
Nemotron multimodal model for visual reasoning and agentic AI workflows
nvidia/nemotron-3-super-120b-a12b
Nemotron middle tier for collaborative agents and high-volume reasoning workloads
xai/grok-4.20-0309-non-reasoning
Grok model for agentic tool use, reasoning, coding, and live assistance
xai/grok-4.20-0309-reasoning
Reasoning Grok for document-heavy analysis and long-horizon tool use
openai/gpt-5.4-pro
More exact GPT-5.4 tier for demanding professional reasoning and agent tasks
openai/gpt-5.4
Agent-ready GPT for coding and computer-use workflows at a lower cost
google/gemini-3.1-flash-lite-preview
Low-latency Gemini model for high-volume multimodal and agent workloads
openai/gpt-5.3-chat-latest
Chat-tuned GPT model for conversational assistance, writing, and tool workflows
google/gemini-3.1-flash-image-preview
Image model for prompt-driven generation, editing, and visual design workflows
alibaba/qwen3.5-27b
Qwen vision-language model for visual reasoning, documents, and agent tasks
alibaba/qwen3.5-35b-a3b
Qwen vision-language model for visual reasoning, documents, and agent tasks
alibaba/qwen3.5-9b
Qwen instruction model for multilingual chat, reasoning, and tool use
alibaba/qwen3.5-122b-a10b
Qwen vision-language model for visual reasoning, documents, and agent tasks
google/gemini-3.1-pro-preview-customtools
Advanced Gemini model for complex reasoning, coding, and multimodal analysis
google/gemini-3.1-pro-preview
Reasoning-first Gemini preview for agentic coding and complex problem solving
sarvam/sarvam-30b
Efficient Indian-language reasoning model for chat, coding, and multilingual work
anthropic/claude-sonnet-4-6
Claude workhorse for coding agents, careful analysis, and production cost control
alibaba/qwen3.5-plus
Qwen vision-language model for visual reasoning, documents, and agent tasks
alibaba/qwen3.5-397b-a17b
Large open Qwen multimodal MoE for visual agents and long technical tasks
minimax/MiniMax-M2.5-highspeed
High-speed MiniMax model for low-latency coding and agent workflows
zhipuai/glm-5
General GLM flagship for coding, analysis, and tool-heavy engineering workflows
minimax/MiniMax-M2.5
Prior MiniMax coding model for agent workflows, office edits, and automation
nvidia/llama-nemotron-embed-vl-1b-v2
Embedding model for semantic search, retrieval, clustering, and ranking pipelines
openai/gpt-5.3-codex
Coding-optimized GPT model for repository edits, reviews, and agentic software work
anthropic/claude-opus-4-6
High-end Claude for difficult coding, planning, and slower expert reasoning
stepfun/step-3.5-flash
StepFun flash lane for quick multimodal reasoning and coding assistance
nvidia/nemotron-content-safety-reasoning-4b
Safety model for policy screening, moderation, and risk-aware routing workflows
zhipuai/glm-4.7-flashx
Efficient GLM model for fast reasoning, coding, and agent workflows
zhipuai/glm-4.7-flash
Budget GLM lane for fast coding help, routing, and everyday automation
moonshotai/kimi-k2.5
Earlier Kimi frontier model for long-context agents, coding, and multimodal work
minimax/MiniMax-M2.1
Earlier MiniMax agent model for practical coding and productivity tasks
zhipuai/glm-4.7
Mature GLM model for dependable coding, reasoning, and structured agent tasks
google/gemini-3-flash-preview
New Gemini flash lane bringing frontier-style multimodal reasoning to cheaper runs
xiaomi/mimo-v2-flash
MiMo flash model for fast multimodal assistance and agent workflows
nvidia/nemotron-3-nano-30b-a3b
Small Nemotron 3 MoE for efficient coding, math, and long-context agents
openai/gpt-5.2-codex
Code-specialist GPT for repository edits, reviews, and long-running software agents
openai/gpt-5.2-pro
Higher-accuracy GPT-5.2 variant for tougher reasoning and review workflows
openai/gpt-5.2
Reliable GPT generation for broad coding, writing, and tool-assisted product work
openai/gpt-5.2-chat-latest
Chat-tuned GPT model for conversational assistance, writing, and tool workflows
mistral/devstral-2512
Mistral's coding-agent model for repository work, terminal tasks, and software fixes
zhipuai/glm-4.6v
GLM vision model for visual reasoning, documents, and multimodal agents
mistral/devstral-medium-latest
Mistral coding agent model for repository tasks and software engineering workflows
deepseek/deepseek-chat
DeepSeek chat model for instruction following, coding, and analysis
deepseek/deepseek-reasoner
DeepSeek reasoning model for multi-step analysis, math, coding, and tools
openai/gpt-image-1.5
Image model for prompt-driven generation, editing, and visual design workflows
anthropic/claude-opus-4-5
Flagship Claude model for deep reasoning, coding, and long-horizon agents
google/gemini-3-pro-image-preview
Nano Banana Pro for higher-fidelity image generation and design-heavy edits
google/gemini-3-pro-preview
Preview Gemini flagship for complex reasoning, coding, and rich multimodal prompts
openai/gpt-5.1-codex-max
Coding-optimized GPT model for repository edits, reviews, and agentic software work
openai/gpt-5.1
Sharper GPT-5 generation for coding, product work, and tool-assisted tasks
| Model | Creator | Input types | Context | Input / Output | Released | Compare |
|---|---|---|---|---|---|---|
| Nemotron VoiceChatnvidia/nemotron-voicechat | 128K | - / - | 2026-03-16 | |||
| Nemotron 3 Super 120B A12Bnvidia/nemotron-3-super-120b-a12b | 262.144K | $0.2 / $0.8 | 2026-03-11 | |||
| Grok 4.20 (Non-Reasoning)xai/grok-4.20-0309-non-reasoning | 1M | $1.25 / $2.5 | 2026-03-09 | |||
| Grok 4.20 (Reasoning)xai/grok-4.20-0309-reasoning | 1M | $1.25 / $2.5 | 2026-03-09 | |||
| GPT-5.4 Proopenai/gpt-5.4-pro | 1.05M | $30 / $180 | 2026-03-05 | |||
| GPT-5.4openai/gpt-5.4 | 1.05M | $2.5 / $15 | 2026-03-05 | |||
| Gemini 3.1 Flash Lite Previewgoogle/gemini-3.1-flash-lite-preview | 1.04858M | $0.25 / $1.5 | 2026-03-03 | |||
| GPT-5.3 Chat (latest)openai/gpt-5.3-chat-latest | 128K | $1.75 / $14 | 2026-03-03 | |||
| Nano Banana 2google/gemini-3.1-flash-image-preview | 65.536K | $0.5 / $3 | 2026-02-26 | |||
| Qwen3.5 27Balibaba/qwen3.5-27b | 262.144K | $0.3 / $2.4 | 2026-02-23 | |||
| Qwen3.5 35B-A3Balibaba/qwen3.5-35b-a3b | 262.144K | $0.25 / $2 | 2026-02-23 | |||
| Qwen3.5 9Balibaba/qwen3.5-9b | 262.144K | $0.04 / $0.15 | 2026-02-23 | |||
| Qwen3.5 122B-A10Balibaba/qwen3.5-122b-a10b | 262.144K | $0.4 / $3.2 | 2026-02-23 | |||
| Gemini 3.1 Pro Preview Custom Toolsgoogle/gemini-3.1-pro-preview-customtools | 1.04858M | $2 / $12 | 2026-02-19 | |||
| Gemini 3.1 Pro Previewgoogle/gemini-3.1-pro-preview | 1.04858M | $2 / $12 | 2026-02-19 | |||
| Sarvam 30Bsarvam/sarvam-30b | 128K | $0.02 / $0.1 | 2026-02-18 | |||
| Claude Sonnet 4.6anthropic/claude-sonnet-4-6 | 1M | $3 / $15 | 2026-02-17 | |||
| Qwen3.5 Plusalibaba/qwen3.5-plus | 1M | $0.4 / $2.4 | 2026-02-16 | |||
| Qwen3.5 397B-A17Balibaba/qwen3.5-397b-a17b | 262.144K | $0.6 / $3.6 | 2026-02-15 | |||
| MiniMax-M2.5-highspeedminimax/MiniMax-M2.5-highspeed | 204.8K | $0.6 / $2.4 | 2026-02-13 | |||
| GLM-5zhipuai/glm-5 | 204.8K | $1 / $3.2 | 2026-02-12 | |||
| MiniMax-M2.5minimax/MiniMax-M2.5 | 204.8K | $0.3 / $1.2 | 2026-02-12 | |||
| Llama Nemotron Embed VL 1B v2nvidia/llama-nemotron-embed-vl-1b-v2 | 32.768K | - / - | 2026-02-10 | |||
| GPT-5.3 Codexopenai/gpt-5.3-codex | 400K | $1.75 / $14 | 2026-02-05 | |||
| Claude Opus 4.6anthropic/claude-opus-4-6 | 1M | $5 / $25 | 2026-02-05 | |||
| Step 3.5 Flashstepfun/step-3.5-flash | 256K | $0.1 / $0.3 | 2026-01-29 | |||
| Nemotron Content Safety Reasoning 4Bnvidia/nemotron-content-safety-reasoning-4b | 128K | - / - | 2026-01-22 | |||
| GLM-4.7-FlashXzhipuai/glm-4.7-flashx | 200K | $0.07 / $0.4 | 2026-01-19 | |||
| GLM-4.7-Flashzhipuai/glm-4.7-flash | 200K | $0.04 / $0.3 | 2026-01-19 | |||
| Kimi K2.5moonshotai/kimi-k2.5 | 262.144K | $0.6 / $3 | 2026-01 | |||
| MiniMax-M2.1minimax/MiniMax-M2.1 | 204.8K | $0.3 / $1.2 | 2025-12-23 | |||
| GLM-4.7zhipuai/glm-4.7 | 204.8K | $0.6 / $2.2 | 2025-12-22 | |||
| Gemini 3 Flash Previewgoogle/gemini-3-flash-preview | 1.04858M | $0.5 / $3 | 2025-12-17 | |||
| MiMo-V2-Flashxiaomi/mimo-v2-flash | 262.144K | $0.14 / $0.28 | 2025-12-16 | |||
| Nemotron 3 Nano 30B A3Bnvidia/nemotron-3-nano-30b-a3b | 262.144K | $0.05 / $0.2 | 2025-12-15 | |||
| GPT-5.2 Codexopenai/gpt-5.2-codex | 400K | $0.14 / $1.14 | 2025-12-11 | |||
| GPT-5.2 Proopenai/gpt-5.2-pro | 400K | $21 / $168 | 2025-12-11 | |||
| GPT-5.2openai/gpt-5.2 | 400K | $1.75 / $14 | 2025-12-11 | |||
| GPT-5.2 Chatopenai/gpt-5.2-chat-latest | 128K | $1.75 / $14 | 2025-12-11 | |||
| Devstral 2mistral/devstral-2512 | 262.144K | $0.4 / $2 | 2025-12-09 | |||
| GLM-4.6Vzhipuai/glm-4.6v | 128K | $0.3 / $0.9 | 2025-12-08 | |||
| Devstral 2 (latest)mistral/devstral-medium-latest | 262.144K | $0.4 / $2 | 2025-12-02 | |||
| DeepSeek Chatdeepseek/deepseek-chat | 1M | $0.14 / $0.28 | 2025-12-01 | |||
| DeepSeek Reasonerdeepseek/deepseek-reasoner | 1M | $0.14 / $0.28 | 2025-12-01 | |||
| GPT-Image-1.5openai/gpt-image-1.5 | Not documented | $5 / $32 | 2025-11-25 | |||
| Claude Opus 4.5 (latest)anthropic/claude-opus-4-5 | 200K | $5 / $25 | 2025-11-24 | |||
| Nano Banana Progoogle/gemini-3-pro-image-preview | 65.536K | $2 / $12 | 2025-11-20 | |||
| Gemini 3 Pro Previewgoogle/gemini-3-pro-preview | 1.04858M | $2 / $12 | 2025-11-18 | |||
| GPT-5.1 Codex Maxopenai/gpt-5.1-codex-max | 400K | $1.1 / $9 | 2025-11-13 | |||
| GPT-5.1openai/gpt-5.1 | 400K | $1.25 / $10 | 2025-11-13 |