978 results
2026-03-16
nvidia/nemotron-voicechat

Nemotron multimodal model for visual reasoning and agentic AI workflows

T Tools Open weights
Context
128K
Input
-
Output
-
2026-03-11
nvidia/nemotron-3-super-120b-a12b

Nemotron middle tier for collaborative agents and high-volume reasoning workloads

T Reasoning Tools Open weights
Context
262.144K
Input
$0.2/M
Output
$0.8/M
2026-03-09
xai/grok-4.20-0309-non-reasoning

Grok model for agentic tool use, reasoning, coding, and live assistance

T Tools Structured output Weight access not listed
Context
1M
Input
$1.25/M
Output
$2.5/M
2026-03-09
xai/grok-4.20-0309-reasoning

Reasoning Grok for document-heavy analysis and long-horizon tool use

T Reasoning Tools Structured output Weight access not listed
Context
1M
Input
$1.25/M
Output
$2.5/M
GPT-5.4 Proby OpenAI
2026-03-05
openai/gpt-5.4-pro

More exact GPT-5.4 tier for demanding professional reasoning and agent tasks

T Reasoning Tools Weight access not listed
Context
1.05M
Input
$30/M
Output
$180/M
GPT-5.4by OpenAI
2026-03-05
openai/gpt-5.4

Agent-ready GPT for coding and computer-use workflows at a lower cost

T Reasoning Tools Structured output Weight access not listed
Context
1.05M
Input
$2.5/M
Output
$15/M
2026-03-03
google/gemini-3.1-flash-lite-preview

Low-latency Gemini model for high-volume multimodal and agent workloads

T Reasoning Tools Structured output Weight access not listed
Context
1.04858M
Input
$0.25/M
Output
$1.5/M
2026-03-03
openai/gpt-5.3-chat-latest

Chat-tuned GPT model for conversational assistance, writing, and tool workflows

T Tools Structured output Weight access not listed
Context
128K
Input
$1.75/M
Output
$14/M
Nano Banana 2by Google
2026-02-26
google/gemini-3.1-flash-image-preview

Image model for prompt-driven generation, editing, and visual design workflows

T Reasoning Weight access not listed
Context
65.536K
Input
$0.5/M
Output
$3/M
Qwen3.5 27Bby Alibaba Qwen
2026-02-23
alibaba/qwen3.5-27b

Qwen vision-language model for visual reasoning, documents, and agent tasks

T Tools Structured output Open weights
Context
262.144K
Input
$0.3/M
Output
$2.4/M
Qwen3.5 35B-A3Bby Alibaba Qwen
2026-02-23
alibaba/qwen3.5-35b-a3b

Qwen vision-language model for visual reasoning, documents, and agent tasks

T Tools Structured output Open weights
Context
262.144K
Input
$0.25/M
Output
$2/M
Qwen3.5 9Bby Alibaba Qwen
2026-02-23
alibaba/qwen3.5-9b

Qwen instruction model for multilingual chat, reasoning, and tool use

T Reasoning Tools Structured output Open weights
Context
262.144K
Input
$0.04/M
Output
$0.15/M
Qwen3.5 122B-A10Bby Alibaba Qwen
2026-02-23
alibaba/qwen3.5-122b-a10b

Qwen vision-language model for visual reasoning, documents, and agent tasks

T Reasoning Tools Structured output Open weights
Context
262.144K
Input
$0.4/M
Output
$3.2/M
google/gemini-3.1-pro-preview-customtools

Advanced Gemini model for complex reasoning, coding, and multimodal analysis

T Reasoning Tools Structured output Weight access not listed
Context
1.04858M
Input
$2/M
Output
$12/M
2026-02-19
google/gemini-3.1-pro-preview

Reasoning-first Gemini preview for agentic coding and complex problem solving

T Reasoning Tools Structured output Weight access not listed
Context
1.04858M
Input
$2/M
Output
$12/M
Sarvam 30Bby Sarvam
2026-02-18
sarvam/sarvam-30b

Efficient Indian-language reasoning model for chat, coding, and multilingual work

T Reasoning Tools Open weights
Context
128K
Input
$0.02/M
Output
$0.1/M
Claude Sonnet 4.6by Anthropic
2026-02-17
anthropic/claude-sonnet-4-6

Claude workhorse for coding agents, careful analysis, and production cost control

T Reasoning Tools Weight access not listed
Context
1M
Input
$3/M
Output
$15/M
Qwen3.5 Plusby Alibaba Qwen
2026-02-16
alibaba/qwen3.5-plus

Qwen vision-language model for visual reasoning, documents, and agent tasks

T Reasoning Tools Weight access not listed
Context
1M
Input
$0.4/M
Output
$2.4/M
Qwen3.5 397B-A17Bby Alibaba Qwen
2026-02-15
alibaba/qwen3.5-397b-a17b

Large open Qwen multimodal MoE for visual agents and long technical tasks

T Reasoning Tools Structured output Open weights
Context
262.144K
Input
$0.6/M
Output
$3.6/M
2026-02-13
minimax/MiniMax-M2.5-highspeed

High-speed MiniMax model for low-latency coding and agent workflows

T Reasoning Tools Open weights
Context
204.8K
Input
$0.6/M
Output
$2.4/M
GLM-5by Zhipu AI
2026-02-12
zhipuai/glm-5

General GLM flagship for coding, analysis, and tool-heavy engineering workflows

T Reasoning Tools Open weights
Context
204.8K
Input
$1/M
Output
$3.2/M
MiniMax-M2.5by MiniMax
2026-02-12
minimax/MiniMax-M2.5

Prior MiniMax coding model for agent workflows, office edits, and automation

T Reasoning Tools Open weights
Context
204.8K
Input
$0.3/M
Output
$1.2/M
2026-02-10
nvidia/llama-nemotron-embed-vl-1b-v2

Embedding model for semantic search, retrieval, clustering, and ranking pipelines

T Open weights
Context
32.768K
Input
-
Output
-
GPT-5.3 Codexby OpenAI
2026-02-05
openai/gpt-5.3-codex

Coding-optimized GPT model for repository edits, reviews, and agentic software work

T Reasoning Tools Weight access not listed
Context
400K
Input
$1.75/M
Output
$14/M
Claude Opus 4.6by Anthropic
2026-02-05
anthropic/claude-opus-4-6

High-end Claude for difficult coding, planning, and slower expert reasoning

T Reasoning Tools Weight access not listed
Context
1M
Input
$5/M
Output
$25/M
Step 3.5 Flashby StepFun
2026-01-29
stepfun/step-3.5-flash

StepFun flash lane for quick multimodal reasoning and coding assistance

T Reasoning Tools Open weights
Context
256K
Input
$0.1/M
Output
$0.3/M
nvidia/nemotron-content-safety-reasoning-4b

Safety model for policy screening, moderation, and risk-aware routing workflows

T Reasoning Open weights
Context
128K
Input
-
Output
-
GLM-4.7-FlashXby Zhipu AI
2026-01-19
zhipuai/glm-4.7-flashx

Efficient GLM model for fast reasoning, coding, and agent workflows

T Reasoning Tools Open weights
Context
200K
Input
$0.07/M
Output
$0.4/M
GLM-4.7-Flashby Zhipu AI
2026-01-19
zhipuai/glm-4.7-flash

Budget GLM lane for fast coding help, routing, and everyday automation

T Reasoning Tools Open weights
Context
200K
Input
$0.04/M
Output
$0.3/M
Kimi K2.5by Moonshot AI
2026-01
moonshotai/kimi-k2.5

Earlier Kimi frontier model for long-context agents, coding, and multimodal work

T Reasoning Tools Structured output Open weights
Context
262.144K
Input
$0.6/M
Output
$3/M
MiniMax-M2.1by MiniMax
2025-12-23
minimax/MiniMax-M2.1

Earlier MiniMax agent model for practical coding and productivity tasks

T Reasoning Tools Open weights
Context
204.8K
Input
$0.3/M
Output
$1.2/M
GLM-4.7by Zhipu AI
2025-12-22
zhipuai/glm-4.7

Mature GLM model for dependable coding, reasoning, and structured agent tasks

T Reasoning Tools Open weights
Context
204.8K
Input
$0.6/M
Output
$2.2/M
2025-12-17
google/gemini-3-flash-preview

New Gemini flash lane bringing frontier-style multimodal reasoning to cheaper runs

T Reasoning Tools Structured output Weight access not listed
Context
1.04858M
Input
$0.5/M
Output
$3/M
MiMo-V2-Flashby Xiaomi
2025-12-16
xiaomi/mimo-v2-flash

MiMo flash model for fast multimodal assistance and agent workflows

T Reasoning Open weights
Context
262.144K
Input
$0.14/M
Output
$0.28/M
2025-12-15
nvidia/nemotron-3-nano-30b-a3b

Small Nemotron 3 MoE for efficient coding, math, and long-context agents

T Reasoning Tools Open weights
Context
262.144K
Input
$0.05/M
Output
$0.2/M
GPT-5.2 Codexby OpenAI
2025-12-11
openai/gpt-5.2-codex

Code-specialist GPT for repository edits, reviews, and long-running software agents

T Reasoning Tools Weight access not listed
Context
400K
Input
$0.14/M
Output
$1.14/M
GPT-5.2 Proby OpenAI
2025-12-11
openai/gpt-5.2-pro

Higher-accuracy GPT-5.2 variant for tougher reasoning and review workflows

T Reasoning Tools Weight access not listed
Context
400K
Input
$21/M
Output
$168/M
GPT-5.2by OpenAI
2025-12-11
openai/gpt-5.2

Reliable GPT generation for broad coding, writing, and tool-assisted product work

T Reasoning Tools Structured output Weight access not listed
Context
400K
Input
$1.75/M
Output
$14/M
GPT-5.2 Chatby OpenAI
2025-12-11
openai/gpt-5.2-chat-latest

Chat-tuned GPT model for conversational assistance, writing, and tool workflows

T Reasoning Tools Structured output Weight access not listed
Context
128K
Input
$1.75/M
Output
$14/M
Devstral 2by Mistral AI
2025-12-09
mistral/devstral-2512

Mistral's coding-agent model for repository work, terminal tasks, and software fixes

T Tools Open weights
Context
262.144K
Input
$0.4/M
Output
$2/M
GLM-4.6Vby Zhipu AI
2025-12-08
zhipuai/glm-4.6v

GLM vision model for visual reasoning, documents, and multimodal agents

T Reasoning Tools Open weights
Context
128K
Input
$0.3/M
Output
$0.9/M
Devstral 2 (latest)by Mistral AI
2025-12-02
mistral/devstral-medium-latest

Mistral coding agent model for repository tasks and software engineering workflows

T Tools Open weights
Context
262.144K
Input
$0.4/M
Output
$2/M
DeepSeek Chatby DeepSeek
2025-12-01
deepseek/deepseek-chat

DeepSeek chat model for instruction following, coding, and analysis

T Tools Open weights
Context
1M
Input
$0.14/M
Output
$0.28/M
2025-12-01
deepseek/deepseek-reasoner

DeepSeek reasoning model for multi-step analysis, math, coding, and tools

T Open weights
Context
1M
Input
$0.14/M
Output
$0.28/M
GPT-Image-1.5by OpenAI
2025-11-25
openai/gpt-image-1.5

Image model for prompt-driven generation, editing, and visual design workflows

T Weight access not listed
Context
Not documented
Input
$5/M
Output
$32/M
2025-11-24
anthropic/claude-opus-4-5

Flagship Claude model for deep reasoning, coding, and long-horizon agents

T Reasoning Tools Weight access not listed
Context
200K
Input
$5/M
Output
$25/M
2025-11-20
google/gemini-3-pro-image-preview

Nano Banana Pro for higher-fidelity image generation and design-heavy edits

T Reasoning Weight access not listed
Context
65.536K
Input
$2/M
Output
$12/M
2025-11-18
google/gemini-3-pro-preview

Preview Gemini flagship for complex reasoning, coding, and rich multimodal prompts

T Reasoning Tools Structured output Weight access not listed
Context
1.04858M
Input
$2/M
Output
$12/M
2025-11-13
openai/gpt-5.1-codex-max

Coding-optimized GPT model for repository edits, reviews, and agentic software work

T Reasoning Tools Weight access not listed
Context
400K
Input
$1.1/M
Output
$9/M
GPT-5.1by OpenAI
2025-11-13
openai/gpt-5.1

Sharper GPT-5 generation for coding, product work, and tool-assisted tasks

T Reasoning Tools Structured output Weight access not listed
Context
400K
Input
$1.25/M
Output
$10/M