978 results
Claude Opus 5by Anthropic
2026-07-24
anthropic/claude-opus-5

Strongest Claude Opus model for coding, agents, and professional work

T Reasoning Tools Weight access not listed
Context
1M
Input
$5/M
Output
$25/M
Laguna S 2.1by Poolside
2026-07-21
poolside/laguna-s-2.1

Agentic coding model from Poolside in the XS size class for local deployment

T Reasoning Tools Open weights
Context
1.04858M
Input
$0.1/M
Output
$0.2/M
2026-07-21
google/gemini-3.6-flash

Fast Gemini model balancing multimodal reasoning, tool use, and cost

T Reasoning Tools Structured output Weight access not listed
Context
1.04858M
Input
$1.5/M
Output
$7.5/M
2026-07-21
google/gemini-3.5-flash-lite

Fast Gemini model balancing multimodal reasoning, tool use, and cost

T Reasoning Tools Structured output Weight access not listed
Context
1.04858M
Input
$0.3/M
Output
$2.5/M
Qwen3.8 Max Previewby Alibaba Qwen
2026-07-19
alibaba/qwen3.8-max-preview

Preview Qwen flagship for million-token multimodal reasoning and long-horizon agentic workflows

T Reasoning Tools Structured output Weight access not listed
Context
1M
Input
-
Output
-
Kimi K3by Moonshot AI
2026-07-16
moonshotai/kimi-k3

Multimodal Kimi model with 1M context and toggleable max-effort thinking for long-horizon agent work

T Reasoning Tools Structured output Open weights
Context
1.04858M
Input
$3/M
Output
$15/M
Inklingby Thinkingmachines
2026-07-15
thinkingmachines/inkling

Multimodal MoE reasoning model (975B total, 41B active) for text, image, and audio

T Reasoning Tools Structured output Open weights
Context
1.04858M
Input
$1.87/M
Output
$4.68/M
GPT-5.6 Lunaby OpenAI
2026-07-09
openai/gpt-5.6-luna

Cost-efficient GPT-5.6 model for fast, high-volume workloads

T Reasoning Tools Structured output Weight access not listed
Context
1.05M
Input
$1/M
Output
$6/M
GPT-5.6 Terraby OpenAI
2026-07-09
openai/gpt-5.6-terra

Balanced GPT-5.6 model for capable, cost-efficient everyday work

T Reasoning Tools Structured output Weight access not listed
Context
1.05M
Input
$2.5/M
Output
$15/M
GPT-5.6 Solby OpenAI
2026-07-09
openai/gpt-5.6-sol

Frontier GPT-5.6 model for complex professional work, coding, and agentic workflows

T Reasoning Tools Structured output Weight access not listed
Context
1.05M
Input
$5/M
Output
$30/M
Grok 4.5by xAI
2026-07-08
xai/grok-4.5

xAI's latest Grok for chat, coding, agentic tools, and lower hallucination risk

T Reasoning Tools Structured output Weight access not listed
Context
500K
Input
$2/M
Output
$6/M
Hy3by Tencent
2026-07-06
tencent/hy3

Tencent Hy reasoning model for coding, instruction following, and agent tasks

T Reasoning Tools Open weights
Context
256K
Input
$0.132/M
Output
$0.528/M
2026-07-06
openai/gpt-realtime-2.1

Realtime speech-to-speech model with configurable reasoning, tool use, and robust voice-agent behavior

T Reasoning Weight access not listed
Context
128K
Input
$4/M
Output
$24/M
Laguna XS 2.1by Poolside
2026-07-02
poolside/laguna-xs-2.1

Agentic coding model from Poolside in the XS size class for local deployment

T Reasoning Tools Open weights
Context
262.144K
Input
$0.06/M
Output
$0.12/M
Claude Sonnet 5by Anthropic
2026-06-30
anthropic/claude-sonnet-5

Everyday Claude agent model for coding, planning, browsing, and general work

T Reasoning Tools Weight access not listed
Context
1M
Input
$2/M
Output
$10/M
2026-06-30
google/gemini-3.1-flash-lite-image

Fastest, most cost-efficient Gemini image model for high-volume 1K generation and editing

T Reasoning Tools Weight access not listed
Context
65.536K
Input
$0.25/M
Output
$30/M
2026-06-30
google/gemini-omni-flash-preview

Video generation and editing model for fast, conversational text- and image-to-video workflows

T Reasoning Weight access not listed
Context
1.04858M
Input
$1.5/M
Output
$17.5/M
LongCat-2.0by Meituan
2026-06-30
meituan/longcat-2.0

Meituan LongCat-2.0, a reasoning model with tool calling and a 1M-token context window

T Reasoning Tools Weight access not listed
Context
1M
Input
$0.3/M
Output
$1.2/M
Ornith 1.0 31Bby Deepreinforce
2026-06-25
deepreinforce/ornith-1.0-31b

Open coding-reasoning model for repository tasks and self-improving agents

T Open weights
Context
262.144K
Input
-
Output
-
Ornith 1.0 35Bby Deepreinforce
2026-06-25
deepreinforce/ornith-1.0-35b

Large coding-reasoning model for agentic software tasks and RL search

T Open weights
Context
262.144K
Input
-
Output
-
Ornith 1.0 9Bby Deepreinforce
2026-06-25
deepreinforce/ornith-1.0-9b

Open coding-reasoning model for repository tasks and self-improving agents

T Open weights
Context
262.144K
Input
-
Output
-
Ornith 1.0 397Bby Deepreinforce
2026-06-25
deepreinforce/ornith-1.0-397b

Large coding-reasoning model for agentic software tasks and RL search

T Open weights
Context
262.144K
Input
-
Output
-
Fugu Ultraby Sakana
2026-06-15
sakana/fugu-ultra

Quality-first multi-agent model for hard research, analysis, and competitions

T Reasoning Tools Structured output Weight access not listed
Context
1M
Input
$5/M
Output
$30/M
Fuguby Sakana
2026-06-15
sakana/fugu

Multi-agent model for routing expert agents across complex analytical tasks

T Reasoning Tools Structured output Weight access not listed
Context
1M
Input
-
Output
-
GLM-5.2by Zhipu AI
2026-06-13
zhipuai/glm-5.2

Open flagship GLM for long-horizon coding agents and million-token context work

T Reasoning Tools Structured output Open weights
Context
1M
Input
$1.4/M
Output
$4.4/M
Kimi K2.7 Codeby Moonshot AI
2026-06-12
moonshotai/kimi-k2.7-code

Coding-focused Kimi model, stronger on long-horizon repo work with less overthinking

T Reasoning Tools Open weights
Context
262.144K
Input
$0.95/M
Output
$4/M
2026-06-12
moonshotai/kimi-k2.7-code-highspeed

Lower-latency Kimi Code variant for interactive edits and coding-agent loops

T Reasoning Tools Structured output Open weights
Context
262.144K
Input
$1.9/M
Output
$8/M
Claude Fable 5by Anthropic
2026-06-09
anthropic/claude-fable-5

Claude model for creative writing, analysis, and controlled agent workflows

T Reasoning Tools Weight access not listed
Context
1M
Input
$10/M
Output
$50/M
2026-06-09
cohere/north-mini-code-1-0

Cohere coding model for practical software engineering and agentic edits

T Reasoning Tools Structured output Open weights
Context
256K
Input
-
Output
-
2026-06-09
google/gemini-3.5-live-translate-preview

Low-latency audio-to-audio model for real-time speech translation across 70+ languages

Weight access not listed
Context
131.072K
Input
$3.5/M
Output
$21/M
2026-06-08
xiaomi/mimo-v2.5-pro-ultraspeed

MiMo pro model for strong multimodal reasoning and agent execution

T Reasoning Tools Open weights
Context
1.04858M
Input
$1.305/M
Output
$2.61/M
2026-06-04
nvidia/nemotron-3-ultra-550b-a55b

Largest Nemotron 3 model for maximum open-weight reasoning and agent accuracy

T Reasoning Tools Structured output Open weights
Context
1M
Input
$0.5/M
Output
$2.5/M
2026-06-04
nvidia/nemotron-3.5-content-safety

Safety model for policy screening, moderation, and risk-aware routing workflows

T Open weights
Context
128K
Input
-
Output
-
MAI-Code-1-Flashby Microsoft
2026-06-02
microsoft/mai-code-1-flash

Microsoft coding model built for fast, efficient assistance in everyday developer workflows

T Reasoning Tools Structured output Weight access not listed
Context
256K
Input
$0.75/M
Output
$4.5/M
Qwen3.7 Plusby Alibaba Qwen
2026-06-02
alibaba/qwen3.7-plus

Multimodal Qwen workhorse for long-context agents, visual inputs, and coding

T Reasoning Tools Weight access not listed
Context
1M
Input
$0.5/M
Output
$3/M
MiniMax-M3by MiniMax
2026-06-01
minimax/MiniMax-M3

MiniMax multimodal model for long-context coding, perception, and agent planning

T Reasoning Tools Open weights
Context
512K
Input
$0.3/M
Output
$1.2/M
2026-05-30
xai/grok-imagine-video-1.5

Video model for image-to-video generation, editing, and extension workflows

T Weight access not listed
Context
1.024K
Input
-
Output
-
Step 3.7 Flashby StepFun
2026-05-29
stepfun/step-3.7-flash

Newer StepFun flash model for faster agents, coding, and multimodal prompts

T Reasoning Tools Open weights
Context
256K
Input
$0.185/M
Output
$1.11/M
Nano Banana 2by Google
2026-05-28
google/gemini-3.1-flash-image

Image model for prompt-driven generation, editing, and visual design workflows

T Reasoning Weight access not listed
Context
131.072K
Input
$0.5/M
Output
$60/M
2026-05-28
google/gemini-3-pro-image

Nano Banana Pro for higher-fidelity image generation and design-heavy edits

T Reasoning Weight access not listed
Context
65.536K
Input
$2/M
Output
$120/M
Claude Opus 4.8by Anthropic
2026-05-28
anthropic/claude-opus-4-8

Top Claude Opus tier for the hardest reasoning, coding, and long-horizon agents

T Reasoning Tools Weight access not listed
Context
1M
Input
$5/M
Output
$25/M
Qwen3.7 Maxby Alibaba Qwen
2026-05-21
alibaba/qwen3.7-max

Qwen frontier model tuned for agent frameworks, coding assistants, and long tasks

T Reasoning Tools Weight access not listed
Context
1M
Input
$2.5/M
Output
$7.5/M
2026-05-20
cohere/command-a-plus-05-2026

Cohere's stronger command model for multilingual agents and enterprise workflows

T Reasoning Tools Structured output Open weights
Context
128K
Input
$2.5/M
Output
$10/M
2026-05-19
google/gemini-3.5-flash

Fast Gemini model balancing multimodal reasoning, tool use, and cost

T Reasoning Tools Structured output Weight access not listed
Context
1.04858M
Input
$1.5/M
Output
$9/M
2026-05-19
google/gemini-flash-latest

Fast Gemini model balancing multimodal reasoning, tool use, and cost

T Reasoning Tools Structured output Weight access not listed
Context
1.04858M
Input
$1.5/M
Output
$9/M
2026-05-07
google/gemini-3.1-flash-lite

Low-latency Gemini model for high-volume multimodal and agent workloads

T Reasoning Tools Weight access not listed
Context
1.04858M
Input
$0.25/M
Output
$1.5/M
2026-05-07
google/gemini-flash-lite-latest

Low-latency Gemini model for high-volume multimodal and agent workloads

T Reasoning Tools Structured output Weight access not listed
Context
1.04858M
Input
$0.25/M
Output
$1.5/M
2026-05-07
openai/gpt-realtime-whisper

Streaming speech-to-text model for low-latency transcript deltas from live audio

Weight access not listed
Context
Not documented
Input
-
Output
-
2026-05-05
openai/gpt-5.5-instant

Compact GPT model for low-latency assistance and high-volume workloads

T Reasoning Tools Structured output Weight access not listed
Context
400K
Input
$5/M
Output
$30/M
2026-04-29
mistral/mistral-medium-latest

Balanced Mistral model for enterprise assistants, multilingual work, and tools

T Reasoning Tools Structured output Open weights
Context
262.144K
Input
$1.5/M
Output
$7.5/M