384 results Clear filters
Grok 4.3by xAI
2026-04-17
xai/grok-4.3

xAI's default Grok for chat, coding, agentic tools, and lower hallucination risk

T Reasoning Tools Structured output Weight access not listed
Context
1M
Input
$1.25/M
Output
$2.5/M
Qwen3.6 35B-A3Bby Alibaba Qwen
2026-04-17
alibaba/qwen3.6-35b-a3b

Open multimodal Qwen MoE for local agents that need vision, audio, and code

T Reasoning Tools Structured output Open weights
Context
262.144K
Input
$0.248/M
Output
$1.485/M
2026-04-16
xai/grok-build-0.1

Fast Grok coding model tuned for agentic engineering and iterative edits

T Reasoning Tools Structured output Weight access not listed
Context
256K
Input
$1/M
Output
$2/M
Claude Opus 4.7by Anthropic
2026-04-16
anthropic/claude-opus-4-7

Stronger Opus tier for advanced software work and high-stakes reasoning

T Reasoning Tools Weight access not listed
Context
1M
Input
$5/M
Output
$25/M
2026-04-14
google/gemini-robotics-er-1.6-preview

Vision-language model for embodied reasoning: spatial understanding, task planning, and physical-world agentic robotics

T Reasoning Tools Structured output Weight access not listed
Context
131.072K
Input
$1/M
Output
$5/M
2026-04-08
meta/muse-spark-1.1

Muse Spark is a natively multimodal reasoning model with support for tool-use, visual chain of thought, and multi-agent orchestration.

T Reasoning Tools Structured output Weight access not listed
Context
1M
Input
$1.25/M
Output
$4.25/M
2026-04-02
google/gemma-4-26b-a4b-it

Open Gemma instruction model for efficient chat and self-hosted deployments

T Reasoning Tools Structured output Open weights
Context
262.144K
Input
$0.06/M
Output
$0.33/M
2026-04-02
google/gemma-4-31b-it

Largest Gemma 4 instruction model for open, self-hosted chat and reasoning

T Reasoning Tools Open weights
Context
262.144K
Input
$0.1/M
Output
$0.35/M
2026-04-02
google/gemma-4-E4B-it

Open Gemma instruction model for efficient chat and self-hosted deployments

T Reasoning Tools Structured output Open weights
Context
131.072K
Input
$0.2/M
Output
$0.2/M
2026-04-02
google/gemma-4-E2B-it

Open Gemma instruction model for efficient chat and self-hosted deployments

T Reasoning Tools Structured output Open weights
Context
131.072K
Input
$0.1/M
Output
$0.1/M
Qwen3.6 Plusby Alibaba Qwen
2026-04-02
alibaba/qwen3.6-plus

Earlier Qwen multimodal workhorse for million-token agent and document tasks

T Reasoning Tools Weight access not listed
Context
1M
Input
$0.5/M
Output
$3/M
GLM-5V-Turboby Zhipu AI
2026-04-01
zhipuai/glm-5v-turbo

Fast GLM vision model for screenshots, documents, and multimodal agent tasks

T Reasoning Tools Weight access not listed
Context
200K
Input
$5/M
Output
$22/M
2026-03-31
nvidia/llama-nemotron-rerank-vl-1b-v2

Reranking model for improving retrieval quality in search and recommendation systems

T Open weights
Context
128K
Input
-
Output
-
2026-03-31
google/veo-3.1-lite-generate-preview

Video model for prompt-guided generation, editing, and motion workflows

T Weight access not listed
Context
1.024K
Input
-
Output
-
2026-03-26
google/gemini-3.1-flash-live-preview

High-quality, low-latency Live API model for real-time dialogue and voice-first AI applications

T Reasoning Tools Weight access not listed
Context
131.072K
Input
$0.75/M
Output
$4.5/M
2026-03-25
google/lyria-3-pro-preview

Music generation model for full-length songs from text or images with vocals and structure

T Weight access not listed
Context
131.072K
Input
-
Output
-
2026-03-25
google/lyria-3-clip-preview

Music generation model for short 30-second clips, loops, and previews from text or image prompts

T Weight access not listed
Context
131.072K
Input
-
Output
-
MiMo-V2-Omniby Xiaomi
2026-03-18
xiaomi/mimo-v2-omni

MiMo omni model for text, image, video, audio, and agents

T Reasoning Tools Weight access not listed
Context
262.144K
Input
$0.14/M
Output
$0.28/M
GPT-5.4 nanoby OpenAI
2026-03-17
openai/gpt-5.4-nano

Cheapest GPT-5.4 lane for simple routing, extraction, and bulk automation

T Reasoning Tools Structured output Weight access not listed
Context
400K
Input
$0.2/M
Output
$1.25/M
GPT-5.4 miniby OpenAI
2026-03-17
openai/gpt-5.4-mini

Strong small GPT for coding subagents, quick tool use, and high-volume work

T Reasoning Tools Structured output Weight access not listed
Context
400K
Input
$0.75/M
Output
$4.5/M
2026-03-16
mistral/mistral-small-latest

Efficient Mistral model for fast chat, extraction, and production assistants

T Reasoning Tools Open weights
Context
256K
Input
$0.15/M
Output
$0.6/M
Mistral Small 4by Mistral AI
2026-03-16
mistral/mistral-small-2603

Fast Mistral production model for chat, extraction, and cost-sensitive agents

T Reasoning Tools Open weights
Context
256K
Input
$0.15/M
Output
$0.6/M
2026-03-09
xai/grok-4.20-0309-non-reasoning

Grok model for agentic tool use, reasoning, coding, and live assistance

T Tools Structured output Weight access not listed
Context
1M
Input
$1.25/M
Output
$2.5/M
2026-03-09
xai/grok-4.20-0309-reasoning

Reasoning Grok for document-heavy analysis and long-horizon tool use

T Reasoning Tools Structured output Weight access not listed
Context
1M
Input
$1.25/M
Output
$2.5/M
GPT-5.4 Proby OpenAI
2026-03-05
openai/gpt-5.4-pro

More exact GPT-5.4 tier for demanding professional reasoning and agent tasks

T Reasoning Tools Weight access not listed
Context
1.05M
Input
$30/M
Output
$180/M
GPT-5.4by OpenAI
2026-03-05
openai/gpt-5.4

Agent-ready GPT for coding and computer-use workflows at a lower cost

T Reasoning Tools Structured output Weight access not listed
Context
1.05M
Input
$2.5/M
Output
$15/M
2026-03-03
google/gemini-3.1-flash-lite-preview

Low-latency Gemini model for high-volume multimodal and agent workloads

T Reasoning Tools Structured output Weight access not listed
Context
1.04858M
Input
$0.25/M
Output
$1.5/M
2026-03-03
openai/gpt-5.3-chat-latest

Chat-tuned GPT model for conversational assistance, writing, and tool workflows

T Tools Structured output Weight access not listed
Context
128K
Input
$1.75/M
Output
$14/M
Nano Banana 2by Google
2026-02-26
google/gemini-3.1-flash-image-preview

Image model for prompt-driven generation, editing, and visual design workflows

T Reasoning Weight access not listed
Context
65.536K
Input
$0.5/M
Output
$3/M
Qwen3.5 27Bby Alibaba Qwen
2026-02-23
alibaba/qwen3.5-27b

Qwen vision-language model for visual reasoning, documents, and agent tasks

T Tools Structured output Open weights
Context
262.144K
Input
$0.3/M
Output
$2.4/M
Qwen3.5 35B-A3Bby Alibaba Qwen
2026-02-23
alibaba/qwen3.5-35b-a3b

Qwen vision-language model for visual reasoning, documents, and agent tasks

T Tools Structured output Open weights
Context
262.144K
Input
$0.25/M
Output
$2/M
Qwen3.5 122B-A10Bby Alibaba Qwen
2026-02-23
alibaba/qwen3.5-122b-a10b

Qwen vision-language model for visual reasoning, documents, and agent tasks

T Reasoning Tools Structured output Open weights
Context
262.144K
Input
$0.4/M
Output
$3.2/M
google/gemini-3.1-pro-preview-customtools

Advanced Gemini model for complex reasoning, coding, and multimodal analysis

T Reasoning Tools Structured output Weight access not listed
Context
1.04858M
Input
$2/M
Output
$12/M
2026-02-19
google/gemini-3.1-pro-preview

Reasoning-first Gemini preview for agentic coding and complex problem solving

T Reasoning Tools Structured output Weight access not listed
Context
1.04858M
Input
$2/M
Output
$12/M
Claude Sonnet 4.6by Anthropic
2026-02-17
anthropic/claude-sonnet-4-6

Claude workhorse for coding agents, careful analysis, and production cost control

T Reasoning Tools Weight access not listed
Context
1M
Input
$3/M
Output
$15/M
Qwen3.5 Plusby Alibaba Qwen
2026-02-16
alibaba/qwen3.5-plus

Qwen vision-language model for visual reasoning, documents, and agent tasks

T Reasoning Tools Weight access not listed
Context
1M
Input
$0.4/M
Output
$2.4/M
Qwen3.5 397B-A17Bby Alibaba Qwen
2026-02-15
alibaba/qwen3.5-397b-a17b

Large open Qwen multimodal MoE for visual agents and long technical tasks

T Reasoning Tools Structured output Open weights
Context
262.144K
Input
$0.6/M
Output
$3.6/M
2026-02-10
nvidia/llama-nemotron-embed-vl-1b-v2

Embedding model for semantic search, retrieval, clustering, and ranking pipelines

T Open weights
Context
32.768K
Input
-
Output
-
GPT-5.3 Codexby OpenAI
2026-02-05
openai/gpt-5.3-codex

Coding-optimized GPT model for repository edits, reviews, and agentic software work

T Reasoning Tools Weight access not listed
Context
400K
Input
$1.75/M
Output
$14/M
Claude Opus 4.6by Anthropic
2026-02-05
anthropic/claude-opus-4-6

High-end Claude for difficult coding, planning, and slower expert reasoning

T Reasoning Tools Weight access not listed
Context
1M
Input
$5/M
Output
$25/M
Kimi K2.5by Moonshot AI
2026-01
moonshotai/kimi-k2.5

Earlier Kimi frontier model for long-context agents, coding, and multimodal work

T Reasoning Tools Structured output Open weights
Context
262.144K
Input
$0.6/M
Output
$3/M
2025-12-17
google/gemini-3-flash-preview

New Gemini flash lane bringing frontier-style multimodal reasoning to cheaper runs

T Reasoning Tools Structured output Weight access not listed
Context
1.04858M
Input
$0.5/M
Output
$3/M
GPT-5.2 Codexby OpenAI
2025-12-11
openai/gpt-5.2-codex

Code-specialist GPT for repository edits, reviews, and long-running software agents

T Reasoning Tools Weight access not listed
Context
400K
Input
$0.14/M
Output
$1.14/M
GPT-5.2 Proby OpenAI
2025-12-11
openai/gpt-5.2-pro

Higher-accuracy GPT-5.2 variant for tougher reasoning and review workflows

T Reasoning Tools Weight access not listed
Context
400K
Input
$21/M
Output
$168/M
GPT-5.2by OpenAI
2025-12-11
openai/gpt-5.2

Reliable GPT generation for broad coding, writing, and tool-assisted product work

T Reasoning Tools Structured output Weight access not listed
Context
400K
Input
$1.75/M
Output
$14/M
GPT-5.2 Chatby OpenAI
2025-12-11
openai/gpt-5.2-chat-latest

Chat-tuned GPT model for conversational assistance, writing, and tool workflows

T Reasoning Tools Structured output Weight access not listed
Context
128K
Input
$1.75/M
Output
$14/M
GLM-4.6Vby Zhipu AI
2025-12-08
zhipuai/glm-4.6v

GLM vision model for visual reasoning, documents, and multimodal agents

T Reasoning Tools Open weights
Context
128K
Input
$0.3/M
Output
$0.9/M
GPT-Image-1.5by OpenAI
2025-11-25
openai/gpt-image-1.5

Image model for prompt-driven generation, editing, and visual design workflows

T Weight access not listed
Context
Not documented
Input
$5/M
Output
$32/M
2025-11-24
anthropic/claude-opus-4-5

Flagship Claude model for deep reasoning, coding, and long-horizon agents

T Reasoning Tools Weight access not listed
Context
200K
Input
$5/M
Output
$25/M
2025-11-20
google/gemini-3-pro-image-preview

Nano Banana Pro for higher-fidelity image generation and design-heavy edits

T Reasoning Weight access not listed
Context
65.536K
Input
$2/M
Output
$12/M