986 results
Undated
01-ai/Yi-6B-Chat

No provider description is available for this model yet.

Open weights
Context
Not documented
Input
-
Output
-
Undated
01-ai/Yi-34B-200K

No provider description is available for this model yet.

Open weights
Context
Not documented
Input
-
Output
-
Undated
01-ai/Yi-6B-200K

No provider description is available for this model yet.

Open weights
Context
Not documented
Input
-
Output
-
Undated
01-ai/Yi-9B-200K

No provider description is available for this model yet.

Open weights
Context
Not documented
Input
-
Output
-
01-ai/Yi-6Bby 01.AI
Undated
01-ai/Yi-6B

No provider description is available for this model yet.

Open weights
Context
Not documented
Input
-
Output
-
01-ai/Yi-34Bby 01.AI
Undated
01-ai/Yi-34B

No provider description is available for this model yet.

Open weights
Context
Not documented
Input
-
Output
-
Undated
01-ai/Yi-Coder-9B

No provider description is available for this model yet.

Open weights
Context
Not documented
Input
-
Output
-
Undated
01-ai/Yi-Coder-9B-Chat

No provider description is available for this model yet.

Open weights
Context
Not documented
Input
-
Output
-
Undated
01-ai/Yi-Coder-1.5B-Chat

No provider description is available for this model yet.

Open weights
Context
Not documented
Input
-
Output
-
Undated
01-ai/Yi-1.5-6B-Chat

No provider description is available for this model yet.

Open weights
Context
Not documented
Input
-
Output
-
Undated
01-ai/Yi-1.5-34B-Chat

No provider description is available for this model yet.

Open weights
Context
Not documented
Input
-
Output
-
Undated
01-ai/Yi-VL-6B

No provider description is available for this model yet.

Open weights
Context
Not documented
Input
-
Output
-
Undated
01-ai/Yi-VL-34B

No provider description is available for this model yet.

Open weights
Context
Not documented
Input
-
Output
-
Undated
01-ai/Yi-1.5-9B-Chat-16K

No provider description is available for this model yet.

Open weights
Context
Not documented
Input
-
Output
-
Undated
01-ai/Yi-1.5-9B-32K

No provider description is available for this model yet.

Open weights
Context
Not documented
Input
-
Output
-
Undated
01-ai/Yi-1.5-34B-Chat-16K

No provider description is available for this model yet.

Open weights
Context
Not documented
Input
-
Output
-
Undated
01-ai/Yi-1.5-34B-32K

No provider description is available for this model yet.

Open weights
Context
Not documented
Input
-
Output
-
Undated
01-ai/Yi-1.5-6B

No provider description is available for this model yet.

Open weights
Context
Not documented
Input
-
Output
-
Undated
01-ai/Yi-1.5-9B

No provider description is available for this model yet.

Open weights
Context
Not documented
Input
-
Output
-
Undated
01-ai/Yi-1.5-9B-Chat

No provider description is available for this model yet.

Open weights
Context
Not documented
Input
-
Output
-
Undated
01-ai/Yi-1.5-34B

No provider description is available for this model yet.

Open weights
Context
Not documented
Input
-
Output
-
Auto Router (Beta)by Openrouter
Undated
openrouter/auto-beta

Auto Router (Beta) is a task-aware router from OpenRouter. It classifies each request, then routes it the [most popular model](/rankings#task-spend) for that task based on aggregate spend, filtered by your...

T Weight access not listed
Context
2M
Input
-
Output
-
deepseek-ai/DeepSeek-R1-Distill-Qwen-14B

No provider description is available for this model yet.

Open weights
Context
Not documented
Input
-
Output
-
Undated
poolside/laguna-s-2.1:free

Laguna S 2.1 is the latest coding agent model from [Poolside](<https://poolside.ai/>). Laguna S 2.1 is a 118B total parameter model with 8B active parameters, scoring 70.2% on Terminal-Bench 2.1 and...

T Weight access not listed
Context
262.144K
Input
Free
Output
Free
Undated
microsoft/Mage-Flow-Turbo

No provider description is available for this model yet.

Open weights
Context
Not documented
Input
-
Output
-
Undated
microsoft/Mage-Flow-Base

No provider description is available for this model yet.

Open weights
Context
Not documented
Input
-
Output
-
Undated
microsoft/Mage-Flow

No provider description is available for this model yet.

Open weights
Context
Not documented
Input
-
Output
-
Undated
microsoft/Fara1.5-27B

No provider description is available for this model yet.

Open weights
Context
Not documented
Input
-
Output
-
Undated
microsoft/Fara1.5-4B

No provider description is available for this model yet.

Open weights
Context
Not documented
Input
-
Output
-
Undated
microsoft/Fara1.5-9B

No provider description is available for this model yet.

Open weights
Context
Not documented
Input
-
Output
-
Undated
microsoft/MagenticBrain

No provider description is available for this model yet.

Open weights
Context
Not documented
Input
-
Output
-
Undated
upstage/Solar-Open2-250B

No provider description is available for this model yet.

Weight access not listed
Context
Not documented
Input
-
Output
-
Ling-3.0-flash (free)by Inclusionai
Undated
inclusionai/ling-3.0-flash:free

*Ling-3.0-flash* is a *124B-parameter Mixture-of-Experts (MoE) model*, with approximately *5.1B parameters activated per token*. The model is designed with *token efficiency and production-scale agentic inference* as key priorities, enabling developers...

T Weight access not listed
Context
262.144K
Input
Free
Output
Free
Undated
nvidia/Qwen-Image-Flash

No provider description is available for this model yet.

Weight access not listed
Context
Not documented
Input
-
Output
-
nvidia/Cosmos3-Super-Text2Image-4Step

No provider description is available for this model yet.

Weight access not listed
Context
Not documented
Input
-
Output
-
Undated
anthropic/claude-opus-5-fast

Fast-mode variant of [Opus 5](/anthropic/claude-opus-5) - identical capabilities with higher output speed at 2x pricing relative to regular Opus 5. Learn more in Anthropic's docs: https://platform.claude.com/docs/en/build-with-claude/fast-mode

T Weight access not listed
Context
1M
Input
$10/M
Output
$50/M
Qwen: Qwen3.7 Flashby Alibaba Qwen
Undated
qwen/qwen3.7-flash

Qwen3.7 Flash is a vision-language reasoning model from Alibaba. It is suited for multimodal agents, visual coding, search, and computer interaction, with strengths in object recognition, spatial understanding, and real-world...

T Weight access not listed
Context
1M
Input
$0.03/M
Output
$0.13/M
google/gemini-3.6-flash:batch

Gemini 3.6 Flash is a high-efficiency model from Google for coding, agentic workflows, and web and app development. It is designed to produce polished outputs with fewer unnecessary edits and...

T Weight access not listed
Context
1.04858M
Input
$0.75/M
Output
$3.75/M
google/gemini-3.5-flash-lite:batch

Gemini 3.5 Flash Lite is a high-efficiency model from Google with upgraded agentic capabilities. It is suited for subagents that execute focused tasks within complex, multi-agent workflows.

T Weight access not listed
Context
1.04858M
Input
$0.15/M
Output
$1.25/M
anthropic/claude-sonnet-5:batch

Sonnet 5 is Anthropic's most capable Sonnet-class model, with frontier performance across coding, agents, and professional work. It supports adaptive thinking with selectable reasoning effort levels (low, medium, high, max,...

T Weight access not listed
Context
1M
Input
$1/M
Output
$5/M
Undated
anthropic/claude-fable-5:batch

Claude Fable 5 is a Mythos-class model from Anthropic, built for autonomous knowledge work and coding. It supports text, image, and file inputs with text output, with reasoning support and...

T Weight access not listed
Context
1M
Input
$5/M
Output
$25/M
Undated
minimax/minimax-m3:batch

MiniMax-M3 is a multimodal foundation model from MiniMax. It supports text, image, and video inputs with text output, a 1M-token context window, and is suited for long-horizon agentic work, coding,...

T Weight access not listed
Context
524.288K
Input
$0.15/M
Output
$0.6/M
anthropic/claude-opus-4.8:batch

Claude Opus 4.8 is Anthropic's most capable generally available model in the Opus family. It supports text, image, and file inputs with text output, with reasoning support and a 1M-token...

T Weight access not listed
Context
1M
Input
$2.5/M
Output
$12.5/M
google/gemini-3.5-flash:batch

Gemini 3.5 Flash is Google's high-efficiency multimodal model, bringing near-Pro level coding and reasoning at Flash-tier cost and speed. It is highly optimized for coding proficiency and parallel agentic execution...

T Weight access not listed
Context
1.04858M
Input
$0.75/M
Output
$4.5/M
google/gemini-3.1-flash-lite:batch

Gemini 3.1 Flash Lite is Google’s GA high-efficiency multimodal model optimized for low-latency, high-volume workloads. It supports text, image, video, audio, and PDF inputs, and is designed for lightweight agentic...

T Weight access not listed
Context
1.04858M
Input
$0.125/M
Output
$0.75/M
Undated
openai/gpt-5.5:batch

GPT-5.5 is OpenAI’s frontier model designed for complex professional workloads, building on GPT-5.4 with stronger reasoning, higher reliability, and improved token efficiency on hard tasks. It features a 1M+ token...

T Weight access not listed
Context
1.05M
Input
$2.5/M
Output
$15/M
anthropic/claude-opus-4.7:batch

Opus 4.7 is the next generation of Anthropic's Opus family, built for long-running, asynchronous agents. Building on the coding and agentic strengths of Opus 4.6, it delivers stronger performance on...

T Weight access not listed
Context
1M
Input
$2.5/M
Output
$12.5/M
Undated
openai/gpt-5.4-nano:batch

GPT-5.4 nano is the most lightweight and cost-efficient variant of the GPT-5.4 family, optimized for speed-critical and high-volume tasks. It supports text and image inputs and is designed for low-latency...

T Weight access not listed
Context
400K
Input
$0.1/M
Output
$0.625/M
Undated
openai/gpt-5.4-mini:batch

GPT-5.4 mini brings the core capabilities of GPT-5.4 to a faster, more efficient model optimized for high-throughput workloads. It supports text and image inputs with strong performance across reasoning, coding,...

T Weight access not listed
Context
400K
Input
$0.375/M
Output
$2.25/M