01-ai/Yi-6B-Chat
No provider description is available for this model yet.
- Context
- Not documented
- Input
- -
- Output
- -
Filter 986 source-linked models by creator, price, context, modality, and published benchmark coverage.
01-ai/Yi-6B-Chat
No provider description is available for this model yet.
01-ai/Yi-34B-200K
No provider description is available for this model yet.
01-ai/Yi-6B-200K
No provider description is available for this model yet.
01-ai/Yi-9B-200K
No provider description is available for this model yet.
01-ai/Yi-Coder-9B
No provider description is available for this model yet.
01-ai/Yi-Coder-9B-Chat
No provider description is available for this model yet.
01-ai/Yi-Coder-1.5B-Chat
No provider description is available for this model yet.
01-ai/Yi-1.5-6B-Chat
No provider description is available for this model yet.
01-ai/Yi-1.5-34B-Chat
No provider description is available for this model yet.
01-ai/Yi-VL-6B
No provider description is available for this model yet.
01-ai/Yi-VL-34B
No provider description is available for this model yet.
01-ai/Yi-1.5-9B-Chat-16K
No provider description is available for this model yet.
01-ai/Yi-1.5-9B-32K
No provider description is available for this model yet.
01-ai/Yi-1.5-34B-Chat-16K
No provider description is available for this model yet.
01-ai/Yi-1.5-34B-32K
No provider description is available for this model yet.
01-ai/Yi-1.5-6B
No provider description is available for this model yet.
01-ai/Yi-1.5-9B
No provider description is available for this model yet.
01-ai/Yi-1.5-9B-Chat
No provider description is available for this model yet.
01-ai/Yi-1.5-34B
No provider description is available for this model yet.
openrouter/auto-beta
Auto Router (Beta) is a task-aware router from OpenRouter. It classifies each request, then routes it the [most popular model](/rankings#task-spend) for that task based on aggregate spend, filtered by your...
deepseek-ai/DeepSeek-R1-Distill-Qwen-14B
No provider description is available for this model yet.
poolside/laguna-s-2.1:free
Laguna S 2.1 is the latest coding agent model from [Poolside](<https://poolside.ai/>). Laguna S 2.1 is a 118B total parameter model with 8B active parameters, scoring 70.2% on Terminal-Bench 2.1 and...
microsoft/Mage-Flow-Turbo
No provider description is available for this model yet.
microsoft/Mage-Flow-Base
No provider description is available for this model yet.
microsoft/Mage-Flow
No provider description is available for this model yet.
microsoft/Fara1.5-27B
No provider description is available for this model yet.
microsoft/Fara1.5-4B
No provider description is available for this model yet.
microsoft/Fara1.5-9B
No provider description is available for this model yet.
microsoft/MagenticBrain
No provider description is available for this model yet.
upstage/Solar-Open2-250B
No provider description is available for this model yet.
inclusionai/ling-3.0-flash:free
*Ling-3.0-flash* is a *124B-parameter Mixture-of-Experts (MoE) model*, with approximately *5.1B parameters activated per token*. The model is designed with *token efficiency and production-scale agentic inference* as key priorities, enabling developers...
nvidia/Qwen-Image-Flash
No provider description is available for this model yet.
nvidia/Cosmos3-Super-Text2Image-4Step
No provider description is available for this model yet.
ibm-granite/granite-guardian-3.2-8b-factuality-detection
No provider description is available for this model yet.
anthropic/claude-opus-5-fast
Fast-mode variant of [Opus 5](/anthropic/claude-opus-5) - identical capabilities with higher output speed at 2x pricing relative to regular Opus 5. Learn more in Anthropic's docs: https://platform.claude.com/docs/en/build-with-claude/fast-mode
qwen/qwen3.7-flash
Qwen3.7 Flash is a vision-language reasoning model from Alibaba. It is suited for multimodal agents, visual coding, search, and computer interaction, with strengths in object recognition, spatial understanding, and real-world...
google/gemini-3.6-flash:batch
Gemini 3.6 Flash is a high-efficiency model from Google for coding, agentic workflows, and web and app development. It is designed to produce polished outputs with fewer unnecessary edits and...
google/gemini-3.5-flash-lite:batch
Gemini 3.5 Flash Lite is a high-efficiency model from Google with upgraded agentic capabilities. It is suited for subagents that execute focused tasks within complex, multi-agent workflows.
anthropic/claude-sonnet-5:batch
Sonnet 5 is Anthropic's most capable Sonnet-class model, with frontier performance across coding, agents, and professional work. It supports adaptive thinking with selectable reasoning effort levels (low, medium, high, max,...
anthropic/claude-fable-5:batch
Claude Fable 5 is a Mythos-class model from Anthropic, built for autonomous knowledge work and coding. It supports text, image, and file inputs with text output, with reasoning support and...
minimax/minimax-m3:batch
MiniMax-M3 is a multimodal foundation model from MiniMax. It supports text, image, and video inputs with text output, a 1M-token context window, and is suited for long-horizon agentic work, coding,...
anthropic/claude-opus-4.8:batch
Claude Opus 4.8 is Anthropic's most capable generally available model in the Opus family. It supports text, image, and file inputs with text output, with reasoning support and a 1M-token...
google/gemini-3.5-flash:batch
Gemini 3.5 Flash is Google's high-efficiency multimodal model, bringing near-Pro level coding and reasoning at Flash-tier cost and speed. It is highly optimized for coding proficiency and parallel agentic execution...
google/gemini-3.1-flash-lite:batch
Gemini 3.1 Flash Lite is Google’s GA high-efficiency multimodal model optimized for low-latency, high-volume workloads. It supports text, image, video, audio, and PDF inputs, and is designed for lightweight agentic...
openai/gpt-5.5:batch
GPT-5.5 is OpenAI’s frontier model designed for complex professional workloads, building on GPT-5.4 with stronger reasoning, higher reliability, and improved token efficiency on hard tasks. It features a 1M+ token...
anthropic/claude-opus-4.7:batch
Opus 4.7 is the next generation of Anthropic's Opus family, built for long-running, asynchronous agents. Building on the coding and agentic strengths of Opus 4.6, it delivers stronger performance on...
openai/gpt-5.4-nano:batch
GPT-5.4 nano is the most lightweight and cost-efficient variant of the GPT-5.4 family, optimized for speed-critical and high-volume tasks. It supports text and image inputs and is designed for low-latency...
openai/gpt-5.4-mini:batch
GPT-5.4 mini brings the core capabilities of GPT-5.4 to a faster, more efficient model optimized for high-throughput workloads. It supports text and image inputs with strong performance across reasoning, coding,...
| Model | Creator | Input types | Context | Input / Output | Released | Compare |
|---|---|---|---|---|---|---|
| 01-ai/Yi-6B-Chat01-ai/Yi-6B-Chat | Not documented | - / - | Undated | |||
| 01-ai/Yi-34B-200K01-ai/Yi-34B-200K | Not documented | - / - | Undated | |||
| 01-ai/Yi-6B-200K01-ai/Yi-6B-200K | Not documented | - / - | Undated | |||
| 01-ai/Yi-9B-200K01-ai/Yi-9B-200K | Not documented | - / - | Undated | |||
| 01-ai/Yi-6B01-ai/Yi-6B | Not documented | - / - | Undated | |||
| 01-ai/Yi-34B01-ai/Yi-34B | Not documented | - / - | Undated | |||
| 01-ai/Yi-Coder-9B01-ai/Yi-Coder-9B | Not documented | - / - | Undated | |||
| 01-ai/Yi-Coder-9B-Chat01-ai/Yi-Coder-9B-Chat | Not documented | - / - | Undated | |||
| 01-ai/Yi-Coder-1.5B-Chat01-ai/Yi-Coder-1.5B-Chat | Not documented | - / - | Undated | |||
| 01-ai/Yi-1.5-6B-Chat01-ai/Yi-1.5-6B-Chat | Not documented | - / - | Undated | |||
| 01-ai/Yi-1.5-34B-Chat01-ai/Yi-1.5-34B-Chat | Not documented | - / - | Undated | |||
| 01-ai/Yi-VL-6B01-ai/Yi-VL-6B | Not documented | - / - | Undated | |||
| 01-ai/Yi-VL-34B01-ai/Yi-VL-34B | Not documented | - / - | Undated | |||
| 01-ai/Yi-1.5-9B-Chat-16K01-ai/Yi-1.5-9B-Chat-16K | Not documented | - / - | Undated | |||
| 01-ai/Yi-1.5-9B-32K01-ai/Yi-1.5-9B-32K | Not documented | - / - | Undated | |||
| 01-ai/Yi-1.5-34B-Chat-16K01-ai/Yi-1.5-34B-Chat-16K | Not documented | - / - | Undated | |||
| 01-ai/Yi-1.5-34B-32K01-ai/Yi-1.5-34B-32K | Not documented | - / - | Undated | |||
| 01-ai/Yi-1.5-6B01-ai/Yi-1.5-6B | Not documented | - / - | Undated | |||
| 01-ai/Yi-1.5-9B01-ai/Yi-1.5-9B | Not documented | - / - | Undated | |||
| 01-ai/Yi-1.5-9B-Chat01-ai/Yi-1.5-9B-Chat | Not documented | - / - | Undated | |||
| 01-ai/Yi-1.5-34B01-ai/Yi-1.5-34B | Not documented | - / - | Undated | |||
| Auto Router (Beta)openrouter/auto-beta | 2M | - / - | Undated | |||
| deepseek-ai/DeepSeek-R1-Distill-Qwen-14Bdeepseek-ai/DeepSeek-R1-Distill-Qwen-14B | Not documented | - / - | Undated | |||
| Poolside: Laguna S 2.1 (free)poolside/laguna-s-2.1:free | 262.144K | Free / Free | Undated | |||
| microsoft/Mage-Flow-Turbomicrosoft/Mage-Flow-Turbo | Not documented | - / - | Undated | |||
| microsoft/Mage-Flow-Basemicrosoft/Mage-Flow-Base | Not documented | - / - | Undated | |||
| microsoft/Mage-Flowmicrosoft/Mage-Flow | Not documented | - / - | Undated | |||
| microsoft/Fara1.5-27Bmicrosoft/Fara1.5-27B | Not documented | - / - | Undated | |||
| microsoft/Fara1.5-4Bmicrosoft/Fara1.5-4B | Not documented | - / - | Undated | |||
| microsoft/Fara1.5-9Bmicrosoft/Fara1.5-9B | Not documented | - / - | Undated | |||
| microsoft/MagenticBrainmicrosoft/MagenticBrain | Not documented | - / - | Undated | |||
| upstage/Solar-Open2-250Bupstage/Solar-Open2-250B | Not documented | - / - | Undated | |||
| Ling-3.0-flash (free)inclusionai/ling-3.0-flash:free | 262.144K | Free / Free | Undated | |||
| nvidia/Qwen-Image-Flashnvidia/Qwen-Image-Flash | Not documented | - / - | Undated | |||
| nvidia/Cosmos3-Super-Text2Image-4Stepnvidia/Cosmos3-Super-Text2Image-4Step | Not documented | - / - | Undated | |||
| ibm-granite/granite-guardian-3.2-8b-factuality-detectionibm-granite/granite-guardian-3.2-8b-factuality-detection | Not documented | - / - | Undated | |||
| Claude Opus 5 (Fast)anthropic/claude-opus-5-fast | 1M | $10 / $50 | Undated | |||
| Qwen: Qwen3.7 Flashqwen/qwen3.7-flash | 1M | $0.03 / $0.13 | Undated | |||
| Google: Gemini 3.6 Flash (batch)google/gemini-3.6-flash:batch | 1.04858M | $0.75 / $3.75 | Undated | |||
| Google: Gemini 3.5 Flash Lite (batch)google/gemini-3.5-flash-lite:batch | 1.04858M | $0.15 / $1.25 | Undated | |||
| Anthropic: Claude Sonnet 5 (batch)anthropic/claude-sonnet-5:batch | 1M | $1 / $5 | Undated | |||
| Anthropic: Claude Fable 5 (batch)anthropic/claude-fable-5:batch | 1M | $5 / $25 | Undated | |||
| MiniMax: MiniMax M3 (batch)minimax/minimax-m3:batch | 524.288K | $0.15 / $0.6 | Undated | |||
| Anthropic: Claude Opus 4.8 (batch)anthropic/claude-opus-4.8:batch | 1M | $2.5 / $12.5 | Undated | |||
| Google: Gemini 3.5 Flash (batch)google/gemini-3.5-flash:batch | 1.04858M | $0.75 / $4.5 | Undated | |||
| Google: Gemini 3.1 Flash Lite (batch)google/gemini-3.1-flash-lite:batch | 1.04858M | $0.125 / $0.75 | Undated | |||
| OpenAI: GPT-5.5 (batch)openai/gpt-5.5:batch | 1.05M | $2.5 / $15 | Undated | |||
| Anthropic: Claude Opus 4.7 (batch)anthropic/claude-opus-4.7:batch | 1M | $2.5 / $12.5 | Undated | |||
| OpenAI: GPT-5.4 Nano (batch)openai/gpt-5.4-nano:batch | 400K | $0.1 / $0.625 | Undated | |||
| OpenAI: GPT-5.4 Mini (batch)openai/gpt-5.4-mini:batch | 400K | $0.375 / $2.25 | Undated |