384 results Clear filters
Undated
anthropic/claude-sonnet-4.5

Claude Sonnet 4.5 is Anthropic’s most advanced Sonnet model to date, optimized for real-world agents and coding workflows. It delivers state-of-the-art performance on coding benchmarks such as SWE-bench Verified, with...

T Weight access not listed
Context
1M
Input
$3/M
Output
$15/M
Undated
qwen/qwen3-vl-235b-a22b-thinking

Qwen3-VL-235B-A22B Thinking is a multimodal model that unifies strong text generation with visual understanding across images and video. The Thinking model is optimized for multimodal reasoning in STEM and math....

T Weight access not listed
Context
131.072K
Input
$0.4/M
Output
$4/M
Undated
qwen/qwen3-vl-235b-a22b-instruct

Qwen3-VL-235B-A22B Instruct is an open-weight multimodal model that unifies strong text generation with visual understanding across images and video. The Instruct model targets general vision-language use (VQA, document parsing, chart/table...

T Weight access not listed
Context
131.072K
Input
$0.21/M
Output
$1.9/M
Undated
mistralai/mistral-medium-3.1

Mistral Medium 3.1 is an updated version of Mistral Medium 3, which is a high-performance enterprise-grade language model designed to deliver frontier-level capabilities at significantly reduced operational cost. It balances...

T Weight access not listed
Context
131.072K
Input
$0.4/M
Output
$2/M
Undated
z-ai/glm-4.5v

GLM-4.5V is a vision-language foundation model for multimodal agent applications. Built on a Mixture-of-Experts (MoE) architecture with 106B parameters and 12B activated parameters, it achieves state-of-the-art results in video understanding,...

T Weight access not listed
Context
65.536K
Input
$0.6/M
Output
$1.8/M
Undated
openai/gpt-5-chat

GPT-5 Chat is designed for advanced, natural, multimodal, and context-aware conversations for enterprise applications.

T Weight access not listed
Context
128K
Input
$1.25/M
Output
$10/M
Undated
anthropic/claude-opus-4.1

Claude Opus 4.1 is an updated version of Anthropic’s flagship model, offering improved performance in coding, reasoning, and agentic tasks. It achieves 74.5% on SWE-bench Verified and shows notable gains...

T Weight access not listed
Context
200K
Input
$15/M
Output
$75/M
Undated
bytedance/ui-tars-1.5-7b

UI-TARS-1.5 is a multimodal vision-language agent optimized for GUI-based environments, including desktop interfaces, web browsers, mobile systems, and games. Built by ByteDance, it builds upon the UI-TARS framework with reinforcement...

T Weight access not listed
Context
128K
Input
$0.1/M
Output
$0.2/M
baidu/ernie-4.5-vl-424b-a47b

ERNIE-4.5-VL-424B-A47B is a multimodal Mixture-of-Experts (MoE) model from Baidu’s ERNIE 4.5 series, featuring 424B total parameters with 47B active per token. It is trained jointly on text and image data...

T Weight access not listed
Context
123K
Input
$0.42/M
Output
$1.25/M
Undated
mistralai/mistral-small-3.2-24b-instruct

Mistral-Small-3.2-24B-Instruct-2506 is an updated 24B parameter model from Mistral optimized for instruction following, repetition reduction, and improved function calling. Compared to the 3.1 release, version 3.2 significantly improves accuracy on...

T Weight access not listed
Context
131.072K
Input
$0.1/M
Output
$0.3/M
google/gemini-2.5-pro-preview

Gemini 2.5 Pro is Google’s state-of-the-art AI model designed for advanced reasoning, coding, mathematics, and scientific tasks. It employs “thinking” capabilities, enabling it to reason through responses with enhanced accuracy...

T Weight access not listed
Context
1.04858M
Input
$1.25/M
Output
$10/M
Undated
anthropic/claude-opus-4

Claude Opus 4 is benchmarked as the world’s best coding model, at time of release, bringing sustained performance on complex, long-running tasks and agent workflows. It sets new benchmarks in...

T Weight access not listed
Context
200K
Input
$15/M
Output
$75/M
Undated
anthropic/claude-sonnet-4

Claude Sonnet 4 significantly enhances the capabilities of its predecessor, Sonnet 3.7, excelling in both coding and reasoning tasks with improved precision and controllability. Achieving state-of-the-art performance on SWE-bench (72.7%),...

T Weight access not listed
Context
200K
Input
$3/M
Output
$15/M
Undated
mistralai/mistral-medium-3

Mistral Medium 3 is a high-performance enterprise-grade language model designed to deliver frontier-level capabilities at significantly reduced operational cost. It balances state-of-the-art reasoning and multimodal performance with 8× lower cost...

T Weight access not listed
Context
131.072K
Input
$0.4/M
Output
$2/M
google/gemini-2.5-pro-preview-05-06

Gemini 2.5 Pro is Google’s state-of-the-art AI model designed for advanced reasoning, coding, mathematics, and scientific tasks. It employs “thinking” capabilities, enabling it to reason through responses with enhanced accuracy...

T Weight access not listed
Context
1.04858M
Input
$1.25/M
Output
$10/M
Undated
meta-llama/llama-guard-4-12b

Llama Guard 4 is a Llama 4 Scout-derived multimodal pretrained model, fine-tuned for content safety classification. Similar to previous versions, it can be used to classify content in both LLM...

T Weight access not listed
Context
163.84K
Input
$0.18/M
Output
$0.18/M
Undated
openai/o4-mini-high

OpenAI o4-mini-high is the same model as [o4-mini](/openai/o4-mini) with reasoning_effort set to high. OpenAI o4-mini is a compact reasoning model in the o-series, optimized for fast, cost-efficient performance while retaining...

T Weight access not listed
Context
200K
Input
$1.1/M
Output
$4.4/M
Undated
meta-llama/llama-4-maverick

Llama 4 Maverick 17B Instruct (128E) is a high-capacity multimodal language model from Meta, built on a mixture-of-experts (MoE) architecture with 128 experts and 17 billion active parameters per forward...

T Weight access not listed
Context
1.04858M
Input
$0.2/M
Output
$0.8/M
Undated
meta-llama/llama-4-scout

Llama 4 Scout 17B Instruct (16E) is a mixture-of-experts (MoE) language model developed by Meta, activating 17 billion parameters out of a total of 109B. It supports native multimodal input...

T Weight access not listed
Context
327.68K
Input
$0.1/M
Output
$0.3/M
Undated
mistralai/mistral-small-3.1-24b-instruct

Mistral Small 3.1 24B Instruct is an upgraded variant of Mistral Small 3 (2501), featuring 24 billion parameters with advanced multimodal capabilities. It provides state-of-the-art performance in text-based reasoning and...

T Weight access not listed
Context
128K
Input
$0.351/M
Output
$0.555/M
Undated
google/gemma-3-4b-it

Gemma 3 introduces multimodality, supporting vision-language input and text outputs. It handles context windows up to 128k tokens, understands over 140 languages, and offers improved math, reasoning, and chat capabilities,...

T Weight access not listed
Context
131.072K
Input
$0.05/M
Output
$0.1/M
Undated
google/gemma-3-12b-it

Gemma 3 introduces multimodality, supporting vision-language input and text outputs. It handles context windows up to 128k tokens, understands over 140 languages, and offers improved math, reasoning, and chat capabilities,...

T Weight access not listed
Context
131.072K
Input
$0.05/M
Output
$0.15/M
Undated
google/gemma-3-27b-it

Gemma 3 introduces multimodality, supporting vision-language input and text outputs. It handles context windows up to 128k tokens, understands over 140 languages, and offers improved math, reasoning, and chat capabilities,...

T Weight access not listed
Context
131.072K
Input
$0.08/M
Output
$0.45/M
Undated
qwen/qwen2.5-vl-72b-instruct

Qwen2.5-VL is proficient in recognizing common objects such as flowers, birds, fish, and insects. It is also highly capable of analyzing texts, charts, icons, graphics, and layouts within images.

T Weight access not listed
Context
128K
Input
$0.8/M
Output
$1/M
Undated
minimax/minimax-01

MiniMax-01 is a combines MiniMax-Text-01 for text generation and MiniMax-VL-01 for image understanding. It has 456 billion parameters, with 45.9 billion parameters activated per inference, and can handle a context...

T Weight access not listed
Context
1.00019M
Input
$0.2/M
Output
$1.1/M
Undated
amazon/nova-lite-v1

Amazon Nova Lite 1.0 is a very low-cost multimodal model from Amazon that focused on fast processing of image, video, and text inputs to generate text output. Amazon Nova Lite...

T Weight access not listed
Context
300K
Input
$0.06/M
Output
$0.24/M
Undated
amazon/nova-pro-v1

Amazon Nova Pro 1.0 is a capable multimodal model from Amazon focused on providing a combination of accuracy, speed, and cost for a wide range of tasks. As of December...

T Weight access not listed
Context
300K
Input
$0.8/M
Output
$3.2/M
meta-llama/llama-3.2-11b-vision-instruct

Llama 3.2 11B Vision is a multimodal model with 11 billion parameters, designed to handle tasks combining visual and textual data. It excels in tasks such as image captioning and...

T Weight access not listed
Context
131.072K
Input
$0.345/M
Output
$0.345/M
openai/gpt-4o-mini-2024-07-18

GPT-4o mini is OpenAI's newest model after [GPT-4 Omni](/models/openai/gpt-4o), supporting both text and image inputs with text outputs. As their most advanced small model, it is many multiples more affordable...

T Weight access not listed
Context
128K
Input
$0.15/M
Output
$0.6/M
Undated
anthropic/claude-3-haiku

Claude 3 Haiku is Anthropic's fastest and most compact model for near-instant responsiveness. Quick and accurate targeted performance. See the launch announcement and benchmark results [here](https://www.anthropic.com/news/claude-3-haiku) #multimodal

T Weight access not listed
Context
200K
Input
$0.25/M
Output
$1.25/M
Auto Routerby Openrouter
Undated
openrouter/auto

Your prompt will be processed by a meta-model and routed to one of dozens of models (see below), optimizing for the best possible output. To see which model was used,...

T Weight access not listed
Context
2M
Input
-
Output
-
meta-llama/Llama-4-Maverick-17B-128E-Instruct

No provider description is available for this model yet.

Weight access not listed
Context
Not documented
Input
-
Output
-
meta-llama/Llama-4-Scout-17B-16E-Instruct

No provider description is available for this model yet.

Weight access not listed
Context
Not documented
Input
-
Output
-
meta-llama/Llama-Guard-4-12B

No provider description is available for this model yet.

Weight access not listed
Context
Not documented
Input
-
Output
-
meta-llama/Llama-4-Maverick-17B-128E

No provider description is available for this model yet.

Weight access not listed
Context
Not documented
Input
-
Output
-
meta-llama/Llama-4-Scout-17B-16E

No provider description is available for this model yet.

Weight access not listed
Context
Not documented
Input
-
Output
-
meta-llama/Llama-3.2-90B-Vision-Instruct

No provider description is available for this model yet.

Open weights
Context
Not documented
Input
-
Output
-
meta-llama/Llama-3.2-11B-Vision-Instruct

No provider description is available for this model yet.

Open weights
Context
Not documented
Input
-
Output
-
meta-llama/Llama-Guard-3-11B-Vision

No provider description is available for this model yet.

Open weights
Context
Not documented
Input
-
Output
-
meta-llama/Llama-3.2-90B-Vision

No provider description is available for this model yet.

Open weights
Context
Not documented
Input
-
Output
-
meta-llama/Llama-3.2-11B-Vision

No provider description is available for this model yet.

Open weights
Context
Not documented
Input
-
Output
-
Undated
google/gemma-4-26B-A4B

No provider description is available for this model yet.

Open weights
Context
Not documented
Input
-
Output
-
Undated
google/gemma-4-26B-A4B-it

No provider description is available for this model yet.

Open weights
Context
Not documented
Input
-
Output
-
Undated
google/gemma-4-31B

No provider description is available for this model yet.

Open weights
Context
Not documented
Input
-
Output
-
Undated
google/gemma-4-31B-it

No provider description is available for this model yet.

Open weights
Context
Not documented
Input
-
Output
-
google/diffusiongemma-26B-A4B-it

No provider description is available for this model yet.

Open weights
Context
Not documented
Input
-
Output
-
Undated
google/medgemma-1.5-4b-it

No provider description is available for this model yet.

Weight access not listed
Context
Not documented
Input
-
Output
-
Undated
google/translategemma-4b-it

No provider description is available for this model yet.

Open weights
Context
Not documented
Input
-
Output
-
Undated
google/translategemma-12b-it

No provider description is available for this model yet.

Open weights
Context
Not documented
Input
-
Output
-
Undated
google/translategemma-27b-it

No provider description is available for this model yet.

Open weights
Context
Not documented
Input
-
Output
-