978 results
Undated
thedrummer/unslopnemo-12b

UnslopNemo v4.1 is the latest addition from the creator of Rocinante, designed for adventure writing and role-play scenarios.

T Weight access not listed
Context
32.768K
Input
$0.4/M
Output
$0.4/M
Magnum v4 72Bby Anthracite Org
Undated
anthracite-org/magnum-v4-72b

This is a series of models designed to replicate the prose quality of the Claude 3 models, specifically Sonnet(https://openrouter.ai/anthropic/claude-3.5-sonnet) and Opus(https://openrouter.ai/anthropic/claude-3-opus). The model is fine-tuned on top of [Qwen2.5 72B](https://openrouter.ai/qwen/qwen-2.5-72b-instruct).

T Weight access not listed
Context
16.384K
Input
$3/M
Output
$5/M
Undated
qwen/qwen-2.5-7b-instruct

Qwen2.5 7B is the latest series of Qwen large language models. Qwen2.5 brings the following improvements upon Qwen2: - Significantly more knowledge and has greatly improved capabilities in coding and...

T Weight access not listed
Context
32.768K
Input
$0.1/M
Output
$0.2/M
inflection/inflection-3-productivity

Inflection 3 Productivity is optimized for following instructions. It is better for tasks requiring JSON output or precise adherence to provided guidelines. It has access to recent news. For emotional...

T Weight access not listed
Context
8K
Input
$2.5/M
Output
$10/M
Undated
inflection/inflection-3-pi

Inflection 3 Pi powers Inflection's [Pi](https://pi.ai) chatbot, including backstory, emotional intelligence, productivity, and safety. It has access to recent news, and excels in scenarios like customer support and roleplay. Pi...

T Weight access not listed
Context
8K
Input
$2.5/M
Output
$10/M
Undated
thedrummer/rocinante-12b

Rocinante 12B is designed for engaging storytelling and rich prose. Early testers have reported: - Expanded vocabulary with unique and expressive word choices - Enhanced creativity for vivid narratives -...

T Weight access not listed
Context
65.536K
Input
$0.25/M
Output
$0.5/M
meta-llama/llama-3.2-1b-instruct

Llama 3.2 1B is a 1-billion-parameter language model focused on efficiently performing natural language tasks, such as summarization, dialogue, and multilingual text analysis. Its smaller size allows it to operate...

T Weight access not listed
Context
60K
Input
$0.027/M
Output
$0.201/M
meta-llama/llama-3.2-11b-vision-instruct

Llama 3.2 11B Vision is a multimodal model with 11 billion parameters, designed to handle tasks combining visual and textual data. It excels in tasks such as image captioning and...

T Weight access not listed
Context
131.072K
Input
$0.345/M
Output
$0.345/M
meta-llama/llama-3.2-3b-instruct:free

Llama 3.2 3B is a 3-billion-parameter multilingual large language model, optimized for advanced natural language processing tasks like dialogue generation, reasoning, and summarization. Designed with the latest transformer architecture, it...

T Weight access not listed
Context
131.072K
Input
Free
Output
Free
meta-llama/llama-3.2-3b-instruct

Llama 3.2 3B is a 3-billion-parameter multilingual large language model, optimized for advanced natural language processing tasks like dialogue generation, reasoning, and summarization. Designed with the latest transformer architecture, it...

T Weight access not listed
Context
131.072K
Input
$0.05/M
Output
$0.33/M
Qwen2.5 72B Instructby Alibaba Qwen
Undated
qwen/qwen-2.5-72b-instruct

Qwen2.5 72B is the latest series of Qwen large language models. Qwen2.5 brings the following improvements upon Qwen2: - Significantly more knowledge and has greatly improved capabilities in coding and...

T Weight access not listed
Context
32.768K
Input
$0.36/M
Output
$0.4/M
sao10k/l3.1-euryale-70b

Euryale L3.1 70B v2.2 is a model focused on creative roleplay from [Sao10k](https://ko-fi.com/sao10k). It is the successor of [Euryale L3 70B v2.1](/models/sao10k/l3-euryale-70b).

T Weight access not listed
Context
131.072K
Input
$0.85/M
Output
$0.85/M
Undated
nousresearch/hermes-3-llama-3.1-70b

Hermes 3 is a generalist language model with many improvements over [Hermes 2](/models/nousresearch/nous-hermes-2-mistral-7b-dpo), including advanced agentic capabilities, much better roleplaying, reasoning, multi-turn conversation, long context coherence, and improvements across the...

T Weight access not listed
Context
131.072K
Input
$0.7/M
Output
$0.7/M
Undated
nousresearch/hermes-3-llama-3.1-405b:free

Hermes 3 is a generalist language model with many improvements over Hermes 2, including advanced agentic capabilities, much better roleplaying, reasoning, multi-turn conversation, long context coherence, and improvements across the...

T Weight access not listed
Context
131.072K
Input
Free
Output
Free
Undated
nousresearch/hermes-3-llama-3.1-405b

Hermes 3 is a generalist language model with many improvements over Hermes 2, including advanced agentic capabilities, much better roleplaying, reasoning, multi-turn conversation, long context coherence, and improvements across the...

T Weight access not listed
Context
131.072K
Input
$1/M
Output
$1/M
Undated
sao10k/l3-lunaris-8b

Lunaris 8B is a versatile generalist and roleplaying model based on Llama 3. It's a strategic merge of multiple models, designed to balance creativity with improved logic and general knowledge....

T Weight access not listed
Context
8.192K
Input
$0.04/M
Output
$0.05/M
meta-llama/llama-3.1-8b-instruct

Meta's latest class of model (Llama 3.1) launched with a variety of sizes & flavors. This 8B instruct-tuned version is fast and efficient. It has demonstrated strong performance compared to...

T Weight access not listed
Context
131.072K
Input
$0.05/M
Output
$0.08/M
meta-llama/llama-3.1-70b-instruct

Meta's latest class of model (Llama 3.1) launched with a variety of sizes & flavors. This 70B instruct-tuned version is optimized for high quality dialogue usecases. It has demonstrated strong...

T Weight access not listed
Context
131.072K
Input
$0.4/M
Output
$0.4/M
Undated
mistralai/mistral-nemo

A 12B parameter model with a 128k token context length built by Mistral in collaboration with NVIDIA. The model is multilingual, supporting English, French, German, Spanish, Italian, Portuguese, Chinese, Japanese,...

T Weight access not listed
Context
131.072K
Input
$0.019/M
Output
$0.03/M
openai/gpt-4o-mini-2024-07-18

GPT-4o mini is OpenAI's newest model after [GPT-4 Omni](/models/openai/gpt-4o), supporting both text and image inputs with text outputs. As their most advanced small model, it is many multiples more affordable...

T Weight access not listed
Context
128K
Input
$0.15/M
Output
$0.6/M
Undated
google/gemma-2-27b-it

Gemma 2 27B by Google is an open model built from the same research and technology used to create the [Gemini models](/models?q=gemini). Gemma models are well-suited for a variety of...

T Weight access not listed
Context
8.192K
Input
$0.65/M
Output
$0.65/M
Undated
mistralai/mixtral-8x22b-instruct

Mistral's official instruct fine-tuned version of [Mixtral 8x22B](/models/mistralai/mixtral-8x22b). It uses 39B active parameters out of 141B, offering unparalleled cost efficiency for its size. Its strengths include: - strong math, coding,...

T Weight access not listed
Context
65.536K
Input
$2/M
Output
$6/M
WizardLM-2 8x22Bby Microsoft
Undated
microsoft/wizardlm-2-8x22b

WizardLM-2 8x22B is Microsoft AI's most advanced Wizard model. It demonstrates highly competitive performance compared to leading proprietary models, and it consistently outperforms all existing state-of-the-art opensource models. It is...

T Weight access not listed
Context
65.535K
Input
$0.62/M
Output
$0.62/M
Undated
anthropic/claude-3-haiku

Claude 3 Haiku is Anthropic's fastest and most compact model for near-instant responsiveness. Quick and accurate targeted performance. See the launch announcement and benchmark results [here](https://www.anthropic.com/news/claude-3-haiku) #multimodal

T Weight access not listed
Context
200K
Input
$0.25/M
Output
$1.25/M
Mistral Largeby Mistral AI
Undated
mistralai/mistral-large

This is Mistral AI's flagship model, Mistral Large 2 (version `mistral-large-2407`). It's a proprietary weights-available model and excels at reasoning, code, JSON, chat, and more. Read the launch announcement [here](https://mistral.ai/news/mistral-large-2407/)....

T Weight access not listed
Context
128K
Input
$2/M
Output
$6/M
openai/gpt-3.5-turbo-0613

GPT-3.5 Turbo is OpenAI's fastest model. It can understand and generate natural language or code, and is optimized for chat and traditional completion tasks. Training data up to Sep 2021.

T Weight access not listed
Context
4.095K
Input
$1/M
Output
$2/M
Undated
openai/gpt-4-turbo-preview

The preview GPT-4 model with improved instruction following, JSON mode, reproducible outputs, parallel function calling, and more. Training data: up to Dec 2023. **Note:** heavily rate limited by OpenAI while...

T Weight access not listed
Context
128K
Input
$10/M
Output
$30/M
Auto Routerby Openrouter
Undated
openrouter/auto

Your prompt will be processed by a meta-model and routed to one of dozens of models (see below), optimizing for the best possible output. To see which model was used,...

T Weight access not listed
Context
2M
Input
-
Output
-
openai/gpt-3.5-turbo-instruct

This model is a variant of GPT-3.5 Turbo tuned for instructional prompts and omitting chat-related optimizations. Training data: up to Sep 2021.

T Weight access not listed
Context
4.095K
Input
$1.5/M
Output
$2/M
Undated
openai/gpt-3.5-turbo-16k

This model offers four times the context length of gpt-3.5-turbo, allowing it to support approximately 20 pages of text in a single request at a higher cost. Training data: up...

T Weight access not listed
Context
16.385K
Input
$3/M
Output
$4/M
Undated
mancer/weaver

An attempt to recreate Claude-style verbosity, but don't expect the same level of coherence or memory. Meant for use in roleplay/narrative situations.

T Weight access not listed
Context
8K
Input
$0.5/M
Output
$0.75/M
Undated
undi95/remm-slerp-l2-13b

A recreation trial of the original MythoMax-L2-B13 but with updated models. #merge

T Weight access not listed
Context
6.144K
Input
$0.45/M
Output
$0.65/M
MythoMax 13Bby Gryphe
Undated
gryphe/mythomax-l2-13b

One of the highest performing and most popular fine-tunes of Llama 2 13B, with rich descriptions and roleplay. #merge

T Weight access not listed
Context
4.096K
Input
$0.06/M
Output
$0.06/M
Undated
kwaipilot/kat-coder-air-v2.5

KAT-Coder-Air V2.5 is a flagship-level Agentic Coding model that can directly hand over an entire issue or an entire business workflow to it, allowing it to autonomously locate and make...

T Weight access not listed
Context
256K
Input
$0.15/M
Output
$0.6/M
Undated
kwaipilot/kat-coder-pro-v2.5

KAT-Coder-Pro V2.5 is a flagship-level Agentic Coding model that can directly hand over an entire issue or an entire business workflow to it, allowing it to autonomously locate and make...

T Weight access not listed
Context
256K
Input
$0.74/M
Output
$2.96/M
meta-llama/Meta-Llama-3-8B-Instruct

No provider description is available for this model yet.

Open weights
Context
Not documented
Input
-
Output
-
meta-llama/Meta-Llama-3-70B-Instruct

No provider description is available for this model yet.

Open weights
Context
Not documented
Input
-
Output
-
meta-llama/Llama-4-Maverick-17B-128E-Instruct

No provider description is available for this model yet.

Weight access not listed
Context
Not documented
Input
-
Output
-
meta-llama/Llama-4-Scout-17B-16E-Instruct

No provider description is available for this model yet.

Weight access not listed
Context
Not documented
Input
-
Output
-
meta-llama/Llama-Guard-4-12B

No provider description is available for this model yet.

Weight access not listed
Context
Not documented
Input
-
Output
-
meta-llama/Llama-4-Maverick-17B-128E

No provider description is available for this model yet.

Weight access not listed
Context
Not documented
Input
-
Output
-
meta-llama/Llama-4-Scout-17B-16E

No provider description is available for this model yet.

Weight access not listed
Context
Not documented
Input
-
Output
-
meta-llama/Llama-3.2-90B-Vision-Instruct

No provider description is available for this model yet.

Open weights
Context
Not documented
Input
-
Output
-
meta-llama/Llama-3.3-70B-Instruct

No provider description is available for this model yet.

Open weights
Context
Not documented
Input
-
Output
-
meta-llama/Llama-3.1-70B-Instruct

No provider description is available for this model yet.

Open weights
Context
Not documented
Input
-
Output
-
meta-llama/Llama-3.2-11B-Vision-Instruct

No provider description is available for this model yet.

Open weights
Context
Not documented
Input
-
Output
-