thedrummer/unslopnemo-12b
UnslopNemo v4.1 is the latest addition from the creator of Rocinante, designed for adventure writing and role-play scenarios.
- Context
- 32.768K
- Input
- $0.4/M
- Output
- $0.4/M
Filter 978 source-linked models by creator, price, context, modality, and published benchmark coverage.
thedrummer/unslopnemo-12b
UnslopNemo v4.1 is the latest addition from the creator of Rocinante, designed for adventure writing and role-play scenarios.
anthracite-org/magnum-v4-72b
This is a series of models designed to replicate the prose quality of the Claude 3 models, specifically Sonnet(https://openrouter.ai/anthropic/claude-3.5-sonnet) and Opus(https://openrouter.ai/anthropic/claude-3-opus). The model is fine-tuned on top of [Qwen2.5 72B](https://openrouter.ai/qwen/qwen-2.5-72b-instruct).
qwen/qwen-2.5-7b-instruct
Qwen2.5 7B is the latest series of Qwen large language models. Qwen2.5 brings the following improvements upon Qwen2: - Significantly more knowledge and has greatly improved capabilities in coding and...
inflection/inflection-3-productivity
Inflection 3 Productivity is optimized for following instructions. It is better for tasks requiring JSON output or precise adherence to provided guidelines. It has access to recent news. For emotional...
inflection/inflection-3-pi
Inflection 3 Pi powers Inflection's [Pi](https://pi.ai) chatbot, including backstory, emotional intelligence, productivity, and safety. It has access to recent news, and excels in scenarios like customer support and roleplay. Pi...
thedrummer/rocinante-12b
Rocinante 12B is designed for engaging storytelling and rich prose. Early testers have reported: - Expanded vocabulary with unique and expressive word choices - Enhanced creativity for vivid narratives -...
meta-llama/llama-3.2-1b-instruct
Llama 3.2 1B is a 1-billion-parameter language model focused on efficiently performing natural language tasks, such as summarization, dialogue, and multilingual text analysis. Its smaller size allows it to operate...
meta-llama/llama-3.2-11b-vision-instruct
Llama 3.2 11B Vision is a multimodal model with 11 billion parameters, designed to handle tasks combining visual and textual data. It excels in tasks such as image captioning and...
meta-llama/llama-3.2-3b-instruct:free
Llama 3.2 3B is a 3-billion-parameter multilingual large language model, optimized for advanced natural language processing tasks like dialogue generation, reasoning, and summarization. Designed with the latest transformer architecture, it...
meta-llama/llama-3.2-3b-instruct
Llama 3.2 3B is a 3-billion-parameter multilingual large language model, optimized for advanced natural language processing tasks like dialogue generation, reasoning, and summarization. Designed with the latest transformer architecture, it...
qwen/qwen-2.5-72b-instruct
Qwen2.5 72B is the latest series of Qwen large language models. Qwen2.5 brings the following improvements upon Qwen2: - Significantly more knowledge and has greatly improved capabilities in coding and...
sao10k/l3.1-euryale-70b
Euryale L3.1 70B v2.2 is a model focused on creative roleplay from [Sao10k](https://ko-fi.com/sao10k). It is the successor of [Euryale L3 70B v2.1](/models/sao10k/l3-euryale-70b).
nousresearch/hermes-3-llama-3.1-70b
Hermes 3 is a generalist language model with many improvements over [Hermes 2](/models/nousresearch/nous-hermes-2-mistral-7b-dpo), including advanced agentic capabilities, much better roleplaying, reasoning, multi-turn conversation, long context coherence, and improvements across the...
nousresearch/hermes-3-llama-3.1-405b:free
Hermes 3 is a generalist language model with many improvements over Hermes 2, including advanced agentic capabilities, much better roleplaying, reasoning, multi-turn conversation, long context coherence, and improvements across the...
nousresearch/hermes-3-llama-3.1-405b
Hermes 3 is a generalist language model with many improvements over Hermes 2, including advanced agentic capabilities, much better roleplaying, reasoning, multi-turn conversation, long context coherence, and improvements across the...
sao10k/l3-lunaris-8b
Lunaris 8B is a versatile generalist and roleplaying model based on Llama 3. It's a strategic merge of multiple models, designed to balance creativity with improved logic and general knowledge....
meta-llama/llama-3.1-8b-instruct
Meta's latest class of model (Llama 3.1) launched with a variety of sizes & flavors. This 8B instruct-tuned version is fast and efficient. It has demonstrated strong performance compared to...
meta-llama/llama-3.1-70b-instruct
Meta's latest class of model (Llama 3.1) launched with a variety of sizes & flavors. This 70B instruct-tuned version is optimized for high quality dialogue usecases. It has demonstrated strong...
mistralai/mistral-nemo
A 12B parameter model with a 128k token context length built by Mistral in collaboration with NVIDIA. The model is multilingual, supporting English, French, German, Spanish, Italian, Portuguese, Chinese, Japanese,...
openai/gpt-4o-mini-2024-07-18
GPT-4o mini is OpenAI's newest model after [GPT-4 Omni](/models/openai/gpt-4o), supporting both text and image inputs with text outputs. As their most advanced small model, it is many multiples more affordable...
google/gemma-2-27b-it
Gemma 2 27B by Google is an open model built from the same research and technology used to create the [Gemini models](/models?q=gemini). Gemma models are well-suited for a variety of...
mistralai/mixtral-8x22b-instruct
Mistral's official instruct fine-tuned version of [Mixtral 8x22B](/models/mistralai/mixtral-8x22b). It uses 39B active parameters out of 141B, offering unparalleled cost efficiency for its size. Its strengths include: - strong math, coding,...
microsoft/wizardlm-2-8x22b
WizardLM-2 8x22B is Microsoft AI's most advanced Wizard model. It demonstrates highly competitive performance compared to leading proprietary models, and it consistently outperforms all existing state-of-the-art opensource models. It is...
anthropic/claude-3-haiku
Claude 3 Haiku is Anthropic's fastest and most compact model for near-instant responsiveness. Quick and accurate targeted performance. See the launch announcement and benchmark results [here](https://www.anthropic.com/news/claude-3-haiku) #multimodal
mistralai/mistral-large
This is Mistral AI's flagship model, Mistral Large 2 (version `mistral-large-2407`). It's a proprietary weights-available model and excels at reasoning, code, JSON, chat, and more. Read the launch announcement [here](https://mistral.ai/news/mistral-large-2407/)....
openai/gpt-3.5-turbo-0613
GPT-3.5 Turbo is OpenAI's fastest model. It can understand and generate natural language or code, and is optimized for chat and traditional completion tasks. Training data up to Sep 2021.
openai/gpt-4-turbo-preview
The preview GPT-4 model with improved instruction following, JSON mode, reproducible outputs, parallel function calling, and more. Training data: up to Dec 2023. **Note:** heavily rate limited by OpenAI while...
openrouter/auto
Your prompt will be processed by a meta-model and routed to one of dozens of models (see below), optimizing for the best possible output. To see which model was used,...
openai/gpt-3.5-turbo-instruct
This model is a variant of GPT-3.5 Turbo tuned for instructional prompts and omitting chat-related optimizations. Training data: up to Sep 2021.
openai/gpt-3.5-turbo-16k
This model offers four times the context length of gpt-3.5-turbo, allowing it to support approximately 20 pages of text in a single request at a higher cost. Training data: up...
mancer/weaver
An attempt to recreate Claude-style verbosity, but don't expect the same level of coherence or memory. Meant for use in roleplay/narrative situations.
undi95/remm-slerp-l2-13b
A recreation trial of the original MythoMax-L2-B13 but with updated models. #merge
gryphe/mythomax-l2-13b
One of the highest performing and most popular fine-tunes of Llama 2 13B, with rich descriptions and roleplay. #merge
kwaipilot/kat-coder-air-v2.5
KAT-Coder-Air V2.5 is a flagship-level Agentic Coding model that can directly hand over an entire issue or an entire business workflow to it, allowing it to autonomously locate and make...
kwaipilot/kat-coder-pro-v2.5
KAT-Coder-Pro V2.5 is a flagship-level Agentic Coding model that can directly hand over an entire issue or an entire business workflow to it, allowing it to autonomously locate and make...
meta-llama/Meta-Llama-3-8B-Instruct
No provider description is available for this model yet.
meta-llama/Meta-Llama-3-70B-Instruct
No provider description is available for this model yet.
meta-llama/Llama-4-Maverick-17B-128E-Instruct
No provider description is available for this model yet.
meta-llama/Llama-4-Scout-17B-16E-Instruct
No provider description is available for this model yet.
meta-llama/Llama-Guard-4-12B
No provider description is available for this model yet.
meta-llama/Llama-4-Maverick-17B-128E
No provider description is available for this model yet.
meta-llama/Llama-4-Scout-17B-16E
No provider description is available for this model yet.
meta-llama/Llama-3.2-90B-Vision-Instruct
No provider description is available for this model yet.
meta-llama/Llama-3.3-70B-Instruct
No provider description is available for this model yet.
meta-llama/Llama-3.1-70B-Instruct
No provider description is available for this model yet.
meta-llama/Llama-3.2-11B-Vision-Instruct
No provider description is available for this model yet.
meta-llama/Llama-3.2-3B-Instruct-QLORA_INT4_EO8
No provider description is available for this model yet.
meta-llama/Llama-3.2-3B-Instruct-SpinQuant_INT4_EO8
No provider description is available for this model yet.
meta-llama/Llama-3.2-1B-Instruct-SpinQuant_INT4_EO8
No provider description is available for this model yet.
meta-llama/Llama-3.2-1B-Instruct-QLORA_INT4_EO8
No provider description is available for this model yet.
| Model | Creator | Input types | Context | Input / Output | Released | Compare |
|---|---|---|---|---|---|---|
| TheDrummer: UnslopNemo 12Bthedrummer/unslopnemo-12b | 32.768K | $0.4 / $0.4 | Undated | |||
| Magnum v4 72Banthracite-org/magnum-v4-72b | 16.384K | $3 / $5 | Undated | |||
| Qwen: Qwen2.5 7B Instructqwen/qwen-2.5-7b-instruct | 32.768K | $0.1 / $0.2 | Undated | |||
| Inflection: Inflection 3 Productivityinflection/inflection-3-productivity | 8K | $2.5 / $10 | Undated | |||
| Inflection: Inflection 3 Piinflection/inflection-3-pi | 8K | $2.5 / $10 | Undated | |||
| TheDrummer: Rocinante 12Bthedrummer/rocinante-12b | 65.536K | $0.25 / $0.5 | Undated | |||
| Meta: Llama 3.2 1B Instructmeta-llama/llama-3.2-1b-instruct | 60K | $0.027 / $0.201 | Undated | |||
| Meta: Llama 3.2 11B Vision Instructmeta-llama/llama-3.2-11b-vision-instruct | 131.072K | $0.345 / $0.345 | Undated | |||
| Meta: Llama 3.2 3B Instruct (free)meta-llama/llama-3.2-3b-instruct:free | 131.072K | Free / Free | Undated | |||
| Meta: Llama 3.2 3B Instructmeta-llama/llama-3.2-3b-instruct | 131.072K | $0.05 / $0.33 | Undated | |||
| Qwen2.5 72B Instructqwen/qwen-2.5-72b-instruct | 32.768K | $0.36 / $0.4 | Undated | |||
| Sao10K: Llama 3.1 Euryale 70B v2.2sao10k/l3.1-euryale-70b | 131.072K | $0.85 / $0.85 | Undated | |||
| Nous: Hermes 3 70B Instructnousresearch/hermes-3-llama-3.1-70b | 131.072K | $0.7 / $0.7 | Undated | |||
| Nous: Hermes 3 405B Instruct (free)nousresearch/hermes-3-llama-3.1-405b:free | 131.072K | Free / Free | Undated | |||
| Nous: Hermes 3 405B Instructnousresearch/hermes-3-llama-3.1-405b | 131.072K | $1 / $1 | Undated | |||
| Sao10K: Llama 3 8B Lunarissao10k/l3-lunaris-8b | 8.192K | $0.04 / $0.05 | Undated | |||
| Meta: Llama 3.1 8B Instructmeta-llama/llama-3.1-8b-instruct | 131.072K | $0.05 / $0.08 | Undated | |||
| Meta: Llama 3.1 70B Instructmeta-llama/llama-3.1-70b-instruct | 131.072K | $0.4 / $0.4 | Undated | |||
| Mistral: Mistral Nemomistralai/mistral-nemo | 131.072K | $0.019 / $0.03 | Undated | |||
| OpenAI: GPT-4o-mini (2024-07-18)openai/gpt-4o-mini-2024-07-18 | 128K | $0.15 / $0.6 | Undated | |||
| Google: Gemma 2 27Bgoogle/gemma-2-27b-it | 8.192K | $0.65 / $0.65 | Undated | |||
| Mistral: Mixtral 8x22B Instructmistralai/mixtral-8x22b-instruct | 65.536K | $2 / $6 | Undated | |||
| WizardLM-2 8x22Bmicrosoft/wizardlm-2-8x22b | 65.535K | $0.62 / $0.62 | Undated | |||
| Anthropic: Claude 3 Haikuanthropic/claude-3-haiku | 200K | $0.25 / $1.25 | Undated | |||
| Mistral Largemistralai/mistral-large | 128K | $2 / $6 | Undated | |||
| OpenAI: GPT-3.5 Turbo (older v0613)openai/gpt-3.5-turbo-0613 | 4.095K | $1 / $2 | Undated | |||
| OpenAI: GPT-4 Turbo Previewopenai/gpt-4-turbo-preview | 128K | $10 / $30 | Undated | |||
| Auto Routeropenrouter/auto | 2M | - / - | Undated | |||
| OpenAI: GPT-3.5 Turbo Instructopenai/gpt-3.5-turbo-instruct | 4.095K | $1.5 / $2 | Undated | |||
| OpenAI: GPT-3.5 Turbo 16kopenai/gpt-3.5-turbo-16k | 16.385K | $3 / $4 | Undated | |||
| Mancer: Weaver (alpha)mancer/weaver | 8K | $0.5 / $0.75 | Undated | |||
| ReMM SLERP 13Bundi95/remm-slerp-l2-13b | 6.144K | $0.45 / $0.65 | Undated | |||
| MythoMax 13Bgryphe/mythomax-l2-13b | 4.096K | $0.06 / $0.06 | Undated | |||
| Kwaipilot: KAT-Coder-Air V2.5kwaipilot/kat-coder-air-v2.5 | 256K | $0.15 / $0.6 | Undated | |||
| Kwaipilot: KAT-Coder-Pro V2.5kwaipilot/kat-coder-pro-v2.5 | 256K | $0.74 / $2.96 | Undated | |||
| meta-llama/Meta-Llama-3-8B-Instructmeta-llama/Meta-Llama-3-8B-Instruct | Not documented | - / - | Undated | |||
| meta-llama/Meta-Llama-3-70B-Instructmeta-llama/Meta-Llama-3-70B-Instruct | Not documented | - / - | Undated | |||
| meta-llama/Llama-4-Maverick-17B-128E-Instructmeta-llama/Llama-4-Maverick-17B-128E-Instruct | Not documented | - / - | Undated | |||
| meta-llama/Llama-4-Scout-17B-16E-Instructmeta-llama/Llama-4-Scout-17B-16E-Instruct | Not documented | - / - | Undated | |||
| meta-llama/Llama-Guard-4-12Bmeta-llama/Llama-Guard-4-12B | Not documented | - / - | Undated | |||
| meta-llama/Llama-4-Maverick-17B-128Emeta-llama/Llama-4-Maverick-17B-128E | Not documented | - / - | Undated | |||
| meta-llama/Llama-4-Scout-17B-16Emeta-llama/Llama-4-Scout-17B-16E | Not documented | - / - | Undated | |||
| meta-llama/Llama-3.2-90B-Vision-Instructmeta-llama/Llama-3.2-90B-Vision-Instruct | Not documented | - / - | Undated | |||
| meta-llama/Llama-3.3-70B-Instructmeta-llama/Llama-3.3-70B-Instruct | Not documented | - / - | Undated | |||
| meta-llama/Llama-3.1-70B-Instructmeta-llama/Llama-3.1-70B-Instruct | Not documented | - / - | Undated | |||
| meta-llama/Llama-3.2-11B-Vision-Instructmeta-llama/Llama-3.2-11B-Vision-Instruct | Not documented | - / - | Undated | |||
| meta-llama/Llama-3.2-3B-Instruct-QLORA_INT4_EO8meta-llama/Llama-3.2-3B-Instruct-QLORA_INT4_EO8 | Not documented | - / - | Undated | |||
| meta-llama/Llama-3.2-3B-Instruct-SpinQuant_INT4_EO8meta-llama/Llama-3.2-3B-Instruct-SpinQuant_INT4_EO8 | Not documented | - / - | Undated | |||
| meta-llama/Llama-3.2-1B-Instruct-SpinQuant_INT4_EO8meta-llama/Llama-3.2-1B-Instruct-SpinQuant_INT4_EO8 | Not documented | - / - | Undated | |||
| meta-llama/Llama-3.2-1B-Instruct-QLORA_INT4_EO8meta-llama/Llama-3.2-1B-Instruct-QLORA_INT4_EO8 | Not documented | - / - | Undated |