Guide

2,496 step-by-step guides to setting up AI models from OpenAI, Anthropic, Google, OpenRouter, and more in TypingMind with your own API key.

Search and filter guides

Showing 613-663 of 713 guides

OpenRouter logo

DeepSeek: R1 Distill Llama 70B

OpenRouter · Jan 23, 2025

DeepSeek R1 Distill Llama 70B is a distilled large language model based on [Llama-3.3-70B-Instruct](/meta-llama/llama-3.3-70b-instruct), using outputs from [DeepSeek R1](/deepseek/deepseek-r1). The model combines advanced distillation techniques to achieve high performance across...

8.2K context
View
OpenRouter logo

DeepSeek: R1

OpenRouter · Jan 20, 2025

DeepSeek R1 is here: Performance on par with [OpenAI o1](/openai/o1), but open-sourced and with fully open reasoning tokens. It's 671B parameters in size, with 37B active in an inference pass....

64K context
View
OpenRouter logo

MiniMax: MiniMax-01

OpenRouter · Jan 15, 2025

MiniMax-01 is a combines MiniMax-Text-01 for text generation and MiniMax-VL-01 for image understanding. It has 456 billion parameters, with 45.9 billion parameters activated per inference, and can handle a context...

1M context
View
OpenRouter logo

Microsoft: Phi 4

OpenRouter · Jan 10, 2025

[Microsoft Research](/microsoft) Phi-4 is designed to perform well in complex reasoning tasks and can operate efficiently in situations with limited memory or where quick responses are needed. At 14 billion...

16.4K context
View
OpenRouter logo

Sao10K: Llama 3.1 70B Hanami x1

OpenRouter · Jan 8, 2025

This is [Sao10K](/sao10k)'s experiment over [Euryale v2.2](/sao10k/l3.1-euryale-70b).

16K context
View
OpenRouter logo

DeepSeek: DeepSeek V3

OpenRouter · Dec 26, 2024

DeepSeek-V3 is the latest model from the DeepSeek team, building upon the instruction following and coding abilities of the previous versions. Pre-trained on nearly 15 trillion tokens, the reported evaluations...

163.8K context
View
OpenRouter logo

Sao10K: Llama 3.3 Euryale 70B

OpenRouter · Dec 18, 2024

Euryale L3.3 70B is a model focused on creative roleplay from [Sao10k](https://ko-fi.com/sao10k). It is the successor of [Euryale L3 70B v2.2](/models/sao10k/l3-euryale-70b).

131.1K context
View
OpenRouter logo

OpenAI: o1

OpenRouter · Dec 17, 2024

The latest and strongest model family from OpenAI, o1 is designed to spend more time thinking before responding. The o1 model series is trained with large-scale reinforcement learning to reason...

200K context+1
View
OpenRouter logo

OpenAI: o1 (batch)

OpenRouter · Dec 17, 2024

The latest and strongest model family from OpenAI, o1 is designed to spend more time thinking before responding. The o1 model series is trained with large-scale reinforcement learning to reason...

200K context+1
View
OpenRouter logo

Cohere: Command R7B (12-2024)

OpenRouter · Dec 14, 2024

Command R7B (12-2024) is a small, fast update of the Command R+ model, delivered in December 2024. It excels at RAG, tool use, agents, and similar tasks requiring complex reasoning...

128K context
View
OpenRouter logo

Google: Gemini 2.0 Flash Experimental (free)

OpenRouter · Dec 11, 2024

Gemini Flash 2.0 offers a significantly faster time to first token (TTFT) compared to [Gemini Flash 1.5](/google/gemini-flash-1.5), while maintaining quality on par with larger models like [Gemini Pro 1.5](/google/gemini-pro-1.5). It introduces notable enhancements in multimodal understanding, coding capabilities, complex instruction following, and function calling. These advancements come together to deliver more seamless and robust agentic experiences.

1M context
View
OpenRouter logo

Meta: Llama 3.3 70B Instruct (free)

OpenRouter · Dec 6, 2024

The Meta Llama 3.3 multilingual large language model (LLM) is a pretrained and instruction tuned generative model in 70B (text in/text out). The Llama 3.3 instruction tuned text only model...

131.1K context
View
OpenRouter logo

Meta: Llama 3.3 70B Instruct

OpenRouter · Dec 6, 2024

The Meta Llama 3.3 multilingual large language model (LLM) is a pretrained and instruction tuned generative model in 70B (text in/text out). The Llama 3.3 instruction tuned text only model...

131.1K context
View
OpenRouter logo

Amazon: Nova Lite 1.0

OpenRouter · Dec 5, 2024

Amazon Nova Lite 1.0 is a very low-cost multimodal model from Amazon that focused on fast processing of image, video, and text inputs to generate text output. Amazon Nova Lite...

300K context
View
OpenRouter logo

Amazon: Nova Micro 1.0

OpenRouter · Dec 5, 2024

Amazon Nova Micro 1.0 is a text-only model that delivers the lowest latency responses in the Amazon Nova family of models at a very low cost. With a context length...

128K context
View
OpenRouter logo

Amazon: Nova Pro 1.0

OpenRouter · Dec 5, 2024

Amazon Nova Pro 1.0 is a capable multimodal model from Amazon focused on providing a combination of accuracy, speed, and cost for a wide range of tasks. As of December...

300K context
View
OpenRouter logo

OpenAI: GPT-4o (2024-11-20)

OpenRouter · Nov 20, 2024

The 2024-11-20 version of GPT-4o offers a leveled-up creative writing ability with more natural, engaging, and tailored writing to improve relevance & readability. It’s also better at working with uploaded...

128K context
View
OpenRouter logo

Mistral Large 2411

OpenRouter · Nov 19, 2024

Mistral Large 2 2411 is an update of [Mistral Large 2](/mistralai/mistral-large) released together with [Pixtral Large 2411](/mistralai/pixtral-large-2411) It provides a significant upgrade on the previous [Mistral Large 24.07](/mistralai/mistral-large-2407), with notable...

131.1K context
View
OpenRouter logo

Mistral Large 2407

OpenRouter · Nov 19, 2024

This is Mistral AI's flagship model, Mistral Large 2 (version mistral-large-2407). It's a proprietary weights-available model and excels at reasoning, code, JSON, chat, and more. Read the launch announcement [here](https://mistral.ai/news/mistral-large-2407/)....

131.1K context
View
OpenRouter logo

Mistral: Pixtral Large 2411

OpenRouter · Nov 19, 2024

Pixtral Large is a 124B parameter, open-weight, multimodal model built on top of [Mistral Large 2](/mistralai/mistral-large-2411). The model is able to understand documents, charts and natural images. The model is...

131.1K context
View
OpenRouter logo

Qwen2.5 Coder 32B Instruct

OpenRouter · Nov 11, 2024

Qwen2.5-Coder is the latest series of Code-Specific Qwen large language models (formerly known as CodeQwen). Qwen2.5-Coder brings the following improvements upon CodeQwen1.5: - Significantly improvements in **code generation**, **code reasoning**...

32.8K context
View
OpenRouter logo

SorcererLM 8x22B

OpenRouter · Nov 8, 2024

SorcererLM is an advanced RP and storytelling model, built as a Low-rank 16-bit LoRA fine-tuned on [WizardLM-2 8x22B](/microsoft/wizardlm-2-8x22b). - Advanced reasoning and emotional intelligence for engaging and immersive interactions - Vivid writing capabilities enriched with spatial and contextual awareness - Enhanced narrative depth, promoting creative and dynamic storytelling

16K context
View
OpenRouter logo

TheDrummer: UnslopNemo 12B

OpenRouter · Nov 8, 2024

UnslopNemo v4.1 is the latest addition from the creator of Rocinante, designed for adventure writing and role-play scenarios.

1M context
View
OpenRouter logo

Anthropic: Claude 3.5 Haiku (2024-10-22)

OpenRouter · Nov 4, 2024

Claude 3.5 Haiku features enhancements across all skill sets including coding, tool use, and reasoning. As the fastest model in the Anthropic lineup, it offers rapid response times suitable for applications that require high interactivity and low latency, such as user-facing chatbots and on-the-fly code completions. It also excels in specialized tasks like data extraction and real-time content moderation, making it a versatile tool for a broad range of industries. It does not support image inputs. See the launch announcement and benchmark results [here](https://www.anthropic.com/news/3-5-models-and-computer-use)

200K context
View
OpenRouter logo

Anthropic: Claude 3.5 Haiku

OpenRouter · Nov 4, 2024

Claude 3.5 Haiku features offers enhanced capabilities in speed, coding accuracy, and tool use. Engineered to excel in real-time applications, it delivers quick response times that are essential for dynamic...

200K context
View
OpenRouter logo

Magnum v4 72B

OpenRouter · Oct 22, 2024

This is a series of models designed to replicate the prose quality of the Claude 3 models, specifically Sonnet(https://openrouter.ai/anthropic/claude-3.5-sonnet) and Opus(https://openrouter.ai/anthropic/claude-3-opus). The model is fine-tuned on top of [Qwen2.5 72B](https://openrouter.ai/qwen/qwen-2.5-72b-instruct).

32.8K context
View
OpenRouter logo

Anthropic: Claude 3.5 Sonnet

OpenRouter · Oct 22, 2024

New Claude 3.5 Sonnet delivers better-than-Opus capabilities, faster-than-Sonnet speeds, at the same Sonnet prices. Sonnet is particularly good at: - Coding: Scores ~49% on SWE-Bench Verified, higher than the last best score, and without any fancy prompt scaffolding - Data science: Augments human data science expertise; navigates unstructured data while using multiple tools for insights - Visual processing: excelling at interpreting charts, graphs, and images, accurately transcribing text to derive insights beyond just the text alone - Agentic tasks: exceptional tool use, making it great at agentic tasks (i.e. complex, multi-step problem solving tasks that require engaging with other systems) #multimodal

200K context
View
OpenRouter logo

Mistral: Ministral 8B

OpenRouter · Oct 17, 2024

Ministral 8B is an 8B parameter model featuring a unique interleaved sliding-window attention pattern for faster, memory-efficient inference. Designed for edge use cases, it supports up to 128k context length and excels in knowledge and reasoning tasks. It outperforms peers in the sub-10B category, making it perfect for low-latency, privacy-first applications.

131.1K context
View
OpenRouter logo

Mistral: Ministral 3B

OpenRouter · Oct 17, 2024

Ministral 3B is a 3B parameter model optimized for on-device and edge computing. It excels in knowledge, commonsense reasoning, and function-calling, outperforming larger models like Mistral 7B on most benchmarks. Supporting up to 128k context length, it’s ideal for orchestrating agentic workflows and specialist tasks with efficient inference.

131.1K context
View
OpenRouter logo

Qwen: Qwen2.5 7B Instruct

OpenRouter · Oct 16, 2024

Qwen2.5 7B is the latest series of Qwen large language models. Qwen2.5 brings the following improvements upon Qwen2: - Significantly more knowledge and has greatly improved capabilities in coding and...

32.8K context
View
OpenRouter logo

NVIDIA: Llama 3.1 Nemotron 70B Instruct

OpenRouter · Oct 15, 2024

NVIDIA's Llama 3.1 Nemotron 70B is a language model designed for generating precise and useful responses. Leveraging [Llama 3.1 70B](/models/meta-llama/llama-3.1-70b-instruct) architecture and Reinforcement Learning from Human Feedback (RLHF), it excels...

131.1K context
View
OpenRouter logo

Inflection: Inflection 3 Pi

OpenRouter · Oct 11, 2024

Inflection 3 Pi powers Inflection's [Pi](https://pi.ai) chatbot, including backstory, emotional intelligence, productivity, and safety. It has access to recent news, and excels in scenarios like customer support and roleplay. Pi...

8K context
View
OpenRouter logo

Inflection: Inflection 3 Productivity

OpenRouter · Oct 11, 2024

Inflection 3 Productivity is optimized for following instructions. It is better for tasks requiring JSON output or precise adherence to provided guidelines. It has access to recent news. For emotional...

8K context
View
OpenRouter logo

TheDrummer: Rocinante 12B

OpenRouter · Sep 30, 2024

Rocinante 12B is designed for engaging storytelling and rich prose. Early testers have reported: - Expanded vocabulary with unique and expressive word choices - Enhanced creativity for vivid narratives -...

65.5K context
View
OpenRouter logo

Meta: Llama 3.2 3B Instruct (free)

OpenRouter · Sep 25, 2024

Llama 3.2 3B is a 3-billion-parameter multilingual large language model, optimized for advanced natural language processing tasks like dialogue generation, reasoning, and summarization. Designed with the latest transformer architecture, it...

131.1K context
View
OpenRouter logo

Meta: Llama 3.2 3B Instruct

OpenRouter · Sep 25, 2024

Llama 3.2 3B is a 3-billion-parameter multilingual large language model, optimized for advanced natural language processing tasks like dialogue generation, reasoning, and summarization. Designed with the latest transformer architecture, it...

131.1K context
View
OpenRouter logo

Meta: Llama 3.2 1B Instruct

OpenRouter · Sep 25, 2024

Llama 3.2 1B is a 1-billion-parameter language model focused on efficiently performing natural language tasks, such as summarization, dialogue, and multilingual text analysis. Its smaller size allows it to operate...

60K context
View
OpenRouter logo

Meta: Llama 3.2 90B Vision Instruct

OpenRouter · Sep 25, 2024

The Llama 90B Vision model is a top-tier, 90-billion-parameter multimodal model designed for the most challenging visual reasoning and language tasks. It offers unparalleled accuracy in image captioning, visual question answering, and advanced image-text comprehension. Pre-trained on vast multimodal datasets and fine-tuned with human feedback, the Llama 90B Vision is engineered to handle the most demanding image-based AI tasks. This model is perfect for industries requiring cutting-edge multimodal AI capabilities, particularly those dealing with complex, real-time visual and textual analysis. Click here for the [original model card](https://github.com/meta-llama/llama-models/blob/main/models/llama3_2/MODEL_CARD_VISION.md). Usage of this model is subject to [Meta's Acceptable Use Policy](https://www.llama.com/llama3/use-policy/).

32.8K context
View
OpenRouter logo

Meta: Llama 3.2 11B Vision Instruct

OpenRouter · Sep 25, 2024

Llama 3.2 11B Vision is a multimodal model with 11 billion parameters, designed to handle tasks combining visual and textual data. It excels in tasks such as image captioning and...

131.1K context
View
OpenRouter logo

Qwen2.5 72B Instruct

OpenRouter · Sep 19, 2024

Qwen2.5 72B is the latest series of Qwen large language models. Qwen2.5 brings the following improvements upon Qwen2: - Significantly more knowledge and has greatly improved capabilities in coding and...

32.8K context
View
OpenRouter logo

NeverSleep: Lumimaid v0.2 8B

OpenRouter · Sep 15, 2024

Lumimaid v0.2 8B is a finetune of [Llama 3.1 8B](/models/meta-llama/llama-3.1-8b-instruct) with a "HUGE step up dataset wise" compared to Lumimaid v0.1. Sloppy chats output were purged. Usage of this model is subject to [Meta's Acceptable Use Policy](https://llama.meta.com/llama3/use-policy/).

32.8K context
View
OpenRouter logo

Mistral: Pixtral 12B

OpenRouter · Sep 10, 2024

The first multi-modal, text+image-to-text model from Mistral AI. Its weights were launched via torrent: https://x.com/mistralai/status/1833758285167722836.

32.8K context
View
OpenRouter logo

Cohere: Command R (08-2024)

OpenRouter · Aug 30, 2024

command-r-08-2024 is an update of the [Command R](/models/cohere/command-r) with improved performance for multilingual retrieval-augmented generation (RAG) and tool use. More broadly, it is better at math, code and reasoning and...

128K context
View
OpenRouter logo

Cohere: Command R+ (08-2024)

OpenRouter · Aug 30, 2024

command-r-plus-08-2024 is an update of the [Command R+](/models/cohere/command-r-plus) with roughly 50% higher throughput and 25% lower latencies as compared to the previous Command R+ version, while keeping the hardware footprint...

128K context
View
OpenRouter logo

Sao10K: Llama 3.1 Euryale 70B v2.2

OpenRouter · Aug 28, 2024

Euryale L3.1 70B v2.2 is a model focused on creative roleplay from [Sao10k](https://ko-fi.com/sao10k). It is the successor of [Euryale L3 70B v2.1](/models/sao10k/l3-euryale-70b).

131.1K context
View
OpenRouter logo

Qwen: Qwen2.5-VL 7B Instruct (free)

OpenRouter · Aug 28, 2024

Qwen2.5 VL 7B is a multimodal LLM from the Qwen Team with the following key enhancements: - SoTA understanding of images of various resolution & ratio: Qwen2.5-VL achieves state-of-the-art performance on visual understanding benchmarks, including MathVista, DocVQA, RealWorldQA, MTVQA, etc. - Understanding videos of 20min+: Qwen2.5-VL can understand videos over 20 minutes for high-quality video-based question answering, dialog, content creation, etc. - Agent that can operate your mobiles, robots, etc.: with the abilities of complex reasoning and decision making, Qwen2.5-VL can be integrated with devices like mobile phones, robots, etc., for automatic operation based on visual environment and text instructions. - Multilingual Support: to serve global users, besides English and Chinese, Qwen2.5-VL now supports the understanding of texts in different languages inside images, including most European languages, Japanese, Korean, Arabic, Vietnamese, etc. For more details, see this [blog post](https://qwenlm.github.io/blog/qwen2-vl/) and [GitHub repo](https://github.com/QwenLM/Qwen2-VL). Usage of this model is subject to [Tongyi Qianwen LICENSE AGREEMENT](https://huggingface.co/Qwen/Qwen1.5-110B-Chat/blob/main/LICENSE).

32.8K context
View
OpenRouter logo

Qwen: Qwen2.5-VL 7B Instruct

OpenRouter · Aug 28, 2024

Qwen2.5 VL 7B is a multimodal LLM from the Qwen Team with the following key enhancements: - SoTA understanding of images of various resolution & ratio: Qwen2.5-VL achieves state-of-the-art performance on visual understanding benchmarks, including MathVista, DocVQA, RealWorldQA, MTVQA, etc. - Understanding videos of 20min+: Qwen2.5-VL can understand videos over 20 minutes for high-quality video-based question answering, dialog, content creation, etc. - Agent that can operate your mobiles, robots, etc.: with the abilities of complex reasoning and decision making, Qwen2.5-VL can be integrated with devices like mobile phones, robots, etc., for automatic operation based on visual environment and text instructions. - Multilingual Support: to serve global users, besides English and Chinese, Qwen2.5-VL now supports the understanding of texts in different languages inside images, including most European languages, Japanese, Korean, Arabic, Vietnamese, etc. For more details, see this [blog post](https://qwenlm.github.io/blog/qwen2-vl/) and [GitHub repo](https://github.com/QwenLM/Qwen2-VL). Usage of this model is subject to [Tongyi Qianwen LICENSE AGREEMENT](https://huggingface.co/Qwen/Qwen1.5-110B-Chat/blob/main/LICENSE).

32.8K context
View
OpenRouter logo

Nous: Hermes 3 70B Instruct

OpenRouter · Aug 18, 2024

Hermes 3 is a generalist language model with many improvements over [Hermes 2](/models/nousresearch/nous-hermes-2-mistral-7b-dpo), including advanced agentic capabilities, much better roleplaying, reasoning, multi-turn conversation, long context coherence, and improvements across the...

131.1K context
View
OpenRouter logo

Nous: Hermes 3 405B Instruct (free)

OpenRouter · Aug 16, 2024

Hermes 3 is a generalist language model with many improvements over Hermes 2, including advanced agentic capabilities, much better roleplaying, reasoning, multi-turn conversation, long context coherence, and improvements across the...

131.1K context
View
OpenRouter logo

Nous: Hermes 3 405B Instruct

OpenRouter · Aug 16, 2024

Hermes 3 is a generalist language model with many improvements over Hermes 2, including advanced agentic capabilities, much better roleplaying, reasoning, multi-turn conversation, long context coherence, and improvements across the...

131.1K context
View
OpenRouter logo

OpenAI: ChatGPT-4o

OpenRouter · Aug 14, 2024

OpenAI ChatGPT 4o is continually updated by OpenAI to point to the current version of GPT-4o used by ChatGPT. It therefore differs slightly from the API version of [GPT-4o](/models/openai/gpt-4o) in that it has additional RLHF. It is intended for research and evaluation. OpenAI notes that this model is not suited for production use-cases as it may be removed or redirected to another model in the future.

128K context
View