Guide

2,496 step-by-step guides to setting up AI models from OpenAI, Anthropic, Google, OpenRouter, and more in TypingMind with your own API key.

Search and filter guides

Showing 562-612 of 713 guides

OpenRouter logo

xAI: Grok 3 Mini Beta

OpenRouter · Apr 9, 2025

Grok 3 Mini is a lightweight, smaller thinking model. Unlike traditional models that generate answers immediately, Grok 3 Mini thinks before responding. It’s ideal for reasoning-heavy tasks that don’t demand...

131.1K context
View
OpenRouter logo

xAI: Grok 3 Beta

OpenRouter · Apr 9, 2025

Grok 3 is the latest model from xAI. It's their flagship model that excels at enterprise use cases like data extraction, coding, and text summarization. Possesses deep domain knowledge in...

131.1K context
View
OpenRouter logo

NVIDIA: Llama 3.1 Nemotron Ultra 253B v1

OpenRouter · Apr 8, 2025

Llama-3.1-Nemotron-Ultra-253B-v1 is a large language model (LLM) optimized for advanced reasoning, human-interactive chat, retrieval-augmented generation (RAG), and tool-calling tasks. Derived from Meta’s Llama-3.1-405B-Instruct, it has been significantly customized using Neural...

131.1K context
View
OpenRouter logo

Meta: Llama 4 Maverick

OpenRouter · Apr 5, 2025

Llama 4 Maverick 17B Instruct (128E) is a high-capacity multimodal language model from Meta, built on a mixture-of-experts (MoE) architecture with 128 experts and 17 billion active parameters per forward...

1M context
View
OpenRouter logo

Meta: Llama 4 Scout

OpenRouter · Apr 5, 2025

Llama 4 Scout 17B Instruct (16E) is a mixture-of-experts (MoE) language model developed by Meta, activating 17 billion parameters out of a total of 109B. It supports native multimodal input...

1.3M context
View
OpenRouter logo

Qwen: Qwen2.5 VL 32B Instruct

OpenRouter · Mar 24, 2025

Qwen2.5-VL-32B is a multimodal vision-language model fine-tuned through reinforcement learning for enhanced mathematical reasoning, structured outputs, and visual problem-solving capabilities. It excels at visual analysis tasks, including object recognition, textual...

128K context
View
OpenRouter logo

DeepSeek: DeepSeek V3 0324

OpenRouter · Mar 24, 2025

DeepSeek V3, a 685B-parameter, mixture-of-experts model, is the latest iteration of the flagship chat model family from the DeepSeek team. It succeeds the [DeepSeek V3](/deepseek/deepseek-chat-v3) model and performs really well...

163.8K context
View
OpenRouter logo

OpenAI: o1-pro

OpenRouter · Mar 19, 2025

The o1 series of models are trained with reinforcement learning to think before they answer and perform complex reasoning. The o1-pro model uses more compute to think harder and provide...

200K context
View
OpenRouter logo

OpenAI: o1-pro (batch)

OpenRouter · Mar 19, 2025

The o1 series of models are trained with reinforcement learning to think before they answer and perform complex reasoning. The o1-pro model uses more compute to think harder and provide...

200K context
View
OpenRouter logo

Mistral: Mistral Small 3.1 24B (free)

OpenRouter · Mar 17, 2025

Mistral Small 3.1 24B Instruct is an upgraded variant of Mistral Small 3 (2501), featuring 24 billion parameters with advanced multimodal capabilities. It provides state-of-the-art performance in text-based reasoning and vision tasks, including image analysis, programming, mathematical reasoning, and multilingual support across dozens of languages. Equipped with an extensive 128k token context window and optimized for efficient local inference, it supports use cases such as conversational agents, function calling, long-document comprehension, and privacy-sensitive deployments. The updated version is [Mistral Small 3.2](mistralai/mistral-small-3.2-24b-instruct)

128K context
View
OpenRouter logo

Mistral: Mistral Small 3.1 24B

OpenRouter · Mar 17, 2025

Mistral Small 3.1 24B Instruct is an upgraded variant of Mistral Small 3 (2501), featuring 24 billion parameters with advanced multimodal capabilities. It provides state-of-the-art performance in text-based reasoning and...

128K context
View
OpenRouter logo

AllenAI: Olmo 2 32B Instruct

OpenRouter · Mar 14, 2025

OLMo-2 32B Instruct is a supervised instruction-finetuned variant of the OLMo-2 32B March 2025 base model. It excels in complex reasoning and instruction-following tasks across diverse benchmarks such as GSM8K,...

128K context
View
OpenRouter logo

Google: Gemma 3 4B (free)

OpenRouter · Mar 13, 2025

Gemma 3 introduces multimodality, supporting vision-language input and text outputs. It handles context windows up to 128k tokens, understands over 140 languages, and offers improved math, reasoning, and chat capabilities,...

32.8K context
View
OpenRouter logo

Google: Gemma 3 4B

OpenRouter · Mar 13, 2025

Gemma 3 introduces multimodality, supporting vision-language input and text outputs. It handles context windows up to 128k tokens, understands over 140 languages, and offers improved math, reasoning, and chat capabilities,...

131.1K context
View
OpenRouter logo

Google: Gemma 3 12B (free)

OpenRouter · Mar 13, 2025

Gemma 3 introduces multimodality, supporting vision-language input and text outputs. It handles context windows up to 128k tokens, understands over 140 languages, and offers improved math, reasoning, and chat capabilities,...

32.8K context
View
OpenRouter logo

Google: Gemma 3 12B

OpenRouter · Mar 13, 2025

Gemma 3 introduces multimodality, supporting vision-language input and text outputs. It handles context windows up to 128k tokens, understands over 140 languages, and offers improved math, reasoning, and chat capabilities,...

131.1K context
View
OpenRouter logo

Cohere: Command A

OpenRouter · Mar 13, 2025

Command A is an open-weights 111B parameter model with a 256k context window focused on delivering great performance across agentic, multilingual, and coding use cases. Compared to other leading proprietary...

256K context
View
OpenRouter logo

OpenAI: GPT-4o-mini Search Preview

OpenRouter · Mar 12, 2025

GPT-4o mini Search Preview is a specialized model for web search in Chat Completions. It is trained to understand and execute web search queries.

128K context
View
OpenRouter logo

OpenAI: GPT-4o Search Preview

OpenRouter · Mar 12, 2025

GPT-4o Search Previewis a specialized model for web search in Chat Completions. It is trained to understand and execute web search queries.

128K context
View
OpenRouter logo

Reka Flash 3

OpenRouter · Mar 12, 2025

Reka Flash 3 is a general-purpose, instruction-tuned large language model with 21 billion parameters, developed by Reka. It excels at general chat, coding tasks, instruction-following, and function calling. Featuring a...

65.5K context
View
OpenRouter logo

Google: Gemma 3 27B (free)

OpenRouter · Mar 12, 2025

Gemma 3 introduces multimodality, supporting vision-language input and text outputs. It handles context windows up to 128k tokens, understands over 140 languages, and offers improved math, reasoning, and chat capabilities,...

131.1K context
View
OpenRouter logo

Google: Gemma 3 27B

OpenRouter · Mar 12, 2025

Gemma 3 introduces multimodality, supporting vision-language input and text outputs. It handles context windows up to 128k tokens, understands over 140 languages, and offers improved math, reasoning, and chat capabilities,...

131.1K context
View
OpenRouter logo

TheDrummer: Skyfall 36B V2

OpenRouter · Mar 10, 2025

Skyfall 36B v2 is an enhanced iteration of Mistral Small 2501, specifically fine-tuned for improved creativity, nuanced writing, role-playing, and coherent storytelling.

32.8K context
View
OpenRouter logo

Microsoft: Phi 4 Multimodal Instruct

OpenRouter · Mar 8, 2025

Phi-4 Multimodal Instruct is a versatile 5.6B parameter foundation model that combines advanced reasoning and instruction-following capabilities across both text and visual inputs, providing accurate text outputs. The unified architecture enables efficient, low-latency inference, suitable for edge and mobile deployments. Phi-4 Multimodal Instruct supports text inputs in multiple languages including Arabic, Chinese, English, French, German, Japanese, Spanish, and more, with visual input optimized primarily for English. It delivers impressive performance on multimodal tasks involving mathematical, scientific, and document reasoning, providing developers and enterprises a powerful yet compact model for sophisticated interactive applications. For more information, see the [Phi-4 Multimodal blog post](https://azure.microsoft.com/en-us/blog/empowering-innovation-the-next-generation-of-the-phi-family/).

131.1K context
View
OpenRouter logo

Perplexity: Sonar Reasoning Pro

OpenRouter · Mar 7, 2025

Note: Sonar Pro pricing includes Perplexity search pricing. See [details here](https://docs.perplexity.ai/guides/pricing#detailed-pricing-breakdown-for-sonar-reasoning-pro-and-sonar-pro) Sonar Reasoning Pro is a premier reasoning model powered by DeepSeek R1 with Chain of Thought (CoT). Designed for...

128K context
View
OpenRouter logo

Perplexity: Sonar Pro

OpenRouter · Mar 7, 2025

Note: Sonar Pro pricing includes Perplexity search pricing. See [details here](https://docs.perplexity.ai/guides/pricing#detailed-pricing-breakdown-for-sonar-reasoning-pro-and-sonar-pro) For enterprises seeking more advanced capabilities, the Sonar Pro API can handle in-depth, multi-step queries with added extensibility, like...

200K context
View
OpenRouter logo

Perplexity: Sonar Deep Research

OpenRouter · Mar 7, 2025

Sonar Deep Research is a research-focused model designed for multi-step retrieval, synthesis, and reasoning across complex topics. It autonomously searches, reads, and evaluates sources, refining its approach as it gathers...

128K context
View
OpenRouter logo

Qwen: QwQ 32B

OpenRouter · Mar 5, 2025

QwQ is the reasoning model of the Qwen series. Compared with conventional instruction-tuned models, QwQ, which is capable of thinking and reasoning, can achieve significantly enhanced performance in downstream tasks,...

131.1K context
View
OpenRouter logo

Google: Gemini 2.0 Flash Lite

OpenRouter · Feb 25, 2025

Gemini 2.0 Flash Lite offers a significantly faster time to first token (TTFT) compared to [Gemini Flash 1.5](/google/gemini-flash-1.5), while maintaining quality on par with larger models like [Gemini Pro 1.5](/google/gemini-pro-1.5),...

1M context+1
View
OpenRouter logo

Anthropic: Claude 3.7 Sonnet (thinking)

OpenRouter · Feb 24, 2025

Claude 3.7 Sonnet is an advanced large language model with improved reasoning, coding, and problem-solving capabilities. It introduces a hybrid reasoning approach, allowing users to choose between rapid responses and...

200K context
View
OpenRouter logo

Anthropic: Claude 3.7 Sonnet

OpenRouter · Feb 24, 2025

Claude 3.7 Sonnet is an advanced large language model with improved reasoning, coding, and problem-solving capabilities. It introduces a hybrid reasoning approach, allowing users to choose between rapid responses and...

200K context
View
OpenRouter logo

Mistral: Saba

OpenRouter · Feb 17, 2025

Mistral Saba is a 24B-parameter language model specifically designed for the Middle East and South Asia, delivering accurate and contextually relevant responses while maintaining efficient performance. Trained on curated regional...

32.8K context
View
OpenRouter logo

Llama Guard 3 8B

OpenRouter · Feb 12, 2025

Llama Guard 3 is a Llama-3.1-8B pretrained model, fine-tuned for content safety classification. Similar to previous versions, it can be used to classify content in both LLM inputs (prompt classification)...

131.1K context
View
OpenRouter logo

OpenAI: o3 Mini High

OpenRouter · Feb 12, 2025

OpenAI o3-mini-high is the same model as [o3-mini](/openai/o3-mini) with reasoning_effort set to high. o3-mini is a cost-efficient language model optimized for STEM reasoning tasks, particularly excelling in science, mathematics, and...

200K context
View
OpenRouter logo

OpenAI: o3 Mini High (batch)

OpenRouter · Feb 12, 2025

OpenAI o3-mini-high is the same model as [o3-mini](/openai/o3-mini) with reasoning_effort set to high. o3-mini is a cost-efficient language model optimized for STEM reasoning tasks, particularly excelling in science, mathematics, and...

200K context
View
OpenRouter logo

Google: Gemini 2.0 Flash

OpenRouter · Feb 5, 2025

Gemini Flash 2.0 offers a significantly faster time to first token (TTFT) compared to [Gemini Flash 1.5](/google/gemini-flash-1.5), while maintaining quality on par with larger models like [Gemini Pro 1.5](/google/gemini-pro-1.5). It...

1M context+2
View
OpenRouter logo

Qwen: Qwen VL Plus

OpenRouter · Feb 5, 2025

Qwen's Enhanced Large Visual Language Model. Significantly upgraded for detailed recognition capabilities and text recognition abilities, supporting ultra-high pixel resolutions up to millions of pixels and extreme aspect ratios for...

131.1K context
View
OpenRouter logo

AionLabs: Aion-1.0

OpenRouter · Feb 4, 2025

Aion-1.0 is a multi-model system designed for high performance across various tasks, including reasoning and coding. It is built on DeepSeek-R1, augmented with additional models and techniques such as Tree...

131.1K context
View
OpenRouter logo

AionLabs: Aion-1.0-Mini

OpenRouter · Feb 4, 2025

Aion-1.0-Mini 32B parameter model is a distilled version of the DeepSeek-R1 model, designed for strong performance in reasoning domains such as mathematics, coding, and logic. It is a modified variant...

131.1K context
View
OpenRouter logo

AionLabs: Aion-RP 1.0 (8B)

OpenRouter · Feb 4, 2025

Aion-RP-Llama-3.1-8B ranks the highest in the character evaluation portion of the RPBench-Auto benchmark, a roleplaying-specific variant of Arena-Hard-Auto, where LLMs evaluate each other’s responses. It is a fine-tuned base model...

32.8K context
View
OpenRouter logo

Qwen: Qwen VL Max

OpenRouter · Feb 1, 2025

Qwen VL Max is a visual understanding model with 7500 tokens context length. It excels in delivering optimal performance for a broader spectrum of complex tasks.

131.1K context
View
OpenRouter logo

Qwen: Qwen-Turbo

OpenRouter · Feb 1, 2025

Qwen-Turbo, based on Qwen2.5, is a 1M context model that provides fast speed and low cost, suitable for simple tasks.

131.1K context
View
OpenRouter logo

Qwen: Qwen2.5 VL 72B Instruct

OpenRouter · Feb 1, 2025

Qwen2.5-VL is proficient in recognizing common objects such as flowers, birds, fish, and insects. It is also highly capable of analyzing texts, charts, icons, graphics, and layouts within images.

128K context
View
OpenRouter logo

Qwen: Qwen-Plus

OpenRouter · Feb 1, 2025

Qwen-Plus, based on the Qwen2.5 foundation model, is a 131K context model with a balanced performance, speed, and cost combination.

1M context
View
OpenRouter logo

Qwen: Qwen-Max

OpenRouter · Feb 1, 2025

Qwen-Max, based on Qwen2.5, provides the best inference performance among [Qwen models](/qwen), especially for complex multi-step tasks. It's a large-scale MoE model that has been pretrained on over 20 trillion...

32.8K context
View
OpenRouter logo

OpenAI: o3 Mini

OpenRouter · Jan 31, 2025

OpenAI o3-mini is a cost-efficient language model optimized for STEM reasoning tasks, particularly excelling in science, mathematics, and coding. This model supports the `reasoning_effort` parameter, which can be set to...

200K context
View
OpenRouter logo

OpenAI: o3 Mini (batch)

OpenRouter · Jan 31, 2025

OpenAI o3-mini is a cost-efficient language model optimized for STEM reasoning tasks, particularly excelling in science, mathematics, and coding. This model supports the `reasoning_effort` parameter, which can be set to...

200K context
View
OpenRouter logo

Mistral: Mistral Small 3

OpenRouter · Jan 30, 2025

Mistral Small 3 is a 24B-parameter language model optimized for low-latency performance across common AI tasks. Released under the Apache 2.0 license, it features both pre-trained and instruction-tuned versions designed...

32.8K context
View
OpenRouter logo

DeepSeek: R1 Distill Qwen 32B

OpenRouter · Jan 29, 2025

DeepSeek R1 Distill Qwen 32B is a distilled large language model based on [Qwen 2.5 32B](https://huggingface.co/Qwen/Qwen2.5-32B), using outputs from [DeepSeek R1](/deepseek/deepseek-r1). It outperforms OpenAI's o1-mini across various benchmarks, achieving new...

128K context
View
OpenRouter logo

DeepSeek: R1 Distill Qwen 14B

OpenRouter · Jan 29, 2025

DeepSeek R1 Distill Qwen 14B is a distilled large language model based on [Qwen 2.5 14B](https://huggingface.co/deepseek-ai/DeepSeek-R1-Distill-Qwen-14B), using outputs from [DeepSeek R1](/deepseek/deepseek-r1). It outperforms OpenAI's o1-mini across various benchmarks, achieving new state-of-the-art results for dense models. Other benchmark results include: - AIME 2024 pass@1: 69.7 - MATH-500 pass@1: 93.9 - CodeForces Rating: 1481 The model leverages fine-tuning from DeepSeek R1's outputs, enabling competitive performance comparable to larger frontier models.

32.8K context
View
OpenRouter logo

Perplexity: Sonar

OpenRouter · Jan 27, 2025

Sonar is lightweight, affordable, fast, and simple to use — now featuring citations and the ability to customize sources. It is designed for companies seeking to integrate lightweight question-and-answer features...

127.1K context
View