Guide
2,496 step-by-step guides to setting up AI models from OpenAI, Anthropic, Google, OpenRouter, and more in TypingMind with your own API key.
Search and filter guides
Showing 1,837-1,887 of 2,496 guides

Kimi K2
DeepInfra · Jul 11, 2025
Kimi K2 from DeepInfra - text input, 131,072 token context

Kimi K2 Instruct
Vercel AI Gateway · Jul 11, 2025
Kimi model for long-context chat, coding, and agentic reasoning

Kimi K2 Instruct
Novita · Jul 11, 2025
Kimi model for long-context chat, coding, and agentic reasoning

Mistral: Devstral Medium
OpenRouter · Jul 10, 2025
Devstral Medium is a high-performance code generation and agentic reasoning model developed jointly by Mistral AI and All Hands AI. Positioned as a step up from Devstral Small, it achieves...

Mistral: Devstral Small 1.1
OpenRouter · Jul 10, 2025
Devstral Small 1.1 is a 24B parameter open-weight language model for software engineering agents, developed by Mistral AI in collaboration with All Hands AI. Finetuned from Mistral Small 3.1 and...

Devstral Medium
Mistral · Jul 10, 2025
Legacy model retained for compatibility with older integrations

Devstral Small
Mistral · Jul 10, 2025
Legacy model retained for compatibility with older integrations

Venice: Uncensored (free)
OpenRouter · Jul 9, 2025
Venice Uncensored Dolphin Mistral 24B Venice Edition is a fine-tuned variant of Mistral-Small-24B-Instruct-2501, developed by dphn.ai in collaboration with Venice.ai. This model is designed as an “uncensored” instruct-tuned LLM, preserving...

Venice: Uncensored
OpenRouter · Jul 9, 2025
Venice Uncensored Dolphin Mistral 24B Venice Edition is a fine-tuned variant of Mistral-Small-24B-Instruct-2501, developed by dphn.ai in collaboration with Venice.ai. This model is designed as an “uncensored” instruct-tuned LLM, preserving...

xAI: Grok 4
OpenRouter · Jul 9, 2025
Grok 4 is xAI's latest reasoning model with a 256k context window. It supports parallel tool calling, structured outputs, and both image and text inputs. Note that reasoning is not...

Google: Gemma 3n 2B (free)
OpenRouter · Jul 9, 2025
Gemma 3n E2B IT is a multimodal, instruction-tuned model developed by Google DeepMind, designed to operate efficiently at an effective parameter size of 2B while leveraging a 6B architecture. Based...

Grok 4
xAI · Jul 9, 2025
Grok 4 from xAI - text input, 256,000 token context

Gemma 3n 2B
Google · Jul 9, 2025
Gemma 3n 2B from Google - text input, 8,192 token context

Tencent: Hunyuan A13B Instruct
OpenRouter · Jul 8, 2025
Hunyuan-A13B is a 13B active parameter Mixture-of-Experts (MoE) language model developed by Tencent, with a total parameter count of 80B and support for reasoning via Chain-of-Thought. It offers competitive benchmark...

TNG: DeepSeek R1T2 Chimera (free)
OpenRouter · Jul 8, 2025
DeepSeek-TNG-R1T2-Chimera is the second-generation Chimera model from TNG Tech. It is a 671 B-parameter mixture-of-experts text-generation model assembled from DeepSeek-AI’s R1-0528, R1, and V3-0324 checkpoints with an Assembly-of-Experts merge. The tri-parent design yields strong reasoning performance while running roughly 20 % faster than the original R1 and more than 2× faster than R1-0528 under vLLM, giving a favorable cost-to-intelligence trade-off. The checkpoint supports contexts up to 60 k tokens in standard use (tested to ~130 k) and maintains consistent <think> token behaviour, making it suitable for long-context analysis, dialogue and other open-ended generation tasks.

TNG: DeepSeek R1T2 Chimera
OpenRouter · Jul 8, 2025
DeepSeek-TNG-R1T2-Chimera is the second-generation Chimera model from TNG Tech. It is a 671 B-parameter mixture-of-experts text-generation model assembled from DeepSeek-AI’s R1-0528, R1, and V3-0324 checkpoints with an Assembly-of-Experts merge. The...

Morph: Morph V3 Large
OpenRouter · Jul 7, 2025
Morph's high-accuracy apply model for complex code edits. ~4,500 tokens/sec with 98% accuracy for precise code transformations. The model requires the prompt to be in the following format: <instruction>{instruction}</instruction> <code>{initial_code}</code>...

Morph: Morph V3 Fast
OpenRouter · Jul 7, 2025
Morph's fastest apply model for code edits. ~10,500 tokens/sec with 96% accuracy for rapid code transformations. The model requires the prompt to be in the following format: <instruction>{instruction}</instruction> <code>{initial_code}</code> <update>{edit_snippet}</update>...

Qwen3 235B A22B Thinking 2507 TEE
Chutes · Jul 1, 2025
Qwen reasoning model for deliberate problem solving, math, and coding

Baidu: ERNIE 4.5 VL 424B A47B
OpenRouter · Jun 30, 2025
ERNIE-4.5-VL-424B-A47B is a multimodal Mixture-of-Experts (MoE) model from Baidu’s ERNIE 4.5 series, featuring 424B total parameters with 47B active per token. It is trained jointly on text and image data...

Baidu: ERNIE 4.5 300B A47B
OpenRouter · Jun 30, 2025
ERNIE-4.5-300B-A47B is a 300B parameter Mixture-of-Experts (MoE) language model developed by Baidu as part of the ERNIE 4.5 series. It activates 47B parameters per token and supports text generation in...

ERNIE 4.5 VL 28B A3B
Novita · Jun 30, 2025
Multimodal reasoning model for visual analysis, planning, and tool use

ERNIE 4.5 21B A3B
Novita · Jun 30, 2025
Open-weight instruction model for adaptable chat and self-hosted production workloads

ERNIE 4.5 VL 424B A47B
Novita · Jun 30, 2025
Multimodal reasoning model for visual analysis, planning, and tool use

ERNIE 4.5 300B A47B
Novita · Jun 30, 2025
Open-weight instruction model for adaptable chat and self-hosted production workloads

Inception: Mercury
OpenRouter · Jun 26, 2025
Mercury is the first diffusion large language model (dLLM). Applying a breakthrough discrete diffusion approach, the model runs 5-10x faster than even speed optimized models like GPT-4.1 Nano and Claude...

Mistral: Mistral Small 3.2 24B
OpenRouter · Jun 20, 2025
Mistral-Small-3.2-24B-Instruct-2506 is an updated 24B parameter model from Mistral optimized for instruction following, repetition reduction, and improved function calling. Compared to the 3.1 release, version 3.2 significantly improves accuracy on...

Mistral Small 3.2
Mistral · Jun 20, 2025
Efficient Mistral model for fast chat, extraction, and production assistants

MiniMax: MiniMax M1
OpenRouter · Jun 17, 2025
MiniMax-M1 is a large-scale, open-weight reasoning model designed for extended context and high-efficiency inference. It leverages a hybrid Mixture-of-Experts (MoE) architecture paired with a custom "lightning attention" mechanism, allowing it...

Google: Gemini 2.5 Flash
OpenRouter · Jun 17, 2025
Gemini 2.5 Flash is Google's state-of-the-art workhorse model, specifically designed for advanced reasoning, coding, mathematics, and scientific tasks. It includes built-in "thinking" capabilities, enabling it to provide responses with greater...

Google: Gemini 2.5 Flash (batch)
OpenRouter · Jun 17, 2025
Gemini 2.5 Flash is Google's state-of-the-art workhorse model, specifically designed for advanced reasoning, coding, mathematics, and scientific tasks. It includes built-in "thinking" capabilities, enabling it to provide responses with greater...

Google: Gemini 2.5 Pro
OpenRouter · Jun 17, 2025
Gemini 2.5 Pro is Google’s state-of-the-art AI model designed for advanced reasoning, coding, mathematics, and scientific tasks. It employs “thinking” capabilities, enabling it to reason through responses with enhanced accuracy...

Google: Gemini 2.5 Pro (batch)
OpenRouter · Jun 17, 2025
Gemini 2.5 Pro is Google’s state-of-the-art AI model designed for advanced reasoning, coding, mathematics, and scientific tasks. It employs “thinking” capabilities, enabling it to reason through responses with enhanced accuracy...

Gemini 2.5 Flash
Google · Jun 17, 2025
Fast Gemini workhorse for multimodal apps where latency and price matter

Gemini Live 2.5 Flash Preview Native Audio
Google · Jun 17, 2025
Gemini Live 2.5 Flash Preview Native Audio from Google - text, audio, video input, 131,072 token context

Gemini 2.5 Flash-Lite
Google · Jun 17, 2025
Lean Gemini 2.5 lane for cheap multimodal traffic and quick agents

Gemini 2.5 Flash Lite Preview 06-17
Google · Jun 17, 2025
Gemini 2.5 Flash Lite Preview 06-17 from Google - text, image, audio, video, pdf input, 1,048,576 token context

Gemini 2.5 Pro
Google · Jun 17, 2025
Google's proven reasoning model for coding, math, and multimodal analysis

Gemini 2.5 Flash Lite
Vercel AI Gateway · Jun 17, 2025
Lean Gemini 2.5 lane for cheap multimodal traffic and quick agents

Gemini 2.5 Flash
Vercel AI Gateway · Jun 17, 2025
Fast Gemini workhorse for multimodal apps where latency and price matter

Gemini 2.5 Pro
Vercel AI Gateway · Jun 17, 2025
Google's proven reasoning model for coding, math, and multimodal analysis

MiniMax M1
Novita · Jun 17, 2025
MiniMax model for chat, coding, office work, and agentic tasks

MoonshotAI: Kimi Dev 72B
OpenRouter · Jun 16, 2025
Kimi-Dev-72B is an open-source large language model fine-tuned for software engineering and issue resolution tasks. Based on Qwen2.5-72B, it is optimized using large-scale reinforcement learning that applies code patches in real repositories and validates them via full test suite execution—rewarding only correct, robust completions. The model achieves 60.4% on SWE-bench Verified, setting a new benchmark among open-source models for software bug fixing and code reasoning.

Claude Opus 4
DeepInfra · Jun 12, 2025
Claude Opus 4 from DeepInfra - text, image input, 200,000 token context

Qwen3-32B
Groq · Jun 11, 2025
Qwen instruction model for multilingual chat, reasoning, and tool use

Seedance v1.0 Pro
Vercel AI Gateway · Jun 11, 2025
Video model for prompt-guided generation, editing, and motion workflows

OpenAI: o3 Pro
OpenRouter · Jun 10, 2025
The o-series of models are trained with reinforcement learning to think before they answer and perform complex reasoning. The o3-pro model uses more compute to think harder and provide consistently...

OpenAI: o3 Pro (batch)
OpenRouter · Jun 10, 2025
The o-series of models are trained with reinforcement learning to think before they answer and perform complex reasoning. The o3-pro model uses more compute to think harder and provide consistently...

xAI: Grok 3 Mini
OpenRouter · Jun 10, 2025
A lightweight model that thinks before responding. Fast, smart, and great for logic-based tasks that do not require deep domain knowledge. The raw thinking traces are accessible.

xAI: Grok 3
OpenRouter · Jun 10, 2025
Grok 3 is the latest model from xAI. It's their flagship model that excels at enterprise use cases like data extraction, coding, and text summarization. Possesses deep domain knowledge in...

o3-pro
OpenAI · Jun 10, 2025
High-effort o3 tier for difficult technical reasoning and careful answers

