Guide

2,496 step-by-step guides to setting up AI models from OpenAI, Anthropic, Google, OpenRouter, and more in TypingMind with your own API key.

Search and filter guides

Showing 1,837-1,887 of 2,496 guides

DeepInfra logo

Kimi K2

DeepInfra · Jul 11, 2025

Kimi K2 from DeepInfra - text input, 131,072 token context

131.1K context
View
Vercel AI Gateway logo

Kimi K2 Instruct

Vercel AI Gateway · Jul 11, 2025

Kimi model for long-context chat, coding, and agentic reasoning

131.1K context
View
Novita logo

Kimi K2 Instruct

Novita · Jul 11, 2025

Kimi model for long-context chat, coding, and agentic reasoning

131.1K context
View
OpenRouter logo

Mistral: Devstral Medium

OpenRouter · Jul 10, 2025

Devstral Medium is a high-performance code generation and agentic reasoning model developed jointly by Mistral AI and All Hands AI. Positioned as a step up from Devstral Small, it achieves...

131.1K context
View
OpenRouter logo

Mistral: Devstral Small 1.1

OpenRouter · Jul 10, 2025

Devstral Small 1.1 is a 24B parameter open-weight language model for software engineering agents, developed by Mistral AI in collaboration with All Hands AI. Finetuned from Mistral Small 3.1 and...

131.1K context
View
Mistral logo

Devstral Medium

Mistral · Jul 10, 2025

Legacy model retained for compatibility with older integrations

128K context
View
Mistral logo

Devstral Small

Mistral · Jul 10, 2025

Legacy model retained for compatibility with older integrations

128K context
View
OpenRouter logo

Venice: Uncensored (free)

OpenRouter · Jul 9, 2025

Venice Uncensored Dolphin Mistral 24B Venice Edition is a fine-tuned variant of Mistral-Small-24B-Instruct-2501, developed by dphn.ai in collaboration with Venice.ai. This model is designed as an “uncensored” instruct-tuned LLM, preserving...

32.8K context
View
OpenRouter logo

Venice: Uncensored

OpenRouter · Jul 9, 2025

Venice Uncensored Dolphin Mistral 24B Venice Edition is a fine-tuned variant of Mistral-Small-24B-Instruct-2501, developed by dphn.ai in collaboration with Venice.ai. This model is designed as an “uncensored” instruct-tuned LLM, preserving...

128K context
View
OpenRouter logo

xAI: Grok 4

OpenRouter · Jul 9, 2025

Grok 4 is xAI's latest reasoning model with a 256k context window. It supports parallel tool calling, structured outputs, and both image and text inputs. Note that reasoning is not...

256K context+1
View
OpenRouter logo

Google: Gemma 3n 2B (free)

OpenRouter · Jul 9, 2025

Gemma 3n E2B IT is a multimodal, instruction-tuned model developed by Google DeepMind, designed to operate efficiently at an effective parameter size of 2B while leveraging a 6B architecture. Based...

8.2K context
View
xAI logo

Grok 4

xAI · Jul 9, 2025

Grok 4 from xAI - text input, 256,000 token context

256K context
View
Google logo

Gemma 3n 2B

Google · Jul 9, 2025

Gemma 3n 2B from Google - text input, 8,192 token context

8.2K context
View
OpenRouter logo

Tencent: Hunyuan A13B Instruct

OpenRouter · Jul 8, 2025

Hunyuan-A13B is a 13B active parameter Mixture-of-Experts (MoE) language model developed by Tencent, with a total parameter count of 80B and support for reasoning via Chain-of-Thought. It offers competitive benchmark...

131.1K context
View
OpenRouter logo

TNG: DeepSeek R1T2 Chimera (free)

OpenRouter · Jul 8, 2025

DeepSeek-TNG-R1T2-Chimera is the second-generation Chimera model from TNG Tech. It is a 671 B-parameter mixture-of-experts text-generation model assembled from DeepSeek-AI’s R1-0528, R1, and V3-0324 checkpoints with an Assembly-of-Experts merge. The tri-parent design yields strong reasoning performance while running roughly 20 % faster than the original R1 and more than 2× faster than R1-0528 under vLLM, giving a favorable cost-to-intelligence trade-off. The checkpoint supports contexts up to 60 k tokens in standard use (tested to ~130 k) and maintains consistent <think> token behaviour, making it suitable for long-context analysis, dialogue and other open-ended generation tasks.

163.8K context
View
OpenRouter logo

TNG: DeepSeek R1T2 Chimera

OpenRouter · Jul 8, 2025

DeepSeek-TNG-R1T2-Chimera is the second-generation Chimera model from TNG Tech. It is a 671 B-parameter mixture-of-experts text-generation model assembled from DeepSeek-AI’s R1-0528, R1, and V3-0324 checkpoints with an Assembly-of-Experts merge. The...

163.8K context
View
OpenRouter logo

Morph: Morph V3 Large

OpenRouter · Jul 7, 2025

Morph's high-accuracy apply model for complex code edits. ~4,500 tokens/sec with 98% accuracy for precise code transformations. The model requires the prompt to be in the following format: <instruction>{instruction}</instruction> <code>{initial_code}</code>...

262.1K context
View
OpenRouter logo

Morph: Morph V3 Fast

OpenRouter · Jul 7, 2025

Morph's fastest apply model for code edits. ~10,500 tokens/sec with 96% accuracy for rapid code transformations. The model requires the prompt to be in the following format: <instruction>{instruction}</instruction> <code>{initial_code}</code> <update>{edit_snippet}</update>...

81.9K context
View
Chutes logo

Qwen3 235B A22B Thinking 2507 TEE

Chutes · Jul 1, 2025

Qwen reasoning model for deliberate problem solving, math, and coding

262.1K context
View
OpenRouter logo

Baidu: ERNIE 4.5 VL 424B A47B

OpenRouter · Jun 30, 2025

ERNIE-4.5-VL-424B-A47B is a multimodal Mixture-of-Experts (MoE) model from Baidu’s ERNIE 4.5 series, featuring 424B total parameters with 47B active per token. It is trained jointly on text and image data...

123K context
View
OpenRouter logo

Baidu: ERNIE 4.5 300B A47B

OpenRouter · Jun 30, 2025

ERNIE-4.5-300B-A47B is a 300B parameter Mixture-of-Experts (MoE) language model developed by Baidu as part of the ERNIE 4.5 series. It activates 47B parameters per token and supports text generation in...

131.1K context
View
Novita logo

ERNIE 4.5 VL 28B A3B

Novita · Jun 30, 2025

Multimodal reasoning model for visual analysis, planning, and tool use

30K context
View
Novita logo

ERNIE 4.5 21B A3B

Novita · Jun 30, 2025

Open-weight instruction model for adaptable chat and self-hosted production workloads

120K context
View
Novita logo

ERNIE 4.5 VL 424B A47B

Novita · Jun 30, 2025

Multimodal reasoning model for visual analysis, planning, and tool use

123K context
View
Novita logo

ERNIE 4.5 300B A47B

Novita · Jun 30, 2025

Open-weight instruction model for adaptable chat and self-hosted production workloads

123K context
View
OpenRouter logo

Inception: Mercury

OpenRouter · Jun 26, 2025

Mercury is the first diffusion large language model (dLLM). Applying a breakthrough discrete diffusion approach, the model runs 5-10x faster than even speed optimized models like GPT-4.1 Nano and Claude...

128K context
View
OpenRouter logo

Mistral: Mistral Small 3.2 24B

OpenRouter · Jun 20, 2025

Mistral-Small-3.2-24B-Instruct-2506 is an updated 24B parameter model from Mistral optimized for instruction following, repetition reduction, and improved function calling. Compared to the 3.1 release, version 3.2 significantly improves accuracy on...

256K context
View
Mistral logo

Mistral Small 3.2

Mistral · Jun 20, 2025

Efficient Mistral model for fast chat, extraction, and production assistants

128K context
View
OpenRouter logo

MiniMax: MiniMax M1

OpenRouter · Jun 17, 2025

MiniMax-M1 is a large-scale, open-weight reasoning model designed for extended context and high-efficiency inference. It leverages a hybrid Mixture-of-Experts (MoE) architecture paired with a custom "lightning attention" mechanism, allowing it...

1M context
View
OpenRouter logo

Google: Gemini 2.5 Flash

OpenRouter · Jun 17, 2025

Gemini 2.5 Flash is Google's state-of-the-art workhorse model, specifically designed for advanced reasoning, coding, mathematics, and scientific tasks. It includes built-in "thinking" capabilities, enabling it to provide responses with greater...

1M context+3
View
OpenRouter logo

Google: Gemini 2.5 Flash (batch)

OpenRouter · Jun 17, 2025

Gemini 2.5 Flash is Google's state-of-the-art workhorse model, specifically designed for advanced reasoning, coding, mathematics, and scientific tasks. It includes built-in "thinking" capabilities, enabling it to provide responses with greater...

1M context+3
View
OpenRouter logo

Google: Gemini 2.5 Pro

OpenRouter · Jun 17, 2025

Gemini 2.5 Pro is Google’s state-of-the-art AI model designed for advanced reasoning, coding, mathematics, and scientific tasks. It employs “thinking” capabilities, enabling it to reason through responses with enhanced accuracy...

1M context+3
View
OpenRouter logo

Google: Gemini 2.5 Pro (batch)

OpenRouter · Jun 17, 2025

Gemini 2.5 Pro is Google’s state-of-the-art AI model designed for advanced reasoning, coding, mathematics, and scientific tasks. It employs “thinking” capabilities, enabling it to reason through responses with enhanced accuracy...

1M context+3
View
Google logo

Gemini 2.5 Flash

Google · Jun 17, 2025

Fast Gemini workhorse for multimodal apps where latency and price matter

1M context+3
View
Google logo

Gemini Live 2.5 Flash Preview Native Audio

Google · Jun 17, 2025

Gemini Live 2.5 Flash Preview Native Audio from Google - text, audio, video input, 131,072 token context

131.1K context
View
Google logo

Gemini 2.5 Flash-Lite

Google · Jun 17, 2025

Lean Gemini 2.5 lane for cheap multimodal traffic and quick agents

1M context+3
View
Google logo

Gemini 2.5 Flash Lite Preview 06-17

Google · Jun 17, 2025

Gemini 2.5 Flash Lite Preview 06-17 from Google - text, image, audio, video, pdf input, 1,048,576 token context

1M context+2
View
Google logo

Gemini 2.5 Pro

Google · Jun 17, 2025

Google's proven reasoning model for coding, math, and multimodal analysis

1M context+3
View
Vercel AI Gateway logo

Gemini 2.5 Flash Lite

Vercel AI Gateway · Jun 17, 2025

Lean Gemini 2.5 lane for cheap multimodal traffic and quick agents

1M context+1
View
Vercel AI Gateway logo

Gemini 2.5 Flash

Vercel AI Gateway · Jun 17, 2025

Fast Gemini workhorse for multimodal apps where latency and price matter

1M context+3
View
Vercel AI Gateway logo

Gemini 2.5 Pro

Vercel AI Gateway · Jun 17, 2025

Google's proven reasoning model for coding, math, and multimodal analysis

1M context+3
View
Novita logo

MiniMax M1

Novita · Jun 17, 2025

MiniMax model for chat, coding, office work, and agentic tasks

1M context
View
OpenRouter logo

MoonshotAI: Kimi Dev 72B

OpenRouter · Jun 16, 2025

Kimi-Dev-72B is an open-source large language model fine-tuned for software engineering and issue resolution tasks. Based on Qwen2.5-72B, it is optimized using large-scale reinforcement learning that applies code patches in real repositories and validates them via full test suite execution—rewarding only correct, robust completions. The model achieves 60.4% on SWE-bench Verified, setting a new benchmark among open-source models for software bug fixing and code reasoning.

131.1K context
View
DeepInfra logo

Claude Opus 4

DeepInfra · Jun 12, 2025

Claude Opus 4 from DeepInfra - text, image input, 200,000 token context

200K context
View
Groq logo

Qwen3-32B

Groq · Jun 11, 2025

Qwen instruction model for multilingual chat, reasoning, and tool use

131.1K context
View
Vercel AI Gateway logo

Seedance v1.0 Pro

Vercel AI Gateway · Jun 11, 2025

Video model for prompt-guided generation, editing, and motion workflows

View
OpenRouter logo

OpenAI: o3 Pro

OpenRouter · Jun 10, 2025

The o-series of models are trained with reinforcement learning to think before they answer and perform complex reasoning. The o3-pro model uses more compute to think harder and provide consistently...

200K context
View
OpenRouter logo

OpenAI: o3 Pro (batch)

OpenRouter · Jun 10, 2025

The o-series of models are trained with reinforcement learning to think before they answer and perform complex reasoning. The o3-pro model uses more compute to think harder and provide consistently...

200K context
View
OpenRouter logo

xAI: Grok 3 Mini

OpenRouter · Jun 10, 2025

A lightweight model that thinks before responding. Fast, smart, and great for logic-based tasks that do not require deep domain knowledge. The raw thinking traces are accessible.

131.1K context
View
OpenRouter logo

xAI: Grok 3

OpenRouter · Jun 10, 2025

Grok 3 is the latest model from xAI. It's their flagship model that excels at enterprise use cases like data extraction, coding, and text summarization. Possesses deep domain knowledge in...

131.1K context
View
OpenAI logo

o3-pro

OpenAI · Jun 10, 2025

High-effort o3 tier for difficult technical reasoning and careful answers

200K context
View