Guide
2,496 step-by-step guides to setting up AI models from OpenAI, Anthropic, Google, OpenRouter, and more in TypingMind with your own API key.
Search and filter guides
Showing 1,429-1,479 of 2,496 guides

Grok 4.1 Fast Non-Reasoning
Vercel AI Gateway · Nov 19, 2025
Fast Grok model for responsive chat, reasoning, and tool-assisted work

Google: Gemini 3 Pro Preview
OpenRouter · Nov 18, 2025
Gemini 3 Pro is Google’s flagship frontier model for high-precision multimodal reasoning, combining strong performance across text, image, video, audio, and code with a 1M-token context window. Reasoning Details must be preserved when using multi-turn tool calling, see our docs here: https://openrouter.ai/docs/use-cases/reasoning-tokens#preserving-reasoning-blocks. It delivers state-of-the-art benchmark results in general reasoning, STEM problem solving, factual QA, and multimodal understanding, including leading scores on LMArena, GPQA Diamond, MathArena Apex, MMMU-Pro, and Video-MMMU. Interactions emphasize depth and interpretability: the model is designed to infer intent with minimal prompting and produce direct, insight-focused responses. Built for advanced development and agentic workflows, Gemini 3 Pro provides robust tool-calling, long-horizon planning stability, and strong zero-shot generation for complex UI, visualization, and coding tasks. It excels at agentic coding (SWE-Bench Verified, Terminal-Bench 2.0), multimodal analysis, and structured long-form tasks such as research synthesis, planning, and interactive learning experiences. Suitable applications include autonomous agents, coding assistants, multimodal analytics, scientific reasoning, and high-context information processing.

Gemini 3 Pro Preview
Google · Nov 18, 2025
Preview Gemini flagship for complex reasoning, coding, and rich multimodal prompts

Qwen3 Embedding 0.6B
Vercel AI Gateway · Nov 14, 2025
Embedding model for semantic search, retrieval, clustering, and ranking pipelines

Deep Cogito: Cogito v2.1 671B
OpenRouter · Nov 13, 2025
Cogito v2.1 671B MoE represents one of the strongest open models globally, matching performance of frontier closed and open models. This model is trained using self play with reinforcement learning...

OpenAI: GPT-5.1
OpenRouter · Nov 13, 2025
GPT-5.1 is the latest frontier-grade model in the GPT-5 series, offering stronger general-purpose reasoning, improved instruction adherence, and a more natural conversational style compared to GPT-5. It uses adaptive reasoning...

OpenAI: GPT-5.1 (batch)
OpenRouter · Nov 13, 2025
GPT-5.1 is the latest frontier-grade model in the GPT-5 series, offering stronger general-purpose reasoning, improved instruction adherence, and a more natural conversational style compared to GPT-5. It uses adaptive reasoning...

OpenAI: GPT-5.1 Chat
OpenRouter · Nov 13, 2025
GPT-5.1 Chat (AKA Instant is the fast, lightweight member of the 5.1 family, optimized for low-latency chat while retaining strong general intelligence. It uses adaptive reasoning to selectively “think” on...

OpenAI: GPT-5.1-Codex
OpenRouter · Nov 13, 2025
GPT-5.1-Codex is a specialized version of GPT-5.1 optimized for software engineering and coding workflows. It is designed for both interactive development sessions and long, independent execution of complex engineering tasks....

OpenAI: GPT-5.1-Codex-Mini
OpenRouter · Nov 13, 2025
GPT-5.1-Codex-Mini is a smaller and faster version of GPT-5.1-Codex

GPT-5.1 Codex
OpenAI · Nov 13, 2025
GPT-5.1 Codex from OpenAI - text, image input, 400,000 token context

GPT-5.1 Codex mini
OpenAI · Nov 13, 2025
GPT-5.1 Codex mini from OpenAI - text, image input, 400,000 token context

GPT-5.1
OpenAI · Nov 13, 2025
Sharper GPT-5 generation for coding, product work, and tool-assisted tasks

GPT-5.1 Codex Max
OpenAI · Nov 13, 2025
GPT-5.1 Codex Max from OpenAI - text, image input, 400,000 token context

GPT-5.1 Chat
OpenAI · Nov 13, 2025
GPT-5.1 Chat from OpenAI - text, image input, 128,000 token context

MiniMax M2
DeepInfra · Nov 13, 2025
MiniMax M2 from DeepInfra - text input, 262,144 token context

Cogito v2.1 671B
Together AI · Nov 13, 2025
Reasoning model for deliberate analysis, multi-step problem solving, and tool use

GPT-5.1-Codex
Vercel AI Gateway · Nov 13, 2025
Codex GPT for repository edits, code review, and practical software agents

GPT 5.1 Codex Max
Vercel AI Gateway · Nov 13, 2025
Coding-optimized GPT model for repository edits, reviews, and agentic software work

GPT-5.1 Codex mini
Vercel AI Gateway · Nov 13, 2025
Coding-optimized GPT model for repository edits, reviews, and agentic software work

GPT-5.1
Cloudflare AI Gateway · Nov 13, 2025
Sharper GPT-5 generation for coding, product work, and tool-assisted tasks

GPT 5.1 Thinking (Fast)
Vercel AI Gateway · Nov 12, 2025
Compact GPT model for low-latency assistance and high-volume workloads

GPT 5.1 Thinking
Vercel AI Gateway · Nov 12, 2025
Image model for prompt-driven generation, editing, and visual design workflows

Kwaipilot: KAT-Coder-Pro V1 (free)
OpenRouter · Nov 10, 2025
KAT-Coder-Pro V1 is KwaiKAT's most advanced agentic coding model in the KAT-Coder series. Designed specifically for agentic coding tasks, it excels in real-world software engineering scenarios, achieving 73.4% solve rate on the SWE-Bench Verified benchmark. The model has been optimized for tool-use capability, multi-turn interaction, instruction following, generalization, and comprehensive capabilities through a multi-stage training process, including mid-training, supervised fine-tuning (SFT), reinforcement fine-tuning (RFT), and scalable agentic RL.

Kwaipilot: KAT-Coder-Pro V1
OpenRouter · Nov 10, 2025
KAT-Coder-Pro V1 is KwaiKAT's most advanced agentic coding model in the KAT-Coder series. Designed specifically for agentic coding tasks, it excels in real-world software engineering scenarios, achieving 73.4% solve rate on the SWE-Bench Verified benchmark. The model has been optimized for tool-use capability, multi-turn interaction, instruction following, generalization, and comprehensive capabilities through a multi-stage training process, including mid-training, supervised fine-tuning (SFT), reinforcement fine-tuning (RFT), and scalable agentic RL.

Kimi K2 Thinking
Synthetic · Nov 7, 2025
Kimi K2 Thinking from Synthetic - text input, 262,144 token context

Kimi K2 Thinking
Novita · Nov 7, 2025
Kimi reasoning model for long-horizon research, planning, and tool use

MoonshotAI: Kimi K2 Thinking
OpenRouter · Nov 6, 2025
Kimi K2 Thinking is Moonshot AI’s most advanced open reasoning model to date, extending the K2 series into agentic, long-horizon reasoning. Built on the trillion-parameter Mixture-of-Experts (MoE) architecture introduced in...

Kimi K2 Thinking
Fireworks AI · Nov 6, 2025
Kimi K2 Thinking from Fireworks AI - text input, 256,000 token context

Kimi K2 Thinking Turbo
Moonshot AI · Nov 6, 2025
Kimi reasoning model for long-horizon research, planning, and tool use

Kimi K2 Thinking
Moonshot AI · Nov 6, 2025
Thinking Kimi model for slower research passes, planning, and hard technical questions

OpenAI GPT OSS 120B
Venice AI · Nov 6, 2025
Open GPT reasoning model for self-hosted agents and controllable deployments

Kimi K2 Thinking
AWS Bedrock · Nov 6, 2025
Thinking Kimi model for slower research passes, planning, and hard technical questions

Kimi K2 Thinking
DeepInfra · Nov 6, 2025
Kimi K2 Thinking from DeepInfra - text input, 131,072 token context
Kimi-K2-Thinking
Hugging Face · Nov 6, 2025
Kimi reasoning model for long-horizon research, planning, and tool use

Kimi K2 Thinking
Vercel AI Gateway · Nov 6, 2025
Thinking Kimi model for slower research passes, planning, and hard technical questions

Google Gemma 3 27B Instruct
Venice AI · Nov 4, 2025
Open Gemma instruction model for efficient chat and self-hosted deployments

Claude Opus 4.5 (Global)
AWS Bedrock · Nov 1, 2025
Flagship Claude model for deep reasoning, coding, and long-horizon agents

Claude Opus 4.5 (US)
AWS Bedrock · Nov 1, 2025
Flagship Claude model for deep reasoning, coding, and long-horizon agents

Claude Opus 4.5 (EU)
AWS Bedrock · Nov 1, 2025
Flagship Claude model for deep reasoning, coding, and long-horizon agents

Claude Opus 4.5
AWS Bedrock · Nov 1, 2025
Flagship Claude model for deep reasoning, coding, and long-horizon agents

Amazon: Nova Premier 1.0
OpenRouter · Oct 31, 2025
Amazon Nova Premier is the most capable of Amazon’s multimodal models for complex reasoning tasks and for use as the best teacher for distilling custom models.

Perplexity: Sonar Pro Search
OpenRouter · Oct 30, 2025
Exclusively available on the OpenRouter API, Sonar Pro's new Pro Search mode is Perplexity's most advanced agentic search system. It is designed for deeper reasoning and analysis. Pricing is based...

Mistral: Voxtral Small 24B 2507
OpenRouter · Oct 30, 2025
Voxtral Small is an enhancement of Mistral Small 3, incorporating state-of-the-art audio input capabilities while retaining best-in-class text performance. It excels at speech transcription, translation and audio understanding. Input audio...

OpenAI: gpt-oss-safeguard-20b
OpenRouter · Oct 29, 2025
gpt-oss-safeguard-20b is a safety reasoning model from OpenAI built upon gpt-oss-20b. This open-weight, 21B-parameter Mixture-of-Experts (MoE) model offers lower latency for safety tasks like content classification, LLM filtering, and trust...

Safety GPT OSS 20B
Groq · Oct 29, 2025
Safety model for policy screening, moderation, and risk-aware routing workflows

GPT OSS Safeguard 20B
AWS Bedrock · Oct 29, 2025
Safety model for policy screening, moderation, and risk-aware routing workflows

GPT OSS Safeguard 120B
AWS Bedrock · Oct 29, 2025
Safety model for policy screening, moderation, and risk-aware routing workflows

GPT OSS Safeguard 120B
Vercel AI Gateway · Oct 29, 2025
Safety model for policy screening, moderation, and risk-aware routing workflows

gpt-oss-safeguard-20b
Vercel AI Gateway · Oct 29, 2025
Safety model for policy screening, moderation, and risk-aware routing workflows

NVIDIA: Nemotron Nano 12B 2 VL (free)
OpenRouter · Oct 28, 2025
NVIDIA Nemotron Nano 2 VL is a 12-billion-parameter open multimodal reasoning model designed for video understanding and document intelligence. It introduces a hybrid Transformer-Mamba architecture, combining transformer-level accuracy with Mamba’s...

