Guide
2,496 step-by-step guides to setting up AI models from OpenAI, Anthropic, Google, OpenRouter, and more in TypingMind with your own API key.
Search and filter guides
Showing 1-51 of 87 guides

MiMo-V2.6-Pro
DeepInfra · Sep 22, 2026
Stronger MiMo Pro tier for multimodal reasoning and coding-agent execution

MiMo-V2.6-Flash
DeepInfra · Sep 22, 2026
MiMo Flash model for multimodal coding agents and long-context automation

DeepSeek V4.1 Flash
DeepInfra · Sep 10, 2026
DeepSeek V4.1 Flash model for reasoning and agentic coding

Hy4 preview
DeepInfra · Aug 28, 2026
A next-generation productivity model with significantly enhanced Agent and complex task execution capabilities.

GLM-5.3-Flash
DeepInfra · Aug 26, 2026
Native multimodal GLM model for efficient coding and long-horizon agent tasks

Qwen3.8 Flash
DeepInfra · Aug 26, 2026
Qwen vision-language model for visual reasoning, documents, and agent tasks

DeepSeek V4 Flash Vision Exp
DeepInfra · Aug 21, 2026
Experimental multimodal DeepSeek V4 Flash model for image understanding, coding, and agentic work

GLM-5.3
DeepInfra · Aug 14, 2026
Flagship GLM model for long-horizon coding, agents, and complex project delivery

Qwen3.8 27B
DeepInfra · Aug 14, 2026
Dense 27B vision-language model for coding, agent tasks, and image and video understanding

DeepSeek V4 Pro 0813
DeepInfra · Aug 12, 2026
DeepSeek V4 Pro snapshot with million-token context and support for thinking and non-thinking modes

Qwen3.8 2.4T A95B
DeepInfra · Aug 12, 2026
Open-weight sparse MoE (2.4T total, 95B active), the open-weight twin of Qwen3.8 Max for coding, research, complex reasoning, and agentic workflows

Qwen3.8 Max
DeepInfra · Aug 3, 2026
2.4-trillion-parameter MoE flagship for coding, professional work, multimodal understanding, and long-horizon agentic workflows

DeepSeek V4 Flash 0731
DeepInfra · Jul 31, 2026
Official DeepSeek V4 Flash release with enhanced agentic capabilities and integrated DSpark speculative decoding

Inkling Small
DeepInfra · Jul 30, 2026
Multimodal MoE reasoning model (276B total, 12B active) for text, image, and audio

Kimi K3
DeepInfra · Jul 16, 2026
Multimodal Kimi model with 1M context and toggleable max-effort thinking for long-horizon agent work

Inkling
DeepInfra · Jul 15, 2026
Multimodal MoE reasoning model (975B total, 41B active) for text, image, and audio

Hy3
DeepInfra · Jul 6, 2026
Tencent Hy reasoning model for coding, instruction following, and agent tasks

GLM-5.2
DeepInfra · Jun 13, 2026
Open flagship GLM for long-horizon coding agents and million-token context work

Kimi K2.7 Code
DeepInfra · Jun 12, 2026
Coding-focused Kimi model, stronger on long-horizon repo work with less overthinking

MiniMax-M3
DeepInfra · Jun 1, 2026
MiniMax multimodal model for long-context coding, perception, and agent planning

Step 3.7 Flash
DeepInfra · May 29, 2026
Newer StepFun flash model for faster agents, coding, and multimodal prompts

Qwen3.7 Max
DeepInfra · May 21, 2026
Qwen frontier model tuned for agent frameworks, coding assistants, and long tasks

Nemotron 3 Nano Omni 30B A3B Reasoning
DeepInfra · Apr 28, 2026
Open Nemotron omni model combining reasoning with text, vision, and audio

DeepSeek V4 Flash
DeepInfra · Apr 24, 2026
Fast DeepSeek V4 lane for economical reasoning, coding, and long-context work

DeepSeek V4 Pro
DeepInfra · Apr 24, 2026
Open MoE flagship with million-token context for coding and long agent runs

MiMo-V2.5-Pro
DeepInfra · Apr 22, 2026
MiMo-V2.5-Pro from DeepInfra - text input, 1,048,576 token context

MiMo-V2.5
DeepInfra · Apr 22, 2026
MiMo-V2.5 from DeepInfra - text, image, audio, video input, 262,144 token context

MiMo-V2.5
DeepInfra · Apr 22, 2026
Open MiMo model for multimodal coding agents and long-context automation

MiMo-V2.5-Pro
DeepInfra · Apr 22, 2026
Stronger MiMo Pro tier for multimodal reasoning and coding-agent execution

Qwen3.6 27B
DeepInfra · Apr 22, 2026
Qwen vision-language model for visual reasoning, documents, and agent tasks

Kimi K2.6
DeepInfra · Apr 21, 2026
Kimi multimodal agent model for visual understanding, coding, and planning

GLM-5.1
DeepInfra · Apr 7, 2026
Flagship GLM model for hybrid reasoning, coding, and agentic engineering

Gemma 4 26B A4B IT
DeepInfra · Apr 2, 2026
Open Gemma instruction model for efficient chat and self-hosted deployments

Gemma 4 31B IT
DeepInfra · Apr 2, 2026
Largest Gemma 4 instruction model for open, self-hosted chat and reasoning

Gemma 4 E4B IT
DeepInfra · Apr 2, 2026
Open Gemma instruction model for efficient chat and self-hosted deployments

Qwen3.6 35B A3B
DeepInfra · Apr 1, 2026
Qwen vision-language model for visual reasoning, documents, and agent tasks

MiniMax-M2.7
DeepInfra · Mar 18, 2026
Open MiniMax flagship for coding agents, office automation, and complex environments

Qwen3.5 27B
DeepInfra · Feb 23, 2026
Qwen vision-language model for visual reasoning, documents, and agent tasks

Qwen3.5 9B
DeepInfra · Feb 23, 2026
Qwen instruction model for multilingual chat, reasoning, and tool use

Qwen3.5 122B-A10B
DeepInfra · Feb 23, 2026
Qwen vision-language model for visual reasoning, documents, and agent tasks

Seed 2.0 Code
DeepInfra · Feb 14, 2026
ByteDance Seed coding model for multimodal software engineering and long-running agents

Seed 2.0 Mini
DeepInfra · Feb 14, 2026
Lightweight ByteDance Seed 2.0 model for low-latency multimodal reasoning and high-volume tasks

Seed 2.0 Pro
DeepInfra · Feb 14, 2026
Flagship ByteDance Seed 2.0 model for complex multimodal reasoning and long-horizon agent workflows

GLM-5
DeepInfra · Feb 12, 2026
Flagship GLM model for hybrid reasoning, coding, and agentic engineering

MiniMax M2.5
DeepInfra · Feb 12, 2026
MiniMax model for chat, coding, office work, and agentic tasks

Qwen 3.5 397B A17B
DeepInfra · Feb 1, 2026
Qwen vision-language model for visual reasoning, documents, and agent tasks

Qwen 3.5 35B A3B
DeepInfra · Feb 1, 2026
Qwen vision-language model for visual reasoning, documents, and agent tasks

Kimi K2.5
DeepInfra · Jan 27, 2026
Kimi multimodal agent model for visual understanding, coding, and planning

GLM-4.7-Flash
DeepInfra · Jan 19, 2026
Efficient GLM model for fast reasoning, coding, and agent workflows

MiniMax M2.1
DeepInfra · Dec 23, 2025
MiniMax M2.1 from DeepInfra - text input, 196,608 token context

GLM-4.7
DeepInfra · Dec 22, 2025
Flagship GLM model for hybrid reasoning, coding, and agentic engineering

