Guide
2,496 step-by-step guides to setting up AI models from OpenAI, Anthropic, Google, OpenRouter, and more in TypingMind with your own API key.
Search and filter guides
Showing 1,174-1,224 of 2,496 guides

NVIDIA Nemotron 3 Nano 30B
Venice AI · Jan 27, 2026
Small Nemotron 3 MoE for efficient coding, math, and long-context agents

GLM 4.7 FP8
Chutes · Jan 27, 2026
GLM 4.7 FP8 from Chutes - text input, 202,752 token context

GLM 4.7 Flash
Chutes · Jan 27, 2026
GLM 4.7 Flash from Chutes - text input, 202,752 token context

GLM 4.5 FP8
Chutes · Jan 27, 2026
GLM 4.5 FP8 from Chutes - text input, 131,072 token context

GLM 4.6 FP8
Chutes · Jan 27, 2026
GLM 4.6 FP8 from Chutes - text input, 202,752 token context

Llama 3.2 1B Instruct
Chutes · Jan 27, 2026
Llama 3.2 1B Instruct from Chutes - text input, 16,384 token context

TNG R1T Chimera Turbo
Chutes · Jan 27, 2026
TNG R1T Chimera Turbo from Chutes - text input, 163,840 token context

Kimi K2.5
AWS Bedrock · Jan 27, 2026
Earlier Kimi frontier model for long-context agents, coding, and multimodal work

Kimi K2.5
DeepInfra · Jan 27, 2026
Kimi multimodal agent model for visual understanding, coding, and planning

kimi-k2.5
Ollama Cloud · Jan 27, 2026
Kimi multimodal agent model for visual understanding, coding, and planning

deepseek/deepseek-ocr-2
Novita · Jan 27, 2026
OCR model for extracting structured text from documents and screenshots

Kimi K2.5
Novita · Jan 27, 2026
Kimi multimodal agent model for visual understanding, coding, and planning

MiniMax: MiniMax M2-her
OpenRouter · Jan 23, 2026
MiniMax M2-her is a dialogue-first large language model built for immersive roleplay, character-driven chat, and expressive multi-turn conversations. Designed to stay consistent in tone and personality, it supports rich message...

Qwen 3 Max Thinking
Vercel AI Gateway · Jan 23, 2026
Qwen reasoning model for deliberate problem solving, math, and coding

Writer: Palmyra X5
OpenRouter · Jan 21, 2026
Palmyra X5 is Writer's most advanced model, purpose-built for building and scaling AI agents across the enterprise. It delivers industry-leading speed and efficiency on context windows up to 1 million...

LiquidAI: LFM2.5-1.2B-Thinking (free)
OpenRouter · Jan 20, 2026
LFM2.5-1.2B-Thinking is a lightweight reasoning-focused model optimized for agentic tasks, data extraction, and RAG—while still running comfortably on edge devices. It supports long context (up to 32K tokens) and is...

LiquidAI: LFM2.5-1.2B-Instruct (free)
OpenRouter · Jan 20, 2026
LFM2.5-1.2B-Instruct is a compact, high-performance instruction-tuned model built for fast on-device AI. It delivers strong chat quality in a 1.2B parameter footprint, with efficient edge inference and broad runtime support.

OpenAI: GPT Audio
OpenRouter · Jan 19, 2026
The gpt-audio model is OpenAI's first generally available audio model. The new snapshot features an upgraded decoder for more natural sounding voices and maintains better voice consistency. Audio is priced...

OpenAI: GPT Audio Mini
OpenRouter · Jan 19, 2026
A cost-efficient version of GPT Audio. The new snapshot features an upgraded decoder for more natural sounding voices and maintains better voice consistency. Input is priced at $0.60 per million...

Z.ai: GLM 4.7 Flash
OpenRouter · Jan 19, 2026
As a 30B-class SOTA model, GLM-4.7-Flash offers a new option that balances performance and efficiency. It is further optimized for agentic coding use cases, strengthening coding capabilities, long-horizon task planning,...

GLM-4.7-Flash
Synthetic · Jan 19, 2026
Budget GLM lane for fast coding help, routing, and everyday automation

GLM-4.7-Flash
AWS Bedrock · Jan 19, 2026
Budget GLM lane for fast coding help, routing, and everyday automation

GLM-4.7-Flash
DeepInfra · Jan 19, 2026
Efficient GLM model for fast reasoning, coding, and agent workflows

GLM-4.7-Flash
Z.AI · Jan 19, 2026
Budget GLM lane for fast coding help, routing, and everyday automation

GLM-4.7-FlashX
Z.AI · Jan 19, 2026
Efficient GLM model for fast reasoning, coding, and agent workflows

GLM 4.7 Flash
Vercel AI Gateway · Jan 19, 2026
Budget GLM lane for fast coding help, routing, and everyday automation

GLM 4.7 FlashX
Vercel AI Gateway · Jan 19, 2026
Efficient GLM model for fast reasoning, coding, and agent workflows

GLM-4.7-Flash
Novita · Jan 19, 2026
Efficient GLM model for fast reasoning, coding, and agent workflows

Qwen3 VL 235B
Venice AI · Jan 16, 2026
Multimodal model for analyzing text, images, documents, and rich media

Mistral Small 3.2 24B Instruct
Venice AI · Jan 15, 2026
Efficient Mistral model for fast chat, extraction, and production assistants

voyage-4-large
Vercel AI Gateway · Jan 15, 2026
Flagship model for demanding analysis, coding, and production agent workflows

voyage-4
Vercel AI Gateway · Jan 15, 2026
General-purpose chat model for instruction following, writing, and analysis

voyage-4-lite
Vercel AI Gateway · Jan 15, 2026
Efficient model for low-latency assistance, extraction, and routine automation

FLUX.2 [klein] 4B
Vercel AI Gateway · Jan 15, 2026
Image model for prompt-driven generation, editing, and visual design workflows

FLUX.2 [klein] 9B
Vercel AI Gateway · Jan 15, 2026
Image model for prompt-driven generation, editing, and visual design workflows

OpenAI: GPT-5.2-Codex
OpenRouter · Jan 14, 2026
GPT-5.2-Codex is an upgraded version of GPT-5.1-Codex optimized for software engineering and coding workflows. It is designed for both interactive development sessions and long, independent execution of complex engineering tasks....

Devstral 2 123B Instruct 2512 TEE
Chutes · Jan 10, 2026
Devstral 2 123B Instruct 2512 TEE from Chutes - text input, 262,144 token context

MiroThinker V1.5 235B
Chutes · Jan 10, 2026
MiroThinker V1.5 235B from Chutes - text input, 262,144 token context

AllenAI: Molmo2 8B
OpenRouter · Jan 9, 2026
Molmo2-8B is an open vision-language model developed by the Allen Institute for AI (Ai2) as part of the Molmo2 family, supporting image, video, and multi-image understanding and grounding. It is based on Qwen3-8B and uses SigLIP 2 as its vision backbone, outperforming other open-weight, open-data models on short videos, counting, and captioning, while remaining competitive on long-video tasks.

AllenAI: Olmo 3.1 32B Instruct
OpenRouter · Jan 6, 2026
Olmo 3.1 32B Instruct is a large-scale, 32-billion-parameter instruction-tuned language model engineered for high-performance conversational AI, multi-turn dialogue, and practical instruction following. As part of the Olmo 3.1 family, this...

Kat Coder Pro
Novita · Jan 5, 2026
Coding model for repository understanding, refactors, and agentic engineering tasks

Kimi K2.5
Moonshot AI · Jan 1, 2026
Earlier Kimi frontier model for long-context agents, coding, and multimodal work

Kimi K2.5 TEE
Chutes · Jan 1, 2026
Kimi K2.5 TEE from Chutes - text, image, video input, 262,144 token context

Kimi K2.5
Synthetic · Jan 1, 2026
Kimi K2.5 from Synthetic - text, image input, 262,144 token context

Kimi K2.5 (NVFP4)
Synthetic · Jan 1, 2026
Kimi K2.5 (NVFP4) from Synthetic - text, image input, 262,144 token context
Kimi-K2.5
Hugging Face · Jan 1, 2026
Kimi multimodal agent model for visual understanding, coding, and planning

Kimi K2.5
Vercel AI Gateway · Jan 1, 2026
Earlier Kimi frontier model for long-context agents, coding, and multimodal work

Hermes 4.3 36B
Chutes · Dec 29, 2025
Hermes 4.3 36B from Chutes - text input, 32,768 token context

Hermes 4 70B
Chutes · Dec 29, 2025
Hermes 4 70B from Chutes - text input, 131,072 token context

Hermes 4 14B
Chutes · Dec 29, 2025
Hermes 4 14B from Chutes - text input, 40,960 token context

Hermes 4 405B FP8 TEE
Chutes · Dec 29, 2025
Hermes 4 405B FP8 TEE from Chutes - text input, 131,072 token context

