Guide

2,496 step-by-step guides to setting up AI models from OpenAI, Anthropic, Google, OpenRouter, and more in TypingMind with your own API key.

Search and filter guides

Showing 1,174-1,224 of 2,496 guides

Venice AI logo

NVIDIA Nemotron 3 Nano 30B

Venice AI · Jan 27, 2026

Small Nemotron 3 MoE for efficient coding, math, and long-context agents

128K context
View
Chutes logo

GLM 4.7 FP8

Chutes · Jan 27, 2026

GLM 4.7 FP8 from Chutes - text input, 202,752 token context

202.8K context
View
Chutes logo

GLM 4.7 Flash

Chutes · Jan 27, 2026

GLM 4.7 Flash from Chutes - text input, 202,752 token context

202.8K context
View
Chutes logo

GLM 4.5 FP8

Chutes · Jan 27, 2026

GLM 4.5 FP8 from Chutes - text input, 131,072 token context

131.1K context
View
Chutes logo

GLM 4.6 FP8

Chutes · Jan 27, 2026

GLM 4.6 FP8 from Chutes - text input, 202,752 token context

202.8K context
View
Chutes logo

Llama 3.2 1B Instruct

Chutes · Jan 27, 2026

Llama 3.2 1B Instruct from Chutes - text input, 16,384 token context

16.4K context
View
Chutes logo

TNG R1T Chimera Turbo

Chutes · Jan 27, 2026

TNG R1T Chimera Turbo from Chutes - text input, 163,840 token context

163.8K context
View
AWS Bedrock logo

Kimi K2.5

AWS Bedrock · Jan 27, 2026

Earlier Kimi frontier model for long-context agents, coding, and multimodal work

262.1K context
View
DeepInfra logo

Kimi K2.5

DeepInfra · Jan 27, 2026

Kimi multimodal agent model for visual understanding, coding, and planning

262.1K context+1
View
Ollama Cloud logo

kimi-k2.5

Ollama Cloud · Jan 27, 2026

Kimi multimodal agent model for visual understanding, coding, and planning

262.1K context
View
Novita logo

deepseek/deepseek-ocr-2

Novita · Jan 27, 2026

OCR model for extracting structured text from documents and screenshots

8.2K context
View
Novita logo

Kimi K2.5

Novita · Jan 27, 2026

Kimi multimodal agent model for visual understanding, coding, and planning

262.1K context+1
View
OpenRouter logo

MiniMax: MiniMax M2-her

OpenRouter · Jan 23, 2026

MiniMax M2-her is a dialogue-first large language model built for immersive roleplay, character-driven chat, and expressive multi-turn conversations. Designed to stay consistent in tone and personality, it supports rich message...

65.5K context
View
Vercel AI Gateway logo

Qwen 3 Max Thinking

Vercel AI Gateway · Jan 23, 2026

Qwen reasoning model for deliberate problem solving, math, and coding

256K context
View
OpenRouter logo

Writer: Palmyra X5

OpenRouter · Jan 21, 2026

Palmyra X5 is Writer's most advanced model, purpose-built for building and scaling AI agents across the enterprise. It delivers industry-leading speed and efficiency on context windows up to 1 million...

1M context
View
OpenRouter logo

LiquidAI: LFM2.5-1.2B-Thinking (free)

OpenRouter · Jan 20, 2026

LFM2.5-1.2B-Thinking is a lightweight reasoning-focused model optimized for agentic tasks, data extraction, and RAG—while still running comfortably on edge devices. It supports long context (up to 32K tokens) and is...

32.8K context
View
OpenRouter logo

LiquidAI: LFM2.5-1.2B-Instruct (free)

OpenRouter · Jan 20, 2026

LFM2.5-1.2B-Instruct is a compact, high-performance instruction-tuned model built for fast on-device AI. It delivers strong chat quality in a 1.2B parameter footprint, with efficient edge inference and broad runtime support.

32.8K context
View
OpenRouter logo

OpenAI: GPT Audio

OpenRouter · Jan 19, 2026

The gpt-audio model is OpenAI's first generally available audio model. The new snapshot features an upgraded decoder for more natural sounding voices and maintains better voice consistency. Audio is priced...

128K context
View
OpenRouter logo

OpenAI: GPT Audio Mini

OpenRouter · Jan 19, 2026

A cost-efficient version of GPT Audio. The new snapshot features an upgraded decoder for more natural sounding voices and maintains better voice consistency. Input is priced at $0.60 per million...

128K context
View
OpenRouter logo

Z.ai: GLM 4.7 Flash

OpenRouter · Jan 19, 2026

As a 30B-class SOTA model, GLM-4.7-Flash offers a new option that balances performance and efficiency. It is further optimized for agentic coding use cases, strengthening coding capabilities, long-horizon task planning,...

200K context
View
Synthetic logo

GLM-4.7-Flash

Synthetic · Jan 19, 2026

Budget GLM lane for fast coding help, routing, and everyday automation

196.6K context
View
AWS Bedrock logo

GLM-4.7-Flash

AWS Bedrock · Jan 19, 2026

Budget GLM lane for fast coding help, routing, and everyday automation

202.8K context
View
DeepInfra logo

GLM-4.7-Flash

DeepInfra · Jan 19, 2026

Efficient GLM model for fast reasoning, coding, and agent workflows

202.8K context
View
Z.AI logo

GLM-4.7-Flash

Z.AI · Jan 19, 2026

Budget GLM lane for fast coding help, routing, and everyday automation

200K context
View
Z.AI logo

GLM-4.7-FlashX

Z.AI · Jan 19, 2026

Efficient GLM model for fast reasoning, coding, and agent workflows

200K context
View
Vercel AI Gateway logo

GLM 4.7 Flash

Vercel AI Gateway · Jan 19, 2026

Budget GLM lane for fast coding help, routing, and everyday automation

200K context
View
Vercel AI Gateway logo

GLM 4.7 FlashX

Vercel AI Gateway · Jan 19, 2026

Efficient GLM model for fast reasoning, coding, and agent workflows

200K context
View
Novita logo

GLM-4.7-Flash

Novita · Jan 19, 2026

Efficient GLM model for fast reasoning, coding, and agent workflows

200K context
View
Venice AI logo

Qwen3 VL 235B

Venice AI · Jan 16, 2026

Multimodal model for analyzing text, images, documents, and rich media

128K context
View
Venice AI logo

Mistral Small 3.2 24B Instruct

Venice AI · Jan 15, 2026

Efficient Mistral model for fast chat, extraction, and production assistants

256K context
View
Vercel AI Gateway logo

voyage-4-large

Vercel AI Gateway · Jan 15, 2026

Flagship model for demanding analysis, coding, and production agent workflows

32K context
View
Vercel AI Gateway logo

voyage-4

Vercel AI Gateway · Jan 15, 2026

General-purpose chat model for instruction following, writing, and analysis

32K context
View
Vercel AI Gateway logo

voyage-4-lite

Vercel AI Gateway · Jan 15, 2026

Efficient model for low-latency assistance, extraction, and routine automation

32K context
View
Vercel AI Gateway logo

FLUX.2 [klein] 4B

Vercel AI Gateway · Jan 15, 2026

Image model for prompt-driven generation, editing, and visual design workflows

View
Vercel AI Gateway logo

FLUX.2 [klein] 9B

Vercel AI Gateway · Jan 15, 2026

Image model for prompt-driven generation, editing, and visual design workflows

View
OpenRouter logo

OpenAI: GPT-5.2-Codex

OpenRouter · Jan 14, 2026

GPT-5.2-Codex is an upgraded version of GPT-5.1-Codex optimized for software engineering and coding workflows. It is designed for both interactive development sessions and long, independent execution of complex engineering tasks....

400K context
View
Chutes logo

Devstral 2 123B Instruct 2512 TEE

Chutes · Jan 10, 2026

Devstral 2 123B Instruct 2512 TEE from Chutes - text input, 262,144 token context

262.1K context
View
Chutes logo

MiroThinker V1.5 235B

Chutes · Jan 10, 2026

MiroThinker V1.5 235B from Chutes - text input, 262,144 token context

262.1K context
View
OpenRouter logo

AllenAI: Molmo2 8B

OpenRouter · Jan 9, 2026

Molmo2-8B is an open vision-language model developed by the Allen Institute for AI (Ai2) as part of the Molmo2 family, supporting image, video, and multi-image understanding and grounding. It is based on Qwen3-8B and uses SigLIP 2 as its vision backbone, outperforming other open-weight, open-data models on short videos, counting, and captioning, while remaining competitive on long-video tasks.

36.9K context
View
OpenRouter logo

AllenAI: Olmo 3.1 32B Instruct

OpenRouter · Jan 6, 2026

Olmo 3.1 32B Instruct is a large-scale, 32-billion-parameter instruction-tuned language model engineered for high-performance conversational AI, multi-turn dialogue, and practical instruction following. As part of the Olmo 3.1 family, this...

65.5K context
View
Novita logo

Kat Coder Pro

Novita · Jan 5, 2026

Coding model for repository understanding, refactors, and agentic engineering tasks

256K context
View
Moonshot AI logo

Kimi K2.5

Moonshot AI · Jan 1, 2026

Earlier Kimi frontier model for long-context agents, coding, and multimodal work

262.1K context+1
View
Chutes logo

Kimi K2.5 TEE

Chutes · Jan 1, 2026

Kimi K2.5 TEE from Chutes - text, image, video input, 262,144 token context

262.1K context+1
View
Synthetic logo

Kimi K2.5

Synthetic · Jan 1, 2026

Kimi K2.5 from Synthetic - text, image input, 262,144 token context

262.1K context
View
Synthetic logo

Kimi K2.5 (NVFP4)

Synthetic · Jan 1, 2026

Kimi K2.5 (NVFP4) from Synthetic - text, image input, 262,144 token context

262.1K context
View
Hugging Face logo

Kimi-K2.5

Hugging Face · Jan 1, 2026

Kimi multimodal agent model for visual understanding, coding, and planning

262.1K context
View
Vercel AI Gateway logo

Kimi K2.5

Vercel AI Gateway · Jan 1, 2026

Earlier Kimi frontier model for long-context agents, coding, and multimodal work

256K context
View
Chutes logo

Hermes 4.3 36B

Chutes · Dec 29, 2025

Hermes 4.3 36B from Chutes - text input, 32,768 token context

32.8K context
View
Chutes logo

Hermes 4 70B

Chutes · Dec 29, 2025

Hermes 4 70B from Chutes - text input, 131,072 token context

131.1K context
View
Chutes logo

Hermes 4 14B

Chutes · Dec 29, 2025

Hermes 4 14B from Chutes - text input, 40,960 token context

41K context
View
Chutes logo

Hermes 4 405B FP8 TEE

Chutes · Dec 29, 2025

Hermes 4 405B FP8 TEE from Chutes - text input, 131,072 token context

131.1K context
View