Guide
2,496 step-by-step guides to setting up AI models from OpenAI, Anthropic, Google, OpenRouter, and more in TypingMind with your own API key.
Search and filter guides
Showing 1,480-1,530 of 2,496 guides

NVIDIA: Nemotron Nano 12B 2 VL
OpenRouter · Oct 28, 2025
NVIDIA Nemotron Nano 2 VL is a 12-billion-parameter open multimodal reasoning model designed for video understanding and document intelligence. It introduces a hybrid Transformer-Mamba architecture, combining transformer-level accuracy with Mamba’s...

NVIDIA Nemotron Nano 12B v2 VL BF16
AWS Bedrock · Oct 28, 2025
Nemotron multimodal model for visual reasoning and agentic AI workflows

Nvidia Nemotron Nano 12B V2 VL
Vercel AI Gateway · Oct 28, 2025
Nemotron multimodal model for visual reasoning and agentic AI workflows

MiniMax-M2
Fireworks AI · Oct 27, 2025
MiniMax-M2 from Fireworks AI - text input, 192,000 token context

MiniMax-M2
Synthetic · Oct 27, 2025
MiniMax-M2 from Synthetic - text input, 196,608 token context

MiniMax-M2
AWS Bedrock · Oct 27, 2025
Efficient open MiniMax model built for coding agents and tool-heavy workflows
MiniMax-M2
Hugging Face · Oct 27, 2025
Efficient open MiniMax model built for coding agents and tool-heavy workflows

MiniMax M2
Vercel AI Gateway · Oct 27, 2025
Efficient open MiniMax model built for coding agents and tool-heavy workflows

MiniMax-M2
MiniMax · Oct 27, 2025
Efficient open MiniMax model built for coding agents and tool-heavy workflows

MiniMax-M2
Novita · Oct 27, 2025
MiniMax model for chat, coding, office work, and agentic tasks

Seedance v1.0 Pro Fast
Vercel AI Gateway · Oct 24, 2025
Video model for prompt-guided generation, editing, and motion workflows

DeepSeek-OCR
Novita · Oct 24, 2025
OCR model for extracting structured text from documents and screenshots

MiniMax: MiniMax M2
OpenRouter · Oct 23, 2025
MiniMax-M2 is a compact, high-efficiency large language model optimized for end-to-end coding and agentic workflows. With 10 billion activated parameters (230 billion total), it delivers near-frontier intelligence across general reasoning,...

Qwen: Qwen3 VL 32B Instruct
OpenRouter · Oct 23, 2025
Qwen3-VL-32B-Instruct is a large-scale multimodal vision-language model designed for high-precision understanding and reasoning across text, images, and video. With 32 billion parameters, it combines deep visual perception with advanced text...

PaddleOCR-VL
Novita · Oct 22, 2025
Multimodal model for analyzing text, images, documents, and rich media

LiquidAI: LFM2-8B-A1B
OpenRouter · Oct 20, 2025
LFM2-8B-A1B is an efficient on-device Mixture-of-Experts (MoE) model from Liquid AI’s LFM2 family, built for fast, high-quality inference on edge hardware. It uses 8.3B total parameters with only ~1.5B active per token, delivering strong performance while keeping compute and memory usage low—making it ideal for phones, tablets, and laptops.

LiquidAI: LFM2-2.6B
OpenRouter · Oct 20, 2025
LFM2 is a new generation of hybrid models developed by Liquid AI, specifically designed for edge AI and on-device deployment. It sets a new standard in terms of quality, speed, and memory efficiency.

IBM: Granite 4.0 Micro
OpenRouter · Oct 20, 2025
Granite-4.0-H-Micro is a 3B parameter from the Granite 4 family of models. These models are the latest in a series of models released by IBM. They are fine-tuned for long...

S1
Vercel AI Gateway · Oct 20, 2025
Speech generation model for controllable voice, narration, and audio delivery

Microsoft: Phi 4 Mini Instruct
OpenRouter · Oct 17, 2025
Phi-4-mini-instruct is a lightweight open model built upon synthetic data and filtered publicly available websites - with a focus on high-quality, reasoning dense data. The model belongs to the Phi-4...

Deep Cogito: Cogito V2 Preview Llama 405B
OpenRouter · Oct 17, 2025
Cogito v2 405B is a dense hybrid reasoning model that combines direct answering capabilities with advanced self-reflection. It represents a significant step toward frontier intelligence with dense architecture delivering performance competitive with leading closed models. This advanced reasoning system combines policy improvement with massive scale for exceptional capabilities.

qwen/qwen3-vl-8b-instruct
Novita · Oct 17, 2025
Qwen vision-language model for visual reasoning, documents, and agent tasks

OpenAI: GPT-5 Image Mini
OpenRouter · Oct 16, 2025
GPT-5 Image Mini combines OpenAI's advanced language capabilities, powered by [GPT-5 Mini](https://openrouter.ai/openai/gpt-5-mini), with GPT Image 1 Mini for efficient image generation. This natively multimodal model features superior instruction following, text...

Anthropic: Claude Haiku 4.5
OpenRouter · Oct 15, 2025
Claude Haiku 4.5 is Anthropic’s fastest and most efficient model, delivering near-frontier intelligence at a fraction of the cost and latency of larger Claude models. Matching Claude Sonnet 4’s performance...

Anthropic: Claude Haiku 4.5 (batch)
OpenRouter · Oct 15, 2025
Claude Haiku 4.5 is Anthropic’s fastest and most efficient model, delivering near-frontier intelligence at a fraction of the cost and latency of larger Claude models. Matching Claude Sonnet 4’s performance...

Veo 3.1
Google · Oct 15, 2025
Video model for prompt-guided generation, editing, and motion workflows

Veo 3.1 fast
Google · Oct 15, 2025
Video model for prompt-guided generation, editing, and motion workflows

Claude Haiku 4.5 (latest)
Anthropic · Oct 15, 2025
Fast Claude lane for lightweight agents, office tasks, and responsive chat

Claude Haiku 4.5
Anthropic · Oct 15, 2025
Fast Claude model for responsive assistance, classification, and lightweight agents

Claude Haiku 4.5 (EU)
AWS Bedrock · Oct 15, 2025
Fast Claude model for responsive assistance, classification, and lightweight agents

Claude Haiku 4.5 (Global)
AWS Bedrock · Oct 15, 2025
Fast Claude model for responsive assistance, classification, and lightweight agents

Claude Haiku 4.5 (US)
AWS Bedrock · Oct 15, 2025
Fast Claude model for responsive assistance, classification, and lightweight agents

Claude Haiku 4.5
AWS Bedrock · Oct 15, 2025
Fast Claude model for responsive assistance, classification, and lightweight agents

Claude Haiku 4.5 (AU)
AWS Bedrock · Oct 15, 2025
Fast Claude model for responsive assistance, classification, and lightweight agents

Claude Haiku 4.5 (JP)
AWS Bedrock · Oct 15, 2025
Fast Claude model for responsive assistance, classification, and lightweight agents

Claude Haiku 4.5
Vercel AI Gateway · Oct 15, 2025
Fast Claude lane for lightweight agents, office tasks, and responsive chat

Veo 3.1 Fast Generate
Vercel AI Gateway · Oct 15, 2025
Video model for prompt-guided generation, editing, and motion workflows

Veo 3.1
Vercel AI Gateway · Oct 15, 2025
Video model for prompt-guided generation, editing, and motion workflows

Claude Haiku 4.5 (latest)
Cloudflare AI Gateway · Oct 15, 2025
Fast Claude lane for lightweight agents, office tasks, and responsive chat

Qwen: Qwen3 VL 8B Thinking
OpenRouter · Oct 14, 2025
Qwen3-VL-8B-Thinking is the reasoning-optimized variant of the Qwen3-VL-8B multimodal model, designed for advanced visual and textual reasoning across complex scenes, documents, and temporal sequences. It integrates enhanced multimodal alignment and...

Qwen: Qwen3 VL 8B Instruct
OpenRouter · Oct 14, 2025
Qwen3-VL-8B-Instruct is a multimodal vision-language model from the Qwen3-VL series, built for high-fidelity understanding and reasoning across text, images, and video. It features improved multimodal fusion with Interleaved-MRoPE for long-horizon...

OpenAI: GPT-5 Image
OpenRouter · Oct 14, 2025
[GPT-5](https://openrouter.ai/openai/gpt-5) Image combines OpenAI's GPT-5 model with state-of-the-art image generation capabilities. It offers major improvements in reasoning, code quality, and user experience while incorporating GPT Image 1's superior instruction following,...

GLM 4.5 Air
Novita · Oct 13, 2025
Efficient GLM model for fast reasoning, coding, and agent workflows

qwen/qwen3-vl-30b-a3b-instruct
Novita · Oct 11, 2025
Qwen vision-language model for visual reasoning, documents, and agent tasks

qwen/qwen3-vl-30b-a3b-thinking
Novita · Oct 11, 2025
Qwen vision-language model for visual reasoning, documents, and agent tasks

OpenAI: o3 Deep Research
OpenRouter · Oct 10, 2025
o3-deep-research is OpenAI's advanced model for deep research, designed to tackle complex, multi-step research tasks. Note: This model always uses the 'web_search' tool which adds additional cost.

OpenAI: o4 Mini Deep Research
OpenRouter · Oct 10, 2025
o4-mini-deep-research is OpenAI's faster, more affordable deep research model—ideal for tackling complex, multi-step research tasks. Note: This model always uses the 'web_search' tool which adds additional cost.

NVIDIA: Llama 3.3 Nemotron Super 49B V1.5
OpenRouter · Oct 10, 2025
Llama-3.3-Nemotron-Super-49B-v1.5 is a 49B-parameter, English-centric reasoning/chat model derived from Meta’s Llama-3.3-70B-Instruct with a 128K context. It’s post-trained for agentic workflows (RAG, tool calling) via SFT across math, code, science, and...

GPT-Realtime mini
Vercel AI Gateway · Oct 10, 2025
Speech generation model for controllable voice, narration, and audio delivery

Baidu: ERNIE 4.5 21B A3B Thinking
OpenRouter · Oct 9, 2025
ERNIE-4.5-21B-A3B-Thinking is Baidu's upgraded lightweight MoE model, refined to boost reasoning depth and quality for top-tier performance in logical puzzles, math, science, coding, text generation, and expert-level academic benchmarks.

Qwen3 Coder 30b A3B Instruct
Novita · Oct 9, 2025
Qwen coding model for software agents, repository edits, and code reasoning

