Guide

2,496 step-by-step guides to setting up AI models from OpenAI, Anthropic, Google, OpenRouter, and more in TypingMind with your own API key.

Search and filter guides

Showing 1,480-1,530 of 2,496 guides

OpenRouter logo

NVIDIA: Nemotron Nano 12B 2 VL

OpenRouter · Oct 28, 2025

NVIDIA Nemotron Nano 2 VL is a 12-billion-parameter open multimodal reasoning model designed for video understanding and document intelligence. It introduces a hybrid Transformer-Mamba architecture, combining transformer-level accuracy with Mamba’s...

131.1K context
View
AWS Bedrock logo

NVIDIA Nemotron Nano 12B v2 VL BF16

AWS Bedrock · Oct 28, 2025

Nemotron multimodal model for visual reasoning and agentic AI workflows

131.1K context
View
Vercel AI Gateway logo

Nvidia Nemotron Nano 12B V2 VL

Vercel AI Gateway · Oct 28, 2025

Nemotron multimodal model for visual reasoning and agentic AI workflows

131.1K context
View
Fireworks AI logo

MiniMax-M2

Fireworks AI · Oct 27, 2025

MiniMax-M2 from Fireworks AI - text input, 192,000 token context

192K context
View
Synthetic logo

MiniMax-M2

Synthetic · Oct 27, 2025

MiniMax-M2 from Synthetic - text input, 196,608 token context

196.6K context
View
AWS Bedrock logo

MiniMax-M2

AWS Bedrock · Oct 27, 2025

Efficient open MiniMax model built for coding agents and tool-heavy workflows

204.6K context
View
Hugging Face logo

MiniMax-M2

Hugging Face · Oct 27, 2025

Efficient open MiniMax model built for coding agents and tool-heavy workflows

204.8K context
View
Vercel AI Gateway logo

MiniMax M2

Vercel AI Gateway · Oct 27, 2025

Efficient open MiniMax model built for coding agents and tool-heavy workflows

205K context
View
MiniMax logo

MiniMax-M2

MiniMax · Oct 27, 2025

Efficient open MiniMax model built for coding agents and tool-heavy workflows

204.8K context
View
Novita logo

MiniMax-M2

Novita · Oct 27, 2025

MiniMax model for chat, coding, office work, and agentic tasks

204.8K context
View
Vercel AI Gateway logo

Seedance v1.0 Pro Fast

Vercel AI Gateway · Oct 24, 2025

Video model for prompt-guided generation, editing, and motion workflows

View
Novita logo

DeepSeek-OCR

Novita · Oct 24, 2025

OCR model for extracting structured text from documents and screenshots

8.2K context
View
OpenRouter logo

MiniMax: MiniMax M2

OpenRouter · Oct 23, 2025

MiniMax-M2 is a compact, high-efficiency large language model optimized for end-to-end coding and agentic workflows. With 10 billion activated parameters (230 billion total), it delivers near-frontier intelligence across general reasoning,...

204.8K context
View
OpenRouter logo

Qwen: Qwen3 VL 32B Instruct

OpenRouter · Oct 23, 2025

Qwen3-VL-32B-Instruct is a large-scale multimodal vision-language model designed for high-precision understanding and reasoning across text, images, and video. With 32 billion parameters, it combines deep visual perception with advanced text...

131.1K context
View
Novita logo

PaddleOCR-VL

Novita · Oct 22, 2025

Multimodal model for analyzing text, images, documents, and rich media

16.4K context
View
OpenRouter logo

LiquidAI: LFM2-8B-A1B

OpenRouter · Oct 20, 2025

LFM2-8B-A1B is an efficient on-device Mixture-of-Experts (MoE) model from Liquid AI’s LFM2 family, built for fast, high-quality inference on edge hardware. It uses 8.3B total parameters with only ~1.5B active per token, delivering strong performance while keeping compute and memory usage low—making it ideal for phones, tablets, and laptops.

32.8K context
View
OpenRouter logo

LiquidAI: LFM2-2.6B

OpenRouter · Oct 20, 2025

LFM2 is a new generation of hybrid models developed by Liquid AI, specifically designed for edge AI and on-device deployment. It sets a new standard in terms of quality, speed, and memory efficiency.

32.8K context
View
OpenRouter logo

IBM: Granite 4.0 Micro

OpenRouter · Oct 20, 2025

Granite-4.0-H-Micro is a 3B parameter from the Granite 4 family of models. These models are the latest in a series of models released by IBM. They are fine-tuned for long...

131K context
View
Vercel AI Gateway logo

S1

Vercel AI Gateway · Oct 20, 2025

Speech generation model for controllable voice, narration, and audio delivery

View
OpenRouter logo

Microsoft: Phi 4 Mini Instruct

OpenRouter · Oct 17, 2025

Phi-4-mini-instruct is a lightweight open model built upon synthetic data and filtered publicly available websites - with a focus on high-quality, reasoning dense data. The model belongs to the Phi-4...

131.1K context
View
OpenRouter logo

Deep Cogito: Cogito V2 Preview Llama 405B

OpenRouter · Oct 17, 2025

Cogito v2 405B is a dense hybrid reasoning model that combines direct answering capabilities with advanced self-reflection. It represents a significant step toward frontier intelligence with dense architecture delivering performance competitive with leading closed models. This advanced reasoning system combines policy improvement with massive scale for exceptional capabilities.

32.8K context
View
Novita logo

qwen/qwen3-vl-8b-instruct

Novita · Oct 17, 2025

Qwen vision-language model for visual reasoning, documents, and agent tasks

131.1K context
View
OpenRouter logo

OpenAI: GPT-5 Image Mini

OpenRouter · Oct 16, 2025

GPT-5 Image Mini combines OpenAI's advanced language capabilities, powered by [GPT-5 Mini](https://openrouter.ai/openai/gpt-5-mini), with GPT Image 1 Mini for efficient image generation. This natively multimodal model features superior instruction following, text...

400K context+1
View
OpenRouter logo

Anthropic: Claude Haiku 4.5

OpenRouter · Oct 15, 2025

Claude Haiku 4.5 is Anthropic’s fastest and most efficient model, delivering near-frontier intelligence at a fraction of the cost and latency of larger Claude models. Matching Claude Sonnet 4’s performance...

200K context+1
View
OpenRouter logo

Anthropic: Claude Haiku 4.5 (batch)

OpenRouter · Oct 15, 2025

Claude Haiku 4.5 is Anthropic’s fastest and most efficient model, delivering near-frontier intelligence at a fraction of the cost and latency of larger Claude models. Matching Claude Sonnet 4’s performance...

200K context+1
View
Google logo

Veo 3.1

Google · Oct 15, 2025

Video model for prompt-guided generation, editing, and motion workflows

480 context
View
Google logo

Veo 3.1 fast

Google · Oct 15, 2025

Video model for prompt-guided generation, editing, and motion workflows

480 context
View
Anthropic logo

Claude Haiku 4.5 (latest)

Anthropic · Oct 15, 2025

Fast Claude lane for lightweight agents, office tasks, and responsive chat

200K context+1
View
Anthropic logo

Claude Haiku 4.5

Anthropic · Oct 15, 2025

Fast Claude model for responsive assistance, classification, and lightweight agents

200K context+1
View
AWS Bedrock logo

Claude Haiku 4.5 (EU)

AWS Bedrock · Oct 15, 2025

Fast Claude model for responsive assistance, classification, and lightweight agents

200K context+1
View
AWS Bedrock logo

Claude Haiku 4.5 (Global)

AWS Bedrock · Oct 15, 2025

Fast Claude model for responsive assistance, classification, and lightweight agents

200K context+1
View
AWS Bedrock logo

Claude Haiku 4.5 (US)

AWS Bedrock · Oct 15, 2025

Fast Claude model for responsive assistance, classification, and lightweight agents

200K context+1
View
AWS Bedrock logo

Claude Haiku 4.5

AWS Bedrock · Oct 15, 2025

Fast Claude model for responsive assistance, classification, and lightweight agents

200K context+1
View
AWS Bedrock logo

Claude Haiku 4.5 (AU)

AWS Bedrock · Oct 15, 2025

Fast Claude model for responsive assistance, classification, and lightweight agents

200K context+1
View
AWS Bedrock logo

Claude Haiku 4.5 (JP)

AWS Bedrock · Oct 15, 2025

Fast Claude model for responsive assistance, classification, and lightweight agents

200K context+1
View
Vercel AI Gateway logo

Claude Haiku 4.5

Vercel AI Gateway · Oct 15, 2025

Fast Claude lane for lightweight agents, office tasks, and responsive chat

200K context
View
Vercel AI Gateway logo

Veo 3.1 Fast Generate

Vercel AI Gateway · Oct 15, 2025

Video model for prompt-guided generation, editing, and motion workflows

View
Vercel AI Gateway logo

Veo 3.1

Vercel AI Gateway · Oct 15, 2025

Video model for prompt-guided generation, editing, and motion workflows

View
Cloudflare AI Gateway logo

Claude Haiku 4.5 (latest)

Cloudflare AI Gateway · Oct 15, 2025

Fast Claude lane for lightweight agents, office tasks, and responsive chat

200K context+1
View
OpenRouter logo

Qwen: Qwen3 VL 8B Thinking

OpenRouter · Oct 14, 2025

Qwen3-VL-8B-Thinking is the reasoning-optimized variant of the Qwen3-VL-8B multimodal model, designed for advanced visual and textual reasoning across complex scenes, documents, and temporal sequences. It integrates enhanced multimodal alignment and...

131.1K context
View
OpenRouter logo

Qwen: Qwen3 VL 8B Instruct

OpenRouter · Oct 14, 2025

Qwen3-VL-8B-Instruct is a multimodal vision-language model from the Qwen3-VL series, built for high-fidelity understanding and reasoning across text, images, and video. It features improved multimodal fusion with Interleaved-MRoPE for long-horizon...

262.1K context
View
OpenRouter logo

OpenAI: GPT-5 Image

OpenRouter · Oct 14, 2025

[GPT-5](https://openrouter.ai/openai/gpt-5) Image combines OpenAI's GPT-5 model with state-of-the-art image generation capabilities. It offers major improvements in reasoning, code quality, and user experience while incorporating GPT Image 1's superior instruction following,...

400K context+1
View
Novita logo

GLM 4.5 Air

Novita · Oct 13, 2025

Efficient GLM model for fast reasoning, coding, and agent workflows

131.1K context
View
Novita logo

qwen/qwen3-vl-30b-a3b-instruct

Novita · Oct 11, 2025

Qwen vision-language model for visual reasoning, documents, and agent tasks

131.1K context
View
Novita logo

qwen/qwen3-vl-30b-a3b-thinking

Novita · Oct 11, 2025

Qwen vision-language model for visual reasoning, documents, and agent tasks

131.1K context
View
OpenRouter logo

OpenAI: o3 Deep Research

OpenRouter · Oct 10, 2025

o3-deep-research is OpenAI's advanced model for deep research, designed to tackle complex, multi-step research tasks. Note: This model always uses the 'web_search' tool which adds additional cost.

200K context+1
View
OpenRouter logo

OpenAI: o4 Mini Deep Research

OpenRouter · Oct 10, 2025

o4-mini-deep-research is OpenAI's faster, more affordable deep research model—ideal for tackling complex, multi-step research tasks. Note: This model always uses the 'web_search' tool which adds additional cost.

200K context+1
View
OpenRouter logo

NVIDIA: Llama 3.3 Nemotron Super 49B V1.5

OpenRouter · Oct 10, 2025

Llama-3.3-Nemotron-Super-49B-v1.5 is a 49B-parameter, English-centric reasoning/chat model derived from Meta’s Llama-3.3-70B-Instruct with a 128K context. It’s post-trained for agentic workflows (RAG, tool calling) via SFT across math, code, science, and...

131.1K context
View
Vercel AI Gateway logo

GPT-Realtime mini

Vercel AI Gateway · Oct 10, 2025

Speech generation model for controllable voice, narration, and audio delivery

View
OpenRouter logo

Baidu: ERNIE 4.5 21B A3B Thinking

OpenRouter · Oct 9, 2025

ERNIE-4.5-21B-A3B-Thinking is Baidu's upgraded lightweight MoE model, refined to boost reasoning depth and quality for top-tier performance in logical puzzles, math, science, coding, text generation, and expert-level academic benchmarks.

131.1K context
View
Novita logo

Qwen3 Coder 30b A3B Instruct

Novita · Oct 9, 2025

Qwen coding model for software agents, repository edits, and code reasoning

160K context
View