Guide
2,496 step-by-step guides to setting up AI models from OpenAI, Anthropic, Google, OpenRouter, and more in TypingMind with your own API key.
Search and filter guides
Showing 1-51 of 63 guides

Gemini 3.8 Flash
Google · Sep 2, 2026
Google's most intelligent Flash model, engineered for long-horizon software engineering, autonomous agents, and complex enterprise workflows

Gemini Flash Latest
Google · Aug 13, 2026
High-efficiency Gemini model for agentic workflows, coding, and multimodal reasoning

Gemini 3.7 Flash
Google · Aug 13, 2026
High-efficiency Gemini model for agentic workflows, coding, and multimodal reasoning

Gemini Flash-Lite Latest
Google · Jul 21, 2026
Fast Gemini model balancing multimodal reasoning, tool use, and cost

Gemini 3.5 Flash Lite
Google · Jul 21, 2026
Fast Gemini model balancing multimodal reasoning, tool use, and cost

Gemini 3.6 Flash
Google · Jul 21, 2026
Fast Gemini model balancing multimodal reasoning, tool use, and cost

Gemini Omni Flash Preview
Google · Jun 30, 2026
Video generation and editing model for fast, conversational text- and image-to-video workflows

Nano Banana 2 Lite
Google · Jun 30, 2026
Fastest, most cost-efficient Gemini image model for high-volume 1K generation and editing

Gemini 3.5 Live Translate Preview
Google · Jun 9, 2026
Low-latency audio-to-audio model for real-time speech translation across 70+ languages

Nano Banana 2
Google · May 28, 2026
Image model for prompt-driven generation, editing, and visual design workflows

Nano Banana Pro
Google · May 28, 2026
Nano Banana Pro for higher-fidelity image generation and design-heavy edits

Gemini 3.5 Flash
Google · May 19, 2026
Fast Gemini model balancing multimodal reasoning, tool use, and cost

Gemini 3.1 Flash Lite
Google · May 7, 2026
Low-latency Gemini model for high-volume multimodal and agent workloads

Gemini Embedding 2
Google · Apr 22, 2026
Multimodal embedding model mapping text, images, video, audio, and PDFs into a unified embedding space

Deep Research Preview (Apr-21-2026)
Google · Apr 21, 2026
Agentic model for autonomous multi-step research, synthesis, and cited reports

Deep Research Max Preview (Apr-21-2026)
Google · Apr 21, 2026
Maximum-comprehensiveness agentic researcher for multi-step investigation, synthesis, and cited reports

Gemini 3.1 Flash TTS Preview
Google · Apr 15, 2026
Low-latency speech generation with steerable prompts and expressive audio tags

Gemini Robotics-ER 1.6 Preview
Google · Apr 14, 2026
Vision-language model for embodied reasoning: spatial understanding, task planning, and physical-world agentic robotics

Gemma 4 26B
Google · Apr 2, 2026
Gemma 4 26B from Google - text, image input, 256,000 token context

Gemma 4 31B IT
Google · Apr 2, 2026
Largest Gemma 4 instruction model for open, self-hosted chat and reasoning

Gemma 4 26B A4B IT
Google · Apr 2, 2026
Open Gemma instruction model for efficient chat and self-hosted deployments

Veo 3.1 lite
Google · Mar 31, 2026
Video model for prompt-guided generation, editing, and motion workflows

Gemini 3.1 Flash Live Preview
Google · Mar 26, 2026
High-quality, low-latency Live API model for real-time dialogue and voice-first AI applications

Lyria 3 Pro Preview
Google · Mar 25, 2026
Music generation model for full-length songs from text or images with vocals and structure

Lyria 3 Clip Preview
Google · Mar 25, 2026
Music generation model for short 30-second clips, loops, and previews from text or image prompts

Gemini 3.1 Flash Lite Preview
Google · Mar 3, 2026
Legacy model retained for compatibility with older integrations

Nano Banana 2
Google · Feb 26, 2026
Image model for prompt-driven generation, editing, and visual design workflows

Gemini 3.1 Pro Preview Custom Tools
Google · Feb 19, 2026
Advanced Gemini model for complex reasoning, coding, and multimodal analysis

Gemini 3.1 Pro Preview
Google · Feb 19, 2026
Reasoning-first Gemini preview for agentic coding and complex problem solving

Gemini 3 Flash Preview
Google · Dec 17, 2025
New Gemini flash lane bringing frontier-style multimodal reasoning to cheaper runs

Nano Banana Pro
Google · Nov 20, 2025
Nano Banana Pro for higher-fidelity image generation and design-heavy edits

Gemini 3 Pro Preview
Google · Nov 18, 2025
Preview Gemini flagship for complex reasoning, coding, and rich multimodal prompts

Veo 3.1
Google · Oct 15, 2025
Video model for prompt-guided generation, editing, and motion workflows

Veo 3.1 fast
Google · Oct 15, 2025
Video model for prompt-guided generation, editing, and motion workflows

Gemini 2.5 Computer Use Preview 10-2025
Google · Oct 7, 2025
Specialized Gemini 2.5 model for browser-control agents that automate UI tasks

Gemini 2.5 Flash TTS
Google · Sep 30, 2025
Gemini 2.5 Flash TTS from Google - text input, 32,768 token context

Gemini 2.5 Flash Preview 09-25
Google · Sep 25, 2025
Gemini 2.5 Flash Preview 09-25 from Google - text, image, audio, video, pdf input, 1,048,576 token context

Gemini 2.5 Flash Lite Preview 09-25
Google · Sep 25, 2025
Gemini 2.5 Flash Lite Preview 09-25 from Google - text, image, audio, video, pdf input, 1,048,576 token context

Gemini Live 2.5 Flash
Google · Sep 1, 2025
Gemini Live 2.5 Flash from Google - text, image, audio, video input, 128,000 token context

Nano Banana
Google · Aug 26, 2025
Nano Banana image model for fast generation, edits, and character-consistent assets

Gemini 2.5 Flash Image (Preview)
Google · Aug 26, 2025
Gemini 2.5 Flash Image (Preview) from Google - text, image input, 32,768 token context

Gemma 3n 2B
Google · Jul 9, 2025
Gemma 3n 2B from Google - text input, 8,192 token context

Gemini 2.5 Flash
Google · Jun 17, 2025
Fast Gemini workhorse for multimodal apps where latency and price matter

Gemini Live 2.5 Flash Preview Native Audio
Google · Jun 17, 2025
Gemini Live 2.5 Flash Preview Native Audio from Google - text, audio, video input, 131,072 token context

Gemini 2.5 Flash-Lite
Google · Jun 17, 2025
Lean Gemini 2.5 lane for cheap multimodal traffic and quick agents

Gemini 2.5 Flash Lite Preview 06-17
Google · Jun 17, 2025
Gemini 2.5 Flash Lite Preview 06-17 from Google - text, image, audio, video, pdf input, 1,048,576 token context

Gemini 2.5 Pro
Google · Jun 17, 2025
Google's proven reasoning model for coding, math, and multimodal analysis

Gemini 2.5 Pro Preview 06-05
Google · Jun 5, 2025
Gemini 2.5 Pro Preview 06-05 from Google - text, image, audio, video, pdf input, 1,048,576 token context

Gemini Embedding 001
Google · May 20, 2025
Embedding model for semantic search, retrieval, clustering, and ranking pipelines

Gemini 2.5 Flash Preview 05-20
Google · May 20, 2025
Gemini 2.5 Flash Preview 05-20 from Google - text, image, audio, video, pdf input, 1,048,576 token context

Gemma 3n 4B
Google · May 20, 2025
Gemma 3n 4B from Google - text input, 8,192 token context

