Guide
2,496 step-by-step guides to setting up AI models from OpenAI, Anthropic, Google, OpenRouter, and more in TypingMind with your own API key.
Search and filter guides
Showing 1,888-1,938 of 2,496 guides

o3 Pro
Vercel AI Gateway · Jun 10, 2025
High-effort o3 tier for difficult technical reasoning and careful answers

Google: Gemini 2.5 Pro Preview 06-05
OpenRouter · Jun 5, 2025
Gemini 2.5 Pro is Google’s state-of-the-art AI model designed for advanced reasoning, coding, mathematics, and scientific tasks. It employs “thinking” capabilities, enabling it to reason through responses with enhanced accuracy...

Gemini 2.5 Pro Preview 06-05
Google · Jun 5, 2025
Gemini 2.5 Pro Preview 06-05 from Google - text, image, audio, video, pdf input, 1,048,576 token context

Qwen3 Embedding 4B
Vercel AI Gateway · Jun 5, 2025
Embedding model for semantic search, retrieval, clustering, and ranking pipelines

Qwen3 Embedding 8B
Vercel AI Gateway · Jun 5, 2025
Embedding model for semantic search, retrieval, clustering, and ranking pipelines

DeepSeek: DeepSeek R1 0528 Qwen3 8B
OpenRouter · May 29, 2025
DeepSeek-R1-0528 is a lightly upgraded release of DeepSeek R1 that taps more compute and smarter post-training tricks, pushing its reasoning and inference to the brink of flagship models like O3 and Gemini 2.5 Pro. It now tops math, programming, and logic leaderboards, showcasing a step-change in depth-of-thought. The distilled variant, DeepSeek-R1-0528-Qwen3-8B, transfers this chain-of-thought into an 8 B-parameter form, beating standard Qwen3 8B by +10 pp and tying the 235 B “thinking” giant on AIME 2024.

Llama Prompt Guard 2 22M
Groq · May 29, 2025
Safety model for policy screening, moderation, and risk-aware routing workflows

Prompt Guard 2 86M
Groq · May 29, 2025
Safety model for policy screening, moderation, and risk-aware routing workflows

FLUX.1 Kontext Max
Vercel AI Gateway · May 29, 2025
Image model for prompt-driven generation, editing, and visual design workflows

FLUX.1 Kontext Pro
Vercel AI Gateway · May 29, 2025
Image model for prompt-driven generation, editing, and visual design workflows

DeepSeek R1 0528 Qwen3 8B
Novita · May 29, 2025
DeepSeek reasoning model for multi-step analysis, math, coding, and tools

DeepSeek: R1 0528 (free)
OpenRouter · May 28, 2025
May 28th update to the [original DeepSeek R1](/deepseek/deepseek-r1) Performance on par with [OpenAI o1](/openai/o1), but open-sourced and with fully open reasoning tokens. It's 671B parameters in size, with 37B active in an inference pass. Fully open-source model.

DeepSeek: R1 0528
OpenRouter · May 28, 2025
May 28th update to the [original DeepSeek R1](/deepseek/deepseek-r1) Performance on par with [OpenAI o1](/openai/o1), but open-sourced and with fully open reasoning tokens. It's 671B parameters in size, with 37B active...

Deepseek R1 05/28
Fireworks AI · May 28, 2025
Deepseek R1 05/28 from Fireworks AI - text input, 160,000 token context

DeepSeek-R1-0528
DeepInfra · May 28, 2025
DeepSeek reasoning model for multi-step analysis, math, coding, and tools
DeepSeek-R1-0528
Hugging Face · May 28, 2025
DeepSeek reasoning model for multi-step analysis, math, coding, and tools

Codestral Embed
Vercel AI Gateway · May 28, 2025
Embedding model for semantic search, retrieval, clustering, and ranking pipelines

DeepSeek R1 0528
Novita · May 28, 2025
DeepSeek reasoning model for multi-step analysis, math, coding, and tools

Anthropic: Claude Opus 4
OpenRouter · May 22, 2025
Claude Opus 4 is benchmarked as the world’s best coding model, at time of release, bringing sustained performance on complex, long-running tasks and agent workflows. It sets new benchmarks in...

Anthropic: Claude Sonnet 4
OpenRouter · May 22, 2025
Claude Sonnet 4 significantly enhances the capabilities of its predecessor, Sonnet 3.7, excelling in both coding and reasoning tasks with improved precision and controllability. Achieving state-of-the-art performance on SWE-bench (72.7%),...

Claude Opus 4 (latest)
Anthropic · May 22, 2025
Claude Opus 4 (latest) from Anthropic - text, image, pdf input, 200,000 token context

Claude Sonnet 4
Anthropic · May 22, 2025
Claude Sonnet 4 from Anthropic - text, image, pdf input, 200,000 token context

Claude Opus 4
Anthropic · May 22, 2025
Claude Opus 4 from Anthropic - text, image, pdf input, 200,000 token context

Claude Sonnet 4 (latest)
Anthropic · May 22, 2025
Claude Sonnet 4 (latest) from Anthropic - text, image, pdf input, 200,000 token context

Claude Opus 4 (US)
AWS Bedrock · May 22, 2025
Claude Opus 4 (US) from AWS Bedrock - text, image, pdf input, 200,000 token context

Claude Sonnet 4
AWS Bedrock · May 22, 2025
Claude Sonnet 4 from AWS Bedrock - text, image, pdf input, 200,000 token context

Claude Opus 4
AWS Bedrock · May 22, 2025
Claude Opus 4 from AWS Bedrock - text, image, pdf input, 200,000 token context

Claude Sonnet 4 (US)
AWS Bedrock · May 22, 2025
Balanced Claude model for coding, analysis, agent workflows, and cost control

Claude Sonnet 4 (Global)
AWS Bedrock · May 22, 2025
Balanced Claude model for coding, analysis, agent workflows, and cost control

Claude Sonnet 4 (EU)
AWS Bedrock · May 22, 2025
Balanced Claude model for coding, analysis, agent workflows, and cost control

Claude Sonnet 4 (APAC)
AWS Bedrock · May 22, 2025
Balanced Claude model for coding, analysis, agent workflows, and cost control

Claude Sonnet 4
Vercel AI Gateway · May 22, 2025
Balanced Claude model for coding, analysis, agent workflows, and cost control

Claude Opus 4
Vercel AI Gateway · May 22, 2025
Flagship Claude model for deep reasoning, coding, and long-horizon agents

Mistral: Devstral Small 2505
OpenRouter · May 21, 2025
Devstral-Small-2505 is a 24B parameter agentic LLM fine-tuned from Mistral-Small-3.1, jointly developed by Mistral AI and All Hands AI for advanced software engineering tasks. It is optimized for codebase exploration, multi-file editing, and integration into coding agents, achieving state-of-the-art results on SWE-Bench Verified (46.8%). Devstral supports a 128k context window and uses a custom Tekken tokenizer. It is text-only, with the vision encoder removed, and is suitable for local deployment on high-end consumer hardware (e.g., RTX 4090, 32GB RAM Macs). Devstral is best used in agentic workflows via the OpenHands scaffold and is compatible with inference frameworks like vLLM, Transformers, and Ollama. It is released under the Apache 2.0 license.

Venice
Setup guide · May 21, 2025
Venice.ai is a privacy-first, uncensored AI platform providing access to leading open-source models without data collection, content restrictions, or conversation logging. Key features include private access to multiple models (Llama, DeepSeek, StableDiffusion, GPT, Claude, Gemini), zero data retention with local chat history storage, decentralized GPU infrastructure preventing single-entity data access, uncensored text generation and image creation, web-enabled research and PDF analysis, and a private API with no logging. The platform serves over 1.3 million users and is integrating Web3 capabilities through partnerships like Warden Protocol for on-chain AI and censorship resistance. Venice offers both free and Pro tiers with mobile apps available, focusing on creative freedom and honest answers without guardrails while maintaining complete conversation privacy through decentralized processing.

Google: Gemma 3n 4B (free)
OpenRouter · May 20, 2025
Gemma 3n E4B-it is optimized for efficient execution on mobile and low-resource devices, such as phones, laptops, and tablets. It supports multimodal inputs—including text, visual data, and audio—enabling diverse tasks...

Google: Gemma 3n 4B
OpenRouter · May 20, 2025
Gemma 3n E4B-it is optimized for efficient execution on mobile and low-resource devices, such as phones, laptops, and tablets. It supports multimodal inputs—including text, visual data, and audio—enabling diverse tasks...

Gemini Embedding 001
Google · May 20, 2025
Embedding model for semantic search, retrieval, clustering, and ranking pipelines

Gemini 2.5 Flash Preview 05-20
Google · May 20, 2025
Gemini 2.5 Flash Preview 05-20 from Google - text, image, audio, video, pdf input, 1,048,576 token context

Gemma 3n 4B
Google · May 20, 2025
Gemma 3n 4B from Google - text input, 8,192 token context

Gemma 3N E4B Instruct
Together AI · May 20, 2025
Open Gemma instruction model for efficient chat and self-hosted deployments

voyage-3.5
Vercel AI Gateway · May 20, 2025
General-purpose chat model for instruction following, writing, and analysis

voyage-3.5-lite
Vercel AI Gateway · May 20, 2025
Efficient model for low-latency assistance, extraction, and routine automation

Gemini Embedding 001
Vercel AI Gateway · May 20, 2025
Embedding model for semantic search, retrieval, clustering, and ranking pipelines

Veo 3.0
Vercel AI Gateway · May 20, 2025
Video model for prompt-guided generation, editing, and motion workflows

OpenAI: Codex Mini
OpenRouter · May 16, 2025
codex-mini-latest is a fine-tuned version of o4-mini specifically for use in Codex CLI. For direct use in the API, we recommend starting with gpt-4.1.

Codex Mini
OpenAI · May 16, 2025
Codex Mini from OpenAI - text input, 200,000 token context

Nous: DeepHermes 3 Mistral 24B Preview
OpenRouter · May 9, 2025
DeepHermes 3 (Mistral 24B Preview) is an instruction-tuned language model by Nous Research based on Mistral-Small-24B, designed for chat, function calling, and advanced multi-turn reasoning. It introduces a dual-mode system that toggles between intuitive chat responses and structured “deep reasoning” mode using special system prompts. Fine-tuned via distillation from R1, it supports structured output (JSON mode) and function call syntax for agent-based applications. DeepHermes 3 supports a **reasoning toggle via system prompt**, allowing users to switch between fast, intuitive responses and deliberate, multi-step reasoning. When activated with the following specific system instruction, the model enters a *"deep thinking"* mode—generating extended chains of thought wrapped in `<think></think>` tags before delivering a final answer. System Prompt: You are a deep thinking AI, you may use extremely long chains of thought to deeply consider the problem and deliberate with yourself via systematic reasoning processes to help come to a correct solution prior to answering. You should enclose your thoughts and internal monologue inside <think> </think> tags, and then provide your solution or response to the problem.

Qwen-Omni Turbo Realtime
Alibaba Cloud · May 8, 2025
Qwen omni model for text, vision, audio, and multimodal agent tasks

Mistral: Mistral Medium 3
OpenRouter · May 7, 2025
Mistral Medium 3 is a high-performance enterprise-grade language model designed to deliver frontier-level capabilities at significantly reduced operational cost. It balances state-of-the-art reasoning and multimodal performance with 8× lower cost...

Google: Gemini 2.5 Pro Preview 05-06
OpenRouter · May 7, 2025
Gemini 2.5 Pro is Google’s state-of-the-art AI model designed for advanced reasoning, coding, mathematics, and scientific tasks. It employs “thinking” capabilities, enabling it to reason through responses with enhanced accuracy...

