Guide

2,496 step-by-step guides to setting up AI models from OpenAI, Anthropic, Google, OpenRouter, and more in TypingMind with your own API key.

Search and filter guides

Showing 1,888-1,938 of 2,496 guides

Vercel AI Gateway logo

o3 Pro

Vercel AI Gateway · Jun 10, 2025

High-effort o3 tier for difficult technical reasoning and careful answers

200K context
View
OpenRouter logo

Google: Gemini 2.5 Pro Preview 06-05

OpenRouter · Jun 5, 2025

Gemini 2.5 Pro is Google’s state-of-the-art AI model designed for advanced reasoning, coding, mathematics, and scientific tasks. It employs “thinking” capabilities, enabling it to reason through responses with enhanced accuracy...

1M context+2
View
Google logo

Gemini 2.5 Pro Preview 06-05

Google · Jun 5, 2025

Gemini 2.5 Pro Preview 06-05 from Google - text, image, audio, video, pdf input, 1,048,576 token context

1M context+3
View
Vercel AI Gateway logo

Qwen3 Embedding 4B

Vercel AI Gateway · Jun 5, 2025

Embedding model for semantic search, retrieval, clustering, and ranking pipelines

32.8K context
View
Vercel AI Gateway logo

Qwen3 Embedding 8B

Vercel AI Gateway · Jun 5, 2025

Embedding model for semantic search, retrieval, clustering, and ranking pipelines

32.8K context
View
OpenRouter logo

DeepSeek: DeepSeek R1 0528 Qwen3 8B

OpenRouter · May 29, 2025

DeepSeek-R1-0528 is a lightly upgraded release of DeepSeek R1 that taps more compute and smarter post-training tricks, pushing its reasoning and inference to the brink of flagship models like O3 and Gemini 2.5 Pro. It now tops math, programming, and logic leaderboards, showcasing a step-change in depth-of-thought. The distilled variant, DeepSeek-R1-0528-Qwen3-8B, transfers this chain-of-thought into an 8 B-parameter form, beating standard Qwen3 8B by +10 pp and tying the 235 B “thinking” giant on AIME 2024.

128K context
View
Groq logo

Llama Prompt Guard 2 22M

Groq · May 29, 2025

Safety model for policy screening, moderation, and risk-aware routing workflows

512 context
View
Groq logo

Prompt Guard 2 86M

Groq · May 29, 2025

Safety model for policy screening, moderation, and risk-aware routing workflows

512 context
View
Vercel AI Gateway logo

FLUX.1 Kontext Max

Vercel AI Gateway · May 29, 2025

Image model for prompt-driven generation, editing, and visual design workflows

512 context
View
Vercel AI Gateway logo

FLUX.1 Kontext Pro

Vercel AI Gateway · May 29, 2025

Image model for prompt-driven generation, editing, and visual design workflows

512 context
View
Novita logo

DeepSeek R1 0528 Qwen3 8B

Novita · May 29, 2025

DeepSeek reasoning model for multi-step analysis, math, coding, and tools

128K context
View
OpenRouter logo

DeepSeek: R1 0528 (free)

OpenRouter · May 28, 2025

May 28th update to the [original DeepSeek R1](/deepseek/deepseek-r1) Performance on par with [OpenAI o1](/openai/o1), but open-sourced and with fully open reasoning tokens. It's 671B parameters in size, with 37B active in an inference pass. Fully open-source model.

163.8K context
View
OpenRouter logo

DeepSeek: R1 0528

OpenRouter · May 28, 2025

May 28th update to the [original DeepSeek R1](/deepseek/deepseek-r1) Performance on par with [OpenAI o1](/openai/o1), but open-sourced and with fully open reasoning tokens. It's 671B parameters in size, with 37B active...

163.8K context
View
Fireworks AI logo

Deepseek R1 05/28

Fireworks AI · May 28, 2025

Deepseek R1 05/28 from Fireworks AI - text input, 160,000 token context

160K context
View
DeepInfra logo

DeepSeek-R1-0528

DeepInfra · May 28, 2025

DeepSeek reasoning model for multi-step analysis, math, coding, and tools

163.8K context
View
Hugging Face logo

DeepSeek-R1-0528

Hugging Face · May 28, 2025

DeepSeek reasoning model for multi-step analysis, math, coding, and tools

163.8K context
View
Vercel AI Gateway logo

Codestral Embed

Vercel AI Gateway · May 28, 2025

Embedding model for semantic search, retrieval, clustering, and ranking pipelines

8.2K context
View
Novita logo

DeepSeek R1 0528

Novita · May 28, 2025

DeepSeek reasoning model for multi-step analysis, math, coding, and tools

163.8K context
View
OpenRouter logo

Anthropic: Claude Opus 4

OpenRouter · May 22, 2025

Claude Opus 4 is benchmarked as the world’s best coding model, at time of release, bringing sustained performance on complex, long-running tasks and agent workflows. It sets new benchmarks in...

200K context
View
OpenRouter logo

Anthropic: Claude Sonnet 4

OpenRouter · May 22, 2025

Claude Sonnet 4 significantly enhances the capabilities of its predecessor, Sonnet 3.7, excelling in both coding and reasoning tasks with improved precision and controllability. Achieving state-of-the-art performance on SWE-bench (72.7%),...

200K context
View
Anthropic logo

Claude Opus 4 (latest)

Anthropic · May 22, 2025

Claude Opus 4 (latest) from Anthropic - text, image, pdf input, 200,000 token context

200K context
View
Anthropic logo

Claude Sonnet 4

Anthropic · May 22, 2025

Claude Sonnet 4 from Anthropic - text, image, pdf input, 200,000 token context

200K context
View
Anthropic logo

Claude Opus 4

Anthropic · May 22, 2025

Claude Opus 4 from Anthropic - text, image, pdf input, 200,000 token context

200K context
View
Anthropic logo

Claude Sonnet 4 (latest)

Anthropic · May 22, 2025

Claude Sonnet 4 (latest) from Anthropic - text, image, pdf input, 200,000 token context

200K context
View
AWS Bedrock logo

Claude Opus 4 (US)

AWS Bedrock · May 22, 2025

Claude Opus 4 (US) from AWS Bedrock - text, image, pdf input, 200,000 token context

200K context
View
AWS Bedrock logo

Claude Sonnet 4

AWS Bedrock · May 22, 2025

Claude Sonnet 4 from AWS Bedrock - text, image, pdf input, 200,000 token context

200K context
View
AWS Bedrock logo

Claude Opus 4

AWS Bedrock · May 22, 2025

Claude Opus 4 from AWS Bedrock - text, image, pdf input, 200,000 token context

200K context
View
AWS Bedrock logo

Claude Sonnet 4 (US)

AWS Bedrock · May 22, 2025

Balanced Claude model for coding, analysis, agent workflows, and cost control

200K context
View
AWS Bedrock logo

Claude Sonnet 4 (Global)

AWS Bedrock · May 22, 2025

Balanced Claude model for coding, analysis, agent workflows, and cost control

200K context
View
AWS Bedrock logo

Claude Sonnet 4 (EU)

AWS Bedrock · May 22, 2025

Balanced Claude model for coding, analysis, agent workflows, and cost control

200K context
View
AWS Bedrock logo

Claude Sonnet 4 (APAC)

AWS Bedrock · May 22, 2025

Balanced Claude model for coding, analysis, agent workflows, and cost control

200K context
View
Vercel AI Gateway logo

Claude Sonnet 4

Vercel AI Gateway · May 22, 2025

Balanced Claude model for coding, analysis, agent workflows, and cost control

1M context
View
Vercel AI Gateway logo

Claude Opus 4

Vercel AI Gateway · May 22, 2025

Flagship Claude model for deep reasoning, coding, and long-horizon agents

200K context
View
OpenRouter logo

Mistral: Devstral Small 2505

OpenRouter · May 21, 2025

Devstral-Small-2505 is a 24B parameter agentic LLM fine-tuned from Mistral-Small-3.1, jointly developed by Mistral AI and All Hands AI for advanced software engineering tasks. It is optimized for codebase exploration, multi-file editing, and integration into coding agents, achieving state-of-the-art results on SWE-Bench Verified (46.8%). Devstral supports a 128k context window and uses a custom Tekken tokenizer. It is text-only, with the vision encoder removed, and is suitable for local deployment on high-end consumer hardware (e.g., RTX 4090, 32GB RAM Macs). Devstral is best used in agentic workflows via the OpenHands scaffold and is compatible with inference frameworks like vLLM, Transformers, and Ollama. It is released under the Apache 2.0 license.

128K context
View
Setup guide logo

Venice

Setup guide · May 21, 2025

Venice.ai is a privacy-first, uncensored AI platform providing access to leading open-source models without data collection, content restrictions, or conversation logging. Key features include private access to multiple models (Llama, DeepSeek, StableDiffusion, GPT, Claude, Gemini), zero data retention with local chat history storage, decentralized GPU infrastructure preventing single-entity data access, uncensored text generation and image creation, web-enabled research and PDF analysis, and a private API with no logging. The platform serves over 1.3 million users and is integrating Web3 capabilities through partnerships like Warden Protocol for on-chain AI and censorship resistance. Venice offers both free and Pro tiers with mobile apps available, focusing on creative freedom and honest answers without guardrails while maintaining complete conversation privacy through decentralized processing.

API key setup
View
OpenRouter logo

Google: Gemma 3n 4B (free)

OpenRouter · May 20, 2025

Gemma 3n E4B-it is optimized for efficient execution on mobile and low-resource devices, such as phones, laptops, and tablets. It supports multimodal inputs—including text, visual data, and audio—enabling diverse tasks...

8.2K context
View
OpenRouter logo

Google: Gemma 3n 4B

OpenRouter · May 20, 2025

Gemma 3n E4B-it is optimized for efficient execution on mobile and low-resource devices, such as phones, laptops, and tablets. It supports multimodal inputs—including text, visual data, and audio—enabling diverse tasks...

32.8K context
View
Google logo

Gemini Embedding 001

Google · May 20, 2025

Embedding model for semantic search, retrieval, clustering, and ranking pipelines

2K context
View
Google logo

Gemini 2.5 Flash Preview 05-20

Google · May 20, 2025

Gemini 2.5 Flash Preview 05-20 from Google - text, image, audio, video, pdf input, 1,048,576 token context

1M context+3
View
Google logo

Gemma 3n 4B

Google · May 20, 2025

Gemma 3n 4B from Google - text input, 8,192 token context

8.2K context
View
Together AI logo

Gemma 3N E4B Instruct

Together AI · May 20, 2025

Open Gemma instruction model for efficient chat and self-hosted deployments

32.8K context
View
Vercel AI Gateway logo

voyage-3.5

Vercel AI Gateway · May 20, 2025

General-purpose chat model for instruction following, writing, and analysis

8.2K context
View
Vercel AI Gateway logo

voyage-3.5-lite

Vercel AI Gateway · May 20, 2025

Efficient model for low-latency assistance, extraction, and routine automation

8.2K context
View
Vercel AI Gateway logo

Gemini Embedding 001

Vercel AI Gateway · May 20, 2025

Embedding model for semantic search, retrieval, clustering, and ranking pipelines

8.2K context
View
Vercel AI Gateway logo

Veo 3.0

Vercel AI Gateway · May 20, 2025

Video model for prompt-guided generation, editing, and motion workflows

View
OpenRouter logo

OpenAI: Codex Mini

OpenRouter · May 16, 2025

codex-mini-latest is a fine-tuned version of o4-mini specifically for use in Codex CLI. For direct use in the API, we recommend starting with gpt-4.1.

200K context
View
OpenAI logo

Codex Mini

OpenAI · May 16, 2025

Codex Mini from OpenAI - text input, 200,000 token context

200K context
View
OpenRouter logo

Nous: DeepHermes 3 Mistral 24B Preview

OpenRouter · May 9, 2025

DeepHermes 3 (Mistral 24B Preview) is an instruction-tuned language model by Nous Research based on Mistral-Small-24B, designed for chat, function calling, and advanced multi-turn reasoning. It introduces a dual-mode system that toggles between intuitive chat responses and structured “deep reasoning” mode using special system prompts. Fine-tuned via distillation from R1, it supports structured output (JSON mode) and function call syntax for agent-based applications. DeepHermes 3 supports a **reasoning toggle via system prompt**, allowing users to switch between fast, intuitive responses and deliberate, multi-step reasoning. When activated with the following specific system instruction, the model enters a *"deep thinking"* mode—generating extended chains of thought wrapped in `<think></think>` tags before delivering a final answer. System Prompt: You are a deep thinking AI, you may use extremely long chains of thought to deeply consider the problem and deliberate with yourself via systematic reasoning processes to help come to a correct solution prior to answering. You should enclose your thoughts and internal monologue inside <think> </think> tags, and then provide your solution or response to the problem.

32.8K context
View
Alibaba Cloud logo

Qwen-Omni Turbo Realtime

Alibaba Cloud · May 8, 2025

Qwen omni model for text, vision, audio, and multimodal agent tasks

32.8K context
View
OpenRouter logo

Mistral: Mistral Medium 3

OpenRouter · May 7, 2025

Mistral Medium 3 is a high-performance enterprise-grade language model designed to deliver frontier-level capabilities at significantly reduced operational cost. It balances state-of-the-art reasoning and multimodal performance with 8× lower cost...

131.1K context
View
OpenRouter logo

Google: Gemini 2.5 Pro Preview 05-06

OpenRouter · May 7, 2025

Gemini 2.5 Pro is Google’s state-of-the-art AI model designed for advanced reasoning, coding, mathematics, and scientific tasks. It employs “thinking” capabilities, enabling it to reason through responses with enhanced accuracy...

1M context+3
View