Guide

2,496 step-by-step guides to setting up AI models from OpenAI, Anthropic, Google, OpenRouter, and more in TypingMind with your own API key.

Search and filter guides

Showing 511-561 of 713 guides

OpenRouter logo

Google: Gemini 2.5 Pro

OpenRouter · Jun 17, 2025

Gemini 2.5 Pro is Google’s state-of-the-art AI model designed for advanced reasoning, coding, mathematics, and scientific tasks. It employs “thinking” capabilities, enabling it to reason through responses with enhanced accuracy...

1M context+3
View
OpenRouter logo

Google: Gemini 2.5 Pro (batch)

OpenRouter · Jun 17, 2025

Gemini 2.5 Pro is Google’s state-of-the-art AI model designed for advanced reasoning, coding, mathematics, and scientific tasks. It employs “thinking” capabilities, enabling it to reason through responses with enhanced accuracy...

1M context+3
View
OpenRouter logo

MoonshotAI: Kimi Dev 72B

OpenRouter · Jun 16, 2025

Kimi-Dev-72B is an open-source large language model fine-tuned for software engineering and issue resolution tasks. Based on Qwen2.5-72B, it is optimized using large-scale reinforcement learning that applies code patches in real repositories and validates them via full test suite execution—rewarding only correct, robust completions. The model achieves 60.4% on SWE-bench Verified, setting a new benchmark among open-source models for software bug fixing and code reasoning.

131.1K context
View
OpenRouter logo

OpenAI: o3 Pro

OpenRouter · Jun 10, 2025

The o-series of models are trained with reinforcement learning to think before they answer and perform complex reasoning. The o3-pro model uses more compute to think harder and provide consistently...

200K context
View
OpenRouter logo

OpenAI: o3 Pro (batch)

OpenRouter · Jun 10, 2025

The o-series of models are trained with reinforcement learning to think before they answer and perform complex reasoning. The o3-pro model uses more compute to think harder and provide consistently...

200K context
View
OpenRouter logo

xAI: Grok 3 Mini

OpenRouter · Jun 10, 2025

A lightweight model that thinks before responding. Fast, smart, and great for logic-based tasks that do not require deep domain knowledge. The raw thinking traces are accessible.

131.1K context
View
OpenRouter logo

xAI: Grok 3

OpenRouter · Jun 10, 2025

Grok 3 is the latest model from xAI. It's their flagship model that excels at enterprise use cases like data extraction, coding, and text summarization. Possesses deep domain knowledge in...

131.1K context
View
OpenRouter logo

Google: Gemini 2.5 Pro Preview 06-05

OpenRouter · Jun 5, 2025

Gemini 2.5 Pro is Google’s state-of-the-art AI model designed for advanced reasoning, coding, mathematics, and scientific tasks. It employs “thinking” capabilities, enabling it to reason through responses with enhanced accuracy...

1M context+2
View
OpenRouter logo

DeepSeek: DeepSeek R1 0528 Qwen3 8B

OpenRouter · May 29, 2025

DeepSeek-R1-0528 is a lightly upgraded release of DeepSeek R1 that taps more compute and smarter post-training tricks, pushing its reasoning and inference to the brink of flagship models like O3 and Gemini 2.5 Pro. It now tops math, programming, and logic leaderboards, showcasing a step-change in depth-of-thought. The distilled variant, DeepSeek-R1-0528-Qwen3-8B, transfers this chain-of-thought into an 8 B-parameter form, beating standard Qwen3 8B by +10 pp and tying the 235 B “thinking” giant on AIME 2024.

128K context
View
OpenRouter logo

DeepSeek: R1 0528 (free)

OpenRouter · May 28, 2025

May 28th update to the [original DeepSeek R1](/deepseek/deepseek-r1) Performance on par with [OpenAI o1](/openai/o1), but open-sourced and with fully open reasoning tokens. It's 671B parameters in size, with 37B active in an inference pass. Fully open-source model.

163.8K context
View
OpenRouter logo

DeepSeek: R1 0528

OpenRouter · May 28, 2025

May 28th update to the [original DeepSeek R1](/deepseek/deepseek-r1) Performance on par with [OpenAI o1](/openai/o1), but open-sourced and with fully open reasoning tokens. It's 671B parameters in size, with 37B active...

163.8K context
View
OpenRouter logo

Anthropic: Claude Opus 4

OpenRouter · May 22, 2025

Claude Opus 4 is benchmarked as the world’s best coding model, at time of release, bringing sustained performance on complex, long-running tasks and agent workflows. It sets new benchmarks in...

200K context
View
OpenRouter logo

Anthropic: Claude Sonnet 4

OpenRouter · May 22, 2025

Claude Sonnet 4 significantly enhances the capabilities of its predecessor, Sonnet 3.7, excelling in both coding and reasoning tasks with improved precision and controllability. Achieving state-of-the-art performance on SWE-bench (72.7%),...

200K context
View
OpenRouter logo

Mistral: Devstral Small 2505

OpenRouter · May 21, 2025

Devstral-Small-2505 is a 24B parameter agentic LLM fine-tuned from Mistral-Small-3.1, jointly developed by Mistral AI and All Hands AI for advanced software engineering tasks. It is optimized for codebase exploration, multi-file editing, and integration into coding agents, achieving state-of-the-art results on SWE-Bench Verified (46.8%). Devstral supports a 128k context window and uses a custom Tekken tokenizer. It is text-only, with the vision encoder removed, and is suitable for local deployment on high-end consumer hardware (e.g., RTX 4090, 32GB RAM Macs). Devstral is best used in agentic workflows via the OpenHands scaffold and is compatible with inference frameworks like vLLM, Transformers, and Ollama. It is released under the Apache 2.0 license.

128K context
View
OpenRouter logo

Google: Gemma 3n 4B (free)

OpenRouter · May 20, 2025

Gemma 3n E4B-it is optimized for efficient execution on mobile and low-resource devices, such as phones, laptops, and tablets. It supports multimodal inputs—including text, visual data, and audio—enabling diverse tasks...

8.2K context
View
OpenRouter logo

Google: Gemma 3n 4B

OpenRouter · May 20, 2025

Gemma 3n E4B-it is optimized for efficient execution on mobile and low-resource devices, such as phones, laptops, and tablets. It supports multimodal inputs—including text, visual data, and audio—enabling diverse tasks...

32.8K context
View
OpenRouter logo

OpenAI: Codex Mini

OpenRouter · May 16, 2025

codex-mini-latest is a fine-tuned version of o4-mini specifically for use in Codex CLI. For direct use in the API, we recommend starting with gpt-4.1.

200K context
View
OpenRouter logo

Nous: DeepHermes 3 Mistral 24B Preview

OpenRouter · May 9, 2025

DeepHermes 3 (Mistral 24B Preview) is an instruction-tuned language model by Nous Research based on Mistral-Small-24B, designed for chat, function calling, and advanced multi-turn reasoning. It introduces a dual-mode system that toggles between intuitive chat responses and structured “deep reasoning” mode using special system prompts. Fine-tuned via distillation from R1, it supports structured output (JSON mode) and function call syntax for agent-based applications. DeepHermes 3 supports a **reasoning toggle via system prompt**, allowing users to switch between fast, intuitive responses and deliberate, multi-step reasoning. When activated with the following specific system instruction, the model enters a *"deep thinking"* mode—generating extended chains of thought wrapped in `<think></think>` tags before delivering a final answer. System Prompt: You are a deep thinking AI, you may use extremely long chains of thought to deeply consider the problem and deliberate with yourself via systematic reasoning processes to help come to a correct solution prior to answering. You should enclose your thoughts and internal monologue inside <think> </think> tags, and then provide your solution or response to the problem.

32.8K context
View
OpenRouter logo

Mistral: Mistral Medium 3

OpenRouter · May 7, 2025

Mistral Medium 3 is a high-performance enterprise-grade language model designed to deliver frontier-level capabilities at significantly reduced operational cost. It balances state-of-the-art reasoning and multimodal performance with 8× lower cost...

131.1K context
View
OpenRouter logo

Google: Gemini 2.5 Pro Preview 05-06

OpenRouter · May 7, 2025

Gemini 2.5 Pro is Google’s state-of-the-art AI model designed for advanced reasoning, coding, mathematics, and scientific tasks. It employs “thinking” capabilities, enabling it to reason through responses with enhanced accuracy...

1M context+3
View
OpenRouter logo

Arcee AI: Spotlight

OpenRouter · May 5, 2025

Spotlight is a 7‑billion‑parameter vision‑language model derived from Qwen 2.5‑VL and fine‑tuned by Arcee AI for tight image‑text grounding tasks. It offers a 32 k‑token context window, enabling rich multimodal...

131.1K context
View
OpenRouter logo

Arcee AI: Maestro Reasoning

OpenRouter · May 5, 2025

Maestro Reasoning is Arcee's flagship analysis model: a 32 B‑parameter derivative of Qwen 2.5‑32 B tuned with DPO and chain‑of‑thought RL for step‑by‑step logic. Compared to the earlier 7 B...

131.1K context
View
OpenRouter logo

Arcee AI: Virtuoso Large

OpenRouter · May 5, 2025

Virtuoso‑Large is Arcee's top‑tier general‑purpose LLM at 72 B parameters, tuned to tackle cross‑domain reasoning, creative writing and enterprise QA. Unlike many 70 B peers, it retains the 128 k...

131.1K context
View
OpenRouter logo

Arcee AI: Coder Large

OpenRouter · May 5, 2025

Coder‑Large is a 32 B‑parameter offspring of Qwen 2.5‑Instruct that has been further trained on permissively‑licensed GitHub, CodeSearchNet and synthetic bug‑fix corpora. It supports a 32k context window, enabling multi‑file...

32.8K context
View
OpenRouter logo

Microsoft: Phi 4 Reasoning Plus

OpenRouter · May 1, 2025

Phi-4-reasoning-plus is an enhanced 14B parameter model from Microsoft, fine-tuned from Phi-4 with additional reinforcement learning to boost accuracy on math, science, and code reasoning tasks. It uses the same dense decoder-only transformer architecture as Phi-4, but generates longer, more comprehensive outputs structured into a step-by-step reasoning trace and final answer. While it offers improved benchmark scores over Phi-4-reasoning across tasks like AIME, OmniMath, and HumanEvalPlus, its responses are typically ~50% longer, resulting in higher latency. Designed for English-only applications, it is well-suited for structured reasoning workflows where output quality takes priority over response speed.

32.8K context
View
OpenRouter logo

Inception: Mercury Coder

OpenRouter · Apr 30, 2025

Mercury Coder is the first diffusion large language model (dLLM). Applying a breakthrough discrete diffusion approach, the model runs 5-10x faster than even speed optimized models like Claude 3.5 Haiku...

128K context
View
OpenRouter logo

Qwen: Qwen3 4B (free)

OpenRouter · Apr 30, 2025

Qwen3-4B is a 4 billion parameter dense language model from the Qwen3 series, designed to support both general-purpose and reasoning-intensive tasks. It introduces a dual-mode architecture—thinking and non-thinking—allowing dynamic switching between high-precision logical reasoning and efficient dialogue generation. This makes it well-suited for multi-turn chat, instruction following, and complex agent workflows.

41K context
View
OpenRouter logo

DeepSeek: DeepSeek Prover V2

OpenRouter · Apr 30, 2025

DeepSeek Prover V2 is a 671B parameter model, speculated to be geared towards logic and mathematics. Likely an upgrade from [DeepSeek-Prover-V1.5](https://huggingface.co/deepseek-ai/DeepSeek-Prover-V1.5-RL) Not much is known about the model yet, as DeepSeek released it on Hugging Face without an announcement or description.

163.8K context
View
OpenRouter logo

Meta: Llama Guard 4 12B

OpenRouter · Apr 30, 2025

Llama Guard 4 is a Llama 4 Scout-derived multimodal pretrained model, fine-tuned for content safety classification. Similar to previous versions, it can be used to classify content in both LLM...

163.8K context
View
OpenRouter logo

Qwen: Qwen3 30B A3B

OpenRouter · Apr 28, 2025

Qwen3, the latest generation in the Qwen large language model series, features both dense and mixture-of-experts (MoE) architectures to excel in reasoning, multilingual support, and advanced agent tasks. Its unique...

131.1K context
View
OpenRouter logo

Qwen: Qwen3 8B

OpenRouter · Apr 28, 2025

Qwen3-8B is a dense 8.2B parameter causal language model from the Qwen3 series, designed for both reasoning-heavy tasks and efficient dialogue. It supports seamless switching between "thinking" mode for math,...

131.1K context
View
OpenRouter logo

Qwen: Qwen3 14B

OpenRouter · Apr 28, 2025

Qwen3-14B is a dense 14.8B parameter causal language model from the Qwen3 series, designed for both complex reasoning and efficient dialogue. It supports seamless switching between a "thinking" mode for...

131.1K context
View
OpenRouter logo

Qwen: Qwen3 32B

OpenRouter · Apr 28, 2025

Qwen3-32B is a dense 32.8B parameter causal language model from the Qwen3 series, optimized for both complex reasoning and efficient dialogue. It supports seamless switching between a "thinking" mode for...

131.1K context
View
OpenRouter logo

Qwen: Qwen3 235B A22B

OpenRouter · Apr 28, 2025

Qwen3-235B-A22B is a 235B parameter mixture-of-experts (MoE) model developed by Qwen, activating 22B parameters per forward pass. It supports seamless switching between a "thinking" mode for complex reasoning, math, and...

131.1K context
View
OpenRouter logo

TNG: DeepSeek R1T Chimera (free)

OpenRouter · Apr 27, 2025

DeepSeek-R1T-Chimera is created by merging DeepSeek-R1 and DeepSeek-V3 (0324), combining the reasoning capabilities of R1 with the token efficiency improvements of V3. It is based on a DeepSeek-MoE Transformer architecture and is optimized for general text generation tasks. The model merges pretrained weights from both source models to balance performance across reasoning, efficiency, and instruction-following tasks. It is released under the MIT license and intended for research and commercial use.

163.8K context
View
OpenRouter logo

TNG: DeepSeek R1T Chimera

OpenRouter · Apr 27, 2025

DeepSeek-R1T-Chimera is created by merging DeepSeek-R1 and DeepSeek-V3 (0324), combining the reasoning capabilities of R1 with the token efficiency improvements of V3. It is based on a DeepSeek-MoE Transformer architecture and is optimized for general text generation tasks. The model merges pretrained weights from both source models to balance performance across reasoning, efficiency, and instruction-following tasks. It is released under the MIT license and intended for research and commercial use.

163.8K context
View
OpenRouter logo

OpenAI: o4 Mini High

OpenRouter · Apr 16, 2025

OpenAI o4-mini-high is the same model as [o4-mini](/openai/o4-mini) with reasoning_effort set to high. OpenAI o4-mini is a compact reasoning model in the o-series, optimized for fast, cost-efficient performance while retaining...

200K context+1
View
OpenRouter logo

OpenAI: o4 Mini High (batch)

OpenRouter · Apr 16, 2025

OpenAI o4-mini-high is the same model as [o4-mini](/openai/o4-mini) with reasoning_effort set to high. OpenAI o4-mini is a compact reasoning model in the o-series, optimized for fast, cost-efficient performance while retaining...

200K context+1
View
OpenRouter logo

OpenAI: o3

OpenRouter · Apr 16, 2025

o3 is a well-rounded and powerful model across domains. It sets a new standard for math, science, coding, and visual reasoning tasks. It also excels at technical writing and instruction-following....

200K context+1
View
OpenRouter logo

OpenAI: o3 (batch)

OpenRouter · Apr 16, 2025

o3 is a well-rounded and powerful model across domains. It sets a new standard for math, science, coding, and visual reasoning tasks. It also excels at technical writing and instruction-following....

200K context+1
View
OpenRouter logo

OpenAI: o4 Mini

OpenRouter · Apr 16, 2025

OpenAI o4-mini is a compact reasoning model in the o-series, optimized for fast, cost-efficient performance while retaining strong multimodal and agentic capabilities. It supports tool use and demonstrates competitive reasoning...

200K context+1
View
OpenRouter logo

OpenAI: o4 Mini (batch)

OpenRouter · Apr 16, 2025

OpenAI o4-mini is a compact reasoning model in the o-series, optimized for fast, cost-efficient performance while retaining strong multimodal and agentic capabilities. It supports tool use and demonstrates competitive reasoning...

200K context+1
View
OpenRouter logo

Qwen: Qwen2.5 Coder 7B Instruct

OpenRouter · Apr 15, 2025

Qwen2.5-Coder-7B-Instruct is a 7B parameter instruction-tuned language model optimized for code-related tasks such as code generation, reasoning, and bug fixing. Based on the Qwen2.5 architecture, it incorporates enhancements like RoPE,...

32.8K context
View
OpenRouter logo

OpenAI: GPT-4.1

OpenRouter · Apr 14, 2025

GPT-4.1 is a flagship large language model optimized for advanced instruction following, real-world software engineering, and long-context reasoning. It supports a 1 million token context window and outperforms GPT-4o and...

1M context
View
OpenRouter logo

OpenAI: GPT-4.1 (batch)

OpenRouter · Apr 14, 2025

GPT-4.1 is a flagship large language model optimized for advanced instruction following, real-world software engineering, and long-context reasoning. It supports a 1 million token context window and outperforms GPT-4o and...

1M context
View
OpenRouter logo

OpenAI: GPT-4.1 Mini

OpenRouter · Apr 14, 2025

GPT-4.1 Mini is a mid-sized model delivering performance competitive with GPT-4o at substantially lower latency and cost. It retains a 1 million token context window and scores 45.1% on hard...

1M context
View
OpenRouter logo

OpenAI: GPT-4.1 Mini (batch)

OpenRouter · Apr 14, 2025

GPT-4.1 Mini is a mid-sized model delivering performance competitive with GPT-4o at substantially lower latency and cost. It retains a 1 million token context window and scores 45.1% on hard...

1M context
View
OpenRouter logo

OpenAI: GPT-4.1 Nano

OpenRouter · Apr 14, 2025

For tasks that demand low latency, GPT‑4.1 nano is the fastest and cheapest model in the GPT-4.1 series. It delivers exceptional performance at a small size with its 1 million...

1M context
View
OpenRouter logo

OpenAI: GPT-4.1 Nano (batch)

OpenRouter · Apr 14, 2025

For tasks that demand low latency, GPT‑4.1 nano is the fastest and cheapest model in the GPT-4.1 series. It delivers exceptional performance at a small size with its 1 million...

1M context
View
OpenRouter logo

EleutherAI: Llemma 7b

OpenRouter · Apr 14, 2025

Llemma 7B is a language model for mathematics. It was initialized with Code Llama 7B weights, and trained on the Proof-Pile-2 for 200B tokens. Llemma models are particularly strong at...

4.1K context
View
OpenRouter logo

AlfredPros: CodeLLaMa 7B Instruct Solidity

OpenRouter · Apr 14, 2025

A finetuned 7 billion parameters Code LLaMA - Instruct model to generate Solidity smart contract using 4-bit QLoRA finetuning provided by PEFT library.

4.1K context
View