Guide

2,496 step-by-step guides to setting up AI models from OpenAI, Anthropic, Google, OpenRouter, and more in TypingMind with your own API key.

Search and filter guides

Showing 460-510 of 713 guides

OpenRouter logo

OpenAI: GPT-5 Mini

OpenRouter · Aug 7, 2025

GPT-5 Mini is a compact version of GPT-5, designed to handle lighter-weight reasoning tasks. It provides the same instruction-following and safety-tuning benefits as GPT-5, but with reduced latency and cost....

400K context+1
View
OpenRouter logo

OpenAI: GPT-5 Mini (batch)

OpenRouter · Aug 7, 2025

GPT-5 Mini is a compact version of GPT-5, designed to handle lighter-weight reasoning tasks. It provides the same instruction-following and safety-tuning benefits as GPT-5, but with reduced latency and cost....

400K context+1
View
OpenRouter logo

OpenAI: GPT-5 Nano

OpenRouter · Aug 7, 2025

GPT-5-Nano is the smallest and fastest variant in the GPT-5 system, optimized for developer tools, rapid interactions, and ultra-low latency environments. While limited in reasoning depth compared to its larger...

400K context+1
View
OpenRouter logo

OpenAI: GPT-5 Nano (batch)

OpenRouter · Aug 7, 2025

GPT-5-Nano is the smallest and fastest variant in the GPT-5 system, optimized for developer tools, rapid interactions, and ultra-low latency environments. While limited in reasoning depth compared to its larger...

400K context+1
View
OpenRouter logo

OpenAI: gpt-oss-120b (free)

OpenRouter · Aug 5, 2025

gpt-oss-120b is an open-weight, 117B-parameter Mixture-of-Experts (MoE) language model from OpenAI designed for high-reasoning, agentic, and general-purpose production use cases. It activates 5.1B parameters per forward pass and is optimized...

131.1K context
View
OpenRouter logo

OpenAI: gpt-oss-120b

OpenRouter · Aug 5, 2025

gpt-oss-120b is an open-weight, 117B-parameter Mixture-of-Experts (MoE) language model from OpenAI designed for high-reasoning, agentic, and general-purpose production use cases. It activates 5.1B parameters per forward pass and is optimized...

131.1K context
View
OpenRouter logo

OpenAI: gpt-oss-120b (exacto)

OpenRouter · Aug 5, 2025

gpt-oss-120b is an open-weight, 117B-parameter Mixture-of-Experts (MoE) language model from OpenAI designed for high-reasoning, agentic, and general-purpose production use cases. It activates 5.1B parameters per forward pass and is optimized to run on a single H100 GPU with native MXFP4 quantization. The model supports configurable reasoning depth, full chain-of-thought access, and native tool use, including function calling, browsing, and structured output generation.

131.1K context
View
OpenRouter logo

OpenAI: gpt-oss-120b (batch)

OpenRouter · Aug 5, 2025

gpt-oss-120b is an open-weight, 117B-parameter Mixture-of-Experts (MoE) language model from OpenAI designed for high-reasoning, agentic, and general-purpose production use cases. It activates 5.1B parameters per forward pass and is optimized...

131.1K context
View
OpenRouter logo

OpenAI: gpt-oss-20b (free)

OpenRouter · Aug 5, 2025

gpt-oss-20b is an open-weight 21B parameter model released by OpenAI under the Apache 2.0 license. It uses a Mixture-of-Experts (MoE) architecture with 3.6B active parameters per forward pass, optimized for...

131.1K context
View
OpenRouter logo

OpenAI: gpt-oss-20b

OpenRouter · Aug 5, 2025

gpt-oss-20b is an open-weight 21B parameter model released by OpenAI under the Apache 2.0 license. It uses a Mixture-of-Experts (MoE) architecture with 3.6B active parameters per forward pass, optimized for...

131.1K context
View
OpenRouter logo

OpenAI: gpt-oss-20b (batch)

OpenRouter · Aug 5, 2025

gpt-oss-20b is an open-weight 21B parameter model released by OpenAI under the Apache 2.0 license. It uses a Mixture-of-Experts (MoE) architecture with 3.6B active parameters per forward pass, optimized for...

131.1K context
View
OpenRouter logo

Anthropic: Claude Opus 4.1

OpenRouter · Aug 5, 2025

Claude Opus 4.1 is an updated version of Anthropic’s flagship model, offering improved performance in coding, reasoning, and agentic tasks. It achieves 74.5% on SWE-bench Verified and shows notable gains...

200K context
View
OpenRouter logo

Anthropic: Claude Opus 4.1 (batch)

OpenRouter · Aug 5, 2025

Claude Opus 4.1 is an updated version of Anthropic’s flagship model, offering improved performance in coding, reasoning, and agentic tasks. It achieves 74.5% on SWE-bench Verified and shows notable gains...

200K context+1
View
OpenRouter logo

Mistral: Codestral 2508

OpenRouter · Aug 1, 2025

Mistral's cutting-edge language model for coding released end of July 2025. Codestral specializes in low-latency, high-frequency tasks such as fill-in-the-middle (FIM), code correction and test generation. [Blog Post](https://mistral.ai/news/codestral-25-08)

256K context
View
OpenRouter logo

Mistral: Codestral 2508 (batch)

OpenRouter · Aug 1, 2025

Mistral's cutting-edge language model for coding released end of July 2025. Codestral specializes in low-latency, high-frequency tasks such as fill-in-the-middle (FIM), code correction and test generation. [Blog Post](https://mistral.ai/news/codestral-25-08)

256K context
View
OpenRouter logo

Qwen: Qwen3 Coder 30B A3B Instruct

OpenRouter · Jul 31, 2025

Qwen3-Coder-30B-A3B-Instruct is a 30.5B parameter Mixture-of-Experts (MoE) model with 128 experts (8 active per forward pass), designed for advanced code generation, repository-scale understanding, and agentic tool use. Built on the...

262.1K context
View
OpenRouter logo

Qwen: Qwen3 30B A3B Instruct 2507

OpenRouter · Jul 29, 2025

Qwen3-30B-A3B-Instruct-2507 is a 30.5B-parameter mixture-of-experts language model from Qwen, with 3.3B active parameters per inference. It operates in non-thinking mode and is designed for high-quality instruction following, multilingual understanding, and...

262.1K context
View
OpenRouter logo

Z.ai: GLM 4.5

OpenRouter · Jul 25, 2025

GLM-4.5 is our latest flagship foundation model, purpose-built for agent-based applications. It leverages a Mixture-of-Experts (MoE) architecture and supports a context length of up to 128k tokens. GLM-4.5 delivers significantly...

131.1K context
View
OpenRouter logo

Z.ai: GLM 4.5 Air (free)

OpenRouter · Jul 25, 2025

GLM-4.5-Air is the lightweight variant of our latest flagship model family, also purpose-built for agent-centric applications. Like GLM-4.5, it adopts the Mixture-of-Experts (MoE) architecture but with a more compact parameter...

131.1K context
View
OpenRouter logo

Z.ai: GLM 4.5 Air

OpenRouter · Jul 25, 2025

GLM-4.5-Air is the lightweight variant of our latest flagship model family, also purpose-built for agent-centric applications. Like GLM-4.5, it adopts the Mixture-of-Experts (MoE) architecture but with a more compact parameter...

131.1K context
View
OpenRouter logo

Qwen: Qwen3 235B A22B Thinking 2507

OpenRouter · Jul 25, 2025

Qwen3-235B-A22B-Thinking-2507 is a high-performance, open-weight Mixture-of-Experts (MoE) language model optimized for complex reasoning tasks. It activates 22B of its 235B parameters per forward pass and natively supports up to 262,144...

131.1K context
View
OpenRouter logo

Z.ai: GLM 4 32B

OpenRouter · Jul 24, 2025

GLM 4 32B is a cost-effective foundation language model. It can efficiently perform complex tasks and has significantly enhanced capabilities in tool use, online search, and code-related intelligent tasks. It...

128K context
View
OpenRouter logo

Qwen: Qwen3 Coder 480B A35B (free)

OpenRouter · Jul 23, 2025

Qwen3-Coder-480B-A35B-Instruct is a Mixture-of-Experts (MoE) code generation model developed by the Qwen team. It is optimized for agentic coding tasks such as function calling, tool use, and long-context reasoning over...

1M context
View
OpenRouter logo

Qwen: Qwen3 Coder 480B A35B

OpenRouter · Jul 23, 2025

Qwen3-Coder-480B-A35B-Instruct is a Mixture-of-Experts (MoE) code generation model developed by the Qwen team. It is optimized for agentic coding tasks such as function calling, tool use, and long-context reasoning over...

262.1K context
View
OpenRouter logo

Qwen: Qwen3 Coder 480B A35B (exacto)

OpenRouter · Jul 23, 2025

Qwen3-Coder-480B-A35B-Instruct is a Mixture-of-Experts (MoE) code generation model developed by the Qwen team. It is optimized for agentic coding tasks such as function calling, tool use, and long-context reasoning over repositories. The model features 480 billion total parameters, with 35 billion active per forward pass (8 out of 160 experts). Pricing for the Alibaba endpoints varies by context length. Once a request is greater than 128k input tokens, the higher pricing is used.

262.1K context
View
OpenRouter logo

ByteDance: UI-TARS 7B

OpenRouter · Jul 22, 2025

UI-TARS-1.5 is a multimodal vision-language agent optimized for GUI-based environments, including desktop interfaces, web browsers, mobile systems, and games. Built by ByteDance, it builds upon the UI-TARS framework with reinforcement...

128K context
View
OpenRouter logo

Google: Gemini 2.5 Flash Lite

OpenRouter · Jul 22, 2025

Gemini 2.5 Flash-Lite is a lightweight reasoning model in the Gemini 2.5 family, optimized for ultra-low latency and cost efficiency. It offers improved throughput, faster token generation, and better performance...

1M context+3
View
OpenRouter logo

Google: Gemini 2.5 Flash Lite (batch)

OpenRouter · Jul 22, 2025

Gemini 2.5 Flash-Lite is a lightweight reasoning model in the Gemini 2.5 family, optimized for ultra-low latency and cost efficiency. It offers improved throughput, faster token generation, and better performance...

1M context+3
View
OpenRouter logo

Qwen: Qwen3 235B A22B Instruct 2507

OpenRouter · Jul 21, 2025

Qwen3-235B-A22B-Instruct-2507 is a multilingual, instruction-tuned mixture-of-experts language model based on the Qwen3-235B architecture, with 22B active parameters per forward pass. It is optimized for general-purpose text generation, including instruction following,...

262.1K context
View
OpenRouter logo

Switchpoint Router

OpenRouter · Jul 11, 2025

Switchpoint AI's router instantly analyzes your request and directs it to the optimal AI from an ever-evolving library. As the world of LLMs advances, our router gets smarter, ensuring you...

131.1K context
View
OpenRouter logo

MoonshotAI: Kimi K2 0711 (free)

OpenRouter · Jul 11, 2025

Kimi K2 Instruct is a large-scale Mixture-of-Experts (MoE) language model developed by Moonshot AI, featuring 1 trillion total parameters with 32 billion active per forward pass. It is optimized for agentic capabilities, including advanced tool use, reasoning, and code synthesis. Kimi K2 excels across a broad range of benchmarks, particularly in coding (LiveCodeBench, SWE-bench), reasoning (ZebraLogic, GPQA), and tool-use (Tau2, AceBench) tasks. It supports long-context inference up to 128K tokens and is designed with a novel training stack that includes the MuonClip optimizer for stable large-scale MoE training.

32.8K context
View
OpenRouter logo

MoonshotAI: Kimi K2 0711

OpenRouter · Jul 11, 2025

Kimi K2 Instruct is a large-scale Mixture-of-Experts (MoE) language model developed by Moonshot AI, featuring 1 trillion total parameters with 32 billion active per forward pass. It is optimized for...

131.1K context
View
OpenRouter logo

THUDM: GLM 4.1V 9B Thinking

OpenRouter · Jul 11, 2025

GLM-4.1V-9B-Thinking is a 9B parameter vision-language model developed by THUDM, based on the GLM-4-9B foundation. It introduces a reasoning-centric "thinking paradigm" enhanced with reinforcement learning to improve multimodal reasoning, long-context understanding (up to 64K tokens), and complex problem solving. It achieves state-of-the-art performance among models in its class, outperforming even larger models like Qwen-2.5-VL-72B on a majority of benchmark tasks.

65.5K context
View
OpenRouter logo

Mistral: Devstral Medium

OpenRouter · Jul 10, 2025

Devstral Medium is a high-performance code generation and agentic reasoning model developed jointly by Mistral AI and All Hands AI. Positioned as a step up from Devstral Small, it achieves...

131.1K context
View
OpenRouter logo

Mistral: Devstral Small 1.1

OpenRouter · Jul 10, 2025

Devstral Small 1.1 is a 24B parameter open-weight language model for software engineering agents, developed by Mistral AI in collaboration with All Hands AI. Finetuned from Mistral Small 3.1 and...

131.1K context
View
OpenRouter logo

Venice: Uncensored (free)

OpenRouter · Jul 9, 2025

Venice Uncensored Dolphin Mistral 24B Venice Edition is a fine-tuned variant of Mistral-Small-24B-Instruct-2501, developed by dphn.ai in collaboration with Venice.ai. This model is designed as an “uncensored” instruct-tuned LLM, preserving...

32.8K context
View
OpenRouter logo

Venice: Uncensored

OpenRouter · Jul 9, 2025

Venice Uncensored Dolphin Mistral 24B Venice Edition is a fine-tuned variant of Mistral-Small-24B-Instruct-2501, developed by dphn.ai in collaboration with Venice.ai. This model is designed as an “uncensored” instruct-tuned LLM, preserving...

128K context
View
OpenRouter logo

xAI: Grok 4

OpenRouter · Jul 9, 2025

Grok 4 is xAI's latest reasoning model with a 256k context window. It supports parallel tool calling, structured outputs, and both image and text inputs. Note that reasoning is not...

256K context+1
View
OpenRouter logo

Google: Gemma 3n 2B (free)

OpenRouter · Jul 9, 2025

Gemma 3n E2B IT is a multimodal, instruction-tuned model developed by Google DeepMind, designed to operate efficiently at an effective parameter size of 2B while leveraging a 6B architecture. Based...

8.2K context
View
OpenRouter logo

Tencent: Hunyuan A13B Instruct

OpenRouter · Jul 8, 2025

Hunyuan-A13B is a 13B active parameter Mixture-of-Experts (MoE) language model developed by Tencent, with a total parameter count of 80B and support for reasoning via Chain-of-Thought. It offers competitive benchmark...

131.1K context
View
OpenRouter logo

TNG: DeepSeek R1T2 Chimera (free)

OpenRouter · Jul 8, 2025

DeepSeek-TNG-R1T2-Chimera is the second-generation Chimera model from TNG Tech. It is a 671 B-parameter mixture-of-experts text-generation model assembled from DeepSeek-AI’s R1-0528, R1, and V3-0324 checkpoints with an Assembly-of-Experts merge. The tri-parent design yields strong reasoning performance while running roughly 20 % faster than the original R1 and more than 2× faster than R1-0528 under vLLM, giving a favorable cost-to-intelligence trade-off. The checkpoint supports contexts up to 60 k tokens in standard use (tested to ~130 k) and maintains consistent <think> token behaviour, making it suitable for long-context analysis, dialogue and other open-ended generation tasks.

163.8K context
View
OpenRouter logo

TNG: DeepSeek R1T2 Chimera

OpenRouter · Jul 8, 2025

DeepSeek-TNG-R1T2-Chimera is the second-generation Chimera model from TNG Tech. It is a 671 B-parameter mixture-of-experts text-generation model assembled from DeepSeek-AI’s R1-0528, R1, and V3-0324 checkpoints with an Assembly-of-Experts merge. The...

163.8K context
View
OpenRouter logo

Morph: Morph V3 Large

OpenRouter · Jul 7, 2025

Morph's high-accuracy apply model for complex code edits. ~4,500 tokens/sec with 98% accuracy for precise code transformations. The model requires the prompt to be in the following format: <instruction>{instruction}</instruction> <code>{initial_code}</code>...

262.1K context
View
OpenRouter logo

Morph: Morph V3 Fast

OpenRouter · Jul 7, 2025

Morph's fastest apply model for code edits. ~10,500 tokens/sec with 96% accuracy for rapid code transformations. The model requires the prompt to be in the following format: <instruction>{instruction}</instruction> <code>{initial_code}</code> <update>{edit_snippet}</update>...

81.9K context
View
OpenRouter logo

Baidu: ERNIE 4.5 VL 424B A47B

OpenRouter · Jun 30, 2025

ERNIE-4.5-VL-424B-A47B is a multimodal Mixture-of-Experts (MoE) model from Baidu’s ERNIE 4.5 series, featuring 424B total parameters with 47B active per token. It is trained jointly on text and image data...

123K context
View
OpenRouter logo

Baidu: ERNIE 4.5 300B A47B

OpenRouter · Jun 30, 2025

ERNIE-4.5-300B-A47B is a 300B parameter Mixture-of-Experts (MoE) language model developed by Baidu as part of the ERNIE 4.5 series. It activates 47B parameters per token and supports text generation in...

131.1K context
View
OpenRouter logo

Inception: Mercury

OpenRouter · Jun 26, 2025

Mercury is the first diffusion large language model (dLLM). Applying a breakthrough discrete diffusion approach, the model runs 5-10x faster than even speed optimized models like GPT-4.1 Nano and Claude...

128K context
View
OpenRouter logo

Mistral: Mistral Small 3.2 24B

OpenRouter · Jun 20, 2025

Mistral-Small-3.2-24B-Instruct-2506 is an updated 24B parameter model from Mistral optimized for instruction following, repetition reduction, and improved function calling. Compared to the 3.1 release, version 3.2 significantly improves accuracy on...

256K context
View
OpenRouter logo

MiniMax: MiniMax M1

OpenRouter · Jun 17, 2025

MiniMax-M1 is a large-scale, open-weight reasoning model designed for extended context and high-efficiency inference. It leverages a hybrid Mixture-of-Experts (MoE) architecture paired with a custom "lightning attention" mechanism, allowing it...

1M context
View
OpenRouter logo

Google: Gemini 2.5 Flash

OpenRouter · Jun 17, 2025

Gemini 2.5 Flash is Google's state-of-the-art workhorse model, specifically designed for advanced reasoning, coding, mathematics, and scientific tasks. It includes built-in "thinking" capabilities, enabling it to provide responses with greater...

1M context+3
View
OpenRouter logo

Google: Gemini 2.5 Flash (batch)

OpenRouter · Jun 17, 2025

Gemini 2.5 Flash is Google's state-of-the-art workhorse model, specifically designed for advanced reasoning, coding, mathematics, and scientific tasks. It includes built-in "thinking" capabilities, enabling it to provide responses with greater...

1M context+3
View