Guide

2,496 step-by-step guides to setting up AI models from OpenAI, Anthropic, Google, OpenRouter, and more in TypingMind with your own API key.

Search and filter guides

Showing 409-459 of 713 guides

OpenRouter logo

Z.ai: GLM 4.6

OpenRouter · Sep 30, 2025

Compared with GLM-4.5, this generation brings several key improvements: Longer context window: The context window has been expanded from 128K to 200K tokens, enabling the model to handle more complex...

204.8K context
View
OpenRouter logo

Z.ai: GLM 4.6 (exacto)

OpenRouter · Sep 30, 2025

Compared with GLM-4.5, this generation brings several key improvements: Longer context window: The context window has been expanded from 128K to 200K tokens, enabling the model to handle more complex agentic tasks. Superior coding performance: The model achieves higher scores on code benchmarks and demonstrates better real-world performance in applications such as Claude Code、Cline、Roo Code and Kilo Code, including improvements in generating visually polished front-end pages. Advanced reasoning: GLM-4.6 shows a clear improvement in reasoning performance and supports tool use during inference, leading to stronger overall capability. More capable agents: GLM-4.6 exhibits stronger performance in tool using and search-based agents, and integrates more effectively within agent frameworks. Refined writing: Better aligns with human preferences in style and readability, and performs more naturally in role-playing scenarios.

204.8K context
View
OpenRouter logo

Anthropic: Claude Sonnet 4.5

OpenRouter · Sep 29, 2025

Claude Sonnet 4.5 is Anthropic’s most advanced Sonnet model to date, optimized for real-world agents and coding workflows. It delivers state-of-the-art performance on coding benchmarks such as SWE-bench Verified, with...

1M context+1
View
OpenRouter logo

Anthropic: Claude Sonnet 4.5 (batch)

OpenRouter · Sep 29, 2025

Claude Sonnet 4.5 is Anthropic’s most advanced Sonnet model to date, optimized for real-world agents and coding workflows. It delivers state-of-the-art performance on coding benchmarks such as SWE-bench Verified, with...

1M context+1
View
OpenRouter logo

DeepSeek: DeepSeek V3.2 Exp

OpenRouter · Sep 29, 2025

DeepSeek-V3.2-Exp is an experimental large language model released by DeepSeek as an intermediate step between V3.1 and future architectures. It introduces DeepSeek Sparse Attention (DSA), a fine-grained sparse attention mechanism...

163.8K context
View
OpenRouter logo

TheDrummer: Cydonia 24B V4.1

OpenRouter · Sep 27, 2025

Uncensored and creative writing model based on Mistral Small 3.2 24B with good recall, prompt adherence, and intelligence.

131.1K context
View
OpenRouter logo

Relace: Relace Apply 3

OpenRouter · Sep 26, 2025

Relace Apply 3 is a specialized code-patching LLM that merges AI-suggested edits straight into your source files. It can apply updates from GPT-4o, Claude, and others into your files at...

256K context
View
OpenRouter logo

Google: Gemini 2.5 Flash Preview 09-2025

OpenRouter · Sep 25, 2025

Gemini 2.5 Flash Preview September 2025 Checkpoint is Google's state-of-the-art workhorse model, specifically designed for advanced reasoning, coding, mathematics, and scientific tasks. It includes built-in "thinking" capabilities, enabling it to provide responses with greater accuracy and nuanced context handling. Additionally, Gemini 2.5 Flash is configurable through the "max tokens for reasoning" parameter, as described in the documentation (https://openrouter.ai/docs/use-cases/reasoning-tokens#max-tokens-for-reasoning).

1M context+3
View
OpenRouter logo

Google: Gemini 2.5 Flash Lite Preview 09-2025

OpenRouter · Sep 25, 2025

Gemini 2.5 Flash-Lite is a lightweight reasoning model in the Gemini 2.5 family, optimized for ultra-low latency and cost efficiency. It offers improved throughput, faster token generation, and better performance...

1M context+3
View
OpenRouter logo

Qwen: Qwen3 VL 235B A22B Thinking

OpenRouter · Sep 23, 2025

Qwen3-VL-235B-A22B Thinking is a multimodal model that unifies strong text generation with visual understanding across images and video. The Thinking model is optimized for multimodal reasoning in STEM and math....

131.1K context
View
OpenRouter logo

Qwen: Qwen3 VL 235B A22B Instruct

OpenRouter · Sep 23, 2025

Qwen3-VL-235B-A22B Instruct is an open-weight multimodal model that unifies strong text generation with visual understanding across images and video. The Instruct model targets general vision-language use (VQA, document parsing, chart/table...

262.1K context
View
OpenRouter logo

Qwen: Qwen3 Max

OpenRouter · Sep 23, 2025

Qwen3-Max is an updated release built on the Qwen3 series, offering major improvements in reasoning, instruction following, multilingual support, and long-tail knowledge coverage compared to the January 2025 version. It...

262.1K context
View
OpenRouter logo

Qwen: Qwen3 Coder Plus

OpenRouter · Sep 23, 2025

Qwen3 Coder Plus is Alibaba's proprietary version of the Open Source Qwen3 Coder 480B A35B. It is a powerful coding agent model specializing in autonomous programming via tool calling and...

1M context
View
OpenRouter logo

OpenAI: GPT-5 Codex

OpenRouter · Sep 23, 2025

GPT-5-Codex is a specialized version of GPT-5 optimized for software engineering and coding workflows. It is designed for both interactive development sessions and long, independent execution of complex engineering tasks....

400K context
View
OpenRouter logo

OpenAI: GPT-5 Codex (batch)

OpenRouter · Sep 23, 2025

GPT-5-Codex is a specialized version of GPT-5 optimized for software engineering and coding workflows. It is designed for both interactive development sessions and long, independent execution of complex engineering tasks....

400K context
View
OpenRouter logo

DeepSeek: DeepSeek V3.1 Terminus (exacto)

OpenRouter · Sep 22, 2025

DeepSeek-V3.1 Terminus is an update to [DeepSeek V3.1](/deepseek/deepseek-chat-v3.1) that maintains the model's original capabilities while addressing issues reported by users, including language consistency and agent capabilities, further optimizing the model's performance in coding and search agents. It is a large hybrid reasoning model (671B parameters, 37B active) that supports both thinking and non-thinking modes. It extends the DeepSeek-V3 base with a two-phase long-context training process, reaching up to 128K tokens, and uses FP8 microscaling for efficient inference. Users can control the reasoning behaviour with the `reasoning` `enabled` boolean. [Learn more in our docs](https://openrouter.ai/docs/use-cases/reasoning-tokens#enable-reasoning-with-default-config) The model improves tool use, code generation, and reasoning efficiency, achieving performance comparable to DeepSeek-R1 on difficult benchmarks while responding more quickly. It supports structured tool calling, code agents, and search agents, making it suitable for research, coding, and agentic workflows.

163.8K context
View
OpenRouter logo

DeepSeek: DeepSeek V3.1 Terminus

OpenRouter · Sep 22, 2025

DeepSeek-V3.1 Terminus is an update to [DeepSeek V3.1](/deepseek/deepseek-chat-v3.1) that maintains the model's original capabilities while addressing issues reported by users, including language consistency and agent capabilities, further optimizing the model's...

163.8K context
View
OpenRouter logo

xAI: Grok 4 Fast

OpenRouter · Sep 19, 2025

Grok 4 Fast is xAI's latest multimodal model with SOTA cost-efficiency and a 2M token context window. It comes in two flavors: non-reasoning and reasoning. Read more about the model...

2M context+1
View
OpenRouter logo

Tongyi DeepResearch 30B A3B

OpenRouter · Sep 18, 2025

Tongyi DeepResearch is an agentic large language model developed by Tongyi Lab, with 30 billion total parameters activating only 3 billion per token. It's optimized for long-horizon, deep information-seeking tasks...

131.1K context
View
OpenRouter logo

Qwen: Qwen3 Coder Flash

OpenRouter · Sep 17, 2025

Qwen3 Coder Flash is Alibaba's fast and cost efficient version of their proprietary Qwen3 Coder Plus. It is a powerful coding agent model specializing in autonomous programming via tool calling...

1M context
View
OpenRouter logo

OpenGVLab: InternVL3 78B

OpenRouter · Sep 15, 2025

The InternVL3 series is an advanced multimodal large language model (MLLM). Compared to InternVL 2.5, InternVL3 demonstrates stronger multimodal perception and reasoning capabilities. In addition, InternVL3 is benchmarked against the Qwen2.5 Chat models, whose pre-trained base models serve as the initialization for its language component. Benefiting from Native Multimodal Pre-Training, the InternVL3 series surpasses the Qwen2.5 series in overall text performance.

32.8K context
View
OpenRouter logo

Qwen: Qwen3 Next 80B A3B Thinking

OpenRouter · Sep 11, 2025

Qwen3-Next-80B-A3B-Thinking is a reasoning-first chat model in the Qwen3-Next line that outputs structured “thinking” traces by default. It’s designed for hard multi-step problems; math proofs, code synthesis/debugging, logic, and agentic...

262.1K context
View
OpenRouter logo

Qwen: Qwen3 Next 80B A3B Instruct

OpenRouter · Sep 11, 2025

Qwen3-Next-80B-A3B-Instruct is an instruction-tuned chat model in the Qwen3-Next series optimized for fast, stable responses without “thinking” traces. It targets complex tasks across reasoning, code generation, knowledge QA, and multilingual...

262.1K context
View
OpenRouter logo

Qwen: Qwen3 Next 80B A3B Instruct (free)

OpenRouter · Sep 11, 2025

Qwen3-Next-80B-A3B-Instruct is an instruction-tuned chat model in the Qwen3-Next series optimized for fast, stable responses without “thinking” traces. It targets complex tasks across reasoning, code generation, knowledge QA, and multilingual...

262.1K context
View
OpenRouter logo

Meituan: LongCat Flash Chat

OpenRouter · Sep 9, 2025

LongCat-Flash-Chat is a large-scale Mixture-of-Experts (MoE) model with 560B total parameters, of which 18.6B–31.3B (≈27B on average) are dynamically activated per input. It introduces a shortcut-connected MoE design to reduce...

131.1K context
View
OpenRouter logo

Qwen: Qwen Plus 0728

OpenRouter · Sep 8, 2025

Qwen Plus 0728, based on the Qwen3 foundation model, is a 1 million context hybrid reasoning model with a balanced performance, speed, and cost combination.

1M context
View
OpenRouter logo

Qwen: Qwen Plus 0728 (thinking)

OpenRouter · Sep 8, 2025

Qwen Plus 0728, based on the Qwen3 foundation model, is a 1 million context hybrid reasoning model with a balanced performance, speed, and cost combination.

1M context
View
OpenRouter logo

NVIDIA: Nemotron Nano 9B V2 (free)

OpenRouter · Sep 5, 2025

NVIDIA-Nemotron-Nano-9B-v2 is a large language model (LLM) trained from scratch by NVIDIA, and designed as a unified model for both reasoning and non-reasoning tasks. It responds to user queries and...

128K context
View
OpenRouter logo

NVIDIA: Nemotron Nano 9B V2

OpenRouter · Sep 5, 2025

NVIDIA-Nemotron-Nano-9B-v2 is a large language model (LLM) trained from scratch by NVIDIA, and designed as a unified model for both reasoning and non-reasoning tasks. It responds to user queries and...

131.1K context
View
OpenRouter logo

MoonshotAI: Kimi K2 0905

OpenRouter · Sep 4, 2025

Kimi K2 0905 is the September update of [Kimi K2 0711](moonshotai/kimi-k2). It is a large-scale Mixture-of-Experts (MoE) language model developed by Moonshot AI, featuring 1 trillion total parameters with 32...

262.1K context
View
OpenRouter logo

MoonshotAI: Kimi K2 0905 (exacto)

OpenRouter · Sep 4, 2025

Kimi K2 0905 is the September update of [Kimi K2 0711](moonshotai/kimi-k2). It is a large-scale Mixture-of-Experts (MoE) language model developed by Moonshot AI, featuring 1 trillion total parameters with 32 billion active per forward pass. It supports long-context inference up to 256k tokens, extended from the previous 128k. This update improves agentic coding with higher accuracy and better generalization across scaffolds, and enhances frontend coding with more aesthetic and functional outputs for web, 3D, and related tasks. Kimi K2 is optimized for agentic capabilities, including advanced tool use, reasoning, and code synthesis. It excels across coding (LiveCodeBench, SWE-bench), reasoning (ZebraLogic, GPQA), and tool-use (Tau2, AceBench) benchmarks. The model is trained with a novel stack incorporating the MuonClip optimizer for stable large-scale MoE training.

262.1K context
View
OpenRouter logo

Deep Cogito: Cogito V2 Preview Llama 70B

OpenRouter · Sep 2, 2025

Cogito v2 70B is a dense hybrid reasoning model that combines direct answering capabilities with advanced self-reflection. Built with iterative policy improvement, it delivers strong performance across reasoning tasks while maintaining efficiency through shorter reasoning chains and improved intuition.

32.8K context
View
OpenRouter logo

Cogito V2 Preview Llama 109B

OpenRouter · Sep 2, 2025

An instruction-tuned, hybrid-reasoning Mixture-of-Experts model built on Llama-4-Scout-17B-16E. Cogito v2 can answer directly or engage an extended “thinking” phase, with alignment guided by Iterated Distillation & Amplification (IDA). It targets coding, STEM, instruction following, and general helpfulness, with stronger multilingual, tool-calling, and reasoning performance than size-equivalent baselines. The model supports long-context use (up to 10M tokens) and standard Transformers workflows. Users can control the reasoning behaviour with the `reasoning` `enabled` boolean. [Learn more in our docs](https://openrouter.ai/docs/use-cases/reasoning-tokens#enable-reasoning-with-default-config)

32.8K context
View
OpenRouter logo

StepFun: Step3

OpenRouter · Aug 28, 2025

Step3 is a cutting-edge multimodal reasoning model—built on a Mixture-of-Experts architecture with 321B total parameters and 38B active. It is designed end-to-end to minimize decoding costs while delivering top-tier performance in vision–language reasoning. Through the co-design of Multi-Matrix Factorization Attention (MFA) and Attention-FFN Disaggregation (AFD), Step3 maintains exceptional efficiency across both flagship and low-end accelerators.

65.5K context
View
OpenRouter logo

Qwen: Qwen3 30B A3B Thinking 2507

OpenRouter · Aug 28, 2025

Qwen3-30B-A3B-Thinking-2507 is a 30B parameter Mixture-of-Experts reasoning model optimized for complex tasks requiring extended multi-step thinking. The model is designed specifically for “thinking mode,” where internal reasoning traces are separated...

81.9K context
View
OpenRouter logo

xAI: Grok Code Fast 1

OpenRouter · Aug 26, 2025

Grok Code Fast 1 is a speedy and economical reasoning model that excels at agentic coding. With reasoning traces visible in the response, developers can steer Grok Code for high-quality...

256K context
View
OpenRouter logo

Nous: Hermes 4 70B

OpenRouter · Aug 26, 2025

Hermes 4 70B is a hybrid reasoning model from Nous Research, built on Meta-Llama-3.1-70B. It introduces the same hybrid mode as the larger 405B release, allowing the model to either...

131.1K context
View
OpenRouter logo

Nous: Hermes 4 405B

OpenRouter · Aug 26, 2025

Hermes 4 is a large-scale reasoning model built on Meta-Llama-3.1-405B and released by Nous Research. It introduces a hybrid reasoning mode, where the model can choose to deliberate internally with...

131.1K context
View
OpenRouter logo

Google: Gemini 2.5 Flash Image Preview (Nano Banana)

OpenRouter · Aug 26, 2025

Gemini 2.5 Flash Image Preview, a.k.a. "Nano Banana," is a state of the art image generation model with contextual understanding. It is capable of image generation, edits, and multi-turn conversations.

32.8K context
View
OpenRouter logo

DeepSeek: DeepSeek V3.1

OpenRouter · Aug 21, 2025

DeepSeek-V3.1 is a large hybrid reasoning model (671B parameters, 37B active) that supports both thinking and non-thinking modes via prompt templates. It extends the DeepSeek-V3 base with a two-phase long-context...

163.8K context
View
OpenRouter logo

OpenAI: GPT-4o Audio

OpenRouter · Aug 15, 2025

The gpt-4o-audio-preview model adds support for audio inputs as prompts. This enhancement allows the model to detect nuances within audio recordings and add depth to generated user experiences. Audio outputs...

128K context
View
OpenRouter logo

Mistral: Mistral Medium 3.1

OpenRouter · Aug 13, 2025

Mistral Medium 3.1 is an updated version of Mistral Medium 3, which is a high-performance enterprise-grade language model designed to deliver frontier-level capabilities at significantly reduced operational cost. It balances...

131.1K context
View
OpenRouter logo

Mistral: Mistral Medium 3.1 (batch)

OpenRouter · Aug 13, 2025

Mistral Medium 3.1 is an updated version of Mistral Medium 3, which is a high-performance enterprise-grade language model designed to deliver frontier-level capabilities at significantly reduced operational cost. It balances...

131.1K context
View
OpenRouter logo

Baidu: ERNIE 4.5 21B A3B

OpenRouter · Aug 12, 2025

A sophisticated text-based Mixture-of-Experts (MoE) model featuring 21B total parameters with 3B activated per token, delivering exceptional multimodal understanding and generation through heterogeneous MoE structures and modality-isolated routing. Supporting an...

131.1K context
View
OpenRouter logo

Baidu: ERNIE 4.5 VL 28B A3B

OpenRouter · Aug 12, 2025

A powerful multimodal Mixture-of-Experts chat model featuring 28B total parameters with 3B activated per token, delivering exceptional text and vision understanding through its innovative heterogeneous MoE structure with modality-isolated routing....

131.1K context
View
OpenRouter logo

Z.ai: GLM 4.5V

OpenRouter · Aug 11, 2025

GLM-4.5V is a vision-language foundation model for multimodal agent applications. Built on a Mixture-of-Experts (MoE) architecture with 106B parameters and 12B activated parameters, it achieves state-of-the-art results in video understanding,...

65.5K context
View
OpenRouter logo

AI21: Jamba Mini 1.7

OpenRouter · Aug 8, 2025

Jamba Mini 1.7 is a compact and efficient member of the Jamba open model family, incorporating key improvements in grounding and instruction-following while maintaining the benefits of the SSM-Transformer hybrid architecture and 256K context window. Despite its compact size, it delivers accurate, contextually grounded responses and improved steerability.

256K context
View
OpenRouter logo

AI21: Jamba Large 1.7

OpenRouter · Aug 8, 2025

Jamba Large 1.7 is the latest model in the Jamba open family, offering improvements in grounding, instruction-following, and overall efficiency. Built on a hybrid SSM-Transformer architecture with a 256K context...

256K context
View
OpenRouter logo

OpenAI: GPT-5 Chat

OpenRouter · Aug 7, 2025

GPT-5 Chat is designed for advanced, natural, multimodal, and context-aware conversations for enterprise applications.

128K context
View
OpenRouter logo

OpenAI: GPT-5

OpenRouter · Aug 7, 2025

GPT-5 is OpenAI’s most advanced model, offering major improvements in reasoning, code quality, and user experience. It is optimized for complex tasks that require step-by-step reasoning, instruction following, and accuracy...

400K context+1
View
OpenRouter logo

OpenAI: GPT-5 (batch)

OpenRouter · Aug 7, 2025

GPT-5 is OpenAI’s most advanced model, offering major improvements in reasoning, code quality, and user experience. It is optimized for complex tasks that require step-by-step reasoning, instruction following, and accuracy...

400K context+1
View