Guide

2,496 step-by-step guides to setting up AI models from OpenAI, Anthropic, Google, OpenRouter, and more in TypingMind with your own API key.

Search and filter guides

Showing 1,633-1,683 of 2,496 guides

Alibaba Cloud logo

Qwen3-Omni Flash Realtime

Alibaba Cloud · Sep 15, 2025

Qwen omni model for text, vision, audio, and multimodal agent tasks

65.5K context
View
OpenRouter logo

Qwen: Qwen3 Next 80B A3B Thinking

OpenRouter · Sep 11, 2025

Qwen3-Next-80B-A3B-Thinking is a reasoning-first chat model in the Qwen3-Next line that outputs structured “thinking” traces by default. It’s designed for hard multi-step problems; math proofs, code synthesis/debugging, logic, and agentic...

262.1K context
View
OpenRouter logo

Qwen: Qwen3 Next 80B A3B Instruct

OpenRouter · Sep 11, 2025

Qwen3-Next-80B-A3B-Instruct is an instruction-tuned chat model in the Qwen3-Next series optimized for fast, stable responses without “thinking” traces. It targets complex tasks across reasoning, code generation, knowledge QA, and multilingual...

262.1K context
View
OpenRouter logo

Qwen: Qwen3 Next 80B A3B Instruct (free)

OpenRouter · Sep 11, 2025

Qwen3-Next-80B-A3B-Instruct is an instruction-tuned chat model in the Qwen3-Next series optimized for fast, stable responses without “thinking” traces. It targets complex tasks across reasoning, code generation, knowledge QA, and multilingual...

262.1K context
View
AWS Bedrock logo

Qwen3-Next 80B-A3B Instruct

AWS Bedrock · Sep 11, 2025

Qwen instruction model for multilingual chat, reasoning, and tool use

262.1K context
View
Hugging Face logo

Qwen3-Next-80B-A3B-Instruct

Hugging Face · Sep 11, 2025

Qwen instruction model for multilingual chat, reasoning, and tool use

262.1K context
View
Hugging Face logo

Qwen3-Next-80B-A3B-Thinking

Hugging Face · Sep 11, 2025

Qwen reasoning model for deliberate problem solving, math, and coding

262.1K context
View
Novita logo

Qwen3 Next 80B A3B Thinking

Novita · Sep 10, 2025

Qwen reasoning model for deliberate problem solving, math, and coding

131.1K context
View
Novita logo

Qwen3 Next 80B A3B Instruct

Novita · Sep 10, 2025

Qwen instruction model for multilingual chat, reasoning, and tool use

131.1K context
View
OpenRouter logo

Meituan: LongCat Flash Chat

OpenRouter · Sep 9, 2025

LongCat-Flash-Chat is a large-scale Mixture-of-Experts (MoE) model with 560B total parameters, of which 18.6B–31.3B (≈27B on average) are dynamically activated per input. It introduces a shortcut-connected MoE design to reduce...

131.1K context
View
Vercel AI Gateway logo

Seedream 4.0

Vercel AI Gateway · Sep 9, 2025

Image model for prompt-driven generation, editing, and visual design workflows

View
OpenRouter logo

Qwen: Qwen Plus 0728

OpenRouter · Sep 8, 2025

Qwen Plus 0728, based on the Qwen3 foundation model, is a 1 million context hybrid reasoning model with a balanced performance, speed, and cost combination.

1M context
View
OpenRouter logo

Qwen: Qwen Plus 0728 (thinking)

OpenRouter · Sep 8, 2025

Qwen Plus 0728, based on the Qwen3 foundation model, is a 1 million context hybrid reasoning model with a balanced performance, speed, and cost combination.

1M context
View
Alibaba Cloud logo

Qwen3-ASR Flash

Alibaba Cloud · Sep 8, 2025

Speech transcription model for accurate audio-to-text and captioning workflows

53.2K context
View
OpenRouter logo

NVIDIA: Nemotron Nano 9B V2 (free)

OpenRouter · Sep 5, 2025

NVIDIA-Nemotron-Nano-9B-v2 is a large language model (LLM) trained from scratch by NVIDIA, and designed as a unified model for both reasoning and non-reasoning tasks. It responds to user queries and...

128K context
View
OpenRouter logo

NVIDIA: Nemotron Nano 9B V2

OpenRouter · Sep 5, 2025

NVIDIA-Nemotron-Nano-9B-v2 is a large language model (LLM) trained from scratch by NVIDIA, and designed as a unified model for both reasoning and non-reasoning tasks. It responds to user queries and...

131.1K context
View
Groq logo

Kimi K2 Instruct 0905

Groq · Sep 5, 2025

Kimi K2 Instruct 0905 from Groq - text input, 262,144 token context

262.1K context
View
Moonshot AI logo

Kimi K2 Turbo

Moonshot AI · Sep 5, 2025

Fast Kimi model for responsive chat, coding help, and agent loops

262.1K context
View
Moonshot AI logo

Kimi K2 0905

Moonshot AI · Sep 5, 2025

Kimi model for long-context chat, coding, and agentic reasoning

262.1K context
View
Synthetic logo

Kimi K2 0905

Synthetic · Sep 5, 2025

Kimi K2 0905 from Synthetic - text input, 262,144 token context

262.1K context
View
DeepInfra logo

Kimi K2 0905

DeepInfra · Sep 5, 2025

Kimi K2 0905 from DeepInfra - text input, 262,144 token context

262.1K context
View
Vercel AI Gateway logo

Qwen3 Max Preview

Vercel AI Gateway · Sep 5, 2025

Flagship Qwen model for complex reasoning, coding, and agentic workflows

262.1K context
View
Novita logo

Kimi K2 0905

Novita · Sep 5, 2025

Kimi model for long-context chat, coding, and agentic reasoning

262.1K context
View
OpenRouter logo

MoonshotAI: Kimi K2 0905

OpenRouter · Sep 4, 2025

Kimi K2 0905 is the September update of [Kimi K2 0711](moonshotai/kimi-k2). It is a large-scale Mixture-of-Experts (MoE) language model developed by Moonshot AI, featuring 1 trillion total parameters with 32...

262.1K context
View
OpenRouter logo

MoonshotAI: Kimi K2 0905 (exacto)

OpenRouter · Sep 4, 2025

Kimi K2 0905 is the September update of [Kimi K2 0711](moonshotai/kimi-k2). It is a large-scale Mixture-of-Experts (MoE) language model developed by Moonshot AI, featuring 1 trillion total parameters with 32 billion active per forward pass. It supports long-context inference up to 256k tokens, extended from the previous 128k. This update improves agentic coding with higher accuracy and better generalization across scaffolds, and enhances frontend coding with more aesthetic and functional outputs for web, 3D, and related tasks. Kimi K2 is optimized for agentic capabilities, including advanced tool use, reasoning, and code synthesis. It excels across coding (LiveCodeBench, SWE-bench), reasoning (ZebraLogic, GPQA), and tool-use (Tau2, AceBench) benchmarks. The model is trained with a novel stack incorporating the MuonClip optimizer for stable large-scale MoE training.

262.1K context
View
Groq logo

Compound Mini

Groq · Sep 4, 2025

Efficient model for low-latency assistance, extraction, and routine automation

131.1K context
View
Groq logo

Compound

Groq · Sep 4, 2025

General-purpose chat model for instruction following, writing, and analysis

131.1K context
View
Hugging Face logo

Kimi-K2-Instruct-0905

Hugging Face · Sep 4, 2025

Kimi model for long-context chat, coding, and agentic reasoning

262.1K context
View
Novita logo

Qwen MT Plus

Novita · Sep 3, 2025

Translation model for multilingual conversion, localization, and cross-language workflows

16.4K context
View
OpenRouter logo

Deep Cogito: Cogito V2 Preview Llama 70B

OpenRouter · Sep 2, 2025

Cogito v2 70B is a dense hybrid reasoning model that combines direct answering capabilities with advanced self-reflection. Built with iterative policy improvement, it delivers strong performance across reasoning tasks while maintaining efficiency through shorter reasoning chains and improved intuition.

32.8K context
View
OpenRouter logo

Cogito V2 Preview Llama 109B

OpenRouter · Sep 2, 2025

An instruction-tuned, hybrid-reasoning Mixture-of-Experts model built on Llama-4-Scout-17B-16E. Cogito v2 can answer directly or engage an extended “thinking” phase, with alignment guided by Iterated Distillation & Amplification (IDA). It targets coding, STEM, instruction following, and general helpfulness, with stronger multilingual, tool-calling, and reasoning performance than size-equivalent baselines. The model supports long-context use (up to 10M tokens) and standard Transformers workflows. Users can control the reasoning behaviour with the `reasoning` `enabled` boolean. [Learn more in our docs](https://openrouter.ai/docs/use-cases/reasoning-tokens#enable-reasoning-with-default-config)

32.8K context
View
Setup guide logo

LucidQuery

Setup guide · Sep 1, 2025

LucidQuery is an advanced AI platform pushing toward general intelligence through hybrid multimodal capabilities that seamlessly integrate reasoning, web search, coding, and geolocation intelligence. Key features include multimodal reasoning across text, images, audio, video, code, and structured data for deep contextual understanding, precision geolocation intelligence with spatial analysis and map-based entity search, advanced coding AI for programming problem-solving and automation, real-time web search integration for up-to-date evidence-based responses, and explainable structured outputs combining logical reasoning with generative capabilities. The platform provides API access enabling developers to build sophisticated applications requiring multi-format data processing and real-world spatial context, designed for use cases spanning financial modeling, geographic analysis, research, and complex workflow automation.

API key setup
View
Google logo

Gemini Live 2.5 Flash

Google · Sep 1, 2025

Gemini Live 2.5 Flash from Google - text, image, audio, video input, 128,000 token context

128K context
View
LucidQuery logo

LucidQuery Nexus Coder

LucidQuery · Sep 1, 2025

Coding model for repository understanding, refactors, and agentic engineering tasks

250K context
View
DeepInfra logo

Qwen3-Next 80B-A3B Instruct

DeepInfra · Sep 1, 2025

Qwen instruction model for multilingual chat, reasoning, and tool use

262.1K context
View
Vercel AI Gateway logo

Qwen3 Next 80B A3B Thinking

Vercel AI Gateway · Sep 1, 2025

Efficient Qwen thinking model for local reasoning, math, and coding agents

262.1K context
View
Vercel AI Gateway logo

Qwen3 Next 80B A3B Instruct

Vercel AI Gateway · Sep 1, 2025

Qwen instruction model for multilingual chat, reasoning, and tool use

262.1K context
View
Vercel AI Gateway logo

Seed 1.8

Vercel AI Gateway · Sep 1, 2025

Multimodal reasoning model for visual analysis, planning, and tool use

256K context
View
Vercel AI Gateway logo

Seed 1.6

Vercel AI Gateway · Sep 1, 2025

Multimodal reasoning model for visual analysis, planning, and tool use

256K context
View
Alibaba Cloud logo

Qwen3-Next 80B-A3B (Thinking)

Alibaba Cloud · Sep 1, 2025

Efficient Qwen thinking model for local reasoning, math, and coding agents

131.1K context
View
Alibaba Cloud logo

Qwen3-Next 80B-A3B Instruct

Alibaba Cloud · Sep 1, 2025

Qwen instruction model for multilingual chat, reasoning, and tool use

131.1K context
View
OpenRouter logo

StepFun: Step3

OpenRouter · Aug 28, 2025

Step3 is a cutting-edge multimodal reasoning model—built on a Mixture-of-Experts architecture with 321B total parameters and 38B active. It is designed end-to-end to minimize decoding costs while delivering top-tier performance in vision–language reasoning. Through the co-design of Multi-Matrix Factorization Attention (MFA) and Attention-FFN Disaggregation (AFD), Step3 maintains exceptional efficiency across both flagship and low-end accelerators.

65.5K context
View
OpenRouter logo

Qwen: Qwen3 30B A3B Thinking 2507

OpenRouter · Aug 28, 2025

Qwen3-30B-A3B-Thinking-2507 is a 30B parameter Mixture-of-Experts reasoning model optimized for complex tasks requiring extended multi-step thinking. The model is designed specifically for “thinking mode,” where internal reasoning traces are separated...

81.9K context
View
xAI logo

Grok Code Fast 1

xAI · Aug 28, 2025

Grok Code Fast 1 from xAI - text input, 256,000 token context

256K context
View
Cohere logo

Command A Translate

Cohere · Aug 28, 2025

Translation model for multilingual conversion, localization, and cross-language workflows

8K context
View
OpenRouter logo

xAI: Grok Code Fast 1

OpenRouter · Aug 26, 2025

Grok Code Fast 1 is a speedy and economical reasoning model that excels at agentic coding. With reasoning traces visible in the response, developers can steer Grok Code for high-quality...

256K context
View
OpenRouter logo

Nous: Hermes 4 70B

OpenRouter · Aug 26, 2025

Hermes 4 70B is a hybrid reasoning model from Nous Research, built on Meta-Llama-3.1-70B. It introduces the same hybrid mode as the larger 405B release, allowing the model to either...

131.1K context
View
OpenRouter logo

Nous: Hermes 4 405B

OpenRouter · Aug 26, 2025

Hermes 4 is a large-scale reasoning model built on Meta-Llama-3.1-405B and released by Nous Research. It introduces a hybrid reasoning mode, where the model can choose to deliberate internally with...

131.1K context
View
OpenRouter logo

Google: Gemini 2.5 Flash Image Preview (Nano Banana)

OpenRouter · Aug 26, 2025

Gemini 2.5 Flash Image Preview, a.k.a. "Nano Banana," is a state of the art image generation model with contextual understanding. It is capable of image generation, edits, and multi-turn conversations.

32.8K context
View
Google logo

Nano Banana

Google · Aug 26, 2025

Nano Banana image model for fast generation, edits, and character-consistent assets

32.8K context
View
Google logo

Gemini 2.5 Flash Image (Preview)

Google · Aug 26, 2025

Gemini 2.5 Flash Image (Preview) from Google - text, image input, 32,768 token context

32.8K context
View