Guide
2,496 step-by-step guides to setting up AI models from OpenAI, Anthropic, Google, OpenRouter, and more in TypingMind with your own API key.
Search and filter guides
Showing 1,735-1,785 of 2,496 guides

OpenAI: gpt-oss-120b (free)
OpenRouter · Aug 5, 2025
gpt-oss-120b is an open-weight, 117B-parameter Mixture-of-Experts (MoE) language model from OpenAI designed for high-reasoning, agentic, and general-purpose production use cases. It activates 5.1B parameters per forward pass and is optimized...

OpenAI: gpt-oss-120b
OpenRouter · Aug 5, 2025
gpt-oss-120b is an open-weight, 117B-parameter Mixture-of-Experts (MoE) language model from OpenAI designed for high-reasoning, agentic, and general-purpose production use cases. It activates 5.1B parameters per forward pass and is optimized...

OpenAI: gpt-oss-120b (exacto)
OpenRouter · Aug 5, 2025
gpt-oss-120b is an open-weight, 117B-parameter Mixture-of-Experts (MoE) language model from OpenAI designed for high-reasoning, agentic, and general-purpose production use cases. It activates 5.1B parameters per forward pass and is optimized to run on a single H100 GPU with native MXFP4 quantization. The model supports configurable reasoning depth, full chain-of-thought access, and native tool use, including function calling, browsing, and structured output generation.

OpenAI: gpt-oss-120b (batch)
OpenRouter · Aug 5, 2025
gpt-oss-120b is an open-weight, 117B-parameter Mixture-of-Experts (MoE) language model from OpenAI designed for high-reasoning, agentic, and general-purpose production use cases. It activates 5.1B parameters per forward pass and is optimized...

OpenAI: gpt-oss-20b (free)
OpenRouter · Aug 5, 2025
gpt-oss-20b is an open-weight 21B parameter model released by OpenAI under the Apache 2.0 license. It uses a Mixture-of-Experts (MoE) architecture with 3.6B active parameters per forward pass, optimized for...

OpenAI: gpt-oss-20b
OpenRouter · Aug 5, 2025
gpt-oss-20b is an open-weight 21B parameter model released by OpenAI under the Apache 2.0 license. It uses a Mixture-of-Experts (MoE) architecture with 3.6B active parameters per forward pass, optimized for...

OpenAI: gpt-oss-20b (batch)
OpenRouter · Aug 5, 2025
gpt-oss-20b is an open-weight 21B parameter model released by OpenAI under the Apache 2.0 license. It uses a Mixture-of-Experts (MoE) architecture with 3.6B active parameters per forward pass, optimized for...

Anthropic: Claude Opus 4.1
OpenRouter · Aug 5, 2025
Claude Opus 4.1 is an updated version of Anthropic’s flagship model, offering improved performance in coding, reasoning, and agentic tasks. It achieves 74.5% on SWE-bench Verified and shows notable gains...

Anthropic: Claude Opus 4.1 (batch)
OpenRouter · Aug 5, 2025
Claude Opus 4.1 is an updated version of Anthropic’s flagship model, offering improved performance in coding, reasoning, and agentic tasks. It achieves 74.5% on SWE-bench Verified and shows notable gains...

GPT OSS 20B
Groq · Aug 5, 2025
Open-weight GPT model for self-hosted reasoning and instruction-following workloads

GPT OSS 120B
Groq · Aug 5, 2025
Open GPT reasoning model for self-hosted agents and controllable deployments

Claude Opus 4.1 (latest)
Anthropic · Aug 5, 2025
Flagship Claude model for deep reasoning, coding, and long-horizon agents

Claude Opus 4.1
Anthropic · Aug 5, 2025
Flagship Claude model for deep reasoning, coding, and long-horizon agents

GPT OSS 20B
Fireworks AI · Aug 5, 2025
Open-weight GPT model for self-hosted reasoning and instruction-following workloads

GPT OSS 120B
Fireworks AI · Aug 5, 2025
Open-weight GPT model for self-hosted reasoning and instruction-following workloads

GPT OSS 120B
Synthetic · Aug 5, 2025
Open-weight GPT model for self-hosted reasoning and instruction-following workloads

gpt-oss-120b
AWS Bedrock · Aug 5, 2025
Open GPT reasoning model for self-hosted agents and controllable deployments

Claude Opus 4.1
AWS Bedrock · Aug 5, 2025
Flagship Claude model for deep reasoning, coding, and long-horizon agents

gpt-oss-20b
AWS Bedrock · Aug 5, 2025
Open GPT reasoning model for self-hosted agents and controllable deployments

Claude Opus 4.1 (US)
AWS Bedrock · Aug 5, 2025
Flagship Claude model for deep reasoning, coding, and long-horizon agents

gpt-oss-120b
AWS Bedrock · Aug 5, 2025
Open GPT reasoning model for self-hosted agents and controllable deployments

gpt-oss-20b
AWS Bedrock · Aug 5, 2025
Open GPT reasoning model for self-hosted agents and controllable deployments

gpt-oss-120b (GovCloud)
AWS Bedrock · Aug 5, 2025
Open GPT reasoning model for self-hosted agents and controllable deployments

gpt-oss-20b (GovCloud)
AWS Bedrock · Aug 5, 2025
Open GPT reasoning model for self-hosted agents and controllable deployments

GPT OSS 120B
DeepInfra · Aug 5, 2025
Open-weight GPT model for self-hosted reasoning and instruction-following workloads

GPT OSS 20B
DeepInfra · Aug 5, 2025
Open-weight GPT model for self-hosted reasoning and instruction-following workloads
GPT OSS 120B
Hugging Face · Aug 5, 2025
Open GPT reasoning model for self-hosted agents and controllable deployments
GPT OSS 20B
Hugging Face · Aug 5, 2025
Open-weight GPT model for self-hosted reasoning and instruction-following workloads

GPT OSS 20B
Together AI · Aug 5, 2025
Open-weight GPT model for self-hosted reasoning and instruction-following workloads

GPT OSS 120B
Together AI · Aug 5, 2025
Open-weight GPT model for self-hosted reasoning and instruction-following workloads

GPT OSS 20B
Vercel AI Gateway · Aug 5, 2025
Open-weight GPT model for self-hosted reasoning and instruction-following workloads

GPT OSS 120B
Vercel AI Gateway · Aug 5, 2025
Open GPT reasoning model for self-hosted agents and controllable deployments

gpt-oss:20b
Ollama Cloud · Aug 5, 2025
Open-weight GPT model for self-hosted reasoning and instruction-following workloads

gpt-oss:120b
Ollama Cloud · Aug 5, 2025
Open-weight GPT model for self-hosted reasoning and instruction-following workloads

GPT OSS 120B
Cerebras · Aug 5, 2025
Open GPT reasoning model for self-hosted agents and controllable deployments

Mistral: Codestral 2508
OpenRouter · Aug 1, 2025
Mistral's cutting-edge language model for coding released end of July 2025. Codestral specializes in low-latency, high-frequency tasks such as fill-in-the-middle (FIM), code correction and test generation. [Blog Post](https://mistral.ai/news/codestral-25-08)

Mistral: Codestral 2508 (batch)
OpenRouter · Aug 1, 2025
Mistral's cutting-edge language model for coding released end of July 2025. Codestral specializes in low-latency, high-frequency tasks such as fill-in-the-middle (FIM), code correction and test generation. [Blog Post](https://mistral.ai/news/codestral-25-08)

GLM 4.5 Air
Fireworks AI · Aug 1, 2025
GLM 4.5 Air from Fireworks AI - text input, 131,072 token context

DeepSeek R1 (0528)
Synthetic · Aug 1, 2025
DeepSeek R1 (0528) from Synthetic - text input, 128,000 token context

DeepSeek V3 (0324)
Synthetic · Aug 1, 2025
DeepSeek V3 (0324) from Synthetic - text input, 128,000 token context

Qwen: Qwen3 Coder 30B A3B Instruct
OpenRouter · Jul 31, 2025
Qwen3-Coder-30B-A3B-Instruct is a 30.5B parameter Mixture-of-Experts (MoE) model with 128 experts (8 active per forward pass), designed for advanced code generation, repository-scale understanding, and agentic tool use. Built on the...

Qwen3-Coder 30B-A3B Instruct
AWS Bedrock · Jul 31, 2025
Smaller Qwen coder for efficient local agents and repo-level fixes

Qwen 3 Coder 30B A3B Instruct
Vercel AI Gateway · Jul 31, 2025
Qwen coding model for software agents, repository edits, and code reasoning

Veo 3.0 Fast Generate
Vercel AI Gateway · Jul 31, 2025
Video model for prompt-guided generation, editing, and motion workflows

Command A Vision
Cohere · Jul 31, 2025
Cohere vision model for multilingual document analysis, OCR, and image understanding

Qwen: Qwen3 30B A3B Instruct 2507
OpenRouter · Jul 29, 2025
Qwen3-30B-A3B-Instruct-2507 is a 30.5B-parameter mixture-of-experts language model from Qwen, with 3.3B active parameters per inference. It operates in non-thinking mode and is designed for high-quality instruction following, multilingual understanding, and...

GLM 4.5
Fireworks AI · Jul 29, 2025
GLM 4.5 from Fireworks AI - text input, 131,072 token context

GLM 4.5
Synthetic · Jul 28, 2025
GLM 4.5 from Synthetic - text input, 128,000 token context

GLM-4.5
DeepInfra · Jul 28, 2025
GLM-4.5 from DeepInfra - text input, 131,072 token context
GLM-4.5-Air
Hugging Face · Jul 28, 2025
Lighter GLM-4.5 variant for fast coding assistance and cheaper agents
GLM-4.5
Hugging Face · Jul 28, 2025
Hybrid-reasoning GLM release that made the 4.5 line broadly useful

