Guide

2,496 step-by-step guides to setting up AI models from OpenAI, Anthropic, Google, OpenRouter, and more in TypingMind with your own API key.

Search and filter guides

Showing 1-51 of 2,496 guides

OpenRouter logo

OpenAI: GPT-6.1 Sol Pro

OpenRouter · Sep 29, 2026

GPT-6.1 Sol Pro is the same underlying model as [GPT-6.1 Sol](https://openrouter.ai/openai/gpt-6.1-sol), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks. **Cost note:** pro mode spends far more...

1.1M context+1
View
OpenRouter logo

OpenAI: GPT-6.1 Sol Pro (batch)

OpenRouter · Sep 29, 2026

GPT-6.1 Sol Pro is the same underlying model as [GPT-6.1 Sol](https://openrouter.ai/openai/gpt-6.1-sol), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks. **Cost note:** pro mode spends far more...

1.1M context+1
View
OpenRouter logo

OpenAI: GPT-6.1 Sol

OpenRouter · Sep 29, 2026

GPT-6.1 Sol is an upgrade to GPT-6 Sol from OpenAI, positioned below the flagship GPT-6 Astra in the GPT-6 series. It is suited for agentic coding, computer use, document-heavy professional...

1.1M context+1
View
OpenRouter logo

OpenAI: GPT-6.1 Sol (batch)

OpenRouter · Sep 29, 2026

GPT-6.1 Sol is an upgrade to GPT-6 Sol from OpenAI, positioned below the flagship GPT-6 Astra in the GPT-6 series. It is suited for agentic coding, computer use, document-heavy professional...

1.1M context+1
View
Setup guide logo

Together AI

Setup guide · Sep 29, 2026

Together AI is an AI cloud for running open-source models. Its serverless inference API gives pay-per-token access to models like Llama, DeepSeek, Qwen, GLM, Kimi, MiniMax, and gpt-oss through an OpenAI-compatible endpoint. Beyond serverless inference, Together AI offers dedicated endpoints, fine-tuning, batch inference, and GPU clusters for teams that need more control over performance and cost.

API key setup
View
Setup guide logo

Cohere

Setup guide · Sep 29, 2026

Cohere is an enterprise AI company that builds the Command family of language models, designed for business use cases like retrieval-augmented generation (RAG), tool use, and multilingual chat. The lineup also includes Command A Vision for images, Command A Reasoning for multi-step problems, and the open-weight Aya models covering 23 languages. Cohere offers a free, rate-limited trial key for testing and pay-per-token production keys, with an OpenAI-compatible endpoint that works with any OpenAI-compatible client.

API key setup
View
OpenAI logo

GPT-6.1 Sol

OpenAI · Sep 29, 2026

Near-Astra performance for complex coding, computer use, and professional work at a lower cost.

1.1M context+1
View
Venice AI logo

GPT-6.1 Sol

Venice AI · Sep 29, 2026

GPT model for general reasoning, writing, coding, and tool-assisted tasks

1.1M context
View
Vercel AI Gateway logo

Liquid d1

Vercel AI Gateway · Sep 29, 2026

General-purpose chat model for instruction following, writing, and analysis

32K context
View
Vercel AI Gateway logo

Ling 3.1 Flash (Free)

Vercel AI Gateway · Sep 29, 2026

Efficient model for low-latency assistance, extraction, and routine automation

262.1K context
View
Vercel AI Gateway logo

Ling 3.1 Flash

Vercel AI Gateway · Sep 29, 2026

Efficient model for low-latency assistance, extraction, and routine automation

262.1K context
View
Vercel AI Gateway logo

GPT-6.1 Sol (Fast)

Vercel AI Gateway · Sep 29, 2026

Near-Astra performance for complex coding, computer use, and professional work at a lower cost.

1.1M context+1
View
Vercel AI Gateway logo

GPT-6.1 Sol

Vercel AI Gateway · Sep 29, 2026

Near-Astra performance for complex coding, computer use, and professional work at a lower cost.

1.1M context+1
View
OpenRouter logo

Anthropic: Claude Sonnet 5.5

OpenRouter · Sep 28, 2026

Claude Sonnet 5.5 is Anthropic's Sonnet-class model for well-scoped everyday work, succeeding Claude Sonnet 5 as a direct upgrade. It is especially strong at building features, fixing bugs, and producing...

1M context+1
View
OpenRouter logo

Anthropic: Claude Sonnet 5.5 (batch)

OpenRouter · Sep 28, 2026

Claude Sonnet 5.5 is Anthropic's Sonnet-class model for well-scoped everyday work, succeeding Claude Sonnet 5 as a direct upgrade. It is especially strong at building features, fixing bugs, and producing...

1M context+1
View
Anthropic logo

Claude Sonnet 5.5

Anthropic · Sep 28, 2026

Fast Claude model for everyday coding, agents, and knowledge work

1M context+1
View
Venice AI logo

MiMo-V2.6-Flash

Venice AI · Sep 28, 2026

MiMo flash model for fast multimodal assistance and agent workflows

1M context+2
View
AWS Bedrock logo

Claude Sonnet 5.5 (Global)

AWS Bedrock · Sep 28, 2026

Fast Claude model for everyday coding, agents, and knowledge work

1M context
View
Vercel AI Gateway logo

Claude Sonnet 5.5

Vercel AI Gateway · Sep 28, 2026

Fast Claude model for everyday coding, agents, and knowledge work

1M context
View
Setup guide logo

Z.AI

Setup guide · Sep 27, 2026

Z.AI is the international platform of Zhipu AI, the team behind the GLM model family. Its API gives pay-per-token access to GLM models for chat, reasoning, coding, and vision through an OpenAI-compatible endpoint, including free Flash models for lighter tasks.

API key setup
View
Setup guide logo

OpenCode Go

Setup guide · Sep 27, 2026

OpenCode Go is a model service from the OpenCode team that gives you one API key for 30+ curated coding models, including DeepSeek V4, Qwen3.8, GLM-5.3, Kimi K3, MiMo, MiniMax, and Grok, billed per token. Most models use an OpenAI-compatible Chat Completions endpoint, while the GPT, Grok, and Muse Spark models use the OpenAI Responses API.

API key setup
View
Setup guide logo

Cloudflare AI Gateway

Setup guide · Sep 27, 2026

Cloudflare AI Gateway sits between your apps and AI providers such as OpenAI, Anthropic, Google, xAI, and DeepSeek, adding caching, rate limiting, logging, and analytics across all of them. Its unified OpenAI-compatible endpoint lets one Cloudflare API token reach models from every provider. Store your provider keys in the gateway or use Cloudflare's Unified Billing to pay per token, then connect TypingMind with an authentication token created inside your gateway.

API key setup
View
Setup guide logo

AWS Bedrock

Setup guide · Sep 27, 2026

AWS Bedrock is a fully managed AWS service that gives you access to 100+ foundation models from leading AI providers, including Anthropic Claude, Meta Llama, Mistral, and more, through a single OpenAI-compatible API. Bedrock uses pay-as-you-go pricing with no monthly fees or minimum commitments, and requests run inside your own AWS account. Long-term API keys work with any OpenAI-compatible client, and you can monitor usage in the AWS console.

API key setup
View
Setup guide logo

Xiaomi

Setup guide · Sep 26, 2026

Xiaomi MiMo is Xiaomi's API platform for its MiMo language models. The lineup includes Pro models for complex reasoning and coding, Flash models for fast, low-cost tasks, UltraSpeed variants for faster output, and multimodal models that accept images, audio, and video as input. The MiMo API uses pay-as-you-go pricing and offers an OpenAI-compatible endpoint, so it works with any OpenAI-compatible client.

API key setup
View
Venice AI logo

Claude Opus 5.5 Fast

Venice AI · Sep 26, 2026

Flagship Claude model for deep reasoning, coding, and long-horizon agents

1M context
View
Venice AI logo

Claude Sonnet 5.5

Venice AI · Sep 26, 2026

Balanced Claude model for coding, analysis, agent workflows, and cost control

1M context
View
Vercel AI Gateway logo

LongCat 2.5 Preview

Vercel AI Gateway · Sep 26, 2026

Multimodal reasoning model for visual analysis, planning, and tool use

1M context
View
OpenRouter logo

TypeSafe: Jev Router

OpenRouter · Sep 25, 2026

Jev Router picks the best model and reasoning effort for each request, balancing quality, speed, and cost. It runs on [Jev](https://openrouter.ai/~typesafe/jev-latest), TypeSafe's first System One model, and adapts as your...

1M context
View
OpenRouter logo

Perceptron: Perceptron Mk1.5

OpenRouter · Sep 25, 2026

Perceptron Mk1.5 is Perceptron's embodied reasoning model for physical agents. It accepts text, image, video, and audio input, and answers with text plus optional structured annotations: points, boxes, polygons, tracks,...

36.9K context+1
View
Setup guide logo

Cerebras

Setup guide · Sep 25, 2026

Cerebras runs AI models on its wafer-scale chips, delivering some of the fastest inference available, often thousands of tokens per second. Its cloud offers open models such as gpt-oss 120B and Qwen through a serverless API. Cerebras uses pay-per-token pricing and an OpenAI-compatible Chat Completions endpoint, so it works with any OpenAI-compatible client.

API key setup
View
Vercel AI Gateway logo

Pixel Canary

Vercel AI Gateway · Sep 25, 2026

Multimodal reasoning model for visual analysis, planning, and tool use

262.1K context
View
OpenCode Go logo

LongCat 2.5 Preview Free

OpenCode Go · Sep 25, 2026

Meituan's multimodal reasoning model for coding and agent tasks, with image understanding and a 1M-token context window

1M context
View
OpenRouter logo

Fireworks: Ember-1

OpenRouter · Sep 24, 2026

Ember-1 is a specialized reasoning model from Fireworks Research, built on [Kimi K3](https://openrouter.ai/moonshotai/kimi-k3). It is designed to make every token go further: it produces shorter reasoning traces, using roughly 40%...

1M context
View
Setup guide logo

Alibaba Cloud

Setup guide · Sep 24, 2026

Alibaba Cloud Model Studio is Alibaba Cloud's platform for the Qwen model family, from the flagship Qwen Max models to fast, low-cost Qwen Flash models, plus vision, omni, and reasoning variants. It also hosts third-party models such as DeepSeek and GLM. Model Studio supports the OpenAI Responses API with pay-as-you-go pricing, and offers endpoints in several regions, including China (Hong Kong), Singapore, US (Virginia), and Germany (Frankfurt).

API key setup
View
Setup guide logo

MiniMax

Setup guide · Sep 24, 2026

MiniMax is an AI company that builds the MiniMax M-series language models, designed for coding, agentic workflows, and long-context reasoning at a low price per token. The lineup includes MiniMax-M3 and the M2 family, with highspeed variants for faster output. The MiniMax API uses pay-as-you-go pricing and offers an OpenAI-compatible endpoint, so it works with any OpenAI-compatible client.

API key setup
View
Setup guide logo

Novita

Setup guide · Sep 24, 2026

Novita AI is a cloud platform for running open-source AI models, with 100+ models such as DeepSeek, Kimi, Qwen, GLM, MiniMax, Llama, and gpt-oss available through a serverless API. It uses pay-per-token pricing and an OpenAI-compatible endpoint, so it works with any OpenAI-compatible client. Novita also offers GPU instances and agent sandboxes for teams that need more control.

API key setup
View
OpenRouter logo

Z.ai: GLM 5.3 Prime

OpenRouter · Sep 23, 2026

GLM-5.3-Prime is the high-speed variant of Z.ai's GLM-5.3, inheriting its full capabilities while delivering 1.5–2× the output throughput through inference acceleration. It supports text input and output with a 1M-token...

1M context
View
OpenRouter logo

Qwen: Qwen3.8 Max Prime

OpenRouter · Sep 23, 2026

Qwen3.8 Max Prime is a higher-throughput variant of Qwen3.8 Max from Alibaba's Qwen team, served as a separate SKU at a higher price point. It accepts text, image, and video...

1M context+1
View
OpenRouter logo

Space Bunny Alpha

OpenRouter · Sep 23, 2026

Space Bunny Alpha is an anonymous large model with blazing-fast inference, strong coding capabilities and native multimodal input support. It delivers adjustable reasoning effort, and a 1M-token context window. Space...

1M context
View
OpenRouter logo

AionLabs: Aion 3.5 Mini

OpenRouter · Sep 23, 2026

Aion 3.5 Mini is a multi-model roleplaying and storytelling system from AionLabs, built on the GLM family of models. It is the smaller, lower-cost sibling of Aion 3.5 and uses...

262.1K context
View
OpenRouter logo

AionLabs: Aion 3.5

OpenRouter · Sep 23, 2026

Aion 3.5 is a multi-model roleplaying and storytelling system from AionLabs, built on the GLM family of models. It uses a collaborative generation process in which multiple specialized models each...

262.1K context
View
OpenRouter logo

Upstage: Solar Mini 4

OpenRouter · Sep 23, 2026

Solar Mini 4 is Upstage's compact, cost-efficient language model, a 35B-parameter mixture-of-experts with 3B active parameters and a 524K context window. It is built for agentic use cases where response...

524.3K context
View
Setup guide logo

Ollama Cloud

Setup guide · Sep 23, 2026

Ollama Cloud runs large open models on Ollama's datacenter hardware, so you can use models too big for your own computer, like DeepSeek V4, Kimi K3, GLM-5.3, Qwen3.5 397B, and gpt-oss 120B, with the same Ollama account. Usage is billed per million tokens from your account's usage credits, and the API is OpenAI-compatible, so it works with any OpenAI-compatible client.

API key setup
View
Venice AI logo

Aion 3.5

Venice AI · Sep 23, 2026

Reasoning model for deliberate analysis, multi-step problem solving, and tool use

262.1K context
View
Venice AI logo

Aion 3.5 Mini

Venice AI · Sep 23, 2026

Efficient model for low-latency assistance, extraction, and routine automation

262.1K context
View
Vercel AI Gateway logo

Recraft V4.1 Flash

Vercel AI Gateway · Sep 23, 2026

Image model for prompt-driven generation, editing, and visual design workflows

View
Vercel AI Gateway logo

Qwen 3.8 Max Prime

Vercel AI Gateway · Sep 23, 2026

High-throughput edition of Qwen3.8 Max for coding, professional work, multimodal understanding, and long-running agent workflows

1M context
View
Vercel AI Gateway logo

Gemini 3.8 Flash-Lite TTS

Vercel AI Gateway · Sep 23, 2026

Speech generation model for controllable voice, narration, and audio delivery

View
Vercel AI Gateway logo

Gemini 3.8 Flash TTS

Vercel AI Gateway · Sep 23, 2026

Speech generation model for controllable voice, narration, and audio delivery

View
Vercel AI Gateway logo

Ember-1

Vercel AI Gateway · Sep 23, 2026

Multimodal reasoning model for visual analysis, planning, and tool use

1M context
View
OpenCode Go logo

Space Bunny Free

OpenCode Go · Sep 23, 2026

Anonymous preview reasoning model for coding, agentic tasks, tool use, and multimodal input

1M context
View