Guide
2,496 step-by-step guides to setting up AI models from OpenAI, Anthropic, Google, OpenRouter, and more in TypingMind with your own API key.
Search and filter guides
Showing 1,990-2,040 of 2,496 guides

OpenAI: o4 Mini (batch)
OpenRouter · Apr 16, 2025
OpenAI o4-mini is a compact reasoning model in the o-series, optimized for fast, cost-efficient performance while retaining strong multimodal and agentic capabilities. It supports tool use and demonstrates competitive reasoning...

o4-mini
OpenAI · Apr 16, 2025
Fast o-series model for compact reasoning, coding, and tool use

o3
OpenAI · Apr 16, 2025
Deliberate o-series reasoner for hard math, coding, and multi-step analysis

o3 (Fast)
Vercel AI Gateway · Apr 16, 2025
Deliberate o-series reasoner for hard math, coding, and multi-step analysis

o4-mini (Fast)
Vercel AI Gateway · Apr 16, 2025
Fast o-series model for compact reasoning, coding, and tool use

o3
Vercel AI Gateway · Apr 16, 2025
Deliberate o-series reasoner for hard math, coding, and multi-step analysis

o4-mini
Vercel AI Gateway · Apr 16, 2025
Fast o-series model for compact reasoning, coding, and tool use

Qwen2.5 7B Instruct
Novita · Apr 16, 2025
Qwen instruction model for multilingual chat, reasoning, and tool use

o4-mini
Cloudflare AI Gateway · Apr 16, 2025
Fast o-series model for compact reasoning, coding, and tool use

o3
Cloudflare AI Gateway · Apr 16, 2025
Deliberate o-series reasoner for hard math, coding, and multi-step analysis

Qwen: Qwen2.5 Coder 7B Instruct
OpenRouter · Apr 15, 2025
Qwen2.5-Coder-7B-Instruct is a 7B parameter instruction-tuned language model optimized for code-related tasks such as code generation, reasoning, and bug fixing. Based on the Qwen2.5 architecture, it incorporates enhancements like RoPE,...

Embed v4.0
Vercel AI Gateway · Apr 15, 2025
Embedding model for semantic search, retrieval, clustering, and ranking pipelines

OpenAI: GPT-4.1
OpenRouter · Apr 14, 2025
GPT-4.1 is a flagship large language model optimized for advanced instruction following, real-world software engineering, and long-context reasoning. It supports a 1 million token context window and outperforms GPT-4o and...

OpenAI: GPT-4.1 (batch)
OpenRouter · Apr 14, 2025
GPT-4.1 is a flagship large language model optimized for advanced instruction following, real-world software engineering, and long-context reasoning. It supports a 1 million token context window and outperforms GPT-4o and...

OpenAI: GPT-4.1 Mini
OpenRouter · Apr 14, 2025
GPT-4.1 Mini is a mid-sized model delivering performance competitive with GPT-4o at substantially lower latency and cost. It retains a 1 million token context window and scores 45.1% on hard...

OpenAI: GPT-4.1 Mini (batch)
OpenRouter · Apr 14, 2025
GPT-4.1 Mini is a mid-sized model delivering performance competitive with GPT-4o at substantially lower latency and cost. It retains a 1 million token context window and scores 45.1% on hard...

OpenAI: GPT-4.1 Nano
OpenRouter · Apr 14, 2025
For tasks that demand low latency, GPT‑4.1 nano is the fastest and cheapest model in the GPT-4.1 series. It delivers exceptional performance at a small size with its 1 million...

OpenAI: GPT-4.1 Nano (batch)
OpenRouter · Apr 14, 2025
For tasks that demand low latency, GPT‑4.1 nano is the fastest and cheapest model in the GPT-4.1 series. It delivers exceptional performance at a small size with its 1 million...

EleutherAI: Llemma 7b
OpenRouter · Apr 14, 2025
Llemma 7B is a language model for mathematics. It was initialized with Code Llama 7B weights, and trained on the Proof-Pile-2 for 200B tokens. Llemma models are particularly strong at...

AlfredPros: CodeLLaMa 7B Instruct Solidity
OpenRouter · Apr 14, 2025
A finetuned 7 billion parameters Code LLaMA - Instruct model to generate Solidity smart contract using 4-bit QLoRA finetuning provided by PEFT library.

GPT-4.1 nano
OpenAI · Apr 14, 2025
Tiny GPT-4.1 option for classification, routing, and very high-volume tasks

GPT-4.1 mini
OpenAI · Apr 14, 2025
Affordable GPT-4.1 lane for fast coding help and structured extraction

GPT-4.1
OpenAI · Apr 14, 2025
Long-lived GPT workhorse for coding, instruction following, and production apps

GPT-4.1 mini (Fast)
Vercel AI Gateway · Apr 14, 2025
Affordable GPT-4.1 lane for fast coding help and structured extraction

GPT-4.1 nano (Fast)
Vercel AI Gateway · Apr 14, 2025
Tiny GPT-4.1 option for classification, routing, and very high-volume tasks

GPT-4.1 (Fast)
Vercel AI Gateway · Apr 14, 2025
Long-lived GPT workhorse for coding, instruction following, and production apps

GPT-4.1 mini
Vercel AI Gateway · Apr 14, 2025
Affordable GPT-4.1 lane for fast coding help and structured extraction

GPT-4.1
Vercel AI Gateway · Apr 14, 2025
Long-lived GPT workhorse for coding, instruction following, and production apps

GPT-4.1 nano
Vercel AI Gateway · Apr 14, 2025
Tiny GPT-4.1 option for classification, routing, and very high-volume tasks

GPT-4.1 nano
Cloudflare AI Gateway · Apr 14, 2025
Tiny GPT-4.1 option for classification, routing, and very high-volume tasks

GPT-4.1
Cloudflare AI Gateway · Apr 14, 2025
Long-lived GPT workhorse for coding, instruction following, and production apps

GPT-4.1 mini
Cloudflare AI Gateway · Apr 14, 2025
Affordable GPT-4.1 lane for fast coding help and structured extraction

xAI: Grok 3 Mini Beta
OpenRouter · Apr 9, 2025
Grok 3 Mini is a lightweight, smaller thinking model. Unlike traditional models that generate answers immediately, Grok 3 Mini thinks before responding. It’s ideal for reasoning-heavy tasks that don’t demand...

xAI: Grok 3 Beta
OpenRouter · Apr 9, 2025
Grok 3 is the latest model from xAI. It's their flagship model that excels at enterprise use cases like data extraction, coding, and text summarization. Possesses deep domain knowledge in...

NVIDIA: Llama 3.1 Nemotron Ultra 253B v1
OpenRouter · Apr 8, 2025
Llama-3.1-Nemotron-Ultra-253B-v1 is a large language model (LLM) optimized for advanced reasoning, human-interactive chat, retrieval-augmented generation (RAG), and tool-calling tasks. Derived from Meta’s Llama-3.1-405B-Instruct, it has been significantly customized using Neural...

Pixtral Large (25.02)
AWS Bedrock · Apr 8, 2025
Mistral vision-language model for image understanding and multimodal chat

Pixtral Large (25.02) (EU)
AWS Bedrock · Apr 8, 2025
Mistral vision-language model for image understanding and multimodal chat

Pixtral Large (25.02) (US)
AWS Bedrock · Apr 8, 2025
Mistral vision-language model for image understanding and multimodal chat

Llama 3.3 70B
Venice AI · Apr 6, 2025
Open Llama instruction model for multilingual chat, reasoning, and coding

Llama 4 Maverick Instruct
Novita · Apr 6, 2025
Open multimodal Llama model for strong reasoning and fast responses

Llama 4 Scout Instruct
Novita · Apr 6, 2025
Open multimodal Llama model for long-context analysis and efficient agents

Meta: Llama 4 Maverick
OpenRouter · Apr 5, 2025
Llama 4 Maverick 17B Instruct (128E) is a high-capacity multimodal language model from Meta, built on a mixture-of-experts (MoE) architecture with 128 experts and 17 billion active parameters per forward...

Meta: Llama 4 Scout
OpenRouter · Apr 5, 2025
Llama 4 Scout 17B Instruct (16E) is a mixture-of-experts (MoE) language model developed by Meta, activating 17 billion parameters out of a total of 109B. It supports native multimodal input...

Llama 4 Scout 17B 16E
Groq · Apr 5, 2025
Open multimodal Llama model for long-context analysis and efficient agents

Llama 4 Maverick 17B
Groq · Apr 5, 2025
Llama 4 Maverick 17B from Groq - text, image input, 131,072 token context

Llama Guard 4 12B
Groq · Apr 5, 2025
Llama Guard 4 12B from Groq - text, image input, 131,072 token context

Llama-4-Scout-17B-16E-Instruct
Synthetic · Apr 5, 2025
Llama-4-Scout-17B-16E-Instruct from Synthetic - text, image input, 328,000 token context

Llama-4-Maverick-17B-128E-Instruct-FP8
Synthetic · Apr 5, 2025
Llama-4-Maverick-17B-128E-Instruct-FP8 from Synthetic - text, image input, 524,000 token context

Llama 4 Maverick 17B Instruct
AWS Bedrock · Apr 5, 2025
Open multimodal Llama for strong reasoning with efficient everyday serving

Llama 4 Scout 17B Instruct
AWS Bedrock · Apr 5, 2025
Open Llama with long-context vision for efficient multimodal agents

Llama 4 Maverick 17B Instruct (US)
AWS Bedrock · Apr 5, 2025
Open multimodal Llama for strong reasoning with efficient everyday serving

