Guide
2,496 step-by-step guides to setting up AI models from OpenAI, Anthropic, Google, OpenRouter, and more in TypingMind with your own API key.
Search and filter guides
Showing 2,245-2,295 of 2,496 guides

Qwen QwQ 32B
Groq · Nov 27, 2024
Qwen QwQ 32B from Groq - text input, 131,072 token context

OpenAI: GPT-4o (2024-11-20)
OpenRouter · Nov 20, 2024
The 2024-11-20 version of GPT-4o offers a leveled-up creative writing ability with more natural, engaging, and tailored writing to improve relevance & readability. It’s also better at working with uploaded...

GPT-4o (2024-11-20)
OpenAI · Nov 20, 2024
GPT model for general reasoning, writing, coding, and tool-assisted tasks

Mistral Large 2411
OpenRouter · Nov 19, 2024
Mistral Large 2 2411 is an update of [Mistral Large 2](/mistralai/mistral-large) released together with [Pixtral Large 2411](/mistralai/pixtral-large-2411) It provides a significant upgrade on the previous [Mistral Large 24.07](/mistralai/mistral-large-2407), with notable...

Mistral Large 2407
OpenRouter · Nov 19, 2024
This is Mistral AI's flagship model, Mistral Large 2 (version mistral-large-2407). It's a proprietary weights-available model and excels at reasoning, code, JSON, chat, and more. Read the launch announcement [here](https://mistral.ai/news/mistral-large-2407/)....

Mistral: Pixtral Large 2411
OpenRouter · Nov 19, 2024
Pixtral Large is a 124B parameter, open-weight, multimodal model built on top of [Mistral Large 2](/mistralai/mistral-large-2411). The model is able to understand documents, charts and natural images. The model is...

Mistral Large 2.1
Mistral · Nov 18, 2024
Flagship Mistral model for advanced reasoning, coding, and multilingual work
Qwen2.5-Coder-32B-Instruct
Hugging Face · Nov 12, 2024
Qwen coding model for software agents, repository edits, and code reasoning

Qwen2.5 Coder 32B Instruct
OpenRouter · Nov 11, 2024
Qwen2.5-Coder is the latest series of Code-Specific Qwen large language models (formerly known as CodeQwen). Qwen2.5-Coder brings the following improvements upon CodeQwen1.5: - Significantly improvements in **code generation**, **code reasoning**...

Qwen2.5-Coder-32B-Instruct
Synthetic · Nov 11, 2024
Qwen2.5-Coder-32B-Instruct from Synthetic - text input, 32,768 token context

SorcererLM 8x22B
OpenRouter · Nov 8, 2024
SorcererLM is an advanced RP and storytelling model, built as a Low-rank 16-bit LoRA fine-tuned on [WizardLM-2 8x22B](/microsoft/wizardlm-2-8x22b). - Advanced reasoning and emotional intelligence for engaging and immersive interactions - Vivid writing capabilities enriched with spatial and contextual awareness - Enhanced narrative depth, promoting creative and dynamic storytelling

TheDrummer: UnslopNemo 12B
OpenRouter · Nov 8, 2024
UnslopNemo v4.1 is the latest addition from the creator of Rocinante, designed for adventure writing and role-play scenarios.

Anthropic: Claude 3.5 Haiku (2024-10-22)
OpenRouter · Nov 4, 2024
Claude 3.5 Haiku features enhancements across all skill sets including coding, tool use, and reasoning. As the fastest model in the Anthropic lineup, it offers rapid response times suitable for applications that require high interactivity and low latency, such as user-facing chatbots and on-the-fly code completions. It also excels in specialized tasks like data extraction and real-time content moderation, making it a versatile tool for a broad range of industries. It does not support image inputs. See the launch announcement and benchmark results [here](https://www.anthropic.com/news/3-5-models-and-computer-use)

Anthropic: Claude 3.5 Haiku
OpenRouter · Nov 4, 2024
Claude 3.5 Haiku features offers enhanced capabilities in speed, coding accuracy, and tool use. Engineered to excel in real-time applications, it delivers quick response times that are essential for dynamic...

Grok Vision Beta
xAI · Nov 1, 2024
Grok Vision Beta from xAI - text, image input, 8,192 token context

Grok Beta
xAI · Nov 1, 2024
Grok Beta from xAI - text input, 131,072 token context

Mistral Large 3
Mistral · Nov 1, 2024
Mistral's largest general model for enterprise agents, coding, and multilingual reasoning

Pixtral Large (latest)
Mistral · Nov 1, 2024
Mistral's larger vision model for document-heavy image understanding and chat

Mistral Large (latest)
Mistral · Nov 1, 2024
Flagship Mistral model for advanced reasoning, coding, and multilingual work

FLUX1.1 [pro] Ultra
Vercel AI Gateway · Nov 1, 2024
Image model for prompt-driven generation, editing, and visual design workflows

Qwen Turbo
Alibaba Cloud · Nov 1, 2024
Efficient Qwen model for fast chat, extraction, and high-volume workloads

Recraft V3
Vercel AI Gateway · Oct 30, 2024
Image model for prompt-driven generation, editing, and visual design workflows

Qwen-VL OCR
Alibaba Cloud · Oct 28, 2024
OCR model for extracting structured text from documents and screenshots

Aya Expanse 8B
Cohere · Oct 24, 2024
Compact open multilingual model optimized for generation across 23 languages

Aya Expanse 32B
Cohere · Oct 24, 2024
Open multilingual model optimized for generation across 23 languages

Magnum v4 72B
OpenRouter · Oct 22, 2024
This is a series of models designed to replicate the prose quality of the Claude 3 models, specifically Sonnet(https://openrouter.ai/anthropic/claude-3.5-sonnet) and Opus(https://openrouter.ai/anthropic/claude-3-opus). The model is fine-tuned on top of [Qwen2.5 72B](https://openrouter.ai/qwen/qwen-2.5-72b-instruct).

Anthropic: Claude 3.5 Sonnet
OpenRouter · Oct 22, 2024
New Claude 3.5 Sonnet delivers better-than-Opus capabilities, faster-than-Sonnet speeds, at the same Sonnet prices. Sonnet is particularly good at: - Coding: Scores ~49% on SWE-Bench Verified, higher than the last best score, and without any fancy prompt scaffolding - Data science: Augments human data science expertise; navigates unstructured data while using multiple tools for insights - Visual processing: excelling at interpreting charts, graphs, and images, accurately transcribing text to derive insights beyond just the text alone - Agentic tasks: exceptional tool use, making it great at agentic tasks (i.e. complex, multi-step problem solving tasks that require engaging with other systems) #multimodal

Claude Sonnet 3.5 v2
Anthropic · Oct 22, 2024
Claude Sonnet 3.5 v2 from Anthropic - text, image, pdf input, 200,000 token context

Claude Haiku 3.5 (latest)
Anthropic · Oct 22, 2024
Claude Haiku 3.5 (latest) from Anthropic - text, image, pdf input, 200,000 token context

Claude Haiku 3.5
Anthropic · Oct 22, 2024
Claude Haiku 3.5 from Anthropic - text, image, pdf input, 200,000 token context

Claude Sonnet 3.5 v2
AWS Bedrock · Oct 22, 2024
Claude Sonnet 3.5 v2 from AWS Bedrock - text, image, pdf input, 200,000 token context

Claude Haiku 3.5
AWS Bedrock · Oct 22, 2024
Claude Haiku 3.5 from AWS Bedrock - text, image, pdf input, 200,000 token context

Mistral: Ministral 8B
OpenRouter · Oct 17, 2024
Ministral 8B is an 8B parameter model featuring a unique interleaved sliding-window attention pattern for faster, memory-efficient inference. Designed for edge use cases, it supports up to 128k context length and excels in knowledge and reasoning tasks. It outperforms peers in the sub-10B category, making it perfect for low-latency, privacy-first applications.

Mistral: Ministral 3B
OpenRouter · Oct 17, 2024
Ministral 3B is a 3B parameter model optimized for on-device and edge computing. It excels in knowledge, commonsense reasoning, and function-calling, outperforming larger models like Mistral 7B on most benchmarks. Supporting up to 128k context length, it’s ideal for orchestrating agentic workflows and specialist tasks with efficient inference.

Qwen: Qwen2.5 7B Instruct
OpenRouter · Oct 16, 2024
Qwen2.5 7B is the latest series of Qwen large language models. Qwen2.5 brings the following improvements upon Qwen2: - Significantly more knowledge and has greatly improved capabilities in coding and...

NVIDIA: Llama 3.1 Nemotron 70B Instruct
OpenRouter · Oct 15, 2024
NVIDIA's Llama 3.1 Nemotron 70B is a language model designed for generating precise and useful responses. Leveraging [Llama 3.1 70B](/models/meta-llama/llama-3.1-70b-instruct) architecture and Reinforcement Learning from Human Feedback (RLHF), it excels...

Qwen 2.5 72B Instruct
Novita · Oct 15, 2024
Qwen instruction model for multilingual chat, reasoning, and tool use

Inflection: Inflection 3 Pi
OpenRouter · Oct 11, 2024
Inflection 3 Pi powers Inflection's [Pi](https://pi.ai) chatbot, including backstory, emotional intelligence, productivity, and safety. It has access to recent news, and excels in scenarios like customer support and roleplay. Pi...

Inflection: Inflection 3 Productivity
OpenRouter · Oct 11, 2024
Inflection 3 Productivity is optimized for following instructions. It is better for tasks requiring JSON output or precise adherence to provided guidelines. It has access to recent news. For emotional...

Palmyra X4
AWS Bedrock · Oct 9, 2024
Enterprise language model for workflow automation, coding, data analysis, and tool use

Palmyra X4 (US)
AWS Bedrock · Oct 9, 2024
Enterprise language model for workflow automation, coding, data analysis, and tool use

Gemini 1.5 Flash-8B
Google · Oct 3, 2024
Gemini 1.5 Flash-8B from Google - text, image, audio, video input, 1,000,000 token context

Llama 3.2 3B
Venice AI · Oct 3, 2024
Open Llama instruction model for multilingual chat, reasoning, and coding

FLUX1.1 [pro]
Vercel AI Gateway · Oct 2, 2024
Image model for prompt-driven generation, editing, and visual design workflows

Whisper Large V3 Turbo
Groq · Oct 1, 2024
Speech transcription model for accurate audio-to-text and captioning workflows

Ministral 8B (latest)
Mistral · Oct 1, 2024
Compact Mistral model for edge, latency-sensitive, and cost-efficient workloads

Ministral 3B (latest)
Mistral · Oct 1, 2024
Compact Mistral model for edge, latency-sensitive, and cost-efficient workloads

FLUX.1 Fill [pro]
Vercel AI Gateway · Oct 1, 2024
Image model for prompt-driven generation, editing, and visual design workflows

Ministral 8B (latest)
Vercel AI Gateway · Oct 1, 2024
Compact Mistral model for edge, latency-sensitive, and cost-efficient workloads

Ministral 3B (latest)
Vercel AI Gateway · Oct 1, 2024
Compact Mistral model for edge, latency-sensitive, and cost-efficient workloads

TheDrummer: Rocinante 12B
OpenRouter · Sep 30, 2024
Rocinante 12B is designed for engaging storytelling and rich prose. Early testers have reported: - Expanded vocabulary with unique and expressive word choices - Enhanced creativity for vivid narratives -...

