Guide
2,496 step-by-step guides to setting up AI models from OpenAI, Anthropic, Google, OpenRouter, and more in TypingMind with your own API key.
Search and filter guides
Showing 2,296-2,346 of 2,496 guides

Meta: Llama 3.2 3B Instruct (free)
OpenRouter · Sep 25, 2024
Llama 3.2 3B is a 3-billion-parameter multilingual large language model, optimized for advanced natural language processing tasks like dialogue generation, reasoning, and summarization. Designed with the latest transformer architecture, it...

Meta: Llama 3.2 3B Instruct
OpenRouter · Sep 25, 2024
Llama 3.2 3B is a 3-billion-parameter multilingual large language model, optimized for advanced natural language processing tasks like dialogue generation, reasoning, and summarization. Designed with the latest transformer architecture, it...

Meta: Llama 3.2 1B Instruct
OpenRouter · Sep 25, 2024
Llama 3.2 1B is a 1-billion-parameter language model focused on efficiently performing natural language tasks, such as summarization, dialogue, and multilingual text analysis. Its smaller size allows it to operate...

Meta: Llama 3.2 90B Vision Instruct
OpenRouter · Sep 25, 2024
The Llama 90B Vision model is a top-tier, 90-billion-parameter multimodal model designed for the most challenging visual reasoning and language tasks. It offers unparalleled accuracy in image captioning, visual question answering, and advanced image-text comprehension. Pre-trained on vast multimodal datasets and fine-tuned with human feedback, the Llama 90B Vision is engineered to handle the most demanding image-based AI tasks. This model is perfect for industries requiring cutting-edge multimodal AI capabilities, particularly those dealing with complex, real-time visual and textual analysis. Click here for the [original model card](https://github.com/meta-llama/llama-models/blob/main/models/llama3_2/MODEL_CARD_VISION.md). Usage of this model is subject to [Meta's Acceptable Use Policy](https://www.llama.com/llama3/use-policy/).

Meta: Llama 3.2 11B Vision Instruct
OpenRouter · Sep 25, 2024
Llama 3.2 11B Vision is a multimodal model with 11 billion parameters, designed to handle tasks combining visual and textual data. It excels in tasks such as image captioning and...

Llama 3.2 90B Instruct
AWS Bedrock · Sep 25, 2024
Llama 3.2 90B Instruct from AWS Bedrock - text, image input, 128,000 token context

Llama 3.2 1B Instruct
AWS Bedrock · Sep 25, 2024
Llama 3.2 1B Instruct from AWS Bedrock - text input, 131,000 token context

Llama 3.2 11B Instruct
AWS Bedrock · Sep 25, 2024
Llama 3.2 11B Instruct from AWS Bedrock - text, image input, 128,000 token context

Llama 3.2 3B Instruct
AWS Bedrock · Sep 25, 2024
Llama 3.2 3B Instruct from AWS Bedrock - text input, 131,000 token context

Qwen2.5 72B Instruct
OpenRouter · Sep 19, 2024
Qwen2.5 72B is the latest series of Qwen large language models. Qwen2.5 brings the following improvements upon Qwen2: - Significantly more knowledge and has greatly improved capabilities in coding and...

Qwen 2.5 7B Instruct Turbo
Together AI · Sep 19, 2024
Efficient Qwen model for fast chat, extraction, and high-volume workloads

L31 70B Euryale V2.2
Novita · Sep 19, 2024
Open-weight instruction model for adaptable chat and self-hosted production workloads

Llama 3.2 3B Instruct
Novita · Sep 18, 2024
Open Llama instruction model for multilingual chat, reasoning, and coding

Mistral Small (latest)
Vercel AI Gateway · Sep 17, 2024
Efficient Mistral model for fast chat, extraction, and production assistants

NeverSleep: Lumimaid v0.2 8B
OpenRouter · Sep 15, 2024
Lumimaid v0.2 8B is a finetune of [Llama 3.1 8B](/models/meta-llama/llama-3.1-8b-instruct) with a "HUGE step up dataset wise" compared to Lumimaid v0.1. Sloppy chats output were purged. Usage of this model is subject to [Meta's Acceptable Use Policy](https://llama.meta.com/llama3/use-policy/).

o1-preview
OpenAI · Sep 12, 2024
o1-preview from OpenAI - text input, 128,000 token context

o1-mini
OpenAI · Sep 12, 2024
o1-mini from OpenAI - text input, 128,000 token context

Mistral: Pixtral 12B
OpenRouter · Sep 10, 2024
The first multi-modal, text+image-to-text model from Mistral AI. Its weights were launched via torrent: https://x.com/mistralai/status/1833758285167722836.

Pixtral 12B
Mistral · Sep 1, 2024
Mistral vision-language model for image understanding and multimodal chat

Qwen2.5 32B Instruct
Alibaba Cloud · Sep 1, 2024
Qwen instruction model for multilingual chat, reasoning, and tool use

Qwen2.5-VL 72B Instruct
Alibaba Cloud · Sep 1, 2024
Qwen vision-language model for visual reasoning, documents, and agent tasks

Qwen2.5 72B Instruct
Alibaba Cloud · Sep 1, 2024
Qwen instruction model for multilingual chat, reasoning, and tool use

Qwen2.5 7B Instruct
Alibaba Cloud · Sep 1, 2024
Qwen instruction model for multilingual chat, reasoning, and tool use

Qwen2.5 14B Instruct
Alibaba Cloud · Sep 1, 2024
Qwen instruction model for multilingual chat, reasoning, and tool use

Qwen2.5-VL 7B Instruct
Alibaba Cloud · Sep 1, 2024
Qwen vision-language model for visual reasoning, documents, and agent tasks

Cohere: Command R (08-2024)
OpenRouter · Aug 30, 2024
command-r-08-2024 is an update of the [Command R](/models/cohere/command-r) with improved performance for multilingual retrieval-augmented generation (RAG) and tool use. More broadly, it is better at math, code and reasoning and...

Cohere: Command R+ (08-2024)
OpenRouter · Aug 30, 2024
command-r-plus-08-2024 is an update of the [Command R+](/models/cohere/command-r-plus) with roughly 50% higher throughput and 25% lower latencies as compared to the previous Command R+ version, while keeping the hardware footprint...

Command R
Cohere · Aug 30, 2024
Cohere retrieval model for long-context chat and enterprise RAG workflows

Command R+
Cohere · Aug 30, 2024
Cohere's RAG workhorse for long-context enterprise search and tool use

Sao10K: Llama 3.1 Euryale 70B v2.2
OpenRouter · Aug 28, 2024
Euryale L3.1 70B v2.2 is a model focused on creative roleplay from [Sao10k](https://ko-fi.com/sao10k). It is the successor of [Euryale L3 70B v2.1](/models/sao10k/l3-euryale-70b).

Qwen: Qwen2.5-VL 7B Instruct (free)
OpenRouter · Aug 28, 2024
Qwen2.5 VL 7B is a multimodal LLM from the Qwen Team with the following key enhancements: - SoTA understanding of images of various resolution & ratio: Qwen2.5-VL achieves state-of-the-art performance on visual understanding benchmarks, including MathVista, DocVQA, RealWorldQA, MTVQA, etc. - Understanding videos of 20min+: Qwen2.5-VL can understand videos over 20 minutes for high-quality video-based question answering, dialog, content creation, etc. - Agent that can operate your mobiles, robots, etc.: with the abilities of complex reasoning and decision making, Qwen2.5-VL can be integrated with devices like mobile phones, robots, etc., for automatic operation based on visual environment and text instructions. - Multilingual Support: to serve global users, besides English and Chinese, Qwen2.5-VL now supports the understanding of texts in different languages inside images, including most European languages, Japanese, Korean, Arabic, Vietnamese, etc. For more details, see this [blog post](https://qwenlm.github.io/blog/qwen2-vl/) and [GitHub repo](https://github.com/QwenLM/Qwen2-VL). Usage of this model is subject to [Tongyi Qianwen LICENSE AGREEMENT](https://huggingface.co/Qwen/Qwen1.5-110B-Chat/blob/main/LICENSE).

Qwen: Qwen2.5-VL 7B Instruct
OpenRouter · Aug 28, 2024
Qwen2.5 VL 7B is a multimodal LLM from the Qwen Team with the following key enhancements: - SoTA understanding of images of various resolution & ratio: Qwen2.5-VL achieves state-of-the-art performance on visual understanding benchmarks, including MathVista, DocVQA, RealWorldQA, MTVQA, etc. - Understanding videos of 20min+: Qwen2.5-VL can understand videos over 20 minutes for high-quality video-based question answering, dialog, content creation, etc. - Agent that can operate your mobiles, robots, etc.: with the abilities of complex reasoning and decision making, Qwen2.5-VL can be integrated with devices like mobile phones, robots, etc., for automatic operation based on visual environment and text instructions. - Multilingual Support: to serve global users, besides English and Chinese, Qwen2.5-VL now supports the understanding of texts in different languages inside images, including most European languages, Japanese, Korean, Arabic, Vietnamese, etc. For more details, see this [blog post](https://qwenlm.github.io/blog/qwen2-vl/) and [GitHub repo](https://github.com/QwenLM/Qwen2-VL). Usage of this model is subject to [Tongyi Qianwen LICENSE AGREEMENT](https://huggingface.co/Qwen/Qwen1.5-110B-Chat/blob/main/LICENSE).

Grok 2 Vision
xAI · Aug 20, 2024
Grok 2 Vision from xAI - text, image input, 8,192 token context

Grok 2
xAI · Aug 20, 2024
Grok 2 from xAI - text input, 131,072 token context

Grok 2 Vision (1212)
xAI · Aug 20, 2024
Grok 2 Vision (1212) from xAI - text, image input, 8,192 token context

Grok 2 Latest
xAI · Aug 20, 2024
Grok 2 Latest from xAI - text input, 131,072 token context

Grok 2 Vision Latest
xAI · Aug 20, 2024
Grok 2 Vision Latest from xAI - text, image input, 8,192 token context

Nous: Hermes 3 70B Instruct
OpenRouter · Aug 18, 2024
Hermes 3 is a generalist language model with many improvements over [Hermes 2](/models/nousresearch/nous-hermes-2-mistral-7b-dpo), including advanced agentic capabilities, much better roleplaying, reasoning, multi-turn conversation, long context coherence, and improvements across the...

Nous: Hermes 3 405B Instruct (free)
OpenRouter · Aug 16, 2024
Hermes 3 is a generalist language model with many improvements over Hermes 2, including advanced agentic capabilities, much better roleplaying, reasoning, multi-turn conversation, long context coherence, and improvements across the...

Nous: Hermes 3 405B Instruct
OpenRouter · Aug 16, 2024
Hermes 3 is a generalist language model with many improvements over Hermes 2, including advanced agentic capabilities, much better roleplaying, reasoning, multi-turn conversation, long context coherence, and improvements across the...

Jamba 1.5 Mini
AWS Bedrock · Aug 15, 2024
Jamba 1.5 Mini from AWS Bedrock - text input, 256,000 token context

Jamba 1.5 Large
AWS Bedrock · Aug 15, 2024
Jamba 1.5 Large from AWS Bedrock - text input, 256,000 token context

Morph v3 Fast
Vercel AI Gateway · Aug 15, 2024
Efficient model for low-latency assistance, extraction, and routine automation

Morph v3 Large
Vercel AI Gateway · Aug 15, 2024
Flagship model for demanding analysis, coding, and production agent workflows

OpenAI: ChatGPT-4o
OpenRouter · Aug 14, 2024
OpenAI ChatGPT 4o is continually updated by OpenAI to point to the current version of GPT-4o used by ChatGPT. It therefore differs slightly from the API version of [GPT-4o](/models/openai/gpt-4o) in that it has additional RLHF. It is intended for research and evaluation. OpenAI notes that this model is not suited for production use-cases as it may be removed or redirected to another model in the future.

Sao10K: Llama 3 8B Lunaris
OpenRouter · Aug 13, 2024
Lunaris 8B is a versatile generalist and roleplaying model based on Llama 3. It's a strategic merge of multiple models, designed to balance creativity with improved logic and general knowledge....

OpenAI: GPT-4o (2024-08-06)
OpenRouter · Aug 6, 2024
The 2024-08-06 version of GPT-4o offers improved performance in structured outputs, with the ability to supply a JSON schema in the respone_format. Read more [here](https://openai.com/index/introducing-structured-outputs-in-the-api/). GPT-4o ("o" for "omni") is...

GPT-4o (2024-08-06)
OpenAI · Aug 6, 2024
GPT model for general reasoning, writing, coding, and tool-assisted tasks

Meta: Llama 3.1 405B (base)
OpenRouter · Aug 2, 2024
Meta's latest class of model (Llama 3.1) launched with a variety of sizes & flavors. This is the base 405B pre-trained version. It has demonstrated strong performance compared to leading closed-source models in human evaluations. To read more about the model release, [click here](https://ai.meta.com/blog/meta-llama-3/). Usage of this model is subject to [Meta's Acceptable Use Policy](https://llama.meta.com/llama3/use-policy/).

Flux Schnell
Vercel AI Gateway · Aug 2, 2024
Image model for prompt-driven generation, editing, and visual design workflows

Text Embedding 005
Vercel AI Gateway · Aug 1, 2024
Embedding model for semantic search, retrieval, clustering, and ranking pipelines

