Guide

2,496 step-by-step guides to setting up AI models from OpenAI, Anthropic, Google, OpenRouter, and more in TypingMind with your own API key.

Search and filter guides

Showing 2,296-2,346 of 2,496 guides

OpenRouter logo

Meta: Llama 3.2 3B Instruct (free)

OpenRouter · Sep 25, 2024

Llama 3.2 3B is a 3-billion-parameter multilingual large language model, optimized for advanced natural language processing tasks like dialogue generation, reasoning, and summarization. Designed with the latest transformer architecture, it...

131.1K context
View
OpenRouter logo

Meta: Llama 3.2 3B Instruct

OpenRouter · Sep 25, 2024

Llama 3.2 3B is a 3-billion-parameter multilingual large language model, optimized for advanced natural language processing tasks like dialogue generation, reasoning, and summarization. Designed with the latest transformer architecture, it...

131.1K context
View
OpenRouter logo

Meta: Llama 3.2 1B Instruct

OpenRouter · Sep 25, 2024

Llama 3.2 1B is a 1-billion-parameter language model focused on efficiently performing natural language tasks, such as summarization, dialogue, and multilingual text analysis. Its smaller size allows it to operate...

60K context
View
OpenRouter logo

Meta: Llama 3.2 90B Vision Instruct

OpenRouter · Sep 25, 2024

The Llama 90B Vision model is a top-tier, 90-billion-parameter multimodal model designed for the most challenging visual reasoning and language tasks. It offers unparalleled accuracy in image captioning, visual question answering, and advanced image-text comprehension. Pre-trained on vast multimodal datasets and fine-tuned with human feedback, the Llama 90B Vision is engineered to handle the most demanding image-based AI tasks. This model is perfect for industries requiring cutting-edge multimodal AI capabilities, particularly those dealing with complex, real-time visual and textual analysis. Click here for the [original model card](https://github.com/meta-llama/llama-models/blob/main/models/llama3_2/MODEL_CARD_VISION.md). Usage of this model is subject to [Meta's Acceptable Use Policy](https://www.llama.com/llama3/use-policy/).

32.8K context
View
OpenRouter logo

Meta: Llama 3.2 11B Vision Instruct

OpenRouter · Sep 25, 2024

Llama 3.2 11B Vision is a multimodal model with 11 billion parameters, designed to handle tasks combining visual and textual data. It excels in tasks such as image captioning and...

131.1K context
View
AWS Bedrock logo

Llama 3.2 90B Instruct

AWS Bedrock · Sep 25, 2024

Llama 3.2 90B Instruct from AWS Bedrock - text, image input, 128,000 token context

128K context
View
AWS Bedrock logo

Llama 3.2 1B Instruct

AWS Bedrock · Sep 25, 2024

Llama 3.2 1B Instruct from AWS Bedrock - text input, 131,000 token context

131K context
View
AWS Bedrock logo

Llama 3.2 11B Instruct

AWS Bedrock · Sep 25, 2024

Llama 3.2 11B Instruct from AWS Bedrock - text, image input, 128,000 token context

128K context
View
AWS Bedrock logo

Llama 3.2 3B Instruct

AWS Bedrock · Sep 25, 2024

Llama 3.2 3B Instruct from AWS Bedrock - text input, 131,000 token context

131K context
View
OpenRouter logo

Qwen2.5 72B Instruct

OpenRouter · Sep 19, 2024

Qwen2.5 72B is the latest series of Qwen large language models. Qwen2.5 brings the following improvements upon Qwen2: - Significantly more knowledge and has greatly improved capabilities in coding and...

32.8K context
View
Together AI logo

Qwen 2.5 7B Instruct Turbo

Together AI · Sep 19, 2024

Efficient Qwen model for fast chat, extraction, and high-volume workloads

32.8K context
View
Novita logo

L31 70B Euryale V2.2

Novita · Sep 19, 2024

Open-weight instruction model for adaptable chat and self-hosted production workloads

8.2K context
View
Novita logo

Llama 3.2 3B Instruct

Novita · Sep 18, 2024

Open Llama instruction model for multilingual chat, reasoning, and coding

32.8K context
View
Vercel AI Gateway logo

Mistral Small (latest)

Vercel AI Gateway · Sep 17, 2024

Efficient Mistral model for fast chat, extraction, and production assistants

262.1K context
View
OpenRouter logo

NeverSleep: Lumimaid v0.2 8B

OpenRouter · Sep 15, 2024

Lumimaid v0.2 8B is a finetune of [Llama 3.1 8B](/models/meta-llama/llama-3.1-8b-instruct) with a "HUGE step up dataset wise" compared to Lumimaid v0.1. Sloppy chats output were purged. Usage of this model is subject to [Meta's Acceptable Use Policy](https://llama.meta.com/llama3/use-policy/).

32.8K context
View
OpenAI logo

o1-preview

OpenAI · Sep 12, 2024

o1-preview from OpenAI - text input, 128,000 token context

128K context
View
OpenAI logo

o1-mini

OpenAI · Sep 12, 2024

o1-mini from OpenAI - text input, 128,000 token context

128K context
View
OpenRouter logo

Mistral: Pixtral 12B

OpenRouter · Sep 10, 2024

The first multi-modal, text+image-to-text model from Mistral AI. Its weights were launched via torrent: https://x.com/mistralai/status/1833758285167722836.

32.8K context
View
Mistral logo

Pixtral 12B

Mistral · Sep 1, 2024

Mistral vision-language model for image understanding and multimodal chat

128K context
View
Alibaba Cloud logo

Qwen2.5 32B Instruct

Alibaba Cloud · Sep 1, 2024

Qwen instruction model for multilingual chat, reasoning, and tool use

131.1K context
View
Alibaba Cloud logo

Qwen2.5-VL 72B Instruct

Alibaba Cloud · Sep 1, 2024

Qwen vision-language model for visual reasoning, documents, and agent tasks

131.1K context
View
Alibaba Cloud logo

Qwen2.5 72B Instruct

Alibaba Cloud · Sep 1, 2024

Qwen instruction model for multilingual chat, reasoning, and tool use

131.1K context
View
Alibaba Cloud logo

Qwen2.5 7B Instruct

Alibaba Cloud · Sep 1, 2024

Qwen instruction model for multilingual chat, reasoning, and tool use

131.1K context
View
Alibaba Cloud logo

Qwen2.5 14B Instruct

Alibaba Cloud · Sep 1, 2024

Qwen instruction model for multilingual chat, reasoning, and tool use

131.1K context
View
Alibaba Cloud logo

Qwen2.5-VL 7B Instruct

Alibaba Cloud · Sep 1, 2024

Qwen vision-language model for visual reasoning, documents, and agent tasks

131.1K context
View
OpenRouter logo

Cohere: Command R (08-2024)

OpenRouter · Aug 30, 2024

command-r-08-2024 is an update of the [Command R](/models/cohere/command-r) with improved performance for multilingual retrieval-augmented generation (RAG) and tool use. More broadly, it is better at math, code and reasoning and...

128K context
View
OpenRouter logo

Cohere: Command R+ (08-2024)

OpenRouter · Aug 30, 2024

command-r-plus-08-2024 is an update of the [Command R+](/models/cohere/command-r-plus) with roughly 50% higher throughput and 25% lower latencies as compared to the previous Command R+ version, while keeping the hardware footprint...

128K context
View
Cohere logo

Command R

Cohere · Aug 30, 2024

Cohere retrieval model for long-context chat and enterprise RAG workflows

128K context
View
Cohere logo

Command R+

Cohere · Aug 30, 2024

Cohere's RAG workhorse for long-context enterprise search and tool use

128K context
View
OpenRouter logo

Sao10K: Llama 3.1 Euryale 70B v2.2

OpenRouter · Aug 28, 2024

Euryale L3.1 70B v2.2 is a model focused on creative roleplay from [Sao10k](https://ko-fi.com/sao10k). It is the successor of [Euryale L3 70B v2.1](/models/sao10k/l3-euryale-70b).

131.1K context
View
OpenRouter logo

Qwen: Qwen2.5-VL 7B Instruct (free)

OpenRouter · Aug 28, 2024

Qwen2.5 VL 7B is a multimodal LLM from the Qwen Team with the following key enhancements: - SoTA understanding of images of various resolution & ratio: Qwen2.5-VL achieves state-of-the-art performance on visual understanding benchmarks, including MathVista, DocVQA, RealWorldQA, MTVQA, etc. - Understanding videos of 20min+: Qwen2.5-VL can understand videos over 20 minutes for high-quality video-based question answering, dialog, content creation, etc. - Agent that can operate your mobiles, robots, etc.: with the abilities of complex reasoning and decision making, Qwen2.5-VL can be integrated with devices like mobile phones, robots, etc., for automatic operation based on visual environment and text instructions. - Multilingual Support: to serve global users, besides English and Chinese, Qwen2.5-VL now supports the understanding of texts in different languages inside images, including most European languages, Japanese, Korean, Arabic, Vietnamese, etc. For more details, see this [blog post](https://qwenlm.github.io/blog/qwen2-vl/) and [GitHub repo](https://github.com/QwenLM/Qwen2-VL). Usage of this model is subject to [Tongyi Qianwen LICENSE AGREEMENT](https://huggingface.co/Qwen/Qwen1.5-110B-Chat/blob/main/LICENSE).

32.8K context
View
OpenRouter logo

Qwen: Qwen2.5-VL 7B Instruct

OpenRouter · Aug 28, 2024

Qwen2.5 VL 7B is a multimodal LLM from the Qwen Team with the following key enhancements: - SoTA understanding of images of various resolution & ratio: Qwen2.5-VL achieves state-of-the-art performance on visual understanding benchmarks, including MathVista, DocVQA, RealWorldQA, MTVQA, etc. - Understanding videos of 20min+: Qwen2.5-VL can understand videos over 20 minutes for high-quality video-based question answering, dialog, content creation, etc. - Agent that can operate your mobiles, robots, etc.: with the abilities of complex reasoning and decision making, Qwen2.5-VL can be integrated with devices like mobile phones, robots, etc., for automatic operation based on visual environment and text instructions. - Multilingual Support: to serve global users, besides English and Chinese, Qwen2.5-VL now supports the understanding of texts in different languages inside images, including most European languages, Japanese, Korean, Arabic, Vietnamese, etc. For more details, see this [blog post](https://qwenlm.github.io/blog/qwen2-vl/) and [GitHub repo](https://github.com/QwenLM/Qwen2-VL). Usage of this model is subject to [Tongyi Qianwen LICENSE AGREEMENT](https://huggingface.co/Qwen/Qwen1.5-110B-Chat/blob/main/LICENSE).

32.8K context
View
xAI logo

Grok 2 Vision

xAI · Aug 20, 2024

Grok 2 Vision from xAI - text, image input, 8,192 token context

8.2K context
View
xAI logo

Grok 2

xAI · Aug 20, 2024

Grok 2 from xAI - text input, 131,072 token context

131.1K context
View
xAI logo

Grok 2 Vision (1212)

xAI · Aug 20, 2024

Grok 2 Vision (1212) from xAI - text, image input, 8,192 token context

8.2K context
View
xAI logo

Grok 2 Latest

xAI · Aug 20, 2024

Grok 2 Latest from xAI - text input, 131,072 token context

131.1K context
View
xAI logo

Grok 2 Vision Latest

xAI · Aug 20, 2024

Grok 2 Vision Latest from xAI - text, image input, 8,192 token context

8.2K context
View
OpenRouter logo

Nous: Hermes 3 70B Instruct

OpenRouter · Aug 18, 2024

Hermes 3 is a generalist language model with many improvements over [Hermes 2](/models/nousresearch/nous-hermes-2-mistral-7b-dpo), including advanced agentic capabilities, much better roleplaying, reasoning, multi-turn conversation, long context coherence, and improvements across the...

131.1K context
View
OpenRouter logo

Nous: Hermes 3 405B Instruct (free)

OpenRouter · Aug 16, 2024

Hermes 3 is a generalist language model with many improvements over Hermes 2, including advanced agentic capabilities, much better roleplaying, reasoning, multi-turn conversation, long context coherence, and improvements across the...

131.1K context
View
OpenRouter logo

Nous: Hermes 3 405B Instruct

OpenRouter · Aug 16, 2024

Hermes 3 is a generalist language model with many improvements over Hermes 2, including advanced agentic capabilities, much better roleplaying, reasoning, multi-turn conversation, long context coherence, and improvements across the...

131.1K context
View
AWS Bedrock logo

Jamba 1.5 Mini

AWS Bedrock · Aug 15, 2024

Jamba 1.5 Mini from AWS Bedrock - text input, 256,000 token context

256K context
View
AWS Bedrock logo

Jamba 1.5 Large

AWS Bedrock · Aug 15, 2024

Jamba 1.5 Large from AWS Bedrock - text input, 256,000 token context

256K context
View
Vercel AI Gateway logo

Morph v3 Fast

Vercel AI Gateway · Aug 15, 2024

Efficient model for low-latency assistance, extraction, and routine automation

16K context
View
Vercel AI Gateway logo

Morph v3 Large

Vercel AI Gateway · Aug 15, 2024

Flagship model for demanding analysis, coding, and production agent workflows

32K context
View
OpenRouter logo

OpenAI: ChatGPT-4o

OpenRouter · Aug 14, 2024

OpenAI ChatGPT 4o is continually updated by OpenAI to point to the current version of GPT-4o used by ChatGPT. It therefore differs slightly from the API version of [GPT-4o](/models/openai/gpt-4o) in that it has additional RLHF. It is intended for research and evaluation. OpenAI notes that this model is not suited for production use-cases as it may be removed or redirected to another model in the future.

128K context
View
OpenRouter logo

Sao10K: Llama 3 8B Lunaris

OpenRouter · Aug 13, 2024

Lunaris 8B is a versatile generalist and roleplaying model based on Llama 3. It's a strategic merge of multiple models, designed to balance creativity with improved logic and general knowledge....

8.2K context
View
OpenRouter logo

OpenAI: GPT-4o (2024-08-06)

OpenRouter · Aug 6, 2024

The 2024-08-06 version of GPT-4o offers improved performance in structured outputs, with the ability to supply a JSON schema in the respone_format. Read more [here](https://openai.com/index/introducing-structured-outputs-in-the-api/). GPT-4o ("o" for "omni") is...

128K context
View
OpenAI logo

GPT-4o (2024-08-06)

OpenAI · Aug 6, 2024

GPT model for general reasoning, writing, coding, and tool-assisted tasks

128K context
View
OpenRouter logo

Meta: Llama 3.1 405B (base)

OpenRouter · Aug 2, 2024

Meta's latest class of model (Llama 3.1) launched with a variety of sizes & flavors. This is the base 405B pre-trained version. It has demonstrated strong performance compared to leading closed-source models in human evaluations. To read more about the model release, [click here](https://ai.meta.com/blog/meta-llama-3/). Usage of this model is subject to [Meta's Acceptable Use Policy](https://llama.meta.com/llama3/use-policy/).

32.8K context
View
Vercel AI Gateway logo

Flux Schnell

Vercel AI Gateway · Aug 2, 2024

Image model for prompt-driven generation, editing, and visual design workflows

512 context
View
Vercel AI Gateway logo

Text Embedding 005

Vercel AI Gateway · Aug 1, 2024

Embedding model for semantic search, retrieval, clustering, and ranking pipelines

8.2K context
View