๐Ÿ‘๏ธ Vision

Models that understand images and visual content. 384 models.

0

Yi 34B Chat Model Documentation

01.AI

The Yi series models are large language models trained from scratch by developers at 01.AI. This 34B parameter model has been instruct-tuned specifically for chat applications, providing optimized per...

Vision
0

Yi Vision 34B Model Documentation

01.AI

Yi-VL-34B is the world's first open-source 34 billion parameter vision-language model, combining advanced image understanding with multilingual text generation capabilities. It represents a significan...

Vision
A

AllenAI: Olmo 3.1 32B Think (free)

Allen AI

This is a 32-billion parameter model emphasizing reasoning capabilities. The system excels at deep reasoning, complex multi-step logic, and advanced instruction following. Version 3.1 represents impro...

Vision
A

Goliath 120B Model Documentation

Alpindale

Goliath 120B is a merged model that combines "two fine-tuned Llama 70B models into one 120B model" by merging Xwin and Euryale variants. The model was created using the mergekit framework by @chargodd...

Vision
A

Amazon Nova 2 Lite v1

Amazon

Vision
A

Amazon Nova Pro 1.0

Amazon

Amazon's multimodal model designed to balance "accuracy, speed, and cost for a wide range of tasks." As of December 2024, it demonstrates state-of-the-art performance on visual question answering (Tex...

Vision
A

Magnum v4 72B

Anthracite

Magnum v4 72B is a fine-tuned version of Qwen2.5 72B that aims to replicate the prose quality of Claude 3 models, specifically Sonnet and Opus. This model is designed for creative writing and roleplay...

Vision
A

Anthropic Claude Models on LangMart

Anthropic

Vision
A

Anthropic: Claude 3 Haiku (20240307)

Anthropic

Anthropic's fastest and most compact Claude 3 model. Designed for near-instant responses while maintaining high-quality output. Ideal for high-volume, cost-sensitive applications.

Vision Tools Streaming
A

Anthropic: Claude 3 Opus

Anthropic

Anthropic's most intelligent model with best-in-market performance on highly complex tasks. Navigates open-ended prompts and sight-unseen scenarios with remarkable fluency and human-like understanding...

Vision Tools Reasoning
A

Anthropic: Claude 3 Sonnet

Anthropic

Claude 3 Sonnet is an ideal balance of intelligence and speed for enterprise workloads. Maximum utility at a lower price, dependable, balanced for scaled deployments. It offers excellent performance w...

Vision
A

Anthropic: Claude 3 Sonnet (20240229)

Anthropic

Anthropic's balanced Claude 3 model offering a good combination of capability and speed. Suitable for most general-purpose applications requiring quality responses.

Vision Tools Streaming
A

Anthropic: Claude 3.5 Sonnet (20241022)

Anthropic

Anthropic's Claude 3.5 Sonnet is a balanced AI model that combines strong intelligence with fast response times. This specific version from October 22, 2024 provides a fixed checkpoint for reproducibl...

Vision Tools Streaming
A

Anthropic: Claude Haiku 4.5

Anthropic

Claude Haiku 4.5 is a conversational AI model designed for multi-turn dialogue and interactive tasks. Developed by Anthropic, this cost-effective model is optimized for efficient inference while maint...

Vision Tools Streaming
A

Anthropic: Claude Opus 4

Anthropic

Claude Opus 4 is a conversational AI model designed for multi-turn dialogue and interactive tasks. Developed by Anthropic, this model is optimized for its specific use case category.

Vision Tools Streaming
A

Anthropic: Claude Opus 4.1

Anthropic

Hybrid reasoning model pushing frontier for coding and AI agents with extended thinking capabilities. Achieves 74.5% on SWE-bench Verified with 32K max output tokens.

Vision Tools Reasoning
A

Anthropic: Claude Opus 4.5

Anthropic

Claude Opus 4.5 is a conversational AI model designed for multi-turn dialogue and interactive tasks. Developed by Anthropic, this flagship model represents the latest capabilities and state-of-the-art...

Vision Tools Streaming
A

Anthropic: Claude Sonnet 4

Anthropic

Claude Sonnet 4 is a conversational AI model designed for multi-turn dialogue and interactive tasks. Developed by Anthropic, this model is optimized for its specific use case category.

Vision Tools Streaming
A

Anthropic: Claude Sonnet 4.5

Anthropic

Claude Sonnet 4.5 is a conversational AI model designed for multi-turn dialogue and interactive tasks. Developed by Anthropic, this premium model offers excellent quality and balanced performance acro...

Vision Tools Streaming
A

Claude 3 Haiku

Anthropic

Claude 3 Haiku is Anthropic's fastest and most compact model, designed for near-instant responsiveness with quick and accurate targeted performance. It excels at tasks requiring rapid responses while ...

Vision Tools Streaming
A

Claude 3.5 Haiku

Anthropic

Claude 3.5 Haiku offers enhanced capabilities in speed, coding accuracy, and tool use. It is engineered to excel in real-time applications, delivering quick response times.

Vision
A

Claude Opus 4

Anthropic

Claude Opus 4 is benchmarked as the world's best coding model, at time of release, bringing sustained performance on complex, long-running tasks and agent workflows. It sets new benchmarks in software...

Tools Reasoning Vision
A

Claude Sonnet 4

Anthropic

Claude Sonnet 4 represents a significant upgrade from its predecessor, Claude Sonnet 3.7, with particular strength in coding and reasoning tasks. The model achieves state-of-the-art performance on SWE...

Vision Tools Streaming Reasoning Files
A

Arcee AI: Spotlight

Arcee AI

Vision
B

Baidu: ERNIE 4.5 21B A3B Thinking

Baidu

**Model ID:** `baidu/ernie-4.5-21b-a3b-thinking`

Vision
B

ERNIE 4.5 VL 424B A47B Model Details

Baidu

Vision
B

Black Forest Labs: FLUX.2 Max

Black Forest Labs

FLUX.2 [max] is the new top-tier image model from Black Forest Labs, pushing image quality, prompt understanding, and editing consistency to the highest level yet.

Vision
B

FLUX.1 [schnell] - Black Forest Labs

Black Forest Labs

FLUX.1 [schnell] is Black Forest Labs' fastest image generation model, a 12 billion parameter rectified flow transformer capable of generating high-quality images from text descriptions. Trained using...

Vision
C

Cohere Command R+ (08-2024)

Cohere

Vision
C

Cohere: Command R

Cohere

Command R is a large language model optimized for conversational interaction and long context tasks. It targets the "scalable" category of models that balance high performance with strong accuracy. Co...

Tools Streaming Vision
C

Cohere: Command R7B

Cohere

Command R7B is the smallest and fastest model in Cohere's R family of enterprise-focused large language models (LLMs). With 7-8 billion parameters, it is an open weights model with advanced capabiliti...

Tools Streaming Vision
C

Cohere: Embed 4

Cohere

Embed 4 is Cohere's most performant multilingual multimodal embedding model. It transforms different modalities such as images, texts, and interleaved images and texts into a single vector representat...

Vision
C

Cohere: Rerank 4 Fast

Cohere

Rerank 4 Fast is a AI model for general-purpose tasks. Developed by Cohere, this model is optimized for its specific use case category.

Streaming Vision
C

Cohere: Rerank 4 Pro

Cohere

Rerank 4 Pro is a AI model for general-purpose tasks. Developed by Cohere, this model is optimized for its specific use case category.

Streaming Vision
C

Collections: Flash 2.5

Collection

The **Flash 2.5 Collection** is a curated organization-level collection of fast-responding, lightweight language models optimized for speed and cost-effectiveness. This collection focuses on models th...

Vision
C

Collections: Organization Shared Models

Collection

The **Organization Shared Models** collection is a flexible, team-managed collection that enables organizations to pool and share language models across all members. This collection uses a least-used ...

Vision
D

DeepSeek: DeepSeek Coder V2

DeepSeek

DeepSeek Coder V2 is a code generation and understanding model for programming tasks. Developed by DeepSeek, this model provides solid performance and is suitable for most use cases.

Tools Streaming Vision
D

DeepSeek: DeepSeek V3.2

DeepSeek

**Model ID:** `deepseek/deepseek-v3.2`

Vision
G

Gemma 3 27B (free) - Complete Model Details

Google

Vision
G

Google AI: Gemini Embedding

Google

Gemini Embedding is a text embedding model for semantic search and vector-based tasks. Developed by Google AI, this model is optimized for its specific use case category.

Streaming Vision
G

Google Gemini 1.5 Flash

Google

Gemini 1.5 Flash is a foundation model that performs well at a variety of multimodal tasks such as visual understanding, classification, summarization, and creating content from image, audio and video...

Vision
G

Google Gemini 2.0 Flash Experimental (Free)

Google

Gemini Flash 2.0 offers a significantly faster time to first token (TTFT) compared to Gemini Flash 1.5, while maintaining quality on par with larger models like Gemini Pro 1.5. It introduces notable e...

Vision
G

Google Gemini 2.0 Flash Lite - Complete Model Details

Google

Vision
G

Google: Codey Code Completion

Google

Code completion model.

Vision
G

Google: Gecko Embedding

Google

Lightweight embedding model.

Vision
G

Google: Gemini 1.0 Pro Vision

Google

Gemini 1.0 with vision (deprecated).

Vision Tools
G

Google: Gemini 1.5 Pro

Google

Previous generation pro model.

Vision Tools
G

Google: Gemini 2 Flash Thinking

Google

Gemini 2 Flash with extended reasoning capabilities.

Vision Tools
G

Google: Gemini 2.0 Flash

Google

A cost-effective multimodal model optimized for general-purpose tasks requiring balanced performance. Gemini 2.0 Flash delivers strong capabilities across text, image, and code understanding while mai...

Vision Tools Streaming
G

Google: Gemini 2.0 Flash (Image Generation) Experimental

Google

An experimental version of Gemini 2.0 Flash with image generation capabilities. Combines text understanding with the ability to generate images based on prompts.

Streaming Vision
G

Google: Gemini 2.0 Flash 001

Google

The stable version 001 release of Gemini 2.0 Flash, providing a cost-effective multimodal model for general-purpose tasks. This versioned release ensures consistent behavior and reproducible results f...

Vision Tools Streaming
G

Google: Gemini 2.0 Flash Experimental

Google

An experimental preview of Gemini 2.0 Flash featuring the latest capabilities and improvements. This version provides early access to new features while maintaining the fast inference speeds character...

Vision Tools Streaming
G

Google: Gemini 2.0 Flash Preview Image Generation

Google

A preview version of Gemini 2.0 Flash optimized for image generation tasks. Designed for creating visual content from text descriptions.

Vision
G

Google: Gemini 2.0 Flash Thinking Experimental

Google

An experimental version of Gemini 2.0 Flash with enhanced reasoning capabilities. This model features a "thinking" mode that allows it to work through complex problems step-by-step before providing fi...

Vision Tools Streaming
G

Google: Gemini 2.0 Flash Thinking Preview 01-21

Google

A dated experimental version of Gemini 2.0 Flash with thinking capabilities from January 21, 2025. Features improved reasoning capabilities over earlier thinking model versions.

Vision Tools Streaming
G

Google: Gemini 2.0 Flash Thinking Preview 12-19

Google

A dated experimental version of Gemini 2.0 Flash with thinking capabilities from December 19, 2024. Provides enhanced reasoning through explicit step-by-step problem solving.

Vision Tools Streaming
G

Google: Gemini 2.0 Flash-Lite

Google

Streamlined and ultra-efficient model designed for simple, high-frequency tasks. Gemini 2.0 Flash-Lite prioritizes speed and affordability while maintaining essential multimodal capabilities.

Vision Tools Streaming
G

Google: Gemini 2.0 Flash-Lite Preview

Google

A preview version of Gemini 2.0 Flash-Lite, providing early access to streamlined capabilities optimized for high-frequency, simple tasks. This model supports multimodal capabilities including vision ...

Vision Tools
G

Google: Gemini 2.0 Flash-Lite Preview 02-05

Google

A dated preview version of Gemini 2.0 Flash-Lite from February 5, 2025. Provides a fixed checkpoint for reproducible results. This model supports multimodal capabilities including vision and image und...

Vision Tools
G

Google: Gemini 2.0 Pro

Google

Professional-grade Gemini 2.0 model.

Vision Tools
G

Google: Gemini 2.0 Pro Experimental

Google

An experimental version of Gemini 2.0 Pro offering higher capability than Flash variants. Designed for complex tasks requiring advanced reasoning, coding, and multimodal understanding.

Vision Tools Streaming Reasoning
G

Google: Gemini 2.0 Pro Experimental 02-05

Google

A dated experimental version of Gemini 2.0 Pro from February 5, 2025. Provides high-capability performance for complex tasks with a specific model checkpoint.

Vision Tools Streaming Reasoning
G

Google: Gemini 2.0 Pro Vision

Google

Vision-optimized Gemini 2.0 Pro.

Vision Tools
G

Google: Gemini 2.5 Computer Use Preview 10-2025

Google

A specialized preview model designed for computer use and automation tasks. Enables AI-driven interaction with computer interfaces, including clicking, typing, and navigating applications.

Vision
G

Google: Gemini 2.5 Flash

Google

Lightning-fast and highly capable thinking model that delivers a balance of intelligence and latency. Gemini 2.5 Flash builds upon Gemini 2.0 Flash with upgraded reasoning, hybrid thinking control, an...

Vision Tools Streaming
G

Google: Gemini 2.5 Flash Image (Nano Banana)

Google

A specialized image-focused variant of Gemini 2.5 Flash, codenamed Nano Banana. Optimized for image understanding and generation tasks with fast inference.

Vision
G

Google: Gemini 2.5 Flash Image Preview (Nano Banana)

Google

A preview version of the image-focused Gemini 2.5 Flash variant, codenamed Nano Banana. Provides early access to enhanced image capabilities.

Vision
G

Google: Gemini 2.5 Flash Preview

Google

Preview of Gemini 2.5 Flash.

Vision Tools
G

Google: Gemini 2.5 Flash Preview 05-20

Google

A dated preview of Gemini 2.5 Flash from May 20, 2025. Provides a fixed model checkpoint for reproducible experiments and applications.

Vision Tools Streaming
G

Google: Gemini 2.5 Flash Preview Sep 2025

Google

A September 2025 preview of Gemini 2.5 Flash with the latest capabilities and improvements before stable release. This model supports multimodal capabilities including vision and image understanding.

Vision Tools Streaming
G

Google: Gemini 2.5 Flash-Lite

Google

Built for massive scale, Gemini 2.5 Flash-Lite balances cost and performance for high-throughput tasks. Optimized for efficiency without sacrificing multimodal capabilities.

Vision Tools Streaming
G

Google: Gemini 2.5 Flash-Lite Preview 06-17

Google

A dated preview version of Gemini 2.5 Flash-Lite from June 17, 2025. Optimized for efficiency and high-throughput tasks with a fixed checkpoint.

Vision
G

Google: Gemini 2.5 Flash-Lite Preview Sep 2025

Google

A September 2025 preview of Gemini 2.5 Flash-Lite, optimized for efficiency and cost-effectiveness in high-throughput applications. This model supports multimodal capabilities including vision and ima...

Vision
G

Google: Gemini 2.5 Pro

Google

Vision
G

Google: Gemini 2.5 Pro

Google

Google's high-capability model for complex reasoning and coding. Features adaptive thinking and a 1 million token context window, designed for complex agentic and multimodal challenges. Gemini 2.5 Pro...

Vision Tools Streaming Reasoning
G

Google: Gemini 2.5 Pro Preview 03-25

Google

A dated preview version of Gemini 2.5 Pro from March 25, 2025. Provides access to advanced capabilities with a fixed model checkpoint for reproducibility.

Vision Tools Streaming Reasoning
G

Google: Gemini 2.5 Pro Preview 06-05

Google

Gemini 2.5 Pro is Google's latest and most capable model, featuring a massive 1 million token context window. This preview version (06-05) represents the cutting edge of Google's multimodal AI capabil...

Vision
G

Google: Gemini 3 Flash Preview

Google

Fast Gemini 3 for speed and efficiency.

Vision Tools
G

Google: Gemini 3 Mobile

Google

Lightweight Gemini 3 optimized for mobile devices.

Vision Tools
G

Google: Gemini 3 Opus

Google

High-end model for demanding applications.

Vision Tools
G

Google: Gemini 3 Pro Preview

Google

Latest flagship Gemini model with advanced reasoning and multimodal capabilities.

Vision Tools
G

Google: Gemini 3.5 Sonnet

Google

Balanced mid-tier Gemini model.

Vision Tools
G

Google: Gemini Audio Understanding

Google

Audio analysis model.

Vision
G

Google: Gemini Code Reasoning

Google

Advanced code analysis and generation model.

Tools Vision
G

Google: Gemini Document Understanding

Google

Specialized model for document processing and extraction.

Vision
G

Google: Gemini Experimental 1206

Google

An experimental Gemini model from December 6, 2024. Provides early access to new capabilities and improvements in development. This model supports multimodal capabilities including vision and image un...

Vision Tools Streaming
G

Google: Gemini Flash-Lite Latest

Google

The latest stable version of Gemini Flash-Lite, automatically updated to the most recent stable release. Optimized for efficiency and high-throughput tasks.

Vision
G

Google: Gemini Image Generation 001

Google

Image generation from text.

Vision
G

Google: Gemini Multimodal Live

Google

Real-time streaming multimodal model.

Vision Tools
G

Google: Gemini Nano

Google

Smallest Gemini model.

Vision
G

Google: Gemini Pro Latest

Google

The latest stable version of Gemini Pro, automatically updated to the most recent stable release. Provides high-capability performance for complex tasks.

Vision Tools Reasoning
G

Google: Gemini Reasoning Engine

Google

Experimental advanced reasoning engine for complex tasks.

Vision Tools
G

Google: Gemini Robotics-ER 1.5 Preview

Google

A specialized model for robotics and embodied reasoning (ER) applications. Designed to understand and reason about physical environments, robot actions, and spatial relationships.

Vision
G

Google: Gemini Text Embedding 004

Google

Latest text embedding model.

Vision
G

Google: Gemini Video Understanding

Google

Video analysis and understanding.

Vision
G

Google: Gemma 1.1 7B Instruct-Tuned

Google

Previous Gemma generation 7B model.

Vision
G

Google: Gemma 2 27B Instruct-Tuned

Google

Previous generation 27B model.

Vision
G

Google: Gemma 2 9B Instruct-Tuned

Google

Vision
G

Google: Gemma 3 12B Instruct-Tuned

Google

Compact 12B instruction-tuned model.

Vision
G

Google: Gemma 3 27B Instruct-Tuned

Google

Open-source 27B instruction-tuned model.

Vision
G

Google: Gemma 3 4B Instruct-Tuned

Google

Ultra-lightweight 4B model.

Vision
G

Google: Gemma 3 Long Context 27B

Google

Extended context Gemma 3 27B.

Vision
G

Google: Gemma 3 Long Context 27B

Google

Extended context version of Gemma 3 27B supporting up to 1M token context.

Vision
G

Google: LearnLM 2.0 Flash Experimental

Google

An experimental model designed specifically for educational applications. LearnLM is optimized for tutoring, explanation generation, and adaptive learning interactions. This model supports multimodal ...

Vision
G

Google: Multimodal Understanding Pro

Google

Advanced multimodal model.

Vision Tools
G

Google: Nano Banana Pro

Google

A high-capability image-focused model codenamed Nano Banana Pro. Designed for advanced image understanding and generation with professional-grade quality. This model supports multimodal capabilities i...

Vision
G

Google: Nano Banana Pro (Gemini 3 Pro Image Preview)

Google

A high-capability image-focused model codenamed Nano Banana Pro. Part of the Gemini 3 Pro family with specialized capabilities for advanced image understanding and generation.

Vision
G

Google: PaLM 2 Chat Bison

Google

Legacy PaLM chat model (deprecated).

Vision
G

Google: PaLM 2 Text Bison

Google

Legacy PaLM text model (deprecated).

Vision
G

Groq: Claude 3.5 Sonnet

Groq

Streaming Vision
G

Groq: Command Nightly

Groq

Groq: Command Nightly is a language model provided by the provider. This model offers advanced capabilities for natural language processing tasks.

Streaming Vision
G

Groq: DeepSeek R1 Distill Llama 70B

Groq

Groq: DeepSeek R1 Distill Llama 70B is a language model provided by the provider. This model offers advanced capabilities for natural language processing tasks.

Streaming Vision
G

Groq: Distil Whisper Large V3 EN

Groq

Groq: Distil Whisper Large V3 EN is a language model provided by the provider. This model offers advanced capabilities for natural language processing tasks.

Streaming Vision
G

Groq: Gemma 2 9B IT

Groq

Groq: Gemma 2 9B IT is a language model provided by the provider. This model offers advanced capabilities for natural language processing tasks.

Streaming Vision
G

Groq: Gemma 7B IT

Groq

Groq: Gemma 7B IT is a language model provided by the provider. This model offers advanced capabilities for natural language processing tasks.

Streaming Vision
G

Groq: GPT OSS 120B 128k

Groq

GPT OSS 120B 128k is a conversational AI model designed for multi-turn dialogue and interactive tasks. Developed by Groq, this model is optimized for its specific use case category.

Streaming Vision
G

Groq: GPT OSS 20B 128k

Groq

GPT OSS 20B 128k is a conversational AI model designed for multi-turn dialogue and interactive tasks. Developed by Groq, this model is optimized for its specific use case category.

Streaming Vision
G

Groq: GPT OSS Safeguard 20B

Groq

GPT OSS Safeguard 20B is a content moderation model for safety and policy compliance checking. Developed by Groq, this model is optimized for its specific use case category.

Streaming Vision
G

Groq: GPT-4 Turbo

Groq

Groq: GPT-4 Turbo is a language model provided by the provider. This model offers advanced capabilities for natural language processing tasks.

Streaming Vision
G

Groq: GPT-4 Vision Preview

Groq

Groq: GPT-4 Vision Preview is a language model provided by the provider. This model offers advanced capabilities for natural language processing tasks.

Streaming Vision
G

Groq: GPT-4o Mini

Groq

Groq: GPT-4o Mini is a language model provided by the provider. This model offers advanced capabilities for natural language processing tasks.

Streaming Vision
G

Groq: Kimi K2 1T 256k

Groq

Kimi K2 1T 256k is a conversational AI model designed for multi-turn dialogue and interactive tasks. Developed by Groq, this model is optimized for its specific use case category.

Streaming Vision
G

Groq: Llama 3.1 8B Instant

Groq

Streaming Vision
G

Groq: Llama 3.3 70B Speculative Decoding

Groq

Streaming Vision
G

Groq: Llama 3.3 70B Versatile

Groq

Streaming Vision
G

Groq: Llama 4 Maverick

Groq

Llama 4 Maverick is a conversational AI model designed for multi-turn dialogue and interactive tasks. Developed by Groq, this flagship model represents the latest capabilities and state-of-the-art per...

Tools Streaming Vision
G

Groq: Llama 4 Scout

Groq

Llama 4 Scout is a conversational AI model designed for multi-turn dialogue and interactive tasks. Developed by Groq, this model provides solid performance and is suitable for most use cases.

Tools Streaming Vision
G

Groq: Llama Guard 4 12B

Groq

Llama Guard 4 12B is a content moderation model for safety and policy compliance checking. Developed by Groq, this model is optimized for its specific use case category.

Streaming Vision
G

Groq: Mixtral 8x7B (Extended)

Groq

Groq: Mixtral 8x7B (Extended) is a language model provided by the provider. This model offers advanced capabilities for natural language processing tasks.

Streaming Vision
G

Groq: Mixtral 8x7B 32K

Groq

Groq: Mixtral 8x7B 32K is a language model provided by the provider. This model offers advanced capabilities for natural language processing tasks.

Streaming Vision
G

Groq: Neural Chat 7B v3.1

Groq

Streaming Vision
G

Groq: OpenChat 3.5

Groq

Streaming Vision
G

Groq: PaLM 2 Chat Bison 32K

Groq

Groq: PaLM 2 Chat Bison 32K is a language model provided by the provider. This model offers advanced capabilities for natural language processing tasks.

Streaming Vision
G

Groq: PaLM 2 CodeChat Bison

Groq

Groq: PaLM 2 CodeChat Bison is a language model provided by the provider. This model offers advanced capabilities for natural language processing tasks.

Streaming Vision
G

Groq: Qwen 2 72B 4-bit

Groq

Groq: Qwen 2 72B 4-bit is a language model provided by the provider. This model offers advanced capabilities for natural language processing tasks.

Streaming Vision
G

Groq: Qwen 2 7B 4-bit

Groq

Groq: Qwen 2 7B 4-bit is a language model provided by the provider. This model offers advanced capabilities for natural language processing tasks.

Streaming Vision
G

Groq: Qwen3 32B

Groq

Qwen3 32B is a conversational AI model designed for multi-turn dialogue and interactive tasks. Developed by Groq, this model provides solid performance and is suitable for most use cases.

Tools Streaming Vision
G

Groq: Solar 10.7B Instruct v1

Groq

Streaming Vision
G

Groq: Whisper Large V3 Turbo

Groq

Groq: Whisper Large V3 Turbo is a language model provided by the provider. This model offers advanced capabilities for natural language processing tasks.

Streaming Vision
G

Groq: Whisper V3 Large

Groq

Whisper V3 Large is a audio processing model for speech synthesis and audio understanding. Developed by Groq, this model is optimized for its specific use case category.

Streaming Vision
G

Groq: Yi 34B Chat 3 32K

Groq

Groq: Yi 34B Chat 3 32K is a language model provided by the provider. This model offers advanced capabilities for natural language processing tasks.

Streaming Vision
G

MythoMax 13B

Gryphe

One of the highest performing and most popular fine-tunes of Llama 2 13B, with rich descriptions and roleplay. This is a merged model (#merge) that combines multiple fine-tuning approaches to achieve ...

Vision
I

Inflection 3 Pi

Inflection

Inflection 3 Pi powers Inflection's Pi chatbot, including backstory, emotional intelligence, productivity, and safety. It has access to recent news, and excels in scenarios like customer support and r...

Vision
K

Psyfighter v2 13B

Koboldai

A specialized merged model designed for enhanced fictional storytelling with supplementary medical knowledge. The model combines three base models to balance creative narrative generation with anatomi...

Vision
M

Meta Llama 3.1 405B Instruct

Meta

Llama 3.1 405B Instruct is Meta's largest and most capable open-source language model, representing their flagship offering in the Llama 3.1 series. This 405-billion parameter model features a 128K to...

Vision
M

Meta Llama 3.1 8B Instruct

Meta

Llama 3.1 8B Instruct is part of Meta's latest class of language models, offering a balance between efficiency and capability. This 8-billion parameter instruction-tuned variant emphasizes speed and e...

Vision
M

Meta Llama 3.3 70B Instruct

Meta

Llama 3.3 70B Instruct is a pretrained and instruction-tuned generative model optimized for multilingual dialogue use cases. It outperforms many available open source and closed chat models on common ...

Vision
M

Meta: Llama 2 70B Chat

Meta

Vision
M

Meta: Llama 3.2 1B Instruct

Meta

Llama 3.2 1B is a 1-billion-parameter language model optimized for efficiently performing natural language tasks such as summarization, dialogue, and multilingual text analysis. With its smaller size,...

Vision
M

Meta: Llama 3.2 3B Instruct

Meta

Llama 3.2 3B is a 3-billion-parameter multilingual large language model optimized for advanced natural language processing tasks including dialogue generation, complex reasoning, and text summarizatio...

Vision
M

Meta: Llama 3.2 90B Vision Instruct

Meta

A 90-billion-parameter multimodal model excelling at visual reasoning and language tasks. The model handles image captioning, visual question answering, and advanced image-text comprehension through p...

Vision
M

Meta: Llama 4 Maverick

Meta

A high-capacity multimodal language model from Meta, built on a mixture-of-experts (MoE) architecture with 128 experts and 17 billion active parameters per forward pass (400B total parameters). The mo...

Vision Tools
M

Meta: Llama 4 Scout

Meta

Vision
M

MiniMax M2.1

Minimax

A lightweight, state-of-the-art language model with 10 billion activated parameters, optimized for coding, agentic workflows, and application development. The model delivers cleaner, more concise outp...

Vision
M

Mistral 7B Instruct

Mistral AI

A high-performing, industry-standard 7.3B parameter model, with optimizations for speed and context length. Mistral 7B Instruct has multiple version variants, and this endpoint is intended to be the l...

Vision
M

Mistral AI Codestral

Mistral AI

Codestral is Mistral AI's cutting-edge language model explicitly designed for code generation tasks. It is Mistral's inaugural code-specific generative model, representing an open-weight generative AI...

Vision
M

Mistral AI: Codestral

Mistral AI

Codestral is a code generation and understanding model for programming tasks. Developed by Mistral AI, this model is optimized for its specific use case category.

Streaming Vision
M

Mistral AI: Codestral Embed

Mistral AI

Codestral Embed is a text embedding model for semantic search and vector-based tasks. Developed by Mistral AI, this model is optimized for its specific use case category.

Streaming Vision
M

Mistral AI: Devstral 2

Mistral AI

Devstral 2 is a code generation and understanding model for programming tasks. Developed by Mistral AI, this model is optimized for its specific use case category.

Streaming Vision
M

Mistral AI: Devstral Small 2

Mistral AI

Devstral Small 2 is a code generation and understanding model for programming tasks. Developed by Mistral AI, this model is optimized for its specific use case category.

Streaming Vision
M

Mistral AI: Magistral Medium

Mistral AI

Magistral Medium is a conversational AI model designed for multi-turn dialogue and interactive tasks. Developed by Mistral AI, this model is optimized for its specific use case category.

Streaming Vision
M

Mistral AI: Magistral Small

Mistral AI

Magistral Small is a conversational AI model designed for multi-turn dialogue and interactive tasks. Developed by Mistral AI, this model is optimized for its specific use case category.

Streaming Vision
M

Mistral AI: Ministral 14B

Mistral AI

Ministral 14B is a conversational AI model designed for multi-turn dialogue and interactive tasks. Developed by Mistral AI, this model is optimized for its specific use case category.

Streaming Vision
M

Mistral AI: Ministral 3B

Mistral AI

Ministral 3B is a conversational AI model designed for multi-turn dialogue and interactive tasks. Developed by Mistral AI, this model is optimized for its specific use case category.

Streaming Vision
M

Mistral AI: Ministral 8B

Mistral AI

Ministral 8B is a conversational AI model designed for multi-turn dialogue and interactive tasks. Developed by Mistral AI, this model is optimized for its specific use case category.

Streaming Vision
M

Mistral AI: Mistral Embed

Mistral AI

Mistral Embed is a text embedding model for semantic search and vector-based tasks. Developed by Mistral AI, this model is optimized for its specific use case category.

Streaming Vision
M

Mistral AI: Mistral Large 3

Mistral AI

Mistral Large 3 is a conversational AI model designed for multi-turn dialogue and interactive tasks. Developed by Mistral AI, this flagship model represents the latest capabilities and state-of-the-ar...

Tools Streaming Vision
M

Mistral AI: Mistral Medium 3

Mistral AI

Mistral Medium 3 is a conversational AI model designed for multi-turn dialogue and interactive tasks. Developed by Mistral AI, this premium model offers excellent quality and balanced performance acro...

Tools Streaming Vision
M

Mistral AI: Mistral Moderation

Mistral AI

Mistral Moderation is a content moderation model for safety and policy compliance checking. Developed by Mistral AI, this model is optimized for its specific use case category.

Streaming Vision
M

Mistral AI: Mistral OCR

Mistral AI

Mistral OCR is a AI model for general-purpose tasks. Developed by Mistral AI, this model is optimized for its specific use case category.

Streaming Vision
M

Mistral AI: Mistral Small 3.2

Mistral AI

Mistral Small 3.2 is a conversational AI model designed for multi-turn dialogue and interactive tasks. Developed by Mistral AI, this model provides solid performance and is suitable for most use cases...

Tools Streaming Vision
M

Mistral AI: Mistral Small Creative

Mistral AI

Mistral Small Creative is a conversational AI model designed for multi-turn dialogue and interactive tasks. Developed by Mistral AI, this model is optimized for its specific use case category.

Streaming Vision
M

Mistral AI: Voxtral Mini

Mistral AI

Voxtral Mini is a audio processing model for speech synthesis and audio understanding. Developed by Mistral AI, this model is optimized for its specific use case category.

Streaming Vision
M

Mistral AI: Voxtral Small

Mistral AI

Voxtral Small is a audio processing model for speech synthesis and audio understanding. Developed by Mistral AI, this model is optimized for its specific use case category.

Streaming Vision
M

Mistral Large

Mistral AI

Mistral Large is Mistral AI's flagship offering. The model excels at reasoning, code generation, JSON handling, and chat applications. It is a proprietary model with support for dozens of languages in...

Vision
M

Mistral Medium Model Documentation

Mistral AI

A closed-source, medium-sized model from Mistral AI that excels at reasoning, code, JSON, chat, and more. This model performs comparably to other companies' flagship models and represents Mistral's mi...

Vision
M

Mistral: Devstral 2 2512 (Free)

Mistral AI

Devstral 2 is a state-of-the-art open-source model by Mistral AI specializing in agentic coding. It is a 123B-parameter dense transformer model supporting a 256K context window.

Vision
M

Mistral: Mixtral 8x7B Instruct

Mistral AI

Mixtral 8x7B Instruct is a pretrained generative Sparse Mixture of Experts model with 8 experts totaling 47 billion parameters. It has been fine-tuned by Mistral AI specifically for chat and instructi...

Vision
M

Mistral: Pixtral 12B

Mistral AI

The first multi-modal, text+image-to-text model from Mistral AI. Its weights were launched via torrent, making it openly available for research and commercial use.

Vision
N

Nous: Hermes 2 Vision 7B (Alpha)

Nous Research

Nous: Hermes 2 Vision 7B is an alpha-stage vision-language model that extends the capabilities of OpenHermes-2.5 by incorporating visual perception abilities. The model was developed using a specializ...

Vision
O

Ollama: Llama 3.1 8B Instruct

Ollama

Llama 3.1 8B Instruct is Meta's state-of-the-art instruction-tuned language model with 8 billion parameters. It's a compact yet powerful model designed for general-purpose conversational AI, reasoning...

Vision
O

Ollama: Qwen2.5 7B Instruct

Ollama

Qwen2.5 7B Instruct is Alibaba's latest-generation instruction-tuned language model with 7.6 billion parameters, representing a significant upgrade to the Qwen family. Built on 18 trillion tokens of d...

Vision
O

OpenAI GPT-4 32K Model Documentation

OpenAI

**Source**: OpenRouter (https://langmart.ai/model-docs)

Vision
O

OpenAI GPT-4 Vision Model Specifications

OpenAI

**Last Updated:** December 24, 2025

Vision
O

OpenAI o1-preview

OpenAI

OpenAI o1-preview is a reasoning-focused model designed to "spend more time thinking before responding." It employs chain-of-thought reasoning with self-fact-checking capabilities, making it particula...

Vision
O

OpenAI: Computer Use Preview

OpenAI

Computer Use Preview is a conversational AI model designed for multi-turn dialogue and interactive tasks. Developed by OpenAI, this model is optimized for its specific use case category.

Streaming Vision
O

OpenAI: DALL-E 2

OpenAI

DALL-E 2 is a image generation model for creating visual content from descriptions. Developed by OpenAI, this model is optimized for its specific use case category.

Streaming Vision
O

OpenAI: DALL-E 3

OpenAI

DALL-E 3 is a image generation model for creating visual content from descriptions. Developed by OpenAI, this model is optimized for its specific use case category.

Streaming Vision
O

OpenAI: GPT Image 1

OpenAI

OpenAI's first-generation image generation model integrated with GPT capabilities. Enables text-to-image generation with natural language understanding. This model supports multimodal capabilities inc...

Vision
O

OpenAI: GPT Image 1 Mini

OpenAI

A lightweight version of OpenAI's GPT Image 1, optimized for faster generation and lower cost while maintaining good quality. This model supports multimodal capabilities including vision and image und...

Vision
O

OpenAI: GPT Image 1.5

OpenAI

OpenAI's enhanced image generation model with improved quality, better prompt understanding, and more detailed outputs compared to GPT Image 1.

Vision
O

OpenAI: GPT-4.1

OpenAI

Enhanced version of GPT-4 with improved reasoning and multimodal capabilities. This model supports multimodal capabilities including vision and image understanding. It features advanced reasoning capa...

Vision Tools Streaming Reasoning
O

OpenAI: GPT-4.1 (April 2025)

OpenAI

Latest GPT-4.1 variant with current knowledge cutoff. This model supports multimodal capabilities including vision and image understanding. It features advanced reasoning capabilities for complex prob...

Vision Tools Streaming Reasoning
O

OpenAI: GPT-4o (August 2024)

OpenAI

August 2024 update with improved performance. This model supports multimodal capabilities including vision and image understanding. It features advanced reasoning capabilities for complex problem-solv...

Vision Tools Streaming Reasoning
O

OpenAI: GPT-4o (May 2024)

OpenAI

Initial release of GPT-4o optimized model. This model supports multimodal capabilities including vision and image understanding. It features advanced reasoning capabilities for complex problem-solving...

Vision Tools Streaming Reasoning
O

OpenAI: GPT-4o (November 2024)

OpenAI

Latest GPT-4o with updated knowledge and improved capabilities. This model supports multimodal capabilities including vision and image understanding. It features advanced reasoning capabilities for co...

Vision Tools Streaming Reasoning
O

OpenAI: GPT-4o Audio Preview

OpenAI

Preview of GPT-4o with audio processing capabilities. This model supports multimodal capabilities including vision and image understanding. It features advanced reasoning capabilities for complex prob...

Vision Tools Streaming Reasoning
O

OpenAI: GPT-4o Audio Preview (October 2024)

OpenAI

October release of audio-enabled GPT-4o preview. This model supports multimodal capabilities including vision and image understanding. It features advanced reasoning capabilities for complex problem-s...

Vision Tools Streaming Reasoning
O

OpenAI: GPT-4o Mini

OpenAI

Compact multimodal model for efficient applications. This model supports multimodal capabilities including vision and image understanding. It features advanced reasoning capabilities for complex probl...

Vision Tools Streaming
O

OpenAI: GPT-4o Mini (July 2024)

OpenAI

Initial release of the mini multimodal variant. This model supports multimodal capabilities including vision and image understanding. It features advanced reasoning capabilities for complex problem-so...

Vision Tools Streaming
O

OpenAI: GPT-4o Mini TTS

OpenAI

GPT-4o Mini TTS is a audio processing model for speech synthesis and audio understanding. Developed by OpenAI, this model is optimized for its specific use case category.

Streaming Vision
O

OpenAI: GPT-5.2 Chat (AKA Instant)

OpenAI

GPT-5.2 Chat is the fast, lightweight member of the 5.2 family, optimized for low-latency chat while retaining strong general intelligence. The model uses adaptive reasoning to selectively engage deep...

Vision
O

OpenAI: O3 Deep Research

OpenAI

O3 Deep Research is a conversational AI model designed for multi-turn dialogue and interactive tasks. Developed by OpenAI, this model is optimized for its specific use case category.

Streaming Vision
O

OpenAI: O3 Pro

OpenAI

O3 Pro is a conversational AI model designed for multi-turn dialogue and interactive tasks. Developed by OpenAI, this model is optimized for its specific use case category.

Streaming Vision
O

OpenAI: Omni Moderation

OpenAI

Omni Moderation is a content moderation model for safety and policy compliance checking. Developed by OpenAI, this model is optimized for its specific use case category.

Streaming Vision
O

OpenAI: TTS

OpenAI

TTS is a audio processing model for speech synthesis and audio understanding. Developed by OpenAI, this model is optimized for its specific use case category.

Streaming Vision
O

OpenAI: TTS HD

OpenAI

TTS HD is a audio processing model for speech synthesis and audio understanding. Developed by OpenAI, this model is optimized for its specific use case category.

Streaming Vision
O

OpenAI: Whisper

OpenAI

Whisper is a audio processing model for speech synthesis and audio understanding. Developed by OpenAI, this model is optimized for its specific use case category.

Streaming Vision
O

Openai-compatible: Fake Gpt 4 Vision

Openai Compatible

Fake Gpt 4 Vision with vision capabilities for processing images and visual content. This model supports multimodal capabilities including vision and image understanding. It features advanced reasonin...

Vision Streaming
O

OpenChat 3.5 7B (Free)

Openchat

OpenChat is a library of open-source language models fine-tuned with C-RLFT, a strategy inspired by offline reinforcement learning. The model is trained on mixed-quality data without preference labels...

Reasoning Vision
O

LangMart: Amazon: Nova 2 Lite

Openrouter

Nova 2 Lite is a fast, cost-effective reasoning model for everyday workloads that can process text, images, and videos to generate text.

Vision
O

LangMart: Amazon: Nova Lite 1.0

Openrouter

Amazon Nova Lite 1.0 is a very low-cost multimodal model from Amazon that focused on fast processing of image, video, and text inputs to generate text output. Amazon Nova Lite can handle real-time cus...

Vision
O

LangMart: Amazon: Nova Premier 1.0

Openrouter

Amazon Nova Premier is the most capable of Amazonโ€™s multimodal models for complex reasoning tasks and for use as the best teacher for distilling custom models.

Vision
O

LangMart: Amazon: Nova Pro 1.0

Openrouter

Amazon Nova Pro 1.0 is a capable multimodal model from Amazon focused on providing a combination of accuracy, speed, and cost for a wide range of tasks. As of December 2024, it achieves state-of-the-a...

Vision
O

LangMart: Anthropic Claude 3.5 Sonnet

Openrouter

Anthropic's Claude 3.5 Sonnet accessed via LangMart. A balanced model combining strong intelligence with fast response times, ideal for most use cases.

Vision Tools Streaming
O

LangMart: Anthropic Claude 3.7 Sonnet

Openrouter

Anthropic's Claude 3.7 Sonnet accessed via LangMart. An enhanced version with improved reasoning and capabilities over Claude 3.5 Sonnet. This model supports multimodal capabilities including vision a...

Vision Tools
O

LangMart: Anthropic Claude Haiku 4.5

Openrouter

Anthropic's fastest and most efficient Claude model accessed via LangMart. Designed for high-volume, low-latency applications requiring quick responses. This model supports multimodal capabilities inc...

Vision Streaming
O

LangMart: Anthropic: Claude 3 Haiku

Openrouter

Claude 3 Haiku is Anthropic's fastest and most compact model for

Vision
O

LangMart: Anthropic: Claude 3 Opus

Openrouter

Claude 3 Opus is Anthropic's most powerful model for highly complex tasks. It boasts top-level performance, intelligence, fluency, and understanding.

Vision
O

LangMart: Anthropic: Claude Opus 4

Openrouter

Claude Opus 4 is benchmarked as the worldโ€™s best coding model, at time of release, bringing sustained performance on complex, long-running tasks and agent workflows. It sets new benchmarks in software...

Vision
O

LangMart: Anthropic: Claude Sonnet 4

Openrouter

Claude Sonnet 4 significantly enhances the capabilities of its predecessor, Sonnet 3.7, excelling in both coding and reasoning tasks with improved precision and controllability. Achieving state-of-the...

Vision
O

LangMart: Arcee AI: Spotlight

Openrouter

Spotlight is a 7โ€‘billionโ€‘parameter visionโ€‘language model derived from Qwenโ€ฏ2.5โ€‘VL and fineโ€‘tuned by Arcee AI for tight imageโ€‘text grounding tasks. It offers a 32โ€ฏkโ€‘token context window, enabling rich ...

Vision
O

LangMart: Cogito V2 Preview Llama 109B

Openrouter

An instruction-tuned, hybrid-reasoning Mixture-of-Experts model built on Llama-4-Scout-17B-16E. Cogito v2 can answer directly or engage an extended โ€œthinkingโ€ phase, with alignment guided by Iterated ...

Vision
O

LangMart: EleutherAI: Llemma 7b

Openrouter

EleutherAI: Llemma 7b is a conversational AI model designed for multi-turn dialogue and interactive tasks. Developed by LangMart, this model is optimized for its specific use case category.

Streaming Vision
O

LangMart: Google: Gemini 2.0 Flash

Openrouter

Google: Gemini 2.0 Flash is a image generation model for creating visual content from descriptions. Developed by LangMart, this model is optimized for its specific use case category.

Streaming Vision
O

LangMart: Google: Gemini 2.5 Flash

Openrouter

Google: Gemini 2.5 Flash is a image generation model for creating visual content from descriptions. Developed by LangMart, this model is optimized for its specific use case category.

Streaming Vision
O

LangMart: Google: Gemini 2.5 Flash Image (Nano Banana)

Openrouter

Google: Gemini 2.5 Flash Image (Nano Banana) is a image generation model for creating visual content from descriptions. Developed by LangMart, this model is optimized for its specific use case categor...

Streaming Vision
O

LangMart: Google: Gemini 2.5 Flash Image Preview (Nano Banana)

Openrouter

Google: Gemini 2.5 Flash Image Preview (Nano Banana) is a image generation model for creating visual content from descriptions. Developed by LangMart, this model is optimized for its specific use case...

Streaming Vision
O

LangMart: Google: Gemini 2.5 Flash Lite

Openrouter

Google: Gemini 2.5 Flash Lite is a image generation model for creating visual content from descriptions. Developed by LangMart, this model is optimized for its specific use case category.

Streaming Vision
O

LangMart: Google: Gemini 2.5 Flash Lite Preview 09-2025

Openrouter

Google: Gemini 2.5 Flash Lite Preview 09-2025 is a image generation model for creating visual content from descriptions. Developed by LangMart, this model is optimized for its specific use case catego...

Streaming Vision
O

LangMart: Google: Gemini 2.5 Flash Preview 09-2025

Openrouter

Google: Gemini 2.5 Flash Preview 09-2025 is a image generation model for creating visual content from descriptions. Developed by LangMart, this model is optimized for its specific use case category.

Streaming Vision
O

LangMart: Google: Gemini 2.5 Pro

Openrouter

Google: Gemini 2.5 Pro is a image generation model for creating visual content from descriptions. Developed by LangMart, this model is optimized for its specific use case category.

Streaming Vision
O

LangMart: Google: Gemini 3 Flash Preview

Openrouter

Google: Gemini 3 Flash Preview is a image generation model for creating visual content from descriptions. Developed by LangMart, this model is optimized for its specific use case category.

Streaming Vision
O

LangMart: Google: Gemini 3 Pro Preview

Openrouter

Google: Gemini 3 Pro Preview is a image generation model for creating visual content from descriptions. Developed by LangMart, this model is optimized for its specific use case category.

Streaming Vision
O

LangMart: Google: Gemma 2 27B

Openrouter

Google: Gemma 2 27B is a conversational AI model designed for multi-turn dialogue and interactive tasks. Developed by LangMart, this model is optimized for its specific use case category.

Streaming Vision
O

LangMart: Google: Gemma 2 9B

Openrouter

Google: Gemma 2 9B is a conversational AI model designed for multi-turn dialogue and interactive tasks. Developed by LangMart, this model is optimized for its specific use case category.

Streaming Vision
O

LangMart: Google: Gemma 3 12B

Openrouter

Google: Gemma 3 12B is a image generation model for creating visual content from descriptions. Developed by LangMart, this model is optimized for its specific use case category.

Streaming Vision
O

LangMart: Google: Gemma 3 27B

Openrouter

Google: Gemma 3 27B is a image generation model for creating visual content from descriptions. Developed by LangMart, this model is optimized for its specific use case category.

Streaming Vision
O

LangMart: Google: Gemma 3 4B

Openrouter

Google: Gemma 3 4B is a image generation model for creating visual content from descriptions. Developed by LangMart, this model is optimized for its specific use case category.

Streaming Vision
O

LangMart: Google: Gemma 3n 4B

Openrouter

Google: Gemma 3n 4B is a conversational AI model designed for multi-turn dialogue and interactive tasks. Developed by LangMart, this model is optimized for its specific use case category.

Streaming Vision
O

LangMart: Google: Nano Banana Pro (Gemini 3 Pro Image Preview)

Openrouter

Google: Nano Banana Pro (Gemini 3 Pro Image Preview) is a image generation model for creating visual content from descriptions. Developed by LangMart, this model is optimized for its specific use case...

Streaming Vision
O

LangMart: Meta Llama/llama 3.2 11b Vision Instruct

Openrouter

Meta Llama/llama 3.2 11b Vision Instruct with vision capabilities for processing images and visual content. This model supports multimodal capabilities including vision and image understanding. It fea...

Vision Streaming
O

LangMart: Meta Llama/llama 3.2 90b Vision Instruct

Openrouter

Meta Llama/llama 3.2 90b Vision Instruct with vision capabilities for processing images and visual content. This model supports multimodal capabilities including vision and image understanding. It fea...

Vision Streaming
O

LangMart: Meta: Llama 4 Maverick

Openrouter

Llama 4 Maverick 17B Instruct (128E) is a high-capacity multimodal language model from Meta, built on a mixture-of-experts (MoE) architecture with 128 experts and 17 billion active parameters per forw...

Vision
O

LangMart: Meta: Llama 4 Scout

Openrouter

Llama 4 Scout 17B Instruct (16E) is a mixture-of-experts (MoE) language model developed by Meta, activating 17 billion parameters out of a total of 109B. It supports native multimodal input (text and ...

Vision
O

LangMart: Meta: Llama Guard 4 12B

Openrouter

Llama Guard 4 is a Llama 4 Scout-derived multimodal pretrained model, fine-tuned for content safety classification. Similar to previous versions, it can be used to classify content in both LLM inputs ...

Vision
O

LangMart: Microsoft: Phi 4 Multimodal Instruct

Openrouter

Phi-4 Multimodal Instruct is a versatile 5.6B parameter foundation model that combines advanced reasoning and instruction-following capabilities across both text and visual inputs, providing accurate ...

Vision
O

LangMart: MiniMax: MiniMax-01

Openrouter

MiniMax-01 is a combines MiniMax-Text-01 for text generation and MiniMax-VL-01 for image understanding. It has 456 billion parameters, with 45.9 billion parameters activated per inference, and can han...

Vision
O

LangMart: Mistral: Devstral Small 2505

Openrouter

Devstral-Small-2505 is a 24B parameter agentic LLM fine-tuned from Mistral-Small-3.1, jointly developed by Mistral AI and All Hands AI for advanced software engineering tasks. It is optimized for code...

Vision
O

LangMart: Mistral: Ministral 3 14B 2512

Openrouter

The largest model in the Ministral 3 family, Ministral 3 14B offers frontier capabilities and performance comparable to its larger Mistral Small 3.2 24B counterpart. A powerful and efficient language ...

Vision
O

LangMart: Mistral: Ministral 3 3B 2512

Openrouter

The smallest model in the Ministral 3 family, Ministral 3 3B is a powerful, efficient tiny language model with vision capabilities.

Vision
O

LangMart: Mistral: Ministral 3 8B 2512

Openrouter

A balanced model in the Ministral 3 family, Ministral 3 8B is a powerful, efficient tiny language model with vision capabilities.

Vision
O

LangMart: Mistral: Mistral Large 3 2512

Openrouter

Mistral Large 3 2512 is Mistralโ€™s most capable model to date, featuring a sparse mixture-of-experts architecture with 41B active parameters (675B total), and released under the Apache 2.0 license.

Vision
O

LangMart: Mistral: Mistral Medium 3

Openrouter

Mistral Medium 3 is a high-performance enterprise-grade language model designed to deliver frontier-level capabilities at significantly reduced operational cost. It balances state-of-the-art reasoning...

Vision
O

LangMart: Mistral: Pixtral 12B

Openrouter

The first multi-modal, text+image-to-text model from Mistral AI. Its weights were launched via torrent: https://x.com/mistralai/status/1833758285167722836.

Vision
O

LangMart: Mistral: Pixtral Large 2411

Openrouter

Pixtral Large is a 124B parameter, open-weight, multimodal model built on top of [Mistral Large 2](/mistralai/mistral-large-2411). The model is able to understand documents, charts and natural images....

Vision
O

LangMart: NVIDIA: Nemotron Nano 12B 2 VL

Openrouter

NVIDIA Nemotron Nano 2 VL is a 12-billion-parameter open multimodal reasoning model designed for video understanding and document intelligence. It introduces a hybrid Transformer-Mamba architecture, c...

Vision
O

LangMart: OpenAI: ChatGPT-4o

Openrouter

OpenAI: ChatGPT-4o is a image generation model for creating visual content from descriptions. Developed by LangMart, this model is optimized for its specific use case category.

Streaming Vision
O

LangMart: OpenAI: Codex Mini

Openrouter

codex-mini-latest is a fine-tuned version of o4-mini specifically for use in Codex CLI. For direct use in the API, we recommend starting with gpt-4.1.

Vision
O

LangMart: OpenAI: GPT-3.5 Turbo

Openrouter

OpenAI: GPT-3.5 Turbo is a conversational AI model designed for multi-turn dialogue and interactive tasks. Developed by LangMart, this model is optimized for its specific use case category.

Streaming Vision
O

LangMart: OpenAI: GPT-3.5 Turbo 16k

Openrouter

OpenAI: GPT-3.5 Turbo 16k is a conversational AI model designed for multi-turn dialogue and interactive tasks. Developed by LangMart, this model is optimized for its specific use case category.

Streaming Vision
O

LangMart: OpenAI: GPT-3.5 Turbo Instruct

Openrouter

OpenAI: GPT-3.5 Turbo Instruct is a conversational AI model designed for multi-turn dialogue and interactive tasks. Developed by LangMart, this model is optimized for its specific use case category.

Streaming Vision
O

LangMart: OpenAI: GPT-4

Openrouter

OpenAI: GPT-4 is a conversational AI model designed for multi-turn dialogue and interactive tasks. Developed by LangMart, this model is optimized for its specific use case category.

Streaming Vision
O

LangMart: OpenAI: GPT-4 Turbo

Openrouter

OpenAI: GPT-4 Turbo is a image generation model for creating visual content from descriptions. Developed by LangMart, this model is optimized for its specific use case category.

Streaming Vision
O

LangMart: OpenAI: GPT-4 Turbo (older v1106)

Openrouter

OpenAI: GPT-4 Turbo (older v1106) is a conversational AI model designed for multi-turn dialogue and interactive tasks. Developed by LangMart, this model is optimized for its specific use case category...

Streaming Vision
O

LangMart: OpenAI: GPT-4 Turbo Preview

Openrouter

OpenAI: GPT-4 Turbo Preview is a conversational AI model designed for multi-turn dialogue and interactive tasks. Developed by LangMart, this model is optimized for its specific use case category.

Streaming Vision
O

LangMart: OpenAI: GPT-4.1

Openrouter

OpenAI: GPT-4.1 is a image generation model for creating visual content from descriptions. Developed by LangMart, this model is optimized for its specific use case category.

Streaming Vision
O

LangMart: OpenAI: GPT-4.1 Mini

Openrouter

OpenAI: GPT-4.1 Mini is a image generation model for creating visual content from descriptions. Developed by LangMart, this model is optimized for its specific use case category.

Streaming Vision
O

LangMart: OpenAI: GPT-4.1 Nano

Openrouter

OpenAI: GPT-4.1 Nano is a image generation model for creating visual content from descriptions. Developed by LangMart, this model is optimized for its specific use case category.

Streaming Vision
O

LangMart: OpenAI: GPT-4o

Openrouter

OpenAI: GPT-4o is a image generation model for creating visual content from descriptions. Developed by LangMart, this model is optimized for its specific use case category.

Streaming Vision
O

LangMart: OpenAI: GPT-4o (2024-05-13)

Openrouter

OpenAI: GPT-4o (2024-05-13) is a image generation model for creating visual content from descriptions. Developed by LangMart, this model is optimized for its specific use case category.

Streaming Vision
O

LangMart: OpenAI: GPT-4o (2024-08-06)

Openrouter

OpenAI: GPT-4o (2024-08-06) is a image generation model for creating visual content from descriptions. Developed by LangMart, this model is optimized for its specific use case category.

Streaming Vision
O

LangMart: OpenAI: GPT-4o (2024-11-20)

Openrouter

OpenAI: GPT-4o (2024-11-20) is a image generation model for creating visual content from descriptions. Developed by LangMart, this model is optimized for its specific use case category.

Streaming Vision
O

LangMart: OpenAI: GPT-4o Audio

Openrouter

OpenAI: GPT-4o Audio is a conversational AI model designed for multi-turn dialogue and interactive tasks. Developed by LangMart, this model is optimized for its specific use case category.

Streaming Vision
O

LangMart: OpenAI: GPT-4o Search Preview

Openrouter

OpenAI: GPT-4o Search Preview is a conversational AI model designed for multi-turn dialogue and interactive tasks. Developed by LangMart, this model is optimized for its specific use case category.

Streaming Vision
O

LangMart: OpenAI: GPT-4o-mini

Openrouter

OpenAI: GPT-4o-mini is a image generation model for creating visual content from descriptions. Developed by LangMart, this model is optimized for its specific use case category.

Streaming Vision
O

LangMart: OpenAI: GPT-4o-mini (2024-07-18)

Openrouter

OpenAI: GPT-4o-mini (2024-07-18) is a image generation model for creating visual content from descriptions. Developed by LangMart, this model is optimized for its specific use case category.

Streaming Vision
O

LangMart: OpenAI: GPT-4o-mini Search Preview

Openrouter

OpenAI: GPT-4o-mini Search Preview is a conversational AI model designed for multi-turn dialogue and interactive tasks. Developed by LangMart, this model is optimized for its specific use case categor...

Streaming Vision
O

LangMart: OpenAI: GPT-5

Openrouter

OpenAI: GPT-5 is a image generation model for creating visual content from descriptions. Developed by LangMart, this model is optimized for its specific use case category.

Streaming Vision
O

LangMart: OpenAI: GPT-5 Chat

Openrouter

GPT-5 Chat is designed for advanced, natural, multimodal, and context-aware conversations for enterprise applications.

Vision
O

LangMart: OpenAI: GPT-5 Codex

Openrouter

OpenAI: GPT-5 Codex is a image generation model for creating visual content from descriptions. Developed by LangMart, this model is optimized for its specific use case category.

Streaming Vision
O

LangMart: OpenAI: GPT-5 Image

Openrouter

[GPT-5](https://langmart.ai/model-docs) Image combines OpenAI's GPT-5 model with state-of-the-art image generation capabilities. It offers major improvements in reasoning, code quality, and user exper...

Vision
O

LangMart: OpenAI: GPT-5 Image Mini

Openrouter

GPT-5 Image Mini combines OpenAI's advanced language capabilities, powered by [GPT-5 Mini](https://langmart.ai/model-docs), with GPT Image 1 Mini for efficient image generation. This natively multimod...

Vision
O

LangMart: OpenAI: GPT-5 Mini

Openrouter

OpenAI: GPT-5 Mini is a image generation model for creating visual content from descriptions. Developed by LangMart, this model is optimized for its specific use case category.

Streaming Vision
O

LangMart: OpenAI: GPT-5 Nano

Openrouter

OpenAI: GPT-5 Nano is a image generation model for creating visual content from descriptions. Developed by LangMart, this model is optimized for its specific use case category.

Streaming Vision
O

LangMart: OpenAI: GPT-5 Pro

Openrouter

OpenAI: GPT-5 Pro is a image generation model for creating visual content from descriptions. Developed by LangMart, this model is optimized for its specific use case category.

Streaming Vision
O

LangMart: OpenAI: GPT-5.1

Openrouter

OpenAI: GPT-5.1 is a image generation model for creating visual content from descriptions. Developed by LangMart, this model is optimized for its specific use case category.

Streaming Vision
O

LangMart: OpenAI: GPT-5.1-Codex

Openrouter

OpenAI: GPT-5.1-Codex is a image generation model for creating visual content from descriptions. Developed by LangMart, this model is optimized for its specific use case category.

Streaming Vision
O

LangMart: OpenAI: GPT-5.1-Codex-Max

Openrouter

OpenAI: GPT-5.1-Codex-Max is a image generation model for creating visual content from descriptions. Developed by LangMart, this model is optimized for its specific use case category.

Streaming Vision
O

LangMart: OpenAI: GPT-5.1-Codex-Mini

Openrouter

OpenAI: GPT-5.1-Codex-Mini is a image generation model for creating visual content from descriptions. Developed by LangMart, this model is optimized for its specific use case category.

Streaming Vision
O

LangMart: OpenAI: GPT-5.2

Openrouter

OpenAI: GPT-5.2 is a image generation model for creating visual content from descriptions. Developed by LangMart, this model is optimized for its specific use case category.

Streaming Vision
O

LangMart: OpenAI: GPT-5.2 Pro

Openrouter

OpenAI: GPT-5.2 Pro is a image generation model for creating visual content from descriptions. Developed by LangMart, this model is optimized for its specific use case category.

Streaming Vision
O

LangMart: OpenAI: o1

Openrouter

OpenAI: o1 is a image generation model for creating visual content from descriptions. Developed by LangMart, this model is optimized for its specific use case category.

Streaming Vision
O

LangMart: OpenAI: o1-pro

Openrouter

OpenAI: o1-pro is a image generation model for creating visual content from descriptions. Developed by LangMart, this model is optimized for its specific use case category.

Streaming Vision
O

LangMart: OpenAI: o3

Openrouter

OpenAI: o3 is a image generation model for creating visual content from descriptions. Developed by LangMart, this model is optimized for its specific use case category.

Streaming Vision
O

LangMart: OpenAI: o3 Deep Research

Openrouter

OpenAI: o3 Deep Research is a image generation model for creating visual content from descriptions. Developed by LangMart, this model is optimized for its specific use case category.

Streaming Vision
O

LangMart: OpenAI: o3 Mini

Openrouter

OpenAI: o3 Mini is a conversational AI model designed for multi-turn dialogue and interactive tasks. Developed by LangMart, this model is optimized for its specific use case category.

Streaming Vision
O

LangMart: OpenAI: o3 Mini High

Openrouter

OpenAI o3-mini-high is the same model as [o3-mini](/openai/o3-mini) with reasoning_effort set to high.

Vision
O

LangMart: OpenAI: o3 Pro

Openrouter

OpenAI: o3 Pro is a image generation model for creating visual content from descriptions. Developed by LangMart, this model is optimized for its specific use case category.

Streaming Vision
O

LangMart: OpenAI: o4 Mini

Openrouter

OpenAI: o4 Mini is a image generation model for creating visual content from descriptions. Developed by LangMart, this model is optimized for its specific use case category.

Streaming Vision
O

LangMart: OpenAI: o4 Mini Deep Research

Openrouter

OpenAI: o4 Mini Deep Research is a image generation model for creating visual content from descriptions. Developed by LangMart, this model is optimized for its specific use case category.

Streaming Vision
O

LangMart: OpenAI: o4 Mini High

Openrouter

OpenAI o4-mini-high is the same model as [o4-mini](/openai/o4-mini) with reasoning_effort set to high.

Vision
O

LangMart: OpenGVLab: InternVL3 78B

Openrouter

The InternVL3 series is an advanced multimodal large language model (MLLM). Compared to InternVL 2.5, InternVL3 demonstrates stronger multimodal perception and reasoning capabilities.

Vision
O

LangMart: Perplexity: Sonar

Openrouter

Sonar is lightweight, affordable, fast, and simple to use โ€” now featuring citations and the ability to customize sources. It is designed for companies seeking to integrate lightweight question-and-ans...

Vision
O

LangMart: Perplexity: Sonar Pro

Openrouter

Note: Sonar Pro pricing includes Perplexity search pricing. See [details here](https://docs.perplexity.ai/guides/pricing#detailed-pricing-breakdown-for-sonar-reasoning-pro-and-sonar-pro)

Vision
O

LangMart: Perplexity: Sonar Pro Search

Openrouter

Exclusively available on the LangMart API, Sonar Pro's new Pro Search mode is Perplexity's most advanced agentic search system. It is designed for deeper reasoning and analysis. Pricing is based on to...

Vision
O

LangMart: Perplexity: Sonar Reasoning Pro

Openrouter

Note: Sonar Pro pricing includes Perplexity search pricing. See [details here](https://docs.perplexity.ai/guides/pricing#detailed-pricing-breakdown-for-sonar-reasoning-pro-and-sonar-pro)

Vision
O

LangMart: Qwen: Qwen VL Max

Openrouter

Qwen VL Max is a visual understanding model with 7500 tokens context length. It excels in delivering optimal performance for a broader spectrum of complex tasks.

Vision
O

LangMart: Qwen: Qwen VL Plus

Openrouter

Qwen's Enhanced Large Visual Language Model. Significantly upgraded for detailed recognition capabilities and text recognition abilities, supporting ultra-high pixel resolutions up to millions of pixe...

Vision
O

LangMart: Qwen: Qwen3 VL 235B A22B Instruct

Openrouter

Qwen3-VL-235B-A22B Instruct is an open-weight multimodal model that unifies strong text generation with visual understanding across images and video. The Instruct model targets general vision-language...

Vision
O

LangMart: Qwen: Qwen3 VL 235B A22B Thinking

Openrouter

Qwen3-VL-235B-A22B Thinking is a multimodal model that unifies strong text generation with visual understanding across images and video. The Thinking model is optimized for multimodal reasoning in STE...

Vision
O

LangMart: Qwen: Qwen3 VL 30B A3B Instruct

Openrouter

Qwen3-VL-30B-A3B-Instruct is a multimodal model that unifies strong text generation with visual understanding for images and videos. Its Instruct variant optimizes instruction-following for general mu...

Vision
O

LangMart: Qwen: Qwen3 VL 30B A3B Thinking

Openrouter

Qwen3-VL-30B-A3B-Thinking is a multimodal model that unifies strong text generation with visual understanding for images and videos. Its Thinking variant enhances reasoning in STEM, math, and complex ...

Vision
O

LangMart: Qwen: Qwen3 VL 32B Instruct

Openrouter

Qwen3-VL-32B-Instruct is a large-scale multimodal vision-language model designed for high-precision understanding and reasoning across text, images, and video. With 32 billion parameters, it combines ...

Vision
O

LangMart: Qwen: Qwen3 VL 8B Instruct

Openrouter

Qwen3-VL-8B-Instruct is a multimodal vision-language model from the Qwen3-VL series, built for high-fidelity understanding and reasoning across text, images, and video. It features improved multimodal...

Vision
O

LangMart: Qwen: Qwen3 VL 8B Thinking

Openrouter

Qwen3-VL-8B-Thinking is the reasoning-optimized variant of the Qwen3-VL-8B multimodal model, designed for advanced visual and textual reasoning across complex scenes, documents, and temporal sequences...

Vision
O

LangMart: StepFun: Step3

Openrouter

Step3 is a cutting-edge multimodal reasoning modelโ€”built on a Mixture-of-Experts architecture with 321B total parameters and 38B active. It is designed end-to-end to minimize decoding costs while deli...

Vision
O

LangMart: xAI: Grok 4

Openrouter

Grok 4 is xAI's latest reasoning model with a 256k context window. It supports parallel tool calling, structured outputs, and both image and text inputs. Note that reasoning is not exposed, reasoning ...

Vision
O

LangMart: xAI: Grok 4 Fast

Openrouter

Grok 4 Fast is xAI's latest multimodal model with SOTA cost-efficiency and a 2M token context window. It comes in two flavors: non-reasoning and reasoning. Read more about the model on xAI's [news pos...

Vision
O

OpenAI: GPT-4o

Openrouter

OpenAI: GPT-4o is a capable language model available on LangMart via OpenRouter for general-purpose text generation and analysis tasks.

Vision Tools Streaming Reasoning
O

OpenAI: GPT-4o (2024-05-13)

Openrouter

OpenAI: GPT-4o (2024-05-13) is a capable language model available on LangMart via OpenRouter for general-purpose text generation and analysis tasks.

Vision Tools Streaming Reasoning
O

OpenAI: GPT-4o (2024-08-06)

Openrouter

OpenAI: GPT-4o (2024-08-06) is a capable language model available on LangMart via OpenRouter for general-purpose text generation and analysis tasks.

Vision Tools Streaming Reasoning
O

OpenAI: GPT-4o (2024-11-20)

Openrouter

OpenAI: GPT-4o (2024-11-20) is a capable language model available on LangMart via OpenRouter for general-purpose text generation and analysis tasks.

Vision Tools Streaming Reasoning
O

OpenAI: GPT-4o-mini

Openrouter

OpenAI: GPT-4o-mini is a capable language model available on LangMart via OpenRouter for general-purpose text generation and analysis tasks.

Vision Tools Streaming Reasoning
O

OpenAI: GPT-4o-mini (2024-07-18)

Openrouter

OpenAI: GPT-4o-mini (2024-07-18) is a capable language model available on LangMart via OpenRouter for general-purpose text generation and analysis tasks.

Vision Tools Streaming Reasoning
P

Perplexity PPLX 7B Online

Perplexity

**Model ID**: `perplexity/pplx-7b-online`

Vision
P

Perplexity Sonar Reasoning

Perplexity

**Model ID**: `perplexity/sonar-reasoning`

Vision
P

Perplexity: Sonar Pro

Perplexity

Perplexity Sonar Pro is an advanced search-augmented language model designed for in-depth, multi-step queries with added extensibility. It offers approximately double the citations per search compared...

Vision
Q

Qwen 1.5 14B Chat

Qwen

**Model ID**: `qwen/qwen-1.5-14b-chat`

Vision
Q

Qwen 2.5 72B Instruct

Qwen

**Model ID**: `qwen/qwen-2.5-72b-instruct`

Vision
Q

Qwen 2.5 7B Instruct

Qwen

**Model ID**: `qwen/qwen-2.5-7b-instruct`

Vision
Q

Qwen: Qwen3 VL 8B Instruct

Qwen

**Model ID:** `qwen/qwen3-vl-8b-instruct`

Vision
Q

Qwen2.5 VL 72B Instruct

Qwen

Qwen2.5 VL 72B Instruct is a state-of-the-art multimodal model that excels at visual understanding and reasoning tasks. The model demonstrates exceptional capabilities in:

Vision
R

Reka Flash

Reka

**Model ID**: `reka/reka-flash`

Vision
S

Stable Diffusion 3.5 Large

Stabilityai

Stable Diffusion 3.5 Large is the most powerful model in the Stable Diffusion family, featuring superior quality and prompt adherence. It is a Multimodal Diffusion Transformer (MMDiT) text-to-image ge...

Vision
T

Together AI: Arcee AI Agent

Together AI

Arcee AI Agent is a conversational AI model designed for multi-turn dialogue and interactive tasks. Developed by Together AI, this model offers reliable performance for diverse applications.

Streaming Vision
T

Together AI: Arcee AI Caller Agent

Together AI

Arcee AI Caller Agent is a conversational AI model designed for multi-turn dialogue and interactive tasks. Developed by Together AI, this model offers reliable performance for diverse applications.

Streaming Vision
T

Together AI: Arcee AI Spotlight

Together AI

Arcee AI Spotlight is a conversational AI model designed for multi-turn dialogue and interactive tasks. Developed by Together AI, this model offers reliable performance for diverse applications.

Streaming Vision
T

Together AI: Arcee AI Virtuoso

Together AI

Arcee AI Virtuoso is a conversational AI model designed for multi-turn dialogue and interactive tasks. Developed by Together AI, this model offers reliable performance for diverse applications.

Streaming Vision
T

Together AI: BGE Base EN v1.5

Together AI

BGE Base EN v1.5 is a conversational AI model designed for multi-turn dialogue and interactive tasks. Developed by Together AI, this model offers reliable performance for diverse applications.

Streaming Vision
T

Together AI: BGE Large EN v1.5

Together AI

BGE Large EN v1.5 is a conversational AI model designed for multi-turn dialogue and interactive tasks. Developed by Together AI, this model offers reliable performance for diverse applications.

Streaming Vision
T

Together AI: Cogito v1 Preview 32B

Together AI

Cogito v1 Preview 32B is a conversational AI model designed for multi-turn dialogue and interactive tasks. Developed by Together AI, this model offers reliable performance for diverse applications.

Streaming Vision
T

Together AI: Cogito v1 Preview 70B

Together AI

Cogito v1 Preview 70B is a conversational AI model designed for multi-turn dialogue and interactive tasks. Developed by Together AI, this model offers reliable performance for diverse applications.

Streaming Vision
T

Together AI: Cogito v1 Preview 8B

Together AI

Cogito v1 Preview 8B is a conversational AI model designed for multi-turn dialogue and interactive tasks. Developed by Together AI, this model offers reliable performance for diverse applications.

Streaming Vision
T

Together AI: DeepSeek R1

Together AI

DeepSeek R1 is a conversational AI model designed for multi-turn dialogue and interactive tasks. Developed by Together AI, this model offers reliable performance for diverse applications.

Streaming Vision
T

Together AI: DeepSeek V3

Together AI

DeepSeek V3 is a conversational AI model designed for multi-turn dialogue and interactive tasks. Developed by Together AI, this model offers reliable performance for diverse applications.

Streaming Vision
T

Together AI: DeepSeek V3 0324

Together AI

DeepSeek V3 0324 is a conversational AI model designed for multi-turn dialogue and interactive tasks. Developed by Together AI, this model offers reliable performance for diverse applications.

Streaming Vision
T

Together AI: FLUX.1 Dev

Together AI

FLUX.1 Dev is a conversational AI model designed for multi-turn dialogue and interactive tasks. Developed by Together AI, this model offers reliable performance for diverse applications.

Streaming Vision
T

Together AI: FLUX.1 Pro

Together AI

FLUX.1 Pro is a conversational AI model designed for multi-turn dialogue and interactive tasks. Developed by Together AI, this model offers reliable performance for diverse applications.

Streaming Vision
T

Together AI: FLUX.1.1 Pro

Together AI

FLUX.1.1 Pro is a conversational AI model designed for multi-turn dialogue and interactive tasks. Developed by Together AI, this model offers reliable performance for diverse applications.

Streaming Vision
T

Together AI: FLUX.1.1 Pro Ultra

Together AI

FLUX.1.1 Pro Ultra is a conversational AI model designed for multi-turn dialogue and interactive tasks. Developed by Together AI, this model offers reliable performance for diverse applications.

Streaming Vision
T

Together AI: GLM-4 9B

Together AI

GLM-4 9B is a conversational AI model designed for multi-turn dialogue and interactive tasks. Developed by Together AI, this model offers reliable performance for diverse applications.

Streaming Vision
T

Together AI: GLM-Z1 9B

Together AI

GLM-Z1 9B is a conversational AI model designed for multi-turn dialogue and interactive tasks. Developed by Together AI, this model offers reliable performance for diverse applications.

Streaming Vision
T

Together AI: GTE ModernBERT Base

Together AI

GTE ModernBERT Base is a conversational AI model designed for multi-turn dialogue and interactive tasks. Developed by Together AI, this model offers reliable performance for diverse applications.

Streaming Vision
T

Together AI: Kimi K2 Instruct

Together AI

Kimi K2 Instruct is a conversational AI model designed for multi-turn dialogue and interactive tasks. Developed by Together AI, this model offers reliable performance for diverse applications.

Streaming Vision
T

Together AI: Llama 3.1 405B

Together AI

Llama 3.1 405B is a conversational AI model designed for multi-turn dialogue and interactive tasks. Developed by Together AI, this model offers reliable performance for diverse applications.

Tools Streaming Vision
T

Together AI: Llama 3.1 70B

Together AI

Llama 3.1 70B is a conversational AI model designed for multi-turn dialogue and interactive tasks. Developed by Together AI, this model offers reliable performance for diverse applications.

Tools Streaming Vision
T

Together AI: Llama 3.1 8B

Together AI

Llama 3.1 8B is a conversational AI model designed for multi-turn dialogue and interactive tasks. Developed by Together AI, this model offers reliable performance for diverse applications.

Tools Streaming Vision
T

Together AI: Llama 3.3 70B

Together AI

Llama 3.3 70B is a conversational AI model designed for multi-turn dialogue and interactive tasks. Developed by Together AI, this model offers reliable performance for diverse applications.

Tools Streaming Vision
T

Together AI: Llama 4 Maverick

Together AI

Llama 4 Maverick is a conversational AI model designed for multi-turn dialogue and interactive tasks. Developed by Together AI, this model offers reliable performance for diverse applications.

Streaming Vision
T

Together AI: Llama 4 Scout

Together AI

Llama 4 Scout is a conversational AI model designed for multi-turn dialogue and interactive tasks. Developed by Together AI, this model offers reliable performance for diverse applications.

Streaming Vision
T

Together AI: Llama Guard 2 8B

Together AI

Llama Guard 2 8B is a conversational AI model designed for multi-turn dialogue and interactive tasks. Developed by Together AI, this model offers reliable performance for diverse applications.

Streaming Vision
T

Together AI: Llama Guard 3 11B Vision Turbo

Together AI

Llama Guard 3 11B Vision Turbo is a conversational AI model designed for multi-turn dialogue and interactive tasks. Developed by Together AI, this model offers reliable performance for diverse applica...

Streaming Vision
T

Together AI: Llama Guard 3 8B

Together AI

Llama Guard 3 8B is a conversational AI model designed for multi-turn dialogue and interactive tasks. Developed by Together AI, this model offers reliable performance for diverse applications.

Streaming Vision
T

Together AI: Llama Guard 4 12B

Together AI

Llama Guard 4 12B is a conversational AI model designed for multi-turn dialogue and interactive tasks. Developed by Together AI, this model offers reliable performance for diverse applications.

Streaming Vision
T

Together AI: M2-BERT 80M 32K Retrieval

Together AI

M2-BERT 80M 32K Retrieval is a conversational AI model designed for multi-turn dialogue and interactive tasks. Developed by Together AI, this model offers reliable performance for diverse applications...

Streaming Vision
T

Together AI: Mixtral 8x22B

Together AI

Mixtral 8x22B is a conversational AI model designed for multi-turn dialogue and interactive tasks. Developed by Together AI, this model offers reliable performance for diverse applications.

Tools Streaming Vision
T

Together AI: Multilingual e5 Large Instruct

Together AI

Multilingual e5 Large Instruct is a conversational AI model designed for multi-turn dialogue and interactive tasks. Developed by Together AI, this model offers reliable performance for diverse applica...

Streaming Vision
T

Together AI: Mxbai Rerank Large V2

Together AI

Mxbai Rerank Large V2 is a conversational AI model designed for multi-turn dialogue and interactive tasks. Developed by Together AI, this model offers reliable performance for diverse applications.

Streaming Vision
T

Together AI: Qwen 2.5 72B

Together AI

Qwen 2.5 72B is a conversational AI model designed for multi-turn dialogue and interactive tasks. Developed by Together AI, this model offers reliable performance for diverse applications.

Tools Streaming Vision
T

Together AI: Qwen 2.5 7B

Together AI

Qwen 2.5 7B is a conversational AI model designed for multi-turn dialogue and interactive tasks. Developed by Together AI, this model offers reliable performance for diverse applications.

Streaming Vision
T

Together AI: Qwen 2.5 Coder 32B

Together AI

Qwen 2.5 Coder 32B is a conversational AI model designed for multi-turn dialogue and interactive tasks. Developed by Together AI, this model offers reliable performance for diverse applications.

Streaming Vision
T

Together AI: Qwen QwQ 32B

Together AI

Qwen QwQ 32B is a conversational AI model designed for multi-turn dialogue and interactive tasks. Developed by Together AI, this model offers reliable performance for diverse applications.

Streaming Vision
T

Together AI: Qwen3 235B A22B

Together AI

Qwen3 235B A22B is a conversational AI model designed for multi-turn dialogue and interactive tasks. Developed by Together AI, this model offers reliable performance for diverse applications.

Streaming Vision
T

Together AI: Qwen3 30B A3B

Together AI

Qwen3 30B A3B is a conversational AI model designed for multi-turn dialogue and interactive tasks. Developed by Together AI, this model offers reliable performance for diverse applications.

Streaming Vision
T

Together AI: Qwen3 32B

Together AI

Qwen3 32B is a conversational AI model designed for multi-turn dialogue and interactive tasks. Developed by Together AI, this model offers reliable performance for diverse applications.

Streaming Vision
T

Together AI: Qwen3 8B

Together AI

Qwen3 8B is a conversational AI model designed for multi-turn dialogue and interactive tasks. Developed by Together AI, this model offers reliable performance for diverse applications.

Streaming Vision
T

Together AI: Salesforce Llama Rank V1 8B

Together AI

Salesforce Llama Rank V1 8B is a conversational AI model designed for multi-turn dialogue and interactive tasks. Developed by Together AI, this model offers reliable performance for diverse applicatio...

Streaming Vision
T

Together AI: Stable Diffusion 3.5 Large Turbo

Together AI

Stable Diffusion 3.5 Large Turbo is a conversational AI model designed for multi-turn dialogue and interactive tasks. Developed by Together AI, this model offers reliable performance for diverse appli...

Streaming Vision
T

Together AI: Stable Diffusion XL

Together AI

Stable Diffusion XL is a conversational AI model designed for multi-turn dialogue and interactive tasks. Developed by Together AI, this model offers reliable performance for diverse applications.

Streaming Vision
T

Together AI: VirtueGuard Text Lite

Together AI

VirtueGuard Text Lite is a conversational AI model designed for multi-turn dialogue and interactive tasks. Developed by Together AI, this model offers reliable performance for diverse applications.

Streaming Vision
T

Together AI: Whisper Large v3

Together AI

Whisper Large v3 is a conversational AI model designed for multi-turn dialogue and interactive tasks. Developed by Together AI, this model offers reliable performance for diverse applications.

Streaming Vision
X

xAI Grok 3 Beta

Xai

Grok 3 Beta is xAI's flagship reasoning model, described as their "most advanced model" showcasing superior reasoning capabilities and extensive pretraining knowledge. It excels at enterprise use case...

Vision
Z

Z.AI: GLM 4.7

Z Ai

GLM-4.7 is Z.AI's latest flagship model, featuring upgrades in two key areas: enhanced programming capabilities and more stable multi-step reasoning/execution.

Vision