๐Ÿ†“ Free

Free-tier models with no per-token cost. 105 models.

A

AllenAI: Olmo 3.1 32B Think (free)

Allen AI

This is a 32-billion parameter model emphasizing reasoning capabilities. The system excels at deep reasoning, complex multi-step logic, and advanced instruction following. Version 3.1 represents impro...

Vision
C

Collections: Auto Free

Collection

The **Auto Free Collection** is a dynamic system-level collection that automatically routes inference requests to one of the predefined free models available in the system. This collection provides co...

G

Gemma 3 27B (free) - Complete Model Details

Google

Vision
G

Google Gemini 2.0 Flash Experimental (Free)

Google

Gemini Flash 2.0 offers a significantly faster time to first token (TTFT) compared to Gemini Flash 1.5, while maintaining quality on par with larger models like Gemini Pro 1.5. It introduces notable e...

Vision
G

Google: Gemini 2.5 Flash Preview

Google

Preview of Gemini 2.5 Flash.

Vision Tools
G

Google: Gemini 3 Flash Preview

Google

Fast Gemini 3 for speed and efficiency.

Vision Tools
G

Google: Gemini 3 Mobile

Google

Lightweight Gemini 3 optimized for mobile devices.

Vision Tools
G

Google: Gemini 3 Pro Preview

Google

Latest flagship Gemini model with advanced reasoning and multimodal capabilities.

Vision Tools
G

Google: Gemini Reasoning Engine

Google

Experimental advanced reasoning engine for complex tasks.

Vision Tools
M

Mistral AI: Devstral 2

Mistral AI

Devstral 2 is a code generation and understanding model for programming tasks. Developed by Mistral AI, this model is optimized for its specific use case category.

Streaming Vision
M

Mistral AI: Devstral Small 2

Mistral AI

Devstral Small 2 is a code generation and understanding model for programming tasks. Developed by Mistral AI, this model is optimized for its specific use case category.

Streaming Vision
M

Mistral: Devstral 2 2512 (Free)

Mistral AI

Devstral 2 is a state-of-the-art open-source model by Mistral AI specializing in agentic coding. It is a 123B-parameter dense transformer model supporting a 256K context window.

Vision
O

Ollama: Embeddinggemma:300m

Ollama

Embeddinggemma:300m optimized for generating high-quality embeddings. This model supports multimodal capabilities including vision and image understanding.

Streaming
O

Ollama: Functiongemma:270m

Ollama

Functiongemma:270m is a capable language model from Ollama for general-purpose text generation and analysis tasks. This model supports multimodal capabilities including vision and image understanding....

Tools Streaming
O

Ollama: Gemma3:1b

Ollama

Gemma3:1b is a capable language model from Ollama for general-purpose text generation and analysis tasks. This model supports multimodal capabilities including vision and image understanding.

Streaming
O

Ollama: Llama 3.1 8B Instruct

Ollama

Llama 3.1 8B Instruct is Meta's state-of-the-art instruction-tuned language model with 8 billion parameters. It's a compact yet powerful model designed for general-purpose conversational AI, reasoning...

Vision
O

Ollama: Mistral 7B Instruct

Ollama

Mistral 7B Instruct is Mistral AI's powerful 7-billion parameter instruction-tuned language model, renowned for exceptional efficiency and speed. Despite having only 7B parameters, it achieves perform...

O

Ollama: Qwen2.5 7B Instruct

Ollama

Qwen2.5 7B Instruct is Alibaba's latest-generation instruction-tuned language model with 7.6 billion parameters, representing a significant upgrade to the Qwen family. Built on 18 trillion tokens of d...

Vision
O

OpenAI: Omni Moderation

OpenAI

Omni Moderation is a content moderation model for safety and policy compliance checking. Developed by OpenAI, this model is optimized for its specific use case category.

Streaming Vision
O

OpenAI: Text Embedding 3 Large

OpenAI

Large embedding model for advanced semantic tasks. This model supports multimodal capabilities including vision and image understanding. It features advanced reasoning capabilities for complex problem...

Streaming
O

OpenAI: Text Embedding 3 Small

OpenAI

Efficient text embedding model with high quality representations. This model supports multimodal capabilities including vision and image understanding. It features advanced reasoning capabilities for ...

Streaming
O

OpenAI: Text Embedding Ada 002

OpenAI

Legacy embedding model still widely used. This model supports multimodal capabilities including vision and image understanding. It features advanced reasoning capabilities for complex problem-solving ...

Streaming
O

Openai-compatible: Fake Gpt 4 Vision

Openai Compatible

Fake Gpt 4 Vision with vision capabilities for processing images and visual content. This model supports multimodal capabilities including vision and image understanding. It features advanced reasonin...

Vision Streaming
O

OpenChat 3.5 7B (Free)

Openchat

OpenChat is a library of open-source language models fine-tuned with C-RLFT, a strategy inspired by offline reinforcement learning. The model is trained on mixed-quality data without preference labels...

Reasoning Vision
O

Agentica: Deepcoder 14B Preview (free)

Openrouter

DeepCoder 14B Preview is Agentica's code-focused model optimized for programming tasks and code generation.

Streaming
O

ArliAI: QwQ 32B RpR v1 (free)

Openrouter

QwQ 32B RpR v1 is ArliAI's roleplay-optimized reasoning model combining analytical capabilities with creative character interaction.

Streaming
O

Cohere: North Mini Code (free)

Openrouter

Cohere: North Mini Code (free) is a capable language model available on LangMart via OpenRouter for general-purpose text generation and analysis tasks.

Streaming
O

DeepSeek: DeepSeek R1 0528 Qwen3 8B (free)

Openrouter

DeepSeek R1 0528 Qwen3 8B is a distilled version combining R1 reasoning with efficient Qwen3 8B architecture.

Streaming
O

DeepSeek: DeepSeek V3 0324 (free)

Openrouter

Main chat model with JSON output and tool calling capabilities. Optimized for conversational AI, general-purpose tasks, and integration with external tools and APIs.

Streaming
O

DeepSeek: DeepSeek V3.1 (free)

Openrouter

Main chat model with JSON output and tool calling capabilities. Optimized for conversational AI, general-purpose tasks, and integration with external tools and APIs.

Streaming
O

DeepSeek: R1 (free)

Openrouter

DeepSeek R1 is a reasoning-focused model with extended thinking capabilities for complex multi-step problem solving.

Streaming
O

DeepSeek: R1 Distill Llama 70B (free)

Openrouter

DeepSeek R1 Distill Llama 70B transfers R1 reasoning capabilities to efficient Llama 70B architecture.

Streaming
O

Dots Studio: Dots3-Note Preview (free)

Openrouter

Dots Studio: Dots3-Note Preview (free) is a capable language model available on LangMart via OpenRouter for general-purpose text generation and analysis tasks.

Streaming
O

Google: Gemma 4 26B A4B (free)

Openrouter

Google: Gemma 4 26B A4B (free) is a capable language model available on LangMart via OpenRouter for general-purpose text generation and analysis tasks.

Streaming
O

Google: Gemma 4 31B (free)

Openrouter

Google: Gemma 4 31B (free) is a capable language model available on LangMart via OpenRouter for general-purpose text generation and analysis tasks.

Streaming
O

Google: Lyria 3 Clip Preview

Openrouter

Google: Lyria 3 Clip Preview is a capable language model available on LangMart via OpenRouter for general-purpose text generation and analysis tasks.

Streaming
O

Google: Lyria 3 Pro Preview

Openrouter

Google: Lyria 3 Pro Preview is a capable language model available on LangMart via OpenRouter for general-purpose text generation and analysis tasks.

Streaming
O

LangMart: Alibaba/tongyi Deepresearch 30b A3b:free

Openrouter

Free tier version of Alibaba/tongyi Deepresearch 30b A3b:free. This model supports multimodal capabilities including vision and image understanding.

Streaming
O

LangMart: Allenai/olmo 3 32b Think:free

Openrouter

Free tier version of Allenai/olmo 3 32b Think:free. This model supports multimodal capabilities including vision and image understanding.

Streaming
O

LangMart: Allenai/olmo 3.1 32b Think:free

Openrouter

Free tier version of Allenai/olmo 3.1 32b Think:free. This model supports multimodal capabilities including vision and image understanding.

Streaming
O

LangMart: Arcee Ai/trinity Mini:free

Openrouter

Free tier version of Arcee Ai/trinity Mini:free. This model supports multimodal capabilities including vision and image understanding.

Streaming
O

LangMart: Auto Router

Openrouter

Your prompt will be processed by a meta-model and routed to one of dozens of models (see below), optimizing for the best possible output.

O

LangMart: Cognitivecomputations/dolphin Mistral 24b Venice Edition:free

Openrouter

Free tier version of Cognitivecomputations/dolphin Mistral 24b Venice Edition:free. This model supports multimodal capabilities including vision and image understanding.

Streaming
O

LangMart: Deepseek/deepseek R1 0528:free

Openrouter

Free tier version of Deepseek/deepseek R1 0528:free. This model supports multimodal capabilities including vision and image understanding.

Streaming
O

LangMart: Google/gemini 2.0 Flash Exp:free

Openrouter

Google/gemini 2.0 Flash Exp:free with optimized speed for rapid response generation. This model supports multimodal capabilities including vision and image understanding.

Streaming
O

LangMart: Google/gemma 3 12b It:free

Openrouter

Free tier version of Google/gemma 3 12b It:free. This model supports multimodal capabilities including vision and image understanding.

Streaming
O

LangMart: Google/gemma 3 27b It:free

Openrouter

Free tier version of Google/gemma 3 27b It:free. This model supports multimodal capabilities including vision and image understanding.

Streaming
O

LangMart: Google/gemma 3 4b It:free

Openrouter

Free tier version of Google/gemma 3 4b It:free. This model supports multimodal capabilities including vision and image understanding.

Streaming
O

LangMart: Google/gemma 3n E2b It:free

Openrouter

Free tier version of Google/gemma 3n E2b It:free. This model supports multimodal capabilities including vision and image understanding.

Streaming
O

LangMart: Google/gemma 3n E4b It:free

Openrouter

Free tier version of Google/gemma 3n E4b It:free. This model supports multimodal capabilities including vision and image understanding.

Streaming
O

LangMart: Kwaipilot/kat Coder Pro:free

Openrouter

Free tier version of Kwaipilot/kat Coder Pro:free. This model supports multimodal capabilities including vision and image understanding. The model is optimized for code generation and programming task...

Streaming
O

LangMart: Meta Llama/llama 3.1 405b Instruct:free

Openrouter

Free tier version of Meta Llama/llama 3.1 405b Instruct:free. This model supports multimodal capabilities including vision and image understanding.

Streaming
O

LangMart: Meta Llama/llama 3.2 3b Instruct:free

Openrouter

Free tier version of Meta Llama/llama 3.2 3b Instruct:free. This model supports multimodal capabilities including vision and image understanding.

Streaming
O

LangMart: Meta Llama/llama 3.3 70b Instruct:free

Openrouter

Free tier version of Meta Llama/llama 3.3 70b Instruct:free. This model supports multimodal capabilities including vision and image understanding.

Streaming
O

LangMart: Mistralai/devstral 2512:free

Openrouter

Free tier version of Mistralai/devstral 2512:free. This model supports multimodal capabilities including vision and image understanding.

Streaming
O

LangMart: Mistralai/mistral 7b Instruct:free

Openrouter

Free tier version of Mistralai/mistral 7b Instruct:free. This model supports multimodal capabilities including vision and image understanding.

Streaming
O

LangMart: Mistralai/mistral Small 3.1 24b Instruct:free

Openrouter

Free tier version of Mistralai/mistral Small 3.1 24b Instruct:free. This model supports multimodal capabilities including vision and image understanding.

Streaming
O

LangMart: Moonshotai/kimi K2:free

Openrouter

Free tier version of Moonshotai/kimi K2:free. This model supports multimodal capabilities including vision and image understanding.

Streaming
O

LangMart: Nex Agi/deepseek V3.1 Nex N1:free

Openrouter

Free tier version of Nex Agi/deepseek V3.1 Nex N1:free. This model supports multimodal capabilities including vision and image understanding.

Streaming
O

LangMart: Nousresearch/hermes 3 Llama 3.1 405b:free

Openrouter

Free tier version of Nousresearch/hermes 3 Llama 3.1 405b:free. This model supports multimodal capabilities including vision and image understanding.

Streaming
O

LangMart: Nvidia/nemotron 3 Nano 30b A3b:free

Openrouter

Free tier version of Nvidia/nemotron 3 Nano 30b A3b:free. This model supports multimodal capabilities including vision and image understanding.

Streaming
O

LangMart: Nvidia/nemotron Nano 12b V2 Vl:free

Openrouter

Free tier version of Nvidia/nemotron Nano 12b V2 Vl:free. This model supports multimodal capabilities including vision and image understanding.

Streaming
O

LangMart: Nvidia/nemotron Nano 9b V2:free

Openrouter

Free tier version of Nvidia/nemotron Nano 9b V2:free. This model supports multimodal capabilities including vision and image understanding.

Streaming
O

LangMart: Openai/gpt Oss 120b:free

Openrouter

Free tier version of Openai/gpt Oss 120b:free. This model supports multimodal capabilities including vision and image understanding.

Streaming
O

LangMart: Openai/gpt Oss 20b:free

Openrouter

Free tier version of Openai/gpt Oss 20b:free. This model supports multimodal capabilities including vision and image understanding.

Streaming
O

LangMart: Qwen/qwen 2.5 Vl 7b Instruct:free

Openrouter

Free tier version of Qwen/qwen 2.5 Vl 7b Instruct:free. This model supports multimodal capabilities including vision and image understanding.

Streaming
O

LangMart: Qwen/qwen3 4b:free

Openrouter

Free tier version of Qwen/qwen3 4b:free. This model supports multimodal capabilities including vision and image understanding.

Streaming
O

LangMart: Qwen/qwen3 Coder:free

Openrouter

Free tier version of Qwen/qwen3 Coder:free. This model supports multimodal capabilities including vision and image understanding. The model is optimized for code generation and programming tasks.

Streaming
O

LangMart: Tngtech/deepseek R1t Chimera:free

Openrouter

Free tier version of Tngtech/deepseek R1t Chimera:free. This model supports multimodal capabilities including vision and image understanding.

Streaming
O

LangMart: Tngtech/deepseek R1t2 Chimera:free

Openrouter

Free tier version of Tngtech/deepseek R1t2 Chimera:free. This model supports multimodal capabilities including vision and image understanding.

Streaming
O

LangMart: Tngtech/tng R1t Chimera:free

Openrouter

Free tier version of Tngtech/tng R1t Chimera:free. This model supports multimodal capabilities including vision and image understanding.

Streaming
O

LangMart: Xiaomi/mimo V2 Flash:free

Openrouter

Xiaomi/mimo V2 Flash:free with optimized speed for rapid response generation. This model supports multimodal capabilities including vision and image understanding.

Streaming
O

LangMart: Z Ai/glm 4.5 Air:free

Openrouter

Free tier version of Z Ai/glm 4.5 Air:free. This model supports multimodal capabilities including vision and image understanding.

Streaming
O

LiquidAI: LFM2.5-2.6B (free)

Openrouter

LiquidAI: LFM2.5-2.6B (free) is a capable language model available on LangMart via OpenRouter for general-purpose text generation and analysis tasks.

Streaming
O

Meituan: LongCat Flash Chat (free)

Openrouter

LongCat Flash Chat is Meituan's efficient model with 131K context optimized for Chinese language tasks.

Streaming
O

Meta: Llama 3.3 8B Instruct (free)

Openrouter

Llama 3.3 8B Instruct (free) is Meta's efficient instruction-tuned model for general tasks.

Streaming
O

Meta: Llama 4 Maverick (free)

Openrouter

A high-capacity multimodal language model from Meta, built on a mixture-of-experts (MoE) architecture with 128 experts and 17 billion active parameters per forward pass (400B total parameters). The mo...

Streaming
O

Meta: Llama 4 Scout (free)

Openrouter

Meta's Llama 4 Scout is a mixture-of-experts (MoE) language model that activates 17 billion parameters from a total of 109 billion parameters. The model supports multimodal inputs (text and images) an...

Streaming
O

Microsoft: MAI DS R1 (free)

Openrouter

MAI DS R1 is Microsoft's DeepSeek-based reasoning model with enhanced analytical capabilities.

Streaming
O

MiniMax: MiniMax M2.7 (free)

Openrouter

MiniMax: MiniMax M2.7 (free) is a capable language model available on LangMart via OpenRouter for general-purpose text generation and analysis tasks.

Streaming
O

MiniMax: MiniMax M3 (free)

Openrouter

MiniMax: MiniMax M3 (free) is a capable language model available on LangMart via OpenRouter for general-purpose text generation and analysis tasks.

Streaming
O

Mistral: Mistral Nemo (free)

Openrouter

A 12-billion parameter model featuring a 128k token context window, developed by Mistral in partnership with NVIDIA. The model supports multiple languages including English, French, German, Spanish, I...

Streaming
O

Mistral: Mistral Small 3 (free)

Openrouter

Mistral Small is a 22-billion parameter model serving as a convenient mid-point between smaller and larger Mistral options. It emphasizes reasoning capabilities, code generation, and multilingual supp...

Streaming
O

Mistral: Mistral Small 3.2 24B (free)

Openrouter

Mistral Small is a 22-billion parameter model serving as a convenient mid-point between smaller and larger Mistral options. It emphasizes reasoning capabilities, code generation, and multilingual supp...

Streaming
O

NVIDIA: Nemotron 3 Nano Omni (free)

Openrouter

NVIDIA: Nemotron 3 Nano Omni (free) is a capable language model available on LangMart via OpenRouter for general-purpose text generation and analysis tasks.

Streaming
O

NVIDIA: Nemotron 3 Super (free)

Openrouter

NVIDIA: Nemotron 3 Super (free) is a capable language model available on LangMart via OpenRouter for general-purpose text generation and analysis tasks.

Streaming
O

NVIDIA: Nemotron 3 Ultra (free)

Openrouter

NVIDIA: Nemotron 3 Ultra (free) is a capable language model available on LangMart via OpenRouter for general-purpose text generation and analysis tasks.

Streaming
O

NVIDIA: Nemotron 3.5 Content Safety (free)

Openrouter

NVIDIA: Nemotron 3.5 Content Safety (free) is a capable language model available on LangMart via OpenRouter for general-purpose text generation and analysis tasks.

Streaming
O

NVIDIA: Nemotron 3.5 Lightning (free)

Openrouter

NVIDIA: Nemotron 3.5 Lightning (free) is a capable language model available on LangMart via OpenRouter for general-purpose text generation and analysis tasks.

Streaming
O

OpenRouter: Free Models Router

Openrouter

Free Models Router is a capable language model available on LangMart via OpenRouter for general-purpose text generation and analysis tasks.

Streaming
O

OpenRouter: Ling 3.0 Flash Fin (free)

Openrouter

Ling 3.0 Flash Fin (free) is a capable language model available on LangMart via OpenRouter for general-purpose text generation and analysis tasks.

Streaming
O

OpenRouter: Qwen2.5 72B Instruct (free)

Openrouter

Qwen2.5 72B is the latest series of Qwen large language models. Qwen2.5 brings the following improvements upon Qwen2:

Streaming
O

OpenRouter: Qwen2.5 Coder 32B Instruct (free)

Openrouter

Qwen2.5 Coder 32B Instruct is a code-focused large language model representing the latest iteration in the Qwen coding series. It replaces the earlier CodeQwen1.5 with significantly enhanced capabilit...

Streaming
O

Poolside: Laguna S 2.1 (free)

Openrouter

Poolside: Laguna S 2.1 (free) is a capable language model available on LangMart via OpenRouter for general-purpose text generation and analysis tasks.

Streaming
O

Poolside: Laguna XS 2.1 (free)

Openrouter

Poolside: Laguna XS 2.1 (free) is a capable language model available on LangMart via OpenRouter for general-purpose text generation and analysis tasks.

Streaming
O

Qwen: Qwen2.5 VL 32B Instruct (free)

Openrouter

Qwen2.5 VL 32B Instruct offers enhanced vision-language capabilities in the 32B configuration.

Streaming
O

Qwen: Qwen3 14B (free)

Openrouter

Qwen3 14B is Alibaba's efficient model with strong general capabilities and 40K context.

Streaming
O

Qwen: Qwen3 30B A3B (free)

Openrouter

Qwen3 30B A3B is an efficient MoE model with 3B active parameters for fast inference.

Streaming
O

Thinking Machines: Inkling (free)

Openrouter

Thinking Machines: Inkling (free) is a capable language model available on LangMart via OpenRouter for general-purpose text generation and analysis tasks.

Streaming
O

Thinking Machines: Inkling Small (free)

Openrouter

Thinking Machines: Inkling Small (free) is a capable language model available on LangMart via OpenRouter for general-purpose text generation and analysis tasks.

Streaming
O

Z.ai: GLM 5.2 (free)

Openrouter

Z.ai: GLM 5.2 (free) is a capable language model available on LangMart via OpenRouter for general-purpose text generation and analysis tasks.

Streaming
S

Stable Diffusion 3.5 Large

Stabilityai

Stable Diffusion 3.5 Large is the most powerful model in the Stable Diffusion family, featuring superior quality and prompt adherence. It is a Multimodal Diffusion Transformer (MMDiT) text-to-image ge...

Vision
U

Toppy M 7B

Undi95

Toppy M 7B is a wild 7B parameter model that merges several models using the new `task_arithmetic` merge method from [mergekit](https://github.com/cg123/mergekit). This model combines multiple fine-tu...

U

SOLAR-10.7B-Instruct-v1.0 Model Documentation

Upstage

tokenizer = AutoTokenizer.from_pretrained("upstage/SOLAR-10.7B-Instruct-v1.0")

X

MiMo-V2-Flash (Free)

Xiaomi

MiMo-V2-Flash is an open-source language model developed by Xiaomi featuring a Mixture-of-Experts (MoE) architecture with 309B total parameters and 15B active parameters. It employs hybrid attention m...

Reasoning