AIAIFindr.app

Best AI Chatbots & Large Language Models in 2026

Conversational AI assistants and large language models for chat, reasoning, search and general knowledge work — from ChatGPT and Claude to open-weight models.

368 tools in this category, ranked by quality score.

Cl

Claude

Anthropic

9.6

Claude is Anthropic's family of AI assistants, known for long-context reasoning, careful writing and strong coding…

Chatbots & LLMsCodingWriting
Free tier$20/moView details →
Ch

ChatGPT

OpenAI

9.5

ChatGPT is OpenAI's flagship conversational AI, powering hundreds of millions of weekly users across web, mobile and…

Chatbots & LLMsCodingWriting
Free tier$20/moView details →
Ge

Gemini

Google DeepMind

9.2

Gemini is Google's natively multimodal model family, deeply integrated across Search, Workspace, Android and the Pixel…

Chatbots & LLMsCodingWriting
Free tier$20/moView details →
Me

Meta: Muse Spark 1.2

Meta

9.2

Muse Spark 1.2 is a reasoning model from Meta, designed for complex agentic tasks.

Video & AudioChatbots & LLMs
Paid$1.25/1M tokensView details →
Qw

Qwen: Qwen3.8 Max

Qwen

9.2

Qwen3.8 Max is the flagship model in Alibaba's Qwen3.8 series, the general-availability successor to the Qwen3.8 Max…

Chatbots & LLMs
Paid$2.00/1M tokensView details →
Qw

Qwen: Qwen3.7 Flash

Qwen

9.2

Qwen3.7 Flash is a vision-language reasoning model from Alibaba. It is suited for multimodal agents, visual coding…

Chatbots & LLMs
Paid$0.03/1M tokensView details →
Cl

Claude Opus 5 (Fast)

Anthropic

9.2

Fast-mode variant of [Opus 5](/anthropic/claude-opus-5) - identical capabilities with higher output speed at 2x pricing…

Chatbots & LLMs
Paid$10.00/1M tokensView details →
Cl

Claude Opus 5 (batch)

Anthropic

9.2

Claude Opus 5 is Anthropic’s flagship model for demanding reasoning, coding, and long-horizon agentic work.

Chatbots & LLMs
Paid$2.50/1M tokensView details →
Go

Google: Gemini 3.6 Flash (batch)

Google

9.2

Gemini 3.6 Flash is a high-efficiency model from Google for coding, agentic workflows, and web and app development.

Video & AudioChatbots & LLMs
Paid$0.75/1M tokensView details →
Go

Google: Gemini 3.5 Flash Lite (batch)

Google

9.2

Gemini 3.5 Flash Lite is a high-efficiency model from Google with upgraded agentic capabilities.

Video & AudioChatbots & LLMs
Paid$0.15/1M tokensView details →
Mo

MoonshotAI: Kimi K3

Moonshot AI

9.2

Kimi K3 is a 2.8T parameter open-weight multimodal reasoning model from Moonshot AI.

Chatbots & LLMs
Paid$3.00/1M tokensView details →
Me

Meta: Muse Spark 1.1

Meta

9.2

Muse Spark 1.1 is a multimodal reasoning model from Meta, built for agentic tasks.

Video & AudioChatbots & LLMs
Paid$1.25/1M tokensView details →
Op

OpenAI: GPT-5.6 Luna Pro

OpenAI

9.2

GPT-5.6 Luna Pro is the same underlying model as [GPT-5.6 Luna](https://openrouter.ai/openai/gpt-5.6-luna), served with…

Chatbots & LLMs
Paid$0.10/1M tokensView details →
Op

OpenAI: GPT-5.6 Luna

OpenAI

9.2

GPT-5.6 Luna is a fast, cost-efficient model in OpenAI's GPT-5.6 series. It is suited for high-volume…

Chatbots & LLMs
Paid$0.10/1M tokensView details →
Op

OpenAI: GPT-5.6 Terra Pro

OpenAI

9.2

GPT-5.6 Terra Pro is the same underlying model as [GPT-5.6 Terra](https://openrouter.ai/openai/gpt-5.6-terra), served…

Chatbots & LLMs
Paid$1.00/1M tokensView details →
Op

OpenAI: GPT-5.6 Terra

OpenAI

9.2

GPT-5.6 Terra is a balanced model in OpenAI's GPT-5.6 series, positioned between the flagship Sol tier and the…

Chatbots & LLMs
Paid$1.00/1M tokensView details →
Op

OpenAI: GPT-5.6 Sol Pro (batch)

OpenAI

9.2

GPT-5.6 Sol Pro is the same underlying model as [GPT-5.6 Sol](https://openrouter.ai/openai/gpt-5.6-sol), served with…

Chatbots & LLMs
Paid$2.50/1M tokensView details →
Op

OpenAI: GPT-5.6 Sol (batch)

OpenAI

9.2

GPT-5.6 Sol is the flagship model in OpenAI's GPT-5.6 series. It is suited for complex reasoning, coding, and agentic…

Chatbots & LLMs
Paid$2.50/1M tokensView details →
An

Anthropic: Claude Sonnet 5 (batch)

Anthropic

9.2

Sonnet 5 is Anthropic's most capable Sonnet-class model, with frontier performance across coding, agents, and…

Chatbots & LLMs
Paid$1.00/1M tokensView details →
Sa

Sakana: Fugu Ultra

Sakana AI

9.2

Fugu Ultra is the higher-performance model in Sakana AI's Fugu family. Rather than a single monolithic model, Fugu is a…

Chatbots & LLMs
Paid$5.00/1M tokensView details →
An

Anthropic: Claude Fable Latest

~anthropic

9.2

This model always redirects to the latest model in the Claude Fable family.

Chatbots & LLMs
Paid$10.00/1M tokensView details →
An

Anthropic: Claude Fable 5 (batch)

Anthropic

9.2

Claude Fable 5 is a Mythos-class model from Anthropic, built for autonomous knowledge work and coding.

Chatbots & LLMs
Paid$5.00/1M tokensView details →
Qw

Qwen: Qwen3.7 Plus

Qwen

9.2

Qwen3.7-Plus is a cost-effective model in Alibaba's Qwen3.7 series. It supports text and image input with text output…

Chatbots & LLMs
Paid$0.32/1M tokensView details →
An

Anthropic: Claude Opus 4.8 (Fast)

Anthropic

9.2

Fast-mode variant of [Opus 4.8](/anthropic/claude-opus-4.8) - identical capabilities with higher output speed at 2x…

Chatbots & LLMs
Paid$10.00/1M tokensView details →
An

Anthropic: Claude Opus 4.8 (batch)

Anthropic

9.2

Claude Opus 4.8 is Anthropic's most capable generally available model in the Opus family.

Chatbots & LLMs
Paid$2.50/1M tokensView details →
Go

Google: Gemini 3.5 Flash (batch)

Google

9.2

Gemini 3.5 Flash is Google's high-efficiency multimodal model, bringing near-Pro level coding and reasoning at…

Video & AudioChatbots & LLMs
Paid$0.75/1M tokensView details →
An

Anthropic: Claude Opus 4.7 (Fast)

Anthropic

9.2

Fast-mode variant of [Opus 4.7](/anthropic/claude-opus-4.7) - identical capabilities with higher output speed at…

Chatbots & LLMs
Paid$30.00/1M tokensView details →
Go

Google: Gemini 3.1 Flash Lite (batch)

Google

9.2

Gemini 3.1 Flash Lite is Google’s GA high-efficiency multimodal model optimized for low-latency, high-volume workloads.

Video & AudioChatbots & LLMs
Paid$0.12/1M tokensView details →
Sp

SpaceXAI: Grok 4.3

xAI

9.2

Grok 4.3 is a reasoning model from SpaceXAI. It accepts text and image inputs with text output, and is suited for…

Chatbots & LLMs
Paid$1.25/1M tokensView details →
Go

Google Gemini Pro Latest

~google

9.2

This model always redirects to the latest model in the Google Gemini Pro family.

Video & AudioChatbots & LLMs
Paid$2.00/1M tokensView details →
Mo

MoonshotAI Kimi Latest

~moonshotai

9.2

This model always redirects to the latest model in the MoonshotAI Kimi family.

Chatbots & LLMs
Paid$2.80/1M tokensView details →
Go

Google Gemini Flash Latest

~google

9.2

This model always redirects to the latest model in the Google Gemini Flash family.

Video & AudioChatbots & LLMs
Paid$1.50/1M tokensView details →
An

Anthropic Claude Sonnet Latest

~anthropic

9.2

This model always redirects to the latest model in the Anthropic Claude Sonnet family.

Chatbots & LLMs
Paid$2.00/1M tokensView details →
Op

OpenAI GPT Latest

~openai

9.2

This model always redirects to the latest model in the OpenAI GPT family.

Chatbots & LLMs
Paid$5.00/1M tokensView details →
Qw

Qwen: Qwen3.5 Plus 2026-04-20

Qwen

9.2

Qwen3.5 Plus (April 2026) is a large-scale multimodal language model from Alibaba.

Chatbots & LLMs
Paid$0.30/1M tokensView details →
Qw

Qwen: Qwen3.6 Flash

Qwen

9.2

Qwen3.6 Flash is a fast, efficient language model from Alibaba's Qwen 3.6 series.

Chatbots & LLMs
Paid$0.19/1M tokensView details →
Op

OpenAI: GPT-5.5 Pro (batch)

OpenAI

9.2

GPT-5.5 Pro is OpenAI’s high-capability model optimized for deep reasoning and accuracy on complex, high-stakes…

Chatbots & LLMs
Paid$15.00/1M tokensView details →
Op

OpenAI: GPT-5.5 (batch)

OpenAI

9.2

GPT-5.5 is OpenAI’s frontier model designed for complex professional workloads, building on GPT-5.4 with stronger…

Chatbots & LLMs
Paid$2.50/1M tokensView details →
Xi

Xiaomi: MiMo-V2.5

Xiaomi

9.2

MiMo-V2.5 is a native omnimodal model by Xiaomi. It delivers Pro-level agentic performance at roughly half the…

Video & AudioChatbots & LLMs
Paid$0.14/1M tokensView details →
An

Anthropic: Claude Opus Latest

~anthropic

9.2

This model always redirects to the latest model in the Claude Opus family.

Chatbots & LLMs
Paid$5.00/1M tokensView details →
An

Anthropic: Claude Opus 4.7 (batch)

Anthropic

9.2

Opus 4.7 is the next generation of Anthropic's Opus family, built for long-running, asynchronous agents.

Chatbots & LLMs
Paid$2.50/1M tokensView details →
Qw

Qwen: Qwen3.6 Plus

Qwen

9.2

Qwen 3.6 Plus builds on a hybrid architecture that combines efficient linear attention with sparse mixture-of-experts…

Chatbots & LLMs
Paid$0.33/1M tokensView details →
Sp

SpaceXAI: Grok 4.20

xAI

9.2

Grok 4.20 is a reasoning model from SpaceXAI with industry-leading speed and agentic tool calling capabilities.

Chatbots & LLMs
Paid$1.25/1M tokensView details →
Op

OpenAI: GPT-5.4 Pro (batch)

OpenAI

9.2

GPT-5.4 Pro is OpenAI's most advanced model, building on GPT-5.4's unified architecture with enhanced reasoning…

Chatbots & LLMs
Paid$15.00/1M tokensView details →
Op

OpenAI: GPT-5.4 (batch)

OpenAI

9.2

GPT-5.4 is OpenAI’s latest frontier model, unifying the Codex and GPT lines into a single system.

Chatbots & LLMs
Paid$1.25/1M tokensView details →
Go

Google: Gemini 3.1 Flash Lite Preview

Google

9.2

Gemini 3.1 Flash Lite Preview is Google's high-efficiency model optimized for high-volume use cases.

Video & AudioChatbots & LLMs
Paid$0.25/1M tokensView details →
Qw

Qwen: Qwen3.5-Flash

Qwen

9.2

The Qwen3.5 native vision-language Flash models are built on a hybrid architecture that integrates a linear attention…

Chatbots & LLMs
Paid$0.07/1M tokensView details →
Go

Google: Gemini 3.1 Pro Preview Custom Tools

Google

9.2

Gemini 3.1 Pro Preview Custom Tools is a variant of Gemini 3.1 Pro that improves tool selection behavior by preventing…

Video & AudioChatbots & LLMs
Paid$2.00/1M tokensView details →
Go

Google: Gemini 3.1 Pro Preview (batch)

Google

9.2

Gemini 3.1 Pro Preview is Google’s frontier reasoning model, delivering enhanced software engineering performance…

Video & AudioChatbots & LLMs
Paid$1.00/1M tokensView details →
An

Anthropic: Claude Sonnet 4.6 (batch)

Anthropic

9.2

Sonnet 4.6 is Anthropic's most capable Sonnet-class model yet, with frontier performance across coding, agents, and…

Chatbots & LLMs
Paid$1.50/1M tokensView details →
Qw

Qwen: Qwen3.5 Plus 2026-02-15

Qwen

9.2

The Qwen3.5 native vision-language series Plus models are built on a hybrid architecture that integrates linear…

Chatbots & LLMs
Paid$0.26/1M tokensView details →
An

Anthropic: Claude Opus 4.6 (batch)

Anthropic

9.2

Opus 4.6 is Anthropic’s strongest model for coding and long-running professional tasks.

Chatbots & LLMs
Paid$2.50/1M tokensView details →
Go

Google: Gemini 3 Flash Preview (batch)

Google

9.2

Gemini 3 Flash Preview is a high speed, high value thinking model designed for agentic workflows, multi turn chat, and…

Video & AudioChatbots & LLMs
Paid$0.25/1M tokensView details →
Am

Amazon: Nova 2 Lite

Amazon

9.2

Nova 2 Lite is a fast, cost-effective reasoning model for everyday workloads that can process text, images, and videos…

Chatbots & LLMs
Paid$0.30/1M tokensView details →
An

Anthropic: Claude Sonnet 4.5 (batch)

Anthropic

9.2

Claude Sonnet 4.5 is Anthropic’s most advanced Sonnet model to date, optimized for real-world agents and coding…

Chatbots & LLMs
Paid$1.50/1M tokensView details →
Go

Google: Gemini 2.5 Flash Lite (batch)

Google

9.2

Gemini 2.5 Flash-Lite is a lightweight reasoning model in the Gemini 2.5 family, optimized for ultra-low latency and…

Video & AudioChatbots & LLMs
Paid$0.05/1M tokensView details →
Go

Google: Gemini 2.5 Flash (batch)

Google

9.2

Gemini 2.5 Flash is Google's state-of-the-art workhorse model, specifically designed for advanced reasoning, coding…

Video & AudioChatbots & LLMs
Paid$0.15/1M tokensView details →
Go

Google: Gemini 2.5 Pro (batch)

Google

9.2

Gemini 2.5 Pro is Google’s state-of-the-art AI model designed for advanced reasoning, coding, mathematics, and…

Video & AudioChatbots & LLMs
Paid$0.62/1M tokensView details →
Go

Google: Gemini 2.5 Pro Preview 06-05

Google

9.2

Gemini 2.5 Pro is Google’s state-of-the-art AI model designed for advanced reasoning, coding, mathematics, and…

Video & AudioChatbots & LLMs
Paid$1.25/1M tokensView details →
An

Anthropic: Claude Sonnet 4

Anthropic

9.2

Claude Sonnet 4 significantly enhances the capabilities of its predecessor, Sonnet 3.7, excelling in both coding and…

Chatbots & LLMs
Paid$3.00/1M tokensView details →
Go

Google: Gemini 2.5 Pro Preview 05-06

Google

9.2

Gemini 2.5 Pro is Google’s state-of-the-art AI model designed for advanced reasoning, coding, mathematics, and…

Video & AudioChatbots & LLMs
Paid$1.25/1M tokensView details →
De

DeepSeek

DeepSeek

8.9

DeepSeek is a Chinese AI lab that stunned the industry with frontier-level reasoning models at a fraction of typical…

Chatbots & LLMsCoding
Free tierPay per tokenView details →
Qw

Qwen: Qwen3.8 2.4T A95B

Qwen

8.9

Qwen3.8 2.4T A95B is an open-weight sparse mixture-of-experts model from Qwen and the open-weight variant of [Qwen3.8…

Chatbots & LLMs
Paid$2.00/1M tokensView details →
De

DeepSeek: DeepSeek V4 Pro 0813

DeepSeek

8.9

DeepSeek V4 Pro 0813 is a large-scale mixture-of-experts model from DeepSeek. This is the GA release of DeepSeek V4 Pro.

Chatbots & LLMs
Paid$0.43/1M tokensView details →
NV

NVIDIA: Nemotron 3.5 Lightning (free)

NVIDIA

8.9

NVIDIA Nemotron 3.5 Lightning is an open mixture-of-experts model from NVIDIA, with 3B active parameters out of 30B…

Chatbots & LLMs
Free tierFreeView details →
De

DeepSeek V4 Flash Latest

~deepseek

8.9

This model always redirects to the latest model in the DeepSeek V4 Flash family.

Chatbots & LLMs
Paid$0.08/1M tokensView details →
De

DeepSeek: DeepSeek V4 Flash 0731

DeepSeek

8.9

DeepSeek V4 Flash 0731 is a sparse mixture-of-experts model from DeepSeek, with 13B active parameters out of 284B…

Chatbots & LLMs
Paid$0.08/1M tokensView details →
Me

Meituan: LongCat 2.0

Meituan

8.9

LongCat 2.0 is a sparse mixture-of-experts language model from Meituan, with 48B active parameters out of 1.6T total.

Chatbots & LLMs
Paid$0.30/1M tokensView details →
NV

NVIDIA: Nemotron 3 Ultra (free)

NVIDIA

8.9

NVIDIA Nemotron 3 Ultra is an open frontier-reasoning and orchestration model from NVIDIA, with 55B active parameters…

Chatbots & LLMs
Free tierFreeView details →
Qw

Qwen: Qwen3.7 Max

Qwen

8.9

Qwen3.7-Max is the flagship model in Alibaba's Qwen3.7 series. It supports text input and output and is designed for…

Chatbots & LLMs
Paid$1.48/1M tokensView details →
De

DeepSeek: DeepSeek V4 Pro

DeepSeek

8.9

DeepSeek V4 Pro is a large-scale Mixture-of-Experts model from DeepSeek with 1.6T total parameters and 49B activated…

Chatbots & LLMs
Paid$1.17/1M tokensView details →
De

DeepSeek: DeepSeek V4 Flash 0423

DeepSeek

8.9

DeepSeek V4 Flash is an efficiency-optimized Mixture-of-Experts model from DeepSeek with 284B total parameters and 13B…

Chatbots & LLMs
Paid$0.14/1M tokensView details →
Xi

Xiaomi: MiMo-V2.5-Pro

Xiaomi

8.9

MiMo-V2.5-Pro is Xiaomi’s flagship model, delivering strong performance in general agentic capabilities, complex…

Chatbots & LLMs
Paid$0.43/1M tokensView details →
Qw

Qwen: Qwen3 Coder Plus

Qwen

8.9

Qwen3 Coder Plus is Alibaba's proprietary version of the Open Source Qwen3 Coder 480B A35B.

CodingChatbots & LLMs
Paid$0.65/1M tokensView details →
Qw

Qwen: Qwen Plus 0728

Qwen

8.9

Qwen Plus 0728, based on the Qwen3 foundation model, is a 1 million context hybrid reasoning model with a balanced…

Chatbots & LLMs
Paid$0.26/1M tokensView details →
Mi

MiniMax: MiniMax M1

MiniMax

8.9

MiniMax-M1 is a large-scale, open-weight reasoning model designed for extended context and high-efficiency inference.

Chatbots & LLMs
Paid$0.55/1M tokensView details →
Pe

Perplexity

Perplexity AI

8.8

Perplexity is an AI answer engine that combines live web search with large language models to deliver cited, up-to-date…

Chatbots & LLMsWriting
Free tier$20/moView details →
Sp

SpaceXAI: Grok 4.20 Multi-Agent

xAI

8.8

Grok 4.20 Multi-Agent is a variant of SpaceXAI’s Grok 4.20 designed for collaborative, agent-based workflows.

Chatbots & LLMs
Paid$1.25/1M tokensView details →
Au

Auto Router (Beta)

Openrouter

8.7

Auto Router (Beta) is a task-aware router from OpenRouter. It classifies each request, then routes it the [most popular…

Image GenerationVideo & AudioChatbots & LLMs
PaidVariableView details →
Am

Amazon: Nova Premier 1.0

Amazon

8.7

Amazon Nova Premier is the most capable of Amazon’s multimodal models for complex reasoning tasks and for use as the…

Chatbots & LLMs
Paid$2.50/1M tokensView details →
Op

OpenAI: GPT-4.1 (batch)

OpenAI

8.7

GPT-4.1 is a flagship large language model optimized for advanced instruction following, real-world software…

Chatbots & LLMs
Paid$1.00/1M tokensView details →
Op

OpenAI: GPT-4.1 Mini (batch)

OpenAI

8.7

GPT-4.1 Mini is a mid-sized model delivering performance competitive with GPT-4o at substantially lower latency and…

Chatbots & LLMs
Paid$0.20/1M tokensView details →
Op

OpenAI: GPT-4.1 Nano (batch)

OpenAI

8.7

For tasks that demand low latency, GPT‑4.1 nano is the fastest and cheapest model in the GPT-4.1 series.

Chatbots & LLMs
Paid$0.05/1M tokensView details →
Me

Meta: Llama 4 Maverick

Meta

8.7

Llama 4 Maverick 17B Instruct (128E) is a high-capacity multimodal language model from Meta, built on a…

Chatbots & LLMs
Paid$0.20/1M tokensView details →
Me

Meta: Llama 4 Scout

Meta

8.7

Llama 4 Scout 17B Instruct (16E) is a mixture-of-experts (MoE) language model developed by Meta, activating 17 billion…

Chatbots & LLMs
Paid$0.10/1M tokensView details →
Au

Auto Router

Openrouter

8.7

Your prompt will be processed by a meta-model and routed to one of dozens of models (see below), optimizing for the…

Image GenerationVideo & AudioChatbots & LLMs
PaidVariableView details →
Qw

Qwen

Alibaba Cloud

8.6

Qwen is Alibaba's series of open-weight models spanning chat, coding, vision and math.

Chatbots & LLMsCoding
Free tierPay per tokenView details →
Mi

Mistral

Mistral AI

8.5

Mistral AI is a European lab offering both open-weight and commercial models that punch well above their size.

Chatbots & LLMsCoding
Free tierPay per tokenView details →
Sp

SpaceXAI: Grok 4.6

xAI

8.5

Grok 4.6 is SpaceXAI's smartest model with frontier performance on coding, knowledge work, and STEM.

Chatbots & LLMs
Paid$2.00/1M tokensView details →
Th

Thinking Machines: Inkling Small

Thinkingmachines

8.5

Inkling Small is an open-weight multimodal mixture-of-experts model from Thinking Machines Lab, with 12B active…

Video & AudioChatbots & LLMs
Paid$0.45/1M tokensView details →
Th

Thinking Machines: Inkling (batch)

Thinkingmachines

8.5

Inkling is an open-weight multimodal mixture-of-experts model from Thinking Machines Lab, with 41B active parameters…

Video & AudioChatbots & LLMs
Paid$0.50/1M tokensView details →
Sp

SpaceXAI: Grok 4.5

xAI

8.5

Grok 4.5 is a model from SpaceXAI with frontier performance on coding, knowledge work, and STEM.

Chatbots & LLMs
Paid$2.00/1M tokensView details →
xA

xAI: Grok Latest

~x Ai

8.5

This model always redirects to the latest Grok model from xAI.

Chatbots & LLMs
Paid$2.00/1M tokensView details →
Mi

MiniMax: MiniMax M3 (batch)

MiniMax

8.5

MiniMax-M3 is a multimodal foundation model from MiniMax. It supports text, image, and video inputs with text output, a…

Chatbots & LLMs
Paid$0.15/1M tokensView details →
Gr

Grok

xAI

8.4

Grok is xAI's assistant, tightly integrated with the X platform and trained with a distinctively candid, witty persona.

Chatbots & LLMsWriting
Free tier$8/moView details →
Qw

Qwen: Qwen3 Coder Flash

Qwen

8.4

Qwen3 Coder Flash is Alibaba's fast and cost efficient version of their proprietary Qwen3 Coder Plus.

CodingChatbots & LLMs
Paid$0.20/1M tokensView details →
Qw

Qwen: Qwen-Plus

Qwen

8.4

Qwen-Plus, based on the Qwen2.5 foundation model, is a 131K context model with a balanced performance, speed, and cost…

Chatbots & LLMs
Paid$0.26/1M tokensView details →
Th

TheDrummer: UnslopNemo 12B

TheDrummer

8.4

UnslopNemo v4.1 is the latest addition from the creator of Rocinante, designed for adventure writing and role-play…

Chatbots & LLMs
Paid$0.40/1M tokensView details →
Ll

Llama

Meta

8.3

Llama is Meta's family of open-weight large language models that became the de-facto foundation for the open-source AI…

Chatbots & LLMsCoding
Free tierPay per tokenView details →
Op

OpenAI GPT Mini Latest

~openai

8.3

This model always redirects to the latest model in the OpenAI GPT Mini family.

Chatbots & LLMs
Paid$0.75/1M tokensView details →
Go

Google: Lyria 3 Pro Preview

Google

8.3

Full-length songs are priced at $0.08 per song. Lyria 3 is Google's family of music generation models, available…

Video & AudioChatbots & LLMs
Free tierFreeView details →
Go

Google: Lyria 3 Clip Preview

Google

8.3

30 second duration clips are priced at $0.04 per clip. Lyria 3 is Google's family of music generation models, available…

Video & AudioChatbots & LLMs
Free tierFreeView details →
Op

OpenAI: GPT-5.4 Nano (batch)

OpenAI

8.3

GPT-5.4 nano is the most lightweight and cost-efficient variant of the GPT-5.4 family, optimized for speed-critical and…

Chatbots & LLMs
Paid$0.10/1M tokensView details →
Op

OpenAI: GPT-5.4 Mini (batch)

OpenAI

8.3

GPT-5.4 mini brings the core capabilities of GPT-5.4 to a faster, more efficient model optimized for high-throughput…

Chatbots & LLMs
Paid$0.38/1M tokensView details →
Op

OpenAI: GPT-5.3-Codex

OpenAI

8.3

GPT-5.3-Codex is OpenAI’s most advanced agentic coding model, combining the frontier software engineering performance…

CodingChatbots & LLMs
Paid$1.75/1M tokensView details →
Op

OpenAI: GPT-5.2-Codex

OpenAI

8.3

GPT-5.2-Codex is an upgraded version of GPT-5.1-Codex optimized for software engineering and coding workflows.

CodingChatbots & LLMs
Paid$1.75/1M tokensView details →
Op

OpenAI: GPT-5.2 Pro (batch)

OpenAI

8.3

GPT-5.2 Pro is OpenAI’s most advanced model, offering major improvements in agentic coding and long context performance…

Chatbots & LLMs
Paid$10.50/1M tokensView details →
Op

OpenAI: GPT-5.2 (batch)

OpenAI

8.3

GPT-5.2 is the latest frontier-grade model in the GPT-5 series, offering stronger agentic and long context perfomance…

Chatbots & LLMs
Paid$0.88/1M tokensView details →
Op

OpenAI: GPT-5.1-Codex-Max

OpenAI

8.3

GPT-5.1-Codex-Max is OpenAI’s latest agentic coding model, designed for long-running, high-context software development…

CodingChatbots & LLMs
Paid$1.25/1M tokensView details →
Op

OpenAI: GPT-5.1 (batch)

OpenAI

8.3

GPT-5.1 is the latest frontier-grade model in the GPT-5 series, offering stronger general-purpose reasoning, improved…

Chatbots & LLMs
Paid$0.62/1M tokensView details →
Op

OpenAI: GPT-5.1-Codex

OpenAI

8.3

GPT-5.1-Codex is a specialized version of GPT-5.1 optimized for software engineering and coding workflows.

CodingChatbots & LLMs
Paid$1.25/1M tokensView details →
Op

OpenAI: GPT-5.1-Codex-Mini

OpenAI

8.3

GPT-5.1-Codex-Mini is a smaller and faster version of GPT-5.1-Codex

CodingChatbots & LLMs
Paid$0.25/1M tokensView details →
Op

OpenAI: GPT-5 Pro (batch)

OpenAI

8.3

GPT-5 Pro is OpenAI’s most advanced model, offering major improvements in reasoning, code quality, and user experience.

Chatbots & LLMs
Paid$7.50/1M tokensView details →
Op

OpenAI: GPT-5 Codex (batch)

OpenAI

8.3

GPT-5-Codex is a specialized version of GPT-5 optimized for software engineering and coding workflows.

CodingChatbots & LLMs
Paid$0.62/1M tokensView details →
Op

OpenAI: GPT-5 (batch)

OpenAI

8.3

GPT-5 is OpenAI’s most advanced model, offering major improvements in reasoning, code quality, and user experience.

Chatbots & LLMs
Paid$0.62/1M tokensView details →
Op

OpenAI: GPT-5 Mini (batch)

OpenAI

8.3

GPT-5 Mini is a compact version of GPT-5, designed to handle lighter-weight reasoning tasks.

Chatbots & LLMs
Paid$0.12/1M tokensView details →
Op

OpenAI: GPT-5 Nano (batch)

OpenAI

8.3

GPT-5-Nano is the smallest and fastest variant in the GPT-5 system, optimized for developer tools, rapid interactions…

Chatbots & LLMs
Paid$0.02/1M tokensView details →
Mi

MiniMax: MiniMax-01

MiniMax

8.3

MiniMax-01 is a combines MiniMax-Text-01 for text generation and MiniMax-VL-01 for image understanding.

Chatbots & LLMs
Paid$0.20/1M tokensView details →
Co

Cohere Command

Cohere

8.2

Cohere's Command models are purpose-built for enterprise retrieval-augmented generation and agentic workflows.

Chatbots & LLMs
Free tierPay per tokenView details →
Up

Upstage: Solar Pro 4

Upstage

8.2

Solar Pro 4 is a large language model from Upstage. It is suited for agentic workflows, office productivity…

Chatbots & LLMs
Paid$0.03/1M tokensView details →
Z.

Z.ai: GLM 5.2 (batch)

Z.AI

8.2

GLM 5.2 is a large-scale reasoning model from Z.ai. It supports text input and output with a 1M-token context window…

Chatbots & LLMs
Paid$0.70/1M tokensView details →
By

ByteDance Seed: Seed 2.1 Turbo

Bytedance Seed

8.1

Seed 2.1 Turbo is a multimodal model from ByteDance Seed for coding and long-horizon agent workflows.

Chatbots & LLMs
Paid$0.50/1M tokensView details →
By

ByteDance Seed: Seed-2.0-Code

Bytedance Seed

8.1

Seed 2.0 Code is a model from ByteDance Seed optimized for agentic coding.

CodingChatbots & LLMs
Paid$0.50/1M tokensView details →
Sa

Sakana: Sakana Namazu

Sakana AI

8.1

Sakana Namazu is a Japanese-specialized reasoning model from Sakana AI, based on Kimi K2.6 with additional training for…

Chatbots & LLMs
Paid$0.95/1M tokensView details →
Ne

Nex AGI: Nex-N2-Mini

Nex Agi

8.1

Nex-N2-Mini is an open-source agentic mixture-of-experts model from Nex AGI, the smaller sibling in the Nex-N2 series.

Chatbots & LLMs
Paid$0.02/1M tokensView details →
Mo

MoonshotAI: Kimi K2.7 Code (batch)

Moonshot AI

8.1

MoonshotAI: Kimi K2.7 Code is a coding-focused model in Moonshot AI's Kimi K2 family, built to complete end-to-end…

CodingChatbots & LLMs
Paid$0.47/1M tokensView details →
Ne

Nex AGI: Nex-N2-Pro

Nex Agi

8.1

Nex-N2-Pro is an agentic mixture-of-experts model from Nex AGI, with 17B active parameters out of 397B total.

Chatbots & LLMs
Paid$0.25/1M tokensView details →
St

StepFun: Step 3.7 Flash

Stepfun

8.1

Step 3.7 Flash is StepFun's latest high-efficiency multimodal Mixture-of-Experts model.

Chatbots & LLMs
Paid$0.20/1M tokensView details →
Sp

SpaceXAI: Grok Build 0.1

xAI

8.1

Grok Build 0.1 is SpaceXAI’s fast coding model trained specifically for agentic software engineering workflows.

Chatbots & LLMs
Paid$1.00/1M tokensView details →
Mi

Mistral: Mistral Medium 3.5

Mistral AI

8.1

Mistral Medium 3.5 is a dense 128B instruction-following model from Mistral AI.

Chatbots & LLMs
Paid$1.50/1M tokensView details →
NV

NVIDIA: Nemotron 3 Nano Omni (free)

NVIDIA

8.1

NVIDIA Nemotron™ 3 Nano Omni is a 30B-A3B open multimodal model designed to function as a perception and context…

Video & AudioChatbots & LLMs
Free tierFreeView details →
Qw

Qwen: Qwen3.6 35B A3B

Qwen

8.1

Qwen3.6-35B-A3B is an open-weight multimodal model from Alibaba Cloud with 35 billion total parameters and 3 billion…

Chatbots & LLMs
Paid$0.15/1M tokensView details →
Qw

Qwen: Qwen3.6 27B

Qwen

8.1

Qwen3.6 27B is a dense 27-billion-parameter language model from the Qwen Team at Alibaba, released in April 2026.

Chatbots & LLMs
Paid$0.60/1M tokensView details →
Mo

MoonshotAI: Kimi K2.6

Moonshot AI

8.1

Kimi K2.6 is Moonshot AI's next-generation multimodal model, designed for long-horizon coding, coding-driven UI/UX…

Chatbots & LLMs
Paid$0.95/1M tokensView details →
Go

Google: Gemma 4 26B A4B (free)

Google

8.1

Gemma 4 26B A4B IT is an instruction-tuned Mixture-of-Experts (MoE) model from Google DeepMind.

Chatbots & LLMs
Free tierFreeView details →
Go

Google: Gemma 4 31B (free)

Google

8.1

Gemma 4 31B Instruct is Google DeepMind's 30.7B dense multimodal model supporting text and image input with text…

Chatbots & LLMs
Free tierFreeView details →
Mi

Mistral: Mistral Small 4

Mistral AI

8.1

Mistral Small 4 is the next major release in the Mistral Small family, unifying the capabilities of several flagship…

Chatbots & LLMs
Paid$0.15/1M tokensView details →
By

ByteDance Seed: Seed-2.0-Lite

Bytedance Seed

8.1

Seed-2.0-Lite is a versatile, cost‑efficient enterprise workhorse that delivers strong multimodal and agent…

Chatbots & LLMs
Paid$0.25/1M tokensView details →
Qw

Qwen: Qwen3.5-9B

Qwen

8.1

Qwen3.5-9B is a multimodal foundation model from the Qwen3.5 family, designed to deliver strong reasoning, coding, and…

Chatbots & LLMs
Paid$0.10/1M tokensView details →
By

ByteDance Seed: Seed-2.0-Mini

Bytedance Seed

8.1

Seed-2.0-mini targets latency-sensitive, high-concurrency, and cost-sensitive scenarios, emphasizing fast response and…

Chatbots & LLMs
Paid$0.10/1M tokensView details →
Qw

Qwen: Qwen3.5-35B-A3B

Qwen

8.1

The Qwen3.5 Series 35B-A3B is a native vision-language model designed with a hybrid architecture that integrates linear…

Chatbots & LLMs
Paid$0.25/1M tokensView details →
Qw

Qwen: Qwen3.5-27B

Qwen

8.1

The Qwen3.5 27B native vision-language Dense model incorporates a linear attention mechanism, delivering fast response…

Chatbots & LLMs
Paid$0.20/1M tokensView details →
Qw

Qwen: Qwen3.5-122B-A10B

Qwen

8.1

The Qwen3.5 122B-A10B native vision-language model is built on a hybrid architecture that integrates a linear attention…

Chatbots & LLMs
Paid$0.29/1M tokensView details →
Qw

Qwen: Qwen3.5 397B A17B

Qwen

8.1

The Qwen3.5 series 397B-A17B native vision-language model is built on a hybrid architecture that integrates a linear…

Chatbots & LLMs
Paid$0.45/1M tokensView details →
Mo

MoonshotAI: Kimi K2.5

Moonshot AI

8.1

Kimi K2.5 is Moonshot AI's native multimodal model, delivering state-of-the-art visual coding capability and a…

Chatbots & LLMs
Paid$0.57/1M tokensView details →
By

ByteDance Seed: Seed 1.6 Flash

Bytedance Seed

8.1

Seed 1.6 Flash is an ultra-fast multimodal deep thinking model by ByteDance Seed, supporting both text and visual…

Chatbots & LLMs
Paid$0.07/1M tokensView details →
By

ByteDance Seed: Seed 1.6

Bytedance Seed

8.1

Seed 1.6 is a general-purpose model released by the ByteDance Seed team. It incorporates multimodal capabilities and…

Chatbots & LLMs
Paid$0.25/1M tokensView details →
Qw

Qwen: Qwen3 VL 30B A3B Thinking

Qwen

8.1

Qwen3-VL-30B-A3B-Thinking is a multimodal model that unifies strong text generation with visual understanding for…

Chatbots & LLMs
Paid$0.20/1M tokensView details →
Po

Poe

Quora

8.0

Poe by Quora is an aggregator that gives access to dozens of AI models — including GPT, Claude, Gemini and many image…

Chatbots & LLMs
Free tier$19.99/moView details →
Op

OpenRouter: Fusion

Openrouter

8.0

Fusion turns your prompt into a small multi-model deliberation. A panel of expert models (see below) analyzes your…

Chatbots & LLMs
PaidVariableView details →
An

Anthropic Claude Haiku Latest

~anthropic

8.0

This model always redirects to the latest model in the Anthropic Claude Haiku family.

Chatbots & LLMs
Paid$1.00/1M tokensView details →
Pa

Pareto Code Router

Openrouter

8.0

The Pareto Router maintains a tiered shortlist of strong coding models, ranked by [Artificial…

CodingChatbots & LLMs
PaidVariableView details →
Z.

Z.ai: GLM 5V Turbo

Z.AI

8.0

GLM-5V-Turbo is Z.ai’s first native multimodal agent foundation model, built for vision-based coding and agent-driven…

Chatbots & LLMs
Paid$1.20/1M tokensView details →
Wr

Writer: Palmyra X5

Writer

8.0

Palmyra X5 is Writer's most advanced model, purpose-built for building and scaling AI agents across the enterprise.

Chatbots & LLMs
Paid$0.60/1M tokensView details →
An

Anthropic: Claude Opus 4.5 (batch)

Anthropic

8.0

Claude Opus 4.5 is Anthropic’s frontier reasoning model optimized for complex software engineering, agentic workflows…

Chatbots & LLMs
Paid$2.50/1M tokensView details →
An

Anthropic: Claude Haiku 4.5 (batch)

Anthropic

8.0

Claude Haiku 4.5 is Anthropic’s fastest and most efficient model, delivering near-frontier intelligence at a fraction…

Chatbots & LLMs
Paid$0.50/1M tokensView details →
An

Anthropic: Claude Opus 4.1 (batch)

Anthropic

8.0

Claude Opus 4.1 is an updated version of Anthropic’s flagship model, offering improved performance in coding…

Chatbots & LLMs
Paid$7.50/1M tokensView details →
Op

OpenAI: o3 Pro (batch)

OpenAI

8.0

The o-series of models are trained with reinforcement learning to think before they answer and perform complex…

Chatbots & LLMs
Paid$10.00/1M tokensView details →
An

Anthropic: Claude Opus 4

Anthropic

8.0

Claude Opus 4 is benchmarked as the world’s best coding model, at time of release, bringing sustained performance on…

Chatbots & LLMs
Paid$15.00/1M tokensView details →
Op

OpenAI: o4 Mini High (batch)

OpenAI

8.0

OpenAI o4-mini-high is the same model as [o4-mini](/openai/o4-mini) with reasoning_effort set to high.

Chatbots & LLMs
Paid$0.55/1M tokensView details →
Op

OpenAI: o3 (batch)

OpenAI

8.0

o3 is a well-rounded and powerful model across domains. It sets a new standard for math, science, coding, and visual…

Chatbots & LLMs
Paid$1.00/1M tokensView details →
Op

OpenAI: o4 Mini (batch)

OpenAI

8.0

OpenAI o4-mini is a compact reasoning model in the o-series, optimized for fast, cost-efficient performance while…

Chatbots & LLMs
Paid$0.55/1M tokensView details →
Op

OpenAI: o1 (batch)

OpenAI

8.0

The latest and strongest model family from OpenAI, o1 is designed to spend more time thinking before responding.

Chatbots & LLMs
Paid$7.50/1M tokensView details →
Ch

Character.ai

Character.AI

7.9

Character.ai lets users chat with millions of user-created AI personas, from historical figures to original fictional…

Chatbots & LLMs
Free tier$9.99/moView details →
Me

Meta: Muse Glimmer 30B

Meta

7.9

Muse Glimmer 30B is a dense, open-weight multimodal model from Meta Superintelligence Labs, distilled from Muse Spark…

Chatbots & LLMs
Paid$0.35/1M tokensView details →
Go

Google: Nano Banana Pro (Gemini 3 Pro Image)

Google

7.9

Nano Banana Pro is Google’s most advanced image-generation and editing model, built on Gemini 3 Pro.

Image GenerationChatbots & LLMs
Paid$2.00/1M tokensView details →
Z.

Z.ai: GLM 4.6V

Z.AI

7.9

GLM-4.6V is a large multimodal model designed for high-fidelity visual understanding and long-context reasoning across…

Chatbots & LLMs
Paid$0.30/1M tokensView details →
NV

NVIDIA: Nemotron Nano 12B 2 VL (free)

NVIDIA

7.9

NVIDIA Nemotron Nano 2 VL is a 12-billion-parameter open multimodal reasoning model designed for video understanding…

Chatbots & LLMs
Free tierFreeView details →
Op

OpenAI: GPT-5 Image Mini

OpenAI

7.9

GPT-5 Image Mini combines OpenAI's advanced language capabilities, powered by [GPT-5…

Image GenerationChatbots & LLMs
Paid$2.50/1M tokensView details →
Qw

Qwen: Qwen3 VL 8B Thinking

Qwen

7.9

Qwen3-VL-8B-Thinking is the reasoning-optimized variant of the Qwen3-VL-8B multimodal model, designed for advanced…

Chatbots & LLMs
Paid$0.18/1M tokensView details →
Op

OpenAI: GPT-5 Image

OpenAI

7.9

[GPT-5](https://openrouter.ai/openai/gpt-5) Image combines OpenAI's GPT-5 model with state-of-the-art image generation…

Image GenerationChatbots & LLMs
Paid$10.00/1M tokensView details →
Qw

Qwen: Qwen3 VL 235B A22B Thinking

Qwen

7.9

Qwen3-VL-235B-A22B Thinking is a multimodal model that unifies strong text generation with visual understanding across…

Chatbots & LLMs
Paid$0.40/1M tokensView details →
Pi

Pi

Inflection AI

7.8

Pi (Personal Intelligence) is a conversational assistant designed around empathy, emotional support and natural…

Chatbots & LLMs
Free tierFreeView details →
Yo

You.com

You.com

7.8

You.com is an AI search and productivity platform that blends web search with multi-model chat and research agents.

Chatbots & LLMsWriting
Free tier$15/moView details →
in

inclusionAI: Ling 3.0 Tiny (free)

Inclusionai

7.8

Ling 3.0 Tiny is a mixture-of-experts model from InclusionAI, with 1.3B active parameters out of 7.9B total.

Chatbots & LLMs
Free tierFreeView details →
Li

Ling-3.0-flash

Inclusionai

7.8

*Ling-3.0-flash* is a *124B-parameter Mixture-of-Experts (MoE) model*, with approximately *5.1B parameters activated…

Chatbots & LLMs
Paid$0.02/1M tokensView details →
Po

Poolside: Laguna S 2.1 (free)

Poolside

7.8

Laguna S 2.1 is the latest coding agent model from [Poolside](<https://poolside.ai/>).

Chatbots & LLMs
Free tierFreeView details →
Te

Tencent: Hy3

Tencent

7.8

Hy3 is a 295B-parameter Mixture-of-Experts model from Tencent (21B active, 192 experts with top-8 routing) built for…

Chatbots & LLMs
Paid$0.13/1M tokensView details →
Po

Poolside: Laguna XS 2.1 (free)

Poolside

7.8

Laguna XS 2.1 is the latest coding agent model in the 33B-A3B category from [Poolside](https://poolside.ai/) and a step…

Chatbots & LLMs
Free tierFreeView details →
Co

Cohere: North Mini Code (free)

Cohere

7.8

North Mini Code is Cohere's first agentic coding model and the debut of its North family.

CodingChatbots & LLMs
Free tierFreeView details →
in

inclusionAI: Ring-2.6-1T

Inclusionai

7.8

Ring-2.6-1T is a 1T-parameter-scale thinking model with 63B active parameters, built for real-world agent workflows…

Chatbots & LLMs
Paid$0.07/1M tokensView details →
Op

OpenAI: GPT Chat Latest

OpenAI

7.8

GPT Chat Latest points to OpenAI's stable API alias `chat-latest` that always resolves to the latest Instant chat model…

Chatbots & LLMs
Paid$5.00/1M tokensView details →
Qw

Qwen: Qwen3.6 Max Preview

Qwen

7.8

Qwen3.6-Max-Preview is a proprietary frontier model from Alibaba Cloud built on a sparse mixture-of-experts…

Chatbots & LLMs
Paid$1.03/1M tokensView details →
Te

Tencent: Hy3 preview

Tencent

7.8

Hy3 preview is a high-efficiency Mixture-of-Experts model from Tencent designed for agentic workflows and production…

Chatbots & LLMs
Paid$0.06/1M tokensView details →
Ar

Arcee AI: Trinity Large Thinking

Arcee Ai

7.8

Trinity Large Thinking is a powerful open source reasoning model from the team at Arcee AI.

Chatbots & LLMs
Paid$0.22/1M tokensView details →
NV

NVIDIA: Nemotron 3 Super (free)

NVIDIA

7.8

NVIDIA Nemotron 3 Super is a 120B-parameter open hybrid MoE model, activating just 12B parameters for maximum compute…

Chatbots & LLMs
Free tierFreeView details →
Qw

Qwen: Qwen3 Max Thinking

Qwen

7.8

Qwen3-Max-Thinking is the flagship reasoning model in the Qwen3 series, designed for high-stakes cognitive tasks that…

Chatbots & LLMs
Paid$0.78/1M tokensView details →
St

StepFun: Step 3.5 Flash

Stepfun

7.8

Step 3.5 Flash is StepFun's most capable open-source foundation model. Built on a sparse Mixture of Experts (MoE)…

Chatbots & LLMs
Paid$0.10/1M tokensView details →
NV

NVIDIA: Nemotron 3 Nano 30B A3B (free)

NVIDIA

7.8

NVIDIA Nemotron 3 Nano 30B A3B is a small language MoE model with highest compute efficiency and accuracy for…

Chatbots & LLMs
Free tierFreeView details →
Mo

MoonshotAI: Kimi K2 Thinking

Moonshot AI

7.8

Kimi K2 Thinking is Moonshot AI’s most advanced open reasoning model to date, extending the K2 series into agentic…

Chatbots & LLMs
Paid$0.60/1M tokensView details →
Qw

Qwen: Qwen3 Max

Qwen

7.8

Qwen3-Max is an updated release built on the Qwen3 series, offering major improvements in reasoning, instruction…

Chatbots & LLMs
Paid$0.78/1M tokensView details →
Qw

Qwen: Qwen3 Next 80B A3B Thinking

Qwen

7.8

Qwen3-Next-80B-A3B-Thinking is a reasoning-first chat model in the Qwen3-Next line that outputs structured “thinking”…

Chatbots & LLMs
Paid$0.15/1M tokensView details →
Z.

Z.ai: GLM 4.5V

Z.AI

7.8

GLM-4.5V is a vision-language foundation model for multimodal agent applications.

Chatbots & LLMs
Paid$0.60/1M tokensView details →
Qw

Qwen: Qwen3 235B A22B Thinking 2507

Qwen

7.8

Qwen3-235B-A22B-Thinking-2507 is a high-performance, open-weight Mixture-of-Experts (MoE) language model optimized for…

Chatbots & LLMs
Paid$0.23/1M tokensView details →
Hu

HuggingChat

Hugging Face

7.7

HuggingChat is Hugging Face's free, open-source chat interface for leading open-weight models like Llama, Qwen and…

Chatbots & LLMsCoding
Free tierFreeView details →
Op

OpenAI: GPT-5.4 Image 2

OpenAI

7.7

[GPT-5.4](https://openrouter.ai/openai/gpt-5.4) Image 2 combines OpenAI's GPT-5.4 model with state-of-the-art image…

Image GenerationChatbots & LLMs
Paid$8.00/1M tokensView details →
Z.

Z.ai: GLM 5.1

Z.AI

7.7

GLM-5.1 delivers a major leap in coding capability, with particularly significant gains in handling long-horizon tasks.

Chatbots & LLMs
Paid$1.40/1M tokensView details →
Re

Reka Edge

Rekaai

7.7

Reka Edge is an extremely efficient 7B multimodal vision-language model that accepts image/video+text inputs and…

Chatbots & LLMs
Paid$0.10/1M tokensView details →
Mi

MiniMax: MiniMax M2.7

MiniMax

7.7

MiniMax-M2.7 is a next-generation large language model designed for autonomous, real-world productivity and continuous…

Chatbots & LLMs
Paid$0.30/1M tokensView details →
Z.

Z.ai: GLM 5 Turbo

Z.AI

7.7

GLM-5 Turbo is a new model from Z.ai designed for fast inference and strong performance in agent-driven environments…

Chatbots & LLMs
Paid$1.20/1M tokensView details →
Mi

MiniMax: MiniMax M2.5

MiniMax

7.7

MiniMax-M2.5 is a SOTA large language model designed for real-world productivity.

Chatbots & LLMs
Paid$0.22/1M tokensView details →
Z.

Z.ai: GLM 5

Z.AI

7.7

GLM-5 is Z.ai’s flagship open-source foundation model engineered for complex systems design and long-horizon agent…

Chatbots & LLMs
Paid$0.95/1M tokensView details →
Z.

Z.ai: GLM 4.7 Flash

Z.AI

7.7

As a 30B-class SOTA model, GLM-4.7-Flash offers a new option that balances performance and efficiency.

Chatbots & LLMs
Paid$0.06/1M tokensView details →
Mi

MiniMax: MiniMax M2.1

MiniMax

7.7

MiniMax-M2.1 is a lightweight, state-of-the-art large language model optimized for coding, agentic workflows, and…

Chatbots & LLMs
Paid$0.30/1M tokensView details →
Z.

Z.ai: GLM 4.7

Z.AI

7.7

GLM-4.7 is Z.ai’s latest flagship model, featuring upgrades in two key areas: enhanced programming capabilities and…

Chatbots & LLMs
Paid$0.40/1M tokensView details →
Mi

MiniMax: MiniMax M2

MiniMax

7.7

MiniMax-M2 is a compact, high-efficiency large language model optimized for end-to-end coding and agentic workflows.

Chatbots & LLMs
Paid$0.26/1M tokensView details →
Z.

Z.ai: GLM 4.6

Z.AI

7.7

Compared with GLM-4.5, this generation brings several key improvements: Longer context window: The context window has…

Chatbots & LLMs
Paid$0.50/1M tokensView details →
Op

OpenAI: o3 Mini High (batch)

OpenAI

7.7

OpenAI o3-mini-high is the same model as [o3-mini](/openai/o3-mini) with reasoning_effort set to high.

Chatbots & LLMs
Paid$0.55/1M tokensView details →
Op

OpenAI: o3 Mini (batch)

OpenAI

7.7

OpenAI o3-mini is a cost-efficient language model optimized for STEM reasoning tasks, particularly excelling in…

Chatbots & LLMs
Paid$0.55/1M tokensView details →
Am

Amazon: Nova Lite 1.0

Amazon

7.7

Amazon Nova Lite 1.0 is a very low-cost multimodal model from Amazon that focused on fast processing of image, video…

Chatbots & LLMs
Paid$0.06/1M tokensView details →
Am

Amazon: Nova Pro 1.0

Amazon

7.7

Amazon Nova Pro 1.0 is a capable multimodal model from Amazon focused on providing a combination of accuracy, speed…

Chatbots & LLMs
Paid$0.80/1M tokensView details →
Li

LiquidAI: LFM2.5-2.6B (free)

Liquid AI

7.6

LFM2.5-2.6B is a compact reasoning model from Liquid AI. It is suited for agent workflows, data extraction, RAG, and…

Chatbots & LLMs
Free tierFreeView details →
Ai

AionLabs: Aion-3.0-Mini

Aion Labs

7.6

Aion-3.0 Mini is a multi-model roleplaying and storytelling system from AionLabs, built on the DeepSeek family of…

Chatbots & LLMs
Paid$0.70/1M tokensView details →
Ai

AionLabs: Aion-3.0

Aion Labs

7.6

Aion-3.0 is a multi-model roleplaying and storytelling system from AionLabs, built on the GLM family of models.

Chatbots & LLMs
Paid$3.00/1M tokensView details →
In

Inception: Mercury 2

Inception

7.6

Mercury 2 is an extremely fast reasoning LLM, and the first reasoning diffusion LLM (dLLM).

Chatbots & LLMs
Paid$0.25/1M tokensView details →
Ai

AionLabs: Aion-2.0

Aion Labs

7.6

Aion-2.0 is a variant of DeepSeek V3.2 optimized for immersive roleplaying and storytelling.

Chatbots & LLMs
Paid$0.80/1M tokensView details →
Up

Upstage: Solar Pro 3

Upstage

7.6

Solar Pro 3 is Upstage's powerful Mixture-of-Experts (MoE) language model.

Chatbots & LLMs
Paid$0.15/1M tokensView details →
Mi

Mistral: Ministral 3 14B 2512

Mistral AI

7.6

The largest model in the Ministral 3 family, Ministral 3 14B offers frontier capabilities and performance comparable to…

Chatbots & LLMs
Paid$0.20/1M tokensView details →
Mi

Mistral: Ministral 3 8B 2512

Mistral AI

7.6

A balanced model in the Ministral 3 family, Ministral 3 8B is a powerful, efficient tiny language model with vision…

Chatbots & LLMs
Paid$0.15/1M tokensView details →
Mi

Mistral: Mistral Large 3 2512

Mistral AI

7.6

Mistral Large 3 2512 is Mistral’s most capable model to date, featuring a sparse mixture-of-experts architecture with…

Chatbots & LLMs
Paid$0.50/1M tokensView details →
De

DeepSeek: DeepSeek V3.2

DeepSeek

7.6

DeepSeek-V3.2 is a large language model designed to harmonize high computational efficiency with strong reasoning and…

Chatbots & LLMs
Paid$0.27/1M tokensView details →
Pe

Perplexity: Sonar Pro Search

Perplexity

7.6

Exclusively available on the OpenRouter API, Sonar Pro's new Pro Search mode is Perplexity's most advanced agentic…

Chatbots & LLMs
Paid$3.00/1M tokensView details →
Qw

Qwen: Qwen3 VL 8B Instruct

Qwen

7.6

Qwen3-VL-8B-Instruct is a multimodal vision-language model from the Qwen3-VL series, built for high-fidelity…

Chatbots & LLMs
Paid$0.12/1M tokensView details →
Qw

Qwen: Qwen3 VL 30B A3B Instruct

Qwen

7.6

Qwen3-VL-30B-A3B-Instruct is a multimodal model that unifies strong text generation with visual understanding for…

Chatbots & LLMs
Paid$0.15/1M tokensView details →
De

DeepSeek: DeepSeek V3.2 Exp

DeepSeek

7.6

DeepSeek-V3.2-Exp is an experimental large language model released by DeepSeek as an intermediate step between V3.1 and…

Chatbots & LLMs
Paid$0.27/1M tokensView details →
Qw

Qwen: Qwen3 VL 235B A22B Instruct

Qwen

7.6

Qwen3-VL-235B-A22B Instruct is an open-weight multimodal model that unifies strong text generation with visual…

Chatbots & LLMs
Paid$0.26/1M tokensView details →
De

DeepSeek: DeepSeek V3.1 Terminus

DeepSeek

7.6

DeepSeek-V3.1 Terminus is an update to [DeepSeek V3.1](/deepseek/deepseek-chat-v3.1) that maintains the model's…

Chatbots & LLMs
Paid$0.27/1M tokensView details →
NV

NVIDIA: Nemotron Nano 9B V2 (free)

NVIDIA

7.6

NVIDIA-Nemotron-Nano-9B-v2 is a large language model (LLM) trained from scratch by NVIDIA, and designed as a unified…

Chatbots & LLMs
Free tierFreeView details →
De

DeepSeek: DeepSeek V3.1

DeepSeek

7.6

DeepSeek-V3.1 is a large hybrid reasoning model (671B parameters, 37B active) that supports both thinking and…

Chatbots & LLMs
Paid$0.25/1M tokensView details →
Op

OpenAI: gpt-oss-120b

OpenAI

7.6

gpt-oss-120b is an open-weight, 117B-parameter Mixture-of-Experts (MoE) language model from OpenAI designed for…

Chatbots & LLMs
Paid$0.03/1M tokensView details →
Op

OpenAI: gpt-oss-20b (free)

OpenAI

7.6

gpt-oss-20b is an open-weight 21B parameter model released by OpenAI under the Apache 2.0 license.

Chatbots & LLMs
Free tierFreeView details →
Z.

Z.ai: GLM 4.5

Z.AI

7.6

GLM-4.5 is our latest flagship foundation model, purpose-built for agent-based applications.

Chatbots & LLMs
Paid$0.60/1M tokensView details →
Z.

Z.ai: GLM 4.5 Air

Z.AI

7.6

GLM-4.5-Air is the lightweight variant of our latest flagship model family, also purpose-built for agent-centric…

Chatbots & LLMs
Paid$0.13/1M tokensView details →
Mi

Mistral: Mistral Small 3.2 24B

Mistral AI

7.6

Mistral-Small-3.2-24B-Instruct-2506 is an updated 24B parameter model from Mistral optimized for instruction following…

Chatbots & LLMs
Paid$0.09/1M tokensView details →
De

DeepSeek: R1 0528

DeepSeek

7.6

May 28th update to the [original DeepSeek R1](/deepseek/deepseek-r1) Performance on par with [OpenAI o1](/openai/o1)…

Chatbots & LLMs
Paid$0.50/1M tokensView details →
Qw

Qwen: Qwen3 30B A3B

Qwen

7.6

Qwen3, the latest generation in the Qwen large language model series, features both dense and mixture-of-experts (MoE)…

Chatbots & LLMs
Paid$0.12/1M tokensView details →
Qw

Qwen: Qwen3 8B

Qwen

7.6

Qwen3-8B is a dense 8.2B parameter causal language model from the Qwen3 series, designed for both reasoning-heavy tasks…

Chatbots & LLMs
Paid$0.12/1M tokensView details →
Qw

Qwen: Qwen3 14B

Qwen

7.6

Qwen3-14B is a dense 14.8B parameter causal language model from the Qwen3 series, designed for both complex reasoning…

Chatbots & LLMs
Paid$0.12/1M tokensView details →
Qw

Qwen: Qwen3 32B

Qwen

7.6

Qwen3-32B is a dense 32.8B parameter causal language model from the Qwen3 series, optimized for both complex reasoning…

Chatbots & LLMs
Paid$0.08/1M tokensView details →
Qw

Qwen: Qwen3 235B A22B

Qwen

7.6

Qwen3-235B-A22B is a 235B parameter mixture-of-experts (MoE) model developed by Qwen, activating 22B parameters per…

Chatbots & LLMs
Paid$0.45/1M tokensView details →
Op

OpenAI: o1-pro (batch)

OpenAI

7.6

The o1 series of models are trained with reinforcement learning to think before they answer and perform complex…

Chatbots & LLMs
Paid$75.00/1M tokensView details →
Go

Google: Gemma 3 27B

Google

7.6

Gemma 3 introduces multimodality, supporting vision-language input and text outputs.

Chatbots & LLMs
Paid$0.08/1M tokensView details →
De

DeepSeek: R1

DeepSeek

7.6

DeepSeek R1 is here: Performance on par with [OpenAI o1](/openai/o1), but open-sourced and with fully open reasoning…

Chatbots & LLMs
Paid$0.70/1M tokensView details →
Go

Google: Nano Banana 2 (Gemini 3.1 Flash Image)

Google

7.5

Gemini 3.1 Flash Image, a.k.a. "Nano Banana 2," is Google’s latest state of the art image generation and editing model…

Image GenerationChatbots & LLMs
Paid$0.50/1M tokensView details →
NV

NVIDIA: Nemotron 3.5 Content Safety (free)

NVIDIA

7.5

NVIDIA Nemotron 3.5 Content Safety is a compact 4B-parameter multimodal guardrail model from NVIDIA, fine-tuned from…

Chatbots & LLMs
Free tierFreeView details →
Fr

Free Models Router

Openrouter

7.5

The simplest way to get free inference. openrouter/free is a router that selects free models at random from the models…

Chatbots & LLMs
Free tierFreeView details →
Qw

Qwen: Qwen3 30B A3B Thinking 2507

Qwen

7.5

Qwen3-30B-A3B-Thinking-2507 is a 30B parameter Mixture-of-Experts reasoning model optimized for complex tasks requiring…

Chatbots & LLMs
Paid$0.20/1M tokensView details →
Ba

Baidu: ERNIE 4.5 VL 424B A47B

Baidu

7.5

ERNIE-4.5-VL-424B-A47B is a multimodal Mixture-of-Experts (MoE) model from Baidu’s ERNIE 4.5 series, featuring 424B…

Chatbots & LLMs
Paid$0.42/1M tokensView details →
Pe

Perplexity: Sonar Reasoning Pro

Perplexity

7.5

Note: Sonar Pro pricing includes Perplexity search pricing. See [details…

Chatbots & LLMs
Paid$2.00/1M tokensView details →
An

Anthropic: Claude 3 Haiku

Anthropic

7.5

Claude 3 Haiku is Anthropic's fastest and most compact model for near-instant responsiveness.

Chatbots & LLMs
Paid$0.25/1M tokensView details →
Go

Google: Nano Banana 2 Lite (Gemini 3.1 Flash Lite Image)

Google

7.4

Nano Banana 2 Lite (Gemini 3.1 Flash Lite Image) is Google's fastest, most cost-efficient Gemini image model, built for…

Image GenerationChatbots & LLMs
Paid$0.25/1M tokensView details →
Go

Google: Nano Banana 2 (Gemini 3.1 Flash Image Preview)

Google

7.4

Gemini 3.1 Flash Image Preview, a.k.a. "Nano Banana 2," is Google’s latest state of the art image generation and…

Image GenerationChatbots & LLMs
Paid$0.50/1M tokensView details →
Op

OpenAI: GPT-5.2 Chat

OpenAI

7.4

GPT-5.2 Chat (AKA Instant) is the fast, lightweight member of the 5.2 family, optimized for low-latency chat while…

Chatbots & LLMs
Paid$1.75/1M tokensView details →
Mi

Mistral: Ministral 3 3B 2512

Mistral AI

7.4

The smallest model in the Ministral 3 family, Ministral 3 3B is a powerful, efficient tiny language model with vision…

Chatbots & LLMs
Paid$0.10/1M tokensView details →
Go

Google: Nano Banana Pro (Gemini 3 Pro Image Preview)

Google

7.4

Nano Banana Pro is Google’s most advanced image-generation and editing model, built on Gemini 3 Pro.

Image GenerationChatbots & LLMs
Paid$2.00/1M tokensView details →
Qw

Qwen: Qwen3 VL 32B Instruct

Qwen

7.4

Qwen3-VL-32B-Instruct is a large-scale multimodal vision-language model designed for high-precision understanding and…

Chatbots & LLMs
Paid$0.10/1M tokensView details →
Mi

Mistral: Mistral Medium 3.1

Mistral AI

7.4

Mistral Medium 3.1 is an updated version of Mistral Medium 3, which is a high-performance enterprise-grade language…

Chatbots & LLMs
Paid$0.40/1M tokensView details →
Mi

Mistral: Mistral Medium 3

Mistral AI

7.4

Mistral Medium 3 is a high-performance enterprise-grade language model designed to deliver frontier-level capabilities…

Chatbots & LLMs
Paid$0.40/1M tokensView details →
Go

Google: Gemma 3 12B

Google

7.4

Gemma 3 introduces multimodality, supporting vision-language input and text outputs.

Chatbots & LLMs
Paid$0.05/1M tokensView details →
Op

OpenAI: GPT-4o (2024-11-20)

OpenAI

7.4

The 2024-11-20 version of GPT-4o offers a leveled-up creative writing ability with more natural, engaging, and tailored…

Chatbots & LLMs
Paid$2.50/1M tokensView details →
Op

OpenAI: GPT-4o (2024-08-06)

OpenAI

7.4

The 2024-08-06 version of GPT-4o offers improved performance in structured outputs, with the ability to supply a JSON…

Chatbots & LLMs
Paid$2.50/1M tokensView details →
Op

OpenAI: GPT-4o-mini (batch)

OpenAI

7.4

GPT-4o mini is OpenAI's newest model after [GPT-4 Omni](/models/openai/gpt-4o), supporting both text and image inputs…

Chatbots & LLMs
Paid$0.07/1M tokensView details →
Op

OpenAI: GPT-4o-mini (2024-07-18)

OpenAI

7.4

GPT-4o mini is OpenAI's newest model after [GPT-4 Omni](/models/openai/gpt-4o), supporting both text and image inputs…

Chatbots & LLMs
Paid$0.15/1M tokensView details →
Op

OpenAI: GPT-4o (batch)

OpenAI

7.4

GPT-4o ("o" for "omni") is OpenAI's latest AI model, supporting both text and image inputs with text outputs.

Chatbots & LLMs
Paid$1.25/1M tokensView details →
Op

OpenAI: GPT-4o (2024-05-13)

OpenAI

7.4

GPT-4o ("o" for "omni") is OpenAI's latest AI model, supporting both text and image inputs with text outputs.

Chatbots & LLMs
Paid$5.00/1M tokensView details →
Op

OpenAI: GPT-4 Turbo (batch)

OpenAI

7.4

The latest GPT-4 Turbo model with vision capabilities. Vision requests can now use JSON mode and function…

Chatbots & LLMs
Paid$5.00/1M tokensView details →
Kw

Kwaipilot: KAT-Coder-Air V2.5

Kwaipilot

7.3

KAT-Coder-Air V2.5 is a flagship-level Agentic Coding model that can directly hand over an entire issue or an entire…

CodingChatbots & LLMs
Paid$0.15/1M tokensView details →
Kw

Kwaipilot: KAT-Coder-Pro V2.5

Kwaipilot

7.3

KAT-Coder-Pro V2.5 is a flagship-level Agentic Coding model that can directly hand over an entire issue or an entire…

CodingChatbots & LLMs
Paid$0.74/1M tokensView details →
Pe

Perceptron: Perceptron Mk1

Perceptron

7.3

Perceptron Mk1 (Mark One) is Perceptron's highest-quality vision-language model for video and embodied reasoning.** It…

Chatbots & LLMs
Paid$0.15/1M tokensView details →
in

inclusionAI: Ling-2.6-1T

Inclusionai

7.3

Ling-2.6-1T is an instant (instruct) model from inclusionAI and the company’s trillion-parameter flagship, designed for…

Chatbots & LLMs
Paid$0.07/1M tokensView details →
in

inclusionAI: Ling-2.6-flash

Inclusionai

7.3

Ling-2.6-flash is an instant (instruct) model from inclusionAI with 104B total parameters and 7.4B active parameters…

Chatbots & LLMs
Paid$0.01/1M tokensView details →
Kw

Kwaipilot: KAT-Coder-Pro V2

Kwaipilot

7.3

KAT-Coder-Pro V2 is the latest high-performance model in KwaiKAT’s KAT-Coder series, designed for complex…

CodingChatbots & LLMs
Paid$0.30/1M tokensView details →
Qw

Qwen: Qwen3 Coder Next

Qwen

7.3

Qwen3-Coder-Next is an open-weight causal language model optimized for coding agents and local development workflows.

CodingChatbots & LLMs
Paid$0.12/1M tokensView details →
Re

Relace: Relace Search

Relace

7.3

The relace-search model uses 4-12 `view_file` and `grep` tools in parallel to explore a codebase and return relevant…

Chatbots & LLMs
Paid$1.00/1M tokensView details →
Qw

Qwen: Qwen3 Next 80B A3B Instruct

Qwen

7.3

Qwen3-Next-80B-A3B-Instruct is an instruction-tuned chat model in the Qwen3-Next series optimized for fast, stable…

Chatbots & LLMs
Paid$0.09/1M tokensView details →
Mo

MoonshotAI: Kimi K2 0905

Moonshot AI

7.3

Kimi K2 0905 is the September update of [Kimi K2 0711](moonshotai/kimi-k2).

Chatbots & LLMs
Paid$0.60/1M tokensView details →
AI

AI21: Jamba Large 1.7

AI21 Labs

7.3

Jamba Large 1.7 is the latest model in the Jamba open family, offering improvements in grounding…

Chatbots & LLMs
Paid$2.00/1M tokensView details →
Mi

Mistral: Codestral 2508

Mistral AI

7.3

Mistral's cutting-edge language model for coding released end of July 2025.

CodingChatbots & LLMs
Paid$0.30/1M tokensView details →
Qw

Qwen: Qwen3 Coder 30B A3B Instruct

Qwen

7.3

Qwen3-Coder-30B-A3B-Instruct is a 30.5B parameter Mixture-of-Experts (MoE) model with 128 experts (8 active per forward…

CodingChatbots & LLMs
Paid$0.07/1M tokensView details →
Qw

Qwen: Qwen3 30B A3B Instruct 2507

Qwen

7.3

Qwen3-30B-A3B-Instruct-2507 is a 30.5B-parameter mixture-of-experts language model from Qwen, with 3.3B active…

Chatbots & LLMs
Paid$0.05/1M tokensView details →
Qw

Qwen: Qwen3 Coder 480B A35B

Qwen

7.3

Qwen3-Coder-480B-A35B-Instruct is a Mixture-of-Experts (MoE) code generation model developed by the Qwen team.

CodingChatbots & LLMs
Paid$0.30/1M tokensView details →
Qw

Qwen: Qwen3 235B A22B Instruct 2507

Qwen

7.3

Qwen3-235B-A22B-Instruct-2507 is a multilingual, instruction-tuned mixture-of-experts language model based on the…

Chatbots & LLMs
Paid$0.09/1M tokensView details →
De

Deep Cogito: Cogito v2.1 671B

Deepcogito

7.2

Cogito v2.1 671B MoE represents one of the strongest open models globally, matching performance of frontier closed and…

Chatbots & LLMs
Paid$1.25/1M tokensView details →
No

Nous: Hermes 4 70B

Nous Research

7.2

Hermes 4 70B is a hybrid reasoning model from Nous Research, built on Meta-Llama-3.1-70B.

Chatbots & LLMs
Paid$0.13/1M tokensView details →
No

Nous: Hermes 4 405B

Nous Research

7.2

Hermes 4 is a large-scale reasoning model built on Meta-Llama-3.1-405B and released by Nous Research.

Chatbots & LLMs
Paid$1.00/1M tokensView details →
Te

Tencent: Hunyuan A13B Instruct

Tencent

7.2

Hunyuan-A13B is a 13B active parameter Mixture-of-Experts (MoE) language model developed by Tencent, with a total…

Chatbots & LLMs
Paid$0.14/1M tokensView details →
Pe

Perplexity: Sonar Deep Research

Perplexity

7.2

Sonar Deep Research is a research-focused model designed for multi-step retrieval, synthesis, and reasoning across…

Chatbots & LLMs
Paid$2.00/1M tokensView details →
IB

IBM: Granite 4.1 8B

Ibm Granite

7.1

Granite 4.1 8B is a dense, decoder-only 8-billion-parameter language model from IBM, part of the Granite 4.1 family.

Chatbots & LLMs
Paid$0.05/1M tokensView details →
Op

OpenAI: GPT Audio

OpenAI

7.1

The gpt-audio model is OpenAI's first generally available audio model. The new snapshot features an upgraded decoder…

Video & AudioChatbots & LLMs
Paid$2.50/1M tokensView details →
Op

OpenAI: GPT Audio Mini

OpenAI

7.1

A cost-efficient version of GPT Audio. The new snapshot features an upgraded decoder for more natural sounding voices…

Video & AudioChatbots & LLMs
Paid$0.60/1M tokensView details →
Al

AllenAI: Olmo 3 32B Think

Allenai

7.1

Olmo 3 32B Think is a large-scale, 32-billion-parameter model purpose-built for deep reasoning, complex logic chains…

Chatbots & LLMs
Paid$0.15/1M tokensView details →
Mo

MoonshotAI: Kimi K2 0711

Moonshot AI

7.1

Kimi K2 Instruct is a large-scale Mixture-of-Experts (MoE) language model developed by Moonshot AI, featuring 1…

Chatbots & LLMs
Paid$0.57/1M tokensView details →
Ar

Arcee AI: Virtuoso Large

Arcee Ai

7.1

Virtuoso‑Large is Arcee's top‑tier general‑purpose LLM at 72 B parameters, tuned to tackle cross‑domain reasoning…

Chatbots & LLMs
Paid$0.75/1M tokensView details →
De

DeepSeek: DeepSeek V3 0324

DeepSeek

7.1

DeepSeek V3, a 685B-parameter, mixture-of-experts model, is the latest iteration of the flagship chat model family from…

Chatbots & LLMs
Paid$0.27/1M tokensView details →
Re

Reka Flash 3

Rekaai

7.1

Reka Flash 3 is a general-purpose, instruction-tuned large language model with 21 billion parameters, developed by…

Chatbots & LLMs
Paid$0.10/1M tokensView details →
Pe

Perplexity: Sonar Pro

Perplexity

7.1

Note: Sonar Pro pricing includes Perplexity search pricing. See [details…

Chatbots & LLMs
Paid$3.00/1M tokensView details →
De

DeepSeek: DeepSeek V3

DeepSeek

7.1

DeepSeek-V3 is the latest model from the DeepSeek team, building upon the instruction following and coding abilities of…

Chatbots & LLMs
Paid$0.26/1M tokensView details →
Me

Meta: Llama 3.3 70B Instruct

Meta

7.1

The Meta Llama 3.3 multilingual large language model (LLM) is a pretrained and instruction tuned generative model in…

Chatbots & LLMs
Paid$0.10/1M tokensView details →
Am

Amazon: Nova Micro 1.0

Amazon

7.1

Amazon Nova Micro 1.0 is a text-only model that delivers the lowest latency responses in the Amazon Nova family of…

Chatbots & LLMs
Paid$0.04/1M tokensView details →
Mi

Mistral Large 2407

Mistral AI

7.1

This is Mistral AI's flagship model, Mistral Large 2 (version mistral-large-2407).

Chatbots & LLMs
Paid$2.00/1M tokensView details →
Co

Cohere: Command R (08-2024)

Cohere

7.1

command-r-08-2024 is an update of the [Command R](/models/cohere/command-r) with improved performance for multilingual…

Chatbots & LLMs
Paid$0.15/1M tokensView details →
Co

Cohere: Command R+ (08-2024)

Cohere

7.1

command-r-plus-08-2024 is an update of the [Command R+](/models/cohere/command-r-plus) with roughly 50% higher…

Chatbots & LLMs
Paid$2.50/1M tokensView details →
Sa

Sao10K: Llama 3.1 Euryale 70B v2.2

Sao10K

7.1

Euryale L3.1 70B v2.2 is a model focused on creative roleplay from [Sao10k](https://ko-fi.com/sao10k).

Chatbots & LLMs
Paid$0.85/1M tokensView details →
Me

Meta: Llama 3.1 70B Instruct

Meta

7.1

Meta's latest class of model (Llama 3.1) launched with a variety of sizes & flavors.

Chatbots & LLMs
Paid$0.40/1M tokensView details →
Me

Meta: Llama 3.1 8B Instruct

Meta

7.1

Meta's latest class of model (Llama 3.1) launched with a variety of sizes & flavors.

Chatbots & LLMs
Paid$0.05/1M tokensView details →
Mi

Mistral: Mistral Nemo

Mistral AI

7.1

A 12B parameter model with a 128k token context length built by Mistral in collaboration with NVIDIA.

Chatbots & LLMs
Paid$0.02/1M tokensView details →
Mi

Mistral Large

Mistral AI

7.1

This is Mistral AI's flagship model, Mistral Large 2 (version `mistral-large-2407`).

Chatbots & LLMs
Paid$2.00/1M tokensView details →
Op

OpenAI: GPT-4 Turbo Preview

OpenAI

7.1

The preview GPT-4 model with improved instruction following, JSON mode, reproducible outputs, parallel function…

Chatbots & LLMs
Paid$10.00/1M tokensView details →
By

ByteDance: UI-TARS 7B

Bytedance

7.0

UI-TARS-1.5 is a multimodal vision-language agent optimized for GUI-based environments, including desktop interfaces…

Chatbots & LLMs
Paid$0.10/1M tokensView details →
Mi

Mistral: Mistral Small 3.1 24B

Mistral AI

7.0

Mistral Small 3.1 24B Instruct is an upgraded variant of Mistral Small 3 (2501), featuring 24 billion parameters with…

Chatbots & LLMs
Paid$0.35/1M tokensView details →
Go

Google: Gemma 3 4B

Google

7.0

Gemma 3 introduces multimodality, supporting vision-language input and text outputs.

Chatbots & LLMs
Paid$0.05/1M tokensView details →
Qw

Qwen: Qwen2.5 VL 72B Instruct

Qwen

7.0

Qwen2.5-VL is proficient in recognizing common objects such as flowers, birds, fish, and insects.

Chatbots & LLMs
Paid$0.25/1M tokensView details →
Pe

Perplexity: Sonar

Perplexity

7.0

Sonar is lightweight, affordable, fast, and simple to use — now featuring citations and the ability to customize…

Chatbots & LLMs
Paid$1.00/1M tokensView details →
De

DeepSeek: R1 Distill Llama 70B

DeepSeek

7.0

DeepSeek R1 Distill Llama 70B is a distilled large language model based on…

Chatbots & LLMs
Paid$0.80/1M tokensView details →
Mi

Mistral: Mixtral 8x22B Instruct

Mistral AI

7.0

Mistral's official instruct fine-tuned version of [Mixtral 8x22B](/models/mistralai/mixtral-8x22b).

Chatbots & LLMs
Paid$2.00/1M tokensView details →
Mi

Mistral: Voxtral Small 24B 2507

Mistral AI

6.9

Voxtral Small is an enhancement of Mistral Small 3, incorporating state-of-the-art audio input capabilities while…

Video & AudioChatbots & LLMs
Paid$0.10/1M tokensView details →
Re

Relace: Relace Apply 3

Relace

6.9

Relace Apply 3 is a specialized code-patching LLM that merges AI-suggested edits straight into your source files.

Chatbots & LLMs
Paid$0.85/1M tokensView details →
Mo

Morph: Morph V3 Large

Morph

6.9

Morph's high-accuracy apply model for complex code edits. ~4,500 tokens/sec with 98% accuracy for precise code…

Chatbots & LLMs
Paid$0.90/1M tokensView details →
Co

Cohere: Command A

Cohere

6.9

Command A is an open-weights 111B parameter model with a 256k context window focused on delivering great performance…

Chatbots & LLMs
Paid$2.50/1M tokensView details →
Mi

Mistral: Saba

Mistral AI

6.9

Mistral Saba is a 24B-parameter language model specifically designed for the Middle East and South Asia, delivering…

Chatbots & LLMs
Paid$0.20/1M tokensView details →
Qw

Qwen: Qwen2.5 7B Instruct

Qwen

6.9

Qwen2.5 7B is the latest series of Qwen large language models. Qwen2.5 brings the following improvements upon Qwen2…

Chatbots & LLMs
Paid$0.10/1M tokensView details →
Qw

Qwen2.5 72B Instruct

Qwen

6.9

Qwen2.5 72B is the latest series of Qwen large language models. Qwen2.5 brings the following improvements upon Qwen2…

Chatbots & LLMs
Paid$0.36/1M tokensView details →
Op

OpenAI: GPT-3.5 Turbo (older v0613)

OpenAI

6.9

GPT-3.5 Turbo is OpenAI's fastest model. It can understand and generate natural language or code, and is optimized for…

Chatbots & LLMs
Paid$1.00/1M tokensView details →
Op

OpenAI: GPT-3.5 Turbo 16k

OpenAI

6.9

This model offers four times the context length of gpt-3.5-turbo, allowing it to support approximately 20 pages of text…

Chatbots & LLMs
Paid$3.00/1M tokensView details →
Op

OpenAI: GPT-3.5 Turbo (batch)

OpenAI

6.9

GPT-3.5 Turbo is OpenAI's fastest model. It can understand and generate natural language or code, and is optimized for…

Chatbots & LLMs
Paid$0.25/1M tokensView details →
Op

OpenAI: GPT-4

OpenAI

6.9

OpenAI's flagship model, GPT-4 is a large-scale multimodal language model capable of solving difficult problems with…

Chatbots & LLMs
Paid$30.00/1M tokensView details →
Go

Google: Nano Banana (Gemini 2.5 Flash Image)

Google

6.8

Gemini 2.5 Flash Image, a.k.a. "Nano Banana," is now generally available.

Image GenerationChatbots & LLMs
Paid$0.30/1M tokensView details →
Bo

Body Builder (beta)

Openrouter

6.7

Transform your natural language requests into structured OpenRouter API request objects.

Chatbots & LLMs
PaidVariableView details →
IB

IBM: Granite 4.0 Micro

Ibm Granite

6.7

Granite-4.0-H-Micro is a 3B parameter from the Granite 4 family of models.

Chatbots & LLMs
Paid$0.02/1M tokensView details →
Th

TheDrummer: Cydonia 24B V4.1

TheDrummer

6.7

Uncensored and creative writing model based on Mistral Small 3.2 24B with good recall, prompt adherence, and…

Chatbots & LLMs
Paid$0.30/1M tokensView details →
Ve

Venice: Uncensored

Cognitivecomputations

6.7

Venice Uncensored Dolphin Mistral 24B Venice Edition is a fine-tuned variant of Mistral-Small-24B-Instruct-2501…

Chatbots & LLMs
Paid$0.20/1M tokensView details →
Sa

Sao10K: Llama 3.3 Euryale 70B

Sao10K

6.7

Euryale L3.3 70B is a model focused on creative roleplay from [Sao10k](https://ko-fi.com/sao10k).

Chatbots & LLMs
Paid$0.65/1M tokensView details →
Co

Cohere: Command R7B (12-2024)

Cohere

6.7

Command R7B (12-2024) is a small, fast update of the Command R+ model, delivered in December 2024.

Chatbots & LLMs
Paid$0.04/1M tokensView details →
Me

Meta: Llama 3.2 3B Instruct

Meta

6.7

Llama 3.2 3B is a 3-billion-parameter multilingual large language model, optimized for advanced natural language…

Chatbots & LLMs
Paid$0.05/1M tokensView details →
No

Nous: Hermes 3 70B Instruct

Nous Research

6.7

Hermes 3 is a generalist language model with many improvements over [Hermes…

Chatbots & LLMs
Paid$0.70/1M tokensView details →
No

Nous: Hermes 3 405B Instruct

Nous Research

6.7

Hermes 3 is a generalist language model with many improvements over Hermes 2, including advanced agentic capabilities…

Chatbots & LLMs
Paid$1.00/1M tokensView details →
Mi

MiniMax: MiniMax M2-her

MiniMax

6.6

MiniMax M2-her is a dialogue-first large language model built for immersive roleplay, character-driven chat, and…

Chatbots & LLMs
Paid$0.30/1M tokensView details →
Mo

Morph: Morph V3 Fast

Morph

6.6

Morph's fastest apply model for code edits. ~10,500 tokens/sec with 96% accuracy for rapid code transformations.

Chatbots & LLMs
Paid$0.80/1M tokensView details →
Th

TheDrummer: Rocinante 12B

TheDrummer

6.6

Rocinante 12B is designed for engaging storytelling and rich prose. Early testers have reported: - Expanded vocabulary…

Chatbots & LLMs
Paid$0.25/1M tokensView details →
Me

Meta: Llama 3.2 1B Instruct

Meta

6.6

Llama 3.2 1B is a 1-billion-parameter language model focused on efficiently performing natural language tasks, such as…

Chatbots & LLMs
Paid$0.03/1M tokensView details →
Wi

WizardLM-2 8x22B

Microsoft

6.6

WizardLM-2 8x22B is Microsoft AI's most advanced Wizard model. It demonstrates highly competitive performance compared…

Chatbots & LLMs
Paid$0.62/1M tokensView details →
Go

Google: Gemma 3n 4B

Google

6.5

Gemma 3n E4B-it is optimized for efficient execution on mobile and low-resource devices, such as phones, laptops, and…

Chatbots & LLMs
Paid$0.06/1M tokensView details →
Th

TheDrummer: Skyfall 36B V2

TheDrummer

6.5

Skyfall 36B v2 is an enhanced iteration of Mistral Small 2501, specifically fine-tuned for improved creativity, nuanced…

Chatbots & LLMs
Paid$0.55/1M tokensView details →
Ai

AionLabs: Aion-RP 1.0 (8B)

Aion Labs

6.5

Aion-RP-Llama-3.1-8B ranks the highest in the character evaluation portion of the RPBench-Auto benchmark, a…

Chatbots & LLMs
Paid$0.80/1M tokensView details →
Mi

Mistral: Mistral Small 3

Mistral AI

6.5

Mistral Small 3 is a 24B-parameter language model optimized for low-latency performance across common AI tasks.

Chatbots & LLMs
Paid$0.05/1M tokensView details →
Mi

Microsoft: Phi 4

Microsoft

6.5

[Microsoft Research](/microsoft) Phi-4 is designed to perform well in complex reasoning tasks and can operate…

Chatbots & LLMs
Paid$0.07/1M tokensView details →
Qw

Qwen2.5 Coder 32B Instruct

Qwen

6.5

Qwen2.5-Coder is the latest series of Code-Specific Qwen large language models (formerly known as CodeQwen).

CodingChatbots & LLMs
Paid$0.66/1M tokensView details →
Ma

Magnum v4 72B

Anthracite Org

6.5

This is a series of models designed to replicate the prose quality of the Claude 3 models, specifically…

Chatbots & LLMs
Paid$3.00/1M tokensView details →
Sa

Sao10K: Llama 3 8B Lunaris

Sao10K

6.5

Lunaris 8B is a versatile generalist and roleplaying model based on Llama 3.

Chatbots & LLMs
Paid$0.04/1M tokensView details →
Go

Google: Gemma 2 27B

Google

6.5

Gemma 2 27B by Google is an open model built from the same research and technology used to create the [Gemini…

Chatbots & LLMs
Paid$0.65/1M tokensView details →
Op

OpenAI: GPT-3.5 Turbo Instruct

OpenAI

6.5

This model is a variant of GPT-3.5 Turbo tuned for instructional prompts and omitting chat-related optimizations.

Chatbots & LLMs
Paid$1.50/1M tokensView details →
Ma

Mancer: Weaver (alpha)

Mancer

6.5

An attempt to recreate Claude-style verbosity, but don't expect the same level of coherence or memory.

Chatbots & LLMs
Paid$0.50/1M tokensView details →
Re

ReMM SLERP 13B

Undi95

6.5

A recreation trial of the original MythoMax-L2-B13 but with updated models. #merge

Chatbots & LLMs
Paid$0.45/1M tokensView details →
My

MythoMax 13B

Gryphe

6.5

One of the highest performing and most popular fine-tunes of Llama 2 13B, with rich descriptions and roleplay. #merge

Chatbots & LLMs
Paid$0.06/1M tokensView details →
Li

LibreChat

Librechat

LibreChat is a free and open-source chat interface for assistant AIs. .

Chatbots & LLMs
Free tierFreeView details →
Ch

Chatbot UI

Chatbotui

An open source ChatGPT UI. .

Chatbots & LLMs
Free tierFreeView details →
Ex

Exa

Exa

Language model powered search.

Chatbots & LLMs
PaidSee siteView details →
Ph

Phind

Phind

AI-based search engine.

Chatbots & LLMs
PaidSee siteView details →
Ko

Komo

Komo

An AI-powered search engine.

Chatbots & LLMs
PaidSee siteView details →
Im

Improving GPT‑5.6 Sol in ChatGPT—and expanding access to…

OpenAI

ChatGPT introduces improved GPT-5.6 Sol with better accuracy and consistency, plus expanded access for free users and…

Chatbots & LLMs
PaidSee siteView details →
th

the ChatGPT for small business program

OpenAI

OpenAI launches the ChatGPT for Small Businesses program, helping entrepreneurs build AI skills, automate work, and…

Chatbots & LLMs
PaidSee siteView details →
A

A scorecard for the AI age

OpenAI

Sarah Friar, CFO of OpenAI, introduces a practical AI scorecard to measure ROI through useful work, cost per successful…

Chatbots & LLMs
PaidSee siteView details →
NA

NABTU and Meta Announce New Partnership to Invest In…

Meta

Meta and North America's Building Trades Unions (NABTU) are partnering to invest in skilled trades workers building…

Chatbots & LLMs
PaidSee siteView details →
th

the AI Glasses Impact Grant Recipients: Helping People…

Meta

We've awarded AI Glasses Impact Grants to 30 organizations across the US using Meta AI glasses to help people work…

Chatbots & LLMs
PaidSee siteView details →
Sh

Shieldstral.

Mistral AI

Shieldstral introduces a 3B open-weights multimodal safety classifier that outperforms models up to 7x its size.

Chatbots & LLMs
PaidSee siteView details →
Ro

Robostral Navigate

Mistral AI

Introducing Robostral Navigate: 8B model achieving 76.6% on R2R-CE with just a single RGB camera.

Chatbots & LLMs
PaidSee siteView details →
Ol

OlmoEarth embeddings: Custom embedding exports from…

Hugging Face

OlmoEarth embeddings: Custom embedding exports from… — announced by Hugging Face.

Chatbots & LLMs
PaidSee siteView details →
Me

Meta is back with Muse Glimmer: local, agentic, multimodal,…

Hugging Face

Meta is back with Muse Glimmer: local, agentic, multimodal,… — announced by Hugging Face.

Chatbots & LLMs
PaidSee siteView details →

How to choose a chatbots & llms tool

We track 368 chatbots & llms tools, of which 36 offer a free tier. Claude currently leads on our quality score.

When choosing a chatbots & llms tool, weigh four things: capability on your specific tasks, pricing (free tier vs. monthly vs. per-token), how well it fits your existing workflow and integrations, and how actively it's maintained. Start with a shortlist of two or three, trial them on a real task, and let the results decide.

Explore other categories