Browse 300+ AI Models

Explore the complete directory of AI models from all major providers. Find the perfect AI for coding, writing, analysis, and more.

All Models (401)Ai21 (1)Aion-labs (4)Allenai (1)Amazon (5)Anthracite-org (1)Anthropic (28)Arcee-ai (2)Baidu (1)Bytedance (1)Bytedance-seed (6)Cognitivecomputations (1)Cohere (5)Deepcogito (1)DeepSeek (13)Google (39)
Browse by: Free Cheapest Newest Vision Reasoning Web Search Biggest Context Image Generation
Bytedance-seed

ByteDance Seed: Seed 2.1 Turbo

Seed 2.1 Turbo is a multimodal model from ByteDance Seed for coding and long-horizon agent workflows. It is suited for end-to-end software delivery, multi-step task execution, and understanding visual
262K context standard
Qwen

Qwen: Qwen3.8 2.4T A95B

Qwen3.8 2.4T A95B is an open-weight sparse mixture-of-experts model from Qwen and the open-weight variant of [Qwen3.8 Max](/qwen/qwen3.8-max), with 95 billion active parameters out of 2.4 trillion tot
1000K context standard
Bytedance-seed

ByteDance Seed: Seed-2.0-Code

Seed 2.0 Code is a model from ByteDance Seed optimized for agentic coding. It is suited for frontend development, multilingual programming tasks, and coding-agent workflows in tools such as Claude...
262K context standard
DeepSeek

DeepSeek: DeepSeek V4 Pro 0813

DeepSeek V4 Pro 0813 is a large-scale mixture-of-experts model from DeepSeek. This is the GA release of DeepSeek V4 Pro.
1049K context budget
xAI

SpaceXAI: Grok 4.6

Grok 4.6 is SpaceXAI's smartest model with frontier performance on coding, knowledge work, and STEM.
500K context standard
Liquid

LiquidAI: LFM2.5-2.6B (free)

LFM2.5-2.6B is a compact reasoning model from Liquid AI. It is suited for agent workflows, data extraction, RAG, and long-context processing. Liquid advises against using it for agentic coding or...
128K context 3B budget
NVIDIA

NVIDIA: Nemotron 3.5 Lightning

NVIDIA Nemotron 3.5 Lightning is an open mixture-of-experts model from NVIDIA, with 3B active parameters out of 30B total. It is suited for high-throughput agentic workloads and specialized tasks that
262K context 3B budget
NVIDIA

NVIDIA: Nemotron 3.5 Lightning (free)

NVIDIA Nemotron 3.5 Lightning is an open mixture-of-experts model from NVIDIA, with 3B active parameters out of 30B total. It is suited for high-throughput agentic workloads and specialized tasks that
1000K context 3B budget
Sakana

Sakana: Sakana Namazu

Sakana Namazu is a Japanese-specialized reasoning model from Sakana AI, based on Kimi K2.6 with additional training for Japanese language and business contexts. It is suited for Japanese instruction f
262K context standard
Upstage

Upstage: Solar Pro 4

Solar Pro 4 is Upstage's cost-efficient large language model, featuring a 524K context window. It is built for long-horizon tasks and agentic workflows, with strong capabilities in office productivity
524K context budget
Meta

Meta: Muse Glimmer 30B

Muse Glimmer 30B is a dense, open-weight multimodal model from Meta Superintelligence Labs, distilled from Muse Spark and optimized for autonomous agents on consumer hardware. It is suited for long-ho
131K context 30B standard
Meta

Meta: Muse Spark 1.2

Muse Spark 1.2 is a reasoning model from Meta, designed for complex agentic tasks. It accepts text, images, video, audio, and PDF documents, returns text, and offers a 1M-token context...
1049K context standard
Qwen

Qwen: Qwen3.8 Max

Qwen3.8 Max is the flagship model in Alibaba's Qwen3.8 series, the general-availability successor to the Qwen3.8 Max Preview. It is a multimodal reasoning model intended for complex reasoning, visual
1000K context standard
~deepseek

DeepSeek V4 Flash Latest

This model always redirects to the latest model in the DeepSeek V4 Flash family.
1049K context budget
DeepSeek

DeepSeek: DeepSeek V4 Flash 0731

DeepSeek V4 Flash 0731 is a sparse mixture-of-experts model from DeepSeek, with 13B active parameters out of 284B total. This re-post-trained revision is suited for coding, reasoning, and agent workfl
1049K context 13B budget
Thinkingmachines

Thinking Machines: Inkling Small

Inkling Small is an open-weight multimodal mixture-of-experts model from Thinking Machines Lab, with 12B active parameters out of 276B total. It is positioned as the smaller, more efficient member of.
524K context 12B standard
Qwen

Qwen: Qwen3.7 Flash

Qwen3.7 Flash is a vision-language reasoning model from Alibaba. It is suited for multimodal agents, visual coding, search, and computer interaction, with strengths in object recognition, spatial unde
1000K context budget
Anthropic

Claude Opus 5 (Fast)

Fast-mode variant of [Opus 5](/anthropic/claude-opus-5) - identical capabilities with higher output speed at 2x pricing relative to regular Opus 5. Learn more in Anthropic's docs: https://platform.cl
1000K context premium
Anthropic

Claude Opus 5

Claude Opus 5 is Anthropic’s flagship model for demanding reasoning, coding, and long-horizon agentic work. It is particularly strong at end-to-end software tasks, code review and bug finding, visual
1000K context premium
Anthropic

Claude Opus 5 (batch)

Claude Opus 5 is Anthropic’s flagship model for demanding reasoning, coding, and long-horizon agentic work. It is particularly strong at end-to-end software tasks, code review and bug finding, visual
1000K context premium
Inclusionai

Ling-3.0-flash

*Ling-3.0-flash* is a *124B-parameter Mixture-of-Experts (MoE) model*, with approximately *5.1B parameters activated per token*. The model is designed with *token efficiency and production-scale agent
262K context 124B budget
Poolside

Poolside: Laguna S 2.1

Laguna S 2.1 is the latest coding agent model from [Poolside](). Laguna S 2.1 is a 118B total parameter model with 8B active parameters, scoring 70.2% on Terminal-Bench 2.1 and..
1049K context 118B budget
Poolside

Poolside: Laguna S 2.1 (free)

Laguna S 2.1 is the latest coding agent model from [Poolside](). Laguna S 2.1 is a 118B total parameter model with 8B active parameters, scoring 70.2% on Terminal-Bench 2.1 and..
262K context 118B budget
Google

Google: Gemini 3.6 Flash

Gemini 3.6 Flash is a high-efficiency model from Google for coding, agentic workflows, and web and app development. It is designed to produce polished outputs with fewer unnecessary edits and...
1049K context standard
Google

Google: Gemini 3.6 Flash (batch)

Gemini 3.6 Flash is a high-efficiency model from Google for coding, agentic workflows, and web and app development. It is designed to produce polished outputs with fewer unnecessary edits and...
1049K context standard
Google

Google: Gemini 3.5 Flash Lite

Gemini 3.5 Flash Lite is a high-efficiency model from Google with upgraded agentic capabilities. It is suited for subagents that execute focused tasks within complex, multi-agent workflows.
1049K context standard
Google

Google: Gemini 3.5 Flash Lite (batch)

Gemini 3.5 Flash Lite is a high-efficiency model from Google with upgraded agentic capabilities. It is suited for subagents that execute focused tasks within complex, multi-agent workflows.
1049K context standard
Meituan

Meituan: LongCat 2.0

LongCat 2.0 is a sparse mixture-of-experts language model from Meituan, with 48B active parameters out of 1.6T total. It is suited for coding, repository-level changes, long-horizon problem solving, a
1049K context 48B standard
Thinkingmachines

Thinking Machines: Inkling

Inkling is an open-weight multimodal mixture-of-experts model from Thinking Machines Lab, with 41B active parameters out of 975B total. It is designed for general-purpose reasoning, coding, agentic an
524K context 41B standard
Thinkingmachines

Thinking Machines: Inkling (batch)

Inkling is an open-weight multimodal mixture-of-experts model from Thinking Machines Lab, with 41B active parameters out of 975B total. It is designed for general-purpose reasoning, coding, agentic an
524K context 41B standard
Moonshotai

MoonshotAI: Kimi K3

Kimi K3 is a 2.8T parameter open-weight multimodal reasoning model from Moonshot AI. It is suited for complex coding, knowledge work, and long-horizon agentic workflows, and is particularly strong at.
1049K context premium
Meta

Meta: Muse Spark 1.1

Muse Spark 1.1 is a multimodal reasoning model from Meta, built for agentic tasks. It accepts text, images, video, audio, and PDF documents and returns text, with a 1M-token context...
1049K context standard
Kwaipilot

Kwaipilot: KAT-Coder-Air V2.5

KAT-Coder-Air V2.5 is a flagship-level Agentic Coding model that can directly hand over an entire issue or an entire business workflow to it, allowing it to autonomously locate and make...
256K context budget
Kwaipilot

Kwaipilot: KAT-Coder-Pro V2.5

KAT-Coder-Pro V2.5 is a flagship-level Agentic Coding model that can directly hand over an entire issue or an entire business workflow to it, allowing it to autonomously locate and make...
256K context standard
OpenAI

OpenAI: GPT-5.6 Luna Pro

GPT-5.6 Luna Pro is the same underlying model as [GPT-5.6 Luna](https://ask-coreai.com), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks. Learn more in OpenAI'
1050K context budget
OpenAI

OpenAI: GPT-5.6 Luna Pro (batch)

GPT-5.6 Luna Pro is the same underlying model as [GPT-5.6 Luna](https://ask-coreai.com), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks. Learn more in OpenAI'
1050K context budget
OpenAI

OpenAI: GPT-5.6 Luna

GPT-5.6 Luna is a fast, cost-efficient model in OpenAI's GPT-5.6 series. It is suited for high-volume, latency-sensitive tasks such as chat, classification, and lightweight agentic workflows, providin
1050K context budget
OpenAI

OpenAI: GPT-5.6 Luna (batch)

GPT-5.6 Luna is a fast, cost-efficient model in OpenAI's GPT-5.6 series. It is suited for high-volume, latency-sensitive tasks such as chat, classification, and lightweight agentic workflows, providin
1050K context budget
OpenAI

OpenAI: GPT-5.6 Terra Pro

GPT-5.6 Terra Pro is the same underlying model as [GPT-5.6 Terra](https://ask-coreai.com), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks. Learn more in OpenA
1050K context standard
OpenAI

OpenAI: GPT-5.6 Terra Pro (batch)

GPT-5.6 Terra Pro is the same underlying model as [GPT-5.6 Terra](https://ask-coreai.com), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks. Learn more in OpenA
1050K context standard
OpenAI

OpenAI: GPT-5.6 Terra

GPT-5.6 Terra is a balanced model in OpenAI's GPT-5.6 series, positioned between the flagship Sol tier and the cost-efficient Luna tier. It is suited for everyday coding, reasoning, and agentic...
1050K context standard
OpenAI

OpenAI: GPT-5.6 Terra (batch)

GPT-5.6 Terra is a balanced model in OpenAI's GPT-5.6 series, positioned between the flagship Sol tier and the cost-efficient Luna tier. It is suited for everyday coding, reasoning, and agentic...
1050K context standard
OpenAI

OpenAI: GPT-5.6 Sol Pro

GPT-5.6 Sol Pro is the same underlying model as [GPT-5.6 Sol](https://ask-coreai.com), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks. Learn more in OpenAI's
1050K context premium
OpenAI

OpenAI: GPT-5.6 Sol Pro (batch)

GPT-5.6 Sol Pro is the same underlying model as [GPT-5.6 Sol](https://ask-coreai.com), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks. Learn more in OpenAI's
1050K context premium
OpenAI

OpenAI: GPT-5.6 Sol

GPT-5.6 Sol is the flagship model in OpenAI's GPT-5.6 series. It is suited for complex reasoning, coding, and agentic workflows, and is particularly strong at command-line and multi-step coding tasks.
1050K context premium
OpenAI

OpenAI: GPT-5.6 Sol (batch)

GPT-5.6 Sol is the flagship model in OpenAI's GPT-5.6 series. It is suited for complex reasoning, coding, and agentic workflows, and is particularly strong at command-line and multi-step coding tasks.
1050K context premium
xAI

SpaceXAI: Grok 4.5

Grok 4.5 is a model from SpaceXAI with frontier performance on coding, knowledge work, and STEM.
500K context standard
~x-ai

xAI: Grok Latest

This model always redirects to the latest Grok model from xAI.
500K context standard

Popular AI Model Comparisons

Try Any AI Model Instantly

Chat with GPT-5, Claude, Gemini, and 300+ models — all in one app. Compare responses side-by-side.

Use Web App Download App →