Browse 300+ AI Models

Explore the complete directory of AI models from all major providers. Find the perfect AI for coding, writing, analysis, and more.

All Models (450)Aion-labs (6)Amazon (5)Anthracite-org (1)Anthropic (27)Arcee-ai (1)Baidu (1)Bytedance (1)Bytedance-seed (6)Cognitivecomputations (1)Cohere (6)DeepSeek (16)Dots-studio (1)Fireworks (1)Google (41)Gryphe (1)
Browse by: Free Cheapest Newest Vision Reasoning Web Search Biggest Context Image Generation
Poolside

Poolside: Laguna S 2.1 (free)

Laguna S 2.1 is the latest coding agent model from [Poolside](). Laguna S 2.1 is a 118B total parameter model with 8B active parameters, scoring 70.2% on Terminal-Bench 2.1 and..
262K context 118B budget
Google

Google: Gemini 3.6 Flash

Gemini 3.6 Flash is a high-efficiency model from Google for coding, agentic workflows, and web and app development. It is designed to produce polished outputs with fewer unnecessary edits and...
1049K context standard
Google

Google: Gemini 3.6 Flash (batch)

Gemini 3.6 Flash is a high-efficiency model from Google for coding, agentic workflows, and web and app development. It is designed to produce polished outputs with fewer unnecessary edits and...
1049K context standard
Google

Google: Gemini 3.5 Flash Lite

Gemini 3.5 Flash Lite is a high-efficiency model from Google with upgraded agentic capabilities. It is suited for subagents that execute focused tasks within complex, multi-agent workflows.
1049K context standard
Google

Google: Gemini 3.5 Flash Lite (batch)

Gemini 3.5 Flash Lite is a high-efficiency model from Google with upgraded agentic capabilities. It is suited for subagents that execute focused tasks within complex, multi-agent workflows.
1049K context standard
Meituan

Meituan: LongCat 2.0

LongCat 2.0 is a sparse mixture-of-experts language model from Meituan, with 48B active parameters out of 1.6T total. It is suited for coding, repository-level changes, long-horizon problem solving, a
1049K context 48B standard
Thinkingmachines

Thinking Machines: Inkling

Inkling is an open-weight multimodal mixture-of-experts model from Thinking Machines Lab, with 41B active parameters out of 975B total. It is designed for general-purpose reasoning, coding, agentic an
524K context 41B standard
Thinkingmachines

Thinking Machines: Inkling (free)

Inkling is an open-weight multimodal mixture-of-experts model from Thinking Machines Lab, with 41B active parameters out of 975B total. It is designed for general-purpose reasoning, coding, agentic an
1049K context 41B budget
Moonshotai

MoonshotAI: Kimi K3

Kimi K3 is a 2.8T parameter open-weight multimodal reasoning model from Moonshot AI. It is suited for complex coding, knowledge work, and long-horizon agentic workflows, and is particularly strong at.
1049K context premium
Moonshotai

MoonshotAI: Kimi K3 (batch)

Kimi K3 is a 2.8T parameter open-weight multimodal reasoning model from Moonshot AI. It is suited for complex coding, knowledge work, and long-horizon agentic workflows, and is particularly strong at.
1049K context premium
Meta

Meta: Muse Spark 1.1

Muse Spark 1.1 is a multimodal reasoning model from Meta, built for agentic tasks. It accepts text, images, video, audio, and PDF documents and returns text, with a 1M-token context...
1049K context standard
Kwaipilot

Kwaipilot: KAT-Coder-Pro V2.5

KAT-Coder-Pro V2.5 is a flagship-level Agentic Coding model that can directly hand over an entire issue or an entire business workflow to it, allowing it to autonomously locate and make...
262K context standard
OpenAI

OpenAI: GPT-5.6 Luna Pro

GPT-5.6 Luna Pro is the same underlying model as [GPT-5.6 Luna](https://ask-coreai.com), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks. Learn more in OpenAI'
1050K context standard
OpenAI

OpenAI: GPT-5.6 Luna Pro (batch)

GPT-5.6 Luna Pro is the same underlying model as [GPT-5.6 Luna](https://ask-coreai.com), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks. Learn more in OpenAI'
1050K context budget
OpenAI

OpenAI: GPT-5.6 Luna

GPT-5.6 Luna is a fast, cost-efficient model in OpenAI's GPT-5.6 series. It is suited for high-volume, latency-sensitive tasks such as chat, classification, and lightweight agentic workflows, providin
1050K context standard
OpenAI

OpenAI: GPT-5.6 Luna (batch)

GPT-5.6 Luna is a fast, cost-efficient model in OpenAI's GPT-5.6 series. It is suited for high-volume, latency-sensitive tasks such as chat, classification, and lightweight agentic workflows, providin
1050K context budget
OpenAI

OpenAI: GPT-5.6 Terra Pro

GPT-5.6 Terra Pro is the same underlying model as [GPT-5.6 Terra](https://ask-coreai.com), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks. Learn more in OpenA
1050K context premium
OpenAI

OpenAI: GPT-5.6 Terra Pro (batch)

GPT-5.6 Terra Pro is the same underlying model as [GPT-5.6 Terra](https://ask-coreai.com), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks. Learn more in OpenA
1050K context standard
OpenAI

OpenAI: GPT-5.6 Terra

GPT-5.6 Terra is a balanced model in OpenAI's GPT-5.6 series, positioned between the flagship Sol tier and the cost-efficient Luna tier. It is suited for everyday coding, reasoning, and agentic...
1050K context premium
OpenAI

OpenAI: GPT-5.6 Terra (batch)

GPT-5.6 Terra is a balanced model in OpenAI's GPT-5.6 series, positioned between the flagship Sol tier and the cost-efficient Luna tier. It is suited for everyday coding, reasoning, and agentic...
1050K context standard
OpenAI

OpenAI: GPT-5.6 Sol Pro

GPT-5.6 Sol Pro is the same underlying model as [GPT-5.6 Sol](https://ask-coreai.com), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks. Learn more in OpenAI's
1050K context standard
OpenAI

OpenAI: GPT-5.6 Sol Pro (batch)

GPT-5.6 Sol Pro is the same underlying model as [GPT-5.6 Sol](https://ask-coreai.com), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks. Learn more in OpenAI's
1050K context standard
OpenAI

OpenAI: GPT-5.6 Sol

GPT-5.6 Sol is the flagship model in OpenAI's GPT-5.6 series. It is suited for complex reasoning, coding, and agentic workflows, and is particularly strong at command-line and multi-step coding tasks.
1050K context standard
OpenAI

OpenAI: GPT-5.6 Sol (batch)

GPT-5.6 Sol is the flagship model in OpenAI's GPT-5.6 series. It is suited for complex reasoning, coding, and agentic workflows, and is particularly strong at command-line and multi-step coding tasks.
1050K context standard
xAI

SpaceXAI: Grok 4.5

Grok 4.5 is a model from SpaceXAI with frontier performance on coding, knowledge work, and STEM.
500K context standard
~x-ai

xAI: Grok Latest

This model always redirects to the latest Grok model from xAI.
500K context standard
Aion-labs

AionLabs: Aion-3.0-Mini

Aion-3.0 Mini is a multi-model roleplaying and storytelling system from AionLabs, built on the DeepSeek family of models. It uses a collaborative generation process in which multiple specialized model
131K context standard
Aion-labs

AionLabs: Aion-3.0

Aion-3.0 is a multi-model roleplaying and storytelling system from AionLabs, built on the GLM family of models. It uses a collaborative generation process in which multiple specialized models each con
131K context standard
Tencent

Tencent: Hy3

Hy3 is a 295B-parameter Mixture-of-Experts model from Tencent (21B active, 192 experts with top-8 routing) built for reasoning, agentic workflows, and real-world production use. It supports a configur
262K context 295B budget
Poolside

Poolside: Laguna XS 2.1

Laguna XS 2.1 is the latest coding agent model in the 33B-A3B category from [Poolside](https://poolside.ai/) and a step forward from their Laguna XS.2 model (released in April 2026). It combines...
262K context 33B budget
Poolside

Poolside: Laguna XS 2.1 (free)

Laguna XS 2.1 is the latest coding agent model in the 33B-A3B category from [Poolside](https://poolside.ai/) and a step forward from their Laguna XS.2 model (released in April 2026). It combines...
262K context 33B budget
Anthropic

Anthropic: Claude Sonnet 5

Sonnet 5 is Anthropic's most capable Sonnet-class model, with frontier performance across coding, agents, and professional work. It supports adaptive thinking with selectable reasoning effort levels (
1000K context standard
Anthropic

Anthropic: Claude Sonnet 5 (batch)

Sonnet 5 is Anthropic's most capable Sonnet-class model, with frontier performance across coding, agents, and professional work. It supports adaptive thinking with selectable reasoning effort levels (
1000K context standard
Sakana

Sakana: Fugu Ultra

Fugu Ultra is the higher-performance model in Sakana AI's Fugu family. Rather than a single monolithic model, Fugu is a learned multi-agent orchestration system: a language model trained to route...
1000K context premium
Cohere

Cohere: North Mini Code (free)

North Mini Code is Cohere's first agentic coding model and the debut of its North family. A sparse mixture-of-experts model with 30B total parameters and 3B active, it is optimized...
256K context 30B budget
Z-ai

Z.ai: GLM 5.2

GLM 5.2 is a large-scale reasoning model from Z.ai. It supports text input and output with a 1M-token context window, and is suited for long-horizon agent workflows, project-level software engineering
1049K context standard
Moonshotai

MoonshotAI: Kimi K2.7 Code

MoonshotAI: Kimi K2.7 Code is a coding-focused model in Moonshot AI's Kimi K2 family, built to complete end-to-end programming tasks reliably over long contexts. It uses a native multimodal mixture-of
262K context standard
~anthropic

Anthropic: Claude Fable Latest

This model always redirects to the latest model in the Claude Fable family.
1000K context premium
Anthropic

Anthropic: Claude Fable 5

Claude Fable 5 is a Mythos-class model from Anthropic, built for autonomous knowledge work and coding. It supports text, image, and file inputs with text output, with reasoning support and...
1000K context premium
Anthropic

Anthropic: Claude Fable 5 (batch)

Claude Fable 5 is a Mythos-class model from Anthropic, built for autonomous knowledge work and coding. It supports text, image, and file inputs with text output, with reasoning support and...
1000K context premium
NVIDIA

NVIDIA: Nemotron 3.5 Content Safety

NVIDIA Nemotron 3.5 Content Safety is a compact 4B-parameter multimodal guardrail model from NVIDIA, fine-tuned from Google Gemma-3-4B. It moderates both inputs to and responses from LLMs and VLMs, ac
131K context 4B budget
NVIDIA

NVIDIA: Nemotron 3.5 Content Safety (free)

NVIDIA Nemotron 3.5 Content Safety is a compact 4B-parameter multimodal guardrail model from NVIDIA, fine-tuned from Google Gemma-3-4B. It moderates both inputs to and responses from LLMs and VLMs, ac
128K context 4B budget
NVIDIA

NVIDIA: Nemotron 3 Ultra

NVIDIA Nemotron 3 Ultra is an open frontier-reasoning and orchestration model from NVIDIA, with 55B active parameters out of 550B total (MoE). Built on a hybrid Transformer-Mamba mixture-of-experts ar
203K context 55B standard
NVIDIA

NVIDIA: Nemotron 3 Ultra (free)

NVIDIA Nemotron 3 Ultra is an open frontier-reasoning and orchestration model from NVIDIA, with 55B active parameters out of 550B total (MoE). Built on a hybrid Transformer-Mamba mixture-of-experts ar
1000K context 55B budget
Qwen

Qwen: Qwen3.7 Plus

Qwen3.7-Plus is a cost-effective model in Alibaba's Qwen3.7 series. It supports text and image input with text output, building on the series' text capabilities with a comprehensive upgrade to its...
1000K context standard
Minimax

MiniMax: MiniMax M3

MiniMax-M3 is a multimodal foundation model from MiniMax. It supports text, image, and video inputs with text output, a 1M-token context window, and is suited for long-horizon agentic work, coding,...
524K context standard
Stepfun

StepFun: Step 3.7 Flash

Step 3.7 Flash is StepFun's latest high-efficiency multimodal Mixture-of-Experts model. It pairs a 196B-parameter language backbone with a vision encoder for native image and video understanding, acti
256K context 196B standard
Anthropic

Anthropic: Claude Opus 4.8

Claude Opus 4.8 is Anthropic's most capable generally available model in the Opus family. It supports text, image, and file inputs with text output, with reasoning support and a 1M-token...
1000K context premium

Popular AI Model Comparisons

Try Any AI Model Instantly

Chat with GPT-5, Claude, Gemini, and 300+ models — all in one app. Compare responses side-by-side.

Use Web App Download App →