*Ling-3.0-flash* is a *124B-parameter Mixture-of-Experts (MoE) model*, with approximately *5.1B parameters activated per token*. The model is designed with *token efficiency and production-scale agentic inference* as key priorities, enabling developers...
With CoreAI, you can start chatting with Ling-3.0-flash instantly — no separate subscription needed. CoreAI lets you chat with Ling-3.0-flash alongside 300+ other AI models from Inclusionai and other providers like OpenAI, Anthropic, Google, Meta, and more.
Measured against all 386 chat models CoreAI serves, Ling-3.0-flash sits at #155 for context capacity — 262K tokens, roughly 196,608 words or 393 pages of text in a single conversation. On price, it is the #4 cheapest of 125 budget-tier models by input cost. Released 2026-07-23, it is one of the 15 newest additions to the catalog. It supports step-by-step reasoning (thinking mode).
Every model in this table is available on CoreAI — open the web app and run the same prompt through Ling-3.0-flash and inclusionAI: Ling-2.6-flash side by side to judge with your own eyes.
A typical question (about 1,000 tokens in, 500 out) runs $0.00005 on Ling-3.0-flash's raw per-token rates; a heavy session with a 50K-token document costs about $0.001. On CoreAI you don't juggle these numbers — one subscription covers Ling-3.0-flash and 300+ other models under a plan usage limit, with no API keys. See plans.
Ling-3.0-flash handles 262K tokens of context — roughly 196,608 words, or about 393 pages of text in one conversation.
Ling-3.0-flash's underlying rates are $0.021 per 1M input tokens and $0.063 per 1M output tokens. On CoreAI it is included in one subscription alongside 300+ models, so there is no per-token billing to manage.
No — Ling-3.0-flash is text-only. If you need image understanding, its budget-tier peers with vision are listed in the comparison table above.
Yes. Ling-3.0-flash can work through problems step by step before answering; on CoreAI you can toggle thinking mode per message.
Open the CoreAI web app or the iOS/Android app, pick Ling-3.0-flash in the model selector, and chat — no API key, no per-token billing. One subscription covers Ling-3.0-flash and 300+ other models.
The closest alternatives on CoreAI are inclusionAI: Ling-2.6-flash, Poolside: Laguna S 2.1 (free), IBM: Granite 4.0 Micro — similar tier and capability. The fastest way to choose is running one prompt through them side by side in CoreAI's Compare view.
Chat with Ling-3.0-flash and 300+ other AI models — all in one app.