Skip to main content
AI Choice EngineAI Choice Engine

InclusionAI

Ant Ling 3.0 Flash

Ant Ling 3.0 Flash is a InclusionAI model with 262K context, Not verified in / Not verified out per million tokens, last verified September 2, 2026. Open weights: yes. Public-board scores are linked to the publisher; vendor-only figures stay in the watch-out, not on the board.

Ling 3.0 Flash is Ant Group's open-weight (announced) trillion-class agentic MoE from a fintech lab.

Catalog record checked September 2, 2026Individual provider fields may change

AI model specification details
SpecificationAnt Ling 3.0 Flash
ProviderInclusionAI
TierBudget
Context window262K
Max outputNot verifiedUnverified
Input / 1M tokensNot verifiedUnverified
Output / 1M tokensNot verifiedUnverified
WeightsOpen
Parameters124B total / 5.1B active (MoE)
Reasoning levelslow, medium, high
Modalitiestext
LicenseApache 2.0
API model idling-3.0-flash
ReleasedJuly 23, 2026

Pricing tiers: Announced Apache 2.0 (weights status per coverage — verify). 124B/5.1B active MoE. Free on OpenRouter through 2026-08-03.

Verified evidence

Published benchmark results

Each result keeps its source and measurement date visible. A missing benchmark is not treated as a zero.

Published benchmark results for Ant Ling 3.0 Flash
BenchmarkScoreMeasuredSource
Artificial Analysis Intelligence Index462026-08-14Artificial Analysis · View source

Best for

  • Open-weight deployments
  • Agentic work
  • Cost-sensitive

Watch out

Confirm weights are actually published before citing as open.

Source receipts

Catalog figures for this model were checked against the following sources.

Common questions

Ant Ling 3.0 Flash

Answered from the verified figures on this page rather than general guidance.

What is Ant Ling 3.0 Flash's context window?

Ant Ling 3.0 Flash accepts about 262K tokens of context. That only matters if you routinely send very long documents, large codebases, or multi-turn histories that approach that limit.

What is Ant Ling 3.0 Flash best for?

Ant Ling 3.0 Flash is a budget tier from InclusionAI. It suits open-weight deployments, agentic work, cost-sensitive. Confirm weights are actually published before citing as open.

Can I self-host Ant Ling 3.0 Flash?

Ant Ling 3.0 Flash publishes open weights, but self-hosting depends on the licence, hardware footprint, quantisation quality, and serving stack. A hosted API is often cheaper until you have measured throughput and concurrency on your own hardware.