I

InclusionAI: ling-2.6-flash:free

ling-2.6-flash:free
262.1K contextTools
Released Apr 28, 2026Updated Aug 10, 2026

Ling-2.6-flash is an instant (instruct) model from inclusionAI with 104B total parameters and 7.4B active parameters, designed for real-world agents that require fast responses, strong execution, and high token efficiency. It delivers performance comparable to state-of-the-art models at a similar scale while significantly reducing token usage across coding, document processing, and lightweight agent workflows.

All providers for this model are busy right now

Every upstream provider has hit its rate limit. The model comes back automatically once limits lift, usually within hours. Try again in a little while or switch to another model.

Request this model on Discord

Pricing

Input price
$0.00/ 1M tokens
Output price
$0.00/ 1M tokens
Context window 262.1K tokensCompatible endpoints openaiVendor InclusionAI

Uptime

Performance

Loading performance data...

Usage & Ranking

Loading usage...

Frequently asked questions

Similar models