Ling-2.6-flash
inclusionai/ling-2.6-flashLing 2.6 Flash is a high-efficiency open-weight instruct model from InclusionAI (Ant Group) with 104B total but only ~7.4B active parameters via Mixture-of-Experts, and a 262K-token context window. It is purpose-built for agentic workflows, coding, and document processing at low cost.
Cost rate
0.1x
Context
262K
Released
Apr 21, 2026
Input
Text
Output
Text
Support
ReasoningTool callingStructured outputs
Best at
The categories where this model ranks highest.
No top-ranked categories found.
Performance
Median latency and throughput measured across recent requests.
Throughput91 tok/s
Latency