Original Reddit post

A new execution-focused MoE from Ant’s inclusionAI: Ling-3.0-flash. 124B total, ~5.1B active (sparse MoE), 256K context, sub-100ms TTFT, and a hybrid reasoning mode you can toggle. Tuned for agent workflows — stable long-horizon tool calling and reliable instruction following. The positioning is deliberately narrow: a fast, low-cost execution node meant to pair with a larger planner, not a frontier reasoning system on its own. The trade-offs are what you’d expect for the class — obscure deep-domain knowledge and native multimodal aren’t where it’s strong. Availability: it’s on OpenRouter, API only, no open weights for this release. Free to use through August 3 submitted by /u/Loose_Bank1709

Originally posted by u/Loose_Bank1709 on r/ArtificialInteligence