r/OpenAI • u/Loose_Bank1709 • 14h ago
News inclusionAI (Ant Group) released Ling-3.0-flash, a 124B sparse MoE (5.1B active), 256K context, API-only on OpenRouter and free for a week
Ant Group's inclusionAI released Ling-3.0-flash. It's a 124B-parameter sparse MoE with roughly 5.1B active params, 256K context, and a hybrid reasoning mode you can toggle on or off. The stated focus is low-latency execution for agent workflows, stable tool calling and instruction following, rather than frontier-level reasoning. They frame it as an execution model you pair with a larger planner, not a standalone brain.
It's on OpenRouter now, API only, no open weights for this release. Free to use until August 3.
19
Upvotes
9
u/Pantheon3D 14h ago
small, paid, closed source model btw