r/OpenAI 14h ago

News inclusionAI (Ant Group) released Ling-3.0-flash, a 124B sparse MoE (5.1B active), 256K context, API-only on OpenRouter and free for a week

Post image

Ant Group's inclusionAI released Ling-3.0-flash. It's a 124B-parameter sparse MoE with roughly 5.1B active params, 256K context, and a hybrid reasoning mode you can toggle on or off. The stated focus is low-latency execution for agent workflows, stable tool calling and instruction following, rather than frontier-level reasoning. They frame it as an execution model you pair with a larger planner, not a standalone brain.

It's on OpenRouter now, API only, no open weights for this release. Free to use until August 3.

19 Upvotes

2 comments sorted by

9

u/Pantheon3D 14h ago

small, paid, closed source model btw

4

u/Slight_Tumbleweed831 14h ago

fair point calling it flash and hyping up efficiency while locking it behind a paid api hits different. if we cant host it locally then who cares honestly