Ling-3.0-flash hits OpenRouter via Novita AI
Novita AI has brought Ling-3.0-flash to OpenRouter, offering developer access to the 124B-parameter Mixture-of-Experts (MoE) model with ~5.1B active parameters per token for token-efficient agentic inference. To celebrate the rollout, the model is available for free on OpenRouter through August 3.
High-efficiency MoE architectures are becoming crucial for scaling production AI agents without ballooning compute costs.
- –Activating only ~5.1B out of 124B parameters per token yields fast, cost-effective inference suitable for real-time agent workloads.
- –OpenRouter deployment via Novita AI makes it easily testable and integrable across existing developer stacks.
- –Free availability through August 3 lowers the adoption barrier for teams experimenting with new agentic backends.
DISCOVERED
3h ago
2026-07-23
PUBLISHED
3h ago
2026-07-23
RELEVANCE
AUTHOR
OpenRouter