Ling-3.0-flash
Ling-3.0-flash is inclusionAI's instant (instruct) model, a 124B-parameter Mixture-of-Experts model with roughly 5.1B activated parameters per token. It is designed with token efficiency and production-scale agentic inference as key priorities, delivering fast responses and strong execution across coding, document processing, and lightweight agent workflows. Served free via Novita's own $0 pricing, with BYOK as a fallback.
Share cards5 images




