inclusionai
Models published by inclusionai, available through the AnyRouter API. Each can route across multiple upstream providers for availability and price.
Ling-2.6-1T is inclusionAI's trillion-parameter flagship model, designed for real-world agents that require fast execution and high efficiency at scale.
Ling-2.6-Flash is a fast, efficient instant model from inclusionAI for advanced coding, reasoning, and agentic workflows.
Ling-3.0-flash-Fin is inclusionAI's finance-focused mixture-of-experts model, built on Ling-3.0-flash (124B total / 5.1B active). It is designed for real-world investment research, market analysis, and financial reasoning, while retaining general reasoning, coding, and agentic skills. Distinct from text-only inclusionai/ling-3.0-flash, the VL listing, and the health/medicine Sante SKU. Served free via the platform free pool, with BYOK as a fallback.
Ling-3.0-flash-Sante is inclusionAI's health and medicine-focused mixture-of-experts model, built on Ling-3.0-flash (124B total / 5.1B active). It is designed for medical knowledge reasoning, clinical safety, evidence-based retrieval, and long-horizon medical tasks, while retaining general reasoning, coding, and agentic skills. Distinct from text-only inclusionai/ling-3.0-flash and the VL listing. Served free via the platform free pool, with BYOK as a fallback. Command Code lists the same product as a free-while-it-lasts promo — that is the upstream's credit, not AnyRouter Free monthly credits.
Ling-3.0-flash-VL is inclusionAI's native multimodal instruct model, a 124B-parameter Mixture-of-Experts model with roughly 5.5B activated parameters per token. Built on Ling-3.0-flash, it adds native image and video understanding (up to 256K context) for visual reasoning, document and chart reading, and agentic GUI tasks. Distinct from the text-only inclusionai/ling-3.0-flash listing. Served free via the platform free pool, with BYOK as a fallback.
Ling-3.0-flash is inclusionAI's instant (instruct) model, a 124B-parameter Mixture-of-Experts model with roughly 5.1B activated parameters per token. It is designed with token efficiency and production-scale agentic inference as key priorities, delivering fast responses and strong execution across coding, document processing, and lightweight agent workflows. Served free via Novita's own $0 pricing, with BYOK as a fallback.
Ling 3.0 Tiny is a mixture-of-experts model from inclusionAI, with 1.3B active parameters out of 7.9B total. It is designed for responsive agents, instruction following, and multi-turn conversations, with switchable thinking and instant modes. Served free via the platform free pool, with BYOK as a fallback.
Ring-2.6-1T is inclusionAI's 1T-parameter-scale thinking model63B active parameters, built for real-world agent workflows that need both strong capability and operational efficiency. It is optimized for coding agents, tool use, and long-horizon task execution,adaptive reasoning effort (high and xhigh modes) that dynamically allocates the reasoning budget to task complexity for stronger results at lower token overhead., which lists a limited-time discount of up to 90% off July 31