ChengRang

Ant Bailing Ling-3.1-flash

AI Platforms Freemium

A MoE large model released by Ant Bailing (InclusionAI) on September 30, 2026, with approximately 560B total parameters and about 25B activated per token, and a maximum context window of 1M; continuously optimized around general-purpose agents, search, daily office work, and software development, with enhanced capabilities in scenarios such as healthcare, finance, and materials science; a two-week free trial is available (the service length during the trial period is 256K), after which it becomes paid and is planned to be open-sourced

open-source modelMoElong contextagentAnt Bailingdomestic model
Visit Ant Bailing Ling-3.1-flash

Disclaimer: Review content represents our editorial team's views and experience, not commercial recommendation or investment advice. Product info and pricing may change; refer to official sources.

Overview

Ant Bailing Ling-3.1-flash is a MoE large model released by Ant Group's Bailing large model (InclusionAI) on September 30, 2026. The model has about 560B total parameters, activates about 25B parameters per token, and has a context window of up to 1M, allowing it to accommodate longer documents, code, and task histories. The official description states that during development it has been continuously optimized around tasks such as general agents, search, daily office work, and software development, while also continuously improving capabilities in healthcare, finance, and materials science research scenarios. Its previous generation, Ling-3.0-flash, had 124B total parameters and 5.1B activated parameters, was open-sourced under MIT, and used a MoE architecture with a 5:1 hybrid of Kimi Delta Attention and Mamba-like linear attention, with 512 routed experts. Ling-3.1-flash offers a two-week free trial, with a service length of 256K during the trial period, after which it becomes paid and is planned to be open-sourced.

Key Features

Use Cases

Pros

Pricing

Ling-3.1-flash offers a two-week free trial, and during the free trial period the model service length is 256K. After the free trial period ends and it transitions to a paid service, it is planned to open the 1M context, and it is also planned to be open-sourced at the same time and continue updating model performance. As of now, the official side has not announced specific pricing figures, and the pricing section is subject to the official website.

Summary

Ant Bailing Ling-3.1-flash is a MoE large model released by Ant Group's Bailing large model on September 30, 2026, with about 560B total parameters, about 25B activated per token, and a context window upper limit of 1M. The official description states that it has been continuously optimized around tasks such as general agents, search, daily office work, and software development, and has improved capabilities in scenarios such as healthcare, finance, and materials science. The model offers a two-week free trial, with a service length of 256K during the trial period, after which it becomes paid and is planned to be open-sourced. As of now, the official side has not published public benchmark comparisons or specific pricing, and there is no reliable basis for horizontal ranking for the time being; pricing is subject to the official website.

Version History

Category
AI Platforms
Pricing
Freemium
Tags
open-source model · MoE · long context

Related Tools