█

AI/TLDR

Ling-3.1-flash

Ant Group's roughly 560B-parameter Mixture-of-Experts model for work, coding and healthcare agents, announced 30 September 2026 with about 25B parameters active per token.

Ling (open-weight)API onlyPreview — free hosted trial; weights not yet published
Released
30 Sep 2026
Context
Up to 1M tokens (designed); 262,144 tokens served during the trial
Parameters
~560B total · ~25B activated per token (Mixture-of-Experts)
License
Not yet published

Overview

Ling-3.1-flash is a language model from inclusionAI, the team behind Ant Group's artificial-general-intelligence programme. Ant Ling announced it on 30 September 2026 as a model of about 560 billion total parameters with about 25 billion activated per token and a context window of up to 1 million tokens, and said it plans to open-source the model soon. It is the successor in the Ling line to Ling-3.0-flash, a 124B model with 5.1B active parameters, so it is several times larger in both total and active parameters.

Unlike earlier Ling releases, Ling-3.1-flash did not ship with downloadable weights. It launched as a free hosted trial: OpenRouter lists it as inclusionai/ling-3.1-flash, a text-to-text hybrid reasoning model with reasoning on by default, served by Novita at a 262,144-token context, a 32,768-token output cap and a price of zero. TechNode reported that the free trial runs for two weeks at a 256,000-token context, and that inclusionAI plans to enable the larger window and release the model as open source after the trial.

Ant Ling positions the model across work, coding and healthcare. Its launch post headlines 1,673 Elo on GDPval-AA v2.1, 75.16 on FrontierSWE and 65.35 on HealthBench Professional, and its published chart sets Ling-3.1-flash against GPT-5.6 Sol, Claude Opus 5, Kimi K3, GLM 5.3, GLM 5.3 flash and DeepSeek-V4.1-Flash on twelve benchmarks spanning agentic work, terminal and software engineering, cybersecurity, finance, healthcare and multi-turn instruction following. Ant Ling notes that HealthBench Professional was evaluated in its AQ environment, and that Ling-3.1-flash's healthcare capabilities can only be experienced in AQ.

Released2026-09-30
LicenseNot yet published
WeightsAPI only
Parameters~560B total · ~25B activated per token (Mixture-of-Experts)
ContextUp to 1M tokens (designed); 262,144 tokens served during the trial
Max output32,768 tokens (OpenRouter trial endpoint)
ArchitectureHybrid reasoning Mixture-of-Experts
ModalitiesText
StatusPreview — free hosted trial; weights not yet published

Benchmarks

Twelve bar charts comparing Ling-3.1-flash with GPT-5.6 Sol, Claude Opus 5, Kimi K3, GLM 5.3, GLM 5.3 flash and DeepSeek-V4.1-Flash on GDPval-AA v2.1, τ³-Banking, SkillsBench, Automationbench (public), Terminal-Bench 4.0, CyberGym, FrontierSWE, SWE Atlas Codebase QnA, Finance Agent v2, HealthBench Professional, DRACO and MultiChallenge, with Ling-3.1-flash's bar highlighted in blue.
Ant Ling's launch chart. HealthBench Professional was evaluated in Ant's AQ environment. — Ant Ling (inclusionAI)

Ling-3.1-flash against the models in Ant Ling's launch chart. A dash means the model was not shown for that benchmark. GDPval-AA is an Elo rating; all other rows are scores.

BenchmarkLing-3.1-flashGPT-5.6 SolClaude Opus 5Kimi K3GLM 5.3GLM 5.3 flashDeepSeek-V4.1-Flash
GDPval-AA v2.11673 Elo1588 Elo1708 Elo1524 Elo1644 Elo1641 Elo1600 Elo
τ³-Banking47.844.342.14650.347.242.5
SkillsBench68.773.563.763.659.461.371.3
Automationbench (public)52.545.850.346.748.248.854.8
Terminal-Bench 4.040.439.94912.641.932.831.2
CyberGym87.984.5—8084.5—88.1
FrontierSWE75.16——73.5274.3458.55—
SWE Atlas - Codebase QnA55.925462665961.2954.03
Finance Agent v257.8756.4862.9661.1160.4959.2661.57
HealthBench Professional65.3560.7659.849.5649.5649.0750.37
DRACO85.4977.788.677.578.178.5579.85
MultiChallenge69.7866.3264.9855.9863.8962.672.12

Comparison source ↗

This model's scores

  1. FrontierSWE75.16%
  2. HealthBench Professional65.35%
  3. CyberGym87.9%
  4. DRACO85.49%
  5. Terminal-Bench 4.040.4%

Scores on a 0–100 scale (25-point gridlines); higher is better. Each benchmark links to its published source.

Strengths

  • Roughly 560B total parameters with only about 25B active per token, according to Ant Ling
  • Designed for up to a 1M-token context; the trial serves 262,144 tokens
  • Led every model in Ant Ling's published chart on HealthBench Professional (65.35) and FrontierSWE (75.16)
  • Free to try during the launch trial, including through OpenRouter
  • Ant Ling says it plans to open-source the model

Best for

  • Agentic office and knowledge work of the kind GDPval-AA measures
  • Software-engineering agents working in terminals and real codebases
  • Healthcare question answering, which Ant Ling says is available through its AQ environment
  • Finance and banking agent workflows

How to access

ProviderModel ID
OpenRouter ↗inclusionai/ling-3.1-flash

Ling (open-weight) — every version

The full lineage of the Ling (open-weight) line, newest first. Every version has its own page — click any to compare specs, benchmarks and pricing.

VersionReleasedContextLicense
Ling-3.1-flash2026-09-30256K (trial)Not yet published
Ling-3.0-tiny2026-08—MIT
Ling-3.0-flashcurrent2026-08-02256KMIT

FAQ

Is Ling-3.1-flash open source?

Not at launch. Ant Ling's 30 September 2026 announcement says it plans to open-source the model soon, but no weights or license had been published when it launched as a free hosted trial. Earlier Ling models such as Ling-3.0-flash were released under the MIT License.

How many parameters does Ling-3.1-flash have?

About 560 billion total parameters, with about 25 billion activated per token, according to Ant Ling. It is a Mixture-of-Experts model, so only a fraction of the weights run for each token.

How long a context does Ling-3.1-flash support?

Ant Ling says it supports up to a 1-million-token context. During the launch trial it is served at a shorter window: OpenRouter lists a 262,144-token context and a 32,768-token output cap.

How can I try Ling-3.1-flash?

At launch it was available free through a hosted trial, including on OpenRouter as inclusionai/ling-3.1-flash. TechNode reported that the free trial lasts two weeks.

How does Ling-3.1-flash compare with other models?

In Ant Ling's own launch chart it scored highest of the seven models shown on HealthBench Professional (65.35) and FrontierSWE (75.16), while Claude Opus 5 led on GDPval-AA v2.1, Terminal-Bench 4.0 and DRACO, and GPT-5.6 Sol led on SkillsBench. These are the maker's figures.