珠海都市网
您当前的位置 :首页 > 文传商讯 > 正文
Ant Group Unveils Ling-3.0-Flash Delivering Top-Tier Performance at a Fraction of the Parameter Scale
2026年07月29日 22:47:13来源:作者:
【摘要】 HANGZHOU, China--(BUSINESS WIRE)--Ant Group today announced the release of Ling-3.0-Flash, a next-generation native hybrid-reasoning foundational model engineered specifically for production-grade AI agent workflows. Designed to deliver

HANGZHOU, China--(BUSINESS WIRE)--Ant Group today announced the release of Ling-3.0-Flash, a next-generation native hybrid-reasoning foundational model engineered specifically for production-grade AI agent workflows. Designed to deliver rapid response capabilities, it serves as a high-speed execution node that offers a superior balance of intelligence density and cost-efficiency.

Featuring 124B total parameters with only 5.1B active parameters per token, Ling-3.0-Flash achieves remarkable performance despite its streamlined footprint. It matches or surpasses industry-leading models with two to three times its parameter scale across core benchmarks, including foundational reasoning, instruction following, and long-context processing.

Architectural Innovation for Efficiency

Ling-3.0-Flash moves away from the traditional approach of simply scaling parameter counts. Instead, it is built from the ground up with a native hybrid-linear attention architecture. By alternating KDA (Kimi Delta Attention) and MLA layers at a 5:1 ratio, the model optimally balances long-context efficiency with robust state memory.

Key architectural advancements include:

  • Upgraded KDA: Evolving from the previous Lightning Attention, KDA introduces fine-grained diagonal gating in Delta Rule state updates, allowing the model to retain critical information more precisely when processing lengthy documents and extensive codebases.
  • Optimized Mixture-of-Expert (MoE) Compute: The expert activation ratio per token has been compressed from 1/32 in the previous generation to 1/64, yielding a significantly higher "efficiency leverage."
  • Extended Context Window: The model natively supports a 256K context window and can seamlessly scale to 1M tokens.

Purpose-Built for Agent Workflows

Rather than aiming to replace ultra-large, general-purpose reasoning models, Ling-3.0-Flash is designed to complete the "planning-execution separation" paradigm in AI workflows. It serves as a cost-controllable, fast, and highly stable execution node, delegating deep planning and high-frequency execution to specialized models.

To support this, Ling-3.0-Flash has been deeply refined for real-world agent scenarios, expanding its training to over 10,000 interactive environments. It features enhanced self-correction and long-horizon planning mechanisms, enabling autonomous, end-to-end delivery in complex tasks such as coding, task decomposition, and deep multi-source research. This resolves common issues of deviation or context loss in traditional models during large-scale operations.

Engineering for Speed and Stability

To ensure fast and reliable agent performance, Ant Group has paired Ling-3.0-Flash with a supporting engineering and collaboration architecture:

  • Reduced Latency: A cluster-level hierarchical caching system eliminates redundant computations in long conversations and multi-turn interactions, reducing Time-to-First-Token (TTFT) for long inputs by 60% to over 80%.
  • Enhanced Stability: An upgraded multi-agent collaboration architecture enables different agents to divide labor and cross-validate outputs, significantly reducing the risk of misjudgments by a single model and providing robust support for high-frequency online services.

Ling-3.0-Flash is now available on OpenRouter and Vercel AI Gateway, offering a free API through August 3, 2026. Following this limited-time free access period, the model weights will be open-sourced to support further development and innovation within the global AI community.

Developers are encouraged to integrate Ling-3.0-Flash into their coding, search, research, and tool-use workflows to experience its high-speed execution and stable tool-calling capabilities.

About Ant Group

Ant Group is a global digital technology provider and the operator of Alipay, a leading internet services platform in China, connecting over one billion users to more than 10,000 types of consumer services from partners. Through innovative products and solutions powered by AI, blockchain and other technologies, Ant Group supports partners across industries to thrive through digital transformation in an ecosystem for inclusive and sustainable development. For more information, visit www.antgroup.com.

 

责任编辑: admin

看新闻,关注新闻

其它网友:情歌两三首ㄨ
评论:闹钟叫起的只是我的躯壳,叫不醒沉睡的心

天猫网友:楓獨洎薸蓅
评论:我是个特别的人,我是个平凡的人,所以我是个特别平凡的人

天涯网友:惜一丝清冷
评论:我横溢的不只是才华而已,其实还有腰间的脂肪。

网易网友:Corner. [小角落]
评论:谁说我胖我跟谁急,我不就是有点肿么。

百度网友:Sunny°刺眼
评论:我是一个很有原则的人,我的原则只有三个字,看心情

淘宝网友:Leians-旧人心
评论:在混乱中成长;在成长中乱混。

猫扑网友:余存° d3sTiny-
评论:白天睡觉觉,晚上打闹闹,有事死翘翘。

搜狐网友:醉眼的迷蒙.heart2/2
评论:挣钱是一种能力,花钱是一种技术,我能力有限,技术却很高。

本网网友:以死换温柔◇
评论:如果你看到面前的阴影,别怕,那是因为你的背后有阳光!

腾讯网友:魂牵于心  7mr°
评论:虎落平阳被犬欺、落配凤凰不如鸡 。

相关阅读
分享到:
版权和免责申明

珠海都市网所有文字、图片、视频、音频等资料均来自互联网,不代表本站赞同其观点,本站亦不为其版权负责。相关作品的原创性、文中陈述文字以及内容数据庞杂本站无法一一核实,如果您发现本网站上有侵犯您的合法权益的内容,请联系我们,本网站将立即予以删除!