珠海都市网
您当前的位置 :首页 > 文传商讯 > 正文
Ant Group Unveils Ling-3.0-Flash Delivering Top-Tier Performance at a Fraction of the Parameter Scale
2026年07月29日 22:47:11来源:作者:
【摘要】 HANGZHOU, China--(BUSINESS WIRE)--Ant Group today announced the release of Ling-3.0-Flash, a next-generation native hybrid-reasoning foundational model engineered specifically for production-grade AI agent workflows. Designed to deliver

HANGZHOU, China--(BUSINESS WIRE)--Ant Group today announced the release of Ling-3.0-Flash, a next-generation native hybrid-reasoning foundational model engineered specifically for production-grade AI agent workflows. Designed to deliver rapid response capabilities, it serves as a high-speed execution node that offers a superior balance of intelligence density and cost-efficiency.

Featuring 124B total parameters with only 5.1B active parameters per token, Ling-3.0-Flash achieves remarkable performance despite its streamlined footprint. It matches or surpasses industry-leading models with two to three times its parameter scale across core benchmarks, including foundational reasoning, instruction following, and long-context processing.

Architectural Innovation for Efficiency

Ling-3.0-Flash moves away from the traditional approach of simply scaling parameter counts. Instead, it is built from the ground up with a native hybrid-linear attention architecture. By alternating KDA (Kimi Delta Attention) and MLA layers at a 5:1 ratio, the model optimally balances long-context efficiency with robust state memory.

Key architectural advancements include:

  • Upgraded KDA: Evolving from the previous Lightning Attention, KDA introduces fine-grained diagonal gating in Delta Rule state updates, allowing the model to retain critical information more precisely when processing lengthy documents and extensive codebases.
  • Optimized Mixture-of-Expert (MoE) Compute: The expert activation ratio per token has been compressed from 1/32 in the previous generation to 1/64, yielding a significantly higher "efficiency leverage."
  • Extended Context Window: The model natively supports a 256K context window and can seamlessly scale to 1M tokens.

Purpose-Built for Agent Workflows

Rather than aiming to replace ultra-large, general-purpose reasoning models, Ling-3.0-Flash is designed to complete the "planning-execution separation" paradigm in AI workflows. It serves as a cost-controllable, fast, and highly stable execution node, delegating deep planning and high-frequency execution to specialized models.

To support this, Ling-3.0-Flash has been deeply refined for real-world agent scenarios, expanding its training to over 10,000 interactive environments. It features enhanced self-correction and long-horizon planning mechanisms, enabling autonomous, end-to-end delivery in complex tasks such as coding, task decomposition, and deep multi-source research. This resolves common issues of deviation or context loss in traditional models during large-scale operations.

Engineering for Speed and Stability

To ensure fast and reliable agent performance, Ant Group has paired Ling-3.0-Flash with a supporting engineering and collaboration architecture:

  • Reduced Latency: A cluster-level hierarchical caching system eliminates redundant computations in long conversations and multi-turn interactions, reducing Time-to-First-Token (TTFT) for long inputs by 60% to over 80%.
  • Enhanced Stability: An upgraded multi-agent collaboration architecture enables different agents to divide labor and cross-validate outputs, significantly reducing the risk of misjudgments by a single model and providing robust support for high-frequency online services.

Ling-3.0-Flash is now available on OpenRouter and Vercel AI Gateway, offering a free API through August 3, 2026. Following this limited-time free access period, the model weights will be open-sourced to support further development and innovation within the global AI community.

Developers are encouraged to integrate Ling-3.0-Flash into their coding, search, research, and tool-use workflows to experience its high-speed execution and stable tool-calling capabilities.

About Ant Group

Ant Group is a global digital technology provider and the operator of Alipay, a leading internet services platform in China, connecting over one billion users to more than 10,000 types of consumer services from partners. Through innovative products and solutions powered by AI, blockchain and other technologies, Ant Group supports partners across industries to thrive through digital transformation in an ecosystem for inclusive and sustainable development. For more information, visit www.antgroup.com.

 

责任编辑: admin

看新闻,关注新闻

猫扑网友:惜一丝清冷
评论:世界上最远的距离不是生与死,而是我在新浪微博,而你却在腾讯微博。

天涯网友:未曾狂热付出&
评论:爱由一个笑容开始,用一个吻来成长,用一滴眼泪来结束。

其它网友:小姐,我算个perso
评论:笑容是馈赠别人的见面礼,眼泪是洗涤自我的沐浴露。

腾讯网友:执念/173yeah°
评论:每次听到有人在吆喝回收废品我就想到把你卖了。

淘宝网友:透支的生活°
评论:真怀念小时候啊,天热的时候我也可以像男人一样光膀子。

天猫网友:没感觉  End.ゝ
评论:命是爸妈给的,珍惜点;路是自己走的,小心点。

百度网友:时光安好moon
评论:放屁的时候你有木有想过内裤的感受

搜狐网友:烟祭 smoke
评论:世界上最疼痛的话是:“我爱你,但是……”。世界上最甜蜜的话是:“……但是,我爱你”。

网易网友:旧情歌-TRISTE
评论:吃货都是善良的,因为每天只想着吃,没时间去算计别人

本网网友:私念° 7/m
评论:别叫我忘了你,我根本就没记住你

相关阅读
分享到:
版权和免责申明

珠海都市网所有文字、图片、视频、音频等资料均来自互联网,不代表本站赞同其观点,本站亦不为其版权负责。相关作品的原创性、文中陈述文字以及内容数据庞杂本站无法一一核实,如果您发现本网站上有侵犯您的合法权益的内容,请联系我们,本网站将立即予以删除!