Ant Group Unveils Ling-3.0-Flash Delivering Top-Tier Performance at a Fraction of the Parameter Scal
Ant Group today announced the release of Ling-3.0-Flash, a next-generation native hybrid-reasoning foundational model engineered specifically for production-grade AI agent workflows. Designed to deliver rapid response capabilities, it serves as a high-speed execution node that offers a superior balance of intelligence density and cost-efficiency.
Featuring 124B total parameters with only 5.1B active parameters per token, Ling-3.0-Flash achieves remarkable performance despite its streamlined footprint. It matches or surpasses industry-leading models with two to three times its parameter scale across core benchmarks, including foundational reasoning, instruction following, and long-context processing.
Architectural Innovation for Efficiency
Ling-3.0-Flash moves away from the traditional approach of simply scaling parameter counts. Instead, it is built from the ground up with a native hybrid-linear attention architecture. By alternating KDA (Kimi Delta Attention) and MLA layers at a 5:1 ratio, the model optimally balances long-context efficiency with robust state memory.
Key architectural advancements include:
- Upgraded KDA: Evolving from the previous Lightning Attention, KDA introduces fine-grained diagonal gating in Delta Rule state updates, allowing the model to retain critical information more precisely when processing lengthy documents and extensive codebases.
- Optimized Mixture-of-Expert (MoE) Compute: The expert activation ratio per token has been compressed from 1/32 in the previous generation to 1/64, yielding a significantly higher "efficiency leverage."
- Extended Context Window: The model natively supports a 256K context window and can seamlessly scale to 1M tokens.
Purpose-Built for Agent Workflows
Rather than aiming to replace ultra-large, general-purpose reasoning models, Ling-3.0-Flash is designed to complete the "planning-execution separation" paradigm in AI workflows. It serves as a cost-controllable, fast, and highly stable execution node, delegating deep planning and high-frequency execution to specialized models.
To support this, Ling-3.0-Flash has been deeply refined for real-world agent scenarios, expanding its training to over 10,000 interactive environments. It features enhanced self-correction and long-horizon planning mechanisms, enabling autonomous, end-to-end delivery in complex tasks such as coding, task decomposition, and deep multi-source research. This resolves common issues of deviation or context loss in traditional models during large-scale operations.
Engineering for Speed and Stability
To ensure fast and reliable agent performance, Ant Group has paired Ling-3.0-Flash with a supporting engineering and collaboration architecture:
- Reduced Latency: A cluster-level hierarchical caching system eliminates redundant computations in long conversations and multi-turn interactions, reducing Time-to-First-Token (TTFT) for long inputs by 60% to over 80%.
- Enhanced Stability: An upgraded multi-agent collaboration architecture enables different agents to divide labor and cross-validate outputs, significantly reducing the risk of misjudgments by a single model and providing robust support for high-frequency online services.
Ling-3.0-Flash is now available on OpenRouter and Vercel AI Gateway, offering a free API through August 3, 2026. Following this limited-time free access period, the model weights will be open-sourced to support further development and innovation within the global AI community.
Developers are encouraged to integrate Ling-3.0-Flash into their coding, search, research, and tool-use workflows to experience its high-speed execution and stable tool-calling capabilities.
About Ant Group
Ant Group is a global digital technology provider and the operator of Alipay, a leading internet services platform in China, connecting over one billion users to more than 10,000 types of consumer services from partners. Through innovative products and solutions powered by AI, blockchain and other technologies, Ant Group supports partners across industries to thrive through digital transformation in an ecosystem for inclusive and sustainable development.

Ling-3.0-Flash delivers strong performance across multiple core benchmarks.
- 2024年度百强艺术家榜单——顾彤春作品鉴赏
- 中秋国庆双节同欢乡村振兴专列同庆 佛山市姜标酒业有限公司品牌献礼双节助力乡村振兴
- 冠君产业信托ESG Gala举行共创明“Teen”电影放映会
- 翔创科技携手农行可克达拉兵团分行,建“智慧粮仓”,筑粮食安全基石
- 稳正科技荣获ISO体系认证,致力于为客户提供优异的产品和服务
- 共享学术盛宴 |余萍院长受邀参加中华医学会整形外科分会第二十次全国学术交流会
- 慧商智慧:解读高慧商人的领导力秘笈!
- 3月28日,潮庭食品与您相约武汉良之隆
- 全球首创:KFSH Jeddah 单次干细胞采集方案,树立兼顾供者安全与治疗效率的全球新标杆
- 绳舞羊城,逐梦少年!2025“奔跑吧・少年”全国跳绳段位挑战赛(广东广州站)圆满落幕
- 墨香传情润中尼,刘文军书画作品成文化互鉴纽带
- 北京肾病专家排名前十
- 重磅!方芯半导体推出国产EtherCAT从站控制芯片,原位替代Microchip LAN9252/9253/9254
- 港大工程学者开发革命性的钻石制备技术
- 桐乡耀华2025-2026学年奖学金申请开启啦
- 传承中式汤品文化,苏州好得睐亮相 2026 江苏国际服务贸易展览会(马来西亚)
- 数智赋能・精准破题 临研通打通罕见病 RWS 研发前置新路径
- 全球运动品牌 U.S. Polo Assn. 在阿根廷推出男装系列
- 国寿财险苏仙区支公司召开2026年五盖山镇农业保险启动会
- 《惜花芷》热播掀热潮,田淼张婧仪母女情深引热议
- 第五届斑彩螺奖揭晓:中国香港演员郑则仕斩获最受欢迎男演员
- 艺绘骏姿 馥启新章 Creed恺芮得携手艺术家蔡赟骅共贺新年
- 锚定“大西安”新中心,西安沣东格兰云天国际酒店签约 ——焕新城市高端接待名片,构建商旅融合新高地
- 总部位于瑞典的全球最大商用车制造商之一选择HCLTech为其提供AI驱动的数字基础服务,续签并扩大原协议
- WEEX 美股狂欢:0 手续费交易,瓜分 50,000 USDT 空投
- Organon 将在 ISPOR 2026 大会公布医疗可及性与价值领域最新研究成果
- 临商银行九州支行推进金融知识普及常态化,守护金融秩序稳定
- 临商银行兰陵支行打造“适老”金融服务 助力老年群体跨越“数字鸿沟”
- 专访李明洋:以全链路运营逻辑,破解品牌全国化增长密码
- 2024亚太国际乳房整形新技术研讨会圆满落幕,集结整形权威大咖共探技术高峰





