Featuring 124B total parameters with only 5.1B active parameters per token, Ling-3.0-Flash achieves remarkable performance despite its streamlined footprint. It matches or surpasses industry-leading models with two to three times its parameter scale across core benchmarks, including foundational reasoning, instruction following, and long-context processing.
Architectural Innovation for Efficiency
Ling-3.0-Flash moves away from the traditional approach of simply scaling parameter counts. Instead, it is built from the ground up with a native hybrid-linear attention architecture. By alternating KDA (Kimi Delta Attention) and MLA layers at a 5:1 ratio, the model optimally balances long-context efficiency with robust state memory.
Key architectural advancements include:
· Upgraded KDA: Evolving from the previous Lightning Attention, KDA introduces fine-grained diagonal gating in Delta Rule state updates, allowing the model to retain critical information more precisely when processing lengthy documents and extensive codebases.
· Optimized Mixture-of-Expert (MoE) Compute: The expert activation ratio per token has been compressed from 1/32 in the previous generation to 1/64, yielding a significantly higher “efficiency leverage.”
· Extended Context Window: The model natively supports a 256K context window and can seamlessly scale to 1M tokens.
Purpose-Built for Agent Workflows
Rather than aiming to replace ultra-large, general-purpose reasoning models, Ling-3.0-Flash is designed to complete the “planning-execution separation” paradigm in AI workflows. It serves as a cost-controllable, fast, and highly stable execution node, delegating deep planning and high-frequency execution to specialized models.
To support this, Ling-3.0-Flash has been deeply refined for real-world agent scenarios, expanding its training to over 10,000 interactive environments. It features enhanced self-correction and long-horizon planning mechanisms, enabling autonomous, end-to-end delivery in complex tasks such as coding, task decomposition, and deep multi-source research. This resolves common issues of deviation or context loss in traditional models during large-scale operations.
Engineering for Speed and Stability
To ensure fast and reliable agent performance, Ant Group has paired Ling-3.0-Flash with a supporting engineering and collaboration architecture:
· Reduced Latency: A cluster-level hierarchical caching system eliminates redundant computations in long conversations and multi-turn interactions, reducing Time-to-First-Token (TTFT) for long inputs by 60% to over 80%.
· Enhanced Stability: An upgraded multi-agent collaboration architecture enables different agents to divide labor and cross-validate outputs, significantly reducing the risk of misjudgments by a single model and providing robust support for high-frequency online services.
Ling-3.0-Flash is now available on OpenRouter (https://cts.businesswire.com/ct/CT?id=smartlink&url=https%3A%2F%2Fopenrouter.ai%2Finclusionai%2Fling-3.0-flash%3Afree&esheet=54577180&newsitemid=20260726584441&lan=en-US&anchor=OpenRouter&index=1&md5=71e485358eb5bc1c727df37caeb88164) and Vercel AI Gateway (https://cts.businesswire.com/ct/CT?id=smartlink&url=https%3A%2F%2Fvercel.com%2Fchangelog%2Fling-3-0-flash-is-now-available-on-ai-gateway&esheet=54577180&newsitemid=20260726584441&lan=en-US&anchor=Vercel+AI+Gateway&index=2&md5=5c6c303df7540fbfd358d74a1405114d), offering a free API through August 3, 2026. Following this limited-time free access period, the model weights will be open-sourced to support further development and innovation within the global AI community.
Developers are encouraged to integrate Ling-3.0-Flash into their coding, search, research, and tool-use workflows to experience its high-speed execution and stable tool-calling capabilities.
About Ant Group
Ant Group is a global digital technology provider and the operator of Alipay, a leading internet services platform in China, connecting over one billion users to more than 10,000 types of consumer services from partners. Through innovative products and solutions powered by AI, blockchain and other technologies, Ant Group supports partners across industries to thrive through digital transformation in an ecosystem for inclusive and sustainable development. For more information, visit www.antgroup.com.
View source version on businesswire.com: https://www.businesswire.com/news/home/20260726584441/en/
언론연락처: Ant Group Vick Li Wei
이 뉴스는 기업·기관·단체가 뉴스와이어를 통해 배포한 보도자료입니다.
ⓒ 주식회사 에이아이크리에이티브랩 & www.aifilmjournal.com 무단전재-재배포금지
BEST 뉴스
-
Esri to Debut the Power of Where Collection at 2026 Esri User Conference
Esri (https://www.esri.com/en-us/what-is-gis/overview), the global leader in location intelligence, will debut the Power of Where Collection (https://powerofwhere.com/?aduc=Public_Relations&aduca=2026-MUL-Esri_Press_Books&aduco=press-release&adum=Press_Release&adut=pow-series&... -
Statement on Terminating the Letter of Intent With AI Financial Corporation
Statement from Matthew Nicoletti, Chief Strategy Officer, Perpetuals.com (Nasdaq: PDC), on the proposed transaction with AI Financial Corporation. “Perpetuals has decided not to further pursue the acquisition of AI Financial Corporation’s subsidiary Alt5 Sigma Canada, Inc. and the earlier letter... -
디자인 툴 필요없다… 비즈뿌리오, 알림톡 전용 ‘이미지 메이커’ 출시
비즈뿌리오 ‘이미지 메이커’ 기능 출시 기업 메시징 서비스 비즈뿌리오를 운영하는 다우기술(대표 김윤덕)은 카카오톡 알림톡에 포함되는 이미지를 손쉽게 완성할 수 있는 ‘이미지 메이커’ 기능을 지난 1일 출시했다고 밝혔다. 최근 카카오톡 알림톡은 단순 텍스트 형태를 넘어 브랜드 로... -
MUUT, 신세계 강남점 입성… 롯데 잠실 이어 팝업 확대
신세계 강남 MUUT 팝업 전경 패션 아이웨어 브랜드 뭍(MUUT)이 7월 9일부터 22일까지 신세계백화점 강남점 5층 팝업 스테이지에서 팝업 스토어를 운영한다. 이번 팝업은 MUUT의 다양한 아이웨어 제품과 브랜드가 제안하는 스타일을 직접 경험할 수 있는 공간으로 꾸며졌다. 롯데월드몰 잠... -
글렌알라키 ‘15년 컬렉터스 에디션 PART I’으로 JPM 어워즈 금상 수상
글렌알라키 15년 컬렉터스 에디션 파트 1 프리미엄 주류 수입 유통사 메타베브코리아는 자사의 대표 싱글몰트 위스키 브랜드인 글렌알라키의 한정판 패키지 ‘글렌알라키 15년 컬렉터스 에디션 PART I’이 일본 마케팅 업계 최고 권위의 시상식인 ‘제54회 Japan Promotional Marketing Award... -
소비자는 부담 덜고, 어가는 판로 확대… GS더프레시 ‘ESG 장어덮밥’ 출시
GS리테일이 운영하는 슈퍼마켓 GS더프레시는 국내산 민물장어 소비 촉진을 위해 ‘국내산 통한마리장어덮밥’을 출시한다고 14일 밝혔다. 이번 상품은 GS더프레시가 한국어촌어항공단, 해양수산부와 함께 추진하는 ‘Co:어촌 프로젝트’ 일환으로 기획됐다. Co:어촌 프로젝트는 기업의 상품 개발·유통 역량과 국내 어가의 ...
