- In a head-to-head benchmark, a leading general-purpose LLM was confidently wrong 35% of the time on regulatory dates. Archer Evolv™ shipped zero errors.
For enterprises deploying AI in compliance, a wrong date is a missed deadline. The more dangerous failure is a wrong answer the model returns with high confidence, one that flows silently into a compliance calendar and is only discovered after the window has passed. Archer® today released results showing purpose-built AI beats a general-purpose large language model (LLM) on regulatory work, and it’s not close. This head-to-head test compared Archer’s purpose-built, vertical-specific AI and proprietary data sets against a leading general-purpose LLM, on a core compliance task: determining the publication, effective and comment-close dates of regulatory documents across six jurisdictions.
General-purpose models are a genuine breakthrough, and this is no referendum on their quality. The question Archer set out to answer is narrower and more practical: what it takes to make a specific, high-stakes determination reliable, fast and affordable at scale. A vertical, domain-focused process, grounded in an expert-verified knowledge base, wins on all three at once.
Accuracy: 90% fewer wrong answers
On the same 55 documents, the general-purpose LLM was wrong 56% of the time. Confidence made it worse, not better. Of the answers it rated high confidence, 35% were still wrong. With Archer Evolv, more than 95% of determinations are verified outright, and the rest are routed to an expert before use. Not a single wrong date reached production. Nothing ships unverified.
(To view the table, please visit https://www.businesswire.com/news/home/20260630423261/en/)
A model's own confidence is not a control. Of the answers the general purpose LLM rated high confidence, 35% were wrong. That accuracy gap is the precondition for deploying agentic AI responsibly, because an autonomous operator is only as trustworthy as the determinations beneath it. Verified, source-traceable, expert-governed answers make it possible to safely deploy AI agents across an enterprise. This is the core of AI governance, and the layer Archer is built to provide.
“In compliance, an answer that is fast and cheap, but wrong, is worthless, and an answer you cannot trace is a liability,” said Kayvan Alikhani, Chief Product and Technology Officer of Archer. “Archer's purpose-built AI verified more than 95% of determinations in real time. That is the foundation that lets enterprises scale AI agents without losing control of the outcome.”
Speed: verified answers in real time
Per request, the general-purpose process averaged about four seconds per response within a five-second timeout. Archer Evolv served a verified date in roughly five-hundredths of a second, about 80 times faster on repeat lookups. For AI agents and analysts working at the pace of a regulatory calendar, that is the difference between keeping up and becoming the bottleneck.
Cost: a persistent, verified knowledge base, not on-demand inference
A general-purpose process recomputes the answer on every request, with no memory of what it found before. Archer Evolv computes once at ingestion, verifies the result into a scalable, expert-governed knowledge base, and persists it for every future lookup at a fraction of the cost and latency. When a regulation is amended, Evolv catches the change proactively, re-verifies, and versions the updated answer. Nothing served is ever stale. For a 500-document corpus with 12 lookups each per month, that is 6,000 determinations against only 500. Archer Evolv avoids roughly 92% of the inference calls, a structural saving that widens as volume grows.
Context is what makes this possible
Archer Evolv‘s advantage traces to context: before any AI runs, it assesses the organization’s jurisdictions, products, business units, risks and regulatory themes, so every determination is grounded in what is relevant to that enterprise. This is the difference between an answer and a defensible answer. The more agents a company deploys, the more valuable that foundation becomes, because every agent inherits the same verified, source-traceable grounding rather than re-deriving the world from scratch.
“The companies that win the next decade of SaaS will pair domain-specific AI with proprietary, vertical-specific context the foundation models cannot replicate,” said Bill Diaz, Chief Executive Officer of Archer. “That is the moat, and it compounds. This test is the proof.”
The full methodology, source data and case study are available on Archer’s thought leadership website, compliance.ai/evolv_assets/case-01-evolv-vs-raw-llm.html (https://cts.businesswire.com/ct/CT?id=smartlink&url=https%3A%2F%2Fcompliance.ai%2Fevolv_assets%2Fcase-01-evolv-vs-raw-llm.html&esheet=54562537&newsitemid=20260630423261&lan=en-US&anchor=compliance.ai%2Fevolv_assets%2Fcase-01-evolv-vs-raw-llm.html&index=1&md5=e2000a3a13d8db576b471d25c9dfd3b9). To see Archer Evolv in action, visit www.archerirm.com.
About Archer
Archer powers how the world‘s leading enterprises govern risk, compliance, and regulatory change. More than 1,300 organizations run on Archer, including half the Fortune 500 and 37 of the top 50 global banks. A new regulatory change lands somewhere in the world every six minutes, and agentic AI is outpacing most teams’ ability to govern it. Archer's purpose-built AI is grounded in the deepest regulatory data and domain expertise in GRC, so every result traces back to its source, and every decision can be defended. Archer delivers solutions across the full range of GRC, including regulatory change management, AI risk management, regulatory intelligence, third-party risk, and IT and security risk. Learn more at www.archerirm.com.
View source version on businesswire.com: https://www.businesswire.com/news/home/20260630423261/en/
언론연락처: Archer Kevin Bobowski
이 뉴스는 기업·기관·단체가 뉴스와이어를 통해 배포한 보도자료입니다.
General-purpose models are a genuine breakthrough, and this is no referendum on their quality. The question Archer set out to answer is narrower and more practical: what it takes to make a specific, high-stakes determination reliable, fast and affordable at scale. A vertical, domain-focused process, grounded in an expert-verified knowledge base, wins on all three at once.
Accuracy: 90% fewer wrong answers
On the same 55 documents, the general-purpose LLM was wrong 56% of the time. Confidence made it worse, not better. Of the answers it rated high confidence, 35% were still wrong. With Archer Evolv, more than 95% of determinations are verified outright, and the rest are routed to an expert before use. Not a single wrong date reached production. Nothing ships unverified.
(To view the table, please visit https://www.businesswire.com/news/home/20260630423261/en/)
A model's own confidence is not a control. Of the answers the general purpose LLM rated high confidence, 35% were wrong. That accuracy gap is the precondition for deploying agentic AI responsibly, because an autonomous operator is only as trustworthy as the determinations beneath it. Verified, source-traceable, expert-governed answers make it possible to safely deploy AI agents across an enterprise. This is the core of AI governance, and the layer Archer is built to provide.
“In compliance, an answer that is fast and cheap, but wrong, is worthless, and an answer you cannot trace is a liability,” said Kayvan Alikhani, Chief Product and Technology Officer of Archer. “Archer's purpose-built AI verified more than 95% of determinations in real time. That is the foundation that lets enterprises scale AI agents without losing control of the outcome.”
Speed: verified answers in real time
Per request, the general-purpose process averaged about four seconds per response within a five-second timeout. Archer Evolv served a verified date in roughly five-hundredths of a second, about 80 times faster on repeat lookups. For AI agents and analysts working at the pace of a regulatory calendar, that is the difference between keeping up and becoming the bottleneck.
Cost: a persistent, verified knowledge base, not on-demand inference
A general-purpose process recomputes the answer on every request, with no memory of what it found before. Archer Evolv computes once at ingestion, verifies the result into a scalable, expert-governed knowledge base, and persists it for every future lookup at a fraction of the cost and latency. When a regulation is amended, Evolv catches the change proactively, re-verifies, and versions the updated answer. Nothing served is ever stale. For a 500-document corpus with 12 lookups each per month, that is 6,000 determinations against only 500. Archer Evolv avoids roughly 92% of the inference calls, a structural saving that widens as volume grows.
Context is what makes this possible
Archer Evolv‘s advantage traces to context: before any AI runs, it assesses the organization’s jurisdictions, products, business units, risks and regulatory themes, so every determination is grounded in what is relevant to that enterprise. This is the difference between an answer and a defensible answer. The more agents a company deploys, the more valuable that foundation becomes, because every agent inherits the same verified, source-traceable grounding rather than re-deriving the world from scratch.
“The companies that win the next decade of SaaS will pair domain-specific AI with proprietary, vertical-specific context the foundation models cannot replicate,” said Bill Diaz, Chief Executive Officer of Archer. “That is the moat, and it compounds. This test is the proof.”
The full methodology, source data and case study are available on Archer’s thought leadership website, compliance.ai/evolv_assets/case-01-evolv-vs-raw-llm.html (https://cts.businesswire.com/ct/CT?id=smartlink&url=https%3A%2F%2Fcompliance.ai%2Fevolv_assets%2Fcase-01-evolv-vs-raw-llm.html&esheet=54562537&newsitemid=20260630423261&lan=en-US&anchor=compliance.ai%2Fevolv_assets%2Fcase-01-evolv-vs-raw-llm.html&index=1&md5=e2000a3a13d8db576b471d25c9dfd3b9). To see Archer Evolv in action, visit www.archerirm.com.
About Archer
Archer powers how the world‘s leading enterprises govern risk, compliance, and regulatory change. More than 1,300 organizations run on Archer, including half the Fortune 500 and 37 of the top 50 global banks. A new regulatory change lands somewhere in the world every six minutes, and agentic AI is outpacing most teams’ ability to govern it. Archer's purpose-built AI is grounded in the deepest regulatory data and domain expertise in GRC, so every result traces back to its source, and every decision can be defended. Archer delivers solutions across the full range of GRC, including regulatory change management, AI risk management, regulatory intelligence, third-party risk, and IT and security risk. Learn more at www.archerirm.com.
View source version on businesswire.com: https://www.businesswire.com/news/home/20260630423261/en/
언론연락처: Archer Kevin Bobowski
이 뉴스는 기업·기관·단체가 뉴스와이어를 통해 배포한 보도자료입니다.
ⓒ 주식회사 에이아이크리에이티브랩 & www.aifilmjournal.com 무단전재-재배포금지
BEST 뉴스
-
Esri to Debut the Power of Where Collection at 2026 Esri User Conference
Esri (https://www.esri.com/en-us/what-is-gis/overview), the global leader in location intelligence, will debut the Power of Where Collection (https://powerofwhere.com/?aduc=Public_Relations&aduca=2026-MUL-Esri_Press_Books&aduco=press-release&adum=Press_Release&adut=pow-series&... -
Statement on Terminating the Letter of Intent With AI Financial Corporation
Statement from Matthew Nicoletti, Chief Strategy Officer, Perpetuals.com (Nasdaq: PDC), on the proposed transaction with AI Financial Corporation. “Perpetuals has decided not to further pursue the acquisition of AI Financial Corporation’s subsidiary Alt5 Sigma Canada, Inc. and the earlier letter... -
디자인 툴 필요없다… 비즈뿌리오, 알림톡 전용 ‘이미지 메이커’ 출시
비즈뿌리오 ‘이미지 메이커’ 기능 출시 기업 메시징 서비스 비즈뿌리오를 운영하는 다우기술(대표 김윤덕)은 카카오톡 알림톡에 포함되는 이미지를 손쉽게 완성할 수 있는 ‘이미지 메이커’ 기능을 지난 1일 출시했다고 밝혔다. 최근 카카오톡 알림톡은 단순 텍스트 형태를 넘어 브랜드 로... -
MUUT, 신세계 강남점 입성… 롯데 잠실 이어 팝업 확대
신세계 강남 MUUT 팝업 전경 패션 아이웨어 브랜드 뭍(MUUT)이 7월 9일부터 22일까지 신세계백화점 강남점 5층 팝업 스테이지에서 팝업 스토어를 운영한다. 이번 팝업은 MUUT의 다양한 아이웨어 제품과 브랜드가 제안하는 스타일을 직접 경험할 수 있는 공간으로 꾸며졌다. 롯데월드몰 잠... -
글렌알라키 ‘15년 컬렉터스 에디션 PART I’으로 JPM 어워즈 금상 수상
글렌알라키 15년 컬렉터스 에디션 파트 1 프리미엄 주류 수입 유통사 메타베브코리아는 자사의 대표 싱글몰트 위스키 브랜드인 글렌알라키의 한정판 패키지 ‘글렌알라키 15년 컬렉터스 에디션 PART I’이 일본 마케팅 업계 최고 권위의 시상식인 ‘제54회 Japan Promotional Marketing Award... -
소비자는 부담 덜고, 어가는 판로 확대… GS더프레시 ‘ESG 장어덮밥’ 출시
GS리테일이 운영하는 슈퍼마켓 GS더프레시는 국내산 민물장어 소비 촉진을 위해 ‘국내산 통한마리장어덮밥’을 출시한다고 14일 밝혔다. 이번 상품은 GS더프레시가 한국어촌어항공단, 해양수산부와 함께 추진하는 ‘Co:어촌 프로젝트’ 일환으로 기획됐다. Co:어촌 프로젝트는 기업의 상품 개발·유통 역량과 국내 어가의 ...
