{"id":144136,"date":"2026-08-18T21:56:08","date_gmt":"2026-08-18T21:56:08","guid":{"rendered":"https:\/\/www.europesays.com\/ai\/144136\/"},"modified":"2026-08-18T21:56:08","modified_gmt":"2026-08-18T21:56:08","slug":"government-is-measuring-how-fast-its-ai-works-not-whether-it-actually-worked","status":"publish","type":"post","link":"https:\/\/www.europesays.com\/ai\/144136\/","title":{"rendered":"Government is measuring how fast its AI works, not whether it actually worked"},"content":{"rendered":"<p>In 2013, my family had to fight a disability claim on behalf of my father, a disabled veteran, and I\u2019ve managed his care ever since. Back then, the process was manual and it took years: delay after delay, denial after denial, most of them simple mistakes from the people answering the phones. We finally hired a lawyer, who took a cut of benefits he was already owed. The part that still stays with me is that the information proving his eligibility was sitting in a database inside the same system that kept denying him. The answer was there the whole time. No one on the other end of the line owned the job of finding it.<\/p>\n<p>I don\u2019t share that to relitigate the Department of Veterans Affairs, which, like much of the federal government, has worked hard to modernize in the years since. I share it because it taught me something I\u2019ve watched hold true across a career in customer experience analytics, in and out of government: Most service systems, federal ones included, are built to do one job well, and it isn\u2019t solving your problem. It\u2019s ending your call. You reach a federal agency, an agent verifies your information, tells you what\u2019s wrong and what to do next, and moves you along. More often than not, you call back because the information was incomplete or you were routed to the wrong place the first time. The goal of the interaction is speed. It is not resolution.<\/p>\n<p>I\u2019ve named the missing piece \u201ccall ownership.\u201d When no one owns a problem end to end, the work gets passed around, agent to agent, department to department. Each handoff is another chance to lose context, repeat a question, or hand off an answer that\u2019s fast but wrong. The veteran ends up interacting with what should be a lifeline far more than they ever should.<\/p>\n<p>Here is why it continues: We measure the activity, not the outcome. A customer satisfaction score can look excellent on a dashboard while the person behind it called four or more times, re-explained the same issue to four agents and two systems, and got the wrong answer twice. A call gets routed to the wrong place, labeled resolved and closed. Ten days later the veteran discovers the information was wrong and calls back, but because the measurement window was only three to seven days, the original failure was never recorded as a failure at all. The dashboard stays green. Multiply that across millions of interactions and you get a system that reports success while people quietly call back, again and again. Because we measure activity, we miss the goal of an actual resolution.<\/p>\n<p>]]><\/p>\n<p>Now, agencies are moving artificial intelligence into these same pipelines, claims intake, contact centers, case triage. Done right, automation is exactly the fix: an AI that pulls the eligibility record my father\u2019s case needed, that carries context across a handoff, that takes ownership a human queue couldn\u2019t. Done the way we measure today, it does the opposite. It automates the bounce. It closes a fast wrong answer before anyone checks whether it was right, at a scale and speed no call center ever could. The same blind spot, now running at machine speed.<\/p>\n<p>Federal agencies are making real, hard-won progress on speed. The VA, for instance, <a href=\"https:\/\/news.va.gov\/press-room\/va-benefits-claims-backlog-drops-below-70k-for-first-time-since-2020\/\" target=\"_blank\" rel=\"noopener nofollow\">cut its claims backlog to roughly 70,000<\/a> after processing a record 3 million-plus claims in fiscal 2025. But speed is not the same as a correct answer, and most agencies still cannot show, end to end, that a given case was decided correctly.<\/p>\n<p>Federal policy already gestures at the gap but stops short of closing it. <a href=\"https:\/\/www.whitehouse.gov\/wp-content\/uploads\/2025\/02\/M-25-22-Driving-Efficient-Acquisition-of-Artificial-Intelligence-in-Government.pdf\" target=\"_blank\" rel=\"noopener nofollow\">The Office of Management and Budget\u2019s Memorandum-25-22<\/a> directs agencies to monitor AI performance and safeguard taxpayer dollars when they buy these systems. <a href=\"https:\/\/www.gao.gov\/products\/gao-21-519sp\" target=\"_blank\" rel=\"noopener nofollow\">The Government Accountability Office\u2019s AI Accountability Framework<\/a> names performance and monitoring among its core principles. <a href=\"https:\/\/www.whitehouse.gov\/wp-content\/uploads\/2025\/12\/M-26-04-Increasing-Public-Trust-in-Artificial-Intelligence-Through-Unbiased-AI-Principles-1.pdf\" target=\"_blank\" rel=\"noopener nofollow\">OMB\u2019s M-26-04<\/a> ties federal AI to public trust. Each requires \u201cperformance.\u201d None defines it as the only thing that matters to a veteran: Did the problem actually get solved?<\/p>\n<p>Agencies can close the gap in the contracts they are writing right now, while the standards are still being set. Two requirements would do most of the work.<\/p>\n<p>First, measure resolution, not activity. Pay for first-contact resolution, for claims that don\u2019t reopen, for determinations that hold up on appeal, and extend the measurement window past seven days, so a wrong answer that surfaces on day 10 counts against the system instead of vanishing from it. The layman\u2019s version belongs in the contract itself: The AI must be an asset to the citizen and to the human resolving the issue, accurate and reliable enough for a true resolution, not just a faster call.<\/p>\n<p>Second, require provenance. Every agentic system should log which model made or recommended which decision, on what data. Without that record, accountability is a guess; with it, a wrong determination can be audited, appealed and traced. That is what turns GAO\u2019s \u201cmonitoring\u201d principle and M-26-04\u2019s \u201cpublic trust\u201d from aspiration into something an agency can enforce.<\/p>\n<p>A friend of mine would put it more plainly than any of this: We\u2019re shooting at the wrong hoop. We are dealing with veterans and citizens who deserve the highest possible resolution, and for them, speed was never the measure of service, accuracy is. Some of these cases are matters of life and death \u2014 most aren\u2019t \u2014 all of them deserve a system built to solve the root of the problem rather than to get you off the phone.<\/p>\n<p>]]><\/p>\n<p>My father\u2019s answer was sitting in the database the whole time. No veteran, or anyone else, should need years and a lawyer to reach an answer the government already has.<\/p>\n<p>Brandon Burdin is the founder of MarginSignal OS. He is the caregiver to his father, a disabled veteran.<\/p>\n<p class=\"article-copyright\">Copyright<br \/>\n                            \u00a9\u00a02026 Federal News Network. All rights reserved. This website is not intended for users located within the European Economic Area.\n                    <\/p>\n","protected":false},"excerpt":{"rendered":"In 2013, my family had to fight a disability claim on behalf of my father, a disabled veteran,&hellip;\n","protected":false},"author":2,"featured_media":144137,"comment_status":"","ping_status":"","sticky":false,"template":"","format":"standard","meta":{"footnotes":""},"categories":[2],"tags":[24,25,70962,3673,59837,70963,4770,70964,4036,45255],"class_list":["post-144136","post","type-post","status-publish","format-standard","has-post-thumbnail","category-ai","tag-ai","tag-artificial-intelligence","tag-brandon-burdin","tag-customer-experience","tag-department-of-veterans-affairs","tag-federal-customer-experience","tag-government-accountability-office","tag-marginsignal-os","tag-office-of-management-and-budget","tag-veterans-benefits"],"_links":{"self":[{"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/posts\/144136","targetHints":{"allow":["GET"]}}],"collection":[{"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/posts"}],"about":[{"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/types\/post"}],"author":[{"embeddable":true,"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/users\/2"}],"replies":[{"embeddable":true,"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/comments?post=144136"}],"version-history":[{"count":0,"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/posts\/144136\/revisions"}],"wp:featuredmedia":[{"embeddable":true,"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/media\/144137"}],"wp:attachment":[{"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/media?parent=144136"}],"wp:term":[{"taxonomy":"category","embeddable":true,"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/categories?post=144136"},{"taxonomy":"post_tag","embeddable":true,"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/tags?post=144136"}],"curies":[{"name":"wp","href":"https:\/\/api.w.org\/{rel}","templated":true}]}}