Policy mandateCognitive automation2025-04-03
US federal agencies must test high-impact AI before deployment under OMB M-25-21 — querying the service and observing outputs if they lack the model — and keep testing after
Software tester / QA engineeroccupation page →Event date / reported
2025-04-03
Evidence stage
Policy mandateRegulation, subsidy or public procurement is requiring or funding adoption — the mirror of a constraint. It shows adoption is being required, not that it has happened, so one mandate is never enough on its own; two independent ones are.
Tasks this bears on
Testing systems with a model inside
Evaluating software whose output is not deterministic — building evaluation sets, catching regressions in behaviour, testing for harmful outputs.
New task✓ Evidence-backed
Where this applies
A federal memorandum binding US executive agencies. Among the minimum risk management practices for high-impact AI, agencies must develop pre-deployment testing and risk mitigation plans reflecting expected real-world outcomes; where they cannot access the source code, models or data, they must use alternative test methods such as querying the AI service and observing outputs or providing evaluation data to the vendor; and they must conduct testing and periodic human review of AI use cases where feasible after deployment. It binds agencies, not testers by name, and covers high-impact uses only.
What this means
For high-impact government AI, testing the model's behaviour before and after deployment is now a written requirement — including the black-box testing of querying a service and checking what comes back. That is exactly the new testing work this task describes.
What it does not yet show
One government's rule for its own agencies; it does not show who does the testing or how well.
What you can check
Open OMB Memorandum M-25-21 and find "Agencies must develop pre-deployment testing".
Does it change the assessment?
No. The impact index is never moved by a single event. Nor did this record change a layer: all 1 linked judgement above already rested on earlier evidence. This one adds to them.
Source
US Office of Management and Budget — Memorandum M-25-21, Accelerating Federal Use of AI through Innovation, Governance, and Public Trust (April 3, 2025), section 4(b), minimum risk management practices · verified 2026-09-27 · Claude (VOLO agent) · interpreted 2026-09-27 · Claude (VOLO agent)
Primary source — published by the party that did this, or the authority of record. No co-signature needed.