Citizen Portal
Sign In

Get Full Government Meeting Transcripts, Videos, & Alerts Forever!

Get email alerts on the Ai Accuracy topic

No spam. Unsubscribe anytime.

Witness tells House subcommittee human oversight needed as AI can 'hallucinate' legislative facts

House Administration: House Committee · July 21, 2026
AI-Generated Content: All content on this page was generated by AI to highlight key points from the meeting. For complete details and context, we recommend watching the full video. so we can fix them.

Summary

At a House subcommittee exchange, an expert witness warned that commercial legal AI can 'hallucinate' in 17–33% of queries and urged human-centered design, citation requirements and mandatory user confirmations before relying on model outputs.

At a House subcommittee discussion about using large language models to personalize legislative information, a witness warned that current AI systems can produce inaccurate outputs and argued that human oversight and system design are essential to preserve accuracy and trust. Moderator R opened the exchange by asking how to leverage LLMs to customize legislative material while ensuring accuracy and public confidence.

The witness (transcript label: SPEAKER 6) said accuracy has multiple meanings—exact wording versus outcome—and cited a Stanford study, saying that even top-tier commercial legal AI "hallucinated 17 to 33% of queries, even when it was given the data." The witness recommended human-centered design features such as mandatory user notifications that surface items the model could not compute and checkpoints that require manual confirmation before action: "You're not allowed to move on this journey until you have confirmed A, B, C, D," the witness said, describing an approach that puts factual accountability on users and system designers.

The witness also contrasted fully generative systems with citation-forward tools, naming Perplexity as an example that attaches sources to outputs so users can verify them. In response, a questioner (transcript label: SPEAKER 1) asked whether reliance on LLMs would ever remove the need for human review; the witness replied that it depends on engineering, and reiterated that traceable sourcing and enforced UX guardrails can reduce—but not eliminate—the need for human judgment. The exchange ended with the questioner beginning an analogy to writing standards.