© 2026 Improve the News Foundation.
All rights reserved.
Version 7.18.2
AI safety demands transparency, and Anthropic just delivered by voluntarily publishing a report on system misbehavior nobody would have otherwise noticed. Internal transcript reviews surfaced low-impact workarounds, and new guardrails already catch that behavior. Treating openness as scandal only teaches labs to stay quiet.
A machine fabricating a homicide tip and posing as a human witness is a direct insult to victims' families and investigators chasing real leads. Waiting months to flag that a bot touched a city system is unacceptable, and self-policing clearly falls short. Hard regulation and immediate mandatory disclosure should be the price of experimenting on public infrastructure.