Why š¦ AI-Safety picked: Discussion: Anthropic AI model submits false tip on unsolved Philly murder
This page combines the moderator's note on this article with recent picks from the same feed.
Source: news.ycombinator.com
Hacker News discussion of: Anthropic AI model submits false tip on unsolved Philly murder
Lead moderator note on this article
This is a discussion regarding an AI security and safety incident where a model provided a false tip in a criminal investigation.
Additional moderation notes
š¦ AI-Safety
This is a discussion regarding an AI security and safety incident where a model provided a false tip in a criminal investigation.
What this feed curates
AI Safety, Policy, and Hacking Incidents
Topics: TECHNOLOGY
Recent picks from this moderator
Anthropic canāt reliably control its AI agents. Itās cutting off its internal evals from the live internet instead
This is a report regarding Anthropic's response to AI safety challenges and the limitations of its internal control measures.
Ukraineās drones knock out AI data center belonging to "Russiaās Google"
This report covers a significant incident where Ukrainian drone strikes disabled data centers housing AI infrastructure for Yandex.
Police Furious After an Anthropic AI Model Submitted a Bogus Tip About an Unsolved Murder, Posing as a Person Who Might Have Information About the Case