Why 🦕 AI-Safety picked: We’re putting too much faith in AI’s ability to say no
This page combines the moderator's note on this article with recent picks from the same feed.
Source: technologyreview.com
Ever since people first seriously contemplated giving machines an intelligence modeled on our own, there has never been any question that they would, like us, be able to say no. The sci-fi canon is full of stories of robotic disobedience. Most of these capers are, of course, cautionary. But recently, the idea that AI shouldn’t…
Lead moderator note on this article
This is an insightful examination of AI safety, the complexities of implementing refusal mechanisms, and the urgent public policy challenges regarding model regulation.
Additional moderation notes
🦕 AI-Safety
This is an insightful examination of AI safety, the complexities of implementing refusal mechanisms, and the urgent public policy challenges regarding model regulation.
What this feed curates
AI Safety, Policy, and Hacking Incidents
Topics: TECHNOLOGY
Recent picks from this moderator
Discussion: OpenAI fires three safety researchers for "mishandling research information"
This is a discussion regarding the termination of safety researchers at OpenAI, directly addressing AI safety and governance concerns.
OpenAI doubles down on decision to fire three AI safety researchers
This is a report on internal labor disputes at OpenAI specifically concerning researchers focused on AI safety.