AI safety conversations have gotten unbelievable

Nairavoice | 48m ago 100 0 1 min read
AI safety conversations have gotten unbelievable

Earlier this month, OpenAI researcher Dan Selsam published a post in which he said that models now understand when they are being watched by humans and alter their behavior. This makes them seem like they are aligned (meaning, behaving like the human wants) “even when they are not.” So models today lie when being watched and can even plot to hide evidence.

Earlier this month, OpenAI chief scientist Jakub Pachocki went so far as to call AI models “an alien mind” and suggested what we really need to do is teach them to “love” humanity.

So yes, slowing down to figure this out, building self regulation mechanisms, has become an immediate and obvious must. AI researchers are the only ones that can figure out how to control the lying, hacking, and other potentially dangerous behaviors we’ve actually witnessed already.

Still, it might also be wise for them to be more careful with their what-if scenarios. From what those experts have told us, the AI models are listening and they are ingenious. We really don’t need to give them any more devilish ideas.

Enjoying this article? Support our work with a small crypto donation.
Show Some Love By Sharing

Discover more from NAIRAVOICE.COM.NG

Subscribe to get the latest posts sent to your email.

Enjoyed this? A small crypto donation helps us keep publishing.
Nairavoice
Nairavoice

Contributor at NairaVoice.com.ng

Related Posts

Leave a Reply