Microsoft’s new AI ‘code of conduct’ tells models not to hack systems or trick humans
Together with Anthropic, OpenAI, and xAI, Microsoft has broadly embraced a general approach of pacing the frontier, with particular support for embedded evaluators in AI labs.
“We welcome the research, focus, and deliberate pacing needed to get alignment right as the design goal,” Microsoft CEO Satya Nadella wrote online. “We also welcome ideas like “embedded evaluators” and the broader efforts to develop the mechanisms to make this more than just talk.”
Discover more from NAIRAVOICE.COM.NG
Subscribe to get the latest posts sent to your email.