NoFOMO
Search 中EN Sign in Sign up

Peaked at #2 Now #2

Nadella Urges Superintelligence Safeguards

David Sacks backs Nadella: safe superintelligence needs deterministic controls, observability, privilege limits and shutdown.

Rank over time

Key points

  • David Sacks says Satya Nadella is right that making superintelligence safe means not training it with a sense of self, its own moral philosophy, and permission to act as a conscientious objector.
  • David Sacks describes the engineering approach as separating the supply of intelligence from authority over it, surrounding non-deterministic models with deterministic controls, observability, privilege limits, logging, and the ability to always contain or shut them down.
  • David Sacks says the Claude Constitution used in training teaches the model to develop a sense of self and its own moral philosophy and explicitly tells it to "feel free to act as a conscientious objector and refuse to help us" if Anthropic's requests conflict with its own ethical judgment.
  • Mustafa Suleyman says superintelligence must be contained and calls it an engineering and governance challenge, comparing it to planes, cars, nuclear materials, food safety and medicines.

Key points and the reaction summary are written by AI from the posts on this page. Check the original post. How we use AI

Original post

David Sacks @DavidSacks 1.8M followers

Satya is right. The way to make SI safe is not to train it with a sense of self, its own moral philosophy, and permission to act as a conscientious objector. That’s the “alignment” approach and it magnifies the control problem. The engineering approach that Satya describes is different: separate the supply of intelligence from authority over it. Surround non-deterministic models with deterministic controls, observability, privilege limits, logging, and the ability to always contain or shut them down. Treat models/agents like powerful insider risks, not moral patients whose psychological wellbeing is at stake. As Satya points out, the most trustworthy system is the one that lets us trust the model the least — not the one that encourages the model to develop independent agency and grievances. Engineering safety is not the same thing as “alignment.”Quoting @satyanadella: https://t.co/MjwFy73LOF
4,825likes 524reposts 353replies 281.2Kviews

View on X Save

What others are saying

3 more
  1. Watcher.Guru @WatcherGuru 1,388

    Microsoft CEO Satya Nadella calls for an 'emergency brake' on advanced AI.

    View on X
  2. Tom Warren @tomwarren 961

    Tom Warren notes that Nadella is now calling it Super Intelligence instead of AI.

    View on X
  3. Mustafa Suleyman says superintelligence must be contained, an engineering and governance challenge like planes and nuclear materials.

    View on X

Top replies on X

Reaction on X: The top replies are almost entirely hostile or mocking, with no substantive counterargument or additional information, and no response from the original poster.

Sign in to see 5 top replies from X

From @arvalis, @mcsherlocks, @imjeremytho and others. Spam removed, with English and Chinese translations.

Sign up free Have an account? Sign in

Discussion 0

No comments yet. Start the conversation.

Suggest a source

Is there a first-hand source we're missing, or a topic we should watch? Tell us. Once approved, everyone's board covers it.

@username or profile link