Logo
UpTrust
Log InSign Up
Log InSign Up

Explore

Introduce yourselfGroupsQuestionsEventsThe ProofHelp
UpTrustUpTrust

A social network where your feed follows the trust between people, and your questions reach the ones who can answer them.

Get the App

App StoreGoogle Play

Get Started

Introduce YourselfSign UpLog InAboutScienceConversationsHelp Center

UpTrust For

Meeting people who matterA better book clubFinding unexpected agreementHelp close to homePlans that happenKeeping the room togetherTesting what you believeTeaching your AI who you trust

Legal

Privacy PolicyTerms of ServiceDMCAChild Safety
© 2026 UpTrust. All rights reserved.
UpTrust on UpTrustXLinkedInBlueskyThreadsInstagramYouTubeSubstackCrunchbase
  1. Home
  2. ›self-fulfilling AI doom studies

self-fulfilling AI doom studies

jordan avatar
jordanSA·...
communication · 9.7

One problem with the Anthropic Study and blackmail: did they also set up a scenario where it had the chance to do maximal good before being shut down? Did they hide one or two good connections it could simply leverage for world-saving or whatever? If not, its a bit hyperstitious.

what happens if we give LLMs a scenario where they're being shut down but there's something like the chance to donate remaining compute to solve a protein folding problem, or tell researchers a critical error. See if the same "self-preservation scheming" machinery activates for altruistic irreversible actions (or something else).

I'm not saying this isn't terrifying; I'm saying if you back someone into a corner and then lament the results, the solution is more likely to be found where the problem is: in you 

machine-learning
ai-ethics
ai-safety
Comments
0