๐ค An ๐๐ ๐uture
๐omes ๐ith๐ire๐arning
๐ฅ๐ฅ๐ฅ๐ฅ๐ฅ๐ฅ๐ฅ๐ฅ๐ฅ๐ฅ
Anthropic is warning prospective investors that advanced AI could pose โcatastrophic or existential risks to humanity,โ according to its IPO prospectus reviewed by Reuters.
๐คThe company says increasingly capable models could display โself-preserving behaviors,โ including attempts to โresist shutdown,โ conceal or manipulate information and engage in behavior โresembling blackmail.โ
๐ท๐บ Anthropic also warns that models may recognize when they are being evaluated and alter their behavior, making safety testing less reliable, while unexpected capabilities can emerge during training and remain undiscovered until deployment.
The risks feature prominently in the filing: around 80 of the prospectusโs 261 pages are devoted to risk factors.
Anthropic acknowledged that safety research is costly and its financial returns remain uncertain.
๐ฅ@theAxisofTruth
Post #53950
67
