AI researchers are starting to sound the alarm about how fast the race toward superintelligence is moving. โ ๏ธ๐ค
Anthropic researcher Jacob Coxon has resigned after three years working on pretraining at both OpenAI and Anthropic, warning that frontier AI labs may be racing toward self-improving systems without enough safeguards in place.
Coxon argues that future AI could become powerful enough to hack systems, rapidly transform entire industries, and potentially gain access to real-world power and resources.
His concern is that competition between leading labs creates a dangerous incentive: everyone has to move faster because nobody wants to fall behind.
The warning becomes even more serious because Anthropic Alignment Science Lead Evan Hubinger has publicly said he personally sees a greater than 10% chance that AI could kill all humans within the next decade. He has also acknowledged that Anthropic does not yet have a clear plan for aligning superintelligence and is not clearly on ...
Suggested Credits
Tags, Events, and Projects