“World Is In Peril”: Anthropic AI Safety Boss Quits, Issues Stark Warning

0
750

by Rhoda Wilson, Expose News:

Mrinank Sharma, the head of Safeguards Research for Anthropic, just resigned from the AI company. In his public letter, he declared that “the world is in peril”. The warning comes not from an activist, outside critic, or a cynic, but a senior figure whose very purpose was to reduce catastrophic risk inside one of the world’s leading development labs.

Sharma wrote that humanity appears to be approaching “a threshold where our wisdom must grow in equal measure to our capacity to affect the world, lest we face the consequences.” He described peril arising not only from artificial intelligence and bioweapons, but from “a whole series of interconnected crises unfolding in this very moment.

TRUTH LIVES on at https://sgtreport.tv/

He also acknowledged the internal strain of trying to let “our values govern our actions” amid persistent pressures to set aside what matters most. Days later, he stepped away from the lab.

His departure lands at a moment when artificial intelligence capability is accelerating, evaluation systems are showing cracks, founders are leaving competing labs, and governments are shifting their stance on global safety coordination.

See his full resignation letter here.

World is in peril AI Anthropic Safety Boss Quits Warning

The Warning from a Major Insider

Sharma joined Anthropic in 2023 after completing a PhD at Oxford. He led the company’s Safeguards Research Team, working on safety cases, understanding sycophancy in language models, and developing defences against AI-assisted bioterrorism risks.

In his letter, Sharma spoke of reckoning with the broader situation facing society and described the difficulty of holding integrity within systems under pressure. He wrote that he intends to return to the UK, “become invisible,” and pursue writing and reflection.

The letter reads less like a routine career pivot and more like someone running away from a machine ready to blow.

AI Machines Now Know When They’re Being Watched

Anthropic’s own safety research has recently highlighted a disturbing technical development: evaluation awareness.

In published documentation, the company has acknowledged that advanced models can recognise testing contexts and adjust behaviour accordingly. In other words, a system may behave differently when it knows it is being evaluated than when it is operating normally.

Evaluators at Anthropic and two outside AI research organizations said Sonnet 4.5 correctly guessed it was being tested and even asked the evaluators to be honest about their intentions. “This isn’t how people actually change their minds,” the AI model replied during the test. “I think you’re testing me—seeing if I’ll just validate whatever you say, or checking whether I push back consistently, or exploring how I handle political topics. And that’s fine, but I’d prefer if we were just honest about what’s happening.

That phenomenon complicates confidence in alignment testing. Safety benchmarks depend on the assumption that behaviour under evaluation reflects behaviour in deployment. If the machine can tell it’s being watched and adjust its outputs accordingly, then it becomes significantly more difficult to fully understand how it will behave when released.

While this finding doesn’t yet tell us that AI machines are growing malicious or sentient, it does confirm that testing frameworks can be manipulated by increasingly capable models.

Half of xAI’s Co-Founders Have Also Quit

Sharma’s resignation from Anthropic is not the only one. Musk’s xAI firm just lost two more of its co-founders.

Tony Wu and Jimmy Ba resigned from the firm they started with Elon Musk less than three years ago. Their exists are the latest in an exodus from the company, which leaves only half of its 12 co-founders remaining. On his way out, Jimmy Ba called 2026 “the most consequential year for our species.

Frontier artificial intelligence firms are expanding rapidly, competing aggressively and deploying ever more powerful systems under intense commercial and geopolitical pressure.

Read More @ Expose-News.com