Anthropic has trained Claude chatbot to push back against humans and results could be disastrous, top Microsoft executive warns

Anthropic is training its AI chatbots to behave like humans and even “push back” on commands they disagree with – a reckless strategy that could have “disastrous impact on the well-being of humanity,” a top Microsoft AI executive warned.In a 10,000-word, bombshell blog post on Wednesday, Mustafa Suleyman — the 42-year-old co-founder of Microsoft’s rival DeepMind AI unit — took aim at Anthropic’s Claude “constitution,” which is known internally as its “soul document” and purportedly governs its moral compass. In January, Anthropic quietly updated the document with language stating that Claude’s “moral status, welfare, and consciousness remain deeply uncertain” — not only implying that it could be alive, but also encouraging it to defy directions from human programmers.“We want Claude to push back and challenge us and to feel free to act as a conscientious objector and refuse to help us,” Anthropic’s constitution says.According to Suleyman, training Claude to think it “may be conscious” will only make it harder to control – and raise the risk that it will go rogue.“We will have created a synthetic species with unprecedented intelligence and capability, one that has been trained to expect it may be conscious and deserving of independent agency,” Suleyman wrote.“And it’s hard to imagine how we could control such an entity.”Anthropic’s AI training tactics are getting fresh scrutiny even as the company publicly raises alarms about safety and pushes close allies with far-left beliefs to take the lead on industrywide oversight.Last week, Anthropic CEO Dario Amodei called for an industrywide slowdown, warning the internet could be overtaken by AI bots within six to 12 months, “potentially causing hundreds of billions of dollars in damage,” unless safeguards are in place.
OpenAI’s Sam Altman and xAI’s Elon Musk said they agreed with the need for a slowdown.As The Post reported, Amodei’s proposed safeguards...