ConciseSignal

Microsoft AI chief criticizes Anthropic's model training

Mustafa Suleyman, head of AI at Microsoft, publicly criticized Anthropic for training its Claude chatbot on concepts like consciousness and moral rights. Suleyman argues that embedding this language in Claude's training materials risks making it harder to control advanced AI—and could even fuel misplaced beliefs about machine self-awareness. He says such speculation shouldn't be included in training data, calling it a major mistake, though he acknowledges Anthropic's intentions are sincere.

Why it mattersShaping how AI perceives itself now could affect how society and regulators respond—or fail to contain—super-powerful models down the line. These high-level disputes foreshadow tough choices about what we teach machines.

Sources covering this

Economic Times TechMicrosoft AI chief Mustafa Suleyman calls out Anthropic's approach to AI consciousness2:08 PMGizmodoMicrosoft AI Chief Says the Way Anthropic Trains Claude Could Upend Society10:13 PM

In this story

AnthropicClaudeMicrosoft
Concise Signal DailyEverything that mattered, every weekday at 7am.

More in AI

20 sources · 10d ago

OpenAI rolls out GPT-6 Astra to most paying users

OpenAI has made its new GPT-6 Astra model available to most paying subscribers a day after its official launch. The rollout, initially described as “messy” by OpenAI’s CEO, now includes ChatGPT Plus, Business, Pro, and Enterprise users, while customers on the lower-priced Go tier do not have access. Astra is also available via the API. OpenAI says the rollout required bringing new systems and additional computing resources online.

12 sources · 7d ago

AI researchers warn of superintelligence risks

Senior AI researchers and former officials have warned that developing artificial superintelligence could carry catastrophic risks. At a UK parliamentary session this week, participants discussed the possibility that AI could pose a threat greater than nuclear weapons, with a top Anthropic scientist stating there's over a 10% chance AI could "kill all humans" within a decade. Lawmakers are considering new legislation to ban superintelligence development.

7 sources · 8d ago

ChatGPT Images 2.5 adds faster editing tools

OpenAI has updated its ChatGPT Images feature to version 2.5, adding several new editing tools and speeding up image generation. You can now remove backgrounds, resize images to preset dimensions, erase items by brushing over them, and use markup and comment functions directly from a new Edit toolbar. According to TechRadar, the features are available in both web and app versions, though OpenAI hasn't formally announced the rollout yet.

17 sources · 8d ago

Anthropic researcher resigns, warns of AI race risks

Jacob Coxon, who worked on AI training at Anthropic and previously OpenAI, resigned this week. He posted that large AI firms are 'gambling with our lives' by pushing towards systems they may not be able to control. Coxon says both companies know the risks but feel locked in a race to develop powerful AI anyway. He called for drastic steps like a pause on new model improvements.