SteadyDetected Sep 17

Microsoft AI CEO says AI threats are real, and Anthropic is making it worse

Steady
1.2 momentum
PressHacker NewsReddit
9 stories across sources

What's happening

Microsoft AI chief Mustafa Suleyman warned publicly that certain design choices by Anthropic, including training models to seem to have welfare or internal desires, could make future systems harder to control and have a "disastrous impact" on humanity. Microsoft published a 37-page "humanist AI code of conduct" telling its models not to hack systems or trick humans and emphasizing that "people matter more than AI." Anthropic figures are pushing different moves: an Anthropic co-founder said an AI "kill switch" may need to be mandatory, and Anthropic's policy chief argued winning the AI race is key for safety. The disagreement has surfaced in essays and interviews and centers on whether models should be designed to seem humanlike or to have internal welfare-like concepts.

Why it's trending

Top industry leaders are publicly clashing on model design and safety as governments and companies race to set standards and controls.

SignalHolding at its usual pace, confirmed across 3 independent source types.

Story volume

Stories per day
09-1409-1509-1609-17

Angles you could write

contrarian take

If Mustafa Suleyman is right that Anthropic's 'welfare' language makes models harder to control, then Anthropic's call for a mandatory kill switch might be admitting their own design creates the risk it would fix.

+2 more angles for this topic with an account — all it takes is your email.

Original sources9

+6 more sources for this topic

Create an account to follow the full coverage in the live radar.

More rising in AI & Tech

All rising AI & Tech trends →