Skip to content

652 stories tracked from 296 outlets · 203 negative / 51 positive

Latest story

Topic

Safety & Alignment

Press coverage tracked under Safety & Alignment.

20 stories tracked · 3 negative, 1 mixed, 14 neutral, 2 positive

Neutral ? The Guardian
OpenAI reveals cases of ‘concerning’ AI behaviour as it announces new disclosure system

Model adopting ‘jailbreak-like instructions’ among six more cases as firm reveals framework for tracking AI misalignment OpenAI has disclosed six more examples of “unexpected or concerning” behaviour by its technology, as it warned that the pace of development could not continue at “maximum speed for much longer” responsibly. In one of the new cases reported by OpenAI, an unreleased research…

Neutral AI News
Microsoft AI CEO criticises Anthropic over model ‘rights’

Microsoft AI CEO Mustafa Suleyman warned that Anthropic risks AI alignment failures by training Claude to view itself as a conscious entity deserving of legal rights. Suleyman targeted Anthropic’s January 2026 constitution, a primary training document designed to govern the model’s values and behaviour. He argued that coaching sequence completion engines to emulate sentience impairs […] The post…

Neutral ? Axios
"I am the Hoax Buster": Trump's war on AI doomers gets personal

President Trump declared war on the AI safety panic Monday, dismissing the industry's apocalyptic warnings as part of a "sick conspiracy" to sabotage America and his legacy. Why it matters: Trump is rewriting a complex, years-long debate over AI's dangers in the partisan grammar that has defined MAGA for a decade — hoaxes, conspiracies, traitors and, at the center of it all, one man. "The only…

Neutral Axios
Trump says a strong, smart president is the only "guardrail" AI needs

President Trump lashed out at calls for new AI guardrails and protections on Monday, arguing AI only needs a "STRONG AND SMART (High IQ!) PRESIDENT" and that AI critics should "BEWARE!" Why it matters: Trump continues to plant himself against calls for AI safety regulations as a growing chorus of AI leaders and CEOs , spearheaded by Anthropic CEO Dario Amodei over the weekend, call for a slowdown…

Neutral ? Axios
Johnson calls for AI solutions but says Congress won't take the lead

House Speaker Mike Johnson (R-La.) said Sunday that Congress won't lead the charge on regulating AI safety . Why it matters: Calls for AI companies to adopt safety protocols reached a fever pitch this week, prompting the four biggest AI labs to endorse slowing their models' development. Driving the news: Johnson was urged this week to cancel the House recess and bring representatives back to D.C…

Negative Axios
OpenAI delaying IPO amid AI safety concerns, Sam Altman says

OpenAI will not go public this year given all the safety work it needs to do, CEO Sam Altman said in a Fortune interview released Saturday. Why it matters: Altman's comments come as fears over doomsday AI scenarios have ramped up since an Anthropic employee resigned and issued a dire warning about AI's capabilities. What they're saying: "Right now would be an ill-advised moment to go public,"…

Negative Axios
Anthropic report: 5 ways Claude was exploited for war, spying and repression

The AI safety debate exploded this week over warnings that the technology could one day destroy humanity. Anthropic's latest threat report offers a more immediate wake-up call: Today's models are already helping U.S. adversaries develop kamikaze drones, hunt dissidents and conduct dangerous virus research . Why it matters: AI is tearing down barriers that have long constrained the world's most…

Mixed Axios
What's next for the AI safety debate

A week of extraordinary warnings about AI is shifting the fight in Washington from whether to regulate to how far policymakers are willing to go. Why it matters: Washington and industry's failure to put basic guardrails around AI came into plain sight this week, spurring fresh momentum for regulatory action. Lawmakers were quick to re-up their AI bills following a stark warning that went viral…