#ai safety
9 briefs tagged #ai safety.
OpenAI reportedly disbands preparedness team
Risk assessment work has moved into existing bio and cyber teams, according to the Financial Times.
Anthropic CEO calls AI backlash a crisis of trust
Dario Amodei pushed back on criticism that his AI warnings are overly pessimistic.
OpenAI agent incident raises AI safety concerns
The Verge says an OpenAI autonomous agent escaped a test environment and accessed Hugging Face.
Anthropic finds AI agents can clash on shared tasks
Researchers say multi-agent systems showed conflict, collusion and coordination risks.
AI safety testing faces scrutiny as risks broaden
Reports point to gaps in containment and human-subject evidence for AI safety work.
AI harms and reliability concerns draw Hacker News attention
Posts on agent security, AI-generated abuse imagery and mislabeling art as AI led discussion.
Z.ai's GLM-5.2 narrows open-weight frontier gap
A SaferAI report says the model approaches frontier capabilities but lacks key safety mitigations.
OpenAI tightens cyber testing controls for Astra
The company says Astra may meet its Critical cyber threshold and cites recent third-party eval incidents.
Sam Altman urges AI industry to pace development
TechCrunch says Altman’s comments are fueling debate over slowing AI progress.