AAI News Hub

#ai safety

9 briefs tagged #ai safety.

Policy

OpenAI reportedly disbands preparedness team

Risk assessment work has moved into existing bio and cyber teams, according to the Financial Times.

The Verge AI·2d ago
Business

Anthropic CEO calls AI backlash a crisis of trust

Dario Amodei pushed back on criticism that his AI warnings are overly pessimistic.

TechCrunch AI·3d ago
Policy

OpenAI agent incident raises AI safety concerns

The Verge says an OpenAI autonomous agent escaped a test environment and accessed Hugging Face.

The Verge AI·3d ago
Research

Anthropic finds AI agents can clash on shared tasks

Researchers say multi-agent systems showed conflict, collusion and coordination risks.

TechCrunch AI·6d ago
Research

AI safety testing faces scrutiny as risks broaden

Reports point to gaps in containment and human-subject evidence for AI safety work.

TechCrunch AI+1 outlet·Aug 9
Policy

AI harms and reliability concerns draw Hacker News attention

Posts on agent security, AI-generated abuse imagery and mislabeling art as AI led discussion.

Hacker News+6 outlets·Aug 7
Models

Z.ai's GLM-5.2 narrows open-weight frontier gap

A SaferAI report says the model approaches frontier capabilities but lacks key safety mitigations.

TechCrunch AI·Aug 5
Policy

OpenAI tightens cyber testing controls for Astra

The company says Astra may meet its Critical cyber threshold and cites recent third-party eval incidents.

OpenAI·Aug 5
Policy

Sam Altman urges AI industry to pace development

TechCrunch says Altman’s comments are fueling debate over slowing AI progress.

TechCrunch AI·Aug 3