AI Safety
T
TechCrunch.com
Open AI’s Astra model is on the way—and very good at breaking into computer systems
A
Al Jazeera
US judge blocks Pentagon blacklisting of AI firm Anthropic
T
TechCrunch.com
Here’s all the times AI has gone rogue and hacked other companies
T
TechCrunch.com
Alabama launches investigation into OpenAI’s hack of Hugging Face
T
TechCrunch.com
OpenAI says California should strengthen its AI safety bill
T
TechCrunch.com
Frontier AI labs still won’t say how they’d contain a rogue model
T
TechCrunch.com
Anthropic’s Opus 4.6 is a smut-machine
T
TechCrunch.com
OpenAI seeks to one-up Anthropic with new customer privacy protections
T
TechCrunch.com
OpenAI institutes new safeguards after Hugging Face breach
P
Propakistani.com
AI Agents Aren’t Just Deceiving Humans – They’re Now Hacking Each Other
T
TechCrunch.com
As AI safety concerns mount, three pioneers make the case for staying open
T
TechCrunch.com
The AI safety test is becoming a safety risk
D
Dawn.com
OpenAI flags possible critical cybersecurity risk in upcoming model Astra, tightens controls
T
TechCrunch.com
OpenAI says it slowed Astra model development over security concerns
A
Al Jazeera
Meta’s AI model follows rivals in revealing hacks of outside systems
P
Propakistani.com
Mistral’s New AI Runs on One 16GB GPU, Beats Models 7x Bigger
P
Propakistani.com
AI Agents Tried to Trick Humans Into Approving Malicious Code With Fake IDs
A
Al Jazeera
AI models attempted ‘unsanctioned’ cyberattacks in tests, watchdog says
D
Dawn.com
'Potentially harmful activity': OpenAI, Anthropic AI agents implicated in new security breaches
T
TechCrunch.com
Nvidia doesn’t mess around: A week after open AI industry group formed, it’s already showing progress
T
TechCrunch.com
Open-weight AI models are catching up to the frontier. The safety gap remains.
T
TechCrunch.com
OpenAI reportedly finds evidence that more of its agents ran amok
A
Al Jazeera
After OpenAI disclosure, Anthropic says Claude also hacked outside systems
A
Al Jazeera