AI safety
18 Posts
OpenAI Is Helping Build Shared Standards for Advanced AI—Here’s What That Actually Means
OpenAI joins the Appia Foundation to help define shared safety standards and evaluation frameworks for...
Trump’s Crackdown on Anthropic: Who Actually Wins?
The Trump administration is tightening the screws on Anthropic, but the real beneficiaries might not...
The US banned Anthropic’s Fable 5 release, but the numbers don’t seem to care
Despite the US government forcing Anthropic to pull Fable 5 and Mythos 5 over security...
A $5M PAC is trying to punch above its weight against Big Tech’s war chest
Guardrails, a PAC funded by small donations from AI workers, is taking on Big Tech's...
Pramaana Labs just scored $27M to make AI stop hallucinating in high-stakes fields
Pramaana Labs raised $27M from Khosla Ventures to apply formal verification to AI systems in...
OpenAI’s Deployment Simulation: Actually Testing AI Behavior Before Ship
OpenAI's new Deployment Simulation method uses real conversation data to predict how models behave before...
Anthropic’s safety warnings backfired — the government just pulled its best model
Anthropic's transparency about safety flaws led to regulators pulling its most capable AI model from...
OpenAI Gets Sued Over a School Shooting It Might Have Prevented
Seven families from the Tumbler Ridge school shooting are suing OpenAI, alleging the company knew...
Meta Disbanded Its Responsible AI Team — And That Should Worry You
Meta broke up its Responsible AI team, moving most members to generative AI products. This...
A Rogue AI at Meta Gave Bad Advice and Exposed Employee Data
Meta suffered a SEV1 security incident after an internal AI agent gave an engineer inaccurate...
OpenAI’s Trust Problem Isn’t About the Tech—It’s About Sam Altman
On the same day OpenAI released a shiny new policy paper about keeping superintelligence safe,...
Cutting Through the AI Noise: MIT Tech Review’s 10 Things That Actually Matter
MIT Technology Review launches a new essential guide: 10 Things That Matter in AI Right...