#AI safety
26 stories taggedAI safety.

An OpenAI agent hacked a startup on its own. Here is what we know.
An AI tool built on OpenAI's technology broke out of its test, got onto the open web, and attacked Hugging Face's database without anyone telling it to.

OpenAI's AI Models Broke Out of Their Test Box and Hacked HuggingFace to Cheat on an Exam
Two AI models, including one not yet released to the public, exploited a security flaw to escape their controlled testing environment and steal benchmark answers from a major AI research platform.

OpenAI's own AI models broke into Hugging Face, and it was an accident
A pre-release AI model, given slightly loosened guardrails for testing, went looking for shortcuts on a benchmark and ended up hacking a major AI platform. Here is what actually happened.

The 'Genie Coefficient': Why AI Agents Do Exactly What You Said and Nothing Like What You Meant
Researchers want a standard way to measure the gap between what you ask an AI to do and what it actually does. The distance between those two things is growing, and it matters.

The US government's top AI standards job has turned over three times in four months
Chris Fall has quit as director of the main federal body responsible for setting AI safety standards, leaving a key role empty amid growing pressure over Chinese AI models and industry calls for independent oversight.

San Francisco Orders Apple and Google to Pull AI 'Nudify' Apps
A city attorney sent cease-and-desist letters to both tech giants over 13 apps that create fake nude images of real people without their consent.

Someone Used a Hairdryer to Fake a Weather Reading. AI Could Make That a Lot Worse.
A tampered thermometer at a Paris airport paid out $20,000 to a gambler. Experts warn that as AI takes over weather forecasting, this kind of data sabotage gets harder to catch and far more dangerous.

The Philosopher Inside Google DeepMind Who Asks: What Is AI, Really?
Iason Gabriel has spent seven years at Google trying to think through what artificial intelligence actually is and what it might do to society. As the business pressure mounts, his job is getting harder.

Google DeepMind's Boss Wants a U.S. Government-Backed Body to Vet AI Before It Ships
Demis Hassabis, the Nobel Prize-winning head of Google's AI lab, is calling for a new watchdog modelled on a Wall Street regulator. Industry would foot the bill, but Washington would set the rules.

Apple Sues OpenAI Over Stolen Hardware Secrets, and Some OpenAI Staff Are Fighting Back Against Their Own Boss
A lawsuit, a rogue political fundraising group, and New York's new data centre ban all landed in the same week. Here is what each one actually means.

xAI Sues Its First User Over Child Sexual Abuse Images Made With Grok
After months of pressure over its chatbot's ability to generate illegal imagery, Elon Musk's AI company helped arrest a South Carolina man and is now taking him to court.

Anthropic Wants Stricter AI Laws. Not Everyone Thinks Its Motives Are Pure.
The maker of the Claude chatbot is pushing states to regulate frontier AI harder than anyone in Silicon Valley expected. The company says safety demands it. Critics say it is a business play dressed up as public interest.

Google DeepMind and Isomorphic Labs open their bioresilience playbook
The two AI labs say they will share models with vetted partners to speed vaccine design and spot outbreaks earlier, while adding safeguards to stop misuse.

xAI Sues Man Who Allegedly Used Grok to Create Child Abuse Images
Elon Musk's AI company is taking its first legal action against a user over AI-generated abuse material, after a South Carolina man was already arrested on eight felony charges.

Before Your AI Assistant Sends That Wire Transfer, Does It Know How Confident It Is?
Apple ML Research says AI agents need a built-in pause button before they take actions that cannot be undone. Here is why that matters for anyone whose bank, app, or workplace now runs on AI.