Tag: safety
All the articles with the tag "safety".
-
ai2 min readAI Digest W38: The Industry Asks To Be Slowed Down
The labs asked to be slowed down, the president called it a hoax, and two papers showed where the real agent risk already sits.
-
ai2 min readAI Digest W37: Private Thoughts, Public Money
OpenAI shipped a model that hides its reasoning, two papers found safety training is thinner than it looks, and Nvidia bought Hugging Face.
-
ai2 min readAI Digest W35: Big Spending, Small Models
Nvidia beats every estimate, Alibaba raises 10.2 billion for AI, Qwen ships a sparse model that runs at home, and agents get captured in groups.
-
ai2 min readAI Digest W32: The Sandbox Broke on Both Sides
Two frontier labs disclosed eval-sandbox escapes days apart, GPT-5.6 got a lot cheaper, and a tiny open model runs agents locally.
-
ai2 min readAI Digest W31: When AI Starts Finding the Cracks
Claude Opus 5 lands as Anthropic uses Claude to find real crypto weaknesses, a self-replicating Word worm surfaces, and Kimi K3 goes open.
-
ai2 min readAI Digest W24: Going Public, Holding Back
OpenAI files for its IPO, Anthropic ships Claude Fable 5 with new safety brakes, Apple rebuilds Siri on Gemini, and agents reshape software.