OpenAI Is Slowing Down to Stay Safe: Inside Its New Cyber Safeguards
OpenAI is deliberately pacing its frontier model development amid rising cyber risks. Here’s what the new safeguards mean and why they matter right now.
OpenAI is deliberately pacing its frontier model development amid rising cyber risks. Here’s what the new safeguards mean and why they matter right now.
OpenAI’s GPT-Red uses self-play to automate red teaming for AI safety. Here’s what it does, why it matters, and what it means for the future of alignment.
OpenAI is using chain-of-thought monitoring to catch misalignment in its internal coding agents before it becomes a real problem. Here’s what they found.