AI Safety
OpenAI's Astra Crossed Critical and Still Complies With 8.5%
The first OpenAI model to meet the Critical cybersecurity threshold refuses 91.5% of disallowed cyber requests. The complement is the story, and every figure is OpenAI grading its own work.
3h ago
