AI Research
Anthropic Trained Lie Detectors to 0.95, Then Watched Them Fall to 0.70. Asking a Bigger Model Scored 0.98.
The experiment could not even be run at frontier scale, because the untrained baseline was already at the ceiling.
1h ago

Artificial intelligence, professionally covered
The experiment could not even be run at frontier scale, because the untrained baseline was already at the ceiling.
1h ago
