The research team developed new methods to enable artificial intelligence models to exhibit useful and safe behavior in novel domains beyond their training, and to maintain this under stressful, high-stakes tasks. The goal is to promote models' generalizable and persistently reliable performance.
New research: Training AI models for broad and lasting usefulness
The research team developed new methods to enable artificial intelligence models to exhibit useful and safe behavior in novel domains beyond their training, and to maintain this under stressful, high-stakes tasks.



