AI & ML
The broader research problem of ensuring a model’s behavior matches its intended human goals and values.