0 alternative definitions
The process of ensuring AI systems pursue goals and exhibit behaviors consistent with human values, intentions, and safety.