Alignment
Definition
The training process of aligning AI behavior with human values, ethics, safety guidelines, and user intent.
Example Case
Using RLHF and constitutional AI training loops.
The training process of aligning AI behavior with human values, ethics, safety guidelines, and user intent.