keyword
human preference alignment
Human preference alignment is the process of training or adjusting an AI model so its responses better reflect what people prefer, often using human judgments or comparisons of possible responses as feedback. In language models, this feedback can guide fine-tuning so the model is more likely to produce preferred outputs.
1 item

