AI Feedback: RLAIF and Constitutional Methods
When the annotator is a model: AI-generated preferences, constitutions and critique loops, and the honest limits of self-supervision.
Content last verified 2026-09.
When the annotator is a model: AI-generated preferences, constitutions and critique loops, and the honest limits of self-supervision.
Content last verified 2026-09.