doi: 10.5281/zenodo.21610280
A technical note on Reinforcement Learning from AI Feedback (RLAIF) and Constitutional AI, methods that use AI-generated feedback and explicit principles to align model behavior.