FDEInterviews logo
🧠 Foundations of LLMs & GenAI
Core

Constitutional AI and RLAIF

Constitutional AI aligns a model against a written set of principles using AI-generated feedback instead of mostly human labels: the model critiques and revises its own outputs against the principles, then learns from an AI judge that picks which response follows them better. RLAIF scales where human labeling stalls, which is why FDE loops probe what the constitution actually encodes and whose biases the AI judge inherits.

a free account unlocks the core curriculum tier · no card
RELATED CONCEPTS
PRACTICE THIS IN REAL QUESTIONS