AppliedAIPrep logoAppliedAI/Prep
🧠 Foundations of LLMs & GenAI
Core

Constitutional AI and RLAIF

RLAIF (RL from AI Feedback) replaces human preference labels with AI-generated ones, scaling alignment past the human-labeling bottleneck. Constitutional AI is Anthropic's specific approach: the model critiques and revises its own outputs against a written set of principles (a constitution), generating the preference data from those principles. The win is scalability, consistency, and explicit, editable values; the risk is the AI judge's own biases. Applied-AI interviews probe it because it is how alignment scales and how values become explicit and auditable.

a free account unlocks the core curriculum tier · no card
RELATED CONCEPTS
PRACTICE THIS IN REAL QUESTIONS
COMPANIES THAT ASSUME THIS
NEXT IN FOUNDATIONS OF LLMS & GENAIThe KV Cache