0
I’ve seen Anthropic’s Constitutional AI described as a way to guide model behavior using a set of principles. How does that approach differ from training that relies mainly on human feedback or preference rankings?
I’m interested in what the distinction means in practice: who or what evaluates the model’s responses, and where do human judgments still fit into the process?