Canonical event

Anthropic publishes Constitutional AI and reinforcement learning from AI feedback

Anthropic published a method in which models critique and revise outputs according to written principles and use AI-generated preferences for reinforcement learning.

Artifacts

Signals

View canonical JSON