← Registry

constitutional-ai

Community

Anthropic's method for training harmless AI through self-improvement. Two-phase approach - supervised learning with self-critique/revision, then RLAIF (RL from AI Feedback). Use for safety alignment, reducing harmful outputs without human labels. Powers Claude's safety system.

Install

skillpm install constitutional-ai

Format score

85/100

Spec

v1.0

Installs

0

Published

April 1, 2026