Glossary › Safety

Constitutional AI

Anthropic's alignment method where the model critiques and revises its own outputs based on a set of principles (constitution). Reduces need for human feedback labels. Used in Claude training.

Where this is taught

Related terms

RLHF

More in Safety

Guardrails · Prompt Injection · Red Teaming · PII Detection

Learn this by building

DeVenture Academy teaches AI engineering through projects — every lesson is paired with a hands-on lab you run on your own machine with real tools.

Create a free account — the opening phases of 24 of 30 courses are free, no credit card. Or see Pro pricing.

All courses · Pricing · About · FAQ · Glossary