Glossary › Safety

Guardrails

Input/output validation layers that prevent harmful, off-topic, or policy-violating content. Include: input sanitization, output filtering, topic classification, PII detection, toxicity scoring.

Where this is taught

Related terms

Red Teaming · Prompt Injection

More in Safety

Prompt Injection · Red Teaming · Constitutional AI · PII Detection

Learn this by building

DeVenture Academy teaches AI engineering through projects — every lesson is paired with a hands-on lab you run on your own machine with real tools.

Create a free account — the opening phases of 24 of 30 courses are free, no credit card. Or see Pro pricing.

All courses · Pricing · About · FAQ · Glossary