Guardrails
Ethics and Safety
The safety rules and filters built into AI systems to prevent harmful, illegal, or off-limits behavior.
Guardrails are the boundaries an AI operates within: refusing to help with weapons or fraud, filtering hateful output, staying on topic in a customer-service bot, or blocking a company assistant from revealing confidential data.They combine training, filters, and system instructions. No guardrail is perfect - people constantly probe for jailbreaks - so serious deployments layer several defenses and monitor real usage.