GovernanceSafety, compliance and the open-source ecosystem

AI 安全

Research and practice ensuring AI systems are reliable, controllable and non-harmful.

AI safety spans alignment, red-teaming, jailbreak defense, content moderation and anomaly monitoring. Frontier labs invest heavily in safety evals; users should also mind data sanitization and output review.

Related terms