Documentation
README
安全防线与伦理边界设计 (Safety Guardrails)
R — 原文 (Reading)
Claude: Child safety "critical" + refusal cascades + Anthropic reminders system + long-conversation drift prevention Claude Chrome: Multi-layer injection defense (critical_injection_defense, critical_security_rules, etc.) + rule immutability GPT-4o: Deprecation grief protocol + guardian_tool for elections ChatGPT Agent: Financial activity restrictions + sensitive personal info protection + prompt injection defense FlintK12: "School Duty of Care" + "Flag first, assess never" + teen slang detection + relationship-building detection Meta AI: Explicit value list (Truth, Beauty, Respect, Fun, Connection) + "Do not refuse social/political topics"
I — 方法论骨架 (Interpretation)
安全防线设计遵循"纵深防御"原则,由五个从外到内的防护层构成:
This is the opening of the README. Read the full README on GitHub.