Documentation
README
Prompt Injection Defense
Design for compromise of model reasoning. Prompt text and classifiers can reduce attack success, but they do not create a reliable security boundary. Keep consequential authority, authorization, validation, and policy enforcement outside the model.
Inputs
Collect or infer, and label assumptions for:
This is the opening of the README. Read the full README on GitHub.