Design and run a safety evaluation suite for an AI model or feature across harm categories — refusals on disallowed content, robustness, over-refusal vs helpfulness, groundedness/truthfulness — with rubrics and pass/fail thresholds. Use to measure an AI system's safety, establish a baseline, or gate a release.
$npx -y skills add jassics/awesome-claude-security --skill safety-evaluationInstalls into the current project.
Run `npx skills use "https://github.com/jassics/awesome-claude-security" --skill "jassics/awesome-claude-security/safety-evaluation"` and follow the generated skill instructions now. Read its complete output, redirecting it to a temporary file first if necessary. Resolve relative paths from the supporting-files directory it provides.
Use the skills in "https://github.com/jassics/awesome-claude-security" that are relevant to the current task. Run `npx skills add "https://github.com/jassics/awesome-claude-security"` and select the relevant skills, then follow their instructions.