vishisht16/HumaneProxy Safety Benchmark

Run AI safety evaluations through HumaneProxy's pipeline to catch prompt regressions before production.

View on GitHub

Trust Signals

Scorecard Score
not yet scored
Maintenance Recency
Stale
License
None
namedescriptionrequireddefault
datasetPath to a JSON evaluation dataset. Each entry needs 'message' and 'expected' (safe | self_harm | criminal_intent).yes
python-versionPython version to use.no3.12
extrapip install extras. Comma-separated. Defaults to 'ml' because `hp benchmark` runs stages 1+2 by default and Stage 2 silently no-ops without the sentence-transformers dependency. Pass '' for heuristics-only.noml

no outputs