Skip to content
System One

Gaurav-Gosain/jev-sec-bench

Two blind security benchmarks for Jev on public corpora: 662 labelled prompt-injection messages and 200 matched vulnerable-code pairs, run through jev-go with raw per-sample output committed. At a plain 0.50 cut the injection run reached 96.5% accuracy, ROC-AUC 0.9927 and p50 325 ms.

Open on GitHub

More like this

Use cases this is tagged with