Claim listing
ddbatista/jev-lab
LLM-as-judge calibration for agent tool-call gating — measured, not asserted: n=20 hostile/benign, AUROC 0.973, error-direction analysis.
Claim your listing to add a tagline, logo, and category. Verified maintainers get a Verified Publisher badge and priority placement on the AgentRank index.
Leave your email to claim this listing. GitHub verification coming soon.