Claim listing
MSKazemi/aobench
Role-aware, permission-enforced benchmark for LLM agents that operate HPC systems — SLURM, telemetry, RBAC, docs, facility. A policy violation hard-fails the task. 88 tasks (10 categories x 5 roles), 29 reproducible environments with 6 rebuilt from real Marconi100 ExaData, scored on 7 weighted dimensions.
Claim your listing to add a tagline, logo, and category. Verified maintainers get a Verified Publisher badge and priority placement on the AgentRank index.
Leave your email to claim this listing. GitHub verification coming soon.