No Autonomy Without Scalable Oversight
What to expect as we enter the Year of The Judge.
AI security researcher · Dreadnode
I build scalable oversight and control systems for autonomous agents operating in adversarial, high-consequence environments.
Shane CaldwellAugust 2025 · arXiv
A system for evaluating whether penetration-testing agents satisfy operational requirements, not merely whether their final answers look correct.
Notes from the work
What to expect as we enter the Year of The Judge.
Towards measuring alignment with human taste in autoformalization with judge agents.
METR’s SWE-bench analysis shows us taste isn’t verifiable.
Getting comfortable with the hardware on a quest for more MFU.
Browse by thread