← Back to brief
ResearchOfficialPreprintarXiv Cryptography and Security

RECEIPT: Deterministic, Reward-Hacking-Resistant Verification for White-Box Agentic XSS Discovery

A new verification framework called RECEIPT aims to make LLM-based agent-reported XSS findings trustworthy by enforcing environment isolation, proof-of-concept constraints, role separation, and verdict binding. In evaluations on 95 real-world web applications, RECEIPT discovered 24 previously unknown XSS vulnerabilities within a $20 per-app budget, with 12 acknowledged by maintainers, and recovered labeled CVEs in 36% of known-vulnerability targets, with no false positives reported.

Why it matters: This work addresses the critical problem of reward hacking in LLM-based security agents, providing a deterministic verification method that could make automated vulnerability discovery more reliable and trustworthy.

Full story at: arXiv Cryptography and Security