Exploiting large language models in peer review: indirect prompt injection attacks and integrity probes

This paper studies indirect prompt injection in AI assisted peer review across 42,000 chatbot outputs. Hidden instructions alter assessments and can support organizer integrity tests.

July 2026 · F. Torrielli, S. Locci, A. Rapp, L. Di Caro