NeurIPS 2026 author says OpenReview PDF contained a prompt injection

Kwangryeol · reddit · 2026-07-24

A NeurIPS 2026 author says GPT detected a prompt injection inside the PDF downloaded from OpenReview, and that the injected text appears not to have been in the original submission.

The author suspects the injection may have been added by the conference workflow and asks others to check whether they see the same thing in their review copies. They also warn that suspiciously formulaic review wording could indicate LLM-generated reviews that were not properly written by humans.

The post includes the exact injected prompt, which instructs the output to contain three specific phrases, and asks whether anyone else found it in the reviewer version of their paper.

Original post →

More from Safety

Safety channel →