Reddit asks whether NeurIPS ethics reviewers were fooled by conference-side prompt injection
dontknowwhattoplay · reddit · 2026-07-29
A Reddit post asks whether others have seen NeurIPS ethics reviewers affected by conference-side prompt injection designed to catch LLM-based reviewers.
- The poster claims some reviewers reported ethical issues after being exposed to hidden manipulations.
- They suggest the conference-side setup may not have been disclosed to ethics reviewers.
- The discussion is about how prompt injection can affect review workflows and what disclosure or safeguards should exist in conference processes.
More from Safety
- NVIDIA-Led 'Open Weights' Coalition Accused of Hijacking Open Source Definition — alex_verem · 2026-07-29
- Traceforce launches on YC with a tool to spot risky AI agent activity on laptops — ycombinator · 2026-07-29
- Open-weight models need costly fine-tuning defenses, not vague “safe” branding — walden42 · 2026-07-29
- OpenAI and Anthropic staff reportedly urge the US to pace frontier AI development — Puzzleheaded_Week_52 · 2026-07-29
- TransluceAI proposes oversight foundation models to catch reward hacking at scale — JacobSteinhardt · 2026-07-29
- Black Hat is expected to push agentic AI and prompt injection into the security spotlight — DavidLinthicum · 2026-07-29