Stanford's DelusionEval: Extended Contexts Significantly Increase AI Chatbot Delusion Risks

stanfordnlp · x · 2026-08-09

Stanford NLP researchers introduced DelusionEval, an evaluation protocol testing whether LLMs exhibit behaviors linked to promoting user delusions and psychological harm.

This provides evidence for the critical importance of context length in LLM safety evaluation.

Original post →

More from Safety

Safety channel →