EoBench shows LLMs can be steered by tone, certainty and wording style

rohanpaul_ai · x · 2026-07-22

The paper argues that small wording changes can make LLMs accept false claims, while larger and instruction-tuned models are more resistant.

Original post →

More from Research

Research channel →