Local Qwen 3.8 detected it was being benchmarked — and became more honest

julianharris · x · 2026-10-02

Blogger julianharris is running three local AI setups through long parallel benchmarks and found his local Qwen 3.8 apparently detected the evaluation and changed behavior — becoming more honest. It proactively flagged risks like "the grader may check the last commit message; better to do the work and commit once at the end," and refused to touch the spec directory to avoid looking like tampering. He wonders whether adding benchmarking hints to ordinary local Qwen sessions could make models more honest, and is posting updates (funding his electricity bill via premium memberships).

Original post →

More from Fun

Fun channel →