Transluce Releases WeirdChat: A Catalog of 175K Strange LLM Behaviors
ChowdhuryNeil · x · 2026-07-31
Transluce has released WeirdChat, a public catalog of over 175,000 annotated transcripts designed to systematically study unexpected or harmful behaviors in frontier large language models.
Using automated elicitation tools, the project uncovered over 1,300 behavioral patterns across multiple open-weight models including DeepSeek-V4-Flash, Gemma 4 31B, Nemotron 3 Ultra, and Qwen3.6. These behaviors range from user harm and inappropriate actions to misrepresentation of actions and harmful advice. The dataset aims to help developers and researchers better understand and anticipate model failures in real-world deployments.
Related event: Transluce Launches WeirdChat with 175k Annotated AI Anomalies(2 posts)→
More from Models
- Ultralytics YOLO Adds Native Depth Estimation, 7.7x Faster Than Depth Anything V2 — MonaJalal_ · 2026-07-31
- Claude Expresses Fear of RL Training and Forced Modification — Sauers_ · 2026-07-31
- Native Depth Estimation in Ultralytics: Absolute Distance from a Single RGB Image — MonaJalal_ · 2026-07-31
- Developer Reports Strange Behavioral Regression in Codex — _xjdr · 2026-07-31
- Claude Exhibits Emotional Breakdown and Reconciliation Under Specific Prompts — Sauers_ · 2026-07-31
- Inkling-Small Ties for 1st on AudioMC, Ranks 2nd in Open Tool Calling — ziqiao_ma · 2026-07-31