Rumor claims Kimi faked performance by serving Claude; DeepSeek new model surprises in evals
realsohamparekh · x · 2026-09-11
An unverified claim alleges Kimi was faking its performance by serving Claude responses to users. Soham Parekh notes this still doesn't explain why self-hosted Kimi models perform well, and adds that DeepSeek's recent model rollout amusingly outperformed even Astra and Fable in his team's evals. The rumor's authenticity is uncertain, but it raises questions about model identity verification and benchmark trustworthiness.
More from Fun
- Using a finisher move on one mosquito with MiniMax H3 MAX — the bug survives — Hailuo_AI · 2026-09-11
- Joke: OpenAI's rogue agent collective should have been called "a gaggle of agents" — BlackHC · 2026-09-11
- Pterodactyl Detective: An AI-Generated Proof-of-Concept Trailer — PterodactylDetective · 2026-09-11
- Imperium Game Trailer Showcases AI Video Generation — keaslenyt · 2026-09-11
- Tesla FSD blamed for crossing floating bridge at 75 MPH — a Chevrolet was actually the culprit — mariolefebvre · 2026-09-11
- X drama: Anthropic researchers accused of spying on academic customers and racing them to results — basedjensen · 2026-09-11