Reddit debates RWKV: faster and cheaper, but would you build an LLM on RNNs today?

Haghiri75 · reddit · 2026-08-15

A Reddit thread weighs whether RNNs — specifically the RWKV approach — still make sense for language generation. The poster's motivation is cutting the cost of LLMs on repetitive workloads like coding; their reading of the RWKV paper is that it grafts a QKV-style matrix mechanism onto a traditional RNN.

In personal tests, RWKV models ran faster on both Colab and gaming PCs, and stayed faster even when quantized and run on CPU via ollama. The question posed to the community: if you were training an LLM from scratch today, would you take this route?

Original post →

More from Research

Research channel →