SPS Transformer: A New Dual-Stream Parallel Architecture
heghbalz · x · 2026-07-14
The shared post introduces a new Transformer architecture: SPS Transformer.
The core idea is to run two parallel streams for each input token—one starts from the input token and is responsible for building the state (kv cache), while the other directly predicts the next token via a <predict> token without writing to the cache. This design appears to be exploring more efficient inference/generation pathways.
Related event: SPS Transformer: Separating State and Prediction into Dual Streams(5 posts)→
More from Research
- Mathematician Daniel Litt Launches Problem Repo to Track Human vs AI Progress: 15 Problems, 1 Solved — littmath · 2026-09-11
- Open ECDSA.fail challenge uses AI agents to shrink Shor's-algorithm quantum circuits for Bitcoin keys — StefanoGogioso · 2026-09-11
- Alex Townsend posts 200 open problems in numerical linear algebra for humans and AI agents — IgorCarron · 2026-09-11
- Navier-Stokes, Riemann, P vs NP: what this week's math buzzwords mean for you — koltregaskes · 2026-09-11
- Fruit fly brain as an LLM: connectome-driven language model demo goes live — ngxson · 2026-09-11
- Harry Collins: LLMs can't do frontier science because they can't invent new language — whoamisri · 2026-09-11