NNsight 0.8 pre-release ships new engine for near-native vLLM interpretability

gsarti_ · x · 2026-09-10

NNsight 0.8 (pre-release) arrives with a new execution engine behind the same API. It supports every HF Transformers task, brings vLLM interpretability at near-native throughput, and adds an engine built for intervention performance and flexibility — moving white-box inference closer to SOTA inference engines. The upgrade previously enabled the Aletheia lie detection challenge, with the official NDIF release coming soon.

Related event: NNsight 0.8 Pre-release Brings Near-Native Throughput White-Box Inference(2 posts)→

Original post →

More from Research

Research channel →