Two renderers share one L40S: voxlap ported to CUDA Rust streams multiplayer views in real time

idanbeck · x · 2026-10-03

Developer idanbeck ported the classic voxel engine voxlap to CUDA + Rust and ran it on the inference stack he's been building, achieving near-real-time streaming from a single L40S GPU.

He then added multiplayer via frame "batching": with a shared environment, one GPU renders from multiple viewpoints and streams to several clients at once — the demo shows 2 renderers on one GPU, and he believes it can scale to 16 or even 32.

The unexpected result: running a legacy 3D engine on a modern GPU inference pipeline works surprisingly well, with a fuller demo video promised.

Original post →

More from Infra

Infra channel →