A demo trains and visualizes RL policies inside tldraw, fully offline

max__drake · x · 2026-07-25

A demo shows training and visualizing RL policies inside tldraw, offline.

The post is short, but the core idea is a workflow where reinforcement-learning policies are developed and inspected in a familiar canvas-style interface rather than a traditional ML notebook. That makes it interesting as both a research-engineering prototype and a tooling experiment.

Original post →

More from coding & agent

coding & agent channel →