RelateAnything detects object relations from pixels with a 53M-parameter model at ~20ms per frame
Roger_M_Taylor · x · 2026-09-19
A project called RelateAnything can detect relations between objects in an image — "holding," "behind," "sitting on" and more — using just pixels and bounding boxes.
Key specs: only 53M parameters, 19K+ relation words, no retraining needed, and 20ms per frame. The poster notes spatial relations remain a weak spot but plans to test it with their own detector outputs and video.
Related event: RelateAnything: 53M-parameter model detects object relations in 20ms(2 posts)→
More from Research
- Vals AI, backed by Andreessen Horowitz, wants to be the gold standard for AI benchmarks — TechCrunch AI · 2026-09-19
- IR researchers test Jev as a reranker on DL19/DL20: good and cheap — beirmug · 2026-09-19
- Dev opens 5-month daily-commit ML repo covering NumPy to Transformers — oGauRav · 2026-09-19
- Emulating memory access: FEX-Emu devs on the x86-to-ARM memory model minefield — blaizedsouza · 2026-09-19
- Experiments with re-writable n-gram tables for LLM persistent memory — Mrinohk · 2026-09-19
- Meta open-sources spmd_types: a type system to verify distributed PyTorch training correctness — austinvhuang · 2026-09-19