TAPe+ML: sub-100K-parameter vision system hits 84.7 mAP50 on COCO detection

Comexp · hf · 2026-09-22

Researchers present TAPe+ML v3, built on the Theory of Active Perception (TAPe): a structured representation encoding relations among perceptual elements before recognition, instead of operating on raw pixel tensors, paired with a modular recognition architecture for classification, detection, and instance segmentation.

Core claim: shifting part of the modeling burden from network parameters to structured input representations can support compact multi-task vision systems with reduced data, memory, and compute.

Original post →

More from Research

Research channel →