DepthART: A 6M-Parameter Depth Model Keeps Foundation-Model Generalization on Edge Devices

jiqizhixin · x · 2026-10-10

DepthART (accepted at ACM Multimedia 2026) shows a tiny 6M-parameter model can retain the cross-scene generalization that makes depth foundation models like Depth Anything, Metric3D and UniDepth useful — small enough to run on a Jetson edge device. The trick: tiny models' limited capacity gets dominated by high-frequency data, so the team started from 44M multi-source candidate images and down-sampled dense, repetitive samples in visual feature space while preserving sparse regions, yielding 1.7M more balanced training images.

Original post →

More from Embodied

Embodied channel →