Capturing data…
Capturing data…
THU, 24 SEPT · 98 ITEMS
"Estimating accurate hand pose in camera space with a vision transformer" — arXiv cs.GR · Graphics & VFX
This is a new arXiv preprint proposing a -based method for estimation that outputs joint positions in —that is, the hand's actual 3D location relative to the camera, not just its shape in normalized coordinates. The paper introduces two components, TIS and PIE, and reports a 7.8% improvement over state-of-the-art methods for camera-space hand pose estimation.
The claim is the authors' own self-reported result, coming directly from the primary source: the arXiv preprint. It is solid as a description of what the paper proposes, but the preprint has not been peer-reviewed and the 7.8% improvement figure has not been independently replicated.
Genuinely new: the preprint was posted to arXiv in the 2026-09-22 to 2026-09-23 window, days before this digest.