arXiv · 2504.19032
VISUALCENT: Visual Human Analysis using Dynamic Centroid Representation
Abstract
We introduce VISUALCENT, a unified human pose and instance segmentation framework to address generalizability and scalability limitations to multi person visual human analysis. VISUALCENT leverages centroid based bottom up keypoint detection paradigm and uses Keypoint Heatmap incorporating Disk Representation and KeyCentroid to identify the optimal keypoint coordinates. For the unified segmentation task, an explicit keypoint is defined as a dynamic centroid called MaskCentroid to swiftly cluster pixels to specific human instance during rapid changes in human body movement or significantly occluded environment. Experimental results on COCO and OCHuman datasets demonstrate VISUALCENTs accuracy and real time performance advantages, outperforming existing methods in mAP scores and execution frame rate per second. The implementation is available on the project page.
Explore related subjects
Keep this discovery
Explore connections, maps & timelines
Niaz Ahmad, Youngmoon Lee, Guanghui Wang. 2025-04-26. VISUALCENT: Visual Human Analysis using Dynamic Centroid Representation. https://arxiv.org/abs/2504.19032
Cite the original work for its findings. Save a collection to share your selection of sources.