Finding Optimal Viewpoints for Monocular 3D Human Pose Estimation in Dynamic 3D Gaussian Splatting Space

Citations

WEB OF SCIENCE

0
Citations

SCOPUS

0

초록

Monocular 3D human pose estimation (HPE) focuses on locating 3D joint positions from a single viewpoint. However, it often suffers from viewpoint-dependent depth ambiguities and occlusions, emphasizing that next-best-viewpoint (NBV) selection becomes crucial. Prior work has explored NBV selection, but typically relies on datasets with fixed, discrete camera setups that do not support continuous viewpoint selection or on synthetic environments that lack photorealism. To address these shortcomings, we propose a novel pipeline for NBV selection in monocular 3D HPE. We construct a continuous and photorealistic environment, using 3D Gaussian Splatting (3DGS), which provides continuous novel-view rendering both spatially and temporally. Within this environment, we enable a reinforcement learning (RL) agent to actively select NBVs that minimize pose estimation error. Our approach is evaluated against fixed, random, and rotating camera trajectories, achieving over 24% improvement in pose estimation error. The code is available at https://github.com/knu-vis/FOV.

제목
Finding Optimal Viewpoints for Monocular 3D Human Pose Estimation in Dynamic 3D Gaussian Splatting Space
저자
Bae, Wan-Gi; Lee, Sanghyeon; Lee, Jong Taek
DOI
10.1109/AVSS65446.2025.11149906
발행일
2025
유형
Proceedings Paper
저널명
2025 IEEE INTERNATIONAL CONFERENCE ON ADVANCED VISUAL AND SIGNAL-BASED SYSTEMS, AVSS
호
2025