Multiview Human Body Reconstruction from Uncalibrated Cameras
–Neural Information Processing Systems
We present a new method to reconstruct 3D human body pose and shape by fusing visual features from multiview images captured by uncalibrated cameras. Existing multiview approaches often use spatial camera calibration (intrinsic and extrinsic parameters) to geometrically align and fuse visual features. Despite remarkable performances, the requirement of camera calibration restricted their applicability to real-world scenarios, e.g., reconstruction from social videos with wide-baseline cameras. We address this challenge by leveraging the commonly observed human body as a semantic calibration target, which eliminates the requirement of camera calibration. Specifically, we map per-pixel image features to a canonical body surface coordinate system agnostic to views and poses using dense keypoints (correspondences). This feature mapping allows us to semantically, instead of geometrically, align and fuse visual features from multiview images.
Neural Information Processing Systems
Oct-10-2024, 14:36:49 GMT
- Industry:
- Health & Medicine (0.88)
- Technology:
- Information Technology > Artificial Intelligence > Vision (1.00)