Thanks to the development of 2D keypoint detectors, monocular 3D human pose estimation (HPE) via 2D-to-3D uplifting approaches have achieved remarkable improvements. Still, monocular 3D HPE is a challenging problem du...
详细信息
ISBN:
(纸本)9781665491907
Thanks to the development of 2D keypoint detectors, monocular 3D human pose estimation (HPE) via 2D-to-3D uplifting approaches have achieved remarkable improvements. Still, monocular 3D HPE is a challenging problem due to the inherent depth ambiguities and occlusions. To handle this problem, many previous works exploit temporal information to mitigate such difficulties. However, there are many real-world applications where frame sequences are not accessible. This paper focuses on reconstructing a 3D pose from a single 2D keypoint detection. Rather than exploiting temporal information, we alleviate the depth ambiguity by generating multiple 3D pose candidates which can be mapped to an identical 2D keypoint. We build a novel diffusion-based framework to effectively sample diverse 3D poses from an off-the-shelf 2D detector. By considering the correlation between human joints by replacing the conventional denoising U-Net with graph convolutional network, our approach accomplishes further performance improvements. We evaluate our method on the widely adopted Human3.6M and HumanEva-I datasets. Comprehensive experiments are conducted to prove the efficacy of the proposed method, and they confirm that our model outperforms state-of-the-art multi-hypothesis 3D HPE methods.
Sarcasm is often used to convey the opposite of what is actually said (positive words, negative meaning). It is a type of humor that relies on the listener or reader to understand the intended meaning of the words. Sa...
详细信息
Gesture recognition is the revolutionary change in the world of Virtual and Augmented reality development allowing seamless non-physical control of computerized devices to create a highly interactive and flexible hybr...
详细信息
The diffusion model has been widely applied in various aspects of artificial intelligence due to its flexible and diverse generative performance. However, there is a lack of research on applying diffusion models in th...
详细信息
In Cooperative Learning (CL), students work together in small groups to achieve shared learning goals. Several studies have proved that CL promotes active learning, social skills development, inclusion, and well-being...
详细信息
In the domain of software development, making informed decisions about the utilization of large language models (LLMs) requires a thorough examination of their advantages, disadvantages, and associated risks. This pap...
详细信息
Risk management is an activity that will be carried out by various organizations and agencies so that the organization's sustainability will continue. Various parts of the organization can face risks and threats. ...
详细信息
Big data is being used by public sector entities to improve services to cities. Many government agencies are in unfamiliar territory as they strive to incorporate and integrate one novel kind of technology (big data) ...
详细信息
Accurate 3D vascular segmentation is essential for diagnosing and treating vascular diseases. This task remains challenging due to the complexity of the 3D data and the morphological diversity of blood vessels. In rec...
详细信息
Providing users with quality information is essential for the success of mobile application development. Hajj mobile application development is complex due to user satisfaction level, age and cultural difference, lack...
详细信息
暂无评论