20 results for “Understanding of 6D pose estimation”
CS papers onlyHybrid search: Keyword + semantic, ranked by combined score.ⓘ
Want pure semantic search? Try claim verification →
The paper introduces PIXIE, a zero-shot framework for estimating 6D pose of an object from an RGB image using only an untextured 3D model.
Panfei Cheng, Hongshan Yu, Wenrui Chen, Xiaojun Tang +2 more
The paper proposes a novel symmetry-aware, category-level method for 9D object pose estimation that accurately estimates translation and size first, followed by rotation, achieving state-of-the-art re…
The paper introduces ProxyPose, a method for six-degree-of-freedom (6-DoF) pose tracking using a video diffusion model and a single marked pixel in the first frame.
The paper introduces MetricScenes, a new large-scale, in-the-wild dataset, and demonstrates that fine-tuning existing geometry models on this dataset significantly mitigates the scale-collapse problem…
This paper introduces VIDAR, a framework for metric dense monocular reconstruction using visual-inertial odometry and Depth Anything 3.
This paper presents a lightweight network for omnidirectional human detection and relative 2D pose estimation from planar LiDAR sequences using Space-Time Blocks.
CIPER proposes a unified transformer framework to simultaneously perform cross-view image retrieval and precise 3-DoF pose estimation, overcoming the limitations of cascaded, separate methods.
This paper proposes Point Cloud Upsampling through Patch-based Frequency Superposition (PUtPFS), an optimization-based approach for uniform point cloud upsampling without data dependency or training.
PRIMA is a framework that significantly improves 3D quadruped mesh recovery by integrating biological knowledge and a test-time adaptation strategy, achieving state-of-the-art results on diverse and c…
Jiangwei Ren, Xingyu Jiang, Zijie Song, Wei Xu +3 more
This paper proposes Wat3R, a cross-domain semi-supervised learning framework for adapting 3D reconstruction models from air to underwater scenes using unlabeled real underwater video footage and a tea…
Wenhao Li, Xueying Jiang, Quanhao Qian, Deli Zhao +3 more
This paper introduces Camera-Centric VLA, a new model for Vision-Language-Action policies that predicts camera-centric actions and hand-eye matrix, allowing the policy to figure out camera geometry on…
The paper introduces S2MDF, a plug-and-play module that enforces a hard constraint to eliminate interpenetrations in multi-object Signed Distance Field (SDF) representations, significantly improving p…
The paper reframes industrial visual sim-to-real transfer as a domain-gap problem categorized by the availability of explicit object geometry (CAD), arguing that the required prior evidence dictates t…
TROPHIES introduces a unified framework to jointly reconstruct dynamic humans, static scenes, and camera poses from multi-view videos, achieving globally consistent and physically plausible 4D reconst…
Fangzhou Zhao, Yao Sun, Xuesong Liu, Runze Cheng +2 more
This paper proposes a SemCom framework for real-time mobile 3D reconstruction, which includes a semantic transceiver and a confidence-guided geometric estimation method.
This paper proposes Metric-DROID, an end-to-end recurrent architecture for precise metric depth estimation in monocular autonomous robot navigation, using proprioceptive odometry and a LSTM Update Ope…