ArXivCSExplorer
☆☆Bookmarks🏆RSSHow to UseFAQ
Built with and by Teycir Ben Soltane•
How to Use•FAQ•GitHub•arXiv.org•
Share:

20 results for “visual-inertial odometry”

CS papers only

Hybrid search: Keyword + semantic, ranked by combined score.ⓘ

Want pure semantic search? Try claim verification →

cs.ROEmpiricalRecentJul 19, 2026

VIDAR: Visual-Inertial Dense Alignment and Reconstruction via a Geometric Foundation Model

Diyari Mohammed Salih, Lingxiang Hu, Naima AitOufroukh-Mammar, Fabien Bonardi

This paper introduces VIDAR, a framework for metric dense monocular reconstruction using visual-inertial odometry and Depth Anything 3.

View →
cs.ROcs.LGcs.SEEmpiricalRecentJul 13, 2026

Self-Healing Visual Recovery for Autonomous Ground Vehicles Using Camera-Only Visual Odometry

Jakob Solberg Berntzen, Safia Fatima, Leon Moonen

This paper presents a two-stage recovery approach for camera-only low-cost unmanned ground vehicles to restore guideline tracking when lines are lost.

View →
cs.ROeess.SYEmpiricalRecentJul 3, 2026

Closed-loop vs. Open-loop Kalman Filter Architectures in Airborne Aided Inertial Navigation

Antonia Hager, Torleiv H. Bryne

This paper compares the performance of open-loop and closed-loop filters in inertial navigation systems using simulations.

View →
cs.ROcs.CVEmpiricalRecentJul 19, 2026

DROID-ANCHOR: Odometry-Anchored Recurrent Metric Depth Estimation

Yuxuan Chen, Brook Du

This paper proposes Metric-DROID, an end-to-end recurrent architecture for precise metric depth estimation in monocular autonomous robot navigation, using proprioceptive odometry and a LSTM Update Ope…

View →
cs.CVRecentJun 2, 2026

PixVOD: Pixel-Distributed Direct Visual Odometry and Depth Estimation

Shinjeong Kim, Ignacio Alzugaray, Callum Rhodes, Paul H. J. Kelly +1 more

PixVOD proposes a fully parallelizable, pixel-distributed framework for visual odometry and depth estimation that performs computations directly on the sensor using Gaussian Belief Propagation.

View →
cs.ROcs.CVEmpiricalRecentJul 23, 2026

GLAM-SLAM: Real-time Gaussian Large-scale Mapping via Flow Densification and Spatial Decomposition

Panagiotis Mermigkas, Argyris Manetas, Petros Maragos

GLAM-SLAM is a real-time, decoupled Gaussian-splatting SLAM system for large-scale outdoor scenes with a robust feature-based frontend and structured sparse mapping representation.

View →
cs.ROEmpiricalRecentJul 17, 2026

Vision-Language-Motion Maps: An Open-Vocabulary, Uncertainty-Aware, Queryable Motion Attribute for 3D Scene Maps

Dibyendu Ghosh, Ayushi Shakya

This paper introduces Vision-Language-Motion Maps (VLMM), an open-vocabulary, natural-language-queryable 3D map with fused motion attributes and per-element uncertainty, which outperforms semantic-onl…

View →
cs.ROEmpiricalRecentJul 17, 2026

A New Implementation of NeoSLAM and a Comparative Evaluation with RatSLAM

Joao Victor T. Borges, Fabio Coelho, Paulo Padrao, Jose Fuentes +3 more

This paper presents a new modular architecture for NeoSLAM using modern frameworks, achieving real-time execution and minimal data discarding. It also compares NeoSLAM and RatSLAM across three dataset…

View →
eess.SPTutorialRecentJul 17, 2026

Inertial Human Motion Capture: From Biomechanics to Recent Sensor Fusion Methods and Back

Manon Kok, Ive Weygers, Hassan Osman, Daniel Weber +3 more

This tutorial-style review focuses on kinematics of inertial measurement units (IMUs) for human motion capture and introduces methods to determine adequate formulations for sensor measurements.

View →
cs.CVRecentJun 1, 2026

Honey, I Shrunk the Arc de Triomphe!

Yuanbo Xiangli, Hanyu Chen, Xueqing Tsang, Noah Snavely

The paper introduces MetricScenes, a new large-scale, in-the-wild dataset, and demonstrates that fine-tuning existing geometry models on this dataset significantly mitigates the scale-collapse problem…

View →
cs.ROEmpiricalRecentJul 1, 2026

Where Am I? Semantic Map Grounding via Vision-Language Models for Multi-Modal Localization

Suraj Borate, Aarav Shah, Madhu Vadali

This paper addresses robot localization in GPS-denied indoor environments using a semantic reasoning approach with a vision-language model, achieving high accuracy with a composite loss and curriculum…

View →
cs.ROEmpiricalRecentJul 1, 2026

FastBridge: Closing the Model-Based Realization Gap in Safety Filters on 3D Gaussian Splatting for Fast Quadrotor Flight

Tscholl Dario, Nakka Yashwanth Kumar, Gunter Brian

This paper introduces a nonlinear, actuator-aware safety filter for 3D Gaussian Splatting (3DGS) based on full quadrotor dynamics, reducing trajectory jerk by 47% and running 2.25 times faster than ex…

View →
cs.CVcs.AIcs.LGEmpiricalRecentJul 6, 2026

From Fixed to Free Cameras: Calibration-Free View-Robust Vision-Language-Action Model

Wenhao Li, Xueying Jiang, Quanhao Qian, Deli Zhao +3 more

This paper introduces Camera-Centric VLA, a new model for Vision-Language-Action policies that predicts camera-centric actions and hand-eye matrix, allowing the policy to figure out camera geometry on…

View →
cs.CVcs.GRRecentJun 1, 2026

Ultra Diffusion Poser: Diffusion-Based Human Motion Tracking From Sparse Inertial Sensors and Ranging-Based Between-Sensor Distances

Dominik Hollidt, Tommaso Bendinelli, Christian Holz

Ultra Diffusion Poser is a novel diffusion model that improves human motion tracking from sparse IMUs and UWB ranging by explicitly modeling the geometric constraints imposed by inter-sensor distances…

View →
cs.ROEmpiricalRecentJul 16, 2026

Learning Agile Navigation in Crowded Environments for Quadruped Robots

Shuyu Wu, Zeyu Liu, Tianbao Zhang, Fanxing Li +5 more

This paper proposes VOP-Nav, a novel navigation system for quadruped robots that combines the geometric safety of Velocity Obstacles with the agile adaptability of end-to-end learning.

View →