|
arxiv.org
|
NEXUS: Neural Energy Fields for Physically Consistent Contact-Rich 3D Object Dynamics
|
|
arxiv.org
|
PhySPRING: Structure-Preserving Reduction of Physics-Informed Twins via GNN
|
|
arxiv.org
|
Robot Learning from Human Videos: A Survey
|
|
arxiv.org
|
Towards the Vision-Sound-Language-Action Paradigm: The HEAR Framework for Sound-Centric Manipulation
|
|
arxiv.org
|
OGScene3D: Incremental Open-Vocabulary 3D Gaussian Scene Graph Mapping for Scene Understanding
|
|
arxiv.org
|
RegFormer++: An Efficient Large-Scale 3D LiDAR Point Registration Network with Projection-Aware 2D Transformer
|
|
arxiv.org
|
Continually Evolving Skill Knowledge in Vision Language Action Model
|
|
arxiv.org
|
ERMV: Editing 4D Robotic Multi-view Images to Enhance Embodied Agents
|
|
arxiv.org
|
CAD-SLAM: Consistency-Aware Dynamic SLAM with Dynamic-Static Decoupled Mapping
|
|
arxiv.org
|
3D Gaussian Splatting in Robotics: A Survey
|
|
arxiv.org
|
SC-NeRF: Self-Correcting Neural Radiance Field with Sparse Views
|
|
arxiv.org
|
FFPA-Net: Efficient Feature Fusion with Projection Awareness for 3D Object Detection
|
|
arxiv.org
|
DetFlowTrack: 3D Multi-object Tracking Based on Simultaneous Optimization of Object Detection and Scene Flow Estimation
|
|
scholar.google.com.hk
|
Person Parametric Physics-Informed Representation for mmWave-Based Human Pose Estimation
|
|
scholar.google.com.hk
|
RoboFlow4D: A Lightweight Flow World Model Toward Real-Time Flow-Guided Robotic Manipulation
|
|
scholar.google.com.hk
|
NeRFs in Robotics: A Survey
|
|
scholar.google.com.hk
|
TransSeg: 3-D Semantic Segmentation Based on 2D-3D Trans-Modal Fusion for Autonomous Driving
|
|
scholar.google.com.hk
|
Zero-Shot Depth Map Restoration from Sparse Infrastructure Point Clouds Using Diffusion Models and Prompted Segmentation
|
|
scholar.google.com.hk
|
CoMA-SLAM: Collaborative Multi-Agent Gaussian SLAM with Geometric Consistency
|
|
arxiv.org
|
ActionReasoning: Robot Action Reasoning in 3D Space with LLM for Robotic Brick Stacking
|
|
scholar.google.com.hk
|
Towards Mitigating Modality Bias in Vision-Language Models for Temporal Action Localization
|
|
scholar.google.com.hk
|
SNI-SLAM++: Tightly-Coupled Semantic Neural Implicit SLAM
|
|
scholar.google.com.hk
|
DifFlow3D: Hierarchical Diffusion Models for Uncertainty-Aware 3D Scene Flow Estimation
|
|
scholar.google.com.hk
|
MovSAM: A Single-Image Moving Object Segmentation Framework Based on Deep Thinking
|
|
scholar.google.com.hk
|
SemGauss-SLAM: Dense Semantic Gaussian Splatting SLAM
|
|
scholar.google.com.hk
|
MonoPCFlow: Enabling Efficient Scene Flow Estimation from Monocular View
|
|
scholar.google.com.hk
|
End-to-End 2D-3D Registration between Image and LiDAR Point Cloud for Vehicle Localization
|
|
scholar.google.com.hk
|
Offboard Occupancy Refinement with Hybrid Propagation for Autonomous Driving
|
|
scholar.google.com.hk
|
Unsupervised Learning of 3D Scene Flow With LiDAR Odometry Assistance
|
|
scholar.google.com.hk
|
G2-Mapping: General Gaussian Mapping for Monocular, RGB-D, and LiDAR-Inertial-Visual Systems
|
|
scholar.google.com.hk
|
RL-GSBridge: 3D Gaussian Splatting Based Real2Sim2Real Method for Robotic Manipulation Learning
|
|
scholar.google.com.hk
|
DVN-SLAM: Dynamic Visual Neural SLAM Based on Local-Global Encoding
|
|
scholar.google.com.hk
|
CODE: COllaborative Visual-UWB SLAM for Online Large-Scale Metric DEnse Mapping
|
|
scholar.google.com.hk
|
Digital Twin-Based HVAC Model Generation Using Point Cloud and Image Data for Building Retrofitting
|
|
scholar.google.com.hk
|
Geometric Model of Hydropower Plants: A Foundational Step Towards Creating a Hydropower Digital Twin
|
|
scholar.google.com.hk
|
DnFPlane for Efficient and High-Quality 4D Reconstruction of Deformable Tissues
|
|
scholar.google.com.hk
|
Spherical Frustum Sparse Convolution Network for LiDAR Point Cloud Semantic Segmentation
|
|
scholar.google.com.hk
|
DSLO: Deep Sequence LiDAR Odometry Based on Inconsistent Spatio-Temporal Propagation
|
|
scholar.google.com.hk
|
EMIE-MAP: Large-Scale Road Surface Reconstruction Based on Explicit Mesh and Implicit Encoding
|
|
scholar.google.com.hk
|
LHMap-loc: Cross-Modal Monocular Localization Using LiDAR Point Cloud Heat Map
|
|
scholar.google.com.hk
|
3DSFLabelling: Boosting 3D Scene Flow Estimation by Pseudo Auto-Labelling
|
|
scholar.google.com.hk
|
SNI-SLAM: Semantic Neural Implicit SLAM
|
|
scholar.google.com.hk
|
DifFlow3D: Toward Robust Uncertainty-Aware Scene Flow Estimation with Iterative Diffusion-Based Refinement
|
|
scholar.google.com.hk
|
Review of Multimodal Data and Their Applications for Road Maintenance
|
|
scholar.google.com.hk
|
3D Multi-target Tracking Based on Joint Optimisation of Object Detection and Scene Flow Estimation
|
|
scholar.google.com.hk
|
Dense 3D Neural Map Reconstruction Only Using a Low-Cost LiDAR
|
|
scholar.google.com.hk
|
An Integrated Solution for Automatic 3D Object-Based Information Retrieval
|
|
scholar.google.com.hk
|
Automated Generation of Geometric Digital Twin of Roof for Building Retrofitting
|
|
scholar.google.com.hk
|
Digital Twins in Construction: Leveraging Point Cloud Data and BIM for Monitoring Project Schedule
|
|
scholar.google.com.hk
|
Pseudo-LiDAR for Visual Odometry
|
|
scholar.google.com.hk
|
RLSAC: Reinforcement Learning Enhanced Sample Consensus for End-to-End Robust Estimation
|
|
scholar.google.com.hk
|
DELFlow: Dense Efficient Learning of Scene Flow for Large-Scale Point Clouds
|
|
scholar.google.com.hk
|
TransLO: A Window-Based Masked Point Transformer Framework for Large-Scale LiDAR Odometry
|
|
scholar.google.com.hk
|
Unsupervised Learning of Depth and Pose Based on Monocular Camera and Inertial Measurement Unit
|
|
scholar.google.com.hk
|
Anomaly Detection for Robust Autonomous Navigation
|
|
scholar.google.com.hk
|
Self-Supervised Multi-Frame Monocular Depth Estimation with Pseudo-LiDAR Pose Enhancement
|
|
scholar.google.com.hk
|
Interactive Multi-Scale Fusion of 2D and 3D Features for Multi-Object Vehicle Tracking
|
|
scholar.google.com.hk
|
Graph Relational Reinforcement Learning for Mobile Robot Navigation in Large-Scale Crowded Environments
|
|
scholar.google.com.hk
|
RegFormer: An Efficient Projection-Aware Transformer Network for Large-Scale Point Cloud Registration
|
|
scholar.google.com.hk
|
3D Hierarchical Refinement and Augmentation for Unsupervised Learning of Depth and Pose from Monocular Video
|
|
scholar.google.com.hk
|
3D Scene Flow Estimation on Pseudo-LiDAR: Bridging the Gap on Estimating Point Motion
|
|
scholar.google.com.hk
|
Efficient 3D Deep LiDAR Odometry
|
|
scholar.google.com.hk
|
Learning of Long-Horizon Sparse-Reward Robotic Manipulator Tasks with Base Controllers
|
|
scholar.google.com.hk
|
What Matters for 3D Scene Flow Network
|
|
scholar.google.com.hk
|
Unsupervised Learning of Optical Flow with Non-Occlusion from Geometry
|
|
scholar.google.com.hk
|
Motion Projection Consistency Based 3D Human Pose Estimation with Virtual Bones from Monocular Videos
|
|
scholar.google.com.hk
|
FusionNet: Coarse-to-Fine Extrinsic Calibration Network of LiDAR and Camera with Hierarchical Point-Pixel Fusion
|
|
scholar.google.com.hk
|
Residual 3D Scene Flow Learning with Context-Aware Feature Extraction
|
|
scholar.google.com.hk
|
SFGAN: Unsupervised Generative Adversarial Learning of 3D Scene Flow from the 3D Scene Self
|
|
scholar.google.com.hk
|
Spherical Interpolated Convolutional Network with Distance-Feature Density for 3D Semantic Segmentation of Point Clouds
|
|
scholar.google.com.hk
|
Anchor-Based Spatio-Temporal Attention 3D Convolutional Networks for Dynamic 3D Point Cloud Sequences
|
|
scholar.google.com.hk
|
Unsupervised Learning of 3D Scene Flow from Monocular Camera
|
|
scholar.google.com.hk
|
Hierarchical Attention Learning of Scene Flow in 3D Point Clouds
|
|
scholar.google.com.hk
|
PWCLO-Net: Deep LiDAR Odometry in 3D Point Clouds Using Hierarchical Embedding Mask Optimization
|
|
scholar.google.com.hk
|
Unsupervised Learning of Depth, Optical Flow and Pose with Occlusion from 3D Geometry
|
|
scholar.google.com.hk
|
Unsupervised Learning of Monocular Depth and Ego-Motion Using Multiple Masks
|