Robotics
Papers filed under cs.RO on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
121 to 180 of 3,242
Embodied Agent Interface: Benchmarking LLMs for Embodied Decision Making
Manling Li, Shiyu Zhao, Qineng Wang +12
cs.CLcs.AIcs.LGarXiv:2410.07166v32024Adaptive Vision-Language Grasping via Composable Foundation Priors and Generalizable Grasp Synthesis
Sixu Yan, Shikang Wang, Binhua Huang +11
cs.ROcs.AIcs.CVarXiv:2609.04096v12026A System for Fast, Resilient, and Adaptable Loco-Manipulation Behaviors on Humanoid Robots
Duncan Calvert, Luigi Penco, Dexton Anderson +4
cs.ROarXiv:2609.01518v12026FFRob: Leveraging Symbolic Planning for Efficient Task and Motion Planning
Caelan Reed Garrett, Tomas Lozano-Perez, Leslie Pack Kaelbling
cs.ROarXiv:1608.01335v22016RGBD Datasets: Past, Present and Future
Michael Firman
cs.CVcs.ROarXiv:1604.00999v22016Enhanced Robot Audition Based on Microphone Array Source Separation with Post-Filter
Jean-Marc Valin, Jean Rouat, François Michaud
cs.ROcs.SDarXiv:1603.02341v12016Diffusion Policy: Visuomotor Policy Learning via Action Diffusion
Cheng Chi, Zhenjia Xu, Siyuan Feng +5
cs.ROarXiv:2303.04137v52023Remote Human and Robot Interaction for Greenhouse Gardening Using Virtual Reality
Daniel Udekwe, Hasan Seyyedhasani
cs.ROeess.SYarXiv:2608.27545v12026MeshPriorDiT: Hierarchical Modeling for Action-Conditioned Cloth Dynamics
Zihang Wang, Jianming Hu, Shang Su +4
cs.ROarXiv:2608.26766v12026RoG-DAgger: Rollout-Guided Post-Training for End-to-End Driving
Liangyu Zhong, Joachim Sicking, Fabian Hueger +1
cs.ROarXiv:2608.24525v12026Hierarchical Skill Retrieval for Data-Efficient Adaptation of Vision-Language-Action Models
Haoran Hao, Shahram Najam Syed, Jeff Schneider +1
cs.ROcs.AIcs.LGarXiv:2608.24042v12026DELE-w0.5: Inferring Action from Future Latent State for Robotic Manipulation
Fenghao Lei, Zhixiong Huang, Long Yang +5
cs.ROcs.AIcs.CVarXiv:2608.22067v32026AMZ Driverless: The Full Autonomous Racing System
Juraj Kabzan, Miguel de la Iglesia Valls, Victor Reijgwart +19
cs.ROarXiv:1905.05150v12019Depth from Videos in the Wild: Unsupervised Monocular Depth Learning from Unknown Cameras
Ariel Gordon, Hanhan Li, Rico Jonschkowski +1
cs.CVcs.GRcs.LGarXiv:1904.04998v12019tinyDSM: A Framework for Skill Modeling and Development for Resource-Constrained Millirobots
Markus D. Kobelrausch, Michael Miedler, Axel Jantsch
cs.ROcs.AIarXiv:2608.17596v12026Planning Optimal Paths for Multiple Robots on Graphs
Jingjin Yu, Steven M. LaValle
cs.ROcs.AIeess.SYarXiv:1204.3830v42012Learning agile and dynamic motor skills for legged robots
Jemin Hwangbo, Joonho Lee, Alexey Dosovitskiy +4
cs.ROcs.LGstat.MLarXiv:1901.08652v12019Language Models as Zero-Shot Planners: Extracting Actionable Knowledge for Embodied Agents
Wenlong Huang, Pieter Abbeel, Deepak Pathak +1
cs.LGcs.AIcs.CLarXiv:2201.07207v22022Perception and Sensing for Autonomous Vehicles Under Adverse Weather Conditions: A Survey
Yuxiao Zhang, Alexander Carballo, Hanting Yang +1
cs.ROarXiv:2112.08936v22021Image2Sim: Scaling Embodied Navigation via Generative Neural Simulator
Zihan Wang, Seungjun Lee, Yinghao Xu +1
cs.CVcs.ROarXiv:2607.05765v12026A Persistent Spatial Semantic Representation for High-level Natural Language Instruction Execution
Valts Blukis, Chris Paxton, Dieter Fox +2
cs.ROcs.AIcs.CLarXiv:2107.05612v32021The Road Ahead in Autonomous Driving: The KITScenes Multimodal Dataset
Richard Schwarzkopf, Fabian Immel, Alexander Blumberg +21
cs.CVcs.LGcs.ROarXiv:2606.02956v12026The EuroCity Persons Dataset: A Novel Benchmark for Object Detection
Markus Braun, Sebastian Krebs, Fabian Flohr +1
cs.CVcs.AIcs.LGarXiv:1805.07193v22018Multi-objective path planning of an autonomous mobile robot using hybrid PSO-MFB optimization algorithm
Fatin H. Ajeil, Ibraheem Kasim Ibraheem, Mouayad A. Sahib +1
cs.ROarXiv:1805.00224v32018Towards Generalizable Robotic Manipulation in Dynamic Environments
Heng Fang, Shangru Li, Shuhan Wang +3
cs.CVcs.ROarXiv:2603.15620v32026Steering Vision-Language-Action Models as Anti-Exploration: A Test-Time Scaling Approach
Siyuan Yang, Yang Zhang, Haoran He +4
cs.ROcs.AIarXiv:2512.02834v12025Maximum Entropy RL (Provably) Solves Some Robust RL Problems
Benjamin Eysenbach, Sergey Levine
cs.LGcs.ROarXiv:2103.06257v22021Virtual Testing of Automated Driving Systems through Credible Simulations
Riccardo Dona, Espedito Rusciano, Biagio Ciuffo
cs.ROcs.SEarXiv:2609.03760v12026TRaIL-Odom: Tightly Coupled Continuous Time Radar-IMU-LiDAR Odometry with Adaptive Doppler Weighting
Chiyun Noh, Turcan Tuna, William Talbot +3
cs.ROarXiv:2609.03561v12026DensePhysNet: Learning Dense Physical Object Representations via Multi-step Dynamic Interactions
Zhenjia Xu, Jiajun Wu, Andy Zeng +2
cs.ROcs.AIcs.CVarXiv:1906.03853v22019A Survey on Efficient Vision-Language-Action Models
Zhaoshu Yu, Bo Wang, Pengpeng Zeng +7
cs.CVcs.AIcs.LGarXiv:2510.24795v22025SwiftVLA: Unlocking Spatiotemporal Dynamics for Lightweight VLA Models at Minimal Overhead
Chaojun Ni, Cheng Chen, Xiaofeng Wang +12
cs.CVcs.ROarXiv:2512.00903v12025NORA-1.5: A Vision-Language-Action Model Trained using World Model- and Action-based Preference Rewards
Chia-Yu Hung, Navonil Majumder, Haoyuan Deng +7
cs.ROcs.AIarXiv:2511.14659v12025Equilibria for Networks of Linear Translational Springs
Luke Oeding, Ethan Clayton, Jackson Elsea +2
math.AGcs.ROarXiv:2609.03143v12026Mixture of Horizons in Action Chunking
Dong Jing, Gang Wang, Jiaqi Liu +7
cs.ROcs.AIcs.CVarXiv:2511.19433v22025Vision-Language-Action Models for Autonomous Driving: Past, Present, and Future
Tianshuai Hu, Xiaolu Liu, Song Wang +17
cs.ROarXiv:2512.16760v22025Fast-FoundationStereo: Real-Time Zero-Shot Stereo Matching
Bowen Wen, Shaurya Dewan, Stan Birchfield
cs.CVcs.ROarXiv:2512.11130v22025Distributed Mapping with Privacy and Communication Constraints: Lightweight Algorithms and Object-based Models
Siddharth Choudhary, Luca Carlone, Carlos Nieto +3
cs.ROcs.CVarXiv:1702.03435v12017RoboChallenge: Large-scale Real-robot Evaluation of Embodied Policies
Adina Yakefu, Bin Xie, Chongyang Xu +34
cs.ROarXiv:2510.17950v12025KeyPose: Multi-View 3D Labeling and Keypoint Estimation for Transparent Objects
Xingyu Liu, Rico Jonschkowski, Anelia Angelova +1
cs.CVcs.LGcs.ROarXiv:1912.02805v22019HiF-VLA: Hindsight, Insight and Foresight through Motion Representation for Vision-Language-Action Models
Minghui Lin, Pengxiang Ding, Shu Wang +7
cs.ROarXiv:2512.09928v22025See, Point, Fly: A Learning-Free VLM Framework for Universal Unmanned Aerial Navigation
Chih Yao Hu, Yang-Sen Lin, Yuna Lee +7
cs.ROcs.AIcs.CLarXiv:2509.22653v12025DEXOP: A Device for Robotic Transfer of Dexterous Human Manipulation
Hao-Shu Fang, Branden Romero, Yichen Xie +9
cs.ROcs.AIcs.CVarXiv:2509.04441v22025A Mathematical Theory of Pragmatic Information
Kai Niu, Ping Zhang
cs.ITcs.AIcs.ROarXiv:2609.10986v12026EVPeriscope: Extended Perception across Aerial and Ground Vehicles with Event-based Propeller Tracking
Dexter Ong, Vijay Kumar, Pratik Chaudhari
cs.ROarXiv:2609.11920v12026UniMPA: A Unified Memory-Prediction-Action Model via Action-Grounded Transition Modeling
Wei Li, Rui Shao, Jie He +3
cs.ROarXiv:2609.11875v12026Learning Agent-based Model Predictive Control for Holistic Vehicle Performance
Jiaming Zhong, Reza Valiollahi Mehrizi, Mohammad Pirani +4
cs.ROarXiv:2609.11871v12026Rapid Learning of Dexterous In-Hand Pen Writing through Real-Time Jacobian Estimation
Kai Stewart, Yasunori Toshimitsu, Robert K. Katzschmann
cs.ROarXiv:2609.11775v12026SEED-UMI: Sharing the Exoskeleton between human and robot for onE-to-one Dexterous demonstration
Tengbo Yu, Jiahao Wu, Daohan Li +4
cs.ROarXiv:2609.11753v12026Aerodynamic Prior-Free Coordinated Trajectory Generation and Tracking Control for a Tail-Sitter UAV
Erchao Rong, Zihao Liu, Junning Liang +5
cs.ROarXiv:2609.11698v12026Visual-SLAM for the detection of hidden tomatoes in greenhouses by Hierarchical Localization and GLOMAPfor robotized harvesting
Fernando Cañadas-Aránega, José C. Moreno, José L. Blanco-Claraco +1
cs.ROarXiv:2609.11766v12026ActSafeGuard: Differentiable and Training-Aligned Constraint Enforcement for Flow-Matching Policies
Jianming Ma, Rongjun Jin, Xiaxi Si +3
cs.ROcs.AIarXiv:2609.11697v12026Contact-Aware Incremental Model Predictive Control for an Underactuated Aerial Manipulator
Darwin Liu, Tamas Keviczky, Sihao Sun
cs.ROarXiv:2609.11661v12026Using Automated Vehicles Operational Data to Confirm Safety and Anticipate Threats
Riccardo Donà, Espedito Rusciano, Germana Trentadue +2
cs.ROarXiv:2609.11549v12026CAP: Continuously Adaptive Perception-Blind Humanoid Locomotion via Learned Denoising
Hongjin Chen, Zijun Xu, Shihao Ma +8
cs.ROarXiv:2609.11553v12026Quasi-static analysis of passive stability in a novel underactuated multi-finger hand
Léonie Plancoulaine, Sylvain Guégan, Franck Plestan +1
cs.ROarXiv:2609.11579v12026CARLAverse: A Highly Modular, Distributed, and Multimodal Framework for Human-in-the-Loop Simulation
Patrick Rebling, Philipp Nenninger, Reiner Kriesten
cs.ROcs.HCarXiv:2609.11478v12026GigaBrain-0: A World Model-Powered Vision-Language-Action Model
GigaBrain Team, Angen Ye, Boyuan Wang +24
cs.ROcs.CVarXiv:2510.19430v320253D Euler-Angle Orientation Control for Two-Ray Fading Mitigation in Maritime Air-to-Sea Communications
Mohammed Bajja, Abdoul Karim A. H. Saliah, Hajar El Hammouti +2
cs.ROarXiv:2609.11476v12026Autonomous Recharging and Flight Mission Planning for Battery-operated Autonomous Drones
Rashid Alyassi, Majid Khonji, Areg Karapetyan +3
cs.ROcs.DSarXiv:1703.10049v52017