Robotics

Papers filed under cs.RO on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

121 to 180 of 3,242

  1. Embodied Agent Interface: Benchmarking LLMs for Embodied Decision Making

    Manling Li, Shiyu Zhao, Qineng Wang +12

    cs.CLcs.AIcs.LGarXiv:2410.07166v32024
  2. Adaptive Vision-Language Grasping via Composable Foundation Priors and Generalizable Grasp Synthesis

    Sixu Yan, Shikang Wang, Binhua Huang +11

    cs.ROcs.AIcs.CVarXiv:2609.04096v12026
  3. A System for Fast, Resilient, and Adaptable Loco-Manipulation Behaviors on Humanoid Robots

    Duncan Calvert, Luigi Penco, Dexton Anderson +4

    cs.ROarXiv:2609.01518v12026
  4. FFRob: Leveraging Symbolic Planning for Efficient Task and Motion Planning

    Caelan Reed Garrett, Tomas Lozano-Perez, Leslie Pack Kaelbling

    cs.ROarXiv:1608.01335v22016
  5. RGBD Datasets: Past, Present and Future

    Michael Firman

    cs.CVcs.ROarXiv:1604.00999v22016
  6. Enhanced Robot Audition Based on Microphone Array Source Separation with Post-Filter

    Jean-Marc Valin, Jean Rouat, François Michaud

    cs.ROcs.SDarXiv:1603.02341v12016
  7. Diffusion Policy: Visuomotor Policy Learning via Action Diffusion

    Cheng Chi, Zhenjia Xu, Siyuan Feng +5

    cs.ROarXiv:2303.04137v52023
  8. Remote Human and Robot Interaction for Greenhouse Gardening Using Virtual Reality

    Daniel Udekwe, Hasan Seyyedhasani

    cs.ROeess.SYarXiv:2608.27545v12026
  9. MeshPriorDiT: Hierarchical Modeling for Action-Conditioned Cloth Dynamics

    Zihang Wang, Jianming Hu, Shang Su +4

    cs.ROarXiv:2608.26766v12026
  10. RoG-DAgger: Rollout-Guided Post-Training for End-to-End Driving

    Liangyu Zhong, Joachim Sicking, Fabian Hueger +1

    cs.ROarXiv:2608.24525v12026
  11. Hierarchical Skill Retrieval for Data-Efficient Adaptation of Vision-Language-Action Models

    Haoran Hao, Shahram Najam Syed, Jeff Schneider +1

    cs.ROcs.AIcs.LGarXiv:2608.24042v12026
  12. DELE-w0.5: Inferring Action from Future Latent State for Robotic Manipulation

    Fenghao Lei, Zhixiong Huang, Long Yang +5

    cs.ROcs.AIcs.CVarXiv:2608.22067v32026
  13. AMZ Driverless: The Full Autonomous Racing System

    Juraj Kabzan, Miguel de la Iglesia Valls, Victor Reijgwart +19

    cs.ROarXiv:1905.05150v12019
  14. Depth from Videos in the Wild: Unsupervised Monocular Depth Learning from Unknown Cameras

    Ariel Gordon, Hanhan Li, Rico Jonschkowski +1

    cs.CVcs.GRcs.LGarXiv:1904.04998v12019
  15. tinyDSM: A Framework for Skill Modeling and Development for Resource-Constrained Millirobots

    Markus D. Kobelrausch, Michael Miedler, Axel Jantsch

    cs.ROcs.AIarXiv:2608.17596v12026
  16. Planning Optimal Paths for Multiple Robots on Graphs

    Jingjin Yu, Steven M. LaValle

    cs.ROcs.AIeess.SYarXiv:1204.3830v42012
  17. Learning agile and dynamic motor skills for legged robots

    Jemin Hwangbo, Joonho Lee, Alexey Dosovitskiy +4

    cs.ROcs.LGstat.MLarXiv:1901.08652v12019
  18. Language Models as Zero-Shot Planners: Extracting Actionable Knowledge for Embodied Agents

    Wenlong Huang, Pieter Abbeel, Deepak Pathak +1

    cs.LGcs.AIcs.CLarXiv:2201.07207v22022
  19. Perception and Sensing for Autonomous Vehicles Under Adverse Weather Conditions: A Survey

    Yuxiao Zhang, Alexander Carballo, Hanting Yang +1

    cs.ROarXiv:2112.08936v22021
  20. Image2Sim: Scaling Embodied Navigation via Generative Neural Simulator

    Zihan Wang, Seungjun Lee, Yinghao Xu +1

    cs.CVcs.ROarXiv:2607.05765v12026
  21. A Persistent Spatial Semantic Representation for High-level Natural Language Instruction Execution

    Valts Blukis, Chris Paxton, Dieter Fox +2

    cs.ROcs.AIcs.CLarXiv:2107.05612v32021
  22. The Road Ahead in Autonomous Driving: The KITScenes Multimodal Dataset

    Richard Schwarzkopf, Fabian Immel, Alexander Blumberg +21

    cs.CVcs.LGcs.ROarXiv:2606.02956v12026
  23. The EuroCity Persons Dataset: A Novel Benchmark for Object Detection

    Markus Braun, Sebastian Krebs, Fabian Flohr +1

    cs.CVcs.AIcs.LGarXiv:1805.07193v22018
  24. Multi-objective path planning of an autonomous mobile robot using hybrid PSO-MFB optimization algorithm

    Fatin H. Ajeil, Ibraheem Kasim Ibraheem, Mouayad A. Sahib +1

    cs.ROarXiv:1805.00224v32018
  25. Towards Generalizable Robotic Manipulation in Dynamic Environments

    Heng Fang, Shangru Li, Shuhan Wang +3

    cs.CVcs.ROarXiv:2603.15620v32026
  26. Steering Vision-Language-Action Models as Anti-Exploration: A Test-Time Scaling Approach

    Siyuan Yang, Yang Zhang, Haoran He +4

    cs.ROcs.AIarXiv:2512.02834v12025
  27. Maximum Entropy RL (Provably) Solves Some Robust RL Problems

    Benjamin Eysenbach, Sergey Levine

    cs.LGcs.ROarXiv:2103.06257v22021
  28. Virtual Testing of Automated Driving Systems through Credible Simulations

    Riccardo Dona, Espedito Rusciano, Biagio Ciuffo

    cs.ROcs.SEarXiv:2609.03760v12026
  29. TRaIL-Odom: Tightly Coupled Continuous Time Radar-IMU-LiDAR Odometry with Adaptive Doppler Weighting

    Chiyun Noh, Turcan Tuna, William Talbot +3

    cs.ROarXiv:2609.03561v12026
  30. DensePhysNet: Learning Dense Physical Object Representations via Multi-step Dynamic Interactions

    Zhenjia Xu, Jiajun Wu, Andy Zeng +2

    cs.ROcs.AIcs.CVarXiv:1906.03853v22019
  31. A Survey on Efficient Vision-Language-Action Models

    Zhaoshu Yu, Bo Wang, Pengpeng Zeng +7

    cs.CVcs.AIcs.LGarXiv:2510.24795v22025
  32. SwiftVLA: Unlocking Spatiotemporal Dynamics for Lightweight VLA Models at Minimal Overhead

    Chaojun Ni, Cheng Chen, Xiaofeng Wang +12

    cs.CVcs.ROarXiv:2512.00903v12025
  33. NORA-1.5: A Vision-Language-Action Model Trained using World Model- and Action-based Preference Rewards

    Chia-Yu Hung, Navonil Majumder, Haoyuan Deng +7

    cs.ROcs.AIarXiv:2511.14659v12025
  34. Equilibria for Networks of Linear Translational Springs

    Luke Oeding, Ethan Clayton, Jackson Elsea +2

    math.AGcs.ROarXiv:2609.03143v12026
  35. Mixture of Horizons in Action Chunking

    Dong Jing, Gang Wang, Jiaqi Liu +7

    cs.ROcs.AIcs.CVarXiv:2511.19433v22025
  36. Vision-Language-Action Models for Autonomous Driving: Past, Present, and Future

    Tianshuai Hu, Xiaolu Liu, Song Wang +17

    cs.ROarXiv:2512.16760v22025
  37. Fast-FoundationStereo: Real-Time Zero-Shot Stereo Matching

    Bowen Wen, Shaurya Dewan, Stan Birchfield

    cs.CVcs.ROarXiv:2512.11130v22025
  38. Distributed Mapping with Privacy and Communication Constraints: Lightweight Algorithms and Object-based Models

    Siddharth Choudhary, Luca Carlone, Carlos Nieto +3

    cs.ROcs.CVarXiv:1702.03435v12017
  39. RoboChallenge: Large-scale Real-robot Evaluation of Embodied Policies

    Adina Yakefu, Bin Xie, Chongyang Xu +34

    cs.ROarXiv:2510.17950v12025
  40. KeyPose: Multi-View 3D Labeling and Keypoint Estimation for Transparent Objects

    Xingyu Liu, Rico Jonschkowski, Anelia Angelova +1

    cs.CVcs.LGcs.ROarXiv:1912.02805v22019
  41. HiF-VLA: Hindsight, Insight and Foresight through Motion Representation for Vision-Language-Action Models

    Minghui Lin, Pengxiang Ding, Shu Wang +7

    cs.ROarXiv:2512.09928v22025
  42. See, Point, Fly: A Learning-Free VLM Framework for Universal Unmanned Aerial Navigation

    Chih Yao Hu, Yang-Sen Lin, Yuna Lee +7

    cs.ROcs.AIcs.CLarXiv:2509.22653v12025
  43. DEXOP: A Device for Robotic Transfer of Dexterous Human Manipulation

    Hao-Shu Fang, Branden Romero, Yichen Xie +9

    cs.ROcs.AIcs.CVarXiv:2509.04441v22025
  44. A Mathematical Theory of Pragmatic Information

    Kai Niu, Ping Zhang

    cs.ITcs.AIcs.ROarXiv:2609.10986v12026
  45. EVPeriscope: Extended Perception across Aerial and Ground Vehicles with Event-based Propeller Tracking

    Dexter Ong, Vijay Kumar, Pratik Chaudhari

    cs.ROarXiv:2609.11920v12026
  46. UniMPA: A Unified Memory-Prediction-Action Model via Action-Grounded Transition Modeling

    Wei Li, Rui Shao, Jie He +3

    cs.ROarXiv:2609.11875v12026
  47. Learning Agent-based Model Predictive Control for Holistic Vehicle Performance

    Jiaming Zhong, Reza Valiollahi Mehrizi, Mohammad Pirani +4

    cs.ROarXiv:2609.11871v12026
  48. Rapid Learning of Dexterous In-Hand Pen Writing through Real-Time Jacobian Estimation

    Kai Stewart, Yasunori Toshimitsu, Robert K. Katzschmann

    cs.ROarXiv:2609.11775v12026
  49. SEED-UMI: Sharing the Exoskeleton between human and robot for onE-to-one Dexterous demonstration

    Tengbo Yu, Jiahao Wu, Daohan Li +4

    cs.ROarXiv:2609.11753v12026
  50. Aerodynamic Prior-Free Coordinated Trajectory Generation and Tracking Control for a Tail-Sitter UAV

    Erchao Rong, Zihao Liu, Junning Liang +5

    cs.ROarXiv:2609.11698v12026
  51. Visual-SLAM for the detection of hidden tomatoes in greenhouses by Hierarchical Localization and GLOMAPfor robotized harvesting

    Fernando Cañadas-Aránega, José C. Moreno, José L. Blanco-Claraco +1

    cs.ROarXiv:2609.11766v12026
  52. ActSafeGuard: Differentiable and Training-Aligned Constraint Enforcement for Flow-Matching Policies

    Jianming Ma, Rongjun Jin, Xiaxi Si +3

    cs.ROcs.AIarXiv:2609.11697v12026
  53. Contact-Aware Incremental Model Predictive Control for an Underactuated Aerial Manipulator

    Darwin Liu, Tamas Keviczky, Sihao Sun

    cs.ROarXiv:2609.11661v12026
  54. Using Automated Vehicles Operational Data to Confirm Safety and Anticipate Threats

    Riccardo Donà, Espedito Rusciano, Germana Trentadue +2

    cs.ROarXiv:2609.11549v12026
  55. CAP: Continuously Adaptive Perception-Blind Humanoid Locomotion via Learned Denoising

    Hongjin Chen, Zijun Xu, Shihao Ma +8

    cs.ROarXiv:2609.11553v12026
  56. Quasi-static analysis of passive stability in a novel underactuated multi-finger hand

    Léonie Plancoulaine, Sylvain Guégan, Franck Plestan +1

    cs.ROarXiv:2609.11579v12026
  57. CARLAverse: A Highly Modular, Distributed, and Multimodal Framework for Human-in-the-Loop Simulation

    Patrick Rebling, Philipp Nenninger, Reiner Kriesten

    cs.ROcs.HCarXiv:2609.11478v12026
  58. GigaBrain-0: A World Model-Powered Vision-Language-Action Model

    GigaBrain Team, Angen Ye, Boyuan Wang +24

    cs.ROcs.CVarXiv:2510.19430v32025
  59. 3D Euler-Angle Orientation Control for Two-Ray Fading Mitigation in Maritime Air-to-Sea Communications

    Mohammed Bajja, Abdoul Karim A. H. Saliah, Hajar El Hammouti +2

    cs.ROarXiv:2609.11476v12026
  60. Autonomous Recharging and Flight Mission Planning for Battery-operated Autonomous Drones

    Rashid Alyassi, Majid Khonji, Areg Karapetyan +3

    cs.ROcs.DSarXiv:1703.10049v52017