Robotics

Papers filed under cs.RO on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

1,501 to 1,560 of 3,243

  1. Robotic Detection of Marine Litter Using Deep Visual Detection Models

    Michael Fulton, Jungseok Hong, Md Jahidul Islam +1

    cs.ROarXiv:1804.01079v22018
  2. Real-Time Execution of Action Chunking Flow Policies

    Kevin Black, Manuel Y. Galliker, Sergey Levine

    cs.ROcs.AIcs.LGarXiv:2506.07339v22025
  3. DiffuSearch: How Hybrid Trajectory Planning Benefits from Aligned Objectives in Diffusion and Action Space

    Steffen Hagedorn, Aron Distelzweig, Alexandru P. Condurache

    cs.ROcs.AIarXiv:2609.02252v12026
  4. End-to-end Driving via Conditional Imitation Learning

    Felipe Codevilla, Matthias Müller, Antonio López +2

    cs.ROcs.CVcs.LGarXiv:1710.02410v22017
  5. Learning Complex Dexterous Manipulation with Deep Reinforcement Learning and Demonstrations

    Aravind Rajeswaran, Vikash Kumar, Abhishek Gupta +4

    cs.LGcs.AIcs.ROarXiv:1709.10087v22017
  6. VINS-Mono: A Robust and Versatile Monocular Visual-Inertial State Estimator

    Tong Qin, Peiliang Li, Shaojie Shen

    cs.ROarXiv:1708.03852v12017
  7. ORB-SLAM2: an Open-Source SLAM System for Monocular, Stereo and RGB-D Cameras

    Raul Mur-Artal, Juan D. Tardos

    cs.ROcs.CVarXiv:1610.06475v22016
  8. Visual-Inertial Monocular SLAM with Map Reuse

    Raul Mur-Artal, Juan D. Tardos

    cs.ROcs.CVarXiv:1610.05949v22016
  9. Deep Visual Foresight for Planning Robot Motion

    Chelsea Finn, Sergey Levine

    cs.LGcs.AIcs.CVarXiv:1610.00696v22016
  10. From Perception to Decision: A Data-driven Approach to End-to-end Motion Planning for Autonomous Ground Robots

    Mark Pfeiffer, Michael Schaeuble, Juan Nieto +2

    cs.ROarXiv:1609.07910v32016
  11. Past, Present, and Future of Simultaneous Localization And Mapping: Towards the Robust-Perception Age

    Cesar Cadena, Luca Carlone, Henry Carrillo +5

    cs.ROarXiv:1606.05830v42016
  12. A Survey of Motion Planning and Control Techniques for Self-driving Urban Vehicles

    Brian Paden, Michal Cap, Sze Zheng Yong +2

    cs.ROarXiv:1604.07446v12016
  13. On-Manifold Preintegration for Real-Time Visual-Inertial Odometry

    Christian Forster, Luca Carlone, Frank Dellaert +1

    cs.ROarXiv:1512.02363v32015
  14. Optimal Multi-Robot Path Planning on Graphs: Complete Algorithms and Effective Heuristics

    Jingjin Yu, Steven M. LaValle

    cs.ROarXiv:1507.03290v12015
  15. ORB-SLAM: a Versatile and Accurate Monocular SLAM System

    Raul Mur-Artal, J. M. M. Montiel, Juan D. Tardos

    cs.ROcs.CVarXiv:1502.00956v22015
  16. Data-Driven Grasp Synthesis - A Survey

    Jeannette Bohg, Antonio Morales, Tamim Asfour +1

    cs.ROarXiv:1309.2660v22013
  17. Deep Learning for Detecting Robotic Grasps

    Ian Lenz, Honglak Lee, Ashutosh Saxena

    cs.LGcs.CVcs.ROarXiv:1301.3592v62013
  18. Learning Human Activities and Object Affordances from RGB-D Videos

    Hema Swetha Koppula, Rudhir Gupta, Ashutosh Saxena

    cs.ROcs.AIcs.CVarXiv:1210.1207v22012
  19. Sampling-based Algorithms for Optimal Motion Planning

    Sertac Karaman, Emilio Frazzoli

    cs.ROarXiv:1105.1186v12011
  20. StarVLA: A Lego-like Codebase for Vision-Language-Action Model Developing

    StarVLA Community

    cs.ROcs.AIcs.CVarXiv:2604.05014v12026
  21. VILENS: Visual, Inertial, Lidar, and Leg Odometry for All-Terrain Legged Robots

    David Wisth, Marco Camurri, Maurice Fallon

    cs.ROcs.CVarXiv:2107.07243v32021
  22. Learning to Rearrange Deformable Cables, Fabrics, and Bags with Goal-Conditioned Transporter Networks

    Daniel Seita, Pete Florence, Jonathan Tompson +4

    cs.ROcs.LGarXiv:2012.03385v42020
  23. Direct LiDAR Odometry: Fast Localization with Dense Point Clouds

    Kenny Chen, Brett T. Lopez, Ali-akbar Agha-mohammadi +1

    cs.ROarXiv:2110.00605v32021
  24. FoundationStereo: Zero-Shot Stereo Matching

    Bowen Wen, Matthew Trepte, Joseph Aribido +3

    cs.CVcs.LGcs.ROarXiv:2501.09898v42025
  25. SG-AMP: Scene-Graph-Guided Active Perception and Semantics-Aware Motion Planning for Pepper Plants

    Rohit Menon, Shiva Rudra Lolla, Niklas Mueller-Goldingen +3

    cs.ROarXiv:2609.01579v12026
  26. Anomaly Detection in Autonomous Driving: A Survey

    Daniel Bogdoll, Maximilian Nitsche, J. Marius Zöllner

    cs.ROarXiv:2204.07974v22022
  27. Fast3R: Towards 3D Reconstruction of 1000+ Images in One Forward Pass

    Jianing Yang, Alexander Sax, Kevin J. Liang +6

    cs.CVcs.AIcs.GRarXiv:2501.13928v22025
  28. RoboDreamer: Learning Compositional World Models for Robot Imagination

    Siyuan Zhou, Yilun Du, Jiaben Chen +3

    cs.ROarXiv:2404.12377v12024
  29. Motus: A Unified Latent Action World Model

    Hongzhe Bi, Hengkai Tan, Shenghao Xie +13

    cs.CVcs.LGcs.ROarXiv:2512.13030v22025
  30. BeyondMimic: From Motion Tracking to Versatile Humanoid Control via Guided Diffusion

    Qiayuan Liao, Takara E. Truong, Xiaoyu Huang +4

    cs.ROarXiv:2508.08241v42025
  31. Learning Decentralized Controllers for Robot Swarms with Graph Neural Networks

    Ekaterina Tolstaya, Fernando Gama, James Paulos +3

    cs.ROarXiv:1903.10527v42019
  32. WorldVLA: Towards Autoregressive Action World Model

    Jun Cen, Chaohui Yu, Hangjie Yuan +9

    cs.ROcs.AIarXiv:2506.21539v12025
  33. X-VLA: Soft-Prompted Transformer as Scalable Cross-Embodiment Vision-Language-Action Model

    Jinliang Zheng, Jianxiong Li, Zhihao Wang +12

    cs.ROcs.AIcs.CVarXiv:2510.10274v12025
  34. The Regretful Agent: Heuristic-Aided Navigation through Progress Estimation

    Chih-Yao Ma, Zuxuan Wu, Ghassan AlRegib +2

    cs.AIcs.CVcs.ROarXiv:1903.01602v12019
  35. LIBERO-Plus: In-depth Robustness Analysis of Vision-Language-Action Models

    Senyu Fei, Siyin Wang, Junhao Shi +10

    cs.ROcs.CLcs.CVarXiv:2510.13626v32025
  36. Multi-agent Path Planning and Network Flow

    Jingjin Yu, Steven M. LaValle

    cs.DScs.ROeess.SYarXiv:1204.5717v42012
  37. Modeling Human Motion with Quaternion-based Neural Networks

    Dario Pavllo, Christoph Feichtenhofer, Michael Auli +1

    cs.CVcs.AIcs.ROarXiv:1901.07677v22019
  38. Automatic Extrinsic Calibration Method for LiDAR and Camera Sensor Setups

    Jorge Beltrán, Carlos Guindel, Arturo de la Escalera +1

    cs.ROcs.CVarXiv:2101.04431v22021
  39. Deep Neural Networks for Multiple Speaker Detection and Localization

    Weipeng He, Petr Motlicek, Jean-Marc Odobez

    cs.SDcs.AIcs.MMarXiv:1711.11565v32017
  40. $π^{*}_{0.6}$: a VLA That Learns From Experience

    Physical Intelligence, Ali Amin, Raichelle Aniceto +53

    cs.LGcs.ROarXiv:2511.14759v22025
  41. Recursive Value Learning for Long-Horizon Offline Goal-Conditioned RL

    Hyeonseong Jeon, Youngwoon Lee

    cs.LGcs.ROarXiv:2609.02237v12026
  42. AgiBot World Colosseo: A Large-scale Manipulation Platform for Scalable and Intelligent Embodied Systems

    AgiBot-World-Contributors, Qingwen Bu, Jisong Cai +49

    cs.ROcs.CVcs.LGarXiv:2503.06669v42025
  43. CoT-VLA: Visual Chain-of-Thought Reasoning for Vision-Language-Action Models

    Qingqing Zhao, Yao Lu, Moo Jin Kim +12

    cs.CVcs.AIcs.LGarXiv:2503.22020v12025
  44. Designing Versatile Samples for Learned Trajectory Scoring

    Yaguang Li, Jiaru Zhang, Chuheng Wei +2

    cs.ROcs.CVarXiv:2609.01799v12026
  45. LookStep: Efficient Vision-Language Navigation with Linguistic Foresight and Event Driven Memory

    Kun-Yang Yu, Yingzhe Li, Hongyu Xu +8

    cs.CVcs.ROarXiv:2609.02350v12026
  46. Towards MRI-Based Autonomous Robotic US Acquisitions: A First Feasibility Study

    Christoph Hennersperger, Bernhard Fuerst, Salvatore Virga +4

    cs.ROarXiv:1607.08371v12016
  47. CrashDiffuser: VLM-Guided Collision Intent Reasoning for Fine-Grained Safety-Critical Traffic Scenario Generation

    Shucheng Zhang, Yuang Zhang, Bingzhang Wang +3

    cs.ROcs.AIarXiv:2609.02270v12026
  48. In-Hand Object Rotation via Rapid Motor Adaptation

    Haozhi Qi, Ashish Kumar, Roberto Calandra +2

    cs.ROcs.AIcs.CVarXiv:2210.04887v12022
  49. SpatialVLA: Exploring Spatial Representations for Visual-Language-Action Model

    Delin Qu, Haoming Song, Qizhi Chen +8

    cs.ROcs.AIarXiv:2501.15830v52025
  50. UniVLA: Learning to Act Anywhere with Task-centric Latent Actions

    Qingwen Bu, Yanting Yang, Jisong Cai +5

    cs.ROcs.AIcs.LGarXiv:2505.06111v32025
  51. Consistency Policy: Accelerated Visuomotor Policies via Consistency Distillation

    Aaditya Prasad, Kevin Lin, Jimmy Wu +2

    cs.ROcs.AIarXiv:2405.07503v22024
  52. FAST: Efficient Action Tokenization for Vision-Language-Action Models

    Karl Pertsch, Kyle Stachowicz, Brian Ichter +6

    cs.ROcs.LGarXiv:2501.09747v12025
  53. SNE-RoadSeg: Incorporating Surface Normal Information into Semantic Segmentation for Accurate Freespace Detection

    Rui Fan, Hengli Wang, Peide Cai +1

    cs.CVcs.ROeess.IVarXiv:2008.11351v12020
  54. RoboTwin 2.0: A Scalable Data Generator and Benchmark with Strong Domain Randomization for Robust Bimanual Robotic Manipulation

    Tianxing Chen, Zanxin Chen, Baijun Chen +23

    cs.ROcs.AIcs.CLarXiv:2506.18088v22025
  55. V-JEPA 2: Self-Supervised Video Models Enable Understanding, Prediction and Planning

    Mido Assran, Adrien Bardes, David Fan +27

    cs.AIcs.CVcs.LGarXiv:2506.09985v12025
  56. Conditional Generative Neural System for Probabilistic Trajectory Prediction

    Jiachen Li, Hengbo Ma, Masayoshi Tomizuka

    cs.CVcs.AIcs.LGarXiv:1905.01631v22019
  57. Automated Speed and Lane Change Decision Making using Deep Reinforcement Learning

    Carl-Johan Hoel, Krister Wolff, Leo Laine

    cs.ROcs.AIcs.LGarXiv:1803.10056v22018
  58. Towards Human-Level Bimanual Dexterous Manipulation with Reinforcement Learning

    Yuanpei Chen, Tianhao Wu, Shengjie Wang +8

    cs.ROcs.AIcs.LGarXiv:2206.08686v22022
  59. AM-Bench: A Modular Simulation Suite and Benchmark for Aerial Manipulation Policy Learning

    Yutong Wang, Dongjae Lee, Xiaofeng Guo +9

    cs.ROarXiv:2609.00641v12026
  60. Safe Exploration in Finite Markov Decision Processes with Gaussian Processes

    Matteo Turchetta, Felix Berkenkamp, Andreas Krause

    cs.LGcs.AIcs.ROarXiv:1606.04753v22016