Robotics

Papers filed under cs.RO on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

1,081 to 1,140 of 3,243

  1. ApexNav: An Adaptive Exploration Strategy for Zero-Shot Object Navigation with Target-centric Semantic Fusion

    Mingjie Zhang, Yuheng Du, Chengkai Wu +4

    cs.ROarXiv:2504.14478v32025
  2. A Survey on Vision-Language-Action Models for Autonomous Driving

    Sicong Jiang, Zilin Huang, Kangan Qian +17

    cs.CVcs.AIcs.ROarXiv:2506.24044v12025
  3. Motubrain: An Advanced World Action Model for Robot Control

    Motubrain Team, Chendong Xiang, Fan Bao +17

    cs.ROarXiv:2604.27792v52026
  4. UAVs Meet LLMs: Overviews and Perspectives Toward Agentic Low-Altitude Mobility

    Yonglin Tian, Fei Lin, Yiduo Li +11

    cs.ROcs.AIarXiv:2501.02341v22025
  5. Humanoid Policy ~ Human Policy

    Ri-Zhao Qiu, Shiqi Yang, Xuxin Cheng +12

    cs.ROcs.AIcs.CVarXiv:2503.13441v32025
  6. Deep-Learned Collision Avoidance Policy for Distributed Multi-Agent Navigation

    Pinxin Long, Wenxi Liu, Jia Pan

    cs.AIcs.CVcs.ROarXiv:1609.06838v22016
  7. MoLe-VLA: Dynamic Layer-skipping Vision Language Action Model via Mixture-of-Layers for Efficient Robot Manipulation

    Rongyu Zhang, Menghang Dong, Yuan Zhang +6

    cs.ROcs.AIarXiv:2503.20384v22025
  8. Benchmarking Reinforcement Learning Algorithms on Real-World Robots

    A. Rupam Mahmood, Dmytro Korenkevych, Gautham Vasan +2

    cs.LGcs.AIcs.ROarXiv:1809.07731v12018
  9. Adaptive Depth-Map-Guided Bundle Adjustment for Correspondence-Free Multi-View Point Cloud Registration

    Yiran Zhou, Yingyu Wang, Shoudong Huang +1

    cs.ROarXiv:2609.01089v12026
  10. FLOWER: Democratizing Generalist Robot Policies with Efficient Vision-Language-Action Flow Policies

    Moritz Reuss, Hongyi Zhou, Marcel Rühle +3

    cs.ROarXiv:2509.04996v12025
  11. ChatVLA-2: Vision-Language-Action Model with Open-World Embodied Reasoning from Pretrained Knowledge

    Zhongyi Zhou, Yichen Zhu, Junjie Wen +2

    cs.ROcs.AIcs.CVarXiv:2505.21906v22025
  12. P3Depth: Monocular Depth Estimation with a Piecewise Planarity Prior

    Vaishakh Patil, Christos Sakaridis, Alexander Liniger +1

    cs.CVcs.AIcs.LGarXiv:2204.02091v12022
  13. Improving Vision-Language-Action Model with Online Reinforcement Learning

    Yanjiang Guo, Jianke Zhang, Xiaoyu Chen +4

    cs.ROcs.CVcs.LGarXiv:2501.16664v12025
  14. WoVR: World Models as Reliable Simulators for Post-Training VLA Policies with RL

    Zhennan Jiang, Shangqing Zhou, Yutong Jiang +11

    cs.ROcs.AIarXiv:2602.13977v22026
  15. A soft robot that adapts to environments through shape change

    Dylan S. Shah, Joshua P. Powers, Liana G. Tilton +3

    cs.ROarXiv:2008.06397v52020
  16. LDA-1B: Scaling Latent Dynamics Action Model via Universal Embodied Data Ingestion

    Jiangran Lyu, Kai Liu, Xuheng Zhang +20

    cs.ROarXiv:2602.12215v22026
  17. Beyond Textual Chain-of-Thought: A Survey on Action-Grounded Reasoning in Autonomous Driving

    Zhengxu Tang, Xiaozhou Zhang, Guofeng Cui +10

    cs.CVcs.CLcs.ROarXiv:2609.01659v12026
  18. ReinFlow: Fine-tuning Flow Matching Policy with Online Reinforcement Learning

    Tonghe Zhang, Chao Yu, Sichang Su +1

    cs.ROcs.LGarXiv:2505.22094v72025
  19. CoLT-Drive: Counterfactual Long-Tail Benchmarking and Knowledge-Preserving Adaptation for Driving Affordance Prediction

    Zhengxu Tang, Guofeng Cui, Ziyu Gong +8

    cs.CVcs.AIcs.CLarXiv:2609.00242v12026
  20. RAD: Training an End-to-End Driving Policy via Large-Scale 3DGS-based Reinforcement Learning

    Hao Gao, Shaoyu Chen, Bo Jiang +11

    cs.CVcs.ROarXiv:2502.13144v22025
  21. Humanoid Parkour Learning

    Ziwen Zhuang, Shenzhe Yao, Hang Zhao

    cs.ROarXiv:2406.10759v22024
  22. Neural probabilistic motor primitives for humanoid control

    Josh Merel, Leonard Hasenclever, Alexandre Galashov +5

    cs.LGcs.AIcs.ROarXiv:1811.11711v22018
  23. Variable Compliance Control for Robotic Peg-in-Hole Assembly: A Deep Reinforcement Learning Approach

    Cristian C. Beltran-Hernandez, Damien Petit, Ixchel G. Ramirez-Alpizar +1

    cs.ROcs.LGarXiv:2008.10224v32020
  24. DTC: Deep Tracking Control

    Fabian Jenelten, Junzhe He, Farbod Farshidian +1

    cs.ROcs.LGeess.SYarXiv:2309.15462v22023
  25. InternVLA-A1: Unifying Understanding, Generation and Action for Robotic Manipulation

    Junhao Cai, Zetao Cai, Jiafei Cao +39

    cs.ROarXiv:2601.02456v22026
  26. Data-Driven Model Predictive Control of Autonomous Mobility-on-Demand Systems

    Ramon Iglesias, Federico Rossi, Kevin Wang +3

    cs.ROcs.MAeess.SYarXiv:1709.07032v12017
  27. FASTER: Fast and Safe Trajectory Planner for Flights in Unknown Environments

    Jesus Tordesillas, Brett T. Lopez, Jonathan P. How

    cs.ROarXiv:1903.03558v32019
  28. Video Generators are Robot Policies

    Junbang Liang, Pavel Tokmakov, Ruoshi Liu +4

    cs.ROarXiv:2508.00795v12025
  29. Finding a needle in an exponential haystack: Discrete RRT for exploration of implicit roadmaps in multi-robot motion planning

    Kiril Solovey, Oren Salzman, Dan Halperin

    cs.ROarXiv:1305.2889v32013
  30. Imitation Learning as $f$-Divergence Minimization

    Liyiming Ke, Sanjiban Choudhury, Matt Barnes +3

    cs.LGcs.ITcs.ROarXiv:1905.12888v22019
  31. Koopman-Based Robust Model Predictive Control for Nonlinear Systems with Stochastic Intermittent Measurements

    Guanhua Liu, Tong Wu, Lixian Zhang +2

    cs.ROeess.SYarXiv:2609.02079v12026
  32. Patchwork++: Fast and Robust Ground Segmentation Solving Partial Under-Segmentation Using 3D Point Cloud

    Seungjae Lee, Hyungtae Lim, Hyun Myung

    cs.ROcs.CVarXiv:2207.11919v22022
  33. VLA-Cache: Efficient Vision-Language-Action Manipulation via Adaptive Token Caching

    Siyu Xu, Yunke Wang, Chenghao Xia +3

    cs.ROcs.CVcs.LGarXiv:2502.02175v22025
  34. Track Any Motions under Any Disturbances

    Zhikai Zhang, Jun Guo, Chao Chen +10

    cs.ROarXiv:2509.13833v32025
  35. CogVLA: Cognition-Aligned Vision-Language-Action Model via Instruction-Driven Routing & Sparsification

    Wei Li, Renshan Zhang, Rui Shao +2

    cs.CVcs.ROarXiv:2508.21046v32025
  36. Don't Shake the Wheel: Momentum-Aware Planning in End-to-End Autonomous Driving

    Ziying Song, Caiyan Jia, Lin Liu +7

    cs.ROarXiv:2503.03125v32025
  37. RL-100: Performant Robotic Manipulation with Real-World Reinforcement Learning

    Kun Lei, Huanyu Li, Dongjie Yu +6

    cs.ROcs.AIcs.LGarXiv:2510.14830v42025
  38. GPU-Accelerated Astrodynamics World Models for Spacecraft Rendezvous and Proximity Operations

    Duncan Eddy, Isaac R. Ward, Grace Ra Kim +1

    cs.ROeess.SYarXiv:2609.03067v12026
  39. LiDAR-Camera Calibration using 3D-3D Point correspondences

    Ankit Dhall, Kunal Chelani, Vishnu Radhakrishnan +1

    cs.ROcs.CVarXiv:1705.09785v12017
  40. End-to-End Race Driving with Deep Reinforcement Learning

    Maximilian Jaritz, Raoul de Charette, Marin Toromanoff +2

    cs.CVcs.ROarXiv:1807.02371v22018
  41. VideoVLA: Video Generators Can Be Generalizable Robot Manipulators

    Yichao Shen, Fangyun Wei, Zhiying Du +5

    cs.ROcs.AIcs.CVarXiv:2512.06963v12025
  42. From Pixels to Torques: Policy Learning with Deep Dynamical Models

    Niklas Wahlström, Thomas B. Schön, Marc Peter Deisenroth

    stat.MLcs.LGcs.ROarXiv:1502.02251v32015
  43. ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver

    Wenxuan Song, Ziyang Zhou, Han Zhao +7

    cs.ROcs.CVarXiv:2508.10333v12025
  44. Spatially Aware World Action Model via Geometric Latent Diffusion

    Javier Alejandro Lopetegui Gonzalez, Paul Pacaud, Cordelia Schmid

    cs.CVcs.ROarXiv:2609.02531v12026
  45. Soft robotic suits: State of the art, core technologies and open challenges

    Michele Xiloyannis, Ryan Alicea, Anna-Maria Georgarakis +4

    cs.ROeess.SYarXiv:2105.10588v22021
  46. VLAW: Iterative Co-Improvement of Vision-Language-Action Policy and World Model

    Yanjiang Guo, Tony Lee, Lucy Xiaoyang Shi +3

    cs.ROarXiv:2602.12063v22026
  47. Do Better Imagined Rollouts Mean Better Robot Control? A Controlled Study of World-Model Evaluation Under Feedback

    Dharini Raghavan, Amritpal Singh

    cs.ROarXiv:2609.02811v12026
  48. From Proxy Learning to Driving Decisions: A Transfer-Based Framework for Evaluating Future-Aware Autonomous Driving Planners

    Yikai Wu

    cs.ROarXiv:2609.02688v12026
  49. DriveAdapter: Breaking the Coupling Barrier of Perception and Planning in End-to-End Autonomous Driving

    Xiaosong Jia, Yulu Gao, Li Chen +3

    cs.ROcs.CVarXiv:2308.00398v22023
  50. Embodied Navigation Foundation Model

    Jiazhao Zhang, Anqi Li, Yunpeng Qi +14

    cs.ROarXiv:2509.12129v22025
  51. Adversarial Motion Priors Make Good Substitutes for Complex Reward Functions

    Alejandro Escontrela, Xue Bin Peng, Wenhao Yu +4

    cs.AIcs.ROarXiv:2203.15103v12022
  52. Learning Language-Conditioned Robot Behavior from Offline Data and Crowd-Sourced Annotation

    Suraj Nair, Eric Mitchell, Kevin Chen +3

    cs.ROcs.AIcs.LGarXiv:2109.01115v22021
  53. Recent Advances in Imitation Learning from Observation

    Faraz Torabi, Garrett Warnell, Peter Stone

    cs.ROcs.AIcs.LGarXiv:1905.13566v22019
  54. InterMimic: Towards Universal Whole-Body Control for Physics-Based Human-Object Interactions

    Sirui Xu, Hung Yu Ling, Yu-Xiong Wang +1

    cs.CVcs.GRcs.ROarXiv:2502.20390v22025
  55. Universal Actions for Enhanced Embodied Foundation Models

    Jinliang Zheng, Jianxiong Li, Dongxiu Liu +7

    cs.ROcs.AIcs.CVarXiv:2501.10105v22025
  56. RynnVLA-002: A Unified Vision-Language-Action and World Model

    Jun Cen, Siteng Huang, Yuqian Yuan +11

    cs.ROarXiv:2511.17502v32025
  57. F1: A Vision-Language-Action Model Bridging Understanding and Generation to Actions

    Qi Lv, Weijie Kong, Hao Li +7

    cs.ROcs.CVarXiv:2509.06951v22025
  58. RSL-RL: A Learning Library for Robotics Research

    Clemens Schwarke, Mayank Mittal, Nikita Rudin +2

    cs.ROcs.LGarXiv:2509.10771v12025
  59. Geometric analysis of generic 3R robots, and necessary and sufficient conditions for a class of orthogonal robots to have four IKS

    Durgesh Haribhau Salunkhe, Abhilash Nayak

    cs.ROarXiv:2609.00316v12026
  60. Dita: Scaling Diffusion Transformer for Generalist Vision-Language-Action Policy

    Zhi Hou, Tianyi Zhang, Yuwen Xiong +8

    cs.ROcs.CVarXiv:2503.19757v22025