Robotics
Papers filed under cs.RO on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
1,081 to 1,140 of 3,243
ApexNav: An Adaptive Exploration Strategy for Zero-Shot Object Navigation with Target-centric Semantic Fusion
Mingjie Zhang, Yuheng Du, Chengkai Wu +4
cs.ROarXiv:2504.14478v32025A Survey on Vision-Language-Action Models for Autonomous Driving
Sicong Jiang, Zilin Huang, Kangan Qian +17
cs.CVcs.AIcs.ROarXiv:2506.24044v12025Motubrain: An Advanced World Action Model for Robot Control
Motubrain Team, Chendong Xiang, Fan Bao +17
cs.ROarXiv:2604.27792v52026UAVs Meet LLMs: Overviews and Perspectives Toward Agentic Low-Altitude Mobility
Yonglin Tian, Fei Lin, Yiduo Li +11
cs.ROcs.AIarXiv:2501.02341v22025Humanoid Policy ~ Human Policy
Ri-Zhao Qiu, Shiqi Yang, Xuxin Cheng +12
cs.ROcs.AIcs.CVarXiv:2503.13441v32025Deep-Learned Collision Avoidance Policy for Distributed Multi-Agent Navigation
Pinxin Long, Wenxi Liu, Jia Pan
cs.AIcs.CVcs.ROarXiv:1609.06838v22016MoLe-VLA: Dynamic Layer-skipping Vision Language Action Model via Mixture-of-Layers for Efficient Robot Manipulation
Rongyu Zhang, Menghang Dong, Yuan Zhang +6
cs.ROcs.AIarXiv:2503.20384v22025Benchmarking Reinforcement Learning Algorithms on Real-World Robots
A. Rupam Mahmood, Dmytro Korenkevych, Gautham Vasan +2
cs.LGcs.AIcs.ROarXiv:1809.07731v12018Adaptive Depth-Map-Guided Bundle Adjustment for Correspondence-Free Multi-View Point Cloud Registration
Yiran Zhou, Yingyu Wang, Shoudong Huang +1
cs.ROarXiv:2609.01089v12026FLOWER: Democratizing Generalist Robot Policies with Efficient Vision-Language-Action Flow Policies
Moritz Reuss, Hongyi Zhou, Marcel Rühle +3
cs.ROarXiv:2509.04996v12025ChatVLA-2: Vision-Language-Action Model with Open-World Embodied Reasoning from Pretrained Knowledge
Zhongyi Zhou, Yichen Zhu, Junjie Wen +2
cs.ROcs.AIcs.CVarXiv:2505.21906v22025P3Depth: Monocular Depth Estimation with a Piecewise Planarity Prior
Vaishakh Patil, Christos Sakaridis, Alexander Liniger +1
cs.CVcs.AIcs.LGarXiv:2204.02091v12022Improving Vision-Language-Action Model with Online Reinforcement Learning
Yanjiang Guo, Jianke Zhang, Xiaoyu Chen +4
cs.ROcs.CVcs.LGarXiv:2501.16664v12025WoVR: World Models as Reliable Simulators for Post-Training VLA Policies with RL
Zhennan Jiang, Shangqing Zhou, Yutong Jiang +11
cs.ROcs.AIarXiv:2602.13977v22026A soft robot that adapts to environments through shape change
Dylan S. Shah, Joshua P. Powers, Liana G. Tilton +3
cs.ROarXiv:2008.06397v52020LDA-1B: Scaling Latent Dynamics Action Model via Universal Embodied Data Ingestion
Jiangran Lyu, Kai Liu, Xuheng Zhang +20
cs.ROarXiv:2602.12215v22026Beyond Textual Chain-of-Thought: A Survey on Action-Grounded Reasoning in Autonomous Driving
Zhengxu Tang, Xiaozhou Zhang, Guofeng Cui +10
cs.CVcs.CLcs.ROarXiv:2609.01659v12026ReinFlow: Fine-tuning Flow Matching Policy with Online Reinforcement Learning
Tonghe Zhang, Chao Yu, Sichang Su +1
cs.ROcs.LGarXiv:2505.22094v72025CoLT-Drive: Counterfactual Long-Tail Benchmarking and Knowledge-Preserving Adaptation for Driving Affordance Prediction
Zhengxu Tang, Guofeng Cui, Ziyu Gong +8
cs.CVcs.AIcs.CLarXiv:2609.00242v12026RAD: Training an End-to-End Driving Policy via Large-Scale 3DGS-based Reinforcement Learning
Hao Gao, Shaoyu Chen, Bo Jiang +11
cs.CVcs.ROarXiv:2502.13144v22025Humanoid Parkour Learning
Ziwen Zhuang, Shenzhe Yao, Hang Zhao
cs.ROarXiv:2406.10759v22024Neural probabilistic motor primitives for humanoid control
Josh Merel, Leonard Hasenclever, Alexandre Galashov +5
cs.LGcs.AIcs.ROarXiv:1811.11711v22018Variable Compliance Control for Robotic Peg-in-Hole Assembly: A Deep Reinforcement Learning Approach
Cristian C. Beltran-Hernandez, Damien Petit, Ixchel G. Ramirez-Alpizar +1
cs.ROcs.LGarXiv:2008.10224v32020DTC: Deep Tracking Control
Fabian Jenelten, Junzhe He, Farbod Farshidian +1
cs.ROcs.LGeess.SYarXiv:2309.15462v22023InternVLA-A1: Unifying Understanding, Generation and Action for Robotic Manipulation
Junhao Cai, Zetao Cai, Jiafei Cao +39
cs.ROarXiv:2601.02456v22026Data-Driven Model Predictive Control of Autonomous Mobility-on-Demand Systems
Ramon Iglesias, Federico Rossi, Kevin Wang +3
cs.ROcs.MAeess.SYarXiv:1709.07032v12017FASTER: Fast and Safe Trajectory Planner for Flights in Unknown Environments
Jesus Tordesillas, Brett T. Lopez, Jonathan P. How
cs.ROarXiv:1903.03558v32019Video Generators are Robot Policies
Junbang Liang, Pavel Tokmakov, Ruoshi Liu +4
cs.ROarXiv:2508.00795v12025Finding a needle in an exponential haystack: Discrete RRT for exploration of implicit roadmaps in multi-robot motion planning
Kiril Solovey, Oren Salzman, Dan Halperin
cs.ROarXiv:1305.2889v32013Imitation Learning as $f$-Divergence Minimization
Liyiming Ke, Sanjiban Choudhury, Matt Barnes +3
cs.LGcs.ITcs.ROarXiv:1905.12888v22019Koopman-Based Robust Model Predictive Control for Nonlinear Systems with Stochastic Intermittent Measurements
Guanhua Liu, Tong Wu, Lixian Zhang +2
cs.ROeess.SYarXiv:2609.02079v12026Patchwork++: Fast and Robust Ground Segmentation Solving Partial Under-Segmentation Using 3D Point Cloud
Seungjae Lee, Hyungtae Lim, Hyun Myung
cs.ROcs.CVarXiv:2207.11919v22022VLA-Cache: Efficient Vision-Language-Action Manipulation via Adaptive Token Caching
Siyu Xu, Yunke Wang, Chenghao Xia +3
cs.ROcs.CVcs.LGarXiv:2502.02175v22025Track Any Motions under Any Disturbances
Zhikai Zhang, Jun Guo, Chao Chen +10
cs.ROarXiv:2509.13833v32025CogVLA: Cognition-Aligned Vision-Language-Action Model via Instruction-Driven Routing & Sparsification
Wei Li, Renshan Zhang, Rui Shao +2
cs.CVcs.ROarXiv:2508.21046v32025Don't Shake the Wheel: Momentum-Aware Planning in End-to-End Autonomous Driving
Ziying Song, Caiyan Jia, Lin Liu +7
cs.ROarXiv:2503.03125v32025RL-100: Performant Robotic Manipulation with Real-World Reinforcement Learning
Kun Lei, Huanyu Li, Dongjie Yu +6
cs.ROcs.AIcs.LGarXiv:2510.14830v42025GPU-Accelerated Astrodynamics World Models for Spacecraft Rendezvous and Proximity Operations
Duncan Eddy, Isaac R. Ward, Grace Ra Kim +1
cs.ROeess.SYarXiv:2609.03067v12026LiDAR-Camera Calibration using 3D-3D Point correspondences
Ankit Dhall, Kunal Chelani, Vishnu Radhakrishnan +1
cs.ROcs.CVarXiv:1705.09785v12017End-to-End Race Driving with Deep Reinforcement Learning
Maximilian Jaritz, Raoul de Charette, Marin Toromanoff +2
cs.CVcs.ROarXiv:1807.02371v22018VideoVLA: Video Generators Can Be Generalizable Robot Manipulators
Yichao Shen, Fangyun Wei, Zhiying Du +5
cs.ROcs.AIcs.CVarXiv:2512.06963v12025From Pixels to Torques: Policy Learning with Deep Dynamical Models
Niklas Wahlström, Thomas B. Schön, Marc Peter Deisenroth
stat.MLcs.LGcs.ROarXiv:1502.02251v32015ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver
Wenxuan Song, Ziyang Zhou, Han Zhao +7
cs.ROcs.CVarXiv:2508.10333v12025Spatially Aware World Action Model via Geometric Latent Diffusion
Javier Alejandro Lopetegui Gonzalez, Paul Pacaud, Cordelia Schmid
cs.CVcs.ROarXiv:2609.02531v12026Soft robotic suits: State of the art, core technologies and open challenges
Michele Xiloyannis, Ryan Alicea, Anna-Maria Georgarakis +4
cs.ROeess.SYarXiv:2105.10588v22021VLAW: Iterative Co-Improvement of Vision-Language-Action Policy and World Model
Yanjiang Guo, Tony Lee, Lucy Xiaoyang Shi +3
cs.ROarXiv:2602.12063v22026Do Better Imagined Rollouts Mean Better Robot Control? A Controlled Study of World-Model Evaluation Under Feedback
Dharini Raghavan, Amritpal Singh
cs.ROarXiv:2609.02811v12026From Proxy Learning to Driving Decisions: A Transfer-Based Framework for Evaluating Future-Aware Autonomous Driving Planners
Yikai Wu
cs.ROarXiv:2609.02688v12026DriveAdapter: Breaking the Coupling Barrier of Perception and Planning in End-to-End Autonomous Driving
Xiaosong Jia, Yulu Gao, Li Chen +3
cs.ROcs.CVarXiv:2308.00398v22023Embodied Navigation Foundation Model
Jiazhao Zhang, Anqi Li, Yunpeng Qi +14
cs.ROarXiv:2509.12129v22025Adversarial Motion Priors Make Good Substitutes for Complex Reward Functions
Alejandro Escontrela, Xue Bin Peng, Wenhao Yu +4
cs.AIcs.ROarXiv:2203.15103v12022Learning Language-Conditioned Robot Behavior from Offline Data and Crowd-Sourced Annotation
Suraj Nair, Eric Mitchell, Kevin Chen +3
cs.ROcs.AIcs.LGarXiv:2109.01115v22021Recent Advances in Imitation Learning from Observation
Faraz Torabi, Garrett Warnell, Peter Stone
cs.ROcs.AIcs.LGarXiv:1905.13566v22019InterMimic: Towards Universal Whole-Body Control for Physics-Based Human-Object Interactions
Sirui Xu, Hung Yu Ling, Yu-Xiong Wang +1
cs.CVcs.GRcs.ROarXiv:2502.20390v22025Universal Actions for Enhanced Embodied Foundation Models
Jinliang Zheng, Jianxiong Li, Dongxiu Liu +7
cs.ROcs.AIcs.CVarXiv:2501.10105v22025RynnVLA-002: A Unified Vision-Language-Action and World Model
Jun Cen, Siteng Huang, Yuqian Yuan +11
cs.ROarXiv:2511.17502v32025F1: A Vision-Language-Action Model Bridging Understanding and Generation to Actions
Qi Lv, Weijie Kong, Hao Li +7
cs.ROcs.CVarXiv:2509.06951v22025RSL-RL: A Learning Library for Robotics Research
Clemens Schwarke, Mayank Mittal, Nikita Rudin +2
cs.ROcs.LGarXiv:2509.10771v12025Geometric analysis of generic 3R robots, and necessary and sufficient conditions for a class of orthogonal robots to have four IKS
Durgesh Haribhau Salunkhe, Abhilash Nayak
cs.ROarXiv:2609.00316v12026Dita: Scaling Diffusion Transformer for Generalist Vision-Language-Action Policy
Zhi Hou, Tianyi Zhang, Yuwen Xiong +8
cs.ROcs.CVarXiv:2503.19757v22025