Robotics

Papers filed under cs.RO on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

1,441 to 1,500 of 3,243

  1. $χ_{0}$: Resource-Aware Robust Manipulation via Taming Distributional Inconsistencies

    Checheng Yu, Chonghao Sima, Gangcheng Jiang +14

    cs.ROcs.CVarXiv:2602.09021v32026
  2. AnyTeleop: A General Vision-Based Dexterous Robot Arm-Hand Teleoperation System

    Yuzhe Qin, Wei Yang, Binghao Huang +5

    cs.ROcs.CVcs.LGarXiv:2307.04577v32023
  3. Goal-Conditioned Imitation Learning using Score-based Diffusion Policies

    Moritz Reuss, Maximilian Li, Xiaogang Jia +1

    cs.LGcs.ROarXiv:2304.02532v22023
  4. TACTO: A Fast, Flexible, and Open-source Simulator for High-Resolution Vision-based Tactile Sensors

    Shaoxiong Wang, Mike Lambeta, Po-Wei Chou +1

    cs.ROcs.LGstat.MLarXiv:2012.08456v22020
  5. Meta-World: A Benchmark and Evaluation for Multi-Task and Meta Reinforcement Learning

    Tianhe Yu, Deirdre Quillen, Zhanpeng He +7

    cs.LGcs.AIcs.ROarXiv:1910.10897v22019
  6. Soft Actor-Critic Algorithms and Applications

    Tuomas Haarnoja, Aurick Zhou, Kristian Hartikainen +8

    cs.LGcs.AIcs.ROarXiv:1812.05905v22018
  7. Quaternion kinematics for the error-state Kalman filter

    Joan Solà

    cs.ROarXiv:1711.02508v12017
  8. SE-Sync: A Certifiably Correct Algorithm for Synchronization over the Special Euclidean Group

    David M. Rosen, Luca Carlone, Afonso S. Bandeira +1

    cs.ROarXiv:1612.07386v22016
  9. Distributed Machine Learning in Materials that Couple Sensing, Actuation, Computation and Communication

    Dana Hughes, Nikolaus Correll

    cs.LGcs.ROarXiv:1606.03508v12016
  10. Integrating Generic Sensor Fusion Algorithms with Sound State Representations through Encapsulation of Manifolds

    Christoph Hertzberg, René Wagner, Udo Frese +1

    cs.ROcs.CVcs.MSarXiv:1107.1119v12011
  11. Deep Reinforcement Learning for the Control of Robotic Manipulation: A Focussed Mini-Review

    Rongrong Liu, Florent Nageotte, Philippe Zanne +2

    cs.ROarXiv:2102.04148v12021
  12. HitMem: Hierarchical Temporal 3D Memory with Multi-Modal Context-Aware Retrieval for Dynamic Environments

    Ruijie Tang, Chenye Zou, Guoquan Wu +3

    cs.ROarXiv:2609.00950v12026
  13. MemoryVLA: Perceptual-Cognitive Memory in Vision-Language-Action Models for Robotic Manipulation

    Hao Shi, Bin Xie, Yingfei Liu +7

    cs.ROcs.CVarXiv:2508.19236v22025
  14. Hi Robot: Open-Ended Instruction Following with Hierarchical Vision-Language-Action Models

    Lucy Xiaoyang Shi, Brian Ichter, Michael Equi +12

    cs.ROcs.AIcs.LGarXiv:2502.19417v22025
  15. MolmoAct: Action Reasoning Models that can Reason in Space

    Jason Lee, Jiafei Duan, Haoquan Fang +16

    cs.ROarXiv:2508.07917v42025
  16. ASAP: Aligning Simulation and Real-World Physics for Learning Agile Humanoid Whole-Body Skills

    Tairan He, Jiawei Gao, Wenli Xiao +15

    cs.ROcs.AIcs.LGarXiv:2502.01143v32025
  17. DreamVLA: A Vision-Language-Action Model Dreamed with Comprehensive World Knowledge

    Wenyao Zhang, Hongsi Liu, Zekun Qi +11

    cs.CVcs.ROarXiv:2507.04447v32025
  18. A Collaborative Multi-Modality Interaction for VLA-based End-to-End Autonomous Driving

    Jingtao Sun, Xiaohai He, Yike Zhang +4

    cs.CVcs.ROarXiv:2608.20890v12026
  19. Closing the Affective Loop: Multimodal Speaker-Listener Emotion-Dynamics-Aware Empathetic Social Robots

    Zi Haur Pang, Casey Kennington, Tatsuya Kawahara

    cs.HCcs.CLcs.ROarXiv:2608.16686v12026
  20. Temporal Logic Guided Universal Task Representations for Reinforcement Learning

    Hao Zhang, Zhangli Zhou, Zhen Kan

    cs.ROcs.FLcs.LGarXiv:2608.15509v12026
  21. Reward Machines for Signal Temporal Logic

    Alper Kamil Bozkurt, Shangtong Zhang, Yuichi Motai

    cs.AIcs.LGcs.ROarXiv:2608.13625v12026
  22. RynnValue: Scaling Robotic Value Foundation Models with Temporal Distance

    Dongchi Huang, Hongyin Zhang, Bohan Hou +12

    cs.ROcs.CVcs.LGarXiv:2608.09853v12026
  23. Learning Locomotion Skills Using DeepRL: Does the Choice of Action Space Matter?

    Xue Bin Peng, Michiel van de Panne

    cs.LGcs.GRcs.ROarXiv:1611.01055v12016
  24. Alpamayo-R1: Bridging Reasoning and Action Prediction for Generalizable Autonomous Driving in the Long Tail

    NVIDIA, :, Yan Wang +41

    cs.ROcs.AIcs.LGarXiv:2511.00088v22025
  25. EgoDex: Learning Dexterous Manipulation from Large-Scale Egocentric Video

    Ryan Hoque, Peide Huang, David J. Yoon +2

    cs.CVcs.LGcs.ROarXiv:2505.11709v32025
  26. BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation

    Peiyan Li, Yuze Zhu, Yixiang Chen +10

    cs.ROarXiv:2608.05042v12026
    Summaries:한국어
  27. Quo Vadis, World Modeling?

    Yu Yang, Xuemeng Yang, Licheng Wen +17

    cs.CVcs.AIcs.ROarXiv:2608.02713v12026
  28. Xiaomi-Robotics-1: Scaling Vision-Language-Action Models with over 100K Hours of Real-World Trajectories

    Xiaomi Robotics Team, Jun Guo, Piaopiao Jin +31

    cs.ROcs.CVarXiv:2607.15330v22026
    Summaries:한국어
  29. RoboTTT: Context Scaling for Robot Policies

    Yunfan Jiang, Yevgen Chebotar, Ruijie Zheng +8

    cs.ROcs.AIcs.LGarXiv:2607.15275v12026
  30. BadWAM: When World-Action Models Dream Right but Act Wrong

    Qi Li, Xingyi Yang, Xinchao Wang

    cs.LGcs.ROarXiv:2607.15207v12026
    Summaries:한국어
  31. EmbodiedSkills: A Unified Framework for Orchestrating, Training, and Deploying VLA Agents

    Wei Wang, Wenqiao Zhang, Yutong Lin +14

    cs.ROcs.AIarXiv:2609.01281v12026
  32. Ctrl-World: A Controllable Generative World Model for Robot Manipulation

    Yanjiang Guo, Lucy Xiaoyang Shi, Jianyu Chen +1

    cs.ROcs.AIarXiv:2510.10125v32025
  33. VLA-Corrector: Lightweight Detect-and-Correct Inference for Adaptive Action Horizon

    Yi Pan, Miao Pan, Qi Lu +8

    cs.ROarXiv:2607.01804v12026
  34. World Value Models for Robotic Manipulation

    Zhihao Wang, Jianxiong Li, Yu Cui +4

    cs.ROarXiv:2606.24742v12026
  35. SONIC: Supersizing Motion Tracking for Natural Humanoid Whole-Body Control

    Zhengyi Luo, Ye Yuan, Tingwu Wang +25

    cs.ROcs.AIcs.CVarXiv:2511.07820v42025
  36. AffordanceVLA: A Vision-Language-Action Model Empowering Action Generation through Affordance-Aware Understanding

    Qize Yu, Jiadi You, Yuran Wang +10

    cs.ROcs.CVcs.MMarXiv:2606.06155v12026
  37. GRAIL: Generating Humanoid Loco-Manipulation from 3D Assets and Video Priors

    Tianyi Xie, Haotian Zhang, Jinhyung Park +17

    cs.ROarXiv:2606.05160v12026
  38. Text2Action: Generative Adversarial Synthesis from Language to Action

    Hyemin Ahn, Timothy Ha, Yunho Choi +2

    cs.LGcs.CLcs.ROarXiv:1710.05298v22017
  39. Learning from Trials and Errors: Reflective Test-Time Planning for Embodied LLMs

    Yining Hong, Huang Huang, Manling Li +4

    cs.LGcs.AIcs.CLarXiv:2602.21198v32026
  40. ECO: Energy-Constrained Optimization with Reinforcement Learning for Humanoid Walking

    Weidong Huang, Jingwen Zhang, Jiongye Li +6

    cs.ROarXiv:2602.06445v12026
  41. Diffusion Policy Policy Optimization

    Allen Z. Ren, Justin Lidard, Lars L. Ankile +6

    cs.ROcs.LGarXiv:2409.00588v32024
  42. Radar-Camera Fusion for Object Detection and Semantic Segmentation in Autonomous Driving: A Comprehensive Review

    Shanliang Yao, Runwei Guan, Xiaoyu Huang +8

    cs.CVcs.AIcs.ROarXiv:2304.10410v22023
  43. Perceptive Locomotion through Nonlinear Model Predictive Control

    Ruben Grandia, Fabian Jenelten, Shaohui Yang +2

    cs.ROarXiv:2208.08373v12022
  44. ProgPrompt: Generating Situated Robot Task Plans using Large Language Models

    Ishika Singh, Valts Blukis, Arsalan Mousavian +6

    cs.ROcs.AIcs.CLarXiv:2209.11302v12022
  45. Safe Control with Learned Certificates: A Survey of Neural Lyapunov, Barrier, and Contraction methods

    Charles Dawson, Sicun Gao, Chuchu Fan

    cs.ROeess.SYarXiv:2202.11762v22022
  46. PandaSet: Advanced Sensor Suite Dataset for Autonomous Driving

    Pengchuan Xiao, Zhenlei Shao, Steven Hao +9

    cs.CVcs.ROarXiv:2112.12610v12021
  47. FAST-LIO2: Fast Direct LiDAR-inertial Odometry

    Wei Xu, Yixi Cai, Dongjiao He +2

    cs.ROarXiv:2107.06829v12021
  48. OverlapNet: Loop Closing for LiDAR-based SLAM

    Xieyuanli Chen, Thomas Läbe, Andres Milioto +5

    cs.ROarXiv:2105.11344v12021
  49. SuMa++: Efficient LiDAR-based Semantic SLAM

    Xieyuanli Chen, Andres Milioto, Emanuele Palazzolo +3

    cs.ROarXiv:2105.11320v12021
  50. RADIATE: A Radar Dataset for Automotive Perception in Bad Weather

    Marcel Sheeny, Emanuele De Pellegrin, Saptarshi Mukherjee +3

    cs.CVcs.ROarXiv:2010.09076v32020
  51. Pre-Lane-change Signal in Transitional Autonomous Vehicles: Results from Controlled Experiments

    Zeyu Mu, Danjue Chen, Abhinav Sharma +1

    cs.ROeess.SYarXiv:2609.02575v12026
  52. TLIO: Tight Learned Inertial Odometry

    Wenxin Liu, David Caruso, Eddy Ilg +5

    cs.ROcs.CVcs.LGarXiv:2007.01867v32020
  53. TEASER: Fast and Certifiable Point Cloud Registration

    Heng Yang, Jingnan Shi, Luca Carlone

    cs.ROcs.CVmath.OCarXiv:2001.07715v22020
  54. Grasping in the Wild:Learning 6DoF Closed-Loop Grasping from Low-Cost Demonstrations

    Shuran Song, Andy Zeng, Johnny Lee +1

    cs.CVcs.ROarXiv:1912.04344v22019
  55. Contact-Aided Invariant Extended Kalman Filtering for Robot State Estimation

    Ross Hartley, Maani Ghaffari, Ryan M. Eustice +1

    cs.ROarXiv:1904.09251v22019
  56. Neural Lander: Stable Drone Landing Control using Learned Dynamics

    Guanya Shi, Xichen Shi, Michael O'Connell +5

    cs.ROcs.LGarXiv:1811.08027v22018
  57. Unmanned Aerial Vehicles: A Survey on Civil Applications and Key Research Challenges

    Hazim Shakhatreh, Ahmad Sawalmeh, Ala Al-Fuqaha +6

    cs.ROarXiv:1805.00881v12018
  58. VLocNet++: Deep Multitask Learning for Semantic Visual Localization and Odometry

    Noha Radwan, Abhinav Valada, Wolfram Burgard

    cs.ROcs.CVarXiv:1804.08366v62018
  59. Direct Sparse Visual-Inertial Odometry using Dynamic Marginalization

    Lukas von Stumberg, Vladyslav Usenko, Daniel Cremers

    cs.CVcs.ROarXiv:1804.05625v12018
  60. Deep Auxiliary Learning for Visual Localization and Odometry

    Abhinav Valada, Noha Radwan, Wolfram Burgard

    cs.ROcs.LGarXiv:1803.03642v12018