Robotics

Papers filed under cs.RO on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

2,641 to 2,700 of 3,242

  1. Learning Agile Robotic Locomotion Skills by Imitating Animals

    Xue Bin Peng, Erwin Coumans, Tingnan Zhang +3

    cs.ROcs.LGarXiv:2004.00784v32020
  2. SplaTAM: Splat, Track & Map 3D Gaussians for Dense RGB-D SLAM

    Nikhil Keetha, Jay Karhade, Krishna Murthy Jatavallabhula +4

    cs.CVcs.AIcs.ROarXiv:2312.02126v32023
  3. Contact-Anchored Proprioceptive Odometry for Legged and Wheel-Legged Robots

    Minxing Sun, Yao Mao

    cs.ROeess.SParXiv:2602.17393v32026
  4. HERO: Learning Humanoid End-Effector Control for Visual Whole-Body Open-Vocabulary Object Grasping

    Runpei Dong, Ziyan Li, Arjun Gupta +2

    cs.ROcs.CVarXiv:2602.16705v32026
  5. Bridging Teacher Expectations and Robot Learning via Coupling Dynamics

    Evan Dallas, Sean Dallas, Wing-Yue Geoffrey Louie

    cs.ROcs.HCarXiv:2608.23994v12026
  6. CARE: Camera-Residual Reserves for First Sightings in Adaptive LiDAR Sensing

    Jiachen Gong, Yun Li, Ehsan Javanmardi +2

    cs.CVcs.ROarXiv:2608.24282v12026
  7. EgoPush: Learning End-to-End Egocentric Multi-Object Rearrangement for Mobile Robots

    Boyuan An, Zhexiong Wang, Yipeng Wang +4

    cs.ROarXiv:2602.18071v12026
  8. SIREN-Bench: Behavior-Driven Generation and Evaluation of Emergency-Vehicle Interactions

    Yicheng Zhu, Tianmu Zhao, Haoxin Leng +3

    cs.ROarXiv:2608.24094v12026
  9. Gripper-aware Vision Language Action Models

    Hanyi Zhang, Zihong Luo, Tianyu Li +16

    cs.ROarXiv:2608.24603v12026
  10. X-MULTI: VLM-based Imaging Factor Disentanglement for Factor-Aware Image Synthesis

    Sonali Godavarthy, Matthias Neuwirth-Trapp, Tim-Felix Faasch +4

    cs.CVcs.ROarXiv:2608.24563v12026
  11. One-Shot Learning from Demonstration of Contact-Rich Robotic Manipulation by Identifying Physical Interactions

    A. H. G. Overbeek, H. van der Kooij, M. Vlutters

    cs.ROarXiv:2608.24741v12026
  12. Challenges of Real-World Reinforcement Learning

    Gabriel Dulac-Arnold, Daniel Mankowitz, Todd Hester

    cs.LGcs.AIcs.ROarXiv:1904.12901v12019
  13. PointCLIP: Point Cloud Understanding by CLIP

    Renrui Zhang, Ziyu Guo, Wei Zhang +6

    cs.CVcs.AIcs.ROarXiv:2112.02413v12021
  14. Multimodal Deep Learning for Robust RGB-D Object Recognition

    Andreas Eitel, Jost Tobias Springenberg, Luciano Spinello +2

    cs.CVcs.LGcs.NEarXiv:1507.06821v22015
  15. Lightweight Visual Reasoning for Socially-Aware Robots

    Alessio Galatolo, Ronald Cumbal, Alexandros Rouchitsas +3

    cs.ROarXiv:2603.03942v12026
  16. TeamHOI: Learning a Unified Policy for Cooperative Human-Object Interactions with Any Team Size

    Stefan Lionar, Gim Hee Lee

    cs.CVcs.GRcs.MAarXiv:2603.07988v12026
  17. NeurRAFT: Robot Motion Planning via Anchor-Level Flow Matching with Clearance-Aware Preference Tuning

    Sibo Tian, Chang Liu, Minghui Zheng +1

    cs.ROarXiv:2608.24026v12026
  18. Design-to-Plan: A Large Language Model-Based Multi-Agent Framework for Manufacturing Process Planning from 3D CAD Models and 2D Engineering Drawings

    Muhammad Tayyab Khan, Lequn Chen, Wenhe Feng +1

    cs.ROcs.AIarXiv:2608.24039v12026
  19. Parameter Space Noise for Exploration

    Matthias Plappert, Rein Houthooft, Prafulla Dhariwal +6

    cs.LGcs.AIcs.NEarXiv:1706.01905v22017
  20. COT-FM: Cluster-wise Optimal Transport Flow Matching

    Chiensheng Chiang, Kuan-Hsun Tu, Jia-Wei Liao +2

    cs.CVcs.LGcs.ROarXiv:2603.13395v12026
  21. OPV2V: An Open Benchmark Dataset and Fusion Pipeline for Perception with Vehicle-to-Vehicle Communication

    Runsheng Xu, Hao Xiang, Xin Xia +3

    cs.CVcs.ROarXiv:2109.07644v52021
  22. CounterAlign: Counterfactual Supervision for Vision-Language-Action Models

    Haru Kondoh, Kei Ota, Asako Kanezaki +1

    cs.ROarXiv:2608.21740v12026
  23. Model-Free Adaptive Parameter Tuning for Efficient Multi-Robot Warehouse Operations

    Pratap Tokekar, Mouhacine Benosman, Rahul Chandan +3

    cs.ROeess.SYarXiv:2608.21533v12026
  24. How to Train Your Robot with Deep Reinforcement Learning; Lessons We've Learned

    Julian Ibarz, Jie Tan, Chelsea Finn +3

    cs.ROcs.LGarXiv:2102.02915v12021
  25. Pixel-level Scene Understanding in One Token: Visual States Need What-is-Where Composition

    Seokmin Lee, Yunghee Lee, Byeonghyun Pak +1

    cs.CVcs.AIcs.LGarXiv:2603.13904v22026
  26. Robust Bimanual Vision-Language-Action Models via Embarrassingly Simple Modality Masking

    Dongzhou Cheng, Ziang Li, Yixiao Zhou +6

    cs.ROcs.CVarXiv:2608.22419v12026
  27. GOLEM: Modular Humanoid Autonomy Towards Electric Vehicle Battery Disassembly

    Max Conway, William Xie, Allen Devaraj +14

    cs.ROarXiv:2608.21550v12026
  28. The Imitator Game: Benchmarking Robot Imitative Ability Beyond Action Prediction

    Xunzhe Zhou, Yiyang Cai, Fengyi Wang +9

    cs.ROcs.AIarXiv:2608.22301v12026
  29. OpenSCvx: An Open-Source Modular and Extensible Nonlinear Trajectory Planning Package

    Christopher R. Hayner, Griffin J. Norris, Fabio Spada +4

    cs.ROmath.OCarXiv:2608.21631v12026
  30. Event-Based Motion Estimation via Oriented Distance Fields

    Lei Sun, Yuqin Ma, Weilun Li +5

    cs.CVcs.ROarXiv:2608.24223v12026
  31. Robotic Pick-and-Place of Novel Objects in Clutter with Multi-Affordance Grasping and Cross-Domain Image Matching

    Andy Zeng, Shuran Song, Kuan-Ting Yu +18

    cs.ROcs.CVarXiv:1710.01330v52017
  32. 6-DOF GraspNet: Variational Grasp Generation for Object Manipulation

    Arsalan Mousavian, Clemens Eppner, Dieter Fox

    cs.CVcs.ROarXiv:1905.10520v22019
  33. Energy-Aware Performance Evaluation of Nonlinear Mechatronic Systems Under Matched-Tracking Conditions

    Bhanuka Dayawansa

    eess.SYcs.ROarXiv:2608.23578v12026
  34. Safety-aware Model Predictive Path Integral Control with Signal Temporal Logic

    Yiqi Zhao, Taekyung Kim, Hideki Okamoto +4

    cs.ROeess.SYarXiv:2608.23972v12026
  35. Pattern-Derived Visual Swarm Games: Multi-Scale Drone-Vision States for Interception and Sustainability Audits

    Faruk Alpay, Levent Sarioglu

    cs.ROcs.GTeess.IVarXiv:2608.23575v12026
  36. Spinning Quadrotor: Hover Thrust Augmentation with Passive Lifting Surfaces

    Aniketh Parkala, Harikumar Kandath

    cs.ROarXiv:2608.23163v12026
  37. CALVIN: A Benchmark for Language-Conditioned Policy Learning for Long-Horizon Robot Manipulation Tasks

    Oier Mees, Lukas Hermann, Erick Rosete-Beas +1

    cs.ROcs.AIcs.CLarXiv:2112.03227v42021
  38. Learning by Cheating

    Dian Chen, Brady Zhou, Vladlen Koltun +1

    cs.ROcs.AIcs.CVarXiv:1912.12294v12019
  39. On the Optimized Use of Non-Orthonormality Constraints for the Quasi-Static INS Alignment of Autonomous Underwater and Surface Vehicles

    Carlos Renato C. Durao, Felipe O. Silva, Itzik Klein +4

    cs.ROeess.SYarXiv:2608.21390v12026
  40. Eureka: Human-Level Reward Design via Coding Large Language Models

    Yecheng Jason Ma, William Liang, Guanzhi Wang +6

    cs.ROcs.AIcs.LGarXiv:2310.12931v22023
  41. RoboShape: Information-Theoretic Point Cloud Representations for Privacy-Aware Robot Perception

    Oguzhan Baser, Mirac Sozen, Kaan Kale +2

    cs.ROcs.AIcs.CVarXiv:2608.21380v12026
  42. Towards insect-like distributed proprioception in actuators and appendages for flapping-wing insect-scale aerial robots

    Alexander Hedrick, Arvind Gupta, Kaushik Jayaram

    cs.ROarXiv:2608.21699v12026
  43. Gimbal-Based Human Tracking for Companion Robots Using Continual Learning

    Cong-Thanh Vu, Ching-Chieh Liu, Yen-Chen Liu

    cs.ROeess.SYarXiv:2608.21388v12026
  44. Multimodal-Language-Model-Driven Interaction and Companionship for Service Robots in Elderly-Care Facilities

    Ching-Chieh Liu, Cong-Thanh Vu, Yen-Chen Liu

    cs.ROeess.SYarXiv:2608.21387v12026
  45. robosuite: A Modular Simulation Framework and Benchmark for Robot Learning

    Yuke Zhu, Josiah Wong, Ajay Mandlekar +6

    cs.ROcs.AIcs.LGarXiv:2009.12293v32020
  46. CMX: Cross-Modal Fusion for RGB-X Semantic Segmentation with Transformers

    Jiaming Zhang, Huayao Liu, Kailun Yang +3

    cs.CVcs.ROeess.IVarXiv:2203.04838v52022
  47. IVRA: Improving Visual-Token Relations for Robot Action Policy with Training-Free Hint-Based Guidance

    Jongwoo Park, Kanchana Ranasinghe, Jinhyeok Jang +3

    cs.ROarXiv:2601.16207v22026
  48. FoundationPose: Unified 6D Pose Estimation and Tracking of Novel Objects

    Bowen Wen, Wei Yang, Jan Kautz +1

    cs.CVcs.AIcs.ROarXiv:2312.08344v22023
  49. DELTA: Deformable Elevation-Based Local Terrain Attention Encoder for Sparse-Terrain Quadrupedal Locomotion

    Sanghyun Park, Moonkyu Jung, Jemin Hwangbo

    cs.ROarXiv:2608.22033v12026
  50. FARE: Fast-Slow Agentic Robotic Exploration

    Shuhao Liao, Xuxin Lv, Jeric Lew +6

    cs.ROarXiv:2601.14681v12026
  51. STORM: Slot-based Task-aware Object-centric Representation for robotic Manipulation

    Alexandre Chapin, Emmanuel Dellandréa, Liming Chen

    cs.ROarXiv:2601.20381v32026
  52. ChatGPT for Robotics: Design Principles and Model Abilities

    Sai Vemprala, Rogerio Bonatti, Arthur Bucker +1

    cs.AIcs.CLcs.HCarXiv:2306.17582v22023
  53. LD4WAM: Learning Latent Dynamics from Human Videos for World Action Models

    Zhenhao Shen, Jiaqi Liang, Jasper Lu +11

    cs.ROarXiv:2608.22403v12026
  54. MotionDLO: Hybrid Event- and Frame-Based Tracking of Deformable Linear Objects

    Annalena Hartmann, Priyamvada Ajithkumar, Patrick Bründl +1

    cs.ROcs.CVarXiv:2608.22398v12026
  55. A Unified Neural-Aided Alignment and Calibration Method for AUVs

    Guy Damari, Zeev Yampolsky, Itzik Klein

    cs.ROarXiv:2608.22496v12026
  56. ST-BiBench: Benchmarking Multi-Stream Multimodal Coordination in Bimanual Embodied Tasks for MLLMs

    Xin Wu, Zhixuan Liang, Yue Ma +3

    cs.ROcs.AIcs.CVarXiv:2602.08392v22026
  57. One-2-3-45: Any Single Image to 3D Mesh in 45 Seconds without Per-Shape Optimization

    Minghua Liu, Chao Xu, Haian Jin +4

    cs.CVcs.AIcs.ROarXiv:2306.16928v12023
  58. Social Attention: Modeling Attention in Human Crowds

    Anirudh Vemula, Katharina Muelling, Jean Oh

    cs.ROcs.LGarXiv:1710.04689v22017
  59. DIGIT: A Novel Design for a Low-Cost Compact High-Resolution Tactile Sensor with Application to In-Hand Manipulation

    Mike Lambeta, Po-Wei Chou, Stephen Tian +9

    cs.ROcs.LGeess.SYarXiv:2005.14679v12020
  60. Using Simulation and Domain Adaptation to Improve Efficiency of Deep Robotic Grasping

    Konstantinos Bousmalis, Alex Irpan, Paul Wohlhart +9

    cs.LGcs.AIcs.CVarXiv:1709.07857v22017