Robotics
Papers filed under cs.RO on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
1,321 to 1,380 of 3,243
Deep Drone Racing: From Simulation to Reality with Domain Randomization
Antonio Loquercio, Elia Kaufmann, René Ranftl +3
cs.ROarXiv:1905.09727v22019ORB-SLAM3: An Accurate Open-Source Library for Visual, Visual-Inertial and Multi-Map SLAM
Carlos Campos, Richard Elvira, Juan J. Gómez Rodríguez +2
cs.ROarXiv:2007.11898v22020Flow: A Modular Learning Framework for Mixed Autonomy Traffic
Cathy Wu, Aboudy Kreidieh, Kanaad Parvate +2
cs.AIcs.ROeess.SYarXiv:1710.05465v42017Data Scaling Laws in Imitation Learning for Robotic Manipulation
Fanqi Lin, Yingdong Hu, Pingyue Sheng +3
cs.ROarXiv:2410.18647v42024RoboCasa365: A Large-Scale Simulation Framework for Training and Benchmarking Generalist Robots
Soroush Nasiriany, Sepehr Nasiriany, Abhiram Maddukuri +1
cs.ROcs.AIcs.LGarXiv:2603.04356v12026TWIST2: Scalable, Portable, and Holistic Humanoid Data Collection System
Yanjie Ze, Siheng Zhao, Weizhuo Wang +6
cs.ROcs.CVcs.LGarXiv:2511.02832v12025Safe Reinforcement Learning with Model Uncertainty Estimates
Björn Lütjens, Michael Everett, Jonathan P. How
cs.ROcs.AIcs.LGarXiv:1810.08700v22018GuardianBench: A Same-Scene Instruction-Contrastive Benchmark for Latent Contextual Risk in Embodied AI
Zhesheng Zhang, Jiahao Lu, Wei Liu +8
cs.AIcs.CLcs.ROarXiv:2608.21928v12026FF-MPCC: High-speed Agile Formation Flight with Model Predictive Contouring Control
Aditya Dandwate, Vit Kratky, Parakh M. Gupta +2
cs.ROarXiv:2608.21056v12026Coverage Aware Active Evaluation for Failure Discovery with Paired Systems
Anjali Parashar, Rachel Luo, Apoorva Sharma +6
cs.AIcs.ROarXiv:2608.13719v12026StyleVLA: Driving Style-Aware Vision Language Action Model for Autonomous Driving
Yuan Gao, Dengyuan Hua, Mattia Piccinini +4
cs.ROarXiv:2603.09482v12026GeneralVLA: Generalizable Vision-Language-Action Models with Knowledge-Guided Trajectory Planning
Guoqing Ma, Siheng Wang, Zeyu Zhang +2
cs.ROcs.CVarXiv:2602.04315v12026FAST-LIVO2: Fast, Direct LiDAR-Inertial-Visual Odometry
Chunran Zheng, Wei Xu, Zuhao Zou +11
cs.ROcs.CVarXiv:2408.14035v22024Magma: A Foundation Model for Multimodal AI Agents
Jianwei Yang, Reuben Tan, Qianhui Wu +10
cs.CVcs.AIcs.HCarXiv:2502.13130v12025RoboTok: An Internet-Scale Data Engine for Human Demonstration Retrieval and Dexterous Manipulation Learning
Howard Qian, Yiting Chen, Yunfei Xie +6
cs.CVcs.ROarXiv:2609.03199v12026A Careful Examination of Large Behavior Models for Multitask Dexterous Manipulation
TRI LBM Team, Jose Barreiros, Andrew Beaulieu +79
cs.ROarXiv:2507.05331v12025Learning Plannable Representations with Causal InfoGAN
Thanard Kurutach, Aviv Tamar, Ge Yang +2
cs.LGcs.AIcs.CVarXiv:1807.09341v12018PLAS: Latent Action Space for Offline Reinforcement Learning
Wenxuan Zhou, Sujay Bajracharya, David Held
cs.ROcs.AIcs.LGarXiv:2011.07213v12020PhysBench: Benchmarking and Enhancing Vision-Language Models for Physical World Understanding
Wei Chow, Jiageng Mao, Boyi Li +3
cs.CVcs.AIcs.CLarXiv:2501.16411v22025Air-Ground Collaborative Vision-and-Language Navigation via Shared Bird's-Eye Maps
Shuning Zhang, Liang Li, Yunheng Wang +3
cs.ROcs.AIarXiv:2609.03483v12026R2S-Eval: Robot Evaluation with Real-to-Sim Calibration via Vision-Language Models
Yidi Wang, Feixiang Ruan, Ruoqu Chen +4
cs.ROarXiv:2609.03276v12026Swarm-SLAM : Sparse Decentralized Collaborative Simultaneous Localization and Mapping Framework for Multi-Robot Systems
Pierre-Yves Lajoie, Giovanni Beltrame
cs.ROcs.CVarXiv:2301.06230v32023ConRFT: A Reinforced Fine-tuning Method for VLA Models via Consistency Policy
Yuhui Chen, Shuai Tian, Shugao Liu +3
cs.ROcs.AIarXiv:2502.05450v22025LeRobot: An Open-Source Library for End-to-End Robot Learning
Remi Cadene, Simon Aliberts, Francesco Capuano +14
cs.ROarXiv:2602.22818v12026GraspVLA: a Grasping Foundation Model Pre-trained on Billion-scale Synthetic Action Data
Shengliang Deng, Mi Yan, Songlin Wei +10
cs.ROarXiv:2505.03233v32025Knowledge Insulating Vision-Language-Action Models: Train Fast, Run Fast, Generalize Better
Danny Driess, Jost Tobias Springenberg, Brian Ichter +8
cs.LGcs.ROarXiv:2505.23705v12025Autonomous Drone Racing with Deep Reinforcement Learning
Yunlong Song, Mats Steinweg, Elia Kaufmann +1
cs.ROcs.AIarXiv:2103.08624v22021On Global Regulatability of Robot Manipulators by Classical PID
Cheng Zhao, Jingru Zhu, Lei Guo
cs.ROeess.SYarXiv:2609.01207v22026Importance and methods to control, vary, and characterize mud strength for studying locomotion
Divya Ramesh, Gargi Sadalgekar, Qiyuan Fu +3
physics.bio-phcond-mat.softcs.ROarXiv:2609.00563v12026Asymptotically near-optimal RRT for fast, high-quality, motion planning
Oren Salzman, Dan Halperin
cs.ROarXiv:1308.0189v42013Cosmos-Reason1: From Physical Common Sense To Embodied Reasoning
NVIDIA, :, Alisson Azzolini +51
cs.AIcs.CVcs.LGarXiv:2503.15558v32025Mudskippers use tail thrusting to help crutching to move on mud of various wetness
Divya Ramesh, Gargi Sadalgekar, Jiangqi Tan +1
physics.bio-phcond-mat.softcs.ROarXiv:2609.00564v12026DexVLA: Vision-Language Model with Plug-In Diffusion Expert for General Robot Control
Junjie Wen, Yichen Zhu, Jinming Li +3
cs.ROcs.CVarXiv:2502.05855v32025Beyond Imitation: Self-Improving Robot Policies via Off-Policy Q-Planning
Varun Giridhar, Anant Khandelwal, Jeremy A. Collins +2
cs.ROcs.LGarXiv:2608.21204v12026AURA: Action-Gated Memory for Robot Policies at Constant VRAM
Josef Chen
cs.AIcs.ARcs.DCarXiv:2606.02775v12026Good Token Hunting: A Hitchhiker's Guide to Token Selection for Visual Geometry Transformers
Shuhong Zheng, Michael Oechsle, Erik Sandström +3
cs.CVcs.AIcs.GRarXiv:2605.23892v12026$π_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities
Physical Intelligence, Bo Ai, Ali Amin +85
cs.LGcs.ROarXiv:2604.15483v22026Trust as indicator of robot functional and social acceptance. An experimental study on user conformation to the iCub's answers
Ilaria Gaudiello, Elisabetta Zibetti, Sebastien Lefort +2
cs.ROcs.CYcs.HCarXiv:1510.03678v12015RoboRefer: Towards Spatial Referring with Reasoning in Vision-Language Models for Robotics
Enshen Zhou, Jingkun An, Cheng Chi +8
cs.ROcs.AIcs.CVarXiv:2506.04308v42025DreamGen: Unlocking Generalization in Robot Learning through Video World Models
Joel Jang, Seonghyeon Ye, Zongyu Lin +25
cs.ROcs.AIcs.LGarXiv:2505.12705v22025EgoHumanoid: Unlocking In-the-Wild Loco-Manipulation with Robot-Free Egocentric Demonstration
Modi Shi, Shijia Peng, Jin Chen +6
cs.ROarXiv:2602.10106v22026$π_0$: A Vision-Language-Action Flow Model for General Robot Control
Kevin Black, Noah Brown, Danny Driess +21
cs.LGcs.ROarXiv:2410.24164v42024Retargeting Matters: General Motion Retargeting for Humanoid Motion Tracking
Joao Pedro Araujo, Yanjie Ze, Pei Xu +2
cs.ROarXiv:2510.02252v12025Gaussian Splatting SLAM
Hidenobu Matsuki, Riku Murai, Paul H. J. Kelly +1
cs.CVcs.ROarXiv:2312.06741v22023Reactive Diffusion Policy: Slow-Fast Visual-Tactile Policy Learning for Contact-Rich Manipulation
Han Xue, Jieji Ren, Wendi Chen +5
cs.ROcs.AIcs.LGarXiv:2503.02881v32025GMT: General Motion Tracking for Humanoid Whole-Body Control
Zixuan Chen, Mazeyu Ji, Xuxin Cheng +3
cs.ROarXiv:2506.14770v22025Genie Envisioner: A Unified World Foundation Platform for Robotic Manipulation
Yue Liao, Pengfei Zhou, Siyuan Huang +11
cs.ROcs.CVarXiv:2508.05635v32025Pseudo-Simulation for Autonomous Driving
Wei Cao, Marcel Hallgarten, Tianyu Li +11
cs.ROcs.AIcs.CVarXiv:2506.04218v32025ChainSplat: A Physics-Inspired Screw-Theoretic Model for Learning Deformable Linear Object Dynamics from Multi-View RGB Videos
Seungyeon Kim, Noémie Jaquier
cs.ROarXiv:2608.28570v12026SonicNudge: Controlled Displacement of Hovering UAVs via Estimator-Controller Coupling
Shaocheng Luo, Ashir Raza, Haocheng Meng +2
eess.SYcs.ROarXiv:2608.25319v12026Exploring Nonlinear Body Oscillations for Natural Quadruped Gaits
Annika Schmidt, Davide Calzolari, Arne Sachtler +13
cs.ROarXiv:2609.00539v22026Lambda-Hold Control: Human-Like Movement Emerges from a Minimal Task Reward in Predictive Musculoskeletal Simulation
Jun Hyuk Lee, Chihyeong Lee, Jooeun Ahn
cs.ROcs.GRcs.LGarXiv:2608.17030v12026PRM-as-a-Judge 1.5: A Toolkit for Robot Process Assessment
Yuyang Liu, Yanqing Shen, Ruike Chen +22
cs.ROcs.CVarXiv:2608.14284v12026Summaries:한국어Unified Video Action Model
Shuang Li, Yihuai Gao, Dorsa Sadigh +1
cs.ROcs.CVarXiv:2503.00200v32025HumanTracker: Towards Comprehensive and Human-Aligned Motion Tracking Benchmark
Dairu Liu, Zekun Qi, Jiayu Zeng +11
cs.ROcs.AIcs.CVarXiv:2608.13555v12026Spatial Forcing: Implicit Spatial Representation Alignment for Vision-language-action Model
Fuhao Li, Wenxuan Song, Han Zhao +5
cs.ROarXiv:2510.12276v22025StreamVLN: Streaming Vision-and-Language Navigation via SlowFast Context Modeling
Meng Wei, Chenyang Wan, Xiqian Yu +9
cs.ROcs.CVarXiv:2507.05240v22025VisualPatchWorld: Code World Models as Latent Structured Representations for Planning
Jiaxin Bai, Jiaxuan Xiong
cs.CLcs.ROarXiv:2607.25236v120262018 Robotic Scene Segmentation Challenge
Max Allan, Satoshi Kondo, Sebastian Bodenstedt +38
cs.CVcs.ROarXiv:2001.11190v32020Imagined Rollouts are Kinematic, Not Dynamic: A Diagnosis of Long-Horizon World-Model Failure
Finn Rasmus Schäfer, Korbinian Moller, Yuan Gao +3
cs.ROarXiv:2607.05966v12026