Robotics

Papers filed under cs.RO on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

2,221 to 2,280 of 3,243

  1. Driving with LLMs: Fusing Object-Level Vector Modality for Explainable Autonomous Driving

    Long Chen, Oleg Sinavski, Jan Hünermann +5

    cs.ROcs.AIcs.CLarXiv:2310.01957v22023
  2. Bridge Data: Boosting Generalization of Robotic Skills with Cross-Domain Datasets

    Frederik Ebert, Yanlai Yang, Karl Schmeckpeper +5

    cs.ROcs.AIarXiv:2109.13396v12021
  3. Extreme Parkour with Legged Robots

    Xuxin Cheng, Kexin Shi, Ananye Agarwal +1

    cs.ROcs.AIcs.CVarXiv:2309.14341v12023
  4. Multimodal Virtual Point 3D Detection

    Tianwei Yin, Xingyi Zhou, Philipp Krähenbühl

    cs.CVcs.LGcs.ROarXiv:2111.06881v12021
  5. GVINS: Tightly Coupled GNSS-Visual-Inertial Fusion for Smooth and Consistent State Estimation

    Shaozu Cao, Xiuyuan Lu, Shaojie Shen

    cs.ROarXiv:2103.07899v32021
  6. AffordanceNet: An End-to-End Deep Learning Approach for Object Affordance Detection

    Thanh-Toan Do, Anh Nguyen, Ian Reid

    cs.CVcs.ROarXiv:1709.07326v32017
  7. RTM3D: Real-time Monocular 3D Detection from Object Keypoints for Autonomous Driving

    Peixuan Li, Huaici Zhao, Pengfei Liu +1

    cs.CVcs.ROeess.IVarXiv:2001.03343v12020
  8. PDDLStream: Integrating Symbolic Planners and Blackbox Samplers via Optimistic Adaptive Planning

    Caelan Reed Garrett, Tomás Lozano-Pérez, Leslie Pack Kaelbling

    cs.AIcs.ROarXiv:1802.08705v52018
  9. 3D Diffuser Actor: Policy Diffusion with 3D Scene Representations

    Tsung-Wei Ke, Nikolaos Gkanatsios, Katerina Fragkiadaki

    cs.ROcs.AIcs.CVarXiv:2402.10885v32024
  10. Kimera: from SLAM to Spatial Perception with 3D Dynamic Scene Graphs

    Antoni Rosinol, Andrew Violette, Marcus Abate +5

    cs.ROcs.CVarXiv:2101.06894v32021
  11. FUEL: Fast UAV Exploration using Incremental Frontier Structure and Hierarchical Planning

    Boyu Zhou, Yichen Zhang, Xinyi Chen +1

    cs.ROarXiv:2010.11561v22020
  12. Towards Learning a Generic Agent for Vision-and-Language Navigation via Pre-training

    Weituo Hao, Chunyuan Li, Xiujun Li +2

    cs.CVcs.CLcs.LGarXiv:2002.10638v22020
  13. VectorMapNet: End-to-end Vectorized HD Map Learning

    Yicheng Liu, Tianyuan Yuan, Yue Wang +2

    cs.CVcs.ROarXiv:2206.08920v62022
  14. Appearance-Based Loop Closure Detection for Online Large-Scale and Long-Term Operation

    Mathieu Labbé, François Michaud

    cs.ROcs.CVarXiv:2407.15304v12024
  15. Stochastic Neural Networks for Hierarchical Reinforcement Learning

    Carlos Florensa, Yan Duan, Pieter Abbeel

    cs.AIcs.LGcs.NEarXiv:1704.03012v12017
  16. Not All Points Are Equal: Learning Highly Efficient Point-based Detectors for 3D LiDAR Point Clouds

    Yifan Zhang, Qingyong Hu, Guoquan Xu +3

    cs.CVcs.ROarXiv:2203.11139v12022
  17. Robotic Control via Embodied Chain-of-Thought Reasoning

    Michał Zawalski, William Chen, Karl Pertsch +3

    cs.ROcs.LGarXiv:2407.08693v32024
  18. Safety-Enhanced Autonomous Driving Using Interpretable Sensor Fusion Transformer

    Hao Shao, Letian Wang, RuoBing Chen +2

    cs.CVcs.AIcs.LGarXiv:2207.14024v52022
  19. Your Diffusion Model is Secretly a Zero-Shot Classifier

    Alexander C. Li, Mihir Prabhudesai, Shivam Duggal +2

    cs.LGcs.AIcs.CVarXiv:2303.16203v32023
  20. Visual Semantic Navigation using Scene Priors

    Wei Yang, Xiaolong Wang, Ali Farhadi +2

    cs.CVcs.AIcs.ROarXiv:1810.06543v12018
  21. ObjectNav Revisited: On Evaluation of Embodied Agents Navigating to Objects

    Dhruv Batra, Aaron Gokaslan, Aniruddha Kembhavi +5

    cs.CVcs.ROarXiv:2006.13171v22020
  22. FastDepth: Fast Monocular Depth Estimation on Embedded Systems

    Diana Wofk, Fangchang Ma, Tien-Ju Yang +2

    cs.CVcs.ROarXiv:1903.03273v12019
  23. Efficient Informative Sensing using Multiple Robots

    Amarjeet Singh, Andreas Krause, Carlos Guestrin +1

    cs.ROcs.AIarXiv:1401.3462v12014
  24. Deep Stereo using Adaptive Thin Volume Representation with Uncertainty Awareness

    Shuo Cheng, Zexiang Xu, Shilin Zhu +4

    cs.CVcs.LGcs.ROarXiv:1911.12012v22019
  25. NAVSIM: Data-Driven Non-Reactive Autonomous Vehicle Simulation and Benchmarking

    Daniel Dauner, Marcel Hallgarten, Tianyu Li +9

    cs.CVcs.AIcs.LGarXiv:2406.15349v22024
  26. 3D-VLA: A 3D Vision-Language-Action Generative World Model

    Haoyu Zhen, Xiaowen Qiu, Peihao Chen +5

    cs.CVcs.AIcs.CLarXiv:2403.09631v12024
  27. UniSim: A Neural Closed-Loop Sensor Simulator

    Ze Yang, Yun Chen, Jingkang Wang +4

    cs.CVcs.ROarXiv:2308.01898v12023
  28. Generative Semantic Scene Completion

    Shi Chen, Weifeng Ge

    cs.CVcs.LGcs.ROarXiv:2608.26737v12026
  29. Multiple Futures Prediction

    Yichuan Charlie Tang, Ruslan Salakhutdinov

    cs.LGcs.CVcs.MAarXiv:1911.00997v22019
  30. VLFM: Vision-Language Frontier Maps for Zero-Shot Semantic Navigation

    Naoki Yokoyama, Sehoon Ha, Dhruv Batra +2

    cs.ROcs.AIarXiv:2312.03275v12023
  31. Crocoddyl: An Efficient and Versatile Framework for Multi-Contact Optimal Control

    Carlos Mastalli, Rohan Budhiraja, Wolfgang Merkt +7

    cs.ROmath.OCarXiv:1909.04947v22019
  32. Fast Optical Flow using Dense Inverse Search

    Till Kroeger, Radu Timofte, Dengxin Dai +1

    cs.CVcs.ROarXiv:1603.03590v12016
  33. FlashVLA: Streaming Action Decoding for Fast and Asynchronous VLA Inference

    Zekai Li, Jiaming Tang, Zhijian Liu

    cs.ROarXiv:2608.27384v12026
  34. Neural Topological SLAM for Visual Navigation

    Devendra Singh Chaplot, Ruslan Salakhutdinov, Abhinav Gupta +1

    cs.CVcs.AIcs.LGarXiv:2005.12256v22020
  35. IONet: Learning to Cure the Curse of Drift in Inertial Odometry

    Changhao Chen, Xiaoxuan Lu, Andrew Markham +1

    cs.ROcs.AIcs.CVarXiv:1802.02209v12018
  36. LaserNet: An Efficient Probabilistic 3D Object Detector for Autonomous Driving

    Gregory P. Meyer, Ankit Laddha, Eric Kee +2

    cs.CVcs.LGcs.ROarXiv:1903.08701v12019
  37. System Design and Control of an Apple Harvesting Robot

    Kaixiang Zhang, Kyle Lammers, Pengyu Chu +2

    cs.ROeess.SYarXiv:2010.11296v12020
  38. One-Shot Imitation from Observing Humans via Domain-Adaptive Meta-Learning

    Tianhe Yu, Chelsea Finn, Annie Xie +4

    cs.LGcs.AIcs.CVarXiv:1802.01557v12018
  39. HG-DAgger: Interactive Imitation Learning with Human Experts

    Michael Kelly, Chelsea Sidrane, Katherine Driggs-Campbell +1

    cs.ROarXiv:1810.02890v22018
  40. Learning Object Bounding Boxes for 3D Instance Segmentation on Point Clouds

    Bo Yang, Jianan Wang, Ronald Clark +4

    cs.CVcs.AIcs.LGarXiv:1906.01140v22019
  41. RoboTurk: A Crowdsourcing Platform for Robotic Skill Learning through Imitation

    Ajay Mandlekar, Yuke Zhu, Animesh Garg +9

    cs.ROcs.AIcs.LGarXiv:1811.02790v12018
  42. ThreeDWorld: A Platform for Interactive Multi-Modal Physical Simulation

    Chuang Gan, Jeremy Schwartz, Seth Alter +21

    cs.CVcs.GRcs.LGarXiv:2007.04954v22020
  43. Rapid Locomotion via Reinforcement Learning

    Gabriel B Margolis, Ge Yang, Kartik Paigwar +2

    cs.ROcs.AIcs.LGarXiv:2205.02824v12022
  44. Control Barrier Functions for Systems with High Relative Degree

    Wei Xiao, Calin Belta

    eess.SYcs.ROarXiv:1903.04706v22019
  45. ManiSkill2: A Unified Benchmark for Generalizable Manipulation Skills

    Jiayuan Gu, Fanbo Xiang, Xuanlin Li +12

    cs.ROcs.AIarXiv:2302.04659v12023
  46. TADP: Task-Aware Deformable Prediction for Single-Stage 3D Object Detection

    Su Wang, Yaochen Li, Min Yang +3

    cs.CVcs.AIcs.ROarXiv:2608.27282v12026
  47. Benchmarking Model-Based Reinforcement Learning

    Tingwu Wang, Xuchan Bao, Ignasi Clavera +7

    cs.LGcs.AIcs.ROarXiv:1907.02057v12019
  48. F-LOAM: Fast LiDAR Odometry And Mapping

    Han Wang, Chen Wang, Chun-Lin Chen +1

    cs.ROarXiv:2107.00822v12021
  49. NeuralRecon: Real-Time Coherent 3D Reconstruction from Monocular Video

    Jiaming Sun, Yiming Xie, Linghao Chen +2

    cs.CVcs.ROarXiv:2104.00681v12021
  50. R3LIVE: A Robust, Real-time, RGB-colored, LiDAR-Inertial-Visual tightly-coupled state Estimation and mapping package

    Jiarong Lin, Fu Zhang

    cs.ROcs.CVarXiv:2109.07982v12021
  51. ReKep: Spatio-Temporal Reasoning of Relational Keypoint Constraints for Robotic Manipulation

    Wenlong Huang, Chen Wang, Yunzhu Li +2

    cs.ROcs.AIcs.CVarXiv:2409.01652v22024
  52. Constraint-Aware Physics-Informed Neural Networks for Static Shape Estimation of Co-Manipulative Continuum Robots

    Rana Danesh, Pari Qarehdaghi, Farrokh Janabi-Sharifi

    cs.ROcs.LGarXiv:2608.26273v12026
  53. Dispersive Forward Tree Search for Optimal Control: Coverage, Complexity, and Computation

    Shashank A. Deshpande, Jonathan P. How

    cs.ROmath.OCarXiv:2608.26314v12026
  54. WALL-SS: Scaling Long-horizon World Models via Next-Scale Autoregression

    Maeve Zhang, Rain Sun, Xiang Wang +22

    cs.ROarXiv:2608.26239v12026
  55. RoNIN: Robust Neural Inertial Navigation in the Wild: Benchmark, Evaluations, and New Methods

    Hang Yan, Sachini Herath, Yasutaka Furukawa

    cs.CVcs.ROarXiv:1905.12853v12019
  56. Exact Finite-Length Theory of Uniform Car Parking: Spatial Laws, Absorption, and Aggregation

    Ganesh P Kumar

    cs.ROcs.DScs.SCarXiv:2608.22671v12026
  57. NoMaD: Goal Masked Diffusion Policies for Navigation and Exploration

    Ajay Sridhar, Dhruv Shah, Catherine Glossop +1

    cs.ROcs.CVcs.LGarXiv:2310.07896v12023
  58. Arbitrary-Order Hermite Interpolation of Rigid-Motion Jets via Hyper-Multidual Quaternions

    Daniel Condurache

    cs.ROarXiv:2608.27000v12026
  59. Revisiting Active Perception

    Ruzena Bajcsy, Yiannis Aloimonos, John K. Tsotsos

    cs.CVcs.ROarXiv:1603.02729v22016
  60. Task-space model-based control of pneumatic soft actuators

    Nithin S. Kumar, Joshua Gaston, D. Caleb Rucker +1

    cs.ROarXiv:2608.27186v12026