Robotics

Papers filed under cs.RO on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

2,341 to 2,400 of 3,243

  1. TrapVLA: Trapping Vision-Language-Action Models in Configured Failure Modes

    Jun-Hui Liu, Kun-Yu Lin, Yi-Lin Wei +9

    cs.ROcs.CVarXiv:2608.26578v12026
  2. Active sensing to characterize the heterogeneity of plant stress

    Ayman Laaroussi, Peter Hanappe, David Colliaux

    cs.ROcs.AIarXiv:2608.27088v12026
  3. MapTR: Structured Modeling and Learning for Online Vectorized HD Map Construction

    Bencheng Liao, Shaoyu Chen, Xinggang Wang +4

    cs.CVcs.ROarXiv:2208.14437v22022
  4. Pass the Bucket: Efficient, Robust, Local Load Balancing for Teams of Heterogeneous Robots

    Tobias Wallner, Dominik Krupke, Arne Schmidt +1

    cs.ROarXiv:2608.27085v12026
  5. PPF-FoldNet: Unsupervised Learning of Rotation Invariant 3D Local Descriptors

    Haowen Deng, Tolga Birdal, Slobodan Ilic

    cs.CVcs.CGcs.LGarXiv:1808.10322v12018
  6. Language to Rewards for Robotic Skill Synthesis

    Wenhao Yu, Nimrod Gileadi, Chuyuan Fu +17

    cs.ROcs.AIcs.LGarXiv:2306.08647v22023
  7. MaskFusion: Real-Time Recognition, Tracking and Reconstruction of Multiple Moving Objects

    Martin Rünz, Maud Buffier, Lourdes Agapito

    cs.CVcs.ROarXiv:1804.09194v22018
  8. ANYmal Parkour: Learning Agile Navigation for Quadrupedal Robots

    David Hoeller, Nikita Rudin, Dhionis Sako +1

    cs.ROarXiv:2306.14874v12023
    Summaries:한국어
  9. Real-time Semantic Segmentation of Crop and Weed for Precision Agriculture Robots Leveraging Background Knowledge in CNNs

    Andres Milioto, Philipp Lottes, Cyrill Stachniss

    cs.CVcs.ROarXiv:1709.06764v22017
  10. Vision-Language Foundation Models as Effective Robot Imitators

    Xinghang Li, Minghuan Liu, Hanbo Zhang +9

    cs.ROcs.AIcs.LGarXiv:2311.01378v32023
  11. CLAP: Cross-Embodiment Video World Models are Zero-Shot Physical Simulators

    Kechen Liu, Ola Shorinwa

    cs.ROcs.AIcs.CVarXiv:2608.27406v12026
  12. A Unifying Contrast Maximization Framework for Event Cameras, with Applications to Motion, Depth, and Optical Flow Estimation

    Guillermo Gallego, Henri Rebecq, Davide Scaramuzza

    cs.CVcs.ROarXiv:1804.01306v12018
  13. SpatialCrafter: Single Image World Modeling with Generative 3D Proxies

    Chuan Fang, Lingteng Qiu, Yixun Liang +8

    cs.CVcs.ROarXiv:2608.27073v12026
    Summaries:한국어
  14. Geometrically Constrained Trajectory Optimization for Multicopters

    Zhepei Wang, Xin Zhou, Chao Xu +1

    cs.ROarXiv:2103.00190v42021
  15. Information Theoretic Model Predictive Control: Theory and Applications to Autonomous Driving

    Grady Williams, Paul Drews, Brian Goldfain +2

    cs.ROarXiv:1707.02342v12017
  16. Explaining How a Deep Neural Network Trained with End-to-End Learning Steers a Car

    Mariusz Bojarski, Philip Yeres, Anna Choromanska +4

    cs.CVcs.LGcs.NEarXiv:1704.07911v12017
  17. More Than a Feeling: Learning to Grasp and Regrasp using Vision and Touch

    Roberto Calandra, Andrew Owens, Dinesh Jayaraman +5

    cs.ROcs.LGstat.MLarXiv:1805.11085v22018
  18. STEP: State-Aware Task Estimation and Planning with Multi-Modal LLMs for Human-Robot Collaboration

    Maitrey Gramopadhye, Prakash Baskaran, Xiao Liu +2

    cs.ROcs.AIarXiv:2608.27225v12026
  19. EmbodiedGPT: Vision-Language Pre-Training via Embodied Chain of Thought

    Yao Mu, Qinglong Zhang, Mengkang Hu +7

    cs.ROcs.AIcs.CVarXiv:2305.15021v22023
  20. Behavior Transformers: Cloning $k$ modes with one stone

    Nur Muhammad Mahi Shafiullah, Zichen Jeff Cui, Ariuntuya Altanzaya +1

    cs.LGcs.AIcs.CVarXiv:2206.11251v22022
  21. Learning Distilled Collaboration Graph for Multi-Agent Perception

    Yiming Li, Shunli Ren, Pengxiang Wu +3

    cs.CVcs.ROarXiv:2111.00643v22021
  22. Semi-parametric Topological Memory for Navigation

    Nikolay Savinov, Alexey Dosovitskiy, Vladlen Koltun

    cs.LGcs.AIcs.CVarXiv:1803.00653v12018
  23. NavGPT: Explicit Reasoning in Vision-and-Language Navigation with Large Language Models

    Gengze Zhou, Yicong Hong, Qi Wu

    cs.CVcs.AIcs.CLarXiv:2305.16986v32023
  24. Learning-based Model Predictive Control for Safe Exploration

    Torsten Koller, Felix Berkenkamp, Matteo Turchetta +1

    eess.SYcs.AIcs.LGarXiv:1803.08287v32018
  25. PointNetGPD: Detecting Grasp Configurations from Point Sets

    Hongzhuo Liang, Xiaojian Ma, Shuang Li +5

    cs.ROarXiv:1809.06267v42018
  26. MultiPath++: Efficient Information Fusion and Trajectory Aggregation for Behavior Prediction

    Balakrishnan Varadarajan, Ahmed Hefny, Avikalp Srivastava +8

    cs.CVcs.AIcs.LGarXiv:2111.14973v32021
  27. Learning Deep Control Policies for Autonomous Aerial Vehicles with MPC-Guided Policy Search

    Tianhao Zhang, Gregory Kahn, Sergey Levine +1

    cs.LGcs.ROarXiv:1509.06791v22015
  28. Highly Dynamic Quadruped Locomotion via Whole-Body Impulse Control and Model Predictive Control

    Donghyun Kim, Jared Di Carlo, Benjamin Katz +2

    cs.ROarXiv:1909.06586v12019
  29. Semantic Flow for Fast and Accurate Scene Parsing

    Xiangtai Li, Ansheng You, Zhen Zhu +4

    cs.CVcs.ROarXiv:2002.10120v32020
  30. RoboNet: Large-Scale Multi-Robot Learning

    Sudeep Dasari, Frederik Ebert, Stephen Tian +6

    cs.ROcs.CVcs.LGarXiv:1910.11215v22019
  31. Decoupling Planning and Control for Instructable Agents

    Zineng Tang, Kelsey R. Allen, Sjoerd van Steenkiste +2

    cs.AIcs.CLcs.MAarXiv:2608.26788v12026
  32. The Marathon 2: A Navigation System

    Steve Macenski, Francisco Martín, Ruffin White +1

    cs.ROarXiv:2003.00368v22020
  33. Omnidata: A Scalable Pipeline for Making Multi-Task Mid-Level Vision Datasets from 3D Scans

    Ainaz Eftekhar, Alexander Sax, Roman Bachmann +2

    cs.CVcs.AIcs.GRarXiv:2110.04994v12021
  34. Embodied Intelligence via Learning and Evolution

    Agrim Gupta, Silvio Savarese, Surya Ganguli +1

    cs.LGcs.NEcs.ROarXiv:2102.02202v12021
  35. Image Segmentation for Fruit Detection and Yield Estimation in Apple Orchards

    Suchet Bargoti, James Underwood

    cs.ROcs.CVcs.LGarXiv:1610.08120v12016
  36. Learning Monocular Reactive UAV Control in Cluttered Natural Environments

    Stephane Ross, Narek Melik-Barkhudarov, Kumar Shaurya Shankar +4

    cs.ROcs.CVcs.LGarXiv:1211.1690v12012
  37. Why did My Robot Just Change Personality? Prompting Guidelines for a Grounded Robot Persona in LLM-Based HRI

    Ashita Ashok, Franziska Babel, Patrick Holthaus +7

    cs.AIcs.HCcs.ROarXiv:2608.26182v12026
  38. TemporalFlow-VLA: Learning Physically Grounded Execution History for Long-Horizon Robot Manipulation

    Jiarui Yang, Yehao Lu, Yuning Su +9

    cs.ROarXiv:2608.26821v12026
  39. A Review of Verbal and Non-Verbal Human-Robot Interactive Communication

    Nikolaos Mavridis

    cs.ROcs.CLarXiv:1401.4994v12014
  40. MetaDrive: Composing Diverse Driving Scenarios for Generalizable Reinforcement Learning

    Quanyi Li, Zhenghao Peng, Lan Feng +3

    cs.LGcs.ROarXiv:2109.12674v32021
  41. Learning Modular Neural Network Policies for Multi-Task and Multi-Robot Transfer

    Coline Devin, Abhishek Gupta, Trevor Darrell +2

    cs.LGcs.ROarXiv:1609.07088v12016
  42. CGS-SLAM: Collaborative Gaussian Splatting based SLAM for Multi-Agent Reconstruction

    Jean-Daniel de Ambrogi, Aladine Chetouani, Vincent Nguyen +1

    cs.CVcs.ROarXiv:2608.26868v12026
  43. RoboCasa: Large-Scale Simulation of Everyday Tasks for Generalist Robots

    Soroush Nasiriany, Abhiram Maddukuri, Lance Zhang +5

    cs.ROcs.AIcs.LGarXiv:2406.02523v12024
  44. Analysis and Observations from the First Amazon Picking Challenge

    Nikolaus Correll, Kostas E. Bekris, Dmitry Berenson +7

    cs.ROarXiv:1601.05484v32016
  45. Mining beyond Earth with Space Robots: Exploration, Sampling, and Extraction

    Dong Li, Dujun Nie, Xiaotong Zhang +11

    cs.ROarXiv:2608.21358v12026
  46. Zero-WAM: In-Context World-Action Modeling from Human Videos for Open-Ended Task Generalization

    Jiaming Zhou, Qihang Zhang, Gangwei Xu +9

    cs.ROcs.CVarXiv:2608.26103v22026
    Summaries:한국어
  47. Active Learning of Inverse Models with Intrinsically Motivated Goal Exploration in Robots

    Adrien Baranes, Pierre-Yves Oudeyer

    cs.LGcs.AIcs.CVarXiv:1301.4862v12013
  48. Self-Supervised Exploration via Disagreement

    Deepak Pathak, Dhiraj Gandhi, Abhinav Gupta

    cs.LGcs.AIcs.CVarXiv:1906.04161v12019
  49. Memory Anchors for Continual Robot Learning

    Maximilian Du, Zhanyi Sun, Chen Xu +3

    cs.ROarXiv:2608.26545v12026
  50. TidyBot: Personalized Robot Assistance with Large Language Models

    Jimmy Wu, Rika Antonova, Adam Kan +6

    cs.ROcs.AIcs.CLarXiv:2305.05658v22023
  51. Self-supervised Sparse-to-Dense: Self-supervised Depth Completion from LiDAR and Monocular Camera

    Fangchang Ma, Guilherme Venturelli Cavalheiro, Sertac Karaman

    cs.CVcs.AIcs.LGarXiv:1807.00275v22018
  52. TossingBot: Learning to Throw Arbitrary Objects with Residual Physics

    Andy Zeng, Shuran Song, Johnny Lee +2

    cs.ROcs.AIcs.CVarXiv:1903.11239v32019
  53. CubeSLAM: Monocular 3D Object SLAM

    Shichao Yang, Sebastian Scherer

    cs.ROcs.CVarXiv:1806.00557v22018
  54. Safety-Critical Model Predictive Control with Discrete-Time Control Barrier Function

    Jun Zeng, Bike Zhang, Koushil Sreenath

    eess.SYcs.ROarXiv:2007.11718v32020
  55. Beyond the Proving Ground: Independent Public-Road Testing of Assisted Lane Change Systems using LiDAR

    Marcello Cellina, Akos Kriston, Antonio Migneco +7

    cs.ROcs.CVarXiv:2608.26669v12026
  56. Deep Imitation Learning for Complex Manipulation Tasks from Virtual Reality Teleoperation

    Tianhao Zhang, Zoe McCarthy, Owen Jow +4

    cs.LGcs.ROarXiv:1710.04615v22017
    Summaries:한국어
  57. FLARE: A Failure-Aware Framework for Autonomous Correction and Recovery in Visual-Language Robotic Manipulation

    Ganlong Zhao, Zijia Tang, Xingping Chen +3

    cs.ROarXiv:2608.26645v12026
  58. GPT-Driver: Learning to Drive with GPT

    Jiageng Mao, Yuxi Qian, Junjie Ye +2

    cs.CVcs.AIcs.CLarXiv:2310.01415v32023
  59. Making Sense of Vision and Touch: Self-Supervised Learning of Multimodal Representations for Contact-Rich Tasks

    Michelle A. Lee, Yuke Zhu, Krishnan Srinivasan +5

    cs.ROcs.AIcs.LGarXiv:1810.10191v22018
  60. Rapid On-Robot Learning for Dynamic Manipulation Skills: Robot Juggling

    Taeyoon Lee, Chunpeng Wang, Christopher G. Atkeson +2

    cs.ROarXiv:2608.26800v12026