Robotics
Papers filed under cs.RO on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
2,341 to 2,400 of 3,243
TrapVLA: Trapping Vision-Language-Action Models in Configured Failure Modes
Jun-Hui Liu, Kun-Yu Lin, Yi-Lin Wei +9
cs.ROcs.CVarXiv:2608.26578v12026Active sensing to characterize the heterogeneity of plant stress
Ayman Laaroussi, Peter Hanappe, David Colliaux
cs.ROcs.AIarXiv:2608.27088v12026MapTR: Structured Modeling and Learning for Online Vectorized HD Map Construction
Bencheng Liao, Shaoyu Chen, Xinggang Wang +4
cs.CVcs.ROarXiv:2208.14437v22022Pass the Bucket: Efficient, Robust, Local Load Balancing for Teams of Heterogeneous Robots
Tobias Wallner, Dominik Krupke, Arne Schmidt +1
cs.ROarXiv:2608.27085v12026PPF-FoldNet: Unsupervised Learning of Rotation Invariant 3D Local Descriptors
Haowen Deng, Tolga Birdal, Slobodan Ilic
cs.CVcs.CGcs.LGarXiv:1808.10322v12018Language to Rewards for Robotic Skill Synthesis
Wenhao Yu, Nimrod Gileadi, Chuyuan Fu +17
cs.ROcs.AIcs.LGarXiv:2306.08647v22023MaskFusion: Real-Time Recognition, Tracking and Reconstruction of Multiple Moving Objects
Martin Rünz, Maud Buffier, Lourdes Agapito
cs.CVcs.ROarXiv:1804.09194v22018ANYmal Parkour: Learning Agile Navigation for Quadrupedal Robots
David Hoeller, Nikita Rudin, Dhionis Sako +1
cs.ROarXiv:2306.14874v12023Summaries:한국어Real-time Semantic Segmentation of Crop and Weed for Precision Agriculture Robots Leveraging Background Knowledge in CNNs
Andres Milioto, Philipp Lottes, Cyrill Stachniss
cs.CVcs.ROarXiv:1709.06764v22017Vision-Language Foundation Models as Effective Robot Imitators
Xinghang Li, Minghuan Liu, Hanbo Zhang +9
cs.ROcs.AIcs.LGarXiv:2311.01378v32023CLAP: Cross-Embodiment Video World Models are Zero-Shot Physical Simulators
Kechen Liu, Ola Shorinwa
cs.ROcs.AIcs.CVarXiv:2608.27406v12026A Unifying Contrast Maximization Framework for Event Cameras, with Applications to Motion, Depth, and Optical Flow Estimation
Guillermo Gallego, Henri Rebecq, Davide Scaramuzza
cs.CVcs.ROarXiv:1804.01306v12018SpatialCrafter: Single Image World Modeling with Generative 3D Proxies
Chuan Fang, Lingteng Qiu, Yixun Liang +8
cs.CVcs.ROarXiv:2608.27073v12026Summaries:한국어Geometrically Constrained Trajectory Optimization for Multicopters
Zhepei Wang, Xin Zhou, Chao Xu +1
cs.ROarXiv:2103.00190v42021Information Theoretic Model Predictive Control: Theory and Applications to Autonomous Driving
Grady Williams, Paul Drews, Brian Goldfain +2
cs.ROarXiv:1707.02342v12017Explaining How a Deep Neural Network Trained with End-to-End Learning Steers a Car
Mariusz Bojarski, Philip Yeres, Anna Choromanska +4
cs.CVcs.LGcs.NEarXiv:1704.07911v12017More Than a Feeling: Learning to Grasp and Regrasp using Vision and Touch
Roberto Calandra, Andrew Owens, Dinesh Jayaraman +5
cs.ROcs.LGstat.MLarXiv:1805.11085v22018STEP: State-Aware Task Estimation and Planning with Multi-Modal LLMs for Human-Robot Collaboration
Maitrey Gramopadhye, Prakash Baskaran, Xiao Liu +2
cs.ROcs.AIarXiv:2608.27225v12026EmbodiedGPT: Vision-Language Pre-Training via Embodied Chain of Thought
Yao Mu, Qinglong Zhang, Mengkang Hu +7
cs.ROcs.AIcs.CVarXiv:2305.15021v22023Behavior Transformers: Cloning $k$ modes with one stone
Nur Muhammad Mahi Shafiullah, Zichen Jeff Cui, Ariuntuya Altanzaya +1
cs.LGcs.AIcs.CVarXiv:2206.11251v22022Learning Distilled Collaboration Graph for Multi-Agent Perception
Yiming Li, Shunli Ren, Pengxiang Wu +3
cs.CVcs.ROarXiv:2111.00643v22021Semi-parametric Topological Memory for Navigation
Nikolay Savinov, Alexey Dosovitskiy, Vladlen Koltun
cs.LGcs.AIcs.CVarXiv:1803.00653v12018NavGPT: Explicit Reasoning in Vision-and-Language Navigation with Large Language Models
Gengze Zhou, Yicong Hong, Qi Wu
cs.CVcs.AIcs.CLarXiv:2305.16986v32023Learning-based Model Predictive Control for Safe Exploration
Torsten Koller, Felix Berkenkamp, Matteo Turchetta +1
eess.SYcs.AIcs.LGarXiv:1803.08287v32018PointNetGPD: Detecting Grasp Configurations from Point Sets
Hongzhuo Liang, Xiaojian Ma, Shuang Li +5
cs.ROarXiv:1809.06267v42018MultiPath++: Efficient Information Fusion and Trajectory Aggregation for Behavior Prediction
Balakrishnan Varadarajan, Ahmed Hefny, Avikalp Srivastava +8
cs.CVcs.AIcs.LGarXiv:2111.14973v32021Learning Deep Control Policies for Autonomous Aerial Vehicles with MPC-Guided Policy Search
Tianhao Zhang, Gregory Kahn, Sergey Levine +1
cs.LGcs.ROarXiv:1509.06791v22015Highly Dynamic Quadruped Locomotion via Whole-Body Impulse Control and Model Predictive Control
Donghyun Kim, Jared Di Carlo, Benjamin Katz +2
cs.ROarXiv:1909.06586v12019Semantic Flow for Fast and Accurate Scene Parsing
Xiangtai Li, Ansheng You, Zhen Zhu +4
cs.CVcs.ROarXiv:2002.10120v32020RoboNet: Large-Scale Multi-Robot Learning
Sudeep Dasari, Frederik Ebert, Stephen Tian +6
cs.ROcs.CVcs.LGarXiv:1910.11215v22019Decoupling Planning and Control for Instructable Agents
Zineng Tang, Kelsey R. Allen, Sjoerd van Steenkiste +2
cs.AIcs.CLcs.MAarXiv:2608.26788v12026The Marathon 2: A Navigation System
Steve Macenski, Francisco Martín, Ruffin White +1
cs.ROarXiv:2003.00368v22020Omnidata: A Scalable Pipeline for Making Multi-Task Mid-Level Vision Datasets from 3D Scans
Ainaz Eftekhar, Alexander Sax, Roman Bachmann +2
cs.CVcs.AIcs.GRarXiv:2110.04994v12021Embodied Intelligence via Learning and Evolution
Agrim Gupta, Silvio Savarese, Surya Ganguli +1
cs.LGcs.NEcs.ROarXiv:2102.02202v12021Image Segmentation for Fruit Detection and Yield Estimation in Apple Orchards
Suchet Bargoti, James Underwood
cs.ROcs.CVcs.LGarXiv:1610.08120v12016Learning Monocular Reactive UAV Control in Cluttered Natural Environments
Stephane Ross, Narek Melik-Barkhudarov, Kumar Shaurya Shankar +4
cs.ROcs.CVcs.LGarXiv:1211.1690v12012Why did My Robot Just Change Personality? Prompting Guidelines for a Grounded Robot Persona in LLM-Based HRI
Ashita Ashok, Franziska Babel, Patrick Holthaus +7
cs.AIcs.HCcs.ROarXiv:2608.26182v12026TemporalFlow-VLA: Learning Physically Grounded Execution History for Long-Horizon Robot Manipulation
Jiarui Yang, Yehao Lu, Yuning Su +9
cs.ROarXiv:2608.26821v12026A Review of Verbal and Non-Verbal Human-Robot Interactive Communication
Nikolaos Mavridis
cs.ROcs.CLarXiv:1401.4994v12014MetaDrive: Composing Diverse Driving Scenarios for Generalizable Reinforcement Learning
Quanyi Li, Zhenghao Peng, Lan Feng +3
cs.LGcs.ROarXiv:2109.12674v32021Learning Modular Neural Network Policies for Multi-Task and Multi-Robot Transfer
Coline Devin, Abhishek Gupta, Trevor Darrell +2
cs.LGcs.ROarXiv:1609.07088v12016CGS-SLAM: Collaborative Gaussian Splatting based SLAM for Multi-Agent Reconstruction
Jean-Daniel de Ambrogi, Aladine Chetouani, Vincent Nguyen +1
cs.CVcs.ROarXiv:2608.26868v12026RoboCasa: Large-Scale Simulation of Everyday Tasks for Generalist Robots
Soroush Nasiriany, Abhiram Maddukuri, Lance Zhang +5
cs.ROcs.AIcs.LGarXiv:2406.02523v12024Analysis and Observations from the First Amazon Picking Challenge
Nikolaus Correll, Kostas E. Bekris, Dmitry Berenson +7
cs.ROarXiv:1601.05484v32016Mining beyond Earth with Space Robots: Exploration, Sampling, and Extraction
Dong Li, Dujun Nie, Xiaotong Zhang +11
cs.ROarXiv:2608.21358v12026Zero-WAM: In-Context World-Action Modeling from Human Videos for Open-Ended Task Generalization
Jiaming Zhou, Qihang Zhang, Gangwei Xu +9
cs.ROcs.CVarXiv:2608.26103v22026Summaries:한국어Active Learning of Inverse Models with Intrinsically Motivated Goal Exploration in Robots
Adrien Baranes, Pierre-Yves Oudeyer
cs.LGcs.AIcs.CVarXiv:1301.4862v12013Self-Supervised Exploration via Disagreement
Deepak Pathak, Dhiraj Gandhi, Abhinav Gupta
cs.LGcs.AIcs.CVarXiv:1906.04161v12019Memory Anchors for Continual Robot Learning
Maximilian Du, Zhanyi Sun, Chen Xu +3
cs.ROarXiv:2608.26545v12026TidyBot: Personalized Robot Assistance with Large Language Models
Jimmy Wu, Rika Antonova, Adam Kan +6
cs.ROcs.AIcs.CLarXiv:2305.05658v22023Self-supervised Sparse-to-Dense: Self-supervised Depth Completion from LiDAR and Monocular Camera
Fangchang Ma, Guilherme Venturelli Cavalheiro, Sertac Karaman
cs.CVcs.AIcs.LGarXiv:1807.00275v22018TossingBot: Learning to Throw Arbitrary Objects with Residual Physics
Andy Zeng, Shuran Song, Johnny Lee +2
cs.ROcs.AIcs.CVarXiv:1903.11239v32019CubeSLAM: Monocular 3D Object SLAM
Shichao Yang, Sebastian Scherer
cs.ROcs.CVarXiv:1806.00557v22018Safety-Critical Model Predictive Control with Discrete-Time Control Barrier Function
Jun Zeng, Bike Zhang, Koushil Sreenath
eess.SYcs.ROarXiv:2007.11718v32020Beyond the Proving Ground: Independent Public-Road Testing of Assisted Lane Change Systems using LiDAR
Marcello Cellina, Akos Kriston, Antonio Migneco +7
cs.ROcs.CVarXiv:2608.26669v12026Deep Imitation Learning for Complex Manipulation Tasks from Virtual Reality Teleoperation
Tianhao Zhang, Zoe McCarthy, Owen Jow +4
cs.LGcs.ROarXiv:1710.04615v22017Summaries:한국어FLARE: A Failure-Aware Framework for Autonomous Correction and Recovery in Visual-Language Robotic Manipulation
Ganlong Zhao, Zijia Tang, Xingping Chen +3
cs.ROarXiv:2608.26645v12026GPT-Driver: Learning to Drive with GPT
Jiageng Mao, Yuxi Qian, Junjie Ye +2
cs.CVcs.AIcs.CLarXiv:2310.01415v32023Making Sense of Vision and Touch: Self-Supervised Learning of Multimodal Representations for Contact-Rich Tasks
Michelle A. Lee, Yuke Zhu, Krishnan Srinivasan +5
cs.ROcs.AIcs.LGarXiv:1810.10191v22018Rapid On-Robot Learning for Dynamic Manipulation Skills: Robot Juggling
Taeyoon Lee, Chunpeng Wang, Christopher G. Atkeson +2
cs.ROarXiv:2608.26800v12026