Every paper with a summary
Every arXiv paper Paperlayer has summarized so far. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
20,641 to 20,700 of 61,393
Adaptive Gradient Descent without Descent
Yura Malitsky, Konstantin Mishchenko
math.OCcs.LGmath.NAarXiv:1910.09529v22019Interactive Post-Training for Vision-Language-Action Models
Shuhan Tan, Kairan Dou, Yue Zhao +1
cs.LGcs.AIcs.CVarXiv:2505.17016v12025Self-supervised Video Object Segmentation by Motion Grouping
Charig Yang, Hala Lamdouar, Erika Lu +2
cs.CVcs.LGarXiv:2104.07658v22021AgenTracer: Who Is Inducing Failure in the LLM Agentic Systems?
Guibin Zhang, Junhao Wang, Junjie Chen +3
cs.CLcs.MAarXiv:2509.03312v22025Denoising Diffusion Bridge Models
Linqi Zhou, Aaron Lou, Samar Khanna +1
cs.CVcs.AIarXiv:2309.16948v32023Real-Time Shape Control of Multi-Segment Soft Robotic Arms Using Koopman Operators with Global and Local Observables
Jiahe Wang, Eron Ristich, Sultan Haidar Ali +5
cs.ROeess.SYarXiv:2609.03175v12026Exploiting the Benefits of V2B Application on Peak Shaving of Data Center Loads
Arya Joshi, Hamed Haggi, Chinmay Morankar
eess.SYarXiv:2609.00204v12026UniIR: Training and Benchmarking Universal Multimodal Information Retrievers
Cong Wei, Yang Chen, Haonan Chen +5
cs.CVcs.AIcs.CLarXiv:2311.17136v12023RoboBrain 2.0 Technical Report
BAAI RoboBrain Team, Mingyu Cao, Huajie Tan +50
cs.ROarXiv:2507.02029v52025Voyager: Long-Range and World-Consistent Video Diffusion for Explorable 3D Scene Generation
Tianyu Huang, Wangguandong Zheng, Tengfei Wang +8
cs.CVarXiv:2506.04225v12025Lumina-DiMOO: An Omni Diffusion Large Language Model for Multi-Modal Generation and Understanding
Yi Xin, Qi Qin, Siqi Luo +29
cs.CVarXiv:2510.06308v1202514 Examples of How LLMs Can Transform Materials Science and Chemistry: A Reflection on a Large Language Model Hackathon
Kevin Maik Jablonka, Qianxiang Ai, Alexander Al-Feghali +50
cond-mat.mtrl-scics.LGphysics.chem-pharXiv:2306.06283v42023TASED-Net: Temporally-Aggregating Spatial Encoder-Decoder Network for Video Saliency Detection
Kyle Min, Jason J. Corso
cs.CVarXiv:1908.05786v12019Gated Graph Recurrent Neural Networks
Luana Ruiz, Fernando Gama, Alejandro Ribeiro
eess.SPcs.LGarXiv:2002.01038v22020Stride-k Subsampling: Train-Free Audio Token Reduction for Whisper
Chanhee Cho, Junhyuk Choi, Bugeun Kim
cs.SDcs.AIarXiv:2608.30927v12026RealCAD: Towards Real-World Image-to-CAD Reconstruction under Domain Shift and Parameter Bias
Yihe Sun, Ziyu Lu, Kaihua Tang +1
cs.CVarXiv:2608.30617v12026RELIC: Interactive Video World Model with Long-Horizon Memory
Yicong Hong, Yiqun Mei, Chongjian Ge +11
cs.CVarXiv:2512.04040v12025A Large Scale Event-based Detection Dataset for Automotive
Pierre de Tournemire, Davide Nitti, Etienne Perot +2
cs.CVcs.LGcs.ROarXiv:2001.08499v32020A Simple Effective Heuristic for Embedded Mixed-Integer Quadratic Programming
Reza Takapoui, Nicholas Moehle, Stephen Boyd +1
math.OCarXiv:1509.08416v12015Latent Visual Reasoning
Bangzheng Li, Ximeng Sun, Jiang Liu +7
cs.CVcs.CLarXiv:2509.24251v22025Exponential Gaps Between Intuitionistic Linear Extended Frege Systems
Amirhossein Akbar Tabatabai
cs.LOmath.LOarXiv:2609.00422v12026Provably Efficient Safe Exploration via Primal-Dual Policy Optimization
Dongsheng Ding, Xiaohan Wei, Zhuoran Yang +2
cs.LGmath.OCstat.MLarXiv:2003.00534v22020DCFace: Synthetic Face Generation with Dual Condition Diffusion Model
Minchul Kim, Feng Liu, Anil Jain +1
cs.CVarXiv:2304.07060v12023The Price of Remembering: A Calibrated Energy Law for Computation
Mohamed Amine Bergach
cs.PFcs.ARcs.LOarXiv:2609.00744v22026Fast Video Generation with Sliding Tile Attention
Peiyuan Zhang, Yongqi Chen, Runlong Su +4
cs.CVarXiv:2502.04507v32025LayoutTransformer: Layout Generation and Completion with Self-attention
Kamal Gupta, Justin Lazarow, Alessandro Achille +3
cs.CVcs.LGarXiv:2006.14615v22020Jointly Predicting Predicates and Arguments in Neural Semantic Role Labeling
Luheng He, Kenton Lee, Omer Levy +1
cs.CLarXiv:1805.04787v22018OS-Harm: A Benchmark for Measuring Safety of Computer Use Agents
Thomas Kuntz, Agatha Duzan, Hao Zhao +4
cs.SEcs.LGarXiv:2506.14866v22025Compositional Generalization and Natural Language Variation: Can a Semantic Parsing Approach Handle Both?
Peter Shaw, Ming-Wei Chang, Panupong Pasupat +1
cs.CLarXiv:2010.12725v22020OmniHuman-1: Rethinking the Scaling-Up of One-Stage Conditioned Human Animation Models
Gaojie Lin, Jianwen Jiang, Jiaqi Yang +2
cs.CVarXiv:2502.01061v32025Are You Thinking What I am Thinking? : Examining Conceptual Separation in Neural Architectures
Jaee Ponde, Roshni Agarwal, Subhashis Banerjee
cs.LGcs.AIarXiv:2609.00764v12026Beyond Periodicity: Towards a Unifying Framework for Activations in Coordinate-MLPs
Sameera Ramasinghe, Simon Lucey
cs.LGarXiv:2111.15135v22021TempFlow-GRPO: When Timing Matters for GRPO in Flow Models
Xiaoxuan He, Siming Fu, Yuke Zhao +5
cs.CVarXiv:2508.04324v42025Medical Hallucinations in Foundation Models and Their Impact on Healthcare
Yubin Kim, Hyewon Jeong, Shan Chen +24
cs.CLcs.AIcs.CYarXiv:2503.05777v22025Street Scene: A new dataset and evaluation protocol for video anomaly detection
Bharathkumar Ramachandra, Michael Jones
cs.CVarXiv:1902.05872v32019MusGU+: Toward a Musician-Centered Evaluation Framework and Discovery Tool for Generative Music AI
Laura Ibáñez-Martínez, Roser Batlle-Roca, Xavier Serra +1
cs.SDcs.AIcs.CYarXiv:2608.30940v12026DriveTransformer: Unified Transformer for Scalable End-to-End Autonomous Driving
Xiaosong Jia, Junqi You, Zhiyuan Zhang +1
cs.LGcs.CVcs.ROarXiv:2503.07656v22025A Richly Annotated Dataset for Pedestrian Attribute Recognition
Dangwei Li, Zhang Zhang, Xiaotang Chen +2
cs.CVarXiv:1603.07054v32016Line-profile tomography of exoplanet transits -- II. A gas-giant planet transiting a rapidly-rotating A5 star
A. Collier Cameron, E. Guenther, B. Smalley +16
astro-ph.EParXiv:1004.4551v12010GoalFlow: Goal-Driven Flow Matching for Multimodal Trajectories Generation in End-to-End Autonomous Driving
Zebin Xing, Xingyu Zhang, Yang Hu +5
cs.CVarXiv:2503.05689v62025Alleviating Over-segmentation Errors by Detecting Action Boundaries
Yuchi Ishikawa, Seito Kasai, Yoshimitsu Aoki +1
cs.CVarXiv:2007.06866v12020Natural Image Matting via Guided Contextual Attention
Yaoyi Li, Hongtao Lu
cs.CVarXiv:2001.04069v12020Unifying Conformal Language Tasks with In-Context Ensembles
Xiao Shi Huang, Chen-Yuan Lin, Bruce Kuwahara +2
cs.CLcs.LGstat.MLarXiv:2609.03005v12026Roadmap on Atomtronics: State of the art and perspective
L. Amico, M. Boshier, G. Birkl +56
cond-mat.quant-gasquant-pharXiv:2008.04439v52020Are VLMs Ready for Autonomous Driving? An Empirical Study from the Reliability, Data, and Metric Perspectives
Shaoyuan Xie, Lingdong Kong, Yuhao Dong +5
cs.CVcs.ROarXiv:2501.04003v12025Temporal-Relational CrossTransformers for Few-Shot Action Recognition
Toby Perrett, Alessandro Masullo, Tilo Burghardt +2
cs.CVarXiv:2101.06184v32021The Geometry of Ignorance: LLMs Know When to Temper Bayesian Priors
Toni J. B. Liu, Jiajun Bao, Yizhou Liu +4
cs.LGcs.AIcs.CLarXiv:2609.02959v12026Adaptive multiscale model reduction with Generalized Multiscale Finite Element Methods
Eric Chung, Yalchin Efendiev, Thomas Y. Hou
math.NAarXiv:1604.08312v12016Enhanced imaging of microcalcifications in digital breast tomosynthesis through improved image-reconstruction algorithms
Emil Y. Sidky, Xiaochuan Pan, Ingrid S. Reiser +3
physics.med-pharXiv:0904.1016v12009PixelDiT: Pixel Diffusion Transformers for Image Generation
Yongsheng Yu, Wei Xiong, Weili Nie +3
cs.CVarXiv:2511.20645v22025PARTFIELD: Learning 3D Feature Fields for Part Segmentation and Beyond
Minghua Liu, Mikaela Angelina Uy, Donglai Xiang +4
cs.CVarXiv:2504.11451v12025Reinforcement Learning in Economics and Finance
Arthur Charpentier, Romuald Elie, Carl Remlinger
econ.THcs.LGq-fin.CParXiv:2003.10014v12020AA-CLIP: Enhancing Zero-shot Anomaly Detection via Anomaly-Aware CLIP
Wenxin Ma, Xu Zhang, Qingsong Yao +6
cs.CVcs.AIarXiv:2503.06661v12025Advances and Challenges in Foundation Agents: From Brain-Inspired Intelligence to Evolutionary, Collaborative, and Safe Systems
Bang Liu, Xinfeng Li, Jiayi Zhang +45
cs.AIarXiv:2504.01990v22025Machine Learning Methods for Cancer Classification Using Gene Expression Data: A Review
Fadi Alharbi, Aleksandar Vakanski
cs.LGarXiv:2301.12222v12023Evidence-Guided Detection, Localization and Explanation for Text-Centric Image Forensics
Peifeng Liu, Bin Li, Qingsong Zhang +3
cs.CVarXiv:2609.02097v12026Recursive Language Models
Alex L. Zhang, Tim Kraska, Omar Khattab
cs.AIcs.CLarXiv:2512.24601v32025A computationally efficient robust model predictive control framework for uncertain nonlinear systems -- extended version
Johannes Köhler, Raffaele Soloperto, Matthias A. Müller +1
eess.SYarXiv:1910.12081v22019SoftCoT: Soft Chain-of-Thought for Efficient Reasoning with LLMs
Yige Xu, Xu Guo, Zhiwei Zeng +1
cs.CLarXiv:2502.12134v22025Orthogonal Ensembles and Tested Explanations for Performer-Independent Body-Motion Emotion Recognition
Naoto Nishida, Yoshio Ishiguro
cs.CVcs.HCcs.LGarXiv:2609.02510v12026