Every paper with a summary

Every arXiv paper Paperlayer has summarized so far. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

20,641 to 20,700 of 61,393

  1. Adaptive Gradient Descent without Descent

    Yura Malitsky, Konstantin Mishchenko

    math.OCcs.LGmath.NAarXiv:1910.09529v22019
  2. Interactive Post-Training for Vision-Language-Action Models

    Shuhan Tan, Kairan Dou, Yue Zhao +1

    cs.LGcs.AIcs.CVarXiv:2505.17016v12025
  3. Self-supervised Video Object Segmentation by Motion Grouping

    Charig Yang, Hala Lamdouar, Erika Lu +2

    cs.CVcs.LGarXiv:2104.07658v22021
  4. AgenTracer: Who Is Inducing Failure in the LLM Agentic Systems?

    Guibin Zhang, Junhao Wang, Junjie Chen +3

    cs.CLcs.MAarXiv:2509.03312v22025
  5. Denoising Diffusion Bridge Models

    Linqi Zhou, Aaron Lou, Samar Khanna +1

    cs.CVcs.AIarXiv:2309.16948v32023
  6. Real-Time Shape Control of Multi-Segment Soft Robotic Arms Using Koopman Operators with Global and Local Observables

    Jiahe Wang, Eron Ristich, Sultan Haidar Ali +5

    cs.ROeess.SYarXiv:2609.03175v12026
  7. Exploiting the Benefits of V2B Application on Peak Shaving of Data Center Loads

    Arya Joshi, Hamed Haggi, Chinmay Morankar

    eess.SYarXiv:2609.00204v12026
  8. UniIR: Training and Benchmarking Universal Multimodal Information Retrievers

    Cong Wei, Yang Chen, Haonan Chen +5

    cs.CVcs.AIcs.CLarXiv:2311.17136v12023
  9. RoboBrain 2.0 Technical Report

    BAAI RoboBrain Team, Mingyu Cao, Huajie Tan +50

    cs.ROarXiv:2507.02029v52025
  10. Voyager: Long-Range and World-Consistent Video Diffusion for Explorable 3D Scene Generation

    Tianyu Huang, Wangguandong Zheng, Tengfei Wang +8

    cs.CVarXiv:2506.04225v12025
  11. Lumina-DiMOO: An Omni Diffusion Large Language Model for Multi-Modal Generation and Understanding

    Yi Xin, Qi Qin, Siqi Luo +29

    cs.CVarXiv:2510.06308v12025
  12. 14 Examples of How LLMs Can Transform Materials Science and Chemistry: A Reflection on a Large Language Model Hackathon

    Kevin Maik Jablonka, Qianxiang Ai, Alexander Al-Feghali +50

    cond-mat.mtrl-scics.LGphysics.chem-pharXiv:2306.06283v42023
  13. TASED-Net: Temporally-Aggregating Spatial Encoder-Decoder Network for Video Saliency Detection

    Kyle Min, Jason J. Corso

    cs.CVarXiv:1908.05786v12019
  14. Gated Graph Recurrent Neural Networks

    Luana Ruiz, Fernando Gama, Alejandro Ribeiro

    eess.SPcs.LGarXiv:2002.01038v22020
  15. Stride-k Subsampling: Train-Free Audio Token Reduction for Whisper

    Chanhee Cho, Junhyuk Choi, Bugeun Kim

    cs.SDcs.AIarXiv:2608.30927v12026
  16. RealCAD: Towards Real-World Image-to-CAD Reconstruction under Domain Shift and Parameter Bias

    Yihe Sun, Ziyu Lu, Kaihua Tang +1

    cs.CVarXiv:2608.30617v12026
  17. RELIC: Interactive Video World Model with Long-Horizon Memory

    Yicong Hong, Yiqun Mei, Chongjian Ge +11

    cs.CVarXiv:2512.04040v12025
  18. A Large Scale Event-based Detection Dataset for Automotive

    Pierre de Tournemire, Davide Nitti, Etienne Perot +2

    cs.CVcs.LGcs.ROarXiv:2001.08499v32020
  19. A Simple Effective Heuristic for Embedded Mixed-Integer Quadratic Programming

    Reza Takapoui, Nicholas Moehle, Stephen Boyd +1

    math.OCarXiv:1509.08416v12015
  20. Latent Visual Reasoning

    Bangzheng Li, Ximeng Sun, Jiang Liu +7

    cs.CVcs.CLarXiv:2509.24251v22025
  21. Exponential Gaps Between Intuitionistic Linear Extended Frege Systems

    Amirhossein Akbar Tabatabai

    cs.LOmath.LOarXiv:2609.00422v12026
  22. Provably Efficient Safe Exploration via Primal-Dual Policy Optimization

    Dongsheng Ding, Xiaohan Wei, Zhuoran Yang +2

    cs.LGmath.OCstat.MLarXiv:2003.00534v22020
  23. DCFace: Synthetic Face Generation with Dual Condition Diffusion Model

    Minchul Kim, Feng Liu, Anil Jain +1

    cs.CVarXiv:2304.07060v12023
  24. The Price of Remembering: A Calibrated Energy Law for Computation

    Mohamed Amine Bergach

    cs.PFcs.ARcs.LOarXiv:2609.00744v22026
  25. Fast Video Generation with Sliding Tile Attention

    Peiyuan Zhang, Yongqi Chen, Runlong Su +4

    cs.CVarXiv:2502.04507v32025
  26. LayoutTransformer: Layout Generation and Completion with Self-attention

    Kamal Gupta, Justin Lazarow, Alessandro Achille +3

    cs.CVcs.LGarXiv:2006.14615v22020
  27. Jointly Predicting Predicates and Arguments in Neural Semantic Role Labeling

    Luheng He, Kenton Lee, Omer Levy +1

    cs.CLarXiv:1805.04787v22018
  28. OS-Harm: A Benchmark for Measuring Safety of Computer Use Agents

    Thomas Kuntz, Agatha Duzan, Hao Zhao +4

    cs.SEcs.LGarXiv:2506.14866v22025
  29. Compositional Generalization and Natural Language Variation: Can a Semantic Parsing Approach Handle Both?

    Peter Shaw, Ming-Wei Chang, Panupong Pasupat +1

    cs.CLarXiv:2010.12725v22020
  30. OmniHuman-1: Rethinking the Scaling-Up of One-Stage Conditioned Human Animation Models

    Gaojie Lin, Jianwen Jiang, Jiaqi Yang +2

    cs.CVarXiv:2502.01061v32025
  31. Are You Thinking What I am Thinking? : Examining Conceptual Separation in Neural Architectures

    Jaee Ponde, Roshni Agarwal, Subhashis Banerjee

    cs.LGcs.AIarXiv:2609.00764v12026
  32. Beyond Periodicity: Towards a Unifying Framework for Activations in Coordinate-MLPs

    Sameera Ramasinghe, Simon Lucey

    cs.LGarXiv:2111.15135v22021
  33. TempFlow-GRPO: When Timing Matters for GRPO in Flow Models

    Xiaoxuan He, Siming Fu, Yuke Zhao +5

    cs.CVarXiv:2508.04324v42025
  34. Medical Hallucinations in Foundation Models and Their Impact on Healthcare

    Yubin Kim, Hyewon Jeong, Shan Chen +24

    cs.CLcs.AIcs.CYarXiv:2503.05777v22025
  35. Street Scene: A new dataset and evaluation protocol for video anomaly detection

    Bharathkumar Ramachandra, Michael Jones

    cs.CVarXiv:1902.05872v32019
  36. MusGU+: Toward a Musician-Centered Evaluation Framework and Discovery Tool for Generative Music AI

    Laura Ibáñez-Martínez, Roser Batlle-Roca, Xavier Serra +1

    cs.SDcs.AIcs.CYarXiv:2608.30940v12026
  37. DriveTransformer: Unified Transformer for Scalable End-to-End Autonomous Driving

    Xiaosong Jia, Junqi You, Zhiyuan Zhang +1

    cs.LGcs.CVcs.ROarXiv:2503.07656v22025
  38. A Richly Annotated Dataset for Pedestrian Attribute Recognition

    Dangwei Li, Zhang Zhang, Xiaotang Chen +2

    cs.CVarXiv:1603.07054v32016
  39. Line-profile tomography of exoplanet transits -- II. A gas-giant planet transiting a rapidly-rotating A5 star

    A. Collier Cameron, E. Guenther, B. Smalley +16

    astro-ph.EParXiv:1004.4551v12010
  40. GoalFlow: Goal-Driven Flow Matching for Multimodal Trajectories Generation in End-to-End Autonomous Driving

    Zebin Xing, Xingyu Zhang, Yang Hu +5

    cs.CVarXiv:2503.05689v62025
  41. Alleviating Over-segmentation Errors by Detecting Action Boundaries

    Yuchi Ishikawa, Seito Kasai, Yoshimitsu Aoki +1

    cs.CVarXiv:2007.06866v12020
  42. Natural Image Matting via Guided Contextual Attention

    Yaoyi Li, Hongtao Lu

    cs.CVarXiv:2001.04069v12020
  43. Unifying Conformal Language Tasks with In-Context Ensembles

    Xiao Shi Huang, Chen-Yuan Lin, Bruce Kuwahara +2

    cs.CLcs.LGstat.MLarXiv:2609.03005v12026
  44. Roadmap on Atomtronics: State of the art and perspective

    L. Amico, M. Boshier, G. Birkl +56

    cond-mat.quant-gasquant-pharXiv:2008.04439v52020
  45. Are VLMs Ready for Autonomous Driving? An Empirical Study from the Reliability, Data, and Metric Perspectives

    Shaoyuan Xie, Lingdong Kong, Yuhao Dong +5

    cs.CVcs.ROarXiv:2501.04003v12025
  46. Temporal-Relational CrossTransformers for Few-Shot Action Recognition

    Toby Perrett, Alessandro Masullo, Tilo Burghardt +2

    cs.CVarXiv:2101.06184v32021
  47. The Geometry of Ignorance: LLMs Know When to Temper Bayesian Priors

    Toni J. B. Liu, Jiajun Bao, Yizhou Liu +4

    cs.LGcs.AIcs.CLarXiv:2609.02959v12026
  48. Adaptive multiscale model reduction with Generalized Multiscale Finite Element Methods

    Eric Chung, Yalchin Efendiev, Thomas Y. Hou

    math.NAarXiv:1604.08312v12016
  49. Enhanced imaging of microcalcifications in digital breast tomosynthesis through improved image-reconstruction algorithms

    Emil Y. Sidky, Xiaochuan Pan, Ingrid S. Reiser +3

    physics.med-pharXiv:0904.1016v12009
  50. PixelDiT: Pixel Diffusion Transformers for Image Generation

    Yongsheng Yu, Wei Xiong, Weili Nie +3

    cs.CVarXiv:2511.20645v22025
  51. PARTFIELD: Learning 3D Feature Fields for Part Segmentation and Beyond

    Minghua Liu, Mikaela Angelina Uy, Donglai Xiang +4

    cs.CVarXiv:2504.11451v12025
  52. Reinforcement Learning in Economics and Finance

    Arthur Charpentier, Romuald Elie, Carl Remlinger

    econ.THcs.LGq-fin.CParXiv:2003.10014v12020
  53. AA-CLIP: Enhancing Zero-shot Anomaly Detection via Anomaly-Aware CLIP

    Wenxin Ma, Xu Zhang, Qingsong Yao +6

    cs.CVcs.AIarXiv:2503.06661v12025
  54. Advances and Challenges in Foundation Agents: From Brain-Inspired Intelligence to Evolutionary, Collaborative, and Safe Systems

    Bang Liu, Xinfeng Li, Jiayi Zhang +45

    cs.AIarXiv:2504.01990v22025
  55. Machine Learning Methods for Cancer Classification Using Gene Expression Data: A Review

    Fadi Alharbi, Aleksandar Vakanski

    cs.LGarXiv:2301.12222v12023
  56. Evidence-Guided Detection, Localization and Explanation for Text-Centric Image Forensics

    Peifeng Liu, Bin Li, Qingsong Zhang +3

    cs.CVarXiv:2609.02097v12026
  57. Recursive Language Models

    Alex L. Zhang, Tim Kraska, Omar Khattab

    cs.AIcs.CLarXiv:2512.24601v32025
  58. A computationally efficient robust model predictive control framework for uncertain nonlinear systems -- extended version

    Johannes Köhler, Raffaele Soloperto, Matthias A. Müller +1

    eess.SYarXiv:1910.12081v22019
  59. SoftCoT: Soft Chain-of-Thought for Efficient Reasoning with LLMs

    Yige Xu, Xu Guo, Zhiwei Zeng +1

    cs.CLarXiv:2502.12134v22025
  60. Orthogonal Ensembles and Tested Explanations for Performer-Independent Body-Motion Emotion Recognition

    Naoto Nishida, Yoshio Ishiguro

    cs.CVcs.HCcs.LGarXiv:2609.02510v12026