Every paper with a summary

Every arXiv paper Paperlayer has summarized so far. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

20,521 to 20,580 of 61,260

  1. TASED-Net: Temporally-Aggregating Spatial Encoder-Decoder Network for Video Saliency Detection

    Kyle Min, Jason J. Corso

    cs.CVarXiv:1908.05786v12019
  2. Gated Graph Recurrent Neural Networks

    Luana Ruiz, Fernando Gama, Alejandro Ribeiro

    eess.SPcs.LGarXiv:2002.01038v22020
  3. Stride-k Subsampling: Train-Free Audio Token Reduction for Whisper

    Chanhee Cho, Junhyuk Choi, Bugeun Kim

    cs.SDcs.AIarXiv:2608.30927v12026
  4. RealCAD: Towards Real-World Image-to-CAD Reconstruction under Domain Shift and Parameter Bias

    Yihe Sun, Ziyu Lu, Kaihua Tang +1

    cs.CVarXiv:2608.30617v12026
  5. RELIC: Interactive Video World Model with Long-Horizon Memory

    Yicong Hong, Yiqun Mei, Chongjian Ge +11

    cs.CVarXiv:2512.04040v12025
  6. A Large Scale Event-based Detection Dataset for Automotive

    Pierre de Tournemire, Davide Nitti, Etienne Perot +2

    cs.CVcs.LGcs.ROarXiv:2001.08499v32020
  7. A Simple Effective Heuristic for Embedded Mixed-Integer Quadratic Programming

    Reza Takapoui, Nicholas Moehle, Stephen Boyd +1

    math.OCarXiv:1509.08416v12015
  8. Latent Visual Reasoning

    Bangzheng Li, Ximeng Sun, Jiang Liu +7

    cs.CVcs.CLarXiv:2509.24251v22025
  9. Exponential Gaps Between Intuitionistic Linear Extended Frege Systems

    Amirhossein Akbar Tabatabai

    cs.LOmath.LOarXiv:2609.00422v12026
  10. Provably Efficient Safe Exploration via Primal-Dual Policy Optimization

    Dongsheng Ding, Xiaohan Wei, Zhuoran Yang +2

    cs.LGmath.OCstat.MLarXiv:2003.00534v22020
  11. DCFace: Synthetic Face Generation with Dual Condition Diffusion Model

    Minchul Kim, Feng Liu, Anil Jain +1

    cs.CVarXiv:2304.07060v12023
  12. The Price of Remembering: A Calibrated Energy Law for Computation

    Mohamed Amine Bergach

    cs.PFcs.ARcs.LOarXiv:2609.00744v22026
  13. Fast Video Generation with Sliding Tile Attention

    Peiyuan Zhang, Yongqi Chen, Runlong Su +4

    cs.CVarXiv:2502.04507v32025
  14. LayoutTransformer: Layout Generation and Completion with Self-attention

    Kamal Gupta, Justin Lazarow, Alessandro Achille +3

    cs.CVcs.LGarXiv:2006.14615v22020
  15. Jointly Predicting Predicates and Arguments in Neural Semantic Role Labeling

    Luheng He, Kenton Lee, Omer Levy +1

    cs.CLarXiv:1805.04787v22018
  16. OS-Harm: A Benchmark for Measuring Safety of Computer Use Agents

    Thomas Kuntz, Agatha Duzan, Hao Zhao +4

    cs.SEcs.LGarXiv:2506.14866v22025
  17. Compositional Generalization and Natural Language Variation: Can a Semantic Parsing Approach Handle Both?

    Peter Shaw, Ming-Wei Chang, Panupong Pasupat +1

    cs.CLarXiv:2010.12725v22020
  18. OmniHuman-1: Rethinking the Scaling-Up of One-Stage Conditioned Human Animation Models

    Gaojie Lin, Jianwen Jiang, Jiaqi Yang +2

    cs.CVarXiv:2502.01061v32025
  19. Are You Thinking What I am Thinking? : Examining Conceptual Separation in Neural Architectures

    Jaee Ponde, Roshni Agarwal, Subhashis Banerjee

    cs.LGcs.AIarXiv:2609.00764v12026
  20. Beyond Periodicity: Towards a Unifying Framework for Activations in Coordinate-MLPs

    Sameera Ramasinghe, Simon Lucey

    cs.LGarXiv:2111.15135v22021
  21. TempFlow-GRPO: When Timing Matters for GRPO in Flow Models

    Xiaoxuan He, Siming Fu, Yuke Zhao +5

    cs.CVarXiv:2508.04324v42025
  22. Medical Hallucinations in Foundation Models and Their Impact on Healthcare

    Yubin Kim, Hyewon Jeong, Shan Chen +24

    cs.CLcs.AIcs.CYarXiv:2503.05777v22025
  23. Street Scene: A new dataset and evaluation protocol for video anomaly detection

    Bharathkumar Ramachandra, Michael Jones

    cs.CVarXiv:1902.05872v32019
  24. MusGU+: Toward a Musician-Centered Evaluation Framework and Discovery Tool for Generative Music AI

    Laura Ibáñez-Martínez, Roser Batlle-Roca, Xavier Serra +1

    cs.SDcs.AIcs.CYarXiv:2608.30940v12026
  25. DriveTransformer: Unified Transformer for Scalable End-to-End Autonomous Driving

    Xiaosong Jia, Junqi You, Zhiyuan Zhang +1

    cs.LGcs.CVcs.ROarXiv:2503.07656v22025
  26. A Richly Annotated Dataset for Pedestrian Attribute Recognition

    Dangwei Li, Zhang Zhang, Xiaotang Chen +2

    cs.CVarXiv:1603.07054v32016
  27. Line-profile tomography of exoplanet transits -- II. A gas-giant planet transiting a rapidly-rotating A5 star

    A. Collier Cameron, E. Guenther, B. Smalley +16

    astro-ph.EParXiv:1004.4551v12010
  28. GoalFlow: Goal-Driven Flow Matching for Multimodal Trajectories Generation in End-to-End Autonomous Driving

    Zebin Xing, Xingyu Zhang, Yang Hu +5

    cs.CVarXiv:2503.05689v62025
  29. Alleviating Over-segmentation Errors by Detecting Action Boundaries

    Yuchi Ishikawa, Seito Kasai, Yoshimitsu Aoki +1

    cs.CVarXiv:2007.06866v12020
  30. Natural Image Matting via Guided Contextual Attention

    Yaoyi Li, Hongtao Lu

    cs.CVarXiv:2001.04069v12020
  31. Unifying Conformal Language Tasks with In-Context Ensembles

    Xiao Shi Huang, Chen-Yuan Lin, Bruce Kuwahara +2

    cs.CLcs.LGstat.MLarXiv:2609.03005v12026
  32. Roadmap on Atomtronics: State of the art and perspective

    L. Amico, M. Boshier, G. Birkl +56

    cond-mat.quant-gasquant-pharXiv:2008.04439v52020
  33. Are VLMs Ready for Autonomous Driving? An Empirical Study from the Reliability, Data, and Metric Perspectives

    Shaoyuan Xie, Lingdong Kong, Yuhao Dong +5

    cs.CVcs.ROarXiv:2501.04003v12025
  34. Temporal-Relational CrossTransformers for Few-Shot Action Recognition

    Toby Perrett, Alessandro Masullo, Tilo Burghardt +2

    cs.CVarXiv:2101.06184v32021
  35. The Geometry of Ignorance: LLMs Know When to Temper Bayesian Priors

    Toni J. B. Liu, Jiajun Bao, Yizhou Liu +4

    cs.LGcs.AIcs.CLarXiv:2609.02959v12026
  36. Adaptive multiscale model reduction with Generalized Multiscale Finite Element Methods

    Eric Chung, Yalchin Efendiev, Thomas Y. Hou

    math.NAarXiv:1604.08312v12016
  37. Enhanced imaging of microcalcifications in digital breast tomosynthesis through improved image-reconstruction algorithms

    Emil Y. Sidky, Xiaochuan Pan, Ingrid S. Reiser +3

    physics.med-pharXiv:0904.1016v12009
  38. PixelDiT: Pixel Diffusion Transformers for Image Generation

    Yongsheng Yu, Wei Xiong, Weili Nie +3

    cs.CVarXiv:2511.20645v22025
  39. PARTFIELD: Learning 3D Feature Fields for Part Segmentation and Beyond

    Minghua Liu, Mikaela Angelina Uy, Donglai Xiang +4

    cs.CVarXiv:2504.11451v12025
  40. Reinforcement Learning in Economics and Finance

    Arthur Charpentier, Romuald Elie, Carl Remlinger

    econ.THcs.LGq-fin.CParXiv:2003.10014v12020
  41. AA-CLIP: Enhancing Zero-shot Anomaly Detection via Anomaly-Aware CLIP

    Wenxin Ma, Xu Zhang, Qingsong Yao +6

    cs.CVcs.AIarXiv:2503.06661v12025
  42. Advances and Challenges in Foundation Agents: From Brain-Inspired Intelligence to Evolutionary, Collaborative, and Safe Systems

    Bang Liu, Xinfeng Li, Jiayi Zhang +45

    cs.AIarXiv:2504.01990v22025
  43. Machine Learning Methods for Cancer Classification Using Gene Expression Data: A Review

    Fadi Alharbi, Aleksandar Vakanski

    cs.LGarXiv:2301.12222v12023
  44. Evidence-Guided Detection, Localization and Explanation for Text-Centric Image Forensics

    Peifeng Liu, Bin Li, Qingsong Zhang +3

    cs.CVarXiv:2609.02097v12026
  45. Recursive Language Models

    Alex L. Zhang, Tim Kraska, Omar Khattab

    cs.AIcs.CLarXiv:2512.24601v32025
  46. A computationally efficient robust model predictive control framework for uncertain nonlinear systems -- extended version

    Johannes Köhler, Raffaele Soloperto, Matthias A. Müller +1

    eess.SYarXiv:1910.12081v22019
  47. SoftCoT: Soft Chain-of-Thought for Efficient Reasoning with LLMs

    Yige Xu, Xu Guo, Zhiwei Zeng +1

    cs.CLarXiv:2502.12134v22025
  48. Orthogonal Ensembles and Tested Explanations for Performer-Independent Body-Motion Emotion Recognition

    Naoto Nishida, Yoshio Ishiguro

    cs.CVcs.HCcs.LGarXiv:2609.02510v12026
  49. LLMs Can Easily Learn to Reason from Demonstrations Structure, not content, is what matters!

    Dacheng Li, Shiyi Cao, Tyler Griggs +9

    cs.AIarXiv:2502.07374v22025
  50. HAMSTER: Hierarchical Action Models For Open-World Robot Manipulation

    Yi Li, Yuquan Deng, Jesse Zhang +9

    cs.ROcs.AIcs.CVarXiv:2502.05485v42025
  51. Data Science in Statistics Curricula: Preparing Students to "Think with Data"

    Johanna Hardin, Roger Hoerl, Nicholas J. Horton +1

    stat.OTarXiv:1410.3127v32014
  52. SoK: Where Do Flow Labels Come From? Auditing Label Provenance in Encrypted Traffic Benchmarks

    Sizhe Huang, Shujie Yang

    cs.NIcs.LGarXiv:2609.02140v12026
  53. Multimodal Recommender Systems: A Survey

    Qidong Liu, Jiaxi Hu, Yutian Xiao +5

    cs.IRcs.AIarXiv:2302.03883v22023
  54. Learning From Labeled And Unlabeled Data: An Empirical Study Across Techniques And Domains

    N. V. Chawla, Grigoris Karakoulas

    cs.LGarXiv:1109.2047v12011
  55. FIRE4, LiteRed and accompanying tools to solve integration by parts relations

    A. V. Smirnov, V. A. Smirnov

    hep-pharXiv:1302.5885v12013
  56. All Pure Bipartite Entangled States can be Self-Tested

    Andrea Coladangelo, Koon Tong Goh, Valerio Scarani

    quant-pharXiv:1611.08062v22016
  57. Building a Framework for Predictive Science

    Michael M. McKerns, Leif Strand, Tim Sullivan +2

    cs.MScs.DCcs.DMarXiv:1202.1056v12012
  58. Incremental Pooled LLM Evaluation for Cost-Effective Retrieval Model Selection

    Max Nelson, Hanoz Bhathena, Aviral Joshi +1

    cs.IRcs.CLarXiv:2609.02745v12026
  59. Can LLMs Replace Human Evaluators? An Empirical Study of LLM-as-a-Judge in Software Engineering

    Ruiqi Wang, Jiyu Guo, Cuiyun Gao +3

    cs.SEcs.AIarXiv:2502.06193v32025
  60. LLMs Accelerate Annotation for Medical Information Extraction

    Akshay Goel, Almog Gueta, Omry Gilon +10

    cs.CLcs.AIcs.LGarXiv:2312.02296v12023