Every paper with a summary

Every arXiv paper Paperlayer has summarized so far. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

53,461 to 53,520 of 61,271

  1. Spurious Rewards Paradox: Mechanistically Understanding How RLVR Activates Memorization Shortcuts in LLMs

    Lecheng Yan, Ruizhe Li, Guanhua Chen +5

    cs.LGcs.CLarXiv:2601.11061v22026
  2. VGGSound: A Large-scale Audio-Visual Dataset

    Honglie Chen, Weidi Xie, Andrea Vedaldi +1

    cs.CVcs.SDeess.ASarXiv:2004.14368v22020
  3. MA-EgoQA: Question Answering over Egocentric Videos from Multiple Embodied Agents

    Kangsan Kim, Yanlai Yang, Suji Kim +4

    cs.CVcs.AIarXiv:2603.09827v22026
  4. FrankenMotion: Part-level Human Motion Generation and Composition

    Chuqiao Li, Xianghui Xie, Yong Cao +2

    cs.CVarXiv:2601.10909v12026
  5. dVoting: Fast Voting for dLLMs

    Sicheng Feng, Zigeng Chen, Xinyin Ma +2

    cs.CLcs.AIarXiv:2602.12153v12026
  6. Progress measures for grokking via mechanistic interpretability

    Neel Nanda, Lawrence Chan, Tom Lieberum +2

    cs.LGcs.AIarXiv:2301.05217v32023
  7. Attend Before Attention: Efficient and Scalable Video Understanding via Autoregressive Gazing

    Baifeng Shi, Stephanie Fu, Long Lian +10

    cs.CVarXiv:2603.12254v12026
  8. CARAFE: Content-Aware ReAssembly of FEatures

    Jiaqi Wang, Kai Chen, Rui Xu +3

    cs.CVarXiv:1905.02188v32019
  9. Long-term Temporal Convolutions for Action Recognition

    Gül Varol, Ivan Laptev, Cordelia Schmid

    cs.CVarXiv:1604.04494v22016
  10. A Survey on Neural Network Interpretability

    Yu Zhang, Peter Tiňo, Aleš Leonardis +1

    cs.LGcs.AIarXiv:2012.14261v32020
  11. NewsQA: A Machine Comprehension Dataset

    Adam Trischler, Tong Wang, Xingdi Yuan +4

    cs.CLcs.AIarXiv:1611.09830v32016
  12. AgentDS Technical Report: Benchmarking the Future of Human-AI Collaboration in Domain-Specific Data Science

    An Luo, Jin Du, Xun Xian +12

    cs.LGcs.AIstat.MEarXiv:2603.19005v32026
  13. Speeding-up Convolutional Neural Networks Using Fine-tuned CP-Decomposition

    Vadim Lebedev, Yaroslav Ganin, Maksim Rakhuba +2

    cs.CVcs.LGarXiv:1412.6553v32014
  14. Chronos: Learning the Language of Time Series

    Abdul Fatir Ansari, Lorenzo Stella, Caner Turkmen +15

    cs.LGcs.AIarXiv:2403.07815v32024
  15. Efficient Neural Network Robustness Certification with General Activation Functions

    Huan Zhang, Tsui-Wei Weng, Pin-Yu Chen +2

    cs.LGcs.CRstat.MLarXiv:1811.00866v12018
  16. From the Quantum Approximate Optimization Algorithm to a Quantum Alternating Operator Ansatz

    Stuart Hadfield, Zhihui Wang, Bryan O'Gorman +3

    quant-pharXiv:1709.03489v22017
  17. Time-Contrastive Networks: Self-Supervised Learning from Video

    Pierre Sermanet, Corey Lynch, Yevgen Chebotar +4

    cs.CVcs.ROarXiv:1704.06888v32017
  18. Grounding World Simulation Models in a Real-World Metropolis

    Junyoung Seo, Hyunwook Choi, Minkyung Kwon +10

    cs.CVarXiv:2603.15583v12026
  19. CP-nets: A Tool for Representing and Reasoning withConditional Ceteris Paribus Preference Statements

    C. Boutilier, R. I. Brafman, C. Domshlak +2

    cs.AIarXiv:1107.0023v12011
  20. A Deep Relevance Matching Model for Ad-hoc Retrieval

    Jiafeng Guo, Yixing Fan, Qingyao Ai +1

    cs.IRarXiv:1711.08611v12017
  21. Acquisition of Localization Confidence for Accurate Object Detection

    Borui Jiang, Ruixuan Luo, Jiayuan Mao +2

    cs.CVarXiv:1807.11590v12018
  22. ThinkJEPA: Empowering Latent World Models with Large Vision-Language Reasoning Model

    Haichao Zhang, Yijiang Li, Shwai He +5

    cs.CVcs.AIcs.CLarXiv:2603.22281v22026
  23. LongCat-Flash-Prover: Advancing Native Formal Reasoning via Agentic Tool-Integrated Reinforcement Learning

    Jianing Wang, Jianfei Zhang, Qi Guo +24

    cs.AIcs.CLarXiv:2603.21065v12026
  24. RTAB-Map as an Open-Source Lidar and Visual SLAM Library for Large-Scale and Long-Term Online Operation

    Mathieu Labbé, François Michaud

    cs.ROarXiv:2403.06341v12024
  25. C-RADIOv4 (Tech Report)

    Mike Ranzinger, Greg Heinrich, Collin McCarthy +4

    cs.CVarXiv:2601.17237v12026
  26. Memory Fusion Network for Multi-view Sequential Learning

    Amir Zadeh, Paul Pu Liang, Navonil Mazumder +3

    cs.LGcs.AIarXiv:1802.00927v12018
  27. Two at Once: Enhancing Learning and Generalization Capacities via IBN-Net

    Xingang Pan, Ping Luo, Jianping Shi +1

    cs.CVarXiv:1807.09441v32018
  28. Position: Agentic Evolution is the Path to Evolving LLMs

    Minhua Lin, Hanqing Lu, Zhan Shi +11

    cs.AIarXiv:2602.00359v22026
  29. Image Augmentation Is All You Need: Regularizing Deep Reinforcement Learning from Pixels

    Ilya Kostrikov, Denis Yarats, Rob Fergus

    cs.LGcs.CVeess.IVarXiv:2004.13649v42020
  30. BRITS: Bidirectional Recurrent Imputation for Time Series

    Wei Cao, Dong Wang, Jian Li +3

    cs.LGstat.MLarXiv:1805.10572v12018
  31. SinGAN: Learning a Generative Model from a Single Natural Image

    Tamar Rott Shaham, Tali Dekel, Tomer Michaeli

    cs.CVarXiv:1905.01164v22019
  32. AI Fairness 360: An Extensible Toolkit for Detecting, Understanding, and Mitigating Unwanted Algorithmic Bias

    Rachel K. E. Bellamy, Kuntal Dey, Michael Hind +15

    cs.AIarXiv:1810.01943v12018
  33. Imaging Time-Series to Improve Classification and Imputation

    Zhiguang Wang, Tim Oates

    cs.LGcs.NEstat.MLarXiv:1506.00327v12015
  34. The Confidence Dichotomy: Analyzing and Mitigating Miscalibration in Tool-Use Agents

    Weihao Xuan, Qingcheng Zeng, Heli Qi +3

    cs.CLarXiv:2601.07264v12026
  35. Deeper and Wider Siamese Networks for Real-Time Visual Tracking

    Zhipeng Zhang, Houwen Peng

    cs.CVarXiv:1901.01660v32019
  36. Behavior Knowledge Merge in Reinforced Agentic Models

    Xiangchi Yuan, Dachuan Shi, Chunhui Zhang +4

    cs.LGarXiv:2601.13572v12026
  37. PROGRESSLM: Towards Progress Reasoning in Vision-Language Models

    Jianshu Zhang, Chengxuan Qian, Haosen Sun +4

    cs.CVcs.CLarXiv:2601.15224v22026
  38. A Subgoal-driven Framework for Improving Long-Horizon LLM Agents

    Taiyi Wang, Sian Gooding, Florian Hartmann +2

    cs.AIcs.LGcs.MAarXiv:2603.19685v12026
  39. A Simple and Effective Pruning Approach for Large Language Models

    Mingjie Sun, Zhuang Liu, Anna Bair +1

    cs.CLcs.AIcs.LGarXiv:2306.11695v32023
  40. FireRed-OCR Technical Report

    Hao Wu, Haoran Lou, Xinyue Li +19

    cs.CVeess.IVarXiv:2603.01840v12026
  41. WorldCam: Interactive Autoregressive 3D Gaming Worlds with Camera Pose as a Unifying Geometric Representation

    Jisu Nam, Yicong Hong, Chun-Hao Paul Huang +9

    cs.CVarXiv:2603.16871v12026
  42. Understanding Reasoning in LLMs through Strategic Information Allocation under Uncertainty

    Jeonghye Kim, Xufang Luo, Minbeom Kim +3

    cs.AIcs.LGarXiv:2603.15500v22026
  43. LRM: Large Reconstruction Model for Single Image to 3D

    Yicong Hong, Kai Zhang, Jiuxiang Gu +7

    cs.CVcs.AIcs.GRarXiv:2311.04400v22023
  44. A Discriminatively Learned CNN Embedding for Person Re-identification

    Zhedong Zheng, Liang Zheng, Yi Yang

    cs.CVarXiv:1611.05666v22016
  45. Quantum random access memory

    Vittorio Giovannetti, Seth Lloyd, Lorenzo Maccone

    quant-pharXiv:0708.1879v22007
  46. Unified Pre-training for Program Understanding and Generation

    Wasi Uddin Ahmad, Saikat Chakraborty, Baishakhi Ray +1

    cs.CLcs.PLarXiv:2103.06333v22021
  47. Image Generation with a Sphere Encoder

    Kaiyu Yue, Menglin Jia, Ji Hou +1

    cs.CVarXiv:2602.15030v12026
  48. Hierarchical Representations for Efficient Architecture Search

    Hanxiao Liu, Karen Simonyan, Oriol Vinyals +2

    cs.LGcs.CVcs.NEarXiv:1711.00436v22017
  49. A Survey on Large Language Models for Recommendation

    Likang Wu, Zhi Zheng, Zhaopeng Qiu +9

    cs.IRcs.AIarXiv:2305.19860v52023
  50. Graph R-CNN for Scene Graph Generation

    Jianwei Yang, Jiasen Lu, Stefan Lee +2

    cs.CVcs.LGarXiv:1808.00191v12018
  51. U-Mamba: Enhancing Long-range Dependency for Biomedical Image Segmentation

    Jun Ma, Feifei Li, Bo Wang

    eess.IVcs.CVcs.LGarXiv:2401.04722v12024
  52. Ragas: Automated Evaluation of Retrieval Augmented Generation

    Shahul Es, Jithin James, Luis Espinosa-Anke +1

    cs.CLarXiv:2309.15217v22023
  53. Multi-Vector Index Compression in Any Modality

    Hanxiang Qin, Alexander Martin, Rohan Jha +3

    cs.IRcs.CLcs.CVarXiv:2602.21202v12026
  54. OCRVerse: Towards Holistic OCR in End-to-End Vision-Language Models

    Yufeng Zhong, Lei Chen, Xuanle Zhao +7

    cs.CVarXiv:2601.21639v22026
  55. Early Convolutions Help Transformers See Better

    Tete Xiao, Mannat Singh, Eric Mintun +3

    cs.CVarXiv:2106.14881v32021
  56. Adversarial Attacks on Neural Network Policies

    Sandy Huang, Nicolas Papernot, Ian Goodfellow +2

    cs.LGcs.CRstat.MLarXiv:1702.02284v12017
  57. RealMem: Benchmarking LLMs in Real-World Memory-Driven Interaction

    Haonan Bian, Zhiyuan Yao, Sen Hu +7

    cs.CLcs.AIarXiv:2601.06966v12026
  58. Diagnosing the Reliability of LLM-as-a-Judge via Item Response Theory

    Junhyuk Choi, Sohhyung Park, Chanhee Cho +2

    cs.AIarXiv:2602.00521v22026
  59. MOPO: Model-based Offline Policy Optimization

    Tianhe Yu, Garrett Thomas, Lantao Yu +5

    cs.LGcs.AIstat.MLarXiv:2005.13239v62020
  60. Open Problems and Fundamental Limitations of Reinforcement Learning from Human Feedback

    Stephen Casper, Xander Davies, Claudia Shi +29

    cs.AIcs.CLcs.LGarXiv:2307.15217v22023