Every paper with a summary

Every arXiv paper Paperlayer has summarized so far. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

57,001 to 57,060 of 61,217

  1. Apple-$π$: Benchmarking Thinking with Video Towards Law-Grounded Physical Intelligence

    Runmao Yao, Kairui Hu, Yukang Cao +11

    cs.CVarXiv:2607.16401v12026
  2. MeanFlowNFT: Bringing Forward-Process RL to Average-Velocity Generators

    Yushi Huang, Xiangxin Zhou, Jun Zhang +2

    cs.CVcs.LGarXiv:2607.15273v22026
  3. Look Before You Leap: Distilling Tree Search into Action Evaluation for Frozen VLA Models

    Xinyi Xie, Zican Hu, Zhanyu Liu +7

    cs.ROarXiv:2607.03751v12026
  4. Bridging Interleaved Multi-Modal Reasoning as a Unified Decision Process

    Zican Hu, Xuyang Hu, Yiming Liu +10

    cs.AIarXiv:2607.03748v12026
  5. ENPIRE: Agentic Robot Policy Self-Improvement in the Real World

    Wenli Xiao, Jia Xie, Tonghe Zhang +14

    cs.AIarXiv:2606.19980v12026
  6. RefGC-SR$^2$: Reference-guided Super-Resolution and Refinement of AI Generated Content

    Jeahun Sung, Dahyeon Kye, Soo Ye Kim +1

    cs.CVarXiv:2606.15158v22026
  7. AutoLLMResearch: Training Research Agents for Automating LLM Experiment Configuration - Learning from Cheap, Optimizing Expensive

    Taicheng Guo, Nitesh V. Chawla, Olaf Wiest +1

    cs.AIcs.CLcs.LGarXiv:2605.11518v22026
  8. Benchmark Everything Everywhere All at Once

    Shiyun Xiong, Dongming Wu, Peiwen Sun +5

    cs.AIarXiv:2606.06462v12026
  9. Synchronization in complex networks

    Alex Arenas, Albert Diaz-Guilera, Jurgen Kurths +2

    physics.soc-pharXiv:0805.2976v32008
  10. DV-World: Benchmarking Data Visualization Agents in Real-World Scenarios

    Jinxiang Meng, Shaoping Huang, Fangyu Lei +17

    cs.CLarXiv:2604.25914v12026
  11. Rethinking Muon Beyond Pretraining: Spectral Failures and High-Pass Remedies for VLA and RLVR

    Chongyu Fan, Gaowen Liu, Mingyi Hong +2

    cs.LGarXiv:2605.19282v12026
  12. $π$-Bench: Evaluating Proactive Personal Assistant Agents in Long-Horizon Workflows

    Haoran Zhang, Luxin Xu, Zhilin Wang +11

    cs.AIarXiv:2605.14678v32026
  13. MSAVBench: Towards Comprehensive and Reliable Evaluation of Multi-Shot Audio-Video Generation

    Yujie Wei, Yujin Han, Zhekai Chen +20

    cs.CVarXiv:2605.20183v42026
  14. Continual Harness: Online Adaptation for Self-Improving Foundation Agents

    Seth Karten, Joel Zhang, Tersoo Upaa +5

    cs.LGcs.AIarXiv:2605.09998v12026
  15. Can RL Teach Long-Horizon Reasoning to LLMs? Expressiveness Is Key

    Tianle Wang, Zhaoyang Wang, Guangchen Lan +4

    cs.AIcs.CLarXiv:2605.06638v32026
  16. Prompt-Activation Duality: Improving Activation Steering via Attention-Level Interventions

    Diancheng Kang, Zheyuan Liu, Ningshan Ma +3

    cs.CLcs.AIarXiv:2605.10664v22026
  17. A Tutorial on Regularized Partial Correlation Networks

    Sacha Epskamp, Eiko I. Fried

    stat.APstat.MEarXiv:1607.01367v92016
  18. Deep multi-scale video prediction beyond mean square error

    Michael Mathieu, Camille Couprie, Yann LeCun

    cs.LGcs.CVstat.MLarXiv:1511.05440v62015
  19. Ego4D: Around the World in 3,000 Hours of Egocentric Video

    Kristen Grauman, Andrew Westbury, Eugene Byrne +82

    cs.CVcs.AIarXiv:2110.07058v32021
  20. AttnGAN: Fine-Grained Text to Image Generation with Attentional Generative Adversarial Networks

    Tao Xu, Pengchuan Zhang, Qiuyuan Huang +4

    cs.CVarXiv:1711.10485v12017
  21. GLM: General Language Model Pretraining with Autoregressive Blank Infilling

    Zhengxiao Du, Yujie Qian, Xiao Liu +4

    cs.CLcs.AIcs.LGarXiv:2103.10360v22021
  22. Person Transfer GAN to Bridge Domain Gap for Person Re-Identification

    Longhui Wei, Shiliang Zhang, Wen Gao +1

    cs.CVarXiv:1711.08565v22017
  23. cuDNN: Efficient Primitives for Deep Learning

    Sharan Chetlur, Cliff Woolley, Philippe Vandermersch +4

    cs.NEcs.LGcs.MSarXiv:1410.0759v32014
  24. The Rise and Potential of Large Language Model Based Agents: A Survey

    Zhiheng Xi, Wenxiang Chen, Xin Guo +26

    cs.AIcs.CLarXiv:2309.07864v32023
  25. Adversarial Examples Are Not Easily Detected: Bypassing Ten Detection Methods

    Nicholas Carlini, David Wagner

    cs.LGcs.CRcs.CVarXiv:1705.07263v22017
  26. Generalized Focal Loss: Learning Qualified and Distributed Bounding Boxes for Dense Object Detection

    Xiang Li, Wenhai Wang, Lijun Wu +5

    cs.CVarXiv:2006.04388v12020
  27. SAGA: A Fast Incremental Gradient Method With Support for Non-Strongly Convex Composite Objectives

    Aaron Defazio, Francis Bach, Simon Lacoste-Julien

    cs.LGmath.OCstat.MLarXiv:1407.0202v32014
  28. Multitask Prompted Training Enables Zero-Shot Task Generalization

    Victor Sanh, Albert Webson, Colin Raffel +38

    cs.LGcs.CLarXiv:2110.08207v32021
  29. Send a SCOUT First: Pre-hoc Reasoning for Adaptive Detector Allocation in Prompt-Injection Defense

    Shuhao Zhang, Jiarui Li, Qi Cao +2

    cs.CRcs.LGarXiv:2605.30837v22026
  30. On Dynamic Mode Decomposition: Theory and Applications

    Jonathan H. Tu, Clarence W. Rowley, Dirk M. Luchtenburg +2

    math.NAphysics.flu-dynarXiv:1312.0041v12013
  31. Natural Adversarial Examples

    Dan Hendrycks, Kevin Zhao, Steven Basart +2

    cs.LGcs.CVstat.MLarXiv:1907.07174v42019
  32. Distilling LLM Feedback for Lean Theorem Proving

    Gaetan Narozniak, Gérard Biau, Rémi Munos +2

    cs.AIarXiv:2605.30861v12026
  33. BEVFormer: Learning Bird's-Eye-View Representation from Multi-Camera Images via Spatiotemporal Transformers

    Zhiqi Li, Wenhai Wang, Hongyang Li +5

    cs.CVarXiv:2203.17270v22022
  34. Semi-Supervised Noise Adaptation: Transferring Knowledge from Noise Domain

    Yuan Yao, Jin Song, Huixia Li +3

    cs.LGarXiv:2606.00558v22026
  35. Covid-19: Automatic detection from X-Ray images utilizing Transfer Learning with Convolutional Neural Networks

    Ioannis D. Apostolopoulos, Tzani Bessiana

    eess.IVcs.CVcs.LGarXiv:2003.11617v12020
  36. Representation over Routing: Diagnosing Temporal Routing Pathologies in Multi-Timescale PPO

    Jing Sun

    cs.LGcs.AIarXiv:2604.13517v42026
  37. Confidence-Adaptive SwiGLU for Mixture-of-Experts

    Shaohua Li, Xiuchao Sui, Xiaobing Sun +4

    cs.LGcs.CLarXiv:2606.00761v12026
  38. GMAN: A Graph Multi-Attention Network for Traffic Prediction

    Chuanpan Zheng, Xiaoliang Fan, Cheng Wang +1

    eess.SPcs.LGarXiv:1911.08415v22019
  39. Ensemble deep learning: A review

    M. A. Ganaie, Minghui Hu, A. K. Malik +2

    cs.LGcs.AIcs.CVarXiv:2104.02395v32021
  40. MMLU-Pro: A More Robust and Challenging Multi-Task Language Understanding Benchmark

    Yubo Wang, Xueguang Ma, Ge Zhang +14

    cs.CLarXiv:2406.01574v62024
  41. Self-supervised Visual Feature Learning with Deep Neural Networks: A Survey

    Longlong Jing, Yingli Tian

    cs.CVarXiv:1902.06162v12019
  42. Constrained Consensus

    Angelia Nedić, Asuman Ozdaglar, Pablo A. Parrilo

    math.OCarXiv:0802.3922v22008
  43. The Shape of Addition: Geometric Structures of Arithmetic in Large Language Models

    Liuyuan Wen, Xun Zhu, Lihao Huang +2

    cs.LGcs.AIarXiv:2606.03645v12026
  44. Functional Attention: From Pairwise Affinities to Functional Correspondences

    Jiefang Xiao, Maolin Gao, Simon Weber +2

    cs.LGarXiv:2605.31559v12026
  45. Honest Lying: Understanding Memory Confabulation in Reflexive Agents

    Prakhar Dixit, Sadia Kamal, Tim Oates

    cs.LGcs.AIarXiv:2605.29463v22026
  46. GradNorm: Gradient Normalization for Adaptive Loss Balancing in Deep Multitask Networks

    Zhao Chen, Vijay Badrinarayanan, Chen-Yu Lee +1

    cs.CVarXiv:1711.02257v42017
  47. PaintBench: Deterministic Evaluation of Precise Visual Editing

    Kai Xu, Ellis Brown, Shrikar Madhu +3

    cs.GRcs.CVcs.LGarXiv:2606.00188v12026
  48. Relational Knowledge Distillation

    Wonpyo Park, Dongju Kim, Yan Lu +1

    cs.CVcs.LGarXiv:1904.05068v22019
  49. SpatialAct: Probing Spatial Reasoning-to-Action Capabilities of VLM Agents in 3D Scenes

    Tianhui Liu, Jie Feng, Zhiheng Zheng +6

    cs.CVcs.AIcs.CLarXiv:2605.31148v12026
  50. Combinatorial Synthesis: Scaling Code RLVR via Atomic Decomposition and Recombination

    Jiasheng Zheng, Boxi Cao, Boxi Yu +6

    cs.CLcs.SEarXiv:2605.31058v12026
  51. Stacked Attention Networks for Image Question Answering

    Zichao Yang, Xiaodong He, Jianfeng Gao +2

    cs.LGcs.CLcs.CVarXiv:1511.02274v22015
  52. Enhancing the Locality and Breaking the Memory Bottleneck of Transformer on Time Series Forecasting

    Shiyang Li, Xiaoyong Jin, Yao Xuan +4

    cs.LGstat.MLarXiv:1907.00235v32019
  53. Deep Mutual Learning

    Ying Zhang, Tao Xiang, Timothy M. Hospedales +1

    cs.CVarXiv:1706.00384v12017
  54. DRAW: A Recurrent Neural Network For Image Generation

    Karol Gregor, Ivo Danihelka, Alex Graves +2

    cs.CVcs.LGcs.NEarXiv:1502.04623v22015
  55. Simple and Deep Graph Convolutional Networks

    Ming Chen, Zhewei Wei, Zengfeng Huang +2

    cs.LGstat.MLarXiv:2007.02133v12020
  56. Tutorial on Variational Autoencoders

    Carl Doersch

    stat.MLcs.LGarXiv:1606.05908v32016
  57. Learning Latent Dynamics for Planning from Pixels

    Danijar Hafner, Timothy Lillicrap, Ian Fischer +4

    cs.LGcs.AIstat.MLarXiv:1811.04551v52018
  58. OCC-RAG: Optimal Cognitive Core for Faithful Question Answering

    Maksim Savkin, Mikhail Goncharov, Alexander Gambashidze +7

    cs.CLarXiv:2606.00683v12026
  59. A unified approach to mapping and clustering of bibliometric networks

    Ludo Waltman, Nees Jan van Eck, Ed C. M. Noyons

    cs.DLphysics.data-anphysics.soc-pharXiv:1006.1032v12010
  60. Improving Variational Inference with Inverse Autoregressive Flow

    Diederik P. Kingma, Tim Salimans, Rafal Jozefowicz +3

    cs.LGstat.MLarXiv:1606.04934v22016