Every paper with a summary
Every arXiv paper Paperlayer has summarized so far. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
57,001 to 57,060 of 61,217
Apple-$π$: Benchmarking Thinking with Video Towards Law-Grounded Physical Intelligence
Runmao Yao, Kairui Hu, Yukang Cao +11
cs.CVarXiv:2607.16401v12026MeanFlowNFT: Bringing Forward-Process RL to Average-Velocity Generators
Yushi Huang, Xiangxin Zhou, Jun Zhang +2
cs.CVcs.LGarXiv:2607.15273v22026Look Before You Leap: Distilling Tree Search into Action Evaluation for Frozen VLA Models
Xinyi Xie, Zican Hu, Zhanyu Liu +7
cs.ROarXiv:2607.03751v12026Bridging Interleaved Multi-Modal Reasoning as a Unified Decision Process
Zican Hu, Xuyang Hu, Yiming Liu +10
cs.AIarXiv:2607.03748v12026ENPIRE: Agentic Robot Policy Self-Improvement in the Real World
Wenli Xiao, Jia Xie, Tonghe Zhang +14
cs.AIarXiv:2606.19980v12026RefGC-SR$^2$: Reference-guided Super-Resolution and Refinement of AI Generated Content
Jeahun Sung, Dahyeon Kye, Soo Ye Kim +1
cs.CVarXiv:2606.15158v22026AutoLLMResearch: Training Research Agents for Automating LLM Experiment Configuration - Learning from Cheap, Optimizing Expensive
Taicheng Guo, Nitesh V. Chawla, Olaf Wiest +1
cs.AIcs.CLcs.LGarXiv:2605.11518v22026Benchmark Everything Everywhere All at Once
Shiyun Xiong, Dongming Wu, Peiwen Sun +5
cs.AIarXiv:2606.06462v12026Synchronization in complex networks
Alex Arenas, Albert Diaz-Guilera, Jurgen Kurths +2
physics.soc-pharXiv:0805.2976v32008DV-World: Benchmarking Data Visualization Agents in Real-World Scenarios
Jinxiang Meng, Shaoping Huang, Fangyu Lei +17
cs.CLarXiv:2604.25914v12026Rethinking Muon Beyond Pretraining: Spectral Failures and High-Pass Remedies for VLA and RLVR
Chongyu Fan, Gaowen Liu, Mingyi Hong +2
cs.LGarXiv:2605.19282v12026$π$-Bench: Evaluating Proactive Personal Assistant Agents in Long-Horizon Workflows
Haoran Zhang, Luxin Xu, Zhilin Wang +11
cs.AIarXiv:2605.14678v32026MSAVBench: Towards Comprehensive and Reliable Evaluation of Multi-Shot Audio-Video Generation
Yujie Wei, Yujin Han, Zhekai Chen +20
cs.CVarXiv:2605.20183v42026Continual Harness: Online Adaptation for Self-Improving Foundation Agents
Seth Karten, Joel Zhang, Tersoo Upaa +5
cs.LGcs.AIarXiv:2605.09998v12026Can RL Teach Long-Horizon Reasoning to LLMs? Expressiveness Is Key
Tianle Wang, Zhaoyang Wang, Guangchen Lan +4
cs.AIcs.CLarXiv:2605.06638v32026Prompt-Activation Duality: Improving Activation Steering via Attention-Level Interventions
Diancheng Kang, Zheyuan Liu, Ningshan Ma +3
cs.CLcs.AIarXiv:2605.10664v22026A Tutorial on Regularized Partial Correlation Networks
Sacha Epskamp, Eiko I. Fried
stat.APstat.MEarXiv:1607.01367v92016Deep multi-scale video prediction beyond mean square error
Michael Mathieu, Camille Couprie, Yann LeCun
cs.LGcs.CVstat.MLarXiv:1511.05440v62015Ego4D: Around the World in 3,000 Hours of Egocentric Video
Kristen Grauman, Andrew Westbury, Eugene Byrne +82
cs.CVcs.AIarXiv:2110.07058v32021AttnGAN: Fine-Grained Text to Image Generation with Attentional Generative Adversarial Networks
Tao Xu, Pengchuan Zhang, Qiuyuan Huang +4
cs.CVarXiv:1711.10485v12017GLM: General Language Model Pretraining with Autoregressive Blank Infilling
Zhengxiao Du, Yujie Qian, Xiao Liu +4
cs.CLcs.AIcs.LGarXiv:2103.10360v22021Person Transfer GAN to Bridge Domain Gap for Person Re-Identification
Longhui Wei, Shiliang Zhang, Wen Gao +1
cs.CVarXiv:1711.08565v22017cuDNN: Efficient Primitives for Deep Learning
Sharan Chetlur, Cliff Woolley, Philippe Vandermersch +4
cs.NEcs.LGcs.MSarXiv:1410.0759v32014The Rise and Potential of Large Language Model Based Agents: A Survey
Zhiheng Xi, Wenxiang Chen, Xin Guo +26
cs.AIcs.CLarXiv:2309.07864v32023Adversarial Examples Are Not Easily Detected: Bypassing Ten Detection Methods
Nicholas Carlini, David Wagner
cs.LGcs.CRcs.CVarXiv:1705.07263v22017Generalized Focal Loss: Learning Qualified and Distributed Bounding Boxes for Dense Object Detection
Xiang Li, Wenhai Wang, Lijun Wu +5
cs.CVarXiv:2006.04388v12020SAGA: A Fast Incremental Gradient Method With Support for Non-Strongly Convex Composite Objectives
Aaron Defazio, Francis Bach, Simon Lacoste-Julien
cs.LGmath.OCstat.MLarXiv:1407.0202v32014Multitask Prompted Training Enables Zero-Shot Task Generalization
Victor Sanh, Albert Webson, Colin Raffel +38
cs.LGcs.CLarXiv:2110.08207v32021Send a SCOUT First: Pre-hoc Reasoning for Adaptive Detector Allocation in Prompt-Injection Defense
Shuhao Zhang, Jiarui Li, Qi Cao +2
cs.CRcs.LGarXiv:2605.30837v22026On Dynamic Mode Decomposition: Theory and Applications
Jonathan H. Tu, Clarence W. Rowley, Dirk M. Luchtenburg +2
math.NAphysics.flu-dynarXiv:1312.0041v12013Natural Adversarial Examples
Dan Hendrycks, Kevin Zhao, Steven Basart +2
cs.LGcs.CVstat.MLarXiv:1907.07174v42019Distilling LLM Feedback for Lean Theorem Proving
Gaetan Narozniak, Gérard Biau, Rémi Munos +2
cs.AIarXiv:2605.30861v12026BEVFormer: Learning Bird's-Eye-View Representation from Multi-Camera Images via Spatiotemporal Transformers
Zhiqi Li, Wenhai Wang, Hongyang Li +5
cs.CVarXiv:2203.17270v22022Semi-Supervised Noise Adaptation: Transferring Knowledge from Noise Domain
Yuan Yao, Jin Song, Huixia Li +3
cs.LGarXiv:2606.00558v22026Covid-19: Automatic detection from X-Ray images utilizing Transfer Learning with Convolutional Neural Networks
Ioannis D. Apostolopoulos, Tzani Bessiana
eess.IVcs.CVcs.LGarXiv:2003.11617v12020Representation over Routing: Diagnosing Temporal Routing Pathologies in Multi-Timescale PPO
Jing Sun
cs.LGcs.AIarXiv:2604.13517v42026Confidence-Adaptive SwiGLU for Mixture-of-Experts
Shaohua Li, Xiuchao Sui, Xiaobing Sun +4
cs.LGcs.CLarXiv:2606.00761v12026GMAN: A Graph Multi-Attention Network for Traffic Prediction
Chuanpan Zheng, Xiaoliang Fan, Cheng Wang +1
eess.SPcs.LGarXiv:1911.08415v22019Ensemble deep learning: A review
M. A. Ganaie, Minghui Hu, A. K. Malik +2
cs.LGcs.AIcs.CVarXiv:2104.02395v32021MMLU-Pro: A More Robust and Challenging Multi-Task Language Understanding Benchmark
Yubo Wang, Xueguang Ma, Ge Zhang +14
cs.CLarXiv:2406.01574v62024Self-supervised Visual Feature Learning with Deep Neural Networks: A Survey
Longlong Jing, Yingli Tian
cs.CVarXiv:1902.06162v12019Constrained Consensus
Angelia Nedić, Asuman Ozdaglar, Pablo A. Parrilo
math.OCarXiv:0802.3922v22008The Shape of Addition: Geometric Structures of Arithmetic in Large Language Models
Liuyuan Wen, Xun Zhu, Lihao Huang +2
cs.LGcs.AIarXiv:2606.03645v12026Functional Attention: From Pairwise Affinities to Functional Correspondences
Jiefang Xiao, Maolin Gao, Simon Weber +2
cs.LGarXiv:2605.31559v12026Honest Lying: Understanding Memory Confabulation in Reflexive Agents
Prakhar Dixit, Sadia Kamal, Tim Oates
cs.LGcs.AIarXiv:2605.29463v22026GradNorm: Gradient Normalization for Adaptive Loss Balancing in Deep Multitask Networks
Zhao Chen, Vijay Badrinarayanan, Chen-Yu Lee +1
cs.CVarXiv:1711.02257v42017PaintBench: Deterministic Evaluation of Precise Visual Editing
Kai Xu, Ellis Brown, Shrikar Madhu +3
cs.GRcs.CVcs.LGarXiv:2606.00188v12026Relational Knowledge Distillation
Wonpyo Park, Dongju Kim, Yan Lu +1
cs.CVcs.LGarXiv:1904.05068v22019SpatialAct: Probing Spatial Reasoning-to-Action Capabilities of VLM Agents in 3D Scenes
Tianhui Liu, Jie Feng, Zhiheng Zheng +6
cs.CVcs.AIcs.CLarXiv:2605.31148v12026Combinatorial Synthesis: Scaling Code RLVR via Atomic Decomposition and Recombination
Jiasheng Zheng, Boxi Cao, Boxi Yu +6
cs.CLcs.SEarXiv:2605.31058v12026Stacked Attention Networks for Image Question Answering
Zichao Yang, Xiaodong He, Jianfeng Gao +2
cs.LGcs.CLcs.CVarXiv:1511.02274v22015Enhancing the Locality and Breaking the Memory Bottleneck of Transformer on Time Series Forecasting
Shiyang Li, Xiaoyong Jin, Yao Xuan +4
cs.LGstat.MLarXiv:1907.00235v32019Deep Mutual Learning
Ying Zhang, Tao Xiang, Timothy M. Hospedales +1
cs.CVarXiv:1706.00384v12017DRAW: A Recurrent Neural Network For Image Generation
Karol Gregor, Ivo Danihelka, Alex Graves +2
cs.CVcs.LGcs.NEarXiv:1502.04623v22015Simple and Deep Graph Convolutional Networks
Ming Chen, Zhewei Wei, Zengfeng Huang +2
cs.LGstat.MLarXiv:2007.02133v12020Tutorial on Variational Autoencoders
Carl Doersch
stat.MLcs.LGarXiv:1606.05908v32016Learning Latent Dynamics for Planning from Pixels
Danijar Hafner, Timothy Lillicrap, Ian Fischer +4
cs.LGcs.AIstat.MLarXiv:1811.04551v52018OCC-RAG: Optimal Cognitive Core for Faithful Question Answering
Maksim Savkin, Mikhail Goncharov, Alexander Gambashidze +7
cs.CLarXiv:2606.00683v12026A unified approach to mapping and clustering of bibliometric networks
Ludo Waltman, Nees Jan van Eck, Ed C. M. Noyons
cs.DLphysics.data-anphysics.soc-pharXiv:1006.1032v12010Improving Variational Inference with Inverse Autoregressive Flow
Diederik P. Kingma, Tim Salimans, Rafal Jozefowicz +3
cs.LGstat.MLarXiv:1606.04934v22016