Every paper with a summary
Every arXiv paper Paperlayer has summarized so far. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
53,281 to 53,340 of 61,306
VILA: On Pre-training for Visual Language Models
Ji Lin, Hongxu Yin, Wei Ping +7
cs.CVarXiv:2312.07533v42023Up or Down? Adaptive Rounding for Post-Training Quantization
Markus Nagel, Rana Ali Amjad, Mart van Baalen +2
cs.LGcs.CVstat.MLarXiv:2004.10568v22020Survey of the State of the Art in Natural Language Generation: Core tasks, applications and evaluation
Albert Gatt, Emiel Krahmer
cs.CLcs.AIcs.NEarXiv:1703.09902v42017ResMLP: Feedforward networks for image classification with data-efficient training
Hugo Touvron, Piotr Bojanowski, Mathilde Caron +8
cs.CVarXiv:2105.03404v22021Designing an Encoder for StyleGAN Image Manipulation
Omer Tov, Yuval Alaluf, Yotam Nitzan +2
cs.CVarXiv:2102.02766v12021PhysRVG: Physics-Aware Unified Reinforcement Learning for Video Generative Models
Qiyuan Zhang, Biao Gong, Shuai Tan +7
cs.CVarXiv:2601.11087v12026Iterative Bregman Projections for Regularized Transportation Problems
Jean-David Benamou, Guillaume Carlier, Marco Cuturi +2
math.NAmath.AParXiv:1412.5154v12014High-Resolution Representations for Labeling Pixels and Regions
Ke Sun, Yang Zhao, Borui Jiang +7
cs.CVarXiv:1904.04514v12019LLMs Encode Their Failures: Predicting Success from Pre-Generation Activations
William Lugoloobi, Thomas Foster, William Bankes +1
cs.CLcs.AIcs.LGarXiv:2602.09924v42026SemEval-2017 Task 4: Sentiment Analysis in Twitter
Sara Rosenthal, Noura Farra, Preslav Nakov
cs.CLcs.IRcs.LGarXiv:1912.00741v12019WideSeek-R1: Exploring Width Scaling for Broad Information Seeking via Multi-Agent Reinforcement Learning
Zelai Xu, Zhexuan Xu, Ruize Zhang +7
cs.AIcs.LGcs.MAarXiv:2602.04634v32026Kernel Mean Embedding of Distributions: A Review and Beyond
Krikamol Muandet, Kenji Fukumizu, Bharath Sriperumbudur +1
stat.MLcs.LGarXiv:1605.09522v42016Crowdsourcing Multiple Choice Science Questions
Johannes Welbl, Nelson F. Liu, Matt Gardner
cs.HCcs.AIcs.CLarXiv:1707.06209v12017Length-Unbiased Sequence Policy Optimization: Revealing and Controlling Response Length Variation in RLVR
Fanfan Liu, Youyang Yin, Peng Shi +3
cs.CLarXiv:2602.05261v12026Multi-Similarity Loss with General Pair Weighting for Deep Metric Learning
Xun Wang, Xintong Han, Weilin Huang +2
cs.CVarXiv:1904.06627v32019Stable Velocity: A Variance Perspective on Flow Matching
Donglin Yang, Yongxing Zhang, Xin Yu +5
cs.CVarXiv:2602.05435v22026Pseudo Numerical Methods for Diffusion Models on Manifolds
Luping Liu, Yi Ren, Zhijie Lin +1
cs.CVcs.LGmath.NAarXiv:2202.09778v22022Show and Tell: Lessons learned from the 2015 MSCOCO Image Captioning Challenge
Oriol Vinyals, Alexander Toshev, Samy Bengio +1
cs.CVarXiv:1609.06647v12016VideoWorld 2: Learning Transferable Knowledge from Real-world Videos
Zhongwei Ren, Yunchao Wei, Xiao Yu +5
cs.CVarXiv:2602.10102v12026Length-Controlled AlpacaEval: A Simple Way to Debias Automatic Evaluators
Yann Dubois, Balázs Galambosi, Percy Liang +1
cs.LGcs.AIcs.CLarXiv:2404.04475v22024Same Agent, Different Answers: A Repeat-Aware Audit of Corpus-Induced Answer Churn in Retrieval-Augmented QA
Jingjie Ning, Xueqi Li
cs.IRcs.CLarXiv:2608.22856v12026A White Paper on Neural Network Quantization
Markus Nagel, Marios Fournarakis, Rana Ali Amjad +3
cs.LGcs.AIcs.CVarXiv:2106.08295v12021LookaheadKV: Fast and Accurate KV Cache Eviction by Glimpsing into the Future without Generation
Jinwoo Ahn, Ingyu Seong, Akhil Kedia +4
cs.LGcs.AIarXiv:2603.10899v12026Scaffold-GS: Structured 3D Gaussians for View-Adaptive Rendering
Tao Lu, Mulin Yu, Linning Xu +4
cs.CVarXiv:2312.00109v12023Temporal Pattern Attention for Multivariate Time Series Forecasting
Shun-Yao Shih, Fan-Keng Sun, Hung-yi Lee
cs.LGcs.CLstat.MLarXiv:1809.04206v32018BLOCKBENCH: A Framework for Analyzing Private Blockchains
Tien Tuan Anh Dinh, Ji Wang, Gang Chen +3
cs.DBcs.CRcs.DCarXiv:1703.04057v12017Offline Reinforcement Learning as One Big Sequence Modeling Problem
Michael Janner, Qiyang Li, Sergey Levine
cs.LGcs.AIarXiv:2106.02039v42021Rethinking the Value of Agent-Generated Tests for LLM-Based Software Engineering Agents
Zhi Chen, Zhensu Sun, Yuling Shi +4
cs.SEcs.AIarXiv:2602.07900v22026Hypergraph Convolution and Hypergraph Attention
Song Bai, Feihu Zhang, Philip H. S. Torr
cs.LGcs.CVstat.MLarXiv:1901.08150v22019Scaling Vision Transformers to 22 Billion Parameters
Mostafa Dehghani, Josip Djolonga, Basil Mustafa +39
cs.CVcs.AIcs.LGarXiv:2302.05442v12023Classification of COVID-19 in chest X-ray images using DeTraC deep convolutional neural network
Asmaa Abbas, Mohammed M. Abdelsamea, Mohamed Medhat Gaber
eess.IVcs.CVcs.LGarXiv:2003.13815v32020Differentiable Convex Optimization Layers
Akshay Agrawal, Brandon Amos, Shane Barratt +3
cs.LGmath.OCstat.MLarXiv:1910.12430v12019Empty Shelves or Lost Keys? Recall Is the Bottleneck for Parametric Factuality
Nitay Calderon, Eyal Ben-David, Zorik Gekhman +2
cs.CLcs.AIarXiv:2602.14080v22026Metrics for Explainable AI: Challenges and Prospects
Robert R. Hoffman, Shane T. Mueller, Gary Klein +1
cs.AIarXiv:1812.04608v22018Instruction Tuning for Large Language Models: A Survey
Shengyu Zhang, Linfeng Dong, Xiaoya Li +8
cs.CLcs.AIcs.LGarXiv:2308.10792v102023A Survey on Deep Semi-supervised Learning
Xiangli Yang, Zixing Song, Irwin King +1
cs.LGarXiv:2103.00550v22021OmniStream: Mastering Perception, Reconstruction and Action in Continuous Streams
Yibin Yan, Jilan Xu, Shangzhe Di +2
cs.CVarXiv:2603.12265v12026Sanity Checks for Sparse Autoencoders: Do SAEs Beat Random Baselines?
Anton Korznikov, Andrey Galichin, Alexey Dontsov +3
cs.LGarXiv:2602.14111v12026Approximate Bayesian Computational methods
Jean-Michel Marin, Pierre Pudlo, Christian P. Robert +1
stat.COarXiv:1101.0955v22011The Creation and Detection of Deepfakes: A Survey
Yisroel Mirsky, Wenke Lee
cs.CVcs.LGeess.IVarXiv:2004.11138v32020The map equation
M. Rosvall, D. Axelsson, C. T. Bergstrom
physics.soc-pharXiv:0906.1405v22009Human Motion Trajectory Prediction: A Survey
Andrey Rudenko, Luigi Palmieri, Michael Herman +3
cs.ROcs.CVcs.LGarXiv:1905.06113v32019Watching, Reasoning, and Searching: A Video Deep Research Benchmark on Open Web for Agentic Video Reasoning
Chengwen Liu, Xiaomin Yu, Zhuoyue Chang +15
cs.CVcs.AIarXiv:2601.06943v22026TextBoxes: A Fast Text Detector with a Single Deep Neural Network
Minghui Liao, Baoguang Shi, Xiang Bai +2
cs.CVarXiv:1611.06779v12016It Takes Two: A Duet of Periodicity and Directionality for Burst Flicker Removal
Lishen Qu, Shihao Zhou, Jie Liang +3
cs.CVarXiv:2603.22794v12026CityPersons: A Diverse Dataset for Pedestrian Detection
Shanshan Zhang, Rodrigo Benenson, Bernt Schiele
cs.CVarXiv:1702.05693v12017UMEM: Unified Memory Extraction and Management Framework for Generalizable Memory
Yongshi Ye, Hui Jiang, Feihu Jiang +7
cs.CLarXiv:2602.10652v12026MatchTIR: Fine-Grained Supervision for Tool-Integrated Reasoning via Bipartite Matching
Changle Qu, Sunhao Dai, Hengyi Cai +3
cs.CLcs.AIarXiv:2601.10712v12026New Era of Artificial Intelligence in Education: Towards a Sustainable Multifaceted Revolution
Firuz Kamalov, David Santandreu Calong, Ikhlaas Gurrib
cs.CYarXiv:2305.18303v22023Agentic-R: Learning to Retrieve for Agentic Search
Wenhan Liu, Xinyu Ma, Yutao Zhu +4
cs.IRcs.CLarXiv:2601.11888v12026Bottom-up Object Detection by Grouping Extreme and Center Points
Xingyi Zhou, Jiacheng Zhuo, Philipp Krähenbühl
cs.CVarXiv:1901.08043v32019Deep Hidden Physics Models: Deep Learning of Nonlinear Partial Differential Equations
Maziar Raissi
stat.MLcs.LGmath.AParXiv:1801.06637v12018Generalizing to Unseen Domains via Adversarial Data Augmentation
Riccardo Volpi, Hongseok Namkoong, Ozan Sener +3
cs.CVarXiv:1805.12018v22018Feature Selective Anchor-Free Module for Single-Shot Object Detection
Chenchen Zhu, Yihui He, Marios Savvides
cs.CVarXiv:1903.00621v12019SoulX-Singer: Towards High-Quality Zero-Shot Singing Voice Synthesis
Jiale Qian, Hao Meng, Tian Zheng +17
eess.AScs.AIcs.SDarXiv:2602.07803v22026Learning a Convolutional Neural Network for Non-uniform Motion Blur Removal
Jian Sun, Wenfei Cao, Zongben Xu +1
cs.CVarXiv:1503.00593v32015Explainability for Large Language Models: A Survey
Haiyan Zhao, Hanjie Chen, Fan Yang +6
cs.CLcs.AIcs.LGarXiv:2309.01029v32023Bayesian Nonparametric Federated Learning of Neural Networks
Mikhail Yurochkin, Mayank Agarwal, Soumya Ghosh +3
stat.MLcs.LGarXiv:1905.12022v12019PyOD: A Python Toolbox for Scalable Outlier Detection
Yue Zhao, Zain Nasrullah, Zheng Li
cs.LGcs.IRstat.MLarXiv:1901.01588v22019Beyond Inferring Class Representatives: User-Level Privacy Leakage From Federated Learning
Zhibo Wang, Mengkai Song, Zhifei Zhang +3
cs.LGcs.CRcs.CVarXiv:1812.00535v32018