Every paper with a summary
Every arXiv paper Paperlayer has summarized so far. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
56,641 to 56,700 of 61,278
Less is More: Early Stopping Rollout for On-Policy Distillation
Zhou Ziheng, Jiaqi Li, Huacong Tang +2
cs.LGcs.AIarXiv:2605.27028v12026Variational Dropout and the Local Reparameterization Trick
Diederik P. Kingma, Tim Salimans, Max Welling
stat.MLcs.LGstat.COarXiv:1506.02557v22015Triplet-Block Diffusion RWKV
Ke Lin, Yiyang Luo, Zhaolong Su +2
cs.CLarXiv:2605.25969v12026Not only where, But when: Temporal Scheduling for RLVR
Jinghao Zhang, Ruilin Li, Feng Zhao +1
cs.LGarXiv:2605.25381v12026RACE: Large-scale ReAding Comprehension Dataset From Examinations
Guokun Lai, Qizhe Xie, Hanxiao Liu +2
cs.CLcs.AIcs.LGarXiv:1704.04683v52017Cell Detection with Star-convex Polygons
Uwe Schmidt, Martin Weigert, Coleman Broaddus +1
cs.CVarXiv:1806.03535v22018Recursive Flow Matching
Jiahe Huang, Sihan Xu, Sharvaree Vadgama +1
cs.LGcs.AIcs.CVarXiv:2605.26535v12026Neural Network Acceptability Judgments
Alex Warstadt, Amanpreet Singh, Samuel R. Bowman
cs.CLarXiv:1805.12471v32018Text Summarization with Pretrained Encoders
Yang Liu, Mirella Lapata
cs.CLcs.LGarXiv:1908.08345v22019Unified Language Model Pre-training for Natural Language Understanding and Generation
Li Dong, Nan Yang, Wenhui Wang +6
cs.CLarXiv:1905.03197v32019SpatialBench: Is Your Spatial Foundation Model an All-Round Player?
Haosong Peng, Hao Li, Jiaqi Chen +10
cs.CVarXiv:2605.27367v22026MobileMoE: Scaling On-Device Mixture of Experts
Yanbei Chen, Hanxian Huang, Ernie Chang +5
cs.LGcs.AIcs.CLarXiv:2605.27358v12026Measuring the Effects of Non-Identical Data Distribution for Federated Visual Classification
Tzu-Ming Harry Hsu, Hang Qi, Matthew Brown
cs.LGcs.CVstat.MLarXiv:1909.06335v12019Robot Operating System 2: Design, Architecture, and Uses In The Wild
Steve Macenski, Tully Foote, Brian Gerkey +2
cs.ROarXiv:2211.07752v12022EAST: An Efficient and Accurate Scene Text Detector
Xinyu Zhou, Cong Yao, He Wen +4
cs.CVarXiv:1704.03155v22017Balancing Fidelity and Diversity in Diffusion Models via Symmetric Attention Decomposition: Hopfield Perspective
Hyunmin Cho, Woo Kyoung Han, Kyong Hwan Jin
cs.LGcs.AIarXiv:2605.27476v12026BEVFusion: Multi-Task Multi-Sensor Fusion with Unified Bird's-Eye View Representation
Zhijian Liu, Haotian Tang, Alexander Amini +4
cs.CVarXiv:2205.13542v32022JLT: Clean-Latent Prediction in Latent Diffusion Transformers
Funing Fu, Tenghui Wang, Guanyu Zhou +2
cs.CVcs.LGarXiv:2605.27102v22026Approximate Nearest Neighbor Negative Contrastive Learning for Dense Text Retrieval
Lee Xiong, Chenyan Xiong, Ye Li +5
cs.IRcs.CLcs.LGarXiv:2007.00808v22020GQA: Training Generalized Multi-Query Transformer Models from Multi-Head Checkpoints
Joshua Ainslie, James Lee-Thorp, Michiel de Jong +3
cs.CLcs.LGarXiv:2305.13245v32023BatteryMFormer: Multi-level Learning for Battery Degradation Trajectory Forecasting
Ruifeng Tan, Jintao Dong, Weixiang Hong +3
cs.AIarXiv:2605.27044v22026Contrastive Learning for Unpaired Image-to-Image Translation
Taesung Park, Alexei A. Efros, Richard Zhang +1
cs.CVcs.LGarXiv:2007.15651v32020A SIDARTHE Model of COVID-19 Epidemic in Italy
Giulia Giordano, Franco Blanchini, Raffaele Bruno +5
q-bio.PEeess.SYmath.DSarXiv:2003.09861v12020CoAtNet: Marrying Convolution and Attention for All Data Sizes
Zihang Dai, Hanxiao Liu, Quoc V. Le +1
cs.CVcs.LGarXiv:2106.04803v22021AgensFlow: A Coordination-Policy Substrate for Multi-Agent Systems
Nicole Koenigstein
cs.MAcs.AIcs.LGarXiv:2605.27466v12026How far are we from solving the 2D & 3D Face Alignment problem? (and a dataset of 230,000 3D facial landmarks)
Adrian Bulat, Georgios Tzimiropoulos
cs.CVcs.LGarXiv:1703.07332v32017MentorNet: Learning Data-Driven Curriculum for Very Deep Neural Networks on Corrupted Labels
Lu Jiang, Zhengyuan Zhou, Thomas Leung +2
cs.CVarXiv:1712.05055v22017Cascaded Diffusion Models for High Fidelity Image Generation
Jonathan Ho, Chitwan Saharia, William Chan +3
cs.CVcs.AIcs.LGarXiv:2106.15282v32021Model soups: averaging weights of multiple fine-tuned models improves accuracy without increasing inference time
Mitchell Wortsman, Gabriel Ilharco, Samir Yitzhak Gadre +8
cs.LGcs.CLcs.CVarXiv:2203.05482v32022The Roadmap to 6G -- AI Empowered Wireless Networks
Khaled B. Letaief, Wei Chen, Yuanming Shi +2
cs.NIcs.LGarXiv:1904.11686v22019ChatGPT Outperforms Crowd-Workers for Text-Annotation Tasks
Fabrizio Gilardi, Meysam Alizadeh, Maël Kubli
cs.CLcs.CYarXiv:2303.15056v22023Junction Tree Variational Autoencoder for Molecular Graph Generation
Wengong Jin, Regina Barzilay, Tommi Jaakkola
cs.LGcs.NEstat.MLarXiv:1802.04364v42018MELD: A Multimodal Multi-Party Dataset for Emotion Recognition in Conversations
Soujanya Poria, Devamanyu Hazarika, Navonil Majumder +3
cs.CLarXiv:1810.02508v62018M3-Embedding: Multi-Linguality, Multi-Functionality, Multi-Granularity Text Embeddings Through Self-Knowledge Distillation
Jianlv Chen, Shitao Xiao, Peitian Zhang +3
cs.CLcs.AIcs.LGarXiv:2402.03216v52024Pre-trained Models for Natural Language Processing: A Survey
Xipeng Qiu, Tianxiang Sun, Yige Xu +3
cs.CLcs.LGarXiv:2003.08271v42020Deep Facial Expression Recognition: A Survey
Shan Li, Weihong Deng
cs.CVarXiv:1804.08348v22018BiSeNet V2: Bilateral Network with Guided Aggregation for Real-time Semantic Segmentation
Changqian Yu, Changxin Gao, Jingbo Wang +3
cs.CVarXiv:2004.02147v12020DeblurGAN: Blind Motion Deblurring Using Conditional Adversarial Networks
Orest Kupyn, Volodymyr Budzan, Mykola Mykhailych +2
cs.CVarXiv:1711.07064v42017Multi-Agent Reinforcement Learning: A Selective Overview of Theories and Algorithms
Kaiqing Zhang, Zhuoran Yang, Tamer Başar
cs.LGcs.AIcs.MAarXiv:1911.10635v22019A simple neural network module for relational reasoning
Adam Santoro, David Raposo, David G. T. Barrett +4
cs.CLcs.LGarXiv:1706.01427v12017word2vec Explained: deriving Mikolov et al.'s negative-sampling word-embedding method
Yoav Goldberg, Omer Levy
cs.CLcs.LGstat.MLarXiv:1402.3722v12014Making Deep Neural Networks Robust to Label Noise: a Loss Correction Approach
Giorgio Patrini, Alessandro Rozza, Aditya Menon +2
stat.MLcs.LGarXiv:1609.03683v22016Wild Patterns: Ten Years After the Rise of Adversarial Machine Learning
Battista Biggio, Fabio Roli
cs.CVcs.CRcs.GTarXiv:1712.03141v22017Towards the Development of Realistic Botnet Dataset in the Internet of Things for Network Forensic Analytics: Bot-IoT Dataset
Nickolaos Koroniotis, Nour Moustafa, Elena Sitnikova +1
cs.CRarXiv:1811.00701v12018A Theoretically Grounded Application of Dropout in Recurrent Neural Networks
Yarin Gal, Zoubin Ghahramani
stat.MLarXiv:1512.05287v52015Understanding intermediate layers using linear classifier probes
Guillaume Alain, Yoshua Bengio
stat.MLcs.LGarXiv:1610.01644v42016Learning Quadrupedal Locomotion over Challenging Terrain
Joonho Lee, Jemin Hwangbo, Lorenz Wellhausen +2
cs.ROcs.LGeess.SYarXiv:2010.11251v12020ImageBind: One Embedding Space To Bind Them All
Rohit Girdhar, Alaaeldin El-Nouby, Zhuang Liu +4
cs.CVcs.AIcs.LGarXiv:2305.05665v22023IP-Adapter: Text Compatible Image Prompt Adapter for Text-to-Image Diffusion Models
Hu Ye, Jun Zhang, Sibo Liu +2
cs.CVcs.AIarXiv:2308.06721v12023MVSNet: Depth Inference for Unstructured Multi-view Stereo
Yao Yao, Zixin Luo, Shiwei Li +2
cs.CVarXiv:1804.02505v22018A guide to convolution arithmetic for deep learning
Vincent Dumoulin, Francesco Visin
stat.MLcs.LGcs.NEarXiv:1603.07285v22016Multiscale Vision Transformers
Haoqi Fan, Bo Xiong, Karttikeya Mangalam +4
cs.CVcs.AIcs.LGarXiv:2104.11227v12021Expanding Performance Boundaries of Open-Source Multimodal Models with Model, Data, and Test-Time Scaling
Zhe Chen, Weiyun Wang, Yue Cao +39
cs.CVarXiv:2412.05271v52024SDR - half-baked or well done?
Jonathan Le Roux, Scott Wisdom, Hakan Erdogan +1
cs.SDeess.ASarXiv:1811.02508v12018Learning to Reweight Examples for Robust Deep Learning
Mengye Ren, Wenyuan Zeng, Bin Yang +1
cs.LGstat.MLarXiv:1803.09050v32018Generative Adversarial Network in Medical Imaging: A Review
Xin Yi, Ekta Walia, Paul Babyn
cs.CVcs.LGarXiv:1809.07294v42018Deep Gradient Compression: Reducing the Communication Bandwidth for Distributed Training
Yujun Lin, Song Han, Huizi Mao +2
cs.CVcs.DCcs.LGarXiv:1712.01887v32017Quantum repeaters based on atomic ensembles and linear optics
Nicolas Sangouard, Christoph Simon, Hugues de Riedmatten +1
quant-pharXiv:0906.2699v22009Recursive Partitioning for Heterogeneous Causal Effects
Susan Athey, Guido Imbens
stat.MLecon.EMarXiv:1504.01132v32015Networks beyond pairwise interactions: structure and dynamics
Federico Battiston, Giulia Cencetti, Iacopo Iacopini +5
physics.soc-phcond-mat.dis-nncs.SIarXiv:2006.01764v12020