Every paper with a summary

Every arXiv paper Paperlayer has summarized so far. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

56,641 to 56,700 of 61,278

  1. Less is More: Early Stopping Rollout for On-Policy Distillation

    Zhou Ziheng, Jiaqi Li, Huacong Tang +2

    cs.LGcs.AIarXiv:2605.27028v12026
  2. Variational Dropout and the Local Reparameterization Trick

    Diederik P. Kingma, Tim Salimans, Max Welling

    stat.MLcs.LGstat.COarXiv:1506.02557v22015
  3. Triplet-Block Diffusion RWKV

    Ke Lin, Yiyang Luo, Zhaolong Su +2

    cs.CLarXiv:2605.25969v12026
  4. Not only where, But when: Temporal Scheduling for RLVR

    Jinghao Zhang, Ruilin Li, Feng Zhao +1

    cs.LGarXiv:2605.25381v12026
  5. RACE: Large-scale ReAding Comprehension Dataset From Examinations

    Guokun Lai, Qizhe Xie, Hanxiao Liu +2

    cs.CLcs.AIcs.LGarXiv:1704.04683v52017
  6. Cell Detection with Star-convex Polygons

    Uwe Schmidt, Martin Weigert, Coleman Broaddus +1

    cs.CVarXiv:1806.03535v22018
  7. Recursive Flow Matching

    Jiahe Huang, Sihan Xu, Sharvaree Vadgama +1

    cs.LGcs.AIcs.CVarXiv:2605.26535v12026
  8. Neural Network Acceptability Judgments

    Alex Warstadt, Amanpreet Singh, Samuel R. Bowman

    cs.CLarXiv:1805.12471v32018
  9. Text Summarization with Pretrained Encoders

    Yang Liu, Mirella Lapata

    cs.CLcs.LGarXiv:1908.08345v22019
  10. Unified Language Model Pre-training for Natural Language Understanding and Generation

    Li Dong, Nan Yang, Wenhui Wang +6

    cs.CLarXiv:1905.03197v32019
  11. SpatialBench: Is Your Spatial Foundation Model an All-Round Player?

    Haosong Peng, Hao Li, Jiaqi Chen +10

    cs.CVarXiv:2605.27367v22026
  12. MobileMoE: Scaling On-Device Mixture of Experts

    Yanbei Chen, Hanxian Huang, Ernie Chang +5

    cs.LGcs.AIcs.CLarXiv:2605.27358v12026
  13. Measuring the Effects of Non-Identical Data Distribution for Federated Visual Classification

    Tzu-Ming Harry Hsu, Hang Qi, Matthew Brown

    cs.LGcs.CVstat.MLarXiv:1909.06335v12019
  14. Robot Operating System 2: Design, Architecture, and Uses In The Wild

    Steve Macenski, Tully Foote, Brian Gerkey +2

    cs.ROarXiv:2211.07752v12022
  15. EAST: An Efficient and Accurate Scene Text Detector

    Xinyu Zhou, Cong Yao, He Wen +4

    cs.CVarXiv:1704.03155v22017
  16. Balancing Fidelity and Diversity in Diffusion Models via Symmetric Attention Decomposition: Hopfield Perspective

    Hyunmin Cho, Woo Kyoung Han, Kyong Hwan Jin

    cs.LGcs.AIarXiv:2605.27476v12026
  17. BEVFusion: Multi-Task Multi-Sensor Fusion with Unified Bird's-Eye View Representation

    Zhijian Liu, Haotian Tang, Alexander Amini +4

    cs.CVarXiv:2205.13542v32022
  18. JLT: Clean-Latent Prediction in Latent Diffusion Transformers

    Funing Fu, Tenghui Wang, Guanyu Zhou +2

    cs.CVcs.LGarXiv:2605.27102v22026
  19. Approximate Nearest Neighbor Negative Contrastive Learning for Dense Text Retrieval

    Lee Xiong, Chenyan Xiong, Ye Li +5

    cs.IRcs.CLcs.LGarXiv:2007.00808v22020
  20. GQA: Training Generalized Multi-Query Transformer Models from Multi-Head Checkpoints

    Joshua Ainslie, James Lee-Thorp, Michiel de Jong +3

    cs.CLcs.LGarXiv:2305.13245v32023
  21. BatteryMFormer: Multi-level Learning for Battery Degradation Trajectory Forecasting

    Ruifeng Tan, Jintao Dong, Weixiang Hong +3

    cs.AIarXiv:2605.27044v22026
  22. Contrastive Learning for Unpaired Image-to-Image Translation

    Taesung Park, Alexei A. Efros, Richard Zhang +1

    cs.CVcs.LGarXiv:2007.15651v32020
  23. A SIDARTHE Model of COVID-19 Epidemic in Italy

    Giulia Giordano, Franco Blanchini, Raffaele Bruno +5

    q-bio.PEeess.SYmath.DSarXiv:2003.09861v12020
  24. CoAtNet: Marrying Convolution and Attention for All Data Sizes

    Zihang Dai, Hanxiao Liu, Quoc V. Le +1

    cs.CVcs.LGarXiv:2106.04803v22021
  25. AgensFlow: A Coordination-Policy Substrate for Multi-Agent Systems

    Nicole Koenigstein

    cs.MAcs.AIcs.LGarXiv:2605.27466v12026
  26. How far are we from solving the 2D & 3D Face Alignment problem? (and a dataset of 230,000 3D facial landmarks)

    Adrian Bulat, Georgios Tzimiropoulos

    cs.CVcs.LGarXiv:1703.07332v32017
  27. MentorNet: Learning Data-Driven Curriculum for Very Deep Neural Networks on Corrupted Labels

    Lu Jiang, Zhengyuan Zhou, Thomas Leung +2

    cs.CVarXiv:1712.05055v22017
  28. Cascaded Diffusion Models for High Fidelity Image Generation

    Jonathan Ho, Chitwan Saharia, William Chan +3

    cs.CVcs.AIcs.LGarXiv:2106.15282v32021
  29. Model soups: averaging weights of multiple fine-tuned models improves accuracy without increasing inference time

    Mitchell Wortsman, Gabriel Ilharco, Samir Yitzhak Gadre +8

    cs.LGcs.CLcs.CVarXiv:2203.05482v32022
  30. The Roadmap to 6G -- AI Empowered Wireless Networks

    Khaled B. Letaief, Wei Chen, Yuanming Shi +2

    cs.NIcs.LGarXiv:1904.11686v22019
  31. ChatGPT Outperforms Crowd-Workers for Text-Annotation Tasks

    Fabrizio Gilardi, Meysam Alizadeh, Maël Kubli

    cs.CLcs.CYarXiv:2303.15056v22023
  32. Junction Tree Variational Autoencoder for Molecular Graph Generation

    Wengong Jin, Regina Barzilay, Tommi Jaakkola

    cs.LGcs.NEstat.MLarXiv:1802.04364v42018
  33. MELD: A Multimodal Multi-Party Dataset for Emotion Recognition in Conversations

    Soujanya Poria, Devamanyu Hazarika, Navonil Majumder +3

    cs.CLarXiv:1810.02508v62018
  34. M3-Embedding: Multi-Linguality, Multi-Functionality, Multi-Granularity Text Embeddings Through Self-Knowledge Distillation

    Jianlv Chen, Shitao Xiao, Peitian Zhang +3

    cs.CLcs.AIcs.LGarXiv:2402.03216v52024
  35. Pre-trained Models for Natural Language Processing: A Survey

    Xipeng Qiu, Tianxiang Sun, Yige Xu +3

    cs.CLcs.LGarXiv:2003.08271v42020
  36. Deep Facial Expression Recognition: A Survey

    Shan Li, Weihong Deng

    cs.CVarXiv:1804.08348v22018
  37. BiSeNet V2: Bilateral Network with Guided Aggregation for Real-time Semantic Segmentation

    Changqian Yu, Changxin Gao, Jingbo Wang +3

    cs.CVarXiv:2004.02147v12020
  38. DeblurGAN: Blind Motion Deblurring Using Conditional Adversarial Networks

    Orest Kupyn, Volodymyr Budzan, Mykola Mykhailych +2

    cs.CVarXiv:1711.07064v42017
  39. Multi-Agent Reinforcement Learning: A Selective Overview of Theories and Algorithms

    Kaiqing Zhang, Zhuoran Yang, Tamer Başar

    cs.LGcs.AIcs.MAarXiv:1911.10635v22019
  40. A simple neural network module for relational reasoning

    Adam Santoro, David Raposo, David G. T. Barrett +4

    cs.CLcs.LGarXiv:1706.01427v12017
  41. word2vec Explained: deriving Mikolov et al.'s negative-sampling word-embedding method

    Yoav Goldberg, Omer Levy

    cs.CLcs.LGstat.MLarXiv:1402.3722v12014
  42. Making Deep Neural Networks Robust to Label Noise: a Loss Correction Approach

    Giorgio Patrini, Alessandro Rozza, Aditya Menon +2

    stat.MLcs.LGarXiv:1609.03683v22016
  43. Wild Patterns: Ten Years After the Rise of Adversarial Machine Learning

    Battista Biggio, Fabio Roli

    cs.CVcs.CRcs.GTarXiv:1712.03141v22017
  44. Towards the Development of Realistic Botnet Dataset in the Internet of Things for Network Forensic Analytics: Bot-IoT Dataset

    Nickolaos Koroniotis, Nour Moustafa, Elena Sitnikova +1

    cs.CRarXiv:1811.00701v12018
  45. A Theoretically Grounded Application of Dropout in Recurrent Neural Networks

    Yarin Gal, Zoubin Ghahramani

    stat.MLarXiv:1512.05287v52015
  46. Understanding intermediate layers using linear classifier probes

    Guillaume Alain, Yoshua Bengio

    stat.MLcs.LGarXiv:1610.01644v42016
  47. Learning Quadrupedal Locomotion over Challenging Terrain

    Joonho Lee, Jemin Hwangbo, Lorenz Wellhausen +2

    cs.ROcs.LGeess.SYarXiv:2010.11251v12020
  48. ImageBind: One Embedding Space To Bind Them All

    Rohit Girdhar, Alaaeldin El-Nouby, Zhuang Liu +4

    cs.CVcs.AIcs.LGarXiv:2305.05665v22023
  49. IP-Adapter: Text Compatible Image Prompt Adapter for Text-to-Image Diffusion Models

    Hu Ye, Jun Zhang, Sibo Liu +2

    cs.CVcs.AIarXiv:2308.06721v12023
  50. MVSNet: Depth Inference for Unstructured Multi-view Stereo

    Yao Yao, Zixin Luo, Shiwei Li +2

    cs.CVarXiv:1804.02505v22018
  51. A guide to convolution arithmetic for deep learning

    Vincent Dumoulin, Francesco Visin

    stat.MLcs.LGcs.NEarXiv:1603.07285v22016
  52. Multiscale Vision Transformers

    Haoqi Fan, Bo Xiong, Karttikeya Mangalam +4

    cs.CVcs.AIcs.LGarXiv:2104.11227v12021
  53. Expanding Performance Boundaries of Open-Source Multimodal Models with Model, Data, and Test-Time Scaling

    Zhe Chen, Weiyun Wang, Yue Cao +39

    cs.CVarXiv:2412.05271v52024
  54. SDR - half-baked or well done?

    Jonathan Le Roux, Scott Wisdom, Hakan Erdogan +1

    cs.SDeess.ASarXiv:1811.02508v12018
  55. Learning to Reweight Examples for Robust Deep Learning

    Mengye Ren, Wenyuan Zeng, Bin Yang +1

    cs.LGstat.MLarXiv:1803.09050v32018
  56. Generative Adversarial Network in Medical Imaging: A Review

    Xin Yi, Ekta Walia, Paul Babyn

    cs.CVcs.LGarXiv:1809.07294v42018
  57. Deep Gradient Compression: Reducing the Communication Bandwidth for Distributed Training

    Yujun Lin, Song Han, Huizi Mao +2

    cs.CVcs.DCcs.LGarXiv:1712.01887v32017
  58. Quantum repeaters based on atomic ensembles and linear optics

    Nicolas Sangouard, Christoph Simon, Hugues de Riedmatten +1

    quant-pharXiv:0906.2699v22009
  59. Recursive Partitioning for Heterogeneous Causal Effects

    Susan Athey, Guido Imbens

    stat.MLecon.EMarXiv:1504.01132v32015
  60. Networks beyond pairwise interactions: structure and dynamics

    Federico Battiston, Giulia Cencetti, Iacopo Iacopini +5

    physics.soc-phcond-mat.dis-nncs.SIarXiv:2006.01764v12020