Machine Learning

Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

16,321 to 16,380 of 20,193

  1. Data Predictability Shapes Weibull Weight-Scale Growth in Transformer Training

    Tiexin Ding

    cs.LGstat.MLarXiv:2608.23573v12026
  2. Decolonial AI: Decolonial Theory as Sociotechnical Foresight in Artificial Intelligence

    Shakir Mohamed, Marie-Therese Png, William Isaac

    cs.CYcs.AIcs.LGarXiv:2007.04068v12020
  3. InfoDPP-PAC: Principled Patch Selection for Whole Slide Image Analysis

    Prateek Mittal, Ayush Srivastava, Joohi Chauhan

    q-bio.QMcs.CVcs.ITarXiv:2608.23574v12026
  4. A mesh-free multiresolution deep energy method with phase-field modeling of brittle fracture

    Han Zhang, Mehrisadat Makki Alamdari, Babak Shahbodagh +4

    cs.LGmath.NAarXiv:2608.24126v12026
  5. Infant Care Video Dataset for Classification of Interventions Using Transformers

    Igor Bogdanov, James Green

    cs.CVcs.AIcs.LGarXiv:2608.23838v12026
  6. A Comparative Study in Surgical AI: Potential and Limitations of Data, Compute, and Scaling

    Kirill Skobelev, Eric Fithian, Yegor Baranovski +9

    cs.AIcs.CVcs.LGarXiv:2603.27341v42026
  7. Towards Universal Fake Image Detectors that Generalize Across Generative Models

    Utkarsh Ojha, Yuheng Li, Yong Jae Lee

    cs.CVcs.LGarXiv:2302.10174v22023
  8. When and why vision-language models behave like bags-of-words, and what to do about it?

    Mert Yuksekgonul, Federico Bianchi, Pratyusha Kalluri +2

    cs.CVcs.AIcs.CLarXiv:2210.01936v32022
  9. Explainable Machine Learning in Deployment

    Umang Bhatt, Alice Xiang, Shubham Sharma +7

    cs.LGcs.AIcs.CYarXiv:1909.06342v42019
  10. ChorusTIC: Training-Free Multivariate Time Series Classification via Chorus In-Context Learning

    Juntao Fang, Shifeng Xie, Ruichu Cai +6

    cs.LGcs.AIstat.MLarXiv:2608.24033v12026
  11. Equivariant Cellular Sheaves for Molecular Electronic Structure: Bridging Sheaf Cohomology and E(3)-Equivariant Hamiltonian Learning

    Krishna Harish

    cs.LGphysics.chem-pharXiv:2608.23571v12026
  12. A Survey of Deep Learning Applications to Autonomous Vehicle Control

    Sampo Kuutti, Richard Bowden, Yaochu Jin +2

    cs.LGcs.CVeess.SYarXiv:1912.10773v12019
  13. Measuring Calibration in Deep Learning

    Jeremy Nixon, Mike Dusenberry, Ghassen Jerfel +4

    cs.LGstat.MLarXiv:1904.01685v22019
  14. SecureBoost: A Lossless Federated Learning Framework

    Kewei Cheng, Tao Fan, Yilun Jin +4

    cs.LGstat.MLarXiv:1901.08755v32019
  15. LongVideoBench: A Benchmark for Long-context Interleaved Video-Language Understanding

    Haoning Wu, Dongxu Li, Bei Chen +1

    cs.CVcs.CLcs.LGarXiv:2407.15754v12024
  16. Malware Detection by Eating a Whole EXE

    Edward Raff, Jon Barker, Jared Sylvester +3

    stat.MLcs.CRcs.LGarXiv:1710.09435v12017
  17. Learning Deep Generative Models of Graphs

    Yujia Li, Oriol Vinyals, Chris Dyer +2

    cs.LGstat.MLarXiv:1803.03324v12018
  18. Applications of Deep Learning and Reinforcement Learning to Biological Data

    Mufti Mahmud, M. Shamim Kaiser, Amir Hussain +1

    cs.LGstat.MLarXiv:1711.03985v22017
  19. Backdoor Attacks on Decentralised Post-Training

    Oğuzhan Ersoy, Nikolay Blagoev, Jona te Lintelo +3

    cs.CRcs.LGarXiv:2604.02372v12026
  20. Asymmetric Non-local Neural Networks for Semantic Segmentation

    Zhen Zhu, Mengde Xu, Song Bai +2

    cs.CVcs.LGarXiv:1908.07678v52019
  21. The Computational Limits of Deep Learning

    Neil C. Thompson, Kristjan Greenewald, Keeheon Lee +1

    cs.LGstat.MLarXiv:2007.05558v22020
  22. Correcting Variable Importance Scored by Random Forests

    Guancheng Zhou, Haiping Xu, Jason Liu +1

    stat.MEcs.AIcs.LGarXiv:2606.10770v12026
  23. CALVIN: A Benchmark for Language-Conditioned Policy Learning for Long-Horizon Robot Manipulation Tasks

    Oier Mees, Lukas Hermann, Erick Rosete-Beas +1

    cs.ROcs.AIcs.CLarXiv:2112.03227v42021
  24. ReAct: Out-of-distribution Detection With Rectified Activations

    Yiyou Sun, Chuan Guo, Yixuan Li

    cs.LGarXiv:2111.12797v12021
  25. Lipschitz regularity of deep neural networks: analysis and efficient estimation

    Kevin Scaman, Aladin Virmaux

    stat.MLcs.LGarXiv:1805.10965v22018
  26. StyleGAN-XL: Scaling StyleGAN to Large Diverse Datasets

    Axel Sauer, Katja Schwarz, Andreas Geiger

    cs.LGcs.CVarXiv:2202.00273v22022
  27. NeuroPrefetcher: Storage-Aware Sparse LLM Inference via Delta Prefetching

    Nobel Dhar, Md Romyull Islam, Xuechen Zhang +4

    cs.DCcs.LGarXiv:2608.22643v12026
  28. Contrastive learning of global and local features for medical image segmentation with limited annotations

    Krishna Chaitanya, Ertunc Erdil, Neerav Karani +1

    cs.CVcs.LGeess.IVarXiv:2006.10511v22020
  29. Improving Diffusion Models for Inverse Problems using Manifold Constraints

    Hyungjin Chung, Byeongsu Sim, Dohoon Ryu +1

    cs.LGcs.AIcs.CVarXiv:2206.00941v32022
  30. Joint Extraction of Entities and Relations Based on a Novel Tagging Scheme

    Suncong Zheng, Feng Wang, Hongyun Bao +3

    cs.CLcs.AIcs.LGarXiv:1706.05075v12017
  31. Symbolic Classification-Enabled LHC Limits Online BSM Global Fits

    Shehu AbdusSalam

    hep-phcs.LGcs.SCarXiv:2605.22330v12026
  32. A review of machine learning applications in wildfire science and management

    Piyush Jain, Sean C P Coogan, Sriram Ganapathi Subramanian +3

    cs.LGstat.MLarXiv:2003.00646v22020
  33. Benchmarking Composable Compression Techniques in Mixture-of-Experts LLMs

    Afsara Benazir, Chen Chen, Rongxiao Qu +3

    cs.LGarXiv:2608.21693v12026
  34. Contrastive Decoding: Open-ended Text Generation as Optimization

    Xiang Lisa Li, Ari Holtzman, Daniel Fried +5

    cs.CLcs.AIcs.LGarXiv:2210.15097v22022
  35. PC-DARTS: Partial Channel Connections for Memory-Efficient Architecture Search

    Yuhui Xu, Lingxi Xie, Xiaopeng Zhang +4

    cs.CVcs.LGarXiv:1907.05737v42019
  36. Reinforcement Learning on Benign Facts Amplifies Leakage of Memorized Private Data

    Renfei Zhang, Niloofar Mireshghallah

    cs.LGcs.AIarXiv:2608.21727v12026
  37. Learning by Cheating

    Dian Chen, Brady Zhou, Vladlen Koltun +1

    cs.ROcs.AIcs.CVarXiv:1912.12294v12019
  38. Measuring Robustness to Natural Distribution Shifts in Image Classification

    Rohan Taori, Achal Dave, Vaishaal Shankar +3

    cs.LGcs.CVstat.MLarXiv:2007.00644v22020
  39. Eureka: Human-Level Reward Design via Coding Large Language Models

    Yecheng Jason Ma, William Liang, Guanzhi Wang +6

    cs.ROcs.AIcs.LGarXiv:2310.12931v22023
  40. SiT: Exploring Flow and Diffusion-based Generative Models with Scalable Interpolant Transformers

    Nanye Ma, Mark Goldstein, Michael S. Albergo +3

    cs.CVcs.LGarXiv:2401.08740v22024
  41. Dreaming to Distill: Data-free Knowledge Transfer via DeepInversion

    Hongxu Yin, Pavlo Molchanov, Zhizhong Li +5

    cs.LGcs.CVstat.MLarXiv:1912.08795v22019
  42. Model of Models: When Does Emitting a Specialist Beat Attending, Adapting, or Tuning?

    John C. Howell

    cs.LGcs.AIarXiv:2608.21386v12026
  43. InterFaceGAN: Interpreting the Disentangled Face Representation Learned by GANs

    Yujun Shen, Ceyuan Yang, Xiaoou Tang +1

    cs.CVcs.LGeess.IVarXiv:2005.09635v22020
  44. A Multiscale Visualization of Attention in the Transformer Model

    Jesse Vig

    cs.HCcs.CLcs.LGarXiv:1906.05714v12019
  45. ADMIL: Attention-Distilled Multiple Instance Learning for Selective Foundation Model Inference in Pathology

    Duncan Stothers, Ren-Chin Wu, William Lotter

    cs.CVcs.AIcs.LGarXiv:2608.22066v12026
  46. Personalized and Aspiration-Oriented Career Path Recommendation

    Kuleshwar Sahu, Girish Keshav Palshikar, Rajiv Srivastava

    cs.LGarXiv:2608.22056v12026
  47. Adversarial Audio Synthesis

    Chris Donahue, Julian McAuley, Miller Puckette

    cs.SDcs.LGarXiv:1802.04208v32018
  48. Certified Data Removal from Machine Learning Models

    Chuan Guo, Tom Goldstein, Awni Hannun +1

    cs.LGstat.MLarXiv:1911.03030v62019
  49. The Real-World-Weight Cross-Entropy Loss Function: Modeling the Costs of Mislabeling

    Yaoshiang Ho, Samuel Wookey

    cs.LGcs.AIstat.MLarXiv:2001.00570v12020
  50. Agentic Scaffolding Amplifies Sycophantic Behavior in Large Language Models

    Thantham Jittham

    cs.CLcs.AIcs.LGarXiv:2608.21377v12026
  51. First-Principles Atomistic Structure and Dynamics of Polyethylene During High-Pressure Radical Polymerization via Machine Learning Force Fields

    Bharatha K. Gunawardana, Teresa Shah, Bicha Azizova +6

    cond-mat.mtrl-scicond-mat.dis-nncs.LGarXiv:2608.21741v12026
  52. Personalizing Session-based Recommendations with Hierarchical Recurrent Neural Networks

    Massimo Quadrana, Alexandros Karatzoglou, Balázs Hidasi +1

    cs.LGcs.HCcs.IRarXiv:1706.04148v52017
  53. DeepSense: A Unified Deep Learning Framework for Time-Series Mobile Sensing Data Processing

    Shuochao Yao, Shaohan Hu, Yiran Zhao +2

    cs.LGcs.NEcs.NIarXiv:1611.01942v22016
  54. Searching for Activation Functions

    Prajit Ramachandran, Barret Zoph, Quoc V. Le

    cs.NEcs.CVcs.LGarXiv:1710.05941v22017
  55. TPU v4: An Optically Reconfigurable Supercomputer for Machine Learning with Hardware Support for Embeddings

    Norman P. Jouppi, George Kurian, Sheng Li +11

    cs.ARcs.AIcs.LGarXiv:2304.01433v32023
  56. A review and comparison of strategies for multi-step ahead time series forecasting based on the NN5 forecasting competition

    Souhaib Ben Taieb, Gianluca Bontempi, Amir Atiya +1

    stat.MLcs.AIcs.LGarXiv:1108.3259v12011
  57. Stress Testing Unlearning Algorithms

    Noam Diamant, Ethan Fetaya, Neta Glazer

    cs.LGarXiv:2608.22527v12026
  58. MASH-Bench: Diagnosing Cross-Source Failure in Mass-Shooting Risk Classification

    Neha Sharma, Ritesh Sharma

    cs.LGcs.CYarXiv:2608.22460v12026
  59. StocBench: A Benchmark for Generative Modeling of Stochastic Dynamics

    Sebastian Pfister, Benjamin Holzschuh, Nils Thuerey

    cs.LGarXiv:2608.22309v12026
  60. Dual-Scale State-Space Modeling with Speaker-Wise Dynamic CRF for Speech Emotion Recognition in Conversation

    Guan-Hua Wen, Kuan-Yu Chen, Hou-Chiang Tseng

    cs.LGarXiv:2608.22399v12026