Machine Learning

Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

16,561 to 16,620 of 20,454

  1. Professor Forcing: A New Algorithm for Training Recurrent Networks

    Alex Lamb, Anirudh Goyal, Ying Zhang +3

    stat.MLcs.LGarXiv:1610.09038v12016
  2. Machine Learning DDoS Detection for Consumer Internet of Things Devices

    Rohan Doshi, Noah Apthorpe, Nick Feamster

    cs.CRcs.LGarXiv:1804.04159v12018
  3. Scaling DoRA: High-Rank Adaptation via Factored Norms and Fused Kernels

    Alexandra Zelenin, Alexandra Zhuravlyova

    cs.LGstat.MLarXiv:2603.22276v12026
  4. Understanding over-squashing and bottlenecks on graphs via curvature

    Jake Topping, Francesco Di Giovanni, Benjamin Paul Chamberlain +2

    stat.MLcs.LGarXiv:2111.14522v32021
  5. Formal Verification of Piece-Wise Linear Feed-Forward Neural Networks

    Ruediger Ehlers

    cs.LOcs.AIcs.LGarXiv:1705.01320v32017
  6. Merging Models with Fisher-Weighted Averaging

    Michael Matena, Colin Raffel

    cs.LGarXiv:2111.09832v22021
  7. Blockwise Stabilized Adaptive Cubic Regularization with Subsolvers via Recurrence

    Rodion Podorozhny

    cs.LGmath.NAarXiv:2608.22129v22026
  8. Mobile-Former: Bridging MobileNet and Transformer

    Yinpeng Chen, Xiyang Dai, Dongdong Chen +4

    cs.CVcs.LGarXiv:2108.05895v32021
  9. The role of explainability in creating trustworthy artificial intelligence for health care: a comprehensive survey of the terminology, design choices, and evaluation strategies

    Aniek F. Markus, Jan A. Kors, Peter R. Rijnbeek

    cs.AIcs.LGstat.MLarXiv:2007.15911v22020
  10. Right for the Right Reasons: Training Differentiable Models by Constraining their Explanations

    Andrew Slavin Ross, Michael C. Hughes, Finale Doshi-Velez

    cs.LGcs.AIstat.MLarXiv:1703.03717v22017
  11. Explainable Prediction of Medical Codes from Clinical Text

    James Mullenbach, Sarah Wiegreffe, Jon Duke +2

    cs.CLcs.LGstat.MLarXiv:1802.05695v22018
  12. Provably Efficient Reinforcement Learning with Linear Function Approximation

    Chi Jin, Zhuoran Yang, Zhaoran Wang +1

    cs.LGmath.OCstat.MLarXiv:1907.05388v22019
  13. Deep Learning for Time Series Anomaly Detection: A Survey

    Zahra Zamanzadeh Darban, Geoffrey I. Webb, Shirui Pan +2

    cs.LGcs.AIarXiv:2211.05244v32022
  14. RLAIF vs. RLHF: Scaling Reinforcement Learning from Human Feedback with AI Feedback

    Harrison Lee, Samrat Phatale, Hassan Mansoor +8

    cs.CLcs.AIcs.LGarXiv:2309.00267v32023
  15. Deep learning with noisy labels: exploring techniques and remedies in medical image analysis

    Davood Karimi, Haoran Dou, Simon K. Warfield +1

    cs.CVcs.LGeess.IVarXiv:1912.02911v42019
  16. DAG-GNN: DAG Structure Learning with Graph Neural Networks

    Yue Yu, Jie Chen, Tian Gao +1

    cs.LGcs.AIstat.MLarXiv:1904.10098v12019
  17. Optimal Ratio for Data Splitting

    V. Roshan Joseph

    stat.MLcs.LGarXiv:2202.03326v12022
  18. Online Continual Learning with Maximally Interfered Retrieval

    Rahaf Aljundi, Lucas Caccia, Eugene Belilovsky +4

    cs.LGstat.MLarXiv:1908.04742v32019
  19. How is ChatGPT's behavior changing over time?

    Lingjiao Chen, Matei Zaharia, James Zou

    cs.CLcs.AIcs.LGarXiv:2307.09009v32023
  20. "Do Anything Now": Characterizing and Evaluating In-The-Wild Jailbreak Prompts on Large Language Models

    Xinyue Shen, Zeyuan Chen, Michael Backes +2

    cs.CRcs.LGarXiv:2308.03825v22023
  21. VirtualHome: Simulating Household Activities via Programs

    Xavier Puig, Kevin Ra, Marko Boben +4

    cs.CVcs.AIcs.LGarXiv:1806.07011v12018
  22. Unified Training of Universal Time Series Forecasting Transformers

    Gerald Woo, Chenghao Liu, Akshat Kumar +3

    cs.LGcs.AIarXiv:2402.02592v22024
  23. SuperSpike: Supervised learning in multi-layer spiking neural networks

    Friedemann Zenke, Surya Ganguli

    q-bio.NCcs.LGcs.NEarXiv:1705.11146v22017
  24. Data Predictability Shapes Weibull Weight-Scale Growth in Transformer Training

    Tiexin Ding

    cs.LGstat.MLarXiv:2608.23573v12026
  25. Decolonial AI: Decolonial Theory as Sociotechnical Foresight in Artificial Intelligence

    Shakir Mohamed, Marie-Therese Png, William Isaac

    cs.CYcs.AIcs.LGarXiv:2007.04068v12020
  26. InfoDPP-PAC: Principled Patch Selection for Whole Slide Image Analysis

    Prateek Mittal, Ayush Srivastava, Joohi Chauhan

    q-bio.QMcs.CVcs.ITarXiv:2608.23574v12026
  27. A mesh-free multiresolution deep energy method with phase-field modeling of brittle fracture

    Han Zhang, Mehrisadat Makki Alamdari, Babak Shahbodagh +4

    cs.LGmath.NAarXiv:2608.24126v12026
  28. Infant Care Video Dataset for Classification of Interventions Using Transformers

    Igor Bogdanov, James Green

    cs.CVcs.AIcs.LGarXiv:2608.23838v12026
  29. A Comparative Study in Surgical AI: Potential and Limitations of Data, Compute, and Scaling

    Kirill Skobelev, Eric Fithian, Yegor Baranovski +9

    cs.AIcs.CVcs.LGarXiv:2603.27341v42026
  30. Towards Universal Fake Image Detectors that Generalize Across Generative Models

    Utkarsh Ojha, Yuheng Li, Yong Jae Lee

    cs.CVcs.LGarXiv:2302.10174v22023
  31. When and why vision-language models behave like bags-of-words, and what to do about it?

    Mert Yuksekgonul, Federico Bianchi, Pratyusha Kalluri +2

    cs.CVcs.AIcs.CLarXiv:2210.01936v32022
  32. Explainable Machine Learning in Deployment

    Umang Bhatt, Alice Xiang, Shubham Sharma +7

    cs.LGcs.AIcs.CYarXiv:1909.06342v42019
  33. ChorusTIC: Training-Free Multivariate Time Series Classification via Chorus In-Context Learning

    Juntao Fang, Shifeng Xie, Ruichu Cai +6

    cs.LGcs.AIstat.MLarXiv:2608.24033v12026
  34. Equivariant Cellular Sheaves for Molecular Electronic Structure: Bridging Sheaf Cohomology and E(3)-Equivariant Hamiltonian Learning

    Krishna Harish

    cs.LGphysics.chem-pharXiv:2608.23571v12026
  35. A Survey of Deep Learning Applications to Autonomous Vehicle Control

    Sampo Kuutti, Richard Bowden, Yaochu Jin +2

    cs.LGcs.CVeess.SYarXiv:1912.10773v12019
  36. Measuring Calibration in Deep Learning

    Jeremy Nixon, Mike Dusenberry, Ghassen Jerfel +4

    cs.LGstat.MLarXiv:1904.01685v22019
  37. SecureBoost: A Lossless Federated Learning Framework

    Kewei Cheng, Tao Fan, Yilun Jin +4

    cs.LGstat.MLarXiv:1901.08755v32019
  38. LongVideoBench: A Benchmark for Long-context Interleaved Video-Language Understanding

    Haoning Wu, Dongxu Li, Bei Chen +1

    cs.CVcs.CLcs.LGarXiv:2407.15754v12024
  39. Malware Detection by Eating a Whole EXE

    Edward Raff, Jon Barker, Jared Sylvester +3

    stat.MLcs.CRcs.LGarXiv:1710.09435v12017
  40. Learning Deep Generative Models of Graphs

    Yujia Li, Oriol Vinyals, Chris Dyer +2

    cs.LGstat.MLarXiv:1803.03324v12018
  41. Applications of Deep Learning and Reinforcement Learning to Biological Data

    Mufti Mahmud, M. Shamim Kaiser, Amir Hussain +1

    cs.LGstat.MLarXiv:1711.03985v22017
  42. Backdoor Attacks on Decentralised Post-Training

    Oğuzhan Ersoy, Nikolay Blagoev, Jona te Lintelo +3

    cs.CRcs.LGarXiv:2604.02372v12026
  43. Asymmetric Non-local Neural Networks for Semantic Segmentation

    Zhen Zhu, Mengde Xu, Song Bai +2

    cs.CVcs.LGarXiv:1908.07678v52019
  44. The Computational Limits of Deep Learning

    Neil C. Thompson, Kristjan Greenewald, Keeheon Lee +1

    cs.LGstat.MLarXiv:2007.05558v22020
  45. Correcting Variable Importance Scored by Random Forests

    Guancheng Zhou, Haiping Xu, Jason Liu +1

    stat.MEcs.AIcs.LGarXiv:2606.10770v12026
  46. CALVIN: A Benchmark for Language-Conditioned Policy Learning for Long-Horizon Robot Manipulation Tasks

    Oier Mees, Lukas Hermann, Erick Rosete-Beas +1

    cs.ROcs.AIcs.CLarXiv:2112.03227v42021
  47. ReAct: Out-of-distribution Detection With Rectified Activations

    Yiyou Sun, Chuan Guo, Yixuan Li

    cs.LGarXiv:2111.12797v12021
  48. Lipschitz regularity of deep neural networks: analysis and efficient estimation

    Kevin Scaman, Aladin Virmaux

    stat.MLcs.LGarXiv:1805.10965v22018
  49. StyleGAN-XL: Scaling StyleGAN to Large Diverse Datasets

    Axel Sauer, Katja Schwarz, Andreas Geiger

    cs.LGcs.CVarXiv:2202.00273v22022
  50. NeuroPrefetcher: Storage-Aware Sparse LLM Inference via Delta Prefetching

    Nobel Dhar, Md Romyull Islam, Xuechen Zhang +4

    cs.DCcs.LGarXiv:2608.22643v12026
  51. Contrastive learning of global and local features for medical image segmentation with limited annotations

    Krishna Chaitanya, Ertunc Erdil, Neerav Karani +1

    cs.CVcs.LGeess.IVarXiv:2006.10511v22020
  52. Improving Diffusion Models for Inverse Problems using Manifold Constraints

    Hyungjin Chung, Byeongsu Sim, Dohoon Ryu +1

    cs.LGcs.AIcs.CVarXiv:2206.00941v32022
  53. Joint Extraction of Entities and Relations Based on a Novel Tagging Scheme

    Suncong Zheng, Feng Wang, Hongyun Bao +3

    cs.CLcs.AIcs.LGarXiv:1706.05075v12017
  54. Symbolic Classification-Enabled LHC Limits Online BSM Global Fits

    Shehu AbdusSalam

    hep-phcs.LGcs.SCarXiv:2605.22330v12026
  55. A review of machine learning applications in wildfire science and management

    Piyush Jain, Sean C P Coogan, Sriram Ganapathi Subramanian +3

    cs.LGstat.MLarXiv:2003.00646v22020
  56. Benchmarking Composable Compression Techniques in Mixture-of-Experts LLMs

    Afsara Benazir, Chen Chen, Rongxiao Qu +3

    cs.LGarXiv:2608.21693v12026
  57. Contrastive Decoding: Open-ended Text Generation as Optimization

    Xiang Lisa Li, Ari Holtzman, Daniel Fried +5

    cs.CLcs.AIcs.LGarXiv:2210.15097v22022
  58. PC-DARTS: Partial Channel Connections for Memory-Efficient Architecture Search

    Yuhui Xu, Lingxi Xie, Xiaopeng Zhang +4

    cs.CVcs.LGarXiv:1907.05737v42019
  59. Reinforcement Learning on Benign Facts Amplifies Leakage of Memorized Private Data

    Renfei Zhang, Niloofar Mireshghallah

    cs.LGcs.AIarXiv:2608.21727v12026
  60. Learning by Cheating

    Dian Chen, Brady Zhou, Vladlen Koltun +1

    cs.ROcs.AIcs.CVarXiv:1912.12294v12019