Machine Learning
Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
16,561 to 16,620 of 20,454
Professor Forcing: A New Algorithm for Training Recurrent Networks
Alex Lamb, Anirudh Goyal, Ying Zhang +3
stat.MLcs.LGarXiv:1610.09038v12016Machine Learning DDoS Detection for Consumer Internet of Things Devices
Rohan Doshi, Noah Apthorpe, Nick Feamster
cs.CRcs.LGarXiv:1804.04159v12018Scaling DoRA: High-Rank Adaptation via Factored Norms and Fused Kernels
Alexandra Zelenin, Alexandra Zhuravlyova
cs.LGstat.MLarXiv:2603.22276v12026Understanding over-squashing and bottlenecks on graphs via curvature
Jake Topping, Francesco Di Giovanni, Benjamin Paul Chamberlain +2
stat.MLcs.LGarXiv:2111.14522v32021Formal Verification of Piece-Wise Linear Feed-Forward Neural Networks
Ruediger Ehlers
cs.LOcs.AIcs.LGarXiv:1705.01320v32017Merging Models with Fisher-Weighted Averaging
Michael Matena, Colin Raffel
cs.LGarXiv:2111.09832v22021Blockwise Stabilized Adaptive Cubic Regularization with Subsolvers via Recurrence
Rodion Podorozhny
cs.LGmath.NAarXiv:2608.22129v22026Mobile-Former: Bridging MobileNet and Transformer
Yinpeng Chen, Xiyang Dai, Dongdong Chen +4
cs.CVcs.LGarXiv:2108.05895v32021The role of explainability in creating trustworthy artificial intelligence for health care: a comprehensive survey of the terminology, design choices, and evaluation strategies
Aniek F. Markus, Jan A. Kors, Peter R. Rijnbeek
cs.AIcs.LGstat.MLarXiv:2007.15911v22020Right for the Right Reasons: Training Differentiable Models by Constraining their Explanations
Andrew Slavin Ross, Michael C. Hughes, Finale Doshi-Velez
cs.LGcs.AIstat.MLarXiv:1703.03717v22017Explainable Prediction of Medical Codes from Clinical Text
James Mullenbach, Sarah Wiegreffe, Jon Duke +2
cs.CLcs.LGstat.MLarXiv:1802.05695v22018Provably Efficient Reinforcement Learning with Linear Function Approximation
Chi Jin, Zhuoran Yang, Zhaoran Wang +1
cs.LGmath.OCstat.MLarXiv:1907.05388v22019Deep Learning for Time Series Anomaly Detection: A Survey
Zahra Zamanzadeh Darban, Geoffrey I. Webb, Shirui Pan +2
cs.LGcs.AIarXiv:2211.05244v32022RLAIF vs. RLHF: Scaling Reinforcement Learning from Human Feedback with AI Feedback
Harrison Lee, Samrat Phatale, Hassan Mansoor +8
cs.CLcs.AIcs.LGarXiv:2309.00267v32023Deep learning with noisy labels: exploring techniques and remedies in medical image analysis
Davood Karimi, Haoran Dou, Simon K. Warfield +1
cs.CVcs.LGeess.IVarXiv:1912.02911v42019DAG-GNN: DAG Structure Learning with Graph Neural Networks
Yue Yu, Jie Chen, Tian Gao +1
cs.LGcs.AIstat.MLarXiv:1904.10098v12019Optimal Ratio for Data Splitting
V. Roshan Joseph
stat.MLcs.LGarXiv:2202.03326v12022Online Continual Learning with Maximally Interfered Retrieval
Rahaf Aljundi, Lucas Caccia, Eugene Belilovsky +4
cs.LGstat.MLarXiv:1908.04742v32019How is ChatGPT's behavior changing over time?
Lingjiao Chen, Matei Zaharia, James Zou
cs.CLcs.AIcs.LGarXiv:2307.09009v32023"Do Anything Now": Characterizing and Evaluating In-The-Wild Jailbreak Prompts on Large Language Models
Xinyue Shen, Zeyuan Chen, Michael Backes +2
cs.CRcs.LGarXiv:2308.03825v22023VirtualHome: Simulating Household Activities via Programs
Xavier Puig, Kevin Ra, Marko Boben +4
cs.CVcs.AIcs.LGarXiv:1806.07011v12018Unified Training of Universal Time Series Forecasting Transformers
Gerald Woo, Chenghao Liu, Akshat Kumar +3
cs.LGcs.AIarXiv:2402.02592v22024SuperSpike: Supervised learning in multi-layer spiking neural networks
Friedemann Zenke, Surya Ganguli
q-bio.NCcs.LGcs.NEarXiv:1705.11146v22017Data Predictability Shapes Weibull Weight-Scale Growth in Transformer Training
Tiexin Ding
cs.LGstat.MLarXiv:2608.23573v12026Decolonial AI: Decolonial Theory as Sociotechnical Foresight in Artificial Intelligence
Shakir Mohamed, Marie-Therese Png, William Isaac
cs.CYcs.AIcs.LGarXiv:2007.04068v12020InfoDPP-PAC: Principled Patch Selection for Whole Slide Image Analysis
Prateek Mittal, Ayush Srivastava, Joohi Chauhan
q-bio.QMcs.CVcs.ITarXiv:2608.23574v12026A mesh-free multiresolution deep energy method with phase-field modeling of brittle fracture
Han Zhang, Mehrisadat Makki Alamdari, Babak Shahbodagh +4
cs.LGmath.NAarXiv:2608.24126v12026Infant Care Video Dataset for Classification of Interventions Using Transformers
Igor Bogdanov, James Green
cs.CVcs.AIcs.LGarXiv:2608.23838v12026A Comparative Study in Surgical AI: Potential and Limitations of Data, Compute, and Scaling
Kirill Skobelev, Eric Fithian, Yegor Baranovski +9
cs.AIcs.CVcs.LGarXiv:2603.27341v42026Towards Universal Fake Image Detectors that Generalize Across Generative Models
Utkarsh Ojha, Yuheng Li, Yong Jae Lee
cs.CVcs.LGarXiv:2302.10174v22023When and why vision-language models behave like bags-of-words, and what to do about it?
Mert Yuksekgonul, Federico Bianchi, Pratyusha Kalluri +2
cs.CVcs.AIcs.CLarXiv:2210.01936v32022Explainable Machine Learning in Deployment
Umang Bhatt, Alice Xiang, Shubham Sharma +7
cs.LGcs.AIcs.CYarXiv:1909.06342v42019ChorusTIC: Training-Free Multivariate Time Series Classification via Chorus In-Context Learning
Juntao Fang, Shifeng Xie, Ruichu Cai +6
cs.LGcs.AIstat.MLarXiv:2608.24033v12026Equivariant Cellular Sheaves for Molecular Electronic Structure: Bridging Sheaf Cohomology and E(3)-Equivariant Hamiltonian Learning
Krishna Harish
cs.LGphysics.chem-pharXiv:2608.23571v12026A Survey of Deep Learning Applications to Autonomous Vehicle Control
Sampo Kuutti, Richard Bowden, Yaochu Jin +2
cs.LGcs.CVeess.SYarXiv:1912.10773v12019Measuring Calibration in Deep Learning
Jeremy Nixon, Mike Dusenberry, Ghassen Jerfel +4
cs.LGstat.MLarXiv:1904.01685v22019SecureBoost: A Lossless Federated Learning Framework
Kewei Cheng, Tao Fan, Yilun Jin +4
cs.LGstat.MLarXiv:1901.08755v32019LongVideoBench: A Benchmark for Long-context Interleaved Video-Language Understanding
Haoning Wu, Dongxu Li, Bei Chen +1
cs.CVcs.CLcs.LGarXiv:2407.15754v12024Malware Detection by Eating a Whole EXE
Edward Raff, Jon Barker, Jared Sylvester +3
stat.MLcs.CRcs.LGarXiv:1710.09435v12017Learning Deep Generative Models of Graphs
Yujia Li, Oriol Vinyals, Chris Dyer +2
cs.LGstat.MLarXiv:1803.03324v12018Applications of Deep Learning and Reinforcement Learning to Biological Data
Mufti Mahmud, M. Shamim Kaiser, Amir Hussain +1
cs.LGstat.MLarXiv:1711.03985v22017Backdoor Attacks on Decentralised Post-Training
Oğuzhan Ersoy, Nikolay Blagoev, Jona te Lintelo +3
cs.CRcs.LGarXiv:2604.02372v12026Asymmetric Non-local Neural Networks for Semantic Segmentation
Zhen Zhu, Mengde Xu, Song Bai +2
cs.CVcs.LGarXiv:1908.07678v52019The Computational Limits of Deep Learning
Neil C. Thompson, Kristjan Greenewald, Keeheon Lee +1
cs.LGstat.MLarXiv:2007.05558v22020Correcting Variable Importance Scored by Random Forests
Guancheng Zhou, Haiping Xu, Jason Liu +1
stat.MEcs.AIcs.LGarXiv:2606.10770v12026CALVIN: A Benchmark for Language-Conditioned Policy Learning for Long-Horizon Robot Manipulation Tasks
Oier Mees, Lukas Hermann, Erick Rosete-Beas +1
cs.ROcs.AIcs.CLarXiv:2112.03227v42021ReAct: Out-of-distribution Detection With Rectified Activations
Yiyou Sun, Chuan Guo, Yixuan Li
cs.LGarXiv:2111.12797v12021Lipschitz regularity of deep neural networks: analysis and efficient estimation
Kevin Scaman, Aladin Virmaux
stat.MLcs.LGarXiv:1805.10965v22018StyleGAN-XL: Scaling StyleGAN to Large Diverse Datasets
Axel Sauer, Katja Schwarz, Andreas Geiger
cs.LGcs.CVarXiv:2202.00273v22022NeuroPrefetcher: Storage-Aware Sparse LLM Inference via Delta Prefetching
Nobel Dhar, Md Romyull Islam, Xuechen Zhang +4
cs.DCcs.LGarXiv:2608.22643v12026Contrastive learning of global and local features for medical image segmentation with limited annotations
Krishna Chaitanya, Ertunc Erdil, Neerav Karani +1
cs.CVcs.LGeess.IVarXiv:2006.10511v22020Improving Diffusion Models for Inverse Problems using Manifold Constraints
Hyungjin Chung, Byeongsu Sim, Dohoon Ryu +1
cs.LGcs.AIcs.CVarXiv:2206.00941v32022Joint Extraction of Entities and Relations Based on a Novel Tagging Scheme
Suncong Zheng, Feng Wang, Hongyun Bao +3
cs.CLcs.AIcs.LGarXiv:1706.05075v12017Symbolic Classification-Enabled LHC Limits Online BSM Global Fits
Shehu AbdusSalam
hep-phcs.LGcs.SCarXiv:2605.22330v12026A review of machine learning applications in wildfire science and management
Piyush Jain, Sean C P Coogan, Sriram Ganapathi Subramanian +3
cs.LGstat.MLarXiv:2003.00646v22020Benchmarking Composable Compression Techniques in Mixture-of-Experts LLMs
Afsara Benazir, Chen Chen, Rongxiao Qu +3
cs.LGarXiv:2608.21693v12026Contrastive Decoding: Open-ended Text Generation as Optimization
Xiang Lisa Li, Ari Holtzman, Daniel Fried +5
cs.CLcs.AIcs.LGarXiv:2210.15097v22022PC-DARTS: Partial Channel Connections for Memory-Efficient Architecture Search
Yuhui Xu, Lingxi Xie, Xiaopeng Zhang +4
cs.CVcs.LGarXiv:1907.05737v42019Reinforcement Learning on Benign Facts Amplifies Leakage of Memorized Private Data
Renfei Zhang, Niloofar Mireshghallah
cs.LGcs.AIarXiv:2608.21727v12026Learning by Cheating
Dian Chen, Brady Zhou, Vladlen Koltun +1
cs.ROcs.AIcs.CVarXiv:1912.12294v12019