Machine Learning
Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
16,321 to 16,380 of 20,193
Data Predictability Shapes Weibull Weight-Scale Growth in Transformer Training
Tiexin Ding
cs.LGstat.MLarXiv:2608.23573v12026Decolonial AI: Decolonial Theory as Sociotechnical Foresight in Artificial Intelligence
Shakir Mohamed, Marie-Therese Png, William Isaac
cs.CYcs.AIcs.LGarXiv:2007.04068v12020InfoDPP-PAC: Principled Patch Selection for Whole Slide Image Analysis
Prateek Mittal, Ayush Srivastava, Joohi Chauhan
q-bio.QMcs.CVcs.ITarXiv:2608.23574v12026A mesh-free multiresolution deep energy method with phase-field modeling of brittle fracture
Han Zhang, Mehrisadat Makki Alamdari, Babak Shahbodagh +4
cs.LGmath.NAarXiv:2608.24126v12026Infant Care Video Dataset for Classification of Interventions Using Transformers
Igor Bogdanov, James Green
cs.CVcs.AIcs.LGarXiv:2608.23838v12026A Comparative Study in Surgical AI: Potential and Limitations of Data, Compute, and Scaling
Kirill Skobelev, Eric Fithian, Yegor Baranovski +9
cs.AIcs.CVcs.LGarXiv:2603.27341v42026Towards Universal Fake Image Detectors that Generalize Across Generative Models
Utkarsh Ojha, Yuheng Li, Yong Jae Lee
cs.CVcs.LGarXiv:2302.10174v22023When and why vision-language models behave like bags-of-words, and what to do about it?
Mert Yuksekgonul, Federico Bianchi, Pratyusha Kalluri +2
cs.CVcs.AIcs.CLarXiv:2210.01936v32022Explainable Machine Learning in Deployment
Umang Bhatt, Alice Xiang, Shubham Sharma +7
cs.LGcs.AIcs.CYarXiv:1909.06342v42019ChorusTIC: Training-Free Multivariate Time Series Classification via Chorus In-Context Learning
Juntao Fang, Shifeng Xie, Ruichu Cai +6
cs.LGcs.AIstat.MLarXiv:2608.24033v12026Equivariant Cellular Sheaves for Molecular Electronic Structure: Bridging Sheaf Cohomology and E(3)-Equivariant Hamiltonian Learning
Krishna Harish
cs.LGphysics.chem-pharXiv:2608.23571v12026A Survey of Deep Learning Applications to Autonomous Vehicle Control
Sampo Kuutti, Richard Bowden, Yaochu Jin +2
cs.LGcs.CVeess.SYarXiv:1912.10773v12019Measuring Calibration in Deep Learning
Jeremy Nixon, Mike Dusenberry, Ghassen Jerfel +4
cs.LGstat.MLarXiv:1904.01685v22019SecureBoost: A Lossless Federated Learning Framework
Kewei Cheng, Tao Fan, Yilun Jin +4
cs.LGstat.MLarXiv:1901.08755v32019LongVideoBench: A Benchmark for Long-context Interleaved Video-Language Understanding
Haoning Wu, Dongxu Li, Bei Chen +1
cs.CVcs.CLcs.LGarXiv:2407.15754v12024Malware Detection by Eating a Whole EXE
Edward Raff, Jon Barker, Jared Sylvester +3
stat.MLcs.CRcs.LGarXiv:1710.09435v12017Learning Deep Generative Models of Graphs
Yujia Li, Oriol Vinyals, Chris Dyer +2
cs.LGstat.MLarXiv:1803.03324v12018Applications of Deep Learning and Reinforcement Learning to Biological Data
Mufti Mahmud, M. Shamim Kaiser, Amir Hussain +1
cs.LGstat.MLarXiv:1711.03985v22017Backdoor Attacks on Decentralised Post-Training
Oğuzhan Ersoy, Nikolay Blagoev, Jona te Lintelo +3
cs.CRcs.LGarXiv:2604.02372v12026Asymmetric Non-local Neural Networks for Semantic Segmentation
Zhen Zhu, Mengde Xu, Song Bai +2
cs.CVcs.LGarXiv:1908.07678v52019The Computational Limits of Deep Learning
Neil C. Thompson, Kristjan Greenewald, Keeheon Lee +1
cs.LGstat.MLarXiv:2007.05558v22020Correcting Variable Importance Scored by Random Forests
Guancheng Zhou, Haiping Xu, Jason Liu +1
stat.MEcs.AIcs.LGarXiv:2606.10770v12026CALVIN: A Benchmark for Language-Conditioned Policy Learning for Long-Horizon Robot Manipulation Tasks
Oier Mees, Lukas Hermann, Erick Rosete-Beas +1
cs.ROcs.AIcs.CLarXiv:2112.03227v42021ReAct: Out-of-distribution Detection With Rectified Activations
Yiyou Sun, Chuan Guo, Yixuan Li
cs.LGarXiv:2111.12797v12021Lipschitz regularity of deep neural networks: analysis and efficient estimation
Kevin Scaman, Aladin Virmaux
stat.MLcs.LGarXiv:1805.10965v22018StyleGAN-XL: Scaling StyleGAN to Large Diverse Datasets
Axel Sauer, Katja Schwarz, Andreas Geiger
cs.LGcs.CVarXiv:2202.00273v22022NeuroPrefetcher: Storage-Aware Sparse LLM Inference via Delta Prefetching
Nobel Dhar, Md Romyull Islam, Xuechen Zhang +4
cs.DCcs.LGarXiv:2608.22643v12026Contrastive learning of global and local features for medical image segmentation with limited annotations
Krishna Chaitanya, Ertunc Erdil, Neerav Karani +1
cs.CVcs.LGeess.IVarXiv:2006.10511v22020Improving Diffusion Models for Inverse Problems using Manifold Constraints
Hyungjin Chung, Byeongsu Sim, Dohoon Ryu +1
cs.LGcs.AIcs.CVarXiv:2206.00941v32022Joint Extraction of Entities and Relations Based on a Novel Tagging Scheme
Suncong Zheng, Feng Wang, Hongyun Bao +3
cs.CLcs.AIcs.LGarXiv:1706.05075v12017Symbolic Classification-Enabled LHC Limits Online BSM Global Fits
Shehu AbdusSalam
hep-phcs.LGcs.SCarXiv:2605.22330v12026A review of machine learning applications in wildfire science and management
Piyush Jain, Sean C P Coogan, Sriram Ganapathi Subramanian +3
cs.LGstat.MLarXiv:2003.00646v22020Benchmarking Composable Compression Techniques in Mixture-of-Experts LLMs
Afsara Benazir, Chen Chen, Rongxiao Qu +3
cs.LGarXiv:2608.21693v12026Contrastive Decoding: Open-ended Text Generation as Optimization
Xiang Lisa Li, Ari Holtzman, Daniel Fried +5
cs.CLcs.AIcs.LGarXiv:2210.15097v22022PC-DARTS: Partial Channel Connections for Memory-Efficient Architecture Search
Yuhui Xu, Lingxi Xie, Xiaopeng Zhang +4
cs.CVcs.LGarXiv:1907.05737v42019Reinforcement Learning on Benign Facts Amplifies Leakage of Memorized Private Data
Renfei Zhang, Niloofar Mireshghallah
cs.LGcs.AIarXiv:2608.21727v12026Learning by Cheating
Dian Chen, Brady Zhou, Vladlen Koltun +1
cs.ROcs.AIcs.CVarXiv:1912.12294v12019Measuring Robustness to Natural Distribution Shifts in Image Classification
Rohan Taori, Achal Dave, Vaishaal Shankar +3
cs.LGcs.CVstat.MLarXiv:2007.00644v22020Eureka: Human-Level Reward Design via Coding Large Language Models
Yecheng Jason Ma, William Liang, Guanzhi Wang +6
cs.ROcs.AIcs.LGarXiv:2310.12931v22023SiT: Exploring Flow and Diffusion-based Generative Models with Scalable Interpolant Transformers
Nanye Ma, Mark Goldstein, Michael S. Albergo +3
cs.CVcs.LGarXiv:2401.08740v22024Dreaming to Distill: Data-free Knowledge Transfer via DeepInversion
Hongxu Yin, Pavlo Molchanov, Zhizhong Li +5
cs.LGcs.CVstat.MLarXiv:1912.08795v22019Model of Models: When Does Emitting a Specialist Beat Attending, Adapting, or Tuning?
John C. Howell
cs.LGcs.AIarXiv:2608.21386v12026InterFaceGAN: Interpreting the Disentangled Face Representation Learned by GANs
Yujun Shen, Ceyuan Yang, Xiaoou Tang +1
cs.CVcs.LGeess.IVarXiv:2005.09635v22020A Multiscale Visualization of Attention in the Transformer Model
Jesse Vig
cs.HCcs.CLcs.LGarXiv:1906.05714v12019ADMIL: Attention-Distilled Multiple Instance Learning for Selective Foundation Model Inference in Pathology
Duncan Stothers, Ren-Chin Wu, William Lotter
cs.CVcs.AIcs.LGarXiv:2608.22066v12026Personalized and Aspiration-Oriented Career Path Recommendation
Kuleshwar Sahu, Girish Keshav Palshikar, Rajiv Srivastava
cs.LGarXiv:2608.22056v12026Adversarial Audio Synthesis
Chris Donahue, Julian McAuley, Miller Puckette
cs.SDcs.LGarXiv:1802.04208v32018Certified Data Removal from Machine Learning Models
Chuan Guo, Tom Goldstein, Awni Hannun +1
cs.LGstat.MLarXiv:1911.03030v62019The Real-World-Weight Cross-Entropy Loss Function: Modeling the Costs of Mislabeling
Yaoshiang Ho, Samuel Wookey
cs.LGcs.AIstat.MLarXiv:2001.00570v12020Agentic Scaffolding Amplifies Sycophantic Behavior in Large Language Models
Thantham Jittham
cs.CLcs.AIcs.LGarXiv:2608.21377v12026First-Principles Atomistic Structure and Dynamics of Polyethylene During High-Pressure Radical Polymerization via Machine Learning Force Fields
Bharatha K. Gunawardana, Teresa Shah, Bicha Azizova +6
cond-mat.mtrl-scicond-mat.dis-nncs.LGarXiv:2608.21741v12026Personalizing Session-based Recommendations with Hierarchical Recurrent Neural Networks
Massimo Quadrana, Alexandros Karatzoglou, Balázs Hidasi +1
cs.LGcs.HCcs.IRarXiv:1706.04148v52017DeepSense: A Unified Deep Learning Framework for Time-Series Mobile Sensing Data Processing
Shuochao Yao, Shaohan Hu, Yiran Zhao +2
cs.LGcs.NEcs.NIarXiv:1611.01942v22016Searching for Activation Functions
Prajit Ramachandran, Barret Zoph, Quoc V. Le
cs.NEcs.CVcs.LGarXiv:1710.05941v22017TPU v4: An Optically Reconfigurable Supercomputer for Machine Learning with Hardware Support for Embeddings
Norman P. Jouppi, George Kurian, Sheng Li +11
cs.ARcs.AIcs.LGarXiv:2304.01433v32023A review and comparison of strategies for multi-step ahead time series forecasting based on the NN5 forecasting competition
Souhaib Ben Taieb, Gianluca Bontempi, Amir Atiya +1
stat.MLcs.AIcs.LGarXiv:1108.3259v12011Stress Testing Unlearning Algorithms
Noam Diamant, Ethan Fetaya, Neta Glazer
cs.LGarXiv:2608.22527v12026MASH-Bench: Diagnosing Cross-Source Failure in Mass-Shooting Risk Classification
Neha Sharma, Ritesh Sharma
cs.LGcs.CYarXiv:2608.22460v12026StocBench: A Benchmark for Generative Modeling of Stochastic Dynamics
Sebastian Pfister, Benjamin Holzschuh, Nils Thuerey
cs.LGarXiv:2608.22309v12026Dual-Scale State-Space Modeling with Speaker-Wise Dynamic CRF for Speech Emotion Recognition in Conversation
Guan-Hua Wen, Kuan-Yu Chen, Hou-Chiang Tseng
cs.LGarXiv:2608.22399v12026