Machine Learning
Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
1,741 to 1,800 of 20,192
Lag-Llama: Towards Foundation Models for Probabilistic Time Series Forecasting
Kashif Rasul, Arjun Ashok, Andrew Robert Williams +15
cs.LGcs.AIarXiv:2310.08278v32023Training of Physical Neural Networks
Ali Momeni, Babak Rahmani, Benjamin Scellier +25
physics.app-phcs.LGarXiv:2406.03372v12024Not All Experts are Equal: Efficient Expert Pruning and Skipping for Mixture-of-Experts Large Language Models
Xudong Lu, Qi Liu, Yuhui Xu +5
cs.CLcs.AIcs.LGarXiv:2402.14800v22024Conformal Prediction with Large Language Models for Multi-Choice Question Answering
Bhawesh Kumar, Charlie Lu, Gauri Gupta +4
cs.CLcs.LGstat.MLarXiv:2305.18404v32023Learning-to-Cache: Accelerating Diffusion Transformer via Layer Caching
Xinyin Ma, Gongfan Fang, Michael Bi Mi +1
cs.LGcs.CVarXiv:2406.01733v22024Translation Artifacts in Cross-lingual Transfer Learning
Mikel Artetxe, Gorka Labaka, Eneko Agirre
cs.CLcs.LGarXiv:2004.04721v42020PriSTI: A Conditional Diffusion Framework for Spatiotemporal Imputation
Mingzhe Liu, Han Huang, Hao Feng +3
cs.LGarXiv:2302.09746v12023Correlated initialization of deep residual networks
Felix Benning, Ivan Nourdin, Giovanni Peccati
math.PRcs.LGstat.MLarXiv:2609.03589v12026BrepGen: A B-rep Generative Diffusion Model with Structured Latent Geometry
Xiang Xu, Joseph G. Lambourne, Pradeep Kumar Jayaraman +3
cs.CVcs.LGarXiv:2401.15563v32024Decoupling KL and Trajectories: A Unified Perspective for SFT, DAgger, Offline RL, and OPD in LLM Distillation
Anhao Zhao, Haoran Xin, Yingqi Fan +3
cs.LGcs.AIcs.CLarXiv:2605.16826v12026CycleResearcher: Improving Automated Research via Automated Review
Yixuan Weng, Minjun Zhu, Guangsheng Bao +4
cs.CLcs.AIcs.CYarXiv:2411.00816v32024GeoShapley: A Game Theory Approach to Measuring Spatial Effects in Machine Learning Models
Ziqi Li
cs.LGstat.MLarXiv:2312.03675v22023We're Different, We're the Same: Creative Homogeneity Across LLMs
Emily Wenger, Yoed Kenett
cs.CYcs.AIcs.CLarXiv:2501.19361v12025A Review of Deep Transfer Learning and Recent Advancements
Mohammadreza Iman, Khaled Rasheed, Hamid R. Arabnia
cs.LGcs.AIcs.CVarXiv:2201.09679v22022CMA-ES for Hyperparameter Optimization of Deep Neural Networks
Ilya Loshchilov, Frank Hutter
cs.NEcs.LGarXiv:1604.07269v12016Towards Bayesian Deep Learning: A Framework and Some Existing Methods
Hao Wang, Dit-Yan Yeung
stat.MLcs.CVcs.LGarXiv:1608.06884v22016Challenges in Benchmarking Stream Learning Algorithms with Real-world Data
Vinicius M. A. Souza, Denis M. dos Reis, Andre G. Maletzke +1
cs.LGstat.MLarXiv:2005.00113v22020The AI Scientist-v2: Workshop-Level Automated Scientific Discovery via Agentic Tree Search
Yutaro Yamada, Robert Tjarko Lange, Cong Lu +5
cs.AIcs.CLcs.LGarXiv:2504.08066v12025Summaries:한국어Sparse-RS: a versatile framework for query-efficient sparse black-box adversarial attacks
Francesco Croce, Maksym Andriushchenko, Naman D. Singh +2
cs.LGcs.CRcs.CVarXiv:2006.12834v32020Membership Inference Attack on Graph Neural Networks
Iyiola E. Olatunji, Wolfgang Nejdl, Megha Khosla
cs.LGcs.CRarXiv:2101.06570v32021BioT5: Enriching Cross-modal Integration in Biology with Chemical Knowledge and Natural Language Associations
Qizhi Pei, Wei Zhang, Jinhua Zhu +5
cs.CLcs.AIcs.LGarXiv:2310.07276v32023Exploring Contrastive Learning in Human Activity Recognition for Healthcare
Chi Ian Tang, Ignacio Perez-Pozuelo, Dimitris Spathis +1
cs.LGeess.SParXiv:2011.11542v32020L2CS-Net: Fine-Grained Gaze Estimation in Unconstrained Environments
Ahmed A. Abdelrahman, Thorsten Hempel, Aly Khalifa +1
cs.CVcs.LGcs.ROarXiv:2203.03339v12022Tensor Methods and Recommender Systems
Evgeny Frolov, Ivan Oseledets
cs.LGcs.IRstat.MLarXiv:1603.06038v22016The Causal-Neural Connection: Expressiveness, Learnability, and Inference
Kevin Xia, Kai-Zhan Lee, Yoshua Bengio +1
cs.LGcs.AIarXiv:2107.00793v32021On the Unreasonable Effectiveness of Feature propagation in Learning on Graphs with Missing Node Features
Emanuele Rossi, Henry Kenlay, Maria I. Gorinova +3
cs.LGarXiv:2111.12128v32021Soft-Attention Improves Skin Cancer Classification Performance
Soumyya Kanti Datta, Seyed Mohammad Abuzar Hashemi, Sargur N. Srihari +1
eess.IVcs.CVcs.LGarXiv:2105.03358v42021Learning Energy-Based Models by Diffusion Recovery Likelihood
Ruiqi Gao, Yang Song, Ben Poole +2
cs.LGstat.MLarXiv:2012.08125v22020Convex Tensor Decomposition via Structured Schatten Norm Regularization
Ryota Tomioka, Taiji Suzuki
stat.MLcs.LGmath.NAarXiv:1303.6370v12013Consistency Models Made Easy
Zhengyang Geng, Ashwini Pokle, William Luo +2
cs.LGcs.CVarXiv:2406.14548v22024B-Pref: Benchmarking Preference-Based Reinforcement Learning
Kimin Lee, Laura Smith, Anca Dragan +1
cs.LGcs.AIcs.HCarXiv:2111.03026v12021Supervised Raw Video Denoising with a Benchmark Dataset on Dynamic Scenes
Huanjing Yue, Cong Cao, Lei Liao +2
eess.IVcs.CVcs.LGarXiv:2003.14013v12020Learning Posterior Predictive Distributions for Node Classification from Synthetic Graph Priors
Jeongwhan Choi, Jongwoo Kim, Woosung Kang +1
cs.LGarXiv:2604.19028v12026On Accurate and Reliable Anomaly Detection for Gas Turbine Combustors: A Deep Learning Approach
Weizhong Yan, Lijie Yu
cs.LGstat.MLarXiv:1908.09238v12019Dobi-SVD: Differentiable SVD for LLM Compression and Some New Perspectives
Qinsi Wang, Jinghan Ke, Masayoshi Tomizuka +3
cs.LGarXiv:2502.02723v12025Large Language Models and the Reverse Turing Test
Terrence Sejnowski
cs.CLcs.AIcs.LGarXiv:2207.14382v92022Memory-Efficient Fine-Tuning of Compressed Large Language Models via sub-4-bit Integer Quantization
Jeonghoon Kim, Jung Hyun Lee, Sungdong Kim +4
cs.LGcs.AIarXiv:2305.14152v22023Deep learning for in vitro prediction of pharmaceutical formulations
Yilong Yang, Zhuyifan Ye, Yan Su +3
cs.LGstat.MLarXiv:1809.02069v12018Do-PFN: In-Context Learning for Causal Effect Estimation
Jake Robertson, Arik Reuter, Siyuan Guo +3
cs.LGarXiv:2506.06039v32025Assuring the Machine Learning Lifecycle: Desiderata, Methods, and Challenges
Rob Ashmore, Radu Calinescu, Colin Paterson
cs.LGcs.SEstat.MLarXiv:1905.04223v12019Fully-Connected Spatial-Temporal Graph for Multivariate Time-Series Data
Yucheng Wang, Yuecong Xu, Jianfei Yang +4
cs.LGarXiv:2309.05305v32023Fine Perceptive GANs for Brain MR Image Super-Resolution in Wavelet Domain
Senrong You, Yong Liu, Baiying Lei +1
eess.IVcs.CVcs.LGarXiv:2011.04145v12020A Multi-Objective Deep Reinforcement Learning Framework
Thanh Thi Nguyen, Ngoc Duy Nguyen, Peter Vamplew +3
cs.LGcs.AIstat.MLarXiv:1803.02965v32018UNR-Explainer: Counterfactual Explanations for Unsupervised Node Representation Learning Models
Hyunju Kang, Geonhee Han, Hogun Park
cs.LGcs.AIarXiv:2605.17285v12026Multiscale modeling of inelastic materials with Thermodynamics-based Artificial Neural Networks (TANN)
Filippo Masi, Ioannis Stefanou
cond-mat.mtrl-scics.CEcs.LGarXiv:2108.13137v32021A spelling correction model for end-to-end speech recognition
Jinxi Guo, Tara N. Sainath, Ron J. Weiss
eess.AScs.AIcs.CLarXiv:1902.07178v12019Semi-Supervised and Task-Driven Data Augmentation
Krishna Chaitanya, Neerav Karani, Christian Baumgartner +3
cs.CVcs.LGstat.MLarXiv:1902.05396v22019DeepScientist: Advancing Frontier-Pushing Scientific Findings Progressively
Yixuan Weng, Minjun Zhu, Qiujie Xie +4
cs.CLcs.LGarXiv:2509.26603v12025Theory-guided hard constraint projection (HCP): a knowledge-based data-driven scientific machine learning method
Yuntian Chen, Dou Huang, Dongxiao Zhang +4
cs.LGcs.AIarXiv:2012.06148v22020SceneGen: Learning to Generate Realistic Traffic Scenes
Shuhan Tan, Kelvin Wong, Shenlong Wang +3
cs.CVcs.AIcs.LGarXiv:2101.06541v12021What is a meaningful representation of protein sequences?
Nicki Skafte Detlefsen, Søren Hauberg, Wouter Boomsma
q-bio.BMcs.LGq-bio.QMarXiv:2012.02679v42020Towards a Theoretical Framework of Out-of-Distribution Generalization
Haotian Ye, Chuanlong Xie, Tianle Cai +3
cs.LGarXiv:2106.04496v32021PEORL: Integrating Symbolic Planning and Hierarchical Reinforcement Learning for Robust Decision-Making
Fangkai Yang, Daoming Lyu, Bo Liu +1
cs.LGcs.AIstat.MLarXiv:1804.07779v32018T1: Terminal Agent Reinforcement Learning for Long-Horizon Tasks
Junyao Yang, Yucheng Shi, Zhongzhi Li +4
cs.LGcs.AIarXiv:2609.11042v12026Machine Learning for Intrusion Detection in Industrial Control Systems: Applications, Challenges, and Recommendations
Muhammad Azmi Umer, Khurum Nazir Junejo, Muhammad Taha Jilani +1
cs.CRcs.LGarXiv:2202.11917v12022Autonomous Discovery of Unknown Reaction Pathways from Data by Chemical Reaction Neural Network
Weiqi Ji, Sili Deng
q-bio.MNcs.LGphysics.chem-pharXiv:2002.09062v22020Deep Learning for Free-Hand Sketch: A Survey
Peng Xu, Timothy M. Hospedales, Qiyue Yin +3
cs.CVcs.GRcs.LGarXiv:2001.02600v32020Training Deep Convolutional Neural Networks with Resistive Cross-Point Devices
Tayfun Gokmen, O. Murat Onen, Wilfried Haensch
cs.LGcs.NEstat.MLarXiv:1705.08014v12017Fairness risk measures
Robert C. Williamson, Aditya Krishna Menon
cs.LGstat.MLarXiv:1901.08665v12019Limitations of Lazy Training of Two-layers Neural Networks
Behrooz Ghorbani, Song Mei, Theodor Misiakiewicz +1
stat.MLcs.LGmath.STarXiv:1906.08899v12019