Machine Learning

Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

1,741 to 1,800 of 20,192

  1. Lag-Llama: Towards Foundation Models for Probabilistic Time Series Forecasting

    Kashif Rasul, Arjun Ashok, Andrew Robert Williams +15

    cs.LGcs.AIarXiv:2310.08278v32023
  2. Training of Physical Neural Networks

    Ali Momeni, Babak Rahmani, Benjamin Scellier +25

    physics.app-phcs.LGarXiv:2406.03372v12024
  3. Not All Experts are Equal: Efficient Expert Pruning and Skipping for Mixture-of-Experts Large Language Models

    Xudong Lu, Qi Liu, Yuhui Xu +5

    cs.CLcs.AIcs.LGarXiv:2402.14800v22024
  4. Conformal Prediction with Large Language Models for Multi-Choice Question Answering

    Bhawesh Kumar, Charlie Lu, Gauri Gupta +4

    cs.CLcs.LGstat.MLarXiv:2305.18404v32023
  5. Learning-to-Cache: Accelerating Diffusion Transformer via Layer Caching

    Xinyin Ma, Gongfan Fang, Michael Bi Mi +1

    cs.LGcs.CVarXiv:2406.01733v22024
  6. Translation Artifacts in Cross-lingual Transfer Learning

    Mikel Artetxe, Gorka Labaka, Eneko Agirre

    cs.CLcs.LGarXiv:2004.04721v42020
  7. PriSTI: A Conditional Diffusion Framework for Spatiotemporal Imputation

    Mingzhe Liu, Han Huang, Hao Feng +3

    cs.LGarXiv:2302.09746v12023
  8. Correlated initialization of deep residual networks

    Felix Benning, Ivan Nourdin, Giovanni Peccati

    math.PRcs.LGstat.MLarXiv:2609.03589v12026
  9. BrepGen: A B-rep Generative Diffusion Model with Structured Latent Geometry

    Xiang Xu, Joseph G. Lambourne, Pradeep Kumar Jayaraman +3

    cs.CVcs.LGarXiv:2401.15563v32024
  10. Decoupling KL and Trajectories: A Unified Perspective for SFT, DAgger, Offline RL, and OPD in LLM Distillation

    Anhao Zhao, Haoran Xin, Yingqi Fan +3

    cs.LGcs.AIcs.CLarXiv:2605.16826v12026
  11. CycleResearcher: Improving Automated Research via Automated Review

    Yixuan Weng, Minjun Zhu, Guangsheng Bao +4

    cs.CLcs.AIcs.CYarXiv:2411.00816v32024
  12. GeoShapley: A Game Theory Approach to Measuring Spatial Effects in Machine Learning Models

    Ziqi Li

    cs.LGstat.MLarXiv:2312.03675v22023
  13. We're Different, We're the Same: Creative Homogeneity Across LLMs

    Emily Wenger, Yoed Kenett

    cs.CYcs.AIcs.CLarXiv:2501.19361v12025
  14. A Review of Deep Transfer Learning and Recent Advancements

    Mohammadreza Iman, Khaled Rasheed, Hamid R. Arabnia

    cs.LGcs.AIcs.CVarXiv:2201.09679v22022
  15. CMA-ES for Hyperparameter Optimization of Deep Neural Networks

    Ilya Loshchilov, Frank Hutter

    cs.NEcs.LGarXiv:1604.07269v12016
  16. Towards Bayesian Deep Learning: A Framework and Some Existing Methods

    Hao Wang, Dit-Yan Yeung

    stat.MLcs.CVcs.LGarXiv:1608.06884v22016
  17. Challenges in Benchmarking Stream Learning Algorithms with Real-world Data

    Vinicius M. A. Souza, Denis M. dos Reis, Andre G. Maletzke +1

    cs.LGstat.MLarXiv:2005.00113v22020
  18. The AI Scientist-v2: Workshop-Level Automated Scientific Discovery via Agentic Tree Search

    Yutaro Yamada, Robert Tjarko Lange, Cong Lu +5

    cs.AIcs.CLcs.LGarXiv:2504.08066v12025
    Summaries:한국어
  19. Sparse-RS: a versatile framework for query-efficient sparse black-box adversarial attacks

    Francesco Croce, Maksym Andriushchenko, Naman D. Singh +2

    cs.LGcs.CRcs.CVarXiv:2006.12834v32020
  20. Membership Inference Attack on Graph Neural Networks

    Iyiola E. Olatunji, Wolfgang Nejdl, Megha Khosla

    cs.LGcs.CRarXiv:2101.06570v32021
  21. BioT5: Enriching Cross-modal Integration in Biology with Chemical Knowledge and Natural Language Associations

    Qizhi Pei, Wei Zhang, Jinhua Zhu +5

    cs.CLcs.AIcs.LGarXiv:2310.07276v32023
  22. Exploring Contrastive Learning in Human Activity Recognition for Healthcare

    Chi Ian Tang, Ignacio Perez-Pozuelo, Dimitris Spathis +1

    cs.LGeess.SParXiv:2011.11542v32020
  23. L2CS-Net: Fine-Grained Gaze Estimation in Unconstrained Environments

    Ahmed A. Abdelrahman, Thorsten Hempel, Aly Khalifa +1

    cs.CVcs.LGcs.ROarXiv:2203.03339v12022
  24. Tensor Methods and Recommender Systems

    Evgeny Frolov, Ivan Oseledets

    cs.LGcs.IRstat.MLarXiv:1603.06038v22016
  25. The Causal-Neural Connection: Expressiveness, Learnability, and Inference

    Kevin Xia, Kai-Zhan Lee, Yoshua Bengio +1

    cs.LGcs.AIarXiv:2107.00793v32021
  26. On the Unreasonable Effectiveness of Feature propagation in Learning on Graphs with Missing Node Features

    Emanuele Rossi, Henry Kenlay, Maria I. Gorinova +3

    cs.LGarXiv:2111.12128v32021
  27. Soft-Attention Improves Skin Cancer Classification Performance

    Soumyya Kanti Datta, Seyed Mohammad Abuzar Hashemi, Sargur N. Srihari +1

    eess.IVcs.CVcs.LGarXiv:2105.03358v42021
  28. Learning Energy-Based Models by Diffusion Recovery Likelihood

    Ruiqi Gao, Yang Song, Ben Poole +2

    cs.LGstat.MLarXiv:2012.08125v22020
  29. Convex Tensor Decomposition via Structured Schatten Norm Regularization

    Ryota Tomioka, Taiji Suzuki

    stat.MLcs.LGmath.NAarXiv:1303.6370v12013
  30. Consistency Models Made Easy

    Zhengyang Geng, Ashwini Pokle, William Luo +2

    cs.LGcs.CVarXiv:2406.14548v22024
  31. B-Pref: Benchmarking Preference-Based Reinforcement Learning

    Kimin Lee, Laura Smith, Anca Dragan +1

    cs.LGcs.AIcs.HCarXiv:2111.03026v12021
  32. Supervised Raw Video Denoising with a Benchmark Dataset on Dynamic Scenes

    Huanjing Yue, Cong Cao, Lei Liao +2

    eess.IVcs.CVcs.LGarXiv:2003.14013v12020
  33. Learning Posterior Predictive Distributions for Node Classification from Synthetic Graph Priors

    Jeongwhan Choi, Jongwoo Kim, Woosung Kang +1

    cs.LGarXiv:2604.19028v12026
  34. On Accurate and Reliable Anomaly Detection for Gas Turbine Combustors: A Deep Learning Approach

    Weizhong Yan, Lijie Yu

    cs.LGstat.MLarXiv:1908.09238v12019
  35. Dobi-SVD: Differentiable SVD for LLM Compression and Some New Perspectives

    Qinsi Wang, Jinghan Ke, Masayoshi Tomizuka +3

    cs.LGarXiv:2502.02723v12025
  36. Large Language Models and the Reverse Turing Test

    Terrence Sejnowski

    cs.CLcs.AIcs.LGarXiv:2207.14382v92022
  37. Memory-Efficient Fine-Tuning of Compressed Large Language Models via sub-4-bit Integer Quantization

    Jeonghoon Kim, Jung Hyun Lee, Sungdong Kim +4

    cs.LGcs.AIarXiv:2305.14152v22023
  38. Deep learning for in vitro prediction of pharmaceutical formulations

    Yilong Yang, Zhuyifan Ye, Yan Su +3

    cs.LGstat.MLarXiv:1809.02069v12018
  39. Do-PFN: In-Context Learning for Causal Effect Estimation

    Jake Robertson, Arik Reuter, Siyuan Guo +3

    cs.LGarXiv:2506.06039v32025
  40. Assuring the Machine Learning Lifecycle: Desiderata, Methods, and Challenges

    Rob Ashmore, Radu Calinescu, Colin Paterson

    cs.LGcs.SEstat.MLarXiv:1905.04223v12019
  41. Fully-Connected Spatial-Temporal Graph for Multivariate Time-Series Data

    Yucheng Wang, Yuecong Xu, Jianfei Yang +4

    cs.LGarXiv:2309.05305v32023
  42. Fine Perceptive GANs for Brain MR Image Super-Resolution in Wavelet Domain

    Senrong You, Yong Liu, Baiying Lei +1

    eess.IVcs.CVcs.LGarXiv:2011.04145v12020
  43. A Multi-Objective Deep Reinforcement Learning Framework

    Thanh Thi Nguyen, Ngoc Duy Nguyen, Peter Vamplew +3

    cs.LGcs.AIstat.MLarXiv:1803.02965v32018
  44. UNR-Explainer: Counterfactual Explanations for Unsupervised Node Representation Learning Models

    Hyunju Kang, Geonhee Han, Hogun Park

    cs.LGcs.AIarXiv:2605.17285v12026
  45. Multiscale modeling of inelastic materials with Thermodynamics-based Artificial Neural Networks (TANN)

    Filippo Masi, Ioannis Stefanou

    cond-mat.mtrl-scics.CEcs.LGarXiv:2108.13137v32021
  46. A spelling correction model for end-to-end speech recognition

    Jinxi Guo, Tara N. Sainath, Ron J. Weiss

    eess.AScs.AIcs.CLarXiv:1902.07178v12019
  47. Semi-Supervised and Task-Driven Data Augmentation

    Krishna Chaitanya, Neerav Karani, Christian Baumgartner +3

    cs.CVcs.LGstat.MLarXiv:1902.05396v22019
  48. DeepScientist: Advancing Frontier-Pushing Scientific Findings Progressively

    Yixuan Weng, Minjun Zhu, Qiujie Xie +4

    cs.CLcs.LGarXiv:2509.26603v12025
  49. Theory-guided hard constraint projection (HCP): a knowledge-based data-driven scientific machine learning method

    Yuntian Chen, Dou Huang, Dongxiao Zhang +4

    cs.LGcs.AIarXiv:2012.06148v22020
  50. SceneGen: Learning to Generate Realistic Traffic Scenes

    Shuhan Tan, Kelvin Wong, Shenlong Wang +3

    cs.CVcs.AIcs.LGarXiv:2101.06541v12021
  51. What is a meaningful representation of protein sequences?

    Nicki Skafte Detlefsen, Søren Hauberg, Wouter Boomsma

    q-bio.BMcs.LGq-bio.QMarXiv:2012.02679v42020
  52. Towards a Theoretical Framework of Out-of-Distribution Generalization

    Haotian Ye, Chuanlong Xie, Tianle Cai +3

    cs.LGarXiv:2106.04496v32021
  53. PEORL: Integrating Symbolic Planning and Hierarchical Reinforcement Learning for Robust Decision-Making

    Fangkai Yang, Daoming Lyu, Bo Liu +1

    cs.LGcs.AIstat.MLarXiv:1804.07779v32018
  54. T1: Terminal Agent Reinforcement Learning for Long-Horizon Tasks

    Junyao Yang, Yucheng Shi, Zhongzhi Li +4

    cs.LGcs.AIarXiv:2609.11042v12026
  55. Machine Learning for Intrusion Detection in Industrial Control Systems: Applications, Challenges, and Recommendations

    Muhammad Azmi Umer, Khurum Nazir Junejo, Muhammad Taha Jilani +1

    cs.CRcs.LGarXiv:2202.11917v12022
  56. Autonomous Discovery of Unknown Reaction Pathways from Data by Chemical Reaction Neural Network

    Weiqi Ji, Sili Deng

    q-bio.MNcs.LGphysics.chem-pharXiv:2002.09062v22020
  57. Deep Learning for Free-Hand Sketch: A Survey

    Peng Xu, Timothy M. Hospedales, Qiyue Yin +3

    cs.CVcs.GRcs.LGarXiv:2001.02600v32020
  58. Training Deep Convolutional Neural Networks with Resistive Cross-Point Devices

    Tayfun Gokmen, O. Murat Onen, Wilfried Haensch

    cs.LGcs.NEstat.MLarXiv:1705.08014v12017
  59. Fairness risk measures

    Robert C. Williamson, Aditya Krishna Menon

    cs.LGstat.MLarXiv:1901.08665v12019
  60. Limitations of Lazy Training of Two-layers Neural Networks

    Behrooz Ghorbani, Song Mei, Theodor Misiakiewicz +1

    stat.MLcs.LGmath.STarXiv:1906.08899v12019