Machine Learning

Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

4,261 to 4,320 of 20,454

  1. Small Language Models: Survey, Measurements, and Insights

    Zhenyan Lu, Xiang Li, Dongqi Cai +5

    cs.CLcs.AIcs.LGarXiv:2409.15790v32024
  2. Nested Slice Sampling: Vectorized Nested Sampling for GPU-Accelerated Inference

    David Yallup, Namu Kroupa, Will Handley

    stat.COcs.LGstat.MLarXiv:2601.23252v22026
  3. Do Latent-CoT Models Think Step-by-Step? A Mechanistic Study on Sequential Reasoning Tasks

    Jia Liang, Liangming Pan

    cs.AIcs.LGarXiv:2602.00449v12026
  4. Inductive Biases and Variable Creation in Self-Attention Mechanisms

    Benjamin L. Edelman, Surbhi Goel, Sham Kakade +1

    cs.LGstat.MLarXiv:2110.10090v22021
  5. Penalizing Gradient Norm for Efficiently Improving Generalization in Deep Learning

    Yang Zhao, Hao Zhang, Xiuyuan Hu

    cs.LGcs.AIarXiv:2202.03599v32022
  6. Between Pure and Approximate Differential Privacy

    Thomas Steinke, Jonathan Ullman

    cs.DScs.CRcs.LGarXiv:1501.06095v12015
  7. NeRS: Neural Reflectance Surfaces for Sparse-view 3D Reconstruction in the Wild

    Jason Y. Zhang, Gengshan Yang, Shubham Tulsiani +1

    cs.CVcs.LGarXiv:2110.07604v32021
  8. Training verified learners with learned verifiers

    Krishnamurthy Dvijotham, Sven Gowal, Robert Stanforth +4

    cs.LGstat.MLarXiv:1805.10265v22018
  9. Online 3D Bin Packing with Constrained Deep Reinforcement Learning

    Hang Zhao, Qijin She, Chenyang Zhu +2

    cs.LGstat.MLarXiv:2006.14978v52020
  10. Towards minimax policies for online linear optimization with bandit feedback

    Sébastien Bubeck, Nicolò Cesa-Bianchi, Sham M. Kakade

    cs.LGstat.MLarXiv:1202.3079v12012
  11. From LLMs to LRMs: Rethinking Pruning for Reasoning-Centric Models

    Longwei Ding, Anhao Zhao, Fanghua Ye +2

    cs.LGarXiv:2601.18091v12026
  12. BalDRO: A Distributionally Robust Optimization based Framework for Large Language Model Unlearning

    Pengyang Shao, Naixin Zhai, Lei Chen +4

    cs.LGarXiv:2601.09172v32026
  13. Temporal Multimodal Fusion for Video Emotion Classification in the Wild

    Valentin Vielzeuf, Stéphane Pateux, Frédéric Jurie

    cs.CVcs.LGcs.MMarXiv:1709.07200v12017
  14. A Signal Propagation Perspective for Pruning Neural Networks at Initialization

    Namhoon Lee, Thalaiyasingam Ajanthan, Stephen Gould +1

    cs.LGcs.CVstat.MLarXiv:1906.06307v22019
  15. Reasoning in Trees: Improving Retrieval-Augmented Generation for Multi-Hop Question Answering

    Yuling Shi, Maolin Sun, Zijun Liu +4

    cs.CLcs.LGarXiv:2601.11255v12026
  16. A Framework for Evaluating Gradient Leakage Attacks in Federated Learning

    Wenqi Wei, Ling Liu, Margaret Loper +4

    cs.LGcs.CRstat.MLarXiv:2004.10397v22020
  17. Image Generators with Conditionally-Independent Pixel Synthesis

    Ivan Anokhin, Kirill Demochkin, Taras Khakhulin +3

    cs.CVcs.AIcs.LGarXiv:2011.13775v12020
  18. Hadamard Response: Estimating Distributions Privately, Efficiently, and with Little Communication

    Jayadev Acharya, Ziteng Sun, Huanyu Zhang

    cs.LGcs.DScs.ITarXiv:1802.04705v22018
  19. Gaussian Process Prior Variational Autoencoders

    Francesco Paolo Casale, Adrian V Dalca, Luca Saglietti +2

    cs.LGstat.MLarXiv:1810.11738v22018
  20. Signed Graph Attention Networks

    Junjie Huang, Huawei Shen, Liang Hou +1

    cs.SIcs.LGphysics.soc-pharXiv:1906.10958v32019
  21. Quantization-Aware Collaborative Inference for Large Embodied AI Models

    Zhonghao Lyu, Ming Xiao, Mikael Skoglund +2

    cs.LGeess.SParXiv:2602.13052v12026
  22. Cast-R1: Learning Tool-Augmented Sequential Decision Policies for Time Series Forecasting

    Xiaoyu Tao, Mingyue Cheng, Chuang Jiang +3

    cs.LGarXiv:2602.13802v12026
  23. Hierarchical Decomposition of Prompt-Based Continual Learning: Rethinking Obscured Sub-optimality

    Liyuan Wang, Jingyi Xie, Xingxing Zhang +3

    cs.LGarXiv:2310.07234v12023
  24. Semi-Supervised Learning of Visual Features by Non-Parametrically Predicting View Assignments with Support Samples

    Mahmoud Assran, Mathilde Caron, Ishan Misra +4

    cs.CVcs.AIcs.LGarXiv:2104.13963v32021
  25. Just-In-Time Reinforcement Learning: Continual Learning in LLM Agents Without Gradient Updates

    Yibo Li, Zijie Lin, Ailin Deng +5

    cs.LGcs.AIarXiv:2601.18510v32026
  26. MNL-Bandit: A Dynamic Learning Approach to Assortment Selection

    Shipra Agrawal, Vashist Avadhanula, Vineet Goyal +1

    cs.LGarXiv:1706.03880v22017
  27. HiPER: Hierarchical Reinforcement Learning with Explicit Credit Assignment for Large Language Model Agents

    Jiangweizhi Peng, Yuanxin Liu, Ruida Zhou +4

    cs.LGcs.AIarXiv:2602.16165v22026
  28. Low-Dimensional and Transversely Curved Optimization Dynamics in Grokking

    Yongzhong Xu

    cs.LGcs.AIarXiv:2602.16746v32026
  29. Co-RedTeam: Orchestrated Security Discovery and Exploitation with LLM Agents

    Pengfei He, Ash Fox, Lesly Miculicich +7

    cs.LGcs.CRarXiv:2602.02164v22026
  30. Weakly-Supervised Video Moment Retrieval via Semantic Completion Network

    Zhijie Lin, Zhou Zhao, Zhu Zhang +2

    cs.CVcs.LGcs.MMarXiv:1911.08199v32019
  31. From Brute Force to Semantic Insight: Performance-Guided Data Transformation Design with LLMs

    Usha Shrestha, Dmitry Ignatov, Radu Timofte

    cs.CVcs.LGarXiv:2601.03808v22026
  32. Privately Learning High-Dimensional Distributions

    Gautam Kamath, Jerry Li, Vikrant Singhal +1

    cs.DScs.CRcs.LGarXiv:1805.00216v32018
  33. Tight Analyses for Non-Smooth Stochastic Gradient Descent

    Nicholas J. A. Harvey, Christopher Liaw, Yaniv Plan +1

    cs.LGmath.OCstat.MLarXiv:1812.05217v12018
  34. Neural Arabic Question Answering

    Hussein Mozannar, Karl El Hajal, Elie Maamary +1

    cs.CLcs.LGarXiv:1906.05394v12019
  35. Graph-Structured Deep Learning Framework for Multi-task Contention Identification with High-dimensional Metrics

    Xiao Yang, Yinan Ni, Yuqi Tang +3

    cs.DCcs.LGarXiv:2601.20389v12026
  36. Improving aircraft performance using machine learning: a review

    Soledad Le Clainche, Esteban Ferrer, Sam Gibson +3

    cs.LGphysics.data-anphysics.flu-dynarXiv:2210.11481v12022
  37. Evaluating explainable artificial intelligence methods for multi-label deep learning classification tasks in remote sensing

    Ioannis Kakogeorgiou, Konstantinos Karantzalos

    cs.LGcs.CVarXiv:2104.01375v22021
  38. Tackling Data Heterogeneity in Federated Learning with Class Prototypes

    Yutong Dai, Zeyuan Chen, Junnan Li +3

    cs.LGcs.AIarXiv:2212.02758v22022
  39. Does a Technique for Building Multimodal Representation Matter? -- Comparative Analysis

    Maciej Pawłowski, Anna Wróblewska, Sylwia Sysko-Romańczuk

    cs.LGarXiv:2206.06367v12022
  40. Understanding Neural Networks via Feature Visualization: A survey

    Anh Nguyen, Jason Yosinski, Jeff Clune

    cs.LGcs.AIcs.CVarXiv:1904.08939v12019
  41. HFedMoE: Resource-aware Heterogeneous Federated Learning with Mixture-of-Experts

    Zihan Fang, Zheng Lin, Senkang Hu +5

    cs.LGcs.AIcs.NIarXiv:2601.00583v12026
  42. Escaping Saddles with Stochastic Gradients

    Hadi Daneshmand, Jonas Kohler, Aurelien Lucchi +1

    cs.LGmath.OCstat.MLarXiv:1803.05999v22018
  43. Online Change-point Detection for Cooperative Multi-Agent Reinforcement Learning

    Fatemeh Saberi Khomami, Julita Vassileva

    cs.MAcs.LGarXiv:2609.05298v12026
  44. PAC-Bayesian Reconstruction Guarantees for Time Series Variational Autoencoders

    Chloé Hashimoto-Cullen, Ghislain Agoua, Benjamin Guedj +1

    stat.MLcs.LGarXiv:2609.05212v12026
  45. STEM: Scaling Transformers with Embedding Modules

    Ranajoy Sadhukhan, Sheng Cao, Harry Dong +5

    cs.LGarXiv:2601.10639v12026
  46. Deep learning for temporal data representation in electronic health records: A systematic review of challenges and methodologies

    Feng Xie, Han Yuan, Yilin Ning +5

    cs.LGarXiv:2107.09951v12021
  47. On the origin of neural scaling laws: from random graphs to natural language

    Maissam Barkeshli, Alberto Alfarano, Andrey Gromov

    cs.LGcond-mat.dis-nncs.AIarXiv:2601.10684v12026
  48. Camera Measurement of Physiological Vital Signs

    Daniel McDuff

    cs.CVcs.LGeess.IVarXiv:2111.11547v12021
  49. The Ladder: A Reliable Leaderboard for Machine Learning Competitions

    Avrim Blum, Moritz Hardt

    cs.LGarXiv:1502.04585v12015
  50. Adaptive Gated Deepfake Detection for Low-Resolution and Resource-Constrained Environments

    Vaishnavi Sen, Cody Laurie, Rashida Hasan

    cs.CVcs.LGarXiv:2609.05320v12026
  51. DeepSteal: Advanced Model Extractions Leveraging Efficient Weight Stealing in Memories

    Adnan Siraj Rakin, Md Hafizul Islam Chowdhuryy, Fan Yao +1

    cs.CRcs.AIcs.CVarXiv:2111.04625v12021
  52. Overcoming Oscillations in Quantization-Aware Training

    Markus Nagel, Marios Fournarakis, Yelysei Bondarenko +1

    cs.LGarXiv:2203.11086v22022
  53. Shallow neural network approximation in mixed Sobolev spaces

    Yuwen Li, Guozhi Zhang

    math.NAcs.LGarXiv:2609.05263v12026
  54. Confounding-Robust Policy Improvement

    Nathan Kallus, Angela Zhou

    cs.LGstat.MLarXiv:1805.08593v32018
  55. Proton Irradiation Characterization of an Open-Source ML Accelerator on a Zynq UltraScale+ MPSoC

    Saad Memon, Rafal Graczyk, Jan Swakoń +3

    cs.ARcs.ETcs.LGarXiv:2609.05249v12026
  56. Pneumonia Detection on chest X-ray images Using Ensemble of Deep Convolutional Neural Networks

    Alhassan Mabrouk, Rebeca P. Díaz Redondo, Abdelghani Dahou +2

    eess.IVcs.CVcs.LGarXiv:2312.07965v12023
  57. Coupled Control and Wireless World Models for Resilient Remote Robotic Control

    H. P. Madushanka, Sumudu Samarakoon, Mehdi Bennis

    cs.ROcs.LGarXiv:2609.04851v12026
  58. SMILE: Self-Explainable Multimodal Information Bottleneck for Medical Diagnosis

    Yuqing Yang, Alexander Schmatz, Zhaozhao Ma +3

    cs.CVcs.LGarXiv:2609.05174v12026
  59. Minimax Lower Bound for Estimating Diffusion-based Local Intrinsic Dimension

    Jaehee Seo, Wontae Jeong, Jisu Kim

    stat.MLcs.LGmath.STarXiv:2609.04822v12026
  60. LookThere! Sparse Vision by Reinforced Selection

    Sreehari Rammohan, Yousef Yassin, Anthony Fuller +3

    cs.CVcs.LGarXiv:2609.04698v12026