Machine Learning

Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

4,021 to 4,080 of 20,217

  1. A Signal Propagation Perspective for Pruning Neural Networks at Initialization

    Namhoon Lee, Thalaiyasingam Ajanthan, Stephen Gould +1

    cs.LGcs.CVstat.MLarXiv:1906.06307v22019
  2. Reasoning in Trees: Improving Retrieval-Augmented Generation for Multi-Hop Question Answering

    Yuling Shi, Maolin Sun, Zijun Liu +4

    cs.CLcs.LGarXiv:2601.11255v12026
  3. A Framework for Evaluating Gradient Leakage Attacks in Federated Learning

    Wenqi Wei, Ling Liu, Margaret Loper +4

    cs.LGcs.CRstat.MLarXiv:2004.10397v22020
  4. Image Generators with Conditionally-Independent Pixel Synthesis

    Ivan Anokhin, Kirill Demochkin, Taras Khakhulin +3

    cs.CVcs.AIcs.LGarXiv:2011.13775v12020
  5. Hadamard Response: Estimating Distributions Privately, Efficiently, and with Little Communication

    Jayadev Acharya, Ziteng Sun, Huanyu Zhang

    cs.LGcs.DScs.ITarXiv:1802.04705v22018
  6. Gaussian Process Prior Variational Autoencoders

    Francesco Paolo Casale, Adrian V Dalca, Luca Saglietti +2

    cs.LGstat.MLarXiv:1810.11738v22018
  7. Signed Graph Attention Networks

    Junjie Huang, Huawei Shen, Liang Hou +1

    cs.SIcs.LGphysics.soc-pharXiv:1906.10958v32019
  8. Quantization-Aware Collaborative Inference for Large Embodied AI Models

    Zhonghao Lyu, Ming Xiao, Mikael Skoglund +2

    cs.LGeess.SParXiv:2602.13052v12026
  9. Cast-R1: Learning Tool-Augmented Sequential Decision Policies for Time Series Forecasting

    Xiaoyu Tao, Mingyue Cheng, Chuang Jiang +3

    cs.LGarXiv:2602.13802v12026
  10. Hierarchical Decomposition of Prompt-Based Continual Learning: Rethinking Obscured Sub-optimality

    Liyuan Wang, Jingyi Xie, Xingxing Zhang +3

    cs.LGarXiv:2310.07234v12023
  11. Semi-Supervised Learning of Visual Features by Non-Parametrically Predicting View Assignments with Support Samples

    Mahmoud Assran, Mathilde Caron, Ishan Misra +4

    cs.CVcs.AIcs.LGarXiv:2104.13963v32021
  12. Just-In-Time Reinforcement Learning: Continual Learning in LLM Agents Without Gradient Updates

    Yibo Li, Zijie Lin, Ailin Deng +5

    cs.LGcs.AIarXiv:2601.18510v32026
  13. MNL-Bandit: A Dynamic Learning Approach to Assortment Selection

    Shipra Agrawal, Vashist Avadhanula, Vineet Goyal +1

    cs.LGarXiv:1706.03880v22017
  14. HiPER: Hierarchical Reinforcement Learning with Explicit Credit Assignment for Large Language Model Agents

    Jiangweizhi Peng, Yuanxin Liu, Ruida Zhou +4

    cs.LGcs.AIarXiv:2602.16165v22026
  15. Low-Dimensional and Transversely Curved Optimization Dynamics in Grokking

    Yongzhong Xu

    cs.LGcs.AIarXiv:2602.16746v32026
  16. Co-RedTeam: Orchestrated Security Discovery and Exploitation with LLM Agents

    Pengfei He, Ash Fox, Lesly Miculicich +7

    cs.LGcs.CRarXiv:2602.02164v22026
  17. Weakly-Supervised Video Moment Retrieval via Semantic Completion Network

    Zhijie Lin, Zhou Zhao, Zhu Zhang +2

    cs.CVcs.LGcs.MMarXiv:1911.08199v32019
  18. From Brute Force to Semantic Insight: Performance-Guided Data Transformation Design with LLMs

    Usha Shrestha, Dmitry Ignatov, Radu Timofte

    cs.CVcs.LGarXiv:2601.03808v22026
  19. Privately Learning High-Dimensional Distributions

    Gautam Kamath, Jerry Li, Vikrant Singhal +1

    cs.DScs.CRcs.LGarXiv:1805.00216v32018
  20. Tight Analyses for Non-Smooth Stochastic Gradient Descent

    Nicholas J. A. Harvey, Christopher Liaw, Yaniv Plan +1

    cs.LGmath.OCstat.MLarXiv:1812.05217v12018
  21. Neural Arabic Question Answering

    Hussein Mozannar, Karl El Hajal, Elie Maamary +1

    cs.CLcs.LGarXiv:1906.05394v12019
  22. Graph-Structured Deep Learning Framework for Multi-task Contention Identification with High-dimensional Metrics

    Xiao Yang, Yinan Ni, Yuqi Tang +3

    cs.DCcs.LGarXiv:2601.20389v12026
  23. Improving aircraft performance using machine learning: a review

    Soledad Le Clainche, Esteban Ferrer, Sam Gibson +3

    cs.LGphysics.data-anphysics.flu-dynarXiv:2210.11481v12022
  24. Evaluating explainable artificial intelligence methods for multi-label deep learning classification tasks in remote sensing

    Ioannis Kakogeorgiou, Konstantinos Karantzalos

    cs.LGcs.CVarXiv:2104.01375v22021
  25. Tackling Data Heterogeneity in Federated Learning with Class Prototypes

    Yutong Dai, Zeyuan Chen, Junnan Li +3

    cs.LGcs.AIarXiv:2212.02758v22022
  26. Does a Technique for Building Multimodal Representation Matter? -- Comparative Analysis

    Maciej Pawłowski, Anna Wróblewska, Sylwia Sysko-Romańczuk

    cs.LGarXiv:2206.06367v12022
  27. Understanding Neural Networks via Feature Visualization: A survey

    Anh Nguyen, Jason Yosinski, Jeff Clune

    cs.LGcs.AIcs.CVarXiv:1904.08939v12019
  28. HFedMoE: Resource-aware Heterogeneous Federated Learning with Mixture-of-Experts

    Zihan Fang, Zheng Lin, Senkang Hu +5

    cs.LGcs.AIcs.NIarXiv:2601.00583v12026
  29. Escaping Saddles with Stochastic Gradients

    Hadi Daneshmand, Jonas Kohler, Aurelien Lucchi +1

    cs.LGmath.OCstat.MLarXiv:1803.05999v22018
  30. Online Change-point Detection for Cooperative Multi-Agent Reinforcement Learning

    Fatemeh Saberi Khomami, Julita Vassileva

    cs.MAcs.LGarXiv:2609.05298v12026
  31. PAC-Bayesian Reconstruction Guarantees for Time Series Variational Autoencoders

    Chloé Hashimoto-Cullen, Ghislain Agoua, Benjamin Guedj +1

    stat.MLcs.LGarXiv:2609.05212v12026
  32. STEM: Scaling Transformers with Embedding Modules

    Ranajoy Sadhukhan, Sheng Cao, Harry Dong +5

    cs.LGarXiv:2601.10639v12026
  33. Deep learning for temporal data representation in electronic health records: A systematic review of challenges and methodologies

    Feng Xie, Han Yuan, Yilin Ning +5

    cs.LGarXiv:2107.09951v12021
  34. On the origin of neural scaling laws: from random graphs to natural language

    Maissam Barkeshli, Alberto Alfarano, Andrey Gromov

    cs.LGcond-mat.dis-nncs.AIarXiv:2601.10684v12026
  35. Camera Measurement of Physiological Vital Signs

    Daniel McDuff

    cs.CVcs.LGeess.IVarXiv:2111.11547v12021
  36. The Ladder: A Reliable Leaderboard for Machine Learning Competitions

    Avrim Blum, Moritz Hardt

    cs.LGarXiv:1502.04585v12015
  37. Adaptive Gated Deepfake Detection for Low-Resolution and Resource-Constrained Environments

    Vaishnavi Sen, Cody Laurie, Rashida Hasan

    cs.CVcs.LGarXiv:2609.05320v12026
  38. DeepSteal: Advanced Model Extractions Leveraging Efficient Weight Stealing in Memories

    Adnan Siraj Rakin, Md Hafizul Islam Chowdhuryy, Fan Yao +1

    cs.CRcs.AIcs.CVarXiv:2111.04625v12021
  39. Overcoming Oscillations in Quantization-Aware Training

    Markus Nagel, Marios Fournarakis, Yelysei Bondarenko +1

    cs.LGarXiv:2203.11086v22022
  40. Shallow neural network approximation in mixed Sobolev spaces

    Yuwen Li, Guozhi Zhang

    math.NAcs.LGarXiv:2609.05263v12026
  41. Confounding-Robust Policy Improvement

    Nathan Kallus, Angela Zhou

    cs.LGstat.MLarXiv:1805.08593v32018
  42. Proton Irradiation Characterization of an Open-Source ML Accelerator on a Zynq UltraScale+ MPSoC

    Saad Memon, Rafal Graczyk, Jan Swakoń +3

    cs.ARcs.ETcs.LGarXiv:2609.05249v12026
  43. Pneumonia Detection on chest X-ray images Using Ensemble of Deep Convolutional Neural Networks

    Alhassan Mabrouk, Rebeca P. Díaz Redondo, Abdelghani Dahou +2

    eess.IVcs.CVcs.LGarXiv:2312.07965v12023
  44. Coupled Control and Wireless World Models for Resilient Remote Robotic Control

    H. P. Madushanka, Sumudu Samarakoon, Mehdi Bennis

    cs.ROcs.LGarXiv:2609.04851v12026
  45. SMILE: Self-Explainable Multimodal Information Bottleneck for Medical Diagnosis

    Yuqing Yang, Alexander Schmatz, Zhaozhao Ma +3

    cs.CVcs.LGarXiv:2609.05174v12026
  46. Minimax Lower Bound for Estimating Diffusion-based Local Intrinsic Dimension

    Jaehee Seo, Wontae Jeong, Jisu Kim

    stat.MLcs.LGmath.STarXiv:2609.04822v12026
  47. LookThere! Sparse Vision by Reinforced Selection

    Sreehari Rammohan, Yousef Yassin, Anthony Fuller +3

    cs.CVcs.LGarXiv:2609.04698v12026
  48. Same Request, Different Answer: Quantization Amplifies Cache-Induced Divergence in LLM Serving

    Aditi Patodiya

    cs.SEcs.DCcs.LGarXiv:2609.04748v12026
  49. Low Level Control of a Quadrotor with Deep Model-Based Reinforcement Learning

    Nathan O. Lambert, Daniel S. Drew, Joseph Yaconelli +3

    cs.ROcs.LGarXiv:1901.03737v22019
  50. AQAScore: Evaluating Semantic Alignment in Text-to-Audio Generation via Audio Question Answering

    Chun-Yi Kuan, Kai-Wei Chang, Hung-yi Lee

    eess.AScs.AIcs.CLarXiv:2601.14728v12026
  51. FluxDisco: Symbolic Regression for Stoichiometric Dynamical Systems via Monte Carlo Graph Search

    Cassandra Durr, Alvaro Köhn-Luque, Chris Jewell +1

    stat.MLcs.LGphysics.data-anarXiv:2609.05207v12026
  52. Designing Interpretable ML System to Enhance Trust in Healthcare: A Systematic Review to Proposed Responsible Clinician-AI-Collaboration Framework

    Elham Nasarian, Roohallah Alizadehsani, U. Rajendra Acharya +1

    cs.AIcs.HCcs.LGarXiv:2311.11055v22023
  53. An Alternative Probabilistic Interpretation of the Huber Loss

    Gregory P. Meyer

    stat.MLcs.CVcs.LGarXiv:1911.02088v32019
  54. Impact of Data Loss in Postprocessing on Training and Inference of Quantum Neural Networks

    Soraya V. Panambalom, Edoardo Altamura, Nick Chancellor +1

    quant-phcs.ETcs.LGarXiv:2609.05060v12026
  55. Terahertz-Band Joint Ultra-Massive MIMO Radar-Communications: Model-Based and Model-Free Hybrid Beamforming

    Ahmet M. Elbir, Kumar Vijay Mishra, Symeon Chatzinotas

    eess.SPcs.ITcs.LGarXiv:2103.00328v22021
  56. An Analysis of Self-supervised Pre-training with Dependent Samples

    Maximilian Fleissner, Debarghya Ghoshdastidar, Samory Kpotufe

    stat.MLcs.LGarXiv:2609.05031v12026
  57. On the Generalization Capacities of MLLMs for Spatial Intelligence

    Gongjie Zhang, Wenhao Li, Quanhao Qian +4

    cs.CVcs.LGarXiv:2603.06704v12026
  58. Stress Field Prediction in Cantilevered Structures Using Convolutional Neural Networks

    Zhenguo Nie, Haoliang Jiang, Levent Burak Kara

    cs.LGstat.MLarXiv:1808.08914v32018
  59. IGenBench: Benchmarking the Reliability of Text-to-Infographic Generation

    Yinghao Tang, Xueding Liu, Boyuan Zhang +13

    cs.LGcs.CVarXiv:2601.04498v22026
  60. LLaTTE: Scaling Laws for Multi-Stage Sequence Modeling in Large-Scale Ads Recommendation

    Lee Xiong, Zhirong Chen, Rahul Mayuranath +17

    cs.IRcs.AIcs.LGarXiv:2601.20083v12026