Machine Learning
Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
4,021 to 4,080 of 20,217
A Signal Propagation Perspective for Pruning Neural Networks at Initialization
Namhoon Lee, Thalaiyasingam Ajanthan, Stephen Gould +1
cs.LGcs.CVstat.MLarXiv:1906.06307v22019Reasoning in Trees: Improving Retrieval-Augmented Generation for Multi-Hop Question Answering
Yuling Shi, Maolin Sun, Zijun Liu +4
cs.CLcs.LGarXiv:2601.11255v12026A Framework for Evaluating Gradient Leakage Attacks in Federated Learning
Wenqi Wei, Ling Liu, Margaret Loper +4
cs.LGcs.CRstat.MLarXiv:2004.10397v22020Image Generators with Conditionally-Independent Pixel Synthesis
Ivan Anokhin, Kirill Demochkin, Taras Khakhulin +3
cs.CVcs.AIcs.LGarXiv:2011.13775v12020Hadamard Response: Estimating Distributions Privately, Efficiently, and with Little Communication
Jayadev Acharya, Ziteng Sun, Huanyu Zhang
cs.LGcs.DScs.ITarXiv:1802.04705v22018Gaussian Process Prior Variational Autoencoders
Francesco Paolo Casale, Adrian V Dalca, Luca Saglietti +2
cs.LGstat.MLarXiv:1810.11738v22018Signed Graph Attention Networks
Junjie Huang, Huawei Shen, Liang Hou +1
cs.SIcs.LGphysics.soc-pharXiv:1906.10958v32019Quantization-Aware Collaborative Inference for Large Embodied AI Models
Zhonghao Lyu, Ming Xiao, Mikael Skoglund +2
cs.LGeess.SParXiv:2602.13052v12026Cast-R1: Learning Tool-Augmented Sequential Decision Policies for Time Series Forecasting
Xiaoyu Tao, Mingyue Cheng, Chuang Jiang +3
cs.LGarXiv:2602.13802v12026Hierarchical Decomposition of Prompt-Based Continual Learning: Rethinking Obscured Sub-optimality
Liyuan Wang, Jingyi Xie, Xingxing Zhang +3
cs.LGarXiv:2310.07234v12023Semi-Supervised Learning of Visual Features by Non-Parametrically Predicting View Assignments with Support Samples
Mahmoud Assran, Mathilde Caron, Ishan Misra +4
cs.CVcs.AIcs.LGarXiv:2104.13963v32021Just-In-Time Reinforcement Learning: Continual Learning in LLM Agents Without Gradient Updates
Yibo Li, Zijie Lin, Ailin Deng +5
cs.LGcs.AIarXiv:2601.18510v32026MNL-Bandit: A Dynamic Learning Approach to Assortment Selection
Shipra Agrawal, Vashist Avadhanula, Vineet Goyal +1
cs.LGarXiv:1706.03880v22017HiPER: Hierarchical Reinforcement Learning with Explicit Credit Assignment for Large Language Model Agents
Jiangweizhi Peng, Yuanxin Liu, Ruida Zhou +4
cs.LGcs.AIarXiv:2602.16165v22026Low-Dimensional and Transversely Curved Optimization Dynamics in Grokking
Yongzhong Xu
cs.LGcs.AIarXiv:2602.16746v32026Co-RedTeam: Orchestrated Security Discovery and Exploitation with LLM Agents
Pengfei He, Ash Fox, Lesly Miculicich +7
cs.LGcs.CRarXiv:2602.02164v22026Weakly-Supervised Video Moment Retrieval via Semantic Completion Network
Zhijie Lin, Zhou Zhao, Zhu Zhang +2
cs.CVcs.LGcs.MMarXiv:1911.08199v32019From Brute Force to Semantic Insight: Performance-Guided Data Transformation Design with LLMs
Usha Shrestha, Dmitry Ignatov, Radu Timofte
cs.CVcs.LGarXiv:2601.03808v22026Privately Learning High-Dimensional Distributions
Gautam Kamath, Jerry Li, Vikrant Singhal +1
cs.DScs.CRcs.LGarXiv:1805.00216v32018Tight Analyses for Non-Smooth Stochastic Gradient Descent
Nicholas J. A. Harvey, Christopher Liaw, Yaniv Plan +1
cs.LGmath.OCstat.MLarXiv:1812.05217v12018Neural Arabic Question Answering
Hussein Mozannar, Karl El Hajal, Elie Maamary +1
cs.CLcs.LGarXiv:1906.05394v12019Graph-Structured Deep Learning Framework for Multi-task Contention Identification with High-dimensional Metrics
Xiao Yang, Yinan Ni, Yuqi Tang +3
cs.DCcs.LGarXiv:2601.20389v12026Improving aircraft performance using machine learning: a review
Soledad Le Clainche, Esteban Ferrer, Sam Gibson +3
cs.LGphysics.data-anphysics.flu-dynarXiv:2210.11481v12022Evaluating explainable artificial intelligence methods for multi-label deep learning classification tasks in remote sensing
Ioannis Kakogeorgiou, Konstantinos Karantzalos
cs.LGcs.CVarXiv:2104.01375v22021Tackling Data Heterogeneity in Federated Learning with Class Prototypes
Yutong Dai, Zeyuan Chen, Junnan Li +3
cs.LGcs.AIarXiv:2212.02758v22022Does a Technique for Building Multimodal Representation Matter? -- Comparative Analysis
Maciej Pawłowski, Anna Wróblewska, Sylwia Sysko-Romańczuk
cs.LGarXiv:2206.06367v12022Understanding Neural Networks via Feature Visualization: A survey
Anh Nguyen, Jason Yosinski, Jeff Clune
cs.LGcs.AIcs.CVarXiv:1904.08939v12019HFedMoE: Resource-aware Heterogeneous Federated Learning with Mixture-of-Experts
Zihan Fang, Zheng Lin, Senkang Hu +5
cs.LGcs.AIcs.NIarXiv:2601.00583v12026Escaping Saddles with Stochastic Gradients
Hadi Daneshmand, Jonas Kohler, Aurelien Lucchi +1
cs.LGmath.OCstat.MLarXiv:1803.05999v22018Online Change-point Detection for Cooperative Multi-Agent Reinforcement Learning
Fatemeh Saberi Khomami, Julita Vassileva
cs.MAcs.LGarXiv:2609.05298v12026PAC-Bayesian Reconstruction Guarantees for Time Series Variational Autoencoders
Chloé Hashimoto-Cullen, Ghislain Agoua, Benjamin Guedj +1
stat.MLcs.LGarXiv:2609.05212v12026STEM: Scaling Transformers with Embedding Modules
Ranajoy Sadhukhan, Sheng Cao, Harry Dong +5
cs.LGarXiv:2601.10639v12026Deep learning for temporal data representation in electronic health records: A systematic review of challenges and methodologies
Feng Xie, Han Yuan, Yilin Ning +5
cs.LGarXiv:2107.09951v12021On the origin of neural scaling laws: from random graphs to natural language
Maissam Barkeshli, Alberto Alfarano, Andrey Gromov
cs.LGcond-mat.dis-nncs.AIarXiv:2601.10684v12026Camera Measurement of Physiological Vital Signs
Daniel McDuff
cs.CVcs.LGeess.IVarXiv:2111.11547v12021The Ladder: A Reliable Leaderboard for Machine Learning Competitions
Avrim Blum, Moritz Hardt
cs.LGarXiv:1502.04585v12015Adaptive Gated Deepfake Detection for Low-Resolution and Resource-Constrained Environments
Vaishnavi Sen, Cody Laurie, Rashida Hasan
cs.CVcs.LGarXiv:2609.05320v12026DeepSteal: Advanced Model Extractions Leveraging Efficient Weight Stealing in Memories
Adnan Siraj Rakin, Md Hafizul Islam Chowdhuryy, Fan Yao +1
cs.CRcs.AIcs.CVarXiv:2111.04625v12021Overcoming Oscillations in Quantization-Aware Training
Markus Nagel, Marios Fournarakis, Yelysei Bondarenko +1
cs.LGarXiv:2203.11086v22022Shallow neural network approximation in mixed Sobolev spaces
Yuwen Li, Guozhi Zhang
math.NAcs.LGarXiv:2609.05263v12026Confounding-Robust Policy Improvement
Nathan Kallus, Angela Zhou
cs.LGstat.MLarXiv:1805.08593v32018Proton Irradiation Characterization of an Open-Source ML Accelerator on a Zynq UltraScale+ MPSoC
Saad Memon, Rafal Graczyk, Jan Swakoń +3
cs.ARcs.ETcs.LGarXiv:2609.05249v12026Pneumonia Detection on chest X-ray images Using Ensemble of Deep Convolutional Neural Networks
Alhassan Mabrouk, Rebeca P. Díaz Redondo, Abdelghani Dahou +2
eess.IVcs.CVcs.LGarXiv:2312.07965v12023Coupled Control and Wireless World Models for Resilient Remote Robotic Control
H. P. Madushanka, Sumudu Samarakoon, Mehdi Bennis
cs.ROcs.LGarXiv:2609.04851v12026SMILE: Self-Explainable Multimodal Information Bottleneck for Medical Diagnosis
Yuqing Yang, Alexander Schmatz, Zhaozhao Ma +3
cs.CVcs.LGarXiv:2609.05174v12026Minimax Lower Bound for Estimating Diffusion-based Local Intrinsic Dimension
Jaehee Seo, Wontae Jeong, Jisu Kim
stat.MLcs.LGmath.STarXiv:2609.04822v12026LookThere! Sparse Vision by Reinforced Selection
Sreehari Rammohan, Yousef Yassin, Anthony Fuller +3
cs.CVcs.LGarXiv:2609.04698v12026Same Request, Different Answer: Quantization Amplifies Cache-Induced Divergence in LLM Serving
Aditi Patodiya
cs.SEcs.DCcs.LGarXiv:2609.04748v12026Low Level Control of a Quadrotor with Deep Model-Based Reinforcement Learning
Nathan O. Lambert, Daniel S. Drew, Joseph Yaconelli +3
cs.ROcs.LGarXiv:1901.03737v22019AQAScore: Evaluating Semantic Alignment in Text-to-Audio Generation via Audio Question Answering
Chun-Yi Kuan, Kai-Wei Chang, Hung-yi Lee
eess.AScs.AIcs.CLarXiv:2601.14728v12026FluxDisco: Symbolic Regression for Stoichiometric Dynamical Systems via Monte Carlo Graph Search
Cassandra Durr, Alvaro Köhn-Luque, Chris Jewell +1
stat.MLcs.LGphysics.data-anarXiv:2609.05207v12026Designing Interpretable ML System to Enhance Trust in Healthcare: A Systematic Review to Proposed Responsible Clinician-AI-Collaboration Framework
Elham Nasarian, Roohallah Alizadehsani, U. Rajendra Acharya +1
cs.AIcs.HCcs.LGarXiv:2311.11055v22023An Alternative Probabilistic Interpretation of the Huber Loss
Gregory P. Meyer
stat.MLcs.CVcs.LGarXiv:1911.02088v32019Impact of Data Loss in Postprocessing on Training and Inference of Quantum Neural Networks
Soraya V. Panambalom, Edoardo Altamura, Nick Chancellor +1
quant-phcs.ETcs.LGarXiv:2609.05060v12026Terahertz-Band Joint Ultra-Massive MIMO Radar-Communications: Model-Based and Model-Free Hybrid Beamforming
Ahmet M. Elbir, Kumar Vijay Mishra, Symeon Chatzinotas
eess.SPcs.ITcs.LGarXiv:2103.00328v22021An Analysis of Self-supervised Pre-training with Dependent Samples
Maximilian Fleissner, Debarghya Ghoshdastidar, Samory Kpotufe
stat.MLcs.LGarXiv:2609.05031v12026On the Generalization Capacities of MLLMs for Spatial Intelligence
Gongjie Zhang, Wenhao Li, Quanhao Qian +4
cs.CVcs.LGarXiv:2603.06704v12026Stress Field Prediction in Cantilevered Structures Using Convolutional Neural Networks
Zhenguo Nie, Haoliang Jiang, Levent Burak Kara
cs.LGstat.MLarXiv:1808.08914v32018IGenBench: Benchmarking the Reliability of Text-to-Infographic Generation
Yinghao Tang, Xueding Liu, Boyuan Zhang +13
cs.LGcs.CVarXiv:2601.04498v22026LLaTTE: Scaling Laws for Multi-Stage Sequence Modeling in Large-Scale Ads Recommendation
Lee Xiong, Zhirong Chen, Rahul Mayuranath +17
cs.IRcs.AIcs.LGarXiv:2601.20083v12026