Machine Learning
Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
5,161 to 5,220 of 20,175
Overfitting Mechanism and Avoidance in Deep Neural Networks
Shaeke Salman, Xiuwen Liu
cs.LGcs.NEstat.MLarXiv:1901.06566v12019Wireless Communications for Collaborative Federated Learning
Mingzhe Chen, H. Vincent Poor, Walid Saad +1
cs.ITcs.LGarXiv:2006.02499v22020A Unified Framework for Sparse Relaxed Regularized Regression: SR3
Peng Zheng, Travis Askham, Steven L. Brunton +2
stat.MLcs.LGmath.OCarXiv:1807.05411v42018Detection of Coronavirus (COVID-19) Associated Pneumonia based on Generative Adversarial Networks and a Fine-Tuned Deep Transfer Learning Model using Chest X-ray Dataset
Nour Eldeen M. Khalifa, Mohamed Hamed N. Taha, Aboul Ella Hassanien +1
eess.IVcs.CVcs.LGarXiv:2004.01184v12020Deep Reinforcement Learning for Solving the Heterogeneous Capacitated Vehicle Routing Problem
Jingwen Li, Yining Ma, Ruize Gao +4
cs.LGmath.OCarXiv:2110.02629v22021Parameter-Efficient Fine-Tuning for Foundation Models
Dan Zhang, Tao Feng, Lilong Xue +3
cs.CLcs.AIcs.LGarXiv:2501.13787v12025NVIDIA FLARE: Federated Learning from Simulation to Real-World
Holger R. Roth, Yan Cheng, Yuhong Wen +20
cs.LGcs.AIcs.CVarXiv:2210.13291v32022TIPS: Turn-Level Information-Potential Reward Shaping for Search-Augmented LLMs
Yutao Xie, Nathaniel Thomas, Nicklas Hansen +3
cs.CLcs.AIcs.LGarXiv:2603.22293v12026Incompressible Knowledge Probes: Estimating Black-Box LLM Parameter Counts via Factual Capacity
Bojie Li
cs.LGcs.AIarXiv:2604.24827v22026Machine learning approach for early detection of autism by combining questionnaire and home video screening
Halim Abbas, Ford Garberson, Eric Glover +1
cs.CYcs.LGarXiv:1703.06076v12017Optimizing Large Language Model Training Using FP4 Quantization
Ruizhe Wang, Yeyun Gong, Xiao Liu +5
cs.LGcs.CLarXiv:2501.17116v22025Astra: A Multi-Agent System for GPU Kernel Performance Optimization
Anjiang Wei, Tianran Sun, Yogesh Seenichamy +5
cs.DCcs.AIcs.CLarXiv:2509.07506v22025General In-Hand Object Rotation with Vision and Touch
Haozhi Qi, Brent Yi, Sudharshan Suresh +4
cs.ROcs.AIcs.CVarXiv:2309.09979v22023Position: Graph Learning Will Lose Relevance Due To Poor Benchmarks
Maya Bechler-Speicher, Ben Finkelshtein, Fabrizio Frasca +9
cs.LGcs.AIcs.NEarXiv:2502.14546v12025Variational Federated Multi-Task Learning
Luca Corinzia, Ami Beuret, Joachim M. Buhmann
cs.LGstat.MLarXiv:1906.06268v22019Towards Efficient Large Language Model Serving: A Survey on System-Aware KV Cache Optimization
Jiantong Jiang, Peiyu Yang, Rui Zhang +1
cs.LGcs.AIcs.CLarXiv:2607.08057v12026A Generative Deep Learning Approach to Stochastic Downscaling of Precipitation Forecasts
Lucy Harris, Andrew T. T. McRae, Matthew Chantry +2
physics.ao-phcs.AIcs.CVarXiv:2204.02028v22022A Survey of Research in Large Language Models for Electronic Design Automation
Jingyu Pan, Guanglei Zhou, Chen-Chia Chang +3
cs.LGarXiv:2501.09655v12025Graph Neural Networks in Modern AI-aided Drug Discovery
Odin Zhang, Haitao Lin, Xujun Zhang +9
q-bio.BMcs.LGarXiv:2506.06915v12025Oculi: A Conversational Agentic Platform for Automated Credit Risk Analysis
Vennise Ho, Kristian Diana, Sandy Mourad +3
cs.AIcs.LGarXiv:2608.28944v12026Adaptive and Safe Bayesian Optimization in High Dimensions via One-Dimensional Subspaces
Johannes Kirschner, Mojmír Mutný, Nicole Hiller +2
cs.LGstat.MLarXiv:1902.03229v22019Linear Mode Connectivity in Multitask and Continual Learning
Seyed Iman Mirzadeh, Mehrdad Farajtabar, Dilan Gorur +2
cs.LGcs.AIcs.CVarXiv:2010.04495v12020Robust Ensemble Clustering Using Probability Trajectories
Dong Huang, Jian-Huang Lai, Chang-Dong Wang
stat.MLcs.LGarXiv:1606.01160v12016V2TATC: A Joint Voice-Trajectory Embedding Framework and Dataset for Air Traffic Controller Situational Awareness
Louis Brusset, Mathurin Petit, Jordan Kam +1
cs.LGeess.ASarXiv:2608.28981v12026OCGQuant: Outlier-Companion Grouping for NVFP4 Quantization
Yishan Yao, Binjun Li, Hanling Yi +5
cs.CLcs.AIcs.LGarXiv:2609.00066v12026Poly-YOLO: higher speed, more precise detection and instance segmentation for YOLOv3
Petr Hurtik, Vojtech Molek, Jan Hula +3
cs.CVcs.LGeess.IVarXiv:2005.13243v22020Riemannian Continuous Normalizing Flows
Emile Mathieu, Maximilian Nickel
stat.MLcs.LGarXiv:2006.10605v22020A Federated Learning Approach to Anomaly Detection in Smart Buildings
Raed Abdel Sater, A. Ben Hamza
cs.LGarXiv:2010.10293v32020LLaVE: Large Language and Vision Embedding Models with Hardness-Weighted Contrastive Learning
Zhibin Lan, Liqiang Niu, Fandong Meng +2
cs.CVcs.AIcs.CLarXiv:2503.04812v22025PET-MAD, a lightweight universal interatomic potential for advanced materials modeling
Arslan Mazitov, Filippo Bigi, Matthias Kellner +6
cond-mat.mtrl-scics.LGphysics.chem-pharXiv:2503.14118v22025Test-time regression: a unifying framework for designing sequence models with associative memory
Ke Alexander Wang, Jiaxin Shi, Emily B. Fox
cs.LGcs.AIcs.NEarXiv:2501.12352v32025Self-Supervised Learning of State Estimation for Manipulating Deformable Linear Objects
Mengyuan Yan, Yilin Zhu, Ning Jin +1
cs.ROcs.CVcs.LGarXiv:1911.06283v32019Generative Teaching Networks: Accelerating Neural Architecture Search by Learning to Generate Synthetic Training Data
Felipe Petroski Such, Aditya Rawal, Joel Lehman +2
cs.LGstat.MLarXiv:1912.07768v12019Confidence Is All You Need: Few-Shot RL Fine-Tuning of Language Models
Pengyi Li, Matvey Skripkin, Alexander Zubrey +2
cs.CLcs.LGarXiv:2506.06395v32025Multi-Agent Reinforcement Learning via Double Averaging Primal-Dual Optimization
Hoi-To Wai, Zhuoran Yang, Zhaoran Wang +1
cs.LGmath.OCstat.MLarXiv:1806.00877v42018Reducing SO(3) Convolutions to SO(2) for Efficient Equivariant GNNs
Saro Passaro, C. Lawrence Zitnick
cs.LGphysics.chem-phphysics.comp-pharXiv:2302.03655v22023Hardness-Aware Deep Metric Learning
Wenzhao Zheng, Zhaodong Chen, Jiwen Lu +1
cs.CVcs.LGarXiv:1903.05503v22019Single Model Deep Learning on Imbalanced Small Datasets for Skin Lesion Classification
Peng Yao, Shuwei Shen, Mengjuan Xu +6
cs.CVcs.LGarXiv:2102.01284v22021Large Language Models to Enhance Bayesian Optimization
Tennison Liu, Nicolás Astorga, Nabeel Seedat +1
cs.LGcs.AIarXiv:2402.03921v22024UserBench: An Interactive Gym Environment for User-Centric Agents
Cheng Qian, Zuxin Liu, Akshara Prabhakar +9
cs.AIcs.CLcs.LGarXiv:2507.22034v12025Categorical Flow Maps
Daan Roos, Oscar Davis, Floor Eijkelboom +5
cs.LGarXiv:2602.12233v12026PruneShift: A Framework for Evaluating Decision Reliability in Structured Pruning
Hao Ye, Gaopeng Zhang
cs.LGcs.NEarXiv:2608.29765v12026Natural Compression for Distributed Deep Learning
Samuel Horvath, Chen-Yu Ho, Ludovit Horvath +3
cs.LGmath.OCstat.MLarXiv:1905.10988v32019Unsupervised Learning by Competing Hidden Units
Dmitry Krotov, John Hopfield
cs.LGcs.CVcs.NEarXiv:1806.10181v22018Purified OPSD: On-Policy Self-Distillation Without Losing How to Think
Zhanming Shen, Jintao Tong, Shaotian Yan +9
cs.AIcs.LGarXiv:2607.02234v12026Tsunami: A Learned Multi-dimensional Index for Correlated Data and Skewed Workloads
Jialin Ding, Vikram Nathan, Mohammad Alizadeh +1
cs.DBcs.LGarXiv:2006.13282v12020Are Reasoning Models More Prone to Hallucination?
Zijun Yao, Yantao Liu, Yanxu Chen +5
cs.CLcs.LGarXiv:2505.23646v12025Robust Prompt Optimization for Defending Language Models Against Jailbreaking Attacks
Andy Zhou, Bo Li, Haohan Wang
cs.LGcs.AIcs.CLarXiv:2401.17263v52024Generative Pre-Training for Speech with Autoregressive Predictive Coding
Yu-An Chung, James Glass
eess.AScs.CLcs.LGarXiv:1910.12607v22019Is Chain-of-Thought Reasoning of LLMs a Mirage? A Data Distribution Lens
Chengshuai Zhao, Zhen Tan, Pingchuan Ma +5
cs.AIcs.CLcs.LGarXiv:2508.01191v62025LGGNet: Learning from Local-Global-Graph Representations for Brain-Computer Interface
Yi Ding, Neethu Robinson, Chengxuan Tong +2
cs.NEcs.LGeess.SParXiv:2105.02786v32021DeepEMD: Differentiable Earth Mover's Distance for Few-Shot Learning
Chi Zhang, Yujun Cai, Guosheng Lin +1
cs.CVcs.LGeess.IVarXiv:2003.06777v52020Steering Large Language Model Activations in Sparse Spaces
Reza Bayat, Ali Rahimi-Kalahroudi, Mohammad Pezeshki +2
cs.LGcs.AIarXiv:2503.00177v12025BEACON: Behavioral and Semantic Enrichment of AlphaEarth Embeddings through Tri-Modal Contrastive Learning
Hao Tian, Heng Cai, Yifan Yang
cs.LGarXiv:2608.29553v12026Contrastive Code Representation Learning
Paras Jain, Ajay Jain, Tianjun Zhang +3
cs.LGcs.AIcs.PLarXiv:2007.04973v42020The HSIC Bottleneck: Deep Learning without Back-Propagation
Wan-Duo Kurt Ma, J. P. Lewis, W. Bastiaan Kleijn
cs.LGstat.MLarXiv:1908.01580v32019SPA-RL: Reinforcing LLM Agents via Stepwise Progress Attribution
Hanlin Wang, Chak Tou Leong, Jiashuo Wang +2
cs.CLcs.LGarXiv:2505.20732v12025Robustness of Graph Neural Networks at Scale
Simon Geisler, Tobias Schmidt, Hakan Şirin +3
cs.LGstat.MLarXiv:2110.14038v42021Steer LLM Latents for Hallucination Detection
Seongheon Park, Xuefeng Du, Min-Hsuan Yeh +2
cs.LGcs.AIcs.CLarXiv:2503.01917v22025A Note on Shumailov et al. (2024): `AI Models Collapse When Trained on Recursively Generated Data'
Ali Borji
cs.LGcs.AIarXiv:2410.12954v22024