Machine Learning
Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
16,441 to 16,500 of 20,193
Efficiently Scaling Transformer Inference
Reiner Pope, Sholto Douglas, Aakanksha Chowdhery +7
cs.LGcs.CLarXiv:2211.05102v12022High Frequency Component Helps Explain the Generalization of Convolutional Neural Networks
Haohan Wang, Xindi Wu, Zeyi Huang +1
cs.CVcs.LGarXiv:1905.13545v32019Predictive Entropy Search for Efficient Global Optimization of Black-box Functions
José Miguel Hernández-Lobato, Matthew W. Hoffman, Zoubin Ghahramani
stat.MLcs.LGarXiv:1406.2541v12014Optimal Distributed Online Prediction using Mini-Batches
Ofer Dekel, Ran Gilad-Bachrach, Ohad Shamir +1
cs.LGcs.DCmath.OCarXiv:1012.1367v22010BPDQ: Bit-Plane Decomposition Quantization on a Variable Grid for Large Language Models
Junyu Chen, Jungang Li, Jing Xiong +11
cs.LGarXiv:2602.04163v22026Temporal Pair Consistency for Variance-Reduced Flow Matching
Chika Maduabuchi, Jindong Wang
cs.LGcs.AIcs.CVarXiv:2602.04908v22026No One-Size-Fits-All: Building Systems For Translation to Bashkir, Kazakh, Kyrgyz, Tatar and Chuvash Using Synthetic And Original Data
Dmitry Karpov
cs.CLcs.AIcs.LGarXiv:2602.04442v12026SEM: Sparse Embedding Modulation for Post-Hoc Debiasing of Vision-Language Models
Quentin Guimard, Federico Bartsch, Simone Caldarella +3
cs.CVcs.AIcs.LGarXiv:2603.19028v12026FedML: A Research Library and Benchmark for Federated Machine Learning
Chaoyang He, Songze Li, Jinhyun So +17
cs.LGstat.MLarXiv:2007.13518v42020MathVerse: Does Your Multi-modal LLM Truly See the Diagrams in Visual Math Problems?
Renrui Zhang, Dongzhi Jiang, Yichi Zhang +8
cs.CVcs.AIcs.CLarXiv:2403.14624v22024Late-to-Early Training: LET LLMs Learn Earlier, So Faster and Better
Ji Zhao, Yufei Gu, Shitong Shao +3
cs.CLcs.LGarXiv:2602.05393v12026VAE with a VampPrior
Jakub M. Tomczak, Max Welling
cs.LGcs.AIstat.MLarXiv:1705.07120v52017Precision-Aware Variable Bit Processing Elements for Hardware-Efficient Systolic Array Designs
Dantu Nandini Devi, Madhav Rao
cs.ARcs.ETcs.LGarXiv:2608.22378v12026Social Attention: Modeling Attention in Human Crowds
Anirudh Vemula, Katharina Muelling, Jean Oh
cs.ROcs.LGarXiv:1710.04689v22017HashNet: Deep Learning to Hash by Continuation
Zhangjie Cao, Mingsheng Long, Jianmin Wang +1
cs.LGcs.CVarXiv:1702.00758v42017Editing Factual Knowledge in Language Models
Nicola De Cao, Wilker Aziz, Ivan Titov
cs.CLcs.AIcs.LGarXiv:2104.08164v22021Dreaming in Code for Curriculum Learning in Open-Ended Worlds
Konstantinos Mitsides, Maxence Faldor, Antoine Cully
cs.LGcs.AIcs.CLarXiv:2602.08194v12026Label-Efficient Semantic Segmentation with Diffusion Models
Dmitry Baranchuk, Ivan Rubachev, Andrey Voynov +2
cs.CVcs.LGarXiv:2112.03126v32021FlexMoRE: A Flexible Mixture of Rank-heterogeneous Experts for Efficient Federatedly-trained Large Language Models
Annemette Brok Pirchert, Jacob Nielsen, Mogens Henrik From +2
cs.LGarXiv:2602.08818v12026On the Optimal Reasoning Length for RL-Trained Language Models
Daisuke Nohara, Taishi Nakamura, Rio Yokota
cs.CLcs.AIcs.LGarXiv:2602.09591v32026Deep neural network models for computational histopathology: A survey
Chetan L. Srinidhi, Ozan Ciga, Anne L. Martel
eess.IVcs.CVcs.LGarXiv:1912.12378v22019A Comprehensive Survey of Few-shot Learning: Evolution, Applications, Challenges, and Opportunities
Yisheng Song, Ting Wang, Subrota K Mondal +1
cs.LGarXiv:2205.06743v22022Global Convergence of Policy Gradient Methods for the Linear Quadratic Regulator
Maryam Fazel, Rong Ge, Sham M. Kakade +1
cs.LGstat.MLarXiv:1801.05039v32018Leveraging Procedural Generation to Benchmark Reinforcement Learning
Karl Cobbe, Christopher Hesse, Jacob Hilton +1
cs.LGstat.MLarXiv:1912.01588v22019Internalizing Meta-Experience into Memory for Guided Reinforcement Learning in Large Language Models
Shiting Huang, Zecheng Li, Yu Zeng +7
cs.LGcs.AIarXiv:2602.10224v12026Pervasive Label Errors in Test Sets Destabilize Machine Learning Benchmarks
Curtis G. Northcutt, Anish Athalye, Jonas Mueller
stat.MLcs.AIcs.LGarXiv:2103.14749v42021Deep Evidential Regression
Alexander Amini, Wilko Schwarting, Ava Soleimany +1
cs.LGcs.NEstat.MLarXiv:1910.02600v22019How Architecture and Training Affect TPC Representations Across Experiments
Tyler Wheeler, Michelle P. Kuchera, Raghuram Ramanujan +11
cs.LGcs.CVnucl-exarXiv:2608.21756v12026Implicit Quantile Networks for Distributional Reinforcement Learning
Will Dabney, Georg Ostrovski, David Silver +1
cs.LGcs.AIstat.MLarXiv:1806.06923v12018DIGIT: A Novel Design for a Low-Cost Compact High-Resolution Tactile Sensor with Application to In-Hand Manipulation
Mike Lambeta, Po-Wei Chou, Stephen Tian +9
cs.ROcs.LGeess.SYarXiv:2005.14679v12020Using Simulation and Domain Adaptation to Improve Efficiency of Deep Robotic Grasping
Konstantinos Bousmalis, Alex Irpan, Paul Wohlhart +9
cs.LGcs.AIcs.CVarXiv:1709.07857v22017UberNet: Training a `Universal' Convolutional Neural Network for Low-, Mid-, and High-Level Vision using Diverse Datasets and Limited Memory
Iasonas Kokkinos
cs.CVcs.AIcs.LGarXiv:1609.02132v12016Frontier AI Risk Management Framework in Practice: A Risk Analysis Technical Report v1.5
Dongrui Liu, Yi Yu, Jie Zhang +18
cs.AIcs.CLcs.CVarXiv:2602.14457v12026Convolutional Neural Networks for Classification of Alzheimer's Disease: Overview and Reproducible Evaluation
Junhao Wen, Elina Thibeau-Sutre, Mauricio Diaz-Melo +7
cs.LGeess.IVstat.MLarXiv:1904.07773v62019OPBench: A Graph Benchmark to Combat the Opioid Crisis
Tianyi Ma, Yiyang Li, Yiyue Qian +4
cs.LGcs.AIarXiv:2602.14602v12026Generating Multi-label Discrete Patient Records using Generative Adversarial Networks
Edward Choi, Siddharth Biswal, Bradley Malin +3
cs.LGcs.NEarXiv:1703.06490v32017More accurate behavioral predictions with hybrid Bayesian-connectionist models
Brenden M. Lake, Akshay K. Jagadish, Guangyuan Jiang
cs.LGarXiv:2608.22154v12026A comparative study of fairness-enhancing interventions in machine learning
Sorelle A. Friedler, Carlos Scheidegger, Suresh Venkatasubramanian +3
stat.MLcs.CYcs.LGarXiv:1802.04422v12018Performance Metrics (Error Measures) in Machine Learning Regression, Forecasting and Prognostics: Properties and Typology
Alexei Botchkarev
stat.MEcs.LGstat.MLarXiv:1809.03006v12018Intent Laundering: AI Safety Datasets Are Not What They Seem
Shahriar Golchin, Marc Wetter
cs.CRcs.AIcs.CLarXiv:2602.16729v32026TAROT: Test-driven and Capability-adaptive Curriculum Reinforcement Fine-tuning for Code Generation with Large Language Models
Chansung Park, Juyong Jiang, Fan Wang +4
cs.CLcs.LGcs.SEarXiv:2602.15449v12026Deep Learning-Based Channel Estimation
Mehran Soltani, Vahid Pourahmadi, Ali Mirzaei +1
cs.ITcs.LGeess.SParXiv:1810.05893v42018VATEX: A Large-Scale, High-Quality Multilingual Dataset for Video-and-Language Research
Xin Wang, Jiawei Wu, Junkun Chen +3
cs.CVcs.CLcs.LGarXiv:1904.03493v32019Multimodal Injury Risk and Performance Prediction in Tennis Using Weighted Ensemble Learning
Weihao Qu, Dongyang Wang, Ling Zheng +3
cs.LGarXiv:2608.21530v12026A Semantic Matching Energy Function for Learning with Multi-relational Data
Xavier Glorot, Antoine Bordes, Jason Weston +1
cs.LGarXiv:1301.3485v22013AAVGen: Precision Engineering of Adeno-associated Viral Capsids for Renal Selective Targeting
Mohammadreza Ghaffarzadeh-Esfahani, Yousof Gheisari
q-bio.QMcs.AIcs.CLarXiv:2602.18915v12026Semi-Supervised Learning with Generative Adversarial Networks
Augustus Odena
stat.MLcs.LGarXiv:1606.01583v22016Ani3DHuman: Photorealistic 3D Human Animation with Self-guided Stochastic Sampling
Qi Sun, Can Wang, Jiaxiang Shang +2
cs.CVcs.GRcs.LGarXiv:2602.19089v12026Label Propagation for Deep Semi-supervised Learning
Ahmet Iscen, Giorgos Tolias, Yannis Avrithis +1
cs.CVcs.LGarXiv:1904.04717v12019Taming Throughput-Latency Tradeoff in LLM Inference with Sarathi-Serve
Amey Agrawal, Nitin Kedia, Ashish Panwar +5
cs.LGcs.DCarXiv:2403.02310v32024NanoKnow: How to Know What Your Language Model Knows
Lingwei Gu, Nour Jedidi, Jimmy Lin
cs.CLcs.AIcs.IRarXiv:2602.20122v22026MRMAD: A Multi-Round Multi-Audio Benchmark for Evaluating Acoustic Degradation Perception in Large Audio-Language Models
Yize Li, Ningyuan Yang, Sile Yin +6
cs.SDcs.LGeess.ASarXiv:2608.22236v12026Just Train Twice: Improving Group Robustness without Training Group Information
Evan Zheran Liu, Behzad Haghgoo, Annie S. Chen +5
cs.LGcs.AIcs.CYarXiv:2107.09044v22021QEDBENCH: Quantifying the Alignment Gap in Automated Evaluation of University-Level Mathematical Proofs
Santiago Gonzalez, Alireza Amiri Bavandpour, Peter Ye +48
cs.LGarXiv:2602.20629v32026Shared Nature, Unique Nurture: PRISM for Pluralistic Reasoning via In-context Structure Modeling
Guancheng Tu, Shiyang Zhang, Tianyu Zhang +2
cs.LGarXiv:2602.21317v12026Easy to Learn, Yet Hard to Forget: Towards Robust Unlearning Under Bias
JuneHyoung Kwon, MiHyeon Kim, Eunju Lee +3
cs.LGcs.CVarXiv:2602.21773v12026Token-Level Likelihood-Array Regression for Membership Inference and AI-Generated Text Detection
Jiajun Sun, Zhanrui Cai
stat.MLcs.LGstat.MEarXiv:2608.22179v12026Mobile ALOHA: Learning Bimanual Mobile Manipulation with Low-Cost Whole-Body Teleoperation
Zipeng Fu, Tony Z. Zhao, Chelsea Finn
cs.ROcs.AIcs.CVarXiv:2401.02117v12024Multi-Head Low-Rank Attention
Songtao Liu, Hongwu Peng, Zhiwei Zhang +2
cs.LGarXiv:2603.02188v12026Reward Constrained Policy Optimization
Chen Tessler, Daniel J. Mankowitz, Shie Mannor
cs.LGcs.AIstat.MLarXiv:1805.11074v32018