Machine Learning

Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

16,441 to 16,500 of 20,193

  1. Efficiently Scaling Transformer Inference

    Reiner Pope, Sholto Douglas, Aakanksha Chowdhery +7

    cs.LGcs.CLarXiv:2211.05102v12022
  2. High Frequency Component Helps Explain the Generalization of Convolutional Neural Networks

    Haohan Wang, Xindi Wu, Zeyi Huang +1

    cs.CVcs.LGarXiv:1905.13545v32019
  3. Predictive Entropy Search for Efficient Global Optimization of Black-box Functions

    José Miguel Hernández-Lobato, Matthew W. Hoffman, Zoubin Ghahramani

    stat.MLcs.LGarXiv:1406.2541v12014
  4. Optimal Distributed Online Prediction using Mini-Batches

    Ofer Dekel, Ran Gilad-Bachrach, Ohad Shamir +1

    cs.LGcs.DCmath.OCarXiv:1012.1367v22010
  5. BPDQ: Bit-Plane Decomposition Quantization on a Variable Grid for Large Language Models

    Junyu Chen, Jungang Li, Jing Xiong +11

    cs.LGarXiv:2602.04163v22026
  6. Temporal Pair Consistency for Variance-Reduced Flow Matching

    Chika Maduabuchi, Jindong Wang

    cs.LGcs.AIcs.CVarXiv:2602.04908v22026
  7. No One-Size-Fits-All: Building Systems For Translation to Bashkir, Kazakh, Kyrgyz, Tatar and Chuvash Using Synthetic And Original Data

    Dmitry Karpov

    cs.CLcs.AIcs.LGarXiv:2602.04442v12026
  8. SEM: Sparse Embedding Modulation for Post-Hoc Debiasing of Vision-Language Models

    Quentin Guimard, Federico Bartsch, Simone Caldarella +3

    cs.CVcs.AIcs.LGarXiv:2603.19028v12026
  9. FedML: A Research Library and Benchmark for Federated Machine Learning

    Chaoyang He, Songze Li, Jinhyun So +17

    cs.LGstat.MLarXiv:2007.13518v42020
  10. MathVerse: Does Your Multi-modal LLM Truly See the Diagrams in Visual Math Problems?

    Renrui Zhang, Dongzhi Jiang, Yichi Zhang +8

    cs.CVcs.AIcs.CLarXiv:2403.14624v22024
  11. Late-to-Early Training: LET LLMs Learn Earlier, So Faster and Better

    Ji Zhao, Yufei Gu, Shitong Shao +3

    cs.CLcs.LGarXiv:2602.05393v12026
  12. VAE with a VampPrior

    Jakub M. Tomczak, Max Welling

    cs.LGcs.AIstat.MLarXiv:1705.07120v52017
  13. Precision-Aware Variable Bit Processing Elements for Hardware-Efficient Systolic Array Designs

    Dantu Nandini Devi, Madhav Rao

    cs.ARcs.ETcs.LGarXiv:2608.22378v12026
  14. Social Attention: Modeling Attention in Human Crowds

    Anirudh Vemula, Katharina Muelling, Jean Oh

    cs.ROcs.LGarXiv:1710.04689v22017
  15. HashNet: Deep Learning to Hash by Continuation

    Zhangjie Cao, Mingsheng Long, Jianmin Wang +1

    cs.LGcs.CVarXiv:1702.00758v42017
  16. Editing Factual Knowledge in Language Models

    Nicola De Cao, Wilker Aziz, Ivan Titov

    cs.CLcs.AIcs.LGarXiv:2104.08164v22021
  17. Dreaming in Code for Curriculum Learning in Open-Ended Worlds

    Konstantinos Mitsides, Maxence Faldor, Antoine Cully

    cs.LGcs.AIcs.CLarXiv:2602.08194v12026
  18. Label-Efficient Semantic Segmentation with Diffusion Models

    Dmitry Baranchuk, Ivan Rubachev, Andrey Voynov +2

    cs.CVcs.LGarXiv:2112.03126v32021
  19. FlexMoRE: A Flexible Mixture of Rank-heterogeneous Experts for Efficient Federatedly-trained Large Language Models

    Annemette Brok Pirchert, Jacob Nielsen, Mogens Henrik From +2

    cs.LGarXiv:2602.08818v12026
  20. On the Optimal Reasoning Length for RL-Trained Language Models

    Daisuke Nohara, Taishi Nakamura, Rio Yokota

    cs.CLcs.AIcs.LGarXiv:2602.09591v32026
  21. Deep neural network models for computational histopathology: A survey

    Chetan L. Srinidhi, Ozan Ciga, Anne L. Martel

    eess.IVcs.CVcs.LGarXiv:1912.12378v22019
  22. A Comprehensive Survey of Few-shot Learning: Evolution, Applications, Challenges, and Opportunities

    Yisheng Song, Ting Wang, Subrota K Mondal +1

    cs.LGarXiv:2205.06743v22022
  23. Global Convergence of Policy Gradient Methods for the Linear Quadratic Regulator

    Maryam Fazel, Rong Ge, Sham M. Kakade +1

    cs.LGstat.MLarXiv:1801.05039v32018
  24. Leveraging Procedural Generation to Benchmark Reinforcement Learning

    Karl Cobbe, Christopher Hesse, Jacob Hilton +1

    cs.LGstat.MLarXiv:1912.01588v22019
  25. Internalizing Meta-Experience into Memory for Guided Reinforcement Learning in Large Language Models

    Shiting Huang, Zecheng Li, Yu Zeng +7

    cs.LGcs.AIarXiv:2602.10224v12026
  26. Pervasive Label Errors in Test Sets Destabilize Machine Learning Benchmarks

    Curtis G. Northcutt, Anish Athalye, Jonas Mueller

    stat.MLcs.AIcs.LGarXiv:2103.14749v42021
  27. Deep Evidential Regression

    Alexander Amini, Wilko Schwarting, Ava Soleimany +1

    cs.LGcs.NEstat.MLarXiv:1910.02600v22019
  28. How Architecture and Training Affect TPC Representations Across Experiments

    Tyler Wheeler, Michelle P. Kuchera, Raghuram Ramanujan +11

    cs.LGcs.CVnucl-exarXiv:2608.21756v12026
  29. Implicit Quantile Networks for Distributional Reinforcement Learning

    Will Dabney, Georg Ostrovski, David Silver +1

    cs.LGcs.AIstat.MLarXiv:1806.06923v12018
  30. DIGIT: A Novel Design for a Low-Cost Compact High-Resolution Tactile Sensor with Application to In-Hand Manipulation

    Mike Lambeta, Po-Wei Chou, Stephen Tian +9

    cs.ROcs.LGeess.SYarXiv:2005.14679v12020
  31. Using Simulation and Domain Adaptation to Improve Efficiency of Deep Robotic Grasping

    Konstantinos Bousmalis, Alex Irpan, Paul Wohlhart +9

    cs.LGcs.AIcs.CVarXiv:1709.07857v22017
  32. UberNet: Training a `Universal' Convolutional Neural Network for Low-, Mid-, and High-Level Vision using Diverse Datasets and Limited Memory

    Iasonas Kokkinos

    cs.CVcs.AIcs.LGarXiv:1609.02132v12016
  33. Frontier AI Risk Management Framework in Practice: A Risk Analysis Technical Report v1.5

    Dongrui Liu, Yi Yu, Jie Zhang +18

    cs.AIcs.CLcs.CVarXiv:2602.14457v12026
  34. Convolutional Neural Networks for Classification of Alzheimer's Disease: Overview and Reproducible Evaluation

    Junhao Wen, Elina Thibeau-Sutre, Mauricio Diaz-Melo +7

    cs.LGeess.IVstat.MLarXiv:1904.07773v62019
  35. OPBench: A Graph Benchmark to Combat the Opioid Crisis

    Tianyi Ma, Yiyang Li, Yiyue Qian +4

    cs.LGcs.AIarXiv:2602.14602v12026
  36. Generating Multi-label Discrete Patient Records using Generative Adversarial Networks

    Edward Choi, Siddharth Biswal, Bradley Malin +3

    cs.LGcs.NEarXiv:1703.06490v32017
  37. More accurate behavioral predictions with hybrid Bayesian-connectionist models

    Brenden M. Lake, Akshay K. Jagadish, Guangyuan Jiang

    cs.LGarXiv:2608.22154v12026
  38. A comparative study of fairness-enhancing interventions in machine learning

    Sorelle A. Friedler, Carlos Scheidegger, Suresh Venkatasubramanian +3

    stat.MLcs.CYcs.LGarXiv:1802.04422v12018
  39. Performance Metrics (Error Measures) in Machine Learning Regression, Forecasting and Prognostics: Properties and Typology

    Alexei Botchkarev

    stat.MEcs.LGstat.MLarXiv:1809.03006v12018
  40. Intent Laundering: AI Safety Datasets Are Not What They Seem

    Shahriar Golchin, Marc Wetter

    cs.CRcs.AIcs.CLarXiv:2602.16729v32026
  41. TAROT: Test-driven and Capability-adaptive Curriculum Reinforcement Fine-tuning for Code Generation with Large Language Models

    Chansung Park, Juyong Jiang, Fan Wang +4

    cs.CLcs.LGcs.SEarXiv:2602.15449v12026
  42. Deep Learning-Based Channel Estimation

    Mehran Soltani, Vahid Pourahmadi, Ali Mirzaei +1

    cs.ITcs.LGeess.SParXiv:1810.05893v42018
  43. VATEX: A Large-Scale, High-Quality Multilingual Dataset for Video-and-Language Research

    Xin Wang, Jiawei Wu, Junkun Chen +3

    cs.CVcs.CLcs.LGarXiv:1904.03493v32019
  44. Multimodal Injury Risk and Performance Prediction in Tennis Using Weighted Ensemble Learning

    Weihao Qu, Dongyang Wang, Ling Zheng +3

    cs.LGarXiv:2608.21530v12026
  45. A Semantic Matching Energy Function for Learning with Multi-relational Data

    Xavier Glorot, Antoine Bordes, Jason Weston +1

    cs.LGarXiv:1301.3485v22013
  46. AAVGen: Precision Engineering of Adeno-associated Viral Capsids for Renal Selective Targeting

    Mohammadreza Ghaffarzadeh-Esfahani, Yousof Gheisari

    q-bio.QMcs.AIcs.CLarXiv:2602.18915v12026
  47. Semi-Supervised Learning with Generative Adversarial Networks

    Augustus Odena

    stat.MLcs.LGarXiv:1606.01583v22016
  48. Ani3DHuman: Photorealistic 3D Human Animation with Self-guided Stochastic Sampling

    Qi Sun, Can Wang, Jiaxiang Shang +2

    cs.CVcs.GRcs.LGarXiv:2602.19089v12026
  49. Label Propagation for Deep Semi-supervised Learning

    Ahmet Iscen, Giorgos Tolias, Yannis Avrithis +1

    cs.CVcs.LGarXiv:1904.04717v12019
  50. Taming Throughput-Latency Tradeoff in LLM Inference with Sarathi-Serve

    Amey Agrawal, Nitin Kedia, Ashish Panwar +5

    cs.LGcs.DCarXiv:2403.02310v32024
  51. NanoKnow: How to Know What Your Language Model Knows

    Lingwei Gu, Nour Jedidi, Jimmy Lin

    cs.CLcs.AIcs.IRarXiv:2602.20122v22026
  52. MRMAD: A Multi-Round Multi-Audio Benchmark for Evaluating Acoustic Degradation Perception in Large Audio-Language Models

    Yize Li, Ningyuan Yang, Sile Yin +6

    cs.SDcs.LGeess.ASarXiv:2608.22236v12026
  53. Just Train Twice: Improving Group Robustness without Training Group Information

    Evan Zheran Liu, Behzad Haghgoo, Annie S. Chen +5

    cs.LGcs.AIcs.CYarXiv:2107.09044v22021
  54. QEDBENCH: Quantifying the Alignment Gap in Automated Evaluation of University-Level Mathematical Proofs

    Santiago Gonzalez, Alireza Amiri Bavandpour, Peter Ye +48

    cs.LGarXiv:2602.20629v32026
  55. Shared Nature, Unique Nurture: PRISM for Pluralistic Reasoning via In-context Structure Modeling

    Guancheng Tu, Shiyang Zhang, Tianyu Zhang +2

    cs.LGarXiv:2602.21317v12026
  56. Easy to Learn, Yet Hard to Forget: Towards Robust Unlearning Under Bias

    JuneHyoung Kwon, MiHyeon Kim, Eunju Lee +3

    cs.LGcs.CVarXiv:2602.21773v12026
  57. Token-Level Likelihood-Array Regression for Membership Inference and AI-Generated Text Detection

    Jiajun Sun, Zhanrui Cai

    stat.MLcs.LGstat.MEarXiv:2608.22179v12026
  58. Mobile ALOHA: Learning Bimanual Mobile Manipulation with Low-Cost Whole-Body Teleoperation

    Zipeng Fu, Tony Z. Zhao, Chelsea Finn

    cs.ROcs.AIcs.CVarXiv:2401.02117v12024
  59. Multi-Head Low-Rank Attention

    Songtao Liu, Hongwu Peng, Zhiwei Zhang +2

    cs.LGarXiv:2603.02188v12026
  60. Reward Constrained Policy Optimization

    Chen Tessler, Daniel J. Mankowitz, Shie Mannor

    cs.LGcs.AIstat.MLarXiv:1805.11074v32018