Machine Learning

Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

481 to 540 of 20,192

  1. Measuring Semantic Similarity by Latent Relational Analysis

    Peter D. Turney

    cs.LGcs.CLcs.IRarXiv:cs/0508053v12005
  2. SyGuS-Comp 2016: Results and Analysis

    Rajeev Alur, Dana Fisman, Rishabh Singh +1

    cs.SEcs.LGcs.LOarXiv:1611.07627v12016
  3. TinyV: Reducing False Negatives in Verification Improves RL for LLM Reasoning

    Zhangchen Xu, Yuetai Li, Fengqing Jiang +4

    cs.LGcs.AIcs.CLarXiv:2505.14625v22025
  4. Sortformer: A Novel Approach for Permutation-Resolved Speaker Supervision in Speech-to-Text Systems

    Taejin Park, Ivan Medennikov, Kunal Dhawan +6

    eess.AScs.CLcs.LGarXiv:2409.06656v32024
  5. PairFlow: Closed-Form Source-Target Coupling for Few-Step Generation in Discrete Flow Models

    Mingue Park, Jisung Hwang, Seungwoo Yoo +2

    cs.LGarXiv:2512.20063v32025
  6. Multi-Task Federated Reinforcement Learning with Adversaries

    Aqeel Anwar, Arijit Raychowdhury

    cs.LGcs.AIarXiv:2103.06473v12021
  7. SpreadGNN: Serverless Multi-task Federated Learning for Graph Neural Networks

    Chaoyang He, Emir Ceyani, Keshav Balasubramanian +2

    cs.LGarXiv:2106.02743v12021
  8. Pointwise Complexity for Gaussian Fields: Upper Envelopes, Algorithmic Lower Bounds, and Separation

    Yunbei Xu

    math.PRcond-mat.stat-mechcs.ITarXiv:2606.07931v22026
  9. Predictable Scale: Part I, Step Law -- Optimal Hyperparameter Scaling Law in Large Language Model Pretraining

    Houyi Li, Wenzhen Zheng, Qiufeng Wang +10

    cs.LGcs.AIarXiv:2503.04715v72025
  10. Score Centering Stabilizes Off-policy Reinforcement Learning

    Martin Marek, Max Ryabinin

    cs.LGarXiv:2609.20807v12026
    Summaries:简体中文
  11. Active Learning for Convolutional Neural Networks: A Core-Set Approach

    Ozan Sener, Silvio Savarese

    stat.MLcs.CVcs.LGarXiv:1708.00489v42017
  12. Personalized Federated Learning with Gaussian Processes

    Idan Achituve, Aviv Shamsian, Aviv Navon +2

    cs.LGstat.MLarXiv:2106.15482v22021
  13. MAS-PromptBench: When Does Prompt Optimization Improve Multi-Agent LLM Systems?

    Juyang Bai, Laixi Shi

    cs.LGcs.MAarXiv:2606.23664v12026
  14. q-Learning in Continuous Time

    Yanwei Jia, Xun Yu Zhou

    cs.LGcs.AIq-fin.CParXiv:2207.00713v42022
  15. AgentsNet: Coordination and Collaborative Reasoning in Multi-Agent LLMs

    Florian Grötschla, Luis Müller, Jan Tönshoff +2

    cs.MAcs.LGarXiv:2507.08616v12025
  16. ResearchTown: Simulator of Human Research Community

    Haofei Yu, Zhaochen Hong, Zirui Cheng +5

    cs.CLcs.LGarXiv:2412.17767v22024
  17. STABLEVAL: Disagreement-Aware and Stable Evaluation of AI Systems

    Akash Bonagiri, Gerard Janno Anderias, Saee Patil +6

    cs.LGcs.AIarXiv:2605.02122v22026
  18. Order in the Court: Explainable AI Methods Prone to Disagreement

    Michael Neely, Stefan F. Schouten, Maurits J. R. Bleeker +1

    cs.LGcs.CLarXiv:2105.03287v32021
  19. RepMLPNet: Hierarchical Vision MLP with Re-parameterized Locality

    Xiaohan Ding, Honghao Chen, Xiangyu Zhang +2

    cs.CVcs.AIcs.LGarXiv:2112.11081v22021
  20. When to Show a Suggestion? Integrating Human Feedback in AI-Assisted Programming

    Hussein Mozannar, Gagan Bansal, Adam Fourney +1

    cs.HCcs.LGcs.SEarXiv:2306.04930v32023
    Summaries:한국어
  21. State of the Art Control of Atari Games Using Shallow Reinforcement Learning

    Yitao Liang, Marlos C. Machado, Erik Talvitie +1

    cs.LGarXiv:1512.01563v22015
  22. SoundReactor: Frame-level Online Video-to-Audio Generation

    Koichi Saito, Julian Tanke, Christian Simon +7

    cs.SDcs.LGeess.ASarXiv:2510.02110v12025
  23. Fair and Diverse DPP-based Data Summarization

    L. Elisa Celis, Vijay Keswani, Damian Straszak +3

    cs.LGcs.CYcs.IRarXiv:1802.04023v12018
  24. Progressive Distillation for Fast Sampling of Diffusion Models

    Tim Salimans, Jonathan Ho

    cs.LGcs.AIstat.MLarXiv:2202.00512v22022
  25. Entropy-SGD: Biasing Gradient Descent Into Wide Valleys

    Pratik Chaudhari, Anna Choromanska, Stefano Soatto +6

    cs.LGarXiv:1611.01838v42016
  26. Learning to Reason for Hallucination Span Detection

    Hsuan Su, Ting-Yao Hu, Hema Swetha Koppula +7

    cs.CLcs.AIcs.LGarXiv:2510.02173v22025
  27. R-Transformer: Recurrent Neural Network Enhanced Transformer

    Zhiwei Wang, Yao Ma, Zitao Liu +1

    cs.LGcs.CLcs.CVarXiv:1907.05572v12019
  28. A Near-Linear Time Algorithm for the Chamfer Distance

    Ainesh Bakshi, Piotr Indyk, Rajesh Jayaram +2

    cs.DScs.CGcs.GRarXiv:2307.03043v12023
  29. Lipschitz Bandits with Stochastic Delayed Feedback

    Zhongxuan Liu, Yue Kang, Thomas C. M. Lee

    cs.LGstat.MLarXiv:2510.00309v22025
  30. Measuring training variability from stochastic optimization using robust nonparametric testing

    Sinjini Banerjee, Tim Marrinan, Reilly Cannon +2

    stat.MLcs.LGarXiv:2406.08307v22024
  31. Normalization Propagation: A Parametric Technique for Removing Internal Covariate Shift in Deep Networks

    Devansh Arpit, Yingbo Zhou, Bhargava U. Kota +1

    stat.MLcs.LGarXiv:1603.01431v62016
  32. DiagrammerGPT: Generating Open-Domain, Open-Platform Diagrams via LLM Planning

    Abhay Zala, Han Lin, Jaemin Cho +1

    cs.CVcs.AIcs.CLarXiv:2310.12128v22023
  33. Weight decay induces low-rank attention layers

    Seijin Kobayashi, Yassir Akram, Johannes Von Oswald

    cs.LGarXiv:2410.23819v12024
  34. Think in English, Answer in Korean: Efficient Adaptation of Multilingual Tool-Using Agents

    Utsav Garg, Sungjin Hong, Jason Jung +6

    cs.AIcs.LGarXiv:2606.31648v12026
  35. Online convex optimization in the bandit setting: gradient descent without a gradient

    Abraham D. Flaxman, Adam Tauman Kalai, H. Brendan McMahan

    cs.LGcs.CCarXiv:cs/0408007v12004
  36. Tight Differential Privacy for Discrete-Valued Mechanisms and for the Subsampled Gaussian Mechanism Using FFT

    Antti Koskela, Joonas Jälkö, Lukas Prediger +1

    stat.MLcs.CRcs.LGarXiv:2006.07134v32020
  37. Addressing Some Limitations of Transformers with Feedback Memory

    Angela Fan, Thibaut Lavril, Edouard Grave +2

    cs.LGcs.CLstat.MLarXiv:2002.09402v32020
  38. Fair Adversarial Gradient Tree Boosting

    Vincent Grari, Boris Ruf, Sylvain Lamprier +1

    cs.LGcs.AIcs.CYarXiv:1911.05369v22019
  39. Estimating individual treatment effect: generalization bounds and algorithms

    Uri Shalit, Fredrik D. Johansson, David Sontag

    stat.MLcs.AIcs.LGarXiv:1606.03976v52016
  40. Reinforcement Learning with Verifiable yet Noisy Rewards under Imperfect Verifiers

    Xin-Qiang Cai, Wei Wang, Feng Liu +3

    cs.LGcs.AIarXiv:2510.00915v42025
  41. Geometric Operator Learning with Optimal Transport

    Xinyi Li, Zongyi Li, Nikola Kovachki +1

    cs.LGarXiv:2507.20065v12025
  42. OceanLight: Efficient Global Ocean Forecasting via Geometry-Adaptive Unstructured Mesh Representation

    Wei Wu, Xiang Wang, Hongze Leng +3

    cs.LGcs.AIarXiv:2608.16070v12026
  43. Paper2Agent: Reimagining Research Papers As Interactive and Reliable AI Agents

    Jiacheng Miao, Joe R. Davis, Yaohui Zhang +2

    cs.AIcs.CLcs.LGarXiv:2509.06917v22025
  44. Multimodal Whole Slide Foundation Model for Pathology

    Tong Ding, Sophia J. Wagner, Andrew H. Song +20

    eess.IVcs.AIcs.CVarXiv:2411.19666v12024
  45. DeBERTa: Decoding-enhanced BERT with Disentangled Attention

    Pengcheng He, Xiaodong Liu, Jianfeng Gao +1

    cs.CLcs.LGarXiv:2006.03654v62020
  46. On the Value of Out-of-Distribution Testing: An Example of Goodhart's Law

    Damien Teney, Kushal Kafle, Robik Shrestha +3

    cs.CVcs.LGarXiv:2005.09241v12020
  47. Smoothed Dilated Convolutions for Improved Dense Prediction

    Zhengyang Wang, Shuiwang Ji

    cs.CVcs.LGarXiv:1808.08931v22018
  48. Sleeper Agents: Training Deceptive LLMs that Persist Through Safety Training

    Evan Hubinger, Carson Denison, Jesse Mu +36

    cs.CRcs.AIcs.CLarXiv:2401.05566v32024
  49. Addressing the Item Cold-start Problem by Attribute-driven Active Learning

    Yu Zhu, Jinhao Lin, Shibi He +4

    cs.IRcs.LGstat.MLarXiv:1805.09023v12018
  50. MotifNet: a motif-based Graph Convolutional Network for directed graphs

    Federico Monti, Karl Otness, Michael M. Bronstein

    cs.LGarXiv:1802.01572v12018
  51. Wav2Letter: an End-to-End ConvNet-based Speech Recognition System

    Ronan Collobert, Christian Puhrsch, Gabriel Synnaeve

    cs.LGcs.AIcs.CLarXiv:1609.03193v22016
  52. Sparse MoEs meet Efficient Ensembles

    James Urquhart Allingham, Florian Wenzel, Zelda E Mariet +10

    cs.LGcs.CVstat.MLarXiv:2110.03360v22021
  53. Graph of Thoughts: Solving Elaborate Problems with Large Language Models

    Maciej Besta, Nils Blach, Ales Kubicek +8

    cs.CLcs.AIcs.LGarXiv:2308.09687v42023
  54. The Impact of Positional Encoding on Length Generalization in Transformers

    Amirhossein Kazemnejad, Inkit Padhi, Karthikeyan Natesan Ramamurthy +2

    cs.CLcs.AIcs.LGarXiv:2305.19466v22023
  55. ProlificDreamer: High-Fidelity and Diverse Text-to-3D Generation with Variational Score Distillation

    Zhengyi Wang, Cheng Lu, Yikai Wang +4

    cs.LGcs.CVarXiv:2305.16213v22023
  56. Learning Fine-Grained Bimanual Manipulation with Low-Cost Hardware

    Tony Z. Zhao, Vikash Kumar, Sergey Levine +1

    cs.ROcs.LGarXiv:2304.13705v12023
  57. StoRM: A Diffusion-based Stochastic Regeneration Model for Speech Enhancement and Dereverberation

    Jean-Marie Lemercier, Julius Richter, Simon Welker +1

    eess.AScs.LGcs.SDarXiv:2212.11851v22022
  58. Beyond neural scaling laws: beating power law scaling via data pruning

    Ben Sorscher, Robert Geirhos, Shashank Shekhar +2

    cs.LGcs.AIcs.CVarXiv:2206.14486v62022
  59. FedBABU: Towards Enhanced Representation for Federated Image Classification

    Jaehoon Oh, Sangmook Kim, Se-Young Yun

    cs.LGarXiv:2106.06042v32021
  60. Back into Plato's Cave: Examining Cross-modal Representational Convergence at Scale

    A. Sophia Koepke, Daniil Zverev, Shiry Ginosar +1

    cs.CVcs.AIcs.LGarXiv:2604.18572v22026