Machine Learning

Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

5,941 to 6,000 of 20,198

  1. Gossip Learning with Linear Models on Fully Distributed Data

    Róbert Ormándi, István Hegedüs, Márk Jelasity

    cs.LGcs.DCarXiv:1109.1396v32011
  2. AceReason-Nemotron 1.1: Advancing Math and Code Reasoning through SFT and RL Synergy

    Zihan Liu, Zhuolin Yang, Yang Chen +4

    cs.CLcs.AIcs.LGarXiv:2506.13284v12025
  3. When Models Manipulate Manifolds: The Geometry of a Counting Task

    Wes Gurnee, Emmanuel Ameisen, Isaac Kauvar +4

    cs.LGarXiv:2601.04480v12026
  4. Kernel-based Reconstruction of Graph Signals

    Daniel Romero, Meng Ma, Georgios B. Giannakis

    stat.MLcs.LGarXiv:1605.07174v12016
  5. When Decodability Is Not Enough: Logical Validity Representations, Behavioral Dissociation, and Causal Tests in Language Models

    Smitha Muthya Sudheendra, Jaideep Srivastava

    cs.CLcs.LGarXiv:2609.02438v12026
  6. Compressed Sensing with Deep Image Prior and Learned Regularization

    Dave Van Veen, Ajil Jalal, Mahdi Soltanolkotabi +3

    stat.MLcs.ITcs.LGarXiv:1806.06438v42018
  7. Residual LSTM: Design of a Deep Recurrent Architecture for Distant Speech Recognition

    Jaeyoung Kim, Mostafa El-Khamy, Jungwon Lee

    cs.LGcs.AIcs.SDarXiv:1701.03360v32017
  8. Do Multilingual LLMs Think In English?

    Lisa Schut, Yarin Gal, Sebastian Farquhar

    cs.CLcs.AIcs.LGarXiv:2502.15603v12025
  9. A Context-Aware Citation Recommendation Model with BERT and Graph Convolutional Networks

    Chanwoo Jeong, Sion Jang, Hyuna Shin +2

    cs.CLcs.IRcs.LGarXiv:1903.06464v12019
  10. SMart: A Multi-source Multi-phase Time Series Representation Transfer Framework

    Fang He, Wang-chien Lee

    cs.LGcs.AIarXiv:2609.02203v12026
  11. Schrödinger Bridges on Lie Group Manifolds for Probabilistic Intrinsic Generation

    Shizhe Zhang, Mingyang Zhao, Lei Ma

    stat.MLcs.AIcs.LGarXiv:2609.02196v12026
  12. Anonymous Walk Embeddings

    Sergey Ivanov, Evgeny Burnaev

    cs.LGstat.MLarXiv:1805.11921v32018
  13. A Biologically Plausible Supervised Learning Method for Spiking Neural Networks Using the Symmetric STDP Rule

    Yunzhe Hao, Xuhui Huang, Meng Dong +1

    cs.NEcs.AIcs.LGarXiv:1812.06574v32018
  14. Few-Shot Learning with Embedded Class Models and Shot-Free Meta Training

    Avinash Ravichandran, Rahul Bhotika, Stefano Soatto

    cs.LGcs.CVstat.MLarXiv:1905.04398v22019
  15. RSL-RL: A Learning Library for Robotics Research

    Clemens Schwarke, Mayank Mittal, Nikita Rudin +2

    cs.ROcs.LGarXiv:2509.10771v12025
  16. Visual Planning: Let's Think Only with Images

    Yi Xu, Chengzu Li, Han Zhou +4

    cs.LGcs.AIcs.CLarXiv:2505.11409v32025
  17. Implicit Latent Variable Model for Scene-Consistent Motion Forecasting

    Sergio Casas, Cole Gulino, Simon Suo +3

    cs.CVcs.LGcs.ROarXiv:2007.12036v12020
  18. Variational Bayesian Unlearning

    Quoc Phong Nguyen, Bryan Kian Hsiang Low, Patrick Jaillet

    cs.LGstat.MLarXiv:2010.12883v12020
  19. Tell me about yourself: LLMs are aware of their learned behaviors

    Jan Betley, Xuchan Bao, Martín Soto +3

    cs.CLcs.AIcs.CRarXiv:2501.11120v12025
  20. Subspace Learning and Imputation for Streaming Big Data Matrices and Tensors

    Morteza Mardani, Gonzalo Mateos, Georgios B. Giannakis

    stat.MLcs.ITcs.LGarXiv:1404.4667v12014
  21. Deep Learning for Time Series Forecasting: A Survey

    Xiangjie Kong, Zhenghao Chen, Weiyao Liu +6

    cs.LGcs.AIarXiv:2503.10198v12025
  22. Certified Robustness to Label-Flipping Attacks via Randomized Smoothing

    Elan Rosenfeld, Ezra Winston, Pradeep Ravikumar +1

    cs.LGcs.AIcs.CRarXiv:2002.03018v42020
  23. AMA-Bench: Evaluating Long-Horizon Memory for Agentic Applications

    Yujie Zhao, Boqin Yuan, Junbo Huang +9

    cs.AIcs.LGarXiv:2602.22769v42026
  24. Influence-Balanced Loss for Imbalanced Visual Classification

    Seulki Park, Jongin Lim, Younghan Jeon +1

    cs.CVcs.LGarXiv:2110.02444v12021
  25. Words or Vision: Do Vision-Language Models Have Blind Faith in Text?

    Ailin Deng, Tri Cao, Zhirui Chen +1

    cs.CVcs.AIcs.CLarXiv:2503.02199v12025
  26. BooookScore: A systematic exploration of book-length summarization in the era of LLMs

    Yapei Chang, Kyle Lo, Tanya Goyal +1

    cs.CLcs.AIcs.LGarXiv:2310.00785v42023
  27. SE(3)-Stochastic Flow Matching for Protein Backbone Generation

    Avishek Joey Bose, Tara Akhound-Sadegh, Guillaume Huguet +7

    cs.LGcs.AIarXiv:2310.02391v42023
  28. Multiresolution Recurrent Neural Networks: An Application to Dialogue Response Generation

    Iulian Vlad Serban, Tim Klinger, Gerald Tesauro +4

    cs.CLcs.AIcs.LGarXiv:1606.00776v22016
  29. Sparse Autoencoders Trained on the Same Data Learn Different Features

    Gonçalo Paulo, Nora Belrose

    cs.LGarXiv:2501.16615v22025
  30. To Talk or to Work: Flexible Communication Compression for Energy Efficient Federated Learning over Heterogeneous Mobile Edge Devices

    Liang Li, Dian Shi, Ronghui Hou +3

    cs.LGcs.AIarXiv:2012.11804v12020
  31. A Convergent Gradient Descent Algorithm for Rank Minimization and Semidefinite Programming from Random Linear Measurements

    Qinqing Zheng, John Lafferty

    stat.MLcs.LGarXiv:1506.06081v32015
  32. Imitation Learning from Imperfect Demonstration

    Yueh-Hua Wu, Nontawat Charoenphakdee, Han Bao +2

    cs.LGcs.AIstat.MLarXiv:1901.09387v32019
  33. Multi-Contrast Super-Resolution MRI Through a Progressive Network

    Qing Lyu, Hongming Shan, Ge Wang

    eess.IVcs.LGphysics.med-pharXiv:1908.01612v22019
  34. STP: Self-play LLM Theorem Provers with Iterative Conjecturing and Proving

    Kefan Dong, Tengyu Ma

    cs.LGcs.AIcs.LOarXiv:2502.00212v42025
  35. QTEA: Ternary LLMs with Sparse Residual Salient Weight and By-Column Optimization

    Yipin Guo, Arun M George, Jie Fu +3

    cs.LGcs.AIarXiv:2609.00224v22026
  36. Superintelligent Agents Pose Catastrophic Risks: Can Scientist AI Offer a Safer Path?

    Yoshua Bengio, Michael Cohen, Damiano Fornasiere +10

    cs.AIcs.LGarXiv:2502.15657v22025
  37. Neural Pruning via Growing Regularization

    Huan Wang, Can Qin, Yulun Zhang +1

    cs.CVcs.AIcs.LGarXiv:2012.09243v22020
  38. ArtGS: Building Interactable Replicas of Complex Articulated Objects via Gaussian Splatting

    Yu Liu, Baoxiong Jia, Ruijie Lu +3

    cs.CVcs.GRcs.LGarXiv:2502.19459v22025
  39. SymmCD: Symmetry-Preserving Crystal Generation with Diffusion Models

    Daniel Levy, Siba Smarak Panigrahi, Sékou-Oumar Kaba +5

    cond-mat.mtrl-scics.LGarXiv:2502.03638v32025
  40. Curriculum Reinforcement Learning from Easy to Hard Tasks Improves LLM Reasoning

    Shubham Parashar, Shurui Gui, Xiner Li +8

    cs.LGcs.AIcs.CLarXiv:2506.06632v32025
  41. LimiX: Unleashing Structured-Data Modeling Capability for Generalist Intelligence

    Xingxuan Zhang, Gang Ren, Han Yu +35

    cs.LGcs.AIcs.CLarXiv:2509.03505v22025
  42. Poisoning Attacks on LLMs Require a Near-constant Number of Poison Samples

    Alexandra Souly, Javier Rando, Ed Chapman +10

    cs.LGarXiv:2510.07192v12025
  43. WebSailor-V2: Bridging the Chasm to Proprietary Agents via Synthetic Data and Scalable Reinforcement Learning

    Kuan Li, Zhongwang Zhang, Huifeng Yin +14

    cs.LGcs.CLarXiv:2509.13305v12025
  44. On the Anatomy of MCMC-Based Maximum Likelihood Learning of Energy-Based Models

    Erik Nijkamp, Mitch Hill, Tian Han +2

    stat.MLcs.CVcs.LGarXiv:1903.12370v42019
  45. Graph Normalizing Flows

    Jenny Liu, Aviral Kumar, Jimmy Ba +2

    cs.LGstat.MLarXiv:1905.13177v12019
  46. Training High-Performance Low-Latency Spiking Neural Networks by Differentiation on Spike Representation

    Qingyan Meng, Mingqing Xiao, Shen Yan +3

    cs.NEcs.LGarXiv:2205.00459v22022
  47. Differentiable Learning-to-Normalize via Switchable Normalization

    Ping Luo, Jiamin Ren, Zhanglin Peng +2

    cs.CVcs.LGarXiv:1806.10779v52018
  48. WiSDoM: Wireless Sparse Decision Transformer with Mixture-of-Experts for Multi-Task Mobile Network Optimization

    Fatih Temiz, Shavbo Salehi, Melike Erol-Kantarci

    cs.NIcs.AIcs.LGarXiv:2609.00284v12026
  49. DeepAgent: A General Reasoning Agent with Scalable Toolsets

    Xiaoxi Li, Wenxiang Jiao, Jiarui Jin +8

    cs.AIcs.CLcs.IRarXiv:2510.21618v32025
  50. Faster Than Flash: Exploiting Attention Sparsity for Efficient Long-Context Decoding

    Zhigeng Liu, Zhiyuan Ning, Ruixiao Li +5

    cs.LGcs.AIarXiv:2609.00097v12026
  51. Data-Dependent Stability of Stochastic Gradient Descent

    Ilja Kuzborskij, Christoph H. Lampert

    cs.LGarXiv:1703.01678v42017
  52. Detecting Harmful Memes and Their Targets

    Shraman Pramanick, Dimitar Dimitrov, Rituparna Mukherjee +4

    cs.CLcs.LGcs.MMarXiv:2110.00413v12021
  53. Which Algorithmic Choices Matter at Which Batch Sizes? Insights From a Noisy Quadratic Model

    Guodong Zhang, Lala Li, Zachary Nado +5

    cs.LGstat.MLarXiv:1907.04164v22019
  54. Deep Learning Advancements in Anomaly Detection: A Comprehensive Survey

    Haoqi Huang, Ping Wang, Jianhua Pei +3

    cs.LGarXiv:2503.13195v12025
  55. Specifying Object Attributes and Relations in Interactive Scene Generation

    Oron Ashual, Lior Wolf

    cs.CVcs.LGarXiv:1909.05379v22019
  56. The Amazon Nova Family of Models: Technical Report and Model Card

    Amazon AGI, Aaron Langford, Aayush Shah +783

    cs.AIcs.CYcs.LGarXiv:2506.12103v12025
  57. The Illusion of State in State-Space Models

    William Merrill, Jackson Petty, Ashish Sabharwal

    cs.LGcs.CCcs.CLarXiv:2404.08819v32024
  58. Learning with Feature-Dependent Label Noise: A Progressive Approach

    Yikai Zhang, Songzhu Zheng, Pengxiang Wu +2

    cs.LGcs.CVstat.AParXiv:2103.07756v32021
  59. Temporal Convolution for Real-time Keyword Spotting on Mobile Devices

    Seungwoo Choi, Seokjun Seo, Beomjun Shin +5

    cs.SDcs.LGcs.NEarXiv:1904.03814v22019
  60. SWEET-RL: Training Multi-Turn LLM Agents on Collaborative Reasoning Tasks

    Yifei Zhou, Song Jiang, Yuandong Tian +4

    cs.LGarXiv:2503.15478v12025