Machine Learning

Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

6,361 to 6,420 of 20,454

  1. Mutual information for symmetric rank-one matrix estimation: A proof of the replica formula

    Jean Barbier, Mohamad Dia, Nicolas Macris +3

    cs.ITcond-mat.dis-nncs.LGarXiv:1606.04142v12016
  2. Predicting Citywide Crowd Flows in Irregular Regions Using Multi-View Graph Convolutional Networks

    Junkai Sun, Junbo Zhang, Qiaofei Li +3

    cs.CVcs.LGarXiv:1903.07789v22019
  3. Learning ReLUs via Gradient Descent

    Mahdi Soltanolkotabi

    cs.LGcs.ITmath.OCarXiv:1705.04591v22017
  4. Modeling Sentiment Dependencies with Graph Convolutional Networks for Aspect-level Sentiment Classification

    Pinlong Zhaoa, Linlin Houb, Ou Wua

    cs.CLcs.LGarXiv:1906.04501v12019
  5. SWE-Lancer: Can Frontier LLMs Earn $1 Million from Real-World Freelance Software Engineering?

    Samuel Miserendino, Michele Wang, Tejal Patwardhan +1

    cs.LGcs.SEarXiv:2502.12115v42025
  6. Neural Rough Differential Equations for Long Time Series

    James Morrill, Cristopher Salvi, Patrick Kidger +2

    cs.LGcs.AImath.DSarXiv:2009.08295v42020
  7. CTAB-GAN+: Enhancing Tabular Data Synthesis

    Zilong Zhao, Aditya Kunar, Robert Birke +1

    cs.LGarXiv:2204.00401v12022
  8. On the Adversarial Robustness of Vision Transformers

    Rulin Shao, Zhouxing Shi, Jinfeng Yi +2

    cs.CVcs.AIcs.LGarXiv:2103.15670v32021
  9. Towards Efficient Model Compression via Learned Global Ranking

    Ting-Wu Chin, Ruizhou Ding, Cha Zhang +1

    cs.CVcs.LGarXiv:1904.12368v22019
  10. Sparse Autoencoders Learn Monosemantic Features in Vision-Language Models

    Mateusz Pach, Shyamgopal Karthik, Quentin Bouniot +2

    cs.CVcs.AIcs.LGarXiv:2504.02821v32025
  11. PointVLA: Injecting the 3D World into Vision-Language-Action Models

    Chengmeng Li, Junjie Wen, Yan Peng +3

    cs.ROcs.CVcs.LGarXiv:2503.07511v12025
  12. Assessing Alignment and Stability of Feature Importance Explanations via Weight of Evidence

    Eddie Conti, Claudio Daka, Álvaro Parafita +3

    cs.LGcs.AIarXiv:2609.00090v12026
  13. On the Reliable Detection of Concept Drift from Streaming Unlabeled Data

    Tegjyot Singh Sethi, Mehmed Kantardzic

    stat.MLcs.AIcs.LGarXiv:1704.00023v12017
  14. iFair: Learning Individually Fair Data Representations for Algorithmic Decision Making

    Preethi Lahoti, Krishna P. Gummadi, Gerhard Weikum

    cs.LGcs.IRstat.MLarXiv:1806.01059v22018
  15. Unsupervised Control Through Non-Parametric Discriminative Rewards

    David Warde-Farley, Tom Van de Wiele, Tejas Kulkarni +3

    cs.LGcs.AIstat.MLarXiv:1811.11359v12018
  16. Feynman-Kac Correctors in Diffusion: Annealing, Guidance, and Product of Experts

    Marta Skreta, Tara Akhound-Sadegh, Viktor Ohanesian +6

    cs.LGarXiv:2503.02819v22025
  17. Contrastive learning, multi-view redundancy, and linear models

    Christopher Tosh, Akshay Krishnamurthy, Daniel Hsu

    cs.LGstat.MLarXiv:2008.10150v22020
  18. Lingua Franca or Probing Artifact? Rethinking Latent Language in Multilingual LLMs

    Deniz Bayazit, Badr AlKhamissi, Antoine Bosselut

    cs.CLcs.AIcs.LGarXiv:2609.00155v12026
  19. mmBERT: A Modern Multilingual Encoder with Annealed Language Learning

    Marc Marone, Orion Weller, William Fleshman +3

    cs.CLcs.IRcs.LGarXiv:2509.06888v12025
  20. Structural Temporal Graph Neural Networks for Anomaly Detection in Dynamic Graphs

    Lei Cai, Zhengzhang Chen, Chen Luo +4

    cs.LGcs.SIstat.MLarXiv:2005.07427v22020
  21. PyKEEN 1.0: A Python Library for Training and Evaluating Knowledge Graph Embeddings

    Mehdi Ali, Max Berrendorf, Charles Tapley Hoyt +4

    cs.LGcs.AIstat.MLarXiv:2007.14175v22020
  22. Are Sparse Autoencoders Useful? A Case Study in Sparse Probing

    Subhash Kantamneni, Joshua Engels, Senthooran Rajamanoharan +2

    cs.LGcs.AIarXiv:2502.16681v12025
  23. Dish-TS: A General Paradigm for Alleviating Distribution Shift in Time Series Forecasting

    Wei Fan, Pengyang Wang, Dongkun Wang +3

    cs.LGcs.AIarXiv:2302.14829v32023
  24. Non-square matrix sensing without spurious local minima via the Burer-Monteiro approach

    Dohyung Park, Anastasios Kyrillidis, Constantine Caramanis +1

    stat.MLcs.ITcs.LGarXiv:1609.03240v22016
  25. RL Token: Bootstrapping Online RL with Vision-Language-Action Models

    Charles Xu, Jost Tobias Springenberg, Michael Equi +4

    cs.LGcs.ROarXiv:2604.23073v22026
  26. BACON: Band-limited Coordinate Networks for Multiscale Scene Representation

    David B. Lindell, Dave Van Veen, Jeong Joon Park +1

    cs.CVcs.GRcs.LGarXiv:2112.04645v22021
  27. Sequence Parallelism: Long Sequence Training from System Perspective

    Shenggui Li, Fuzhao Xue, Chaitanya Baranwal +2

    cs.LGcs.DCarXiv:2105.13120v32021
  28. Deformable ProtoPNet: An Interpretable Image Classifier Using Deformable Prototypes

    Jon Donnelly, Alina Jade Barnett, Chaofan Chen

    cs.CVcs.AIcs.LGarXiv:2111.15000v32021
  29. Mutual information is copula entropy

    Jian Ma, Zengqi Sun

    cs.ITcs.LGmath.STarXiv:0808.0845v12008
  30. FedGAN: Federated Generative Adversarial Networks for Distributed Data

    Mohammad Rasouli, Tao Sun, Ram Rajagopal

    cs.LGcs.CVcs.MAarXiv:2006.07228v22020
  31. DiTAR: Diffusion Transformer Autoregressive Modeling for Speech Generation

    Dongya Jia, Zhuo Chen, Jiawei Chen +8

    eess.AScs.AIcs.CLarXiv:2502.03930v42025
  32. AdaCoT: Pareto-Optimal Adaptive Chain-of-Thought Triggering via Reinforcement Learning

    Chenwei Lou, Zewei Sun, Xinnian Liang +6

    cs.LGcs.AIarXiv:2505.11896v22025
  33. PQ-NET: A Generative Part Seq2Seq Network for 3D Shapes

    Rundi Wu, Yixin Zhuang, Kai Xu +2

    cs.CVcs.GRcs.LGarXiv:1911.10949v32019
  34. SafeArena: Evaluating the Safety of Autonomous Web Agents

    Ada Defne Tur, Nicholas Meade, Xing Han Lù +6

    cs.LGcs.AIcs.CLarXiv:2503.04957v12025
  35. Token Assorted: Mixing Latent and Text Tokens for Improved Language Model Reasoning

    DiJia Su, Hanlin Zhu, Yingchen Xu +3

    cs.CLcs.AIcs.LGarXiv:2502.03275v22025
  36. Do RNN and LSTM have Long Memory?

    Jingyu Zhao, Feiqing Huang, Jia Lv +4

    stat.MLcs.LGarXiv:2006.03860v22020
  37. Natural Backdoor Attacks on Speech Recognition Models

    Jinwen Xin, Xixiang Lyu, Jing Ma

    cs.CRcs.LGcs.SDarXiv:2607.15724v12026
  38. Generalisation error in learning with random features and the hidden manifold model

    Federica Gerace, Bruno Loureiro, Florent Krzakala +2

    math.STcs.LGmath.PRarXiv:2002.09339v22020
  39. S*: Test Time Scaling for Code Generation

    Dacheng Li, Shiyi Cao, Chengkun Cao +6

    cs.LGcs.AIarXiv:2502.14382v12025
  40. BeamDojo: Learning Agile Humanoid Locomotion on Sparse Footholds

    Huayi Wang, Zirui Wang, Junli Ren +4

    cs.ROcs.AIcs.LGarXiv:2502.10363v32025
  41. Learning Linear-Quadratic Regulators Efficiently with only $\sqrt{T}$ Regret

    Alon Cohen, Tomer Koren, Yishay Mansour

    cs.LGstat.MLarXiv:1902.06223v22019
  42. Reinforcement Learning for Long-Horizon Interactive LLM Agents

    Kevin Chen, Marco Cusumano-Towner, Brody Huval +4

    cs.LGcs.AIarXiv:2502.01600v32025
  43. Improved Speech Enhancement with the Wave-U-Net

    Craig Macartney, Tillman Weyde

    cs.SDcs.LGcs.NEarXiv:1811.11307v12018
  44. Beyond Reverse KL: Generalizing Direct Preference Optimization with Diverse Divergence Constraints

    Chaoqi Wang, Yibo Jiang, Chenghao Yang +2

    cs.LGcs.AIstat.MLarXiv:2309.16240v12023
  45. DK-GBMKKM: Dynamic Kernel-Space Granular-Ball Multiple Kernel $k$-Means Clustering

    Xiaoyu Lian, Yuchao Zhang, Shuyin Xia +2

    cs.LGarXiv:2609.00647v12026
  46. Sample More to Think Less: Group Filtered Policy Optimization for Concise Reasoning

    Vaishnavi Shrivastava, Ahmed Awadallah, Vidhisha Balachandran +3

    cs.CLcs.LGarXiv:2508.09726v12025
  47. AgentEvolver: Towards Efficient Self-Evolving Agent System

    Yunpeng Zhai, Shuchang Tao, Cheng Chen +10

    cs.LGcs.AIcs.CLarXiv:2511.10395v12025
  48. Deep Neural Network Fingerprinting by Conferrable Adversarial Examples

    Nils Lukas, Yuxuan Zhang, Florian Kerschbaum

    cs.LGcs.CRstat.MLarXiv:1912.00888v42019
  49. A Sober Look at Progress in Language Model Reasoning: Pitfalls and Paths to Reproducibility

    Andreas Hochlehnert, Hardik Bhatnagar, Vishaal Udandarao +3

    cs.LGcs.CLarXiv:2504.07086v22025
  50. Poisson-Gamma Dynamical Systems with Time-varying Transition Dynamics

    Jiahao Wang, Yijun Wang, Nan Fang +1

    cs.LGarXiv:2609.00896v12026
  51. Variational Gaussian Process State-Space Models

    Roger Frigola, Yutian Chen, Carl E. Rasmussen

    cs.LGcs.ROeess.SYarXiv:1406.4905v22014
  52. Multi-Agent Design: Optimizing Agents with Better Prompts and Topologies

    Han Zhou, Xingchen Wan, Ruoxi Sun +5

    cs.LGcs.AIcs.CLarXiv:2502.02533v22025
  53. Verdict Instability of OOD Scores under Reference Resampling

    Donghoon Lee, Shinjin Kang

    cs.LGstat.MLarXiv:2609.00691v12026
  54. Optimization with Non-Differentiable Constraints with Applications to Fairness, Recall, Churn, and Other Goals

    Andrew Cotter, Heinrich Jiang, Serena Wang +4

    cs.LGcs.AIcs.GTarXiv:1809.04198v12018
  55. MLGym: A New Framework and Benchmark for Advancing AI Research Agents

    Deepak Nathani, Lovish Madaan, Nicholas Roberts +14

    cs.CLcs.AIcs.LGarXiv:2502.14499v12025
  56. Tree Search for LLM Agent Reinforcement Learning

    Yuxiang Ji, Ziyu Ma, Yong Wang +3

    cs.LGcs.AIarXiv:2509.21240v32025
  57. Improving Adversarial Transferability via Neuron Attribution-Based Attacks

    Jianping Zhang, Weibin Wu, Jen-tse Huang +4

    cs.LGcs.CRarXiv:2204.00008v12022
  58. The Geometry of Refusal in Large Language Models: Concept Cones and Representational Independence

    Tom Wollschläger, Jannes Elstner, Simon Geisler +3

    cs.LGcs.AIcs.CLarXiv:2502.17420v22025
  59. Memp: Exploring Agent Procedural Memory

    Runnan Fang, Yuan Liang, Xiaobin Wang +6

    cs.CLcs.AIcs.LGarXiv:2508.06433v42025
  60. Diffusion Beats Autoregressive in Data-Constrained Settings

    Mihir Prabhudesai, Mengning Wu, Amir Zadeh +2

    cs.LGcs.AIcs.CVarXiv:2507.15857v72025