Machine Learning

Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

5,401 to 5,460 of 20,133

  1. Workload Identification with Physical Side Channels for AI Governance

    Simone Gargiulo, Gabriel Kulp

    cs.CRcs.AIcs.CYarXiv:2609.00309v12026
  2. Agent0: Unleashing Self-Evolving Agents from Zero Data via Tool-Integrated Reasoning

    Peng Xia, Kaide Zeng, Jiaqi Liu +5

    cs.LGarXiv:2511.16043v12025
  3. MoFlow: One-Step Flow Matching for Human Trajectory Forecasting via Implicit Maximum Likelihood Estimation based Distillation

    Yuxiang Fu, Qi Yan, Lele Wang +2

    cs.CVcs.AIcs.LGarXiv:2503.09950v12025
  4. Lost in Simulation: LLM-Simulated Users are Unreliable Proxies for Human Users in Agentic Evaluations

    Preethi Seshadri, Samuel Cahyawijaya, Ayomide Odumakinde +2

    cs.HCcs.AIcs.CYarXiv:2601.17087v22026
  5. The CoT Collection: Improving Zero-shot and Few-shot Learning of Language Models via Chain-of-Thought Fine-Tuning

    Seungone Kim, Se June Joo, Doyoung Kim +4

    cs.CLcs.AIcs.LGarXiv:2305.14045v22023
  6. Modeling Documents with Deep Boltzmann Machines

    Nitish Srivastava, Ruslan R Salakhutdinov, Geoffrey E. Hinton

    cs.LGcs.IRstat.MLarXiv:1309.6865v12013
  7. Explainable Artificial Intelligence (XAI) on TimeSeries Data: A Survey

    Thomas Rojat, Raphaël Puget, David Filliat +3

    cs.LGcs.AIarXiv:2104.00950v12021
  8. Good Memory Has ECC: Evaluating the Memory of Vision-Language Models Beyond Accuracy

    Shmuel Berman, Jia Deng

    cs.LGcs.AIarXiv:2609.00103v12026
  9. SWE-Debate: Competitive Multi-Agent Debate for Software Issue Resolution

    Han Li, Yuling Shi, Shaoxin Lin +6

    cs.SEcs.CLcs.LGarXiv:2507.23348v12025
  10. Motion Tracks: A Unified Representation for Human-Robot Transfer in Few-Shot Imitation Learning

    Juntao Ren, Priya Sundaresan, Dorsa Sadigh +2

    cs.ROcs.AIcs.LGarXiv:2501.06994v22025
  11. Do Language Models Use Their Depth Efficiently?

    Róbert Csordás, Christopher D. Manning, Christopher Potts

    cs.LGcs.AIcs.NEarXiv:2505.13898v32025
  12. Dataset Distillation via Factorization

    Songhua Liu, Kai Wang, Xingyi Yang +2

    cs.CVcs.LGarXiv:2210.16774v12022
  13. The Disparate Effects of Strategic Manipulation

    Lily Hu, Nicole Immorlica, Jennifer Wortman Vaughan

    cs.LGcs.GTstat.MLarXiv:1808.08646v42018
  14. Topology of deep neural networks

    Gregory Naitzat, Andrey Zhitnikov, Lek-Heng Lim

    cs.LGmath.ATstat.MLarXiv:2004.06093v12020
  15. Online Coreset Selection for Rehearsal-based Continual Learning

    Jaehong Yoon, Divyam Madaan, Eunho Yang +1

    cs.LGcs.CVarXiv:2106.01085v42021
  16. UFT: Unifying Supervised and Reinforcement Fine-Tuning

    Mingyang Liu, Gabriele Farina, Asuman Ozdaglar

    cs.LGcs.CLarXiv:2505.16984v22025
  17. Multi-Sensor Prognostics using an Unsupervised Health Index based on LSTM Encoder-Decoder

    Pankaj Malhotra, Vishnu TV, Anusha Ramakrishnan +4

    cs.LGcs.AIarXiv:1608.06154v12016
  18. Deep Nearest Neighbor Anomaly Detection

    Liron Bergman, Niv Cohen, Yedid Hoshen

    cs.LGcs.CVstat.MLarXiv:2002.10445v12020
  19. Artificial Intelligence Should Genuinely Support Clinical Reasoning and Decision Making To Bridge the Translational Gap

    Kacper Sokol, James Fackler, Julia E Vogt

    cs.HCcs.AIcs.CYarXiv:2506.05030v12025
  20. Dual Mixup Regularized Learning for Adversarial Domain Adaptation

    Yuan Wu, Diana Inkpen, Ahmed El-Roby

    cs.LGcs.CVstat.MLarXiv:2007.03141v22020
  21. Explaining in Style: Training a GAN to explain a classifier in StyleSpace

    Oran Lang, Yossi Gandelsman, Michal Yarom +8

    cs.CVcs.LGcs.NEarXiv:2104.13369v22021
  22. Rethinking LLM Unlearning Objectives: A Gradient Perspective and Go Beyond

    Qizhou Wang, Jin Peng Zhou, Zhanke Zhou +3

    cs.LGarXiv:2502.19301v12025
  23. CyCLIP: Cyclic Contrastive Language-Image Pretraining

    Shashank Goel, Hritik Bansal, Sumit Bhatia +3

    cs.CVcs.LGarXiv:2205.14459v22022
  24. Direct Nash Optimization: Teaching Language Models to Self-Improve with General Preferences

    Corby Rosset, Ching-An Cheng, Arindam Mitra +3

    cs.LGcs.AIcs.CLarXiv:2404.03715v12024
  25. Concentrated Differentially Private Gradient Descent with Adaptive per-Iteration Privacy Budget

    Jaewoo Lee, Daniel Kifer

    cs.LGstat.MLarXiv:1808.09501v12018
  26. Gotta Learn Fast: A New Benchmark for Generalization in RL

    Alex Nichol, Vicki Pfau, Christopher Hesse +2

    cs.LGstat.MLarXiv:1804.03720v22018
  27. A Meta-Analysis of the Anomaly Detection Problem

    Andrew Emmott, Shubhomoy Das, Thomas Dietterich +2

    cs.AIcs.LGstat.MLarXiv:1503.01158v22015
  28. If Only We Had Better Counterfactual Explanations: Five Key Deficits to Rectify in the Evaluation of Counterfactual XAI Techniques

    Mark T Keane, Eoin M Kenny, Eoin Delaney +1

    cs.LGcs.AIarXiv:2103.01035v12021
  29. Defeating the Training-Inference Mismatch via FP16

    Penghui Qi, Zichen Liu, Xiangxin Zhou +4

    cs.LGcs.AIcs.CLarXiv:2510.26788v12025
  30. Watch-And-Help: A Challenge for Social Perception and Human-AI Collaboration

    Xavier Puig, Tianmin Shu, Shuang Li +5

    cs.AIcs.LGcs.MAarXiv:2010.09890v22020
  31. Generative Translation Priors: Bayesian Imaging with Cross-Modality Image Translation

    Evan Bell, Jiaming Liu, Yifan Chen +1

    eess.IVcs.CVcs.LGarXiv:2608.28872v12026
  32. Dynamics of stochastic gradient descent for two-layer neural networks in the teacher-student setup

    Sebastian Goldt, Madhu S. Advani, Andrew M. Saxe +2

    stat.MLcond-mat.dis-nncond-mat.stat-mecharXiv:1906.08632v22019
  33. Collaboration Challenges in Building ML-Enabled Systems: Communication, Documentation, Engineering, and Process

    Nadia Nahar, Shurui Zhou, Grace Lewis +1

    cs.SEcs.LGarXiv:2110.10234v42021
  34. Act3D: 3D Feature Field Transformers for Multi-Task Robotic Manipulation

    Theophile Gervet, Zhou Xian, Nikolaos Gkanatsios +1

    cs.ROcs.AIcs.LGarXiv:2306.17817v22023
  35. Towards Agentic Cloud Engineering: Graph and Loop Engineering with a Zero-Trust Agent Harness

    Sagar Srinivas Sakhinana, Venkataramana Runkana

    cs.SEcs.AIcs.LGarXiv:2609.00050v12026
  36. Horizon Reduction Makes RL Scalable

    Seohong Park, Kevin Frans, Deepinder Mann +3

    cs.LGcs.AIarXiv:2506.04168v32025
  37. Incremental Few-Shot Learning with Attention Attractor Networks

    Mengye Ren, Renjie Liao, Ethan Fetaya +1

    cs.LGcs.CVstat.MLarXiv:1810.07218v32018
  38. Analyzing Differentiable Fuzzy Logic Operators

    Emile van Krieken, Erman Acar, Frank van Harmelen

    cs.AIcs.LGcs.LOarXiv:2002.06100v22020
  39. Domain Adaptation with Auxiliary Target Domain-Oriented Classifier

    Jian Liang, Dapeng Hu, Jiashi Feng

    cs.CVcs.LGarXiv:2007.04171v52020
  40. Flex Attention: A Programming Model for Generating Optimized Attention Kernels

    Juechu Dong, Boyuan Feng, Driss Guessous +2

    cs.LGcs.PFcs.PLarXiv:2412.05496v12024
  41. ReMA: Learning to Meta-think for LLMs with Multi-Agent Reinforcement Learning

    Ziyu Wan, Yunxiang Li, Xiaoyu Wen +8

    cs.AIcs.CLcs.LGarXiv:2503.09501v32025
  42. DexMachina: Functional Retargeting for Bimanual Dexterous Manipulation

    Zhao Mandi, Yifan Hou, Dieter Fox +3

    cs.ROcs.AIcs.LGarXiv:2505.24853v12025
  43. Deep Learning for Screening COVID-19 using Chest X-Ray Images

    Sanhita Basu, Sushmita Mitra, Nilanjan Saha

    eess.IVcs.CVcs.LGarXiv:2004.10507v42020
  44. Self-Adapting Language Models

    Adam Zweiger, Jyothish Pari, Han Guo +3

    cs.LGarXiv:2506.10943v22025
  45. UniGeo: Unifying Geometry Logical Reasoning via Reformulating Mathematical Expression

    Jiaqi Chen, Tong Li, Jinghui Qin +4

    cs.AIcs.LGarXiv:2212.02746v12022
  46. Global Sensitivity Analysis with Dependence Measures

    Sébastien Da Veiga

    math.STcs.LGstat.MLarXiv:1311.2483v12013
  47. SWE-Exp: Experience-Driven Software Issue Resolution

    Silin Chen, Shaoxin Lin, Yuling Shi +8

    cs.SEcs.CLcs.LGarXiv:2507.23361v22025
  48. ResMimic: From General Motion Tracking to Humanoid Whole-body Loco-Manipulation via Residual Learning

    Siheng Zhao, Yanjie Ze, Yue Wang +4

    cs.ROcs.LGarXiv:2510.05070v22025
  49. Flexible Isosurface Extraction for Gradient-Based Mesh Optimization

    Tianchang Shen, Jacob Munkberg, Jon Hasselgren +7

    cs.GRcs.CVcs.LGarXiv:2308.05371v12023
  50. MagicMotion: Controllable Video Generation with Dense-to-Sparse Trajectory Guidance

    Quanhao Li, Zhen Xing, Rui Wang +3

    cs.CVcs.AIcs.LGarXiv:2503.16421v32025
  51. ServerlessLLM: Low-Latency Serverless Inference for Large Language Models

    Yao Fu, Leyang Xue, Yeqi Huang +4

    cs.LGcs.DCarXiv:2401.14351v22024
  52. How To Make the Gradients Small Stochastically: Even Faster Convex and Nonconvex SGD

    Zeyuan Allen-Zhu

    cs.LGcs.DSmath.OCarXiv:1801.02982v32018
  53. Robustly Disentangled Causal Mechanisms: Validating Deep Representations for Interventional Robustness

    Raphael Suter, Đorđe Miladinović, Bernhard Schölkopf +1

    stat.MLcs.LGarXiv:1811.00007v22018
  54. BREAD: Branched Rollouts from Expert Anchors Bridge SFT & RL for Reasoning

    Xuechen Zhang, Zijian Huang, Yingcong Li +3

    cs.LGarXiv:2506.17211v12025
  55. In-context Autoencoder for Context Compression in a Large Language Model

    Tao Ge, Jing Hu, Lei Wang +3

    cs.CLcs.AIcs.LGarXiv:2307.06945v42023
  56. Ithemal: Accurate, Portable and Fast Basic Block Throughput Estimation using Deep Neural Networks

    Charith Mendis, Alex Renda, Saman Amarasinghe +1

    cs.DCcs.LGstat.MLarXiv:1808.07412v22018
  57. Segment Policy Optimization: Effective Segment-Level Credit Assignment in RL for Large Language Models

    Yiran Guo, Lijie Xu, Jie Liu +2

    cs.LGcs.AIcs.CLarXiv:2505.23564v22025
  58. Neural Network Matrix Factorization

    Gintare Karolina Dziugaite, Daniel M. Roy

    cs.LGstat.MLarXiv:1511.06443v22015
  59. Automatic Anomaly Detection in the Cloud Via Statistical Learning

    Jordan Hochenbaum, Owen S. Vallis, Arun Kejariwal

    cs.LGarXiv:1704.07706v12017
  60. Real-Time Decision-Making for Digital Twin in Additive Manufacturing with Model Predictive Control using Time-Series Deep Neural Networks

    Yi-Ping Chen, Vispi Karkaria, Ying-Kuan Tsai +5

    cs.LGcs.AIeess.SYarXiv:2501.07601v52025