Machine Learning
Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
5,401 to 5,460 of 20,133
Workload Identification with Physical Side Channels for AI Governance
Simone Gargiulo, Gabriel Kulp
cs.CRcs.AIcs.CYarXiv:2609.00309v12026Agent0: Unleashing Self-Evolving Agents from Zero Data via Tool-Integrated Reasoning
Peng Xia, Kaide Zeng, Jiaqi Liu +5
cs.LGarXiv:2511.16043v12025MoFlow: One-Step Flow Matching for Human Trajectory Forecasting via Implicit Maximum Likelihood Estimation based Distillation
Yuxiang Fu, Qi Yan, Lele Wang +2
cs.CVcs.AIcs.LGarXiv:2503.09950v12025Lost in Simulation: LLM-Simulated Users are Unreliable Proxies for Human Users in Agentic Evaluations
Preethi Seshadri, Samuel Cahyawijaya, Ayomide Odumakinde +2
cs.HCcs.AIcs.CYarXiv:2601.17087v22026The CoT Collection: Improving Zero-shot and Few-shot Learning of Language Models via Chain-of-Thought Fine-Tuning
Seungone Kim, Se June Joo, Doyoung Kim +4
cs.CLcs.AIcs.LGarXiv:2305.14045v22023Modeling Documents with Deep Boltzmann Machines
Nitish Srivastava, Ruslan R Salakhutdinov, Geoffrey E. Hinton
cs.LGcs.IRstat.MLarXiv:1309.6865v12013Explainable Artificial Intelligence (XAI) on TimeSeries Data: A Survey
Thomas Rojat, Raphaël Puget, David Filliat +3
cs.LGcs.AIarXiv:2104.00950v12021Good Memory Has ECC: Evaluating the Memory of Vision-Language Models Beyond Accuracy
Shmuel Berman, Jia Deng
cs.LGcs.AIarXiv:2609.00103v12026SWE-Debate: Competitive Multi-Agent Debate for Software Issue Resolution
Han Li, Yuling Shi, Shaoxin Lin +6
cs.SEcs.CLcs.LGarXiv:2507.23348v12025Motion Tracks: A Unified Representation for Human-Robot Transfer in Few-Shot Imitation Learning
Juntao Ren, Priya Sundaresan, Dorsa Sadigh +2
cs.ROcs.AIcs.LGarXiv:2501.06994v22025Do Language Models Use Their Depth Efficiently?
Róbert Csordás, Christopher D. Manning, Christopher Potts
cs.LGcs.AIcs.NEarXiv:2505.13898v32025Dataset Distillation via Factorization
Songhua Liu, Kai Wang, Xingyi Yang +2
cs.CVcs.LGarXiv:2210.16774v12022The Disparate Effects of Strategic Manipulation
Lily Hu, Nicole Immorlica, Jennifer Wortman Vaughan
cs.LGcs.GTstat.MLarXiv:1808.08646v42018Topology of deep neural networks
Gregory Naitzat, Andrey Zhitnikov, Lek-Heng Lim
cs.LGmath.ATstat.MLarXiv:2004.06093v12020Online Coreset Selection for Rehearsal-based Continual Learning
Jaehong Yoon, Divyam Madaan, Eunho Yang +1
cs.LGcs.CVarXiv:2106.01085v42021UFT: Unifying Supervised and Reinforcement Fine-Tuning
Mingyang Liu, Gabriele Farina, Asuman Ozdaglar
cs.LGcs.CLarXiv:2505.16984v22025Multi-Sensor Prognostics using an Unsupervised Health Index based on LSTM Encoder-Decoder
Pankaj Malhotra, Vishnu TV, Anusha Ramakrishnan +4
cs.LGcs.AIarXiv:1608.06154v12016Deep Nearest Neighbor Anomaly Detection
Liron Bergman, Niv Cohen, Yedid Hoshen
cs.LGcs.CVstat.MLarXiv:2002.10445v12020Artificial Intelligence Should Genuinely Support Clinical Reasoning and Decision Making To Bridge the Translational Gap
Kacper Sokol, James Fackler, Julia E Vogt
cs.HCcs.AIcs.CYarXiv:2506.05030v12025Dual Mixup Regularized Learning for Adversarial Domain Adaptation
Yuan Wu, Diana Inkpen, Ahmed El-Roby
cs.LGcs.CVstat.MLarXiv:2007.03141v22020Explaining in Style: Training a GAN to explain a classifier in StyleSpace
Oran Lang, Yossi Gandelsman, Michal Yarom +8
cs.CVcs.LGcs.NEarXiv:2104.13369v22021Rethinking LLM Unlearning Objectives: A Gradient Perspective and Go Beyond
Qizhou Wang, Jin Peng Zhou, Zhanke Zhou +3
cs.LGarXiv:2502.19301v12025CyCLIP: Cyclic Contrastive Language-Image Pretraining
Shashank Goel, Hritik Bansal, Sumit Bhatia +3
cs.CVcs.LGarXiv:2205.14459v22022Direct Nash Optimization: Teaching Language Models to Self-Improve with General Preferences
Corby Rosset, Ching-An Cheng, Arindam Mitra +3
cs.LGcs.AIcs.CLarXiv:2404.03715v12024Concentrated Differentially Private Gradient Descent with Adaptive per-Iteration Privacy Budget
Jaewoo Lee, Daniel Kifer
cs.LGstat.MLarXiv:1808.09501v12018Gotta Learn Fast: A New Benchmark for Generalization in RL
Alex Nichol, Vicki Pfau, Christopher Hesse +2
cs.LGstat.MLarXiv:1804.03720v22018A Meta-Analysis of the Anomaly Detection Problem
Andrew Emmott, Shubhomoy Das, Thomas Dietterich +2
cs.AIcs.LGstat.MLarXiv:1503.01158v22015If Only We Had Better Counterfactual Explanations: Five Key Deficits to Rectify in the Evaluation of Counterfactual XAI Techniques
Mark T Keane, Eoin M Kenny, Eoin Delaney +1
cs.LGcs.AIarXiv:2103.01035v12021Defeating the Training-Inference Mismatch via FP16
Penghui Qi, Zichen Liu, Xiangxin Zhou +4
cs.LGcs.AIcs.CLarXiv:2510.26788v12025Watch-And-Help: A Challenge for Social Perception and Human-AI Collaboration
Xavier Puig, Tianmin Shu, Shuang Li +5
cs.AIcs.LGcs.MAarXiv:2010.09890v22020Generative Translation Priors: Bayesian Imaging with Cross-Modality Image Translation
Evan Bell, Jiaming Liu, Yifan Chen +1
eess.IVcs.CVcs.LGarXiv:2608.28872v12026Dynamics of stochastic gradient descent for two-layer neural networks in the teacher-student setup
Sebastian Goldt, Madhu S. Advani, Andrew M. Saxe +2
stat.MLcond-mat.dis-nncond-mat.stat-mecharXiv:1906.08632v22019Collaboration Challenges in Building ML-Enabled Systems: Communication, Documentation, Engineering, and Process
Nadia Nahar, Shurui Zhou, Grace Lewis +1
cs.SEcs.LGarXiv:2110.10234v42021Act3D: 3D Feature Field Transformers for Multi-Task Robotic Manipulation
Theophile Gervet, Zhou Xian, Nikolaos Gkanatsios +1
cs.ROcs.AIcs.LGarXiv:2306.17817v22023Towards Agentic Cloud Engineering: Graph and Loop Engineering with a Zero-Trust Agent Harness
Sagar Srinivas Sakhinana, Venkataramana Runkana
cs.SEcs.AIcs.LGarXiv:2609.00050v12026Horizon Reduction Makes RL Scalable
Seohong Park, Kevin Frans, Deepinder Mann +3
cs.LGcs.AIarXiv:2506.04168v32025Incremental Few-Shot Learning with Attention Attractor Networks
Mengye Ren, Renjie Liao, Ethan Fetaya +1
cs.LGcs.CVstat.MLarXiv:1810.07218v32018Analyzing Differentiable Fuzzy Logic Operators
Emile van Krieken, Erman Acar, Frank van Harmelen
cs.AIcs.LGcs.LOarXiv:2002.06100v22020Domain Adaptation with Auxiliary Target Domain-Oriented Classifier
Jian Liang, Dapeng Hu, Jiashi Feng
cs.CVcs.LGarXiv:2007.04171v52020Flex Attention: A Programming Model for Generating Optimized Attention Kernels
Juechu Dong, Boyuan Feng, Driss Guessous +2
cs.LGcs.PFcs.PLarXiv:2412.05496v12024ReMA: Learning to Meta-think for LLMs with Multi-Agent Reinforcement Learning
Ziyu Wan, Yunxiang Li, Xiaoyu Wen +8
cs.AIcs.CLcs.LGarXiv:2503.09501v32025DexMachina: Functional Retargeting for Bimanual Dexterous Manipulation
Zhao Mandi, Yifan Hou, Dieter Fox +3
cs.ROcs.AIcs.LGarXiv:2505.24853v12025Deep Learning for Screening COVID-19 using Chest X-Ray Images
Sanhita Basu, Sushmita Mitra, Nilanjan Saha
eess.IVcs.CVcs.LGarXiv:2004.10507v42020Self-Adapting Language Models
Adam Zweiger, Jyothish Pari, Han Guo +3
cs.LGarXiv:2506.10943v22025UniGeo: Unifying Geometry Logical Reasoning via Reformulating Mathematical Expression
Jiaqi Chen, Tong Li, Jinghui Qin +4
cs.AIcs.LGarXiv:2212.02746v12022Global Sensitivity Analysis with Dependence Measures
Sébastien Da Veiga
math.STcs.LGstat.MLarXiv:1311.2483v12013SWE-Exp: Experience-Driven Software Issue Resolution
Silin Chen, Shaoxin Lin, Yuling Shi +8
cs.SEcs.CLcs.LGarXiv:2507.23361v22025ResMimic: From General Motion Tracking to Humanoid Whole-body Loco-Manipulation via Residual Learning
Siheng Zhao, Yanjie Ze, Yue Wang +4
cs.ROcs.LGarXiv:2510.05070v22025Flexible Isosurface Extraction for Gradient-Based Mesh Optimization
Tianchang Shen, Jacob Munkberg, Jon Hasselgren +7
cs.GRcs.CVcs.LGarXiv:2308.05371v12023MagicMotion: Controllable Video Generation with Dense-to-Sparse Trajectory Guidance
Quanhao Li, Zhen Xing, Rui Wang +3
cs.CVcs.AIcs.LGarXiv:2503.16421v32025ServerlessLLM: Low-Latency Serverless Inference for Large Language Models
Yao Fu, Leyang Xue, Yeqi Huang +4
cs.LGcs.DCarXiv:2401.14351v22024How To Make the Gradients Small Stochastically: Even Faster Convex and Nonconvex SGD
Zeyuan Allen-Zhu
cs.LGcs.DSmath.OCarXiv:1801.02982v32018Robustly Disentangled Causal Mechanisms: Validating Deep Representations for Interventional Robustness
Raphael Suter, Đorđe Miladinović, Bernhard Schölkopf +1
stat.MLcs.LGarXiv:1811.00007v22018BREAD: Branched Rollouts from Expert Anchors Bridge SFT & RL for Reasoning
Xuechen Zhang, Zijian Huang, Yingcong Li +3
cs.LGarXiv:2506.17211v12025In-context Autoencoder for Context Compression in a Large Language Model
Tao Ge, Jing Hu, Lei Wang +3
cs.CLcs.AIcs.LGarXiv:2307.06945v42023Ithemal: Accurate, Portable and Fast Basic Block Throughput Estimation using Deep Neural Networks
Charith Mendis, Alex Renda, Saman Amarasinghe +1
cs.DCcs.LGstat.MLarXiv:1808.07412v22018Segment Policy Optimization: Effective Segment-Level Credit Assignment in RL for Large Language Models
Yiran Guo, Lijie Xu, Jie Liu +2
cs.LGcs.AIcs.CLarXiv:2505.23564v22025Neural Network Matrix Factorization
Gintare Karolina Dziugaite, Daniel M. Roy
cs.LGstat.MLarXiv:1511.06443v22015Automatic Anomaly Detection in the Cloud Via Statistical Learning
Jordan Hochenbaum, Owen S. Vallis, Arun Kejariwal
cs.LGarXiv:1704.07706v12017Real-Time Decision-Making for Digital Twin in Additive Manufacturing with Model Predictive Control using Time-Series Deep Neural Networks
Yi-Ping Chen, Vispi Karkaria, Ying-Kuan Tsai +5
cs.LGcs.AIeess.SYarXiv:2501.07601v52025