Machine Learning
Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
1,021 to 1,080 of 20,192
Optimizing Dialogue Management with Reinforcement Learning: Experiments with the NJFun System
M. Kearns, D. Litman, S. Singh +1
cs.LGcs.AIarXiv:1106.0676v12011Leveraging Automated Unit Tests for Unsupervised Code Translation
Baptiste Roziere, Jie M. Zhang, Francois Charton +3
cs.SEcs.CLcs.LGarXiv:2110.06773v22021Self-supervised Learning is More Robust to Dataset Imbalance
Hong Liu, Jeff Z. HaoChen, Adrien Gaidon +1
cs.LGcs.CVstat.MLarXiv:2110.05025v22021A Review of Physics-based Machine Learning in Civil Engineering
Shashank Reddy Vadyala, Sai Nethra Betgeri1, John C. Matthews +1
cs.LGcs.AIarXiv:2110.04600v22021ReRound: Reconstructive Rounding to Resolve Midpoint Ambiguity in Calibration-Free LLM Quantization
He-Yen Hsieh, H. T. Kung
cs.LGcs.CLarXiv:2608.11045v12026Towards a Unified View of Parameter-Efficient Transfer Learning
Junxian He, Chunting Zhou, Xuezhe Ma +2
cs.CLcs.LGarXiv:2110.04366v32021LLMs Get Lost in Evolving User Intent
Jihoon Tack, Philippe Laban, Jennifer Neville
cs.LGarXiv:2607.20734v12026Traffic Flow Forecasting with Spatial-Temporal Graph Diffusion Network
Xiyue Zhang, Chao Huang, Yong Xu +5
cs.LGcs.AIarXiv:2110.04038v12021Deep Neural Networks and Tabular Data: A Survey
Vadim Borisov, Tobias Leemann, Kathrin Seßler +3
cs.LGarXiv:2110.01889v32021Touchdown: Natural Language Navigation and Spatial Reasoning in Visual Street Environments
Howard Chen, Alane Suhr, Dipendra Misra +2
cs.CVcs.AIcs.CLarXiv:1811.12354v72018An End-to-End Transformer Model for 3D Object Detection
Ishan Misra, Rohit Girdhar, Armand Joulin
cs.CVcs.AIcs.LGarXiv:2109.08141v12021Evolution Fine-Tuning: Learning to Discover Across 371 Optimization Tasks
Young-Jun Lee, Seungone Kim, Minki Kang +5
cs.CLcs.LGarXiv:2606.29082v12026Fused Gromov-Wasserstein distance for structured objects: theoretical foundations and mathematical properties
Titouan Vayer, Laetita Chapel, Rémi Flamary +2
stat.MLcs.LGarXiv:1811.02834v12018Effective and Efficient Graph Learning for Multi-view Clustering
Quanxue Gao, Wei Xia, Xinbo Gao +3
cs.LGarXiv:2108.06734v22021REVES: REvision and VErification--Augmented Training for Test-Time Scaling
Yuanxin Liu, Ruida Zhou, Xinyan Zhao +6
cs.LGcs.CLarXiv:2606.18910v12026Post-hoc Interpretability for Neural NLP: A Survey
Andreas Madsen, Siva Reddy, Sarath Chandar
cs.CLcs.LGcs.NEarXiv:2108.04840v52021Fashionable Modelling with Flux
Michael Innes, Elliot Saba, Keno Fischer +6
cs.PLcs.LGarXiv:1811.01457v32018A Persistent Spatial Semantic Representation for High-level Natural Language Instruction Execution
Valts Blukis, Chris Paxton, Dieter Fox +2
cs.ROcs.AIcs.CLarXiv:2107.05612v32021Meta-learning PINN loss functions
Apostolos F Psaros, Kenji Kawaguchi, George Em Karniadakis
cs.LGarXiv:2107.05544v12021Learning to Delegate for Large-scale Vehicle Routing
Sirui Li, Zhongxia Yan, Cathy Wu
cs.LGcs.AIarXiv:2107.04139v22021SNIP: Single-shot Network Pruning based on Connection Sensitivity
Namhoon Lee, Thalaiyasingam Ajanthan, Philip H. S. Torr
cs.CVcs.LGarXiv:1810.02340v22018Analysis of Diffractive Optical Neural Networks and Their Integration with Electronic Neural Networks
Deniz Mengu, Yi Luo, Yair Rivenson +1
cs.NEcs.LGphysics.opticsarXiv:1810.01916v42018The Road Ahead in Autonomous Driving: The KITScenes Multimodal Dataset
Richard Schwarzkopf, Fabian Immel, Alexander Blumberg +21
cs.CVcs.LGcs.ROarXiv:2606.02956v12026Scenic: A Language for Scenario Specification and Scene Generation
Daniel J. Fremont, Tommaso Dreossi, Shromona Ghosh +3
cs.PLcs.CVcs.LGarXiv:1809.09310v22018Decentralized Instruction Tuning: Conflict-Aware Splitting and Weight Merging
Minsik Choi, Geewook Kim
cs.LGarXiv:2606.01717v12026GANs for Medical Image Analysis
Salome Kazeminia, Christoph Baur, Arjan Kuijper +4
cs.CVcs.LGstat.MLarXiv:1809.06222v32018Don't Use Large Mini-Batches, Use Local SGD
Tao Lin, Sebastian U. Stich, Kumar Kshitij Patel +1
cs.LGstat.MLarXiv:1808.07217v62018Summaries:한국어QuAC : Question Answering in Context
Eunsol Choi, He He, Mohit Iyyer +5
cs.CLcs.AIcs.LGarXiv:1808.07036v32018Identifying Implementation Bugs in Machine Learning based Image Classifiers using Metamorphic Testing
Anurag Dwarakanath, Manish Ahuja, Samarth Sikand +4
cs.SEcs.LGarXiv:1808.05353v12018On the limits and opportunities of AI reviewers: Reviewing the reviews of Nature-family papers with 45 expert scientists
Seungone Kim, Dongkeun Yoon, Kiril Gashteovski +55
cs.CLcs.AIcs.LGarXiv:2605.20668v12026Supporting Very Large Models using Automatic Dataflow Graph Partitioning
Minjie Wang, Chien-chin Huang, Jinyang Li
cs.DCcs.LGarXiv:1807.08887v22018CurveBench: A Benchmark for Exact Topological Reasoning over Nested Jordan Curves
Amirreza Mohseni, Mona Mohammadi, Morteza Saghafian +1
cs.CVcs.LGarXiv:2605.14068v22026Visual Domain Adaptation with Manifold Embedded Distribution Alignment
Jindong Wang, Wenjie Feng, Yiqiang Chen +3
cs.CVcs.LGarXiv:1807.07258v22018IndicMedDialog: A Parallel Multi-Turn Medical Dialogue Dataset for Accessible Healthcare in Indic Languages
Shubham Kumar Nigam, Suparnojit Sarkar, Piyush Patel
cs.CLcs.AIcs.IRarXiv:2605.13292v12026From Frequency to Meaning: Vector Space Models of Semantics
Peter D. Turney, Patrick Pantel
cs.CLcs.IRcs.LGarXiv:1003.1141v12010DeepAffinity: Interpretable Deep Learning of Compound-Protein Affinity through Unified Recurrent and Convolutional Neural Networks
Mostafa Karimi, Di Wu, Zhangyang Wang +1
q-bio.BMcs.LGstat.MLarXiv:1806.07537v22018Neural Code Comprehension: A Learnable Representation of Code Semantics
Tal Ben-Nun, Alice Shoshana Jakobovits, Torsten Hoefler
cs.LGcs.NEcs.PLarXiv:1806.07336v32018AI scientists produce results without reasoning scientifically
Martiño Ríos-García, Nawaf Alampara, Chandan Gupta +5
cs.AIcond-mat.mtrl-scics.LGarXiv:2604.18805v12026Unsupervised Alignment of Embeddings with Wasserstein Procrustes
Edouard Grave, Armand Joulin, Quentin Berthet
cs.LGcs.CLstat.MLarXiv:1805.11222v12018SCOPE: Signal-Calibrated On-Policy Distillation Enhancement with Dual-Path Adaptive Weighting
Binbin Zheng, Xing Ma, Yiheng Liang +6
cs.LGcs.AIcs.CLarXiv:2604.10688v22026The EuroCity Persons Dataset: A Novel Benchmark for Object Detection
Markus Braun, Sebastian Krebs, Fabian Flohr +1
cs.CVcs.AIcs.LGarXiv:1805.07193v22018Teaching an Agent to Sketch One Part at a Time
Xiaodan Du, Ruize Xu, David Yunis +2
cs.AIcs.CVcs.GRarXiv:2603.19500v22026MDM-Prime-v2: Binary Encoding and Index Shuffling Enable Scaling of Diffusion Language Models
Chen-Hao Chao, Wei-Fang Sun, Junwei Quan +2
cs.LGarXiv:2603.16077v32026Evaluating Large Language Models Trained on Code
Mark Chen, Jerry Tworek, Heewoo Jun +55
cs.LGarXiv:2107.03374v22021Is Automated Topic Model Evaluation Broken?: The Incoherence of Coherence
Alexander Hoyle, Pranav Goel, Denis Peskov +3
cs.CLcs.LGarXiv:2107.02173v32021MixStyle Neural Networks for Domain Generalization and Adaptation
Kaiyang Zhou, Yongxin Yang, Yu Qiao +1
cs.CVcs.AIcs.LGarXiv:2107.02053v22021Sleeper Agent: Scalable Hidden Trigger Backdoors for Neural Networks Trained from Scratch
Hossein Souri, Liam Fowl, Rama Chellappa +2
cs.LGcs.CRcs.CVarXiv:2106.08970v32021Thinking Like Transformers
Gail Weiss, Yoav Goldberg, Eran Yahav
cs.LGcs.CLarXiv:2106.06981v22021A Variational Perspective on Diffusion-Based Generative Models and Score Matching
Chin-Wei Huang, Jae Hyun Lim, Aaron Courville
cs.LGarXiv:2106.02808v22021Anticipative Video Transformer
Rohit Girdhar, Kristen Grauman
cs.CVcs.AIcs.LGarXiv:2106.02036v22021FedScale: Benchmarking Model and System Performance of Federated Learning at Scale
Fan Lai, Yinwei Dai, Sanjay S. Singapuram +4
cs.LGcs.AIcs.DCarXiv:2105.11367v52021GATES: Self-Distillation under Privileged Context with Consensus Gating
Alex Stein, Furong Huang, Tom Goldstein
cs.LGcs.CLarXiv:2602.20574v12026Intriguing Properties of Vision Transformers
Muzammal Naseer, Kanchana Ranasinghe, Salman Khan +3
cs.CVcs.AIcs.LGarXiv:2105.10497v32021Measuring Coding Challenge Competence With APPS
Dan Hendrycks, Steven Basart, Saurav Kadavath +8
cs.SEcs.CLcs.LGarXiv:2105.09938v32021A tutorial on conformal prediction
Glenn Shafer, Vladimir Vovk
cs.LGstat.MLarXiv:0706.3188v12007Detecting cognitive decline using speech only: The ADReSSo Challenge
Saturnino Luz, Fasih Haider, Sofia de la Fuente +2
eess.AScs.CLcs.LGarXiv:2104.09356v12021FedNLP: Benchmarking Federated Learning Methods for Natural Language Processing Tasks
Bill Yuchen Lin, Chaoyang He, Zihang Zeng +7
cs.CLcs.AIcs.LGarXiv:2104.08815v32021Accurate Failure Prediction in Agents Does Not Imply Effective Failure Prevention
Rakshith Vasudev, Melisa Russak, Dan Bikel +1
cs.CLcs.LGarXiv:2602.03338v12026Sample Complexity of Multi-task Reinforcement Learning
Emma Brunskill, Lihong Li
cs.LGstat.MLarXiv:1309.6821v12013What Characterizes Effective Reasoning? Revisiting Length, Review, and Structure of CoT
Yunzhen Feng, Julia Kempe, Cheng Zhang +2
cs.LGarXiv:2509.19284v12025