Machine Learning

Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

1,021 to 1,080 of 20,192

  1. Optimizing Dialogue Management with Reinforcement Learning: Experiments with the NJFun System

    M. Kearns, D. Litman, S. Singh +1

    cs.LGcs.AIarXiv:1106.0676v12011
  2. Leveraging Automated Unit Tests for Unsupervised Code Translation

    Baptiste Roziere, Jie M. Zhang, Francois Charton +3

    cs.SEcs.CLcs.LGarXiv:2110.06773v22021
  3. Self-supervised Learning is More Robust to Dataset Imbalance

    Hong Liu, Jeff Z. HaoChen, Adrien Gaidon +1

    cs.LGcs.CVstat.MLarXiv:2110.05025v22021
  4. A Review of Physics-based Machine Learning in Civil Engineering

    Shashank Reddy Vadyala, Sai Nethra Betgeri1, John C. Matthews +1

    cs.LGcs.AIarXiv:2110.04600v22021
  5. ReRound: Reconstructive Rounding to Resolve Midpoint Ambiguity in Calibration-Free LLM Quantization

    He-Yen Hsieh, H. T. Kung

    cs.LGcs.CLarXiv:2608.11045v12026
  6. Towards a Unified View of Parameter-Efficient Transfer Learning

    Junxian He, Chunting Zhou, Xuezhe Ma +2

    cs.CLcs.LGarXiv:2110.04366v32021
  7. LLMs Get Lost in Evolving User Intent

    Jihoon Tack, Philippe Laban, Jennifer Neville

    cs.LGarXiv:2607.20734v12026
  8. Traffic Flow Forecasting with Spatial-Temporal Graph Diffusion Network

    Xiyue Zhang, Chao Huang, Yong Xu +5

    cs.LGcs.AIarXiv:2110.04038v12021
  9. Deep Neural Networks and Tabular Data: A Survey

    Vadim Borisov, Tobias Leemann, Kathrin Seßler +3

    cs.LGarXiv:2110.01889v32021
  10. Touchdown: Natural Language Navigation and Spatial Reasoning in Visual Street Environments

    Howard Chen, Alane Suhr, Dipendra Misra +2

    cs.CVcs.AIcs.CLarXiv:1811.12354v72018
  11. An End-to-End Transformer Model for 3D Object Detection

    Ishan Misra, Rohit Girdhar, Armand Joulin

    cs.CVcs.AIcs.LGarXiv:2109.08141v12021
  12. Evolution Fine-Tuning: Learning to Discover Across 371 Optimization Tasks

    Young-Jun Lee, Seungone Kim, Minki Kang +5

    cs.CLcs.LGarXiv:2606.29082v12026
  13. Fused Gromov-Wasserstein distance for structured objects: theoretical foundations and mathematical properties

    Titouan Vayer, Laetita Chapel, Rémi Flamary +2

    stat.MLcs.LGarXiv:1811.02834v12018
  14. Effective and Efficient Graph Learning for Multi-view Clustering

    Quanxue Gao, Wei Xia, Xinbo Gao +3

    cs.LGarXiv:2108.06734v22021
  15. REVES: REvision and VErification--Augmented Training for Test-Time Scaling

    Yuanxin Liu, Ruida Zhou, Xinyan Zhao +6

    cs.LGcs.CLarXiv:2606.18910v12026
  16. Post-hoc Interpretability for Neural NLP: A Survey

    Andreas Madsen, Siva Reddy, Sarath Chandar

    cs.CLcs.LGcs.NEarXiv:2108.04840v52021
  17. Fashionable Modelling with Flux

    Michael Innes, Elliot Saba, Keno Fischer +6

    cs.PLcs.LGarXiv:1811.01457v32018
  18. A Persistent Spatial Semantic Representation for High-level Natural Language Instruction Execution

    Valts Blukis, Chris Paxton, Dieter Fox +2

    cs.ROcs.AIcs.CLarXiv:2107.05612v32021
  19. Meta-learning PINN loss functions

    Apostolos F Psaros, Kenji Kawaguchi, George Em Karniadakis

    cs.LGarXiv:2107.05544v12021
  20. Learning to Delegate for Large-scale Vehicle Routing

    Sirui Li, Zhongxia Yan, Cathy Wu

    cs.LGcs.AIarXiv:2107.04139v22021
  21. SNIP: Single-shot Network Pruning based on Connection Sensitivity

    Namhoon Lee, Thalaiyasingam Ajanthan, Philip H. S. Torr

    cs.CVcs.LGarXiv:1810.02340v22018
  22. Analysis of Diffractive Optical Neural Networks and Their Integration with Electronic Neural Networks

    Deniz Mengu, Yi Luo, Yair Rivenson +1

    cs.NEcs.LGphysics.opticsarXiv:1810.01916v42018
  23. The Road Ahead in Autonomous Driving: The KITScenes Multimodal Dataset

    Richard Schwarzkopf, Fabian Immel, Alexander Blumberg +21

    cs.CVcs.LGcs.ROarXiv:2606.02956v12026
  24. Scenic: A Language for Scenario Specification and Scene Generation

    Daniel J. Fremont, Tommaso Dreossi, Shromona Ghosh +3

    cs.PLcs.CVcs.LGarXiv:1809.09310v22018
  25. Decentralized Instruction Tuning: Conflict-Aware Splitting and Weight Merging

    Minsik Choi, Geewook Kim

    cs.LGarXiv:2606.01717v12026
  26. GANs for Medical Image Analysis

    Salome Kazeminia, Christoph Baur, Arjan Kuijper +4

    cs.CVcs.LGstat.MLarXiv:1809.06222v32018
  27. Don't Use Large Mini-Batches, Use Local SGD

    Tao Lin, Sebastian U. Stich, Kumar Kshitij Patel +1

    cs.LGstat.MLarXiv:1808.07217v62018
    Summaries:한국어
  28. QuAC : Question Answering in Context

    Eunsol Choi, He He, Mohit Iyyer +5

    cs.CLcs.AIcs.LGarXiv:1808.07036v32018
  29. Identifying Implementation Bugs in Machine Learning based Image Classifiers using Metamorphic Testing

    Anurag Dwarakanath, Manish Ahuja, Samarth Sikand +4

    cs.SEcs.LGarXiv:1808.05353v12018
  30. On the limits and opportunities of AI reviewers: Reviewing the reviews of Nature-family papers with 45 expert scientists

    Seungone Kim, Dongkeun Yoon, Kiril Gashteovski +55

    cs.CLcs.AIcs.LGarXiv:2605.20668v12026
  31. Supporting Very Large Models using Automatic Dataflow Graph Partitioning

    Minjie Wang, Chien-chin Huang, Jinyang Li

    cs.DCcs.LGarXiv:1807.08887v22018
  32. CurveBench: A Benchmark for Exact Topological Reasoning over Nested Jordan Curves

    Amirreza Mohseni, Mona Mohammadi, Morteza Saghafian +1

    cs.CVcs.LGarXiv:2605.14068v22026
  33. Visual Domain Adaptation with Manifold Embedded Distribution Alignment

    Jindong Wang, Wenjie Feng, Yiqiang Chen +3

    cs.CVcs.LGarXiv:1807.07258v22018
  34. IndicMedDialog: A Parallel Multi-Turn Medical Dialogue Dataset for Accessible Healthcare in Indic Languages

    Shubham Kumar Nigam, Suparnojit Sarkar, Piyush Patel

    cs.CLcs.AIcs.IRarXiv:2605.13292v12026
  35. From Frequency to Meaning: Vector Space Models of Semantics

    Peter D. Turney, Patrick Pantel

    cs.CLcs.IRcs.LGarXiv:1003.1141v12010
  36. DeepAffinity: Interpretable Deep Learning of Compound-Protein Affinity through Unified Recurrent and Convolutional Neural Networks

    Mostafa Karimi, Di Wu, Zhangyang Wang +1

    q-bio.BMcs.LGstat.MLarXiv:1806.07537v22018
  37. Neural Code Comprehension: A Learnable Representation of Code Semantics

    Tal Ben-Nun, Alice Shoshana Jakobovits, Torsten Hoefler

    cs.LGcs.NEcs.PLarXiv:1806.07336v32018
  38. AI scientists produce results without reasoning scientifically

    Martiño Ríos-García, Nawaf Alampara, Chandan Gupta +5

    cs.AIcond-mat.mtrl-scics.LGarXiv:2604.18805v12026
  39. Unsupervised Alignment of Embeddings with Wasserstein Procrustes

    Edouard Grave, Armand Joulin, Quentin Berthet

    cs.LGcs.CLstat.MLarXiv:1805.11222v12018
  40. SCOPE: Signal-Calibrated On-Policy Distillation Enhancement with Dual-Path Adaptive Weighting

    Binbin Zheng, Xing Ma, Yiheng Liang +6

    cs.LGcs.AIcs.CLarXiv:2604.10688v22026
  41. The EuroCity Persons Dataset: A Novel Benchmark for Object Detection

    Markus Braun, Sebastian Krebs, Fabian Flohr +1

    cs.CVcs.AIcs.LGarXiv:1805.07193v22018
  42. Teaching an Agent to Sketch One Part at a Time

    Xiaodan Du, Ruize Xu, David Yunis +2

    cs.AIcs.CVcs.GRarXiv:2603.19500v22026
  43. MDM-Prime-v2: Binary Encoding and Index Shuffling Enable Scaling of Diffusion Language Models

    Chen-Hao Chao, Wei-Fang Sun, Junwei Quan +2

    cs.LGarXiv:2603.16077v32026
  44. Evaluating Large Language Models Trained on Code

    Mark Chen, Jerry Tworek, Heewoo Jun +55

    cs.LGarXiv:2107.03374v22021
  45. Is Automated Topic Model Evaluation Broken?: The Incoherence of Coherence

    Alexander Hoyle, Pranav Goel, Denis Peskov +3

    cs.CLcs.LGarXiv:2107.02173v32021
  46. MixStyle Neural Networks for Domain Generalization and Adaptation

    Kaiyang Zhou, Yongxin Yang, Yu Qiao +1

    cs.CVcs.AIcs.LGarXiv:2107.02053v22021
  47. Sleeper Agent: Scalable Hidden Trigger Backdoors for Neural Networks Trained from Scratch

    Hossein Souri, Liam Fowl, Rama Chellappa +2

    cs.LGcs.CRcs.CVarXiv:2106.08970v32021
  48. Thinking Like Transformers

    Gail Weiss, Yoav Goldberg, Eran Yahav

    cs.LGcs.CLarXiv:2106.06981v22021
  49. A Variational Perspective on Diffusion-Based Generative Models and Score Matching

    Chin-Wei Huang, Jae Hyun Lim, Aaron Courville

    cs.LGarXiv:2106.02808v22021
  50. Anticipative Video Transformer

    Rohit Girdhar, Kristen Grauman

    cs.CVcs.AIcs.LGarXiv:2106.02036v22021
  51. FedScale: Benchmarking Model and System Performance of Federated Learning at Scale

    Fan Lai, Yinwei Dai, Sanjay S. Singapuram +4

    cs.LGcs.AIcs.DCarXiv:2105.11367v52021
  52. GATES: Self-Distillation under Privileged Context with Consensus Gating

    Alex Stein, Furong Huang, Tom Goldstein

    cs.LGcs.CLarXiv:2602.20574v12026
  53. Intriguing Properties of Vision Transformers

    Muzammal Naseer, Kanchana Ranasinghe, Salman Khan +3

    cs.CVcs.AIcs.LGarXiv:2105.10497v32021
  54. Measuring Coding Challenge Competence With APPS

    Dan Hendrycks, Steven Basart, Saurav Kadavath +8

    cs.SEcs.CLcs.LGarXiv:2105.09938v32021
  55. A tutorial on conformal prediction

    Glenn Shafer, Vladimir Vovk

    cs.LGstat.MLarXiv:0706.3188v12007
  56. Detecting cognitive decline using speech only: The ADReSSo Challenge

    Saturnino Luz, Fasih Haider, Sofia de la Fuente +2

    eess.AScs.CLcs.LGarXiv:2104.09356v12021
  57. FedNLP: Benchmarking Federated Learning Methods for Natural Language Processing Tasks

    Bill Yuchen Lin, Chaoyang He, Zihang Zeng +7

    cs.CLcs.AIcs.LGarXiv:2104.08815v32021
  58. Accurate Failure Prediction in Agents Does Not Imply Effective Failure Prevention

    Rakshith Vasudev, Melisa Russak, Dan Bikel +1

    cs.CLcs.LGarXiv:2602.03338v12026
  59. Sample Complexity of Multi-task Reinforcement Learning

    Emma Brunskill, Lihong Li

    cs.LGstat.MLarXiv:1309.6821v12013
  60. What Characterizes Effective Reasoning? Revisiting Length, Review, and Structure of CoT

    Yunzhen Feng, Julia Kempe, Cheng Zhang +2

    cs.LGarXiv:2509.19284v12025