Machine Learning

Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

961 to 1,020 of 20,151

  1. SGHA: Evidence-Grounded Research Problem Discovery with Local Language Models

    Sarvesh Gharat, Junpei Komiyama

    cs.AIcs.LGarXiv:2608.17501v12026
  2. Competition-Level Code Generation with AlphaCode

    Yujia Li, David Choi, Junyoung Chung +23

    cs.PLcs.AIcs.LGarXiv:2203.07814v12022
  3. Pathology Transport: Optimal-Transport Explanations for Clinical Data, and When Their Heatmaps (Fail to) Localize Disease

    Lalit Kumar

    cs.LGarXiv:2608.17370v12026
  4. Combining Physically-Based Modeling and Deep Learning for Fusing GRACE Satellite Data: Can We Learn from Mismatch?

    Alexander Y. Sun, Bridget R. Scanlon, Zizhan Zhang +4

    physics.geo-phcs.LGstat.MLarXiv:1902.01933v12019
  5. Global Simulation-Guided Dynamic Operator Scheduling for Efficient Multi-Tenant Model Serving

    Weinan Liu, Zeyuan Ding, Dian Ding +5

    cs.OScs.DCcs.LGarXiv:2608.15762v12026
  6. ABAW: Valence-Arousal Estimation, Expression Recognition, Action Unit Detection & Multi-Task Learning Challenges

    Dimitrios Kollias

    cs.CVcs.LGarXiv:2202.10659v22022
  7. Handling Distribution Shifts on Graphs: An Invariance Perspective

    Qitian Wu, Hengrui Zhang, Junchi Yan +1

    cs.LGcs.AIarXiv:2202.02466v52022
  8. Learning agile and dynamic motor skills for legged robots

    Jemin Hwangbo, Joonho Lee, Alexey Dosovitskiy +4

    cs.ROcs.LGstat.MLarXiv:1901.08652v12019
  9. Sparse Subspace Clustering: Algorithm, Theory, and Applications

    Ehsan Elhamifar, Rene Vidal

    cs.CVcs.IRcs.ITarXiv:1203.1005v32012
  10. Generalized Fisher Score for Feature Selection

    Quanquan Gu, Zhenhui Li, Jiawei Han

    cs.LGstat.MLarXiv:1202.3725v12012
  11. Practical Lossless Compression with Latent Variables using Bits Back Coding

    James Townsend, Tom Bird, David Barber

    cs.LGcs.AIcs.ITarXiv:1901.04866v12019
  12. FastGRNN: A Fast, Accurate, Stable and Tiny Kilobyte Sized Gated Recurrent Neural Network

    Aditya Kusupati, Manish Singh, Kush Bhatia +3

    cs.LGcs.AIcs.NEarXiv:1901.02358v12019
  13. Patches Are All You Need?

    Asher Trockman, J. Zico Kolter

    cs.CVcs.AIcs.LGarXiv:2201.09792v12022
  14. Language Models as Zero-Shot Planners: Extracting Actionable Knowledge for Embodied Agents

    Wenlong Huang, Pieter Abbeel, Deepak Pathak +1

    cs.LGcs.AIcs.CLarXiv:2201.07207v22022
  15. Graph Neural Networks: A Review of Methods and Applications

    Jie Zhou, Ganqu Cui, Shengding Hu +6

    cs.LGcs.AIstat.MLarXiv:1812.08434v62018
  16. Statistical Topic Models for Multi-Label Document Classification

    Timothy N. Rubin, America Chambers, Padhraic Smyth +1

    stat.MLcs.LGarXiv:1107.2462v22011
  17. Synthetic Persona Pretraining: Alignment from Token Zero

    Julian Minder, Viktor Moskvoretskii, Raghav Singhal +12

    cs.LGcs.AIcs.CLarXiv:2608.13482v12026
  18. A fast accurate fine-grain object detection model based on YOLOv4 deep neural network

    Arunabha M. Roy, Rikhi Bose, Jayabrata Bhaduri

    cs.CVcs.LGarXiv:2111.00298v12021
  19. Improving Robustness using Generated Data

    Sven Gowal, Sylvestre-Alvise Rebuffi, Olivia Wiles +3

    cs.LGcs.CVstat.MLarXiv:2110.09468v22021
  20. Optimizing Dialogue Management with Reinforcement Learning: Experiments with the NJFun System

    M. Kearns, D. Litman, S. Singh +1

    cs.LGcs.AIarXiv:1106.0676v12011
  21. Leveraging Automated Unit Tests for Unsupervised Code Translation

    Baptiste Roziere, Jie M. Zhang, Francois Charton +3

    cs.SEcs.CLcs.LGarXiv:2110.06773v22021
  22. Self-supervised Learning is More Robust to Dataset Imbalance

    Hong Liu, Jeff Z. HaoChen, Adrien Gaidon +1

    cs.LGcs.CVstat.MLarXiv:2110.05025v22021
  23. A Review of Physics-based Machine Learning in Civil Engineering

    Shashank Reddy Vadyala, Sai Nethra Betgeri1, John C. Matthews +1

    cs.LGcs.AIarXiv:2110.04600v22021
  24. ReRound: Reconstructive Rounding to Resolve Midpoint Ambiguity in Calibration-Free LLM Quantization

    He-Yen Hsieh, H. T. Kung

    cs.LGcs.CLarXiv:2608.11045v12026
  25. Towards a Unified View of Parameter-Efficient Transfer Learning

    Junxian He, Chunting Zhou, Xuezhe Ma +2

    cs.CLcs.LGarXiv:2110.04366v32021
  26. LLMs Get Lost in Evolving User Intent

    Jihoon Tack, Philippe Laban, Jennifer Neville

    cs.LGarXiv:2607.20734v12026
  27. Traffic Flow Forecasting with Spatial-Temporal Graph Diffusion Network

    Xiyue Zhang, Chao Huang, Yong Xu +5

    cs.LGcs.AIarXiv:2110.04038v12021
  28. Deep Neural Networks and Tabular Data: A Survey

    Vadim Borisov, Tobias Leemann, Kathrin Seßler +3

    cs.LGarXiv:2110.01889v32021
  29. Touchdown: Natural Language Navigation and Spatial Reasoning in Visual Street Environments

    Howard Chen, Alane Suhr, Dipendra Misra +2

    cs.CVcs.AIcs.CLarXiv:1811.12354v72018
  30. An End-to-End Transformer Model for 3D Object Detection

    Ishan Misra, Rohit Girdhar, Armand Joulin

    cs.CVcs.AIcs.LGarXiv:2109.08141v12021
  31. Evolution Fine-Tuning: Learning to Discover Across 371 Optimization Tasks

    Young-Jun Lee, Seungone Kim, Minki Kang +5

    cs.CLcs.LGarXiv:2606.29082v12026
  32. Fused Gromov-Wasserstein distance for structured objects: theoretical foundations and mathematical properties

    Titouan Vayer, Laetita Chapel, Rémi Flamary +2

    stat.MLcs.LGarXiv:1811.02834v12018
  33. Effective and Efficient Graph Learning for Multi-view Clustering

    Quanxue Gao, Wei Xia, Xinbo Gao +3

    cs.LGarXiv:2108.06734v22021
  34. REVES: REvision and VErification--Augmented Training for Test-Time Scaling

    Yuanxin Liu, Ruida Zhou, Xinyan Zhao +6

    cs.LGcs.CLarXiv:2606.18910v12026
  35. Post-hoc Interpretability for Neural NLP: A Survey

    Andreas Madsen, Siva Reddy, Sarath Chandar

    cs.CLcs.LGcs.NEarXiv:2108.04840v52021
  36. Fashionable Modelling with Flux

    Michael Innes, Elliot Saba, Keno Fischer +6

    cs.PLcs.LGarXiv:1811.01457v32018
  37. A Persistent Spatial Semantic Representation for High-level Natural Language Instruction Execution

    Valts Blukis, Chris Paxton, Dieter Fox +2

    cs.ROcs.AIcs.CLarXiv:2107.05612v32021
  38. Meta-learning PINN loss functions

    Apostolos F Psaros, Kenji Kawaguchi, George Em Karniadakis

    cs.LGarXiv:2107.05544v12021
  39. Learning to Delegate for Large-scale Vehicle Routing

    Sirui Li, Zhongxia Yan, Cathy Wu

    cs.LGcs.AIarXiv:2107.04139v22021
  40. SNIP: Single-shot Network Pruning based on Connection Sensitivity

    Namhoon Lee, Thalaiyasingam Ajanthan, Philip H. S. Torr

    cs.CVcs.LGarXiv:1810.02340v22018
  41. Analysis of Diffractive Optical Neural Networks and Their Integration with Electronic Neural Networks

    Deniz Mengu, Yi Luo, Yair Rivenson +1

    cs.NEcs.LGphysics.opticsarXiv:1810.01916v42018
  42. The Road Ahead in Autonomous Driving: The KITScenes Multimodal Dataset

    Richard Schwarzkopf, Fabian Immel, Alexander Blumberg +21

    cs.CVcs.LGcs.ROarXiv:2606.02956v12026
  43. Scenic: A Language for Scenario Specification and Scene Generation

    Daniel J. Fremont, Tommaso Dreossi, Shromona Ghosh +3

    cs.PLcs.CVcs.LGarXiv:1809.09310v22018
  44. Decentralized Instruction Tuning: Conflict-Aware Splitting and Weight Merging

    Minsik Choi, Geewook Kim

    cs.LGarXiv:2606.01717v12026
  45. GANs for Medical Image Analysis

    Salome Kazeminia, Christoph Baur, Arjan Kuijper +4

    cs.CVcs.LGstat.MLarXiv:1809.06222v32018
  46. Don't Use Large Mini-Batches, Use Local SGD

    Tao Lin, Sebastian U. Stich, Kumar Kshitij Patel +1

    cs.LGstat.MLarXiv:1808.07217v62018
    Summaries:한국어
  47. QuAC : Question Answering in Context

    Eunsol Choi, He He, Mohit Iyyer +5

    cs.CLcs.AIcs.LGarXiv:1808.07036v32018
  48. Identifying Implementation Bugs in Machine Learning based Image Classifiers using Metamorphic Testing

    Anurag Dwarakanath, Manish Ahuja, Samarth Sikand +4

    cs.SEcs.LGarXiv:1808.05353v12018
  49. On the limits and opportunities of AI reviewers: Reviewing the reviews of Nature-family papers with 45 expert scientists

    Seungone Kim, Dongkeun Yoon, Kiril Gashteovski +55

    cs.CLcs.AIcs.LGarXiv:2605.20668v12026
  50. Supporting Very Large Models using Automatic Dataflow Graph Partitioning

    Minjie Wang, Chien-chin Huang, Jinyang Li

    cs.DCcs.LGarXiv:1807.08887v22018
  51. CurveBench: A Benchmark for Exact Topological Reasoning over Nested Jordan Curves

    Amirreza Mohseni, Mona Mohammadi, Morteza Saghafian +1

    cs.CVcs.LGarXiv:2605.14068v22026
  52. Visual Domain Adaptation with Manifold Embedded Distribution Alignment

    Jindong Wang, Wenjie Feng, Yiqiang Chen +3

    cs.CVcs.LGarXiv:1807.07258v22018
  53. IndicMedDialog: A Parallel Multi-Turn Medical Dialogue Dataset for Accessible Healthcare in Indic Languages

    Shubham Kumar Nigam, Suparnojit Sarkar, Piyush Patel

    cs.CLcs.AIcs.IRarXiv:2605.13292v12026
  54. From Frequency to Meaning: Vector Space Models of Semantics

    Peter D. Turney, Patrick Pantel

    cs.CLcs.IRcs.LGarXiv:1003.1141v12010
  55. DeepAffinity: Interpretable Deep Learning of Compound-Protein Affinity through Unified Recurrent and Convolutional Neural Networks

    Mostafa Karimi, Di Wu, Zhangyang Wang +1

    q-bio.BMcs.LGstat.MLarXiv:1806.07537v22018
  56. Neural Code Comprehension: A Learnable Representation of Code Semantics

    Tal Ben-Nun, Alice Shoshana Jakobovits, Torsten Hoefler

    cs.LGcs.NEcs.PLarXiv:1806.07336v32018
  57. AI scientists produce results without reasoning scientifically

    Martiño Ríos-García, Nawaf Alampara, Chandan Gupta +5

    cs.AIcond-mat.mtrl-scics.LGarXiv:2604.18805v12026
  58. Unsupervised Alignment of Embeddings with Wasserstein Procrustes

    Edouard Grave, Armand Joulin, Quentin Berthet

    cs.LGcs.CLstat.MLarXiv:1805.11222v12018
  59. SCOPE: Signal-Calibrated On-Policy Distillation Enhancement with Dual-Path Adaptive Weighting

    Binbin Zheng, Xing Ma, Yiheng Liang +6

    cs.LGcs.AIcs.CLarXiv:2604.10688v22026
  60. The EuroCity Persons Dataset: A Novel Benchmark for Object Detection

    Markus Braun, Sebastian Krebs, Fabian Flohr +1

    cs.CVcs.AIcs.LGarXiv:1805.07193v22018