Machine Learning

Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

5,581 to 5,640 of 20,199

  1. Creation begins with understanding: LLMs as strategy designers for privacy-preserving tabular data synthesis

    Jinmeng Li, Quan Zhang, Hangting Ye +4

    cs.LGarXiv:2608.29674v12026
  2. International AI Safety Report

    Yoshua Bengio, Sören Mindermann, Daniel Privitera +93

    cs.CYcs.AIcs.LGarXiv:2501.17805v12025
  3. Steering Language Models With Activation Engineering

    Alexander Matt Turner, Lisa Thiergart, Gavin Leech +4

    cs.CLcs.LGarXiv:2308.10248v52023
  4. When BERT Plays the Lottery, All Tickets Are Winning

    Sai Prasanna, Anna Rogers, Anna Rumshisky

    cs.CLcs.LGarXiv:2005.00561v22020
  5. Physics Informed Extreme Learning Machine (PIELM) -- A rapid method for the numerical solution of partial differential equations

    Vikas Dwivedi, Balaji Srinivasan

    cs.LGphysics.comp-phstat.MLarXiv:1907.03507v12019
  6. Crop Yield Prediction Integrating Genotype and Weather Variables Using Deep Learning

    Johnathon Shook, Tryambak Gangopadhyay, Linjiang Wu +3

    cs.LGstat.MLarXiv:2006.13847v12020
  7. Can We Detect Failures Without Failure Data? Uncertainty-Aware Runtime Failure Detection for Imitation Learning Policies

    Chen Xu, Tony Khuong Nguyen, Emma Dixon +7

    cs.ROcs.AIcs.LGarXiv:2503.08558v32025
  8. Sim-to-Real Reinforcement Learning for Vision-Based Dexterous Manipulation on Humanoids

    Toru Lin, Kartik Sachdev, Linxi Fan +2

    cs.ROcs.AIcs.CVarXiv:2502.20396v22025
  9. Estimating the Prediction Performance of Spatial Models via Spatial k-Fold Cross Validation

    Jonne Pohjankukka, Tapio Pahikkala, Paavo Nevalainen +1

    stat.APcs.LGarXiv:2005.14263v12020
  10. POP909: A Pop-song Dataset for Music Arrangement Generation

    Ziyu Wang, Ke Chen, Junyan Jiang +5

    cs.SDcs.IRcs.LGarXiv:2008.07142v12020
  11. Provably Efficient Federated Reinforcement Learning with Linear Function Approximation and Logarithmic Communication Cost

    Zihang Liang, Haochen Zhang, Lingzhou Xue

    stat.MLcs.AIcs.LGarXiv:2609.00193v12026
  12. SGD Learns the Conjugate Kernel Class of the Network

    Amit Daniely

    cs.LGcs.DSstat.MLarXiv:1702.08503v22017
  13. Dynamic Weighted Learning for Unsupervised Domain Adaptation

    Ni Xiao, Lei Zhang

    cs.LGarXiv:2103.13814v12021
  14. Sustainable LLM Inference for Edge AI: Evaluating Quantized LLMs for Energy Efficiency, Output Accuracy, and Inference Latency

    Erik Johannes Husom, Arda Goknil, Merve Astekin +5

    cs.CYcs.AIcs.CLarXiv:2504.03360v12025
  15. When Does Self-supervision Improve Few-shot Learning?

    Jong-Chyi Su, Subhransu Maji, Bharath Hariharan

    cs.CVcs.LGarXiv:1910.03560v22019
  16. ECA-BLS: An Efficient Complex-Augmented Broad Learning System

    A. Rahaman, A. Quadir, M. Sajid +2

    cs.LGarXiv:2608.29763v12026
  17. Reward-guided Fine-Tuning of One-Step Generative Models via Wasserstein Gradient Flow

    Hoseong Hwang, Woorim Han, Joungin Chun +2

    cs.LGarXiv:2608.29647v12026
  18. StarCraft Micromanagement with Reinforcement Learning and Curriculum Transfer Learning

    Kun Shao, Yuanheng Zhu, Dongbin Zhao

    cs.AIcs.LGcs.MAarXiv:1804.00810v12018
  19. WirelessGPT: A Generative Pre-trained Multi-task Learning Framework for Wireless Communication

    Tingting Yang, Ping Zhang, Mengfan Zheng +4

    cs.LGarXiv:2502.06877v12025
  20. Process Reward Models for LLM Agents: Practical Framework and Directions

    Sanjiban Choudhury

    cs.LGcs.AIarXiv:2502.10325v12025
  21. Multi-Head Attention: Collaborate Instead of Concatenate

    Jean-Baptiste Cordonnier, Andreas Loukas, Martin Jaggi

    cs.LGcs.CLstat.MLarXiv:2006.16362v22020
  22. Intelligent Zero Trust Architecture for 5G/6G Networks: Principles, Challenges, and the Role of Machine Learning in the context of O-RAN

    Keyvan Ramezanpour, Jithin Jagannath

    cs.NIcs.LGarXiv:2105.01478v32021
  23. Hardware-aware training for large-scale and diverse deep learning inference workloads using in-memory computing-based accelerators

    Malte J. Rasch, Charles Mackin, Manuel Le Gallo +10

    cs.LGcs.ETarXiv:2302.08469v12023
  24. Gradient Alignment in Physics-informed Neural Networks: A Second-Order Optimization Perspective

    Sifan Wang, Ananyae Kumar Bhartari, Bowen Li +1

    cs.LGcs.AIphysics.comp-pharXiv:2502.00604v22025
  25. Universal Statistics of Fisher Information in Deep Neural Networks: Mean Field Approach

    Ryo Karakida, Shotaro Akaho, Shun-ichi Amari

    stat.MLcond-mat.dis-nncs.LGarXiv:1806.01316v32018
  26. Music transcription modelling and composition using deep learning

    Bob L. Sturm, João Felipe Santos, Oded Ben-Tal +1

    cs.SDcs.LGarXiv:1604.08723v12016
  27. MasRouter: Learning to Route LLMs for Multi-Agent Systems

    Yanwei Yue, Guibin Zhang, Boyang Liu +4

    cs.LGcs.MAarXiv:2502.11133v12025
  28. AssemblyNet: A large ensemble of CNNs for 3D Whole Brain MRI Segmentation

    Pierrick Coupé, Boris Mansencal, Michaël Clément +5

    eess.IVcs.CVcs.LGarXiv:1911.09098v12019
  29. Small-scale proxies for large-scale Transformer training instabilities

    Mitchell Wortsman, Peter J. Liu, Lechao Xiao +13

    cs.LGarXiv:2309.14322v22023
  30. Generative Feature Replay For Class-Incremental Learning

    Xialei Liu, Chenshen Wu, Mikel Menta +5

    cs.CVcs.LGarXiv:2004.09199v12020
  31. Robust Broad Learning System with Wave Loss for Classification under Data Uncertainty

    Mushir Akhtar, A. Varshney, A. Quadir +3

    cs.LGarXiv:2608.29983v12026
  32. Zero-Shot Whole-Body Humanoid Control via Behavioral Foundation Models

    Andrea Tirinzoni, Ahmed Touati, Jesse Farebrother +5

    cs.LGarXiv:2504.11054v12025
  33. GraM-Diff: A Unified Graph-Mamba Diffusion Framework for EEG-Based Alzheimer's Disease Data Generation and Diagnosis

    M. Tanveer, Ayush Singh Rana, Sanskriti Jain +5

    cs.LGarXiv:2608.29755v12026
  34. TRACE: Retrospective Streaming Generation of Physical Fields under Sparse Structured Sensing

    Xinyu Zhang, Lihao Chen, Panqi Chen +4

    stat.MLcs.LGarXiv:2608.26219v12026
  35. Local Implicit Grid Representations for 3D Scenes

    Chiyu Max Jiang, Avneesh Sud, Ameesh Makadia +3

    cs.CVcs.CGcs.LGarXiv:2003.08981v12020
  36. Optimization Methods for Large-Scale Machine Learning

    Léon Bottou, Frank E. Curtis, Jorge Nocedal

    stat.MLcs.LGmath.OCarXiv:1606.04838v32016
  37. El Agente: An Autonomous Agent for Quantum Chemistry

    Yunheng Zou, Austin H. Cheng, Abdulrahman Aldossary +13

    cs.AIcs.LGcs.MAarXiv:2505.02484v22025
  38. Multi-Objective Bayesian Optimization over High-Dimensional Search Spaces

    Samuel Daulton, David Eriksson, Maximilian Balandat +1

    cs.LGcs.AImath.OCarXiv:2109.10964v42021
  39. Knowledge Distillation under Teacher Misspecification: An Order-Parameter Analysis of the Gap between Teacher Mimicry and Task Performance

    Kazuyuki Hara, Hideitsu Hino

    cs.LGcs.AIarXiv:2608.29472v12026
  40. Autellix: An Efficient Serving Engine for LLM Agents as General Programs

    Michael Luo, Xiaoxiang Shi, Colin Cai +8

    cs.LGcs.AIcs.DCarXiv:2502.13965v12025
  41. ViewAL: Active Learning with Viewpoint Entropy for Semantic Segmentation

    Yawar Siddiqui, Julien Valentin, Matthias Nießner

    cs.CVcs.LGarXiv:1911.11789v22019
  42. Rethinking Experience Replay: a Bag of Tricks for Continual Learning

    Pietro Buzzega, Matteo Boschini, Angelo Porrello +1

    cs.LGstat.MLarXiv:2010.05595v12020
  43. Generalized Interpolating Discrete Diffusion

    Dimitri von Rütte, Janis Fluri, Yuhui Ding +3

    cs.CLcs.AIcs.LGarXiv:2503.04482v22025
  44. Generalizing to unseen domains via distribution matching

    Isabela Albuquerque, João Monteiro, Mohammad Darvishi +2

    cs.LGstat.MLarXiv:1911.00804v62019
  45. On-Policy Distillation Meets Off-Policy GRPO: Training Compact Instruction-Following Rerankers

    Vignesh Prabhakar, Jialing Pan, Anil Babu Ankisettipalli

    cs.LGcs.AIarXiv:2609.01947v12026
  46. OutageDiT: A Generative Foundation Model for Power Outage Forecasting and Scenario Simulation

    Yunqin Zhu, Feng Qiu, Yao Xie

    cs.LGcs.AIarXiv:2609.01896v12026
  47. Import What You Need: Learning When and How to Augment EHR Graphs with External Knowledge

    Chen Chen, Mohsen Nayebi Kerdabadi, Dongjie Wang +2

    cs.LGcs.AIarXiv:2609.01839v12026
  48. SumGNN: Multi-typed Drug Interaction Prediction via Efficient Knowledge Graph Summarization

    Yue Yu, Kexin Huang, Chao Zhang +3

    cs.LGcs.CLcs.IRarXiv:2010.01450v22020
  49. Path Planning for Masked Diffusion Model Sampling

    Fred Zhangzhi Peng, Zachary Bezemek, Sawan Patel +5

    cs.LGcs.AIarXiv:2502.03540v52025
  50. A Recipe for Watermarking Diffusion Models

    Yunqing Zhao, Tianyu Pang, Chao Du +3

    cs.CVcs.CRcs.LGarXiv:2303.10137v22023
  51. LiDAR Snowfall Simulation for Robust 3D Object Detection

    Martin Hahner, Christos Sakaridis, Mario Bijelic +4

    cs.CVcs.LGarXiv:2203.15118v22022
  52. LLMs4OL: Large Language Models for Ontology Learning

    Hamed Babaei Giglou, Jennifer D'Souza, Sören Auer

    cs.AIcs.CLcs.ITarXiv:2307.16648v22023
  53. Heterogeneous Ensemble Knowledge Transfer for Training Large Models in Federated Learning

    Yae Jee Cho, Andre Manoel, Gauri Joshi +2

    cs.LGarXiv:2204.12703v12022
  54. MMRL: Multi-Modal Representation Learning for Vision-Language Models

    Yuncheng Guo, Xiaodong Gu

    cs.LGcs.CVarXiv:2503.08497v22025
  55. Federated Learning via Intelligent Reflecting Surface

    Zhibin Wang, Jiahang Qiu, Yong Zhou +4

    cs.ITcs.LGeess.SParXiv:2011.05051v22020
  56. Harnessing the Universal Geometry of Embeddings

    Rishi Jha, Collin Zhang, Vitaly Shmatikov +1

    cs.LGarXiv:2505.12540v42025
  57. Hybrid Macro/Micro Level Backpropagation for Training Deep Spiking Neural Networks

    Yingyezhe Jin, Wenrui Zhang, Peng Li

    cs.NEcs.LGarXiv:1805.07866v62018
  58. PromptTTS: Controllable Text-to-Speech with Text Descriptions

    Zhifang Guo, Yichong Leng, Yihan Wu +2

    eess.AScs.CLcs.LGarXiv:2211.12171v12022
  59. IGRF-RFE: A Hybrid Feature Selection Method for MLP-based Network Intrusion Detection on UNSW-NB15 Dataset

    Yuhua Yin, Julian Jang-Jaccard, Wen Xu +4

    cs.LGcs.CRarXiv:2203.16365v22022
  60. RLVMR: Reinforcement Learning with Verifiable Meta-Reasoning Rewards for Robust Long-Horizon Agents

    Zijing Zhang, Ziyang Chen, Mingxiao Li +2

    cs.LGcs.AIarXiv:2507.22844v12025