Machine Learning
Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
5,581 to 5,640 of 20,199
Creation begins with understanding: LLMs as strategy designers for privacy-preserving tabular data synthesis
Jinmeng Li, Quan Zhang, Hangting Ye +4
cs.LGarXiv:2608.29674v12026International AI Safety Report
Yoshua Bengio, Sören Mindermann, Daniel Privitera +93
cs.CYcs.AIcs.LGarXiv:2501.17805v12025Steering Language Models With Activation Engineering
Alexander Matt Turner, Lisa Thiergart, Gavin Leech +4
cs.CLcs.LGarXiv:2308.10248v52023When BERT Plays the Lottery, All Tickets Are Winning
Sai Prasanna, Anna Rogers, Anna Rumshisky
cs.CLcs.LGarXiv:2005.00561v22020Physics Informed Extreme Learning Machine (PIELM) -- A rapid method for the numerical solution of partial differential equations
Vikas Dwivedi, Balaji Srinivasan
cs.LGphysics.comp-phstat.MLarXiv:1907.03507v12019Crop Yield Prediction Integrating Genotype and Weather Variables Using Deep Learning
Johnathon Shook, Tryambak Gangopadhyay, Linjiang Wu +3
cs.LGstat.MLarXiv:2006.13847v12020Can We Detect Failures Without Failure Data? Uncertainty-Aware Runtime Failure Detection for Imitation Learning Policies
Chen Xu, Tony Khuong Nguyen, Emma Dixon +7
cs.ROcs.AIcs.LGarXiv:2503.08558v32025Sim-to-Real Reinforcement Learning for Vision-Based Dexterous Manipulation on Humanoids
Toru Lin, Kartik Sachdev, Linxi Fan +2
cs.ROcs.AIcs.CVarXiv:2502.20396v22025Estimating the Prediction Performance of Spatial Models via Spatial k-Fold Cross Validation
Jonne Pohjankukka, Tapio Pahikkala, Paavo Nevalainen +1
stat.APcs.LGarXiv:2005.14263v12020POP909: A Pop-song Dataset for Music Arrangement Generation
Ziyu Wang, Ke Chen, Junyan Jiang +5
cs.SDcs.IRcs.LGarXiv:2008.07142v12020Provably Efficient Federated Reinforcement Learning with Linear Function Approximation and Logarithmic Communication Cost
Zihang Liang, Haochen Zhang, Lingzhou Xue
stat.MLcs.AIcs.LGarXiv:2609.00193v12026SGD Learns the Conjugate Kernel Class of the Network
Amit Daniely
cs.LGcs.DSstat.MLarXiv:1702.08503v22017Dynamic Weighted Learning for Unsupervised Domain Adaptation
Ni Xiao, Lei Zhang
cs.LGarXiv:2103.13814v12021Sustainable LLM Inference for Edge AI: Evaluating Quantized LLMs for Energy Efficiency, Output Accuracy, and Inference Latency
Erik Johannes Husom, Arda Goknil, Merve Astekin +5
cs.CYcs.AIcs.CLarXiv:2504.03360v12025When Does Self-supervision Improve Few-shot Learning?
Jong-Chyi Su, Subhransu Maji, Bharath Hariharan
cs.CVcs.LGarXiv:1910.03560v22019ECA-BLS: An Efficient Complex-Augmented Broad Learning System
A. Rahaman, A. Quadir, M. Sajid +2
cs.LGarXiv:2608.29763v12026Reward-guided Fine-Tuning of One-Step Generative Models via Wasserstein Gradient Flow
Hoseong Hwang, Woorim Han, Joungin Chun +2
cs.LGarXiv:2608.29647v12026StarCraft Micromanagement with Reinforcement Learning and Curriculum Transfer Learning
Kun Shao, Yuanheng Zhu, Dongbin Zhao
cs.AIcs.LGcs.MAarXiv:1804.00810v12018WirelessGPT: A Generative Pre-trained Multi-task Learning Framework for Wireless Communication
Tingting Yang, Ping Zhang, Mengfan Zheng +4
cs.LGarXiv:2502.06877v12025Process Reward Models for LLM Agents: Practical Framework and Directions
Sanjiban Choudhury
cs.LGcs.AIarXiv:2502.10325v12025Multi-Head Attention: Collaborate Instead of Concatenate
Jean-Baptiste Cordonnier, Andreas Loukas, Martin Jaggi
cs.LGcs.CLstat.MLarXiv:2006.16362v22020Intelligent Zero Trust Architecture for 5G/6G Networks: Principles, Challenges, and the Role of Machine Learning in the context of O-RAN
Keyvan Ramezanpour, Jithin Jagannath
cs.NIcs.LGarXiv:2105.01478v32021Hardware-aware training for large-scale and diverse deep learning inference workloads using in-memory computing-based accelerators
Malte J. Rasch, Charles Mackin, Manuel Le Gallo +10
cs.LGcs.ETarXiv:2302.08469v12023Gradient Alignment in Physics-informed Neural Networks: A Second-Order Optimization Perspective
Sifan Wang, Ananyae Kumar Bhartari, Bowen Li +1
cs.LGcs.AIphysics.comp-pharXiv:2502.00604v22025Universal Statistics of Fisher Information in Deep Neural Networks: Mean Field Approach
Ryo Karakida, Shotaro Akaho, Shun-ichi Amari
stat.MLcond-mat.dis-nncs.LGarXiv:1806.01316v32018Music transcription modelling and composition using deep learning
Bob L. Sturm, João Felipe Santos, Oded Ben-Tal +1
cs.SDcs.LGarXiv:1604.08723v12016MasRouter: Learning to Route LLMs for Multi-Agent Systems
Yanwei Yue, Guibin Zhang, Boyang Liu +4
cs.LGcs.MAarXiv:2502.11133v12025AssemblyNet: A large ensemble of CNNs for 3D Whole Brain MRI Segmentation
Pierrick Coupé, Boris Mansencal, Michaël Clément +5
eess.IVcs.CVcs.LGarXiv:1911.09098v12019Small-scale proxies for large-scale Transformer training instabilities
Mitchell Wortsman, Peter J. Liu, Lechao Xiao +13
cs.LGarXiv:2309.14322v22023Generative Feature Replay For Class-Incremental Learning
Xialei Liu, Chenshen Wu, Mikel Menta +5
cs.CVcs.LGarXiv:2004.09199v12020Robust Broad Learning System with Wave Loss for Classification under Data Uncertainty
Mushir Akhtar, A. Varshney, A. Quadir +3
cs.LGarXiv:2608.29983v12026Zero-Shot Whole-Body Humanoid Control via Behavioral Foundation Models
Andrea Tirinzoni, Ahmed Touati, Jesse Farebrother +5
cs.LGarXiv:2504.11054v12025GraM-Diff: A Unified Graph-Mamba Diffusion Framework for EEG-Based Alzheimer's Disease Data Generation and Diagnosis
M. Tanveer, Ayush Singh Rana, Sanskriti Jain +5
cs.LGarXiv:2608.29755v12026TRACE: Retrospective Streaming Generation of Physical Fields under Sparse Structured Sensing
Xinyu Zhang, Lihao Chen, Panqi Chen +4
stat.MLcs.LGarXiv:2608.26219v12026Local Implicit Grid Representations for 3D Scenes
Chiyu Max Jiang, Avneesh Sud, Ameesh Makadia +3
cs.CVcs.CGcs.LGarXiv:2003.08981v12020Optimization Methods for Large-Scale Machine Learning
Léon Bottou, Frank E. Curtis, Jorge Nocedal
stat.MLcs.LGmath.OCarXiv:1606.04838v32016El Agente: An Autonomous Agent for Quantum Chemistry
Yunheng Zou, Austin H. Cheng, Abdulrahman Aldossary +13
cs.AIcs.LGcs.MAarXiv:2505.02484v22025Multi-Objective Bayesian Optimization over High-Dimensional Search Spaces
Samuel Daulton, David Eriksson, Maximilian Balandat +1
cs.LGcs.AImath.OCarXiv:2109.10964v42021Knowledge Distillation under Teacher Misspecification: An Order-Parameter Analysis of the Gap between Teacher Mimicry and Task Performance
Kazuyuki Hara, Hideitsu Hino
cs.LGcs.AIarXiv:2608.29472v12026Autellix: An Efficient Serving Engine for LLM Agents as General Programs
Michael Luo, Xiaoxiang Shi, Colin Cai +8
cs.LGcs.AIcs.DCarXiv:2502.13965v12025ViewAL: Active Learning with Viewpoint Entropy for Semantic Segmentation
Yawar Siddiqui, Julien Valentin, Matthias Nießner
cs.CVcs.LGarXiv:1911.11789v22019Rethinking Experience Replay: a Bag of Tricks for Continual Learning
Pietro Buzzega, Matteo Boschini, Angelo Porrello +1
cs.LGstat.MLarXiv:2010.05595v12020Generalized Interpolating Discrete Diffusion
Dimitri von Rütte, Janis Fluri, Yuhui Ding +3
cs.CLcs.AIcs.LGarXiv:2503.04482v22025Generalizing to unseen domains via distribution matching
Isabela Albuquerque, João Monteiro, Mohammad Darvishi +2
cs.LGstat.MLarXiv:1911.00804v62019On-Policy Distillation Meets Off-Policy GRPO: Training Compact Instruction-Following Rerankers
Vignesh Prabhakar, Jialing Pan, Anil Babu Ankisettipalli
cs.LGcs.AIarXiv:2609.01947v12026OutageDiT: A Generative Foundation Model for Power Outage Forecasting and Scenario Simulation
Yunqin Zhu, Feng Qiu, Yao Xie
cs.LGcs.AIarXiv:2609.01896v12026Import What You Need: Learning When and How to Augment EHR Graphs with External Knowledge
Chen Chen, Mohsen Nayebi Kerdabadi, Dongjie Wang +2
cs.LGcs.AIarXiv:2609.01839v12026SumGNN: Multi-typed Drug Interaction Prediction via Efficient Knowledge Graph Summarization
Yue Yu, Kexin Huang, Chao Zhang +3
cs.LGcs.CLcs.IRarXiv:2010.01450v22020Path Planning for Masked Diffusion Model Sampling
Fred Zhangzhi Peng, Zachary Bezemek, Sawan Patel +5
cs.LGcs.AIarXiv:2502.03540v52025A Recipe for Watermarking Diffusion Models
Yunqing Zhao, Tianyu Pang, Chao Du +3
cs.CVcs.CRcs.LGarXiv:2303.10137v22023LiDAR Snowfall Simulation for Robust 3D Object Detection
Martin Hahner, Christos Sakaridis, Mario Bijelic +4
cs.CVcs.LGarXiv:2203.15118v22022LLMs4OL: Large Language Models for Ontology Learning
Hamed Babaei Giglou, Jennifer D'Souza, Sören Auer
cs.AIcs.CLcs.ITarXiv:2307.16648v22023Heterogeneous Ensemble Knowledge Transfer for Training Large Models in Federated Learning
Yae Jee Cho, Andre Manoel, Gauri Joshi +2
cs.LGarXiv:2204.12703v12022MMRL: Multi-Modal Representation Learning for Vision-Language Models
Yuncheng Guo, Xiaodong Gu
cs.LGcs.CVarXiv:2503.08497v22025Federated Learning via Intelligent Reflecting Surface
Zhibin Wang, Jiahang Qiu, Yong Zhou +4
cs.ITcs.LGeess.SParXiv:2011.05051v22020Harnessing the Universal Geometry of Embeddings
Rishi Jha, Collin Zhang, Vitaly Shmatikov +1
cs.LGarXiv:2505.12540v42025Hybrid Macro/Micro Level Backpropagation for Training Deep Spiking Neural Networks
Yingyezhe Jin, Wenrui Zhang, Peng Li
cs.NEcs.LGarXiv:1805.07866v62018PromptTTS: Controllable Text-to-Speech with Text Descriptions
Zhifang Guo, Yichong Leng, Yihan Wu +2
eess.AScs.CLcs.LGarXiv:2211.12171v12022IGRF-RFE: A Hybrid Feature Selection Method for MLP-based Network Intrusion Detection on UNSW-NB15 Dataset
Yuhua Yin, Julian Jang-Jaccard, Wen Xu +4
cs.LGcs.CRarXiv:2203.16365v22022RLVMR: Reinforcement Learning with Verifiable Meta-Reasoning Rewards for Robust Long-Horizon Agents
Zijing Zhang, Ziyang Chen, Mingxiao Li +2
cs.LGcs.AIarXiv:2507.22844v12025