Machine Learning
Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
4,321 to 4,380 of 20,454
Same Request, Different Answer: Quantization Amplifies Cache-Induced Divergence in LLM Serving
Aditi Patodiya
cs.SEcs.DCcs.LGarXiv:2609.04748v12026Low Level Control of a Quadrotor with Deep Model-Based Reinforcement Learning
Nathan O. Lambert, Daniel S. Drew, Joseph Yaconelli +3
cs.ROcs.LGarXiv:1901.03737v22019AQAScore: Evaluating Semantic Alignment in Text-to-Audio Generation via Audio Question Answering
Chun-Yi Kuan, Kai-Wei Chang, Hung-yi Lee
eess.AScs.AIcs.CLarXiv:2601.14728v12026FluxDisco: Symbolic Regression for Stoichiometric Dynamical Systems via Monte Carlo Graph Search
Cassandra Durr, Alvaro Köhn-Luque, Chris Jewell +1
stat.MLcs.LGphysics.data-anarXiv:2609.05207v12026Designing Interpretable ML System to Enhance Trust in Healthcare: A Systematic Review to Proposed Responsible Clinician-AI-Collaboration Framework
Elham Nasarian, Roohallah Alizadehsani, U. Rajendra Acharya +1
cs.AIcs.HCcs.LGarXiv:2311.11055v22023An Alternative Probabilistic Interpretation of the Huber Loss
Gregory P. Meyer
stat.MLcs.CVcs.LGarXiv:1911.02088v32019Impact of Data Loss in Postprocessing on Training and Inference of Quantum Neural Networks
Soraya V. Panambalom, Edoardo Altamura, Nick Chancellor +1
quant-phcs.ETcs.LGarXiv:2609.05060v12026Terahertz-Band Joint Ultra-Massive MIMO Radar-Communications: Model-Based and Model-Free Hybrid Beamforming
Ahmet M. Elbir, Kumar Vijay Mishra, Symeon Chatzinotas
eess.SPcs.ITcs.LGarXiv:2103.00328v22021An Analysis of Self-supervised Pre-training with Dependent Samples
Maximilian Fleissner, Debarghya Ghoshdastidar, Samory Kpotufe
stat.MLcs.LGarXiv:2609.05031v12026On the Generalization Capacities of MLLMs for Spatial Intelligence
Gongjie Zhang, Wenhao Li, Quanhao Qian +4
cs.CVcs.LGarXiv:2603.06704v12026Stress Field Prediction in Cantilevered Structures Using Convolutional Neural Networks
Zhenguo Nie, Haoliang Jiang, Levent Burak Kara
cs.LGstat.MLarXiv:1808.08914v32018IGenBench: Benchmarking the Reliability of Text-to-Infographic Generation
Yinghao Tang, Xueding Liu, Boyuan Zhang +13
cs.LGcs.CVarXiv:2601.04498v22026LLaTTE: Scaling Laws for Multi-Stage Sequence Modeling in Large-Scale Ads Recommendation
Lee Xiong, Zhirong Chen, Rahul Mayuranath +17
cs.IRcs.AIcs.LGarXiv:2601.20083v12026Sustainable Edge Vision via Empirically Calibrated DVFS: Eliminating Thermal Throttling on Passively Cooled Hardware
Aayush Marasini, Zhaoxian Zhou
cs.ARcs.CVcs.LGarXiv:2609.04705v12026IT-OSE: Exploring Optimal Sample Size for Industrial Data Augmentation
Mingchun Sun, Rongqiang Zhao, Zhennan Huang +2
cs.LGcs.AIarXiv:2602.15878v12026RF-GPT: Teaching AI to See the Wireless World
Hang Zou, Yu Tian, Bohao Wang +4
eess.SPcs.LGarXiv:2602.14833v12026An improvement of the convergence proof of the ADAM-Optimizer
Sebastian Bock, Josef Goppold, Martin Weiß
cs.LGcs.AIstat.MLarXiv:1804.10587v12018Distill Globally, Adapt Locally: Reasoning Distillation and Product-Type Test-Time Training for Scalable Trade-Up Recommendation
Siliang Liu, Mohammad Ghasemi, Sapan Patel +1
cs.LGarXiv:2609.05363v12026Embedded Graph Flows for Categorical Graph Generation
Ethan Ma, Zihan Wang, Chris Siu Yeung Chow +5
cs.LGarXiv:2609.05328v12026Variational Continuation for Double Pendulum Periodic Orbits
Leo Yao, Ziming Liu, Max Tegmark
cs.LGmath-phnlin.CDarXiv:2609.05337v12026Learning from VAE Errors to support ECG-based Differential Diagnosis of Myocardial Scar
Shayan Sharifi, Riccardo Treu, Ilaria Gandin +3
cs.LGstat.MLarXiv:2609.05294v12026How to Speculate about Uncertainty in Agentic Coding? A Draft-Model Gate Method
Konstantin Grotov, Valentin Malykh
cs.LGarXiv:2609.05274v12026A Differentiable Neural Surrogate for Photon Propagation in Neutrino Telescopes
Felix J. Yu, Berthy T. Feng, Nicholas Kamp +1
astro-ph.HEcs.LGhep-exarXiv:2609.04695v12026Towards Automated Deep Learning: Efficient Joint Neural Architecture and Hyperparameter Search
Arber Zela, Aaron Klein, Stefan Falkner +1
cs.LGcs.AIcs.CVarXiv:1807.06906v12018Hidden In Plain Gaze: Gaze Representations as Privacy Controls for Utility and Re-identification Risk in XR
Cory Ilo, Brendan-David John, Doug A. Bowman
cs.CVcs.ETcs.HCarXiv:2609.04592v12026Centered Permutation Prefixes for SGD with Random Reshuffling: Sharp Rates, Hölder Geometry, and Composite Proximal Extensions
Jiaxiang Li
math.OCcs.LGarXiv:2609.04578v12026Adversarial attacks and defenses in explainable artificial intelligence: A survey
Hubert Baniecki, Przemyslaw Biecek
cs.CRcs.AIcs.CVarXiv:2306.06123v42023Data-driven Flood Emulation: Speeding up Urban Flood Predictions by Deep Convolutional Neural Networks
Zifeng Guo, Joao P. Leitao, Nuno E. Simoes +1
cs.CVcs.CYcs.LGarXiv:2004.08340v22020MURAL: Multimodal Uncertainty-aware Recommendation via Adaptive edge Learning
Ahmad Mousavi, Majid Alikhani, Yeon-Chang Lee +2
cs.IRcs.LGcs.SIarXiv:2609.04574v12026Machine Unlearning under Retain-Forget Entanglement
Jingpu Cheng, Ping Liu, Qianxiao Li +1
cs.LGarXiv:2603.26569v12026Sensor-based Continuous Authentication of Smartphones' Users Using Behavioral Biometrics: A Contemporary Survey
Mohammed Abuhamad, Ahmed Abusnaina, DaeHun Nyang +1
cs.CRcs.HCcs.LGarXiv:2001.08578v22020Improving Flow Matching by Aligning Flow Divergence
Yuhao Huang, Taos Transue, Shih-Hsin Wang +3
cs.LGcs.AImath.NAarXiv:2602.00869v12026High-Fidelity Image Generation With Fewer Labels
Mario Lucic, Michael Tschannen, Marvin Ritter +3
cs.LGcs.CVstat.MLarXiv:1903.02271v22019Practical One-Shot Federated Learning for Cross-Silo Setting
Qinbin Li, Bingsheng He, Dawn Song
cs.LGstat.MLarXiv:2010.01017v22020Reward (Mis)design for Autonomous Driving
W. Bradley Knox, Alessandro Allievi, Holger Banzhaf +2
cs.LGarXiv:2104.13906v22021Co-learning: Learning from Noisy Labels with Self-supervision
Cheng Tan, Jun Xia, Lirong Wu +1
cs.LGarXiv:2108.04063v42021WS-GRPO: Weakly-Supervised Group-Relative Policy Optimization for Rollout-Efficient Reasoning
Gagan Mundada, Zihan Huang, Rohan Surana +8
cs.LGarXiv:2602.17025v12026Optimal Rates for Agentic Networked Information Aggregation
MohammadHossein Bateni, Zahra Hadizadeh, MohammadTaghi Hajiaghayi +2
cs.LGcs.GTecon.THarXiv:2609.05318v12026Parallelizing Linear Recurrent Neural Nets Over Sequence Length
Eric Martin, Chris Cundy
cs.NEcs.AIcs.LGarXiv:1709.04057v22017MomentQuant: an even more minimalist interval method with linear time complexity for time series classification
Johann Faouzi
cs.LGarXiv:2609.05136v12026FedDRAW: Federated Dual Reputation Annealing Weighting for Heterogeneous Multi-Institutional Chest Radiograph Classification
Maryam Moradpour, Anne-Christin Hauschild
cs.LGarXiv:2609.05223v12026Towards Execution-Grounded Automated AI Research
Chenglei Si, Zitong Yang, Yejin Choi +3
cs.CLcs.AIcs.LGarXiv:2601.14525v12026From 80x to 385x: A Best-Matching-Unit Search at the L2 Roof, Measured Against a Symmetrically Tuned Baseline
Andrew James Amos
cs.LGarXiv:2609.05138v12026Latent Multi-task Architecture Learning
Sebastian Ruder, Joachim Bingel, Isabelle Augenstein +1
stat.MLcs.AIcs.CLarXiv:1705.08142v32017State Backdoor: Towards Stealthy Real-world Poisoning Attack on Vision-Language-Action Model in State Space
Ji Guo, Wenbo Jiang, Yansong Lin +6
cs.CRcs.LGarXiv:2601.04266v22026Single-Query Black-Box Calibration Auditing via Logit Bias
Roman Plaud, Antoine Saillenfest, Matthieu Labeau +2
cs.LGarXiv:2609.05125v12026Quantum noise protects quantum classifiers against adversaries
Yuxuan Du, Min-Hsiu Hsieh, Tongliang Liu +2
quant-phcs.LGarXiv:2003.09416v12020Deep Microcompression: Structured Pruning and Bit-packed Quantization for Microcontrollers
Opegbemi Matthias Busoye, Tolulope Matthew Busoye, Eghonghon-aye Eigbe
cs.LGarXiv:2609.05081v12026GLASS: Graph-Language Alignment with Spherical Scoring for Transferable Graph-Level Anomaly Detection
Xudong Wang, Chris Ding, Tongxin Li +1
cs.LGarXiv:2609.05253v12026Hessian-based molecular conformation augmentation for a scalable and efficient strategy of machine learning interatomic potentials
Bumju Kwak, Jeonghee Jo
cs.LGphysics.chem-pharXiv:2609.05233v12026Listening while Speaking: Speech Chain by Deep Learning
Andros Tjandra, Sakriani Sakti, Satoshi Nakamura
cs.CLcs.LGcs.SDarXiv:1707.04879v12017Geoopt: Riemannian Optimization in PyTorch
Max Kochurov, Rasul Karimov, Serge Kozlukov
cs.CGcs.LGarXiv:2005.02819v52020Dimension-Adaptive Batched Lipschitz Narrowing Without Knowing the Zooming Dimension
Yasong Feng
cs.LGarXiv:2609.05214v12026Measure and Improve Robustness in NLP Models: A Survey
Xuezhi Wang, Haohan Wang, Diyi Yang
cs.CLcs.LGarXiv:2112.08313v22021Birth of a Transformer: A Memory Viewpoint
Alberto Bietti, Vivien Cabannes, Diane Bouchacourt +2
stat.MLcs.CLcs.LGarXiv:2306.00802v22023Feature Purification: How Adversarial Training Performs Robust Deep Learning
Zeyuan Allen-Zhu, Yuanzhi Li
cs.LGcs.NEmath.OCarXiv:2005.10190v42020ScienceAgentBench: Toward Rigorous Assessment of Language Agents for Data-Driven Scientific Discovery
Ziru Chen, Shijie Chen, Yuting Ning +17
cs.CLcs.AIcs.LGarXiv:2410.05080v32024Coarse-Graining Hidden Representations: Unsupervised Neuron Selection via Mapping Entropy
Margherita Mele, Andrea Castagna, Roberto Menichetti +2
cs.LGcond-mat.dis-nncond-mat.stat-mecharXiv:2609.05126v12026Sign-Based Optimizers Are Effective Under Heavy-Tailed Noise
Dingzhi Yu, Hongyi Tao, Yuanyu Wan +2
cs.LGcs.CLmath.OCarXiv:2602.07425v22026TriMap: Large-scale Dimensionality Reduction Using Triplets
Ehsan Amid, Manfred K. Warmuth
cs.LGstat.MLarXiv:1910.00204v22019