Machine Learning
Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
17,881 to 17,940 of 20,193
Primal Acceleration of Newton's Method
Nikita Doikov
math.OCcs.AIcs.LGarXiv:2608.21359v12026MolGAN: An implicit generative model for small molecular graphs
Nicola De Cao, Thomas Kipf
stat.MLcs.LGarXiv:1805.11973v22018AgentDecarbonizer: Carbon-Aware Execution for AI Agents
Leyi Yan, Shuangning Li, Sihang Liu
cs.LGarXiv:2608.20566v12026Faults That Fortify: CNN Adversarial Robustness via GPU Undervolting
Behnam Omidi, Ahmad Tahmasivand, Husam Alsyouri +4
cs.LGcs.ARcs.CRarXiv:2608.20572v12026Learning Exact NVIDIA SASS Encoders with $\mathbb{F}_2$ Linear Algebra
Jiading Gai
cs.LGarXiv:2608.20532v12026Metag: A dataset to build agentic meta-reviewing capabilities
Anirudh Sundar, Min Chen, Divya Tadimeti +11
cs.LGarXiv:2608.20488v12026When Clean Data Hurts: Learning with Monotone Corruptions Beyond Binary Classification
Julian Asilis, Shaddin Dughmi, Chirag Pabbaraju
cs.LGstat.MLarXiv:2608.20480v12026Amortized Bandwidth Learning for Kernel Density Estimation under Logarithmic Score
Junyi Liang, Hailiang Du
cs.LGarXiv:2608.20445v12026Mutual information and sensitivity analysis for feature selection in customer targeting: a comparative study
Nestor Barraza, Sergio Moro, Marcelo Ferreyra +1
cs.LGmath.PRarXiv:2608.20447v12026Wrong-Physics Backdoors in Neural PDE Operators
Hanbing Liang, Fujun Liu
cs.LGphysics.comp-pharXiv:2608.20439v12026Machine Learning and ARIMA Model Averaging for Adaptive Public Health Forecasting: Comparative Evaluation and an Ontario COVID-19 Case Study
Yushu Zou, Ye Li, Johra Moosa +3
cs.LGstat.AParXiv:2608.20406v12026SAC-Copula: Quality-Preserving Watermarking for Diffusion Language Models via Smooth Correlated Gumbel Fields
Baixin Li, Haiyun He
cs.CLcs.CRcs.LGarXiv:2608.20839v12026MIL-BERT: Classification of Arbitrarily Large Text with Performance and Explanatory Guarantees
John Cadigan, Dayne Freitag, Eric Yeh
cs.CLcs.LGarXiv:2608.20636v12026Multilingual Verifier Bias in RLVR: Benchmark, Rollout Diagnosis, and the Cross-Lingual Selection Bottleneck
Chenyu Zhou, Qiliang Jiang, Xu Zhou
cs.CLcs.LGarXiv:2608.20362v12026TriPLU: Bypassing the Gate with Direct Trilinear Product FFNs in Tiny Language Models
He Zhang
cs.CLcs.LGarXiv:2608.20360v12026TurboBias 2.0: Streaming Context-Biasing for Production-Efficient ASR Systems
Vladimir Bataev, Lilit Grigoryan, Andrei Andrusenko +3
eess.AScs.AIcs.CLarXiv:2608.21343v12026Jacobian-guided Noise Injection for Quantization Robustness in Large Language Models
Deepanshu Pandey, Arnav Chavan, Nahush Lele +2
cs.LGcs.AIarXiv:2608.20988v12026Quantization-Aware Healing: A Practical Recipe for Recovering Compressed, 4-Bit LLMs
Bakbergen Ryskulov, Iker García-Ferrero, David Montero +5
cs.CLcs.AIcs.LGarXiv:2608.20953v12026Fuzzy-MoE: Interpretable Regime-Conditioned Expert Routing for Non-Stationary Multivariate Time Series Forecasting
Lan Guo, Jie Xiao, Zhao Su +5
cs.LGcs.AIarXiv:2608.20761v12026PSK at WMT 2026 MIST: Task-Specialized QLoRA Adapters for Multilingual Summarization and Question Answering
Srikar Kashyap Pulipaka
cs.CLcs.AIcs.LGarXiv:2608.20757v12026Temporal Validity on Real Software Histories: Eliminating Stale-Fact Errors in Code-Assistant Memory over GitHub Fixes
Neeraj Yadav
cs.SEcs.AIcs.CLarXiv:2608.20685v12026No PUN Intended: Plausible Unknown Names for Person-Centred LLM Evaluation
Dimitri Staufer, David Hartmann, Ibrahim Baroud
cs.CLcs.AIcs.LGarXiv:2608.21206v12026Curriculum-Aware Interpolate-then-Refine: Learned Physiological Time-Series Imputation under Realistic Missingness
Yu-Chao Huang, Haochen Zhang, Nicholas Konz +1
cs.LGcs.AIarXiv:2608.21207v12026TracingFlow: A Simulation-Free Trajectory Inference Framework Based on Second-Order Dynamics
Yuhao Sun, Zekun Wu, Zixun Huang +1
cs.LGcs.AIq-bio.GNarXiv:2608.21070v12026Scaling Muon for Diffusion Transformers
Chenghao Li, Xiao Han, Xinxin Huang +22
cs.LGcs.AIcs.CVarXiv:2608.20818v12026Lightweight Adaptive ReduNet via Hyperspherical Manifold Learning
Zhenglin Huang, Qifa Yan, Bin Dai +1
cs.LGcs.AIarXiv:2608.20668v12026Beyond English-Centric Multilingual Machine Translation
Angela Fan, Shruti Bhosale, Holger Schwenk +14
cs.CLcs.LGarXiv:2010.11125v12020JuryProbe: An Empirical Consensus-Risk Diagnostic for Routing Reference-Free Factuality Judge Panels to Grounded Verification
Tianxin Zhou, Ruixi Lin
cs.CLcs.AIcs.LGarXiv:2608.20607v12026MaxViT: Multi-Axis Vision Transformer
Zhengzhong Tu, Hossein Talebi, Han Zhang +4
cs.CVcs.AIcs.LGarXiv:2204.01697v42022BF1: A Causal Dyadic Sparse-Attention Retrofit for Efficient Long-Context Transformers
Hina Dixit
cs.LGcs.AIarXiv:2608.20427v12026From Thermal Preference Prediction to Adaptive Thermal Intervention: A Reinforcement Learning Approach Using Physiological and Environmental Sensing
Isibor Kennedy Ihianle, Emmanuel Manu, Ehsan Asnaashari +4
cs.LGcs.AIarXiv:2608.20423v12026Rigorous Evaluation of Large Language Models for Malaria Drug Discovery: Trade-offs in Performance, Scale, and Resource Utility
Marvellous O. Ajala, Zainab Ashimiyu-Abdusalam, Comfort Adesina
q-bio.QMcs.AIcs.LGarXiv:2608.20418v12026Amplifying the imaging power of digital sky surveys with space telescopes data and generative AI
Sai Teja Erukude, Lior Shamir
astro-ph.IMastro-ph.GAcs.AIarXiv:2608.20666v12026C-Score: Beyond Accuracy for Robustness Assessment in Semi-Supervised Learning under Open-World Unlabeled Contamination
Tsao-Lun Chen, Chi-Cheng Fu, Han-Yi E. Chou +1
cs.LGcs.AIarXiv:2608.20667v12026Describing Videos by Exploiting Temporal Structure
Li Yao, Atousa Torabi, Kyunghyun Cho +4
stat.MLcs.AIcs.CLarXiv:1502.08029v52015Learned Step Size Quantization
Steven K. Esser, Jeffrey L. McKinstry, Deepika Bablani +2
cs.LGstat.MLarXiv:1902.08153v32019Consistency Models for Fast MRI Reconstruction Using Regularization by Denoising
Merve Gülle, Junno Yun, Yaşar Utku Alçalar +1
eess.IVcs.AIcs.CVarXiv:2608.20561v12026Decision Tree and K-Means Analysis of Raman Spectra for Edible Oils: A Physics-Informed AI Approach
Amrita Shaw, Chandrasekar S. N., Sai Muthukumar V. +2
cs.LGcs.AIarXiv:2608.20440v12026Training Deep Neural Networks on Noisy Labels with Bootstrapping
Scott Reed, Honglak Lee, Dragomir Anguelov +3
cs.CVcs.LGcs.NEarXiv:1412.6596v32014Approximate Homomorphisms and Convergent Representations in Transducers
Santiago Cifuentes
cs.LGcs.AIarXiv:2608.20428v12026Simplified State Space Layers for Sequence Modeling
Jimmy T. H. Smith, Andrew Warrington, Scott W. Linderman
cs.LGarXiv:2208.04933v32022VA-DPO: Valence-Arousal Direct Preference Optimization for Controllable Emotion Generation in Language Models
Hyunwoo Kim
cs.CLcs.AIcs.LGarXiv:2608.20374v12026EviRank: Structured Relevance Evidence for Multimodal Image Re-ranking
Enjun Du, Siyi Liu, Zirong Chen +8
cs.CVcs.LGarXiv:2608.20886v12026Let's Scale Step by Step: Compute-Efficient Hyperparameter Transfer for Large-Scale Mixture-of-Experts
Nayeon Kim, Hojin Lee, Yunju Bak +2
cs.LGcs.AIcs.CLarXiv:2608.20061v12026Are GANs Created Equal? A Large-Scale Study
Mario Lucic, Karol Kurach, Marcin Michalski +2
stat.MLcs.LGarXiv:1711.10337v42017ReCurveflow: A Flow Matching Framework that Learns Curved Reaction Trajectories to Predict Transition State Geometries
Seungheun Baek, Mogan Gim, Jaewoo Kang
cs.AIcs.LGarXiv:2608.20869v12026Differentiable Volumetric Rendering: Learning Implicit 3D Representations without 3D Supervision
Michael Niemeyer, Lars Mescheder, Michael Oechsle +1
cs.CVcs.LGeess.IVarXiv:1912.07372v22019Summaries:한국어NeuroStrata: An Electroencephalographic Connectivity-Aware Deep Representation Learning Framework for Dynamic Brain Network Analysis of Mental Stress
Sayantan Acharya, Hamzeh Asgharnezhad, Abbas Khosravi +3
q-bio.NCcs.AIcs.LGarXiv:2608.20354v12026AutoInt: Automatic Feature Interaction Learning via Self-Attentive Neural Networks
Weiping Song, Chence Shi, Zhiping Xiao +4
cs.IRcs.AIcs.LGarXiv:1810.11921v22018Resnet in Resnet: Generalizing Residual Architectures
Sasha Targ, Diogo Almeida, Kevin Lyman
cs.LGcs.CVcs.NEarXiv:1603.08029v12016Unsupervised Learning for Physical Interaction through Video Prediction
Chelsea Finn, Ian Goodfellow, Sergey Levine
cs.LGcs.AIcs.CVarXiv:1605.07157v42016Personalized Privacy Control in LLMs via Attention Head Intervention
Junseok Kim, Nakyeong Yang, Kyomin Jung
cs.AIcs.CLcs.LGarXiv:2608.21209v12026ROCKET: Exceptionally fast and accurate time series classification using random convolutional kernels
Angus Dempster, François Petitjean, Geoffrey I. Webb
cs.LGstat.MLarXiv:1910.13051v12019TreeWY: Speculative Verification for Gated DeltaNet Hybrids
Sneha Murthy Ghantasala
cs.AIcs.CLcs.DCarXiv:2608.20961v12026Neuro-Geospatial Modelling of EEG Affective States Using Literature-Informed Environmental Context
Utsav Poudel, Jagannath Aryal, Subramaniyaswamy Vairavasundaram
cs.AIcs.HCcs.LGarXiv:2608.20807v12026NVAE: A Deep Hierarchical Variational Autoencoder
Arash Vahdat, Jan Kautz
stat.MLcs.CVcs.LGarXiv:2007.03898v32020PointPainting: Sequential Fusion for 3D Object Detection
Sourabh Vora, Alex H. Lang, Bassam Helou +1
cs.CVcs.LGeess.IVarXiv:1911.10150v22019Dual-Cache Latent Space Communication between Heterogeneous Language Models
Jiyao Liu, Qi Zhang, Yaoyi Jia +2
cs.AIcs.LGarXiv:2608.20617v12026Deep Graph Contrastive Representation Learning
Yanqiao Zhu, Yichen Xu, Feng Yu +3
cs.LGstat.MLarXiv:2006.04131v22020Multimodal Learning with Transformers: A Survey
Peng Xu, Xiatian Zhu, David A. Clifton
cs.CVcs.LGarXiv:2206.06488v22022