Machine Learning
Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
9,781 to 9,840 of 19,974
Learning PDE Time-Stepping with Neural Cellular Automata
Esha Saha, Hao Wang
cs.LGstat.MLarXiv:2608.30328v12026Generative multi-domain transfer learning for fault detection in data-scarce wind turbines
Stefan Jonas, Angela Meyer
cs.LGarXiv:2608.30323v12026Scalable Kernel Methods via Doubly Stochastic Gradients
Bo Dai, Bo Xie, Niao He +4
cs.LGstat.MLarXiv:1407.5599v42014Online Estimation of Dynamic Origin-Destination Matrices Using Reinforcement Learning with Link-Flow Propagation Guidance
Donggyu Min, Dong-Kyu Kim
cs.LGcs.AIarXiv:2608.30317v12026Non-Vacuous Generalization Bounds at the ImageNet Scale: A PAC-Bayesian Compression Approach
Wenda Zhou, Victor Veitch, Morgane Austern +2
stat.MLcs.LGarXiv:1804.05862v32018Tail-Replay: Escaping the Curse of Linear Attention in Prefix Caching for Hybrid LLMs
Yirui Liu, Ruoling Qi, Xuaner Wu +2
cs.LGcs.AIarXiv:2608.30310v12026Data Augmentation Approaches in Natural Language Processing: A Survey
Bohan Li, Yutai Hou, Wanxiang Che
cs.CLcs.AIcs.LGarXiv:2110.01852v32021BCPPO: Bachelier-Inspired Constrained Proximal Policy Optimization for Tail-Risk-Aware Safe Reinforcement Learning
Dongsheng Hou, Yanqiao Chen, Yuhan Rui
cs.LGcs.AIarXiv:2608.30283v12026Partial Sum Minimization of Singular Values in Robust PCA: Algorithm and Applications
Tae-Hyun Oh, Yu-Wing Tai, Jean-Charles Bazin +2
cs.CVcs.AIcs.LGarXiv:1503.01444v22015Multivariate Scientific Data Compression with Learned Cross-Variable Latent Decorrelation and Autoregressive Entropy Modeling
Liangji Zhu, Anand Rangarajan, Sanjay Ranka
cs.LGarXiv:2608.30262v12026Exact Recovery Thresholds for Weighted Data Selection in Vector-Valued Linear Regression
Guangjian Zhang
cs.LGmath.STarXiv:2608.30254v12026Certified Safety Radii in Forecast-Error Space for Wasserstein Distributionally Robust Small Signal Stability-Constrained AC Optimal Power Flow via Lifted Spectrahedral Containment
Ziqi Zhang, Xi Chen
cs.LGarXiv:2608.30201v12026SneakyPrompt: Jailbreaking Text-to-image Generative Models
Yuchen Yang, Bo Hui, Haolin Yuan +2
cs.LGarXiv:2305.12082v32023Reinforcement Learning for Symbolic Equation Solving
Kevin P O Keeffe
cs.LGarXiv:2608.30162v12026Benchmarking Peptide-Protein Affinity Prediction Across Peptide and Target Shifts
Jiaxin Tian, Darren An, Jun Li
cs.LGq-bio.QMarXiv:2608.30175v12026Converse and Collision-Based Achievability for Node Localization with Hybrid Distance-Spectral Graph Positional Encodings
Zimo Yan, Yifan Li, Hao Li +4
cs.LGarXiv:2608.30152v12026Machine Learning Enabled Computational Screening of Inorganic Solid Electrolytes for Dendrite Suppression with Li Metal Anode
Zeeshan Ahmad, Tian Xie, Chinmay Maheshwari +2
cond-mat.mtrl-scics.LGphysics.chem-pharXiv:1804.04651v12018Error bounds for approximations with deep ReLU neural networks in $W^{s,p}$ norms
Ingo Gühring, Gitta Kutyniok, Philipp Petersen
math.FAcs.LGarXiv:1902.07896v12019Learning Deep Disentangled Embeddings with the F-Statistic Loss
Karl Ridgeway, Michael C. Mozer
cs.LGcs.AIstat.MLarXiv:1802.05312v22018Strong Drafts Need Compact Memories: Long-Context Speculative Decoding with Compressed KV Cache
Tong Yuan, Chengxi Liao, Zeyi Wen
cs.LGarXiv:2608.30252v12026The Statistical Complexity of Interactive Decision Making
Dylan J. Foster, Sham M. Kakade, Jian Qian +1
cs.LGmath.OCmath.STarXiv:2112.13487v32021HPLFlowNet: Hierarchical Permutohedral Lattice FlowNet for Scene Flow Estimation on Large-scale Point Clouds
Xiuye Gu, Yijie Wang, Chongruo wu +2
cs.CVcs.LGeess.IVarXiv:1906.05332v12019Mask2Former for Video Instance Segmentation
Bowen Cheng, Anwesa Choudhuri, Ishan Misra +3
cs.CVcs.AIcs.LGarXiv:2112.10764v12021You are AllSet: A Multiset Function Framework for Hypergraph Neural Networks
Eli Chien, Chao Pan, Jianhao Peng +1
cs.LGcs.AIarXiv:2106.13264v42021Optimal Scheduling of Isolated Microgrids Using Automated Reinforcement Learning-based Multi-period Forecasting
Yang Li, Ruinong Wang, Zhen Yang
eess.SPcs.LGeess.SYarXiv:2108.06764v12021Machine learning on small size samples: A synthetic knowledge synthesis
Peter Kokol, Marko Kokol, Sašo Zagoranski
cs.LGcs.AIarXiv:2103.01002v12021TPR-Attention for Combinatorial Generalization
Melisa Civelekoğlu, Isabeau Prémont-Schwarz
cs.LGcs.AIarXiv:2608.30124v12026The Theory Behind Overfitting, Cross Validation, Regularization, Bagging, and Boosting: Tutorial
Benyamin Ghojogh, Mark Crowley
stat.MLcs.LGarXiv:1905.12787v22019Machine Learning for Microcontroller-Class Hardware: A Review
Swapnil Sayan Saha, Sandeep Singh Sandha, Mani Srivastava
cs.LGarXiv:2205.14550v52022The Cost of Training NLP Models: A Concise Overview
Or Sharir, Barak Peleg, Yoav Shoham
cs.CLcs.LGcs.NEarXiv:2004.08900v12020Benchmarking Foundation Models with Language-Model-as-an-Examiner
Yushi Bai, Jiahao Ying, Yixin Cao +10
cs.CLcs.LGarXiv:2306.04181v22023Differentiable, learnable, regionalized process-based models with physical outputs can approach state-of-the-art hydrologic prediction accuracy
Dapeng Feng, Jiangtao Liu, Kathryn Lawson +1
cs.LGarXiv:2203.14827v22022Randomized Dimensionality Reduction for k-means Clustering
Christos Boutsidis, Anastasios Zouzias, Michael W. Mahoney +1
cs.DScs.LGarXiv:1110.2897v32011Hurdles to Progress in Long-form Question Answering
Kalpesh Krishna, Aurko Roy, Mohit Iyyer
cs.CLcs.LGarXiv:2103.06332v22021Controlling Refusal Behavior of LLMs via Stiefel-Constrained Rotation Steering
Kirill Bunin, Dmitry Bylinkin, Vladimir Aletov +3
cs.LGcs.CLarXiv:2608.30986v12026Arcee's MergeKit: A Toolkit for Merging Large Language Models
Charles Goddard, Shamane Siriwardhana, Malikeh Ehghaghi +5
cs.CLcs.AIcs.LGarXiv:2403.13257v32024Supraglacial Lake Fate Is Knowable Long Before the Season Ends
Emam Hossain, Md Osman Gani
cs.LGarXiv:2608.30113v12026Global Guidance Network for Breast Lesion Segmentation in Ultrasound Images
Cheng Xue, Lei Zhu, Huazhu Fu +4
eess.IVcs.CVcs.LGarXiv:2104.01896v12021In-Context Impersonation Reveals Large Language Models' Strengths and Biases
Leonard Salewski, Stephan Alaniz, Isabel Rio-Torto +2
cs.AIcs.CLcs.LGarXiv:2305.14930v22023Graph4BiLO: Graph Neural Network Approximation for Bilevel Mixed-Integer Linear Optimization
Jessica D. Elrefaei, Kaixun Hua, Seungbae Kim +2
cs.LGcs.AIarXiv:2608.30103v12026A Neural Network Architecture Combining Gated Recurrent Unit (GRU) and Support Vector Machine (SVM) for Intrusion Detection in Network Traffic Data
Abien Fred Agarap
cs.NEcs.CRcs.LGarXiv:1709.03082v82017SMOTE-VAR: An Uncertainty-Aware Oversampling Method for Predicting Depression Remission in University Students
Dang Nguyen, Arun Kumar A, Taylor A. Braund +8
cs.LGarXiv:2608.30102v12026Learning Syntactic Program Transformations from Examples
Reudismam Rolim, Gustavo Soares, Loris D'Antoni +5
cs.SEcs.LGcs.PLarXiv:1608.09000v12016N-Gram Graph: Simple Unsupervised Representation for Graphs, with Applications to Molecules
Shengchao Liu, Mehmet Furkan Demirel, Yingyu Liang
cs.LGstat.MLarXiv:1806.09206v22018Spatially-Aware Graph Neural Networks for Relational Behavior Forecasting from Sensor Data
Sergio Casas, Cole Gulino, Renjie Liao +1
cs.CVcs.LGcs.ROarXiv:1910.08233v12019A Dependable Hybrid Machine Learning Model for Network Intrusion Detection
Md. Alamin Talukder, Khondokar Fida Hasan, Md. Manowarul Islam +5
cs.CRcs.LGarXiv:2212.04546v22022DEMO-Net: Degree-specific Graph Neural Networks for Node and Graph Classification
Jun Wu, Jingrui He, Jiejun Xu
cs.LGstat.MLarXiv:1906.02319v12019Learning to Balance Specificity and Invariance for In and Out of Domain Generalization
Prithvijit Chattopadhyay, Yogesh Balaji, Judy Hoffman
cs.CVcs.LGarXiv:2008.12839v12020A Model with No Head and Many Thoughts
Nikita Koriagin, Yaroslav Aksenov, George Bredis +3
cs.LGcs.CLarXiv:2608.31069v12026Implicit Neural Representations for Image Compression
Yannick Strümpler, Janis Postels, Ren Yang +2
eess.IVcs.CVcs.LGarXiv:2112.04267v22021Data2Vis: Automatic Generation of Data Visualizations Using Sequence to Sequence Recurrent Neural Networks
Victor Dibia, Çağatay Demiralp
cs.HCcs.AIcs.LGarXiv:1804.03126v32018Overcoming Catastrophic Forgetting with Unlabeled Data in the Wild
Kibok Lee, Kimin Lee, Jinwoo Shin +1
cs.CVcs.LGstat.MLarXiv:1903.12648v32019Adversarial Continual Learning
Sayna Ebrahimi, Franziska Meier, Roberto Calandra +2
cs.LGcs.AIcs.CVarXiv:2003.09553v22020Uniform Sampling for Matrix Approximation
Michael B. Cohen, Yin Tat Lee, Cameron Musco +3
cs.DScs.LGstat.MLarXiv:1408.5099v12014Communication-Efficient Algorithms for Decentralized and Stochastic Optimization
Guanghui Lan, Soomin Lee, Yi Zhou
math.OCcs.LGarXiv:1701.03961v22017AntisymmetricRNN: A Dynamical System View on Recurrent Neural Networks
Bo Chang, Minmin Chen, Eldad Haber +1
stat.MLcs.LGarXiv:1902.09689v12019E-Commerce Bench: Evaluating LLM Agents on Long-Horizon Autonomous Business Operation
Wei Fan, Xinjie Shen, Xudong Guo +8
cs.LGcs.CLarXiv:2608.30730v12026DeepHammer: Depleting the Intelligence of Deep Neural Networks through Targeted Chain of Bit Flips
Fan Yao, Adnan Siraj Rakin, Deliang Fan
cs.CRcs.LGarXiv:2003.13746v12020CHESS: Contextual Harnessing for Efficient SQL Synthesis
Shayan Talaei, Mohammadreza Pourreza, Yu-Chen Chang +2
cs.LGcs.AIcs.DBarXiv:2405.16755v32024PLC-DPO: Posterior Label Correction in Noisy and Ambiguous Preference Optimization
Boryeong Cho, Sumyeong Ahn, Se-Young Yun
cs.LGcs.CLarXiv:2608.30597v12026