Machine Learning

Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

9,781 to 9,840 of 19,974

  1. Learning PDE Time-Stepping with Neural Cellular Automata

    Esha Saha, Hao Wang

    cs.LGstat.MLarXiv:2608.30328v12026
  2. Generative multi-domain transfer learning for fault detection in data-scarce wind turbines

    Stefan Jonas, Angela Meyer

    cs.LGarXiv:2608.30323v12026
  3. Scalable Kernel Methods via Doubly Stochastic Gradients

    Bo Dai, Bo Xie, Niao He +4

    cs.LGstat.MLarXiv:1407.5599v42014
  4. Online Estimation of Dynamic Origin-Destination Matrices Using Reinforcement Learning with Link-Flow Propagation Guidance

    Donggyu Min, Dong-Kyu Kim

    cs.LGcs.AIarXiv:2608.30317v12026
  5. Non-Vacuous Generalization Bounds at the ImageNet Scale: A PAC-Bayesian Compression Approach

    Wenda Zhou, Victor Veitch, Morgane Austern +2

    stat.MLcs.LGarXiv:1804.05862v32018
  6. Tail-Replay: Escaping the Curse of Linear Attention in Prefix Caching for Hybrid LLMs

    Yirui Liu, Ruoling Qi, Xuaner Wu +2

    cs.LGcs.AIarXiv:2608.30310v12026
  7. Data Augmentation Approaches in Natural Language Processing: A Survey

    Bohan Li, Yutai Hou, Wanxiang Che

    cs.CLcs.AIcs.LGarXiv:2110.01852v32021
  8. BCPPO: Bachelier-Inspired Constrained Proximal Policy Optimization for Tail-Risk-Aware Safe Reinforcement Learning

    Dongsheng Hou, Yanqiao Chen, Yuhan Rui

    cs.LGcs.AIarXiv:2608.30283v12026
  9. Partial Sum Minimization of Singular Values in Robust PCA: Algorithm and Applications

    Tae-Hyun Oh, Yu-Wing Tai, Jean-Charles Bazin +2

    cs.CVcs.AIcs.LGarXiv:1503.01444v22015
  10. Multivariate Scientific Data Compression with Learned Cross-Variable Latent Decorrelation and Autoregressive Entropy Modeling

    Liangji Zhu, Anand Rangarajan, Sanjay Ranka

    cs.LGarXiv:2608.30262v12026
  11. Exact Recovery Thresholds for Weighted Data Selection in Vector-Valued Linear Regression

    Guangjian Zhang

    cs.LGmath.STarXiv:2608.30254v12026
  12. Certified Safety Radii in Forecast-Error Space for Wasserstein Distributionally Robust Small Signal Stability-Constrained AC Optimal Power Flow via Lifted Spectrahedral Containment

    Ziqi Zhang, Xi Chen

    cs.LGarXiv:2608.30201v12026
  13. SneakyPrompt: Jailbreaking Text-to-image Generative Models

    Yuchen Yang, Bo Hui, Haolin Yuan +2

    cs.LGarXiv:2305.12082v32023
  14. Reinforcement Learning for Symbolic Equation Solving

    Kevin P O Keeffe

    cs.LGarXiv:2608.30162v12026
  15. Benchmarking Peptide-Protein Affinity Prediction Across Peptide and Target Shifts

    Jiaxin Tian, Darren An, Jun Li

    cs.LGq-bio.QMarXiv:2608.30175v12026
  16. Converse and Collision-Based Achievability for Node Localization with Hybrid Distance-Spectral Graph Positional Encodings

    Zimo Yan, Yifan Li, Hao Li +4

    cs.LGarXiv:2608.30152v12026
  17. Machine Learning Enabled Computational Screening of Inorganic Solid Electrolytes for Dendrite Suppression with Li Metal Anode

    Zeeshan Ahmad, Tian Xie, Chinmay Maheshwari +2

    cond-mat.mtrl-scics.LGphysics.chem-pharXiv:1804.04651v12018
  18. Error bounds for approximations with deep ReLU neural networks in $W^{s,p}$ norms

    Ingo Gühring, Gitta Kutyniok, Philipp Petersen

    math.FAcs.LGarXiv:1902.07896v12019
  19. Learning Deep Disentangled Embeddings with the F-Statistic Loss

    Karl Ridgeway, Michael C. Mozer

    cs.LGcs.AIstat.MLarXiv:1802.05312v22018
  20. Strong Drafts Need Compact Memories: Long-Context Speculative Decoding with Compressed KV Cache

    Tong Yuan, Chengxi Liao, Zeyi Wen

    cs.LGarXiv:2608.30252v12026
  21. The Statistical Complexity of Interactive Decision Making

    Dylan J. Foster, Sham M. Kakade, Jian Qian +1

    cs.LGmath.OCmath.STarXiv:2112.13487v32021
  22. HPLFlowNet: Hierarchical Permutohedral Lattice FlowNet for Scene Flow Estimation on Large-scale Point Clouds

    Xiuye Gu, Yijie Wang, Chongruo wu +2

    cs.CVcs.LGeess.IVarXiv:1906.05332v12019
  23. Mask2Former for Video Instance Segmentation

    Bowen Cheng, Anwesa Choudhuri, Ishan Misra +3

    cs.CVcs.AIcs.LGarXiv:2112.10764v12021
  24. You are AllSet: A Multiset Function Framework for Hypergraph Neural Networks

    Eli Chien, Chao Pan, Jianhao Peng +1

    cs.LGcs.AIarXiv:2106.13264v42021
  25. Optimal Scheduling of Isolated Microgrids Using Automated Reinforcement Learning-based Multi-period Forecasting

    Yang Li, Ruinong Wang, Zhen Yang

    eess.SPcs.LGeess.SYarXiv:2108.06764v12021
  26. Machine learning on small size samples: A synthetic knowledge synthesis

    Peter Kokol, Marko Kokol, Sašo Zagoranski

    cs.LGcs.AIarXiv:2103.01002v12021
  27. TPR-Attention for Combinatorial Generalization

    Melisa Civelekoğlu, Isabeau Prémont-Schwarz

    cs.LGcs.AIarXiv:2608.30124v12026
  28. The Theory Behind Overfitting, Cross Validation, Regularization, Bagging, and Boosting: Tutorial

    Benyamin Ghojogh, Mark Crowley

    stat.MLcs.LGarXiv:1905.12787v22019
  29. Machine Learning for Microcontroller-Class Hardware: A Review

    Swapnil Sayan Saha, Sandeep Singh Sandha, Mani Srivastava

    cs.LGarXiv:2205.14550v52022
  30. The Cost of Training NLP Models: A Concise Overview

    Or Sharir, Barak Peleg, Yoav Shoham

    cs.CLcs.LGcs.NEarXiv:2004.08900v12020
  31. Benchmarking Foundation Models with Language-Model-as-an-Examiner

    Yushi Bai, Jiahao Ying, Yixin Cao +10

    cs.CLcs.LGarXiv:2306.04181v22023
  32. Differentiable, learnable, regionalized process-based models with physical outputs can approach state-of-the-art hydrologic prediction accuracy

    Dapeng Feng, Jiangtao Liu, Kathryn Lawson +1

    cs.LGarXiv:2203.14827v22022
  33. Randomized Dimensionality Reduction for k-means Clustering

    Christos Boutsidis, Anastasios Zouzias, Michael W. Mahoney +1

    cs.DScs.LGarXiv:1110.2897v32011
  34. Hurdles to Progress in Long-form Question Answering

    Kalpesh Krishna, Aurko Roy, Mohit Iyyer

    cs.CLcs.LGarXiv:2103.06332v22021
  35. Controlling Refusal Behavior of LLMs via Stiefel-Constrained Rotation Steering

    Kirill Bunin, Dmitry Bylinkin, Vladimir Aletov +3

    cs.LGcs.CLarXiv:2608.30986v12026
  36. Arcee's MergeKit: A Toolkit for Merging Large Language Models

    Charles Goddard, Shamane Siriwardhana, Malikeh Ehghaghi +5

    cs.CLcs.AIcs.LGarXiv:2403.13257v32024
  37. Supraglacial Lake Fate Is Knowable Long Before the Season Ends

    Emam Hossain, Md Osman Gani

    cs.LGarXiv:2608.30113v12026
  38. Global Guidance Network for Breast Lesion Segmentation in Ultrasound Images

    Cheng Xue, Lei Zhu, Huazhu Fu +4

    eess.IVcs.CVcs.LGarXiv:2104.01896v12021
  39. In-Context Impersonation Reveals Large Language Models' Strengths and Biases

    Leonard Salewski, Stephan Alaniz, Isabel Rio-Torto +2

    cs.AIcs.CLcs.LGarXiv:2305.14930v22023
  40. Graph4BiLO: Graph Neural Network Approximation for Bilevel Mixed-Integer Linear Optimization

    Jessica D. Elrefaei, Kaixun Hua, Seungbae Kim +2

    cs.LGcs.AIarXiv:2608.30103v12026
  41. A Neural Network Architecture Combining Gated Recurrent Unit (GRU) and Support Vector Machine (SVM) for Intrusion Detection in Network Traffic Data

    Abien Fred Agarap

    cs.NEcs.CRcs.LGarXiv:1709.03082v82017
  42. SMOTE-VAR: An Uncertainty-Aware Oversampling Method for Predicting Depression Remission in University Students

    Dang Nguyen, Arun Kumar A, Taylor A. Braund +8

    cs.LGarXiv:2608.30102v12026
  43. Learning Syntactic Program Transformations from Examples

    Reudismam Rolim, Gustavo Soares, Loris D'Antoni +5

    cs.SEcs.LGcs.PLarXiv:1608.09000v12016
  44. N-Gram Graph: Simple Unsupervised Representation for Graphs, with Applications to Molecules

    Shengchao Liu, Mehmet Furkan Demirel, Yingyu Liang

    cs.LGstat.MLarXiv:1806.09206v22018
  45. Spatially-Aware Graph Neural Networks for Relational Behavior Forecasting from Sensor Data

    Sergio Casas, Cole Gulino, Renjie Liao +1

    cs.CVcs.LGcs.ROarXiv:1910.08233v12019
  46. A Dependable Hybrid Machine Learning Model for Network Intrusion Detection

    Md. Alamin Talukder, Khondokar Fida Hasan, Md. Manowarul Islam +5

    cs.CRcs.LGarXiv:2212.04546v22022
  47. DEMO-Net: Degree-specific Graph Neural Networks for Node and Graph Classification

    Jun Wu, Jingrui He, Jiejun Xu

    cs.LGstat.MLarXiv:1906.02319v12019
  48. Learning to Balance Specificity and Invariance for In and Out of Domain Generalization

    Prithvijit Chattopadhyay, Yogesh Balaji, Judy Hoffman

    cs.CVcs.LGarXiv:2008.12839v12020
  49. A Model with No Head and Many Thoughts

    Nikita Koriagin, Yaroslav Aksenov, George Bredis +3

    cs.LGcs.CLarXiv:2608.31069v12026
  50. Implicit Neural Representations for Image Compression

    Yannick Strümpler, Janis Postels, Ren Yang +2

    eess.IVcs.CVcs.LGarXiv:2112.04267v22021
  51. Data2Vis: Automatic Generation of Data Visualizations Using Sequence to Sequence Recurrent Neural Networks

    Victor Dibia, Çağatay Demiralp

    cs.HCcs.AIcs.LGarXiv:1804.03126v32018
  52. Overcoming Catastrophic Forgetting with Unlabeled Data in the Wild

    Kibok Lee, Kimin Lee, Jinwoo Shin +1

    cs.CVcs.LGstat.MLarXiv:1903.12648v32019
  53. Adversarial Continual Learning

    Sayna Ebrahimi, Franziska Meier, Roberto Calandra +2

    cs.LGcs.AIcs.CVarXiv:2003.09553v22020
  54. Uniform Sampling for Matrix Approximation

    Michael B. Cohen, Yin Tat Lee, Cameron Musco +3

    cs.DScs.LGstat.MLarXiv:1408.5099v12014
  55. Communication-Efficient Algorithms for Decentralized and Stochastic Optimization

    Guanghui Lan, Soomin Lee, Yi Zhou

    math.OCcs.LGarXiv:1701.03961v22017
  56. AntisymmetricRNN: A Dynamical System View on Recurrent Neural Networks

    Bo Chang, Minmin Chen, Eldad Haber +1

    stat.MLcs.LGarXiv:1902.09689v12019
  57. E-Commerce Bench: Evaluating LLM Agents on Long-Horizon Autonomous Business Operation

    Wei Fan, Xinjie Shen, Xudong Guo +8

    cs.LGcs.CLarXiv:2608.30730v12026
  58. DeepHammer: Depleting the Intelligence of Deep Neural Networks through Targeted Chain of Bit Flips

    Fan Yao, Adnan Siraj Rakin, Deliang Fan

    cs.CRcs.LGarXiv:2003.13746v12020
  59. CHESS: Contextual Harnessing for Efficient SQL Synthesis

    Shayan Talaei, Mohammadreza Pourreza, Yu-Chen Chang +2

    cs.LGcs.AIcs.DBarXiv:2405.16755v32024
  60. PLC-DPO: Posterior Label Correction in Noisy and Ambiguous Preference Optimization

    Boryeong Cho, Sumyeong Ahn, Se-Young Yun

    cs.LGcs.CLarXiv:2608.30597v12026