Machine Learning

Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

4,201 to 4,260 of 20,308

  1. From 80x to 385x: A Best-Matching-Unit Search at the L2 Roof, Measured Against a Symmetrically Tuned Baseline

    Andrew James Amos

    cs.LGarXiv:2609.05138v12026
  2. Latent Multi-task Architecture Learning

    Sebastian Ruder, Joachim Bingel, Isabelle Augenstein +1

    stat.MLcs.AIcs.CLarXiv:1705.08142v32017
  3. State Backdoor: Towards Stealthy Real-world Poisoning Attack on Vision-Language-Action Model in State Space

    Ji Guo, Wenbo Jiang, Yansong Lin +6

    cs.CRcs.LGarXiv:2601.04266v22026
  4. Single-Query Black-Box Calibration Auditing via Logit Bias

    Roman Plaud, Antoine Saillenfest, Matthieu Labeau +2

    cs.LGarXiv:2609.05125v12026
  5. Quantum noise protects quantum classifiers against adversaries

    Yuxuan Du, Min-Hsiu Hsieh, Tongliang Liu +2

    quant-phcs.LGarXiv:2003.09416v12020
  6. Deep Microcompression: Structured Pruning and Bit-packed Quantization for Microcontrollers

    Opegbemi Matthias Busoye, Tolulope Matthew Busoye, Eghonghon-aye Eigbe

    cs.LGarXiv:2609.05081v12026
  7. GLASS: Graph-Language Alignment with Spherical Scoring for Transferable Graph-Level Anomaly Detection

    Xudong Wang, Chris Ding, Tongxin Li +1

    cs.LGarXiv:2609.05253v12026
  8. Hessian-based molecular conformation augmentation for a scalable and efficient strategy of machine learning interatomic potentials

    Bumju Kwak, Jeonghee Jo

    cs.LGphysics.chem-pharXiv:2609.05233v12026
  9. Listening while Speaking: Speech Chain by Deep Learning

    Andros Tjandra, Sakriani Sakti, Satoshi Nakamura

    cs.CLcs.LGcs.SDarXiv:1707.04879v12017
  10. Geoopt: Riemannian Optimization in PyTorch

    Max Kochurov, Rasul Karimov, Serge Kozlukov

    cs.CGcs.LGarXiv:2005.02819v52020
  11. Dimension-Adaptive Batched Lipschitz Narrowing Without Knowing the Zooming Dimension

    Yasong Feng

    cs.LGarXiv:2609.05214v12026
  12. Measure and Improve Robustness in NLP Models: A Survey

    Xuezhi Wang, Haohan Wang, Diyi Yang

    cs.CLcs.LGarXiv:2112.08313v22021
  13. Birth of a Transformer: A Memory Viewpoint

    Alberto Bietti, Vivien Cabannes, Diane Bouchacourt +2

    stat.MLcs.CLcs.LGarXiv:2306.00802v22023
  14. Feature Purification: How Adversarial Training Performs Robust Deep Learning

    Zeyuan Allen-Zhu, Yuanzhi Li

    cs.LGcs.NEmath.OCarXiv:2005.10190v42020
  15. ScienceAgentBench: Toward Rigorous Assessment of Language Agents for Data-Driven Scientific Discovery

    Ziru Chen, Shijie Chen, Yuting Ning +17

    cs.CLcs.AIcs.LGarXiv:2410.05080v32024
  16. Coarse-Graining Hidden Representations: Unsupervised Neuron Selection via Mapping Entropy

    Margherita Mele, Andrea Castagna, Roberto Menichetti +2

    cs.LGcond-mat.dis-nncond-mat.stat-mecharXiv:2609.05126v12026
  17. Sign-Based Optimizers Are Effective Under Heavy-Tailed Noise

    Dingzhi Yu, Hongyi Tao, Yuanyu Wan +2

    cs.LGcs.CLmath.OCarXiv:2602.07425v22026
  18. TriMap: Large-scale Dimensionality Reduction Using Triplets

    Ehsan Amid, Manfred K. Warmuth

    cs.LGstat.MLarXiv:1910.00204v22019
  19. A Comparative Study of Counterfactual Explainers for Graph Neural Networks Enabling Multiple Types of Graph Edit

    Maria Myrto Villia, Filippos Gouidis, Theodore Patkos +1

    cs.LGarXiv:2609.05113v12026
  20. Search-P1: Path-Centric Reward Shaping for Stable and Efficient Agentic RAG Training

    Tianle Xia, Ming Xu, Lingxiang Hu +7

    cs.CLcs.IRcs.LGarXiv:2602.22576v12026
  21. AgentSM: Semantic Memory for Agentic Text-to-SQL

    Asim Biswal, Chuan Lei, Xiao Qin +3

    cs.AIcs.DBcs.LGarXiv:2601.15709v12026
  22. Beyond Homoscedasticity: Decoupled Uncertainty Optimization for Deep Imbalanced Regression

    Juncheng Zhou, Jiaxi Lu, Weijing Zeng +3

    cs.LGarXiv:2609.04995v12026
  23. SimPLE: Similar Pseudo Label Exploitation for Semi-Supervised Classification

    Zijian Hu, Zhengyu Yang, Xuefeng Hu +1

    cs.CVcs.LGarXiv:2103.16725v22021
  24. From Memorization to Creativity: LLM as a Designer of Novel Neural Architectures

    Waleed Khalid, Dmitry Ignatov, Radu Timofte

    cs.LGcs.CVarXiv:2601.02997v22026
  25. Fractal basins trap latent reasoning

    Jeffrey Lai, Anthony Bao, John Quinn +1

    cs.LGarXiv:2609.04963v12026
  26. Neo-GNNs: Neighborhood Overlap-aware Graph Neural Networks for Link Prediction

    Seongjun Yun, Seoyoon Kim, Junhyun Lee +2

    cs.LGcs.AIarXiv:2206.04216v12022
  27. AudioMNIST: Exploring Explainable Artificial Intelligence for Audio Analysis on a Simple Benchmark

    Sören Becker, Johanna Vielhaben, Marcel Ackermann +3

    cs.SDcs.AIcs.LGarXiv:1807.03418v32018
  28. Federated Unlearning with Knowledge Distillation

    Chen Wu, Sencun Zhu, Prasenjit Mitra

    cs.LGcs.CRarXiv:2201.09441v12022
  29. From Deep to Shallow: Unconstrained and Efficient Layer Merging Strategy

    Petro Shulzhenko, Gabriele Spadaro, Enzo Tartaglione

    cs.LGarXiv:2609.04881v12026
  30. KVMem: Virtualizing Million-Token Agent Workspaces on a Consumer GPU

    Di Chai, Leye Wang, Zeshen Su +2

    cs.LGarXiv:2609.04852v12026
  31. Solution-space heterogeneity shapes federated learning dynamics across partial differential equations

    Ping Luo, Jiahuan Wang, Ziqing Wen +2

    cs.LGcs.DCarXiv:2609.05012v12026
  32. Pre-training LLM without Learning Rate Decay Enhances Supervised Fine-Tuning

    Kazuki Yano, Shun Kiyono, Sosuke Kobayashi +2

    cs.CLcs.LGarXiv:2603.16127v12026
  33. Confounding-Valid Conformal Inference for Counterfactual KPIs in Wireless Networks

    Abdessamed Qchohi, Jessica Moysen Cortes, Matteo Zecchin

    cs.LGcs.NIeess.SParXiv:2609.05073v12026
  34. Rewards-in-Context: Multi-objective Alignment of Foundation Models with Dynamic Preference Adjustment

    Rui Yang, Xiaoman Pan, Feng Luo +4

    cs.LGcs.AIcs.CLarXiv:2402.10207v62024
  35. Failure-Aware RL: Reliable Offline-to-Online Reinforcement Learning with Self-Recovery for Real-World Manipulation

    Huanyu Li, Kun Lei, Sheng Zang +5

    cs.ROcs.AIcs.LGarXiv:2601.07821v12026
  36. Physics-Aware Random Walk Fingerprints for Scalable Power Grid Graph Classification

    Adnan Anwar

    cs.LGeess.SYarXiv:2609.04943v12026
  37. Fast Gauss Sums via Flash Attention

    Nicolaj Rux, Sebastian Neumayer

    cs.LGmath.NAarXiv:2609.04910v12026
  38. Adaptive Milestone Reward for GUI Agents

    Congmin Zheng, Xiaoyun Mo, Xinbei Ma +10

    cs.LGcs.AIcs.CLarXiv:2602.11524v12026
  39. SpecForge: A Flexible and Efficient Open-Source Training Framework for Speculative Decoding

    Shenggui Li, Chao Wang, Yikai Zhu +14

    cs.LGcs.AIcs.CLarXiv:2603.18567v12026
  40. ML-Decoder: Scalable and Versatile Classification Head

    Tal Ridnik, Gilad Sharir, Avi Ben-Cohen +2

    cs.CVcs.LGarXiv:2111.12933v22021
  41. When Genomic Masking Priors Fail to Transfer: Strong Variant Prediction, Weak Functional Generation

    Susu Hu, Preetam Gattogi, Jens Lehmann +3

    cs.LGarXiv:2609.04861v12026
  42. Demystifying Data-Driven Probabilistic Medium-Range Weather Forecasting

    Jean Kossaifi, Nikola Kovachki, Morteza Mardani +15

    cs.LGcs.AIarXiv:2601.18111v12026
  43. Hyperbolic Vision Transformers: Combining Improvements in Metric Learning

    Aleksandr Ermolov, Leyla Mirvakhabova, Valentin Khrulkov +2

    cs.CVcs.LGarXiv:2203.10833v22022
  44. PACE: Propagation-Aware Collaborative Correction for One-Shot Personalized Federated Graph Learning

    Ruizhe Huang, Chengran Li, Xiaochuan Shi

    cs.LGarXiv:2609.04832v12026
  45. Semi-parametric Image Synthesis

    Xiaojuan Qi, Qifeng Chen, Jiaya Jia +1

    cs.CVcs.AIcs.GRarXiv:1804.10992v12018
  46. Deep Partition Aggregation: Provable Defense against General Poisoning Attacks

    Alexander Levine, Soheil Feizi

    cs.LGstat.MLarXiv:2006.14768v22020
  47. Locating and Steering Refusal Beyond Attention

    Preethi Carmel Bosco, Gopalakrishnan Srinivasan

    cs.LGarXiv:2609.04721v12026
  48. Adaptivity of averaged stochastic gradient descent to local strong convexity for logistic regression

    Francis Bach

    math.STcs.LGmath.OCarXiv:1303.6149v32013
  49. Learning-Augmented Algorithms: Guarantees, Construction Mechanisms, and System-Level Implications

    Hailiang Zhao, Peng Chen, Xueyan Tang +2

    cs.LGcs.DSarXiv:2609.04787v12026
  50. A Second-order Bound with Excess Losses

    Pierre Gaillard, Gilles Stoltz, Tim Van Erven

    stat.MLcs.LGmath.STarXiv:1402.2044v12014
  51. Model-Based Deep Learning: On the Intersection of Deep Learning and Optimization

    Nir Shlezinger, Yonina C. Eldar, Stephen P. Boyd

    eess.SPcs.LGeess.SYarXiv:2205.02640v22022
  52. Communication-Efficient Personalized Federated Learning via Layer-Wise Multi-Threshold Random Sketching

    Xu Zhang, Xingyu Hou, Jiacheng Cheng +2

    cs.LGarXiv:2609.04830v12026
  53. Iterative Refinement Graph Neural Network for Antibody Sequence-Structure Co-design

    Wengong Jin, Jeremy Wohlwend, Regina Barzilay +1

    q-bio.BMcs.LGarXiv:2110.04624v32021
  54. Federated Attack Campaign Detection via Contrastive Encoding of Threat Indicators in Gradient Updates

    Manuel Röder, Bibin Babu, Frank-Michael Schleif

    cs.LGcs.CRarXiv:2609.04815v12026
  55. Compositionality and Generalization in Emergent Languages

    Rahma Chaabouni, Eugene Kharitonov, Diane Bouchacourt +2

    cs.CLcs.AIcs.LGarXiv:2004.09124v12020
  56. How Faithful Is Attribution for Sales Forecasting? A Counterfactual Study

    Glib Kechyn

    cs.LGarXiv:2609.04797v12026
  57. Countdown-Code: A Testbed for Studying The Emergence and Generalization of Reward Hacking in RLVR

    Muhammad Khalifa, Zohaib Khan, Omer Tafveez +2

    cs.LGcs.AIcs.CLarXiv:2603.07084v22026
  58. A Robust Watermark-based Fingerprint Framework for GNNs Ownership Verification

    Han Zhang, Yan Wang, Guanfeng Liu +3

    cs.LGarXiv:2609.04772v12026
  59. Resilience Beyond Stationary Client Unavailability: Unlocking Efficient and Unbiased Federated Learning

    Ming Xiang, Stratis Ioannidis, Edmund Yeh +2

    cs.LGcs.DCmath.OCarXiv:2609.04763v12026
  60. Expandable Subspace Ensemble for Pre-Trained Model-Based Class-Incremental Learning

    Da-Wei Zhou, Hai-Long Sun, Han-Jia Ye +1

    cs.CVcs.LGarXiv:2403.12030v12024