Machine Learning
Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
4,381 to 4,440 of 20,454
A Comparative Study of Counterfactual Explainers for Graph Neural Networks Enabling Multiple Types of Graph Edit
Maria Myrto Villia, Filippos Gouidis, Theodore Patkos +1
cs.LGarXiv:2609.05113v12026Search-P1: Path-Centric Reward Shaping for Stable and Efficient Agentic RAG Training
Tianle Xia, Ming Xu, Lingxiang Hu +7
cs.CLcs.IRcs.LGarXiv:2602.22576v12026AgentSM: Semantic Memory for Agentic Text-to-SQL
Asim Biswal, Chuan Lei, Xiao Qin +3
cs.AIcs.DBcs.LGarXiv:2601.15709v12026Beyond Homoscedasticity: Decoupled Uncertainty Optimization for Deep Imbalanced Regression
Juncheng Zhou, Jiaxi Lu, Weijing Zeng +3
cs.LGarXiv:2609.04995v12026SimPLE: Similar Pseudo Label Exploitation for Semi-Supervised Classification
Zijian Hu, Zhengyu Yang, Xuefeng Hu +1
cs.CVcs.LGarXiv:2103.16725v22021From Memorization to Creativity: LLM as a Designer of Novel Neural Architectures
Waleed Khalid, Dmitry Ignatov, Radu Timofte
cs.LGcs.CVarXiv:2601.02997v22026Fractal basins trap latent reasoning
Jeffrey Lai, Anthony Bao, John Quinn +1
cs.LGarXiv:2609.04963v12026Neo-GNNs: Neighborhood Overlap-aware Graph Neural Networks for Link Prediction
Seongjun Yun, Seoyoon Kim, Junhyun Lee +2
cs.LGcs.AIarXiv:2206.04216v12022AudioMNIST: Exploring Explainable Artificial Intelligence for Audio Analysis on a Simple Benchmark
Sören Becker, Johanna Vielhaben, Marcel Ackermann +3
cs.SDcs.AIcs.LGarXiv:1807.03418v32018Federated Unlearning with Knowledge Distillation
Chen Wu, Sencun Zhu, Prasenjit Mitra
cs.LGcs.CRarXiv:2201.09441v12022From Deep to Shallow: Unconstrained and Efficient Layer Merging Strategy
Petro Shulzhenko, Gabriele Spadaro, Enzo Tartaglione
cs.LGarXiv:2609.04881v12026KVMem: Virtualizing Million-Token Agent Workspaces on a Consumer GPU
Di Chai, Leye Wang, Zeshen Su +2
cs.LGarXiv:2609.04852v12026Solution-space heterogeneity shapes federated learning dynamics across partial differential equations
Ping Luo, Jiahuan Wang, Ziqing Wen +2
cs.LGcs.DCarXiv:2609.05012v12026Pre-training LLM without Learning Rate Decay Enhances Supervised Fine-Tuning
Kazuki Yano, Shun Kiyono, Sosuke Kobayashi +2
cs.CLcs.LGarXiv:2603.16127v12026Confounding-Valid Conformal Inference for Counterfactual KPIs in Wireless Networks
Abdessamed Qchohi, Jessica Moysen Cortes, Matteo Zecchin
cs.LGcs.NIeess.SParXiv:2609.05073v12026Rewards-in-Context: Multi-objective Alignment of Foundation Models with Dynamic Preference Adjustment
Rui Yang, Xiaoman Pan, Feng Luo +4
cs.LGcs.AIcs.CLarXiv:2402.10207v62024Failure-Aware RL: Reliable Offline-to-Online Reinforcement Learning with Self-Recovery for Real-World Manipulation
Huanyu Li, Kun Lei, Sheng Zang +5
cs.ROcs.AIcs.LGarXiv:2601.07821v12026Physics-Aware Random Walk Fingerprints for Scalable Power Grid Graph Classification
Adnan Anwar
cs.LGeess.SYarXiv:2609.04943v12026Fast Gauss Sums via Flash Attention
Nicolaj Rux, Sebastian Neumayer
cs.LGmath.NAarXiv:2609.04910v12026Adaptive Milestone Reward for GUI Agents
Congmin Zheng, Xiaoyun Mo, Xinbei Ma +10
cs.LGcs.AIcs.CLarXiv:2602.11524v12026SpecForge: A Flexible and Efficient Open-Source Training Framework for Speculative Decoding
Shenggui Li, Chao Wang, Yikai Zhu +14
cs.LGcs.AIcs.CLarXiv:2603.18567v12026ML-Decoder: Scalable and Versatile Classification Head
Tal Ridnik, Gilad Sharir, Avi Ben-Cohen +2
cs.CVcs.LGarXiv:2111.12933v22021When Genomic Masking Priors Fail to Transfer: Strong Variant Prediction, Weak Functional Generation
Susu Hu, Preetam Gattogi, Jens Lehmann +3
cs.LGarXiv:2609.04861v12026Demystifying Data-Driven Probabilistic Medium-Range Weather Forecasting
Jean Kossaifi, Nikola Kovachki, Morteza Mardani +15
cs.LGcs.AIarXiv:2601.18111v12026Hyperbolic Vision Transformers: Combining Improvements in Metric Learning
Aleksandr Ermolov, Leyla Mirvakhabova, Valentin Khrulkov +2
cs.CVcs.LGarXiv:2203.10833v22022PACE: Propagation-Aware Collaborative Correction for One-Shot Personalized Federated Graph Learning
Ruizhe Huang, Chengran Li, Xiaochuan Shi
cs.LGarXiv:2609.04832v12026Semi-parametric Image Synthesis
Xiaojuan Qi, Qifeng Chen, Jiaya Jia +1
cs.CVcs.AIcs.GRarXiv:1804.10992v12018Deep Partition Aggregation: Provable Defense against General Poisoning Attacks
Alexander Levine, Soheil Feizi
cs.LGstat.MLarXiv:2006.14768v22020Locating and Steering Refusal Beyond Attention
Preethi Carmel Bosco, Gopalakrishnan Srinivasan
cs.LGarXiv:2609.04721v12026Adaptivity of averaged stochastic gradient descent to local strong convexity for logistic regression
Francis Bach
math.STcs.LGmath.OCarXiv:1303.6149v32013Learning-Augmented Algorithms: Guarantees, Construction Mechanisms, and System-Level Implications
Hailiang Zhao, Peng Chen, Xueyan Tang +2
cs.LGcs.DSarXiv:2609.04787v12026A Second-order Bound with Excess Losses
Pierre Gaillard, Gilles Stoltz, Tim Van Erven
stat.MLcs.LGmath.STarXiv:1402.2044v12014Model-Based Deep Learning: On the Intersection of Deep Learning and Optimization
Nir Shlezinger, Yonina C. Eldar, Stephen P. Boyd
eess.SPcs.LGeess.SYarXiv:2205.02640v22022Communication-Efficient Personalized Federated Learning via Layer-Wise Multi-Threshold Random Sketching
Xu Zhang, Xingyu Hou, Jiacheng Cheng +2
cs.LGarXiv:2609.04830v12026Iterative Refinement Graph Neural Network for Antibody Sequence-Structure Co-design
Wengong Jin, Jeremy Wohlwend, Regina Barzilay +1
q-bio.BMcs.LGarXiv:2110.04624v32021Federated Attack Campaign Detection via Contrastive Encoding of Threat Indicators in Gradient Updates
Manuel Röder, Bibin Babu, Frank-Michael Schleif
cs.LGcs.CRarXiv:2609.04815v12026Compositionality and Generalization in Emergent Languages
Rahma Chaabouni, Eugene Kharitonov, Diane Bouchacourt +2
cs.CLcs.AIcs.LGarXiv:2004.09124v12020How Faithful Is Attribution for Sales Forecasting? A Counterfactual Study
Glib Kechyn
cs.LGarXiv:2609.04797v12026Countdown-Code: A Testbed for Studying The Emergence and Generalization of Reward Hacking in RLVR
Muhammad Khalifa, Zohaib Khan, Omer Tafveez +2
cs.LGcs.AIcs.CLarXiv:2603.07084v22026A Robust Watermark-based Fingerprint Framework for GNNs Ownership Verification
Han Zhang, Yan Wang, Guanfeng Liu +3
cs.LGarXiv:2609.04772v12026Resilience Beyond Stationary Client Unavailability: Unlocking Efficient and Unbiased Federated Learning
Ming Xiang, Stratis Ioannidis, Edmund Yeh +2
cs.LGcs.DCmath.OCarXiv:2609.04763v12026Expandable Subspace Ensemble for Pre-Trained Model-Based Class-Incremental Learning
Da-Wei Zhou, Hai-Long Sun, Han-Jia Ye +1
cs.CVcs.LGarXiv:2403.12030v12024ATOM3D: Tasks On Molecules in Three Dimensions
Raphael J. L. Townshend, Martin Vögele, Patricia Suriana +10
cs.LGphysics.bio-phphysics.comp-pharXiv:2012.04035v42020Tight (Lower) Bounds for the Fixed Budget Best Arm Identification Bandit Problem
Alexandra Carpentier, Andrea Locatelli
stat.MLcs.LGarXiv:1605.09004v12016Semantic Image Inversion and Editing using Rectified Stochastic Differential Equations
Litu Rout, Yujia Chen, Nataniel Ruiz +3
cs.LGcs.CVstat.MLarXiv:2410.10792v12024A Fairness Audit of the Duckworth-Lewis-Stern Method: Format-Specific and Gender-Differential Bias, with an Interpretable Calibration Layer for Cricket Target Revision
Soumyadeep Roy
cs.LGarXiv:2609.04754v12026When Benchmarks Lie: Evaluating Malicious Prompt Classifiers Under True Distribution Shift
Max Fomin
cs.LGarXiv:2602.14161v22026SUMBT: Slot-Utterance Matching for Universal and Scalable Belief Tracking
Hwaran Lee, Jinsik Lee, Tae-Yoon Kim
cs.CLcs.LGarXiv:1907.07421v12019Open World Compositional Zero-Shot Learning
Massimiliano Mancini, Muhammad Ferjad Naeem, Yongqin Xian +1
cs.CVcs.LGarXiv:2101.12609v32021Training Large Language Models for Small-Molecule Design with Synthetic Task Scaling
Frank Hu, Shriram Chennakesavalu, Zichen Wang +5
cs.LGarXiv:2609.04735v12026WEECFP-SuRGE: Wide Embedded Extended Connectivity Fingerprint with Substructure Rotary Graph-distance Encoding
Robert Epps
cs.LGarXiv:2609.04672v12026Limits of End-to-End Learning
Tobias Glasmachers
cs.LGstat.MLarXiv:1704.08305v12017Interpretability for Turing Machines
Billy Snikkers, Rumi Salazar, Daniel Murfet +1
cs.LGcs.FLstat.MLarXiv:2609.04661v12026Optimizer Memory Schedules for Outscaling the Overtraining Axis
Katie Everett, Shikai Qiu
cs.LGarXiv:2609.04577v12026R$^3$L: Reflect-then-Retry Reinforcement Learning with Language-Guided Exploration, Pivotal Credit, and Positive Amplification
Weijie Shi, Yanxi Chen, Zexi Li +5
cs.LGcs.AIarXiv:2601.03715v22026SMILE: Bridging Continuous Optimization and Discrete Symbolic Recovery
Mansooreh Montazerin, Antonio Ortega, Ajitesh Srivastava
cs.LGarXiv:2609.04639v12026Current Agents Fail to Leverage World Model as Tool for Foresight
Cheng Qian, Emre Can Acikgoz, Bingxuan Li +8
cs.AIcs.CLcs.LGarXiv:2601.03905v22026Too Rare to Learn: Prescribed Cyclone Tracks Degrade a Bay of Bengal Ocean Emulator
Sumaiya Islam
cs.LGarXiv:2609.04635v12026Artificial Intelligence in Drug Discovery: Applications and Techniques
Jianyuan Deng, Zhibo Yang, Iwao Ojima +2
cs.LGcs.AIarXiv:2106.05386v42021GNN-Guided Graph Coarsening and Adaptive QUBO Penalties for the Capacitated Vehicle Routing Problem with Time Windows on a Quantum Annealer
Youssef Kamel Rezk, Paweł Gora
cs.LGarXiv:2609.04593v12026