Machine Learning
Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
4,201 to 4,260 of 20,308
From 80x to 385x: A Best-Matching-Unit Search at the L2 Roof, Measured Against a Symmetrically Tuned Baseline
Andrew James Amos
cs.LGarXiv:2609.05138v12026Latent Multi-task Architecture Learning
Sebastian Ruder, Joachim Bingel, Isabelle Augenstein +1
stat.MLcs.AIcs.CLarXiv:1705.08142v32017State Backdoor: Towards Stealthy Real-world Poisoning Attack on Vision-Language-Action Model in State Space
Ji Guo, Wenbo Jiang, Yansong Lin +6
cs.CRcs.LGarXiv:2601.04266v22026Single-Query Black-Box Calibration Auditing via Logit Bias
Roman Plaud, Antoine Saillenfest, Matthieu Labeau +2
cs.LGarXiv:2609.05125v12026Quantum noise protects quantum classifiers against adversaries
Yuxuan Du, Min-Hsiu Hsieh, Tongliang Liu +2
quant-phcs.LGarXiv:2003.09416v12020Deep Microcompression: Structured Pruning and Bit-packed Quantization for Microcontrollers
Opegbemi Matthias Busoye, Tolulope Matthew Busoye, Eghonghon-aye Eigbe
cs.LGarXiv:2609.05081v12026GLASS: Graph-Language Alignment with Spherical Scoring for Transferable Graph-Level Anomaly Detection
Xudong Wang, Chris Ding, Tongxin Li +1
cs.LGarXiv:2609.05253v12026Hessian-based molecular conformation augmentation for a scalable and efficient strategy of machine learning interatomic potentials
Bumju Kwak, Jeonghee Jo
cs.LGphysics.chem-pharXiv:2609.05233v12026Listening while Speaking: Speech Chain by Deep Learning
Andros Tjandra, Sakriani Sakti, Satoshi Nakamura
cs.CLcs.LGcs.SDarXiv:1707.04879v12017Geoopt: Riemannian Optimization in PyTorch
Max Kochurov, Rasul Karimov, Serge Kozlukov
cs.CGcs.LGarXiv:2005.02819v52020Dimension-Adaptive Batched Lipschitz Narrowing Without Knowing the Zooming Dimension
Yasong Feng
cs.LGarXiv:2609.05214v12026Measure and Improve Robustness in NLP Models: A Survey
Xuezhi Wang, Haohan Wang, Diyi Yang
cs.CLcs.LGarXiv:2112.08313v22021Birth of a Transformer: A Memory Viewpoint
Alberto Bietti, Vivien Cabannes, Diane Bouchacourt +2
stat.MLcs.CLcs.LGarXiv:2306.00802v22023Feature Purification: How Adversarial Training Performs Robust Deep Learning
Zeyuan Allen-Zhu, Yuanzhi Li
cs.LGcs.NEmath.OCarXiv:2005.10190v42020ScienceAgentBench: Toward Rigorous Assessment of Language Agents for Data-Driven Scientific Discovery
Ziru Chen, Shijie Chen, Yuting Ning +17
cs.CLcs.AIcs.LGarXiv:2410.05080v32024Coarse-Graining Hidden Representations: Unsupervised Neuron Selection via Mapping Entropy
Margherita Mele, Andrea Castagna, Roberto Menichetti +2
cs.LGcond-mat.dis-nncond-mat.stat-mecharXiv:2609.05126v12026Sign-Based Optimizers Are Effective Under Heavy-Tailed Noise
Dingzhi Yu, Hongyi Tao, Yuanyu Wan +2
cs.LGcs.CLmath.OCarXiv:2602.07425v22026TriMap: Large-scale Dimensionality Reduction Using Triplets
Ehsan Amid, Manfred K. Warmuth
cs.LGstat.MLarXiv:1910.00204v22019A Comparative Study of Counterfactual Explainers for Graph Neural Networks Enabling Multiple Types of Graph Edit
Maria Myrto Villia, Filippos Gouidis, Theodore Patkos +1
cs.LGarXiv:2609.05113v12026Search-P1: Path-Centric Reward Shaping for Stable and Efficient Agentic RAG Training
Tianle Xia, Ming Xu, Lingxiang Hu +7
cs.CLcs.IRcs.LGarXiv:2602.22576v12026AgentSM: Semantic Memory for Agentic Text-to-SQL
Asim Biswal, Chuan Lei, Xiao Qin +3
cs.AIcs.DBcs.LGarXiv:2601.15709v12026Beyond Homoscedasticity: Decoupled Uncertainty Optimization for Deep Imbalanced Regression
Juncheng Zhou, Jiaxi Lu, Weijing Zeng +3
cs.LGarXiv:2609.04995v12026SimPLE: Similar Pseudo Label Exploitation for Semi-Supervised Classification
Zijian Hu, Zhengyu Yang, Xuefeng Hu +1
cs.CVcs.LGarXiv:2103.16725v22021From Memorization to Creativity: LLM as a Designer of Novel Neural Architectures
Waleed Khalid, Dmitry Ignatov, Radu Timofte
cs.LGcs.CVarXiv:2601.02997v22026Fractal basins trap latent reasoning
Jeffrey Lai, Anthony Bao, John Quinn +1
cs.LGarXiv:2609.04963v12026Neo-GNNs: Neighborhood Overlap-aware Graph Neural Networks for Link Prediction
Seongjun Yun, Seoyoon Kim, Junhyun Lee +2
cs.LGcs.AIarXiv:2206.04216v12022AudioMNIST: Exploring Explainable Artificial Intelligence for Audio Analysis on a Simple Benchmark
Sören Becker, Johanna Vielhaben, Marcel Ackermann +3
cs.SDcs.AIcs.LGarXiv:1807.03418v32018Federated Unlearning with Knowledge Distillation
Chen Wu, Sencun Zhu, Prasenjit Mitra
cs.LGcs.CRarXiv:2201.09441v12022From Deep to Shallow: Unconstrained and Efficient Layer Merging Strategy
Petro Shulzhenko, Gabriele Spadaro, Enzo Tartaglione
cs.LGarXiv:2609.04881v12026KVMem: Virtualizing Million-Token Agent Workspaces on a Consumer GPU
Di Chai, Leye Wang, Zeshen Su +2
cs.LGarXiv:2609.04852v12026Solution-space heterogeneity shapes federated learning dynamics across partial differential equations
Ping Luo, Jiahuan Wang, Ziqing Wen +2
cs.LGcs.DCarXiv:2609.05012v12026Pre-training LLM without Learning Rate Decay Enhances Supervised Fine-Tuning
Kazuki Yano, Shun Kiyono, Sosuke Kobayashi +2
cs.CLcs.LGarXiv:2603.16127v12026Confounding-Valid Conformal Inference for Counterfactual KPIs in Wireless Networks
Abdessamed Qchohi, Jessica Moysen Cortes, Matteo Zecchin
cs.LGcs.NIeess.SParXiv:2609.05073v12026Rewards-in-Context: Multi-objective Alignment of Foundation Models with Dynamic Preference Adjustment
Rui Yang, Xiaoman Pan, Feng Luo +4
cs.LGcs.AIcs.CLarXiv:2402.10207v62024Failure-Aware RL: Reliable Offline-to-Online Reinforcement Learning with Self-Recovery for Real-World Manipulation
Huanyu Li, Kun Lei, Sheng Zang +5
cs.ROcs.AIcs.LGarXiv:2601.07821v12026Physics-Aware Random Walk Fingerprints for Scalable Power Grid Graph Classification
Adnan Anwar
cs.LGeess.SYarXiv:2609.04943v12026Fast Gauss Sums via Flash Attention
Nicolaj Rux, Sebastian Neumayer
cs.LGmath.NAarXiv:2609.04910v12026Adaptive Milestone Reward for GUI Agents
Congmin Zheng, Xiaoyun Mo, Xinbei Ma +10
cs.LGcs.AIcs.CLarXiv:2602.11524v12026SpecForge: A Flexible and Efficient Open-Source Training Framework for Speculative Decoding
Shenggui Li, Chao Wang, Yikai Zhu +14
cs.LGcs.AIcs.CLarXiv:2603.18567v12026ML-Decoder: Scalable and Versatile Classification Head
Tal Ridnik, Gilad Sharir, Avi Ben-Cohen +2
cs.CVcs.LGarXiv:2111.12933v22021When Genomic Masking Priors Fail to Transfer: Strong Variant Prediction, Weak Functional Generation
Susu Hu, Preetam Gattogi, Jens Lehmann +3
cs.LGarXiv:2609.04861v12026Demystifying Data-Driven Probabilistic Medium-Range Weather Forecasting
Jean Kossaifi, Nikola Kovachki, Morteza Mardani +15
cs.LGcs.AIarXiv:2601.18111v12026Hyperbolic Vision Transformers: Combining Improvements in Metric Learning
Aleksandr Ermolov, Leyla Mirvakhabova, Valentin Khrulkov +2
cs.CVcs.LGarXiv:2203.10833v22022PACE: Propagation-Aware Collaborative Correction for One-Shot Personalized Federated Graph Learning
Ruizhe Huang, Chengran Li, Xiaochuan Shi
cs.LGarXiv:2609.04832v12026Semi-parametric Image Synthesis
Xiaojuan Qi, Qifeng Chen, Jiaya Jia +1
cs.CVcs.AIcs.GRarXiv:1804.10992v12018Deep Partition Aggregation: Provable Defense against General Poisoning Attacks
Alexander Levine, Soheil Feizi
cs.LGstat.MLarXiv:2006.14768v22020Locating and Steering Refusal Beyond Attention
Preethi Carmel Bosco, Gopalakrishnan Srinivasan
cs.LGarXiv:2609.04721v12026Adaptivity of averaged stochastic gradient descent to local strong convexity for logistic regression
Francis Bach
math.STcs.LGmath.OCarXiv:1303.6149v32013Learning-Augmented Algorithms: Guarantees, Construction Mechanisms, and System-Level Implications
Hailiang Zhao, Peng Chen, Xueyan Tang +2
cs.LGcs.DSarXiv:2609.04787v12026A Second-order Bound with Excess Losses
Pierre Gaillard, Gilles Stoltz, Tim Van Erven
stat.MLcs.LGmath.STarXiv:1402.2044v12014Model-Based Deep Learning: On the Intersection of Deep Learning and Optimization
Nir Shlezinger, Yonina C. Eldar, Stephen P. Boyd
eess.SPcs.LGeess.SYarXiv:2205.02640v22022Communication-Efficient Personalized Federated Learning via Layer-Wise Multi-Threshold Random Sketching
Xu Zhang, Xingyu Hou, Jiacheng Cheng +2
cs.LGarXiv:2609.04830v12026Iterative Refinement Graph Neural Network for Antibody Sequence-Structure Co-design
Wengong Jin, Jeremy Wohlwend, Regina Barzilay +1
q-bio.BMcs.LGarXiv:2110.04624v32021Federated Attack Campaign Detection via Contrastive Encoding of Threat Indicators in Gradient Updates
Manuel Röder, Bibin Babu, Frank-Michael Schleif
cs.LGcs.CRarXiv:2609.04815v12026Compositionality and Generalization in Emergent Languages
Rahma Chaabouni, Eugene Kharitonov, Diane Bouchacourt +2
cs.CLcs.AIcs.LGarXiv:2004.09124v12020How Faithful Is Attribution for Sales Forecasting? A Counterfactual Study
Glib Kechyn
cs.LGarXiv:2609.04797v12026Countdown-Code: A Testbed for Studying The Emergence and Generalization of Reward Hacking in RLVR
Muhammad Khalifa, Zohaib Khan, Omer Tafveez +2
cs.LGcs.AIcs.CLarXiv:2603.07084v22026A Robust Watermark-based Fingerprint Framework for GNNs Ownership Verification
Han Zhang, Yan Wang, Guanfeng Liu +3
cs.LGarXiv:2609.04772v12026Resilience Beyond Stationary Client Unavailability: Unlocking Efficient and Unbiased Federated Learning
Ming Xiang, Stratis Ioannidis, Edmund Yeh +2
cs.LGcs.DCmath.OCarXiv:2609.04763v12026Expandable Subspace Ensemble for Pre-Trained Model-Based Class-Incremental Learning
Da-Wei Zhou, Hai-Long Sun, Han-Jia Ye +1
cs.CVcs.LGarXiv:2403.12030v12024