Machine Learning

Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

1,081 to 1,140 of 20,014

  1. Scaling LLM Multi-turn RL with End-to-end Summarization-based Context Management

    Miao Lu, Weiwei Sun, Weihua Du +4

    cs.CLcs.AIcs.LGarXiv:2510.06727v12025
  2. GRPO-Guard: Mitigating Implicit Over-Optimization in Flow Matching via Regulated Clipping

    Jing Wang, Jiajun Liang, Jie Liu +10

    cs.CVcs.LGarXiv:2510.22319v22025
  3. Premise Selection for Theorem Proving by Deep Graph Embedding

    Mingzhe Wang, Yihe Tang, Jian Wang +1

    cs.AIcs.LGcs.LOarXiv:1709.09994v12017
  4. Predicting online extremism, content adopters, and interaction reciprocity

    Emilio Ferrara, Wen-Qiang Wang, Onur Varol +2

    cs.SIcs.LGphysics.soc-pharXiv:1605.00659v12016
  5. Prosperity before Collapse: How Far Can Off-Policy RL Reach with Stale Data on LLMs?

    Haizhong Zheng, Jiawei Zhao, Beidi Chen

    cs.LGcs.AIarXiv:2510.01161v22025
  6. No Prompt Left Behind: Exploiting Zero-Variance Prompts in LLM Reinforcement Learning via Entropy-Guided Advantage Shaping

    Thanh-Long V. Le, Myeongho Jeon, Kim Vu +2

    cs.CLcs.AIcs.LGarXiv:2509.21880v32025
  7. MIDI-VAE: Modeling Dynamics and Instrumentation of Music with Applications to Style Transfer

    Gino Brunner, Andres Konrad, Yuyi Wang +1

    cs.SDcs.LGeess.ASarXiv:1809.07600v12018
  8. 3D Point Splatting for mmWave Radar Novel View Synthesis

    Adnan Armouti, Yixuan Gao, Rajalakshmi Nandakumar

    cs.CVcs.GRcs.LGarXiv:2609.11894v12026
  9. Explainability Assistant: A Conversational XAI Interface for Interpreting Energy Consumption Models

    Rodion Krjutškov, Eduard Barbu, Nikos Sakkas +1

    cs.AIcs.LGarXiv:2609.11860v12026
  10. Scale-Invariant Convolutional Neural Networks

    Yichong Xu, Tianjun Xiao, Jiaxing Zhang +2

    cs.CVcs.LGcs.NEarXiv:1411.6369v12014
  11. Evaluating Time-Series Foundation Models and Multimodal Dietary Context for CGM Forecasting

    Bowen Zhang, Hsiu-Wen Cheng, Hongyu Yang +9

    stat.MLcs.LGarXiv:2609.11872v12026
  12. Generative Marketing Mix Modeling: A Causal Inference Framework Linking GEO and GEM to Business Impact

    Masahiro Kato, Daiki Honma, Taka Kato

    stat.MLcs.AIcs.LGarXiv:2609.11915v12026
  13. General Cutting Planes for Bound-Propagation-Based Neural Network Verification

    Huan Zhang, Shiqi Wang, Kaidi Xu +5

    cs.LGcs.CRcs.CVarXiv:2208.05740v22022
  14. Language Models are Injective and Hence Invertible

    Giorgos Nikolaou, Tommaso Mencattini, Donato Crisostomi +3

    cs.LGcs.AIarXiv:2510.15511v42025
  15. Near-Optimal Reinforcement Learning with Multi-Step Transition Lookahead

    Corentin Pla, Hugo Richard, Marc Abeille +1

    stat.MLcs.LGarXiv:2609.11807v12026
  16. Large Language Model Hacking: Quantifying the Hidden Risks of Using LLMs for Text Annotation

    Joachim Baumann, Paul Röttger, Aleksandra Urman +4

    cs.CLcs.AIcs.LGarXiv:2509.08825v22025
  17. ORCH: Organizational Principles Enable Collective Intelligence in Embodied AI

    Zhengran Ji, Jonathan Hyun, Boyuan Chen

    cs.MAcs.AIcs.LGarXiv:2609.11737v12026
  18. Building py-kvcache: A Performance Characterization of External KV Caching for vLLM with NVMe SSDs

    Joseph Kanichai, Tiziano De Matteis, Animesh Trivedi

    cs.DCcs.LGarXiv:2609.11744v12026
  19. Logit Refiner: Improving Visual Autoregressive Models via Intra-Scale Dependency Modeling

    Meimingwei Li, Stefan Andreas Baumann, Felix Krause +1

    cs.CVcs.AIcs.LGarXiv:2609.11804v12026
  20. Differentially Private EEG Feature Anonymization: A Privacy-Utility Case Study in Clinical Neurophysiology

    Noman Sadiq, Mohsen Toorani

    cs.CRcs.LGeess.SParXiv:2609.11777v12026
  21. Sparsity Regularized and Robust Mean Variance Portfolio Selection Under Ellipsoidal Uncertainty

    Deniz Akkaya, Emre Can Yayla, Buse Şen +1

    math.OCcs.LGstat.MLarXiv:2609.11749v12026
  22. Plex: Towards Reliability using Pretrained Large Model Extensions

    Dustin Tran, Jeremiah Liu, Michael W. Dusenberry +23

    cs.LGstat.MLarXiv:2207.07411v12022
  23. Bayesian Graph Neural Networks with Adaptive Connection Sampling

    Arman Hasanzadeh, Ehsan Hajiramezanali, Shahin Boluki +4

    cs.LGstat.MLarXiv:2006.04064v32020
  24. Geospatial Foundation Models Capture Health-Relevant Dimensions of Place Beyond Conventional Social Risk Indices

    Nathaniel Hendrix, Carl Y. Zhang, Chris Heitzig +2

    stat.APcs.LGarXiv:2609.11689v12026
  25. Learning structural balance of graphs from quantum spectral features

    Stefano Scali, Oleksandr Kyriienko

    quant-phcond-mat.dis-nncs.LGarXiv:2609.11736v12026
  26. Generalization Analysis of Distributed Kernel-based Robust Gradient Descent Algorithms

    Jun-Yi Meng, Zheng-Chu Guo, Yuan Mao

    stat.MLcs.LGmath.OAarXiv:2609.11712v12026
  27. Evaluating Gemini Robotics Policies in a Veo World Simulator

    Gemini Robotics Team, Krzysztof Choromanski, Coline Devin +20

    cs.ROcs.AIcs.CVarXiv:2512.10675v22025
  28. Reflex-Informed Neuromuscular Reinforcement Learning for Muscle-Driven Locomotion

    Jian Zhou, Xingyu Zhang, Rui Ma +3

    cs.ROcs.GRcs.LGarXiv:2609.11733v12026
  29. Multimodal Taxonomic Conditioning for Generative Plankton Imagery

    Daniela Ivanova, Ozgu Goksu, Nicolas Pugeault

    cs.CVcs.LGarXiv:2609.11673v12026
  30. VeRO: A Harness for Agents to Optimize Agents

    Varun Ursekar, Apaar Shanker, Veronica Chatrath +2

    cs.AIcs.CLcs.LGarXiv:2602.22480v42026
  31. Learning to design drug-like molecules in three-dimensional space using deep generative models

    Yibo Li, Jianfeng Pei, Luhua Lai

    q-bio.QMcs.LGarXiv:2104.08474v12021
  32. Can LLMs Beat Classical Hyperparameter Optimization Algorithms? A Study on autoresearch

    Fabio Ferreira, Lucca Wobbe, Arjun Krishnakumar +2

    cs.LGstat.MLarXiv:2603.24647v52026
  33. Vidu S2: Real-Time Interactive, Editable, and Spatial Video Generation

    Jintao Zhang, Kai Jiang, Jintao Chen +32

    cs.CVcs.LGarXiv:2609.11638v12026
  34. ZipCodec: Ultra-Low-Frame-Rate Streaming Speech Coding

    Luca Della Libera, Cem Subakan, Mirco Ravanelli

    cs.SDcs.AIcs.LGarXiv:2609.11642v12026
  35. Great Models Think Alike and this Undermines AI Oversight

    Shashwat Goel, Joschka Struber, Ilze Amanda Auzina +6

    cs.LGcs.AIcs.CLarXiv:2502.04313v22025
  36. WebInject: Prompt Injection Attack to Web Agents

    Xilong Wang, John Bloch, Zedian Shao +3

    cs.LGcs.AIcs.CLarXiv:2505.11717v42025
  37. Distributed Optimization of Modular Production Systems using Model-based Reinforcement Learning with Inverse Models

    Andreas Schwung, Steve Yuwono, Sofiene Lassoued +1

    cs.AIcs.LGeess.SYarXiv:2609.11615v12026
  38. Identifiability of Nonnegative Tensor Decompositions via Positive Scattering

    Haoming Wang, Ming Yuan

    stat.MLcs.LGmath.COarXiv:2609.11606v12026
  39. A distribution-free certification framework for trustworthy crash-severity prediction

    Amir Rafe, Subasish Das

    stat.MLcs.LGarXiv:2609.11592v12026
  40. Robust breast cancer detection in mammography and digital breast tomosynthesis using annotation-efficient deep learning approach

    William Lotter, Abdul Rahman Diab, Bryan Haslam +10

    eess.IVcs.CVcs.LGarXiv:1912.11027v22019
  41. AdaCliP: Adaptive Clipping for Private SGD

    Venkatadheeraj Pichapati, Ananda Theertha Suresh, Felix X. Yu +2

    cs.LGcs.CRstat.MLarXiv:1908.07643v22019
  42. Fine-Tuning Pre-trained Language Model with Weak Supervision: A Contrastive-Regularized Self-Training Approach

    Yue Yu, Simiao Zuo, Haoming Jiang +3

    cs.CLcs.LGarXiv:2010.07835v32020
  43. ExGRPO: Learning to Reason from Experience

    Runzhe Zhan, Yafu Li, Zhi Wang +5

    cs.LGcs.AIcs.CLarXiv:2510.02245v22025
  44. UME-R1: Exploring Reasoning-Driven Generative Multimodal Embeddings

    Zhibin Lan, Liqiang Niu, Fandong Meng +2

    cs.LGcs.AIarXiv:2511.00405v22025
  45. Published Unlearning Numbers Move Per Checkpoint, and Not Because the Removed Data Survives: An Audit of 263 Released Batch-Normalized Checkpoints

    Junlong Shen Xingyu Li

    cs.AIcs.LGarXiv:2609.11490v12026
  46. Breaking the Central Bias: Spatially Partitioned Experts for Coordinate-Based Neuroevolution

    Romain Claret, Arthur Gygax, Michael O'Neill +3

    cs.NEcs.CVcs.LGarXiv:2609.11518v12026
  47. Enabling Knowledge Graph Understanding at Scale with the EXplore Your Graphs ENgine (EXYGEN)

    Harshdeep Singh, Yurui Zhu, Giovanni Colavizza +1

    cs.AIcs.LGarXiv:2609.11569v12026
  48. Risk-Averse Decision Making with Multi-Level Reliability Guarantees

    Amirmohammad Farzaneh, Osvaldo Simeone

    stat.MLcs.ITcs.LGarXiv:2609.11524v12026
  49. The Path Not Taken: RLVR Provably Learns Off the Principals

    Hanqing Zhu, Zhenyu Zhang, Hanxian Huang +11

    cs.LGcs.AIarXiv:2511.08567v12025
  50. Hologram Representation via Quadratic Phase Gaussian Splatting

    Haolong Wang, Yicheng Zhan, Kaan Akşit +1

    cs.GRcs.CVcs.LGarXiv:2609.11434v12026
  51. Deep operator learning for efficient sampling from invariant measures of stochastic differential equations

    Ling Guo, Lei Li, Jingtong Zhang

    math.NAcs.LGarXiv:2609.11376v12026
  52. Your Model Already Knows Don't Teach It, Learn to Ask It: Soft Prompting for Few-Shot Adaptation of Vision-Language Models

    Gautam Rajendrakumar Gare, Siyi Li, Hewei Wang +5

    cs.CVcs.AIcs.LGarXiv:2609.11310v12026
  53. General Agentic Memory Via Deep Research

    B. Y. Yan, Chaofan Li, Hongjin Qian +2

    cs.CLcs.AIcs.IRarXiv:2511.18423v12025
  54. Improving the Sensitivity of Gravitational Wave Detection with Weighted Conformal Prediction

    Ann-Kristin Malz, Gregory Ashton, Nicolo Colombo

    gr-qccs.LGstat.MLarXiv:2609.11401v12026
  55. GR-RL: Going Dexterous and Precise for Long-Horizon Robotic Manipulation

    Yunfei Li, Xiao Ma, Jiafeng Xu +18

    cs.ROcs.LGarXiv:2512.01801v32025
  56. A Hilbert-Valued Functional Decomposition Framework for Explaining Time-Dependent Outputs

    Sophie Hanna Langbein, Niklas Koenen, Marvin N. Wright +1

    stat.MLcs.LGarXiv:2609.11295v12026
  57. A Two-Mirror Faceted Projection System for EUV Lithography

    Vasiliy A. Es'kin, Egor V. Ivanov, Olga V. Martynova

    physics.opticscs.LGphysics.app-pharXiv:2609.11299v12026
  58. On the Optimal Weighted $\ell_2$ Regularization in Overparameterized Linear Regression

    Denny Wu, Ji Xu

    stat.MLcs.LGmath.STarXiv:2006.05800v42020
  59. Learning Physics-guided Face Relighting under Directional Light

    Thomas Nestmeyer, Jean-François Lalonde, Iain Matthews +1

    cs.CVcs.GRcs.LGarXiv:1906.03355v22019
  60. Cache-to-Cache: Direct Semantic Communication Between Large Language Models

    Tianyu Fu, Zihan Min, Hanling Zhang +4

    cs.CLcs.LGarXiv:2510.03215v22025