Machine Learning
Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
1,081 to 1,140 of 20,014
Scaling LLM Multi-turn RL with End-to-end Summarization-based Context Management
Miao Lu, Weiwei Sun, Weihua Du +4
cs.CLcs.AIcs.LGarXiv:2510.06727v12025GRPO-Guard: Mitigating Implicit Over-Optimization in Flow Matching via Regulated Clipping
Jing Wang, Jiajun Liang, Jie Liu +10
cs.CVcs.LGarXiv:2510.22319v22025Premise Selection for Theorem Proving by Deep Graph Embedding
Mingzhe Wang, Yihe Tang, Jian Wang +1
cs.AIcs.LGcs.LOarXiv:1709.09994v12017Predicting online extremism, content adopters, and interaction reciprocity
Emilio Ferrara, Wen-Qiang Wang, Onur Varol +2
cs.SIcs.LGphysics.soc-pharXiv:1605.00659v12016Prosperity before Collapse: How Far Can Off-Policy RL Reach with Stale Data on LLMs?
Haizhong Zheng, Jiawei Zhao, Beidi Chen
cs.LGcs.AIarXiv:2510.01161v22025No Prompt Left Behind: Exploiting Zero-Variance Prompts in LLM Reinforcement Learning via Entropy-Guided Advantage Shaping
Thanh-Long V. Le, Myeongho Jeon, Kim Vu +2
cs.CLcs.AIcs.LGarXiv:2509.21880v32025MIDI-VAE: Modeling Dynamics and Instrumentation of Music with Applications to Style Transfer
Gino Brunner, Andres Konrad, Yuyi Wang +1
cs.SDcs.LGeess.ASarXiv:1809.07600v120183D Point Splatting for mmWave Radar Novel View Synthesis
Adnan Armouti, Yixuan Gao, Rajalakshmi Nandakumar
cs.CVcs.GRcs.LGarXiv:2609.11894v12026Explainability Assistant: A Conversational XAI Interface for Interpreting Energy Consumption Models
Rodion Krjutškov, Eduard Barbu, Nikos Sakkas +1
cs.AIcs.LGarXiv:2609.11860v12026Scale-Invariant Convolutional Neural Networks
Yichong Xu, Tianjun Xiao, Jiaxing Zhang +2
cs.CVcs.LGcs.NEarXiv:1411.6369v12014Evaluating Time-Series Foundation Models and Multimodal Dietary Context for CGM Forecasting
Bowen Zhang, Hsiu-Wen Cheng, Hongyu Yang +9
stat.MLcs.LGarXiv:2609.11872v12026Generative Marketing Mix Modeling: A Causal Inference Framework Linking GEO and GEM to Business Impact
Masahiro Kato, Daiki Honma, Taka Kato
stat.MLcs.AIcs.LGarXiv:2609.11915v12026General Cutting Planes for Bound-Propagation-Based Neural Network Verification
Huan Zhang, Shiqi Wang, Kaidi Xu +5
cs.LGcs.CRcs.CVarXiv:2208.05740v22022Language Models are Injective and Hence Invertible
Giorgos Nikolaou, Tommaso Mencattini, Donato Crisostomi +3
cs.LGcs.AIarXiv:2510.15511v42025Near-Optimal Reinforcement Learning with Multi-Step Transition Lookahead
Corentin Pla, Hugo Richard, Marc Abeille +1
stat.MLcs.LGarXiv:2609.11807v12026Large Language Model Hacking: Quantifying the Hidden Risks of Using LLMs for Text Annotation
Joachim Baumann, Paul Röttger, Aleksandra Urman +4
cs.CLcs.AIcs.LGarXiv:2509.08825v22025ORCH: Organizational Principles Enable Collective Intelligence in Embodied AI
Zhengran Ji, Jonathan Hyun, Boyuan Chen
cs.MAcs.AIcs.LGarXiv:2609.11737v12026Building py-kvcache: A Performance Characterization of External KV Caching for vLLM with NVMe SSDs
Joseph Kanichai, Tiziano De Matteis, Animesh Trivedi
cs.DCcs.LGarXiv:2609.11744v12026Logit Refiner: Improving Visual Autoregressive Models via Intra-Scale Dependency Modeling
Meimingwei Li, Stefan Andreas Baumann, Felix Krause +1
cs.CVcs.AIcs.LGarXiv:2609.11804v12026Differentially Private EEG Feature Anonymization: A Privacy-Utility Case Study in Clinical Neurophysiology
Noman Sadiq, Mohsen Toorani
cs.CRcs.LGeess.SParXiv:2609.11777v12026Sparsity Regularized and Robust Mean Variance Portfolio Selection Under Ellipsoidal Uncertainty
Deniz Akkaya, Emre Can Yayla, Buse Şen +1
math.OCcs.LGstat.MLarXiv:2609.11749v12026Plex: Towards Reliability using Pretrained Large Model Extensions
Dustin Tran, Jeremiah Liu, Michael W. Dusenberry +23
cs.LGstat.MLarXiv:2207.07411v12022Bayesian Graph Neural Networks with Adaptive Connection Sampling
Arman Hasanzadeh, Ehsan Hajiramezanali, Shahin Boluki +4
cs.LGstat.MLarXiv:2006.04064v32020Geospatial Foundation Models Capture Health-Relevant Dimensions of Place Beyond Conventional Social Risk Indices
Nathaniel Hendrix, Carl Y. Zhang, Chris Heitzig +2
stat.APcs.LGarXiv:2609.11689v12026Learning structural balance of graphs from quantum spectral features
Stefano Scali, Oleksandr Kyriienko
quant-phcond-mat.dis-nncs.LGarXiv:2609.11736v12026Generalization Analysis of Distributed Kernel-based Robust Gradient Descent Algorithms
Jun-Yi Meng, Zheng-Chu Guo, Yuan Mao
stat.MLcs.LGmath.OAarXiv:2609.11712v12026Evaluating Gemini Robotics Policies in a Veo World Simulator
Gemini Robotics Team, Krzysztof Choromanski, Coline Devin +20
cs.ROcs.AIcs.CVarXiv:2512.10675v22025Reflex-Informed Neuromuscular Reinforcement Learning for Muscle-Driven Locomotion
Jian Zhou, Xingyu Zhang, Rui Ma +3
cs.ROcs.GRcs.LGarXiv:2609.11733v12026Multimodal Taxonomic Conditioning for Generative Plankton Imagery
Daniela Ivanova, Ozgu Goksu, Nicolas Pugeault
cs.CVcs.LGarXiv:2609.11673v12026VeRO: A Harness for Agents to Optimize Agents
Varun Ursekar, Apaar Shanker, Veronica Chatrath +2
cs.AIcs.CLcs.LGarXiv:2602.22480v42026Learning to design drug-like molecules in three-dimensional space using deep generative models
Yibo Li, Jianfeng Pei, Luhua Lai
q-bio.QMcs.LGarXiv:2104.08474v12021Can LLMs Beat Classical Hyperparameter Optimization Algorithms? A Study on autoresearch
Fabio Ferreira, Lucca Wobbe, Arjun Krishnakumar +2
cs.LGstat.MLarXiv:2603.24647v52026Vidu S2: Real-Time Interactive, Editable, and Spatial Video Generation
Jintao Zhang, Kai Jiang, Jintao Chen +32
cs.CVcs.LGarXiv:2609.11638v12026ZipCodec: Ultra-Low-Frame-Rate Streaming Speech Coding
Luca Della Libera, Cem Subakan, Mirco Ravanelli
cs.SDcs.AIcs.LGarXiv:2609.11642v12026Great Models Think Alike and this Undermines AI Oversight
Shashwat Goel, Joschka Struber, Ilze Amanda Auzina +6
cs.LGcs.AIcs.CLarXiv:2502.04313v22025WebInject: Prompt Injection Attack to Web Agents
Xilong Wang, John Bloch, Zedian Shao +3
cs.LGcs.AIcs.CLarXiv:2505.11717v42025Distributed Optimization of Modular Production Systems using Model-based Reinforcement Learning with Inverse Models
Andreas Schwung, Steve Yuwono, Sofiene Lassoued +1
cs.AIcs.LGeess.SYarXiv:2609.11615v12026Identifiability of Nonnegative Tensor Decompositions via Positive Scattering
Haoming Wang, Ming Yuan
stat.MLcs.LGmath.COarXiv:2609.11606v12026A distribution-free certification framework for trustworthy crash-severity prediction
Amir Rafe, Subasish Das
stat.MLcs.LGarXiv:2609.11592v12026Robust breast cancer detection in mammography and digital breast tomosynthesis using annotation-efficient deep learning approach
William Lotter, Abdul Rahman Diab, Bryan Haslam +10
eess.IVcs.CVcs.LGarXiv:1912.11027v22019AdaCliP: Adaptive Clipping for Private SGD
Venkatadheeraj Pichapati, Ananda Theertha Suresh, Felix X. Yu +2
cs.LGcs.CRstat.MLarXiv:1908.07643v22019Fine-Tuning Pre-trained Language Model with Weak Supervision: A Contrastive-Regularized Self-Training Approach
Yue Yu, Simiao Zuo, Haoming Jiang +3
cs.CLcs.LGarXiv:2010.07835v32020ExGRPO: Learning to Reason from Experience
Runzhe Zhan, Yafu Li, Zhi Wang +5
cs.LGcs.AIcs.CLarXiv:2510.02245v22025UME-R1: Exploring Reasoning-Driven Generative Multimodal Embeddings
Zhibin Lan, Liqiang Niu, Fandong Meng +2
cs.LGcs.AIarXiv:2511.00405v22025Published Unlearning Numbers Move Per Checkpoint, and Not Because the Removed Data Survives: An Audit of 263 Released Batch-Normalized Checkpoints
Junlong Shen Xingyu Li
cs.AIcs.LGarXiv:2609.11490v12026Breaking the Central Bias: Spatially Partitioned Experts for Coordinate-Based Neuroevolution
Romain Claret, Arthur Gygax, Michael O'Neill +3
cs.NEcs.CVcs.LGarXiv:2609.11518v12026Enabling Knowledge Graph Understanding at Scale with the EXplore Your Graphs ENgine (EXYGEN)
Harshdeep Singh, Yurui Zhu, Giovanni Colavizza +1
cs.AIcs.LGarXiv:2609.11569v12026Risk-Averse Decision Making with Multi-Level Reliability Guarantees
Amirmohammad Farzaneh, Osvaldo Simeone
stat.MLcs.ITcs.LGarXiv:2609.11524v12026The Path Not Taken: RLVR Provably Learns Off the Principals
Hanqing Zhu, Zhenyu Zhang, Hanxian Huang +11
cs.LGcs.AIarXiv:2511.08567v12025Hologram Representation via Quadratic Phase Gaussian Splatting
Haolong Wang, Yicheng Zhan, Kaan Akşit +1
cs.GRcs.CVcs.LGarXiv:2609.11434v12026Deep operator learning for efficient sampling from invariant measures of stochastic differential equations
Ling Guo, Lei Li, Jingtong Zhang
math.NAcs.LGarXiv:2609.11376v12026Your Model Already Knows Don't Teach It, Learn to Ask It: Soft Prompting for Few-Shot Adaptation of Vision-Language Models
Gautam Rajendrakumar Gare, Siyi Li, Hewei Wang +5
cs.CVcs.AIcs.LGarXiv:2609.11310v12026General Agentic Memory Via Deep Research
B. Y. Yan, Chaofan Li, Hongjin Qian +2
cs.CLcs.AIcs.IRarXiv:2511.18423v12025Improving the Sensitivity of Gravitational Wave Detection with Weighted Conformal Prediction
Ann-Kristin Malz, Gregory Ashton, Nicolo Colombo
gr-qccs.LGstat.MLarXiv:2609.11401v12026GR-RL: Going Dexterous and Precise for Long-Horizon Robotic Manipulation
Yunfei Li, Xiao Ma, Jiafeng Xu +18
cs.ROcs.LGarXiv:2512.01801v32025A Hilbert-Valued Functional Decomposition Framework for Explaining Time-Dependent Outputs
Sophie Hanna Langbein, Niklas Koenen, Marvin N. Wright +1
stat.MLcs.LGarXiv:2609.11295v12026A Two-Mirror Faceted Projection System for EUV Lithography
Vasiliy A. Es'kin, Egor V. Ivanov, Olga V. Martynova
physics.opticscs.LGphysics.app-pharXiv:2609.11299v12026On the Optimal Weighted $\ell_2$ Regularization in Overparameterized Linear Regression
Denny Wu, Ji Xu
stat.MLcs.LGmath.STarXiv:2006.05800v42020Learning Physics-guided Face Relighting under Directional Light
Thomas Nestmeyer, Jean-François Lalonde, Iain Matthews +1
cs.CVcs.GRcs.LGarXiv:1906.03355v22019Cache-to-Cache: Direct Semantic Communication Between Large Language Models
Tianyu Fu, Zihan Min, Hanling Zhang +4
cs.CLcs.LGarXiv:2510.03215v22025