Machine Learning
Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
19,861 to 19,920 of 19,961
Any-OPD: Heterogeneous On-Policy Distillation for Flow-Matching Models via Representation-Space Bridging
Siming Fu, Zheming Fu, Ruizhe He +7
cs.LGcs.CVarXiv:2608.03316v12026DeaMoE: Efficient MoE Structure for Fast Small-Batch Decoding
Zewen Jin, Shen Fu, Zeping Duan +8
cs.LGcs.AIarXiv:2608.14385v12026SPEAR: Structure Property Explainability with Attention Regularization
Aditya Raghavan, Utkarsh Pratiush, Dalton A. Pearl +4
cond-mat.mtrl-scics.LGarXiv:2608.13826v12026Conditional Neural Optimal Transport for Predicting Cellular Phenotypes from Molecular Structure
Gauthier Avité, Maxime Sanchez-Renauld, Nicolas Bourriez +1
cs.CVcs.LGarXiv:2608.14293v12026Approximate Muon with low-rank adapters
Ben Anson, Conor Houghton, Edward Milsom
cs.LGarXiv:2608.14492v12026CytoBERT: A Foundation Model for Cytometry Data
Syed Abdul Haseeb Qadri, Bjarne C. Hiller, Felix Blanke +7
cs.LGarXiv:2608.14414v12026LP-NAS: Linear Programming-based Neural Architecture Search
Abhishek Shukla, Ankur Sinha, Faiz Hamid
cs.LGcs.AIarXiv:2608.14472v12026A Four-Axis Trustworthiness Benchmark for LLM-as-Judge in Principle-Based Regulation
Dipankar Sarkar
cs.CRcs.AIcs.CLarXiv:2608.14329v12026A Graph-Based Reinforcement Learning Framework for Structured Drift Diagnosis and Recovery in Autonomous LLM Agents
Ismail El Hamraoui, Sagar Jose, Nicolas Bureau +1
cs.AIcs.LGcs.MAarXiv:2608.14109v12026CoinRAG: Contextualized Information Nugget KV Cache Reuse for Long-Context RAG
Gyuwan Kim, Cheoneum Park, Tao Yang
cs.CLcs.AIcs.IRarXiv:2608.07458v12026VoiceDesigner: Text-to-Voice Generation and Editing via Unified Diffusion Modeling and Data Augmentation
Jiarui Hai, Karan Thakkar, Ke Chen +5
eess.AScs.LGarXiv:2608.13613v12026What to Preserve, Where to Adapt: A Depth-Wise Analysis of Forgetting in Continual Gynecological Image Segmentation
Amal Saqib, Tausifa Jan Saleem, Numan Saeed +1
cs.CVcs.LGarXiv:2608.13660v12026Invisible Shortcuts: Why Vision Encoders Know Your Camera
Vladan Stojnić, Ryan Ramos, Giorgos Kordopatis-Zilos +2
cs.CVcs.LGarXiv:2608.05424v12026ContinualSkillBench: Can LLM Agents Truly Evolve Their Capabilities?
Tianyi Guan, Yiding Wang, Haotong Yang +5
cs.AIcs.CLcs.LGarXiv:2608.03874v12026GradCuit: Credit-Assigned Gradient Flow Enables Robust and Interpretable Test-Time Latent Reasoning
Zhaoxin Yu, Qi Shen, Hengli Li +4
cs.LGcs.CLarXiv:2608.02585v22026CutClean: Neural Network Pruning for Privacy-Preserving Inference
Leonardo Magliolo, Vito Paolo Pastore, Giuseppe Valenzise +1
cs.LGcs.AIarXiv:2608.13773v12026Intelligent Detection of Mechanical, Electrical, and Plumbing (MEP) Metrics Based on 2D Floor Plans
Tarandeep Singh Mandhiratta, ANK Zaman, Abdul-Rahman Mawlood-Yunis
cs.CVcs.AIcs.HCarXiv:2608.14317v12026Wrong but Useful: Trajectory Value Beyond Answer Correctness in Multi-Agent Messages
Chih-Hsuan Yang, Anjir Ahmed Chowdhury, Cheng-Hau Yang +7
cs.AIcs.CLcs.LGarXiv:2608.14375v12026RecipeNet: A Hierarchical Transformer for Recipe Data
Pin-Yen Huang, Sachin Chhabra, Prasanth Sai Gouripeddi +2
cs.LGcs.AIarXiv:2608.14505v12026Architecture and Affordances of PLAUD: Performative Latents and Unsupervised DDSP
Błażej Kotowski, Frederic Font
cs.SDcs.HCcs.LGarXiv:2608.13724v12026Convex losses and their applications to SVM, SVR, and Shallow Neural Networks
Filippo Portera
cs.LGarXiv:2608.14288v12026Building AI-Intensive Software with AI: Early Results and a Cautionary Tale on Measuring Development Cost
Victor Barros de Miranda Neves, Kiev Santos da Gama, Vinicius Cardoso Garcia
cs.SEcs.AIcs.LGarXiv:2608.13730v12026Relevant but Incomplete: Referential Dangling as a Paradigm-Level Failure Mode in Hard Prompt Compression
Zhengpei Hu, Kai Li, Dapeng Fu +5
cs.CLcs.LGarXiv:2608.04569v12026Unknown Unknowns: Model Misspecification in Machine Learning for Physics
Juan Cruz-Martinez, Carolina Cuesta-Lazaro, Alexander Held +1
physics.data-anastro-ph.COastro-ph.GAarXiv:2608.13633v12026Learning Unsteady Aneurysm Hemodynamics with Physics-Informed DeepONets
Oscar L. Cruz-Gonzalez, Valérie Deplano, Badih Ghattas
stat.MLcs.LGphysics.flu-dynarXiv:2608.13629v12026AgilePE: Autonomous UAV Pursuit-Evasion via Self-Play Reinforcement Learning
Wenhao Tang, Tianyang Chen, Zhejun Cui +9
cs.ROcs.LGarXiv:2608.14135v12026Interpretable MEG Decoding of Perceived Speech: Cortical Sources and the Stimulus Features That Drive Retrieval
Ilia Semenkov, Daria Kleeva, Ivan Dakhtin +2
cs.LGcs.SDq-bio.NCarXiv:2608.01481v12026On the Brittleness of Maximum Likelihood Estimation for Gaussian Process Hyperparameter Optimization
Tyler R. Johnson, Kian Ben-Jacob, Christopher P. Muller +1
stat.MLcs.LGstat.MEarXiv:2608.13793v12026Continual Learning in Transition
Zhiyan Hou, Dan Zhang, Tao Feng +11
cs.LGcs.AIarXiv:2608.06216v22026An AI4AI Framework for Visual Token Pruning
Zhen Liu, Wenli Huang, Wei Song +3
cs.LGcs.CVarXiv:2608.07193v12026Expected Free Energy-based Informative Path Planning for Robotic Mars Exploration
Ajith Anil Meera, Pablo Lanillos, Wouter Kouw
cs.ROcs.ITcs.LGarXiv:2608.14466v12026Non-Parametric Spatiotemporal Trajectory Prediction via State-Conditioned Transition Sampling
Michael Fore, Akshay Jain, Justin Downes +2
cs.LGarXiv:2608.14349v12026Deep Reinforcement Learning solution for pickup and delivery routing problems with time window and capacity constraints
Andrew Soroka, Alex Meshcheryakov, Sergey Gerasimov
cs.LGarXiv:2608.14156v12026HarnessOpt-Bench: Evaluating LLMs at Harness Optimization
Varun Ursekar, Apaar Shanker, Yash Maurya +4
cs.AIcs.CLcs.LGarXiv:2608.06301v12026Mind the Long Tail: Understanding the Difficulty of Delay Detection in Business Processes
Keyvan Amiri Elyasi, Lukas Kirchdorfer, Heiner Stuckenschmidt
cs.LGcs.AIarXiv:2608.14367v12026OneDayAgent: Towards a Long-Horizon Harness for Autonomous Agents
Jingsheng Zheng, Xinyuan Fang, Jintian Zhang +3
cs.CLcs.AIcs.HCarXiv:2608.05013v12026Rollplex: Cross-Phase GPU Spatial Sharing for Vision Language Model Post-Training
Hanfeng Lu, Tianyu Feng, Suyi Li +8
cs.LGcs.DCarXiv:2608.14498v12026Buy the Rumor, Sell the News: When Is News Priced In?
Alireza Kargarzadeh, Nariman Khaledian, Navid Parvini +2
cs.AIcs.LGq-fin.STarXiv:2608.14014v12026On-Policy Delta Distillation for Multilingual Math Reasoning
Byeongho Heo, Jaehui Hwang, Sangdoo Yun +1
cs.CLcs.LGarXiv:2608.05802v12026SPOT: Sparse Probing and Outcome Calibration for On-Policy Distillation
Zikun Qu, Min Zhang, Mingze Kong +4
cs.LGcs.AIarXiv:2608.04419v12026Progressive Agent Skill Generation via Reinforcement Learning
Junhao Shen, Zhanqiu Zhang, Yiwen Guo +1
cs.LGcs.CLarXiv:2608.01678v12026TSDS-Toolbox: A Toolbox for Measuring Time-Series Dataset Similarity
Yen-Ku Liu, Hongjie Chen, Ryan A. Rossi +1
cs.LGarXiv:2608.08119v12026From Passive Delegates to Strategic Negotiators: Reinforcing Social Reasoning in Small Language Models with SocialRL
Wenyue Hua, Zachary Huang, Tyler Payne +3
cs.AIcs.CLcs.LGarXiv:2608.13787v12026Identifiability and Order-Dimension Limits of In-Context Learning on Partial Orders
Faizanuddin Ansari, Debanjan Dutta, Swagatam Das
cs.LGarXiv:2608.14004v12026Fixed-Budget Gaussian Volume Encoding with Structure-Aware Allocation
Michael R. Martin, Joseph Insley, Victor A. Mateevitsi +2
cs.CVcs.AIcs.CEarXiv:2608.14112v12026L-FNO: Lorentzian Fourier Neural Operator for Stochastic Event Dynamics
Songhee Kang, Jihoon Kang
cs.LGstat.MLarXiv:2608.13562v12026Multi-Objective Bayesian Optimization for Model Merging
Utkarsh Agarwal, Vamshi Bonagiri, Raul Astudillo +1
cs.LGcs.AIarXiv:2608.14264v12026Post-training Quantization for Hybrid Iterative Generative Models
Jing Gao, Junyi Wu, Wei Wang +2
cs.LGarXiv:2608.13932v12026MINT: A Universal Zero-Shot Predictor for Transaction Data
Parameswaran Kamalaruban, Viktor Drobnyi, Maeve Madigan +3
cs.LGcs.CLarXiv:2608.14198v12026Clearing the Fog: Towards Installing and Refining Proactive Exploration Capabilities in LLM Agents
Zhizhao Guan, Chen Huang, Ziming Liu +5
cs.AIcs.LGarXiv:2608.14339v12026Designing Compact Neural Architectures via Neuron Gating and Mixed Activation
Abhishek Shukla, Ankur Sinha, Faiz Hamid
cs.LGcs.AIarXiv:2608.14443v12026ATLAS: Discovering Agent Strategies through LLM-Guided Abstraction and Automata Learning
Ignacio D. Lopez-Miguel, Andreas Happe, Jürgen Cito +3
cs.SEcs.LGarXiv:2608.14352v12026Probabilistic indirect models for undrained shear strength: addressing significant data missing and variability with advanced imputation and machine learning techniques
Haibin Xiong, Shaoheng Dai, Peng Lan +4
cs.LGcs.DBarXiv:2608.13934v12026On-Policy Delta Distillation
Byeongho Heo, Jaehui Hwang, Sangdoo Yun +1
cs.LGcs.CLarXiv:2607.15161v12026Summaries:한국어Weak-to-Strong Generalization via Direct On-Policy Distillation
Shiyuan Feng, Huan-ang Gao, Haohan Chi +7
cs.LGcs.AIcs.CLarXiv:2607.05394v22026Summaries:한국어Qwen-RobotManip Technical Report: Alignment Unlocks Scale for Robotic Manipulation Foundation Models
Haoqi Yuan, Zhixuan Liang, Anzhe Chen +20
cs.ROcs.CVcs.LGarXiv:2606.17846v22026Summaries:한국어Attention Is All You Need
Ashish Vaswani, Noam Shazeer, Niki Parmar +5
cs.CLcs.LGarXiv:1706.03762v72017Densely Connected Convolutional Networks
Gao Huang, Zhuang Liu, Laurens van der Maaten +1
cs.CVcs.LGarXiv:1608.06993v52016Summaries:한국어Regime-Conditional Verification: Correctness Estimation for Adapting and Monitoring Safety Classifiers
Thiago Sandoval, Ufuk Topcu
cs.AIcs.CLcs.CRarXiv:2608.14089v12026CForce: Boosting Parallel Decoding for dLLMs via Consistency Forcing
Yuji Ren, Chenkai Xu, Zhuocheng Gong +2
cs.LGcs.AIcs.CLarXiv:2608.13925v12026