Machine Learning

Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

19,861 to 19,920 of 19,961

  1. Any-OPD: Heterogeneous On-Policy Distillation for Flow-Matching Models via Representation-Space Bridging

    Siming Fu, Zheming Fu, Ruizhe He +7

    cs.LGcs.CVarXiv:2608.03316v12026
  2. DeaMoE: Efficient MoE Structure for Fast Small-Batch Decoding

    Zewen Jin, Shen Fu, Zeping Duan +8

    cs.LGcs.AIarXiv:2608.14385v12026
  3. SPEAR: Structure Property Explainability with Attention Regularization

    Aditya Raghavan, Utkarsh Pratiush, Dalton A. Pearl +4

    cond-mat.mtrl-scics.LGarXiv:2608.13826v12026
  4. Conditional Neural Optimal Transport for Predicting Cellular Phenotypes from Molecular Structure

    Gauthier Avité, Maxime Sanchez-Renauld, Nicolas Bourriez +1

    cs.CVcs.LGarXiv:2608.14293v12026
  5. Approximate Muon with low-rank adapters

    Ben Anson, Conor Houghton, Edward Milsom

    cs.LGarXiv:2608.14492v12026
  6. CytoBERT: A Foundation Model for Cytometry Data

    Syed Abdul Haseeb Qadri, Bjarne C. Hiller, Felix Blanke +7

    cs.LGarXiv:2608.14414v12026
  7. LP-NAS: Linear Programming-based Neural Architecture Search

    Abhishek Shukla, Ankur Sinha, Faiz Hamid

    cs.LGcs.AIarXiv:2608.14472v12026
  8. A Four-Axis Trustworthiness Benchmark for LLM-as-Judge in Principle-Based Regulation

    Dipankar Sarkar

    cs.CRcs.AIcs.CLarXiv:2608.14329v12026
  9. A Graph-Based Reinforcement Learning Framework for Structured Drift Diagnosis and Recovery in Autonomous LLM Agents

    Ismail El Hamraoui, Sagar Jose, Nicolas Bureau +1

    cs.AIcs.LGcs.MAarXiv:2608.14109v12026
  10. CoinRAG: Contextualized Information Nugget KV Cache Reuse for Long-Context RAG

    Gyuwan Kim, Cheoneum Park, Tao Yang

    cs.CLcs.AIcs.IRarXiv:2608.07458v12026
  11. VoiceDesigner: Text-to-Voice Generation and Editing via Unified Diffusion Modeling and Data Augmentation

    Jiarui Hai, Karan Thakkar, Ke Chen +5

    eess.AScs.LGarXiv:2608.13613v12026
  12. What to Preserve, Where to Adapt: A Depth-Wise Analysis of Forgetting in Continual Gynecological Image Segmentation

    Amal Saqib, Tausifa Jan Saleem, Numan Saeed +1

    cs.CVcs.LGarXiv:2608.13660v12026
  13. Invisible Shortcuts: Why Vision Encoders Know Your Camera

    Vladan Stojnić, Ryan Ramos, Giorgos Kordopatis-Zilos +2

    cs.CVcs.LGarXiv:2608.05424v12026
  14. ContinualSkillBench: Can LLM Agents Truly Evolve Their Capabilities?

    Tianyi Guan, Yiding Wang, Haotong Yang +5

    cs.AIcs.CLcs.LGarXiv:2608.03874v12026
  15. GradCuit: Credit-Assigned Gradient Flow Enables Robust and Interpretable Test-Time Latent Reasoning

    Zhaoxin Yu, Qi Shen, Hengli Li +4

    cs.LGcs.CLarXiv:2608.02585v22026
  16. CutClean: Neural Network Pruning for Privacy-Preserving Inference

    Leonardo Magliolo, Vito Paolo Pastore, Giuseppe Valenzise +1

    cs.LGcs.AIarXiv:2608.13773v12026
  17. Intelligent Detection of Mechanical, Electrical, and Plumbing (MEP) Metrics Based on 2D Floor Plans

    Tarandeep Singh Mandhiratta, ANK Zaman, Abdul-Rahman Mawlood-Yunis

    cs.CVcs.AIcs.HCarXiv:2608.14317v12026
  18. Wrong but Useful: Trajectory Value Beyond Answer Correctness in Multi-Agent Messages

    Chih-Hsuan Yang, Anjir Ahmed Chowdhury, Cheng-Hau Yang +7

    cs.AIcs.CLcs.LGarXiv:2608.14375v12026
  19. RecipeNet: A Hierarchical Transformer for Recipe Data

    Pin-Yen Huang, Sachin Chhabra, Prasanth Sai Gouripeddi +2

    cs.LGcs.AIarXiv:2608.14505v12026
  20. Architecture and Affordances of PLAUD: Performative Latents and Unsupervised DDSP

    Błażej Kotowski, Frederic Font

    cs.SDcs.HCcs.LGarXiv:2608.13724v12026
  21. Convex losses and their applications to SVM, SVR, and Shallow Neural Networks

    Filippo Portera

    cs.LGarXiv:2608.14288v12026
  22. Building AI-Intensive Software with AI: Early Results and a Cautionary Tale on Measuring Development Cost

    Victor Barros de Miranda Neves, Kiev Santos da Gama, Vinicius Cardoso Garcia

    cs.SEcs.AIcs.LGarXiv:2608.13730v12026
  23. Relevant but Incomplete: Referential Dangling as a Paradigm-Level Failure Mode in Hard Prompt Compression

    Zhengpei Hu, Kai Li, Dapeng Fu +5

    cs.CLcs.LGarXiv:2608.04569v12026
  24. Unknown Unknowns: Model Misspecification in Machine Learning for Physics

    Juan Cruz-Martinez, Carolina Cuesta-Lazaro, Alexander Held +1

    physics.data-anastro-ph.COastro-ph.GAarXiv:2608.13633v12026
  25. Learning Unsteady Aneurysm Hemodynamics with Physics-Informed DeepONets

    Oscar L. Cruz-Gonzalez, Valérie Deplano, Badih Ghattas

    stat.MLcs.LGphysics.flu-dynarXiv:2608.13629v12026
  26. AgilePE: Autonomous UAV Pursuit-Evasion via Self-Play Reinforcement Learning

    Wenhao Tang, Tianyang Chen, Zhejun Cui +9

    cs.ROcs.LGarXiv:2608.14135v12026
  27. Interpretable MEG Decoding of Perceived Speech: Cortical Sources and the Stimulus Features That Drive Retrieval

    Ilia Semenkov, Daria Kleeva, Ivan Dakhtin +2

    cs.LGcs.SDq-bio.NCarXiv:2608.01481v12026
  28. On the Brittleness of Maximum Likelihood Estimation for Gaussian Process Hyperparameter Optimization

    Tyler R. Johnson, Kian Ben-Jacob, Christopher P. Muller +1

    stat.MLcs.LGstat.MEarXiv:2608.13793v12026
  29. Continual Learning in Transition

    Zhiyan Hou, Dan Zhang, Tao Feng +11

    cs.LGcs.AIarXiv:2608.06216v22026
  30. An AI4AI Framework for Visual Token Pruning

    Zhen Liu, Wenli Huang, Wei Song +3

    cs.LGcs.CVarXiv:2608.07193v12026
  31. Expected Free Energy-based Informative Path Planning for Robotic Mars Exploration

    Ajith Anil Meera, Pablo Lanillos, Wouter Kouw

    cs.ROcs.ITcs.LGarXiv:2608.14466v12026
  32. Non-Parametric Spatiotemporal Trajectory Prediction via State-Conditioned Transition Sampling

    Michael Fore, Akshay Jain, Justin Downes +2

    cs.LGarXiv:2608.14349v12026
  33. Deep Reinforcement Learning solution for pickup and delivery routing problems with time window and capacity constraints

    Andrew Soroka, Alex Meshcheryakov, Sergey Gerasimov

    cs.LGarXiv:2608.14156v12026
  34. HarnessOpt-Bench: Evaluating LLMs at Harness Optimization

    Varun Ursekar, Apaar Shanker, Yash Maurya +4

    cs.AIcs.CLcs.LGarXiv:2608.06301v12026
  35. Mind the Long Tail: Understanding the Difficulty of Delay Detection in Business Processes

    Keyvan Amiri Elyasi, Lukas Kirchdorfer, Heiner Stuckenschmidt

    cs.LGcs.AIarXiv:2608.14367v12026
  36. OneDayAgent: Towards a Long-Horizon Harness for Autonomous Agents

    Jingsheng Zheng, Xinyuan Fang, Jintian Zhang +3

    cs.CLcs.AIcs.HCarXiv:2608.05013v12026
  37. Rollplex: Cross-Phase GPU Spatial Sharing for Vision Language Model Post-Training

    Hanfeng Lu, Tianyu Feng, Suyi Li +8

    cs.LGcs.DCarXiv:2608.14498v12026
  38. Buy the Rumor, Sell the News: When Is News Priced In?

    Alireza Kargarzadeh, Nariman Khaledian, Navid Parvini +2

    cs.AIcs.LGq-fin.STarXiv:2608.14014v12026
  39. On-Policy Delta Distillation for Multilingual Math Reasoning

    Byeongho Heo, Jaehui Hwang, Sangdoo Yun +1

    cs.CLcs.LGarXiv:2608.05802v12026
  40. SPOT: Sparse Probing and Outcome Calibration for On-Policy Distillation

    Zikun Qu, Min Zhang, Mingze Kong +4

    cs.LGcs.AIarXiv:2608.04419v12026
  41. Progressive Agent Skill Generation via Reinforcement Learning

    Junhao Shen, Zhanqiu Zhang, Yiwen Guo +1

    cs.LGcs.CLarXiv:2608.01678v12026
  42. TSDS-Toolbox: A Toolbox for Measuring Time-Series Dataset Similarity

    Yen-Ku Liu, Hongjie Chen, Ryan A. Rossi +1

    cs.LGarXiv:2608.08119v12026
  43. From Passive Delegates to Strategic Negotiators: Reinforcing Social Reasoning in Small Language Models with SocialRL

    Wenyue Hua, Zachary Huang, Tyler Payne +3

    cs.AIcs.CLcs.LGarXiv:2608.13787v12026
  44. Identifiability and Order-Dimension Limits of In-Context Learning on Partial Orders

    Faizanuddin Ansari, Debanjan Dutta, Swagatam Das

    cs.LGarXiv:2608.14004v12026
  45. Fixed-Budget Gaussian Volume Encoding with Structure-Aware Allocation

    Michael R. Martin, Joseph Insley, Victor A. Mateevitsi +2

    cs.CVcs.AIcs.CEarXiv:2608.14112v12026
  46. L-FNO: Lorentzian Fourier Neural Operator for Stochastic Event Dynamics

    Songhee Kang, Jihoon Kang

    cs.LGstat.MLarXiv:2608.13562v12026
  47. Multi-Objective Bayesian Optimization for Model Merging

    Utkarsh Agarwal, Vamshi Bonagiri, Raul Astudillo +1

    cs.LGcs.AIarXiv:2608.14264v12026
  48. Post-training Quantization for Hybrid Iterative Generative Models

    Jing Gao, Junyi Wu, Wei Wang +2

    cs.LGarXiv:2608.13932v12026
  49. MINT: A Universal Zero-Shot Predictor for Transaction Data

    Parameswaran Kamalaruban, Viktor Drobnyi, Maeve Madigan +3

    cs.LGcs.CLarXiv:2608.14198v12026
  50. Clearing the Fog: Towards Installing and Refining Proactive Exploration Capabilities in LLM Agents

    Zhizhao Guan, Chen Huang, Ziming Liu +5

    cs.AIcs.LGarXiv:2608.14339v12026
  51. Designing Compact Neural Architectures via Neuron Gating and Mixed Activation

    Abhishek Shukla, Ankur Sinha, Faiz Hamid

    cs.LGcs.AIarXiv:2608.14443v12026
  52. ATLAS: Discovering Agent Strategies through LLM-Guided Abstraction and Automata Learning

    Ignacio D. Lopez-Miguel, Andreas Happe, Jürgen Cito +3

    cs.SEcs.LGarXiv:2608.14352v12026
  53. Probabilistic indirect models for undrained shear strength: addressing significant data missing and variability with advanced imputation and machine learning techniques

    Haibin Xiong, Shaoheng Dai, Peng Lan +4

    cs.LGcs.DBarXiv:2608.13934v12026
  54. On-Policy Delta Distillation

    Byeongho Heo, Jaehui Hwang, Sangdoo Yun +1

    cs.LGcs.CLarXiv:2607.15161v12026
    Summaries:한국어
  55. Weak-to-Strong Generalization via Direct On-Policy Distillation

    Shiyuan Feng, Huan-ang Gao, Haohan Chi +7

    cs.LGcs.AIcs.CLarXiv:2607.05394v22026
    Summaries:한국어
  56. Qwen-RobotManip Technical Report: Alignment Unlocks Scale for Robotic Manipulation Foundation Models

    Haoqi Yuan, Zhixuan Liang, Anzhe Chen +20

    cs.ROcs.CVcs.LGarXiv:2606.17846v22026
    Summaries:한국어
  57. Attention Is All You Need

    Ashish Vaswani, Noam Shazeer, Niki Parmar +5

    cs.CLcs.LGarXiv:1706.03762v72017
  58. Densely Connected Convolutional Networks

    Gao Huang, Zhuang Liu, Laurens van der Maaten +1

    cs.CVcs.LGarXiv:1608.06993v52016
    Summaries:한국어
  59. Regime-Conditional Verification: Correctness Estimation for Adapting and Monitoring Safety Classifiers

    Thiago Sandoval, Ufuk Topcu

    cs.AIcs.CLcs.CRarXiv:2608.14089v12026
  60. CForce: Boosting Parallel Decoding for dLLMs via Consistency Forcing

    Yuji Ren, Chenkai Xu, Zhuocheng Gong +2

    cs.LGcs.AIcs.CLarXiv:2608.13925v12026