Machine Learning

Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

20,101 to 20,160 of 20,199

  1. SPEAR: Structure Property Explainability with Attention Regularization

    Aditya Raghavan, Utkarsh Pratiush, Dalton A. Pearl +4

    cond-mat.mtrl-scics.LGarXiv:2608.13826v12026
  2. Conditional Neural Optimal Transport for Predicting Cellular Phenotypes from Molecular Structure

    Gauthier Avité, Maxime Sanchez-Renauld, Nicolas Bourriez +1

    cs.CVcs.LGarXiv:2608.14293v12026
  3. Approximate Muon with low-rank adapters

    Ben Anson, Conor Houghton, Edward Milsom

    cs.LGarXiv:2608.14492v12026
  4. CytoBERT: A Foundation Model for Cytometry Data

    Syed Abdul Haseeb Qadri, Bjarne C. Hiller, Felix Blanke +7

    cs.LGarXiv:2608.14414v12026
  5. LP-NAS: Linear Programming-based Neural Architecture Search

    Abhishek Shukla, Ankur Sinha, Faiz Hamid

    cs.LGcs.AIarXiv:2608.14472v12026
  6. A Four-Axis Trustworthiness Benchmark for LLM-as-Judge in Principle-Based Regulation

    Dipankar Sarkar

    cs.CRcs.AIcs.CLarXiv:2608.14329v12026
  7. A Graph-Based Reinforcement Learning Framework for Structured Drift Diagnosis and Recovery in Autonomous LLM Agents

    Ismail El Hamraoui, Sagar Jose, Nicolas Bureau +1

    cs.AIcs.LGcs.MAarXiv:2608.14109v12026
  8. CoinRAG: Contextualized Information Nugget KV Cache Reuse for Long-Context RAG

    Gyuwan Kim, Cheoneum Park, Tao Yang

    cs.CLcs.AIcs.IRarXiv:2608.07458v12026
  9. VoiceDesigner: Text-to-Voice Generation and Editing via Unified Diffusion Modeling and Data Augmentation

    Jiarui Hai, Karan Thakkar, Ke Chen +5

    eess.AScs.LGarXiv:2608.13613v12026
  10. What to Preserve, Where to Adapt: A Depth-Wise Analysis of Forgetting in Continual Gynecological Image Segmentation

    Amal Saqib, Tausifa Jan Saleem, Numan Saeed +1

    cs.CVcs.LGarXiv:2608.13660v12026
  11. Invisible Shortcuts: Why Vision Encoders Know Your Camera

    Vladan Stojnić, Ryan Ramos, Giorgos Kordopatis-Zilos +2

    cs.CVcs.LGarXiv:2608.05424v12026
  12. ContinualSkillBench: Can LLM Agents Truly Evolve Their Capabilities?

    Tianyi Guan, Yiding Wang, Haotong Yang +5

    cs.AIcs.CLcs.LGarXiv:2608.03874v12026
  13. GradCuit: Credit-Assigned Gradient Flow Enables Robust and Interpretable Test-Time Latent Reasoning

    Zhaoxin Yu, Qi Shen, Hengli Li +4

    cs.LGcs.CLarXiv:2608.02585v22026
  14. CutClean: Neural Network Pruning for Privacy-Preserving Inference

    Leonardo Magliolo, Vito Paolo Pastore, Giuseppe Valenzise +1

    cs.LGcs.AIarXiv:2608.13773v12026
  15. Intelligent Detection of Mechanical, Electrical, and Plumbing (MEP) Metrics Based on 2D Floor Plans

    Tarandeep Singh Mandhiratta, ANK Zaman, Abdul-Rahman Mawlood-Yunis

    cs.CVcs.AIcs.HCarXiv:2608.14317v12026
  16. Wrong but Useful: Trajectory Value Beyond Answer Correctness in Multi-Agent Messages

    Chih-Hsuan Yang, Anjir Ahmed Chowdhury, Cheng-Hau Yang +7

    cs.AIcs.CLcs.LGarXiv:2608.14375v12026
  17. RecipeNet: A Hierarchical Transformer for Recipe Data

    Pin-Yen Huang, Sachin Chhabra, Prasanth Sai Gouripeddi +2

    cs.LGcs.AIarXiv:2608.14505v12026
  18. Architecture and Affordances of PLAUD: Performative Latents and Unsupervised DDSP

    Błażej Kotowski, Frederic Font

    cs.SDcs.HCcs.LGarXiv:2608.13724v12026
  19. Convex losses and their applications to SVM, SVR, and Shallow Neural Networks

    Filippo Portera

    cs.LGarXiv:2608.14288v12026
  20. Building AI-Intensive Software with AI: Early Results and a Cautionary Tale on Measuring Development Cost

    Victor Barros de Miranda Neves, Kiev Santos da Gama, Vinicius Cardoso Garcia

    cs.SEcs.AIcs.LGarXiv:2608.13730v12026
  21. Relevant but Incomplete: Referential Dangling as a Paradigm-Level Failure Mode in Hard Prompt Compression

    Zhengpei Hu, Kai Li, Dapeng Fu +5

    cs.CLcs.LGarXiv:2608.04569v12026
  22. Unknown Unknowns: Model Misspecification in Machine Learning for Physics

    Juan Cruz-Martinez, Carolina Cuesta-Lazaro, Alexander Held +1

    physics.data-anastro-ph.COastro-ph.GAarXiv:2608.13633v12026
  23. Learning Unsteady Aneurysm Hemodynamics with Physics-Informed DeepONets

    Oscar L. Cruz-Gonzalez, Valérie Deplano, Badih Ghattas

    stat.MLcs.LGphysics.flu-dynarXiv:2608.13629v12026
  24. AgilePE: Autonomous UAV Pursuit-Evasion via Self-Play Reinforcement Learning

    Wenhao Tang, Tianyang Chen, Zhejun Cui +9

    cs.ROcs.LGarXiv:2608.14135v12026
  25. Interpretable MEG Decoding of Perceived Speech: Cortical Sources and the Stimulus Features That Drive Retrieval

    Ilia Semenkov, Daria Kleeva, Ivan Dakhtin +2

    cs.LGcs.SDq-bio.NCarXiv:2608.01481v12026
  26. On the Brittleness of Maximum Likelihood Estimation for Gaussian Process Hyperparameter Optimization

    Tyler R. Johnson, Kian Ben-Jacob, Christopher P. Muller +1

    stat.MLcs.LGstat.MEarXiv:2608.13793v12026
  27. Continual Learning in Transition

    Zhiyan Hou, Dan Zhang, Tao Feng +11

    cs.LGcs.AIarXiv:2608.06216v22026
  28. An AI4AI Framework for Visual Token Pruning

    Zhen Liu, Wenli Huang, Wei Song +3

    cs.LGcs.CVarXiv:2608.07193v12026
  29. Expected Free Energy-based Informative Path Planning for Robotic Mars Exploration

    Ajith Anil Meera, Pablo Lanillos, Wouter Kouw

    cs.ROcs.ITcs.LGarXiv:2608.14466v12026
  30. Non-Parametric Spatiotemporal Trajectory Prediction via State-Conditioned Transition Sampling

    Michael Fore, Akshay Jain, Justin Downes +2

    cs.LGarXiv:2608.14349v12026
  31. Deep Reinforcement Learning solution for pickup and delivery routing problems with time window and capacity constraints

    Andrew Soroka, Alex Meshcheryakov, Sergey Gerasimov

    cs.LGarXiv:2608.14156v12026
  32. HarnessOpt-Bench: Evaluating LLMs at Harness Optimization

    Varun Ursekar, Apaar Shanker, Yash Maurya +4

    cs.AIcs.CLcs.LGarXiv:2608.06301v12026
  33. Mind the Long Tail: Understanding the Difficulty of Delay Detection in Business Processes

    Keyvan Amiri Elyasi, Lukas Kirchdorfer, Heiner Stuckenschmidt

    cs.LGcs.AIarXiv:2608.14367v12026
  34. OneDayAgent: Towards a Long-Horizon Harness for Autonomous Agents

    Jingsheng Zheng, Xinyuan Fang, Jintian Zhang +3

    cs.CLcs.AIcs.HCarXiv:2608.05013v12026
  35. Rollplex: Cross-Phase GPU Spatial Sharing for Vision Language Model Post-Training

    Hanfeng Lu, Tianyu Feng, Suyi Li +8

    cs.LGcs.DCarXiv:2608.14498v12026
  36. Buy the Rumor, Sell the News: When Is News Priced In?

    Alireza Kargarzadeh, Nariman Khaledian, Navid Parvini +2

    cs.AIcs.LGq-fin.STarXiv:2608.14014v12026
  37. On-Policy Delta Distillation for Multilingual Math Reasoning

    Byeongho Heo, Jaehui Hwang, Sangdoo Yun +1

    cs.CLcs.LGarXiv:2608.05802v12026
  38. SPOT: Sparse Probing and Outcome Calibration for On-Policy Distillation

    Zikun Qu, Min Zhang, Mingze Kong +4

    cs.LGcs.AIarXiv:2608.04419v12026
  39. Progressive Agent Skill Generation via Reinforcement Learning

    Junhao Shen, Zhanqiu Zhang, Yiwen Guo +1

    cs.LGcs.CLarXiv:2608.01678v12026
  40. TSDS-Toolbox: A Toolbox for Measuring Time-Series Dataset Similarity

    Yen-Ku Liu, Hongjie Chen, Ryan A. Rossi +1

    cs.LGarXiv:2608.08119v12026
  41. From Passive Delegates to Strategic Negotiators: Reinforcing Social Reasoning in Small Language Models with SocialRL

    Wenyue Hua, Zachary Huang, Tyler Payne +3

    cs.AIcs.CLcs.LGarXiv:2608.13787v12026
  42. Identifiability and Order-Dimension Limits of In-Context Learning on Partial Orders

    Faizanuddin Ansari, Debanjan Dutta, Swagatam Das

    cs.LGarXiv:2608.14004v12026
  43. Fixed-Budget Gaussian Volume Encoding with Structure-Aware Allocation

    Michael R. Martin, Joseph Insley, Victor A. Mateevitsi +2

    cs.CVcs.AIcs.CEarXiv:2608.14112v12026
  44. L-FNO: Lorentzian Fourier Neural Operator for Stochastic Event Dynamics

    Songhee Kang, Jihoon Kang

    cs.LGstat.MLarXiv:2608.13562v12026
  45. Multi-Objective Bayesian Optimization for Model Merging

    Utkarsh Agarwal, Vamshi Bonagiri, Raul Astudillo +1

    cs.LGcs.AIarXiv:2608.14264v12026
  46. Post-training Quantization for Hybrid Iterative Generative Models

    Jing Gao, Junyi Wu, Wei Wang +2

    cs.LGarXiv:2608.13932v12026
  47. MINT: A Universal Zero-Shot Predictor for Transaction Data

    Parameswaran Kamalaruban, Viktor Drobnyi, Maeve Madigan +3

    cs.LGcs.CLarXiv:2608.14198v12026
  48. Clearing the Fog: Towards Installing and Refining Proactive Exploration Capabilities in LLM Agents

    Zhizhao Guan, Chen Huang, Ziming Liu +5

    cs.AIcs.LGarXiv:2608.14339v12026
  49. Designing Compact Neural Architectures via Neuron Gating and Mixed Activation

    Abhishek Shukla, Ankur Sinha, Faiz Hamid

    cs.LGcs.AIarXiv:2608.14443v12026
  50. ATLAS: Discovering Agent Strategies through LLM-Guided Abstraction and Automata Learning

    Ignacio D. Lopez-Miguel, Andreas Happe, Jürgen Cito +3

    cs.SEcs.LGarXiv:2608.14352v12026
  51. Probabilistic indirect models for undrained shear strength: addressing significant data missing and variability with advanced imputation and machine learning techniques

    Haibin Xiong, Shaoheng Dai, Peng Lan +4

    cs.LGcs.DBarXiv:2608.13934v12026
  52. On-Policy Delta Distillation

    Byeongho Heo, Jaehui Hwang, Sangdoo Yun +1

    cs.LGcs.CLarXiv:2607.15161v12026
    Summaries:한국어
  53. Weak-to-Strong Generalization via Direct On-Policy Distillation

    Shiyuan Feng, Huan-ang Gao, Haohan Chi +7

    cs.LGcs.AIcs.CLarXiv:2607.05394v22026
    Summaries:한국어
  54. Qwen-RobotManip Technical Report: Alignment Unlocks Scale for Robotic Manipulation Foundation Models

    Haoqi Yuan, Zhixuan Liang, Anzhe Chen +20

    cs.ROcs.CVcs.LGarXiv:2606.17846v22026
    Summaries:한국어
  55. Attention Is All You Need

    Ashish Vaswani, Noam Shazeer, Niki Parmar +5

    cs.CLcs.LGarXiv:1706.03762v72017
  56. Densely Connected Convolutional Networks

    Gao Huang, Zhuang Liu, Laurens van der Maaten +1

    cs.CVcs.LGarXiv:1608.06993v52016
    Summaries:한국어
  57. Regime-Conditional Verification: Correctness Estimation for Adapting and Monitoring Safety Classifiers

    Thiago Sandoval, Ufuk Topcu

    cs.AIcs.CLcs.CRarXiv:2608.14089v12026
  58. CForce: Boosting Parallel Decoding for dLLMs via Consistency Forcing

    Yuji Ren, Chenkai Xu, Zhuocheng Gong +2

    cs.LGcs.AIcs.CLarXiv:2608.13925v12026
  59. Don't Claim Benchmark-Oriented Optimization Improves General Coding Capability -- Diverse Evaluation Is Required

    Egor Shibaev, Vera Kudrevskaia, Timur Galimzyanov +9

    cs.LGcs.AIcs.SEarXiv:2608.13566v12026
  60. When Denoising Hurts: Rethinking the Terminal Step of Diffusion Time Series Forecasters -- Extended Version

    Dat Nguyen-Cong, Luong Tran, Tung Kieu

    cs.LGarXiv:2608.14067v12026