Machine Learning

Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

20,341 to 20,400 of 20,454

  1. Model-agnostic Retrieval-Augmented Extended Forecasting for time series

    Juan Pablo Villa Serna, Rohan Asthana, Vasileios Belagiannis

    cs.LGarXiv:2608.14054v12026
  2. Geometric Filtering of LLM-Generated Samples for Few-Shot Text Classification

    Benjamín Schindler, Gonzalo A. Ruz

    cs.LGcs.CLarXiv:2608.13866v12026
  3. Structure-Guided Spatiotemporal Attention Graph Neural Network for Traffic Flow Prediction

    Xuanmian He, Can Li, Wanjing Ma

    cs.LGcs.AIarXiv:2608.14177v12026
  4. RestoreKV: Recovering Full-Cache Behavior Under Aggressive Query-Agnostic KV Cache Eviction

    Changwoo Baek, Seungjun Shin, Kyeongbo Kong

    cs.CLcs.LGarXiv:2608.01247v12026
  5. ARCHead: Activation-Metric Residual Correction for Large Language Model Output Heads

    Şuayp Talha Kocabay, Talha Rüzgar Akkuş, Kamer Ali Yuksel

    cs.CLcs.LGarXiv:2608.02703v12026
  6. Round-Trip Consistency: Bidirectional Diffusion Models Can Predict Their Own Rollout Errors

    Alexander Scheinker

    stat.MLcs.LGphysics.comp-pharXiv:2608.00675v22026
  7. SAGE: Surrogate-gradient Adaptation via Attention-Guided Entropy for Spiking Transformers

    Kiran Nair, Rodrigue Rizk, KC Santosh

    cs.LGcs.AIcs.CVarXiv:2608.13702v12026
  8. AgentStream: How Well Do Self-Evolving LLM Agents Perform Under Streaming Tasks?

    Dong Yan, Jian Liang, Dapeng Hu +4

    cs.AIcs.LGarXiv:2608.00155v12026
  9. TRUE-Colon: Exposing a Consistent Transfer Asymmetry in Real-Time Polyp Detection

    Sebastian Doerrich, Andreas Franz Schwab, Francesco Di Salvo +3

    eess.IVcs.CVcs.LGarXiv:2608.13711v12026
  10. Resource-Adaptive Primal-Dual Learning for One-Warehouse Multi-Store Systems with Censored Demand

    Jiameng Lyu

    cs.LGmath.OCarXiv:2608.14096v12026
  11. AutoSchema: Live Schema Grounding for Agentic Text-to-Sparql over Heterogeneous Knowledge Graphs

    Yiming Zhang, Koji Tsuda

    cs.LGarXiv:2608.14228v12026
  12. 3DZip: Spatial-Aware Feature Diversity-Guided Token Compression for 3D Question Answering

    Changwoo Baek, Kyeongbo Kong

    cs.CVcs.LGarXiv:2608.01185v12026
  13. Maglev: Sliding Recurrent Memory

    Bo Liu, Qiang Liu

    cs.LGarXiv:2608.02870v22026
  14. Any-OPD: Heterogeneous On-Policy Distillation for Flow-Matching Models via Representation-Space Bridging

    Siming Fu, Zheming Fu, Ruizhe He +7

    cs.LGcs.CVarXiv:2608.03316v12026
  15. DeaMoE: Efficient MoE Structure for Fast Small-Batch Decoding

    Zewen Jin, Shen Fu, Zeping Duan +8

    cs.LGcs.AIarXiv:2608.14385v12026
  16. SPEAR: Structure Property Explainability with Attention Regularization

    Aditya Raghavan, Utkarsh Pratiush, Dalton A. Pearl +4

    cond-mat.mtrl-scics.LGarXiv:2608.13826v12026
  17. Conditional Neural Optimal Transport for Predicting Cellular Phenotypes from Molecular Structure

    Gauthier Avité, Maxime Sanchez-Renauld, Nicolas Bourriez +1

    cs.CVcs.LGarXiv:2608.14293v12026
  18. Approximate Muon with low-rank adapters

    Ben Anson, Conor Houghton, Edward Milsom

    cs.LGarXiv:2608.14492v12026
  19. CytoBERT: A Foundation Model for Cytometry Data

    Syed Abdul Haseeb Qadri, Bjarne C. Hiller, Felix Blanke +7

    cs.LGarXiv:2608.14414v12026
  20. LP-NAS: Linear Programming-based Neural Architecture Search

    Abhishek Shukla, Ankur Sinha, Faiz Hamid

    cs.LGcs.AIarXiv:2608.14472v12026
  21. A Four-Axis Trustworthiness Benchmark for LLM-as-Judge in Principle-Based Regulation

    Dipankar Sarkar

    cs.CRcs.AIcs.CLarXiv:2608.14329v12026
  22. A Graph-Based Reinforcement Learning Framework for Structured Drift Diagnosis and Recovery in Autonomous LLM Agents

    Ismail El Hamraoui, Sagar Jose, Nicolas Bureau +1

    cs.AIcs.LGcs.MAarXiv:2608.14109v12026
  23. CoinRAG: Contextualized Information Nugget KV Cache Reuse for Long-Context RAG

    Gyuwan Kim, Cheoneum Park, Tao Yang

    cs.CLcs.AIcs.IRarXiv:2608.07458v12026
  24. VoiceDesigner: Text-to-Voice Generation and Editing via Unified Diffusion Modeling and Data Augmentation

    Jiarui Hai, Karan Thakkar, Ke Chen +5

    eess.AScs.LGarXiv:2608.13613v12026
  25. What to Preserve, Where to Adapt: A Depth-Wise Analysis of Forgetting in Continual Gynecological Image Segmentation

    Amal Saqib, Tausifa Jan Saleem, Numan Saeed +1

    cs.CVcs.LGarXiv:2608.13660v12026
  26. Invisible Shortcuts: Why Vision Encoders Know Your Camera

    Vladan Stojnić, Ryan Ramos, Giorgos Kordopatis-Zilos +2

    cs.CVcs.LGarXiv:2608.05424v12026
  27. ContinualSkillBench: Can LLM Agents Truly Evolve Their Capabilities?

    Tianyi Guan, Yiding Wang, Haotong Yang +5

    cs.AIcs.CLcs.LGarXiv:2608.03874v12026
  28. GradCuit: Credit-Assigned Gradient Flow Enables Robust and Interpretable Test-Time Latent Reasoning

    Zhaoxin Yu, Qi Shen, Hengli Li +4

    cs.LGcs.CLarXiv:2608.02585v22026
  29. CutClean: Neural Network Pruning for Privacy-Preserving Inference

    Leonardo Magliolo, Vito Paolo Pastore, Giuseppe Valenzise +1

    cs.LGcs.AIarXiv:2608.13773v12026
  30. Intelligent Detection of Mechanical, Electrical, and Plumbing (MEP) Metrics Based on 2D Floor Plans

    Tarandeep Singh Mandhiratta, ANK Zaman, Abdul-Rahman Mawlood-Yunis

    cs.CVcs.AIcs.HCarXiv:2608.14317v12026
  31. Wrong but Useful: Trajectory Value Beyond Answer Correctness in Multi-Agent Messages

    Chih-Hsuan Yang, Anjir Ahmed Chowdhury, Cheng-Hau Yang +7

    cs.AIcs.CLcs.LGarXiv:2608.14375v12026
  32. RecipeNet: A Hierarchical Transformer for Recipe Data

    Pin-Yen Huang, Sachin Chhabra, Prasanth Sai Gouripeddi +2

    cs.LGcs.AIarXiv:2608.14505v12026
  33. Architecture and Affordances of PLAUD: Performative Latents and Unsupervised DDSP

    Błażej Kotowski, Frederic Font

    cs.SDcs.HCcs.LGarXiv:2608.13724v12026
  34. Convex losses and their applications to SVM, SVR, and Shallow Neural Networks

    Filippo Portera

    cs.LGarXiv:2608.14288v12026
  35. Building AI-Intensive Software with AI: Early Results and a Cautionary Tale on Measuring Development Cost

    Victor Barros de Miranda Neves, Kiev Santos da Gama, Vinicius Cardoso Garcia

    cs.SEcs.AIcs.LGarXiv:2608.13730v12026
  36. Relevant but Incomplete: Referential Dangling as a Paradigm-Level Failure Mode in Hard Prompt Compression

    Zhengpei Hu, Kai Li, Dapeng Fu +5

    cs.CLcs.LGarXiv:2608.04569v12026
  37. Unknown Unknowns: Model Misspecification in Machine Learning for Physics

    Juan Cruz-Martinez, Carolina Cuesta-Lazaro, Alexander Held +1

    physics.data-anastro-ph.COastro-ph.GAarXiv:2608.13633v12026
  38. Learning Unsteady Aneurysm Hemodynamics with Physics-Informed DeepONets

    Oscar L. Cruz-Gonzalez, Valérie Deplano, Badih Ghattas

    stat.MLcs.LGphysics.flu-dynarXiv:2608.13629v12026
  39. AgilePE: Autonomous UAV Pursuit-Evasion via Self-Play Reinforcement Learning

    Wenhao Tang, Tianyang Chen, Zhejun Cui +9

    cs.ROcs.LGarXiv:2608.14135v12026
  40. Interpretable MEG Decoding of Perceived Speech: Cortical Sources and the Stimulus Features That Drive Retrieval

    Ilia Semenkov, Daria Kleeva, Ivan Dakhtin +2

    cs.LGcs.SDq-bio.NCarXiv:2608.01481v12026
  41. On the Brittleness of Maximum Likelihood Estimation for Gaussian Process Hyperparameter Optimization

    Tyler R. Johnson, Kian Ben-Jacob, Christopher P. Muller +1

    stat.MLcs.LGstat.MEarXiv:2608.13793v12026
  42. Continual Learning in Transition

    Zhiyan Hou, Dan Zhang, Tao Feng +11

    cs.LGcs.AIarXiv:2608.06216v22026
  43. An AI4AI Framework for Visual Token Pruning

    Zhen Liu, Wenli Huang, Wei Song +3

    cs.LGcs.CVarXiv:2608.07193v12026
  44. Expected Free Energy-based Informative Path Planning for Robotic Mars Exploration

    Ajith Anil Meera, Pablo Lanillos, Wouter Kouw

    cs.ROcs.ITcs.LGarXiv:2608.14466v12026
  45. Non-Parametric Spatiotemporal Trajectory Prediction via State-Conditioned Transition Sampling

    Michael Fore, Akshay Jain, Justin Downes +2

    cs.LGarXiv:2608.14349v12026
  46. Deep Reinforcement Learning solution for pickup and delivery routing problems with time window and capacity constraints

    Andrew Soroka, Alex Meshcheryakov, Sergey Gerasimov

    cs.LGarXiv:2608.14156v12026
  47. HarnessOpt-Bench: Evaluating LLMs at Harness Optimization

    Varun Ursekar, Apaar Shanker, Yash Maurya +4

    cs.AIcs.CLcs.LGarXiv:2608.06301v12026
  48. Mind the Long Tail: Understanding the Difficulty of Delay Detection in Business Processes

    Keyvan Amiri Elyasi, Lukas Kirchdorfer, Heiner Stuckenschmidt

    cs.LGcs.AIarXiv:2608.14367v12026
  49. OneDayAgent: Towards a Long-Horizon Harness for Autonomous Agents

    Jingsheng Zheng, Xinyuan Fang, Jintian Zhang +3

    cs.CLcs.AIcs.HCarXiv:2608.05013v12026
  50. Rollplex: Cross-Phase GPU Spatial Sharing for Vision Language Model Post-Training

    Hanfeng Lu, Tianyu Feng, Suyi Li +8

    cs.LGcs.DCarXiv:2608.14498v12026
  51. Buy the Rumor, Sell the News: When Is News Priced In?

    Alireza Kargarzadeh, Nariman Khaledian, Navid Parvini +2

    cs.AIcs.LGq-fin.STarXiv:2608.14014v12026
  52. On-Policy Delta Distillation for Multilingual Math Reasoning

    Byeongho Heo, Jaehui Hwang, Sangdoo Yun +1

    cs.CLcs.LGarXiv:2608.05802v12026
  53. SPOT: Sparse Probing and Outcome Calibration for On-Policy Distillation

    Zikun Qu, Min Zhang, Mingze Kong +4

    cs.LGcs.AIarXiv:2608.04419v12026
  54. Progressive Agent Skill Generation via Reinforcement Learning

    Junhao Shen, Zhanqiu Zhang, Yiwen Guo +1

    cs.LGcs.CLarXiv:2608.01678v12026
  55. TSDS-Toolbox: A Toolbox for Measuring Time-Series Dataset Similarity

    Yen-Ku Liu, Hongjie Chen, Ryan A. Rossi +1

    cs.LGarXiv:2608.08119v12026
  56. From Passive Delegates to Strategic Negotiators: Reinforcing Social Reasoning in Small Language Models with SocialRL

    Wenyue Hua, Zachary Huang, Tyler Payne +3

    cs.AIcs.CLcs.LGarXiv:2608.13787v12026
  57. Identifiability and Order-Dimension Limits of In-Context Learning on Partial Orders

    Faizanuddin Ansari, Debanjan Dutta, Swagatam Das

    cs.LGarXiv:2608.14004v12026
  58. Fixed-Budget Gaussian Volume Encoding with Structure-Aware Allocation

    Michael R. Martin, Joseph Insley, Victor A. Mateevitsi +2

    cs.CVcs.AIcs.CEarXiv:2608.14112v12026
  59. L-FNO: Lorentzian Fourier Neural Operator for Stochastic Event Dynamics

    Songhee Kang, Jihoon Kang

    cs.LGstat.MLarXiv:2608.13562v12026
  60. Multi-Objective Bayesian Optimization for Model Merging

    Utkarsh Agarwal, Vamshi Bonagiri, Raul Astudillo +1

    cs.LGcs.AIarXiv:2608.14264v12026