Machine Learning

Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

18,061 to 18,120 of 20,454

  1. (1D) Ordered Tokens Enable Efficient Test-Time Search

    Zhitong Gao, Parham Rezaei, Ali Cy +7

    cs.CVcs.AIcs.LGarXiv:2604.15453v12026
  2. AccelOpt: A Self-Improving LLM Agentic System for AI Accelerator Kernel Optimization

    Genghan Zhang, Shaowei Zhu, Anjiang Wei +6

    cs.LGcs.CLarXiv:2511.15915v22025
  3. Reward Hacking in the Era of Large Models: Mechanisms, Emergent Misalignment, Challenges

    Xiaohua Wang, Muzhao Tian, Yuqi Zeng +20

    cs.LGarXiv:2604.13602v12026
  4. Where does output diversity collapse in post-training?

    Constantinos Karouzos, Xingwei Tan, Nikolaos Aletras

    cs.CLcs.AIcs.LGarXiv:2604.16027v12026
  5. GFT: From Imitation to Reward Fine-Tuning with Unbiased Group Advantages and Dynamic Coefficient Rectification

    Wangjie Gan, Miao Pan, Linbo Xi +4

    cs.AIcs.LGarXiv:2604.14258v32026
  6. An Optimal Transport-driven Approach for Cultivating Latent Space in Online Incremental Learning

    Quyen Tran, Hai Nguyen, Hoang Phan +6

    cs.LGcs.CVarXiv:2211.16780v42022
  7. LongAct: Harnessing Intrinsic Activation Patterns for Long-Context Reinforcement Learning

    Bowen Ping, Zijun Chen, Tingfeng Hui +4

    cs.LGcs.CLarXiv:2604.14922v12026
  8. PRL-Bench: A Comprehensive Benchmark Evaluating LLMs' Capabilities in Frontier Physics Research

    Tingjia Miao, Wenkai Jin, Muhua Zhang +19

    cs.LGcs.AIphysics.data-anarXiv:2604.15411v12026
  9. Maximal Brain Damage Without Data or Optimization: Disrupting Neural Networks via Sign-Bit Flips

    Ido Galil, Moshe Kimhi, Ran El-Yaniv

    cs.LGcs.AIcs.CVarXiv:2502.07408v22025
  10. PerturbRx: Learning Treatment-Conditioned Latent Transitions for Patient Drug Response Prediction

    Yoshitaka Inoue, Minoh Jeong, Alfred Hero +2

    q-bio.QMcs.LGarXiv:2608.21349v12026
  11. TwinTrack: Post-hoc Multi-Rater Calibration for Medical Image Segmentation

    Tristan Kirscher, Alexandra Ertl, Klaus Maier-Hein +3

    cs.LGarXiv:2604.15950v22026
  12. Stargazer: A Scalable Model-Fitting Benchmark Environment for AI Agents under Astrophysical Constraints

    Xinge Liu, Terry Jingchen Zhang, Bernhard Schölkopf +2

    cs.LGcs.AIarXiv:2604.15664v22026
  13. Structured Scaling of AI Discovery Across Diverse Scientific Domains

    Haotian Ye, Haowei Lin, Jingyi Tang +30

    cs.LGcs.AIarXiv:2604.19341v22026
  14. Advanced Linear Algebra with Applications - Part I (Numerical linear algebra for PDEs, machine learning, and data assimilation)

    Victorita Dolean, Jemima Tabeart

    math.NAcs.LGarXiv:2608.21234v12026
  15. The Exceedance Design Effect: Effective Sample Size for Thresholds under Clustering

    Adam Noonan

    stat.MLcs.LGarXiv:2608.21262v12026
  16. EasyVideoR1: Easier RL for Video Understanding

    Chuanyu Qin, Chenxu Yang, Qingyi Si +6

    cs.CVcs.LGarXiv:2604.16893v12026
  17. Truthful Calibration Measures for Sequential Prediction

    Anagha Gokul, Jason Hartline, Lunjia Hu +2

    cs.DScs.GTcs.LGarXiv:2608.21348v12026
  18. Test-Time Adaptation for EEG Foundation Models: A Systematic Study under Real-World Distribution Shifts

    Gabriel Jason Lee, Jathurshan Pradeepkumar, Jimeng Sun

    cs.LGcs.AIeess.SParXiv:2604.16926v22026
  19. Agents Explore but Agents Ignore: LLMs Lack Environmental Curiosity

    Leon Engländer, Sophia Althammer, Ahmet Üstün +2

    cs.CLcs.LGarXiv:2604.17609v12026
  20. Just Repair: A Minimal Denoising Network for Time Series Anomaly Detection

    Kadir-Kaan Özer, René Ebeling, Markus Enzweiler

    cs.LGcs.AIarXiv:2604.17388v32026
  21. On the Transferability of Agricultural Weed Detection Under Cross-Field Distribution Shift

    Nikhilesh Prabhakar, Pranuthi Tenali, Wilfredo Abudeye Fernandez +5

    cs.CVcs.LGarXiv:2608.21254v12026
  22. When Can LLMs Learn to Reason with Weak Supervision?

    Salman Rahman, Jingyan Shen, Anna Mordvina +3

    cs.LGcs.AIarXiv:2604.18574v12026
  23. MathNet: a Global Multimodal Benchmark for Mathematical Reasoning and Retrieval

    Shaden Alshammari, Kevin Wen, Abrar Zainal +5

    cs.AIcs.DLcs.IRarXiv:2604.18584v22026
  24. Event-triggered Implicit Perturbation for Zeroth-Order Fine-Tuning of Spiking Transformers

    Tengteng Lei, Prabodh Katti, Rashi Dutt +5

    cs.ARcs.LGcs.NEarXiv:2608.21223v12026
  25. UDM-GRPO: Stable and Efficient Group Relative Policy Optimization for Uniform Discrete Diffusion Models

    Jiaqi Wang, Haoge Deng, Ting Pan +5

    cs.CVcs.LGarXiv:2604.18518v42026
  26. Daedalus-150M: A Convolution-Attention Hybrid Designed for CPU Inference

    Christos Koutsiaris

    cs.IRcs.AIcs.CLarXiv:2608.20210v12026
  27. Accurate and scalable exchange-correlation with deep learning

    Giulia Luise, Chin-Wei Huang, Thijs Vogels +25

    physics.chem-phcs.AIcs.CEarXiv:2506.14665v62025
  28. From a Static Multi-Level Small Semantic Codebook to a Dynamic Single-Level Large Semantic Codebook for Generative Recommendation

    Tianlu Xie, Xin Ku, Mingjie Sun +8

    cs.IRcs.LGarXiv:2608.21012v12026
  29. CubicSplat: Differentiable Vector Graphics via Error-Bounded Forward Relaxation

    Chenglong Liu, Xin Zhang, Yimeng Zhu +5

    cs.GRcs.CVcs.LGarXiv:2608.20803v12026
  30. Chat2Workflow: A Benchmark for Generating Executable Visual Workflows with Natural Language

    Yi Zhong, Buqiang Xu, Yijun Wang +4

    cs.CLcs.AIcs.CVarXiv:2604.19667v22026
  31. Training DeepFilterNet with Accurate Room Acoustic Simulations Improves Single-Channel Speech Enhancement

    Alessia Milo, Georg Götz, Steinar Guðjónsson +3

    eess.AScs.LGphysics.comp-pharXiv:2608.20971v12026
  32. Rethinking Demonstration Unlearning in Imitation Learning for Robotics

    Jiazhuo Li, Yu Zhang, Yiming Fei +4

    cs.ROcs.LGarXiv:2608.20784v12026
  33. Fine-tuning LLMs for Tourist Trajectory Prediction using Field Experiment Data

    Tatsuya Amano, Hirozumi Yamaguchi

    cs.CYcs.LGarXiv:2608.20830v12026
  34. Learning Prostate Anatomy at Test Time for Cancer Detection in Micro-Ultrasound

    Obed Korshie Dzikunu, Mohammad Mahdi Abootorabi, Mohamed Harmanani +7

    cs.CVcs.LGarXiv:2608.20557v12026
  35. Keyed Provenance Watermarking with Complementary Lattice-Based Secure Aggregation for Federated Learning

    Xinyun Liu, Zhi Lu, Yu Chen +1

    cs.CRcs.LGarXiv:2608.20580v12026
  36. Interpretable Information-Decomposed Brain Graph Learning for fMRI-based Disease Diagnosis

    Dengyi Zhao, Zhiheng Zhou, Zihan Wang +2

    q-bio.NCcs.LGarXiv:2608.20380v12026
  37. Keep Your Friends Close, and the Right Neighbours Closer: Disaster-Conditioned Kernel-Regularized Graph Attention for Building Damage Classification

    Fuad Hasan, Chul Min Yeum

    cs.CVcs.LGarXiv:2608.20548v12026
  38. Rethinking Expressivity and Efficiency in Test-Time Training

    Zeyun Zhong, Joya Chen, Manuel Martin +3

    cs.LGarXiv:2608.21308v12026
  39. Asymmetric Capacity Allocation in Self-Refinement Pipelines

    Zhuoyi Yang, Ian G. Harris, Salar Hashemitaheri +7

    cs.LGarXiv:2608.21345v12026
  40. Time-Aware Tranformer-Based Prediction Model for AECOPD

    Weihao Qu, Ling Zheng, Dongyang Wang +2

    cs.LGarXiv:2608.21324v12026
  41. Human-JEPA: A Human-Centric Vision Model that Perceives and Anticipates

    Hui Wei, Licai Sun, Guoying Zhao

    cs.CVcs.LGarXiv:2608.21160v12026
  42. RDP LoRA: Geometry-Driven Identification for Parameter-Efficient Adaptation in Large Language Models

    Yusuf Çelebi, Yağız Asker, Özay Ezerceli +4

    cs.LGcs.AIcs.CLarXiv:2604.19321v12026
  43. AudioWorldSim: Realistic Binaural Audio Datasets For World Models

    Luis Vitor Zerkowski, Luiz Velho

    cs.SDcs.LGarXiv:2608.21075v12026
  44. Sharing the Control Authority Between Deep Reinforcement Learning and Model Predictive Control: Application to Multi-Class Transportation Networks

    Giray Onur, Azita Dabiri, Bart De Schutter

    eess.SYcs.LGarXiv:2608.20858v12026
  45. TEMPO: Scaling Test-time Training for Large Reasoning Models

    Qingyang Zhang, Xinke Kong, Haitao Wu +7

    cs.LGarXiv:2604.19295v12026
  46. Uncertainty propagation in auto-regressive random neural network models

    Janice Adams, Daniele Venturi

    stat.MLcs.LGcs.NEarXiv:2608.20483v12026
  47. Predicting Resource Efficient Hamiltonian Decomposition for Continuous-Time Quantum Walk Simulations

    Mostafa Atallah, Rebekah Herrman, Zain H. Saleem

    quant-phcs.LGarXiv:2608.20660v12026
  48. MMCORE: MultiModal COnnection with Representation Aligned Latent Embeddings

    Zijie Li, Yichun Shi, Jingxiang Sun +8

    cs.CVcs.AIcs.LGarXiv:2604.19902v12026
  49. aiXamine: Unified Black-Box Evaluation of Cross-Dimensional Trade-offs in LLM Safety, Security, and Privacy

    Fatih Deniz, Yazan Boshmaf, Dorde Popovic +1

    cs.CRcs.LGarXiv:2608.20554v12026
  50. Expert Upcycling: Shifting the Compute-Efficient Frontier of Mixture-of-Experts

    Chaitanya Dwivedi, Binxuan Huang, Himanshu Gupta +3

    cs.LGcs.AIarXiv:2604.19835v22026
  51. DR-Venus: Towards Frontier Edge-Scale Deep Research Agents with Only 10K Open Data

    Venus Team, Sunhao Dai, Yong Deng +10

    cs.LGcs.AIcs.CLarXiv:2604.19859v12026
  52. Robust Discovery of Coarse-Grained Continuum Equations from Microscopic Dynamics

    Partha Sarathi Mondal, Manav Kumar Jalan, Anish Kumar +1

    cond-mat.softcond-mat.stat-mechcs.LGarXiv:2608.20404v12026
  53. SkillLearnBench: Benchmarking Continual Learning Methods for Agent Skill Generation on Real-World Tasks

    Shanshan Zhong, Yi Lu, Jingjie Ning +7

    cs.CLcs.LGarXiv:2604.20087v12026
  54. COMPASS: COntinual Multilingual PEFT with Adaptive Semantic Sampling

    Noah Flynn

    cs.LGcs.AIcs.CLarXiv:2604.20720v12026
  55. Harmonic Torsional Diffusion for Protein-Ligand Flexible Docking

    Maksim Zhdanov, Pavel Strashnov, Vladislav Kurenkov

    q-bio.BMcs.LGarXiv:2608.20366v12026
  56. Across-Design Uncertainty in Short Pricing Panels: Evidence from Simulated Price Trajectories

    Pedro Cadahia Delgado

    cs.LGecon.EMarXiv:2608.21334v12026
  57. SPARCL: Spectral Partitioned Analytic Continual Learning

    James Hartley, Zeropy Surio, Daniel Whitmore +2

    cs.LGarXiv:2608.21307v12026
  58. Temporally Extended Mixture-of-Experts Models

    Zeyu Shen, Peter Henderson

    cs.LGarXiv:2604.20156v12026
  59. Tydra: An Efficient Hybrid Model for Tabular Data

    Mieszko Komisarczyk, Saurabh Mathur, Maurice Kraus +2

    cs.LGarXiv:2608.21199v12026
  60. Capturing Cardiac Cyclicity through Phase-Equivariant Self-Supervised Learning

    Blaise Delaney, Dominic Dootson, Juan Jose Juan Castella +5

    cs.LGarXiv:2608.21147v12026