Machine Learning
Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
18,061 to 18,120 of 20,454
(1D) Ordered Tokens Enable Efficient Test-Time Search
Zhitong Gao, Parham Rezaei, Ali Cy +7
cs.CVcs.AIcs.LGarXiv:2604.15453v12026AccelOpt: A Self-Improving LLM Agentic System for AI Accelerator Kernel Optimization
Genghan Zhang, Shaowei Zhu, Anjiang Wei +6
cs.LGcs.CLarXiv:2511.15915v22025Reward Hacking in the Era of Large Models: Mechanisms, Emergent Misalignment, Challenges
Xiaohua Wang, Muzhao Tian, Yuqi Zeng +20
cs.LGarXiv:2604.13602v12026Where does output diversity collapse in post-training?
Constantinos Karouzos, Xingwei Tan, Nikolaos Aletras
cs.CLcs.AIcs.LGarXiv:2604.16027v12026GFT: From Imitation to Reward Fine-Tuning with Unbiased Group Advantages and Dynamic Coefficient Rectification
Wangjie Gan, Miao Pan, Linbo Xi +4
cs.AIcs.LGarXiv:2604.14258v32026An Optimal Transport-driven Approach for Cultivating Latent Space in Online Incremental Learning
Quyen Tran, Hai Nguyen, Hoang Phan +6
cs.LGcs.CVarXiv:2211.16780v42022LongAct: Harnessing Intrinsic Activation Patterns for Long-Context Reinforcement Learning
Bowen Ping, Zijun Chen, Tingfeng Hui +4
cs.LGcs.CLarXiv:2604.14922v12026PRL-Bench: A Comprehensive Benchmark Evaluating LLMs' Capabilities in Frontier Physics Research
Tingjia Miao, Wenkai Jin, Muhua Zhang +19
cs.LGcs.AIphysics.data-anarXiv:2604.15411v12026Maximal Brain Damage Without Data or Optimization: Disrupting Neural Networks via Sign-Bit Flips
Ido Galil, Moshe Kimhi, Ran El-Yaniv
cs.LGcs.AIcs.CVarXiv:2502.07408v22025PerturbRx: Learning Treatment-Conditioned Latent Transitions for Patient Drug Response Prediction
Yoshitaka Inoue, Minoh Jeong, Alfred Hero +2
q-bio.QMcs.LGarXiv:2608.21349v12026TwinTrack: Post-hoc Multi-Rater Calibration for Medical Image Segmentation
Tristan Kirscher, Alexandra Ertl, Klaus Maier-Hein +3
cs.LGarXiv:2604.15950v22026Stargazer: A Scalable Model-Fitting Benchmark Environment for AI Agents under Astrophysical Constraints
Xinge Liu, Terry Jingchen Zhang, Bernhard Schölkopf +2
cs.LGcs.AIarXiv:2604.15664v22026Structured Scaling of AI Discovery Across Diverse Scientific Domains
Haotian Ye, Haowei Lin, Jingyi Tang +30
cs.LGcs.AIarXiv:2604.19341v22026Advanced Linear Algebra with Applications - Part I (Numerical linear algebra for PDEs, machine learning, and data assimilation)
Victorita Dolean, Jemima Tabeart
math.NAcs.LGarXiv:2608.21234v12026The Exceedance Design Effect: Effective Sample Size for Thresholds under Clustering
Adam Noonan
stat.MLcs.LGarXiv:2608.21262v12026EasyVideoR1: Easier RL for Video Understanding
Chuanyu Qin, Chenxu Yang, Qingyi Si +6
cs.CVcs.LGarXiv:2604.16893v12026Truthful Calibration Measures for Sequential Prediction
Anagha Gokul, Jason Hartline, Lunjia Hu +2
cs.DScs.GTcs.LGarXiv:2608.21348v12026Test-Time Adaptation for EEG Foundation Models: A Systematic Study under Real-World Distribution Shifts
Gabriel Jason Lee, Jathurshan Pradeepkumar, Jimeng Sun
cs.LGcs.AIeess.SParXiv:2604.16926v22026Agents Explore but Agents Ignore: LLMs Lack Environmental Curiosity
Leon Engländer, Sophia Althammer, Ahmet Üstün +2
cs.CLcs.LGarXiv:2604.17609v12026Just Repair: A Minimal Denoising Network for Time Series Anomaly Detection
Kadir-Kaan Özer, René Ebeling, Markus Enzweiler
cs.LGcs.AIarXiv:2604.17388v32026On the Transferability of Agricultural Weed Detection Under Cross-Field Distribution Shift
Nikhilesh Prabhakar, Pranuthi Tenali, Wilfredo Abudeye Fernandez +5
cs.CVcs.LGarXiv:2608.21254v12026When Can LLMs Learn to Reason with Weak Supervision?
Salman Rahman, Jingyan Shen, Anna Mordvina +3
cs.LGcs.AIarXiv:2604.18574v12026MathNet: a Global Multimodal Benchmark for Mathematical Reasoning and Retrieval
Shaden Alshammari, Kevin Wen, Abrar Zainal +5
cs.AIcs.DLcs.IRarXiv:2604.18584v22026Event-triggered Implicit Perturbation for Zeroth-Order Fine-Tuning of Spiking Transformers
Tengteng Lei, Prabodh Katti, Rashi Dutt +5
cs.ARcs.LGcs.NEarXiv:2608.21223v12026UDM-GRPO: Stable and Efficient Group Relative Policy Optimization for Uniform Discrete Diffusion Models
Jiaqi Wang, Haoge Deng, Ting Pan +5
cs.CVcs.LGarXiv:2604.18518v42026Daedalus-150M: A Convolution-Attention Hybrid Designed for CPU Inference
Christos Koutsiaris
cs.IRcs.AIcs.CLarXiv:2608.20210v12026Accurate and scalable exchange-correlation with deep learning
Giulia Luise, Chin-Wei Huang, Thijs Vogels +25
physics.chem-phcs.AIcs.CEarXiv:2506.14665v62025From a Static Multi-Level Small Semantic Codebook to a Dynamic Single-Level Large Semantic Codebook for Generative Recommendation
Tianlu Xie, Xin Ku, Mingjie Sun +8
cs.IRcs.LGarXiv:2608.21012v12026CubicSplat: Differentiable Vector Graphics via Error-Bounded Forward Relaxation
Chenglong Liu, Xin Zhang, Yimeng Zhu +5
cs.GRcs.CVcs.LGarXiv:2608.20803v12026Chat2Workflow: A Benchmark for Generating Executable Visual Workflows with Natural Language
Yi Zhong, Buqiang Xu, Yijun Wang +4
cs.CLcs.AIcs.CVarXiv:2604.19667v22026Training DeepFilterNet with Accurate Room Acoustic Simulations Improves Single-Channel Speech Enhancement
Alessia Milo, Georg Götz, Steinar Guðjónsson +3
eess.AScs.LGphysics.comp-pharXiv:2608.20971v12026Rethinking Demonstration Unlearning in Imitation Learning for Robotics
Jiazhuo Li, Yu Zhang, Yiming Fei +4
cs.ROcs.LGarXiv:2608.20784v12026Fine-tuning LLMs for Tourist Trajectory Prediction using Field Experiment Data
Tatsuya Amano, Hirozumi Yamaguchi
cs.CYcs.LGarXiv:2608.20830v12026Learning Prostate Anatomy at Test Time for Cancer Detection in Micro-Ultrasound
Obed Korshie Dzikunu, Mohammad Mahdi Abootorabi, Mohamed Harmanani +7
cs.CVcs.LGarXiv:2608.20557v12026Keyed Provenance Watermarking with Complementary Lattice-Based Secure Aggregation for Federated Learning
Xinyun Liu, Zhi Lu, Yu Chen +1
cs.CRcs.LGarXiv:2608.20580v12026Interpretable Information-Decomposed Brain Graph Learning for fMRI-based Disease Diagnosis
Dengyi Zhao, Zhiheng Zhou, Zihan Wang +2
q-bio.NCcs.LGarXiv:2608.20380v12026Keep Your Friends Close, and the Right Neighbours Closer: Disaster-Conditioned Kernel-Regularized Graph Attention for Building Damage Classification
Fuad Hasan, Chul Min Yeum
cs.CVcs.LGarXiv:2608.20548v12026Rethinking Expressivity and Efficiency in Test-Time Training
Zeyun Zhong, Joya Chen, Manuel Martin +3
cs.LGarXiv:2608.21308v12026Asymmetric Capacity Allocation in Self-Refinement Pipelines
Zhuoyi Yang, Ian G. Harris, Salar Hashemitaheri +7
cs.LGarXiv:2608.21345v12026Time-Aware Tranformer-Based Prediction Model for AECOPD
Weihao Qu, Ling Zheng, Dongyang Wang +2
cs.LGarXiv:2608.21324v12026Human-JEPA: A Human-Centric Vision Model that Perceives and Anticipates
Hui Wei, Licai Sun, Guoying Zhao
cs.CVcs.LGarXiv:2608.21160v12026RDP LoRA: Geometry-Driven Identification for Parameter-Efficient Adaptation in Large Language Models
Yusuf Çelebi, Yağız Asker, Özay Ezerceli +4
cs.LGcs.AIcs.CLarXiv:2604.19321v12026AudioWorldSim: Realistic Binaural Audio Datasets For World Models
Luis Vitor Zerkowski, Luiz Velho
cs.SDcs.LGarXiv:2608.21075v12026Sharing the Control Authority Between Deep Reinforcement Learning and Model Predictive Control: Application to Multi-Class Transportation Networks
Giray Onur, Azita Dabiri, Bart De Schutter
eess.SYcs.LGarXiv:2608.20858v12026TEMPO: Scaling Test-time Training for Large Reasoning Models
Qingyang Zhang, Xinke Kong, Haitao Wu +7
cs.LGarXiv:2604.19295v12026Uncertainty propagation in auto-regressive random neural network models
Janice Adams, Daniele Venturi
stat.MLcs.LGcs.NEarXiv:2608.20483v12026Predicting Resource Efficient Hamiltonian Decomposition for Continuous-Time Quantum Walk Simulations
Mostafa Atallah, Rebekah Herrman, Zain H. Saleem
quant-phcs.LGarXiv:2608.20660v12026MMCORE: MultiModal COnnection with Representation Aligned Latent Embeddings
Zijie Li, Yichun Shi, Jingxiang Sun +8
cs.CVcs.AIcs.LGarXiv:2604.19902v12026aiXamine: Unified Black-Box Evaluation of Cross-Dimensional Trade-offs in LLM Safety, Security, and Privacy
Fatih Deniz, Yazan Boshmaf, Dorde Popovic +1
cs.CRcs.LGarXiv:2608.20554v12026Expert Upcycling: Shifting the Compute-Efficient Frontier of Mixture-of-Experts
Chaitanya Dwivedi, Binxuan Huang, Himanshu Gupta +3
cs.LGcs.AIarXiv:2604.19835v22026DR-Venus: Towards Frontier Edge-Scale Deep Research Agents with Only 10K Open Data
Venus Team, Sunhao Dai, Yong Deng +10
cs.LGcs.AIcs.CLarXiv:2604.19859v12026Robust Discovery of Coarse-Grained Continuum Equations from Microscopic Dynamics
Partha Sarathi Mondal, Manav Kumar Jalan, Anish Kumar +1
cond-mat.softcond-mat.stat-mechcs.LGarXiv:2608.20404v12026SkillLearnBench: Benchmarking Continual Learning Methods for Agent Skill Generation on Real-World Tasks
Shanshan Zhong, Yi Lu, Jingjie Ning +7
cs.CLcs.LGarXiv:2604.20087v12026COMPASS: COntinual Multilingual PEFT with Adaptive Semantic Sampling
Noah Flynn
cs.LGcs.AIcs.CLarXiv:2604.20720v12026Harmonic Torsional Diffusion for Protein-Ligand Flexible Docking
Maksim Zhdanov, Pavel Strashnov, Vladislav Kurenkov
q-bio.BMcs.LGarXiv:2608.20366v12026Across-Design Uncertainty in Short Pricing Panels: Evidence from Simulated Price Trajectories
Pedro Cadahia Delgado
cs.LGecon.EMarXiv:2608.21334v12026SPARCL: Spectral Partitioned Analytic Continual Learning
James Hartley, Zeropy Surio, Daniel Whitmore +2
cs.LGarXiv:2608.21307v12026Temporally Extended Mixture-of-Experts Models
Zeyu Shen, Peter Henderson
cs.LGarXiv:2604.20156v12026Tydra: An Efficient Hybrid Model for Tabular Data
Mieszko Komisarczyk, Saurabh Mathur, Maurice Kraus +2
cs.LGarXiv:2608.21199v12026Capturing Cardiac Cyclicity through Phase-Equivariant Self-Supervised Learning
Blaise Delaney, Dominic Dootson, Juan Jose Juan Castella +5
cs.LGarXiv:2608.21147v12026