Machine Learning
Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
19,561 to 19,620 of 20,132
ThriftAttention: Selective Mixed Precision for Long-Context FP4 Attention
Joe Sharratt
cs.LGarXiv:2605.23081v12026The Distributional View of Knowledge Distillation
Gordei Verbii, Juho Lee
stat.MLcs.LGarXiv:2608.15215v12026Summaries:한국어Interpretable Cross-Lingual Alignment in Small Language Models: Probing Cultural and Pragmatic Reasoning in Japanese-English Bilingual LLMs
Florian Braun
cs.CLcs.LGarXiv:2608.14896v12026A Counterexample to the Tang Zhang Schatten Norm Conjecture and Sharp Positive Results
Zijian Zeng, Houde Liu, Kurunathan Ratnavelu
math.COcs.LGmath.FAarXiv:2608.15558v12026Generative Learning of Separatrices
Ellis R. Crabtree, Dimitris G. Giovanis, Anastasia Georgiou +2
cs.LGmath.DSstat.MLarXiv:2608.14743v12026Distribution-free false-alarm calibration and chance-corrected spatial evaluation for industrial anomaly detection
Jie Deng
cs.CVcs.LGarXiv:2608.15090v12026Shape Operator PCA: Curvature-Aware Projections for Geometric Machine Learning
Alexandre L. M. Levada
cs.LGcs.AIcs.CVarXiv:2608.15313v12026What Makes a Good Layer? Assessing the Layer-Wise Intrinsic Properties of Music Foundation Models
Angelos-Nikolaos Kanatas, Yuexuan Kong, Pablo Alonso-Jiménez +2
cs.SDcs.LGeess.ASarXiv:2608.14819v12026LLMs Can Predict Failure Risk, But Struggle to Predict Which Collaboration Protocol Pays Off: Cost-Aware Protocol Routing Across Reasoning Tasks
Chih-Hsuan Yang, Jingyan Jiang, Cheng-Hau Yang +4
cs.AIcs.CLcs.LGarXiv:2608.14927v12026Uncovering Hidden Leptonic Correlations with Flow Matching and Autoencoders
Haruto Kitagawa, Satsuki Nishimura, Hajime Otsuka
hep-phcs.LGhep-tharXiv:2608.15042v12026IP Protection in the Era of Visual Generative AI: A Survey
Zhuan Shi, Shunchang Liu, Alireza Dehghanpour Farashah +8
cs.CVcs.CRcs.LGarXiv:2608.14730v12026Workspace Topology as an Attack Vector in Agentic Coding Assistants
Alexandre G. R. Day, Pradeep Yadlapalli, Sriram Venkatapathy +9
cs.CRcs.AIcs.CLarXiv:2608.14876v12026Prompting is not enough: supervised baselines and leakage control for measuring shared decision-making with LLMs in pediatric encounters
Bernardo Modenesi, Jody Lin, Kimberly Kaphingst +4
cs.CLcs.AIcs.LGarXiv:2608.14792v12026L3Cube-IndicQuest v2: A Large-Scale Multilingual Benchmark for Evaluating Factual Knowledge of Large Language Models Across Indic Languages
Rinit Jain, Tirthraj Mahajan, Advait Joshi +1
cs.CLcs.LGarXiv:2608.15535v12026LLM Safety Alignment in Low-Resource Languages: A Systematic Literature Review
Valdini Douglace Lemofouet, Blessing Ngozi Uzor, Paula Chikaodinaka Anyanwu +9
cs.CLcs.AIcs.LGarXiv:2608.14626v12026Adaptive surrogate modeling for high-dimensional spatio-temporal output
Berkcan Kapusuzoglu, Shunsaku Matsumoto, Yoshitomo Miyagi +2
cs.CEcs.AIcs.LGarXiv:2608.17250v12026Position: Fairness Failure in Generative Models is an Evaluation Problem
Mariia Vladimirova, Jean-Yves Franceschi, Thibaut Issenhuth
cs.LGcs.AIarXiv:2608.16974v12026A Deep Learning Model for Spatially Clustered Data via Differentiable Cluster Assignment
Kexuan Li, Weidong Ma
stat.MLcs.LGarXiv:2608.14968v12026Earth Observation Foundation Models for Terrestrial Ecohydrology: From Representation Learning to Process Inference
Yi Yu, Jian Peng, Yucheng Lin +2
cs.LGcs.CVphysics.bio-pharXiv:2608.15282v12026Real-Time State-of-Health Estimation and Online Degradation Prognosis from Partial Battery Discharge Using Physics-Informed Neural Networks
Begoña Ispizua, Serio Gil-López, Leire Arrizabalaga +1
cs.LGarXiv:2608.14764v12026Zero-Shot Adaptation of Medical Vision Foundation Models for High-Frequency Micro-Ultrasound Prostate Segmentation
Ayusha Abbas, Saram Abbas, Kabita Adhikari
cs.CVcs.LGarXiv:2608.14796v12026Cross-Modal Ultrasound-MRI Learning for Fetal Brain Ventricular Volumetry and Abnormality Screening
Yuhao Huang, Yuanji Zhang, Yuhuan Lu +3
eess.IVcs.AIcs.CVarXiv:2608.14763v12026Uncertainty Identifies Difficult Samples Across Methods: A Multi-Task Study on a Heterogeneous Skin Lesion Dataset
Leon Koole, Jiapan Guo, Matias Valdenegro-Toro
cs.CVcs.LGarXiv:2608.14768v12026Degeneracy Counting Quantum Algorithm using Decoherence
Malay Marut Das, Mark A. Novotny, Yaroslav Koshka
cs.LGquant-pharXiv:2608.14941v12026NPU Offloading of a Frozen Visual Encoder for Robot Policy Training
Hyojun Yun, Seungjae Won, Hyungpil Moon
cs.ROcs.ARcs.LGarXiv:2608.15002v12026Spinning Conformal Correlators from Neural Networks
Manas Dogra, James Halverson, Joydeep Naskar
hep-thcs.LGarXiv:2608.15001v12026Teach and Grow: An Agent-Centered Architecture for General Robot Learning
Chang Nie, Zhe Liu, Hesheng Wang
cs.ROcs.AIcs.CVarXiv:2608.17209v12026Probing the Prefill: Detecting Code Vulnerabilities via Latent Activations
Alizishaan Khatri
cs.CRcs.AIcs.LGarXiv:2608.16970v12026Iterative tensor network transformations for element-wise evaluation of elementary and filtering functions
Xiao Wang, Tomohiro Hashizume, Pia Siegl +1
cs.LGcond-mat.stat-mechcs.AIarXiv:2608.17135v12026J-Miner: Recovering Executable Decision Knowledge from Language-Model Classifiers
Yunfan Gao, Xinyi Huang, Tao Sheng +3
cs.LGcs.CLarXiv:2608.17063v12026From Abductive Explanations to Global Logical Rules for Node Classification in SGCs
Bryan Lima Cavalcante, Thiago Alves Rocha
cs.LGcs.AIcs.LOarXiv:2608.17103v12026Domain-Adapted Molecular Language Models for Efficient Search of Make-on-Demand Libraries
Henrik Wille, Luis-Finley Schütz, Felix Strieth-Kalthoff
cs.LGcs.AIcs.CLarXiv:2608.17567v12026Integrating Novelty and Surprise for Experience Prioritization and Exploration in Image-Based Reinforcement Learning
Hoda Yamani, Henry Williams, Bruce A. MacDonald
cs.LGcs.AIarXiv:2608.17373v12026FedPref: Federated Preference Learning for Structured Radiology Report Extraction
Flint Xiaofeng Fan, Cheston Tan, Yew-Soon Ong +1
cs.AIcs.LGarXiv:2608.16971v12026Leveraging generative hallucination and biophysics-informed modeling for unified biomolecular sequence-structure co-design
Xuefeng Liu, Mingxuan Cao, Xiao Luo +5
q-bio.QMcs.AIcs.LGarXiv:2608.17381v12026No Gaussian Required: Contrastive Inverse Dynamics for JEPA World Models
Jack Boylan, Chris Hokamp
cs.LGcs.AIarXiv:2608.17542v12026When AI Designs AI: Innovation or Imitation?
Yikang Yang, Zhengxin Yang, Luzhou Peng +4
cs.AIcs.LGarXiv:2608.17471v12026Q-Learning With World Models
Perry Dong, Yueru Jia, Chelsea Finn +1
cs.LGcs.AIarXiv:2608.17163v12026When to Review: Spaced Repetition for Continual Pre-Training of Language Models
Alankar Atreya, Devesh Batra, Yoages Kumar Mantri +3
cs.AIcs.LGarXiv:2608.17530v12026The Role of Feedback Alignment in Self-Distillation
Semih Kara, Oğuzhan Ersoy
cs.AIcs.LGarXiv:2606.11173v12026ARISE: An adaptive residual-informed stability ensemble for feature selection in small-sample biomedical omics
Zardad Khan, Amjad Ali, Naz Gul +2
stat.MLcs.LGarXiv:2608.14866v12026Rethinking Continual Experience Internalization for Self-Evolving LLM Agents
Jingwen Chen, Wenkai Yang, Shengda Fan +7
cs.CLcs.LGarXiv:2606.04703v12026Echo-Memory: A Controlled Study of Memory in Action World Models
Wayne King, Zeyue Xue, Yuxuan Bian +13
cs.CVcs.GRcs.LGarXiv:2606.09803v12026Hardening Agent Benchmarks with Adversarial Hacker-Fixer Loops
Ziqian Zhong, Ivgeni Segal, Ivan Bercovich +3
cs.CRcs.AIcs.LGarXiv:2606.08960v12026Routing Divergence Is Not Evidence of Behavioral Influence in Same-Weight MoE Self-Distillation
Cedric Caruzzo, Donggeun Yoo, Tae Soo Kim
cs.LGcs.AIcs.CLarXiv:2608.15787v12026Deep Embedded Multiplicative DMD for Algebra-Preserving Koopman Learning
Kelan Gray, Finlay Brown, Nicolas Boullé +1
cs.LGmath.DSmath.NAarXiv:2606.05131v12026Latent Reasoning with Normalizing Flows
Guancheng Tu, Xiangjun Fu, Suhao Yu +5
cs.CLcs.LGarXiv:2606.06447v12026ICA Lens: Interpreting Language Models Without Training Another Dictionary
Sida Liu, Feijiang Han
cs.LGcs.AIcs.CLarXiv:2606.11722v12026TuneJury: An Open Metric for Improving Music Generation Preference Alignment
Yonghyun Kim, Junwon Lee, Haiwen Xia +5
cs.SDcs.AIcs.LGarXiv:2606.17006v12026TokenPilot: Cache-Efficient Context Management for LLM Agents
Buqiang Xu, Zirui Xue, Dianmou Chen +12
cs.CLcs.AIcs.LGarXiv:2606.17016v12026BRAID: Learning Equilibrium Maps in Interdependent Security Games via Weight-Tied Iterative Graph Neural Networks
Elnaz Nowrouzi, Zhiqun Zuo, Xueru Zhang +1
cs.GTcs.LGarXiv:2608.14856v12026How Does Reasoning Flow? Tracing Attention-Induced Information Flow for Targeted RL in LLMs
Zhichen Dong, Yang Li, Yuhan Sun +9
cs.LGcs.CLarXiv:2606.10646v12026Time-Series Foundation Model Embeddings for Remaining Useful Life Estimation
Amir El-Ghoussani, Michele De Vita, Ronald Naumann +1
cs.LGcs.AIarXiv:2606.11990v32026From AGI to ASI
Tim Genewein, Matija Franklin, Alexander Lerchner +11
cs.AIcs.CYcs.LGarXiv:2606.12683v12026Quickest Detection of Hallucination Onset: Delay Bounds and Learned CUSUM Statistics
Igor Itkin
cs.LGcs.AIcs.CLarXiv:2606.12476v32026LabVLA: Grounding Vision-Language-Action Models in Scientific Laboratories
Baochang Ren, Xinjie Liu, Xi Chen +15
cs.CLcs.AIcs.LGarXiv:2606.13578v22026Dense Supervision, Sparse Updates: On the Sparsity and Geometry of On-Policy Distillation
Guo Yu, Wenlin Liu, Yulan Hu +3
cs.LGarXiv:2606.13657v32026Human Universal Grasping
Kevin Yuanbo Wu, Tianxing Zhou, Isaac Tu +5
cs.ROcs.AIcs.CVarXiv:2606.17054v12026GD$^2$PO: Mitigating Multi-Reward Conflicts via Group-Dynamic reward-Decoupled Policy Optimization
Haotian Liu, Yihao Liu, Jingwei Ni +11
cs.LGarXiv:2606.16771v12026LoopCoder-v2: Only Loop Once for Efficient Test-Time Computation Scaling
Jian Yang, Shawn Guo, Wei Zhang +16
cs.LGcs.AIarXiv:2606.18023v12026