Machine Learning

Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

19,561 to 19,620 of 20,132

  1. ThriftAttention: Selective Mixed Precision for Long-Context FP4 Attention

    Joe Sharratt

    cs.LGarXiv:2605.23081v12026
  2. The Distributional View of Knowledge Distillation

    Gordei Verbii, Juho Lee

    stat.MLcs.LGarXiv:2608.15215v12026
    Summaries:한국어
  3. Interpretable Cross-Lingual Alignment in Small Language Models: Probing Cultural and Pragmatic Reasoning in Japanese-English Bilingual LLMs

    Florian Braun

    cs.CLcs.LGarXiv:2608.14896v12026
  4. A Counterexample to the Tang Zhang Schatten Norm Conjecture and Sharp Positive Results

    Zijian Zeng, Houde Liu, Kurunathan Ratnavelu

    math.COcs.LGmath.FAarXiv:2608.15558v12026
  5. Generative Learning of Separatrices

    Ellis R. Crabtree, Dimitris G. Giovanis, Anastasia Georgiou +2

    cs.LGmath.DSstat.MLarXiv:2608.14743v12026
  6. Distribution-free false-alarm calibration and chance-corrected spatial evaluation for industrial anomaly detection

    Jie Deng

    cs.CVcs.LGarXiv:2608.15090v12026
  7. Shape Operator PCA: Curvature-Aware Projections for Geometric Machine Learning

    Alexandre L. M. Levada

    cs.LGcs.AIcs.CVarXiv:2608.15313v12026
  8. What Makes a Good Layer? Assessing the Layer-Wise Intrinsic Properties of Music Foundation Models

    Angelos-Nikolaos Kanatas, Yuexuan Kong, Pablo Alonso-Jiménez +2

    cs.SDcs.LGeess.ASarXiv:2608.14819v12026
  9. LLMs Can Predict Failure Risk, But Struggle to Predict Which Collaboration Protocol Pays Off: Cost-Aware Protocol Routing Across Reasoning Tasks

    Chih-Hsuan Yang, Jingyan Jiang, Cheng-Hau Yang +4

    cs.AIcs.CLcs.LGarXiv:2608.14927v12026
  10. Uncovering Hidden Leptonic Correlations with Flow Matching and Autoencoders

    Haruto Kitagawa, Satsuki Nishimura, Hajime Otsuka

    hep-phcs.LGhep-tharXiv:2608.15042v12026
  11. IP Protection in the Era of Visual Generative AI: A Survey

    Zhuan Shi, Shunchang Liu, Alireza Dehghanpour Farashah +8

    cs.CVcs.CRcs.LGarXiv:2608.14730v12026
  12. Workspace Topology as an Attack Vector in Agentic Coding Assistants

    Alexandre G. R. Day, Pradeep Yadlapalli, Sriram Venkatapathy +9

    cs.CRcs.AIcs.CLarXiv:2608.14876v12026
  13. Prompting is not enough: supervised baselines and leakage control for measuring shared decision-making with LLMs in pediatric encounters

    Bernardo Modenesi, Jody Lin, Kimberly Kaphingst +4

    cs.CLcs.AIcs.LGarXiv:2608.14792v12026
  14. L3Cube-IndicQuest v2: A Large-Scale Multilingual Benchmark for Evaluating Factual Knowledge of Large Language Models Across Indic Languages

    Rinit Jain, Tirthraj Mahajan, Advait Joshi +1

    cs.CLcs.LGarXiv:2608.15535v12026
  15. LLM Safety Alignment in Low-Resource Languages: A Systematic Literature Review

    Valdini Douglace Lemofouet, Blessing Ngozi Uzor, Paula Chikaodinaka Anyanwu +9

    cs.CLcs.AIcs.LGarXiv:2608.14626v12026
  16. Adaptive surrogate modeling for high-dimensional spatio-temporal output

    Berkcan Kapusuzoglu, Shunsaku Matsumoto, Yoshitomo Miyagi +2

    cs.CEcs.AIcs.LGarXiv:2608.17250v12026
  17. Position: Fairness Failure in Generative Models is an Evaluation Problem

    Mariia Vladimirova, Jean-Yves Franceschi, Thibaut Issenhuth

    cs.LGcs.AIarXiv:2608.16974v12026
  18. A Deep Learning Model for Spatially Clustered Data via Differentiable Cluster Assignment

    Kexuan Li, Weidong Ma

    stat.MLcs.LGarXiv:2608.14968v12026
  19. Earth Observation Foundation Models for Terrestrial Ecohydrology: From Representation Learning to Process Inference

    Yi Yu, Jian Peng, Yucheng Lin +2

    cs.LGcs.CVphysics.bio-pharXiv:2608.15282v12026
  20. Real-Time State-of-Health Estimation and Online Degradation Prognosis from Partial Battery Discharge Using Physics-Informed Neural Networks

    Begoña Ispizua, Serio Gil-López, Leire Arrizabalaga +1

    cs.LGarXiv:2608.14764v12026
  21. Zero-Shot Adaptation of Medical Vision Foundation Models for High-Frequency Micro-Ultrasound Prostate Segmentation

    Ayusha Abbas, Saram Abbas, Kabita Adhikari

    cs.CVcs.LGarXiv:2608.14796v12026
  22. Cross-Modal Ultrasound-MRI Learning for Fetal Brain Ventricular Volumetry and Abnormality Screening

    Yuhao Huang, Yuanji Zhang, Yuhuan Lu +3

    eess.IVcs.AIcs.CVarXiv:2608.14763v12026
  23. Uncertainty Identifies Difficult Samples Across Methods: A Multi-Task Study on a Heterogeneous Skin Lesion Dataset

    Leon Koole, Jiapan Guo, Matias Valdenegro-Toro

    cs.CVcs.LGarXiv:2608.14768v12026
  24. Degeneracy Counting Quantum Algorithm using Decoherence

    Malay Marut Das, Mark A. Novotny, Yaroslav Koshka

    cs.LGquant-pharXiv:2608.14941v12026
  25. NPU Offloading of a Frozen Visual Encoder for Robot Policy Training

    Hyojun Yun, Seungjae Won, Hyungpil Moon

    cs.ROcs.ARcs.LGarXiv:2608.15002v12026
  26. Spinning Conformal Correlators from Neural Networks

    Manas Dogra, James Halverson, Joydeep Naskar

    hep-thcs.LGarXiv:2608.15001v12026
  27. Teach and Grow: An Agent-Centered Architecture for General Robot Learning

    Chang Nie, Zhe Liu, Hesheng Wang

    cs.ROcs.AIcs.CVarXiv:2608.17209v12026
  28. Probing the Prefill: Detecting Code Vulnerabilities via Latent Activations

    Alizishaan Khatri

    cs.CRcs.AIcs.LGarXiv:2608.16970v12026
  29. Iterative tensor network transformations for element-wise evaluation of elementary and filtering functions

    Xiao Wang, Tomohiro Hashizume, Pia Siegl +1

    cs.LGcond-mat.stat-mechcs.AIarXiv:2608.17135v12026
  30. J-Miner: Recovering Executable Decision Knowledge from Language-Model Classifiers

    Yunfan Gao, Xinyi Huang, Tao Sheng +3

    cs.LGcs.CLarXiv:2608.17063v12026
  31. From Abductive Explanations to Global Logical Rules for Node Classification in SGCs

    Bryan Lima Cavalcante, Thiago Alves Rocha

    cs.LGcs.AIcs.LOarXiv:2608.17103v12026
  32. Domain-Adapted Molecular Language Models for Efficient Search of Make-on-Demand Libraries

    Henrik Wille, Luis-Finley Schütz, Felix Strieth-Kalthoff

    cs.LGcs.AIcs.CLarXiv:2608.17567v12026
  33. Integrating Novelty and Surprise for Experience Prioritization and Exploration in Image-Based Reinforcement Learning

    Hoda Yamani, Henry Williams, Bruce A. MacDonald

    cs.LGcs.AIarXiv:2608.17373v12026
  34. FedPref: Federated Preference Learning for Structured Radiology Report Extraction

    Flint Xiaofeng Fan, Cheston Tan, Yew-Soon Ong +1

    cs.AIcs.LGarXiv:2608.16971v12026
  35. Leveraging generative hallucination and biophysics-informed modeling for unified biomolecular sequence-structure co-design

    Xuefeng Liu, Mingxuan Cao, Xiao Luo +5

    q-bio.QMcs.AIcs.LGarXiv:2608.17381v12026
  36. No Gaussian Required: Contrastive Inverse Dynamics for JEPA World Models

    Jack Boylan, Chris Hokamp

    cs.LGcs.AIarXiv:2608.17542v12026
  37. When AI Designs AI: Innovation or Imitation?

    Yikang Yang, Zhengxin Yang, Luzhou Peng +4

    cs.AIcs.LGarXiv:2608.17471v12026
  38. Q-Learning With World Models

    Perry Dong, Yueru Jia, Chelsea Finn +1

    cs.LGcs.AIarXiv:2608.17163v12026
  39. When to Review: Spaced Repetition for Continual Pre-Training of Language Models

    Alankar Atreya, Devesh Batra, Yoages Kumar Mantri +3

    cs.AIcs.LGarXiv:2608.17530v12026
  40. The Role of Feedback Alignment in Self-Distillation

    Semih Kara, Oğuzhan Ersoy

    cs.AIcs.LGarXiv:2606.11173v12026
  41. ARISE: An adaptive residual-informed stability ensemble for feature selection in small-sample biomedical omics

    Zardad Khan, Amjad Ali, Naz Gul +2

    stat.MLcs.LGarXiv:2608.14866v12026
  42. Rethinking Continual Experience Internalization for Self-Evolving LLM Agents

    Jingwen Chen, Wenkai Yang, Shengda Fan +7

    cs.CLcs.LGarXiv:2606.04703v12026
  43. Echo-Memory: A Controlled Study of Memory in Action World Models

    Wayne King, Zeyue Xue, Yuxuan Bian +13

    cs.CVcs.GRcs.LGarXiv:2606.09803v12026
  44. Hardening Agent Benchmarks with Adversarial Hacker-Fixer Loops

    Ziqian Zhong, Ivgeni Segal, Ivan Bercovich +3

    cs.CRcs.AIcs.LGarXiv:2606.08960v12026
  45. Routing Divergence Is Not Evidence of Behavioral Influence in Same-Weight MoE Self-Distillation

    Cedric Caruzzo, Donggeun Yoo, Tae Soo Kim

    cs.LGcs.AIcs.CLarXiv:2608.15787v12026
  46. Deep Embedded Multiplicative DMD for Algebra-Preserving Koopman Learning

    Kelan Gray, Finlay Brown, Nicolas Boullé +1

    cs.LGmath.DSmath.NAarXiv:2606.05131v12026
  47. Latent Reasoning with Normalizing Flows

    Guancheng Tu, Xiangjun Fu, Suhao Yu +5

    cs.CLcs.LGarXiv:2606.06447v12026
  48. ICA Lens: Interpreting Language Models Without Training Another Dictionary

    Sida Liu, Feijiang Han

    cs.LGcs.AIcs.CLarXiv:2606.11722v12026
  49. TuneJury: An Open Metric for Improving Music Generation Preference Alignment

    Yonghyun Kim, Junwon Lee, Haiwen Xia +5

    cs.SDcs.AIcs.LGarXiv:2606.17006v12026
  50. TokenPilot: Cache-Efficient Context Management for LLM Agents

    Buqiang Xu, Zirui Xue, Dianmou Chen +12

    cs.CLcs.AIcs.LGarXiv:2606.17016v12026
  51. BRAID: Learning Equilibrium Maps in Interdependent Security Games via Weight-Tied Iterative Graph Neural Networks

    Elnaz Nowrouzi, Zhiqun Zuo, Xueru Zhang +1

    cs.GTcs.LGarXiv:2608.14856v12026
  52. How Does Reasoning Flow? Tracing Attention-Induced Information Flow for Targeted RL in LLMs

    Zhichen Dong, Yang Li, Yuhan Sun +9

    cs.LGcs.CLarXiv:2606.10646v12026
  53. Time-Series Foundation Model Embeddings for Remaining Useful Life Estimation

    Amir El-Ghoussani, Michele De Vita, Ronald Naumann +1

    cs.LGcs.AIarXiv:2606.11990v32026
  54. From AGI to ASI

    Tim Genewein, Matija Franklin, Alexander Lerchner +11

    cs.AIcs.CYcs.LGarXiv:2606.12683v12026
  55. Quickest Detection of Hallucination Onset: Delay Bounds and Learned CUSUM Statistics

    Igor Itkin

    cs.LGcs.AIcs.CLarXiv:2606.12476v32026
  56. LabVLA: Grounding Vision-Language-Action Models in Scientific Laboratories

    Baochang Ren, Xinjie Liu, Xi Chen +15

    cs.CLcs.AIcs.LGarXiv:2606.13578v22026
  57. Dense Supervision, Sparse Updates: On the Sparsity and Geometry of On-Policy Distillation

    Guo Yu, Wenlin Liu, Yulan Hu +3

    cs.LGarXiv:2606.13657v32026
  58. Human Universal Grasping

    Kevin Yuanbo Wu, Tianxing Zhou, Isaac Tu +5

    cs.ROcs.AIcs.CVarXiv:2606.17054v12026
  59. GD$^2$PO: Mitigating Multi-Reward Conflicts via Group-Dynamic reward-Decoupled Policy Optimization

    Haotian Liu, Yihao Liu, Jingwei Ni +11

    cs.LGarXiv:2606.16771v12026
  60. LoopCoder-v2: Only Loop Once for Efficient Test-Time Computation Scaling

    Jian Yang, Shawn Guo, Wei Zhang +16

    cs.LGcs.AIarXiv:2606.18023v12026