Machine Learning

Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

20,161 to 20,193 of 20,193

  1. Adversarial Learning of Classifier-Free Guidance Schedules

    Ashwini Pokle, Alexandre Galashov, Arnaud Doucet +2

    cs.LGarXiv:2608.14038v12026
  2. Training Fair Tabular Foundation Models

    Patrik Kenfack, Jesse C. Cresswell, Anthony L. Caterini +2

    cs.LGcs.AIarXiv:2608.14211v12026
  3. BCMT: Blockwise Causal Memory Transformer

    Rachid Arezki

    cs.CLcs.AIcs.LGarXiv:2608.13578v12026
  4. Hybrid Quantum-inspired Kolmogorov-Arnold Networks for Privacy-Aware Federated Biosignal Learning

    Chun-Hua Lin, Samuel Yen-Chi Chen, Yu-Chao Hsu +7

    cs.LGcs.AIcs.DCarXiv:2608.13914v12026
  5. From Prediction to Intervention: Personalized Meal-Level Glucose Regulation via an LLM Agent

    Mingyu Huang, Weiqing Min, Ying Jin +2

    cs.HCcs.AIcs.LGarXiv:2608.13581v12026
  6. Adaptive Protection for Evolutionary Feature Construction in Symbolic Regression with Application to Credit Classification

    Hengzhe Zhang, Qi Chen, Bing Xue +3

    cs.LGcs.NEarXiv:2608.14209v12026
  7. Amplified Does Not Mean Predictive: Reasoning Behaviors in Thinking Models

    Jean de Dieu Nyandwi, Leena Mathur, Yonatan Bisk +2

    cs.CLcs.AIcs.CVarXiv:2608.13760v12026
  8. TailBooster: A Dual-Layer Generative Framework for Extreme Value Augmentation with Operational Validity Enforcement

    Karim Aly, Alexei Sharpanskykh, Jacco Hoekstra

    cs.LGcs.AIarXiv:2608.11951v12026
  9. AI Evaluation Should Work With Humans

    Jan Kulveit, Gavin Leech, Tomáš Gavenčiak +1

    cs.AIcs.LGarXiv:2608.13577v12026
  10. Knowledge-guided Pattern Discovery via Coupled Tensor Factorizations

    Gaute Johannessen, Geert Roelof van der Ploeg, Evrim Acar

    cs.LGarXiv:2608.13234v12026
  11. Improved Large Language Diffusion Models

    Shen Nie, Qiyang Min, Shaoxuan Xu +7

    cs.CLcs.AIcs.LGarXiv:2606.25331v12026
  12. Causal-rCM: A Unified Teacher-Forcing and Self-Forcing Open Recipe for Autoregressive Diffusion Distillation in Streaming Video Generation and Interactive World Models

    Kaiwen Zheng, Guande He, Min Zhao +7

    cs.CVcs.LGarXiv:2606.25473v12026
  13. Sequence to Sequence Learning with Neural Networks

    Ilya Sutskever, Oriol Vinyals, Quoc V. Le

    cs.CLcs.LGarXiv:1409.3215v32014
  14. Sparse Orthogonal Regression Technique: A Spectral Framework for Equation Discovery, Approximation, and Integration

    Sabin Roman, Ljupco Todorovski, Saso Dzeroski

    cs.LGarXiv:2608.13504v12026
  15. Playing Atari with Deep Reinforcement Learning

    Volodymyr Mnih, Koray Kavukcuoglu, David Silver +4

    cs.LGarXiv:1312.5602v12013
  16. Vidu S1: A Real-Time Interactive Video Generation Model

    Jintao Zhang, Kai Jiang, Jintao Chen +24

    cs.CVcs.LGarXiv:2607.03118v22026
  17. KVAE: Family of Tokenizers for Multimodal Generative Models

    Andrey Shutkin, Denis Parkhomenko, Ivan Kirillov +11

    cs.CVcs.LGcs.SDarXiv:2608.05798v12026
  18. RoFormer: Enhanced Transformer with Rotary Position Embedding

    Jianlin Su, Yu Lu, Shengfeng Pan +3

    cs.CLcs.AIcs.LGarXiv:2104.09864v52021
  19. Demystifying On-Policy Distillation: Roles, Pathologies, and Regulations

    Rui Wang, Hongru Wang, Yi Chen +4

    cs.CLcs.LGarXiv:2607.13399v12026
  20. On the global feature importance for interpretable and trustworthy heat demand forecasting

    Milan Zdravković

    cs.LGeess.SYarXiv:2608.13039v12026
  21. Forecast Collapse in Time-Series Foundation Models

    Shu Wan, Miles Ma, Hank Zhu +4

    cs.LGcs.AIcs.CEarXiv:2608.14106v12026
  22. DanceOPD: On-Policy Generative Field Distillation

    Wei Zhou, Xiongwei Zhu, Zelin Xu +8

    cs.CVcs.CLcs.LGarXiv:2606.27377v22026
  23. MobileMem: Learning from a Year of Mobile Experiences

    Xinle Deng, Yida Xue, Xiangyuan Ru +14

    cs.AIcs.CLcs.LGarXiv:2608.13606v12026
  24. On-Policy Self-Distillation without Any Supervision

    Yijiang Li, Bingyang Wang, Yijun Liang +3

    cs.LGarXiv:2608.06296v22026
  25. Distilling the Knowledge in a Neural Network

    Geoffrey Hinton, Oriol Vinyals, Jeff Dean

    stat.MLcs.LGcs.NEarXiv:1503.02531v12015
  26. SkillOpt-Lite: Better and Faster Agent Self-evolution via One Line of Vibe

    Yifei Shen, Bo Li, Xinjie Zhang

    cs.SEcs.AIcs.LGarXiv:2607.03451v12026
  27. Generative Adversarial Networks

    Ian J. Goodfellow, Jean Pouget-Abadie, Mehdi Mirza +5

    stat.MLcs.LGarXiv:1406.2661v12014
  28. MOPD: Multi-Teacher On-Policy Distillation for Capability Integration in LLM Post-Training

    Wenhan Ma, Jianyu Wei, Liang Zhao +10

    cs.CLcs.LGarXiv:2606.30406v12026
  29. Proximal Policy Optimization Algorithms

    John Schulman, Filip Wolski, Prafulla Dhariwal +2

    cs.LGarXiv:1707.06347v22017
  30. Intern-S2-Preview: Scientific Agentic Foundation Model

    Lei Bai, Jiaqi Cao, Chiyu Chen +122

    cs.LGcs.CLcs.CVarXiv:2608.13505v12026
  31. AgenticDataBench: A Comprehensive Benchmark for Data Agents

    Zhaoyan Sun, Shan Zhong, Daizhou Wen +10

    cs.DBcs.AIcs.CLarXiv:2607.01647v12026
  32. The State-Prediction Separation Hypothesis

    Giovanni Monea, Nathan Godey, Kianté Brantley +1

    cs.CLcs.AIcs.LGarXiv:2607.01218v12026
  33. WARP: Weight-Space Analysis for Recovering Training Data Portfolios

    Tzu-Heng Huang, Aditya Goyal, John Cooper +1

    cs.LGarXiv:2607.01686v12026