Machine Learning
Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
19,921 to 19,948 of 19,948
Adaptive Protection for Evolutionary Feature Construction in Symbolic Regression with Application to Credit Classification
Hengzhe Zhang, Qi Chen, Bing Xue +3
cs.LGcs.NEarXiv:2608.14209v12026Amplified Does Not Mean Predictive: Reasoning Behaviors in Thinking Models
Jean de Dieu Nyandwi, Leena Mathur, Yonatan Bisk +2
cs.CLcs.AIcs.CVarXiv:2608.13760v12026TailBooster: A Dual-Layer Generative Framework for Extreme Value Augmentation with Operational Validity Enforcement
Karim Aly, Alexei Sharpanskykh, Jacco Hoekstra
cs.LGcs.AIarXiv:2608.11951v12026AI Evaluation Should Work With Humans
Jan Kulveit, Gavin Leech, Tomáš Gavenčiak +1
cs.AIcs.LGarXiv:2608.13577v12026Knowledge-guided Pattern Discovery via Coupled Tensor Factorizations
Gaute Johannessen, Geert Roelof van der Ploeg, Evrim Acar
cs.LGarXiv:2608.13234v12026Improved Large Language Diffusion Models
Shen Nie, Qiyang Min, Shaoxuan Xu +7
cs.CLcs.AIcs.LGarXiv:2606.25331v12026Causal-rCM: A Unified Teacher-Forcing and Self-Forcing Open Recipe for Autoregressive Diffusion Distillation in Streaming Video Generation and Interactive World Models
Kaiwen Zheng, Guande He, Min Zhao +7
cs.CVcs.LGarXiv:2606.25473v12026Sequence to Sequence Learning with Neural Networks
Ilya Sutskever, Oriol Vinyals, Quoc V. Le
cs.CLcs.LGarXiv:1409.3215v32014Sparse Orthogonal Regression Technique: A Spectral Framework for Equation Discovery, Approximation, and Integration
Sabin Roman, Ljupco Todorovski, Saso Dzeroski
cs.LGarXiv:2608.13504v12026Playing Atari with Deep Reinforcement Learning
Volodymyr Mnih, Koray Kavukcuoglu, David Silver +4
cs.LGarXiv:1312.5602v12013Vidu S1: A Real-Time Interactive Video Generation Model
Jintao Zhang, Kai Jiang, Jintao Chen +24
cs.CVcs.LGarXiv:2607.03118v22026KVAE: Family of Tokenizers for Multimodal Generative Models
Andrey Shutkin, Denis Parkhomenko, Ivan Kirillov +11
cs.CVcs.LGcs.SDarXiv:2608.05798v12026RoFormer: Enhanced Transformer with Rotary Position Embedding
Jianlin Su, Yu Lu, Shengfeng Pan +3
cs.CLcs.AIcs.LGarXiv:2104.09864v52021Demystifying On-Policy Distillation: Roles, Pathologies, and Regulations
Rui Wang, Hongru Wang, Yi Chen +4
cs.CLcs.LGarXiv:2607.13399v12026On the global feature importance for interpretable and trustworthy heat demand forecasting
Milan Zdravković
cs.LGeess.SYarXiv:2608.13039v12026Forecast Collapse in Time-Series Foundation Models
Shu Wan, Miles Ma, Hank Zhu +4
cs.LGcs.AIcs.CEarXiv:2608.14106v12026DanceOPD: On-Policy Generative Field Distillation
Wei Zhou, Xiongwei Zhu, Zelin Xu +8
cs.CVcs.CLcs.LGarXiv:2606.27377v22026MobileMem: Learning from a Year of Mobile Experiences
Xinle Deng, Yida Xue, Xiangyuan Ru +14
cs.AIcs.CLcs.LGarXiv:2608.13606v12026On-Policy Self-Distillation without Any Supervision
Yijiang Li, Bingyang Wang, Yijun Liang +3
cs.LGarXiv:2608.06296v22026Distilling the Knowledge in a Neural Network
Geoffrey Hinton, Oriol Vinyals, Jeff Dean
stat.MLcs.LGcs.NEarXiv:1503.02531v12015SkillOpt-Lite: Better and Faster Agent Self-evolution via One Line of Vibe
Yifei Shen, Bo Li, Xinjie Zhang
cs.SEcs.AIcs.LGarXiv:2607.03451v12026Generative Adversarial Networks
Ian J. Goodfellow, Jean Pouget-Abadie, Mehdi Mirza +5
stat.MLcs.LGarXiv:1406.2661v12014MOPD: Multi-Teacher On-Policy Distillation for Capability Integration in LLM Post-Training
Wenhan Ma, Jianyu Wei, Liang Zhao +10
cs.CLcs.LGarXiv:2606.30406v12026Proximal Policy Optimization Algorithms
John Schulman, Filip Wolski, Prafulla Dhariwal +2
cs.LGarXiv:1707.06347v22017Intern-S2-Preview: Scientific Agentic Foundation Model
Lei Bai, Jiaqi Cao, Chiyu Chen +122
cs.LGcs.CLcs.CVarXiv:2608.13505v12026AgenticDataBench: A Comprehensive Benchmark for Data Agents
Zhaoyan Sun, Shan Zhong, Daizhou Wen +10
cs.DBcs.AIcs.CLarXiv:2607.01647v12026The State-Prediction Separation Hypothesis
Giovanni Monea, Nathan Godey, Kianté Brantley +1
cs.CLcs.AIcs.LGarXiv:2607.01218v12026WARP: Weight-Space Analysis for Recovering Training Data Portfolios
Tzu-Heng Huang, Aditya Goyal, John Cooper +1
cs.LGarXiv:2607.01686v12026