Machine Learning

Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

20,401 to 20,454 of 20,454

  1. Post-training Quantization for Hybrid Iterative Generative Models

    Jing Gao, Junyi Wu, Wei Wang +2

    cs.LGarXiv:2608.13932v12026
  2. MINT: A Universal Zero-Shot Predictor for Transaction Data

    Parameswaran Kamalaruban, Viktor Drobnyi, Maeve Madigan +3

    cs.LGcs.CLarXiv:2608.14198v12026
  3. Clearing the Fog: Towards Installing and Refining Proactive Exploration Capabilities in LLM Agents

    Zhizhao Guan, Chen Huang, Ziming Liu +5

    cs.AIcs.LGarXiv:2608.14339v12026
  4. Designing Compact Neural Architectures via Neuron Gating and Mixed Activation

    Abhishek Shukla, Ankur Sinha, Faiz Hamid

    cs.LGcs.AIarXiv:2608.14443v12026
  5. ATLAS: Discovering Agent Strategies through LLM-Guided Abstraction and Automata Learning

    Ignacio D. Lopez-Miguel, Andreas Happe, Jürgen Cito +3

    cs.SEcs.LGarXiv:2608.14352v12026
  6. Probabilistic indirect models for undrained shear strength: addressing significant data missing and variability with advanced imputation and machine learning techniques

    Haibin Xiong, Shaoheng Dai, Peng Lan +4

    cs.LGcs.DBarXiv:2608.13934v12026
  7. On-Policy Delta Distillation

    Byeongho Heo, Jaehui Hwang, Sangdoo Yun +1

    cs.LGcs.CLarXiv:2607.15161v12026
    Summaries:한국어
  8. Weak-to-Strong Generalization via Direct On-Policy Distillation

    Shiyuan Feng, Huan-ang Gao, Haohan Chi +7

    cs.LGcs.AIcs.CLarXiv:2607.05394v22026
    Summaries:한국어
  9. Qwen-RobotManip Technical Report: Alignment Unlocks Scale for Robotic Manipulation Foundation Models

    Haoqi Yuan, Zhixuan Liang, Anzhe Chen +20

    cs.ROcs.CVcs.LGarXiv:2606.17846v22026
    Summaries:한국어
  10. Attention Is All You Need

    Ashish Vaswani, Noam Shazeer, Niki Parmar +5

    cs.CLcs.LGarXiv:1706.03762v72017
  11. Densely Connected Convolutional Networks

    Gao Huang, Zhuang Liu, Laurens van der Maaten +1

    cs.CVcs.LGarXiv:1608.06993v52016
    Summaries:한국어
  12. Regime-Conditional Verification: Correctness Estimation for Adapting and Monitoring Safety Classifiers

    Thiago Sandoval, Ufuk Topcu

    cs.AIcs.CLcs.CRarXiv:2608.14089v12026
  13. CForce: Boosting Parallel Decoding for dLLMs via Consistency Forcing

    Yuji Ren, Chenkai Xu, Zhuocheng Gong +2

    cs.LGcs.AIcs.CLarXiv:2608.13925v12026
  14. Don't Claim Benchmark-Oriented Optimization Improves General Coding Capability -- Diverse Evaluation Is Required

    Egor Shibaev, Vera Kudrevskaia, Timur Galimzyanov +9

    cs.LGcs.AIcs.SEarXiv:2608.13566v12026
  15. When Denoising Hurts: Rethinking the Terminal Step of Diffusion Time Series Forecasters -- Extended Version

    Dat Nguyen-Cong, Luong Tran, Tung Kieu

    cs.LGarXiv:2608.14067v12026
  16. Think in Latent, Explain in Language: Self-Explainable Latent Reasoning

    Dayuan Zhao, Shengcao Cao, Yu-Xiong Wang +1

    cs.CLcs.AIcs.LGarXiv:2608.13570v12026
  17. Nanbeige4.2-3B on Apple Silicon: Fixing Deployment Bugs and Decreasing Looped Transformer Memory Overhead

    John T. Halloran

    cs.AIcs.LGarXiv:2608.13987v12026
  18. Concept Guidance: Precise, Training-Free Latent Control for Text-to-Image Generation

    Nikolai Röhrich, Isabell Hans, Felix Krause +1

    cs.CVcs.AIcs.LGarXiv:2608.14172v12026
  19. Overcoming Shortcut Learning in Graph Neural Networks through Active Explanation Guidance

    Taraneh Younesian, Steve Azzolin, Antonio Longa +3

    cs.LGcs.AIarXiv:2608.14121v12026
  20. Interactive Analysis of Global Explanations using Aggregated Class Activation Maps for Network Data

    Igor Cherepanov, David Sessler, Alex Ulmer +3

    cs.HCcs.AIcs.LGarXiv:2608.13575v12026
  21. When Does More Correct Data Hurt? Insertion-Stability and the Limits of Dimension-Based Theory

    Joseph Sankoorikal Johny

    cs.LGstat.MLarXiv:2608.14020v12026
  22. Adversarial Learning of Classifier-Free Guidance Schedules

    Ashwini Pokle, Alexandre Galashov, Arnaud Doucet +2

    cs.LGarXiv:2608.14038v12026
  23. Training Fair Tabular Foundation Models

    Patrik Kenfack, Jesse C. Cresswell, Anthony L. Caterini +2

    cs.LGcs.AIarXiv:2608.14211v12026
  24. BCMT: Blockwise Causal Memory Transformer

    Rachid Arezki

    cs.CLcs.AIcs.LGarXiv:2608.13578v12026
  25. Hybrid Quantum-inspired Kolmogorov-Arnold Networks for Privacy-Aware Federated Biosignal Learning

    Chun-Hua Lin, Samuel Yen-Chi Chen, Yu-Chao Hsu +7

    cs.LGcs.AIcs.DCarXiv:2608.13914v12026
  26. From Prediction to Intervention: Personalized Meal-Level Glucose Regulation via an LLM Agent

    Mingyu Huang, Weiqing Min, Ying Jin +2

    cs.HCcs.AIcs.LGarXiv:2608.13581v12026
  27. Adaptive Protection for Evolutionary Feature Construction in Symbolic Regression with Application to Credit Classification

    Hengzhe Zhang, Qi Chen, Bing Xue +3

    cs.LGcs.NEarXiv:2608.14209v12026
  28. Amplified Does Not Mean Predictive: Reasoning Behaviors in Thinking Models

    Jean de Dieu Nyandwi, Leena Mathur, Yonatan Bisk +2

    cs.CLcs.AIcs.CVarXiv:2608.13760v12026
  29. TailBooster: A Dual-Layer Generative Framework for Extreme Value Augmentation with Operational Validity Enforcement

    Karim Aly, Alexei Sharpanskykh, Jacco Hoekstra

    cs.LGcs.AIarXiv:2608.11951v12026
  30. AI Evaluation Should Work With Humans

    Jan Kulveit, Gavin Leech, Tomáš Gavenčiak +1

    cs.AIcs.LGarXiv:2608.13577v12026
  31. Knowledge-guided Pattern Discovery via Coupled Tensor Factorizations

    Gaute Johannessen, Geert Roelof van der Ploeg, Evrim Acar

    cs.LGarXiv:2608.13234v12026
  32. Improved Large Language Diffusion Models

    Shen Nie, Qiyang Min, Shaoxuan Xu +7

    cs.CLcs.AIcs.LGarXiv:2606.25331v12026
  33. Causal-rCM: A Unified Teacher-Forcing and Self-Forcing Open Recipe for Autoregressive Diffusion Distillation in Streaming Video Generation and Interactive World Models

    Kaiwen Zheng, Guande He, Min Zhao +7

    cs.CVcs.LGarXiv:2606.25473v12026
  34. Sequence to Sequence Learning with Neural Networks

    Ilya Sutskever, Oriol Vinyals, Quoc V. Le

    cs.CLcs.LGarXiv:1409.3215v32014
  35. Sparse Orthogonal Regression Technique: A Spectral Framework for Equation Discovery, Approximation, and Integration

    Sabin Roman, Ljupco Todorovski, Saso Dzeroski

    cs.LGarXiv:2608.13504v12026
  36. Playing Atari with Deep Reinforcement Learning

    Volodymyr Mnih, Koray Kavukcuoglu, David Silver +4

    cs.LGarXiv:1312.5602v12013
  37. Vidu S1: A Real-Time Interactive Video Generation Model

    Jintao Zhang, Kai Jiang, Jintao Chen +24

    cs.CVcs.LGarXiv:2607.03118v22026
  38. KVAE: Family of Tokenizers for Multimodal Generative Models

    Andrey Shutkin, Denis Parkhomenko, Ivan Kirillov +11

    cs.CVcs.LGcs.SDarXiv:2608.05798v12026
  39. RoFormer: Enhanced Transformer with Rotary Position Embedding

    Jianlin Su, Yu Lu, Shengfeng Pan +3

    cs.CLcs.AIcs.LGarXiv:2104.09864v52021
  40. Demystifying On-Policy Distillation: Roles, Pathologies, and Regulations

    Rui Wang, Hongru Wang, Yi Chen +4

    cs.CLcs.LGarXiv:2607.13399v12026
  41. On the global feature importance for interpretable and trustworthy heat demand forecasting

    Milan Zdravković

    cs.LGeess.SYarXiv:2608.13039v12026
  42. Forecast Collapse in Time-Series Foundation Models

    Shu Wan, Miles Ma, Hank Zhu +4

    cs.LGcs.AIcs.CEarXiv:2608.14106v12026
  43. DanceOPD: On-Policy Generative Field Distillation

    Wei Zhou, Xiongwei Zhu, Zelin Xu +8

    cs.CVcs.CLcs.LGarXiv:2606.27377v22026
  44. MobileMem: Learning from a Year of Mobile Experiences

    Xinle Deng, Yida Xue, Xiangyuan Ru +14

    cs.AIcs.CLcs.LGarXiv:2608.13606v12026
  45. On-Policy Self-Distillation without Any Supervision

    Yijiang Li, Bingyang Wang, Yijun Liang +3

    cs.LGarXiv:2608.06296v22026
  46. Distilling the Knowledge in a Neural Network

    Geoffrey Hinton, Oriol Vinyals, Jeff Dean

    stat.MLcs.LGcs.NEarXiv:1503.02531v12015
  47. SkillOpt-Lite: Better and Faster Agent Self-evolution via One Line of Vibe

    Yifei Shen, Bo Li, Xinjie Zhang

    cs.SEcs.AIcs.LGarXiv:2607.03451v12026
  48. Generative Adversarial Networks

    Ian J. Goodfellow, Jean Pouget-Abadie, Mehdi Mirza +5

    stat.MLcs.LGarXiv:1406.2661v12014
  49. MOPD: Multi-Teacher On-Policy Distillation for Capability Integration in LLM Post-Training

    Wenhan Ma, Jianyu Wei, Liang Zhao +10

    cs.CLcs.LGarXiv:2606.30406v12026
  50. Proximal Policy Optimization Algorithms

    John Schulman, Filip Wolski, Prafulla Dhariwal +2

    cs.LGarXiv:1707.06347v22017
  51. Intern-S2-Preview: Scientific Agentic Foundation Model

    Lei Bai, Jiaqi Cao, Chiyu Chen +122

    cs.LGcs.CLcs.CVarXiv:2608.13505v12026
  52. AgenticDataBench: A Comprehensive Benchmark for Data Agents

    Zhaoyan Sun, Shan Zhong, Daizhou Wen +10

    cs.DBcs.AIcs.CLarXiv:2607.01647v12026
  53. The State-Prediction Separation Hypothesis

    Giovanni Monea, Nathan Godey, Kianté Brantley +1

    cs.CLcs.AIcs.LGarXiv:2607.01218v12026
  54. WARP: Weight-Space Analysis for Recovering Training Data Portfolios

    Tzu-Heng Huang, Aditya Goyal, John Cooper +1

    cs.LGarXiv:2607.01686v12026