Machine Learning

Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

6,601 to 6,660 of 20,219

  1. Learning Dexterous Manipulation for a Soft Robotic Hand from Human Demonstration

    Abhishek Gupta, Clemens Eppner, Sergey Levine +1

    cs.LGcs.ROarXiv:1603.06348v32016
  2. Multi-Scale Adaptive Graph Neural Network for Multivariate Time Series Forecasting

    Ling Chen, Donghui Chen, Zongjiang Shang +4

    cs.LGarXiv:2201.04828v22022
  3. Group Adaptive Clipping Policy Optimization

    Sheng Jia, Xiao Wang, Shiva Prasad Kasiviswanathan +1

    cs.LGcs.CLarXiv:2609.00444v12026
  4. A Minimalist Approach to LLM Reasoning: from Rejection Sampling to Reinforce

    Wei Xiong, Jiarui Yao, Yuhui Xu +8

    cs.LGcs.AIcs.CLarXiv:2504.11343v22025
  5. dLLM: Simple Diffusion Language Modeling

    Zhanhui Zhou, Lingjie Chen, Hanghang Tong +1

    cs.CLcs.AIcs.LGarXiv:2602.22661v12026
  6. DR Tulu: Reinforcement Learning with Evolving Rubrics for Deep Research

    Rulin Shao, Akari Asai, Shannon Zejiang Shen +18

    cs.CLcs.AIcs.LGarXiv:2511.19399v32025
  7. Towards Understanding Regularization in Batch Normalization

    Ping Luo, Xinjiang Wang, Wenqi Shao +1

    cs.LGcs.CVeess.SYarXiv:1809.00846v42018
  8. Learning to Win by Reading Manuals in a Monte-Carlo Framework

    S. R. K. Branavan, David Silver, Regina Barzilay

    cs.CLcs.AIcs.LGarXiv:1401.5390v12014
  9. Understanding graph embedding methods and their applications

    Mengjia Xu

    cs.LGcs.ITcs.SIarXiv:2012.08019v12020
  10. Large Models for Time Series and Spatio-Temporal Data: A Survey and Outlook

    Ming Jin, Yaxuan Kong, Yuxuan Liang +13

    cs.LGcs.AIarXiv:2310.10196v32023
  11. Graph Neural Networks: Taxonomy, Advances and Trends

    Yu Zhou, Haixia Zheng, Xin Huang +3

    cs.LGarXiv:2012.08752v42020
  12. Exploration in Deep Reinforcement Learning: From Single-Agent to Multiagent Domain

    Jianye Hao, Tianpei Yang, Hongyao Tang +5

    cs.AIcs.LGcs.MAarXiv:2109.06668v62021
  13. Bandwidth-Agile Image Transmission with Deep Joint Source-Channel Coding

    David Burth Kurka, Deniz Gündüz

    cs.ITcs.LGeess.IVarXiv:2009.12480v22020
  14. Day-Ahead Hourly Forecasting of Power Generation from Photovoltaic Plants

    Lorenzo Gigoni, Alessandro Betti, Emanuele Crisostomi +4

    cs.LGstat.MLarXiv:1903.06800v12019
  15. Diffusion LLMs Can Do Faster-Than-AR Inference via Discrete Diffusion Forcing

    Xu Wang, Chenkai Xu, Yijie Jin +3

    cs.LGcs.AIarXiv:2508.09192v12025
  16. From LLM Reasoning to Autonomous AI Agents: A Comprehensive Review

    Mohamed Amine Ferrag, Norbert Tihanyi, Merouane Debbah

    cs.AIcs.LGarXiv:2504.19678v22025
  17. Rethinking Classifier-Free Guidance in On-Policy Diffusion Distillation

    Bingnan Li, Haozhe Wang, Haozhong Xiong +5

    cs.CVcs.AIcs.LGarXiv:2607.24731v22026
  18. Committee neural network potentials control generalization errors and enable active learning

    Christoph Schran, Krystof Brezina, Ondrej Marsalek

    physics.chem-phcs.LGphysics.comp-pharXiv:2006.01541v22020
  19. Introduction to Online Convex Optimization

    Elad Hazan

    cs.LGmath.OCstat.MLarXiv:1909.05207v32019
  20. The Next Decade in AI: Four Steps Towards Robust Artificial Intelligence

    Gary Marcus

    cs.AIcs.LGarXiv:2002.06177v32020
  21. Are Language Models Actually Useful for Time Series Forecasting?

    Mingtian Tan, Mike A. Merrill, Vinayak Gupta +2

    cs.LGcs.AIarXiv:2406.16964v22024
  22. Flow-DPPO: Divergence Proximal Policy Optimization for Flow Matching Models

    Bowen Ping, Xiangxin Zhou, Penghui Qi +3

    cs.LGarXiv:2606.11025v22026
  23. Performance Foundations of Parallel & Distributed Reasoning Language Models

    Maciej Besta, Leonard Schmidt, Lara Nonino +7

    cs.LGcs.AIcs.DCarXiv:2608.27046v12026
  24. Scaling Model-Generated Distillation Data Can Make Latent Teacher Traits More Recoverable

    Zhichen Dong, Zhixuan Liu, Yuyu Fan +3

    cs.LGcs.CLarXiv:2608.26958v12026
  25. Cross-Temperature Defect Identification in Atomistic Simulations via Multi-Level Domain Alignment

    Yating Fang, Jungmin Kim, Qian Qian Zhao +4

    cond-mat.mtrl-scics.LGphysics.comp-pharXiv:2608.22074v12026
  26. Which Negatives Matter? Ask Your Text Encoder: Adaptive Similarity Margins for Dense-Caption Retrieval

    Haoyue Liu, Ye Chen, Zhichao Wang +1

    cs.AIcs.LGarXiv:2608.18521v12026
  27. Evaluating and improving crop-yield forecasting methods during extreme drought

    Shrey Gupta, Yi Ming, George Mohler

    cs.LGarXiv:2608.17971v12026
  28. Reflex-Guard: A Low-Latency Guardrail for LLM Prompt Safety Using Dense Semantic Embeddings

    Istiaque Ahmed, Afia Anjum Borsha, Ranat Das Prangon +2

    cs.CRcs.CLcs.LGarXiv:2608.17556v12026
  29. A Method for Representing Periodic Functions and Enforcing Exactly Periodic Boundary Conditions with Deep Neural Networks

    Suchuan Dong, Naxian Ni

    physics.comp-phcs.LGmath.NAarXiv:2007.07442v12020
  30. H3DNAS: Hardware-Aware ONNX-Native 3D Point Cloud Model Compression

    Anchit Mulye, Rhythm Baghel, Sujay Kumar Ingle +1

    cs.LGcs.ARcs.NEarXiv:2609.02684v12026
  31. The Hitchhiker's Guide to Agentic AI: From Foundations to Systems

    Haggai Roitman

    cs.AIcs.CLcs.IRarXiv:2606.24937v22026
    Summaries:한국어
  32. Sparse Competition during Training For the Emergence of Specialized Modules

    Baptiste Rossigneux, Karim Haroun

    cs.LGarXiv:2608.30978v12026
  33. Informative Label Missingness in Multiclass Classification Information Geometry and Excess Risk

    Fariborz Setoudehtazang, Geoffrey J. McLachlan

    stat.MLcs.LGarXiv:2608.30561v12026
  34. Agentic Large Language Models, a survey

    Aske Plaat, Max van Duijn, Niki van Stein +3

    cs.AIcs.CLcs.LGarXiv:2503.23037v32025
  35. Twin Worlds: Equivariance-Based Abstention for Evidence-Grounded Reasoning

    Vy Nguyen, Ziqi Xu, Jeffrey Chan +5

    cs.CLcs.AIcs.LGarXiv:2608.28018v12026
  36. TEMPLAR Wales: A georeferenced environmental and toponymic dataset of Welsh settlements

    Oktay Karakuş, Can Eyupoglu

    cs.LGcs.CLarXiv:2608.26970v12026
  37. Multiscale Community-Based Fingerprinting of Signed Functional Networks

    Sema Athamnah, Selin Aviyente

    q-bio.NCcs.LGeess.SParXiv:2608.27483v12026
  38. Systematic Literature Review of Machine Learning Models and Applications for Text Recognition

    Nuzhat Khan, Ab Al-Hadi Ab Rahman, Shahriyar Masud Rizvi +5

    cs.CVcs.LGarXiv:2608.26500v12026
  39. Geometry-Constrained Kolmogorov-Arnold Networks: Learning Edge Geometry via Banach Duality

    K S Sesh Kumar

    cs.LGstat.MLarXiv:2608.25807v12026
  40. On Scope Classification and Current Knowledge-Editing Benchmarks: A Negative Result, with INLAY as a Gradient-Free Case Study

    Aditya Pratap Singh

    cs.CLcs.AIcs.LGarXiv:2608.26292v12026
  41. VINCENT: Validated Interaction Network for Cross-drug Explanation of Therapeutics

    Fan-Sheng Chuang, Xuchen Li, Yujing Bian +1

    cs.LGcs.AIarXiv:2608.25841v12026
  42. CropCop: An Auditable 120-Class Plant-Health Model from Benchmark Reconstruction to a Quantised Runtime Artifact

    Rana Muhammad Ahmed, Sabahat Abbas

    cs.CVcs.LGarXiv:2608.25539v12026
  43. Individual Fairness in Hierarchical Clustering

    Binita Maity, Shrutimoy Das

    cs.LGarXiv:2608.25586v12026
  44. HBQ: Hierarchical Scaling Block Quantization with Hardware-Efficiency-Aware Design for Accurate LLM Inference

    Chun-Ting Chen, Dongmin Han, Hangyeol Mun +6

    cs.LGcs.AIcs.ARarXiv:2609.00450v12026
  45. Evolving Curricula with Regret-Based Environment Design

    Jack Parker-Holder, Minqi Jiang, Michael Dennis +4

    cs.LGarXiv:2203.01302v32022
  46. The 'Problem' of Human Label Variation: On Ground Truth in Data, Modeling and Evaluation

    Barbara Plank

    cs.CLcs.LGarXiv:2211.02570v12022
  47. Improving Few-Step Language Flows with Untied Self-Conditioning

    Bocheng Li, Linli Xu

    cs.CLcs.AIcs.LGarXiv:2608.22244v12026
  48. RpBERT: A Text-image Relation Propagation-based BERT Model for Multimodal NER

    Lin Sun, Jiquan Wang, Kai Zhang +2

    cs.CLcs.LGarXiv:2102.02967v12021
  49. Structured Learning on Mapper Representations

    George Babus, Farzana Nasrin

    stat.MLcs.LGarXiv:2608.22044v12026
  50. Decoupling Policy Extraction for Offline Reinforcement Learning

    Xuyao Lin, Yixiang Shan, Jinru Duan +7

    cs.LGcs.ROarXiv:2608.20909v12026
  51. FastKernels: Benchmarking GPU Kernel Generation in Production

    Gabriele Oliaro, Yichao Fu, May Jiang +5

    cs.LGcs.AIcs.CLarXiv:2605.23215v12026
  52. From Storage to Access: Verifiable Activation of Parametric Knowledge in LLMs via Explicit Priming and Implicit Reasoning

    Zuocheng Ying, Yang Yang, Yumou Wu +6

    cs.CLcs.AIcs.LGarXiv:2608.18581v12026
  53. Compress and Forget: bitsandbytes Quantization Amplifies Proactive Interference in LLMs

    Shayan Shahrabi-Farahani, Dara Rahmati

    cs.CLcs.LGarXiv:2608.18578v12026
  54. Study-Strategy Clusters from EdNet Logs Track Engagement, Not Mastery

    Qingchuan Lyu, Yingxin Li, Albert Yang

    cs.LGcs.CYstat.AParXiv:2608.16963v12026
  55. Paired Exact-Reset Evaluation of a Prediction-Derived Medium-to-Full World-Model Cascade

    Malo de Pastor

    cs.LGcs.ROarXiv:2608.14650v12026
  56. Scaling Long-Horizon LLM Agent via Context-Folding

    Weiwei Sun, Miao Lu, Zhan Ling +4

    cs.CLcs.LGarXiv:2510.11967v12025
  57. TriPlay-RL: Tri-Role Self-Play Reinforcement Learning for LLM Safety Alignment

    Zhewen Tan, Wenhan Yu, Jianfeng Si +9

    cs.LGcs.AIarXiv:2601.18292v22026
  58. KV Packet: Recomputation-Free Context-Independent KV Caching for LLMs

    Chuangtao Chen, Grace Li Zhang, Xunzhao Yin +3

    cs.LGcs.AIarXiv:2604.13226v22026
  59. LightMem: Lightweight and Efficient Memory-Augmented Generation

    Jizhan Fang, Xinle Deng, Haoming Xu +9

    cs.CLcs.AIcs.CVarXiv:2510.18866v42025
  60. S0 Tuning: Zero-Overhead Adaptation of Hybrid Recurrent-Attention Models

    Jack Young

    cs.CLcs.LGarXiv:2604.01168v22026