Machine Learning
Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
6,601 to 6,660 of 20,219
Learning Dexterous Manipulation for a Soft Robotic Hand from Human Demonstration
Abhishek Gupta, Clemens Eppner, Sergey Levine +1
cs.LGcs.ROarXiv:1603.06348v32016Multi-Scale Adaptive Graph Neural Network for Multivariate Time Series Forecasting
Ling Chen, Donghui Chen, Zongjiang Shang +4
cs.LGarXiv:2201.04828v22022Group Adaptive Clipping Policy Optimization
Sheng Jia, Xiao Wang, Shiva Prasad Kasiviswanathan +1
cs.LGcs.CLarXiv:2609.00444v12026A Minimalist Approach to LLM Reasoning: from Rejection Sampling to Reinforce
Wei Xiong, Jiarui Yao, Yuhui Xu +8
cs.LGcs.AIcs.CLarXiv:2504.11343v22025dLLM: Simple Diffusion Language Modeling
Zhanhui Zhou, Lingjie Chen, Hanghang Tong +1
cs.CLcs.AIcs.LGarXiv:2602.22661v12026DR Tulu: Reinforcement Learning with Evolving Rubrics for Deep Research
Rulin Shao, Akari Asai, Shannon Zejiang Shen +18
cs.CLcs.AIcs.LGarXiv:2511.19399v32025Towards Understanding Regularization in Batch Normalization
Ping Luo, Xinjiang Wang, Wenqi Shao +1
cs.LGcs.CVeess.SYarXiv:1809.00846v42018Learning to Win by Reading Manuals in a Monte-Carlo Framework
S. R. K. Branavan, David Silver, Regina Barzilay
cs.CLcs.AIcs.LGarXiv:1401.5390v12014Understanding graph embedding methods and their applications
Mengjia Xu
cs.LGcs.ITcs.SIarXiv:2012.08019v12020Large Models for Time Series and Spatio-Temporal Data: A Survey and Outlook
Ming Jin, Yaxuan Kong, Yuxuan Liang +13
cs.LGcs.AIarXiv:2310.10196v32023Graph Neural Networks: Taxonomy, Advances and Trends
Yu Zhou, Haixia Zheng, Xin Huang +3
cs.LGarXiv:2012.08752v42020Exploration in Deep Reinforcement Learning: From Single-Agent to Multiagent Domain
Jianye Hao, Tianpei Yang, Hongyao Tang +5
cs.AIcs.LGcs.MAarXiv:2109.06668v62021Bandwidth-Agile Image Transmission with Deep Joint Source-Channel Coding
David Burth Kurka, Deniz Gündüz
cs.ITcs.LGeess.IVarXiv:2009.12480v22020Day-Ahead Hourly Forecasting of Power Generation from Photovoltaic Plants
Lorenzo Gigoni, Alessandro Betti, Emanuele Crisostomi +4
cs.LGstat.MLarXiv:1903.06800v12019Diffusion LLMs Can Do Faster-Than-AR Inference via Discrete Diffusion Forcing
Xu Wang, Chenkai Xu, Yijie Jin +3
cs.LGcs.AIarXiv:2508.09192v12025From LLM Reasoning to Autonomous AI Agents: A Comprehensive Review
Mohamed Amine Ferrag, Norbert Tihanyi, Merouane Debbah
cs.AIcs.LGarXiv:2504.19678v22025Rethinking Classifier-Free Guidance in On-Policy Diffusion Distillation
Bingnan Li, Haozhe Wang, Haozhong Xiong +5
cs.CVcs.AIcs.LGarXiv:2607.24731v22026Committee neural network potentials control generalization errors and enable active learning
Christoph Schran, Krystof Brezina, Ondrej Marsalek
physics.chem-phcs.LGphysics.comp-pharXiv:2006.01541v22020Introduction to Online Convex Optimization
Elad Hazan
cs.LGmath.OCstat.MLarXiv:1909.05207v32019The Next Decade in AI: Four Steps Towards Robust Artificial Intelligence
Gary Marcus
cs.AIcs.LGarXiv:2002.06177v32020Are Language Models Actually Useful for Time Series Forecasting?
Mingtian Tan, Mike A. Merrill, Vinayak Gupta +2
cs.LGcs.AIarXiv:2406.16964v22024Flow-DPPO: Divergence Proximal Policy Optimization for Flow Matching Models
Bowen Ping, Xiangxin Zhou, Penghui Qi +3
cs.LGarXiv:2606.11025v22026Performance Foundations of Parallel & Distributed Reasoning Language Models
Maciej Besta, Leonard Schmidt, Lara Nonino +7
cs.LGcs.AIcs.DCarXiv:2608.27046v12026Scaling Model-Generated Distillation Data Can Make Latent Teacher Traits More Recoverable
Zhichen Dong, Zhixuan Liu, Yuyu Fan +3
cs.LGcs.CLarXiv:2608.26958v12026Cross-Temperature Defect Identification in Atomistic Simulations via Multi-Level Domain Alignment
Yating Fang, Jungmin Kim, Qian Qian Zhao +4
cond-mat.mtrl-scics.LGphysics.comp-pharXiv:2608.22074v12026Which Negatives Matter? Ask Your Text Encoder: Adaptive Similarity Margins for Dense-Caption Retrieval
Haoyue Liu, Ye Chen, Zhichao Wang +1
cs.AIcs.LGarXiv:2608.18521v12026Evaluating and improving crop-yield forecasting methods during extreme drought
Shrey Gupta, Yi Ming, George Mohler
cs.LGarXiv:2608.17971v12026Reflex-Guard: A Low-Latency Guardrail for LLM Prompt Safety Using Dense Semantic Embeddings
Istiaque Ahmed, Afia Anjum Borsha, Ranat Das Prangon +2
cs.CRcs.CLcs.LGarXiv:2608.17556v12026A Method for Representing Periodic Functions and Enforcing Exactly Periodic Boundary Conditions with Deep Neural Networks
Suchuan Dong, Naxian Ni
physics.comp-phcs.LGmath.NAarXiv:2007.07442v12020H3DNAS: Hardware-Aware ONNX-Native 3D Point Cloud Model Compression
Anchit Mulye, Rhythm Baghel, Sujay Kumar Ingle +1
cs.LGcs.ARcs.NEarXiv:2609.02684v12026The Hitchhiker's Guide to Agentic AI: From Foundations to Systems
Haggai Roitman
cs.AIcs.CLcs.IRarXiv:2606.24937v22026Summaries:한국어Sparse Competition during Training For the Emergence of Specialized Modules
Baptiste Rossigneux, Karim Haroun
cs.LGarXiv:2608.30978v12026Informative Label Missingness in Multiclass Classification Information Geometry and Excess Risk
Fariborz Setoudehtazang, Geoffrey J. McLachlan
stat.MLcs.LGarXiv:2608.30561v12026Agentic Large Language Models, a survey
Aske Plaat, Max van Duijn, Niki van Stein +3
cs.AIcs.CLcs.LGarXiv:2503.23037v32025Twin Worlds: Equivariance-Based Abstention for Evidence-Grounded Reasoning
Vy Nguyen, Ziqi Xu, Jeffrey Chan +5
cs.CLcs.AIcs.LGarXiv:2608.28018v12026TEMPLAR Wales: A georeferenced environmental and toponymic dataset of Welsh settlements
Oktay Karakuş, Can Eyupoglu
cs.LGcs.CLarXiv:2608.26970v12026Multiscale Community-Based Fingerprinting of Signed Functional Networks
Sema Athamnah, Selin Aviyente
q-bio.NCcs.LGeess.SParXiv:2608.27483v12026Systematic Literature Review of Machine Learning Models and Applications for Text Recognition
Nuzhat Khan, Ab Al-Hadi Ab Rahman, Shahriyar Masud Rizvi +5
cs.CVcs.LGarXiv:2608.26500v12026Geometry-Constrained Kolmogorov-Arnold Networks: Learning Edge Geometry via Banach Duality
K S Sesh Kumar
cs.LGstat.MLarXiv:2608.25807v12026On Scope Classification and Current Knowledge-Editing Benchmarks: A Negative Result, with INLAY as a Gradient-Free Case Study
Aditya Pratap Singh
cs.CLcs.AIcs.LGarXiv:2608.26292v12026VINCENT: Validated Interaction Network for Cross-drug Explanation of Therapeutics
Fan-Sheng Chuang, Xuchen Li, Yujing Bian +1
cs.LGcs.AIarXiv:2608.25841v12026CropCop: An Auditable 120-Class Plant-Health Model from Benchmark Reconstruction to a Quantised Runtime Artifact
Rana Muhammad Ahmed, Sabahat Abbas
cs.CVcs.LGarXiv:2608.25539v12026Individual Fairness in Hierarchical Clustering
Binita Maity, Shrutimoy Das
cs.LGarXiv:2608.25586v12026HBQ: Hierarchical Scaling Block Quantization with Hardware-Efficiency-Aware Design for Accurate LLM Inference
Chun-Ting Chen, Dongmin Han, Hangyeol Mun +6
cs.LGcs.AIcs.ARarXiv:2609.00450v12026Evolving Curricula with Regret-Based Environment Design
Jack Parker-Holder, Minqi Jiang, Michael Dennis +4
cs.LGarXiv:2203.01302v32022The 'Problem' of Human Label Variation: On Ground Truth in Data, Modeling and Evaluation
Barbara Plank
cs.CLcs.LGarXiv:2211.02570v12022Improving Few-Step Language Flows with Untied Self-Conditioning
Bocheng Li, Linli Xu
cs.CLcs.AIcs.LGarXiv:2608.22244v12026RpBERT: A Text-image Relation Propagation-based BERT Model for Multimodal NER
Lin Sun, Jiquan Wang, Kai Zhang +2
cs.CLcs.LGarXiv:2102.02967v12021Structured Learning on Mapper Representations
George Babus, Farzana Nasrin
stat.MLcs.LGarXiv:2608.22044v12026Decoupling Policy Extraction for Offline Reinforcement Learning
Xuyao Lin, Yixiang Shan, Jinru Duan +7
cs.LGcs.ROarXiv:2608.20909v12026FastKernels: Benchmarking GPU Kernel Generation in Production
Gabriele Oliaro, Yichao Fu, May Jiang +5
cs.LGcs.AIcs.CLarXiv:2605.23215v12026From Storage to Access: Verifiable Activation of Parametric Knowledge in LLMs via Explicit Priming and Implicit Reasoning
Zuocheng Ying, Yang Yang, Yumou Wu +6
cs.CLcs.AIcs.LGarXiv:2608.18581v12026Compress and Forget: bitsandbytes Quantization Amplifies Proactive Interference in LLMs
Shayan Shahrabi-Farahani, Dara Rahmati
cs.CLcs.LGarXiv:2608.18578v12026Study-Strategy Clusters from EdNet Logs Track Engagement, Not Mastery
Qingchuan Lyu, Yingxin Li, Albert Yang
cs.LGcs.CYstat.AParXiv:2608.16963v12026Paired Exact-Reset Evaluation of a Prediction-Derived Medium-to-Full World-Model Cascade
Malo de Pastor
cs.LGcs.ROarXiv:2608.14650v12026Scaling Long-Horizon LLM Agent via Context-Folding
Weiwei Sun, Miao Lu, Zhan Ling +4
cs.CLcs.LGarXiv:2510.11967v12025TriPlay-RL: Tri-Role Self-Play Reinforcement Learning for LLM Safety Alignment
Zhewen Tan, Wenhan Yu, Jianfeng Si +9
cs.LGcs.AIarXiv:2601.18292v22026KV Packet: Recomputation-Free Context-Independent KV Caching for LLMs
Chuangtao Chen, Grace Li Zhang, Xunzhao Yin +3
cs.LGcs.AIarXiv:2604.13226v22026LightMem: Lightweight and Efficient Memory-Augmented Generation
Jizhan Fang, Xinle Deng, Haoming Xu +9
cs.CLcs.AIcs.CVarXiv:2510.18866v42025S0 Tuning: Zero-Overhead Adaptation of Hybrid Recurrent-Attention Models
Jack Young
cs.CLcs.LGarXiv:2604.01168v22026