Machine Learning
Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
6,541 to 6,600 of 20,205
Domain and Function: A Dual-Space Model of Semantic Relations and Compositions
Peter D. Turney
cs.CLcs.AIcs.LGarXiv:1309.4035v12013Higher Structures in Deep Learning
Michael L. Roberts, Carlos Zapata Carratalá. Nicholas J. Cooper, Lijun Chen +2
cs.LGcs.AIarXiv:2609.00472v12026TimeMixer++: A General Time Series Pattern Machine for Universal Predictive Analysis
Shiyu Wang, Jiawei Li, Xiaoming Shi +6
cs.LGcs.AIarXiv:2410.16032v52024Accelerating Chemical Kinetics for Exoplanet Atmospheres using Neural Networks
Isaac Malsky, Xi Zhang, Tiffany Kataria +5
astro-ph.EPcs.LGarXiv:2609.00428v12026Flow Matching Policy Gradients
David McAllister, Songwei Ge, Brent Yi +5
cs.LGcs.ROarXiv:2507.21053v22025Scaling Reasoning Efficiently via Relaxed On-Policy Distillation
Jongwoo Ko, Sara Abdali, Young Jin Kim +2
cs.LGcs.CLarXiv:2603.11137v12026Safety Tax: Safety Alignment Makes Your Large Reasoning Models Less Reasonable
Tiansheng Huang, Sihao Hu, Fatih Ilhan +4
cs.CRcs.AIcs.LGarXiv:2503.00555v22025MDP Homomorphic Networks: Group Symmetries in Reinforcement Learning
Elise van der Pol, Daniel E. Worrall, Herke van Hoof +2
cs.LGstat.MLarXiv:2006.16908v22020BREEDS: Benchmarks for Subpopulation Shift
Shibani Santurkar, Dimitris Tsipras, Aleksander Madry
cs.CVcs.LGstat.MLarXiv:2008.04859v12020Gated Transformer Networks for Multivariate Time Series Classification
Minghao Liu, Shengqi Ren, Siyuan Ma +4
cs.LGarXiv:2103.14438v12021PAC Bounds for Discounted MDPs
Tor Lattimore, Marcus Hutter
cs.LGarXiv:1202.3890v12012Diffusion-Based Voice Conversion with Fast Maximum Likelihood Sampling Scheme
Vadim Popov, Ivan Vovk, Vladimir Gogoryan +3
cs.SDcs.LGstat.MLarXiv:2109.13821v22021Learning Robust Visual-Semantic Embeddings
Yao-Hung Hubert Tsai, Liang-Kang Huang, Ruslan Salakhutdinov
cs.CVcs.CLcs.LGarXiv:1703.05908v22017Implicit Regularization of Discrete Gradient Dynamics in Linear Neural Networks
Gauthier Gidel, Francis Bach, Simon Lacoste-Julien
cs.LGmath.OCstat.MLarXiv:1904.13262v22019Colorization Transformer
Manoj Kumar, Dirk Weissenborn, Nal Kalchbrenner
cs.CVcs.AIcs.LGarXiv:2102.04432v22021Z-Forcing: Training Stochastic Recurrent Networks
Anirudh Goyal, Alessandro Sordoni, Marc-Alexandre Côté +2
stat.MLcs.LGarXiv:1711.05411v22017Open Problems in Mechanistic Interpretability
Lee Sharkey, Bilal Chughtai, Joshua Batson +26
cs.LGarXiv:2501.16496v12025Self-Supervised Policy Adaptation during Deployment
Nicklas Hansen, Rishabh Jangir, Yu Sun +5
cs.LGcs.CVcs.ROarXiv:2007.04309v32020Analyzing and Improving Representations with the Soft Nearest Neighbor Loss
Nicholas Frosst, Nicolas Papernot, Geoffrey Hinton
stat.MLcs.LGarXiv:1902.01889v12019Quantifying social organization and political polarization in online platforms
Isaac Waller, Ashton Anderson
cs.SIcs.AIcs.CYarXiv:2010.00590v32020Neural Collapse Under MSE Loss: Proximity to and Dynamics on the Central Path
X. Y. Han, Vardan Papyan, David L. Donoho
cs.LGcs.AImath.DGarXiv:2106.02073v42021VoiceCraft: Zero-Shot Speech Editing and Text-to-Speech in the Wild
Puyuan Peng, Po-Yao Huang, Shang-Wen Li +2
eess.AScs.AIcs.CLarXiv:2403.16973v32024CycleNet: Enhancing Time Series Forecasting through Modeling Periodic Patterns
Shengsheng Lin, Weiwei Lin, Xinyi Hu +3
cs.LGarXiv:2409.18479v22024Robometer: Scaling General-Purpose Robotic Reward Models via Trajectory Comparisons
Anthony Liang, Yigit Korkmaz, Jiahui Zhang +14
cs.ROcs.AIcs.LGarXiv:2603.02115v22026Conditional Affordance Learning for Driving in Urban Environments
Axel Sauer, Nikolay Savinov, Andreas Geiger
cs.ROcs.LGeess.SYarXiv:1806.06498v32018Nemotron 3 Nano: Open, Efficient Mixture-of-Experts Hybrid Mamba-Transformer Model for Agentic Reasoning
NVIDIA, :, Aaron Blakeman +311
cs.CLcs.AIcs.LGarXiv:2512.20848v12025WebAgent-R1: Training Web Agents via End-to-End Multi-Turn Reinforcement Learning
Zhepei Wei, Wenlin Yao, Yao Liu +9
cs.CLcs.LGarXiv:2505.16421v22025DeepCorr: Strong Flow Correlation Attacks on Tor Using Deep Learning
Milad Nasr, Alireza Bahramali, Amir Houmansadr
cs.CRcs.LGarXiv:1808.07285v12018TiRex: Zero-Shot Forecasting Across Long and Short Horizons with Enhanced In-Context Learning
Andreas Auer, Patrick Podest, Daniel Klotz +3
cs.LGarXiv:2505.23719v22025SparseTSF: Modeling Long-term Time Series Forecasting with 1k Parameters
Shengsheng Lin, Weiwei Lin, Wentai Wu +2
cs.LGarXiv:2405.00946v22024Sponge Examples: Energy-Latency Attacks on Neural Networks
Ilia Shumailov, Yiren Zhao, Daniel Bates +3
cs.LGcs.CLcs.CRarXiv:2006.03463v22020Training Agents Inside of Scalable World Models
Danijar Hafner, Wilson Yan, Timothy Lillicrap
cs.AIcs.LGcs.ROarXiv:2509.24527v12025The Polar Express: Optimal Matrix Sign Methods and Their Application to the Muon Algorithm
Noah Amsel, David Persson, Christopher Musco +1
cs.LGcs.AIcs.CLarXiv:2505.16932v52025SAEBench: A Comprehensive Benchmark for Sparse Autoencoders in Language Model Interpretability
Adam Karvonen, Can Rager, Johnny Lin +12
cs.LGcs.CLarXiv:2503.09532v42025Understanding plasticity in neural networks
Clare Lyle, Zeyu Zheng, Evgenii Nikishin +3
cs.LGarXiv:2303.01486v42023Programming Refusal with Conditional Activation Steering
Bruce W. Lee, Inkit Padhi, Karthikeyan Natesan Ramamurthy +4
cs.LGcs.AIcs.CLarXiv:2409.05907v32024Controllable Image Captioning with Prompt-Conditioned Scene Rewards
Jongyeop Hyun, Taeyoung Kim, Hyounghun Kim
cs.CVcs.CLcs.LGarXiv:2609.00709v12026OPAL: Offline Primitive Discovery for Accelerating Offline Reinforcement Learning
Anurag Ajay, Aviral Kumar, Pulkit Agrawal +2
cs.LGarXiv:2010.13611v32020Vision-OPD: Learning to See Fine Details for Multimodal LLMs via On-Policy Self-Distillation
Qianhao Yuan, Jie Lou, Xing Yu +4
cs.CVcs.AIcs.CLarXiv:2605.18740v42026PixMix: Dreamlike Pictures Comprehensively Improve Safety Measures
Dan Hendrycks, Andy Zou, Mantas Mazeika +4
cs.LGcs.CVarXiv:2112.05135v32021GOOD: A Graph Out-of-Distribution Benchmark
Shurui Gui, Xiner Li, Limei Wang +1
cs.LGcs.AIarXiv:2206.08452v22022REVE: A Foundation Model for EEG -- Adapting to Any Setup with Large-Scale Pretraining on 25,000 Subjects
Yassine El Ouahidi, Jonathan Lys, Philipp Thölke +5
cs.LGq-bio.NCarXiv:2510.21585v12025Deep Exemplar-based Video Colorization
Bo Zhang, Mingming He, Jing Liao +4
cs.CVcs.AIcs.LGarXiv:1906.09909v12019Act Only When It Pays: Efficient Reinforcement Learning for LLM Reasoning via Selective Rollouts
Haizhong Zheng, Yang Zhou, Brian R. Bartoldson +4
cs.AIcs.LGarXiv:2506.02177v12025FILM: Following Instructions in Language with Modular Methods
So Yeon Min, Devendra Singh Chaplot, Pradeep Ravikumar +2
cs.CLcs.LGarXiv:2110.07342v32021TurboQuant: Online Vector Quantization with Near-optimal Distortion Rate
Amir Zandieh, Majid Daliri, Majid Hadian +1
cs.LGcs.AIcs.DBarXiv:2504.19874v12025Learning Dexterous Manipulation for a Soft Robotic Hand from Human Demonstration
Abhishek Gupta, Clemens Eppner, Sergey Levine +1
cs.LGcs.ROarXiv:1603.06348v32016Multi-Scale Adaptive Graph Neural Network for Multivariate Time Series Forecasting
Ling Chen, Donghui Chen, Zongjiang Shang +4
cs.LGarXiv:2201.04828v22022Group Adaptive Clipping Policy Optimization
Sheng Jia, Xiao Wang, Shiva Prasad Kasiviswanathan +1
cs.LGcs.CLarXiv:2609.00444v12026A Minimalist Approach to LLM Reasoning: from Rejection Sampling to Reinforce
Wei Xiong, Jiarui Yao, Yuhui Xu +8
cs.LGcs.AIcs.CLarXiv:2504.11343v22025dLLM: Simple Diffusion Language Modeling
Zhanhui Zhou, Lingjie Chen, Hanghang Tong +1
cs.CLcs.AIcs.LGarXiv:2602.22661v12026DR Tulu: Reinforcement Learning with Evolving Rubrics for Deep Research
Rulin Shao, Akari Asai, Shannon Zejiang Shen +18
cs.CLcs.AIcs.LGarXiv:2511.19399v32025Towards Understanding Regularization in Batch Normalization
Ping Luo, Xinjiang Wang, Wenqi Shao +1
cs.LGcs.CVeess.SYarXiv:1809.00846v42018Learning to Win by Reading Manuals in a Monte-Carlo Framework
S. R. K. Branavan, David Silver, Regina Barzilay
cs.CLcs.AIcs.LGarXiv:1401.5390v12014Understanding graph embedding methods and their applications
Mengjia Xu
cs.LGcs.ITcs.SIarXiv:2012.08019v12020Large Models for Time Series and Spatio-Temporal Data: A Survey and Outlook
Ming Jin, Yaxuan Kong, Yuxuan Liang +13
cs.LGcs.AIarXiv:2310.10196v32023Graph Neural Networks: Taxonomy, Advances and Trends
Yu Zhou, Haixia Zheng, Xin Huang +3
cs.LGarXiv:2012.08752v42020Exploration in Deep Reinforcement Learning: From Single-Agent to Multiagent Domain
Jianye Hao, Tianpei Yang, Hongyao Tang +5
cs.AIcs.LGcs.MAarXiv:2109.06668v62021Bandwidth-Agile Image Transmission with Deep Joint Source-Channel Coding
David Burth Kurka, Deniz Gündüz
cs.ITcs.LGeess.IVarXiv:2009.12480v22020Day-Ahead Hourly Forecasting of Power Generation from Photovoltaic Plants
Lorenzo Gigoni, Alessandro Betti, Emanuele Crisostomi +4
cs.LGstat.MLarXiv:1903.06800v12019