Machine Learning

Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

6,541 to 6,600 of 20,205

  1. Domain and Function: A Dual-Space Model of Semantic Relations and Compositions

    Peter D. Turney

    cs.CLcs.AIcs.LGarXiv:1309.4035v12013
  2. Higher Structures in Deep Learning

    Michael L. Roberts, Carlos Zapata Carratalá. Nicholas J. Cooper, Lijun Chen +2

    cs.LGcs.AIarXiv:2609.00472v12026
  3. TimeMixer++: A General Time Series Pattern Machine for Universal Predictive Analysis

    Shiyu Wang, Jiawei Li, Xiaoming Shi +6

    cs.LGcs.AIarXiv:2410.16032v52024
  4. Accelerating Chemical Kinetics for Exoplanet Atmospheres using Neural Networks

    Isaac Malsky, Xi Zhang, Tiffany Kataria +5

    astro-ph.EPcs.LGarXiv:2609.00428v12026
  5. Flow Matching Policy Gradients

    David McAllister, Songwei Ge, Brent Yi +5

    cs.LGcs.ROarXiv:2507.21053v22025
  6. Scaling Reasoning Efficiently via Relaxed On-Policy Distillation

    Jongwoo Ko, Sara Abdali, Young Jin Kim +2

    cs.LGcs.CLarXiv:2603.11137v12026
  7. Safety Tax: Safety Alignment Makes Your Large Reasoning Models Less Reasonable

    Tiansheng Huang, Sihao Hu, Fatih Ilhan +4

    cs.CRcs.AIcs.LGarXiv:2503.00555v22025
  8. MDP Homomorphic Networks: Group Symmetries in Reinforcement Learning

    Elise van der Pol, Daniel E. Worrall, Herke van Hoof +2

    cs.LGstat.MLarXiv:2006.16908v22020
  9. BREEDS: Benchmarks for Subpopulation Shift

    Shibani Santurkar, Dimitris Tsipras, Aleksander Madry

    cs.CVcs.LGstat.MLarXiv:2008.04859v12020
  10. Gated Transformer Networks for Multivariate Time Series Classification

    Minghao Liu, Shengqi Ren, Siyuan Ma +4

    cs.LGarXiv:2103.14438v12021
  11. PAC Bounds for Discounted MDPs

    Tor Lattimore, Marcus Hutter

    cs.LGarXiv:1202.3890v12012
  12. Diffusion-Based Voice Conversion with Fast Maximum Likelihood Sampling Scheme

    Vadim Popov, Ivan Vovk, Vladimir Gogoryan +3

    cs.SDcs.LGstat.MLarXiv:2109.13821v22021
  13. Learning Robust Visual-Semantic Embeddings

    Yao-Hung Hubert Tsai, Liang-Kang Huang, Ruslan Salakhutdinov

    cs.CVcs.CLcs.LGarXiv:1703.05908v22017
  14. Implicit Regularization of Discrete Gradient Dynamics in Linear Neural Networks

    Gauthier Gidel, Francis Bach, Simon Lacoste-Julien

    cs.LGmath.OCstat.MLarXiv:1904.13262v22019
  15. Colorization Transformer

    Manoj Kumar, Dirk Weissenborn, Nal Kalchbrenner

    cs.CVcs.AIcs.LGarXiv:2102.04432v22021
  16. Z-Forcing: Training Stochastic Recurrent Networks

    Anirudh Goyal, Alessandro Sordoni, Marc-Alexandre Côté +2

    stat.MLcs.LGarXiv:1711.05411v22017
  17. Open Problems in Mechanistic Interpretability

    Lee Sharkey, Bilal Chughtai, Joshua Batson +26

    cs.LGarXiv:2501.16496v12025
  18. Self-Supervised Policy Adaptation during Deployment

    Nicklas Hansen, Rishabh Jangir, Yu Sun +5

    cs.LGcs.CVcs.ROarXiv:2007.04309v32020
  19. Analyzing and Improving Representations with the Soft Nearest Neighbor Loss

    Nicholas Frosst, Nicolas Papernot, Geoffrey Hinton

    stat.MLcs.LGarXiv:1902.01889v12019
  20. Quantifying social organization and political polarization in online platforms

    Isaac Waller, Ashton Anderson

    cs.SIcs.AIcs.CYarXiv:2010.00590v32020
  21. Neural Collapse Under MSE Loss: Proximity to and Dynamics on the Central Path

    X. Y. Han, Vardan Papyan, David L. Donoho

    cs.LGcs.AImath.DGarXiv:2106.02073v42021
  22. VoiceCraft: Zero-Shot Speech Editing and Text-to-Speech in the Wild

    Puyuan Peng, Po-Yao Huang, Shang-Wen Li +2

    eess.AScs.AIcs.CLarXiv:2403.16973v32024
  23. CycleNet: Enhancing Time Series Forecasting through Modeling Periodic Patterns

    Shengsheng Lin, Weiwei Lin, Xinyi Hu +3

    cs.LGarXiv:2409.18479v22024
  24. Robometer: Scaling General-Purpose Robotic Reward Models via Trajectory Comparisons

    Anthony Liang, Yigit Korkmaz, Jiahui Zhang +14

    cs.ROcs.AIcs.LGarXiv:2603.02115v22026
  25. Conditional Affordance Learning for Driving in Urban Environments

    Axel Sauer, Nikolay Savinov, Andreas Geiger

    cs.ROcs.LGeess.SYarXiv:1806.06498v32018
  26. Nemotron 3 Nano: Open, Efficient Mixture-of-Experts Hybrid Mamba-Transformer Model for Agentic Reasoning

    NVIDIA, :, Aaron Blakeman +311

    cs.CLcs.AIcs.LGarXiv:2512.20848v12025
  27. WebAgent-R1: Training Web Agents via End-to-End Multi-Turn Reinforcement Learning

    Zhepei Wei, Wenlin Yao, Yao Liu +9

    cs.CLcs.LGarXiv:2505.16421v22025
  28. DeepCorr: Strong Flow Correlation Attacks on Tor Using Deep Learning

    Milad Nasr, Alireza Bahramali, Amir Houmansadr

    cs.CRcs.LGarXiv:1808.07285v12018
  29. TiRex: Zero-Shot Forecasting Across Long and Short Horizons with Enhanced In-Context Learning

    Andreas Auer, Patrick Podest, Daniel Klotz +3

    cs.LGarXiv:2505.23719v22025
  30. SparseTSF: Modeling Long-term Time Series Forecasting with 1k Parameters

    Shengsheng Lin, Weiwei Lin, Wentai Wu +2

    cs.LGarXiv:2405.00946v22024
  31. Sponge Examples: Energy-Latency Attacks on Neural Networks

    Ilia Shumailov, Yiren Zhao, Daniel Bates +3

    cs.LGcs.CLcs.CRarXiv:2006.03463v22020
  32. Training Agents Inside of Scalable World Models

    Danijar Hafner, Wilson Yan, Timothy Lillicrap

    cs.AIcs.LGcs.ROarXiv:2509.24527v12025
  33. The Polar Express: Optimal Matrix Sign Methods and Their Application to the Muon Algorithm

    Noah Amsel, David Persson, Christopher Musco +1

    cs.LGcs.AIcs.CLarXiv:2505.16932v52025
  34. SAEBench: A Comprehensive Benchmark for Sparse Autoencoders in Language Model Interpretability

    Adam Karvonen, Can Rager, Johnny Lin +12

    cs.LGcs.CLarXiv:2503.09532v42025
  35. Understanding plasticity in neural networks

    Clare Lyle, Zeyu Zheng, Evgenii Nikishin +3

    cs.LGarXiv:2303.01486v42023
  36. Programming Refusal with Conditional Activation Steering

    Bruce W. Lee, Inkit Padhi, Karthikeyan Natesan Ramamurthy +4

    cs.LGcs.AIcs.CLarXiv:2409.05907v32024
  37. Controllable Image Captioning with Prompt-Conditioned Scene Rewards

    Jongyeop Hyun, Taeyoung Kim, Hyounghun Kim

    cs.CVcs.CLcs.LGarXiv:2609.00709v12026
  38. OPAL: Offline Primitive Discovery for Accelerating Offline Reinforcement Learning

    Anurag Ajay, Aviral Kumar, Pulkit Agrawal +2

    cs.LGarXiv:2010.13611v32020
  39. Vision-OPD: Learning to See Fine Details for Multimodal LLMs via On-Policy Self-Distillation

    Qianhao Yuan, Jie Lou, Xing Yu +4

    cs.CVcs.AIcs.CLarXiv:2605.18740v42026
  40. PixMix: Dreamlike Pictures Comprehensively Improve Safety Measures

    Dan Hendrycks, Andy Zou, Mantas Mazeika +4

    cs.LGcs.CVarXiv:2112.05135v32021
  41. GOOD: A Graph Out-of-Distribution Benchmark

    Shurui Gui, Xiner Li, Limei Wang +1

    cs.LGcs.AIarXiv:2206.08452v22022
  42. REVE: A Foundation Model for EEG -- Adapting to Any Setup with Large-Scale Pretraining on 25,000 Subjects

    Yassine El Ouahidi, Jonathan Lys, Philipp Thölke +5

    cs.LGq-bio.NCarXiv:2510.21585v12025
  43. Deep Exemplar-based Video Colorization

    Bo Zhang, Mingming He, Jing Liao +4

    cs.CVcs.AIcs.LGarXiv:1906.09909v12019
  44. Act Only When It Pays: Efficient Reinforcement Learning for LLM Reasoning via Selective Rollouts

    Haizhong Zheng, Yang Zhou, Brian R. Bartoldson +4

    cs.AIcs.LGarXiv:2506.02177v12025
  45. FILM: Following Instructions in Language with Modular Methods

    So Yeon Min, Devendra Singh Chaplot, Pradeep Ravikumar +2

    cs.CLcs.LGarXiv:2110.07342v32021
  46. TurboQuant: Online Vector Quantization with Near-optimal Distortion Rate

    Amir Zandieh, Majid Daliri, Majid Hadian +1

    cs.LGcs.AIcs.DBarXiv:2504.19874v12025
  47. Learning Dexterous Manipulation for a Soft Robotic Hand from Human Demonstration

    Abhishek Gupta, Clemens Eppner, Sergey Levine +1

    cs.LGcs.ROarXiv:1603.06348v32016
  48. Multi-Scale Adaptive Graph Neural Network for Multivariate Time Series Forecasting

    Ling Chen, Donghui Chen, Zongjiang Shang +4

    cs.LGarXiv:2201.04828v22022
  49. Group Adaptive Clipping Policy Optimization

    Sheng Jia, Xiao Wang, Shiva Prasad Kasiviswanathan +1

    cs.LGcs.CLarXiv:2609.00444v12026
  50. A Minimalist Approach to LLM Reasoning: from Rejection Sampling to Reinforce

    Wei Xiong, Jiarui Yao, Yuhui Xu +8

    cs.LGcs.AIcs.CLarXiv:2504.11343v22025
  51. dLLM: Simple Diffusion Language Modeling

    Zhanhui Zhou, Lingjie Chen, Hanghang Tong +1

    cs.CLcs.AIcs.LGarXiv:2602.22661v12026
  52. DR Tulu: Reinforcement Learning with Evolving Rubrics for Deep Research

    Rulin Shao, Akari Asai, Shannon Zejiang Shen +18

    cs.CLcs.AIcs.LGarXiv:2511.19399v32025
  53. Towards Understanding Regularization in Batch Normalization

    Ping Luo, Xinjiang Wang, Wenqi Shao +1

    cs.LGcs.CVeess.SYarXiv:1809.00846v42018
  54. Learning to Win by Reading Manuals in a Monte-Carlo Framework

    S. R. K. Branavan, David Silver, Regina Barzilay

    cs.CLcs.AIcs.LGarXiv:1401.5390v12014
  55. Understanding graph embedding methods and their applications

    Mengjia Xu

    cs.LGcs.ITcs.SIarXiv:2012.08019v12020
  56. Large Models for Time Series and Spatio-Temporal Data: A Survey and Outlook

    Ming Jin, Yaxuan Kong, Yuxuan Liang +13

    cs.LGcs.AIarXiv:2310.10196v32023
  57. Graph Neural Networks: Taxonomy, Advances and Trends

    Yu Zhou, Haixia Zheng, Xin Huang +3

    cs.LGarXiv:2012.08752v42020
  58. Exploration in Deep Reinforcement Learning: From Single-Agent to Multiagent Domain

    Jianye Hao, Tianpei Yang, Hongyao Tang +5

    cs.AIcs.LGcs.MAarXiv:2109.06668v62021
  59. Bandwidth-Agile Image Transmission with Deep Joint Source-Channel Coding

    David Burth Kurka, Deniz Gündüz

    cs.ITcs.LGeess.IVarXiv:2009.12480v22020
  60. Day-Ahead Hourly Forecasting of Power Generation from Photovoltaic Plants

    Lorenzo Gigoni, Alessandro Betti, Emanuele Crisostomi +4

    cs.LGstat.MLarXiv:1903.06800v12019