Machine Learning

Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

781 to 840 of 20,193

  1. Information-Induced Training Geometry: Exact Reduction, Canonical Completion, and Structured Expressivity

    Zavier Li

    cs.LGmath.OCarXiv:2609.12991v12026
    Summaries:한국어
  2. AttnLRP: Attention-Aware Layer-Wise Relevance Propagation for Transformers

    Reduan Achtibat, Sayed Mohammad Vakilzadeh Hatefi, Maximilian Dreyer +4

    cs.CLcs.AIcs.CVarXiv:2402.05602v22024
  3. It Takes Two: Your GRPO Is Secretly DPO

    Yihong Wu, Liheng Ma, Lei Ding +9

    cs.LGcs.CLarXiv:2510.00977v32025
  4. Task-Embedded Control Networks for Few-Shot Imitation Learning

    Stephen James, Michael Bloesch, Andrew J. Davison

    cs.ROcs.AIcs.CVarXiv:1810.03237v12018
  5. Verifier-free Test-Time Sampling for Vision-Language-Action Models

    Suhyeok Jang, Dongyoung Kim, Changyeon Kim +2

    cs.ROcs.AIcs.LGarXiv:2510.05681v22025
  6. RLAD: Training LLMs to Discover Abstractions for Solving Reasoning Problems

    Yuxiao Qu, Anikait Singh, Yoonho Lee +4

    cs.AIcs.CLcs.LGarXiv:2510.02263v12025
  7. Pretraining Representations for Data-Efficient Reinforcement Learning

    Max Schwarzer, Nitarshan Rajkumar, Michael Noukhovitch +5

    cs.LGarXiv:2106.04799v12021
  8. Prospective Coding Improves Learning in Deep Continuous-Time Recurrent Networks

    Shivang Rawat, Mirko Morello, Flaviano Morone +1

    cs.LGcs.NEq-bio.NCarXiv:2609.04134v12026
  9. Machine learning methods to detect money laundering in the Bitcoin blockchain in the presence of label scarcity

    Joana Lorenz, Maria Inês Silva, David Aparício +2

    cs.LGstat.MLarXiv:2005.14635v22020
  10. WeatherNext 3: Increasing resolution and performance of global weather models with raw observations

    Stephan Rasp, Boris Babenko, Dominic Masters +22

    cs.LGarXiv:2609.03582v12026
    Summaries:한국어
  11. Pushing the (Decision) Boundaries: Dynamically Calibrating Differentially Private Noise to Explainability in Federated Learning

    Michael Khavkin, Kichang Lee, Jaeho Jin +2

    cs.LGarXiv:2609.03851v12026
  12. SWIM: Student Writing Simulation via Proficiency-Conditioned Generation

    Heejin Do, Jakub Kontak, Mrinmaya Sachan

    cs.CLcs.LGarXiv:2609.03215v12026
  13. Local Updates, Global Learning (LUGL): Playing Games with non-incremental Learners

    David Milec, Spyridon Samothrakis, Michael Fairbank +1

    cs.LGcs.AIarXiv:2609.03660v12026
  14. EEG-FM-Compass: Progress, Benchmarking, and Future Directions for EEG Foundation Models

    Dingkun Liu, Yuheng Chen, Zhu Chen +5

    cs.LGcs.CVarXiv:2601.17883v32026
  15. LLM Maybe LongLM: Self-Extend LLM Context Window Without Tuning

    Hongye Jin, Xiaotian Han, Jingfeng Yang +5

    cs.CLcs.AIcs.LGarXiv:2401.01325v32024
  16. Mamba: Linear-Time Sequence Modeling with Selective State Spaces

    Albert Gu, Tri Dao

    cs.LGcs.AIarXiv:2312.00752v22023
    Summaries:한국어
  17. Symbolic Discovery of Optimization Algorithms

    Xiangning Chen, Chen Liang, Da Huang +9

    cs.LGcs.AIcs.CLarXiv:2302.06675v42023
  18. Matryoshka Representation Learning

    Aditya Kusupati, Gantavya Bhatt, Aniket Rege +8

    cs.LGcs.CVarXiv:2205.13147v42022
  19. ST3D: Self-training for Unsupervised Domain Adaptation on 3D Object Detection

    Jihan Yang, Shaoshuai Shi, Zhe Wang +2

    cs.CVcs.LGarXiv:2103.05346v22021
  20. Domain Generalization using Causal Matching

    Divyat Mahajan, Shruti Tople, Amit Sharma

    cs.LGcs.AIstat.MLarXiv:2006.07500v32020
  21. Rethinking Class-Balanced Methods for Long-Tailed Visual Recognition from a Domain Adaptation Perspective

    Muhammad Abdullah Jamal, Matthew Brown, Ming-Hsuan Yang +2

    cs.CVcs.LGstat.MLarXiv:2003.10780v12020
  22. Combating noisy labels by agreement: A joint training method with co-regularization

    Hongxin Wei, Lei Feng, Xiangyu Chen +1

    cs.CVcs.LGstat.MLarXiv:2003.02752v32020
  23. Knowledge Graph Embedding for Link Prediction: A Comparative Analysis

    Andrea Rossi, Donatella Firmani, Antonio Matinata +2

    cs.LGcs.DBstat.MLarXiv:2002.00819v42020
  24. Adversarial Domain Adaptation with Domain Mixup

    Minghao Xu, Jian Zhang, Bingbing Ni +4

    cs.CVcs.LGarXiv:1912.01805v12019
  25. Variational Graph Recurrent Neural Networks

    Ehsan Hajiramezanali, Arman Hasanzadeh, Nick Duffield +3

    cs.LGstat.MLarXiv:1908.09710v32019
  26. Uncertainty-based Continual Learning with Adaptive Regularization

    Hongjoon Ahn, Sungmin Cha, Donggyu Lee +1

    cs.LGstat.MLarXiv:1905.11614v32019
  27. Simplifying Graph Convolutional Networks

    Felix Wu, Tianyi Zhang, Amauri Holanda de Souza +3

    cs.LGstat.MLarXiv:1902.07153v22019
  28. Formal Limitations on the Measurement of Mutual Information

    David McAllester, Karl Stratos

    cs.ITcs.LGstat.MLarXiv:1811.04251v42018
  29. BOHB: Robust and Efficient Hyperparameter Optimization at Scale

    Stefan Falkner, Aaron Klein, Frank Hutter

    cs.LGstat.MLarXiv:1807.01774v12018
  30. SoK: The Faults in our ASRs: An Overview of Attacks against Automatic Speech Recognition and Speaker Identification Systems

    Hadi Abdullah, Kevin Warren, Vincent Bindschaedler +2

    cs.CRcs.LGcs.SDarXiv:2007.06622v32020
  31. Multi-Agent Actor-Critic with Hierarchical Graph Attention Network

    Heechang Ryu, Hayong Shin, Jinkyoo Park

    cs.LGcs.AIcs.MAarXiv:1909.12557v22019
  32. A Large Open Multi-Energy Corpus of Soil Compaction Tests, with Machine-Learning Baselines

    Sompote Youwai, Chana Phutthananon, Warat Kongkitkul

    cs.LGarXiv:2609.03337v12026
  33. Shaping capabilities with token-level data filtering

    Neil Rathi, Alec Radford

    cs.LGcs.AIcs.CLarXiv:2601.21571v22026
  34. Reinforcement Learning via Self-Distillation

    Jonas Hübotter, Frederike Lübeck, Lejs Behric +8

    cs.LGcs.AIarXiv:2601.20802v22026
  35. Stronger Normalization-Free Transformers

    Mingzhi Chen, Taiming Lu, Jiachen Zhu +2

    cs.LGcs.AIcs.CLarXiv:2512.10938v22025
  36. SceneWeaver: All-in-One 3D Scene Synthesis with an Extensible and Self-Reflective Agent

    Yandan Yang, Baoxiong Jia, Shujie Zhang +1

    cs.GRcs.CVcs.LGarXiv:2509.20414v22025
  37. SPIRAL: Self-Play on Zero-Sum Games Incentivizes Reasoning via Multi-Agent Multi-Turn Reinforcement Learning

    Bo Liu, Leon Guertler, Simon Yu +9

    cs.AIcs.CLcs.LGarXiv:2506.24119v32025
  38. Improving and Simplifying Pattern Exploiting Training

    Derek Tam, Rakesh R Menon, Mohit Bansal +2

    cs.CLcs.AIcs.LGarXiv:2103.11955v32021
  39. Measuring Mathematical Problem Solving With the MATH Dataset

    Dan Hendrycks, Collin Burns, Saurav Kadavath +5

    cs.LGcs.AIcs.CLarXiv:2103.03874v22021
  40. Improving GANs Using Optimal Transport

    Tim Salimans, Han Zhang, Alec Radford +1

    cs.LGstat.MLarXiv:1803.05573v12018
  41. Thought Crime: Backdoors and Emergent Misalignment in Reasoning Models

    James Chua, Jan Betley, Mia Taylor +1

    cs.LGcs.AIcs.CLarXiv:2506.13206v22025
  42. Neural Architecture Search with Bayesian Optimisation and Optimal Transport

    Kirthevasan Kandasamy, Willie Neiswanger, Jeff Schneider +2

    cs.LGstat.MLarXiv:1802.07191v32018
  43. SmolVLA: A Vision-Language-Action Model for Affordable and Efficient Robotics

    Mustafa Shukor, Dana Aubakirova, Francesco Capuano +11

    cs.LGcs.ROarXiv:2506.01844v12025
  44. Deterministic Non-Autoregressive Neural Sequence Modeling by Iterative Refinement

    Jason Lee, Elman Mansimov, Kyunghyun Cho

    cs.LGcs.CLstat.MLarXiv:1802.06901v32018
  45. Neural Voice Cloning with a Few Samples

    Sercan O. Arik, Jitong Chen, Kainan Peng +2

    cs.CLcs.LGcs.SDarXiv:1802.06006v32018
  46. Multi-Task Reinforcement Learning with Context-based Representations

    Shagun Sodhani, Amy Zhang, Joelle Pineau

    cs.LGcs.AIcs.ROarXiv:2102.06177v22021
  47. Vision Language Models are Biased

    An Vo, Khai-Nguyen Nguyen, Mohammad Reza Taesiri +3

    cs.LGcs.CVarXiv:2505.23941v42025
  48. Efficient Exploration through Bayesian Deep Q-Networks

    Kamyar Azizzadenesheli, Animashree Anandkumar

    cs.AIcs.LGstat.MLarXiv:1802.04412v42018
  49. Explicit Inductive Bias for Transfer Learning with Convolutional Networks

    Xuhong Li, Yves Grandvalet, Franck Davoine

    cs.LGarXiv:1802.01483v22018
  50. Structured Prediction as Translation between Augmented Natural Languages

    Giovanni Paolini, Ben Athiwaratkun, Jason Krone +6

    cs.LGcs.CLarXiv:2101.05779v32021
  51. Learning coordinated badminton skills for legged manipulators

    Yuntao Ma, Andrei Cramariuc, Farbod Farshidian +1

    cs.ROcs.LGarXiv:2505.22974v22025
  52. Deep Neural Networks for Survival Analysis Based on a Multi-Task Framework

    Stephane Fotso

    stat.MLcs.LGarXiv:1801.05512v12018
  53. Hardware and Software Optimizations for Accelerating Deep Neural Networks: Survey of Current Trends, Challenges, and the Road Ahead

    Maurizio Capra, Beatrice Bussolino, Alberto Marchisio +3

    cs.ARcs.LGarXiv:2012.11233v12020
  54. Multi-expert learning of adaptive legged locomotion

    Chuanyu Yang, Kai Yuan, Qiuguo Zhu +2

    cs.ROcs.AIcs.LGarXiv:2012.05810v12020
  55. On the Binding Problem in Artificial Neural Networks

    Klaus Greff, Sjoerd van Steenkiste, Jürgen Schmidhuber

    cs.NEcs.AIcs.LGarXiv:2012.05208v12020
  56. Deep Learning for Medical Anomaly Detection -- A Survey

    Tharindu Fernando, Harshala Gammulle, Simon Denman +2

    cs.LGcs.CVeess.IVarXiv:2012.02364v22020
  57. Improved Contrastive Divergence Training of Energy Based Models

    Yilun Du, Shuang Li, Joshua Tenenbaum +1

    cs.LGarXiv:2012.01316v42020
  58. Reinforcement Learning for Reasoning in Large Language Models with One Training Example

    Yiping Wang, Qing Yang, Zhiyuan Zeng +11

    cs.LGcs.AIcs.CLarXiv:2504.20571v32025
  59. Gradient Starvation: A Learning Proclivity in Neural Networks

    Mohammad Pezeshki, Sékou-Oumar Kaba, Yoshua Bengio +3

    cs.LGmath.DSstat.MLarXiv:2011.09468v42020
  60. DARLA: Improving Zero-Shot Transfer in Reinforcement Learning

    Irina Higgins, Arka Pal, Andrei A. Rusu +6

    stat.MLcs.AIcs.LGarXiv:1707.08475v22017