Machine Learning
Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
781 to 840 of 20,193
Information-Induced Training Geometry: Exact Reduction, Canonical Completion, and Structured Expressivity
Zavier Li
cs.LGmath.OCarXiv:2609.12991v12026Summaries:한국어AttnLRP: Attention-Aware Layer-Wise Relevance Propagation for Transformers
Reduan Achtibat, Sayed Mohammad Vakilzadeh Hatefi, Maximilian Dreyer +4
cs.CLcs.AIcs.CVarXiv:2402.05602v22024It Takes Two: Your GRPO Is Secretly DPO
Yihong Wu, Liheng Ma, Lei Ding +9
cs.LGcs.CLarXiv:2510.00977v32025Task-Embedded Control Networks for Few-Shot Imitation Learning
Stephen James, Michael Bloesch, Andrew J. Davison
cs.ROcs.AIcs.CVarXiv:1810.03237v12018Verifier-free Test-Time Sampling for Vision-Language-Action Models
Suhyeok Jang, Dongyoung Kim, Changyeon Kim +2
cs.ROcs.AIcs.LGarXiv:2510.05681v22025RLAD: Training LLMs to Discover Abstractions for Solving Reasoning Problems
Yuxiao Qu, Anikait Singh, Yoonho Lee +4
cs.AIcs.CLcs.LGarXiv:2510.02263v12025Pretraining Representations for Data-Efficient Reinforcement Learning
Max Schwarzer, Nitarshan Rajkumar, Michael Noukhovitch +5
cs.LGarXiv:2106.04799v12021Prospective Coding Improves Learning in Deep Continuous-Time Recurrent Networks
Shivang Rawat, Mirko Morello, Flaviano Morone +1
cs.LGcs.NEq-bio.NCarXiv:2609.04134v12026Machine learning methods to detect money laundering in the Bitcoin blockchain in the presence of label scarcity
Joana Lorenz, Maria Inês Silva, David Aparício +2
cs.LGstat.MLarXiv:2005.14635v22020WeatherNext 3: Increasing resolution and performance of global weather models with raw observations
Stephan Rasp, Boris Babenko, Dominic Masters +22
cs.LGarXiv:2609.03582v12026Summaries:한국어Pushing the (Decision) Boundaries: Dynamically Calibrating Differentially Private Noise to Explainability in Federated Learning
Michael Khavkin, Kichang Lee, Jaeho Jin +2
cs.LGarXiv:2609.03851v12026SWIM: Student Writing Simulation via Proficiency-Conditioned Generation
Heejin Do, Jakub Kontak, Mrinmaya Sachan
cs.CLcs.LGarXiv:2609.03215v12026Local Updates, Global Learning (LUGL): Playing Games with non-incremental Learners
David Milec, Spyridon Samothrakis, Michael Fairbank +1
cs.LGcs.AIarXiv:2609.03660v12026EEG-FM-Compass: Progress, Benchmarking, and Future Directions for EEG Foundation Models
Dingkun Liu, Yuheng Chen, Zhu Chen +5
cs.LGcs.CVarXiv:2601.17883v32026LLM Maybe LongLM: Self-Extend LLM Context Window Without Tuning
Hongye Jin, Xiaotian Han, Jingfeng Yang +5
cs.CLcs.AIcs.LGarXiv:2401.01325v32024Mamba: Linear-Time Sequence Modeling with Selective State Spaces
Albert Gu, Tri Dao
cs.LGcs.AIarXiv:2312.00752v22023Summaries:한국어Symbolic Discovery of Optimization Algorithms
Xiangning Chen, Chen Liang, Da Huang +9
cs.LGcs.AIcs.CLarXiv:2302.06675v42023Matryoshka Representation Learning
Aditya Kusupati, Gantavya Bhatt, Aniket Rege +8
cs.LGcs.CVarXiv:2205.13147v42022ST3D: Self-training for Unsupervised Domain Adaptation on 3D Object Detection
Jihan Yang, Shaoshuai Shi, Zhe Wang +2
cs.CVcs.LGarXiv:2103.05346v22021Domain Generalization using Causal Matching
Divyat Mahajan, Shruti Tople, Amit Sharma
cs.LGcs.AIstat.MLarXiv:2006.07500v32020Rethinking Class-Balanced Methods for Long-Tailed Visual Recognition from a Domain Adaptation Perspective
Muhammad Abdullah Jamal, Matthew Brown, Ming-Hsuan Yang +2
cs.CVcs.LGstat.MLarXiv:2003.10780v12020Combating noisy labels by agreement: A joint training method with co-regularization
Hongxin Wei, Lei Feng, Xiangyu Chen +1
cs.CVcs.LGstat.MLarXiv:2003.02752v32020Knowledge Graph Embedding for Link Prediction: A Comparative Analysis
Andrea Rossi, Donatella Firmani, Antonio Matinata +2
cs.LGcs.DBstat.MLarXiv:2002.00819v42020Adversarial Domain Adaptation with Domain Mixup
Minghao Xu, Jian Zhang, Bingbing Ni +4
cs.CVcs.LGarXiv:1912.01805v12019Variational Graph Recurrent Neural Networks
Ehsan Hajiramezanali, Arman Hasanzadeh, Nick Duffield +3
cs.LGstat.MLarXiv:1908.09710v32019Uncertainty-based Continual Learning with Adaptive Regularization
Hongjoon Ahn, Sungmin Cha, Donggyu Lee +1
cs.LGstat.MLarXiv:1905.11614v32019Simplifying Graph Convolutional Networks
Felix Wu, Tianyi Zhang, Amauri Holanda de Souza +3
cs.LGstat.MLarXiv:1902.07153v22019Formal Limitations on the Measurement of Mutual Information
David McAllester, Karl Stratos
cs.ITcs.LGstat.MLarXiv:1811.04251v42018BOHB: Robust and Efficient Hyperparameter Optimization at Scale
Stefan Falkner, Aaron Klein, Frank Hutter
cs.LGstat.MLarXiv:1807.01774v12018SoK: The Faults in our ASRs: An Overview of Attacks against Automatic Speech Recognition and Speaker Identification Systems
Hadi Abdullah, Kevin Warren, Vincent Bindschaedler +2
cs.CRcs.LGcs.SDarXiv:2007.06622v32020Multi-Agent Actor-Critic with Hierarchical Graph Attention Network
Heechang Ryu, Hayong Shin, Jinkyoo Park
cs.LGcs.AIcs.MAarXiv:1909.12557v22019A Large Open Multi-Energy Corpus of Soil Compaction Tests, with Machine-Learning Baselines
Sompote Youwai, Chana Phutthananon, Warat Kongkitkul
cs.LGarXiv:2609.03337v12026Shaping capabilities with token-level data filtering
Neil Rathi, Alec Radford
cs.LGcs.AIcs.CLarXiv:2601.21571v22026Reinforcement Learning via Self-Distillation
Jonas Hübotter, Frederike Lübeck, Lejs Behric +8
cs.LGcs.AIarXiv:2601.20802v22026Stronger Normalization-Free Transformers
Mingzhi Chen, Taiming Lu, Jiachen Zhu +2
cs.LGcs.AIcs.CLarXiv:2512.10938v22025SceneWeaver: All-in-One 3D Scene Synthesis with an Extensible and Self-Reflective Agent
Yandan Yang, Baoxiong Jia, Shujie Zhang +1
cs.GRcs.CVcs.LGarXiv:2509.20414v22025SPIRAL: Self-Play on Zero-Sum Games Incentivizes Reasoning via Multi-Agent Multi-Turn Reinforcement Learning
Bo Liu, Leon Guertler, Simon Yu +9
cs.AIcs.CLcs.LGarXiv:2506.24119v32025Improving and Simplifying Pattern Exploiting Training
Derek Tam, Rakesh R Menon, Mohit Bansal +2
cs.CLcs.AIcs.LGarXiv:2103.11955v32021Measuring Mathematical Problem Solving With the MATH Dataset
Dan Hendrycks, Collin Burns, Saurav Kadavath +5
cs.LGcs.AIcs.CLarXiv:2103.03874v22021Improving GANs Using Optimal Transport
Tim Salimans, Han Zhang, Alec Radford +1
cs.LGstat.MLarXiv:1803.05573v12018Thought Crime: Backdoors and Emergent Misalignment in Reasoning Models
James Chua, Jan Betley, Mia Taylor +1
cs.LGcs.AIcs.CLarXiv:2506.13206v22025Neural Architecture Search with Bayesian Optimisation and Optimal Transport
Kirthevasan Kandasamy, Willie Neiswanger, Jeff Schneider +2
cs.LGstat.MLarXiv:1802.07191v32018SmolVLA: A Vision-Language-Action Model for Affordable and Efficient Robotics
Mustafa Shukor, Dana Aubakirova, Francesco Capuano +11
cs.LGcs.ROarXiv:2506.01844v12025Deterministic Non-Autoregressive Neural Sequence Modeling by Iterative Refinement
Jason Lee, Elman Mansimov, Kyunghyun Cho
cs.LGcs.CLstat.MLarXiv:1802.06901v32018Neural Voice Cloning with a Few Samples
Sercan O. Arik, Jitong Chen, Kainan Peng +2
cs.CLcs.LGcs.SDarXiv:1802.06006v32018Multi-Task Reinforcement Learning with Context-based Representations
Shagun Sodhani, Amy Zhang, Joelle Pineau
cs.LGcs.AIcs.ROarXiv:2102.06177v22021Vision Language Models are Biased
An Vo, Khai-Nguyen Nguyen, Mohammad Reza Taesiri +3
cs.LGcs.CVarXiv:2505.23941v42025Efficient Exploration through Bayesian Deep Q-Networks
Kamyar Azizzadenesheli, Animashree Anandkumar
cs.AIcs.LGstat.MLarXiv:1802.04412v42018Explicit Inductive Bias for Transfer Learning with Convolutional Networks
Xuhong Li, Yves Grandvalet, Franck Davoine
cs.LGarXiv:1802.01483v22018Structured Prediction as Translation between Augmented Natural Languages
Giovanni Paolini, Ben Athiwaratkun, Jason Krone +6
cs.LGcs.CLarXiv:2101.05779v32021Learning coordinated badminton skills for legged manipulators
Yuntao Ma, Andrei Cramariuc, Farbod Farshidian +1
cs.ROcs.LGarXiv:2505.22974v22025Deep Neural Networks for Survival Analysis Based on a Multi-Task Framework
Stephane Fotso
stat.MLcs.LGarXiv:1801.05512v12018Hardware and Software Optimizations for Accelerating Deep Neural Networks: Survey of Current Trends, Challenges, and the Road Ahead
Maurizio Capra, Beatrice Bussolino, Alberto Marchisio +3
cs.ARcs.LGarXiv:2012.11233v12020Multi-expert learning of adaptive legged locomotion
Chuanyu Yang, Kai Yuan, Qiuguo Zhu +2
cs.ROcs.AIcs.LGarXiv:2012.05810v12020On the Binding Problem in Artificial Neural Networks
Klaus Greff, Sjoerd van Steenkiste, Jürgen Schmidhuber
cs.NEcs.AIcs.LGarXiv:2012.05208v12020Deep Learning for Medical Anomaly Detection -- A Survey
Tharindu Fernando, Harshala Gammulle, Simon Denman +2
cs.LGcs.CVeess.IVarXiv:2012.02364v22020Improved Contrastive Divergence Training of Energy Based Models
Yilun Du, Shuang Li, Joshua Tenenbaum +1
cs.LGarXiv:2012.01316v42020Reinforcement Learning for Reasoning in Large Language Models with One Training Example
Yiping Wang, Qing Yang, Zhiyuan Zeng +11
cs.LGcs.AIcs.CLarXiv:2504.20571v32025Gradient Starvation: A Learning Proclivity in Neural Networks
Mohammad Pezeshki, Sékou-Oumar Kaba, Yoshua Bengio +3
cs.LGmath.DSstat.MLarXiv:2011.09468v42020DARLA: Improving Zero-Shot Transfer in Reinforcement Learning
Irina Higgins, Arka Pal, Andrei A. Rusu +6
stat.MLcs.AIcs.LGarXiv:1707.08475v22017