Machine Learning
Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
901 to 960 of 20,192
Self-Play Fine-Tuning Converts Weak Language Models to Strong Language Models
Zixiang Chen, Yihe Deng, Huizhuo Yuan +2
cs.LGcs.AIcs.CLarXiv:2401.01335v32024Synthetic Worlds for Temporal Evaluation and Knowledge Updating in LLMs
Jonathan Zheng, Zirui Shao, Alan Ritter +1
cs.CLcs.LGarXiv:2609.00184v12026Backprop KF: Learning Discriminative Deterministic State Estimators
Tuomas Haarnoja, Anurag Ajay, Sergey Levine +1
cs.LGcs.AIarXiv:1605.07148v42016Different representation learning objectives recover distinct latent structures from the same psychometric data
Cong Cao, Tassos C. Kyriakides, Pambos Vrasidas
cs.AIcs.LGstat.MEarXiv:2609.00100v12026Align Your Gaussians: Text-to-4D with Dynamic 3D Gaussians and Composed Diffusion Models
Huan Ling, Seung Wook Kim, Antonio Torralba +2
cs.CVcs.LGarXiv:2312.13763v22023Learning Dynamics of Logits Debiasing for Long-Tailed Semi-Supervised Learning
Yue Cheng, Jiajun Zhang, Xiaohui Gao +2
cs.LGcs.AIarXiv:2608.30699v12026Steering Llama 2 via Contrastive Activation Addition
Nina Panickssery, Nick Gabrieli, Julian Schulz +3
cs.CLcs.AIcs.LGarXiv:2312.06681v42023Baseline Defenses for Adversarial Attacks Against Aligned Language Models
Neel Jain, Avi Schwarzschild, Yuxin Wen +7
cs.LGcs.CLcs.CRarXiv:2309.00614v22023D4: Improving LLM Pretraining via Document De-Duplication and Diversification
Kushal Tirumala, Daniel Simig, Armen Aghajanyan +1
cs.CLcs.AIcs.LGarXiv:2308.12284v12023PMET: Precise Model Editing in a Transformer
Xiaopeng Li, Shasha Li, Shezheng Song +3
cs.CLcs.AIcs.LGarXiv:2308.08742v62023On the use of deep learning for phase recovery
Kaiqiang Wang, Li Song, Chutian Wang +8
physics.opticscs.LGeess.IVarXiv:2308.00942v12023Aligning Multi-Trajectory Supervision with Policy Optimization for VLA Driving
Tian Zhang, Zhuo Huang, Hongrui Ye +3
cs.CVcs.AIcs.LGarXiv:2608.30122v12026A Survey of Techniques for Optimizing Transformer Inference
Krishna Teja Chitty-Venkata, Sparsh Mittal, Murali Emani +2
cs.LGcs.ARcs.CLarXiv:2307.07982v12023Generating images with recurrent adversarial networks
Daniel Jiwoong Im, Chris Dongjoo Kim, Hui Jiang +1
cs.LGcs.CVarXiv:1602.05110v52016A-MADiff: Attention-Guided Multi-Agent DRL with Diffusion Policies for Memory-Aware Task Orchestration in Mobile AIGC Networks
Chongzhi Wu, Zhengtao Li, Jiawen Kang +4
cs.NIcs.LGarXiv:2608.29255v12026Non-Autoregressive Machine Translation with Latent Alignments
Chitwan Saharia, William Chan, Saurabh Saxena +1
cs.CLcs.LGarXiv:2004.07437v32020Training with Quantization Noise for Extreme Model Compression
Angela Fan, Pierre Stock, Benjamin Graham +4
cs.LGstat.MLarXiv:2004.07320v32020Knowledge Distillation and Student-Teacher Learning for Visual Intelligence: A Review and New Outlooks
Lin Wang, Kuk-Jin Yoon
cs.CVcs.AIcs.LGarXiv:2004.05937v72020Self-Refine: Iterative Refinement with Self-Feedback
Aman Madaan, Niket Tandon, Prakhar Gupta +13
cs.CLcs.AIcs.LGarXiv:2303.17651v22023Summaries:한국어Diffusion Schrödinger Bridge Matching
Yuyang Shi, Valentin De Bortoli, Andrew Campbell +1
stat.MLcs.LGarXiv:2303.16852v32023Off-Policy Evaluation for Semantic ID Recommenders: Does the Model's Own Code Hierarchy Help?
Artem Betlei
cs.LGarXiv:2608.28905v12026Foundation Models and Fair Use
Peter Henderson, Xuechen Li, Dan Jurafsky +3
cs.CYcs.AIcs.LGarXiv:2303.15715v12023Neural Programmer-Interpreters
Scott Reed, Nando de Freitas
cs.LGcs.NEarXiv:1511.06279v42015Stabilizing Transformer Training by Preventing Attention Entropy Collapse
Shuangfei Zhai, Tatiana Likhomanenko, Etai Littwin +5
cs.LGcs.AIcs.CLarXiv:2303.06296v22023Comprehensive Review of Deep Reinforcement Learning Methods and Applications in Economics
Amir Mosavi, Pedram Ghamisi, Yaser Faghan +1
q-fin.STcs.LGecon.GNarXiv:2004.01509v12020Deep learning for Stock Market Prediction
Mojtaba Nabipour, Pooyan Nayyeri, Hamed Jabani +1
q-fin.STcs.LGarXiv:2004.01497v12020Timing-Aware Repurchase Prediction for Web-Scale E-Commerce: Survival Models for Multi-Surface Grocery Recommendation
Akshay Kekuda, Shreeranjani Srirangamsridharan, Ishan Bhatt +5
cs.AIcs.LGarXiv:2608.28393v12026Dynamic Multiscale Graph Neural Networks for 3D Skeleton-Based Human Motion Prediction
Maosen Li, Siheng Chen, Yangheng Zhao +3
cs.CVcs.LGstat.MLarXiv:2003.08802v12020COEVOLVE: A Joint Point Process Model for Information Diffusion and Network Co-evolution
Mehrdad Farajtabar, Yichen Wang, Manuel Gomez Rodriguez +3
cs.SIcs.LGphysics.soc-pharXiv:1507.02293v22015Rapid AI Development Cycle for the Coronavirus (COVID-19) Pandemic: Initial Results for Automated Detection & Patient Monitoring using Deep Learning CT Image Analysis
Ophir Gozes, Maayan Frid-Adar, Hayit Greenspan +5
eess.IVcs.CVcs.LGarXiv:2003.05037v32020A Survey on The Expressive Power of Graph Neural Networks
Ryoma Sato
cs.LGstat.MLarXiv:2003.04078v420203D Equivariant Diffusion for Target-Aware Molecule Generation and Affinity Prediction
Jiaqi Guan, Wesley Wei Qian, Xingang Peng +3
q-bio.BMcs.LGarXiv:2303.03543v12023Knowledge Graphs
Aidan Hogan, Eva Blomqvist, Michael Cochez +15
cs.AIcs.DBcs.LGarXiv:2003.02320v62020Statistical power for cluster analysis
E. S. Dalmaijer, C. L. Nord, D. E. Astle
stat.MLcs.LGq-bio.QMarXiv:2003.00381v32020Training BatchNorm and Only BatchNorm: On the Expressive Power of Random Features in CNNs
Jonathan Frankle, David J. Schwab, Ari S. Morcos
cs.LGcs.AIcs.NEarXiv:2003.00152v32020Calibrating Deep Neural Networks using Focal Loss
Jishnu Mukhoti, Viveka Kulharia, Amartya Sanyal +3
cs.LGcs.CVstat.MLarXiv:2002.09437v22020Do We Really Need Complicated Model Architectures For Temporal Networks?
Weilin Cong, Si Zhang, Jian Kang +5
cs.LGcs.AIarXiv:2302.11636v12023Do We Really Need to Access the Source Data? Source Hypothesis Transfer for Unsupervised Domain Adaptation
Jian Liang, Dapeng Hu, Jiashi Feng
cs.CVcs.LGarXiv:2002.08546v62020Twitter Sentiment Analysis: Lexicon Method, Machine Learning Method and Their Combination
Olga Kolchyna, Tharsis T. P. Souza, Philip Treleaven +1
cs.CLcs.IRcs.LGarXiv:1507.00955v32015pymoo: Multi-objective Optimization in Python
Julian Blank, Kalyanmoy Deb
cs.NEcs.LGcs.MSarXiv:2002.04504v12020Data-Free Adversarial Distillation
Gongfan Fang, Jie Song, Chengchao Shen +3
cs.LGcs.CVstat.MLarXiv:1912.11006v32019Text Understanding from Scratch
Xiang Zhang, Yann LeCun
cs.LGcs.CLarXiv:1502.01710v52015Orthogonal Gradient Descent for Continual Learning
Mehrdad Farajtabar, Navid Azizan, Alex Mott +1
cs.LGstat.MLarXiv:1910.07104v12019Bayesian Flow Networks for Offline Trajectory Planning
Ludvig Killingberg, Helge Langseth
cs.LGcs.AIarXiv:2608.25163v12026Improving and generalizing flow-based generative models with minibatch optimal transport
Alexander Tong, Kilian Fatras, Nikolay Malkin +5
cs.LGarXiv:2302.00482v42023The Dialect Tax: Dialectal Biases Persist throughout the Language Modeling Pipeline
Elle
cs.CLcs.AIcs.LGarXiv:2608.24952v12026Soft-Label Dataset Distillation and Text Dataset Distillation
Ilia Sucholutsky, Matthias Schonlau
cs.LGcs.AIstat.MLarXiv:1910.02551v32019Human-Timescale Adaptation in an Open-Ended Task Space
Adaptive Agent Team, Jakob Bauer, Kate Baumli +25
cs.LGcs.AIcs.NEarXiv:2301.07608v12023Hamiltonian Generative Networks
Peter Toth, Danilo Jimenez Rezende, Andrew Jaegle +3
cs.LGstat.MLarXiv:1909.13789v22019Source-Free Unsupervised Domain Adaptation: A Survey
Yuqi Fang, Pew-Thian Yap, Weili Lin +2
cs.CVcs.AIcs.LGarXiv:2301.00265v22022Guaranteed Matrix Completion via Non-convex Factorization
Ruoyu Sun, Zhi-Quan Luo
cs.LGarXiv:1411.8003v32014Synthetic Data for Deep Learning
Sergey I. Nikolenko
cs.LGcs.CRcs.CVarXiv:1909.11512v12019Acoustic Scene Classification
Daniele Barchiesi, Dimitrios Giannoulis, Dan Stowell +1
cs.SDcs.LGarXiv:1411.3715v12014DE-FAKE: Detection and Attribution of Fake Images Generated by Text-to-Image Generation Models
Zeyang Sha, Zheng Li, Ning Yu +1
cs.CRcs.CVcs.LGarXiv:2210.06998v22022Addressing the Rare Word Problem in Neural Machine Translation
Minh-Thang Luong, Ilya Sutskever, Quoc V. Le +2
cs.CLcs.LGcs.NEarXiv:1410.8206v42014Deep Equilibrium Models
Shaojie Bai, J. Zico Kolter, Vladlen Koltun
cs.LGstat.MLarXiv:1909.01377v22019GLM-130B: An Open Bilingual Pre-trained Model
Aohan Zeng, Xiao Liu, Zhengxiao Du +15
cs.CLcs.AIcs.LGarXiv:2210.02414v22022Lookahead Optimizer: k steps forward, 1 step back
Michael R. Zhang, James Lucas, Geoffrey Hinton +1
cs.LGcs.NEstat.MLarXiv:1907.08610v22019Hierarchical Skill Retrieval for Data-Efficient Adaptation of Vision-Language-Action Models
Haoran Hao, Shahram Najam Syed, Jeff Schneider +1
cs.ROcs.AIcs.LGarXiv:2608.24042v12026Interpretable Counterfactual Explanations Guided by Prototypes
Arnaud Van Looveren, Janis Klaise
cs.LGstat.MLarXiv:1907.02584v22019