Machine Learning

Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

901 to 960 of 20,192

  1. Self-Play Fine-Tuning Converts Weak Language Models to Strong Language Models

    Zixiang Chen, Yihe Deng, Huizhuo Yuan +2

    cs.LGcs.AIcs.CLarXiv:2401.01335v32024
  2. Synthetic Worlds for Temporal Evaluation and Knowledge Updating in LLMs

    Jonathan Zheng, Zirui Shao, Alan Ritter +1

    cs.CLcs.LGarXiv:2609.00184v12026
  3. Backprop KF: Learning Discriminative Deterministic State Estimators

    Tuomas Haarnoja, Anurag Ajay, Sergey Levine +1

    cs.LGcs.AIarXiv:1605.07148v42016
  4. Different representation learning objectives recover distinct latent structures from the same psychometric data

    Cong Cao, Tassos C. Kyriakides, Pambos Vrasidas

    cs.AIcs.LGstat.MEarXiv:2609.00100v12026
  5. Align Your Gaussians: Text-to-4D with Dynamic 3D Gaussians and Composed Diffusion Models

    Huan Ling, Seung Wook Kim, Antonio Torralba +2

    cs.CVcs.LGarXiv:2312.13763v22023
  6. Learning Dynamics of Logits Debiasing for Long-Tailed Semi-Supervised Learning

    Yue Cheng, Jiajun Zhang, Xiaohui Gao +2

    cs.LGcs.AIarXiv:2608.30699v12026
  7. Steering Llama 2 via Contrastive Activation Addition

    Nina Panickssery, Nick Gabrieli, Julian Schulz +3

    cs.CLcs.AIcs.LGarXiv:2312.06681v42023
  8. Baseline Defenses for Adversarial Attacks Against Aligned Language Models

    Neel Jain, Avi Schwarzschild, Yuxin Wen +7

    cs.LGcs.CLcs.CRarXiv:2309.00614v22023
  9. D4: Improving LLM Pretraining via Document De-Duplication and Diversification

    Kushal Tirumala, Daniel Simig, Armen Aghajanyan +1

    cs.CLcs.AIcs.LGarXiv:2308.12284v12023
  10. PMET: Precise Model Editing in a Transformer

    Xiaopeng Li, Shasha Li, Shezheng Song +3

    cs.CLcs.AIcs.LGarXiv:2308.08742v62023
  11. On the use of deep learning for phase recovery

    Kaiqiang Wang, Li Song, Chutian Wang +8

    physics.opticscs.LGeess.IVarXiv:2308.00942v12023
  12. Aligning Multi-Trajectory Supervision with Policy Optimization for VLA Driving

    Tian Zhang, Zhuo Huang, Hongrui Ye +3

    cs.CVcs.AIcs.LGarXiv:2608.30122v12026
  13. A Survey of Techniques for Optimizing Transformer Inference

    Krishna Teja Chitty-Venkata, Sparsh Mittal, Murali Emani +2

    cs.LGcs.ARcs.CLarXiv:2307.07982v12023
  14. Generating images with recurrent adversarial networks

    Daniel Jiwoong Im, Chris Dongjoo Kim, Hui Jiang +1

    cs.LGcs.CVarXiv:1602.05110v52016
  15. A-MADiff: Attention-Guided Multi-Agent DRL with Diffusion Policies for Memory-Aware Task Orchestration in Mobile AIGC Networks

    Chongzhi Wu, Zhengtao Li, Jiawen Kang +4

    cs.NIcs.LGarXiv:2608.29255v12026
  16. Non-Autoregressive Machine Translation with Latent Alignments

    Chitwan Saharia, William Chan, Saurabh Saxena +1

    cs.CLcs.LGarXiv:2004.07437v32020
  17. Training with Quantization Noise for Extreme Model Compression

    Angela Fan, Pierre Stock, Benjamin Graham +4

    cs.LGstat.MLarXiv:2004.07320v32020
  18. Knowledge Distillation and Student-Teacher Learning for Visual Intelligence: A Review and New Outlooks

    Lin Wang, Kuk-Jin Yoon

    cs.CVcs.AIcs.LGarXiv:2004.05937v72020
  19. Self-Refine: Iterative Refinement with Self-Feedback

    Aman Madaan, Niket Tandon, Prakhar Gupta +13

    cs.CLcs.AIcs.LGarXiv:2303.17651v22023
    Summaries:한국어
  20. Diffusion Schrödinger Bridge Matching

    Yuyang Shi, Valentin De Bortoli, Andrew Campbell +1

    stat.MLcs.LGarXiv:2303.16852v32023
  21. Off-Policy Evaluation for Semantic ID Recommenders: Does the Model's Own Code Hierarchy Help?

    Artem Betlei

    cs.LGarXiv:2608.28905v12026
  22. Foundation Models and Fair Use

    Peter Henderson, Xuechen Li, Dan Jurafsky +3

    cs.CYcs.AIcs.LGarXiv:2303.15715v12023
  23. Neural Programmer-Interpreters

    Scott Reed, Nando de Freitas

    cs.LGcs.NEarXiv:1511.06279v42015
  24. Stabilizing Transformer Training by Preventing Attention Entropy Collapse

    Shuangfei Zhai, Tatiana Likhomanenko, Etai Littwin +5

    cs.LGcs.AIcs.CLarXiv:2303.06296v22023
  25. Comprehensive Review of Deep Reinforcement Learning Methods and Applications in Economics

    Amir Mosavi, Pedram Ghamisi, Yaser Faghan +1

    q-fin.STcs.LGecon.GNarXiv:2004.01509v12020
  26. Deep learning for Stock Market Prediction

    Mojtaba Nabipour, Pooyan Nayyeri, Hamed Jabani +1

    q-fin.STcs.LGarXiv:2004.01497v12020
  27. Timing-Aware Repurchase Prediction for Web-Scale E-Commerce: Survival Models for Multi-Surface Grocery Recommendation

    Akshay Kekuda, Shreeranjani Srirangamsridharan, Ishan Bhatt +5

    cs.AIcs.LGarXiv:2608.28393v12026
  28. Dynamic Multiscale Graph Neural Networks for 3D Skeleton-Based Human Motion Prediction

    Maosen Li, Siheng Chen, Yangheng Zhao +3

    cs.CVcs.LGstat.MLarXiv:2003.08802v12020
  29. COEVOLVE: A Joint Point Process Model for Information Diffusion and Network Co-evolution

    Mehrdad Farajtabar, Yichen Wang, Manuel Gomez Rodriguez +3

    cs.SIcs.LGphysics.soc-pharXiv:1507.02293v22015
  30. Rapid AI Development Cycle for the Coronavirus (COVID-19) Pandemic: Initial Results for Automated Detection & Patient Monitoring using Deep Learning CT Image Analysis

    Ophir Gozes, Maayan Frid-Adar, Hayit Greenspan +5

    eess.IVcs.CVcs.LGarXiv:2003.05037v32020
  31. A Survey on The Expressive Power of Graph Neural Networks

    Ryoma Sato

    cs.LGstat.MLarXiv:2003.04078v42020
  32. 3D Equivariant Diffusion for Target-Aware Molecule Generation and Affinity Prediction

    Jiaqi Guan, Wesley Wei Qian, Xingang Peng +3

    q-bio.BMcs.LGarXiv:2303.03543v12023
  33. Knowledge Graphs

    Aidan Hogan, Eva Blomqvist, Michael Cochez +15

    cs.AIcs.DBcs.LGarXiv:2003.02320v62020
  34. Statistical power for cluster analysis

    E. S. Dalmaijer, C. L. Nord, D. E. Astle

    stat.MLcs.LGq-bio.QMarXiv:2003.00381v32020
  35. Training BatchNorm and Only BatchNorm: On the Expressive Power of Random Features in CNNs

    Jonathan Frankle, David J. Schwab, Ari S. Morcos

    cs.LGcs.AIcs.NEarXiv:2003.00152v32020
  36. Calibrating Deep Neural Networks using Focal Loss

    Jishnu Mukhoti, Viveka Kulharia, Amartya Sanyal +3

    cs.LGcs.CVstat.MLarXiv:2002.09437v22020
  37. Do We Really Need Complicated Model Architectures For Temporal Networks?

    Weilin Cong, Si Zhang, Jian Kang +5

    cs.LGcs.AIarXiv:2302.11636v12023
  38. Do We Really Need to Access the Source Data? Source Hypothesis Transfer for Unsupervised Domain Adaptation

    Jian Liang, Dapeng Hu, Jiashi Feng

    cs.CVcs.LGarXiv:2002.08546v62020
  39. Twitter Sentiment Analysis: Lexicon Method, Machine Learning Method and Their Combination

    Olga Kolchyna, Tharsis T. P. Souza, Philip Treleaven +1

    cs.CLcs.IRcs.LGarXiv:1507.00955v32015
  40. pymoo: Multi-objective Optimization in Python

    Julian Blank, Kalyanmoy Deb

    cs.NEcs.LGcs.MSarXiv:2002.04504v12020
  41. Data-Free Adversarial Distillation

    Gongfan Fang, Jie Song, Chengchao Shen +3

    cs.LGcs.CVstat.MLarXiv:1912.11006v32019
  42. Text Understanding from Scratch

    Xiang Zhang, Yann LeCun

    cs.LGcs.CLarXiv:1502.01710v52015
  43. Orthogonal Gradient Descent for Continual Learning

    Mehrdad Farajtabar, Navid Azizan, Alex Mott +1

    cs.LGstat.MLarXiv:1910.07104v12019
  44. Bayesian Flow Networks for Offline Trajectory Planning

    Ludvig Killingberg, Helge Langseth

    cs.LGcs.AIarXiv:2608.25163v12026
  45. Improving and generalizing flow-based generative models with minibatch optimal transport

    Alexander Tong, Kilian Fatras, Nikolay Malkin +5

    cs.LGarXiv:2302.00482v42023
  46. The Dialect Tax: Dialectal Biases Persist throughout the Language Modeling Pipeline

    Elle

    cs.CLcs.AIcs.LGarXiv:2608.24952v12026
  47. Soft-Label Dataset Distillation and Text Dataset Distillation

    Ilia Sucholutsky, Matthias Schonlau

    cs.LGcs.AIstat.MLarXiv:1910.02551v32019
  48. Human-Timescale Adaptation in an Open-Ended Task Space

    Adaptive Agent Team, Jakob Bauer, Kate Baumli +25

    cs.LGcs.AIcs.NEarXiv:2301.07608v12023
  49. Hamiltonian Generative Networks

    Peter Toth, Danilo Jimenez Rezende, Andrew Jaegle +3

    cs.LGstat.MLarXiv:1909.13789v22019
  50. Source-Free Unsupervised Domain Adaptation: A Survey

    Yuqi Fang, Pew-Thian Yap, Weili Lin +2

    cs.CVcs.AIcs.LGarXiv:2301.00265v22022
  51. Guaranteed Matrix Completion via Non-convex Factorization

    Ruoyu Sun, Zhi-Quan Luo

    cs.LGarXiv:1411.8003v32014
  52. Synthetic Data for Deep Learning

    Sergey I. Nikolenko

    cs.LGcs.CRcs.CVarXiv:1909.11512v12019
  53. Acoustic Scene Classification

    Daniele Barchiesi, Dimitrios Giannoulis, Dan Stowell +1

    cs.SDcs.LGarXiv:1411.3715v12014
  54. DE-FAKE: Detection and Attribution of Fake Images Generated by Text-to-Image Generation Models

    Zeyang Sha, Zheng Li, Ning Yu +1

    cs.CRcs.CVcs.LGarXiv:2210.06998v22022
  55. Addressing the Rare Word Problem in Neural Machine Translation

    Minh-Thang Luong, Ilya Sutskever, Quoc V. Le +2

    cs.CLcs.LGcs.NEarXiv:1410.8206v42014
  56. Deep Equilibrium Models

    Shaojie Bai, J. Zico Kolter, Vladlen Koltun

    cs.LGstat.MLarXiv:1909.01377v22019
  57. GLM-130B: An Open Bilingual Pre-trained Model

    Aohan Zeng, Xiao Liu, Zhengxiao Du +15

    cs.CLcs.AIcs.LGarXiv:2210.02414v22022
  58. Lookahead Optimizer: k steps forward, 1 step back

    Michael R. Zhang, James Lucas, Geoffrey Hinton +1

    cs.LGcs.NEstat.MLarXiv:1907.08610v22019
  59. Hierarchical Skill Retrieval for Data-Efficient Adaptation of Vision-Language-Action Models

    Haoran Hao, Shahram Najam Syed, Jeff Schneider +1

    cs.ROcs.AIcs.LGarXiv:2608.24042v12026
  60. Interpretable Counterfactual Explanations Guided by Prototypes

    Arnaud Van Looveren, Janis Klaise

    cs.LGstat.MLarXiv:1907.02584v22019