Machine Learning

Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

14,101 to 14,160 of 20,454

  1. A Gradient Descent Algorithm on the Grassman Manifold for Matrix Completion

    Raghunandan H. Keshavan, Sewoong Oh

    math.NAcs.LGarXiv:0910.5260v22009
  2. Compositionality decomposed: how do neural networks generalise?

    Dieuwke Hupkes, Verna Dankers, Mathijs Mul +1

    cs.CLcs.AIcs.LGarXiv:1908.08351v22019
  3. Variance Reduction for Faster Non-Convex Optimization

    Zeyuan Allen-Zhu, Elad Hazan

    math.OCcs.DScs.LGarXiv:1603.05643v22016
  4. DoReMi: Optimizing Data Mixtures Speeds Up Language Model Pretraining

    Sang Michael Xie, Hieu Pham, Xuanyi Dong +7

    cs.CLcs.LGarXiv:2305.10429v42023
  5. Dose-PlanNet: Physics Based Radiotherapy Dose Prediction with Deep Learning

    Ankit Bhattacharjee, Sougata Maity, Santam Chakraborty +1

    physics.med-phcs.CVcs.LGarXiv:2608.26901v12026
  6. A Survey on Anomaly Detection for Technical Systems using LSTM Networks

    Benjamin Lindemann, Benjamin Maschler, Nada Sahlab +1

    cs.LGcs.AIstat.MLarXiv:2105.13810v12021
  7. VeriGen: A Large Language Model for Verilog Code Generation

    Shailja Thakur, Baleegh Ahmad, Hammond Pearce +4

    cs.PLcs.LGcs.SEarXiv:2308.00708v12023
  8. How Much Position Information Do Convolutional Neural Networks Encode?

    Md Amirul Islam, Sen Jia, Neil D. B. Bruce

    cs.CVcs.LGarXiv:2001.08248v12020
  9. Interpretable Preferences via Multi-Objective Reward Modeling and Mixture-of-Experts

    Haoxiang Wang, Wei Xiong, Tengyang Xie +2

    cs.LGcs.CLarXiv:2406.12845v12024
  10. GP-VAE: Deep Probabilistic Time Series Imputation

    Vincent Fortuin, Dmitry Baranchuk, Gunnar Rätsch +1

    stat.MLcs.LGarXiv:1907.04155v52019
  11. Melding the Data-Decisions Pipeline: Decision-Focused Learning for Combinatorial Optimization

    Bryan Wilder, Bistra Dilkina, Milind Tambe

    cs.LGcs.AIstat.MLarXiv:1809.05504v22018
  12. A Finite Time Analysis of Temporal Difference Learning With Linear Function Approximation

    Jalaj Bhandari, Daniel Russo, Raghav Singal

    cs.LGstat.MLarXiv:1806.02450v22018
  13. NetGAN: Generating Graphs via Random Walks

    Aleksandar Bojchevski, Oleksandr Shchur, Daniel Zügner +1

    stat.MLcs.LGcs.SIarXiv:1803.00816v22018
  14. Learning Disentangled Representations for Recommendation

    Jianxin Ma, Chang Zhou, Peng Cui +2

    cs.LGcs.IRstat.MLarXiv:1910.14238v12019
  15. Stochastic Trajectory Prediction via Motion Indeterminacy Diffusion

    Tianpei Gu, Guangyi Chen, Junlong Li +4

    cs.CVcs.LGarXiv:2203.13777v12022
  16. Equivalence Between Policy Gradients and Soft Q-Learning

    John Schulman, Xi Chen, Pieter Abbeel

    cs.LGarXiv:1704.06440v42017
  17. Efficient Graph Generation with Graph Recurrent Attention Networks

    Renjie Liao, Yujia Li, Yang Song +6

    cs.LGstat.MLarXiv:1910.00760v32019
  18. Text2Motion: From Natural Language Instructions to Feasible Plans

    Kevin Lin, Christopher Agia, Toki Migimatsu +2

    cs.ROcs.AIcs.LGarXiv:2303.12153v52023
  19. MM-Spectrum: Multimodal Multi-spectral Molecular Structural Elucidation with a Stable MoE Framework

    Hai-tao Yu, Nan Min, Zheng Fang +4

    cs.LGarXiv:2608.27286v12026
  20. Why Deep Neural Networks for Function Approximation?

    Shiyu Liang, R. Srikant

    cs.LGcs.NEarXiv:1610.04161v22016
  21. Vid2Seq: Large-Scale Pretraining of a Visual Language Model for Dense Video Captioning

    Antoine Yang, Arsha Nagrani, Paul Hongsuck Seo +5

    cs.CVcs.AIcs.CLarXiv:2302.14115v22023
  22. Revisiting k-means: New Algorithms via Bayesian Nonparametrics

    Brian Kulis, Michael I. Jordan

    cs.LGstat.MLarXiv:1111.0352v22011
  23. A Variational Inequality Perspective on Generative Adversarial Networks

    Gauthier Gidel, Hugo Berard, Gaëtan Vignoud +2

    cs.LGmath.OCstat.MLarXiv:1802.10551v52018
  24. Emotional Preferences as Goal-Priority Regulation

    Shiqi Liu, Yihua Tan, Hu Fu +1

    cs.LGcs.AIarXiv:2608.27072v12026
  25. Fixup Initialization: Residual Learning Without Normalization

    Hongyi Zhang, Yann N. Dauphin, Tengyu Ma

    cs.LGcs.CVstat.MLarXiv:1901.09321v22019
  26. Playing with Duality: An Overview of Recent Primal-Dual Approaches for Solving Large-Scale Optimization Problems

    Nikos Komodakis, Jean-Christophe Pesquet

    math.NAcs.CVcs.LGarXiv:1406.5429v22014
  27. SimGNN: A Neural Network Approach to Fast Graph Similarity Computation

    Yunsheng Bai, Hao Ding, Song Bian +3

    cs.LGstat.MLarXiv:1808.05689v42018
  28. GRU-ODE-Bayes: Continuous modeling of sporadically-observed time series

    Edward De Brouwer, Jaak Simm, Adam Arany +1

    cs.LGstat.MLarXiv:1905.12374v22019
  29. Scalable and Generalizable Social Bot Detection through Data Selection

    Kai-Cheng Yang, Onur Varol, Pik-Mai Hui +1

    cs.CYcs.LGcs.SIarXiv:1911.09179v12019
  30. The Platonic Representation Hypothesis

    Minyoung Huh, Brian Cheung, Tongzhou Wang +1

    cs.LGcs.AIcs.CVarXiv:2405.07987v52024
  31. TERA: Self-Supervised Learning of Transformer Encoder Representation for Speech

    Andy T. Liu, Shang-Wen Li, Hung-yi Lee

    eess.AScs.CLcs.LGarXiv:2007.06028v32020
  32. Expression, Affect, Action Unit Recognition: Aff-Wild2, Multi-Task Learning and ArcFace

    Dimitrios Kollias, Stefanos Zafeiriou

    cs.CVcs.HCcs.LGarXiv:1910.04855v12019
  33. Dataset Condensation with Differentiable Siamese Augmentation

    Bo Zhao, Hakan Bilen

    cs.LGcs.CVarXiv:2102.08259v22021
  34. Mapping the Landscape of Artificial Intelligence Applications against COVID-19

    Joseph Bullock, Alexandra Luccioni, Katherine Hoffmann Pham +2

    cs.CYcs.AIcs.LGarXiv:2003.11336v32020
  35. Self-Supervised Learning of Graph Neural Networks: A Unified Review

    Yaochen Xie, Zhao Xu, Jingtun Zhang +2

    cs.LGarXiv:2102.10757v52021
  36. Convolutional Networks with Adaptive Inference Graphs

    Andreas Veit, Serge Belongie

    cs.CVcs.LGarXiv:1711.11503v32017
  37. MHFormer: Multi-Hypothesis Transformer for 3D Human Pose Estimation

    Wenhao Li, Hong Liu, Hao Tang +2

    cs.CVcs.AIcs.LGarXiv:2111.12707v42021
  38. 3D Hand Shape and Pose from Images in the Wild

    Adnane Boukhayma, Rodrigo de Bem, Philip H. S. Torr

    cs.CVcs.AIcs.LGarXiv:1902.03451v12019
  39. MixText: Linguistically-Informed Interpolation of Hidden Space for Semi-Supervised Text Classification

    Jiaao Chen, Zichao Yang, Diyi Yang

    cs.CLcs.LGarXiv:2004.12239v12020
  40. Learning to Navigate Unseen Environments: Back Translation with Environmental Dropout

    Hao Tan, Licheng Yu, Mohit Bansal

    cs.CLcs.CVcs.LGarXiv:1904.04195v12019
  41. A Perspective on Deep Imaging

    Ge Wang

    q-bio.QMcs.CVcs.LGarXiv:1609.04375v22016
  42. Image-based 3D Object Reconstruction: State-of-the-Art and Trends in the Deep Learning Era

    Xian-Feng Han, Hamid Laga, Mohammed Bennamoun

    cs.CVcs.CGcs.GRarXiv:1906.06543v32019
  43. Global Optimality of Local Search for Low Rank Matrix Recovery

    Srinadh Bhojanapalli, Behnam Neyshabur, Nathan Srebro

    stat.MLcs.LGmath.OCarXiv:1605.07221v22016
  44. Aequitas: A Bias and Fairness Audit Toolkit

    Pedro Saleiro, Benedict Kuester, Loren Hinkson +5

    cs.LGcs.AIcs.CYarXiv:1811.05577v22018
  45. C2AE: Class Conditioned Auto-Encoder for Open-set Recognition

    Poojan Oza, Vishal M Patel

    cs.CVcs.LGarXiv:1904.01198v12019
  46. Deep Learning for Time Series Forecasting: Tutorial and Literature Survey

    Konstantinos Benidis, Syama Sundar Rangapuram, Valentin Flunkert +10

    cs.LGstat.MLarXiv:2004.10240v22020
  47. Unsupervised Representation Learning for Time Series with Temporal Neighborhood Coding

    Sana Tonekaboni, Danny Eytan, Anna Goldenberg

    cs.LGstat.MLarXiv:2106.00750v12021
  48. Rethinking Graph Neural Networks for Anomaly Detection

    Jianheng Tang, Jiajin Li, Ziqi Gao +1

    cs.LGeess.SParXiv:2205.15508v12022
  49. How AI Experiences Art: Emergent Aesthetic Structure in a Self-Supervised Multimodal Embedding Space

    Corey D. C. Heath

    cs.MMcs.CVcs.LGarXiv:2608.27121v12026
  50. Toward Supervised Anomaly Detection

    Nico Goernitz, Marius Micha Kloft, Konrad Rieck +1

    cs.LGarXiv:1401.6424v12014
  51. A Dual Approach to Scalable Verification of Deep Networks

    Krishnamurthy, Dvijotham, Robert Stanforth +3

    cs.LGstat.MLarXiv:1803.06567v22018
  52. PRECOG: PREdiction Conditioned On Goals in Visual Multi-Agent Settings

    Nicholas Rhinehart, Rowan McAllister, Kris Kitani +1

    cs.CVcs.AIcs.LGarXiv:1905.01296v32019
  53. Theoretical Limitations of Self-Attention in Neural Sequence Models

    Michael Hahn

    cs.CLcs.FLcs.LGarXiv:1906.06755v22019
  54. FastPitch: Parallel Text-to-speech with Pitch Prediction

    Adrian Łańcucki

    eess.AScs.CLcs.LGarXiv:2006.06873v22020
  55. CryptoDL: Deep Neural Networks over Encrypted Data

    Ehsan Hesamifard, Hassan Takabi, Mehdi Ghasemi

    cs.CRcs.LGarXiv:1711.05189v12017
  56. COVID-Twitter-BERT: A Natural Language Processing Model to Analyse COVID-19 Content on Twitter

    Martin Müller, Marcel Salathé, Per E Kummervold

    cs.CLcs.LGcs.SIarXiv:2005.07503v12020
  57. Improving Online Algorithms via ML Predictions

    Ravi Kumar, Manish Purohit, Zoya Svitkina

    cs.DScs.LGarXiv:2407.17712v12024
  58. Two models of double descent for weak features

    Mikhail Belkin, Daniel Hsu, Ji Xu

    cs.LGstat.MLarXiv:1903.07571v22019
  59. Verifiable Reinforcement Learning via Policy Extraction

    Osbert Bastani, Yewen Pu, Armando Solar-Lezama

    cs.LGstat.MLarXiv:1805.08328v22018
  60. Generative Adversarial Perturbations

    Omid Poursaeed, Isay Katsman, Bicheng Gao +1

    cs.CVcs.CRcs.LGarXiv:1712.02328v32017