Machine Learning

Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

13,081 to 13,140 of 20,198

  1. A comprehensive deep learning-based approach to reduced order modeling of nonlinear time-dependent parametrized PDEs

    Stefania Fresca, Luca Dede, Andrea Manzoni

    math.NAcs.LGarXiv:2001.04001v12020
  2. Supermasks in Superposition

    Mitchell Wortsman, Vivek Ramanujan, Rosanne Liu +4

    cs.LGcs.AIstat.MLarXiv:2006.14769v32020
  3. Graph Information Bottleneck

    Tailin Wu, Hongyu Ren, Pan Li +1

    cs.LGstat.MLarXiv:2010.12811v12020
    Summaries:한국어
  4. SE(3) diffusion model with application to protein backbone generation

    Jason Yim, Brian L. Trippe, Valentin De Bortoli +4

    cs.LGq-bio.QMstat.MLarXiv:2302.02277v32023
  5. DiffNet++: A Neural Influence and Interest Diffusion Network for Social Recommendation

    Le Wu, Junwei Li, Peijie Sun +3

    cs.SIcs.IRcs.LGarXiv:2002.00844v42020
  6. BABEL: Bodies, Action and Behavior with English Labels

    Abhinanda R. Punnakkal, Arjun Chandrasekaran, Nikos Athanasiou +2

    cs.CVcs.GRcs.LGarXiv:2106.09696v22021
  7. Fast Image Processing with Fully-Convolutional Networks

    Qifeng Chen, Jia Xu, Vladlen Koltun

    cs.CVcs.GRcs.LGarXiv:1709.00643v12017
  8. OminiControl: Minimal and Universal Control for Diffusion Transformer

    Zhenxiong Tan, Songhua Liu, Xingyi Yang +2

    cs.CVcs.AIcs.LGarXiv:2411.15098v62024
  9. Exploring the Encoding Layer and Loss Function in End-to-End Speaker and Language Recognition System

    Weicheng Cai, Jinkun Chen, Ming Li

    eess.AScs.LGcs.SDarXiv:1804.05160v12018
  10. Does Audio Deepfake Detection Generalize?

    Nicolas M. Müller, Pavel Czempin, Franziska Dieckmann +2

    cs.SDcs.LGeess.ASarXiv:2203.16263v52022
  11. Latent ODEs for Irregularly-Sampled Time Series

    Yulia Rubanova, Ricky T. Q. Chen, David Duvenaud

    cs.LGstat.MLarXiv:1907.03907v12019
  12. Understanding Dataset Difficulty with $\mathcal{V}$-Usable Information

    Kawin Ethayarajh, Yejin Choi, Swabha Swayamdipta

    cs.CLcs.AIcs.LGarXiv:2110.08420v32021
  13. TVM: An Automated End-to-End Optimizing Compiler for Deep Learning

    Tianqi Chen, Thierry Moreau, Ziheng Jiang +9

    cs.LGcs.AIcs.PLarXiv:1802.04799v32018
  14. Distribution Matching Losses Can Hallucinate Features in Medical Image Translation

    Joseph Paul Cohen, Margaux Luck, Sina Honari

    cs.CVcs.LGarXiv:1805.08841v32018
  15. Better Mixing via Deep Representations

    Yoshua Bengio, Grégoire Mesnil, Yann Dauphin +1

    cs.LGarXiv:1207.4404v12012
  16. Ghost in the Minecraft: Generally Capable Agents for Open-World Environments via Large Language Models with Text-based Knowledge and Memory

    Xizhou Zhu, Yuntao Chen, Hao Tian +10

    cs.AIcs.CLcs.CVarXiv:2305.17144v22023
  17. Adversarial Attacks and Defences Competition

    Alexey Kurakin, Ian Goodfellow, Samy Bengio +20

    cs.CVcs.CRcs.LGarXiv:1804.00097v12018
  18. A Bayesian Approach to Discovering Truth from Conflicting Sources for Data Integration

    Bo Zhao, Benjamin I. P. Rubinstein, Jim Gemmell +1

    cs.DBcs.LGarXiv:1203.0058v12012
  19. CoMatch: Semi-supervised Learning with Contrastive Graph Regularization

    Junnan Li, Caiming Xiong, Steven Hoi

    cs.LGcs.CVarXiv:2011.11183v22020
  20. Benchmarking and Survey of Explanation Methods for Black Box Models

    Francesco Bodria, Fosca Giannotti, Riccardo Guidotti +3

    cs.AIcs.CYcs.LGarXiv:2102.13076v12021
  21. Deep learning for molecular design - a review of the state of the art

    Daniel C. Elton, Zois Boukouvalas, Mark D. Fuge +1

    cs.LGphysics.chem-phstat.MLarXiv:1903.04388v32019
  22. Rapid Convergence of the Unadjusted Langevin Algorithm: Isoperimetry Suffices

    Santosh S. Vempala, Andre Wibisono

    cs.DScs.LGmath.PRarXiv:1903.08568v42019
  23. Mining Cross-Image Semantics for Weakly Supervised Semantic Segmentation

    Guolei Sun, Wenguan Wang, Jifeng Dai +1

    cs.CVcs.LGeess.IVarXiv:2007.01947v22020
  24. Learning to Remember: A Synaptic Plasticity Driven Framework for Continual Learning

    Oleksiy Ostapenko, Mihai Puscas, Tassilo Klein +2

    cs.NEcs.CVcs.LGarXiv:1904.03137v42019
  25. Differentiation of Blackbox Combinatorial Solvers

    Marin Vlastelica, Anselm Paulus, Vít Musil +2

    cs.LGstat.MLarXiv:1912.02175v22019
  26. FedVision: An Online Visual Object Detection Platform Powered by Federated Learning

    Yang Liu, Anbu Huang, Yun Luo +7

    cs.LGcs.CVstat.MLarXiv:2001.06202v12020
  27. SELF: Learning to Filter Noisy Labels with Self-Ensembling

    Duc Tam Nguyen, Chaithanya Kumar Mummadi, Thi Phuong Nhung Ngo +3

    cs.CVcs.LGstat.MLarXiv:1910.01842v12019
  28. Federated Learning: Opportunities and Challenges

    Priyanka Mary Mammen

    cs.LGcs.DCarXiv:2101.05428v12021
  29. Deep Learning Methods for Parallel Magnetic Resonance Image Reconstruction

    Florian Knoll, Kerstin Hammernik, Chi Zhang +4

    eess.SPcs.CVcs.LGarXiv:1904.01112v12019
  30. ChatGPT is not all you need. A State of the Art Review of large Generative AI models

    Roberto Gozalo-Brizuela, Eduardo C. Garrido-Merchan

    cs.LGcs.AIarXiv:2301.04655v12023
  31. Backdoor Embedding in Convolutional Neural Network Models via Invisible Perturbation

    Cong Liao, Haoti Zhong, Anna Squicciarini +2

    cs.CRcs.LGstat.MLarXiv:1808.10307v12018
  32. Learning with Symmetric Label Noise: The Importance of Being Unhinged

    Brendan van Rooyen, Aditya Krishna Menon, Robert C. Williamson

    cs.LGarXiv:1505.07634v12015
  33. Commonsense Knowledge Mining from Pretrained Models

    Joshua Feldman, Joe Davison, Alexander M. Rush

    cs.CLcs.AIcs.LGarXiv:1909.00505v12019
  34. Transformers are Sample-Efficient World Models

    Vincent Micheli, Eloi Alonso, François Fleuret

    cs.LGcs.AIcs.CVarXiv:2209.00588v22022
  35. Constructing Unrestricted Adversarial Examples with Generative Models

    Yang Song, Rui Shu, Nate Kushman +1

    cs.LGcs.AIcs.CRarXiv:1805.07894v42018
  36. Improving Topic Models with Latent Feature Word Representations

    Dat Quoc Nguyen, Richard Billingsley, Lan Du +1

    cs.CLcs.IRcs.LGarXiv:1810.06306v12018
  37. Alzheimer's Dementia Recognition through Spontaneous Speech: The ADReSS Challenge

    Saturnino Luz, Fasih Haider, Sofia de la Fuente +2

    eess.AScs.LGstat.MLarXiv:2004.06833v32020
  38. Alphazero-like Tree-Search can Guide Large Language Model Decoding and Training

    Xidong Feng, Ziyu Wan, Muning Wen +4

    cs.LGcs.AIcs.CLarXiv:2309.17179v22023
  39. Cooperative Perception for 3D Object Detection in Driving Scenarios using Infrastructure Sensors

    Eduardo Arnold, Mehrdad Dianati, Robert de Temple +1

    cs.CVcs.LGcs.MAarXiv:1912.12147v22019
  40. Joint Device Scheduling and Resource Allocation for Latency Constrained Wireless Federated Learning

    Wenqi Shi, Sheng Zhou, Zhisheng Niu +2

    cs.ITcs.LGcs.NIarXiv:2007.07174v12020
  41. Very high resolution canopy height maps from RGB imagery using self-supervised vision transformer and convolutional decoder trained on Aerial Lidar

    Jamie Tolan, Hung-I Yang, Ben Nosarzewski +13

    cs.CVcs.LGarXiv:2304.07213v32023
  42. Neural Autoregressive Distribution Estimation

    Benigno Uria, Marc-Alexandre Côté, Karol Gregor +2

    cs.LGarXiv:1605.02226v32016
  43. Combating Fake News: A Survey on Identification and Mitigation Techniques

    Karishma Sharma, Feng Qian, He Jiang +3

    cs.LGcs.AIcs.SIarXiv:1901.06437v12019
  44. Relation Classification via Recurrent Neural Network

    Dongxu Zhang, Dong Wang

    cs.CLcs.LGcs.NEarXiv:1508.01006v22015
  45. A Tail-Index Analysis of Stochastic Gradient Noise in Deep Neural Networks

    Umut Simsekli, Levent Sagun, Mert Gurbuzbalaban

    cs.LGstat.MLarXiv:1901.06053v12019
  46. Syntax-Directed Variational Autoencoder for Structured Data

    Hanjun Dai, Yingtao Tian, Bo Dai +2

    cs.LGcs.CLarXiv:1802.08786v12018
  47. Parameter Efficient Training of Deep Convolutional Neural Networks by Dynamic Sparse Reparameterization

    Hesham Mostafa, Xin Wang

    cs.LGstat.MLarXiv:1902.05967v32019
  48. Atom: Low-bit Quantization for Efficient and Accurate LLM Serving

    Yilong Zhao, Chien-Yu Lin, Kan Zhu +7

    cs.LGarXiv:2310.19102v32023
  49. Flipout: Efficient Pseudo-Independent Weight Perturbations on Mini-Batches

    Yeming Wen, Paul Vicol, Jimmy Ba +2

    cs.LGstat.MLarXiv:1803.04386v22018
  50. Learning with Augmented Features for Heterogeneous Domain Adaptation

    Lixin Duan, Dong Xu, Ivor Tsang

    cs.LGarXiv:1206.4660v12012
  51. Predicting materials properties without crystal structure: Deep representation learning from stoichiometry

    Rhys E. A. Goodall, Alpha A. Lee

    physics.comp-phcond-mat.mtrl-scics.LGarXiv:1910.00617v42019
  52. BAGAN: Data Augmentation with Balancing GAN

    Giovanni Mariani, Florian Scheidegger, Roxana Istrate +2

    cs.CVcs.LGstat.MLarXiv:1803.09655v22018
  53. Asynchronous Stochastic Gradient Descent with Delay Compensation

    Shuxin Zheng, Qi Meng, Taifeng Wang +4

    cs.LGcs.DCarXiv:1609.08326v62016
  54. Search on the Replay Buffer: Bridging Planning and Reinforcement Learning

    Benjamin Eysenbach, Ruslan Salakhutdinov, Sergey Levine

    cs.AIcs.LGcs.ROarXiv:1906.05253v12019
  55. MLPerf Tiny Benchmark

    Colby Banbury, Vijay Janapa Reddi, Peter Torelli +19

    cs.LGcs.ARarXiv:2106.07597v42021
  56. Offline Reinforcement Learning with Fisher Divergence Critic Regularization

    Ilya Kostrikov, Jonathan Tompson, Rob Fergus +1

    cs.LGarXiv:2103.08050v12021
  57. A mathematical theory of semantic development in deep neural networks

    Andrew M. Saxe, James L. McClelland, Surya Ganguli

    cs.LGcs.AIq-bio.NCarXiv:1810.10531v12018
  58. Self-attention for raw optical Satellite Time Series Classification

    Marc Rußwurm, Marco Körner

    cs.LGeess.IVstat.MLarXiv:1910.10536v32019
  59. i-RevNet: Deep Invertible Networks

    Jörn-Henrik Jacobsen, Arnold Smeulders, Edouard Oyallon

    cs.LGcs.CVstat.MLarXiv:1802.07088v12018
  60. Hand Pose Estimation via Latent 2.5D Heatmap Regression

    Umar Iqbal, Pavlo Molchanov, Thomas Breuel +2

    cs.CVcs.LGarXiv:1804.09534v12018