Machine Learning

Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

13,741 to 13,800 of 20,199

  1. Privacy Without Regret: Differentially Private Inference-Time Alignment

    Ishi Jain, Nandini Bhattad, Sayak Ray Chowdhury

    cs.LGarXiv:2608.26324v12026
  2. A Cookbook of Self-Supervised Learning

    Randall Balestriero, Mark Ibrahim, Vlad Sobal +16

    cs.LGcs.CVarXiv:2304.12210v22023
  3. Explain Images with Multimodal Recurrent Neural Networks

    Junhua Mao, Wei Xu, Yi Yang +2

    cs.CVcs.CLcs.LGarXiv:1410.1090v12014
  4. EPOpt: Learning Robust Neural Network Policies Using Model Ensembles

    Aravind Rajeswaran, Sarvjeet Ghotra, Balaraman Ravindran +1

    cs.LGcs.AIcs.ROarXiv:1610.01283v42016
  5. Understanding Attention and Generalization in Graph Neural Networks

    Boris Knyazev, Graham W. Taylor, Mohamed R. Amer

    cs.LGcs.AIstat.MLarXiv:1905.02850v32019
  6. TensorFlow Distributions

    Joshua V. Dillon, Ian Langmore, Dustin Tran +7

    cs.LGcs.AIcs.PLarXiv:1711.10604v12017
  7. GNNGuard: Defending Graph Neural Networks against Adversarial Attacks

    Xiang Zhang, Marinka Zitnik

    cs.LGstat.MLarXiv:2006.08149v32020
  8. Whitening for Self-Supervised Representation Learning

    Aleksandr Ermolov, Aliaksandr Siarohin, Enver Sangineto +1

    cs.LGcs.CVstat.MLarXiv:2007.06346v52020
  9. Improving zero-shot learning by mitigating the hubness problem

    Georgiana Dinu, Angeliki Lazaridou, Marco Baroni

    cs.CLcs.LGarXiv:1412.6568v32014
  10. Very Deep VAEs Generalize Autoregressive Models and Can Outperform Them on Images

    Rewon Child

    cs.LGcs.CVarXiv:2011.10650v22020
  11. Levenshtein Transformer

    Jiatao Gu, Changhan Wang, Jake Zhao

    cs.CLcs.LGarXiv:1905.11006v22019
  12. Global optimization of dielectric metasurfaces using a physics-driven neural network

    Jiaqi Jiang, Jonathan A. Fan

    cs.LGphysics.comp-phphysics.opticsarXiv:1906.04157v22019
  13. Federated Multi-Task Learning under a Mixture of Distributions

    Othmane Marfoq, Giovanni Neglia, Aurélien Bellet +2

    cs.LGcs.AImath.OCarXiv:2108.10252v42021
  14. Shared Actors Need Not Share Critics: Effects of Value Mismatch in Parallel Reinforcement Learning

    Zhenya Liu, Yang Meng, Zhuokai Zhao +2

    cs.LGarXiv:2608.26481v12026
  15. Multiaccuracy: Black-Box Post-Processing for Fairness in Classification

    Michael P. Kim, Amirata Ghorbani, James Zou

    cs.LGstat.MLarXiv:1805.12317v22018
  16. Perception Prioritized Training of Diffusion Models

    Jooyoung Choi, Jungbeom Lee, Chaehun Shin +3

    cs.CVcs.LGarXiv:2204.00227v12022
  17. Sparse Sinkhorn Attention

    Yi Tay, Dara Bahri, Liu Yang +2

    cs.LGcs.CLarXiv:2002.11296v12020
  18. AI and Memory Wall

    Amir Gholami, Zhewei Yao, Sehoon Kim +3

    cs.LGcs.ARcs.DCarXiv:2403.14123v12024
  19. Diffusion Self-Guidance for Controllable Image Generation

    Dave Epstein, Allan Jabri, Ben Poole +2

    cs.CVcs.LGstat.MLarXiv:2306.00986v32023
  20. A Statistical Perspective on Algorithmic Leveraging

    Ping Ma, Michael W. Mahoney, Bin Yu

    stat.MEcs.LGstat.MLarXiv:1306.5362v12013
  21. Fast Patch-based Style Transfer of Arbitrary Style

    Tian Qi Chen, Mark Schmidt

    cs.CVcs.GRcs.LGarXiv:1612.04337v12016
  22. Randomized Ensembled Double Q-Learning: Learning Fast Without a Model

    Xinyue Chen, Che Wang, Zijian Zhou +1

    cs.LGcs.AIarXiv:2101.05982v22021
  23. Speech Enhancement and Dereverberation with Diffusion-based Generative Models

    Julius Richter, Simon Welker, Jean-Marie Lemercier +2

    eess.AScs.LGcs.SDarXiv:2208.05830v32022
  24. A Review of Deep Learning with Special Emphasis on Architectures, Applications and Recent Trends

    Saptarshi Sengupta, Sanchita Basak, Pallabi Saikia +5

    cs.LGstat.MLarXiv:1905.13294v32019
  25. Data Augmentation using Random Image Cropping and Patching for Deep CNNs

    Ryo Takahashi, Takashi Matsubara, Kuniaki Uehara

    cs.CVcs.LGarXiv:1811.09030v22018
  26. A Real-World WebAgent with Planning, Long Context Understanding, and Program Synthesis

    Izzeddin Gur, Hiroki Furuta, Austin Huang +4

    cs.LGcs.AIcs.CLarXiv:2307.12856v42023
  27. NaturalSpeech 2: Latent Diffusion Models are Natural and Zero-Shot Speech and Singing Synthesizers

    Kai Shen, Zeqian Ju, Xu Tan +6

    eess.AScs.AIcs.CLarXiv:2304.09116v32023
  28. Transfer Learning for EEG-Based Brain-Computer Interfaces: A Review of Progress Made Since 2016

    Dongrui Wu, Yifan Xu, Bao-Liang Lu

    cs.HCcs.LGeess.SParXiv:2004.06286v42020
  29. Time-lagged autoencoders: Deep learning of slow collective variables for molecular kinetics

    Christoph Wehmeyer, Frank Noé

    stat.MLcs.LGphysics.bio-pharXiv:1710.11239v12017
  30. Prompt Sensitivity of Generative Agents: Evidence from an Epidemic Model

    Ross Williams, Niyousha Hosseinichimeh

    physics.soc-phcs.AIcs.LGarXiv:2608.26221v12026
  31. Universal Source-Free Domain Adaptation

    Jogendra Nath Kundu, Naveen Venkat, Rahul M +1

    cs.CVcs.LGarXiv:2004.04393v12020
  32. Estimating Uncertainty and Interpretability in Deep Learning for Coronavirus (COVID-19) Detection

    Biraja Ghoshal, Allan Tucker

    eess.IVcs.CVcs.LGarXiv:2003.10769v22020
  33. Federated Learning over Wireless Networks: Convergence Analysis and Resource Allocation

    Canh T. Dinh, Nguyen H. Tran, Minh N. H. Nguyen +4

    cs.LGcs.DCcs.NIarXiv:1910.13067v42019
  34. Tying Word Vectors and Word Classifiers: A Loss Framework for Language Modeling

    Hakan Inan, Khashayar Khosravi, Richard Socher

    cs.LGcs.CLstat.MLarXiv:1611.01462v32016
  35. Malicious URL Detection using Machine Learning: A Survey

    Doyen Sahoo, Chenghao Liu, Steven C. H. Hoi

    cs.LGcs.CRarXiv:1701.07179v32017
  36. Blockwise Parallel Decoding for Deep Autoregressive Models

    Mitchell Stern, Noam Shazeer, Jakob Uszkoreit

    cs.LGcs.CLstat.MLarXiv:1811.03115v12018
  37. Analogical Inference for Multi-Relational Embeddings

    Hanxiao Liu, Yuexin Wu, Yiming Yang

    cs.LGcs.AIcs.CLarXiv:1705.02426v22017
  38. Gemma Scope: Open Sparse Autoencoders Everywhere All At Once on Gemma 2

    Tom Lieberum, Senthooran Rajamanoharan, Arthur Conmy +7

    cs.LGcs.AIcs.CLarXiv:2408.05147v22024
  39. Provable Guarantees for Self-Supervised Deep Learning with Spectral Contrastive Loss

    Jeff Z. HaoChen, Colin Wei, Adrien Gaidon +1

    cs.LGstat.MLarXiv:2106.04156v72021
  40. D'ya like DAGs? A Survey on Structure Learning and Causal Discovery

    Matthew J. Vowels, Necati Cihan Camgoz, Richard Bowden

    cs.LGstat.MEstat.MLarXiv:2103.02582v22021
  41. Towards Understanding Knowledge Distillation

    Mary Phuong, Christoph H. Lampert

    cs.LGstat.MLarXiv:2105.13093v12021
  42. A Survey on Recent Approaches for Natural Language Processing in Low-Resource Scenarios

    Michael A. Hedderich, Lukas Lange, Heike Adel +2

    cs.CLcs.LGarXiv:2010.12309v32020
  43. Explanations can be manipulated and geometry is to blame

    Ann-Kathrin Dombrowski, Maximilian Alber, Christopher J. Anders +3

    stat.MLcs.CRcs.LGarXiv:1906.07983v22019
  44. A Unified Framework for Fair and Personalized Decentralized Learning under Communication Constraints

    Krishnendu S. Tharakan, Carlo Fischione

    cs.LGarXiv:2608.26493v12026
  45. Recurrence is required to capture the representational dynamics of the human visual system

    Tim C Kietzmann, Courtney J Spoerer, Lynn Sörensen +3

    q-bio.NCcs.CVcs.LGarXiv:1903.05946v22019
  46. When Interference Graphs Evolve: Doubly Robust Estimation of Dynamic Peer Effects

    Xiaojing Du

    cs.LGcs.SIarXiv:2608.27187v12026
  47. Dynamical phase selection controls compute scaling in looped transformers

    Gunn Kim

    cond-mat.dis-nncond-mat.stat-mechcs.LGarXiv:2608.26556v12026
  48. Spectral Representations for Convolutional Neural Networks

    Oren Rippel, Jasper Snoek, Ryan P. Adams

    stat.MLcs.LGarXiv:1506.03767v12015
  49. Neural Renormalization Group Flow for Percolation

    Anaclara Alvez, Luca Camagna, Sergio Chibbaro +4

    cond-mat.dis-nncs.LGarXiv:2608.26764v12026
  50. Speech Model Pre-training for End-to-End Spoken Language Understanding

    Loren Lugosch, Mirco Ravanelli, Patrick Ignoto +2

    eess.AScs.CLcs.LGarXiv:1904.03670v22019
  51. Chart2SVG: Editable SVG Generation from Raster Chart Images

    Jinning Cui, Lu Chen, Haoyan Shi +5

    cs.LGarXiv:2608.26544v12026
  52. Fully-adaptive Feature Sharing in Multi-Task Networks with Applications in Person Attribute Classification

    Yongxi Lu, Abhishek Kumar, Shuangfei Zhai +3

    cs.CVcs.LGarXiv:1611.05377v12016
  53. Neural Episodic Control

    Alexander Pritzel, Benigno Uria, Sriram Srinivasan +5

    cs.LGstat.MLarXiv:1703.01988v12017
  54. Graph Neural Networks for Scalable Radio Resource Management: Architecture Design and Theoretical Analysis

    Yifei Shen, Yuanming Shi, Jun Zhang +1

    cs.ITcs.LGeess.SParXiv:2007.07632v22020
  55. Knowledge distillation: A good teacher is patient and consistent

    Lucas Beyer, Xiaohua Zhai, Amélie Royer +3

    cs.CVcs.AIcs.LGarXiv:2106.05237v22021
  56. Unsupervised Learning of 3D Structure from Images

    Danilo Jimenez Rezende, S. M. Ali Eslami, Shakir Mohamed +3

    cs.CVcs.LGstat.MLarXiv:1607.00662v22016
  57. Measuring abstract reasoning in neural networks

    David G. T. Barrett, Felix Hill, Adam Santoro +2

    cs.LGstat.MLarXiv:1807.04225v12018
  58. Transformer Hawkes Process

    Simiao Zuo, Haoming Jiang, Zichong Li +2

    cs.LGstat.MLarXiv:2002.09291v52020
  59. The Debate Over Understanding in AI's Large Language Models

    Melanie Mitchell, David C. Krakauer

    cs.LGcs.AIarXiv:2210.13966v32022
  60. Molecular generative model based on conditional variational autoencoder for de novo molecular design

    Jaechang Lim, Seongok Ryu, Jin Woo Kim +1

    cs.LGstat.MLarXiv:1806.05805v12018