Machine Learning
Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
13,741 to 13,800 of 20,199
Privacy Without Regret: Differentially Private Inference-Time Alignment
Ishi Jain, Nandini Bhattad, Sayak Ray Chowdhury
cs.LGarXiv:2608.26324v12026A Cookbook of Self-Supervised Learning
Randall Balestriero, Mark Ibrahim, Vlad Sobal +16
cs.LGcs.CVarXiv:2304.12210v22023Explain Images with Multimodal Recurrent Neural Networks
Junhua Mao, Wei Xu, Yi Yang +2
cs.CVcs.CLcs.LGarXiv:1410.1090v12014EPOpt: Learning Robust Neural Network Policies Using Model Ensembles
Aravind Rajeswaran, Sarvjeet Ghotra, Balaraman Ravindran +1
cs.LGcs.AIcs.ROarXiv:1610.01283v42016Understanding Attention and Generalization in Graph Neural Networks
Boris Knyazev, Graham W. Taylor, Mohamed R. Amer
cs.LGcs.AIstat.MLarXiv:1905.02850v32019TensorFlow Distributions
Joshua V. Dillon, Ian Langmore, Dustin Tran +7
cs.LGcs.AIcs.PLarXiv:1711.10604v12017GNNGuard: Defending Graph Neural Networks against Adversarial Attacks
Xiang Zhang, Marinka Zitnik
cs.LGstat.MLarXiv:2006.08149v32020Whitening for Self-Supervised Representation Learning
Aleksandr Ermolov, Aliaksandr Siarohin, Enver Sangineto +1
cs.LGcs.CVstat.MLarXiv:2007.06346v52020Improving zero-shot learning by mitigating the hubness problem
Georgiana Dinu, Angeliki Lazaridou, Marco Baroni
cs.CLcs.LGarXiv:1412.6568v32014Very Deep VAEs Generalize Autoregressive Models and Can Outperform Them on Images
Rewon Child
cs.LGcs.CVarXiv:2011.10650v22020Levenshtein Transformer
Jiatao Gu, Changhan Wang, Jake Zhao
cs.CLcs.LGarXiv:1905.11006v22019Global optimization of dielectric metasurfaces using a physics-driven neural network
Jiaqi Jiang, Jonathan A. Fan
cs.LGphysics.comp-phphysics.opticsarXiv:1906.04157v22019Federated Multi-Task Learning under a Mixture of Distributions
Othmane Marfoq, Giovanni Neglia, Aurélien Bellet +2
cs.LGcs.AImath.OCarXiv:2108.10252v42021Shared Actors Need Not Share Critics: Effects of Value Mismatch in Parallel Reinforcement Learning
Zhenya Liu, Yang Meng, Zhuokai Zhao +2
cs.LGarXiv:2608.26481v12026Multiaccuracy: Black-Box Post-Processing for Fairness in Classification
Michael P. Kim, Amirata Ghorbani, James Zou
cs.LGstat.MLarXiv:1805.12317v22018Perception Prioritized Training of Diffusion Models
Jooyoung Choi, Jungbeom Lee, Chaehun Shin +3
cs.CVcs.LGarXiv:2204.00227v12022Sparse Sinkhorn Attention
Yi Tay, Dara Bahri, Liu Yang +2
cs.LGcs.CLarXiv:2002.11296v12020AI and Memory Wall
Amir Gholami, Zhewei Yao, Sehoon Kim +3
cs.LGcs.ARcs.DCarXiv:2403.14123v12024Diffusion Self-Guidance for Controllable Image Generation
Dave Epstein, Allan Jabri, Ben Poole +2
cs.CVcs.LGstat.MLarXiv:2306.00986v32023A Statistical Perspective on Algorithmic Leveraging
Ping Ma, Michael W. Mahoney, Bin Yu
stat.MEcs.LGstat.MLarXiv:1306.5362v12013Fast Patch-based Style Transfer of Arbitrary Style
Tian Qi Chen, Mark Schmidt
cs.CVcs.GRcs.LGarXiv:1612.04337v12016Randomized Ensembled Double Q-Learning: Learning Fast Without a Model
Xinyue Chen, Che Wang, Zijian Zhou +1
cs.LGcs.AIarXiv:2101.05982v22021Speech Enhancement and Dereverberation with Diffusion-based Generative Models
Julius Richter, Simon Welker, Jean-Marie Lemercier +2
eess.AScs.LGcs.SDarXiv:2208.05830v32022A Review of Deep Learning with Special Emphasis on Architectures, Applications and Recent Trends
Saptarshi Sengupta, Sanchita Basak, Pallabi Saikia +5
cs.LGstat.MLarXiv:1905.13294v32019Data Augmentation using Random Image Cropping and Patching for Deep CNNs
Ryo Takahashi, Takashi Matsubara, Kuniaki Uehara
cs.CVcs.LGarXiv:1811.09030v22018A Real-World WebAgent with Planning, Long Context Understanding, and Program Synthesis
Izzeddin Gur, Hiroki Furuta, Austin Huang +4
cs.LGcs.AIcs.CLarXiv:2307.12856v42023NaturalSpeech 2: Latent Diffusion Models are Natural and Zero-Shot Speech and Singing Synthesizers
Kai Shen, Zeqian Ju, Xu Tan +6
eess.AScs.AIcs.CLarXiv:2304.09116v32023Transfer Learning for EEG-Based Brain-Computer Interfaces: A Review of Progress Made Since 2016
Dongrui Wu, Yifan Xu, Bao-Liang Lu
cs.HCcs.LGeess.SParXiv:2004.06286v42020Time-lagged autoencoders: Deep learning of slow collective variables for molecular kinetics
Christoph Wehmeyer, Frank Noé
stat.MLcs.LGphysics.bio-pharXiv:1710.11239v12017Prompt Sensitivity of Generative Agents: Evidence from an Epidemic Model
Ross Williams, Niyousha Hosseinichimeh
physics.soc-phcs.AIcs.LGarXiv:2608.26221v12026Universal Source-Free Domain Adaptation
Jogendra Nath Kundu, Naveen Venkat, Rahul M +1
cs.CVcs.LGarXiv:2004.04393v12020Estimating Uncertainty and Interpretability in Deep Learning for Coronavirus (COVID-19) Detection
Biraja Ghoshal, Allan Tucker
eess.IVcs.CVcs.LGarXiv:2003.10769v22020Federated Learning over Wireless Networks: Convergence Analysis and Resource Allocation
Canh T. Dinh, Nguyen H. Tran, Minh N. H. Nguyen +4
cs.LGcs.DCcs.NIarXiv:1910.13067v42019Tying Word Vectors and Word Classifiers: A Loss Framework for Language Modeling
Hakan Inan, Khashayar Khosravi, Richard Socher
cs.LGcs.CLstat.MLarXiv:1611.01462v32016Malicious URL Detection using Machine Learning: A Survey
Doyen Sahoo, Chenghao Liu, Steven C. H. Hoi
cs.LGcs.CRarXiv:1701.07179v32017Blockwise Parallel Decoding for Deep Autoregressive Models
Mitchell Stern, Noam Shazeer, Jakob Uszkoreit
cs.LGcs.CLstat.MLarXiv:1811.03115v12018Analogical Inference for Multi-Relational Embeddings
Hanxiao Liu, Yuexin Wu, Yiming Yang
cs.LGcs.AIcs.CLarXiv:1705.02426v22017Gemma Scope: Open Sparse Autoencoders Everywhere All At Once on Gemma 2
Tom Lieberum, Senthooran Rajamanoharan, Arthur Conmy +7
cs.LGcs.AIcs.CLarXiv:2408.05147v22024Provable Guarantees for Self-Supervised Deep Learning with Spectral Contrastive Loss
Jeff Z. HaoChen, Colin Wei, Adrien Gaidon +1
cs.LGstat.MLarXiv:2106.04156v72021D'ya like DAGs? A Survey on Structure Learning and Causal Discovery
Matthew J. Vowels, Necati Cihan Camgoz, Richard Bowden
cs.LGstat.MEstat.MLarXiv:2103.02582v22021Towards Understanding Knowledge Distillation
Mary Phuong, Christoph H. Lampert
cs.LGstat.MLarXiv:2105.13093v12021A Survey on Recent Approaches for Natural Language Processing in Low-Resource Scenarios
Michael A. Hedderich, Lukas Lange, Heike Adel +2
cs.CLcs.LGarXiv:2010.12309v32020Explanations can be manipulated and geometry is to blame
Ann-Kathrin Dombrowski, Maximilian Alber, Christopher J. Anders +3
stat.MLcs.CRcs.LGarXiv:1906.07983v22019A Unified Framework for Fair and Personalized Decentralized Learning under Communication Constraints
Krishnendu S. Tharakan, Carlo Fischione
cs.LGarXiv:2608.26493v12026Recurrence is required to capture the representational dynamics of the human visual system
Tim C Kietzmann, Courtney J Spoerer, Lynn Sörensen +3
q-bio.NCcs.CVcs.LGarXiv:1903.05946v22019When Interference Graphs Evolve: Doubly Robust Estimation of Dynamic Peer Effects
Xiaojing Du
cs.LGcs.SIarXiv:2608.27187v12026Dynamical phase selection controls compute scaling in looped transformers
Gunn Kim
cond-mat.dis-nncond-mat.stat-mechcs.LGarXiv:2608.26556v12026Spectral Representations for Convolutional Neural Networks
Oren Rippel, Jasper Snoek, Ryan P. Adams
stat.MLcs.LGarXiv:1506.03767v12015Neural Renormalization Group Flow for Percolation
Anaclara Alvez, Luca Camagna, Sergio Chibbaro +4
cond-mat.dis-nncs.LGarXiv:2608.26764v12026Speech Model Pre-training for End-to-End Spoken Language Understanding
Loren Lugosch, Mirco Ravanelli, Patrick Ignoto +2
eess.AScs.CLcs.LGarXiv:1904.03670v22019Chart2SVG: Editable SVG Generation from Raster Chart Images
Jinning Cui, Lu Chen, Haoyan Shi +5
cs.LGarXiv:2608.26544v12026Fully-adaptive Feature Sharing in Multi-Task Networks with Applications in Person Attribute Classification
Yongxi Lu, Abhishek Kumar, Shuangfei Zhai +3
cs.CVcs.LGarXiv:1611.05377v12016Neural Episodic Control
Alexander Pritzel, Benigno Uria, Sriram Srinivasan +5
cs.LGstat.MLarXiv:1703.01988v12017Graph Neural Networks for Scalable Radio Resource Management: Architecture Design and Theoretical Analysis
Yifei Shen, Yuanming Shi, Jun Zhang +1
cs.ITcs.LGeess.SParXiv:2007.07632v22020Knowledge distillation: A good teacher is patient and consistent
Lucas Beyer, Xiaohua Zhai, Amélie Royer +3
cs.CVcs.AIcs.LGarXiv:2106.05237v22021Unsupervised Learning of 3D Structure from Images
Danilo Jimenez Rezende, S. M. Ali Eslami, Shakir Mohamed +3
cs.CVcs.LGstat.MLarXiv:1607.00662v22016Measuring abstract reasoning in neural networks
David G. T. Barrett, Felix Hill, Adam Santoro +2
cs.LGstat.MLarXiv:1807.04225v12018Transformer Hawkes Process
Simiao Zuo, Haoming Jiang, Zichong Li +2
cs.LGstat.MLarXiv:2002.09291v52020The Debate Over Understanding in AI's Large Language Models
Melanie Mitchell, David C. Krakauer
cs.LGcs.AIarXiv:2210.13966v32022Molecular generative model based on conditional variational autoencoder for de novo molecular design
Jaechang Lim, Seongok Ryu, Jin Woo Kim +1
cs.LGstat.MLarXiv:1806.05805v12018