Machine Learning

Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

11,701 to 11,760 of 20,187

  1. A Divergence Minimization Perspective on Imitation Learning Methods

    Seyed Kamyar Seyed Ghasemipour, Richard Zemel, Shixiang Gu

    cs.LGstat.MLarXiv:1911.02256v12019
  2. Neural Sheaf Diffusion: A Topological Perspective on Heterophily and Oversmoothing in GNNs

    Cristian Bodnar, Francesco Di Giovanni, Benjamin Paul Chamberlain +2

    cs.LGmath.ATarXiv:2202.04579v42022
  3. The Disagreement Problem in Explainable Machine Learning: A Practitioner's Perspective

    Satyapriya Krishna, Tessa Han, Alex Gu +3

    cs.LGcs.AIarXiv:2202.01602v62022
  4. Mix-n-Match: Ensemble and Compositional Methods for Uncertainty Calibration in Deep Learning

    Jize Zhang, Bhavya Kailkhura, T. Yong-Jin Han

    cs.LGstat.MLarXiv:2003.07329v22020
  5. The Four Dimensions of Social Network Analysis: An Overview of Research Methods, Applications, and Software Tools

    David Camacho, Àngel Panizo-LLedot, Gema Bello-Orgaz +2

    cs.SIcs.CYcs.LGarXiv:2002.09485v12020
  6. A Survey on Neural Architecture Search

    Martin Wistuba, Ambrish Rawat, Tejaswini Pedapati

    cs.LGcs.CVcs.NEarXiv:1905.01392v22019
  7. Sketch-Guided Text-to-Image Diffusion Models

    Andrey Voynov, Kfir Aberman, Daniel Cohen-Or

    cs.CVcs.GRcs.LGarXiv:2211.13752v12022
  8. Deep Learning for Photoacoustic Tomography from Sparse Data

    Stephan Antholzer, Markus Haltmeier, Johannes Schwab

    cs.CVcs.LGarXiv:1704.04587v32017
  9. Unsupervised Pretraining for Sequence to Sequence Learning

    Prajit Ramachandran, Peter J. Liu, Quoc V. Le

    cs.CLcs.LGcs.NEarXiv:1611.02683v22016
  10. Transfer learning for time series classification

    Hassan Ismail Fawaz, Germain Forestier, Jonathan Weber +2

    cs.LGcs.AIstat.MLarXiv:1811.01533v12018
  11. Stochastic gradient descent for hybrid quantum-classical optimization

    Ryan Sweke, Frederik Wilde, Johannes Meyer +4

    quant-phcs.LGarXiv:1910.01155v32019
  12. Eigenvalues of the Hessian in Deep Learning: Singularity and Beyond

    Levent Sagun, Leon Bottou, Yann LeCun

    cs.LGarXiv:1611.07476v22016
  13. Predicting the Computational Cost of Deep Learning Models

    Daniel Justus, John Brennan, Stephen Bonner +1

    cs.LGcs.AIstat.MLarXiv:1811.11880v12018
  14. Comprehensive Privacy Analysis of Deep Learning: Passive and Active White-box Inference Attacks against Centralized and Federated Learning

    Milad Nasr, Reza Shokri, Amir Houmansadr

    stat.MLcs.CRcs.LGarXiv:1812.00910v22018
  15. Is Local SGD Better than Minibatch SGD?

    Blake Woodworth, Kumar Kshitij Patel, Sebastian U. Stich +5

    cs.LGmath.OCstat.MLarXiv:2002.07839v22020
  16. Deep Reinforcement Learning and the Deadly Triad

    Hado van Hasselt, Yotam Doron, Florian Strub +3

    cs.AIcs.LGarXiv:1812.02648v12018
  17. Tensor Graph Convolutional Networks for Text Classification

    Xien Liu, Xinxin You, Xiao Zhang +2

    cs.CLcs.IRcs.LGarXiv:2001.05313v12020
  18. Invertible Image Rescaling

    Mingqing Xiao, Shuxin Zheng, Chang Liu +6

    eess.IVcs.CVcs.LGarXiv:2005.05650v12020
  19. TransNets: Learning to Transform for Recommendation

    Rose Catherine, William Cohen

    cs.IRcs.CLcs.LGarXiv:1704.02298v22017
  20. Computing Graph Neural Networks: A Survey from Algorithms to Accelerators

    Sergi Abadal, Akshay Jain, Robert Guirado +2

    cs.LGcs.DCstat.MLarXiv:2010.00130v32020
  21. What to talk about and how? Selective Generation using LSTMs with Coarse-to-Fine Alignment

    Hongyuan Mei, Mohit Bansal, Matthew R. Walter

    cs.CLcs.AIcs.LGarXiv:1509.00838v22015
  22. Streaming Graph Neural Networks

    Yao Ma, Ziyi Guo, Zhaochun Ren +3

    cs.LGstat.MLarXiv:1810.10627v22018
  23. Overcoming Forgetting in Federated Learning on Non-IID Data

    Neta Shoham, Tomer Avidor, Aviv Keren +4

    cs.LGcs.CRstat.MLarXiv:1910.07796v12019
  24. Forecasting Global Weather with Graph Neural Networks

    Ryan Keisler

    physics.ao-phcs.LGarXiv:2202.07575v12022
  25. AugGPT: Leveraging ChatGPT for Text Data Augmentation

    Haixing Dai, Zhengliang Liu, Wenxiong Liao +15

    cs.CLcs.AIcs.LGarXiv:2302.13007v32023
  26. Language Models for Image Captioning: The Quirks and What Works

    Jacob Devlin, Hao Cheng, Hao Fang +5

    cs.CLcs.AIcs.CVarXiv:1505.01809v32015
  27. Equivariant 3D-Conditional Diffusion Models for Molecular Linker Design

    Ilia Igashov, Hannes Stärk, Clément Vignac +5

    cs.LGq-bio.BMarXiv:2210.05274v12022
  28. Causality Inspired Representation Learning for Domain Generalization

    Fangrui Lv, Jian Liang, Shuang Li +4

    cs.LGcs.CVarXiv:2203.14237v12022
  29. Information-Theoretic Probing for Linguistic Structure

    Tiago Pimentel, Josef Valvoda, Rowan Hall Maudslay +3

    cs.CLcs.LGarXiv:2004.03061v22020
  30. Reinforcement and Imitation Learning via Interactive No-Regret Learning

    Stephane Ross, J. Andrew Bagnell

    cs.LGstat.MLarXiv:1406.5979v12014
  31. Controlling Overestimation Bias with Truncated Mixture of Continuous Distributional Quantile Critics

    Arsenii Kuznetsov, Pavel Shvechikov, Alexander Grishin +1

    cs.LGcs.AIstat.MLarXiv:2005.04269v12020
  32. Model-Free Episodic Control

    Charles Blundell, Benigno Uria, Alexander Pritzel +6

    stat.MLcs.LGq-bio.NCarXiv:1606.04460v12016
  33. Rethinking the Backdoor Attacks' Triggers: A Frequency Perspective

    Yi Zeng, Won Park, Z. Morley Mao +1

    cs.LGcs.CRarXiv:2104.03413v42021
  34. Online Clustering of Bandits

    Claudio Gentile, Shuai Li, Giovanni Zappella

    cs.LGstat.MLarXiv:1401.8257v32014
  35. Algorithmic Regularization in Learning Deep Homogeneous Models: Layers are Automatically Balanced

    Simon S. Du, Wei Hu, Jason D. Lee

    cs.LGmath.OCstat.MLarXiv:1806.00900v22018
  36. A 3D Generative Model for Structure-Based Drug Design

    Shitong Luo, Jiaqi Guan, Jianzhu Ma +1

    q-bio.BMcs.LGarXiv:2203.10446v22022
  37. Banach Wasserstein GAN

    Jonas Adler, Sebastian Lunz

    cs.CVcs.LGmath.FAarXiv:1806.06621v22018
  38. The Hardware Lottery

    Sara Hooker

    cs.CYcs.AIcs.ARarXiv:2009.06489v22020
  39. Robust Semantic Communications with Masked VQ-VAE Enabled Codebook

    Qiyu Hu, Guangyi Zhang, Zhijin Qin +3

    eess.SPcs.ITcs.LGarXiv:2206.04011v22022
  40. Does Object Recognition Work for Everyone?

    Terrance DeVries, Ishan Misra, Changhan Wang +1

    cs.CVcs.LGarXiv:1906.02659v22019
  41. Three Mechanisms of Weight Decay Regularization

    Guodong Zhang, Chaoqi Wang, Bowen Xu +1

    cs.LGstat.MLarXiv:1810.12281v12018
  42. Label Words are Anchors: An Information Flow Perspective for Understanding In-Context Learning

    Lean Wang, Lei Li, Damai Dai +5

    cs.CLcs.LGarXiv:2305.14160v42023
  43. Rademacher Complexity for Adversarially Robust Generalization

    Dong Yin, Kannan Ramchandran, Peter Bartlett

    cs.LGcs.CRcs.NEarXiv:1810.11914v42018
  44. Counterfactual Fairness in Text Classification through Robustness

    Sahaj Garg, Vincent Perot, Nicole Limtiaco +3

    cs.LGstat.MLarXiv:1809.10610v22018
  45. Rectified Flow: A Marginal Preserving Approach to Optimal Transport

    Qiang Liu

    stat.MLcs.LGarXiv:2209.14577v12022
  46. A Symbolic Approach to Explaining Bayesian Network Classifiers

    Andy Shih, Arthur Choi, Adnan Darwiche

    cs.AIcs.LGarXiv:1805.03364v12018
  47. Augmentor: An Image Augmentation Library for Machine Learning

    Marcus D. Bloice, Christof Stocker, Andreas Holzinger

    cs.CVcs.LGstat.MLarXiv:1708.04680v12017
  48. GR-2: A Generative Video-Language-Action Model with Web-Scale Knowledge for Robot Manipulation

    Chi-Lam Cheang, Guangzeng Chen, Ya Jing +9

    cs.ROcs.CVcs.LGarXiv:2410.06158v12024
  49. A Survey of Zero-shot Generalisation in Deep Reinforcement Learning

    Robert Kirk, Amy Zhang, Edward Grefenstette +1

    cs.LGcs.AIarXiv:2111.09794v62021
  50. Distributed and parallel time series feature extraction for industrial big data applications

    Maximilian Christ, Andreas W. Kempa-Liehr, Michael Feindt

    cs.LGarXiv:1610.07717v32016
  51. Feature Inference Attack on Model Predictions in Vertical Federated Learning

    Xinjian Luo, Yuncheng Wu, Xiaokui Xiao +1

    cs.LGcs.DBarXiv:2010.10152v32020
  52. Episodic Curiosity through Reachability

    Nikolay Savinov, Anton Raichuk, Raphaël Marinier +4

    cs.LGcs.AIcs.CVarXiv:1810.02274v52018
  53. The KFIoU Loss for Rotated Object Detection

    Xue Yang, Yue Zhou, Gefan Zhang +5

    cs.CVcs.AIcs.LGarXiv:2201.12558v62022
  54. Log-based Anomaly Detection with Deep Learning: How Far Are We?

    Van-Hoang Le, Hongyu Zhang

    cs.SEcs.LGarXiv:2202.04301v22022
  55. 3D Infomax improves GNNs for Molecular Property Prediction

    Hannes Stärk, Dominique Beaini, Gabriele Corso +4

    cs.LGcs.AIq-bio.BMarXiv:2110.04126v42021
  56. Get Your Vitamin C! Robust Fact Verification with Contrastive Evidence

    Tal Schuster, Adam Fisch, Regina Barzilay

    cs.CLcs.IRcs.LGarXiv:2103.08541v12021
  57. Function Vectors in Large Language Models

    Eric Todd, Millicent L. Li, Arnab Sen Sharma +3

    cs.CLcs.LGarXiv:2310.15213v22023
  58. Towards Debiasing Sentence Representations

    Paul Pu Liang, Irene Mengze Li, Emily Zheng +3

    cs.CLcs.LGarXiv:2007.08100v12020
  59. Unlocking High-Accuracy Differentially Private Image Classification through Scale

    Soham De, Leonard Berrada, Jamie Hayes +2

    cs.LGcs.CRcs.CVarXiv:2204.13650v22022
  60. Delving into Out-of-Distribution Detection with Vision-Language Representations

    Yifei Ming, Ziyang Cai, Jiuxiang Gu +3

    cs.CVcs.AIcs.LGarXiv:2211.13445v12022