Machine Learning

Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

11,461 to 11,520 of 19,954

  1. Generalized Nonconvex Nonsmooth Low-Rank Minimization

    Canyi Lu, Jinhui Tang, Shuicheng Yan +1

    cs.CVcs.LGstat.MLarXiv:1404.7306v12014
  2. Deliberative Alignment: Reasoning Enables Safer Language Models

    Melody Y. Guan, Manas Joglekar, Eric Wallace +12

    cs.CLcs.AIcs.CYarXiv:2412.16339v22024
  3. Monash Time Series Forecasting Archive

    Rakshitha Godahewa, Christoph Bergmeir, Geoffrey I. Webb +2

    cs.LGstat.MLarXiv:2105.06643v12021
  4. DA-TransUNet: Integrating Spatial and Channel Dual Attention with Transformer U-Net for Medical Image Segmentation

    Guanqun Sun, Yizhi Pan, Weikun Kong +5

    eess.IVcs.CVcs.GRarXiv:2310.12570v22023
  5. Noise Tolerance under Risk Minimization

    Naresh Manwani, P. S. Sastry

    cs.LGarXiv:1109.5231v42011
  6. Adversarial Self-Supervised Contrastive Learning

    Minseon Kim, Jihoon Tack, Sung Ju Hwang

    cs.LGcs.CVstat.MLarXiv:2006.07589v22020
  7. Adversarial Latent Autoencoders

    Stanislav Pidhorskyi, Donald Adjeroh, Gianfranco Doretto

    cs.LGcs.CVarXiv:2004.04467v12020
  8. RUDDER: Return Decomposition for Delayed Rewards

    Jose A. Arjona-Medina, Michael Gillhofer, Michael Widrich +3

    cs.LGcs.AImath.OCarXiv:1806.07857v32018
  9. A Divergence Minimization Perspective on Imitation Learning Methods

    Seyed Kamyar Seyed Ghasemipour, Richard Zemel, Shixiang Gu

    cs.LGstat.MLarXiv:1911.02256v12019
  10. Neural Sheaf Diffusion: A Topological Perspective on Heterophily and Oversmoothing in GNNs

    Cristian Bodnar, Francesco Di Giovanni, Benjamin Paul Chamberlain +2

    cs.LGmath.ATarXiv:2202.04579v42022
  11. The Disagreement Problem in Explainable Machine Learning: A Practitioner's Perspective

    Satyapriya Krishna, Tessa Han, Alex Gu +3

    cs.LGcs.AIarXiv:2202.01602v62022
  12. Mix-n-Match: Ensemble and Compositional Methods for Uncertainty Calibration in Deep Learning

    Jize Zhang, Bhavya Kailkhura, T. Yong-Jin Han

    cs.LGstat.MLarXiv:2003.07329v22020
  13. The Four Dimensions of Social Network Analysis: An Overview of Research Methods, Applications, and Software Tools

    David Camacho, Àngel Panizo-LLedot, Gema Bello-Orgaz +2

    cs.SIcs.CYcs.LGarXiv:2002.09485v12020
  14. A Survey on Neural Architecture Search

    Martin Wistuba, Ambrish Rawat, Tejaswini Pedapati

    cs.LGcs.CVcs.NEarXiv:1905.01392v22019
  15. Sketch-Guided Text-to-Image Diffusion Models

    Andrey Voynov, Kfir Aberman, Daniel Cohen-Or

    cs.CVcs.GRcs.LGarXiv:2211.13752v12022
  16. Deep Learning for Photoacoustic Tomography from Sparse Data

    Stephan Antholzer, Markus Haltmeier, Johannes Schwab

    cs.CVcs.LGarXiv:1704.04587v32017
  17. Unsupervised Pretraining for Sequence to Sequence Learning

    Prajit Ramachandran, Peter J. Liu, Quoc V. Le

    cs.CLcs.LGcs.NEarXiv:1611.02683v22016
  18. Transfer learning for time series classification

    Hassan Ismail Fawaz, Germain Forestier, Jonathan Weber +2

    cs.LGcs.AIstat.MLarXiv:1811.01533v12018
  19. Stochastic gradient descent for hybrid quantum-classical optimization

    Ryan Sweke, Frederik Wilde, Johannes Meyer +4

    quant-phcs.LGarXiv:1910.01155v32019
  20. Eigenvalues of the Hessian in Deep Learning: Singularity and Beyond

    Levent Sagun, Leon Bottou, Yann LeCun

    cs.LGarXiv:1611.07476v22016
  21. Predicting the Computational Cost of Deep Learning Models

    Daniel Justus, John Brennan, Stephen Bonner +1

    cs.LGcs.AIstat.MLarXiv:1811.11880v12018
  22. Comprehensive Privacy Analysis of Deep Learning: Passive and Active White-box Inference Attacks against Centralized and Federated Learning

    Milad Nasr, Reza Shokri, Amir Houmansadr

    stat.MLcs.CRcs.LGarXiv:1812.00910v22018
  23. Is Local SGD Better than Minibatch SGD?

    Blake Woodworth, Kumar Kshitij Patel, Sebastian U. Stich +5

    cs.LGmath.OCstat.MLarXiv:2002.07839v22020
  24. Deep Reinforcement Learning and the Deadly Triad

    Hado van Hasselt, Yotam Doron, Florian Strub +3

    cs.AIcs.LGarXiv:1812.02648v12018
  25. Tensor Graph Convolutional Networks for Text Classification

    Xien Liu, Xinxin You, Xiao Zhang +2

    cs.CLcs.IRcs.LGarXiv:2001.05313v12020
  26. Invertible Image Rescaling

    Mingqing Xiao, Shuxin Zheng, Chang Liu +6

    eess.IVcs.CVcs.LGarXiv:2005.05650v12020
  27. TransNets: Learning to Transform for Recommendation

    Rose Catherine, William Cohen

    cs.IRcs.CLcs.LGarXiv:1704.02298v22017
  28. Computing Graph Neural Networks: A Survey from Algorithms to Accelerators

    Sergi Abadal, Akshay Jain, Robert Guirado +2

    cs.LGcs.DCstat.MLarXiv:2010.00130v32020
  29. What to talk about and how? Selective Generation using LSTMs with Coarse-to-Fine Alignment

    Hongyuan Mei, Mohit Bansal, Matthew R. Walter

    cs.CLcs.AIcs.LGarXiv:1509.00838v22015
  30. Streaming Graph Neural Networks

    Yao Ma, Ziyi Guo, Zhaochun Ren +3

    cs.LGstat.MLarXiv:1810.10627v22018
  31. Overcoming Forgetting in Federated Learning on Non-IID Data

    Neta Shoham, Tomer Avidor, Aviv Keren +4

    cs.LGcs.CRstat.MLarXiv:1910.07796v12019
  32. Forecasting Global Weather with Graph Neural Networks

    Ryan Keisler

    physics.ao-phcs.LGarXiv:2202.07575v12022
  33. AugGPT: Leveraging ChatGPT for Text Data Augmentation

    Haixing Dai, Zhengliang Liu, Wenxiong Liao +15

    cs.CLcs.AIcs.LGarXiv:2302.13007v32023
  34. Language Models for Image Captioning: The Quirks and What Works

    Jacob Devlin, Hao Cheng, Hao Fang +5

    cs.CLcs.AIcs.CVarXiv:1505.01809v32015
  35. Equivariant 3D-Conditional Diffusion Models for Molecular Linker Design

    Ilia Igashov, Hannes Stärk, Clément Vignac +5

    cs.LGq-bio.BMarXiv:2210.05274v12022
  36. Causality Inspired Representation Learning for Domain Generalization

    Fangrui Lv, Jian Liang, Shuang Li +4

    cs.LGcs.CVarXiv:2203.14237v12022
  37. Information-Theoretic Probing for Linguistic Structure

    Tiago Pimentel, Josef Valvoda, Rowan Hall Maudslay +3

    cs.CLcs.LGarXiv:2004.03061v22020
  38. Reinforcement and Imitation Learning via Interactive No-Regret Learning

    Stephane Ross, J. Andrew Bagnell

    cs.LGstat.MLarXiv:1406.5979v12014
  39. Controlling Overestimation Bias with Truncated Mixture of Continuous Distributional Quantile Critics

    Arsenii Kuznetsov, Pavel Shvechikov, Alexander Grishin +1

    cs.LGcs.AIstat.MLarXiv:2005.04269v12020
  40. Model-Free Episodic Control

    Charles Blundell, Benigno Uria, Alexander Pritzel +6

    stat.MLcs.LGq-bio.NCarXiv:1606.04460v12016
  41. Rethinking the Backdoor Attacks' Triggers: A Frequency Perspective

    Yi Zeng, Won Park, Z. Morley Mao +1

    cs.LGcs.CRarXiv:2104.03413v42021
  42. Online Clustering of Bandits

    Claudio Gentile, Shuai Li, Giovanni Zappella

    cs.LGstat.MLarXiv:1401.8257v32014
  43. Algorithmic Regularization in Learning Deep Homogeneous Models: Layers are Automatically Balanced

    Simon S. Du, Wei Hu, Jason D. Lee

    cs.LGmath.OCstat.MLarXiv:1806.00900v22018
  44. A 3D Generative Model for Structure-Based Drug Design

    Shitong Luo, Jiaqi Guan, Jianzhu Ma +1

    q-bio.BMcs.LGarXiv:2203.10446v22022
  45. Banach Wasserstein GAN

    Jonas Adler, Sebastian Lunz

    cs.CVcs.LGmath.FAarXiv:1806.06621v22018
  46. The Hardware Lottery

    Sara Hooker

    cs.CYcs.AIcs.ARarXiv:2009.06489v22020
  47. Robust Semantic Communications with Masked VQ-VAE Enabled Codebook

    Qiyu Hu, Guangyi Zhang, Zhijin Qin +3

    eess.SPcs.ITcs.LGarXiv:2206.04011v22022
  48. Does Object Recognition Work for Everyone?

    Terrance DeVries, Ishan Misra, Changhan Wang +1

    cs.CVcs.LGarXiv:1906.02659v22019
  49. Three Mechanisms of Weight Decay Regularization

    Guodong Zhang, Chaoqi Wang, Bowen Xu +1

    cs.LGstat.MLarXiv:1810.12281v12018
  50. Label Words are Anchors: An Information Flow Perspective for Understanding In-Context Learning

    Lean Wang, Lei Li, Damai Dai +5

    cs.CLcs.LGarXiv:2305.14160v42023
  51. Rademacher Complexity for Adversarially Robust Generalization

    Dong Yin, Kannan Ramchandran, Peter Bartlett

    cs.LGcs.CRcs.NEarXiv:1810.11914v42018
  52. Counterfactual Fairness in Text Classification through Robustness

    Sahaj Garg, Vincent Perot, Nicole Limtiaco +3

    cs.LGstat.MLarXiv:1809.10610v22018
  53. Rectified Flow: A Marginal Preserving Approach to Optimal Transport

    Qiang Liu

    stat.MLcs.LGarXiv:2209.14577v12022
  54. A Symbolic Approach to Explaining Bayesian Network Classifiers

    Andy Shih, Arthur Choi, Adnan Darwiche

    cs.AIcs.LGarXiv:1805.03364v12018
  55. Augmentor: An Image Augmentation Library for Machine Learning

    Marcus D. Bloice, Christof Stocker, Andreas Holzinger

    cs.CVcs.LGstat.MLarXiv:1708.04680v12017
  56. GR-2: A Generative Video-Language-Action Model with Web-Scale Knowledge for Robot Manipulation

    Chi-Lam Cheang, Guangzeng Chen, Ya Jing +9

    cs.ROcs.CVcs.LGarXiv:2410.06158v12024
  57. A Survey of Zero-shot Generalisation in Deep Reinforcement Learning

    Robert Kirk, Amy Zhang, Edward Grefenstette +1

    cs.LGcs.AIarXiv:2111.09794v62021
  58. Distributed and parallel time series feature extraction for industrial big data applications

    Maximilian Christ, Andreas W. Kempa-Liehr, Michael Feindt

    cs.LGarXiv:1610.07717v32016
  59. Feature Inference Attack on Model Predictions in Vertical Federated Learning

    Xinjian Luo, Yuncheng Wu, Xiaokui Xiao +1

    cs.LGcs.DBarXiv:2010.10152v32020
  60. Episodic Curiosity through Reachability

    Nikolay Savinov, Anton Raichuk, Raphaël Marinier +4

    cs.LGcs.AIcs.CVarXiv:1810.02274v52018