Machine Learning

Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

12,241 to 12,300 of 20,193

  1. Two Sides of the Same Coin: Heterophily and Oversmoothing in Graph Convolutional Neural Networks

    Yujun Yan, Milad Hashemi, Kevin Swersky +2

    cs.LGarXiv:2102.06462v82021
  2. Titans: Learning to Memorize at Test Time

    Ali Behrouz, Peilin Zhong, Vahab Mirrokni

    cs.LGcs.AIcs.CLarXiv:2501.00663v12024
  3. Federated Class-Incremental Learning

    Jiahua Dong, Lixu Wang, Zhen Fang +4

    cs.LGarXiv:2203.11473v12022
  4. AutoSDF: Shape Priors for 3D Completion, Reconstruction and Generation

    Paritosh Mittal, Yen-Chi Cheng, Maneesh Singh +1

    cs.CVcs.LGarXiv:2203.09516v32022
  5. Resurrecting the sigmoid in deep learning through dynamical isometry: theory and practice

    Jeffrey Pennington, Samuel S. Schoenholz, Surya Ganguli

    cs.LGstat.MLarXiv:1711.04735v12017
  6. Towards Knowledge-Based Recommender Dialog System

    Qibin Chen, Junyang Lin, Yichang Zhang +4

    cs.CLcs.IRcs.LGarXiv:1908.05391v22019
  7. Ablation Studies in Artificial Neural Networks

    Richard Meyes, Melanie Lu, Constantin Waubert de Puiseau +1

    cs.NEcs.LGq-bio.NCarXiv:1901.08644v22019
  8. Explaining the Success of AdaBoost and Random Forests as Interpolating Classifiers

    Abraham J. Wyner, Matthew Olson, Justin Bleich +1

    stat.MLcs.LGstat.MEarXiv:1504.07676v22015
  9. POD-DL-ROM: enhancing deep learning-based reduced order models for nonlinear parametrized PDEs by proper orthogonal decomposition

    Stefania Fresca, Andrea Manzoni

    math.NAcs.LGarXiv:2101.11845v12021
  10. The relationship between trust in AI and trustworthy machine learning technologies

    Ehsan Toreini, Mhairi Aitken, Kovila Coopamootoo +3

    cs.CYcs.AIcs.LGarXiv:1912.00782v22019
  11. STG2Seq: Spatial-temporal Graph to Sequence Model for Multi-step Passenger Demand Forecasting

    Lei Bai, Lina Yao, Salil. S Kanhere +2

    cs.LGcs.AIstat.MLarXiv:1905.10069v12019
  12. Adversarial GLUE: A Multi-Task Benchmark for Robustness Evaluation of Language Models

    Boxin Wang, Chejian Xu, Shuohang Wang +5

    cs.CLcs.CRcs.LGarXiv:2111.02840v22021
  13. Think before you speak: Training Language Models With Pause Tokens

    Sachin Goyal, Ziwei Ji, Ankit Singh Rawat +3

    cs.CLcs.AIcs.LGarXiv:2310.02226v32023
  14. How to Find Your Friendly Neighborhood: Graph Attention Design with Self-Supervision

    Dongkwan Kim, Alice Oh

    cs.LGcs.AIcs.SIarXiv:2204.04879v12022
  15. Transformers learn to implement preconditioned gradient descent for in-context learning

    Kwangjun Ahn, Xiang Cheng, Hadi Daneshmand +1

    cs.LGcs.AIarXiv:2306.00297v22023
  16. Variational inference for Monte Carlo objectives

    Andriy Mnih, Danilo J. Rezende

    cs.LGstat.MLarXiv:1602.06725v22016
  17. MMA Training: Direct Input Space Margin Maximization through Adversarial Training

    Gavin Weiguang Ding, Yash Sharma, Kry Yik Chau Lui +1

    cs.LGcs.NEstat.MLarXiv:1812.02637v42018
  18. Sionna RT: Differentiable Ray Tracing for Radio Propagation Modeling

    Jakob Hoydis, Fayçal Aït Aoudia, Sebastian Cammerer +4

    cs.ITcs.AIcs.LGarXiv:2303.11103v22023
  19. Deep Reinforcement Learning with Successor Features for Navigation across Similar Environments

    Jingwei Zhang, Jost Tobias Springenberg, Joschka Boedecker +1

    cs.ROcs.AIcs.LGarXiv:1612.05533v32016
  20. Diffusion-based Time Series Imputation and Forecasting with Structured State Space Models

    Juan Miguel Lopez Alcaraz, Nils Strodthoff

    cs.LGstat.MLarXiv:2208.09399v32022
  21. DexCap: Scalable and Portable Mocap Data Collection System for Dexterous Manipulation

    Chen Wang, Haochen Shi, Weizhuo Wang +3

    cs.ROcs.AIcs.CVarXiv:2403.07788v22024
  22. Do NLP Models Know Numbers? Probing Numeracy in Embeddings

    Eric Wallace, Yizhong Wang, Sujian Li +2

    cs.CLcs.LGarXiv:1909.07940v22019
  23. PixelSNAIL: An Improved Autoregressive Generative Model

    Xi Chen, Nikhil Mishra, Mostafa Rohaninejad +1

    cs.LGstat.MLarXiv:1712.09763v12017
  24. Learning how to Active Learn: A Deep Reinforcement Learning Approach

    Meng Fang, Yuan Li, Trevor Cohn

    cs.CLcs.AIcs.LGarXiv:1708.02383v12017
  25. Consistency Regularization for Generative Adversarial Networks

    Han Zhang, Zizhao Zhang, Augustus Odena +1

    cs.LGcs.CVstat.MLarXiv:1910.12027v22019
  26. Adversarial Attacks on Deep-Learning Based Radio Signal Classification

    Meysam Sadeghi, Erik G. Larsson

    cs.ITcs.CRcs.LGarXiv:1808.07713v12018
  27. TRAK: Attributing Model Behavior at Scale

    Sung Min Park, Kristian Georgiev, Andrew Ilyas +2

    stat.MLcs.LGarXiv:2303.14186v22023
  28. SLURP: A Spoken Language Understanding Resource Package

    Emanuele Bastianelli, Andrea Vanzo, Pawel Swietojanski +1

    cs.CLcs.LGarXiv:2011.13205v12020
  29. A Survey of Reinforcement Learning Informed by Natural Language

    Jelena Luketina, Nantas Nardelli, Gregory Farquhar +5

    cs.LGcs.AIcs.CLarXiv:1906.03926v12019
  30. Information-theoretic bounds on quantum advantage in machine learning

    Hsin-Yuan Huang, Richard Kueng, John Preskill

    quant-phcs.ITcs.LGarXiv:2101.02464v22021
  31. A$^3$: Accelerating Attention Mechanisms in Neural Networks with Approximation

    Tae Jun Ham, Sung Jun Jung, Seonghak Kim +8

    cs.DCcs.LGarXiv:2002.10941v12020
  32. Membership Leakage in Label-Only Exposures

    Zheng Li, Yang Zhang

    cs.LGcs.CRstat.MLarXiv:2007.15528v32020
  33. Where are we in the search for an Artificial Visual Cortex for Embodied Intelligence?

    Arjun Majumdar, Karmesh Yadav, Sergio Arnaud +12

    cs.CVcs.AIcs.LGarXiv:2303.18240v22023
  34. Reinforcement Learning in Feature Space: Matrix Bandit, Kernels, and Regret Bound

    Lin F. Yang, Mengdi Wang

    cs.LGstat.MLarXiv:1905.10389v22019
  35. Controllable Invariance through Adversarial Feature Learning

    Qizhe Xie, Zihang Dai, Yulun Du +2

    cs.LGcs.AIcs.CLarXiv:1705.11122v32017
  36. Data Driven Governing Equations Approximation Using Deep Neural Networks

    Tong Qin, Kailiang Wu, Dongbin Xiu

    math.NAcs.LGcs.NEarXiv:1811.05537v12018
  37. MLAgentBench: Evaluating Language Agents on Machine Learning Experimentation

    Qian Huang, Jian Vora, Percy Liang +1

    cs.LGcs.AIarXiv:2310.03302v22023
  38. PassGAN: A Deep Learning Approach for Password Guessing

    Briland Hitaj, Paolo Gasti, Giuseppe Ateniese +1

    cs.CRcs.LGstat.MLarXiv:1709.00440v32017
  39. Diffusion models as plug-and-play priors

    Alexandros Graikos, Nikolay Malkin, Nebojsa Jojic +1

    cs.LGcs.CVarXiv:2206.09012v32022
  40. Universal Adversarial Perturbations Against Semantic Image Segmentation

    Jan Hendrik Metzen, Mummadi Chaithanya Kumar, Thomas Brox +1

    stat.MLcs.AIcs.CVarXiv:1704.05712v32017
  41. JointDNN: An Efficient Training and Inference Engine for Intelligent Mobile Cloud Computing Services

    Amir Erfan Eshratifar, Mohammad Saeed Abrishami, Massoud Pedram

    cs.DCcs.AIcs.LGarXiv:1801.08618v22018
  42. Continual Learning for Robotics: Definition, Framework, Learning Strategies, Opportunities and Challenges

    Timothée Lesort, Vincenzo Lomonaco, Andrei Stoian +3

    cs.LGcs.ROarXiv:1907.00182v32019
  43. Do Convnets Learn Correspondence?

    Jonathan Long, Ning Zhang, Trevor Darrell

    cs.CVcs.LGcs.NEarXiv:1411.1091v12014
  44. Auxiliary Tasks in Multi-task Learning

    Lukas Liebel, Marco Körner

    cs.CVcs.LGarXiv:1805.06334v22018
  45. Factuality Enhanced Language Models for Open-Ended Text Generation

    Nayeon Lee, Wei Ping, Peng Xu +4

    cs.CLcs.AIcs.CYarXiv:2206.04624v32022
  46. A Question-Entailment Approach to Question Answering

    Asma Ben Abacha, Dina Demner-Fushman

    cs.CLcs.AIcs.IRarXiv:1901.08079v12019
  47. Is Homophily a Necessity for Graph Neural Networks?

    Yao Ma, Xiaorui Liu, Neil Shah +1

    cs.LGstat.MLarXiv:2106.06134v42021
  48. Prediction-Powered Inference

    Anastasios N. Angelopoulos, Stephen Bates, Clara Fannjiang +2

    stat.MLcs.AIcs.LGarXiv:2301.09633v42023
  49. Projected GANs Converge Faster

    Axel Sauer, Kashyap Chitta, Jens Müller +1

    cs.CVcs.LGarXiv:2111.01007v12021
  50. Deep Network Approximation for Smooth Functions

    Jianfeng Lu, Zuowei Shen, Haizhao Yang +1

    cs.LGmath.NAstat.MLarXiv:2001.03040v82020
  51. Communication Compression for Decentralized Training

    Hanlin Tang, Shaoduo Gan, Ce Zhang +2

    cs.LGcs.DCeess.SYarXiv:1803.06443v52018
  52. Latent Action Pretraining from Videos

    Seonghyeon Ye, Joel Jang, Byeongguk Jeon +13

    cs.ROcs.CLcs.CVarXiv:2410.11758v22024
  53. LSQ+: Improving low-bit quantization through learnable offsets and better initialization

    Yash Bhalgat, Jinwon Lee, Markus Nagel +2

    cs.CVcs.LGstat.MLarXiv:2004.09576v12020
  54. DeepSentiBank: Visual Sentiment Concept Classification with Deep Convolutional Neural Networks

    Tao Chen, Damian Borth, Trevor Darrell +1

    cs.CVcs.LGcs.MMarXiv:1410.8586v12014
  55. Self-supervised Deep Reinforcement Learning with Generalized Computation Graphs for Robot Navigation

    Gregory Kahn, Adam Villaflor, Bosen Ding +2

    cs.LGcs.AIcs.ROarXiv:1709.10489v32017
  56. Federated Learning for Keyword Spotting

    David Leroy, Alice Coucke, Thibaut Lavril +2

    eess.AScs.CLcs.LGarXiv:1810.05512v42018
  57. Comparing BERT against traditional machine learning text classification

    Santiago González-Carvajal, Eduardo C. Garrido-Merchán

    cs.CLcs.LGstat.MLarXiv:2005.13012v22020
  58. Neural Speech Recognizer: Acoustic-to-Word LSTM Model for Large Vocabulary Speech Recognition

    Hagen Soltau, Hank Liao, Hasim Sak

    cs.CLcs.LGcs.NEarXiv:1610.09975v12016
  59. Robot Parkour Learning

    Ziwen Zhuang, Zipeng Fu, Jianren Wang +4

    cs.ROcs.AIcs.CVarXiv:2309.05665v22023
  60. Coresets via Bilevel Optimization for Continual Learning and Streaming

    Zalán Borsos, Mojmír Mutný, Andreas Krause

    cs.LGstat.MLarXiv:2006.03875v22020