Machine Learning

Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

12,301 to 12,360 of 20,193

  1. Scaling Limits of Wide Neural Networks with Weight Sharing: Gaussian Process Behavior, Gradient Independence, and Neural Tangent Kernel Derivation

    Greg Yang

    cs.NEcond-mat.dis-nncs.LGarXiv:1902.04760v32019
  2. Inference Suboptimality in Variational Autoencoders

    Chris Cremer, Xuechen Li, David Duvenaud

    cs.LGstat.MLarXiv:1801.03558v32018
  3. Skew-Fit: State-Covering Self-Supervised Reinforcement Learning

    Vitchyr H. Pong, Murtaza Dalal, Steven Lin +3

    cs.LGcs.AIcs.ROarXiv:1903.03698v42019
  4. A vector-contraction inequality for Rademacher complexities

    Andreas Maurer

    cs.LGstat.MLarXiv:1605.00251v12016
  5. Trading Regret for Efficiency: Online Convex Optimization with Long Term Constraints

    Mehrdad Mahdavi, Rong Jin, Tianbao Yang

    cs.LGarXiv:1111.6082v32011
  6. Weakly-supervised Disentangling with Recurrent Transformations for 3D View Synthesis

    Jimei Yang, Scott Reed, Ming-Hsuan Yang +1

    cs.LGcs.AIcs.CVarXiv:1601.00706v12016
  7. Depthwise Separable Convolutions for Neural Machine Translation

    Lukasz Kaiser, Aidan N. Gomez, Francois Chollet

    cs.CLcs.LGarXiv:1706.03059v22017
  8. On the Robustness of ChatGPT: An Adversarial and Out-of-distribution Perspective

    Jindong Wang, Xixu Hu, Wenxin Hou +10

    cs.AIcs.CLcs.LGarXiv:2302.12095v52023
  9. IconQA: A New Benchmark for Abstract Diagram Understanding and Visual Language Reasoning

    Pan Lu, Liang Qiu, Jiaqi Chen +6

    cs.CVcs.AIcs.CLarXiv:2110.13214v42021
  10. Efficient Defenses Against Adversarial Attacks

    Valentina Zantedeschi, Maria-Irina Nicolae, Ambrish Rawat

    cs.LGarXiv:1707.06728v22017
  11. Multimodal Sentiment Analysis with Word-Level Fusion and Reinforcement Learning

    Minghai Chen, Sen Wang, Paul Pu Liang +3

    cs.LGcs.AIcs.CLarXiv:1802.00924v12018
  12. Deep Rewiring: Training very sparse deep networks

    Guillaume Bellec, David Kappel, Wolfgang Maass +1

    cs.NEcs.AIcs.DCarXiv:1711.05136v52017
  13. The Mechanics of n-Player Differentiable Games

    David Balduzzi, Sebastien Racaniere, James Martens +3

    cs.LGcs.GTcs.MAarXiv:1802.05642v22018
  14. A Survey of Graph Neural Networks for Social Recommender Systems

    Kartik Sharma, Yeon-Chang Lee, Sivagami Nambi +4

    cs.SIcs.IRcs.LGarXiv:2212.04481v32022
  15. Representer Point Selection for Explaining Deep Neural Networks

    Chih-Kuan Yeh, Joon Sik Kim, Ian E. H. Yen +1

    cs.LGstat.MLarXiv:1811.09720v12018
  16. SteganoGAN: High Capacity Image Steganography with GANs

    Kevin Alex Zhang, Alfredo Cuesta-Infante, Lei Xu +1

    cs.CVcs.LGcs.MMarXiv:1901.03892v22019
  17. Deep Multimodal Fusion by Channel Exchanging

    Yikai Wang, Wenbing Huang, Fuchun Sun +3

    cs.CVcs.LGarXiv:2011.05005v22020
  18. How (not) to Train your Generative Model: Scheduled Sampling, Likelihood, Adversary?

    Ferenc Huszár

    stat.MLcs.AIcs.ITarXiv:1511.05101v12015
  19. oLMpics -- On what Language Model Pre-training Captures

    Alon Talmor, Yanai Elazar, Yoav Goldberg +1

    cs.CLcs.AIcs.LGarXiv:1912.13283v22019
  20. Grounded Language Learning in a Simulated 3D World

    Karl Moritz Hermann, Felix Hill, Simon Green +11

    cs.CLcs.LGstat.MLarXiv:1706.06551v22017
  21. Towards an Appropriate Query, Key, and Value Computation for Knowledge Tracing

    Youngduck Choi, Youngnam Lee, Junghyun Cho +6

    cs.LGcs.AIcs.CYarXiv:2002.07033v52020
  22. Focal Sparse Convolutional Networks for 3D Object Detection

    Yukang Chen, Yanwei Li, Xiangyu Zhang +2

    cs.CVcs.LGarXiv:2204.12463v12022
  23. Evaluating and Aggregating Feature-based Model Explanations

    Umang Bhatt, Adrian Weller, José M. F. Moura

    cs.LGcs.AIcs.CYarXiv:2005.00631v12020
  24. Input complexity and out-of-distribution detection with likelihood-based generative models

    Joan Serrà, David Álvarez, Vicenç Gómez +3

    cs.LGstat.MLarXiv:1909.11480v32019
  25. Backpropagation through the Void: Optimizing control variates for black-box gradient estimation

    Will Grathwohl, Dami Choi, Yuhuai Wu +2

    cs.LGarXiv:1711.00123v32017
  26. Lift & Learn: Physics-informed machine learning for large-scale nonlinear dynamical systems

    Elizabeth Qian, Boris Kramer, Benjamin Peherstorfer +1

    math.NAcs.LGarXiv:1912.08177v52019
  27. BayesOpt: A Bayesian Optimization Library for Nonlinear Optimization, Experimental Design and Bandits

    Ruben Martinez-Cantin

    cs.LGarXiv:1405.7430v12014
  28. LLMLingua: Compressing Prompts for Accelerated Inference of Large Language Models

    Huiqiang Jiang, Qianhui Wu, Chin-Yew Lin +2

    cs.CLcs.LGarXiv:2310.05736v22023
  29. Salvaging Federated Learning by Local Adaptation

    Tao Yu, Eugene Bagdasaryan, Vitaly Shmatikov

    cs.LGcs.AIcs.DCarXiv:2002.04758v32020
  30. MetaFormer Baselines for Vision

    Weihao Yu, Chenyang Si, Pan Zhou +5

    cs.CVcs.AIcs.LGarXiv:2210.13452v42022
  31. CFA: Coupled-hypersphere-based Feature Adaptation for Target-Oriented Anomaly Localization

    Sungwook Lee, Seunghyun Lee, Byung Cheol Song

    cs.CVcs.LGarXiv:2206.04325v12022
  32. BadNL: Backdoor Attacks against NLP Models with Semantic-preserving Improvements

    Xiaoyi Chen, Ahmed Salem, Dingfan Chen +5

    cs.CRcs.LGarXiv:2006.01043v22020
  33. Optimal Clustering Framework for Hyperspectral Band Selection

    Qi Wang, Fahong Zhang, Xuelong Li

    eess.IVcs.LGstat.MLarXiv:1904.13036v12019
  34. Neural Descriptor Fields: SE(3)-Equivariant Object Representations for Manipulation

    Anthony Simeonov, Yilun Du, Andrea Tagliasacchi +4

    cs.ROcs.AIcs.CVarXiv:2112.05124v12021
  35. Sparse Sequence-to-Sequence Models

    Ben Peters, Vlad Niculae, André F. T. Martins

    cs.CLcs.LGarXiv:1905.05702v22019
  36. Detecting Cyberattacks in Industrial Control Systems Using Convolutional Neural Networks

    Moshe Kravchik, Asaf Shabtai

    cs.CRcs.LGarXiv:1806.08110v22018
  37. SafetyNet: Detecting and Rejecting Adversarial Examples Robustly

    Jiajun Lu, Theerasit Issaranon, David Forsyth

    cs.CVcs.LGarXiv:1704.00103v22017
  38. Automated Fact-Checking for Assisting Human Fact-Checkers

    Preslav Nakov, David Corney, Maram Hasanain +6

    cs.AIcs.CLcs.CRarXiv:2103.07769v22021
  39. Interpretable Learning for Self-Driving Cars by Visualizing Causal Attention

    Jinkyu Kim, John Canny

    cs.CVcs.LGarXiv:1703.10631v12017
  40. CausalGAN: Learning Causal Implicit Generative Models with Adversarial Training

    Murat Kocaoglu, Christopher Snyder, Alexandros G. Dimakis +1

    cs.LGcs.AIcs.ITarXiv:1709.02023v22017
  41. Chip-Chat: Challenges and Opportunities in Conversational Hardware Design

    Jason Blocklove, Siddharth Garg, Ramesh Karri +1

    cs.LGcs.ARcs.PLarXiv:2305.13243v22023
  42. Object Contour Detection with a Fully Convolutional Encoder-Decoder Network

    Jimei Yang, Brian Price, Scott Cohen +2

    cs.CVcs.LGarXiv:1603.04530v12016
  43. EDICT: Exact Diffusion Inversion via Coupled Transformations

    Bram Wallace, Akash Gokul, Nikhil Naik

    cs.CVcs.AIcs.LGarXiv:2211.12446v22022
  44. Bayesian Posterior Sampling via Stochastic Gradient Fisher Scoring

    Sungjin Ahn, Anoop Korattikara, Max Welling

    cs.LGstat.COstat.MLarXiv:1206.6380v12012
  45. Stochastic Controlled Averaging for Federated Learning with Communication Compression

    Xinmeng Huang, Ping Li, Xiaoyun Li

    math.OCcs.DCcs.LGarXiv:2308.08165v22023
  46. Unsupervised and Semi-supervised Anomaly Detection with LSTM Neural Networks

    Tolga Ergen, Ali Hassan Mirza, Suleyman Serdar Kozat

    eess.SPcs.LGstat.MLarXiv:1710.09207v12017
  47. Toward Robustness against Label Noise in Training Deep Discriminative Neural Networks

    Arash Vahdat

    cs.LGstat.MLarXiv:1706.00038v22017
  48. Deep Learning-enabled Virtual Histological Staining of Biological Samples

    Bijie Bai, Xilin Yang, Yuzhu Li +3

    physics.med-phcs.CVcs.LGarXiv:2211.06822v12022
  49. Deep Learning for Time Series Forecasting: The Electric Load Case

    Alberto Gasparin, Slobodan Lukovic, Cesare Alippi

    cs.LGstat.MLarXiv:1907.09207v12019
  50. Modern WLAN Fingerprinting Indoor Positioning Methods and Deployment Challenges

    Ali Khalajmehrabadi, Nikolaos Gatsis, David Akopian

    cs.NIcs.LGarXiv:1610.05424v12016
  51. Maximizing acquisition functions for Bayesian optimization

    James T. Wilson, Frank Hutter, Marc Peter Deisenroth

    stat.MLcs.LGarXiv:1805.10196v22018
  52. Globally Optimal Gradient Descent for a ConvNet with Gaussian Inputs

    Alon Brutzkus, Amir Globerson

    cs.LGmath.OCstat.MLarXiv:1702.07966v12017
  53. A physics-informed variational DeepONet for predicting the crack path in brittle materials

    Somdatta Goswami, Minglang Yin, Yue Yu +1

    cs.LGmath.NAarXiv:2108.06905v22021
  54. On the Inductive Bias of Neural Tangent Kernels

    Alberto Bietti, Julien Mairal

    stat.MLcs.LGarXiv:1905.12173v22019
  55. GMNN: Graph Markov Neural Networks

    Meng Qu, Yoshua Bengio, Jian Tang

    cs.LGcs.SIstat.MLarXiv:1905.06214v32019
  56. DyLoRA: Parameter Efficient Tuning of Pre-trained Models using Dynamic Search-Free Low-Rank Adaptation

    Mojtaba Valipour, Mehdi Rezagholizadeh, Ivan Kobyzev +1

    cs.CLcs.LGarXiv:2210.07558v22022
  57. A Closer Look at Deep Learning Heuristics: Learning rate restarts, Warmup and Distillation

    Akhilesh Gotmare, Nitish Shirish Keskar, Caiming Xiong +1

    cs.LGstat.MLarXiv:1810.13243v12018
  58. mPLUG-Owl3: Towards Long Image-Sequence Understanding in Multi-Modal Large Language Models

    Jiabo Ye, Haiyang Xu, Haowei Liu +6

    cs.CVcs.AIcs.CLarXiv:2408.04840v22024
  59. Open-TeleVision: Teleoperation with Immersive Active Visual Feedback

    Xuxin Cheng, Jialong Li, Shiqi Yang +2

    cs.ROcs.HCcs.LGarXiv:2407.01512v22024
  60. Deep Learning with Topological Signatures

    Christoph Hofer, Roland Kwitt, Marc Niethammer +1

    cs.CVcs.LGmath.ATarXiv:1707.04041v32017