Machine Learning

Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

18,301 to 18,360 of 20,454

  1. Turning the TIDE: Cross-Architecture Distillation for Diffusion Large Language Models

    Gongbo Zhang, Wen Wang, Ye Tian +1

    cs.CLcs.AIcs.LGarXiv:2604.26951v12026
  2. ML-Leaks: Model and Data Independent Membership Inference Attacks and Defenses on Machine Learning Models

    Ahmed Salem, Yang Zhang, Mathias Humbert +3

    cs.CRcs.AIcs.LGarXiv:1806.01246v22018
  3. DKN: Deep Knowledge-Aware Network for News Recommendation

    Hongwei Wang, Fuzheng Zhang, Xing Xie +1

    stat.MLcs.LGarXiv:1801.08284v22018
  4. Captum: A unified and generic model interpretability library for PyTorch

    Narine Kokhlikyan, Vivek Miglani, Miguel Martin +8

    cs.LGcs.AIstat.MLarXiv:2009.07896v12020
  5. Manopt, a Matlab toolbox for optimization on manifolds

    Nicolas Boumal, Bamdev Mishra, P. -A. Absil +1

    cs.MScs.LGmath.OCarXiv:1308.5200v12013
  6. Deep Learning for IoT Big Data and Streaming Analytics: A Survey

    Mehdi Mohammadi, Ala Al-Fuqaha, Sameh Sorour +1

    cs.NIcs.DBcs.LGarXiv:1712.04301v22017
  7. End-to-End Attention-based Large Vocabulary Speech Recognition

    Dzmitry Bahdanau, Jan Chorowski, Dmitriy Serdyuk +2

    cs.CLcs.AIcs.LGarXiv:1508.04395v22015
  8. Neural Module Networks

    Jacob Andreas, Marcus Rohrbach, Trevor Darrell +1

    cs.CVcs.CLcs.LGarXiv:1511.02799v42015
  9. Object-Centric Learning with Slot Attention

    Francesco Locatello, Dirk Weissenborn, Thomas Unterthiner +5

    cs.LGcs.CVstat.MLarXiv:2006.15055v22020
  10. Prototypical Contrastive Learning of Unsupervised Representations

    Junnan Li, Pan Zhou, Caiming Xiong +1

    cs.CVcs.LGarXiv:2005.04966v52020
  11. Membership Inference Attacks From First Principles

    Nicholas Carlini, Steve Chien, Milad Nasr +3

    cs.CRcs.LGarXiv:2112.03570v22021
  12. B-PINNs: Bayesian Physics-Informed Neural Networks for Forward and Inverse PDE Problems with Noisy Data

    Liu Yang, Xuhui Meng, George Em Karniadakis

    stat.MLcs.LGarXiv:2003.06097v12020
  13. The LAMBADA dataset: Word prediction requiring a broad discourse context

    Denis Paperno, Germán Kruszewski, Angeliki Lazaridou +6

    cs.CLcs.AIcs.LGarXiv:1606.06031v12016
  14. A General Language Assistant as a Laboratory for Alignment

    Amanda Askell, Yuntao Bai, Anna Chen +19

    cs.CLcs.LGarXiv:2112.00861v32021
  15. Interpretability in the Wild: a Circuit for Indirect Object Identification in GPT-2 small

    Kevin Wang, Alexandre Variengien, Arthur Conmy +2

    cs.LGcs.AIcs.CLarXiv:2211.00593v12022
  16. FINN: A Framework for Fast, Scalable Binarized Neural Network Inference

    Yaman Umuroglu, Nicholas J. Fraser, Giulio Gambardella +4

    cs.CVcs.ARcs.LGarXiv:1612.07119v12016
  17. Audio Adversarial Examples: Targeted Attacks on Speech-to-Text

    Nicholas Carlini, David Wagner

    cs.LGcs.AIcs.CRarXiv:1801.01944v22018
  18. A disciplined approach to neural network hyper-parameters: Part 1 -- learning rate, batch size, momentum, and weight decay

    Leslie N. Smith

    cs.LGcs.CVcs.NEarXiv:1803.09820v22018
  19. Exploiting Shared Representations for Personalized Federated Learning

    Liam Collins, Hamed Hassani, Aryan Mokhtari +1

    cs.LGmath.OCarXiv:2102.07078v32021
  20. Carbon Emissions and Large Neural Network Training

    David Patterson, Joseph Gonzalez, Quoc Le +6

    cs.LGcs.CYarXiv:2104.10350v32021
  21. Bottleneck Transformers for Visual Recognition

    Aravind Srinivas, Tsung-Yi Lin, Niki Parmar +3

    cs.CVcs.AIcs.LGarXiv:2101.11605v22021
  22. FedMD: Heterogenous Federated Learning via Model Distillation

    Daliang Li, Junpu Wang

    cs.LGstat.MLarXiv:1910.03581v12019
  23. BEGAN: Boundary Equilibrium Generative Adversarial Networks

    David Berthelot, Thomas Schumm, Luke Metz

    cs.LGstat.MLarXiv:1703.10717v42017
  24. Generating Images with Perceptual Similarity Metrics based on Deep Networks

    Alexey Dosovitskiy, Thomas Brox

    cs.LGcs.CVcs.NEarXiv:1602.02644v22016
  25. Mean teachers are better role models: Weight-averaged consistency targets improve semi-supervised deep learning results

    Antti Tarvainen, Harri Valpola

    cs.NEcs.LGstat.MLarXiv:1703.01780v62017
  26. Online Self-Calibration Against Hallucination in Vision-Language Models

    Minghui Chen, Chenxu Yang, Hengjie Zhu +3

    cs.CVcs.LGarXiv:2605.00323v12026
  27. Supersizing Self-supervision: Learning to Grasp from 50K Tries and 700 Robot Hours

    Lerrel Pinto, Abhinav Gupta

    cs.LGcs.CVcs.ROarXiv:1509.06825v12015
  28. SplAttN: Bridging 2D and 3D with Gaussian Soft Splatting and Attention for Point Cloud Completion

    Zhaoyang Li, Zhichao You, Tianrui Li

    cs.CVcs.LGarXiv:2605.01466v22026
  29. Learning to Diagnose with LSTM Recurrent Neural Networks

    Zachary C. Lipton, David C. Kale, Charles Elkan +1

    cs.LGarXiv:1511.03677v72015
  30. data2vec: A General Framework for Self-supervised Learning in Speech, Vision and Language

    Alexei Baevski, Wei-Ning Hsu, Qiantong Xu +3

    cs.LGarXiv:2202.03555v32022
  31. BlenderRAG: High-Fidelity 3D Object Generation via Retrieval-Augmented Code Synthesis

    Massimo Rondelli, Francesco Pivi, Maurizio Gabbrielli

    cs.CVcs.AIcs.GRarXiv:2605.00632v12026
  32. A Comparative Study of Efficient Initialization Methods for the K-Means Clustering Algorithm

    M. Emre Celebi, Hassan A. Kingravi, Patricio A. Vela

    cs.LGcs.CVarXiv:1209.1960v12012
  33. HiDDeN: Hiding Data With Deep Networks

    Jiren Zhu, Russell Kaplan, Justin Johnson +1

    cs.CVcs.LGarXiv:1807.09937v12018
  34. Towards Customized Multimodal Role-Play

    Chao Tang, Jianzong Wu, Qingyu Shi +5

    cs.LGarXiv:2605.08129v12026
  35. Devign: Effective Vulnerability Identification by Learning Comprehensive Program Semantics via Graph Neural Networks

    Yaqin Zhou, Shangqing Liu, Jingkai Siow +2

    cs.SEcs.CRcs.LGarXiv:1909.03496v12019
  36. A Minimalist Approach to Offline Reinforcement Learning

    Scott Fujimoto, Shixiang Shane Gu

    cs.LGcs.AIstat.MLarXiv:2106.06860v22021
  37. Dynamic Few-Shot Visual Learning without Forgetting

    Spyros Gidaris, Nikos Komodakis

    cs.CVcs.LGarXiv:1804.09458v12018
  38. Federated Learning Based on Dynamic Regularization

    Durmus Alp Emre Acar, Yue Zhao, Ramon Matas Navarro +3

    cs.LGcs.DCarXiv:2111.04263v22021
  39. Ask Me Anything: Dynamic Memory Networks for Natural Language Processing

    Ankit Kumar, Ozan Irsoy, Peter Ondruska +6

    cs.CLcs.LGcs.NEarXiv:1506.07285v52015
  40. Theano: A Python framework for fast computation of mathematical expressions

    The Theano Development Team, Rami Al-Rfou, Guillaume Alain +110

    cs.SCcs.LGcs.MSarXiv:1605.02688v12016
  41. Compressing Deep Convolutional Networks using Vector Quantization

    Yunchao Gong, Liu Liu, Ming Yang +1

    cs.CVcs.LGcs.NEarXiv:1412.6115v12014
  42. A note on the evaluation of generative models

    Lucas Theis, Aäron van den Oord, Matthias Bethge

    stat.MLcs.LGarXiv:1511.01844v32015
  43. Out-of-Distribution Generalization via Risk Extrapolation (REx)

    David Krueger, Ethan Caballero, Joern-Henrik Jacobsen +5

    cs.LGcs.AIcs.NEarXiv:2003.00688v52020
  44. Tensor field networks: Rotation- and translation-equivariant neural networks for 3D point clouds

    Nathaniel Thomas, Tess Smidt, Steven Kearnes +4

    cs.LGcs.AIcs.CVarXiv:1802.08219v32018
  45. DenseCap: Fully Convolutional Localization Networks for Dense Captioning

    Justin Johnson, Andrej Karpathy, Li Fei-Fei

    cs.CVcs.LGarXiv:1511.07571v12015
  46. DynamicViT: Efficient Vision Transformers with Dynamic Token Sparsification

    Yongming Rao, Wenliang Zhao, Benlin Liu +3

    cs.CVcs.AIcs.LGarXiv:2106.02034v22021
  47. UCTransNet: Rethinking the Skip Connections in U-Net from a Channel-wise Perspective with Transformer

    Haonan Wang, Peng Cao, Jiaqi Wang +1

    cs.CVcs.LGeess.IVarXiv:2109.04335v32021
  48. A Review of Feature Selection Methods Based on Mutual Information

    Jorge R. Vergara, Pablo A. Estévez

    cs.LGstat.MLarXiv:1509.07577v12015
  49. Spatio-Temporal LSTM with Trust Gates for 3D Human Action Recognition

    Jun Liu, Amir Shahroudy, Dong Xu +1

    cs.CVcs.AIcs.LGarXiv:1607.07043v12016
  50. Virtual Worlds as Proxy for Multi-Object Tracking Analysis

    Adrien Gaidon, Qiao Wang, Yohann Cabon +1

    cs.CVcs.LGcs.NEarXiv:1605.06457v12016
  51. Causal inference using invariant prediction: identification and confidence intervals

    Jonas Peters, Peter Bühlmann, Nicolai Meinshausen

    stat.MEcs.LGarXiv:1501.01332v32015
  52. Benchmarking Graph Neural Networks

    Vijay Prakash Dwivedi, Chaitanya K. Joshi, Anh Tuan Luu +3

    cs.LGstat.MLarXiv:2003.00982v52020
  53. Multi-Label Image Recognition with Graph Convolutional Networks

    Zhao-Min Chen, Xiu-Shen Wei, Peng Wang +1

    cs.CVcs.LGarXiv:1904.03582v12019
  54. Compressing Neural Networks with the Hashing Trick

    Wenlin Chen, James T. Wilson, Stephen Tyree +2

    cs.LGcs.NEarXiv:1504.04788v12015
  55. A Lip Sync Expert Is All You Need for Speech to Lip Generation In The Wild

    K R Prajwal, Rudrabha Mukhopadhyay, Vinay Namboodiri +1

    cs.CVcs.LGcs.SDarXiv:2008.10010v12020
  56. Holographic MIMO Surfaces for 6G Wireless Networks: Opportunities, Challenges, and Trends

    Chongwen Huang, Sha Hu, George C. Alexandropoulos +5

    cs.ITcs.LGarXiv:1911.12296v32019
  57. Unmasking Clever Hans Predictors and Assessing What Machines Really Learn

    Sebastian Lapuschkin, Stephan Wäldchen, Alexander Binder +3

    cs.AIcs.CVcs.LGarXiv:1902.10178v12019
  58. Plug and Play Language Models: A Simple Approach to Controlled Text Generation

    Sumanth Dathathri, Andrea Madotto, Janice Lan +5

    cs.CLcs.AIcs.LGarXiv:1912.02164v42019
  59. Holographic Embeddings of Knowledge Graphs

    Maximilian Nickel, Lorenzo Rosasco, Tomaso Poggio

    cs.AIcs.LGstat.MLarXiv:1510.04935v22015
  60. RippleNet: Propagating User Preferences on the Knowledge Graph for Recommender Systems

    Hongwei Wang, Fuzheng Zhang, Jialin Wang +4

    cs.IRcs.LGstat.MLarXiv:1803.03467v42018