Machine Learning
Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
18,301 to 18,360 of 20,454
Turning the TIDE: Cross-Architecture Distillation for Diffusion Large Language Models
Gongbo Zhang, Wen Wang, Ye Tian +1
cs.CLcs.AIcs.LGarXiv:2604.26951v12026ML-Leaks: Model and Data Independent Membership Inference Attacks and Defenses on Machine Learning Models
Ahmed Salem, Yang Zhang, Mathias Humbert +3
cs.CRcs.AIcs.LGarXiv:1806.01246v22018DKN: Deep Knowledge-Aware Network for News Recommendation
Hongwei Wang, Fuzheng Zhang, Xing Xie +1
stat.MLcs.LGarXiv:1801.08284v22018Captum: A unified and generic model interpretability library for PyTorch
Narine Kokhlikyan, Vivek Miglani, Miguel Martin +8
cs.LGcs.AIstat.MLarXiv:2009.07896v12020Manopt, a Matlab toolbox for optimization on manifolds
Nicolas Boumal, Bamdev Mishra, P. -A. Absil +1
cs.MScs.LGmath.OCarXiv:1308.5200v12013Deep Learning for IoT Big Data and Streaming Analytics: A Survey
Mehdi Mohammadi, Ala Al-Fuqaha, Sameh Sorour +1
cs.NIcs.DBcs.LGarXiv:1712.04301v22017End-to-End Attention-based Large Vocabulary Speech Recognition
Dzmitry Bahdanau, Jan Chorowski, Dmitriy Serdyuk +2
cs.CLcs.AIcs.LGarXiv:1508.04395v22015Neural Module Networks
Jacob Andreas, Marcus Rohrbach, Trevor Darrell +1
cs.CVcs.CLcs.LGarXiv:1511.02799v42015Object-Centric Learning with Slot Attention
Francesco Locatello, Dirk Weissenborn, Thomas Unterthiner +5
cs.LGcs.CVstat.MLarXiv:2006.15055v22020Prototypical Contrastive Learning of Unsupervised Representations
Junnan Li, Pan Zhou, Caiming Xiong +1
cs.CVcs.LGarXiv:2005.04966v52020Membership Inference Attacks From First Principles
Nicholas Carlini, Steve Chien, Milad Nasr +3
cs.CRcs.LGarXiv:2112.03570v22021B-PINNs: Bayesian Physics-Informed Neural Networks for Forward and Inverse PDE Problems with Noisy Data
Liu Yang, Xuhui Meng, George Em Karniadakis
stat.MLcs.LGarXiv:2003.06097v12020The LAMBADA dataset: Word prediction requiring a broad discourse context
Denis Paperno, Germán Kruszewski, Angeliki Lazaridou +6
cs.CLcs.AIcs.LGarXiv:1606.06031v12016A General Language Assistant as a Laboratory for Alignment
Amanda Askell, Yuntao Bai, Anna Chen +19
cs.CLcs.LGarXiv:2112.00861v32021Interpretability in the Wild: a Circuit for Indirect Object Identification in GPT-2 small
Kevin Wang, Alexandre Variengien, Arthur Conmy +2
cs.LGcs.AIcs.CLarXiv:2211.00593v12022FINN: A Framework for Fast, Scalable Binarized Neural Network Inference
Yaman Umuroglu, Nicholas J. Fraser, Giulio Gambardella +4
cs.CVcs.ARcs.LGarXiv:1612.07119v12016Audio Adversarial Examples: Targeted Attacks on Speech-to-Text
Nicholas Carlini, David Wagner
cs.LGcs.AIcs.CRarXiv:1801.01944v22018A disciplined approach to neural network hyper-parameters: Part 1 -- learning rate, batch size, momentum, and weight decay
Leslie N. Smith
cs.LGcs.CVcs.NEarXiv:1803.09820v22018Exploiting Shared Representations for Personalized Federated Learning
Liam Collins, Hamed Hassani, Aryan Mokhtari +1
cs.LGmath.OCarXiv:2102.07078v32021Carbon Emissions and Large Neural Network Training
David Patterson, Joseph Gonzalez, Quoc Le +6
cs.LGcs.CYarXiv:2104.10350v32021Bottleneck Transformers for Visual Recognition
Aravind Srinivas, Tsung-Yi Lin, Niki Parmar +3
cs.CVcs.AIcs.LGarXiv:2101.11605v22021FedMD: Heterogenous Federated Learning via Model Distillation
Daliang Li, Junpu Wang
cs.LGstat.MLarXiv:1910.03581v12019BEGAN: Boundary Equilibrium Generative Adversarial Networks
David Berthelot, Thomas Schumm, Luke Metz
cs.LGstat.MLarXiv:1703.10717v42017Generating Images with Perceptual Similarity Metrics based on Deep Networks
Alexey Dosovitskiy, Thomas Brox
cs.LGcs.CVcs.NEarXiv:1602.02644v22016Mean teachers are better role models: Weight-averaged consistency targets improve semi-supervised deep learning results
Antti Tarvainen, Harri Valpola
cs.NEcs.LGstat.MLarXiv:1703.01780v62017Online Self-Calibration Against Hallucination in Vision-Language Models
Minghui Chen, Chenxu Yang, Hengjie Zhu +3
cs.CVcs.LGarXiv:2605.00323v12026Supersizing Self-supervision: Learning to Grasp from 50K Tries and 700 Robot Hours
Lerrel Pinto, Abhinav Gupta
cs.LGcs.CVcs.ROarXiv:1509.06825v12015SplAttN: Bridging 2D and 3D with Gaussian Soft Splatting and Attention for Point Cloud Completion
Zhaoyang Li, Zhichao You, Tianrui Li
cs.CVcs.LGarXiv:2605.01466v22026Learning to Diagnose with LSTM Recurrent Neural Networks
Zachary C. Lipton, David C. Kale, Charles Elkan +1
cs.LGarXiv:1511.03677v72015data2vec: A General Framework for Self-supervised Learning in Speech, Vision and Language
Alexei Baevski, Wei-Ning Hsu, Qiantong Xu +3
cs.LGarXiv:2202.03555v32022BlenderRAG: High-Fidelity 3D Object Generation via Retrieval-Augmented Code Synthesis
Massimo Rondelli, Francesco Pivi, Maurizio Gabbrielli
cs.CVcs.AIcs.GRarXiv:2605.00632v12026A Comparative Study of Efficient Initialization Methods for the K-Means Clustering Algorithm
M. Emre Celebi, Hassan A. Kingravi, Patricio A. Vela
cs.LGcs.CVarXiv:1209.1960v12012HiDDeN: Hiding Data With Deep Networks
Jiren Zhu, Russell Kaplan, Justin Johnson +1
cs.CVcs.LGarXiv:1807.09937v12018Towards Customized Multimodal Role-Play
Chao Tang, Jianzong Wu, Qingyu Shi +5
cs.LGarXiv:2605.08129v12026Devign: Effective Vulnerability Identification by Learning Comprehensive Program Semantics via Graph Neural Networks
Yaqin Zhou, Shangqing Liu, Jingkai Siow +2
cs.SEcs.CRcs.LGarXiv:1909.03496v12019A Minimalist Approach to Offline Reinforcement Learning
Scott Fujimoto, Shixiang Shane Gu
cs.LGcs.AIstat.MLarXiv:2106.06860v22021Dynamic Few-Shot Visual Learning without Forgetting
Spyros Gidaris, Nikos Komodakis
cs.CVcs.LGarXiv:1804.09458v12018Federated Learning Based on Dynamic Regularization
Durmus Alp Emre Acar, Yue Zhao, Ramon Matas Navarro +3
cs.LGcs.DCarXiv:2111.04263v22021Ask Me Anything: Dynamic Memory Networks for Natural Language Processing
Ankit Kumar, Ozan Irsoy, Peter Ondruska +6
cs.CLcs.LGcs.NEarXiv:1506.07285v52015Theano: A Python framework for fast computation of mathematical expressions
The Theano Development Team, Rami Al-Rfou, Guillaume Alain +110
cs.SCcs.LGcs.MSarXiv:1605.02688v12016Compressing Deep Convolutional Networks using Vector Quantization
Yunchao Gong, Liu Liu, Ming Yang +1
cs.CVcs.LGcs.NEarXiv:1412.6115v12014A note on the evaluation of generative models
Lucas Theis, Aäron van den Oord, Matthias Bethge
stat.MLcs.LGarXiv:1511.01844v32015Out-of-Distribution Generalization via Risk Extrapolation (REx)
David Krueger, Ethan Caballero, Joern-Henrik Jacobsen +5
cs.LGcs.AIcs.NEarXiv:2003.00688v52020Tensor field networks: Rotation- and translation-equivariant neural networks for 3D point clouds
Nathaniel Thomas, Tess Smidt, Steven Kearnes +4
cs.LGcs.AIcs.CVarXiv:1802.08219v32018DenseCap: Fully Convolutional Localization Networks for Dense Captioning
Justin Johnson, Andrej Karpathy, Li Fei-Fei
cs.CVcs.LGarXiv:1511.07571v12015DynamicViT: Efficient Vision Transformers with Dynamic Token Sparsification
Yongming Rao, Wenliang Zhao, Benlin Liu +3
cs.CVcs.AIcs.LGarXiv:2106.02034v22021UCTransNet: Rethinking the Skip Connections in U-Net from a Channel-wise Perspective with Transformer
Haonan Wang, Peng Cao, Jiaqi Wang +1
cs.CVcs.LGeess.IVarXiv:2109.04335v32021A Review of Feature Selection Methods Based on Mutual Information
Jorge R. Vergara, Pablo A. Estévez
cs.LGstat.MLarXiv:1509.07577v12015Spatio-Temporal LSTM with Trust Gates for 3D Human Action Recognition
Jun Liu, Amir Shahroudy, Dong Xu +1
cs.CVcs.AIcs.LGarXiv:1607.07043v12016Virtual Worlds as Proxy for Multi-Object Tracking Analysis
Adrien Gaidon, Qiao Wang, Yohann Cabon +1
cs.CVcs.LGcs.NEarXiv:1605.06457v12016Causal inference using invariant prediction: identification and confidence intervals
Jonas Peters, Peter Bühlmann, Nicolai Meinshausen
stat.MEcs.LGarXiv:1501.01332v32015Benchmarking Graph Neural Networks
Vijay Prakash Dwivedi, Chaitanya K. Joshi, Anh Tuan Luu +3
cs.LGstat.MLarXiv:2003.00982v52020Multi-Label Image Recognition with Graph Convolutional Networks
Zhao-Min Chen, Xiu-Shen Wei, Peng Wang +1
cs.CVcs.LGarXiv:1904.03582v12019Compressing Neural Networks with the Hashing Trick
Wenlin Chen, James T. Wilson, Stephen Tyree +2
cs.LGcs.NEarXiv:1504.04788v12015A Lip Sync Expert Is All You Need for Speech to Lip Generation In The Wild
K R Prajwal, Rudrabha Mukhopadhyay, Vinay Namboodiri +1
cs.CVcs.LGcs.SDarXiv:2008.10010v12020Holographic MIMO Surfaces for 6G Wireless Networks: Opportunities, Challenges, and Trends
Chongwen Huang, Sha Hu, George C. Alexandropoulos +5
cs.ITcs.LGarXiv:1911.12296v32019Unmasking Clever Hans Predictors and Assessing What Machines Really Learn
Sebastian Lapuschkin, Stephan Wäldchen, Alexander Binder +3
cs.AIcs.CVcs.LGarXiv:1902.10178v12019Plug and Play Language Models: A Simple Approach to Controlled Text Generation
Sumanth Dathathri, Andrea Madotto, Janice Lan +5
cs.CLcs.AIcs.LGarXiv:1912.02164v42019Holographic Embeddings of Knowledge Graphs
Maximilian Nickel, Lorenzo Rosasco, Tomaso Poggio
cs.AIcs.LGstat.MLarXiv:1510.04935v22015RippleNet: Propagating User Preferences on the Knowledge Graph for Recommender Systems
Hongwei Wang, Fuzheng Zhang, Jialin Wang +4
cs.IRcs.LGstat.MLarXiv:1803.03467v42018