Machine Learning
Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
12,301 to 12,360 of 20,193
Scaling Limits of Wide Neural Networks with Weight Sharing: Gaussian Process Behavior, Gradient Independence, and Neural Tangent Kernel Derivation
Greg Yang
cs.NEcond-mat.dis-nncs.LGarXiv:1902.04760v32019Inference Suboptimality in Variational Autoencoders
Chris Cremer, Xuechen Li, David Duvenaud
cs.LGstat.MLarXiv:1801.03558v32018Skew-Fit: State-Covering Self-Supervised Reinforcement Learning
Vitchyr H. Pong, Murtaza Dalal, Steven Lin +3
cs.LGcs.AIcs.ROarXiv:1903.03698v42019A vector-contraction inequality for Rademacher complexities
Andreas Maurer
cs.LGstat.MLarXiv:1605.00251v12016Trading Regret for Efficiency: Online Convex Optimization with Long Term Constraints
Mehrdad Mahdavi, Rong Jin, Tianbao Yang
cs.LGarXiv:1111.6082v32011Weakly-supervised Disentangling with Recurrent Transformations for 3D View Synthesis
Jimei Yang, Scott Reed, Ming-Hsuan Yang +1
cs.LGcs.AIcs.CVarXiv:1601.00706v12016Depthwise Separable Convolutions for Neural Machine Translation
Lukasz Kaiser, Aidan N. Gomez, Francois Chollet
cs.CLcs.LGarXiv:1706.03059v22017On the Robustness of ChatGPT: An Adversarial and Out-of-distribution Perspective
Jindong Wang, Xixu Hu, Wenxin Hou +10
cs.AIcs.CLcs.LGarXiv:2302.12095v52023IconQA: A New Benchmark for Abstract Diagram Understanding and Visual Language Reasoning
Pan Lu, Liang Qiu, Jiaqi Chen +6
cs.CVcs.AIcs.CLarXiv:2110.13214v42021Efficient Defenses Against Adversarial Attacks
Valentina Zantedeschi, Maria-Irina Nicolae, Ambrish Rawat
cs.LGarXiv:1707.06728v22017Multimodal Sentiment Analysis with Word-Level Fusion and Reinforcement Learning
Minghai Chen, Sen Wang, Paul Pu Liang +3
cs.LGcs.AIcs.CLarXiv:1802.00924v12018Deep Rewiring: Training very sparse deep networks
Guillaume Bellec, David Kappel, Wolfgang Maass +1
cs.NEcs.AIcs.DCarXiv:1711.05136v52017The Mechanics of n-Player Differentiable Games
David Balduzzi, Sebastien Racaniere, James Martens +3
cs.LGcs.GTcs.MAarXiv:1802.05642v22018A Survey of Graph Neural Networks for Social Recommender Systems
Kartik Sharma, Yeon-Chang Lee, Sivagami Nambi +4
cs.SIcs.IRcs.LGarXiv:2212.04481v32022Representer Point Selection for Explaining Deep Neural Networks
Chih-Kuan Yeh, Joon Sik Kim, Ian E. H. Yen +1
cs.LGstat.MLarXiv:1811.09720v12018SteganoGAN: High Capacity Image Steganography with GANs
Kevin Alex Zhang, Alfredo Cuesta-Infante, Lei Xu +1
cs.CVcs.LGcs.MMarXiv:1901.03892v22019Deep Multimodal Fusion by Channel Exchanging
Yikai Wang, Wenbing Huang, Fuchun Sun +3
cs.CVcs.LGarXiv:2011.05005v22020How (not) to Train your Generative Model: Scheduled Sampling, Likelihood, Adversary?
Ferenc Huszár
stat.MLcs.AIcs.ITarXiv:1511.05101v12015oLMpics -- On what Language Model Pre-training Captures
Alon Talmor, Yanai Elazar, Yoav Goldberg +1
cs.CLcs.AIcs.LGarXiv:1912.13283v22019Grounded Language Learning in a Simulated 3D World
Karl Moritz Hermann, Felix Hill, Simon Green +11
cs.CLcs.LGstat.MLarXiv:1706.06551v22017Towards an Appropriate Query, Key, and Value Computation for Knowledge Tracing
Youngduck Choi, Youngnam Lee, Junghyun Cho +6
cs.LGcs.AIcs.CYarXiv:2002.07033v52020Focal Sparse Convolutional Networks for 3D Object Detection
Yukang Chen, Yanwei Li, Xiangyu Zhang +2
cs.CVcs.LGarXiv:2204.12463v12022Evaluating and Aggregating Feature-based Model Explanations
Umang Bhatt, Adrian Weller, José M. F. Moura
cs.LGcs.AIcs.CYarXiv:2005.00631v12020Input complexity and out-of-distribution detection with likelihood-based generative models
Joan Serrà, David Álvarez, Vicenç Gómez +3
cs.LGstat.MLarXiv:1909.11480v32019Backpropagation through the Void: Optimizing control variates for black-box gradient estimation
Will Grathwohl, Dami Choi, Yuhuai Wu +2
cs.LGarXiv:1711.00123v32017Lift & Learn: Physics-informed machine learning for large-scale nonlinear dynamical systems
Elizabeth Qian, Boris Kramer, Benjamin Peherstorfer +1
math.NAcs.LGarXiv:1912.08177v52019BayesOpt: A Bayesian Optimization Library for Nonlinear Optimization, Experimental Design and Bandits
Ruben Martinez-Cantin
cs.LGarXiv:1405.7430v12014LLMLingua: Compressing Prompts for Accelerated Inference of Large Language Models
Huiqiang Jiang, Qianhui Wu, Chin-Yew Lin +2
cs.CLcs.LGarXiv:2310.05736v22023Salvaging Federated Learning by Local Adaptation
Tao Yu, Eugene Bagdasaryan, Vitaly Shmatikov
cs.LGcs.AIcs.DCarXiv:2002.04758v32020MetaFormer Baselines for Vision
Weihao Yu, Chenyang Si, Pan Zhou +5
cs.CVcs.AIcs.LGarXiv:2210.13452v42022CFA: Coupled-hypersphere-based Feature Adaptation for Target-Oriented Anomaly Localization
Sungwook Lee, Seunghyun Lee, Byung Cheol Song
cs.CVcs.LGarXiv:2206.04325v12022BadNL: Backdoor Attacks against NLP Models with Semantic-preserving Improvements
Xiaoyi Chen, Ahmed Salem, Dingfan Chen +5
cs.CRcs.LGarXiv:2006.01043v22020Optimal Clustering Framework for Hyperspectral Band Selection
Qi Wang, Fahong Zhang, Xuelong Li
eess.IVcs.LGstat.MLarXiv:1904.13036v12019Neural Descriptor Fields: SE(3)-Equivariant Object Representations for Manipulation
Anthony Simeonov, Yilun Du, Andrea Tagliasacchi +4
cs.ROcs.AIcs.CVarXiv:2112.05124v12021Sparse Sequence-to-Sequence Models
Ben Peters, Vlad Niculae, André F. T. Martins
cs.CLcs.LGarXiv:1905.05702v22019Detecting Cyberattacks in Industrial Control Systems Using Convolutional Neural Networks
Moshe Kravchik, Asaf Shabtai
cs.CRcs.LGarXiv:1806.08110v22018SafetyNet: Detecting and Rejecting Adversarial Examples Robustly
Jiajun Lu, Theerasit Issaranon, David Forsyth
cs.CVcs.LGarXiv:1704.00103v22017Automated Fact-Checking for Assisting Human Fact-Checkers
Preslav Nakov, David Corney, Maram Hasanain +6
cs.AIcs.CLcs.CRarXiv:2103.07769v22021Interpretable Learning for Self-Driving Cars by Visualizing Causal Attention
Jinkyu Kim, John Canny
cs.CVcs.LGarXiv:1703.10631v12017CausalGAN: Learning Causal Implicit Generative Models with Adversarial Training
Murat Kocaoglu, Christopher Snyder, Alexandros G. Dimakis +1
cs.LGcs.AIcs.ITarXiv:1709.02023v22017Chip-Chat: Challenges and Opportunities in Conversational Hardware Design
Jason Blocklove, Siddharth Garg, Ramesh Karri +1
cs.LGcs.ARcs.PLarXiv:2305.13243v22023Object Contour Detection with a Fully Convolutional Encoder-Decoder Network
Jimei Yang, Brian Price, Scott Cohen +2
cs.CVcs.LGarXiv:1603.04530v12016EDICT: Exact Diffusion Inversion via Coupled Transformations
Bram Wallace, Akash Gokul, Nikhil Naik
cs.CVcs.AIcs.LGarXiv:2211.12446v22022Bayesian Posterior Sampling via Stochastic Gradient Fisher Scoring
Sungjin Ahn, Anoop Korattikara, Max Welling
cs.LGstat.COstat.MLarXiv:1206.6380v12012Stochastic Controlled Averaging for Federated Learning with Communication Compression
Xinmeng Huang, Ping Li, Xiaoyun Li
math.OCcs.DCcs.LGarXiv:2308.08165v22023Unsupervised and Semi-supervised Anomaly Detection with LSTM Neural Networks
Tolga Ergen, Ali Hassan Mirza, Suleyman Serdar Kozat
eess.SPcs.LGstat.MLarXiv:1710.09207v12017Toward Robustness against Label Noise in Training Deep Discriminative Neural Networks
Arash Vahdat
cs.LGstat.MLarXiv:1706.00038v22017Deep Learning-enabled Virtual Histological Staining of Biological Samples
Bijie Bai, Xilin Yang, Yuzhu Li +3
physics.med-phcs.CVcs.LGarXiv:2211.06822v12022Deep Learning for Time Series Forecasting: The Electric Load Case
Alberto Gasparin, Slobodan Lukovic, Cesare Alippi
cs.LGstat.MLarXiv:1907.09207v12019Modern WLAN Fingerprinting Indoor Positioning Methods and Deployment Challenges
Ali Khalajmehrabadi, Nikolaos Gatsis, David Akopian
cs.NIcs.LGarXiv:1610.05424v12016Maximizing acquisition functions for Bayesian optimization
James T. Wilson, Frank Hutter, Marc Peter Deisenroth
stat.MLcs.LGarXiv:1805.10196v22018Globally Optimal Gradient Descent for a ConvNet with Gaussian Inputs
Alon Brutzkus, Amir Globerson
cs.LGmath.OCstat.MLarXiv:1702.07966v12017A physics-informed variational DeepONet for predicting the crack path in brittle materials
Somdatta Goswami, Minglang Yin, Yue Yu +1
cs.LGmath.NAarXiv:2108.06905v22021On the Inductive Bias of Neural Tangent Kernels
Alberto Bietti, Julien Mairal
stat.MLcs.LGarXiv:1905.12173v22019GMNN: Graph Markov Neural Networks
Meng Qu, Yoshua Bengio, Jian Tang
cs.LGcs.SIstat.MLarXiv:1905.06214v32019DyLoRA: Parameter Efficient Tuning of Pre-trained Models using Dynamic Search-Free Low-Rank Adaptation
Mojtaba Valipour, Mehdi Rezagholizadeh, Ivan Kobyzev +1
cs.CLcs.LGarXiv:2210.07558v22022A Closer Look at Deep Learning Heuristics: Learning rate restarts, Warmup and Distillation
Akhilesh Gotmare, Nitish Shirish Keskar, Caiming Xiong +1
cs.LGstat.MLarXiv:1810.13243v12018mPLUG-Owl3: Towards Long Image-Sequence Understanding in Multi-Modal Large Language Models
Jiabo Ye, Haiyang Xu, Haowei Liu +6
cs.CVcs.AIcs.CLarXiv:2408.04840v22024Open-TeleVision: Teleoperation with Immersive Active Visual Feedback
Xuxin Cheng, Jialong Li, Shiqi Yang +2
cs.ROcs.HCcs.LGarXiv:2407.01512v22024Deep Learning with Topological Signatures
Christoph Hofer, Roland Kwitt, Marc Niethammer +1
cs.CVcs.LGmath.ATarXiv:1707.04041v32017