Machine Learning
Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
13,141 to 13,200 of 20,454
Layer-Dependent Importance Sampling for Training Deep and Large Graph Convolutional Networks
Difan Zou, Ziniu Hu, Yewen Wang +3
cs.LGcs.SIstat.MLarXiv:1911.07323v12019Locally Weighted Ensemble Clustering
Dong Huang, Chang-Dong Wang, Jian-Huang Lai
cs.LGarXiv:1605.05011v32016OmniH2O: Universal and Dexterous Human-to-Humanoid Whole-Body Teleoperation and Learning
Tairan He, Zhengyi Luo, Xialin He +6
cs.ROcs.CVcs.LGarXiv:2406.08858v12024CVEfixes: Automated Collection of Vulnerabilities and Their Fixes from Open-Source Software
Guru Prasad Bhandari, Amara Naseer, Leon Moonen
cs.SEcs.AIcs.CRarXiv:2107.08760v12021Why M Heads are Better than One: Training a Diverse Ensemble of Deep Networks
Stefan Lee, Senthil Purushwalkam, Michael Cogswell +2
cs.CVcs.LGcs.NEarXiv:1511.06314v12015The Role of Permutation Invariance in Linear Mode Connectivity of Neural Networks
Rahim Entezari, Hanie Sedghi, Olga Saukh +1
cs.LGarXiv:2110.06296v22021Emergent Complexity and Zero-shot Transfer via Unsupervised Environment Design
Michael Dennis, Natasha Jaques, Eugene Vinitsky +4
cs.LGcs.AIcs.MAarXiv:2012.02096v22020Distributed Online Optimization in Dynamic Environments Using Mirror Descent
Shahin Shahrampour, Ali Jadbabaie
math.OCcs.DCcs.LGarXiv:1609.02845v12016Structured sparsity through convex optimization
Francis Bach, Rodolphe Jenatton, Julien Mairal +1
cs.LGstat.MLarXiv:1109.2397v22011Physics-Informed Neural Networks for Power Systems
George S. Misyris, Andreas Venzke, Spyros Chatzivasileiadis
eess.SYcs.LGeess.SParXiv:1911.03737v32019Machine Learning-Based Heart Disease Diagnosis: A Systematic Literature Review
Md Manjurul Ahsan, Zahed Siddique
cs.LGarXiv:2112.06459v12021Fuzz4All: Universal Fuzzing with Large Language Models
Chunqiu Steven Xia, Matteo Paltenghi, Jia Le Tian +2
cs.SEcs.LGarXiv:2308.04748v32023Using Fast Weights to Attend to the Recent Past
Jimmy Ba, Geoffrey Hinton, Volodymyr Mnih +2
stat.MLcs.LGcs.NEarXiv:1610.06258v32016Time-MoE: Billion-Scale Time Series Foundation Models with Mixture of Experts
Xiaoming Shi, Shiyu Wang, Yuqi Nie +4
cs.LGcs.AIarXiv:2409.16040v42024Graph-to-Sequence Learning using Gated Graph Neural Networks
Daniel Beck, Gholamreza Haffari, Trevor Cohn
cs.CLcs.LGarXiv:1806.09835v12018Rethinking Vision Transformers for MobileNet Size and Speed
Yanyu Li, Ju Hu, Yang Wen +5
cs.CVcs.AIcs.LGarXiv:2212.08059v22022Continuous Inverse Optimal Control with Locally Optimal Examples
Sergey Levine, Vladlen Koltun
cs.LGcs.AIstat.MLarXiv:1206.4617v12012Differentially Private Learning Needs Better Features (or Much More Data)
Florian Tramèr, Dan Boneh
cs.LGcs.CRstat.MLarXiv:2011.11660v32020Asymptotically Exact, Embarrassingly Parallel MCMC
Willie Neiswanger, Chong Wang, Eric Xing
stat.MLcs.DCcs.LGarXiv:1311.4780v22013Neural Radiance Flow for 4D View Synthesis and Video Processing
Yilun Du, Yinan Zhang, Hong-Xing Yu +2
cs.CVcs.LGcs.ROarXiv:2012.09790v22020Interpretable End-to-end Urban Autonomous Driving with Latent Deep Reinforcement Learning
Jianyu Chen, Shengbo Eben Li, Masayoshi Tomizuka
cs.ROcs.CVcs.LGarXiv:2001.08726v32020FengWu: Pushing the Skillful Global Medium-range Weather Forecast beyond 10 Days Lead
Kang Chen, Tao Han, Junchao Gong +11
cs.AIcs.LGphysics.ao-pharXiv:2304.02948v12023Edge YOLO: Real-Time Intelligent Object Detection System Based on Edge-Cloud Cooperation in Autonomous Vehicles
Siyuan Liang, Hao Wu
cs.CVcs.LGeess.SParXiv:2205.14942v12022Learning From Multiple Experts: Self-paced Knowledge Distillation for Long-tailed Classification
Liuyu Xiang, Guiguang Ding, Jungong Han
cs.CVcs.LGstat.MLarXiv:2001.01536v32020Sobolev Training for Neural Networks
Wojciech Marian Czarnecki, Simon Osindero, Max Jaderberg +2
cs.LGarXiv:1706.04859v32017Deep learning: a statistical viewpoint
Peter L. Bartlett, Andrea Montanari, Alexander Rakhlin
math.STcs.LGstat.MLarXiv:2103.09177v12021On the Convergence of Stochastic Gradient Descent with Adaptive Stepsizes
Xiaoyu Li, Francesco Orabona
stat.MLcs.LGmath.OCarXiv:1805.08114v32018WikiHow: A Large Scale Text Summarization Dataset
Mahnaz Koupaee, William Yang Wang
cs.CLcs.IRcs.LGarXiv:1810.09305v12018What graph neural networks cannot learn: depth vs width
Andreas Loukas
cs.LGstat.MLarXiv:1907.03199v22019AAU-net: An Adaptive Attention U-net for Breast Lesions Segmentation in Ultrasound Images
Gongping Chen, Yu Dai, Jianxun Zhang +1
eess.IVcs.CVcs.LGarXiv:2204.12077v32022A Primer on Zeroth-Order Optimization in Signal Processing and Machine Learning
Sijia Liu, Pin-Yu Chen, Bhavya Kailkhura +3
cs.LGeess.SPstat.MLarXiv:2006.06224v22020An exact mapping between the Variational Renormalization Group and Deep Learning
Pankaj Mehta, David J. Schwab
stat.MLcond-mat.stat-mechcs.LGarXiv:1410.3831v12014ASD-DiagNet: A hybrid learning approach for detection of Autism Spectrum Disorder using fMRI data
Taban Eslami, Vahid Mirjalili, Alvis Fong +2
cs.LGeess.IVstat.MLarXiv:1904.07577v12019Dual Discriminator Generative Adversarial Nets
Tu Dinh Nguyen, Trung Le, Hung Vu +1
cs.LGstat.MLarXiv:1709.03831v12017Weisfeiler and Lehman Go Topological: Message Passing Simplicial Networks
Cristian Bodnar, Fabrizio Frasca, Yu Guang Wang +4
cs.LGcs.SIarXiv:2103.03212v22021Label-Free Concept Bottleneck Models
Tuomas Oikarinen, Subhro Das, Lam M. Nguyen +1
cs.LGcs.CVarXiv:2304.06129v22023Discovering Discrete Latent Topics with Neural Variational Inference
Yishu Miao, Edward Grefenstette, Phil Blunsom
cs.CLcs.AIcs.IRarXiv:1706.00359v22017Physics informed deep learning for computational elastodynamics without labeled data
Chengping Rao, Hao Sun, Yang Liu
math.NAcs.AIcs.CEarXiv:2006.08472v12020Machine Learning on Graphs: A Model and Comprehensive Taxonomy
Ines Chami, Sami Abu-El-Haija, Bryan Perozzi +2
cs.LGcs.NEcs.SIarXiv:2005.03675v32020An introduction to domain adaptation and transfer learning
Wouter M. Kouw, Marco Loog
cs.LGcs.CVstat.MLarXiv:1812.11806v22018Detecting and Preventing Hallucinations in Large Vision Language Models
Anisha Gunjal, Jihan Yin, Erhan Bas
cs.CVcs.LGarXiv:2308.06394v32023Aerial Imagery Pile burn detection using Deep Learning: the FLAME dataset
Alireza Shamsoshoara, Fatemeh Afghah, Abolfazl Razi +3
cs.CVcs.AIcs.LGarXiv:2012.14036v12020Robustness via curvature regularization, and vice versa
Seyed-Mohsen Moosavi-Dezfooli, Alhussein Fawzi, Jonathan Uesato +1
cs.LGcs.CVstat.MLarXiv:1811.09716v12018Transformers as Statisticians: Provable In-Context Learning with In-Context Algorithm Selection
Yu Bai, Fan Chen, Huan Wang +2
cs.LGcs.AIcs.CLarXiv:2306.04637v22023MERT: Acoustic Music Understanding Model with Large-Scale Self-supervised Training
Yizhi Li, Ruibin Yuan, Ge Zhang +17
cs.SDcs.AIcs.CLarXiv:2306.00107v52023U-Time: A Fully Convolutional Network for Time Series Segmentation Applied to Sleep Staging
Mathias Perslev, Michael Hejselbak Jensen, Sune Darkner +2
cs.LGeess.SPstat.MLarXiv:1910.11162v12019Evaluating Large Language Models at Evaluating Instruction Following
Zhiyuan Zeng, Jiatong Yu, Tianyu Gao +3
cs.CLcs.LGarXiv:2310.07641v22023Transolver: A Fast Transformer Solver for PDEs on General Geometries
Haixu Wu, Huakun Luo, Haowen Wang +2
cs.LGmath.NAarXiv:2402.02366v22024Deep Learning for Environmentally Robust Speech Recognition: An Overview of Recent Developments
Zixing Zhang, Jürgen Geiger, Jouni Pohjalainen +3
cs.SDcs.CLcs.LGarXiv:1705.10874v32017Aligning Cyber Space with Physical World: A Comprehensive Survey on Embodied AI
Yang Liu, Weixing Chen, Yongjie Bai +4
cs.CVcs.AIcs.LGarXiv:2407.06886v82024Split Computing and Early Exiting for Deep Learning Applications: Survey and Research Challenges
Yoshitomo Matsubara, Marco Levorato, Francesco Restuccia
eess.SPcs.LGarXiv:2103.04505v42021Understanding the Acceleration Phenomenon via High-Resolution Differential Equations
Bin Shi, Simon S. Du, Michael I. Jordan +1
math.OCcs.LGmath.CAarXiv:1810.08907v32018Generative Probabilistic Novelty Detection with Adversarial Autoencoders
Stanislav Pidhorskyi, Ranya Almohsen, Donald A Adjeroh +1
cs.CVcs.LGarXiv:1807.02588v22018Neural Networks with Few Multiplications
Zhouhan Lin, Matthieu Courbariaux, Roland Memisevic +1
cs.LGcs.NEarXiv:1510.03009v32015Beyond Human Data: Scaling Self-Training for Problem-Solving with Language Models
Avi Singh, John D. Co-Reyes, Rishabh Agarwal +38
cs.LGarXiv:2312.06585v42023Learning Robust Representations via Multi-View Information Bottleneck
Marco Federici, Anjan Dutta, Patrick Forré +2
cs.LGstat.MLarXiv:2002.07017v22020Underdamped Langevin MCMC: A non-asymptotic analysis
Xiang Cheng, Niladri S. Chatterji, Peter L. Bartlett +1
stat.MLcs.LGstat.COarXiv:1707.03663v72017Learning without Concentration
Shahar Mendelson
cs.LGstat.MLarXiv:1401.0304v22014Learning Sparse Nonparametric DAGs
Xun Zheng, Chen Dan, Bryon Aragam +2
stat.MLcs.LGstat.MEarXiv:1909.13189v22019First-order Methods for Geodesically Convex Optimization
Hongyi Zhang, Suvrit Sra
math.OCcs.LGstat.MLarXiv:1602.06053v12016