Machine Learning

Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

13,141 to 13,200 of 20,454

  1. Layer-Dependent Importance Sampling for Training Deep and Large Graph Convolutional Networks

    Difan Zou, Ziniu Hu, Yewen Wang +3

    cs.LGcs.SIstat.MLarXiv:1911.07323v12019
  2. Locally Weighted Ensemble Clustering

    Dong Huang, Chang-Dong Wang, Jian-Huang Lai

    cs.LGarXiv:1605.05011v32016
  3. OmniH2O: Universal and Dexterous Human-to-Humanoid Whole-Body Teleoperation and Learning

    Tairan He, Zhengyi Luo, Xialin He +6

    cs.ROcs.CVcs.LGarXiv:2406.08858v12024
  4. CVEfixes: Automated Collection of Vulnerabilities and Their Fixes from Open-Source Software

    Guru Prasad Bhandari, Amara Naseer, Leon Moonen

    cs.SEcs.AIcs.CRarXiv:2107.08760v12021
  5. Why M Heads are Better than One: Training a Diverse Ensemble of Deep Networks

    Stefan Lee, Senthil Purushwalkam, Michael Cogswell +2

    cs.CVcs.LGcs.NEarXiv:1511.06314v12015
  6. The Role of Permutation Invariance in Linear Mode Connectivity of Neural Networks

    Rahim Entezari, Hanie Sedghi, Olga Saukh +1

    cs.LGarXiv:2110.06296v22021
  7. Emergent Complexity and Zero-shot Transfer via Unsupervised Environment Design

    Michael Dennis, Natasha Jaques, Eugene Vinitsky +4

    cs.LGcs.AIcs.MAarXiv:2012.02096v22020
  8. Distributed Online Optimization in Dynamic Environments Using Mirror Descent

    Shahin Shahrampour, Ali Jadbabaie

    math.OCcs.DCcs.LGarXiv:1609.02845v12016
  9. Structured sparsity through convex optimization

    Francis Bach, Rodolphe Jenatton, Julien Mairal +1

    cs.LGstat.MLarXiv:1109.2397v22011
  10. Physics-Informed Neural Networks for Power Systems

    George S. Misyris, Andreas Venzke, Spyros Chatzivasileiadis

    eess.SYcs.LGeess.SParXiv:1911.03737v32019
  11. Machine Learning-Based Heart Disease Diagnosis: A Systematic Literature Review

    Md Manjurul Ahsan, Zahed Siddique

    cs.LGarXiv:2112.06459v12021
  12. Fuzz4All: Universal Fuzzing with Large Language Models

    Chunqiu Steven Xia, Matteo Paltenghi, Jia Le Tian +2

    cs.SEcs.LGarXiv:2308.04748v32023
  13. Using Fast Weights to Attend to the Recent Past

    Jimmy Ba, Geoffrey Hinton, Volodymyr Mnih +2

    stat.MLcs.LGcs.NEarXiv:1610.06258v32016
  14. Time-MoE: Billion-Scale Time Series Foundation Models with Mixture of Experts

    Xiaoming Shi, Shiyu Wang, Yuqi Nie +4

    cs.LGcs.AIarXiv:2409.16040v42024
  15. Graph-to-Sequence Learning using Gated Graph Neural Networks

    Daniel Beck, Gholamreza Haffari, Trevor Cohn

    cs.CLcs.LGarXiv:1806.09835v12018
  16. Rethinking Vision Transformers for MobileNet Size and Speed

    Yanyu Li, Ju Hu, Yang Wen +5

    cs.CVcs.AIcs.LGarXiv:2212.08059v22022
  17. Continuous Inverse Optimal Control with Locally Optimal Examples

    Sergey Levine, Vladlen Koltun

    cs.LGcs.AIstat.MLarXiv:1206.4617v12012
  18. Differentially Private Learning Needs Better Features (or Much More Data)

    Florian Tramèr, Dan Boneh

    cs.LGcs.CRstat.MLarXiv:2011.11660v32020
  19. Asymptotically Exact, Embarrassingly Parallel MCMC

    Willie Neiswanger, Chong Wang, Eric Xing

    stat.MLcs.DCcs.LGarXiv:1311.4780v22013
  20. Neural Radiance Flow for 4D View Synthesis and Video Processing

    Yilun Du, Yinan Zhang, Hong-Xing Yu +2

    cs.CVcs.LGcs.ROarXiv:2012.09790v22020
  21. Interpretable End-to-end Urban Autonomous Driving with Latent Deep Reinforcement Learning

    Jianyu Chen, Shengbo Eben Li, Masayoshi Tomizuka

    cs.ROcs.CVcs.LGarXiv:2001.08726v32020
  22. FengWu: Pushing the Skillful Global Medium-range Weather Forecast beyond 10 Days Lead

    Kang Chen, Tao Han, Junchao Gong +11

    cs.AIcs.LGphysics.ao-pharXiv:2304.02948v12023
  23. Edge YOLO: Real-Time Intelligent Object Detection System Based on Edge-Cloud Cooperation in Autonomous Vehicles

    Siyuan Liang, Hao Wu

    cs.CVcs.LGeess.SParXiv:2205.14942v12022
  24. Learning From Multiple Experts: Self-paced Knowledge Distillation for Long-tailed Classification

    Liuyu Xiang, Guiguang Ding, Jungong Han

    cs.CVcs.LGstat.MLarXiv:2001.01536v32020
  25. Sobolev Training for Neural Networks

    Wojciech Marian Czarnecki, Simon Osindero, Max Jaderberg +2

    cs.LGarXiv:1706.04859v32017
  26. Deep learning: a statistical viewpoint

    Peter L. Bartlett, Andrea Montanari, Alexander Rakhlin

    math.STcs.LGstat.MLarXiv:2103.09177v12021
  27. On the Convergence of Stochastic Gradient Descent with Adaptive Stepsizes

    Xiaoyu Li, Francesco Orabona

    stat.MLcs.LGmath.OCarXiv:1805.08114v32018
  28. WikiHow: A Large Scale Text Summarization Dataset

    Mahnaz Koupaee, William Yang Wang

    cs.CLcs.IRcs.LGarXiv:1810.09305v12018
  29. What graph neural networks cannot learn: depth vs width

    Andreas Loukas

    cs.LGstat.MLarXiv:1907.03199v22019
  30. AAU-net: An Adaptive Attention U-net for Breast Lesions Segmentation in Ultrasound Images

    Gongping Chen, Yu Dai, Jianxun Zhang +1

    eess.IVcs.CVcs.LGarXiv:2204.12077v32022
  31. A Primer on Zeroth-Order Optimization in Signal Processing and Machine Learning

    Sijia Liu, Pin-Yu Chen, Bhavya Kailkhura +3

    cs.LGeess.SPstat.MLarXiv:2006.06224v22020
  32. An exact mapping between the Variational Renormalization Group and Deep Learning

    Pankaj Mehta, David J. Schwab

    stat.MLcond-mat.stat-mechcs.LGarXiv:1410.3831v12014
  33. ASD-DiagNet: A hybrid learning approach for detection of Autism Spectrum Disorder using fMRI data

    Taban Eslami, Vahid Mirjalili, Alvis Fong +2

    cs.LGeess.IVstat.MLarXiv:1904.07577v12019
  34. Dual Discriminator Generative Adversarial Nets

    Tu Dinh Nguyen, Trung Le, Hung Vu +1

    cs.LGstat.MLarXiv:1709.03831v12017
  35. Weisfeiler and Lehman Go Topological: Message Passing Simplicial Networks

    Cristian Bodnar, Fabrizio Frasca, Yu Guang Wang +4

    cs.LGcs.SIarXiv:2103.03212v22021
  36. Label-Free Concept Bottleneck Models

    Tuomas Oikarinen, Subhro Das, Lam M. Nguyen +1

    cs.LGcs.CVarXiv:2304.06129v22023
  37. Discovering Discrete Latent Topics with Neural Variational Inference

    Yishu Miao, Edward Grefenstette, Phil Blunsom

    cs.CLcs.AIcs.IRarXiv:1706.00359v22017
  38. Physics informed deep learning for computational elastodynamics without labeled data

    Chengping Rao, Hao Sun, Yang Liu

    math.NAcs.AIcs.CEarXiv:2006.08472v12020
  39. Machine Learning on Graphs: A Model and Comprehensive Taxonomy

    Ines Chami, Sami Abu-El-Haija, Bryan Perozzi +2

    cs.LGcs.NEcs.SIarXiv:2005.03675v32020
  40. An introduction to domain adaptation and transfer learning

    Wouter M. Kouw, Marco Loog

    cs.LGcs.CVstat.MLarXiv:1812.11806v22018
  41. Detecting and Preventing Hallucinations in Large Vision Language Models

    Anisha Gunjal, Jihan Yin, Erhan Bas

    cs.CVcs.LGarXiv:2308.06394v32023
  42. Aerial Imagery Pile burn detection using Deep Learning: the FLAME dataset

    Alireza Shamsoshoara, Fatemeh Afghah, Abolfazl Razi +3

    cs.CVcs.AIcs.LGarXiv:2012.14036v12020
  43. Robustness via curvature regularization, and vice versa

    Seyed-Mohsen Moosavi-Dezfooli, Alhussein Fawzi, Jonathan Uesato +1

    cs.LGcs.CVstat.MLarXiv:1811.09716v12018
  44. Transformers as Statisticians: Provable In-Context Learning with In-Context Algorithm Selection

    Yu Bai, Fan Chen, Huan Wang +2

    cs.LGcs.AIcs.CLarXiv:2306.04637v22023
  45. MERT: Acoustic Music Understanding Model with Large-Scale Self-supervised Training

    Yizhi Li, Ruibin Yuan, Ge Zhang +17

    cs.SDcs.AIcs.CLarXiv:2306.00107v52023
  46. U-Time: A Fully Convolutional Network for Time Series Segmentation Applied to Sleep Staging

    Mathias Perslev, Michael Hejselbak Jensen, Sune Darkner +2

    cs.LGeess.SPstat.MLarXiv:1910.11162v12019
  47. Evaluating Large Language Models at Evaluating Instruction Following

    Zhiyuan Zeng, Jiatong Yu, Tianyu Gao +3

    cs.CLcs.LGarXiv:2310.07641v22023
  48. Transolver: A Fast Transformer Solver for PDEs on General Geometries

    Haixu Wu, Huakun Luo, Haowen Wang +2

    cs.LGmath.NAarXiv:2402.02366v22024
  49. Deep Learning for Environmentally Robust Speech Recognition: An Overview of Recent Developments

    Zixing Zhang, Jürgen Geiger, Jouni Pohjalainen +3

    cs.SDcs.CLcs.LGarXiv:1705.10874v32017
  50. Aligning Cyber Space with Physical World: A Comprehensive Survey on Embodied AI

    Yang Liu, Weixing Chen, Yongjie Bai +4

    cs.CVcs.AIcs.LGarXiv:2407.06886v82024
  51. Split Computing and Early Exiting for Deep Learning Applications: Survey and Research Challenges

    Yoshitomo Matsubara, Marco Levorato, Francesco Restuccia

    eess.SPcs.LGarXiv:2103.04505v42021
  52. Understanding the Acceleration Phenomenon via High-Resolution Differential Equations

    Bin Shi, Simon S. Du, Michael I. Jordan +1

    math.OCcs.LGmath.CAarXiv:1810.08907v32018
  53. Generative Probabilistic Novelty Detection with Adversarial Autoencoders

    Stanislav Pidhorskyi, Ranya Almohsen, Donald A Adjeroh +1

    cs.CVcs.LGarXiv:1807.02588v22018
  54. Neural Networks with Few Multiplications

    Zhouhan Lin, Matthieu Courbariaux, Roland Memisevic +1

    cs.LGcs.NEarXiv:1510.03009v32015
  55. Beyond Human Data: Scaling Self-Training for Problem-Solving with Language Models

    Avi Singh, John D. Co-Reyes, Rishabh Agarwal +38

    cs.LGarXiv:2312.06585v42023
  56. Learning Robust Representations via Multi-View Information Bottleneck

    Marco Federici, Anjan Dutta, Patrick Forré +2

    cs.LGstat.MLarXiv:2002.07017v22020
  57. Underdamped Langevin MCMC: A non-asymptotic analysis

    Xiang Cheng, Niladri S. Chatterji, Peter L. Bartlett +1

    stat.MLcs.LGstat.COarXiv:1707.03663v72017
  58. Learning without Concentration

    Shahar Mendelson

    cs.LGstat.MLarXiv:1401.0304v22014
  59. Learning Sparse Nonparametric DAGs

    Xun Zheng, Chen Dan, Bryon Aragam +2

    stat.MLcs.LGstat.MEarXiv:1909.13189v22019
  60. First-order Methods for Geodesically Convex Optimization

    Hongyi Zhang, Suvrit Sra

    math.OCcs.LGstat.MLarXiv:1602.06053v12016