Machine Learning

Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

11,521 to 11,580 of 20,193

  1. Online Learning Rate Adaptation with Hypergradient Descent

    Atilim Gunes Baydin, Robert Cornish, David Martinez Rubio +2

    cs.LGstat.MLarXiv:1703.04782v32017
  2. Machine learning with data assimilation and uncertainty quantification for dynamical systems: a review

    Sibo Cheng, Cesar Quilodran-Casas, Said Ouala +14

    cs.LGarXiv:2303.10462v12023
  3. Federated Semi-Supervised Learning with Inter-Client Consistency & Disjoint Learning

    Wonyong Jeong, Jaehong Yoon, Eunho Yang +1

    cs.LGstat.MLarXiv:2006.12097v32020
  4. CLUTRR: A Diagnostic Benchmark for Inductive Reasoning from Text

    Koustuv Sinha, Shagun Sodhani, Jin Dong +2

    cs.LGcs.CLcs.LOarXiv:1908.06177v22019
  5. TeCNO: Surgical Phase Recognition with Multi-Stage Temporal Convolutional Networks

    Tobias Czempiel, Magdalini Paschali, Matthias Keicher +4

    eess.IVcs.CVcs.LGarXiv:2003.10751v12020
  6. Towards Unsupervised Deep Graph Structure Learning

    Yixin Liu, Yu Zheng, Daokun Zhang +3

    cs.LGarXiv:2201.06367v12022
  7. DeepSpeed Ulysses: System Optimizations for Enabling Training of Extreme Long Sequence Transformer Models

    Sam Ade Jacobs, Masahiro Tanaka, Chengming Zhang +4

    cs.LGcs.CLcs.DCarXiv:2309.14509v22023
  8. Selective Question Answering under Domain Shift

    Amita Kamath, Robin Jia, Percy Liang

    cs.CLcs.LGarXiv:2006.09462v12020
  9. A Diagnostic Study of Explainability Techniques for Text Classification

    Pepa Atanasova, Jakob Grue Simonsen, Christina Lioma +1

    cs.CLcs.LGarXiv:2009.13295v12020
  10. Guided Conditional Diffusion for Controllable Traffic Simulation

    Ziyuan Zhong, Davis Rempe, Danfei Xu +5

    cs.ROcs.AIcs.LGarXiv:2210.17366v12022
  11. Hypothesize, Evaluate, Refine: A Scientific Agent for PDE Discovery with Unknown Spatial Coefficient Fields

    YuJie Huang, WenWu He, ZhuoEr Lin +3

    cs.AIcs.LGarXiv:2608.27475v12026
  12. Multi-view Knowledge Graph Embedding for Entity Alignment

    Qingheng Zhang, Zequn Sun, Wei Hu +3

    cs.AIcs.CLcs.LGarXiv:1906.02390v12019
  13. Break-A-Scene: Extracting Multiple Concepts from a Single Image

    Omri Avrahami, Kfir Aberman, Ohad Fried +2

    cs.CVcs.GRcs.LGarXiv:2305.16311v22023
  14. All in One: Multi-task Prompting for Graph Neural Networks

    Xiangguo Sun, Hong Cheng, Jia Li +2

    cs.SIcs.AIcs.LGarXiv:2307.01504v22023
  15. A Comprehensive Survey on Source-free Domain Adaptation

    Zhiqi Yu, Jingjing Li, Zhekai Du +2

    cs.LGcs.CVcs.MMarXiv:2302.11803v12023
  16. Uncovering the structure of clinical EEG signals with self-supervised learning

    Hubert Banville, Omar Chehab, Aapo Hyvärinen +2

    stat.MLcs.LGeess.SParXiv:2007.16104v12020
  17. Deep Learning for Cross-Border Electricity Price Forecasting: A Comparative Study

    Hadeer Elashhab, Sai Srijan Papineni, Marvin Dorn +2

    cs.LGarXiv:2608.17091v12026
  18. Efficient Reasoning on the Edge

    Yelysei Bondarenko, Thomas Hehn, Rob Hesselink +15

    cs.LGcs.CLarXiv:2603.16867v22026
  19. daVinci-Agency: Unlocking Long-Horizon Agency Data-Efficiently

    Mohan Jiang, Dayuan Fu, Junhao Shi +8

    cs.LGcs.AIcs.SEarXiv:2602.02619v22026
  20. TRIP-Bench: A Benchmark for Long-Horizon Interactive Agents in Real-World Scenarios

    Yuanzhe Shen, Zisu Huang, Zhengyuan Wang +14

    cs.AIcs.LGarXiv:2602.01675v12026
  21. Large language models can accurately predict searcher preferences

    Paul Thomas, Seth Spielman, Nick Craswell +1

    cs.IRcs.AIcs.CLarXiv:2309.10621v32023
  22. A Perspective on Explainable Artificial Intelligence Methods: SHAP and LIME

    Ahmed Salih, Zahra Raisi-Estabragh, Ilaria Boscolo Galazzo +4

    stat.MLcs.AIcs.LGarXiv:2305.02012v32023
  23. VoxFormer: Sparse Voxel Transformer for Camera-based 3D Semantic Scene Completion

    Yiming Li, Zhiding Yu, Christopher Choy +5

    cs.CVcs.AIcs.LGarXiv:2302.12251v22023
  24. Can Foundation Models Wrangle Your Data?

    Avanika Narayan, Ines Chami, Laurel Orr +2

    cs.LGcs.AIcs.DBarXiv:2205.09911v22022
  25. CirCNN: Accelerating and Compressing Deep Neural Networks Using Block-CirculantWeight Matrices

    Caiwen Ding, Siyu Liao, Yanzhi Wang +13

    cs.CVcs.AIcs.LGarXiv:1708.08917v12017
  26. Data-driven Advice for Applying Machine Learning to Bioinformatics Problems

    Randal S. Olson, William La Cava, Zairah Mustahsan +2

    q-bio.QMcs.LGstat.MLarXiv:1708.05070v22017
  27. Co$^2$L: Contrastive Continual Learning

    Hyuntak Cha, Jaeho Lee, Jinwoo Shin

    cs.LGcs.CVarXiv:2106.14413v12021
  28. Effect of barren plateaus on gradient-free optimization

    Andrew Arrasmith, M. Cerezo, Piotr Czarnik +2

    quant-phcs.LGstat.MLarXiv:2011.12245v22020
  29. Explainable Artificial Intelligence: a Systematic Review

    Giulia Vilone, Luca Longo

    cs.AIcs.LGarXiv:2006.00093v42020
  30. Single-Stage Semantic Segmentation from Image Labels

    Nikita Araslanov, Stefan Roth

    cs.CVcs.LGarXiv:2005.08104v12020
  31. Drawing Early-Bird Tickets: Towards More Efficient Training of Deep Networks

    Haoran You, Chaojian Li, Pengfei Xu +6

    cs.LGstat.MLarXiv:1909.11957v62019
  32. Deep Learning for Financial Applications : A Survey

    Ahmet Murat Ozbayoglu, Mehmet Ugur Gudelek, Omer Berat Sezer

    q-fin.STcs.LGstat.MLarXiv:2002.05786v12020
  33. Learning Open Set Network with Discriminative Reciprocal Points

    Guangyao Chen, Limeng Qiao, Yemin Shi +5

    cs.CVcs.LGarXiv:2011.00178v12020
  34. Insertion Transformer: Flexible Sequence Generation via Insertion Operations

    Mitchell Stern, William Chan, Jamie Kiros +1

    cs.CLcs.LGstat.MLarXiv:1902.03249v12019
  35. Deep Neural Networks Motivated by Partial Differential Equations

    Lars Ruthotto, Eldad Haber

    cs.LGmath.OCstat.MLarXiv:1804.04272v22018
  36. Learning Control Barrier Functions from Expert Demonstrations

    Alexander Robey, Haimin Hu, Lars Lindemann +4

    eess.SYcs.LGmath.OCarXiv:2004.03315v32020
  37. Playing hard exploration games by watching YouTube

    Yusuf Aytar, Tobias Pfaff, David Budden +3

    cs.LGcs.AIcs.CVarXiv:1805.11592v22018
  38. Freeze-Thaw Bayesian Optimization

    Kevin Swersky, Jasper Snoek, Ryan Prescott Adams

    stat.MLcs.LGarXiv:1406.3896v12014
  39. Multi-Armed Bandit Based Client Scheduling for Federated Learning

    Wenchao Xia, Tony Q. S. Quek, Kun Guo +3

    cs.ITcs.LGarXiv:2007.02315v12020
  40. Symmetric Graph Convolutional Autoencoder for Unsupervised Graph Representation Learning

    Jiwoong Park, Minsik Lee, Hyung Jin Chang +2

    cs.LGcs.CVstat.MLarXiv:1908.02441v12019
  41. Universal Language Model Fine-tuning for Text Classification

    Jeremy Howard, Sebastian Ruder

    cs.CLcs.LGstat.MLarXiv:1801.06146v52018
  42. Feature-Critic Networks for Heterogeneous Domain Generalization

    Yiying Li, Yongxin Yang, Wei Zhou +1

    cs.LGstat.MLarXiv:1901.11448v32019
  43. Guiding Pretraining in Reinforcement Learning with Large Language Models

    Yuqing Du, Olivia Watkins, Zihan Wang +5

    cs.LGcs.AIcs.CLarXiv:2302.06692v22023
  44. Language Modeling Is Compression

    Grégoire Delétang, Anian Ruoss, Paul-Ambroise Duquenne +9

    cs.LGcs.AIcs.CLarXiv:2309.10668v22023
  45. Direct speech-to-speech translation with a sequence-to-sequence model

    Ye Jia, Ron J. Weiss, Fadi Biadsy +4

    cs.CLcs.LGcs.SDarXiv:1904.06037v22019
  46. Understanding and Improving Interpolation in Autoencoders via an Adversarial Regularizer

    David Berthelot, Colin Raffel, Aurko Roy +1

    cs.LGstat.MLarXiv:1807.07543v22018
  47. Complexity of Linear Regions in Deep Networks

    Boris Hanin, David Rolnick

    stat.MLcs.LGmath.PRarXiv:1901.09021v22019
  48. Modified Gaussian Process Regression Models for Cyclic Capacity Prediction of Lithium-ion Batteries

    Kailong Liu, Xiaosong Hu, Zhongbao Wei +2

    cs.LGeess.SYarXiv:2101.00035v12020
  49. End-to-End Model-Free Reinforcement Learning for Urban Driving using Implicit Affordances

    Marin Toromanoff, Emilie Wirbel, Fabien Moutarde

    cs.LGcs.AIcs.CVarXiv:1911.10868v22019
  50. Neural Jump Stochastic Differential Equations

    Junteng Jia, Austin R. Benson

    cs.LGstat.MLarXiv:1905.10403v32019
  51. DiffusionBERT: Improving Generative Masked Language Models with Diffusion Models

    Zhengfu He, Tianxiang Sun, Kuanning Wang +2

    cs.CLcs.AIcs.LGarXiv:2211.15029v22022
  52. Improving Out-of-Distribution Robustness via Selective Augmentation

    Huaxiu Yao, Yu Wang, Sai Li +4

    cs.LGarXiv:2201.00299v32022
  53. Hate Speech Detection and Racial Bias Mitigation in Social Media based on BERT model

    Marzieh Mozafari, Reza Farahbakhsh, Noel Crespi

    cs.SIcs.CLcs.IRarXiv:2008.06460v22020
  54. SGD Learns Over-parameterized Networks that Provably Generalize on Linearly Separable Data

    Alon Brutzkus, Amir Globerson, Eran Malach +1

    cs.LGarXiv:1710.10174v12017
  55. Jumping Ahead: Improving Reconstruction Fidelity with JumpReLU Sparse Autoencoders

    Senthooran Rajamanoharan, Tom Lieberum, Nicolas Sonnerat +4

    cs.LGarXiv:2407.14435v32024
  56. Bike Flow Prediction with Multi-Graph Convolutional Networks

    Di Chai, Leye Wang, Qiang Yang

    cs.LGcs.AIstat.MLarXiv:1807.10934v12018
  57. Spurious Local Minima are Common in Two-Layer ReLU Neural Networks

    Itay Safran, Ohad Shamir

    cs.LGstat.MLarXiv:1712.08968v32017
  58. Towards Understanding Grokking: An Effective Theory of Representation Learning

    Ziming Liu, Ouail Kitouni, Niklas Nolte +3

    cs.LGcond-mat.dis-nncond-mat.stat-mecharXiv:2205.10343v22022
  59. Inductive Matrix Completion Based on Graph Neural Networks

    Muhan Zhang, Yixin Chen

    cs.IRcs.LGstat.MLarXiv:1904.12058v32019
  60. Image Reconstruction: From Sparsity to Data-adaptive Methods and Machine Learning

    Saiprasad Ravishankar, Jong Chul Ye, Jeffrey A. Fessler

    eess.IVcs.LGstat.MLarXiv:1904.02816v32019