Machine Learning
Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
11,521 to 11,580 of 20,193
Online Learning Rate Adaptation with Hypergradient Descent
Atilim Gunes Baydin, Robert Cornish, David Martinez Rubio +2
cs.LGstat.MLarXiv:1703.04782v32017Machine learning with data assimilation and uncertainty quantification for dynamical systems: a review
Sibo Cheng, Cesar Quilodran-Casas, Said Ouala +14
cs.LGarXiv:2303.10462v12023Federated Semi-Supervised Learning with Inter-Client Consistency & Disjoint Learning
Wonyong Jeong, Jaehong Yoon, Eunho Yang +1
cs.LGstat.MLarXiv:2006.12097v32020CLUTRR: A Diagnostic Benchmark for Inductive Reasoning from Text
Koustuv Sinha, Shagun Sodhani, Jin Dong +2
cs.LGcs.CLcs.LOarXiv:1908.06177v22019TeCNO: Surgical Phase Recognition with Multi-Stage Temporal Convolutional Networks
Tobias Czempiel, Magdalini Paschali, Matthias Keicher +4
eess.IVcs.CVcs.LGarXiv:2003.10751v12020Towards Unsupervised Deep Graph Structure Learning
Yixin Liu, Yu Zheng, Daokun Zhang +3
cs.LGarXiv:2201.06367v12022DeepSpeed Ulysses: System Optimizations for Enabling Training of Extreme Long Sequence Transformer Models
Sam Ade Jacobs, Masahiro Tanaka, Chengming Zhang +4
cs.LGcs.CLcs.DCarXiv:2309.14509v22023Selective Question Answering under Domain Shift
Amita Kamath, Robin Jia, Percy Liang
cs.CLcs.LGarXiv:2006.09462v12020A Diagnostic Study of Explainability Techniques for Text Classification
Pepa Atanasova, Jakob Grue Simonsen, Christina Lioma +1
cs.CLcs.LGarXiv:2009.13295v12020Guided Conditional Diffusion for Controllable Traffic Simulation
Ziyuan Zhong, Davis Rempe, Danfei Xu +5
cs.ROcs.AIcs.LGarXiv:2210.17366v12022Hypothesize, Evaluate, Refine: A Scientific Agent for PDE Discovery with Unknown Spatial Coefficient Fields
YuJie Huang, WenWu He, ZhuoEr Lin +3
cs.AIcs.LGarXiv:2608.27475v12026Multi-view Knowledge Graph Embedding for Entity Alignment
Qingheng Zhang, Zequn Sun, Wei Hu +3
cs.AIcs.CLcs.LGarXiv:1906.02390v12019Break-A-Scene: Extracting Multiple Concepts from a Single Image
Omri Avrahami, Kfir Aberman, Ohad Fried +2
cs.CVcs.GRcs.LGarXiv:2305.16311v22023All in One: Multi-task Prompting for Graph Neural Networks
Xiangguo Sun, Hong Cheng, Jia Li +2
cs.SIcs.AIcs.LGarXiv:2307.01504v22023A Comprehensive Survey on Source-free Domain Adaptation
Zhiqi Yu, Jingjing Li, Zhekai Du +2
cs.LGcs.CVcs.MMarXiv:2302.11803v12023Uncovering the structure of clinical EEG signals with self-supervised learning
Hubert Banville, Omar Chehab, Aapo Hyvärinen +2
stat.MLcs.LGeess.SParXiv:2007.16104v12020Deep Learning for Cross-Border Electricity Price Forecasting: A Comparative Study
Hadeer Elashhab, Sai Srijan Papineni, Marvin Dorn +2
cs.LGarXiv:2608.17091v12026Efficient Reasoning on the Edge
Yelysei Bondarenko, Thomas Hehn, Rob Hesselink +15
cs.LGcs.CLarXiv:2603.16867v22026daVinci-Agency: Unlocking Long-Horizon Agency Data-Efficiently
Mohan Jiang, Dayuan Fu, Junhao Shi +8
cs.LGcs.AIcs.SEarXiv:2602.02619v22026TRIP-Bench: A Benchmark for Long-Horizon Interactive Agents in Real-World Scenarios
Yuanzhe Shen, Zisu Huang, Zhengyuan Wang +14
cs.AIcs.LGarXiv:2602.01675v12026Large language models can accurately predict searcher preferences
Paul Thomas, Seth Spielman, Nick Craswell +1
cs.IRcs.AIcs.CLarXiv:2309.10621v32023A Perspective on Explainable Artificial Intelligence Methods: SHAP and LIME
Ahmed Salih, Zahra Raisi-Estabragh, Ilaria Boscolo Galazzo +4
stat.MLcs.AIcs.LGarXiv:2305.02012v32023VoxFormer: Sparse Voxel Transformer for Camera-based 3D Semantic Scene Completion
Yiming Li, Zhiding Yu, Christopher Choy +5
cs.CVcs.AIcs.LGarXiv:2302.12251v22023Can Foundation Models Wrangle Your Data?
Avanika Narayan, Ines Chami, Laurel Orr +2
cs.LGcs.AIcs.DBarXiv:2205.09911v22022CirCNN: Accelerating and Compressing Deep Neural Networks Using Block-CirculantWeight Matrices
Caiwen Ding, Siyu Liao, Yanzhi Wang +13
cs.CVcs.AIcs.LGarXiv:1708.08917v12017Data-driven Advice for Applying Machine Learning to Bioinformatics Problems
Randal S. Olson, William La Cava, Zairah Mustahsan +2
q-bio.QMcs.LGstat.MLarXiv:1708.05070v22017Co$^2$L: Contrastive Continual Learning
Hyuntak Cha, Jaeho Lee, Jinwoo Shin
cs.LGcs.CVarXiv:2106.14413v12021Effect of barren plateaus on gradient-free optimization
Andrew Arrasmith, M. Cerezo, Piotr Czarnik +2
quant-phcs.LGstat.MLarXiv:2011.12245v22020Explainable Artificial Intelligence: a Systematic Review
Giulia Vilone, Luca Longo
cs.AIcs.LGarXiv:2006.00093v42020Single-Stage Semantic Segmentation from Image Labels
Nikita Araslanov, Stefan Roth
cs.CVcs.LGarXiv:2005.08104v12020Drawing Early-Bird Tickets: Towards More Efficient Training of Deep Networks
Haoran You, Chaojian Li, Pengfei Xu +6
cs.LGstat.MLarXiv:1909.11957v62019Deep Learning for Financial Applications : A Survey
Ahmet Murat Ozbayoglu, Mehmet Ugur Gudelek, Omer Berat Sezer
q-fin.STcs.LGstat.MLarXiv:2002.05786v12020Learning Open Set Network with Discriminative Reciprocal Points
Guangyao Chen, Limeng Qiao, Yemin Shi +5
cs.CVcs.LGarXiv:2011.00178v12020Insertion Transformer: Flexible Sequence Generation via Insertion Operations
Mitchell Stern, William Chan, Jamie Kiros +1
cs.CLcs.LGstat.MLarXiv:1902.03249v12019Deep Neural Networks Motivated by Partial Differential Equations
Lars Ruthotto, Eldad Haber
cs.LGmath.OCstat.MLarXiv:1804.04272v22018Learning Control Barrier Functions from Expert Demonstrations
Alexander Robey, Haimin Hu, Lars Lindemann +4
eess.SYcs.LGmath.OCarXiv:2004.03315v32020Playing hard exploration games by watching YouTube
Yusuf Aytar, Tobias Pfaff, David Budden +3
cs.LGcs.AIcs.CVarXiv:1805.11592v22018Freeze-Thaw Bayesian Optimization
Kevin Swersky, Jasper Snoek, Ryan Prescott Adams
stat.MLcs.LGarXiv:1406.3896v12014Multi-Armed Bandit Based Client Scheduling for Federated Learning
Wenchao Xia, Tony Q. S. Quek, Kun Guo +3
cs.ITcs.LGarXiv:2007.02315v12020Symmetric Graph Convolutional Autoencoder for Unsupervised Graph Representation Learning
Jiwoong Park, Minsik Lee, Hyung Jin Chang +2
cs.LGcs.CVstat.MLarXiv:1908.02441v12019Universal Language Model Fine-tuning for Text Classification
Jeremy Howard, Sebastian Ruder
cs.CLcs.LGstat.MLarXiv:1801.06146v52018Feature-Critic Networks for Heterogeneous Domain Generalization
Yiying Li, Yongxin Yang, Wei Zhou +1
cs.LGstat.MLarXiv:1901.11448v32019Guiding Pretraining in Reinforcement Learning with Large Language Models
Yuqing Du, Olivia Watkins, Zihan Wang +5
cs.LGcs.AIcs.CLarXiv:2302.06692v22023Language Modeling Is Compression
Grégoire Delétang, Anian Ruoss, Paul-Ambroise Duquenne +9
cs.LGcs.AIcs.CLarXiv:2309.10668v22023Direct speech-to-speech translation with a sequence-to-sequence model
Ye Jia, Ron J. Weiss, Fadi Biadsy +4
cs.CLcs.LGcs.SDarXiv:1904.06037v22019Understanding and Improving Interpolation in Autoencoders via an Adversarial Regularizer
David Berthelot, Colin Raffel, Aurko Roy +1
cs.LGstat.MLarXiv:1807.07543v22018Complexity of Linear Regions in Deep Networks
Boris Hanin, David Rolnick
stat.MLcs.LGmath.PRarXiv:1901.09021v22019Modified Gaussian Process Regression Models for Cyclic Capacity Prediction of Lithium-ion Batteries
Kailong Liu, Xiaosong Hu, Zhongbao Wei +2
cs.LGeess.SYarXiv:2101.00035v12020End-to-End Model-Free Reinforcement Learning for Urban Driving using Implicit Affordances
Marin Toromanoff, Emilie Wirbel, Fabien Moutarde
cs.LGcs.AIcs.CVarXiv:1911.10868v22019Neural Jump Stochastic Differential Equations
Junteng Jia, Austin R. Benson
cs.LGstat.MLarXiv:1905.10403v32019DiffusionBERT: Improving Generative Masked Language Models with Diffusion Models
Zhengfu He, Tianxiang Sun, Kuanning Wang +2
cs.CLcs.AIcs.LGarXiv:2211.15029v22022Improving Out-of-Distribution Robustness via Selective Augmentation
Huaxiu Yao, Yu Wang, Sai Li +4
cs.LGarXiv:2201.00299v32022Hate Speech Detection and Racial Bias Mitigation in Social Media based on BERT model
Marzieh Mozafari, Reza Farahbakhsh, Noel Crespi
cs.SIcs.CLcs.IRarXiv:2008.06460v22020SGD Learns Over-parameterized Networks that Provably Generalize on Linearly Separable Data
Alon Brutzkus, Amir Globerson, Eran Malach +1
cs.LGarXiv:1710.10174v12017Jumping Ahead: Improving Reconstruction Fidelity with JumpReLU Sparse Autoencoders
Senthooran Rajamanoharan, Tom Lieberum, Nicolas Sonnerat +4
cs.LGarXiv:2407.14435v32024Bike Flow Prediction with Multi-Graph Convolutional Networks
Di Chai, Leye Wang, Qiang Yang
cs.LGcs.AIstat.MLarXiv:1807.10934v12018Spurious Local Minima are Common in Two-Layer ReLU Neural Networks
Itay Safran, Ohad Shamir
cs.LGstat.MLarXiv:1712.08968v32017Towards Understanding Grokking: An Effective Theory of Representation Learning
Ziming Liu, Ouail Kitouni, Niklas Nolte +3
cs.LGcond-mat.dis-nncond-mat.stat-mecharXiv:2205.10343v22022Inductive Matrix Completion Based on Graph Neural Networks
Muhan Zhang, Yixin Chen
cs.IRcs.LGstat.MLarXiv:1904.12058v32019Image Reconstruction: From Sparsity to Data-adaptive Methods and Machine Learning
Saiprasad Ravishankar, Jong Chul Ye, Jeffrey A. Fessler
eess.IVcs.LGstat.MLarXiv:1904.02816v32019