Machine Learning
Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
14,641 to 14,700 of 20,218
Low-Dimensional Hyperbolic Knowledge Graph Embeddings
Ines Chami, Adva Wolf, Da-Cheng Juan +3
cs.LGcs.AIcs.CLarXiv:2005.00545v12020Evaluation of Text Generation: A Survey
Asli Celikyilmaz, Elizabeth Clark, Jianfeng Gao
cs.CLcs.LGarXiv:2006.14799v22020No Spurious Local Minima in Nonconvex Low Rank Problems: A Unified Geometric Analysis
Rong Ge, Chi Jin, Yi Zheng
cs.LGmath.OCstat.MLarXiv:1704.00708v12017Massively Multilingual Neural Machine Translation in the Wild: Findings and Challenges
Naveen Arivazhagan, Ankur Bapna, Orhan Firat +10
cs.CLcs.LGarXiv:1907.05019v12019Cross-Domain Few-Shot Classification via Learned Feature-Wise Transformation
Hung-Yu Tseng, Hsin-Ying Lee, Jia-Bin Huang +1
cs.CVcs.LGarXiv:2001.08735v32020High Accuracy and High Fidelity Extraction of Neural Networks
Matthew Jagielski, Nicholas Carlini, David Berthelot +2
cs.LGcs.CRstat.MLarXiv:1909.01838v22019Max-value Entropy Search for Efficient Bayesian Optimization
Zi Wang, Stefanie Jegelka
stat.MLcs.LGmath.OCarXiv:1703.01968v32017Hyperparameter Search in Machine Learning
Marc Claesen, Bart De Moor
cs.LGstat.MLarXiv:1502.02127v22015Teacher-Student Curriculum Learning
Tambet Matiisen, Avital Oliver, Taco Cohen +1
cs.LGcs.AIarXiv:1707.00183v22017Progressive Feature Alignment for Unsupervised Domain Adaptation
Chaoqi Chen, Weiping Xie, Wenbing Huang +5
cs.CVcs.LGarXiv:1811.08585v22018Distilling Object Detectors with Fine-grained Feature Imitation
Tao Wang, Li Yuan, Xiaopeng Zhang +1
cs.CVcs.AIcs.LGarXiv:1906.03609v12019Sampling is as easy as learning the score: theory for diffusion models with minimal data assumptions
Sitan Chen, Sinho Chewi, Jerry Li +3
cs.LGmath.STarXiv:2209.11215v32022Implicit Bias of Gradient Descent on Linear Convolutional Networks
Suriya Gunasekar, Jason Lee, Daniel Soudry +1
cs.LGstat.MLarXiv:1806.00468v22018A Deep Learning Approach for Brain Tumor Classification and Segmentation Using a Multiscale Convolutional Neural Network
Francisco Javier Díaz-Pernas, Mario Martínez-Zarzuela, Míriam Antón-Rodríguez +1
eess.IVcs.AIcs.CVarXiv:2402.05975v12024Algorithms to estimate Shapley value feature attributions
Hugh Chen, Ian C. Covert, Scott M. Lundberg +1
cs.LGcs.GTarXiv:2207.07605v12022Large Language Models are Effective Text Rankers with Pairwise Ranking Prompting
Zhen Qin, Rolf Jagerman, Kai Hui +9
cs.IRcs.CLcs.LGarXiv:2306.17563v22023SIGN: Scalable Inception Graph Neural Networks
Fabrizio Frasca, Emanuele Rossi, Davide Eynard +3
cs.LGstat.MLarXiv:2004.11198v32020ShapeShifter: Robust Physical Adversarial Attack on Faster R-CNN Object Detector
Shang-Tse Chen, Cory Cornelius, Jason Martin +1
cs.CVcs.CRcs.LGarXiv:1804.05810v32018DeepSeek-Coder-V2: Breaking the Barrier of Closed-Source Models in Code Intelligence
DeepSeek-AI, Qihao Zhu, Daya Guo +37
cs.SEcs.AIcs.LGarXiv:2406.11931v12024Are we done with ImageNet?
Lucas Beyer, Olivier J. Hénaff, Alexander Kolesnikov +2
cs.CVcs.LGarXiv:2006.07159v12020Measuring the Effects of Data Parallelism on Neural Network Training
Christopher J. Shallue, Jaehoon Lee, Joseph Antognini +3
cs.LGstat.MLarXiv:1811.03600v32018Editing Large Language Models: Problems, Methods, and Opportunities
Yunzhi Yao, Peng Wang, Bozhong Tian +5
cs.CLcs.AIcs.CVarXiv:2305.13172v32023SGD: General Analysis and Improved Rates
Robert Mansel Gower, Nicolas Loizou, Xun Qian +3
cs.LGmath.OCstat.MLarXiv:1901.09401v42019Anti-Backdoor Learning: Training Clean Models on Poisoned Data
Yige Li, Xixiang Lyu, Nodens Koren +3
cs.LGcs.AIarXiv:2110.11571v32021LipNet: End-to-End Sentence-level Lipreading
Yannis M. Assael, Brendan Shillingford, Shimon Whiteson +1
cs.LGcs.CLcs.CVarXiv:1611.01599v22016Analyzing the Structure of Attention in a Transformer Language Model
Jesse Vig, Yonatan Belinkov
cs.CLcs.LGstat.MLarXiv:1906.04284v22019To Tune or Not to Tune? Adapting Pretrained Representations to Diverse Tasks
Matthew E. Peters, Sebastian Ruder, Noah A. Smith
cs.CLcs.LGarXiv:1903.05987v22019Fixing the train-test resolution discrepancy
Hugo Touvron, Andrea Vedaldi, Matthijs Douze +1
cs.CVcs.LGarXiv:1906.06423v42019Variational Intrinsic Control
Karol Gregor, Danilo Jimenez Rezende, Daan Wierstra
cs.LGcs.AIarXiv:1611.07507v12016Algorithms for Verifying Deep Neural Networks
Changliu Liu, Tomer Arnon, Christopher Lazarus +3
cs.LGstat.MLarXiv:1903.06758v22019Lower Bounds for Non-Convex Stochastic Optimization
Yossi Arjevani, Yair Carmon, John C. Duchi +3
math.OCcs.ITcs.LGarXiv:1912.02365v22019Retrosynthetic reaction prediction using neural sequence-to-sequence models
Bowen Liu, Bharath Ramsundar, Prasad Kawthekar +7
cs.LGq-bio.QMstat.MLarXiv:1706.01643v12017Benchmarking Multivariate Time Series Classification Algorithms
Alejandro Pasos Ruiz, Michael Flynn, Anthony Bagnall
cs.LGstat.MLarXiv:2007.13156v22020Reinforced Self-Training (ReST) for Language Modeling
Caglar Gulcehre, Tom Le Paine, Srivatsan Srinivasan +11
cs.CLcs.LGarXiv:2308.08998v22023GraphLIME: Local Interpretable Model Explanations for Graph Neural Networks
Qiang Huang, Makoto Yamada, Yuan Tian +3
cs.LGstat.MLarXiv:2001.06216v22020Graph Neural Controlled Differential Equations for Traffic Forecasting
Jeongwhan Choi, Hwangyong Choi, Jeehyun Hwang +1
cs.LGcs.AIarXiv:2112.03558v12021MaskGAN: Better Text Generation via Filling in the______
William Fedus, Ian Goodfellow, Andrew M. Dai
stat.MLcs.AIcs.LGarXiv:1801.07736v32018Model-Ensemble Trust-Region Policy Optimization
Thanard Kurutach, Ignasi Clavera, Yan Duan +2
cs.LGcs.AIcs.ROarXiv:1802.10592v22018Change-Point Detection in Time-Series Data by Relative Density-Ratio Estimation
Song Liu, Makoto Yamada, Nigel Collier +1
stat.MLcs.LGstat.MEarXiv:1203.0453v22012Rasa: Open Source Language Understanding and Dialogue Management
Tom Bocklisch, Joey Faulkner, Nick Pawlowski +1
cs.CLcs.AIcs.LGarXiv:1712.05181v22017A Survey on Neural Speech Synthesis
Xu Tan, Tao Qin, Frank Soong +1
eess.AScs.CLcs.LGarXiv:2106.15561v32021Thinking Fast and Slow with Deep Learning and Tree Search
Thomas Anthony, Zheng Tian, David Barber
cs.AIcs.LGarXiv:1705.08439v42017Big-Data Science in Porous Materials: Materials Genomics and Machine Learning
Kevin Maik Jablonka, Daniele Ongari, Seyed Mohamad Moosavi +1
cond-mat.mtrl-scics.LGarXiv:2001.06728v32020The Numerics of GANs
Lars Mescheder, Sebastian Nowozin, Andreas Geiger
cs.LGarXiv:1705.10461v32017Towards Deep Conversational Recommendations
Raymond Li, Samira Kahou, Hannes Schulz +3
cs.LGcs.CLcs.IRarXiv:1812.07617v22018Automated Algorithm Selection: Survey and Perspectives
Pascal Kerschke, Holger H. Hoos, Frank Neumann +1
cs.LGcs.AIstat.MLarXiv:1811.11597v12018TimeXer: Empowering Transformers for Time Series Forecasting with Exogenous Variables
Yuxuan Wang, Haixu Wu, Jiaxiang Dong +6
cs.LGcs.AIarXiv:2402.19072v42024Remember What You Want to Forget: Algorithms for Machine Unlearning
Ayush Sekhari, Jayadev Acharya, Gautam Kamath +1
cs.LGcs.AIarXiv:2103.03279v22021Understanding Evolution Strategies for LLM Reasoning: Broader Reasoning Coverage than GRPO
Yunpeng Ba, Zhi Zheng, Yue Xie +7
cs.LGarXiv:2608.27351v12026A Field Guide to Federated Optimization
Jianyu Wang, Zachary Charles, Zheng Xu +50
cs.LGarXiv:2107.06917v12021Machine Learning of coarse-grained Molecular Dynamics Force Fields
Jiang Wang, Simon Olsson, Christoph Wehmeyer +5
physics.comp-phcs.LGstat.MLarXiv:1812.01736v32018Problems with Shapley-value-based explanations as feature importance measures
I. Elizabeth Kumar, Suresh Venkatasubramanian, Carlos Scheidegger +1
cs.AIcs.LGstat.MLarXiv:2002.11097v22020ChequeMark: An Ensemble Machine Learning Framework for After-Hours Business Deposit Fraud Detection
Ann Youduo Xu, Emily Yu, Justin Leski +1
cs.LGarXiv:2608.21629v12026A Comprehensive Overview and Comparative Analysis on Deep Learning Models: CNN, RNN, LSTM, GRU
Farhad Mortezapour Shiri, Thinagaran Perumal, Norwati Mustapha +1
cs.LGcs.AIarXiv:2305.17473v42023Not All Patches are What You Need: Expediting Vision Transformers via Token Reorganizations
Youwei Liang, Chongjian Ge, Zhan Tong +3
cs.CVcs.LGarXiv:2202.07800v22022Learning Deep Neural Network Representations for Koopman Operators of Nonlinear Dynamical Systems
Enoch Yeung, Soumya Kundu, Nathan Hodas
cs.LGcs.AImath.DSarXiv:1708.06850v22017Data-centric Artificial Intelligence: A Survey
Daochen Zha, Zaid Pervaiz Bhat, Kwei-Herng Lai +4
cs.LGcs.AIcs.DBarXiv:2303.10158v32023Probabilistic End-to-end Noise Correction for Learning with Noisy Labels
Kun Yi, Jianxin Wu
cs.CVcs.LGarXiv:1903.07788v12019Generalized Zero-Shot Learning via Synthesized Examples
Vinay Kumar Verma, Gundeep Arora, Ashish Mishra +1
cs.LGcs.CVstat.MLarXiv:1712.03878v52017Reaching the Tail: Calibration Diversity Drives Conformal Coverage under Data Scarcity
Donald Aadithiyan
cs.LGarXiv:2608.21591v12026