Machine Learning
Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
2,281 to 2,340 of 20,190
Playing the lottery with rewards and multiple languages: lottery tickets in RL and NLP
Haonan Yu, Sergey Edunov, Yuandong Tian +1
stat.MLcs.AIcs.LGarXiv:1906.02768v32019Generalized Radiograph Representation Learning via Cross-supervision between Images and Free-text Radiology Reports
Hong-Yu Zhou, Xiaoyu Chen, Yinghao Zhang +3
eess.IVcs.CVcs.LGarXiv:2111.03452v22021"What We Can't Measure, We Can't Understand": Challenges to Demographic Data Procurement in the Pursuit of Fairness
McKane Andrus, Elena Spitzer, Jeffrey Brown +1
cs.CYcs.LGarXiv:2011.02282v22020Explicit Sparse Transformer: Concentrated Attention Through Explicit Selection
Guangxiang Zhao, Junyang Lin, Zhiyuan Zhang +3
cs.CLcs.LGarXiv:1912.11637v12019GTC: Guided Training of CTC Towards Efficient and Accurate Scene Text Recognition
Wenyang Hu, Xiaocong Cai, Jun Hou +2
cs.CVcs.LGeess.IVarXiv:2002.01276v12020Large Language Models can Strategically Deceive their Users when Put Under Pressure
Jérémy Scheurer, Mikita Balesni, Marius Hobbhahn
cs.CLcs.AIcs.LGarXiv:2311.07590v42023Generating High Fidelity Images with Subscale Pixel Networks and Multidimensional Upscaling
Jacob Menick, Nal Kalchbrenner
cs.CVcs.GRcs.LGarXiv:1812.01608v12018Propagation Networks for Model-Based Control Under Partial Observation
Yunzhu Li, Jiajun Wu, Jun-Yan Zhu +3
cs.AIcs.LGcs.ROarXiv:1809.11169v22018The Effect of Natural Distribution Shift on Question Answering Models
John Miller, Karl Krauth, Benjamin Recht +1
cs.LGcs.CLstat.MLarXiv:2004.14444v12020Is Model Collapse Inevitable? Breaking the Curse of Recursion by Accumulating Real and Synthetic Data
Matthias Gerstgrasser, Rylan Schaeffer, Apratim Dey +11
cs.LGcs.AIcs.CLarXiv:2404.01413v22024Drug-Drug Interaction Prediction Based on Knowledge Graph Embeddings and Convolutional-LSTM Network
Md. Rezaul Karim, Michael Cochez, Joao Bosco Jares +3
cs.LGcs.AIarXiv:1908.01288v12019SOM-VAE: Interpretable Discrete Representation Learning on Time Series
Vincent Fortuin, Matthias Hüser, Francesco Locatello +2
cs.LGstat.MLarXiv:1806.02199v72018Intrinsic Dimension Estimation for Robust Detection of AI-Generated Texts
Eduard Tulchinskii, Kristian Kuznetsov, Laida Kushnareva +5
cs.CLcs.AIcs.ITarXiv:2306.04723v22023Path Integral Guided Policy Search
Yevgen Chebotar, Mrinal Kalakrishnan, Ali Yahya +3
cs.ROcs.LGarXiv:1610.00529v22016Acceleration for Compressed Gradient Descent in Distributed and Federated Optimization
Zhize Li, Dmitry Kovalev, Xun Qian +1
math.OCcs.DCcs.LGarXiv:2002.11364v22020Agnostic System Identification for Model-Based Reinforcement Learning
Stephane Ross, J. Andrew Bagnell
cs.LGcs.AIeess.SYarXiv:1203.1007v22012Semi-Supervised QA with Generative Domain-Adaptive Nets
Zhilin Yang, Junjie Hu, Ruslan Salakhutdinov +1
cs.CLcs.LGarXiv:1702.02206v22017Leave no Trace: Learning to Reset for Safe and Autonomous Reinforcement Learning
Benjamin Eysenbach, Shixiang Gu, Julian Ibarz +1
cs.LGcs.ROarXiv:1711.06782v12017LSCP: Locally Selective Combination in Parallel Outlier Ensembles
Yue Zhao, Zain Nasrullah, Maciej K. Hryniewicki +1
cs.LGcs.IRstat.MLarXiv:1812.01528v22018Hypothesis Testing Interpretations and Renyi Differential Privacy
Borja Balle, Gilles Barthe, Marco Gaboardi +2
cs.LGstat.MLarXiv:1905.09982v22019Experience Report: Deep Learning-based System Log Analysis for Anomaly Detection
Zhuangbin Chen, Jinyang Liu, Wenwei Gu +2
cs.SEcs.LGarXiv:2107.05908v22021Untargeted Backdoor Watermark: Towards Harmless and Stealthy Dataset Copyright Protection
Yiming Li, Yang Bai, Yong Jiang +3
cs.CRcs.AIcs.CVarXiv:2210.00875v32022Diagnostic Classification Of Lung Nodules Using 3D Neural Networks
Raunak Dey, Zhongjie Lu, Yi Hong
cs.CVcs.LGstat.MLarXiv:1803.07192v12018Probabilistic FastText for Multi-Sense Word Embeddings
Ben Athiwaratkun, Andrew Gordon Wilson, Anima Anandkumar
cs.CLcs.AIcs.LGarXiv:1806.02901v12018Tree Tensor Networks for Generative Modeling
Song Cheng, Lei Wang, Tao Xiang +1
stat.MLcond-mat.stat-mechcs.LGarXiv:1901.02217v12019Federated Learning with Fair Averaging
Zheng Wang, Xiaoliang Fan, Jianzhong Qi +3
cs.LGarXiv:2104.14937v52021ByRDiE: Byzantine-resilient distributed coordinate descent for decentralized learning
Zhixiong Yang, Waheed U. Bajwa
cs.LGcs.DCmath.OCarXiv:1708.08155v42017ProteinNet: a standardized data set for machine learning of protein structure
Mohammed AlQuraishi
q-bio.BMcs.LGq-bio.QMarXiv:1902.00249v12019ReduNet: A White-box Deep Network from the Principle of Maximizing Rate Reduction
Kwan Ho Ryan Chan, Yaodong Yu, Chong You +3
cs.LGcs.CVcs.ITarXiv:2105.10446v32021Neuron Shapley: Discovering the Responsible Neurons
Amirata Ghorbani, James Zou
stat.MLcs.CVcs.LGarXiv:2002.09815v32020Instruction-driven history-aware policies for robotic manipulations
Pierre-Louis Guhur, Shizhe Chen, Ricardo Garcia +3
cs.ROcs.AIcs.CLarXiv:2209.04899v32022Aligning Superhuman AI with Human Behavior: Chess as a Model System
Reid McIlroy-Young, Siddhartha Sen, Jon Kleinberg +1
cs.AIcs.CYcs.LGarXiv:2006.01855v32020Semigroup-JEPA: Latent Dynamics Consistency for Zero-Shot Physics Generalization
Andy Zeyi Liu, Haoran Sun, Lucas Baker +2
cs.LGcs.AIcs.CVarXiv:2609.10464v12026Conjugate-Computation Variational Inference : Converting Variational Inference in Non-Conjugate Models to Inferences in Conjugate Models
Mohammad Emtiyaz Khan, Wu Lin
cs.LGarXiv:1703.04265v22017Coronavirus (COVID-19) Classification using Deep Features Fusion and Ranking Technique
Umut Ozkaya, Saban Ozturk, Mucahid Barstugan
eess.IVcs.CVcs.LGarXiv:2004.03698v12020TEFM: Token-Efficient Faithful Modeling for Structured Data
Zhichao Hou, Lingdao Sha, Xueyu Mao +3
cs.CLcs.LGarXiv:2609.09552v12026On Learning the Geodesic Path for Incremental Learning
Christian Simon, Piotr Koniusz, Mehrtash Harandi
cs.LGcs.CVarXiv:2104.08572v12021Graph Attention Multi-Layer Perceptron
Wentao Zhang, Ziqi Yin, Zeang Sheng +6
cs.LGcs.AIarXiv:2206.04355v12022Representational Strengths and Limitations of Transformers
Clayton Sanford, Daniel Hsu, Matus Telgarsky
cs.LGstat.MLarXiv:2306.02896v22023A Manually-Curated Dataset of Fixes to Vulnerabilities of Open-Source Software
Serena E. Ponta, Henrik Plate, Antonino Sabetta +2
cs.SEcs.CRcs.LGarXiv:1902.02595v32019SpecTr: Fast Speculative Decoding via Optimal Transport
Ziteng Sun, Ananda Theertha Suresh, Jae Hun Ro +3
cs.LGcs.CLcs.DSarXiv:2310.15141v22023Deep Image Translation with an Affinity-Based Change Prior for Unsupervised Multimodal Change Detection
Luigi Tommaso Luppino, Michael Kampffmeyer, Filippo Maria Bianchi +4
cs.LGcs.CVeess.IVarXiv:2001.04271v22020Operator-valued Kernels for Learning from Functional Response Data
Hachem Kadri, Emmanuel Duflos, Philippe Preux +3
cs.LGstat.MLarXiv:1510.08231v32015Breaking the Sample Size Barrier in Model-Based Reinforcement Learning with a Generative Model
Gen Li, Yuting Wei, Yuejie Chi +1
cs.LGcs.ITmath.OCarXiv:2005.12900v82020The Skellam Mechanism for Differentially Private Federated Learning
Naman Agarwal, Peter Kairouz, Ziyu Liu
cs.LGcs.CRcs.DSarXiv:2110.04995v22021X-CoSD: Communication-Efficient Cross-Vocabulary Collaborative Speculative Decoding
Jaeduk Lee, Wan Choi
cs.CLcs.DCcs.LGarXiv:2609.09166v12026NAS-Bench-1Shot1: Benchmarking and Dissecting One-shot Neural Architecture Search
Arber Zela, Julien Siems, Frank Hutter
cs.LGcs.CVcs.NEarXiv:2001.10422v22020IBIB: A Protocol for Measuring Enterprise AI Systems by Serving Route, Not Model Identifier
Blake Stenstrom, Charangan Vasantharajan, Brian Sathianathan
cs.CLcs.AIcs.LGarXiv:2609.10494v12026GNNAutoScale: Scalable and Expressive Graph Neural Networks via Historical Embeddings
Matthias Fey, Jan E. Lenssen, Frank Weichert +1
cs.LGarXiv:2106.05609v12021One Loop, Two Gains: Can Active Learning win the Lottery for Free?
Benedikt Tscheschner, Eduardo Veas, Marc Masana
cs.LGcs.AIcs.CVarXiv:2609.10311v12026A Review on Explainable Artificial Intelligence for Healthcare: Why, How, and When?
Subrato Bharati, M. Rubaiyat Hossain Mondal, Prajoy Podder
cs.LGcs.AIarXiv:2304.04780v12023Accelerated Policy Learning with Parallel Differentiable Simulation
Jie Xu, Viktor Makoviychuk, Yashraj Narang +4
cs.LGcs.AIcs.GRarXiv:2204.07137v12022Combining SchNet and SHARC: The SchNarc machine learning approach for excited-state dynamics
Julia Westermayr, Michael Gastegger, Philipp Marquetand
physics.chem-phcs.LGstat.MLarXiv:2002.07264v12020Forgetting Only What Matters: Layer-Selective Unlearning toward Robust LLMs
Ravi Ranjan, Olivera Kotevska, Agoritsa Polyzou
cs.LGcs.AIarXiv:2609.10439v12026Reducing Dueling Bandits to Cardinal Bandits
Nir Ailon, Thorsten Joachims, Zohar Karnin
cs.LGarXiv:1405.3396v12014RAP: Robustness-Aware Perturbations for Defending against Backdoor Attacks on NLP Models
Wenkai Yang, Yankai Lin, Peng Li +2
cs.CLcs.LGarXiv:2110.07831v12021Orthogonal Recurrent Neural Networks with Scaled Cayley Transform
Kyle Helfrich, Devin Willmott, Qiang Ye
stat.MLcs.LGarXiv:1707.09520v32017Preventing Zero-Shot Transfer Degradation in Continual Learning of Vision-Language Models
Zangwei Zheng, Mingyuan Ma, Kai Wang +3
cs.CVcs.LGarXiv:2303.06628v22023A Survey of Deep Learning for Scientific Discovery
Maithra Raghu, Eric Schmidt
cs.LGstat.MLarXiv:2003.11755v12020OmniMed-FL: A Robust Multimodal Federated Learning Framework for Clinical Diagnosis
Ayush Debnath, Ruelia Saha, Sudip Misra
cs.LGcs.AIarXiv:2609.10364v12026