Machine Learning

Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

2,281 to 2,340 of 20,190

  1. Playing the lottery with rewards and multiple languages: lottery tickets in RL and NLP

    Haonan Yu, Sergey Edunov, Yuandong Tian +1

    stat.MLcs.AIcs.LGarXiv:1906.02768v32019
  2. Generalized Radiograph Representation Learning via Cross-supervision between Images and Free-text Radiology Reports

    Hong-Yu Zhou, Xiaoyu Chen, Yinghao Zhang +3

    eess.IVcs.CVcs.LGarXiv:2111.03452v22021
  3. "What We Can't Measure, We Can't Understand": Challenges to Demographic Data Procurement in the Pursuit of Fairness

    McKane Andrus, Elena Spitzer, Jeffrey Brown +1

    cs.CYcs.LGarXiv:2011.02282v22020
  4. Explicit Sparse Transformer: Concentrated Attention Through Explicit Selection

    Guangxiang Zhao, Junyang Lin, Zhiyuan Zhang +3

    cs.CLcs.LGarXiv:1912.11637v12019
  5. GTC: Guided Training of CTC Towards Efficient and Accurate Scene Text Recognition

    Wenyang Hu, Xiaocong Cai, Jun Hou +2

    cs.CVcs.LGeess.IVarXiv:2002.01276v12020
  6. Large Language Models can Strategically Deceive their Users when Put Under Pressure

    Jérémy Scheurer, Mikita Balesni, Marius Hobbhahn

    cs.CLcs.AIcs.LGarXiv:2311.07590v42023
  7. Generating High Fidelity Images with Subscale Pixel Networks and Multidimensional Upscaling

    Jacob Menick, Nal Kalchbrenner

    cs.CVcs.GRcs.LGarXiv:1812.01608v12018
  8. Propagation Networks for Model-Based Control Under Partial Observation

    Yunzhu Li, Jiajun Wu, Jun-Yan Zhu +3

    cs.AIcs.LGcs.ROarXiv:1809.11169v22018
  9. The Effect of Natural Distribution Shift on Question Answering Models

    John Miller, Karl Krauth, Benjamin Recht +1

    cs.LGcs.CLstat.MLarXiv:2004.14444v12020
  10. Is Model Collapse Inevitable? Breaking the Curse of Recursion by Accumulating Real and Synthetic Data

    Matthias Gerstgrasser, Rylan Schaeffer, Apratim Dey +11

    cs.LGcs.AIcs.CLarXiv:2404.01413v22024
  11. Drug-Drug Interaction Prediction Based on Knowledge Graph Embeddings and Convolutional-LSTM Network

    Md. Rezaul Karim, Michael Cochez, Joao Bosco Jares +3

    cs.LGcs.AIarXiv:1908.01288v12019
  12. SOM-VAE: Interpretable Discrete Representation Learning on Time Series

    Vincent Fortuin, Matthias Hüser, Francesco Locatello +2

    cs.LGstat.MLarXiv:1806.02199v72018
  13. Intrinsic Dimension Estimation for Robust Detection of AI-Generated Texts

    Eduard Tulchinskii, Kristian Kuznetsov, Laida Kushnareva +5

    cs.CLcs.AIcs.ITarXiv:2306.04723v22023
  14. Path Integral Guided Policy Search

    Yevgen Chebotar, Mrinal Kalakrishnan, Ali Yahya +3

    cs.ROcs.LGarXiv:1610.00529v22016
  15. Acceleration for Compressed Gradient Descent in Distributed and Federated Optimization

    Zhize Li, Dmitry Kovalev, Xun Qian +1

    math.OCcs.DCcs.LGarXiv:2002.11364v22020
  16. Agnostic System Identification for Model-Based Reinforcement Learning

    Stephane Ross, J. Andrew Bagnell

    cs.LGcs.AIeess.SYarXiv:1203.1007v22012
  17. Semi-Supervised QA with Generative Domain-Adaptive Nets

    Zhilin Yang, Junjie Hu, Ruslan Salakhutdinov +1

    cs.CLcs.LGarXiv:1702.02206v22017
  18. Leave no Trace: Learning to Reset for Safe and Autonomous Reinforcement Learning

    Benjamin Eysenbach, Shixiang Gu, Julian Ibarz +1

    cs.LGcs.ROarXiv:1711.06782v12017
  19. LSCP: Locally Selective Combination in Parallel Outlier Ensembles

    Yue Zhao, Zain Nasrullah, Maciej K. Hryniewicki +1

    cs.LGcs.IRstat.MLarXiv:1812.01528v22018
  20. Hypothesis Testing Interpretations and Renyi Differential Privacy

    Borja Balle, Gilles Barthe, Marco Gaboardi +2

    cs.LGstat.MLarXiv:1905.09982v22019
  21. Experience Report: Deep Learning-based System Log Analysis for Anomaly Detection

    Zhuangbin Chen, Jinyang Liu, Wenwei Gu +2

    cs.SEcs.LGarXiv:2107.05908v22021
  22. Untargeted Backdoor Watermark: Towards Harmless and Stealthy Dataset Copyright Protection

    Yiming Li, Yang Bai, Yong Jiang +3

    cs.CRcs.AIcs.CVarXiv:2210.00875v32022
  23. Diagnostic Classification Of Lung Nodules Using 3D Neural Networks

    Raunak Dey, Zhongjie Lu, Yi Hong

    cs.CVcs.LGstat.MLarXiv:1803.07192v12018
  24. Probabilistic FastText for Multi-Sense Word Embeddings

    Ben Athiwaratkun, Andrew Gordon Wilson, Anima Anandkumar

    cs.CLcs.AIcs.LGarXiv:1806.02901v12018
  25. Tree Tensor Networks for Generative Modeling

    Song Cheng, Lei Wang, Tao Xiang +1

    stat.MLcond-mat.stat-mechcs.LGarXiv:1901.02217v12019
  26. Federated Learning with Fair Averaging

    Zheng Wang, Xiaoliang Fan, Jianzhong Qi +3

    cs.LGarXiv:2104.14937v52021
  27. ByRDiE: Byzantine-resilient distributed coordinate descent for decentralized learning

    Zhixiong Yang, Waheed U. Bajwa

    cs.LGcs.DCmath.OCarXiv:1708.08155v42017
  28. ProteinNet: a standardized data set for machine learning of protein structure

    Mohammed AlQuraishi

    q-bio.BMcs.LGq-bio.QMarXiv:1902.00249v12019
  29. ReduNet: A White-box Deep Network from the Principle of Maximizing Rate Reduction

    Kwan Ho Ryan Chan, Yaodong Yu, Chong You +3

    cs.LGcs.CVcs.ITarXiv:2105.10446v32021
  30. Neuron Shapley: Discovering the Responsible Neurons

    Amirata Ghorbani, James Zou

    stat.MLcs.CVcs.LGarXiv:2002.09815v32020
  31. Instruction-driven history-aware policies for robotic manipulations

    Pierre-Louis Guhur, Shizhe Chen, Ricardo Garcia +3

    cs.ROcs.AIcs.CLarXiv:2209.04899v32022
  32. Aligning Superhuman AI with Human Behavior: Chess as a Model System

    Reid McIlroy-Young, Siddhartha Sen, Jon Kleinberg +1

    cs.AIcs.CYcs.LGarXiv:2006.01855v32020
  33. Semigroup-JEPA: Latent Dynamics Consistency for Zero-Shot Physics Generalization

    Andy Zeyi Liu, Haoran Sun, Lucas Baker +2

    cs.LGcs.AIcs.CVarXiv:2609.10464v12026
  34. Conjugate-Computation Variational Inference : Converting Variational Inference in Non-Conjugate Models to Inferences in Conjugate Models

    Mohammad Emtiyaz Khan, Wu Lin

    cs.LGarXiv:1703.04265v22017
  35. Coronavirus (COVID-19) Classification using Deep Features Fusion and Ranking Technique

    Umut Ozkaya, Saban Ozturk, Mucahid Barstugan

    eess.IVcs.CVcs.LGarXiv:2004.03698v12020
  36. TEFM: Token-Efficient Faithful Modeling for Structured Data

    Zhichao Hou, Lingdao Sha, Xueyu Mao +3

    cs.CLcs.LGarXiv:2609.09552v12026
  37. On Learning the Geodesic Path for Incremental Learning

    Christian Simon, Piotr Koniusz, Mehrtash Harandi

    cs.LGcs.CVarXiv:2104.08572v12021
  38. Graph Attention Multi-Layer Perceptron

    Wentao Zhang, Ziqi Yin, Zeang Sheng +6

    cs.LGcs.AIarXiv:2206.04355v12022
  39. Representational Strengths and Limitations of Transformers

    Clayton Sanford, Daniel Hsu, Matus Telgarsky

    cs.LGstat.MLarXiv:2306.02896v22023
  40. A Manually-Curated Dataset of Fixes to Vulnerabilities of Open-Source Software

    Serena E. Ponta, Henrik Plate, Antonino Sabetta +2

    cs.SEcs.CRcs.LGarXiv:1902.02595v32019
  41. SpecTr: Fast Speculative Decoding via Optimal Transport

    Ziteng Sun, Ananda Theertha Suresh, Jae Hun Ro +3

    cs.LGcs.CLcs.DSarXiv:2310.15141v22023
  42. Deep Image Translation with an Affinity-Based Change Prior for Unsupervised Multimodal Change Detection

    Luigi Tommaso Luppino, Michael Kampffmeyer, Filippo Maria Bianchi +4

    cs.LGcs.CVeess.IVarXiv:2001.04271v22020
  43. Operator-valued Kernels for Learning from Functional Response Data

    Hachem Kadri, Emmanuel Duflos, Philippe Preux +3

    cs.LGstat.MLarXiv:1510.08231v32015
  44. Breaking the Sample Size Barrier in Model-Based Reinforcement Learning with a Generative Model

    Gen Li, Yuting Wei, Yuejie Chi +1

    cs.LGcs.ITmath.OCarXiv:2005.12900v82020
  45. The Skellam Mechanism for Differentially Private Federated Learning

    Naman Agarwal, Peter Kairouz, Ziyu Liu

    cs.LGcs.CRcs.DSarXiv:2110.04995v22021
  46. X-CoSD: Communication-Efficient Cross-Vocabulary Collaborative Speculative Decoding

    Jaeduk Lee, Wan Choi

    cs.CLcs.DCcs.LGarXiv:2609.09166v12026
  47. NAS-Bench-1Shot1: Benchmarking and Dissecting One-shot Neural Architecture Search

    Arber Zela, Julien Siems, Frank Hutter

    cs.LGcs.CVcs.NEarXiv:2001.10422v22020
  48. IBIB: A Protocol for Measuring Enterprise AI Systems by Serving Route, Not Model Identifier

    Blake Stenstrom, Charangan Vasantharajan, Brian Sathianathan

    cs.CLcs.AIcs.LGarXiv:2609.10494v12026
  49. GNNAutoScale: Scalable and Expressive Graph Neural Networks via Historical Embeddings

    Matthias Fey, Jan E. Lenssen, Frank Weichert +1

    cs.LGarXiv:2106.05609v12021
  50. One Loop, Two Gains: Can Active Learning win the Lottery for Free?

    Benedikt Tscheschner, Eduardo Veas, Marc Masana

    cs.LGcs.AIcs.CVarXiv:2609.10311v12026
  51. A Review on Explainable Artificial Intelligence for Healthcare: Why, How, and When?

    Subrato Bharati, M. Rubaiyat Hossain Mondal, Prajoy Podder

    cs.LGcs.AIarXiv:2304.04780v12023
  52. Accelerated Policy Learning with Parallel Differentiable Simulation

    Jie Xu, Viktor Makoviychuk, Yashraj Narang +4

    cs.LGcs.AIcs.GRarXiv:2204.07137v12022
  53. Combining SchNet and SHARC: The SchNarc machine learning approach for excited-state dynamics

    Julia Westermayr, Michael Gastegger, Philipp Marquetand

    physics.chem-phcs.LGstat.MLarXiv:2002.07264v12020
  54. Forgetting Only What Matters: Layer-Selective Unlearning toward Robust LLMs

    Ravi Ranjan, Olivera Kotevska, Agoritsa Polyzou

    cs.LGcs.AIarXiv:2609.10439v12026
  55. Reducing Dueling Bandits to Cardinal Bandits

    Nir Ailon, Thorsten Joachims, Zohar Karnin

    cs.LGarXiv:1405.3396v12014
  56. RAP: Robustness-Aware Perturbations for Defending against Backdoor Attacks on NLP Models

    Wenkai Yang, Yankai Lin, Peng Li +2

    cs.CLcs.LGarXiv:2110.07831v12021
  57. Orthogonal Recurrent Neural Networks with Scaled Cayley Transform

    Kyle Helfrich, Devin Willmott, Qiang Ye

    stat.MLcs.LGarXiv:1707.09520v32017
  58. Preventing Zero-Shot Transfer Degradation in Continual Learning of Vision-Language Models

    Zangwei Zheng, Mingyuan Ma, Kai Wang +3

    cs.CVcs.LGarXiv:2303.06628v22023
  59. A Survey of Deep Learning for Scientific Discovery

    Maithra Raghu, Eric Schmidt

    cs.LGstat.MLarXiv:2003.11755v12020
  60. OmniMed-FL: A Robust Multimodal Federated Learning Framework for Clinical Diagnosis

    Ayush Debnath, Ruelia Saha, Sudip Misra

    cs.LGcs.AIarXiv:2609.10364v12026