Machine Learning

Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

14,641 to 14,700 of 20,218

  1. Low-Dimensional Hyperbolic Knowledge Graph Embeddings

    Ines Chami, Adva Wolf, Da-Cheng Juan +3

    cs.LGcs.AIcs.CLarXiv:2005.00545v12020
  2. Evaluation of Text Generation: A Survey

    Asli Celikyilmaz, Elizabeth Clark, Jianfeng Gao

    cs.CLcs.LGarXiv:2006.14799v22020
  3. No Spurious Local Minima in Nonconvex Low Rank Problems: A Unified Geometric Analysis

    Rong Ge, Chi Jin, Yi Zheng

    cs.LGmath.OCstat.MLarXiv:1704.00708v12017
  4. Massively Multilingual Neural Machine Translation in the Wild: Findings and Challenges

    Naveen Arivazhagan, Ankur Bapna, Orhan Firat +10

    cs.CLcs.LGarXiv:1907.05019v12019
  5. Cross-Domain Few-Shot Classification via Learned Feature-Wise Transformation

    Hung-Yu Tseng, Hsin-Ying Lee, Jia-Bin Huang +1

    cs.CVcs.LGarXiv:2001.08735v32020
  6. High Accuracy and High Fidelity Extraction of Neural Networks

    Matthew Jagielski, Nicholas Carlini, David Berthelot +2

    cs.LGcs.CRstat.MLarXiv:1909.01838v22019
  7. Max-value Entropy Search for Efficient Bayesian Optimization

    Zi Wang, Stefanie Jegelka

    stat.MLcs.LGmath.OCarXiv:1703.01968v32017
  8. Hyperparameter Search in Machine Learning

    Marc Claesen, Bart De Moor

    cs.LGstat.MLarXiv:1502.02127v22015
  9. Teacher-Student Curriculum Learning

    Tambet Matiisen, Avital Oliver, Taco Cohen +1

    cs.LGcs.AIarXiv:1707.00183v22017
  10. Progressive Feature Alignment for Unsupervised Domain Adaptation

    Chaoqi Chen, Weiping Xie, Wenbing Huang +5

    cs.CVcs.LGarXiv:1811.08585v22018
  11. Distilling Object Detectors with Fine-grained Feature Imitation

    Tao Wang, Li Yuan, Xiaopeng Zhang +1

    cs.CVcs.AIcs.LGarXiv:1906.03609v12019
  12. Sampling is as easy as learning the score: theory for diffusion models with minimal data assumptions

    Sitan Chen, Sinho Chewi, Jerry Li +3

    cs.LGmath.STarXiv:2209.11215v32022
  13. Implicit Bias of Gradient Descent on Linear Convolutional Networks

    Suriya Gunasekar, Jason Lee, Daniel Soudry +1

    cs.LGstat.MLarXiv:1806.00468v22018
  14. A Deep Learning Approach for Brain Tumor Classification and Segmentation Using a Multiscale Convolutional Neural Network

    Francisco Javier Díaz-Pernas, Mario Martínez-Zarzuela, Míriam Antón-Rodríguez +1

    eess.IVcs.AIcs.CVarXiv:2402.05975v12024
  15. Algorithms to estimate Shapley value feature attributions

    Hugh Chen, Ian C. Covert, Scott M. Lundberg +1

    cs.LGcs.GTarXiv:2207.07605v12022
  16. Large Language Models are Effective Text Rankers with Pairwise Ranking Prompting

    Zhen Qin, Rolf Jagerman, Kai Hui +9

    cs.IRcs.CLcs.LGarXiv:2306.17563v22023
  17. SIGN: Scalable Inception Graph Neural Networks

    Fabrizio Frasca, Emanuele Rossi, Davide Eynard +3

    cs.LGstat.MLarXiv:2004.11198v32020
  18. ShapeShifter: Robust Physical Adversarial Attack on Faster R-CNN Object Detector

    Shang-Tse Chen, Cory Cornelius, Jason Martin +1

    cs.CVcs.CRcs.LGarXiv:1804.05810v32018
  19. DeepSeek-Coder-V2: Breaking the Barrier of Closed-Source Models in Code Intelligence

    DeepSeek-AI, Qihao Zhu, Daya Guo +37

    cs.SEcs.AIcs.LGarXiv:2406.11931v12024
  20. Are we done with ImageNet?

    Lucas Beyer, Olivier J. Hénaff, Alexander Kolesnikov +2

    cs.CVcs.LGarXiv:2006.07159v12020
  21. Measuring the Effects of Data Parallelism on Neural Network Training

    Christopher J. Shallue, Jaehoon Lee, Joseph Antognini +3

    cs.LGstat.MLarXiv:1811.03600v32018
  22. Editing Large Language Models: Problems, Methods, and Opportunities

    Yunzhi Yao, Peng Wang, Bozhong Tian +5

    cs.CLcs.AIcs.CVarXiv:2305.13172v32023
  23. SGD: General Analysis and Improved Rates

    Robert Mansel Gower, Nicolas Loizou, Xun Qian +3

    cs.LGmath.OCstat.MLarXiv:1901.09401v42019
  24. Anti-Backdoor Learning: Training Clean Models on Poisoned Data

    Yige Li, Xixiang Lyu, Nodens Koren +3

    cs.LGcs.AIarXiv:2110.11571v32021
  25. LipNet: End-to-End Sentence-level Lipreading

    Yannis M. Assael, Brendan Shillingford, Shimon Whiteson +1

    cs.LGcs.CLcs.CVarXiv:1611.01599v22016
  26. Analyzing the Structure of Attention in a Transformer Language Model

    Jesse Vig, Yonatan Belinkov

    cs.CLcs.LGstat.MLarXiv:1906.04284v22019
  27. To Tune or Not to Tune? Adapting Pretrained Representations to Diverse Tasks

    Matthew E. Peters, Sebastian Ruder, Noah A. Smith

    cs.CLcs.LGarXiv:1903.05987v22019
  28. Fixing the train-test resolution discrepancy

    Hugo Touvron, Andrea Vedaldi, Matthijs Douze +1

    cs.CVcs.LGarXiv:1906.06423v42019
  29. Variational Intrinsic Control

    Karol Gregor, Danilo Jimenez Rezende, Daan Wierstra

    cs.LGcs.AIarXiv:1611.07507v12016
  30. Algorithms for Verifying Deep Neural Networks

    Changliu Liu, Tomer Arnon, Christopher Lazarus +3

    cs.LGstat.MLarXiv:1903.06758v22019
  31. Lower Bounds for Non-Convex Stochastic Optimization

    Yossi Arjevani, Yair Carmon, John C. Duchi +3

    math.OCcs.ITcs.LGarXiv:1912.02365v22019
  32. Retrosynthetic reaction prediction using neural sequence-to-sequence models

    Bowen Liu, Bharath Ramsundar, Prasad Kawthekar +7

    cs.LGq-bio.QMstat.MLarXiv:1706.01643v12017
  33. Benchmarking Multivariate Time Series Classification Algorithms

    Alejandro Pasos Ruiz, Michael Flynn, Anthony Bagnall

    cs.LGstat.MLarXiv:2007.13156v22020
  34. Reinforced Self-Training (ReST) for Language Modeling

    Caglar Gulcehre, Tom Le Paine, Srivatsan Srinivasan +11

    cs.CLcs.LGarXiv:2308.08998v22023
  35. GraphLIME: Local Interpretable Model Explanations for Graph Neural Networks

    Qiang Huang, Makoto Yamada, Yuan Tian +3

    cs.LGstat.MLarXiv:2001.06216v22020
  36. Graph Neural Controlled Differential Equations for Traffic Forecasting

    Jeongwhan Choi, Hwangyong Choi, Jeehyun Hwang +1

    cs.LGcs.AIarXiv:2112.03558v12021
  37. MaskGAN: Better Text Generation via Filling in the______

    William Fedus, Ian Goodfellow, Andrew M. Dai

    stat.MLcs.AIcs.LGarXiv:1801.07736v32018
  38. Model-Ensemble Trust-Region Policy Optimization

    Thanard Kurutach, Ignasi Clavera, Yan Duan +2

    cs.LGcs.AIcs.ROarXiv:1802.10592v22018
  39. Change-Point Detection in Time-Series Data by Relative Density-Ratio Estimation

    Song Liu, Makoto Yamada, Nigel Collier +1

    stat.MLcs.LGstat.MEarXiv:1203.0453v22012
  40. Rasa: Open Source Language Understanding and Dialogue Management

    Tom Bocklisch, Joey Faulkner, Nick Pawlowski +1

    cs.CLcs.AIcs.LGarXiv:1712.05181v22017
  41. A Survey on Neural Speech Synthesis

    Xu Tan, Tao Qin, Frank Soong +1

    eess.AScs.CLcs.LGarXiv:2106.15561v32021
  42. Thinking Fast and Slow with Deep Learning and Tree Search

    Thomas Anthony, Zheng Tian, David Barber

    cs.AIcs.LGarXiv:1705.08439v42017
  43. Big-Data Science in Porous Materials: Materials Genomics and Machine Learning

    Kevin Maik Jablonka, Daniele Ongari, Seyed Mohamad Moosavi +1

    cond-mat.mtrl-scics.LGarXiv:2001.06728v32020
  44. The Numerics of GANs

    Lars Mescheder, Sebastian Nowozin, Andreas Geiger

    cs.LGarXiv:1705.10461v32017
  45. Towards Deep Conversational Recommendations

    Raymond Li, Samira Kahou, Hannes Schulz +3

    cs.LGcs.CLcs.IRarXiv:1812.07617v22018
  46. Automated Algorithm Selection: Survey and Perspectives

    Pascal Kerschke, Holger H. Hoos, Frank Neumann +1

    cs.LGcs.AIstat.MLarXiv:1811.11597v12018
  47. TimeXer: Empowering Transformers for Time Series Forecasting with Exogenous Variables

    Yuxuan Wang, Haixu Wu, Jiaxiang Dong +6

    cs.LGcs.AIarXiv:2402.19072v42024
  48. Remember What You Want to Forget: Algorithms for Machine Unlearning

    Ayush Sekhari, Jayadev Acharya, Gautam Kamath +1

    cs.LGcs.AIarXiv:2103.03279v22021
  49. Understanding Evolution Strategies for LLM Reasoning: Broader Reasoning Coverage than GRPO

    Yunpeng Ba, Zhi Zheng, Yue Xie +7

    cs.LGarXiv:2608.27351v12026
  50. A Field Guide to Federated Optimization

    Jianyu Wang, Zachary Charles, Zheng Xu +50

    cs.LGarXiv:2107.06917v12021
  51. Machine Learning of coarse-grained Molecular Dynamics Force Fields

    Jiang Wang, Simon Olsson, Christoph Wehmeyer +5

    physics.comp-phcs.LGstat.MLarXiv:1812.01736v32018
  52. Problems with Shapley-value-based explanations as feature importance measures

    I. Elizabeth Kumar, Suresh Venkatasubramanian, Carlos Scheidegger +1

    cs.AIcs.LGstat.MLarXiv:2002.11097v22020
  53. ChequeMark: An Ensemble Machine Learning Framework for After-Hours Business Deposit Fraud Detection

    Ann Youduo Xu, Emily Yu, Justin Leski +1

    cs.LGarXiv:2608.21629v12026
  54. A Comprehensive Overview and Comparative Analysis on Deep Learning Models: CNN, RNN, LSTM, GRU

    Farhad Mortezapour Shiri, Thinagaran Perumal, Norwati Mustapha +1

    cs.LGcs.AIarXiv:2305.17473v42023
  55. Not All Patches are What You Need: Expediting Vision Transformers via Token Reorganizations

    Youwei Liang, Chongjian Ge, Zhan Tong +3

    cs.CVcs.LGarXiv:2202.07800v22022
  56. Learning Deep Neural Network Representations for Koopman Operators of Nonlinear Dynamical Systems

    Enoch Yeung, Soumya Kundu, Nathan Hodas

    cs.LGcs.AImath.DSarXiv:1708.06850v22017
  57. Data-centric Artificial Intelligence: A Survey

    Daochen Zha, Zaid Pervaiz Bhat, Kwei-Herng Lai +4

    cs.LGcs.AIcs.DBarXiv:2303.10158v32023
  58. Probabilistic End-to-end Noise Correction for Learning with Noisy Labels

    Kun Yi, Jianxin Wu

    cs.CVcs.LGarXiv:1903.07788v12019
  59. Generalized Zero-Shot Learning via Synthesized Examples

    Vinay Kumar Verma, Gundeep Arora, Ashish Mishra +1

    cs.LGcs.CVstat.MLarXiv:1712.03878v52017
  60. Reaching the Tail: Calibration Diversity Drives Conformal Coverage under Data Scarcity

    Donald Aadithiyan

    cs.LGarXiv:2608.21591v12026