Machine Learning

Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

14,581 to 14,640 of 20,219

  1. Choose a Transformer: Fourier or Galerkin

    Shuhao Cao

    cs.LGmath.NAarXiv:2105.14995v42021
  2. SpinQuant: LLM quantization with learned rotations

    Zechun Liu, Changsheng Zhao, Igor Fedorov +6

    cs.LGcs.AIcs.CLarXiv:2405.16406v42024
  3. Dynamic Graph Convolutional Networks

    Franco Manessi, Alessandro Rozza, Mario Manzo

    cs.LGstat.MLarXiv:1704.06199v12017
  4. Think Globally, Act Locally: A Deep Neural Network Approach to High-Dimensional Time Series Forecasting

    Rajat Sen, Hsiang-Fu Yu, Inderjit Dhillon

    stat.MLcs.LGarXiv:1905.03806v22019
  5. Algorithmic Recourse: from Counterfactual Explanations to Interventions

    Amir-Hossein Karimi, Bernhard Schölkopf, Isabel Valera

    cs.LGcs.AIstat.MLarXiv:2002.06278v42020
  6. MemGuard: Defending against Black-Box Membership Inference Attacks via Adversarial Examples

    Jinyuan Jia, Ahmed Salem, Michael Backes +2

    cs.CRcs.LGarXiv:1909.10594v32019
  7. Evaluation of Deep Convolutional Nets for Document Image Classification and Retrieval

    Adam W. Harley, Alex Ufkes, Konstantinos G. Derpanis

    cs.CVcs.IRcs.LGarXiv:1502.07058v12015
  8. Message Passing Neural PDE Solvers

    Johannes Brandstetter, Daniel Worrall, Max Welling

    cs.LGcs.CVmath.NAarXiv:2202.03376v32022
  9. Verified Uncertainty Calibration

    Ananya Kumar, Percy Liang, Tengyu Ma

    cs.LGstat.MLarXiv:1909.10155v22019
  10. OmniQuant: Omnidirectionally Calibrated Quantization for Large Language Models

    Wenqi Shao, Mengzhao Chen, Zhaoyang Zhang +7

    cs.LGcs.CLarXiv:2308.13137v32023
  11. Lazier Than Lazy Greedy

    Baharan Mirzasoleiman, Ashwinkumar Badanidiyuru, Amin Karbasi +2

    cs.LGcs.DScs.IRarXiv:1409.7938v32014
  12. GameWAM: A World Action Model for Video Games

    Yuncheng Guo, Zhanqiu Zhang, Yiwen Guo +1

    cs.AIcs.CVcs.LGarXiv:2608.26200v12026
  13. Learning Koopman Invariant Subspaces for Dynamic Mode Decomposition

    Naoya Takeishi, Yoshinobu Kawahara, Takehisa Yairi

    cs.LGmath.DSstat.MLarXiv:1710.04340v22017
  14. Time-to-Event Prediction with Neural Networks and Cox Regression

    Håvard Kvamme, Ørnulf Borgan, Ida Scheel

    stat.MLcs.LGarXiv:1907.00825v22019
  15. Deep Anomaly Detection with Deviation Networks

    Guansong Pang, Chunhua Shen, Anton van den Hengel

    cs.LGstat.MLarXiv:1911.08623v12019
  16. Hyperspherical Variational Auto-Encoders

    Tim R. Davidson, Luca Falorsi, Nicola De Cao +2

    stat.MLcs.LGarXiv:1804.00891v32018
  17. Multimodal Intelligence: Representation Learning, Information Fusion, and Applications

    Chao Zhang, Zichao Yang, Xiaodong He +1

    cs.AIcs.CLcs.CVarXiv:1911.03977v32019
  18. A Semi-supervised Graph Attentive Network for Financial Fraud Detection

    Daixin Wang, Jianbin Lin, Peng Cui +7

    cs.SIcs.CRcs.LGarXiv:2003.01171v12020
  19. Deep Learning for Time-Series Analysis

    John Cristian Borges Gamboa

    cs.LGarXiv:1701.01887v12017
  20. A Survey of the State of Explainable AI for Natural Language Processing

    Marina Danilevsky, Kun Qian, Ranit Aharonov +3

    cs.CLcs.AIcs.LGarXiv:2010.00711v12020
  21. An Investigation into Neural Net Optimization via Hessian Eigenvalue Density

    Behrooz Ghorbani, Shankar Krishnan, Ying Xiao

    cs.LGstat.MLarXiv:1901.10159v12019
  22. Towards Understanding Ensemble, Knowledge Distillation and Self-Distillation in Deep Learning

    Zeyuan Allen-Zhu, Yuanzhi Li

    cs.LGcs.NEmath.OCarXiv:2012.09816v32020
  23. The Carbon Footprint of Machine Learning Training Will Plateau, Then Shrink

    David Patterson, Joseph Gonzalez, Urs Hölzle +7

    cs.LGcs.AIcs.GLarXiv:2204.05149v12022
  24. Learning to Dispatch for Job Shop Scheduling via Deep Reinforcement Learning

    Cong Zhang, Wen Song, Zhiguang Cao +3

    cs.LGcs.AIstat.MLarXiv:2010.12367v12020
  25. Aligning Text-to-Image Models using Human Feedback

    Kimin Lee, Hao Liu, Moonkyung Ryu +6

    cs.LGcs.AIcs.CVarXiv:2302.12192v12023
  26. Federated Meta-Learning with Fast Convergence and Efficient Communication

    Fei Chen, Mi Luo, Zhenhua Dong +2

    cs.LGcs.IRarXiv:1802.07876v22018
  27. hoBIT: A Profile-Aware Retrieval-Augmented Chatbot for University Academic Advising

    Yoonseo Kim, Seongmin Lee, Joongheon Kim +1

    cs.IRcs.LGarXiv:2608.26604v12026
  28. Gauge Equivariant Convolutional Networks and the Icosahedral CNN

    Taco S. Cohen, Maurice Weiler, Berkay Kicanaoglu +1

    cs.LGcs.CVcs.NEarXiv:1902.04615v32019
  29. DeepGO: Predicting protein functions from sequence and interactions using a deep ontology-aware classifier

    Maxat Kulmanov, Mohammed Asif Khan, Robert Hoehndorf

    q-bio.GNcs.LGq-bio.QMarXiv:1705.05919v12017
  30. Propagate Yourself: Exploring Pixel-Level Consistency for Unsupervised Visual Representation Learning

    Zhenda Xie, Yutong Lin, Zheng Zhang +3

    cs.CVcs.LGarXiv:2011.10043v22020
  31. Stochastic Adversarial Video Prediction

    Alex X. Lee, Richard Zhang, Frederik Ebert +3

    cs.CVcs.AIcs.LGarXiv:1804.01523v12018
  32. What Are Bayesian Neural Network Posteriors Really Like?

    Pavel Izmailov, Sharad Vikram, Matthew D. Hoffman +1

    cs.LGstat.MLarXiv:2104.14421v12021
  33. Video (language) modeling: a baseline for generative models of natural videos

    MarcAurelio Ranzato, Arthur Szlam, Joan Bruna +3

    cs.LGcs.CVarXiv:1412.6604v52014
  34. Secure and Robust Machine Learning for Healthcare: A Survey

    Adnan Qayyum, Junaid Qadir, Muhammad Bilal +1

    cs.LGeess.IVstat.MLarXiv:2001.08103v12020
  35. Multi-Agent Collaboration: Harnessing the Power of Intelligent LLM Agents

    Yashar Talebirad, Amirhossein Nadiri

    cs.AIcs.LGcs.MAarXiv:2306.03314v12023
  36. Multi-view Self-supervised Deep Learning for 6D Pose Estimation in the Amazon Picking Challenge

    Andy Zeng, Kuan-Ting Yu, Shuran Song +4

    cs.CVcs.LGcs.ROarXiv:1609.09475v32016
  37. Inductive Biases for Deep Learning of Higher-Level Cognition

    Anirudh Goyal, Yoshua Bengio

    cs.LGcs.AIstat.MLarXiv:2011.15091v42020
  38. Visualizing and Measuring the Geometry of BERT

    Andy Coenen, Emily Reif, Ann Yuan +4

    cs.LGcs.CLstat.MLarXiv:1906.02715v22019
  39. Token-Level Advertising

    Hanbing Liu, Bowei Zhang, Changyuan Yu +2

    cs.GTcs.LGarXiv:2608.27382v12026
  40. Generative Adversarial Networks (GANs Survey): Challenges, Solutions, and Future Directions

    Divya Saxena, Jiannong Cao

    cs.LGeess.IVstat.MLarXiv:2005.00065v42020
  41. Machine Learning Advances for Time Series Forecasting

    Ricardo P. Masini, Marcelo C. Medeiros, Eduardo F. Mendes

    econ.EMcs.LGstat.AParXiv:2012.12802v32020
  42. Programming Is Hard -- Or at Least It Used to Be: Educational Opportunities And Challenges of AI Code Generation

    Brett A. Becker, Paul Denny, James Finnie-Ansley +3

    cs.HCcs.AIcs.CYarXiv:2212.01020v12022
  43. Deep Dynamics Models for Learning Dexterous Manipulation

    Anusha Nagabandi, Kurt Konoglie, Sergey Levine +1

    cs.ROcs.LGarXiv:1909.11652v12019
  44. Stabilizing Training of Generative Adversarial Networks through Regularization

    Kevin Roth, Aurelien Lucchi, Sebastian Nowozin +1

    cs.LGstat.MLarXiv:1705.09367v22017
  45. Liquid Time-constant Networks

    Ramin Hasani, Mathias Lechner, Alexander Amini +2

    cs.LGcs.NEstat.MLarXiv:2006.04439v42020
  46. Scatter Component Analysis: A Unified Framework for Domain Adaptation and Domain Generalization

    Muhammad Ghifary, David Balduzzi, W. Bastiaan Kleijn +1

    cs.CVcs.AIcs.LGarXiv:1510.04373v22015
  47. Stochastic model-based minimization of weakly convex functions

    Damek Davis, Dmitriy Drusvyatskiy

    math.OCcs.LGarXiv:1803.06523v32018
  48. Gradient Projection Memory for Continual Learning

    Gobinda Saha, Isha Garg, Kaushik Roy

    cs.LGcs.CVarXiv:2103.09762v12021
  49. Pre-training Molecular Graph Representation with 3D Geometry

    Shengchao Liu, Hanchen Wang, Weiyang Liu +3

    cs.LGcs.CVeess.IVarXiv:2110.07728v22021
  50. Span-based Joint Entity and Relation Extraction with Transformer Pre-training

    Markus Eberts, Adrian Ulges

    cs.CLcs.LGarXiv:1909.07755v42019
  51. Visual Foresight: Model-Based Deep Reinforcement Learning for Vision-Based Robotic Control

    Frederik Ebert, Chelsea Finn, Sudeep Dasari +3

    cs.ROcs.AIcs.CVarXiv:1812.00568v12018
  52. Federated Ensemble Forecasting Under Supply-Chain Market Volatility

    Shunmukha Sagar Puppala

    cs.LGarXiv:2608.21399v12026
  53. Multi-site fMRI Analysis Using Privacy-preserving Federated Learning and Domain Adaptation: ABIDE Results

    Xiaoxiao Li, Yufeng Gu, Nicha Dvornek +3

    cs.LGeess.IVarXiv:2001.05647v32020
  54. Transformer Accelerator (TFA): A Macro-Op INT8 Hardware Chip for Transformer Inference and Machine Translation

    Shashank

    cs.ARcs.CLcs.LGarXiv:2608.23582v12026
  55. Model Reduction and Neural Networks for Parametric PDEs

    Kaushik Bhattacharya, Bamdad Hosseini, Nikola B. Kovachki +1

    math.NAcs.LGstat.MLarXiv:2005.03180v22020
  56. Responsive Safety in Reinforcement Learning by PID Lagrangian Methods

    Adam Stooke, Joshua Achiam, Pieter Abbeel

    math.OCcs.AIcs.LGarXiv:2007.03964v12020
  57. Federated Learning with Buffered Asynchronous Aggregation

    John Nguyen, Kshitiz Malik, Hongyuan Zhan +4

    cs.LGarXiv:2106.06639v42021
  58. Learning Particle Dynamics for Manipulating Rigid Bodies, Deformable Objects, and Fluids

    Yunzhu Li, Jiajun Wu, Russ Tedrake +2

    cs.LGcs.AIcs.ROarXiv:1810.01566v22018
  59. Kubric: A scalable dataset generator

    Klaus Greff, Francois Belletti, Lucas Beyer +32

    cs.CVcs.GRcs.LGarXiv:2203.03570v12022
  60. CyrillicQA: The Influence of Phonetically Encoded Secret Language on LLM Performance

    Erik Thureck

    cs.CLcs.AIcs.LGarXiv:2608.21462v12026