Machine Learning
Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
14,581 to 14,640 of 20,219
Choose a Transformer: Fourier or Galerkin
Shuhao Cao
cs.LGmath.NAarXiv:2105.14995v42021SpinQuant: LLM quantization with learned rotations
Zechun Liu, Changsheng Zhao, Igor Fedorov +6
cs.LGcs.AIcs.CLarXiv:2405.16406v42024Dynamic Graph Convolutional Networks
Franco Manessi, Alessandro Rozza, Mario Manzo
cs.LGstat.MLarXiv:1704.06199v12017Think Globally, Act Locally: A Deep Neural Network Approach to High-Dimensional Time Series Forecasting
Rajat Sen, Hsiang-Fu Yu, Inderjit Dhillon
stat.MLcs.LGarXiv:1905.03806v22019Algorithmic Recourse: from Counterfactual Explanations to Interventions
Amir-Hossein Karimi, Bernhard Schölkopf, Isabel Valera
cs.LGcs.AIstat.MLarXiv:2002.06278v42020MemGuard: Defending against Black-Box Membership Inference Attacks via Adversarial Examples
Jinyuan Jia, Ahmed Salem, Michael Backes +2
cs.CRcs.LGarXiv:1909.10594v32019Evaluation of Deep Convolutional Nets for Document Image Classification and Retrieval
Adam W. Harley, Alex Ufkes, Konstantinos G. Derpanis
cs.CVcs.IRcs.LGarXiv:1502.07058v12015Message Passing Neural PDE Solvers
Johannes Brandstetter, Daniel Worrall, Max Welling
cs.LGcs.CVmath.NAarXiv:2202.03376v32022Verified Uncertainty Calibration
Ananya Kumar, Percy Liang, Tengyu Ma
cs.LGstat.MLarXiv:1909.10155v22019OmniQuant: Omnidirectionally Calibrated Quantization for Large Language Models
Wenqi Shao, Mengzhao Chen, Zhaoyang Zhang +7
cs.LGcs.CLarXiv:2308.13137v32023Lazier Than Lazy Greedy
Baharan Mirzasoleiman, Ashwinkumar Badanidiyuru, Amin Karbasi +2
cs.LGcs.DScs.IRarXiv:1409.7938v32014GameWAM: A World Action Model for Video Games
Yuncheng Guo, Zhanqiu Zhang, Yiwen Guo +1
cs.AIcs.CVcs.LGarXiv:2608.26200v12026Learning Koopman Invariant Subspaces for Dynamic Mode Decomposition
Naoya Takeishi, Yoshinobu Kawahara, Takehisa Yairi
cs.LGmath.DSstat.MLarXiv:1710.04340v22017Time-to-Event Prediction with Neural Networks and Cox Regression
Håvard Kvamme, Ørnulf Borgan, Ida Scheel
stat.MLcs.LGarXiv:1907.00825v22019Deep Anomaly Detection with Deviation Networks
Guansong Pang, Chunhua Shen, Anton van den Hengel
cs.LGstat.MLarXiv:1911.08623v12019Hyperspherical Variational Auto-Encoders
Tim R. Davidson, Luca Falorsi, Nicola De Cao +2
stat.MLcs.LGarXiv:1804.00891v32018Multimodal Intelligence: Representation Learning, Information Fusion, and Applications
Chao Zhang, Zichao Yang, Xiaodong He +1
cs.AIcs.CLcs.CVarXiv:1911.03977v32019A Semi-supervised Graph Attentive Network for Financial Fraud Detection
Daixin Wang, Jianbin Lin, Peng Cui +7
cs.SIcs.CRcs.LGarXiv:2003.01171v12020Deep Learning for Time-Series Analysis
John Cristian Borges Gamboa
cs.LGarXiv:1701.01887v12017A Survey of the State of Explainable AI for Natural Language Processing
Marina Danilevsky, Kun Qian, Ranit Aharonov +3
cs.CLcs.AIcs.LGarXiv:2010.00711v12020An Investigation into Neural Net Optimization via Hessian Eigenvalue Density
Behrooz Ghorbani, Shankar Krishnan, Ying Xiao
cs.LGstat.MLarXiv:1901.10159v12019Towards Understanding Ensemble, Knowledge Distillation and Self-Distillation in Deep Learning
Zeyuan Allen-Zhu, Yuanzhi Li
cs.LGcs.NEmath.OCarXiv:2012.09816v32020The Carbon Footprint of Machine Learning Training Will Plateau, Then Shrink
David Patterson, Joseph Gonzalez, Urs Hölzle +7
cs.LGcs.AIcs.GLarXiv:2204.05149v12022Learning to Dispatch for Job Shop Scheduling via Deep Reinforcement Learning
Cong Zhang, Wen Song, Zhiguang Cao +3
cs.LGcs.AIstat.MLarXiv:2010.12367v12020Aligning Text-to-Image Models using Human Feedback
Kimin Lee, Hao Liu, Moonkyung Ryu +6
cs.LGcs.AIcs.CVarXiv:2302.12192v12023Federated Meta-Learning with Fast Convergence and Efficient Communication
Fei Chen, Mi Luo, Zhenhua Dong +2
cs.LGcs.IRarXiv:1802.07876v22018hoBIT: A Profile-Aware Retrieval-Augmented Chatbot for University Academic Advising
Yoonseo Kim, Seongmin Lee, Joongheon Kim +1
cs.IRcs.LGarXiv:2608.26604v12026Gauge Equivariant Convolutional Networks and the Icosahedral CNN
Taco S. Cohen, Maurice Weiler, Berkay Kicanaoglu +1
cs.LGcs.CVcs.NEarXiv:1902.04615v32019DeepGO: Predicting protein functions from sequence and interactions using a deep ontology-aware classifier
Maxat Kulmanov, Mohammed Asif Khan, Robert Hoehndorf
q-bio.GNcs.LGq-bio.QMarXiv:1705.05919v12017Propagate Yourself: Exploring Pixel-Level Consistency for Unsupervised Visual Representation Learning
Zhenda Xie, Yutong Lin, Zheng Zhang +3
cs.CVcs.LGarXiv:2011.10043v22020Stochastic Adversarial Video Prediction
Alex X. Lee, Richard Zhang, Frederik Ebert +3
cs.CVcs.AIcs.LGarXiv:1804.01523v12018What Are Bayesian Neural Network Posteriors Really Like?
Pavel Izmailov, Sharad Vikram, Matthew D. Hoffman +1
cs.LGstat.MLarXiv:2104.14421v12021Video (language) modeling: a baseline for generative models of natural videos
MarcAurelio Ranzato, Arthur Szlam, Joan Bruna +3
cs.LGcs.CVarXiv:1412.6604v52014Secure and Robust Machine Learning for Healthcare: A Survey
Adnan Qayyum, Junaid Qadir, Muhammad Bilal +1
cs.LGeess.IVstat.MLarXiv:2001.08103v12020Multi-Agent Collaboration: Harnessing the Power of Intelligent LLM Agents
Yashar Talebirad, Amirhossein Nadiri
cs.AIcs.LGcs.MAarXiv:2306.03314v12023Multi-view Self-supervised Deep Learning for 6D Pose Estimation in the Amazon Picking Challenge
Andy Zeng, Kuan-Ting Yu, Shuran Song +4
cs.CVcs.LGcs.ROarXiv:1609.09475v32016Inductive Biases for Deep Learning of Higher-Level Cognition
Anirudh Goyal, Yoshua Bengio
cs.LGcs.AIstat.MLarXiv:2011.15091v42020Visualizing and Measuring the Geometry of BERT
Andy Coenen, Emily Reif, Ann Yuan +4
cs.LGcs.CLstat.MLarXiv:1906.02715v22019Token-Level Advertising
Hanbing Liu, Bowei Zhang, Changyuan Yu +2
cs.GTcs.LGarXiv:2608.27382v12026Generative Adversarial Networks (GANs Survey): Challenges, Solutions, and Future Directions
Divya Saxena, Jiannong Cao
cs.LGeess.IVstat.MLarXiv:2005.00065v42020Machine Learning Advances for Time Series Forecasting
Ricardo P. Masini, Marcelo C. Medeiros, Eduardo F. Mendes
econ.EMcs.LGstat.AParXiv:2012.12802v32020Programming Is Hard -- Or at Least It Used to Be: Educational Opportunities And Challenges of AI Code Generation
Brett A. Becker, Paul Denny, James Finnie-Ansley +3
cs.HCcs.AIcs.CYarXiv:2212.01020v12022Deep Dynamics Models for Learning Dexterous Manipulation
Anusha Nagabandi, Kurt Konoglie, Sergey Levine +1
cs.ROcs.LGarXiv:1909.11652v12019Stabilizing Training of Generative Adversarial Networks through Regularization
Kevin Roth, Aurelien Lucchi, Sebastian Nowozin +1
cs.LGstat.MLarXiv:1705.09367v22017Liquid Time-constant Networks
Ramin Hasani, Mathias Lechner, Alexander Amini +2
cs.LGcs.NEstat.MLarXiv:2006.04439v42020Scatter Component Analysis: A Unified Framework for Domain Adaptation and Domain Generalization
Muhammad Ghifary, David Balduzzi, W. Bastiaan Kleijn +1
cs.CVcs.AIcs.LGarXiv:1510.04373v22015Stochastic model-based minimization of weakly convex functions
Damek Davis, Dmitriy Drusvyatskiy
math.OCcs.LGarXiv:1803.06523v32018Gradient Projection Memory for Continual Learning
Gobinda Saha, Isha Garg, Kaushik Roy
cs.LGcs.CVarXiv:2103.09762v12021Pre-training Molecular Graph Representation with 3D Geometry
Shengchao Liu, Hanchen Wang, Weiyang Liu +3
cs.LGcs.CVeess.IVarXiv:2110.07728v22021Span-based Joint Entity and Relation Extraction with Transformer Pre-training
Markus Eberts, Adrian Ulges
cs.CLcs.LGarXiv:1909.07755v42019Visual Foresight: Model-Based Deep Reinforcement Learning for Vision-Based Robotic Control
Frederik Ebert, Chelsea Finn, Sudeep Dasari +3
cs.ROcs.AIcs.CVarXiv:1812.00568v12018Federated Ensemble Forecasting Under Supply-Chain Market Volatility
Shunmukha Sagar Puppala
cs.LGarXiv:2608.21399v12026Multi-site fMRI Analysis Using Privacy-preserving Federated Learning and Domain Adaptation: ABIDE Results
Xiaoxiao Li, Yufeng Gu, Nicha Dvornek +3
cs.LGeess.IVarXiv:2001.05647v32020Transformer Accelerator (TFA): A Macro-Op INT8 Hardware Chip for Transformer Inference and Machine Translation
Shashank
cs.ARcs.CLcs.LGarXiv:2608.23582v12026Model Reduction and Neural Networks for Parametric PDEs
Kaushik Bhattacharya, Bamdad Hosseini, Nikola B. Kovachki +1
math.NAcs.LGstat.MLarXiv:2005.03180v22020Responsive Safety in Reinforcement Learning by PID Lagrangian Methods
Adam Stooke, Joshua Achiam, Pieter Abbeel
math.OCcs.AIcs.LGarXiv:2007.03964v12020Federated Learning with Buffered Asynchronous Aggregation
John Nguyen, Kshitiz Malik, Hongyuan Zhan +4
cs.LGarXiv:2106.06639v42021Learning Particle Dynamics for Manipulating Rigid Bodies, Deformable Objects, and Fluids
Yunzhu Li, Jiajun Wu, Russ Tedrake +2
cs.LGcs.AIcs.ROarXiv:1810.01566v22018Kubric: A scalable dataset generator
Klaus Greff, Francois Belletti, Lucas Beyer +32
cs.CVcs.GRcs.LGarXiv:2203.03570v12022CyrillicQA: The Influence of Phonetically Encoded Secret Language on LLM Performance
Erik Thureck
cs.CLcs.AIcs.LGarXiv:2608.21462v12026