Machine Learning

Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

1,621 to 1,680 of 20,193

  1. The Lazy Neuron Phenomenon: On Emergence of Activation Sparsity in Transformers

    Zonglin Li, Chong You, Srinadh Bhojanapalli +8

    cs.LGcs.CLcs.CVarXiv:2210.06313v22022
  2. ECG Arrhythmia Classification Using Transfer Learning from 2-Dimensional Deep CNN Features

    Milad Salem, Shayan Taheri, Jiann Shiun-Yuan

    cs.LGcs.CVstat.MLarXiv:1812.04693v12018
  3. Reinforcement Learning Upside Down: Don't Predict Rewards -- Just Map Them to Actions

    Juergen Schmidhuber

    cs.AIcs.LGarXiv:1912.02875v22019
  4. Rapid trial-and-error learning with simulation supports flexible tool use and physical reasoning

    Kelsey R. Allen, Kevin A. Smith, Joshua B. Tenenbaum

    cs.AIcs.LGcs.ROarXiv:1907.09620v32019
  5. DiffiT: Diffusion Vision Transformers for Image Generation

    Ali Hatamizadeh, Jiaming Song, Guilin Liu +2

    cs.CVcs.AIcs.LGarXiv:2312.02139v32023
  6. Clinical Intervention Prediction and Understanding using Deep Networks

    Harini Suresh, Nathan Hunt, Alistair Johnson +3

    cs.LGarXiv:1705.08498v12017
  7. Neural Task Graphs: Generalizing to Unseen Tasks from a Single Video Demonstration

    De-An Huang, Suraj Nair, Danfei Xu +5

    cs.CVcs.AIcs.LGarXiv:1807.03480v22018
  8. Towards a Mathematical Understanding of Neural Network-Based Machine Learning: what we know and what we don't

    Weinan E, Chao Ma, Stephan Wojtowytsch +1

    cs.LGmath.NAstat.MLarXiv:2009.10713v32020
  9. Cross-Subject Transfer Learning in Human Activity Recognition Systems using Generative Adversarial Networks

    Elnaz Soleimani, Ehsan Nazerfard

    cs.LGstat.MLarXiv:1903.12489v12019
  10. Time Series Anomaly Detection Using Convolutional Neural Networks and Transfer Learning

    Tailai Wen, Roy Keyes

    cs.LGcs.CVstat.MLarXiv:1905.13628v12019
  11. A Deep Convolutional Neural Network for COVID-19 Detection Using Chest X-Rays

    Pedro R. A. S. Bassi, Romis Attux

    eess.IVcs.CVcs.LGarXiv:2005.01578v42020
  12. DeepACO: Neural-enhanced Ant Systems for Combinatorial Optimization

    Haoran Ye, Jiarui Wang, Zhiguang Cao +2

    cs.NEcs.AIcs.LGarXiv:2309.14032v22023
  13. Machine Unlearning: Solutions and Challenges

    Jie Xu, Zihan Wu, Cong Wang +1

    cs.LGcs.AIarXiv:2308.07061v32023
  14. Domain-Specific Hallucination Detection in Large Language Models

    Varun Teja Chundru, Debasmita Biswas

    cs.CLcs.AIcs.LGarXiv:2609.11878v12026
  15. AlphaStock: A Buying-Winners-and-Selling-Losers Investment Strategy using Interpretable Deep Reinforcement Attention Networks

    Jingyuan Wang, Yang Zhang, Ke Tang +2

    q-fin.TRcs.LGq-fin.STarXiv:1908.02646v12019
  16. Adapt to Adaptation: Learning Personalization for Cross-Silo Federated Learning

    Jun Luo, Shandong Wu

    cs.LGarXiv:2110.08394v32021
  17. DYffusion: A Dynamics-informed Diffusion Model for Spatiotemporal Forecasting

    Salva Rühling Cachay, Bo Zhao, Hailey Joren +1

    cs.LGcs.AIstat.MLarXiv:2306.01984v22023
  18. Parrot: Efficient Serving of LLM-based Applications with Semantic Variable

    Chaofan Lin, Zhenhua Han, Chengruidong Zhang +4

    cs.LGcs.AIarXiv:2405.19888v12024
  19. Learning Theory and Algorithms for Revenue Optimization in Second-Price Auctions with Reserve

    Mehryar Mohri, Andres Muñoz Medina

    cs.LGarXiv:1310.5665v32013
  20. Label Noise SGD Provably Prefers Flat Global Minimizers

    Alex Damian, Tengyu Ma, Jason D. Lee

    cs.LGcs.ITmath.OCarXiv:2106.06530v22021
  21. AnglE-optimized Text Embeddings

    Xianming Li, Jing Li

    cs.CLcs.AIcs.LGarXiv:2309.12871v92023
  22. Prediction and Clustering in Signed Networks: A Local to Global Perspective

    Kai-Yang Chiang, Cho-Jui Hsieh, Nagarajan Natarajan +2

    cs.SIcs.LGarXiv:1302.5145v22013
  23. Making Neural Programming Architectures Generalize via Recursion

    Jonathon Cai, Richard Shin, Dawn Song

    cs.LGcs.NEcs.PLarXiv:1704.06611v12017
  24. Direct-Manipulation Visualization of Deep Networks

    Daniel Smilkov, Shan Carter, D. Sculley +2

    cs.LGcs.HCstat.MLarXiv:1708.03788v12017
  25. Automated Variational Inference in Probabilistic Programming

    David Wingate, Theophane Weber

    stat.MLcs.AIcs.LGarXiv:1301.1299v12013
  26. LOCUS: Task-Aware Low-Rank Post-Training for Token-Efficient Language Generation

    Dongfang Zhao

    cs.CLcs.AIcs.LGarXiv:2609.11739v12026
  27. Obstacle Tower: A Generalization Challenge in Vision, Control, and Planning

    Arthur Juliani, Ahmed Khalifa, Vincent-Pierre Berges +6

    cs.AIcs.LGarXiv:1902.01378v22019
  28. Towards neural networks that provably know when they don't know

    Alexander Meinke, Matthias Hein

    cs.LGcs.CVstat.MLarXiv:1909.12180v22019
  29. Negative Self-Distillation: Learning to Reason by Avoiding Flaws

    Rongcan Pei, Zhepei Wei, Shuyao Xu +3

    cs.CLcs.LGarXiv:2609.11699v12026
  30. Global Encoding for Abstractive Summarization

    Junyang Lin, Xu Sun, Shuming Ma +1

    cs.CLcs.AIcs.LGarXiv:1805.03989v22018
  31. Generative AI in the Construction Industry: Opportunities & Challenges

    Prashnna Ghimire, Kyungki Kim, Manoj Acharya

    cs.AIcs.LGarXiv:2310.04427v12023
  32. On Learning Sets of Symmetric Elements

    Haggai Maron, Or Litany, Gal Chechik +1

    cs.LGstat.MLarXiv:2002.08599v42020
  33. Pre-gated MoE: An Algorithm-System Co-Design for Fast and Scalable Mixture-of-Expert Inference

    Ranggi Hwang, Jianyu Wei, Shijie Cao +4

    cs.LGcs.AIcs.ARarXiv:2308.12066v32023
  34. Structural priors for data-efficient language learning

    Yana Veitsman, Jonas Mayer Martins, Jonathan Lautenschlager +1

    cs.CLcs.AIcs.LGarXiv:2609.11505v12026
  35. An Analysis of ISO 26262: Using Machine Learning Safely in Automotive Software

    Rick Salay, Rodrigo Queiroz, Krzysztof Czarnecki

    cs.AIcs.LGcs.SEarXiv:1709.02435v12017
  36. FasterViT: Fast Vision Transformers with Hierarchical Attention

    Ali Hatamizadeh, Greg Heinrich, Hongxu Yin +4

    cs.CVcs.AIcs.LGarXiv:2306.06189v22023
  37. E-CONAN (Entailment, CONtradition And Neutral) Benchmarks: Arabic Textual Entailment and Natural Inference Datasets

    Khloud AL Jallad, Nada Ghneim, Ghaida Rebdawi

    cs.CLcs.AIcs.LGarXiv:2609.11334v12026
  38. XGBOD: Improving Supervised Outlier Detection with Unsupervised Representation Learning

    Yue Zhao, Maciej K. Hryniewicki

    cs.LGcs.DBcs.IRarXiv:1912.00290v12019
  39. Block-Recurrent Transformers

    DeLesley Hutchins, Imanol Schlag, Yuhuai Wu +2

    cs.LGcs.AIcs.NEarXiv:2203.07852v32022
  40. VCT: A Video Compression Transformer

    Fabian Mentzer, George Toderici, David Minnen +4

    cs.CVcs.LGeess.IVarXiv:2206.07307v22022
  41. DeepFilterNet: A Low Complexity Speech Enhancement Framework for Full-Band Audio based on Deep Filtering

    Hendrik Schröter, Alberto N. Escalante-B., Tobias Rosenkranz +1

    eess.AScs.LGeess.SParXiv:2110.05588v22021
  42. Performance-Efficiency Trade-offs in Unsupervised Pre-training for Speech Recognition

    Felix Wu, Kwangyoun Kim, Jing Pan +3

    cs.CLcs.LGcs.SDarXiv:2109.06870v12021
  43. Preference-based Online Learning with Dueling Bandits: A Survey

    Viktor Bengs, Robert Busa-Fekete, Adil El Mesaoudi-Paul +1

    cs.LGstat.MLarXiv:1807.11398v22018
  44. Radio Frequency Fingerprint Identification for LoRa Using Spectrogram and CNN

    Guanxiong Shen, Junqing Zhang, Alan Marshall +2

    eess.SPcs.LGarXiv:2101.01668v12020
  45. On Network Design Spaces for Visual Recognition

    Ilija Radosavovic, Justin Johnson, Saining Xie +2

    cs.CVcs.LGarXiv:1905.13214v12019
  46. Interpretable Distribution Features with Maximum Testing Power

    Wittawat Jitkrittum, Zoltan Szabo, Kacper Chwialkowski +1

    stat.MLcs.LGarXiv:1605.06796v22016
  47. AIDE: Fast and Communication Efficient Distributed Optimization

    Sashank J. Reddi, Jakub Konečný, Peter Richtárik +2

    math.OCcs.LGstat.MLarXiv:1608.06879v12016
  48. CKConv: Continuous Kernel Convolution For Sequential Data

    David W. Romero, Anna Kuzina, Erik J. Bekkers +2

    cs.LGarXiv:2102.02611v32021
  49. Improving Generalization via Scalable Neighborhood Component Analysis

    Zhirong Wu, Alexei A. Efros, Stella X. Yu

    cs.CVcs.LGarXiv:1808.04699v12018
  50. Projected Subgradient Methods for Learning Sparse Gaussians

    John Duchi, Stephen Gould, Daphne Koller

    cs.LGstat.MLarXiv:1206.3249v12012
  51. From voxels to pixels and back: Self-supervision in natural-image reconstruction from fMRI

    Roman Beliy, Guy Gaziv, Assaf Hoogi +3

    eess.IVcs.LGq-bio.NCarXiv:1907.02431v12019
  52. Optimal approximate matrix product in terms of stable rank

    Michael B. Cohen, Jelani Nelson, David P. Woodruff

    cs.DScs.LGstat.MLarXiv:1507.02268v32015
  53. DiT-3D: Exploring Plain Diffusion Transformers for 3D Shape Generation

    Shentong Mo, Enze Xie, Ruihang Chu +4

    cs.CVcs.AIcs.LGarXiv:2307.01831v12023
  54. Loss of Plasticity in Continual Deep Reinforcement Learning

    Zaheer Abbas, Rosie Zhao, Joseph Modayil +2

    cs.LGcs.AIarXiv:2303.07507v12023
  55. RiNALMo: General-Purpose RNA Language Models Can Generalize Well on Structure Prediction Tasks

    Rafael Josip Penić, Tin Vlašić, Roland G. Huber +2

    q-bio.BMcs.LGarXiv:2403.00043v22024
  56. A Fragility Spectrum for Recursive Language-Model Training

    Yangze Liu, Zhongyi Han

    cs.CLcs.AIcs.LGarXiv:2609.11149v12026
  57. Self-Distillation as Instance-Specific Label Smoothing

    Zhilu Zhang, Mert R. Sabuncu

    cs.LGstat.MLarXiv:2006.05065v22020
  58. Decentralized Federated Learning: A Survey on Security and Privacy

    Ehsan Hallaji, Roozbeh Razavi-Far, Mehrdad Saif +2

    cs.CRcs.AIcs.LGarXiv:2401.17319v12024
  59. A Survey on Uncertainty Quantification Methods for Deep Learning

    Wenchong He, Zhe Jiang, Tingsong Xiao +2

    cs.LGstat.MLarXiv:2302.13425v72023
  60. Penetrative AI: Making LLMs Comprehend the Physical World

    Huatao Xu, Liying Han, Qirui Yang +2

    cs.AIcs.LGarXiv:2310.09605v32023