Machine Learning
Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
1,621 to 1,680 of 20,193
The Lazy Neuron Phenomenon: On Emergence of Activation Sparsity in Transformers
Zonglin Li, Chong You, Srinadh Bhojanapalli +8
cs.LGcs.CLcs.CVarXiv:2210.06313v22022ECG Arrhythmia Classification Using Transfer Learning from 2-Dimensional Deep CNN Features
Milad Salem, Shayan Taheri, Jiann Shiun-Yuan
cs.LGcs.CVstat.MLarXiv:1812.04693v12018Reinforcement Learning Upside Down: Don't Predict Rewards -- Just Map Them to Actions
Juergen Schmidhuber
cs.AIcs.LGarXiv:1912.02875v22019Rapid trial-and-error learning with simulation supports flexible tool use and physical reasoning
Kelsey R. Allen, Kevin A. Smith, Joshua B. Tenenbaum
cs.AIcs.LGcs.ROarXiv:1907.09620v32019DiffiT: Diffusion Vision Transformers for Image Generation
Ali Hatamizadeh, Jiaming Song, Guilin Liu +2
cs.CVcs.AIcs.LGarXiv:2312.02139v32023Clinical Intervention Prediction and Understanding using Deep Networks
Harini Suresh, Nathan Hunt, Alistair Johnson +3
cs.LGarXiv:1705.08498v12017Neural Task Graphs: Generalizing to Unseen Tasks from a Single Video Demonstration
De-An Huang, Suraj Nair, Danfei Xu +5
cs.CVcs.AIcs.LGarXiv:1807.03480v22018Towards a Mathematical Understanding of Neural Network-Based Machine Learning: what we know and what we don't
Weinan E, Chao Ma, Stephan Wojtowytsch +1
cs.LGmath.NAstat.MLarXiv:2009.10713v32020Cross-Subject Transfer Learning in Human Activity Recognition Systems using Generative Adversarial Networks
Elnaz Soleimani, Ehsan Nazerfard
cs.LGstat.MLarXiv:1903.12489v12019Time Series Anomaly Detection Using Convolutional Neural Networks and Transfer Learning
Tailai Wen, Roy Keyes
cs.LGcs.CVstat.MLarXiv:1905.13628v12019A Deep Convolutional Neural Network for COVID-19 Detection Using Chest X-Rays
Pedro R. A. S. Bassi, Romis Attux
eess.IVcs.CVcs.LGarXiv:2005.01578v42020DeepACO: Neural-enhanced Ant Systems for Combinatorial Optimization
Haoran Ye, Jiarui Wang, Zhiguang Cao +2
cs.NEcs.AIcs.LGarXiv:2309.14032v22023Machine Unlearning: Solutions and Challenges
Jie Xu, Zihan Wu, Cong Wang +1
cs.LGcs.AIarXiv:2308.07061v32023Domain-Specific Hallucination Detection in Large Language Models
Varun Teja Chundru, Debasmita Biswas
cs.CLcs.AIcs.LGarXiv:2609.11878v12026AlphaStock: A Buying-Winners-and-Selling-Losers Investment Strategy using Interpretable Deep Reinforcement Attention Networks
Jingyuan Wang, Yang Zhang, Ke Tang +2
q-fin.TRcs.LGq-fin.STarXiv:1908.02646v12019Adapt to Adaptation: Learning Personalization for Cross-Silo Federated Learning
Jun Luo, Shandong Wu
cs.LGarXiv:2110.08394v32021DYffusion: A Dynamics-informed Diffusion Model for Spatiotemporal Forecasting
Salva Rühling Cachay, Bo Zhao, Hailey Joren +1
cs.LGcs.AIstat.MLarXiv:2306.01984v22023Parrot: Efficient Serving of LLM-based Applications with Semantic Variable
Chaofan Lin, Zhenhua Han, Chengruidong Zhang +4
cs.LGcs.AIarXiv:2405.19888v12024Learning Theory and Algorithms for Revenue Optimization in Second-Price Auctions with Reserve
Mehryar Mohri, Andres Muñoz Medina
cs.LGarXiv:1310.5665v32013Label Noise SGD Provably Prefers Flat Global Minimizers
Alex Damian, Tengyu Ma, Jason D. Lee
cs.LGcs.ITmath.OCarXiv:2106.06530v22021AnglE-optimized Text Embeddings
Xianming Li, Jing Li
cs.CLcs.AIcs.LGarXiv:2309.12871v92023Prediction and Clustering in Signed Networks: A Local to Global Perspective
Kai-Yang Chiang, Cho-Jui Hsieh, Nagarajan Natarajan +2
cs.SIcs.LGarXiv:1302.5145v22013Making Neural Programming Architectures Generalize via Recursion
Jonathon Cai, Richard Shin, Dawn Song
cs.LGcs.NEcs.PLarXiv:1704.06611v12017Direct-Manipulation Visualization of Deep Networks
Daniel Smilkov, Shan Carter, D. Sculley +2
cs.LGcs.HCstat.MLarXiv:1708.03788v12017Automated Variational Inference in Probabilistic Programming
David Wingate, Theophane Weber
stat.MLcs.AIcs.LGarXiv:1301.1299v12013LOCUS: Task-Aware Low-Rank Post-Training for Token-Efficient Language Generation
Dongfang Zhao
cs.CLcs.AIcs.LGarXiv:2609.11739v12026Obstacle Tower: A Generalization Challenge in Vision, Control, and Planning
Arthur Juliani, Ahmed Khalifa, Vincent-Pierre Berges +6
cs.AIcs.LGarXiv:1902.01378v22019Towards neural networks that provably know when they don't know
Alexander Meinke, Matthias Hein
cs.LGcs.CVstat.MLarXiv:1909.12180v22019Negative Self-Distillation: Learning to Reason by Avoiding Flaws
Rongcan Pei, Zhepei Wei, Shuyao Xu +3
cs.CLcs.LGarXiv:2609.11699v12026Global Encoding for Abstractive Summarization
Junyang Lin, Xu Sun, Shuming Ma +1
cs.CLcs.AIcs.LGarXiv:1805.03989v22018Generative AI in the Construction Industry: Opportunities & Challenges
Prashnna Ghimire, Kyungki Kim, Manoj Acharya
cs.AIcs.LGarXiv:2310.04427v12023On Learning Sets of Symmetric Elements
Haggai Maron, Or Litany, Gal Chechik +1
cs.LGstat.MLarXiv:2002.08599v42020Pre-gated MoE: An Algorithm-System Co-Design for Fast and Scalable Mixture-of-Expert Inference
Ranggi Hwang, Jianyu Wei, Shijie Cao +4
cs.LGcs.AIcs.ARarXiv:2308.12066v32023Structural priors for data-efficient language learning
Yana Veitsman, Jonas Mayer Martins, Jonathan Lautenschlager +1
cs.CLcs.AIcs.LGarXiv:2609.11505v12026An Analysis of ISO 26262: Using Machine Learning Safely in Automotive Software
Rick Salay, Rodrigo Queiroz, Krzysztof Czarnecki
cs.AIcs.LGcs.SEarXiv:1709.02435v12017FasterViT: Fast Vision Transformers with Hierarchical Attention
Ali Hatamizadeh, Greg Heinrich, Hongxu Yin +4
cs.CVcs.AIcs.LGarXiv:2306.06189v22023E-CONAN (Entailment, CONtradition And Neutral) Benchmarks: Arabic Textual Entailment and Natural Inference Datasets
Khloud AL Jallad, Nada Ghneim, Ghaida Rebdawi
cs.CLcs.AIcs.LGarXiv:2609.11334v12026XGBOD: Improving Supervised Outlier Detection with Unsupervised Representation Learning
Yue Zhao, Maciej K. Hryniewicki
cs.LGcs.DBcs.IRarXiv:1912.00290v12019Block-Recurrent Transformers
DeLesley Hutchins, Imanol Schlag, Yuhuai Wu +2
cs.LGcs.AIcs.NEarXiv:2203.07852v32022VCT: A Video Compression Transformer
Fabian Mentzer, George Toderici, David Minnen +4
cs.CVcs.LGeess.IVarXiv:2206.07307v22022DeepFilterNet: A Low Complexity Speech Enhancement Framework for Full-Band Audio based on Deep Filtering
Hendrik Schröter, Alberto N. Escalante-B., Tobias Rosenkranz +1
eess.AScs.LGeess.SParXiv:2110.05588v22021Performance-Efficiency Trade-offs in Unsupervised Pre-training for Speech Recognition
Felix Wu, Kwangyoun Kim, Jing Pan +3
cs.CLcs.LGcs.SDarXiv:2109.06870v12021Preference-based Online Learning with Dueling Bandits: A Survey
Viktor Bengs, Robert Busa-Fekete, Adil El Mesaoudi-Paul +1
cs.LGstat.MLarXiv:1807.11398v22018Radio Frequency Fingerprint Identification for LoRa Using Spectrogram and CNN
Guanxiong Shen, Junqing Zhang, Alan Marshall +2
eess.SPcs.LGarXiv:2101.01668v12020On Network Design Spaces for Visual Recognition
Ilija Radosavovic, Justin Johnson, Saining Xie +2
cs.CVcs.LGarXiv:1905.13214v12019Interpretable Distribution Features with Maximum Testing Power
Wittawat Jitkrittum, Zoltan Szabo, Kacper Chwialkowski +1
stat.MLcs.LGarXiv:1605.06796v22016AIDE: Fast and Communication Efficient Distributed Optimization
Sashank J. Reddi, Jakub Konečný, Peter Richtárik +2
math.OCcs.LGstat.MLarXiv:1608.06879v12016CKConv: Continuous Kernel Convolution For Sequential Data
David W. Romero, Anna Kuzina, Erik J. Bekkers +2
cs.LGarXiv:2102.02611v32021Improving Generalization via Scalable Neighborhood Component Analysis
Zhirong Wu, Alexei A. Efros, Stella X. Yu
cs.CVcs.LGarXiv:1808.04699v12018Projected Subgradient Methods for Learning Sparse Gaussians
John Duchi, Stephen Gould, Daphne Koller
cs.LGstat.MLarXiv:1206.3249v12012From voxels to pixels and back: Self-supervision in natural-image reconstruction from fMRI
Roman Beliy, Guy Gaziv, Assaf Hoogi +3
eess.IVcs.LGq-bio.NCarXiv:1907.02431v12019Optimal approximate matrix product in terms of stable rank
Michael B. Cohen, Jelani Nelson, David P. Woodruff
cs.DScs.LGstat.MLarXiv:1507.02268v32015DiT-3D: Exploring Plain Diffusion Transformers for 3D Shape Generation
Shentong Mo, Enze Xie, Ruihang Chu +4
cs.CVcs.AIcs.LGarXiv:2307.01831v12023Loss of Plasticity in Continual Deep Reinforcement Learning
Zaheer Abbas, Rosie Zhao, Joseph Modayil +2
cs.LGcs.AIarXiv:2303.07507v12023RiNALMo: General-Purpose RNA Language Models Can Generalize Well on Structure Prediction Tasks
Rafael Josip Penić, Tin Vlašić, Roland G. Huber +2
q-bio.BMcs.LGarXiv:2403.00043v22024A Fragility Spectrum for Recursive Language-Model Training
Yangze Liu, Zhongyi Han
cs.CLcs.AIcs.LGarXiv:2609.11149v12026Self-Distillation as Instance-Specific Label Smoothing
Zhilu Zhang, Mert R. Sabuncu
cs.LGstat.MLarXiv:2006.05065v22020Decentralized Federated Learning: A Survey on Security and Privacy
Ehsan Hallaji, Roozbeh Razavi-Far, Mehrdad Saif +2
cs.CRcs.AIcs.LGarXiv:2401.17319v12024A Survey on Uncertainty Quantification Methods for Deep Learning
Wenchong He, Zhe Jiang, Tingsong Xiao +2
cs.LGstat.MLarXiv:2302.13425v72023Penetrative AI: Making LLMs Comprehend the Physical World
Huatao Xu, Liying Han, Qirui Yang +2
cs.AIcs.LGarXiv:2310.09605v32023