Machine Learning
Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
1,921 to 1,980 of 20,454
Direct-Manipulation Visualization of Deep Networks
Daniel Smilkov, Shan Carter, D. Sculley +2
cs.LGcs.HCstat.MLarXiv:1708.03788v12017Automated Variational Inference in Probabilistic Programming
David Wingate, Theophane Weber
stat.MLcs.AIcs.LGarXiv:1301.1299v12013LOCUS: Task-Aware Low-Rank Post-Training for Token-Efficient Language Generation
Dongfang Zhao
cs.CLcs.AIcs.LGarXiv:2609.11739v12026Obstacle Tower: A Generalization Challenge in Vision, Control, and Planning
Arthur Juliani, Ahmed Khalifa, Vincent-Pierre Berges +6
cs.AIcs.LGarXiv:1902.01378v22019Towards neural networks that provably know when they don't know
Alexander Meinke, Matthias Hein
cs.LGcs.CVstat.MLarXiv:1909.12180v22019Negative Self-Distillation: Learning to Reason by Avoiding Flaws
Rongcan Pei, Zhepei Wei, Shuyao Xu +3
cs.CLcs.LGarXiv:2609.11699v12026Global Encoding for Abstractive Summarization
Junyang Lin, Xu Sun, Shuming Ma +1
cs.CLcs.AIcs.LGarXiv:1805.03989v22018Generative AI in the Construction Industry: Opportunities & Challenges
Prashnna Ghimire, Kyungki Kim, Manoj Acharya
cs.AIcs.LGarXiv:2310.04427v12023On Learning Sets of Symmetric Elements
Haggai Maron, Or Litany, Gal Chechik +1
cs.LGstat.MLarXiv:2002.08599v42020Pre-gated MoE: An Algorithm-System Co-Design for Fast and Scalable Mixture-of-Expert Inference
Ranggi Hwang, Jianyu Wei, Shijie Cao +4
cs.LGcs.AIcs.ARarXiv:2308.12066v32023Structural priors for data-efficient language learning
Yana Veitsman, Jonas Mayer Martins, Jonathan Lautenschlager +1
cs.CLcs.AIcs.LGarXiv:2609.11505v12026An Analysis of ISO 26262: Using Machine Learning Safely in Automotive Software
Rick Salay, Rodrigo Queiroz, Krzysztof Czarnecki
cs.AIcs.LGcs.SEarXiv:1709.02435v12017FasterViT: Fast Vision Transformers with Hierarchical Attention
Ali Hatamizadeh, Greg Heinrich, Hongxu Yin +4
cs.CVcs.AIcs.LGarXiv:2306.06189v22023E-CONAN (Entailment, CONtradition And Neutral) Benchmarks: Arabic Textual Entailment and Natural Inference Datasets
Khloud AL Jallad, Nada Ghneim, Ghaida Rebdawi
cs.CLcs.AIcs.LGarXiv:2609.11334v12026XGBOD: Improving Supervised Outlier Detection with Unsupervised Representation Learning
Yue Zhao, Maciej K. Hryniewicki
cs.LGcs.DBcs.IRarXiv:1912.00290v12019Block-Recurrent Transformers
DeLesley Hutchins, Imanol Schlag, Yuhuai Wu +2
cs.LGcs.AIcs.NEarXiv:2203.07852v32022VCT: A Video Compression Transformer
Fabian Mentzer, George Toderici, David Minnen +4
cs.CVcs.LGeess.IVarXiv:2206.07307v22022DeepFilterNet: A Low Complexity Speech Enhancement Framework for Full-Band Audio based on Deep Filtering
Hendrik Schröter, Alberto N. Escalante-B., Tobias Rosenkranz +1
eess.AScs.LGeess.SParXiv:2110.05588v22021Performance-Efficiency Trade-offs in Unsupervised Pre-training for Speech Recognition
Felix Wu, Kwangyoun Kim, Jing Pan +3
cs.CLcs.LGcs.SDarXiv:2109.06870v12021Preference-based Online Learning with Dueling Bandits: A Survey
Viktor Bengs, Robert Busa-Fekete, Adil El Mesaoudi-Paul +1
cs.LGstat.MLarXiv:1807.11398v22018Radio Frequency Fingerprint Identification for LoRa Using Spectrogram and CNN
Guanxiong Shen, Junqing Zhang, Alan Marshall +2
eess.SPcs.LGarXiv:2101.01668v12020On Network Design Spaces for Visual Recognition
Ilija Radosavovic, Justin Johnson, Saining Xie +2
cs.CVcs.LGarXiv:1905.13214v12019Interpretable Distribution Features with Maximum Testing Power
Wittawat Jitkrittum, Zoltan Szabo, Kacper Chwialkowski +1
stat.MLcs.LGarXiv:1605.06796v22016AIDE: Fast and Communication Efficient Distributed Optimization
Sashank J. Reddi, Jakub Konečný, Peter Richtárik +2
math.OCcs.LGstat.MLarXiv:1608.06879v12016CKConv: Continuous Kernel Convolution For Sequential Data
David W. Romero, Anna Kuzina, Erik J. Bekkers +2
cs.LGarXiv:2102.02611v32021Improving Generalization via Scalable Neighborhood Component Analysis
Zhirong Wu, Alexei A. Efros, Stella X. Yu
cs.CVcs.LGarXiv:1808.04699v12018Projected Subgradient Methods for Learning Sparse Gaussians
John Duchi, Stephen Gould, Daphne Koller
cs.LGstat.MLarXiv:1206.3249v12012From voxels to pixels and back: Self-supervision in natural-image reconstruction from fMRI
Roman Beliy, Guy Gaziv, Assaf Hoogi +3
eess.IVcs.LGq-bio.NCarXiv:1907.02431v12019Optimal approximate matrix product in terms of stable rank
Michael B. Cohen, Jelani Nelson, David P. Woodruff
cs.DScs.LGstat.MLarXiv:1507.02268v32015DiT-3D: Exploring Plain Diffusion Transformers for 3D Shape Generation
Shentong Mo, Enze Xie, Ruihang Chu +4
cs.CVcs.AIcs.LGarXiv:2307.01831v12023Loss of Plasticity in Continual Deep Reinforcement Learning
Zaheer Abbas, Rosie Zhao, Joseph Modayil +2
cs.LGcs.AIarXiv:2303.07507v12023RiNALMo: General-Purpose RNA Language Models Can Generalize Well on Structure Prediction Tasks
Rafael Josip Penić, Tin Vlašić, Roland G. Huber +2
q-bio.BMcs.LGarXiv:2403.00043v22024A Fragility Spectrum for Recursive Language-Model Training
Yangze Liu, Zhongyi Han
cs.CLcs.AIcs.LGarXiv:2609.11149v12026Self-Distillation as Instance-Specific Label Smoothing
Zhilu Zhang, Mert R. Sabuncu
cs.LGstat.MLarXiv:2006.05065v22020Decentralized Federated Learning: A Survey on Security and Privacy
Ehsan Hallaji, Roozbeh Razavi-Far, Mehrdad Saif +2
cs.CRcs.AIcs.LGarXiv:2401.17319v12024A Survey on Uncertainty Quantification Methods for Deep Learning
Wenchong He, Zhe Jiang, Tingsong Xiao +2
cs.LGstat.MLarXiv:2302.13425v72023Penetrative AI: Making LLMs Comprehend the Physical World
Huatao Xu, Liying Han, Qirui Yang +2
cs.AIcs.LGarXiv:2310.09605v32023LaVR: Scene Latent Conditioned Generative Video Trajectory Re-Rendering using Large 4D Reconstruction Models
Mingyang Xie, Numair Khan, Tianfu Wang +8
cs.CVcs.LGarXiv:2601.14674v22026HittER: Hierarchical Transformers for Knowledge Graph Embeddings
Sanxing Chen, Xiaodong Liu, Jianfeng Gao +3
cs.CLcs.LGarXiv:2008.12813v22020Likely to stop? Predicting Stopout in Massive Open Online Courses
Colin Taylor, Kalyan Veeramachaneni, Una-May O'Reilly
cs.CYcs.LGarXiv:1408.3382v12014Does Neural Machine Translation Benefit from Larger Context?
Sebastien Jean, Stanislas Lauly, Orhan Firat +1
stat.MLcs.CLcs.LGarXiv:1704.05135v12017TableFormer: Table Structure Understanding with Transformers
Ahmed Nassar, Nikolaos Livathinos, Maksym Lysak +1
cs.CVcs.LGarXiv:2203.01017v22022ImageCAS: A Large-Scale Dataset and Benchmark for Coronary Artery Segmentation based on Computed Tomography Angiography Images
An Zeng, Chunbiao Wu, Meiping Huang +10
eess.IVcs.LGarXiv:2211.01607v22022HyperImpute: Generalized Iterative Imputation with Automatic Model Selection
Daniel Jarrett, Bogdan Cebere, Tennison Liu +2
stat.MLcs.LGarXiv:2206.07769v12022Combinatorial Multi-Armed Bandit with General Reward Functions
Wei Chen, Wei Hu, Fu Li +3
cs.LGcs.DSstat.MLarXiv:1610.06603v42016Have Faith in Faithfulness: Going Beyond Circuit Overlap When Finding Model Mechanisms
Michael Hanna, Sandro Pezzelle, Yonatan Belinkov
cs.LGcs.CLarXiv:2403.17806v22024How Useful is Self-Supervised Pretraining for Visual Tasks?
Alejandro Newell, Jia Deng
cs.CVcs.LGarXiv:2003.14323v12020Structured Graph Learning for Clustering and Semi-supervised Classification
Zhao Kang, Chong Peng, Qiang Cheng +4
cs.LGcs.AIcs.CVarXiv:2008.13429v12020Extreme Gradient Boosting for Yield Estimation compared with Deep Learning Approaches
Florian Huber, Artem Yushchenko, Benedikt Stratmann +1
cs.LGarXiv:2208.12633v12022TFAD: A Decomposition Time Series Anomaly Detection Architecture with Time-Frequency Analysis
Chaoli Zhang, Tian Zhou, Qingsong Wen +1
cs.LGcs.AIarXiv:2210.09693v22022Self-supervised Knowledge Distillation Using Singular Value Decomposition
Seung Hyun Lee, Dae Ha Kim, Byung Cheol Song
cs.LGcs.CVstat.MLarXiv:1807.06819v12018Rebalancing Token Importance in Language Models with TF-IDF Weighted Cross-Entropy Loss
Zhijian Li, Stefan Larson, Kevin Leach
cs.CLcs.LGarXiv:2609.11029v12026Time Series Change Point Detection with Self-Supervised Contrastive Predictive Coding
Shohreh Deldari, Daniel V. Smith, Hao Xue +1
cs.LGcs.AIcs.CVarXiv:2011.14097v52020Model-based Exploration of the Frontier of Behaviours for Deep Learning System Testing
Vincenzo Riccio, Paolo Tonella
cs.SEcs.AIcs.LGarXiv:2007.02787v12020DMD: A Large-Scale Multi-Modal Driver Monitoring Dataset for Attention and Alertness Analysis
Juan Diego Ortega, Neslihan Kose, Paola Cañas +5
cs.CVcs.LGeess.IVarXiv:2008.12085v12020Influence-Preserving Proxies for Gradient-Based Data Selection in LLM Fine-tuning
Sirui Chen, Yunzhe Qi, Mengting Ai +4
cs.LGarXiv:2602.17835v12026PEER: A Comprehensive and Multi-Task Benchmark for Protein Sequence Understanding
Minghao Xu, Zuobai Zhang, Jiarui Lu +5
cs.LGarXiv:2206.02096v22022A Positive Case for Faithfulness: LLM Self-Explanations Help Predict Model Behavior
Harry Mayne, Justin Singh Kang, Dewi Gould +3
cs.AIcs.LGarXiv:2602.02639v12026When Noise Fabricates Bias: The Fragility of LLM-as-a-Judge Bias Measurement under Noisy Text
DongHyun Ryu, Jaehyeok Lee, YeongJun Hwang +1
cs.CLcs.LGarXiv:2609.11067v12026Cultural Alignment in Large Language Models: An Explanatory Analysis Based on Hofstede's Cultural Dimensions
Reem I. Masoud, Ziquan Liu, Martin Ferianc +2
cs.CYcs.CLcs.LGarXiv:2309.12342v22023