Machine Learning

Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

1,921 to 1,980 of 20,454

  1. Direct-Manipulation Visualization of Deep Networks

    Daniel Smilkov, Shan Carter, D. Sculley +2

    cs.LGcs.HCstat.MLarXiv:1708.03788v12017
  2. Automated Variational Inference in Probabilistic Programming

    David Wingate, Theophane Weber

    stat.MLcs.AIcs.LGarXiv:1301.1299v12013
  3. LOCUS: Task-Aware Low-Rank Post-Training for Token-Efficient Language Generation

    Dongfang Zhao

    cs.CLcs.AIcs.LGarXiv:2609.11739v12026
  4. Obstacle Tower: A Generalization Challenge in Vision, Control, and Planning

    Arthur Juliani, Ahmed Khalifa, Vincent-Pierre Berges +6

    cs.AIcs.LGarXiv:1902.01378v22019
  5. Towards neural networks that provably know when they don't know

    Alexander Meinke, Matthias Hein

    cs.LGcs.CVstat.MLarXiv:1909.12180v22019
  6. Negative Self-Distillation: Learning to Reason by Avoiding Flaws

    Rongcan Pei, Zhepei Wei, Shuyao Xu +3

    cs.CLcs.LGarXiv:2609.11699v12026
  7. Global Encoding for Abstractive Summarization

    Junyang Lin, Xu Sun, Shuming Ma +1

    cs.CLcs.AIcs.LGarXiv:1805.03989v22018
  8. Generative AI in the Construction Industry: Opportunities & Challenges

    Prashnna Ghimire, Kyungki Kim, Manoj Acharya

    cs.AIcs.LGarXiv:2310.04427v12023
  9. On Learning Sets of Symmetric Elements

    Haggai Maron, Or Litany, Gal Chechik +1

    cs.LGstat.MLarXiv:2002.08599v42020
  10. Pre-gated MoE: An Algorithm-System Co-Design for Fast and Scalable Mixture-of-Expert Inference

    Ranggi Hwang, Jianyu Wei, Shijie Cao +4

    cs.LGcs.AIcs.ARarXiv:2308.12066v32023
  11. Structural priors for data-efficient language learning

    Yana Veitsman, Jonas Mayer Martins, Jonathan Lautenschlager +1

    cs.CLcs.AIcs.LGarXiv:2609.11505v12026
  12. An Analysis of ISO 26262: Using Machine Learning Safely in Automotive Software

    Rick Salay, Rodrigo Queiroz, Krzysztof Czarnecki

    cs.AIcs.LGcs.SEarXiv:1709.02435v12017
  13. FasterViT: Fast Vision Transformers with Hierarchical Attention

    Ali Hatamizadeh, Greg Heinrich, Hongxu Yin +4

    cs.CVcs.AIcs.LGarXiv:2306.06189v22023
  14. E-CONAN (Entailment, CONtradition And Neutral) Benchmarks: Arabic Textual Entailment and Natural Inference Datasets

    Khloud AL Jallad, Nada Ghneim, Ghaida Rebdawi

    cs.CLcs.AIcs.LGarXiv:2609.11334v12026
  15. XGBOD: Improving Supervised Outlier Detection with Unsupervised Representation Learning

    Yue Zhao, Maciej K. Hryniewicki

    cs.LGcs.DBcs.IRarXiv:1912.00290v12019
  16. Block-Recurrent Transformers

    DeLesley Hutchins, Imanol Schlag, Yuhuai Wu +2

    cs.LGcs.AIcs.NEarXiv:2203.07852v32022
  17. VCT: A Video Compression Transformer

    Fabian Mentzer, George Toderici, David Minnen +4

    cs.CVcs.LGeess.IVarXiv:2206.07307v22022
  18. DeepFilterNet: A Low Complexity Speech Enhancement Framework for Full-Band Audio based on Deep Filtering

    Hendrik Schröter, Alberto N. Escalante-B., Tobias Rosenkranz +1

    eess.AScs.LGeess.SParXiv:2110.05588v22021
  19. Performance-Efficiency Trade-offs in Unsupervised Pre-training for Speech Recognition

    Felix Wu, Kwangyoun Kim, Jing Pan +3

    cs.CLcs.LGcs.SDarXiv:2109.06870v12021
  20. Preference-based Online Learning with Dueling Bandits: A Survey

    Viktor Bengs, Robert Busa-Fekete, Adil El Mesaoudi-Paul +1

    cs.LGstat.MLarXiv:1807.11398v22018
  21. Radio Frequency Fingerprint Identification for LoRa Using Spectrogram and CNN

    Guanxiong Shen, Junqing Zhang, Alan Marshall +2

    eess.SPcs.LGarXiv:2101.01668v12020
  22. On Network Design Spaces for Visual Recognition

    Ilija Radosavovic, Justin Johnson, Saining Xie +2

    cs.CVcs.LGarXiv:1905.13214v12019
  23. Interpretable Distribution Features with Maximum Testing Power

    Wittawat Jitkrittum, Zoltan Szabo, Kacper Chwialkowski +1

    stat.MLcs.LGarXiv:1605.06796v22016
  24. AIDE: Fast and Communication Efficient Distributed Optimization

    Sashank J. Reddi, Jakub Konečný, Peter Richtárik +2

    math.OCcs.LGstat.MLarXiv:1608.06879v12016
  25. CKConv: Continuous Kernel Convolution For Sequential Data

    David W. Romero, Anna Kuzina, Erik J. Bekkers +2

    cs.LGarXiv:2102.02611v32021
  26. Improving Generalization via Scalable Neighborhood Component Analysis

    Zhirong Wu, Alexei A. Efros, Stella X. Yu

    cs.CVcs.LGarXiv:1808.04699v12018
  27. Projected Subgradient Methods for Learning Sparse Gaussians

    John Duchi, Stephen Gould, Daphne Koller

    cs.LGstat.MLarXiv:1206.3249v12012
  28. From voxels to pixels and back: Self-supervision in natural-image reconstruction from fMRI

    Roman Beliy, Guy Gaziv, Assaf Hoogi +3

    eess.IVcs.LGq-bio.NCarXiv:1907.02431v12019
  29. Optimal approximate matrix product in terms of stable rank

    Michael B. Cohen, Jelani Nelson, David P. Woodruff

    cs.DScs.LGstat.MLarXiv:1507.02268v32015
  30. DiT-3D: Exploring Plain Diffusion Transformers for 3D Shape Generation

    Shentong Mo, Enze Xie, Ruihang Chu +4

    cs.CVcs.AIcs.LGarXiv:2307.01831v12023
  31. Loss of Plasticity in Continual Deep Reinforcement Learning

    Zaheer Abbas, Rosie Zhao, Joseph Modayil +2

    cs.LGcs.AIarXiv:2303.07507v12023
  32. RiNALMo: General-Purpose RNA Language Models Can Generalize Well on Structure Prediction Tasks

    Rafael Josip Penić, Tin Vlašić, Roland G. Huber +2

    q-bio.BMcs.LGarXiv:2403.00043v22024
  33. A Fragility Spectrum for Recursive Language-Model Training

    Yangze Liu, Zhongyi Han

    cs.CLcs.AIcs.LGarXiv:2609.11149v12026
  34. Self-Distillation as Instance-Specific Label Smoothing

    Zhilu Zhang, Mert R. Sabuncu

    cs.LGstat.MLarXiv:2006.05065v22020
  35. Decentralized Federated Learning: A Survey on Security and Privacy

    Ehsan Hallaji, Roozbeh Razavi-Far, Mehrdad Saif +2

    cs.CRcs.AIcs.LGarXiv:2401.17319v12024
  36. A Survey on Uncertainty Quantification Methods for Deep Learning

    Wenchong He, Zhe Jiang, Tingsong Xiao +2

    cs.LGstat.MLarXiv:2302.13425v72023
  37. Penetrative AI: Making LLMs Comprehend the Physical World

    Huatao Xu, Liying Han, Qirui Yang +2

    cs.AIcs.LGarXiv:2310.09605v32023
  38. LaVR: Scene Latent Conditioned Generative Video Trajectory Re-Rendering using Large 4D Reconstruction Models

    Mingyang Xie, Numair Khan, Tianfu Wang +8

    cs.CVcs.LGarXiv:2601.14674v22026
  39. HittER: Hierarchical Transformers for Knowledge Graph Embeddings

    Sanxing Chen, Xiaodong Liu, Jianfeng Gao +3

    cs.CLcs.LGarXiv:2008.12813v22020
  40. Likely to stop? Predicting Stopout in Massive Open Online Courses

    Colin Taylor, Kalyan Veeramachaneni, Una-May O'Reilly

    cs.CYcs.LGarXiv:1408.3382v12014
  41. Does Neural Machine Translation Benefit from Larger Context?

    Sebastien Jean, Stanislas Lauly, Orhan Firat +1

    stat.MLcs.CLcs.LGarXiv:1704.05135v12017
  42. TableFormer: Table Structure Understanding with Transformers

    Ahmed Nassar, Nikolaos Livathinos, Maksym Lysak +1

    cs.CVcs.LGarXiv:2203.01017v22022
  43. ImageCAS: A Large-Scale Dataset and Benchmark for Coronary Artery Segmentation based on Computed Tomography Angiography Images

    An Zeng, Chunbiao Wu, Meiping Huang +10

    eess.IVcs.LGarXiv:2211.01607v22022
  44. HyperImpute: Generalized Iterative Imputation with Automatic Model Selection

    Daniel Jarrett, Bogdan Cebere, Tennison Liu +2

    stat.MLcs.LGarXiv:2206.07769v12022
  45. Combinatorial Multi-Armed Bandit with General Reward Functions

    Wei Chen, Wei Hu, Fu Li +3

    cs.LGcs.DSstat.MLarXiv:1610.06603v42016
  46. Have Faith in Faithfulness: Going Beyond Circuit Overlap When Finding Model Mechanisms

    Michael Hanna, Sandro Pezzelle, Yonatan Belinkov

    cs.LGcs.CLarXiv:2403.17806v22024
  47. How Useful is Self-Supervised Pretraining for Visual Tasks?

    Alejandro Newell, Jia Deng

    cs.CVcs.LGarXiv:2003.14323v12020
  48. Structured Graph Learning for Clustering and Semi-supervised Classification

    Zhao Kang, Chong Peng, Qiang Cheng +4

    cs.LGcs.AIcs.CVarXiv:2008.13429v12020
  49. Extreme Gradient Boosting for Yield Estimation compared with Deep Learning Approaches

    Florian Huber, Artem Yushchenko, Benedikt Stratmann +1

    cs.LGarXiv:2208.12633v12022
  50. TFAD: A Decomposition Time Series Anomaly Detection Architecture with Time-Frequency Analysis

    Chaoli Zhang, Tian Zhou, Qingsong Wen +1

    cs.LGcs.AIarXiv:2210.09693v22022
  51. Self-supervised Knowledge Distillation Using Singular Value Decomposition

    Seung Hyun Lee, Dae Ha Kim, Byung Cheol Song

    cs.LGcs.CVstat.MLarXiv:1807.06819v12018
  52. Rebalancing Token Importance in Language Models with TF-IDF Weighted Cross-Entropy Loss

    Zhijian Li, Stefan Larson, Kevin Leach

    cs.CLcs.LGarXiv:2609.11029v12026
  53. Time Series Change Point Detection with Self-Supervised Contrastive Predictive Coding

    Shohreh Deldari, Daniel V. Smith, Hao Xue +1

    cs.LGcs.AIcs.CVarXiv:2011.14097v52020
  54. Model-based Exploration of the Frontier of Behaviours for Deep Learning System Testing

    Vincenzo Riccio, Paolo Tonella

    cs.SEcs.AIcs.LGarXiv:2007.02787v12020
  55. DMD: A Large-Scale Multi-Modal Driver Monitoring Dataset for Attention and Alertness Analysis

    Juan Diego Ortega, Neslihan Kose, Paola Cañas +5

    cs.CVcs.LGeess.IVarXiv:2008.12085v12020
  56. Influence-Preserving Proxies for Gradient-Based Data Selection in LLM Fine-tuning

    Sirui Chen, Yunzhe Qi, Mengting Ai +4

    cs.LGarXiv:2602.17835v12026
  57. PEER: A Comprehensive and Multi-Task Benchmark for Protein Sequence Understanding

    Minghao Xu, Zuobai Zhang, Jiarui Lu +5

    cs.LGarXiv:2206.02096v22022
  58. A Positive Case for Faithfulness: LLM Self-Explanations Help Predict Model Behavior

    Harry Mayne, Justin Singh Kang, Dewi Gould +3

    cs.AIcs.LGarXiv:2602.02639v12026
  59. When Noise Fabricates Bias: The Fragility of LLM-as-a-Judge Bias Measurement under Noisy Text

    DongHyun Ryu, Jaehyeok Lee, YeongJun Hwang +1

    cs.CLcs.LGarXiv:2609.11067v12026
  60. Cultural Alignment in Large Language Models: An Explanatory Analysis Based on Hofstede's Cultural Dimensions

    Reem I. Masoud, Ziquan Liu, Martin Ferianc +2

    cs.CYcs.CLcs.LGarXiv:2309.12342v22023