Machine Learning

Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

14,461 to 14,520 of 20,205

  1. A novel time-frequency Transformer based on self-attention mechanism and its application in fault diagnosis of rolling bearings

    Yifei Ding, Minping Jia, Qiuhua Miao +1

    cs.AIcs.LGeess.SParXiv:2104.09079v32021
  2. A Comprehensive Survey on Community Detection with Deep Learning

    Xing Su, Shan Xue, Fanzhen Liu +9

    cs.SIcs.AIcs.LGarXiv:2105.12584v22021
  3. Agentless: Demystifying LLM-based Software Engineering Agents

    Chunqiu Steven Xia, Yinlin Deng, Soren Dunn +1

    cs.SEcs.AIcs.CLarXiv:2407.01489v22024
  4. Differentially Private Learning with Adaptive Clipping

    Galen Andrew, Om Thakkar, H. Brendan McMahan +1

    cs.LGstat.MLarXiv:1905.03871v52019
  5. Lagrange Coded Computing: Optimal Design for Resiliency, Security and Privacy

    Qian Yu, Songze Li, Netanel Raviv +3

    cs.ITcs.DCcs.LGarXiv:1806.00939v42018
  6. Selection Bias Correction in Retail Intelligence

    Spandan Ghose Chowdhury

    cs.AIcs.LGarXiv:2608.26156v12026
  7. Explainable Artificial Intelligence for Customer Churn Prediction in Telecommunications: A Framework for CRM Integration

    Sandeep Gaddamwar

    cs.AIcs.LGarXiv:2608.26151v12026
  8. Neural network models and deep learning - a primer for biologists

    Nikolaus Kriegeskorte, Tal Golan

    q-bio.NCcs.LGcs.NEarXiv:1902.04704v22019
  9. Towards Robust Neural Networks via Random Self-ensemble

    Xuanqing Liu, Minhao Cheng, Huan Zhang +1

    cs.LGcs.CRstat.MLarXiv:1712.00673v22017
  10. Cyclical Annealing Schedule: A Simple Approach to Mitigating KL Vanishing

    Hao Fu, Chunyuan Li, Xiaodong Liu +3

    cs.LGcs.AIcs.CLarXiv:1903.10145v32019
  11. Deep Structured Energy Based Models for Anomaly Detection

    Shuangfei Zhai, Yu Cheng, Weining Lu +1

    cs.LGstat.MLarXiv:1605.07717v22016
  12. CLUE: A Chinese Language Understanding Evaluation Benchmark

    Liang Xu, Hai Hu, Xuanwei Zhang +29

    cs.CLcs.LGarXiv:2004.05986v32020
  13. Refusal Is Not Robustness: Auditing Confident Fabrication in Large Language Models on a Provably Uninformative Clinical Pain Speech Transcript

    Sagnik De, Sreenija Pavuluri

    cs.AIcs.ETcs.LGarXiv:2608.26167v12026
  14. Auxiliary Deep Generative Models

    Lars Maaløe, Casper Kaae Sønderby, Søren Kaae Sønderby +1

    stat.MLcs.AIcs.LGarXiv:1602.05473v42016
  15. Skip-GANomaly: Skip Connected and Adversarially Trained Encoder-Decoder Anomaly Detection

    Samet Akçay, Amir Atapour-Abarghouei, Toby P. Breckon

    cs.CVcs.LGarXiv:1901.08954v12019
  16. A Practical Algorithm for Topic Modeling with Provable Guarantees

    Sanjeev Arora, Rong Ge, Yoni Halpern +5

    cs.LGcs.DSstat.MLarXiv:1212.4777v12012
  17. Swin-UMamba: Mamba-based UNet with ImageNet-based pretraining

    Jiarun Liu, Hao Yang, Hong-Yu Zhou +8

    eess.IVcs.CVcs.LGarXiv:2402.03302v22024
  18. Joint Optimization of Masks and Deep Recurrent Neural Networks for Monaural Source Separation

    Po-Sen Huang, Minje Kim, Mark Hasegawa-Johnson +1

    cs.SDcs.AIcs.LGarXiv:1502.04149v42015
  19. Distributional Smoothing with Virtual Adversarial Training

    Takeru Miyato, Shin-ichi Maeda, Masanori Koyama +2

    stat.MLcs.LGarXiv:1507.00677v92015
  20. Multi-Agent Tensor Fusion for Contextual Trajectory Prediction

    Tianyang Zhao, Yifei Xu, Mathew Monfort +5

    cs.CVcs.LGarXiv:1904.04776v22019
  21. TSMixer: An All-MLP Architecture for Time Series Forecasting

    Si-An Chen, Chun-Liang Li, Nate Yoder +2

    cs.LGcs.AIarXiv:2303.06053v52023
  22. Depth Pro: Sharp Monocular Metric Depth in Less Than a Second

    Aleksei Bochkovskii, Amaël Delaunoy, Hugo Germain +4

    cs.CVcs.LGarXiv:2410.02073v22024
    Summaries:한국어
  23. Pose-Controllable Talking Face Generation by Implicitly Modularized Audio-Visual Representation

    Hang Zhou, Yasheng Sun, Wayne Wu +3

    cs.CVcs.LGcs.MMarXiv:2104.11116v12021
  24. A Compositional Object-Based Approach to Learning Physical Dynamics

    Michael B. Chang, Tomer Ullman, Antonio Torralba +1

    cs.AIcs.LGarXiv:1612.00341v22016
  25. Methodological and Conceptual Framework for 5D Multi-Table Analysis: A Unified Approach for Complex Data Reuse

    Edouard Lansiaux, Hugo Kazzi, Aurélien Loison +2

    cs.AIcs.LGarXiv:2608.26149v12026
  26. Domain Adaptive Neural Networks for Object Recognition

    Muhammad Ghifary, W. Bastiaan Kleijn, Mengjie Zhang

    cs.CVcs.AIcs.LGarXiv:1409.6041v12014
  27. Deep Learning for Joint Source-Channel Coding of Text

    Nariman Farsad, Milind Rao, Andrea Goldsmith

    cs.ITcs.AIcs.LGarXiv:1802.06832v12018
  28. Stacked Generative Adversarial Networks

    Xun Huang, Yixuan Li, Omid Poursaeed +2

    cs.CVcs.LGcs.NEarXiv:1612.04357v42016
  29. When the Canonical Completion Is Wrong: Formalizing and Measuring the Jump in Large Language Models

    Dai Shi, Xiaoyu Li, José Miguel Hernández-Lobato

    cs.CLcs.AIcs.LGarXiv:2608.26187v12026
  30. Separable Self-attention for Mobile Vision Transformers

    Sachin Mehta, Mohammad Rastegari

    cs.CVcs.AIcs.LGarXiv:2206.02680v12022
  31. An Information-Theoretic Analysis of Thompson Sampling

    Daniel Russo, Benjamin Van Roy

    cs.LGarXiv:1403.5341v22014
  32. Data Augmentation by Pairing Samples for Images Classification

    Hiroshi Inoue

    cs.LGcs.CVstat.MLarXiv:1801.02929v22018
  33. AudioGen: Textually Guided Audio Generation

    Felix Kreuk, Gabriel Synnaeve, Adam Polyak +6

    cs.SDcs.CLcs.LGarXiv:2209.15352v22022
  34. Beyond Parallel Blindness: Information Floors and Model Gaps in Block Drafting

    Xinwei Qiang, Xiang Fang, Chang Chen +2

    cs.LGcs.CLcs.ITarXiv:2608.27339v12026
  35. Sequential Short-Text Classification with Recurrent and Convolutional Neural Networks

    Ji Young Lee, Franck Dernoncourt

    cs.CLcs.AIcs.LGarXiv:1603.03827v12016
  36. FUDGE: Controlled Text Generation With Future Discriminators

    Kevin Yang, Dan Klein

    cs.CLcs.LGarXiv:2104.05218v22021
  37. Foundation Models for Time Series Analysis: A Tutorial and Survey

    Yuxuan Liang, Haomin Wen, Yuqi Nie +5

    cs.LGarXiv:2403.14735v32024
  38. Deep Bidirectional and Unidirectional LSTM Recurrent Neural Network for Network-wide Traffic Speed Prediction

    Zhiyong Cui, Ruimin Ke, Ziyuan Pu +1

    cs.LGarXiv:1801.02143v22018
  39. Fully Character-Level Neural Machine Translation without Explicit Segmentation

    Jason Lee, Kyunghyun Cho, Thomas Hofmann

    cs.CLcs.LGarXiv:1610.03017v32016
  40. PRNet: Self-Supervised Learning for Partial-to-Partial Registration

    Yue Wang, Justin M. Solomon

    cs.LGstat.MLarXiv:1910.12240v22019
  41. PMLB: A Large Benchmark Suite for Machine Learning Evaluation and Comparison

    Randal S. Olson, William La Cava, Patryk Orzechowski +2

    cs.LGcs.AIarXiv:1703.00512v12017
  42. GCAN: Graph-aware Co-Attention Networks for Explainable Fake News Detection on Social Media

    Yi-Ju Lu, Cheng-Te Li

    cs.CLcs.LGstat.MLarXiv:2004.11648v12020
  43. Normalization Techniques in Training DNNs: Methodology, Analysis and Application

    Lei Huang, Jie Qin, Yi Zhou +3

    cs.LGcs.CVstat.MLarXiv:2009.12836v12020
  44. Amnesiac Machine Learning

    Laura Graves, Vineel Nagisetty, Vijay Ganesh

    cs.LGcs.AIcs.CRarXiv:2010.10981v12020
  45. A Dynamic Likelihood Approach to Filtering for Advection-Diffusion Dynamics

    Johannes Krotz, Juan M. Restrepo, Jorge Ramirez

    math.DScs.LGmath.STarXiv:2406.06837v22024
  46. LoRA+: Efficient Low Rank Adaptation of Large Models

    Soufiane Hayou, Nikhil Ghosh, Bin Yu

    cs.LGcs.AIcs.CLarXiv:2402.12354v22024
  47. Multi-scale Dynamic Graph Convolutional Network for Hyperspectral Image Classification

    Sheng Wan, Chen Gong, Ping Zhong +3

    eess.IVcs.LGstat.MLarXiv:1905.06133v12019
  48. A Deep Learning Approach to Structured Signal Recovery

    Ali Mousavi, Ankit B. Patel, Richard G. Baraniuk

    cs.LGstat.MLarXiv:1508.04065v12015
  49. Further Optimal Regret Bounds for Thompson Sampling

    Shipra Agrawal, Navin Goyal

    cs.LGcs.DSstat.MLarXiv:1209.3353v12012
  50. A Simple Method for Commonsense Reasoning

    Trieu H. Trinh, Quoc V. Le

    cs.AIcs.CLcs.LGarXiv:1806.02847v22018
  51. Noise or Signal: The Role of Image Backgrounds in Object Recognition

    Kai Xiao, Logan Engstrom, Andrew Ilyas +1

    cs.CVcs.LGarXiv:2006.09994v12020
  52. Improved Techniques for Training Consistency Models

    Yang Song, Prafulla Dhariwal

    cs.LGarXiv:2310.14189v12023
  53. The INTERSPEECH 2020 Deep Noise Suppression Challenge: Datasets, Subjective Testing Framework, and Challenge Results

    Chandan K. A. Reddy, Vishak Gopal, Ross Cutler +10

    eess.AScs.LGcs.SDarXiv:2005.13981v32020
  54. The Instruction Hierarchy: Training LLMs to Prioritize Privileged Instructions

    Eric Wallace, Kai Xiao, Reimar Leike +3

    cs.CRcs.CLcs.LGarXiv:2404.13208v12024
  55. A Comprehensive Review of Deep Learning Applications in Hydrology and Water Resources

    Muhammed Sit, Bekir Z. Demiray, Zhongrun Xiang +3

    physics.geo-phcs.LGstat.MLarXiv:2007.12269v12020
  56. Robust Compressed Sensing MRI with Deep Generative Priors

    Ajil Jalal, Marius Arvinte, Giannis Daras +3

    cs.LGcs.CVcs.ITarXiv:2108.01368v22021
  57. An Empirical Study of Training End-to-End Vision-and-Language Transformers

    Zi-Yi Dou, Yichong Xu, Zhe Gan +9

    cs.CVcs.CLcs.LGarXiv:2111.02387v32021
  58. Proxy Anchor Loss for Deep Metric Learning

    Sungyeon Kim, Dongwon Kim, Minsu Cho +1

    cs.CVcs.LGarXiv:2003.13911v12020
  59. Crystal Diffusion Variational Autoencoder for Periodic Material Generation

    Tian Xie, Xiang Fu, Octavian-Eugen Ganea +2

    cs.LGcond-mat.mtrl-sciphysics.comp-pharXiv:2110.06197v32021
  60. SmaAt-UNet: Precipitation Nowcasting using a Small Attention-UNet Architecture

    Kevin Trebing, Tomasz Stanczyk, Siamak Mehrkanoon

    cs.LGeess.IVarXiv:2007.04417v22020