Machine Learning

Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

3,541 to 3,600 of 20,192

  1. Max-value Entropy Search for Multi-Objective Bayesian Optimization with Constraints

    Syrine Belakaria, Aryan Deshwal, Janardhan Rao Doppa

    cs.LGcs.AIstat.MLarXiv:2009.01721v22020
  2. SimpleMemVLA: A Simple but Effective Native-Video Memory for Vision-Language-Action Models

    Cheng Yin, Wang Xu, Junpeng Yang +8

    cs.CVcs.LGcs.ROarXiv:2609.05533v12026
  3. Federated Unsupervised Representation Learning

    Fengda Zhang, Kun Kuang, Zhaoyang You +6

    cs.LGcs.AIarXiv:2010.08982v12020
  4. Exponential Moving Average Normalization for Self-supervised and Semi-supervised Learning

    Zhaowei Cai, Avinash Ravichandran, Subhransu Maji +3

    cs.LGcs.AIcs.CVarXiv:2101.08482v22021
  5. Actor-Critic Reinforcement Learning for Control with Stability Guarantee

    Minghao Han, Lixian Zhang, Jun Wang +1

    cs.ROcs.LGeess.SYarXiv:2004.14288v32020
  6. Scaling the Scattering Transform: Deep Hybrid Networks

    Edouard Oyallon, Eugene Belilovsky, Sergey Zagoruyko

    cs.CVcs.LGarXiv:1703.08961v22017
  7. Finite Sample Analysis of Stochastic System Identification

    Anastasios Tsiamis, George J. Pappas

    cs.LGeess.SYmath.OCarXiv:1903.09122v12019
  8. Deepcode: Feedback Codes via Deep Learning

    Hyeji Kim, Yihan Jiang, Sreeram Kannan +2

    cs.LGcs.ITstat.MLarXiv:1807.00801v12018
  9. RES: Regularized Stochastic BFGS Algorithm

    Aryan Mokhtari, Alejandro Ribeiro

    cs.LGmath.OCstat.MLarXiv:1401.7625v12014
  10. Influence of Extruded Filament Shape on Buildability in 3D Concrete Printing: A Geometry-Informed Deep Learning-FEM Approach

    Giacomo Rizzieri, Saif-Ur-Rehman, Jörg F. Unger +1

    cs.CEcs.AIcs.LGarXiv:2609.04028v12026
  11. Object Detection for Graphical User Interface: Old Fashioned or Deep Learning or a Combination?

    Jieshan Chen, Mulong Xie, Zhenchang Xing +4

    cs.CVcs.HCcs.LGarXiv:2008.05132v22020
  12. Practical Detection of Trojan Neural Networks: Data-Limited and Data-Free Cases

    Ren Wang, Gaoyuan Zhang, Sijia Liu +3

    cs.LGcs.CRstat.MLarXiv:2007.15802v12020
  13. Personalizing Reinforcement Learning from Human Feedback with Variational Preference Learning

    Sriyash Poddar, Yanming Wan, Hamish Ivison +2

    cs.LGcs.AIcs.CLarXiv:2408.10075v12024
  14. Hypothesis Search: Inductive Reasoning with Language Models

    Ruocheng Wang, Eric Zelikman, Gabriel Poesia +3

    cs.LGcs.AIcs.CLarXiv:2309.05660v22023
  15. Pretraining task diversity and the emergence of non-Bayesian in-context learning for regression

    Allan Raventós, Mansheej Paul, Feng Chen +1

    cs.LGcs.AIcs.CLarXiv:2306.15063v22023
  16. Resiliency of Deep Neural Networks under Quantization

    Wonyong Sung, Sungho Shin, Kyuyeon Hwang

    cs.LGcs.NEarXiv:1511.06488v32015
  17. Subspace Inference Enables Efficient Active Reward Learning from Preferences

    Yutai Zhou, Erdem Bıyık

    cs.LGcs.AIcs.ROarXiv:2609.04066v12026
  18. Deep Attention Recurrent Q-Network

    Ivan Sorokin, Alexey Seleznev, Mikhail Pavlov +2

    cs.LGarXiv:1512.01693v12015
  19. Conservative Safety Critics for Exploration

    Homanga Bharadhwaj, Aviral Kumar, Nicholas Rhinehart +3

    cs.LGcs.AIcs.ROarXiv:2010.14497v22020
  20. KoDF: A Large-scale Korean DeepFake Detection Dataset

    Patrick Kwon, Jaeseong You, Gyuhyeon Nam +2

    cs.CVcs.LGarXiv:2103.10094v22021
  21. Trained Quantization Thresholds for Accurate and Efficient Fixed-Point Inference of Deep Neural Networks

    Sambhav R. Jain, Albert Gural, Michael Wu +1

    cs.CVcs.AIcs.LGarXiv:1903.08066v32019
  22. Score-based Continuous-time Discrete Diffusion Models

    Haoran Sun, Lijun Yu, Bo Dai +2

    cs.LGarXiv:2211.16750v22022
  23. Smart Contract Vulnerability Detection: From Pure Neural Network to Interpretable Graph Feature and Expert Pattern Fusion

    Zhenguang Liu, Peng Qian, Xiang Wang +3

    cs.LGcs.PLarXiv:2106.09282v12021
  24. A Unified Theory of SGD: Variance Reduction, Sampling, Quantization and Coordinate Descent

    Eduard Gorbunov, Filip Hanzely, Peter Richtárik

    math.OCcs.LGmath.NAarXiv:1905.11261v12019
  25. GLiNER: Generalist Model for Named Entity Recognition using Bidirectional Transformer

    Urchade Zaratiana, Nadi Tomeh, Pierre Holat +1

    cs.CLcs.AIcs.LGarXiv:2311.08526v12023
  26. ML-Doctor: Holistic Risk Assessment of Inference Attacks Against Machine Learning Models

    Yugeng Liu, Rui Wen, Xinlei He +6

    cs.CRcs.AIcs.LGarXiv:2102.02551v22021
  27. Counterfactual Risk Minimization: Learning from Logged Bandit Feedback

    Adith Swaminathan, Thorsten Joachims

    cs.LGstat.MLarXiv:1502.02362v22015
  28. Manifold Embedded Knowledge Transfer for Brain-Computer Interfaces

    Wen Zhang, Dongrui Wu

    cs.HCcs.LGarXiv:1910.05878v22019
  29. Contrastive Energy Prediction for Exact Energy-Guided Diffusion Sampling in Offline Reinforcement Learning

    Cheng Lu, Huayu Chen, Jianfei Chen +3

    cs.LGarXiv:2304.12824v22023
  30. Increasing the Action Gap: New Operators for Reinforcement Learning

    Marc G. Bellemare, Georg Ostrovski, Arthur Guez +2

    cs.AIcs.LGarXiv:1512.04860v12015
  31. RATL: Learning from Retrieved Residuals for Robust Multivariate Time-Series Forecasting

    Yuchen He, Yueyang Cang, Zhiyuan Ning +2

    cs.LGcs.AIarXiv:2609.03937v12026
  32. Transformer Neural Processes: Uncertainty-Aware Meta Learning Via Sequence Modeling

    Tung Nguyen, Aditya Grover

    cs.LGcs.AIarXiv:2207.04179v22022
  33. FedLab: A Flexible Federated Learning Framework

    Dun Zeng, Siqi Liang, Xiangjing Hu +2

    cs.LGcs.AIarXiv:2107.11621v42021
  34. Stochastic Optimization with Heavy-Tailed Noise via Accelerated Gradient Clipping

    Eduard Gorbunov, Marina Danilova, Alexander Gasnikov

    math.OCcs.LGarXiv:2005.10785v22020
  35. Continuously Differentiable Exponential Linear Units

    Jonathan T. Barron

    cs.LGarXiv:1704.07483v12017
  36. RARF: Region-Aware Rectified Flows for 3D Brain MRI Inpainting

    Tomas Guija-Valiente, Blanca Rodriguez-Gonzalez, Norberto Malpica +1

    cs.CVcs.AIcs.LGarXiv:2609.03956v12026
  37. AdvPrompter: Fast Adaptive Adversarial Prompting for LLMs

    Anselm Paulus, Arman Zharmagambetov, Chuan Guo +2

    cs.CRcs.AIcs.CLarXiv:2404.16873v22024
  38. Headroom-Drift Replay: A Primitive for Principled Replay Control in GRPO

    Hyun Bin Park, Du-Seong Chang

    cs.LGcs.AIcs.CLarXiv:2609.03941v12026
  39. Maximum-Entropy Fine-Grained Classification

    Abhimanyu Dubey, Otkrist Gupta, Ramesh Raskar +1

    cs.CVcs.LGarXiv:1809.05934v22018
  40. The probability flow ODE is provably fast

    Sitan Chen, Sinho Chewi, Holden Lee +3

    cs.LGmath.STstat.MLarXiv:2305.11798v12023
  41. Fast $ε$-free Inference of Simulation Models with Bayesian Conditional Density Estimation

    George Papamakarios, Iain Murray

    stat.MLcs.LGstat.COarXiv:1605.06376v42016
  42. Heterogeneous Risk Minimization

    Jiashuo Liu, Zheyuan Hu, Peng Cui +2

    cs.LGarXiv:2105.03818v32021
  43. Reducing Information Bottleneck for Weakly Supervised Semantic Segmentation

    Jungbeom Lee, Jooyoung Choi, Jisoo Mok +1

    cs.CVcs.LGarXiv:2110.06530v12021
  44. Towards Debiasing NLU Models from Unknown Biases

    Prasetya Ajie Utama, Nafise Sadat Moosavi, Iryna Gurevych

    cs.CLcs.AIcs.LGarXiv:2009.12303v42020
  45. Speech Emotion Recognition using Self-Supervised Features

    Edmilson Morais, Ron Hoory, Weizhong Zhu +3

    cs.SDcs.AIcs.LGarXiv:2202.03896v12022
  46. Pre-training via Denoising for Molecular Property Prediction

    Sheheryar Zaidi, Michael Schaarschmidt, James Martens +6

    cs.LGq-bio.BMstat.MLarXiv:2206.00133v22022
  47. DecAF: Joint Decoding of Answers and Logical Forms for Question Answering over Knowledge Bases

    Donghan Yu, Sheng Zhang, Patrick Ng +7

    cs.CLcs.AIcs.LGarXiv:2210.00063v22022
  48. Differentiable Interval Bottlenecks for Interpretable Anomaly Detection in Numerical Data

    Lamine Diop, Marc Plantevit

    cs.LGcs.AIarXiv:2609.03878v12026
  49. Witnesses Explain Anomalies

    Lamine Diop

    cs.LGcs.AIarXiv:2609.03826v12026
  50. GIFT-Eval: A Benchmark For General Time Series Forecasting Model Evaluation

    Taha Aksu, Gerald Woo, Juncheng Liu +5

    cs.LGstat.MLarXiv:2410.10393v22024
  51. EPIC-KITCHENS VISOR Benchmark: VIdeo Segmentations and Object Relations

    Ahmad Darkhalil, Dandan Shan, Bin Zhu +6

    cs.CVcs.AIcs.LGarXiv:2209.13064v12022
  52. PREMA: A Predictive Multi-task Scheduling Algorithm For Preemptible Neural Processing Units

    Yujeong Choi, Minsoo Rhu

    cs.DCcs.LGcs.NEarXiv:1909.04548v12019
  53. Accelerated Nuclear Magnetic Resonance Spectroscopy with Deep Learning

    Xiaobo Qu, Yihui Huang, Hengfa Lu +5

    physics.med-phcs.AIcs.LGarXiv:1904.05168v22019
  54. Personalized and Private Peer-to-Peer Machine Learning

    Aurélien Bellet, Rachid Guerraoui, Mahsa Taziki +1

    cs.LGcs.CRcs.DCarXiv:1705.08435v22017
  55. Out-of-Distribution Generalisation with Sequence Models in Offline Multi-Agent Reinforcement Learning

    Oussama Hidaoui, Omer Ebead, Ulrich Armel Mbou Sob +14

    cs.LGcs.AIarXiv:2609.03667v12026
  56. Deep Clustering and Conventional Networks for Music Separation: Stronger Together

    Yi Luo, Zhuo Chen, John R. Hershey +2

    stat.MLcs.LGcs.SDarXiv:1611.06265v22016
  57. RAG vs Fine-tuning: Pipelines, Tradeoffs, and a Case Study on Agriculture

    Angels Balaguer, Vinamra Benara, Renato Luiz de Freitas Cunha +13

    cs.CLcs.LGarXiv:2401.08406v32024
  58. Almost Free State Prediction Separation

    John Langford, Nathan Godey, Giovanni Monea +5

    cs.LGcs.AIarXiv:2609.03807v22026
  59. Broadcasted Residual Learning for Efficient Keyword Spotting

    Byeonggeun Kim, Simyung Chang, Jinkyu Lee +1

    cs.SDcs.LGeess.ASarXiv:2106.04140v42021
  60. Olmo Hybrid: From Theory to Practice and Back

    William Merrill, Yanhong Li, Tyler Romero +19

    cs.LGcs.CLarXiv:2604.03444v42026