Machine Learning
Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
3,541 to 3,600 of 20,192
Max-value Entropy Search for Multi-Objective Bayesian Optimization with Constraints
Syrine Belakaria, Aryan Deshwal, Janardhan Rao Doppa
cs.LGcs.AIstat.MLarXiv:2009.01721v22020SimpleMemVLA: A Simple but Effective Native-Video Memory for Vision-Language-Action Models
Cheng Yin, Wang Xu, Junpeng Yang +8
cs.CVcs.LGcs.ROarXiv:2609.05533v12026Federated Unsupervised Representation Learning
Fengda Zhang, Kun Kuang, Zhaoyang You +6
cs.LGcs.AIarXiv:2010.08982v12020Exponential Moving Average Normalization for Self-supervised and Semi-supervised Learning
Zhaowei Cai, Avinash Ravichandran, Subhransu Maji +3
cs.LGcs.AIcs.CVarXiv:2101.08482v22021Actor-Critic Reinforcement Learning for Control with Stability Guarantee
Minghao Han, Lixian Zhang, Jun Wang +1
cs.ROcs.LGeess.SYarXiv:2004.14288v32020Scaling the Scattering Transform: Deep Hybrid Networks
Edouard Oyallon, Eugene Belilovsky, Sergey Zagoruyko
cs.CVcs.LGarXiv:1703.08961v22017Finite Sample Analysis of Stochastic System Identification
Anastasios Tsiamis, George J. Pappas
cs.LGeess.SYmath.OCarXiv:1903.09122v12019Deepcode: Feedback Codes via Deep Learning
Hyeji Kim, Yihan Jiang, Sreeram Kannan +2
cs.LGcs.ITstat.MLarXiv:1807.00801v12018RES: Regularized Stochastic BFGS Algorithm
Aryan Mokhtari, Alejandro Ribeiro
cs.LGmath.OCstat.MLarXiv:1401.7625v12014Influence of Extruded Filament Shape on Buildability in 3D Concrete Printing: A Geometry-Informed Deep Learning-FEM Approach
Giacomo Rizzieri, Saif-Ur-Rehman, Jörg F. Unger +1
cs.CEcs.AIcs.LGarXiv:2609.04028v12026Object Detection for Graphical User Interface: Old Fashioned or Deep Learning or a Combination?
Jieshan Chen, Mulong Xie, Zhenchang Xing +4
cs.CVcs.HCcs.LGarXiv:2008.05132v22020Practical Detection of Trojan Neural Networks: Data-Limited and Data-Free Cases
Ren Wang, Gaoyuan Zhang, Sijia Liu +3
cs.LGcs.CRstat.MLarXiv:2007.15802v12020Personalizing Reinforcement Learning from Human Feedback with Variational Preference Learning
Sriyash Poddar, Yanming Wan, Hamish Ivison +2
cs.LGcs.AIcs.CLarXiv:2408.10075v12024Hypothesis Search: Inductive Reasoning with Language Models
Ruocheng Wang, Eric Zelikman, Gabriel Poesia +3
cs.LGcs.AIcs.CLarXiv:2309.05660v22023Pretraining task diversity and the emergence of non-Bayesian in-context learning for regression
Allan Raventós, Mansheej Paul, Feng Chen +1
cs.LGcs.AIcs.CLarXiv:2306.15063v22023Resiliency of Deep Neural Networks under Quantization
Wonyong Sung, Sungho Shin, Kyuyeon Hwang
cs.LGcs.NEarXiv:1511.06488v32015Subspace Inference Enables Efficient Active Reward Learning from Preferences
Yutai Zhou, Erdem Bıyık
cs.LGcs.AIcs.ROarXiv:2609.04066v12026Deep Attention Recurrent Q-Network
Ivan Sorokin, Alexey Seleznev, Mikhail Pavlov +2
cs.LGarXiv:1512.01693v12015Conservative Safety Critics for Exploration
Homanga Bharadhwaj, Aviral Kumar, Nicholas Rhinehart +3
cs.LGcs.AIcs.ROarXiv:2010.14497v22020KoDF: A Large-scale Korean DeepFake Detection Dataset
Patrick Kwon, Jaeseong You, Gyuhyeon Nam +2
cs.CVcs.LGarXiv:2103.10094v22021Trained Quantization Thresholds for Accurate and Efficient Fixed-Point Inference of Deep Neural Networks
Sambhav R. Jain, Albert Gural, Michael Wu +1
cs.CVcs.AIcs.LGarXiv:1903.08066v32019Score-based Continuous-time Discrete Diffusion Models
Haoran Sun, Lijun Yu, Bo Dai +2
cs.LGarXiv:2211.16750v22022Smart Contract Vulnerability Detection: From Pure Neural Network to Interpretable Graph Feature and Expert Pattern Fusion
Zhenguang Liu, Peng Qian, Xiang Wang +3
cs.LGcs.PLarXiv:2106.09282v12021A Unified Theory of SGD: Variance Reduction, Sampling, Quantization and Coordinate Descent
Eduard Gorbunov, Filip Hanzely, Peter Richtárik
math.OCcs.LGmath.NAarXiv:1905.11261v12019GLiNER: Generalist Model for Named Entity Recognition using Bidirectional Transformer
Urchade Zaratiana, Nadi Tomeh, Pierre Holat +1
cs.CLcs.AIcs.LGarXiv:2311.08526v12023ML-Doctor: Holistic Risk Assessment of Inference Attacks Against Machine Learning Models
Yugeng Liu, Rui Wen, Xinlei He +6
cs.CRcs.AIcs.LGarXiv:2102.02551v22021Counterfactual Risk Minimization: Learning from Logged Bandit Feedback
Adith Swaminathan, Thorsten Joachims
cs.LGstat.MLarXiv:1502.02362v22015Manifold Embedded Knowledge Transfer for Brain-Computer Interfaces
Wen Zhang, Dongrui Wu
cs.HCcs.LGarXiv:1910.05878v22019Contrastive Energy Prediction for Exact Energy-Guided Diffusion Sampling in Offline Reinforcement Learning
Cheng Lu, Huayu Chen, Jianfei Chen +3
cs.LGarXiv:2304.12824v22023Increasing the Action Gap: New Operators for Reinforcement Learning
Marc G. Bellemare, Georg Ostrovski, Arthur Guez +2
cs.AIcs.LGarXiv:1512.04860v12015RATL: Learning from Retrieved Residuals for Robust Multivariate Time-Series Forecasting
Yuchen He, Yueyang Cang, Zhiyuan Ning +2
cs.LGcs.AIarXiv:2609.03937v12026Transformer Neural Processes: Uncertainty-Aware Meta Learning Via Sequence Modeling
Tung Nguyen, Aditya Grover
cs.LGcs.AIarXiv:2207.04179v22022FedLab: A Flexible Federated Learning Framework
Dun Zeng, Siqi Liang, Xiangjing Hu +2
cs.LGcs.AIarXiv:2107.11621v42021Stochastic Optimization with Heavy-Tailed Noise via Accelerated Gradient Clipping
Eduard Gorbunov, Marina Danilova, Alexander Gasnikov
math.OCcs.LGarXiv:2005.10785v22020Continuously Differentiable Exponential Linear Units
Jonathan T. Barron
cs.LGarXiv:1704.07483v12017RARF: Region-Aware Rectified Flows for 3D Brain MRI Inpainting
Tomas Guija-Valiente, Blanca Rodriguez-Gonzalez, Norberto Malpica +1
cs.CVcs.AIcs.LGarXiv:2609.03956v12026AdvPrompter: Fast Adaptive Adversarial Prompting for LLMs
Anselm Paulus, Arman Zharmagambetov, Chuan Guo +2
cs.CRcs.AIcs.CLarXiv:2404.16873v22024Headroom-Drift Replay: A Primitive for Principled Replay Control in GRPO
Hyun Bin Park, Du-Seong Chang
cs.LGcs.AIcs.CLarXiv:2609.03941v12026Maximum-Entropy Fine-Grained Classification
Abhimanyu Dubey, Otkrist Gupta, Ramesh Raskar +1
cs.CVcs.LGarXiv:1809.05934v22018The probability flow ODE is provably fast
Sitan Chen, Sinho Chewi, Holden Lee +3
cs.LGmath.STstat.MLarXiv:2305.11798v12023Fast $ε$-free Inference of Simulation Models with Bayesian Conditional Density Estimation
George Papamakarios, Iain Murray
stat.MLcs.LGstat.COarXiv:1605.06376v42016Heterogeneous Risk Minimization
Jiashuo Liu, Zheyuan Hu, Peng Cui +2
cs.LGarXiv:2105.03818v32021Reducing Information Bottleneck for Weakly Supervised Semantic Segmentation
Jungbeom Lee, Jooyoung Choi, Jisoo Mok +1
cs.CVcs.LGarXiv:2110.06530v12021Towards Debiasing NLU Models from Unknown Biases
Prasetya Ajie Utama, Nafise Sadat Moosavi, Iryna Gurevych
cs.CLcs.AIcs.LGarXiv:2009.12303v42020Speech Emotion Recognition using Self-Supervised Features
Edmilson Morais, Ron Hoory, Weizhong Zhu +3
cs.SDcs.AIcs.LGarXiv:2202.03896v12022Pre-training via Denoising for Molecular Property Prediction
Sheheryar Zaidi, Michael Schaarschmidt, James Martens +6
cs.LGq-bio.BMstat.MLarXiv:2206.00133v22022DecAF: Joint Decoding of Answers and Logical Forms for Question Answering over Knowledge Bases
Donghan Yu, Sheng Zhang, Patrick Ng +7
cs.CLcs.AIcs.LGarXiv:2210.00063v22022Differentiable Interval Bottlenecks for Interpretable Anomaly Detection in Numerical Data
Lamine Diop, Marc Plantevit
cs.LGcs.AIarXiv:2609.03878v12026Witnesses Explain Anomalies
Lamine Diop
cs.LGcs.AIarXiv:2609.03826v12026GIFT-Eval: A Benchmark For General Time Series Forecasting Model Evaluation
Taha Aksu, Gerald Woo, Juncheng Liu +5
cs.LGstat.MLarXiv:2410.10393v22024EPIC-KITCHENS VISOR Benchmark: VIdeo Segmentations and Object Relations
Ahmad Darkhalil, Dandan Shan, Bin Zhu +6
cs.CVcs.AIcs.LGarXiv:2209.13064v12022PREMA: A Predictive Multi-task Scheduling Algorithm For Preemptible Neural Processing Units
Yujeong Choi, Minsoo Rhu
cs.DCcs.LGcs.NEarXiv:1909.04548v12019Accelerated Nuclear Magnetic Resonance Spectroscopy with Deep Learning
Xiaobo Qu, Yihui Huang, Hengfa Lu +5
physics.med-phcs.AIcs.LGarXiv:1904.05168v22019Personalized and Private Peer-to-Peer Machine Learning
Aurélien Bellet, Rachid Guerraoui, Mahsa Taziki +1
cs.LGcs.CRcs.DCarXiv:1705.08435v22017Out-of-Distribution Generalisation with Sequence Models in Offline Multi-Agent Reinforcement Learning
Oussama Hidaoui, Omer Ebead, Ulrich Armel Mbou Sob +14
cs.LGcs.AIarXiv:2609.03667v12026Deep Clustering and Conventional Networks for Music Separation: Stronger Together
Yi Luo, Zhuo Chen, John R. Hershey +2
stat.MLcs.LGcs.SDarXiv:1611.06265v22016RAG vs Fine-tuning: Pipelines, Tradeoffs, and a Case Study on Agriculture
Angels Balaguer, Vinamra Benara, Renato Luiz de Freitas Cunha +13
cs.CLcs.LGarXiv:2401.08406v32024Almost Free State Prediction Separation
John Langford, Nathan Godey, Giovanni Monea +5
cs.LGcs.AIarXiv:2609.03807v22026Broadcasted Residual Learning for Efficient Keyword Spotting
Byeonggeun Kim, Simyung Chang, Jinkyu Lee +1
cs.SDcs.LGeess.ASarXiv:2106.04140v42021Olmo Hybrid: From Theory to Practice and Back
William Merrill, Yanhong Li, Tyler Romero +19
cs.LGcs.CLarXiv:2604.03444v42026