Machine Learning
Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
3,781 to 3,840 of 20,454
A survey on intrinsic motivation in reinforcement learning
Arthur Aubret, Laetitia Matignon, Salima Hassas
cs.LGcs.AIarXiv:1908.06976v22019AFTer-UNet: Axial Fusion Transformer UNet for Medical Image Segmentation
Xiangyi Yan, Hao Tang, Shanlin Sun +3
eess.IVcs.CVcs.LGarXiv:2110.10403v12021Teach Me to Explain: A Review of Datasets for Explainable Natural Language Processing
Sarah Wiegreffe, Ana Marasović
cs.CLcs.AIcs.LGarXiv:2102.12060v42021Trojaning Language Models for Fun and Profit
Xinyang Zhang, Zheng Zhang, Shouling Ji +1
cs.CRcs.CLcs.LGarXiv:2008.00312v22020BLEU might be Guilty but References are not Innocent
Markus Freitag, David Grangier, Isaac Caswell
cs.CLcs.AIcs.LGarXiv:2004.06063v22020Cluster-to-Conquer: A Framework for End-to-End Multi-Instance Learning for Whole Slide Image Classification
Yash Sharma, Aman Shrivastava, Lubaina Ehsan +3
eess.IVcs.CVcs.LGarXiv:2103.10626v22021Exploiting multi-CNN features in CNN-RNN based Dimensional Emotion Recognition on the OMG in-the-wild Dataset
Dimitrios Kollias, Stefanos Zafeiriou
cs.LGcs.CVstat.MLarXiv:1910.01417v22019Continual Pre-Training of Large Language Models: How to (re)warm your model?
Kshitij Gupta, Benjamin Thérien, Adam Ibrahim +5
cs.CLcs.LGarXiv:2308.04014v22023Multi-level Convolutional Autoencoder Networks for Parametric Prediction of Spatio-temporal Dynamics
Jiayang Xu, Karthik Duraisamy
physics.comp-phcs.LGphysics.flu-dynarXiv:1912.11114v22019Decentralizing Feature Extraction with Quantum Convolutional Neural Network for Automatic Speech Recognition
Chao-Han Huck Yang, Jun Qi, Samuel Yen-Chi Chen +4
cs.SDcs.LGcs.NEarXiv:2010.13309v22020Predicting Clinical Events by Combining Static and Dynamic Information Using Recurrent Neural Networks
Cristóbal Esteban, Oliver Staeck, Yinchong Yang +1
cs.LGcs.AIcs.NEarXiv:1602.02685v22016Taming the Wild: A Unified Analysis of Hogwild!-Style Algorithms
Christopher De Sa, Ce Zhang, Kunle Olukotun +1
cs.LGmath.OCstat.MLarXiv:1506.06438v22015Summaries:한국어Twin Contrastive Learning for Online Clustering
Yunfan Li, Mouxing Yang, Dezhong Peng +3
cs.LGarXiv:2210.11680v12022Survey of Machine Learning Accelerators
Albert Reuther, Peter Michaleas, Michael Jones +3
cs.DCcs.LGarXiv:2009.00993v12020Efficient Algorithms for Outlier-Robust Regression
Adam Klivans, Pravesh K. Kothari, Raghu Meka
cs.LGcs.AIcs.DSarXiv:1803.03241v32018How to Build the Virtual Cell with Artificial Intelligence: Priorities and Opportunities
Charlotte Bunne, Yusuf Roohani, Yanay Rosen +39
q-bio.QMcs.AIcs.LGarXiv:2409.11654v22024Guidelines and Evaluation of Clinical Explainable AI in Medical Image Analysis
Weina Jin, Xiaoxiao Li, Mostafa Fatehi +1
cs.LGcs.AIcs.CVarXiv:2202.10553v32022Stochastic Latent Residual Video Prediction
Jean-Yves Franceschi, Edouard Delasalles, Mickaël Chen +2
cs.CVcs.LGstat.MLarXiv:2002.09219v42020Sequential Beats Joint: On the Interplay between On-Policy Distillation and RLVR
Boyan Li, Bingsen Chen, Chenghao Yang +3
cs.CLcs.AIcs.LGarXiv:2609.04108v22026Language Agents with Reinforcement Learning for Strategic Play in the Werewolf Game
Zelai Xu, Chao Yu, Fei Fang +2
cs.AIcs.LGcs.MAarXiv:2310.18940v42023COVID_MTNet: COVID-19 Detection with Multi-Task Deep Learning Approaches
Md Zahangir Alom, M M Shaifur Rahman, Mst Shamima Nasrin +2
eess.IVcs.CVcs.LGarXiv:2004.03747v32020Continual Learning in Low-rank Orthogonal Subspaces
Arslan Chaudhry, Naeemullah Khan, Puneet K. Dokania +1
cs.LGarXiv:2010.11635v22020TeraPipe: Token-Level Pipeline Parallelism for Training Large-Scale Language Models
Zhuohan Li, Siyuan Zhuang, Shiyuan Guo +4
cs.LGcs.CLcs.DCarXiv:2102.07988v22021Network Enhancement: a general method to denoise weighted biological networks
Bo Wang, Armin Pourshafeie, Marinka Zitnik +4
q-bio.MNcs.LGcs.SIarXiv:1805.03327v22018Structured Adversarial Attack: Towards General Implementation and Better Interpretability
Kaidi Xu, Sijia Liu, Pu Zhao +6
cs.LGcs.AIstat.MLarXiv:1808.01664v32018A Distributed Synchronous SGD Algorithm with Global Top-$k$ Sparsification for Low Bandwidth Networks
Shaohuai Shi, Qiang Wang, Kaiyong Zhao +4
cs.DCcs.LGarXiv:1901.04359v22019Experiments of Federated Learning for COVID-19 Chest X-ray Images
Boyi Liu, Bingjie Yan, Yize Zhou +2
eess.IVcs.CVcs.LGarXiv:2007.05592v12020MAP Estimation, Linear Programming and Belief Propagation with Convex Free Energies
Yair Weiss, Chen Yanover, Talya Meltzer
cs.AIcs.LGstat.MLarXiv:1206.5286v12012Mixed Precision Training of Convolutional Neural Networks using Integer Operations
Dipankar Das, Naveen Mellempudi, Dheevatsa Mudigere +14
cs.NEcs.LGmath.NAarXiv:1802.00930v22018Neural Network Attributions: A Causal Perspective
Aditya Chattopadhyay, Piyushi Manupriya, Anirban Sarkar +1
cs.LGstat.MLarXiv:1902.02302v42019Bayesian Action Decoder for Deep Multi-Agent Reinforcement Learning
Jakob N. Foerster, Francis Song, Edward Hughes +5
cs.MAcs.AIcs.LGarXiv:1811.01458v32018Flexible Clustered Federated Learning for Client-Level Data Distribution Shift
Moming Duan, Duo Liu, Xinyuan Ji +4
cs.LGcs.DCarXiv:2108.09749v12021Semi-Supervised Graph Classification: A Hierarchical Graph Perspective
Jia Li, Yu Rong, Hong Cheng +3
cs.CVcs.LGarXiv:1904.05003v12019PU Learning for Matrix Completion
Cho-Jui Hsieh, Nagarajan Natarajan, Inderjit S. Dhillon
cs.LGmath.NAstat.MLarXiv:1411.6081v12014Deep Network Guided Proof Search
Sarah Loos, Geoffrey Irving, Christian Szegedy +1
cs.AIcs.LGcs.LOarXiv:1701.06972v12017A Low-Cost, Open Platform for End-to-End Autonomous Driving on a Miniature Ackermann Vehicle
Gustavo Claudio Karl Couto, Eric Aislan Antonelo, Gabriel George Zipperer
cs.LGcs.AIcs.ROarXiv:2609.04147v12026Learning to Understand Goal Specifications by Modelling Reward
Dzmitry Bahdanau, Felix Hill, Jan Leike +4
cs.AIcs.LGarXiv:1806.01946v42018How Do Classifiers Induce Agents To Invest Effort Strategically?
Jon Kleinberg, Manish Raghavan
cs.LGcs.CYcs.DSarXiv:1807.05307v52018Max-value Entropy Search for Multi-Objective Bayesian Optimization with Constraints
Syrine Belakaria, Aryan Deshwal, Janardhan Rao Doppa
cs.LGcs.AIstat.MLarXiv:2009.01721v22020SimpleMemVLA: A Simple but Effective Native-Video Memory for Vision-Language-Action Models
Cheng Yin, Wang Xu, Junpeng Yang +8
cs.CVcs.LGcs.ROarXiv:2609.05533v12026Federated Unsupervised Representation Learning
Fengda Zhang, Kun Kuang, Zhaoyang You +6
cs.LGcs.AIarXiv:2010.08982v12020Exponential Moving Average Normalization for Self-supervised and Semi-supervised Learning
Zhaowei Cai, Avinash Ravichandran, Subhransu Maji +3
cs.LGcs.AIcs.CVarXiv:2101.08482v22021Actor-Critic Reinforcement Learning for Control with Stability Guarantee
Minghao Han, Lixian Zhang, Jun Wang +1
cs.ROcs.LGeess.SYarXiv:2004.14288v32020Scaling the Scattering Transform: Deep Hybrid Networks
Edouard Oyallon, Eugene Belilovsky, Sergey Zagoruyko
cs.CVcs.LGarXiv:1703.08961v22017Finite Sample Analysis of Stochastic System Identification
Anastasios Tsiamis, George J. Pappas
cs.LGeess.SYmath.OCarXiv:1903.09122v12019Deepcode: Feedback Codes via Deep Learning
Hyeji Kim, Yihan Jiang, Sreeram Kannan +2
cs.LGcs.ITstat.MLarXiv:1807.00801v12018RES: Regularized Stochastic BFGS Algorithm
Aryan Mokhtari, Alejandro Ribeiro
cs.LGmath.OCstat.MLarXiv:1401.7625v12014Influence of Extruded Filament Shape on Buildability in 3D Concrete Printing: A Geometry-Informed Deep Learning-FEM Approach
Giacomo Rizzieri, Saif-Ur-Rehman, Jörg F. Unger +1
cs.CEcs.AIcs.LGarXiv:2609.04028v12026Object Detection for Graphical User Interface: Old Fashioned or Deep Learning or a Combination?
Jieshan Chen, Mulong Xie, Zhenchang Xing +4
cs.CVcs.HCcs.LGarXiv:2008.05132v22020Practical Detection of Trojan Neural Networks: Data-Limited and Data-Free Cases
Ren Wang, Gaoyuan Zhang, Sijia Liu +3
cs.LGcs.CRstat.MLarXiv:2007.15802v12020Personalizing Reinforcement Learning from Human Feedback with Variational Preference Learning
Sriyash Poddar, Yanming Wan, Hamish Ivison +2
cs.LGcs.AIcs.CLarXiv:2408.10075v12024Hypothesis Search: Inductive Reasoning with Language Models
Ruocheng Wang, Eric Zelikman, Gabriel Poesia +3
cs.LGcs.AIcs.CLarXiv:2309.05660v22023Pretraining task diversity and the emergence of non-Bayesian in-context learning for regression
Allan Raventós, Mansheej Paul, Feng Chen +1
cs.LGcs.AIcs.CLarXiv:2306.15063v22023Resiliency of Deep Neural Networks under Quantization
Wonyong Sung, Sungho Shin, Kyuyeon Hwang
cs.LGcs.NEarXiv:1511.06488v32015Subspace Inference Enables Efficient Active Reward Learning from Preferences
Yutai Zhou, Erdem Bıyık
cs.LGcs.AIcs.ROarXiv:2609.04066v12026Deep Attention Recurrent Q-Network
Ivan Sorokin, Alexey Seleznev, Mikhail Pavlov +2
cs.LGarXiv:1512.01693v12015Conservative Safety Critics for Exploration
Homanga Bharadhwaj, Aviral Kumar, Nicholas Rhinehart +3
cs.LGcs.AIcs.ROarXiv:2010.14497v22020KoDF: A Large-scale Korean DeepFake Detection Dataset
Patrick Kwon, Jaeseong You, Gyuhyeon Nam +2
cs.CVcs.LGarXiv:2103.10094v22021Trained Quantization Thresholds for Accurate and Efficient Fixed-Point Inference of Deep Neural Networks
Sambhav R. Jain, Albert Gural, Michael Wu +1
cs.CVcs.AIcs.LGarXiv:1903.08066v32019Score-based Continuous-time Discrete Diffusion Models
Haoran Sun, Lijun Yu, Bo Dai +2
cs.LGarXiv:2211.16750v22022