Machine Learning
Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
13,561 to 13,620 of 20,219
Meta-Learning Representations for Continual Learning
Khurram Javed, Martha White
cs.LGcs.AIstat.MLarXiv:1905.12588v22019Federated Learning for Malware Detection in IoT Devices
Valerian Rey, Pedro Miguel Sánchez Sánchez, Alberto Huertas Celdrán +2
cs.CRcs.LGarXiv:2104.09994v32021A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation
Runzhe Yang, Xingyuan Sun, Karthik Narasimhan
cs.LGcs.AIarXiv:1908.08342v22019Towards an Automatic Turing Test: Learning to Evaluate Dialogue Responses
Ryan Lowe, Michael Noseworthy, Iulian V. Serban +3
cs.CLcs.AIcs.LGarXiv:1708.07149v22017On Evaluating Adversarial Robustness of Large Vision-Language Models
Yunqing Zhao, Tianyu Pang, Chao Du +4
cs.CVcs.CLcs.CRarXiv:2305.16934v22023Large-scale Classification of Fine-Art Paintings: Learning The Right Metric on The Right Feature
Babak Saleh, Ahmed Elgammal
cs.CVcs.IRcs.LGarXiv:1505.00855v12015Free Lunch for Few-shot Learning: Distribution Calibration
Shuo Yang, Lu Liu, Min Xu
cs.LGcs.CVarXiv:2101.06395v32021District-Level Food Environment Indicators and Social Vulnerability in São Paulo
Pedro Lemes Sixel Lobo, Eric Tokuda, Kuruvilla Joseph Abraham +4
physics.soc-phcs.LGstat.AParXiv:2608.26299v12026Least Ambiguous Set-Valued Classifiers with Bounded Error Levels
Mauricio Sadinle, Jing Lei, Larry Wasserman
stat.MEcs.LGstat.MLarXiv:1609.00451v22016Charting the Right Manifold: Manifold Mixup for Few-shot Learning
Puneet Mangla, Mayank Singh, Abhishek Sinha +3
cs.LGcs.CVstat.MLarXiv:1907.12087v42019Generating 3D Adversarial Point Clouds
Chong Xiang, Charles R. Qi, Bo Li
cs.CRcs.CVcs.LGarXiv:1809.07016v42018One-Shot Imitation from Observing Humans via Domain-Adaptive Meta-Learning
Tianhe Yu, Chelsea Finn, Annie Xie +4
cs.LGcs.AIcs.CVarXiv:1802.01557v12018Iterative Preference Learning from Human Feedback: Bridging Theory and Practice for RLHF under KL-Constraint
Wei Xiong, Hanze Dong, Chenlu Ye +5
cs.LGcs.AIstat.MLarXiv:2312.11456v42023BENDR: using transformers and a contrastive self-supervised learning task to learn from massive amounts of EEG data
Demetres Kostas, Stephane Aroca-Ouellette, Frank Rudzicz
cs.LGcs.NEq-bio.QMarXiv:2101.12037v12021Interpreting Latent Protein Language Model Features with Geometric Annotations
Siddharth Setlur, Djordje Mihajlovic, Darrick Lee
q-bio.QMcs.LGarXiv:2608.26419v12026Applying Deep Learning to Answer Selection: A Study and An Open Task
Minwei Feng, Bing Xiang, Michael R. Glass +2
cs.CLcs.LGarXiv:1508.01585v22015QSGD: Communication-Efficient SGD via Gradient Quantization and Encoding
Dan Alistarh, Demjan Grubic, Jerry Li +2
cs.LGcs.DSarXiv:1610.02132v42016What Makes Multi-modal Learning Better than Single (Provably)
Yu Huang, Chenzhuang Du, Zihui Xue +3
cs.LGcs.AIarXiv:2106.04538v22021Analysis of classifiers' robustness to adversarial perturbations
Alhussein Fawzi, Omar Fawzi, Pascal Frossard
cs.LGcs.CVstat.MLarXiv:1502.02590v42015NeoTriFuse: Reliability-Aware Multimodal Fusion under Missingness Heterogeneity for Neonatal Mortality Risk Prediction
Jiyuan Tian, Qincheng Shen, Ye Lin +2
cs.LGarXiv:2608.26436v12026A Survey on Mixture of Experts in Large Language Models
Weilin Cai, Juyong Jiang, Fan Wang +3
cs.LGcs.CLarXiv:2407.06204v32024Algebraic Multigrid Acceleration for Efficient Label Spreading
Antonia van Betteray, Jonathan Klees, Miriam Schäfers +1
cs.LGarXiv:2608.26309v12026Learning Object Bounding Boxes for 3D Instance Segmentation on Point Clouds
Bo Yang, Jianan Wang, Ronald Clark +4
cs.CVcs.AIcs.LGarXiv:1906.01140v22019RoboTurk: A Crowdsourcing Platform for Robotic Skill Learning through Imitation
Ajay Mandlekar, Yuke Zhu, Animesh Garg +9
cs.ROcs.AIcs.LGarXiv:1811.02790v12018On the Properties of the Softmax Function with Application in Game Theory and Reinforcement Learning
Bolin Gao, Lacra Pavel
math.OCcs.LGarXiv:1704.00805v42017Multi-Dataset Inverse Problem Solving with Distributed Generative AI
Daniel Lersch, Steven Goldenberg, Johann Rudi +5
cs.DCcs.LGarXiv:2608.26283v12026BERTology Meets Biology: Interpreting Attention in Protein Language Models
Jesse Vig, Ali Madani, Lav R. Varshney +3
cs.CLcs.LGq-bio.BMarXiv:2006.15222v32020Supersparse Linear Integer Models for Optimized Medical Scoring Systems
Berk Ustun, Cynthia Rudin
stat.MLcs.DMcs.LGarXiv:1502.04269v32015Pruning Binarized Neural Networks: A Dedicated Framework and Globally Weighted Algorithms
Roan Rubiales, Jean Pierre David
cs.LGarXiv:2608.26233v12026Recommendations with Negative Feedback via Pairwise Deep Reinforcement Learning
Xiangyu Zhao, Liang Zhang, Zhuoye Ding +3
cs.IRcs.LGstat.MLarXiv:1802.06501v32018Vowel Signs Are Not Letters: A Pre-tokenization Ceiling on Multilingual Tokenizer Fertility
Sajal Regmi, Siddhartha Pudasaini, Chetan Phakami Pun
cs.CLcs.LGarXiv:2608.26449v12026Stochastic Optimization with Importance Sampling
Peilin Zhao, Tong Zhang
stat.MLcs.LGarXiv:1401.2753v22014Meta-Reinforcement Learning of Structured Exploration Strategies
Abhishek Gupta, Russell Mendonca, YuXuan Liu +2
cs.LGcs.AIcs.NEarXiv:1802.07245v12018FiLM: Frequency improved Legendre Memory Model for Long-term Time Series Forecasting
Tian Zhou, Ziqing Ma, Xue wang +5
cs.LGstat.MLarXiv:2205.08897v42022EquiBind: Geometric Deep Learning for Drug Binding Structure Prediction
Hannes Stärk, Octavian-Eugen Ganea, Lagnajit Pattanaik +2
q-bio.BMcs.LGarXiv:2202.05146v42022ATOMO: Communication-efficient Learning via Atomic Sparsification
Hongyi Wang, Scott Sievert, Zachary Charles +3
stat.MLcs.DCcs.LGarXiv:1806.04090v32018Transfer Learning with Dynamic Adversarial Adaptation Network
Chaohui Yu, Jindong Wang, Yiqiang Chen +1
cs.LGstat.MLarXiv:1909.08184v12019Hierarchical Federated Learning Across Heterogeneous Cellular Networks
Mehdi Salehi Heydar Abad, Emre Ozfatura, Deniz Gunduz +1
cs.LGcs.DCcs.ITarXiv:1909.02362v12019ThreeDWorld: A Platform for Interactive Multi-Modal Physical Simulation
Chuang Gan, Jeremy Schwartz, Seth Alter +21
cs.CVcs.GRcs.LGarXiv:2007.04954v22020Exploring Randomly Wired Neural Networks for Image Recognition
Saining Xie, Alexander Kirillov, Ross Girshick +1
cs.CVcs.LGarXiv:1904.01569v22019ADAHESSIAN: An Adaptive Second Order Optimizer for Machine Learning
Zhewei Yao, Amir Gholami, Sheng Shen +3
cs.LGmath.NAstat.MLarXiv:2006.00719v32020Sorting out Lipschitz function approximation
Cem Anil, James Lucas, Roger Grosse
cs.LGstat.MLarXiv:1811.05381v22018Recommender Systems with Generative Retrieval
Shashank Rajput, Nikhil Mehta, Anima Singh +10
cs.IRcs.LGarXiv:2305.05065v32023A Review of Graph Neural Networks and Their Applications in Power Systems
Wenlong Liao, Birgitte Bak-Jensen, Jayakrishnan Radhakrishna Pillai +2
cs.LGcs.AIeess.SYarXiv:2101.10025v22021A Survey of Community Detection Approaches: From Statistical Modeling to Deep Learning
Di Jin, Zhizhi Yu, Pengfei Jiao +5
cs.SIcs.AIcs.LGarXiv:2101.01669v32021The Relative Performance of Ensemble Methods with Deep Convolutional Neural Networks for Image Classification
Cheng Ju, Aurélien Bibaut, Mark J. van der Laan
stat.MLcs.CVcs.LGarXiv:1704.01664v12017Rapid Locomotion via Reinforcement Learning
Gabriel B Margolis, Ge Yang, Kartik Paigwar +2
cs.ROcs.AIcs.LGarXiv:2205.02824v12022A Survey on Negative Transfer
Wen Zhang, Lingfei Deng, Lei Zhang +1
cs.LGcs.CVstat.MLarXiv:2009.00909v42020Circuit Condensation: Post-Training that Concentrates a Behavior's Causal Circuit
Sai Adith Senthil Kumar
cs.LGarXiv:2608.27254v12026On the Indistinguishability of Human v/s AI Generated Text
Jaee Ponde, Aritra Das, Mihir More +1
cs.LGarXiv:2608.26797v12026Unifying Detection and Adaptation in Task-Free Continual Learning
Dezheng Han, Anbang Zhang, Zhihao Zhu +1
cs.LGcs.CLarXiv:2608.27070v12026Exploiting Programmatic Behavior of LLMs: Dual-Use Through Standard Security Attacks
Daniel Kang, Xuechen Li, Ion Stoica +3
cs.CRcs.LGarXiv:2302.05733v12023Graph-Bert: Only Attention is Needed for Learning Graph Representations
Jiawei Zhang, Haopeng Zhang, Congying Xia +1
cs.LGcs.NEstat.MLarXiv:2001.05140v22020Vision Transformers are Robust Learners
Sayak Paul, Pin-Yu Chen
cs.CVcs.LGarXiv:2105.07581v32021A Physics-Informed Machine Learning Approach for Solving Heat Transfer Equation in Advanced Manufacturing and Engineering Applications
Navid Zobeiry, Keith D. Humfeld
cs.LGarXiv:2010.02011v12020Review and Comparison of Commonly Used Activation Functions for Deep Neural Networks
Tomasz Szandała
cs.LGcs.NEarXiv:2010.09458v12020Safety by Design: Realized-Cost Constraints for Contextual Bandits with Continuous Actions
Spyros Dragazis, Aldo Pacchiano
cs.LGarXiv:2608.26755v12026Alleviating the Inconsistency Problem of Applying Graph Neural Network to Fraud Detection
Zhiwei Liu, Yingtong Dou, Philip S. Yu +2
cs.SIcs.IRcs.LGarXiv:2005.00625v32020Generating Images with Multimodal Language Models
Jing Yu Koh, Daniel Fried, Ruslan Salakhutdinov
cs.CLcs.CVcs.LGarXiv:2305.17216v32023Dynamical Isometry and a Mean Field Theory of CNNs: How to Train 10,000-Layer Vanilla Convolutional Neural Networks
Lechao Xiao, Yasaman Bahri, Jascha Sohl-Dickstein +2
stat.MLcs.LGarXiv:1806.05393v22018