Machine Learning

Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

13,801 to 13,860 of 20,454

  1. Federated Learning for Malware Detection in IoT Devices

    Valerian Rey, Pedro Miguel Sánchez Sánchez, Alberto Huertas Celdrán +2

    cs.CRcs.LGarXiv:2104.09994v32021
  2. A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation

    Runzhe Yang, Xingyuan Sun, Karthik Narasimhan

    cs.LGcs.AIarXiv:1908.08342v22019
  3. Towards an Automatic Turing Test: Learning to Evaluate Dialogue Responses

    Ryan Lowe, Michael Noseworthy, Iulian V. Serban +3

    cs.CLcs.AIcs.LGarXiv:1708.07149v22017
  4. On Evaluating Adversarial Robustness of Large Vision-Language Models

    Yunqing Zhao, Tianyu Pang, Chao Du +4

    cs.CVcs.CLcs.CRarXiv:2305.16934v22023
  5. Large-scale Classification of Fine-Art Paintings: Learning The Right Metric on The Right Feature

    Babak Saleh, Ahmed Elgammal

    cs.CVcs.IRcs.LGarXiv:1505.00855v12015
  6. Free Lunch for Few-shot Learning: Distribution Calibration

    Shuo Yang, Lu Liu, Min Xu

    cs.LGcs.CVarXiv:2101.06395v32021
  7. District-Level Food Environment Indicators and Social Vulnerability in São Paulo

    Pedro Lemes Sixel Lobo, Eric Tokuda, Kuruvilla Joseph Abraham +4

    physics.soc-phcs.LGstat.AParXiv:2608.26299v12026
  8. Least Ambiguous Set-Valued Classifiers with Bounded Error Levels

    Mauricio Sadinle, Jing Lei, Larry Wasserman

    stat.MEcs.LGstat.MLarXiv:1609.00451v22016
  9. Charting the Right Manifold: Manifold Mixup for Few-shot Learning

    Puneet Mangla, Mayank Singh, Abhishek Sinha +3

    cs.LGcs.CVstat.MLarXiv:1907.12087v42019
  10. Generating 3D Adversarial Point Clouds

    Chong Xiang, Charles R. Qi, Bo Li

    cs.CRcs.CVcs.LGarXiv:1809.07016v42018
  11. One-Shot Imitation from Observing Humans via Domain-Adaptive Meta-Learning

    Tianhe Yu, Chelsea Finn, Annie Xie +4

    cs.LGcs.AIcs.CVarXiv:1802.01557v12018
  12. Iterative Preference Learning from Human Feedback: Bridging Theory and Practice for RLHF under KL-Constraint

    Wei Xiong, Hanze Dong, Chenlu Ye +5

    cs.LGcs.AIstat.MLarXiv:2312.11456v42023
  13. BENDR: using transformers and a contrastive self-supervised learning task to learn from massive amounts of EEG data

    Demetres Kostas, Stephane Aroca-Ouellette, Frank Rudzicz

    cs.LGcs.NEq-bio.QMarXiv:2101.12037v12021
  14. Interpreting Latent Protein Language Model Features with Geometric Annotations

    Siddharth Setlur, Djordje Mihajlovic, Darrick Lee

    q-bio.QMcs.LGarXiv:2608.26419v12026
  15. Applying Deep Learning to Answer Selection: A Study and An Open Task

    Minwei Feng, Bing Xiang, Michael R. Glass +2

    cs.CLcs.LGarXiv:1508.01585v22015
  16. QSGD: Communication-Efficient SGD via Gradient Quantization and Encoding

    Dan Alistarh, Demjan Grubic, Jerry Li +2

    cs.LGcs.DSarXiv:1610.02132v42016
  17. What Makes Multi-modal Learning Better than Single (Provably)

    Yu Huang, Chenzhuang Du, Zihui Xue +3

    cs.LGcs.AIarXiv:2106.04538v22021
  18. Analysis of classifiers' robustness to adversarial perturbations

    Alhussein Fawzi, Omar Fawzi, Pascal Frossard

    cs.LGcs.CVstat.MLarXiv:1502.02590v42015
  19. NeoTriFuse: Reliability-Aware Multimodal Fusion under Missingness Heterogeneity for Neonatal Mortality Risk Prediction

    Jiyuan Tian, Qincheng Shen, Ye Lin +2

    cs.LGarXiv:2608.26436v12026
  20. A Survey on Mixture of Experts in Large Language Models

    Weilin Cai, Juyong Jiang, Fan Wang +3

    cs.LGcs.CLarXiv:2407.06204v32024
  21. Algebraic Multigrid Acceleration for Efficient Label Spreading

    Antonia van Betteray, Jonathan Klees, Miriam Schäfers +1

    cs.LGarXiv:2608.26309v12026
  22. Learning Object Bounding Boxes for 3D Instance Segmentation on Point Clouds

    Bo Yang, Jianan Wang, Ronald Clark +4

    cs.CVcs.AIcs.LGarXiv:1906.01140v22019
  23. RoboTurk: A Crowdsourcing Platform for Robotic Skill Learning through Imitation

    Ajay Mandlekar, Yuke Zhu, Animesh Garg +9

    cs.ROcs.AIcs.LGarXiv:1811.02790v12018
  24. On the Properties of the Softmax Function with Application in Game Theory and Reinforcement Learning

    Bolin Gao, Lacra Pavel

    math.OCcs.LGarXiv:1704.00805v42017
  25. Multi-Dataset Inverse Problem Solving with Distributed Generative AI

    Daniel Lersch, Steven Goldenberg, Johann Rudi +5

    cs.DCcs.LGarXiv:2608.26283v12026
  26. BERTology Meets Biology: Interpreting Attention in Protein Language Models

    Jesse Vig, Ali Madani, Lav R. Varshney +3

    cs.CLcs.LGq-bio.BMarXiv:2006.15222v32020
  27. Supersparse Linear Integer Models for Optimized Medical Scoring Systems

    Berk Ustun, Cynthia Rudin

    stat.MLcs.DMcs.LGarXiv:1502.04269v32015
  28. Pruning Binarized Neural Networks: A Dedicated Framework and Globally Weighted Algorithms

    Roan Rubiales, Jean Pierre David

    cs.LGarXiv:2608.26233v12026
  29. Recommendations with Negative Feedback via Pairwise Deep Reinforcement Learning

    Xiangyu Zhao, Liang Zhang, Zhuoye Ding +3

    cs.IRcs.LGstat.MLarXiv:1802.06501v32018
  30. Vowel Signs Are Not Letters: A Pre-tokenization Ceiling on Multilingual Tokenizer Fertility

    Sajal Regmi, Siddhartha Pudasaini, Chetan Phakami Pun

    cs.CLcs.LGarXiv:2608.26449v12026
  31. Stochastic Optimization with Importance Sampling

    Peilin Zhao, Tong Zhang

    stat.MLcs.LGarXiv:1401.2753v22014
  32. Meta-Reinforcement Learning of Structured Exploration Strategies

    Abhishek Gupta, Russell Mendonca, YuXuan Liu +2

    cs.LGcs.AIcs.NEarXiv:1802.07245v12018
  33. FiLM: Frequency improved Legendre Memory Model for Long-term Time Series Forecasting

    Tian Zhou, Ziqing Ma, Xue wang +5

    cs.LGstat.MLarXiv:2205.08897v42022
  34. EquiBind: Geometric Deep Learning for Drug Binding Structure Prediction

    Hannes Stärk, Octavian-Eugen Ganea, Lagnajit Pattanaik +2

    q-bio.BMcs.LGarXiv:2202.05146v42022
  35. ATOMO: Communication-efficient Learning via Atomic Sparsification

    Hongyi Wang, Scott Sievert, Zachary Charles +3

    stat.MLcs.DCcs.LGarXiv:1806.04090v32018
  36. Transfer Learning with Dynamic Adversarial Adaptation Network

    Chaohui Yu, Jindong Wang, Yiqiang Chen +1

    cs.LGstat.MLarXiv:1909.08184v12019
  37. Hierarchical Federated Learning Across Heterogeneous Cellular Networks

    Mehdi Salehi Heydar Abad, Emre Ozfatura, Deniz Gunduz +1

    cs.LGcs.DCcs.ITarXiv:1909.02362v12019
  38. ThreeDWorld: A Platform for Interactive Multi-Modal Physical Simulation

    Chuang Gan, Jeremy Schwartz, Seth Alter +21

    cs.CVcs.GRcs.LGarXiv:2007.04954v22020
  39. Exploring Randomly Wired Neural Networks for Image Recognition

    Saining Xie, Alexander Kirillov, Ross Girshick +1

    cs.CVcs.LGarXiv:1904.01569v22019
  40. ADAHESSIAN: An Adaptive Second Order Optimizer for Machine Learning

    Zhewei Yao, Amir Gholami, Sheng Shen +3

    cs.LGmath.NAstat.MLarXiv:2006.00719v32020
  41. Sorting out Lipschitz function approximation

    Cem Anil, James Lucas, Roger Grosse

    cs.LGstat.MLarXiv:1811.05381v22018
  42. Recommender Systems with Generative Retrieval

    Shashank Rajput, Nikhil Mehta, Anima Singh +10

    cs.IRcs.LGarXiv:2305.05065v32023
  43. A Review of Graph Neural Networks and Their Applications in Power Systems

    Wenlong Liao, Birgitte Bak-Jensen, Jayakrishnan Radhakrishna Pillai +2

    cs.LGcs.AIeess.SYarXiv:2101.10025v22021
  44. A Survey of Community Detection Approaches: From Statistical Modeling to Deep Learning

    Di Jin, Zhizhi Yu, Pengfei Jiao +5

    cs.SIcs.AIcs.LGarXiv:2101.01669v32021
  45. The Relative Performance of Ensemble Methods with Deep Convolutional Neural Networks for Image Classification

    Cheng Ju, Aurélien Bibaut, Mark J. van der Laan

    stat.MLcs.CVcs.LGarXiv:1704.01664v12017
  46. Rapid Locomotion via Reinforcement Learning

    Gabriel B Margolis, Ge Yang, Kartik Paigwar +2

    cs.ROcs.AIcs.LGarXiv:2205.02824v12022
  47. A Survey on Negative Transfer

    Wen Zhang, Lingfei Deng, Lei Zhang +1

    cs.LGcs.CVstat.MLarXiv:2009.00909v42020
  48. Circuit Condensation: Post-Training that Concentrates a Behavior's Causal Circuit

    Sai Adith Senthil Kumar

    cs.LGarXiv:2608.27254v12026
  49. On the Indistinguishability of Human v/s AI Generated Text

    Jaee Ponde, Aritra Das, Mihir More +1

    cs.LGarXiv:2608.26797v12026
  50. Unifying Detection and Adaptation in Task-Free Continual Learning

    Dezheng Han, Anbang Zhang, Zhihao Zhu +1

    cs.LGcs.CLarXiv:2608.27070v12026
  51. Exploiting Programmatic Behavior of LLMs: Dual-Use Through Standard Security Attacks

    Daniel Kang, Xuechen Li, Ion Stoica +3

    cs.CRcs.LGarXiv:2302.05733v12023
  52. Graph-Bert: Only Attention is Needed for Learning Graph Representations

    Jiawei Zhang, Haopeng Zhang, Congying Xia +1

    cs.LGcs.NEstat.MLarXiv:2001.05140v22020
  53. Vision Transformers are Robust Learners

    Sayak Paul, Pin-Yu Chen

    cs.CVcs.LGarXiv:2105.07581v32021
  54. A Physics-Informed Machine Learning Approach for Solving Heat Transfer Equation in Advanced Manufacturing and Engineering Applications

    Navid Zobeiry, Keith D. Humfeld

    cs.LGarXiv:2010.02011v12020
  55. Review and Comparison of Commonly Used Activation Functions for Deep Neural Networks

    Tomasz Szandała

    cs.LGcs.NEarXiv:2010.09458v12020
  56. Safety by Design: Realized-Cost Constraints for Contextual Bandits with Continuous Actions

    Spyros Dragazis, Aldo Pacchiano

    cs.LGarXiv:2608.26755v12026
  57. Alleviating the Inconsistency Problem of Applying Graph Neural Network to Fraud Detection

    Zhiwei Liu, Yingtong Dou, Philip S. Yu +2

    cs.SIcs.IRcs.LGarXiv:2005.00625v32020
  58. Generating Images with Multimodal Language Models

    Jing Yu Koh, Daniel Fried, Ruslan Salakhutdinov

    cs.CLcs.CVcs.LGarXiv:2305.17216v32023
  59. Dynamical Isometry and a Mean Field Theory of CNNs: How to Train 10,000-Layer Vanilla Convolutional Neural Networks

    Lechao Xiao, Yasaman Bahri, Jascha Sohl-Dickstein +2

    stat.MLcs.LGarXiv:1806.05393v22018
  60. Attributed Network Embedding for Learning in a Dynamic Environment

    Jundong Li, Harsh Dani, Xia Hu +3

    cs.SIcs.LGstat.MLarXiv:1706.01860v22017