Machine Learning

Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

3,781 to 3,840 of 20,454

  1. A survey on intrinsic motivation in reinforcement learning

    Arthur Aubret, Laetitia Matignon, Salima Hassas

    cs.LGcs.AIarXiv:1908.06976v22019
  2. AFTer-UNet: Axial Fusion Transformer UNet for Medical Image Segmentation

    Xiangyi Yan, Hao Tang, Shanlin Sun +3

    eess.IVcs.CVcs.LGarXiv:2110.10403v12021
  3. Teach Me to Explain: A Review of Datasets for Explainable Natural Language Processing

    Sarah Wiegreffe, Ana Marasović

    cs.CLcs.AIcs.LGarXiv:2102.12060v42021
  4. Trojaning Language Models for Fun and Profit

    Xinyang Zhang, Zheng Zhang, Shouling Ji +1

    cs.CRcs.CLcs.LGarXiv:2008.00312v22020
  5. BLEU might be Guilty but References are not Innocent

    Markus Freitag, David Grangier, Isaac Caswell

    cs.CLcs.AIcs.LGarXiv:2004.06063v22020
  6. Cluster-to-Conquer: A Framework for End-to-End Multi-Instance Learning for Whole Slide Image Classification

    Yash Sharma, Aman Shrivastava, Lubaina Ehsan +3

    eess.IVcs.CVcs.LGarXiv:2103.10626v22021
  7. Exploiting multi-CNN features in CNN-RNN based Dimensional Emotion Recognition on the OMG in-the-wild Dataset

    Dimitrios Kollias, Stefanos Zafeiriou

    cs.LGcs.CVstat.MLarXiv:1910.01417v22019
  8. Continual Pre-Training of Large Language Models: How to (re)warm your model?

    Kshitij Gupta, Benjamin Thérien, Adam Ibrahim +5

    cs.CLcs.LGarXiv:2308.04014v22023
  9. Multi-level Convolutional Autoencoder Networks for Parametric Prediction of Spatio-temporal Dynamics

    Jiayang Xu, Karthik Duraisamy

    physics.comp-phcs.LGphysics.flu-dynarXiv:1912.11114v22019
  10. Decentralizing Feature Extraction with Quantum Convolutional Neural Network for Automatic Speech Recognition

    Chao-Han Huck Yang, Jun Qi, Samuel Yen-Chi Chen +4

    cs.SDcs.LGcs.NEarXiv:2010.13309v22020
  11. Predicting Clinical Events by Combining Static and Dynamic Information Using Recurrent Neural Networks

    Cristóbal Esteban, Oliver Staeck, Yinchong Yang +1

    cs.LGcs.AIcs.NEarXiv:1602.02685v22016
  12. Taming the Wild: A Unified Analysis of Hogwild!-Style Algorithms

    Christopher De Sa, Ce Zhang, Kunle Olukotun +1

    cs.LGmath.OCstat.MLarXiv:1506.06438v22015
    Summaries:한국어
  13. Twin Contrastive Learning for Online Clustering

    Yunfan Li, Mouxing Yang, Dezhong Peng +3

    cs.LGarXiv:2210.11680v12022
  14. Survey of Machine Learning Accelerators

    Albert Reuther, Peter Michaleas, Michael Jones +3

    cs.DCcs.LGarXiv:2009.00993v12020
  15. Efficient Algorithms for Outlier-Robust Regression

    Adam Klivans, Pravesh K. Kothari, Raghu Meka

    cs.LGcs.AIcs.DSarXiv:1803.03241v32018
  16. How to Build the Virtual Cell with Artificial Intelligence: Priorities and Opportunities

    Charlotte Bunne, Yusuf Roohani, Yanay Rosen +39

    q-bio.QMcs.AIcs.LGarXiv:2409.11654v22024
  17. Guidelines and Evaluation of Clinical Explainable AI in Medical Image Analysis

    Weina Jin, Xiaoxiao Li, Mostafa Fatehi +1

    cs.LGcs.AIcs.CVarXiv:2202.10553v32022
  18. Stochastic Latent Residual Video Prediction

    Jean-Yves Franceschi, Edouard Delasalles, Mickaël Chen +2

    cs.CVcs.LGstat.MLarXiv:2002.09219v42020
  19. Sequential Beats Joint: On the Interplay between On-Policy Distillation and RLVR

    Boyan Li, Bingsen Chen, Chenghao Yang +3

    cs.CLcs.AIcs.LGarXiv:2609.04108v22026
  20. Language Agents with Reinforcement Learning for Strategic Play in the Werewolf Game

    Zelai Xu, Chao Yu, Fei Fang +2

    cs.AIcs.LGcs.MAarXiv:2310.18940v42023
  21. COVID_MTNet: COVID-19 Detection with Multi-Task Deep Learning Approaches

    Md Zahangir Alom, M M Shaifur Rahman, Mst Shamima Nasrin +2

    eess.IVcs.CVcs.LGarXiv:2004.03747v32020
  22. Continual Learning in Low-rank Orthogonal Subspaces

    Arslan Chaudhry, Naeemullah Khan, Puneet K. Dokania +1

    cs.LGarXiv:2010.11635v22020
  23. TeraPipe: Token-Level Pipeline Parallelism for Training Large-Scale Language Models

    Zhuohan Li, Siyuan Zhuang, Shiyuan Guo +4

    cs.LGcs.CLcs.DCarXiv:2102.07988v22021
  24. Network Enhancement: a general method to denoise weighted biological networks

    Bo Wang, Armin Pourshafeie, Marinka Zitnik +4

    q-bio.MNcs.LGcs.SIarXiv:1805.03327v22018
  25. Structured Adversarial Attack: Towards General Implementation and Better Interpretability

    Kaidi Xu, Sijia Liu, Pu Zhao +6

    cs.LGcs.AIstat.MLarXiv:1808.01664v32018
  26. A Distributed Synchronous SGD Algorithm with Global Top-$k$ Sparsification for Low Bandwidth Networks

    Shaohuai Shi, Qiang Wang, Kaiyong Zhao +4

    cs.DCcs.LGarXiv:1901.04359v22019
  27. Experiments of Federated Learning for COVID-19 Chest X-ray Images

    Boyi Liu, Bingjie Yan, Yize Zhou +2

    eess.IVcs.CVcs.LGarXiv:2007.05592v12020
  28. MAP Estimation, Linear Programming and Belief Propagation with Convex Free Energies

    Yair Weiss, Chen Yanover, Talya Meltzer

    cs.AIcs.LGstat.MLarXiv:1206.5286v12012
  29. Mixed Precision Training of Convolutional Neural Networks using Integer Operations

    Dipankar Das, Naveen Mellempudi, Dheevatsa Mudigere +14

    cs.NEcs.LGmath.NAarXiv:1802.00930v22018
  30. Neural Network Attributions: A Causal Perspective

    Aditya Chattopadhyay, Piyushi Manupriya, Anirban Sarkar +1

    cs.LGstat.MLarXiv:1902.02302v42019
  31. Bayesian Action Decoder for Deep Multi-Agent Reinforcement Learning

    Jakob N. Foerster, Francis Song, Edward Hughes +5

    cs.MAcs.AIcs.LGarXiv:1811.01458v32018
  32. Flexible Clustered Federated Learning for Client-Level Data Distribution Shift

    Moming Duan, Duo Liu, Xinyuan Ji +4

    cs.LGcs.DCarXiv:2108.09749v12021
  33. Semi-Supervised Graph Classification: A Hierarchical Graph Perspective

    Jia Li, Yu Rong, Hong Cheng +3

    cs.CVcs.LGarXiv:1904.05003v12019
  34. PU Learning for Matrix Completion

    Cho-Jui Hsieh, Nagarajan Natarajan, Inderjit S. Dhillon

    cs.LGmath.NAstat.MLarXiv:1411.6081v12014
  35. Deep Network Guided Proof Search

    Sarah Loos, Geoffrey Irving, Christian Szegedy +1

    cs.AIcs.LGcs.LOarXiv:1701.06972v12017
  36. A Low-Cost, Open Platform for End-to-End Autonomous Driving on a Miniature Ackermann Vehicle

    Gustavo Claudio Karl Couto, Eric Aislan Antonelo, Gabriel George Zipperer

    cs.LGcs.AIcs.ROarXiv:2609.04147v12026
  37. Learning to Understand Goal Specifications by Modelling Reward

    Dzmitry Bahdanau, Felix Hill, Jan Leike +4

    cs.AIcs.LGarXiv:1806.01946v42018
  38. How Do Classifiers Induce Agents To Invest Effort Strategically?

    Jon Kleinberg, Manish Raghavan

    cs.LGcs.CYcs.DSarXiv:1807.05307v52018
  39. Max-value Entropy Search for Multi-Objective Bayesian Optimization with Constraints

    Syrine Belakaria, Aryan Deshwal, Janardhan Rao Doppa

    cs.LGcs.AIstat.MLarXiv:2009.01721v22020
  40. SimpleMemVLA: A Simple but Effective Native-Video Memory for Vision-Language-Action Models

    Cheng Yin, Wang Xu, Junpeng Yang +8

    cs.CVcs.LGcs.ROarXiv:2609.05533v12026
  41. Federated Unsupervised Representation Learning

    Fengda Zhang, Kun Kuang, Zhaoyang You +6

    cs.LGcs.AIarXiv:2010.08982v12020
  42. Exponential Moving Average Normalization for Self-supervised and Semi-supervised Learning

    Zhaowei Cai, Avinash Ravichandran, Subhransu Maji +3

    cs.LGcs.AIcs.CVarXiv:2101.08482v22021
  43. Actor-Critic Reinforcement Learning for Control with Stability Guarantee

    Minghao Han, Lixian Zhang, Jun Wang +1

    cs.ROcs.LGeess.SYarXiv:2004.14288v32020
  44. Scaling the Scattering Transform: Deep Hybrid Networks

    Edouard Oyallon, Eugene Belilovsky, Sergey Zagoruyko

    cs.CVcs.LGarXiv:1703.08961v22017
  45. Finite Sample Analysis of Stochastic System Identification

    Anastasios Tsiamis, George J. Pappas

    cs.LGeess.SYmath.OCarXiv:1903.09122v12019
  46. Deepcode: Feedback Codes via Deep Learning

    Hyeji Kim, Yihan Jiang, Sreeram Kannan +2

    cs.LGcs.ITstat.MLarXiv:1807.00801v12018
  47. RES: Regularized Stochastic BFGS Algorithm

    Aryan Mokhtari, Alejandro Ribeiro

    cs.LGmath.OCstat.MLarXiv:1401.7625v12014
  48. Influence of Extruded Filament Shape on Buildability in 3D Concrete Printing: A Geometry-Informed Deep Learning-FEM Approach

    Giacomo Rizzieri, Saif-Ur-Rehman, Jörg F. Unger +1

    cs.CEcs.AIcs.LGarXiv:2609.04028v12026
  49. Object Detection for Graphical User Interface: Old Fashioned or Deep Learning or a Combination?

    Jieshan Chen, Mulong Xie, Zhenchang Xing +4

    cs.CVcs.HCcs.LGarXiv:2008.05132v22020
  50. Practical Detection of Trojan Neural Networks: Data-Limited and Data-Free Cases

    Ren Wang, Gaoyuan Zhang, Sijia Liu +3

    cs.LGcs.CRstat.MLarXiv:2007.15802v12020
  51. Personalizing Reinforcement Learning from Human Feedback with Variational Preference Learning

    Sriyash Poddar, Yanming Wan, Hamish Ivison +2

    cs.LGcs.AIcs.CLarXiv:2408.10075v12024
  52. Hypothesis Search: Inductive Reasoning with Language Models

    Ruocheng Wang, Eric Zelikman, Gabriel Poesia +3

    cs.LGcs.AIcs.CLarXiv:2309.05660v22023
  53. Pretraining task diversity and the emergence of non-Bayesian in-context learning for regression

    Allan Raventós, Mansheej Paul, Feng Chen +1

    cs.LGcs.AIcs.CLarXiv:2306.15063v22023
  54. Resiliency of Deep Neural Networks under Quantization

    Wonyong Sung, Sungho Shin, Kyuyeon Hwang

    cs.LGcs.NEarXiv:1511.06488v32015
  55. Subspace Inference Enables Efficient Active Reward Learning from Preferences

    Yutai Zhou, Erdem Bıyık

    cs.LGcs.AIcs.ROarXiv:2609.04066v12026
  56. Deep Attention Recurrent Q-Network

    Ivan Sorokin, Alexey Seleznev, Mikhail Pavlov +2

    cs.LGarXiv:1512.01693v12015
  57. Conservative Safety Critics for Exploration

    Homanga Bharadhwaj, Aviral Kumar, Nicholas Rhinehart +3

    cs.LGcs.AIcs.ROarXiv:2010.14497v22020
  58. KoDF: A Large-scale Korean DeepFake Detection Dataset

    Patrick Kwon, Jaeseong You, Gyuhyeon Nam +2

    cs.CVcs.LGarXiv:2103.10094v22021
  59. Trained Quantization Thresholds for Accurate and Efficient Fixed-Point Inference of Deep Neural Networks

    Sambhav R. Jain, Albert Gural, Michael Wu +1

    cs.CVcs.AIcs.LGarXiv:1903.08066v32019
  60. Score-based Continuous-time Discrete Diffusion Models

    Haoran Sun, Lijun Yu, Bo Dai +2

    cs.LGarXiv:2211.16750v22022