Machine Learning
Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
17,221 to 17,280 of 20,215
Federated Learning for Smart Healthcare: A Survey
Dinh C. Nguyen, Quoc-Viet Pham, Pubudu N. Pathirana +5
cs.LGeess.SParXiv:2111.08834v12021Adversarial Attack on Graph Structured Data
Hanjun Dai, Hui Li, Tian Tian +4
cs.LGcs.CRcs.SIarXiv:1806.02371v12018Bilevel Programming for Hyperparameter Optimization and Meta-Learning
Luca Franceschi, Paolo Frasconi, Saverio Salzo +2
stat.MLcs.LGarXiv:1806.04910v22018Measuring Catastrophic Forgetting in Neural Networks
Ronald Kemker, Marc McClure, Angelina Abitino +2
cs.AIcs.CVcs.LGarXiv:1708.02072v42017Edge AI: On-Demand Accelerating Deep Neural Network Inference via Edge Computing
En Li, Liekang Zeng, Zhi Zhou +1
cs.NIcs.CVcs.DCarXiv:1910.05316v12019Machine Learning Testing: Survey, Landscapes and Horizons
Jie M. Zhang, Mark Harman, Lei Ma +1
cs.LGcs.AIcs.SEarXiv:1906.10742v22019Overcoming Exploration in Reinforcement Learning with Demonstrations
Ashvin Nair, Bob McGrew, Marcin Andrychowicz +2
cs.LGcs.AIcs.NEarXiv:1709.10089v22017TIES-Merging: Resolving Interference When Merging Models
Prateek Yadav, Derek Tam, Leshem Choshen +2
cs.LGcs.AIcs.CLarXiv:2306.01708v22023Arcee Trinity Large Technical Report
Varun Singh, Lucas Krauss, Sami Jaghouar +23
cs.LGcs.CLarXiv:2602.17004v12026Explaining NonLinear Classification Decisions with Deep Taylor Decomposition
Grégoire Montavon, Sebastian Bach, Alexander Binder +2
cs.LGstat.MLarXiv:1512.02479v12015LIBERO-Para: A Diagnostic Benchmark and Metrics for Paraphrase Robustness in VLA Models
Chanyoung Kim, Minwoo Kim, Minseok Kang +2
cs.LGarXiv:2603.28301v12026Asymmetric Loss For Multi-Label Classification
Emanuel Ben-Baruch, Tal Ridnik, Nadav Zamir +4
cs.CVcs.LGarXiv:2009.14119v42020DualPrompt: Complementary Prompting for Rehearsal-free Continual Learning
Zifeng Wang, Zizhao Zhang, Sayna Ebrahimi +8
cs.LGcs.CVarXiv:2204.04799v22022Gated Feedback Recurrent Neural Networks
Junyoung Chung, Caglar Gulcehre, Kyunghyun Cho +1
cs.NEcs.LGstat.MLarXiv:1502.02367v42015Towards Automated Kernel Generation in the Era of LLMs
Yang Yu, Peiyu Zang, Chi Hsu Tsai +11
cs.LGcs.CLarXiv:2601.15727v32026Autoregressive Image Generation using Residual Quantization
Doyup Lee, Chiheon Kim, Saehoon Kim +2
cs.CVcs.LGarXiv:2203.01941v22022BARF: Bundle-Adjusting Neural Radiance Fields
Chen-Hsuan Lin, Wei-Chiu Ma, Antonio Torralba +1
cs.CVcs.GRcs.LGarXiv:2104.06405v22021DiffusionCLIP: Text-Guided Diffusion Models for Robust Image Manipulation
Gwanghyun Kim, Taesung Kwon, Jong Chul Ye
cs.CVcs.AIcs.LGarXiv:2110.02711v62021Sparsified SGD with Memory
Sebastian U. Stich, Jean-Baptiste Cordonnier, Martin Jaggi
cs.LGcs.DCcs.DSarXiv:1809.07599v22018Variational Dropout Sparsifies Deep Neural Networks
Dmitry Molchanov, Arsenii Ashukha, Dmitry Vetrov
stat.MLcs.LGarXiv:1701.05369v32017Quantifying Language Models' Sensitivity to Spurious Features in Prompt Design or: How I learned to start worrying about prompt formatting
Melanie Sclar, Yejin Choi, Yulia Tsvetkov +1
cs.CLcs.AIcs.LGarXiv:2310.11324v22023Physics-Informed Neural Operator for Learning Partial Differential Equations
Zongyi Li, Hongkai Zheng, Nikola Kovachki +5
cs.LGmath.NAarXiv:2111.03794v42021Reward-free Alignment for Conflicting Objectives
Peter Chen, Xiaopeng Li, Xi Chen +1
cs.CLcs.AIcs.LGarXiv:2602.02495v32026TuckER: Tensor Factorization for Knowledge Graph Completion
Ivana Balažević, Carl Allen, Timothy M. Hospedales
cs.LGstat.MLarXiv:1901.09590v22019A Benchmark for Interpretability Methods in Deep Neural Networks
Sara Hooker, Dumitru Erhan, Pieter-Jan Kindermans +1
cs.LGcs.AIstat.MLarXiv:1806.10758v32018TabPFN: A Transformer That Solves Small Tabular Classification Problems in a Second
Noah Hollmann, Samuel Müller, Katharina Eggensperger +1
cs.LGstat.MLarXiv:2207.01848v62022Self-Supervised Pre-Training of Swin Transformers for 3D Medical Image Analysis
Yucheng Tang, Dong Yang, Wenqi Li +5
cs.CVcs.AIcs.LGarXiv:2111.14791v22021Parseval Networks: Improving Robustness to Adversarial Examples
Moustapha Cisse, Piotr Bojanowski, Edouard Grave +2
stat.MLcs.AIcs.CRarXiv:1704.08847v22017Thinking in Frames: How Visual Context and Test-Time Scaling Empower Video Reasoning
Chengzu Li, Zanyi Wang, Jiaang Li +9
cs.LGcs.AIcs.CLarXiv:2601.21037v12026A Comprehensive Survey of Deep Learning for Image Captioning
Md. Zakir Hossain, Ferdous Sohel, Mohd Fairuz Shiratuddin +1
cs.CVcs.LGstat.MLarXiv:1810.04020v22018CAD2RL: Real Single-Image Flight without a Single Real Image
Fereshteh Sadeghi, Sergey Levine
cs.LGcs.CVcs.ROarXiv:1611.04201v42016Deep Reconstruction-Classification Networks for Unsupervised Domain Adaptation
Muhammad Ghifary, W. Bastiaan Kleijn, Mengjie Zhang +2
cs.CVcs.AIcs.LGarXiv:1607.03516v22016LLaDA-o: An Effective and Length-Adaptive Omni Diffusion Model
Zebin You, Xiaolu Zhang, Jun Zhou +2
cs.CVcs.LGarXiv:2603.01068v12026Deep learning to represent sub-grid processes in climate models
Stephan Rasp, Michael S. Pritchard, Pierre Gentine
physics.ao-phcs.LGstat.MLarXiv:1806.04731v32018Parallel WaveNet: Fast High-Fidelity Speech Synthesis
Aaron van den Oord, Yazhe Li, Igor Babuschkin +19
cs.LGarXiv:1711.10433v12017Linear representations in language models can change dramatically over a conversation
Andrew Kyle Lampinen, Yuxuan Li, Eghbal Hosseini +2
cs.CLcs.LGarXiv:2601.20834v22026Beyond Imitation: Filtering On-Policy Distillation by Reasoning Progress
Chen Yang, Haiyuan Wan, Rengrong Xiong +2
cs.AIcs.LGarXiv:2608.19408v12026Mechanistic Data Attribution: Tracing the Training Origins of Interpretable LLM Units
Jianhui Chen, Yuzhang Luo, Liangming Pan
cs.CLcs.AIcs.LGarXiv:2601.21996v22026Drug discovery with explainable artificial intelligence
José Jiménez-Luna, Francesca Grisoni, Gisbert Schneider
cs.AIcs.LGstat.MLarXiv:2007.00523v22020Effective Use of Word Order for Text Categorization with Convolutional Neural Networks
Rie Johnson, Tong Zhang
cs.CLcs.LGstat.MLarXiv:1412.1058v22014Policy Distillation
Andrei A. Rusu, Sergio Gomez Colmenarejo, Caglar Gulcehre +6
cs.LGarXiv:1511.06295v22015COVID-CT-Dataset: A CT Scan Dataset about COVID-19
Xingyi Yang, Xuehai He, Jinyu Zhao +3
cs.LGcs.CVeess.IVarXiv:2003.13865v32020TAPAS: Weakly Supervised Table Parsing via Pre-training
Jonathan Herzig, Paweł Krzysztof Nowak, Thomas Müller +2
cs.IRcs.AIcs.CLarXiv:2004.02349v22020A Neural Network Approach to Context-Sensitive Generation of Conversational Responses
Alessandro Sordoni, Michel Galley, Michael Auli +6
cs.CLcs.AIcs.LGarXiv:1506.06714v12015Action-Conditional Video Prediction using Deep Networks in Atari Games
Junhyuk Oh, Xiaoxiao Guo, Honglak Lee +2
cs.LGcs.AIcs.CVarXiv:1507.08750v22015Cross-Task Generalization via Natural Language Crowdsourcing Instructions
Swaroop Mishra, Daniel Khashabi, Chitta Baral +1
cs.CLcs.AIcs.CVarXiv:2104.08773v42021Texygen: A Benchmarking Platform for Text Generation Models
Yaoming Zhu, Sidi Lu, Lei Zheng +4
cs.CLcs.IRcs.LGarXiv:1802.01886v12018Understanding LSTM -- a tutorial into Long Short-Term Memory Recurrent Neural Networks
Ralf C. Staudemeyer, Eric Rothstein Morris
cs.NEcs.CLcs.LGarXiv:1909.09586v12019F-GRPO: Don't Let Your Policy Learn the Obvious and Forget the Rare
Daniil Plyusov, Alexey Gorbatovski, Boris Shaposhnikov +4
cs.LGcs.AIarXiv:2602.06717v22026Large Language Monkeys: Scaling Inference Compute with Repeated Sampling
Bradley Brown, Jordan Juravsky, Ryan Ehrlich +4
cs.LGcs.AIarXiv:2407.21787v32024A Theoretical Analysis of Contrastive Unsupervised Representation Learning
Sanjeev Arora, Hrishikesh Khandeparkar, Mikhail Khodak +2
cs.LGcs.AIstat.MLarXiv:1902.09229v12019Traffic Graph Convolutional Recurrent Neural Network: A Deep Learning Framework for Network-Scale Traffic Learning and Forecasting
Zhiyong Cui, Kristian Henrickson, Ruimin Ke +2
cs.LGstat.MLarXiv:1802.07007v32018ImagenWorld: Stress-Testing Image Generation Models with Explainable Human Evaluation on Open-ended Real-World Tasks
Samin Mahdizadeh Sani, Max Ku, Nima Jamali +23
cs.GRcs.AIcs.CVarXiv:2603.27862v12026Embed to Control: A Locally Linear Latent Dynamics Model for Control from Raw Images
Manuel Watter, Jost Tobias Springenberg, Joschka Boedecker +1
cs.LGcs.CVstat.MLarXiv:1506.07365v32015Prism: Efficient Test-Time Scaling via Hierarchical Search and Self-Verification for Discrete Diffusion Language Models
Jinbin Bai, Yixuan Li, Yuchen Zhu +8
cs.LGarXiv:2602.01842v32026Transfer Learning in Deep Reinforcement Learning: A Survey
Zhuangdi Zhu, Kaixiang Lin, Anil K. Jain +1
cs.LGcs.AIstat.MLarXiv:2009.07888v72020Population Based Training of Neural Networks
Max Jaderberg, Valentin Dalibard, Simon Osindero +9
cs.LGcs.NEarXiv:1711.09846v22017Unitary Evolution Recurrent Neural Networks
Martin Arjovsky, Amar Shah, Yoshua Bengio
cs.LGcs.NEstat.MLarXiv:1511.06464v42015TabTransformer: Tabular Data Modeling Using Contextual Embeddings
Xin Huang, Ashish Khetan, Milan Cvitkovic +1
cs.LGcs.AIarXiv:2012.06678v12020Self-Improving World Modelling with Latent Actions
Yifu Qiu, Zheng Zhao, Waylon Li +4
cs.LGcs.AIcs.CLarXiv:2602.06130v22026