Machine Learning

Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

19,381 to 19,440 of 20,199

  1. VideoMDM: Towards 3D Human Motion Generation From 2D Supervision

    Amir Mann, Gal Michael Harari, Merav Keidar +1

    cs.LGcs.CVarXiv:2606.13364v12026
  2. Looped Language Models Improve Compositional Tool Calling

    Andrei Cristian Popescu, Haitz Sáez de Ocáriz Borde, Pietro Liò

    cs.AIcs.LGarXiv:2608.18171v12026
  3. Where A Small Language Model Helps in Invoice Categorisation, Understood Through Embedding Geometry

    Emma Ceccherini, Daniel Lawson, Anjulika Salhan

    stat.MLcs.LGarXiv:2608.18033v12026
  4. Improved Training of Wasserstein GANs

    Ishaan Gulrajani, Faruk Ahmed, Martin Arjovsky +2

    cs.LGstat.MLarXiv:1704.00028v32017
  5. Identity Mappings in Deep Residual Networks

    Kaiming He, Xiangyu Zhang, Shaoqing Ren +1

    cs.CVcs.LGarXiv:1603.05027v32016
  6. Lymphocyte Mimicry Correction via Region-Level Tissue Reasoning and Unbalanced Optimal Transport

    Xiang Li, Yuqi Wang, Casey C. Heirman +2

    cs.CVcs.LGarXiv:2608.17151v12026
  7. Training Verifiers to Solve Math Word Problems

    Karl Cobbe, Vineet Kosaraju, Mohammad Bavarian +9

    cs.LGcs.CLarXiv:2110.14168v22021
  8. Nonlocal Transition Kernel for Efficient Learning of Restricted Boltzmann Machines

    Kaiji Sekimoto, Muneki Yasuda

    stat.MLcond-mat.dis-nncs.LGarXiv:2608.17450v12026
  9. Hybrid ML for Lightweight Pre-Route Delay Estimation in Open-Source IC Design

    Marvin Castro Castro, Erick Carvajal Barboza

    cs.LGarXiv:2608.17914v12026
  10. Prototypical Networks for Few-shot Learning

    Jake Snell, Kevin Swersky, Richard S. Zemel

    cs.LGstat.MLarXiv:1703.05175v22017
  11. WONDER: A Radio World Model-based Negotiation Framework for Multi-Agent UAV Coverage Optimization

    Jiahao Huang, Rongpeng Li, Zhifeng Zhao +2

    cs.MAcs.LGarXiv:2608.16955v12026
  12. Overcoming catastrophic forgetting in neural networks

    James Kirkpatrick, Razvan Pascanu, Neil Rabinowitz +11

    cs.LGcs.AIstat.MLarXiv:1612.00796v22016
  13. Visual Instruction Tuning

    Haotian Liu, Chunyuan Li, Qingyang Wu +1

    cs.CVcs.AIcs.CLarXiv:2304.08485v22023
  14. Digital Twin-Based Intrusion Detection for Vehicle Powertrain CAN Bus Systems

    Araf Rahman, M Sabbir Salek, Mashrur Chowdhury

    cs.CRcs.LGarXiv:2608.17093v12026
  15. Score-Based Generative Modeling through Stochastic Differential Equations

    Yang Song, Jascha Sohl-Dickstein, Diederik P. Kingma +3

    cs.LGstat.MLarXiv:2011.13456v22020
  16. DeepWalk: Online Learning of Social Representations

    Bryan Perozzi, Rami Al-Rfou, Steven Skiena

    cs.SIcs.LGarXiv:1403.6652v22014
  17. Show, Attend and Tell: Neural Image Caption Generation with Visual Attention

    Kelvin Xu, Jimmy Ba, Ryan Kiros +5

    cs.LGcs.CVarXiv:1502.03044v32015
  18. Wasted large language models: A life cycle thinking approach

    Erik Johannes Husom, Maria Emine Nylund, Ophelia Prillard

    cs.CYcs.HCcs.LGarXiv:2608.17055v12026
  19. Enriching Word Vectors with Subword Information

    Piotr Bojanowski, Edouard Grave, Armand Joulin +1

    cs.CLcs.LGarXiv:1607.04606v22016
  20. SGDR: Stochastic Gradient Descent with Warm Restarts

    Ilya Loshchilov, Frank Hutter

    cs.LGcs.NEmath.OCarXiv:1608.03983v52016
  21. Fashion-MNIST: a Novel Image Dataset for Benchmarking Machine Learning Algorithms

    Han Xiao, Kashif Rasul, Roland Vollgraf

    cs.LGcs.CVstat.MLarXiv:1708.07747v22017
  22. Picture the Epsilon: Pursuing Identity-Level Privacy Guarantees for Images

    Arman Zareian Jahromi, Vishnu Bondalakunta, Mohammad Akbar Bin Shah +3

    cs.CRcs.LGarXiv:2608.17147v12026
  23. MAGPIE-Net: Predicting short-duration heavy-rainfall events in station neighborhoods from multitemporal FY-4A AGRI observations

    Xiang Lin, Yunying Li, Chengzhi Ye +2

    cs.LGarXiv:2608.17753v12026
  24. Conformal Prediction for Molecular Properties under Label Shift

    Hyeonsu Lee, Juyeon Kim, Erkhembayar Jadamba +2

    cs.LGarXiv:2608.17678v12026
  25. Perceptual Losses for Real-Time Style Transfer and Super-Resolution

    Justin Johnson, Alexandre Alahi, Li Fei-Fei

    cs.CVcs.LGarXiv:1603.08155v12016
  26. Repetition as Reinforcement: Enhancing Sample Efficiency via Instant Episode Repetition in Reinforcement Learning

    Hoda Yamani, Yuning Xing, Koen van Rijnsoever +2

    cs.LGcs.ROarXiv:2608.17347v12026
  27. Domain-Adversarial Training of Neural Networks

    Yaroslav Ganin, Evgeniya Ustinova, Hana Ajakan +5

    stat.MLcs.LGcs.NEarXiv:1505.07818v42015
  28. Conditional Generative Adversarial Nets

    Mehdi Mirza, Simon Osindero

    cs.LGcs.AIcs.CVarXiv:1411.1784v12014
  29. TabNSM: Neural Sparse Mixer for Tabular Regression

    Ali Eslamian, Qiang Cheng

    cs.LGcs.CEarXiv:2608.18026v12026
  30. Dropout as a Bayesian Approximation: Representing Model Uncertainty in Deep Learning

    Yarin Gal, Zoubin Ghahramani

    stat.MLcs.LGarXiv:1506.02142v62015
  31. Deep Learning in Neural Networks: An Overview

    Juergen Schmidhuber

    cs.NEcs.LGarXiv:1404.7828v42014
  32. Caffe: Convolutional Architecture for Fast Feature Embedding

    Yangqing Jia, Evan Shelhamer, Jeff Donahue +5

    cs.CVcs.LGcs.NEarXiv:1408.5093v12014
  33. Communication Reduction via Semantic-Based Encoding in DMPC Using LSTMs

    Torben Schiz, Pedro H. J. Nardelli, Henrik Ebel

    eess.SYcs.DCcs.LGarXiv:2608.17592v12026
  34. Training-Free Human-in-the-Loop Anomaly Detection via Memory Bank Correction

    Ayusha Abbas, Saram Abbas, Kabita Adhikari

    cs.LGarXiv:2608.17775v12026
  35. Diffusion Models Beat GANs on Image Synthesis

    Prafulla Dhariwal, Alex Nichol

    cs.LGcs.AIcs.CVarXiv:2105.05233v42021
  36. BART: Denoising Sequence-to-Sequence Pre-training for Natural Language Generation, Translation, and Comprehension

    Mike Lewis, Yinhan Liu, Naman Goyal +5

    cs.CLcs.LGstat.MLarXiv:1910.13461v12019
    Summaries:한국어
  37. SegNet: A Deep Convolutional Encoder-Decoder Architecture for Image Segmentation

    Vijay Badrinarayanan, Alex Kendall, Roberto Cipolla

    cs.CVcs.LGcs.NEarXiv:1511.00561v32015
  38. Prism-GRPO: Faster VLA Policy Optimization via Splitting Same-outcome Groups

    Zeyun Deng, Yuzhe Lu, Yawei Wang +6

    cs.ROcs.LGarXiv:2608.17423v12026
  39. mixup: Beyond Empirical Risk Minimization

    Hongyi Zhang, Moustapha Cisse, Yann N. Dauphin +1

    cs.LGstat.MLarXiv:1710.09412v22017
  40. node2vec: Scalable Feature Learning for Networks

    Aditya Grover, Jure Leskovec

    cs.SIcs.LGstat.MLarXiv:1607.00653v12016
  41. Feature Priming in Online Linear Regression: Sparse-Regret Lower Bounds and a Tight Univariate Rate

    Huibo Xu, Shi Fu, Qixin Zhang +1

    stat.MLcs.LGstat.AParXiv:2608.17573v12026
  42. Layer Normalization

    Jimmy Lei Ba, Jamie Ryan Kiros, Geoffrey E. Hinton

    stat.MLcs.LGarXiv:1607.06450v12016
  43. Efficient Resource Optimization for Split Federated Learning

    Wei Wei, Xianhao Chen

    cs.LGarXiv:2608.17849v12026
  44. A Style-Based Generator Architecture for Generative Adversarial Networks

    Tero Karras, Samuli Laine, Timo Aila

    cs.NEcs.LGstat.MLarXiv:1812.04948v32018
  45. Optimize Your Sampling: Tuned Diffusion Sampling with Bayesian Optimization

    Travis Zhang, Christian Belardi, Justin Lovelace +4

    cs.LGcs.CVarXiv:2608.18040v12026
  46. Fourth-Moment Geometry of Rademacher Sums

    Peigan Gao, Jian Qian

    cs.LGmath.PRarXiv:2608.17802v12026
  47. Empirical Evaluation of Gated Recurrent Neural Networks on Sequence Modeling

    Junyoung Chung, Caglar Gulcehre, KyungHyun Cho +1

    cs.NEcs.LGarXiv:1412.3555v12014
  48. Unsupervised Representation Learning with Deep Convolutional Generative Adversarial Networks

    Alec Radford, Luke Metz, Soumith Chintala

    cs.LGcs.CVarXiv:1511.06434v22015
  49. Revisiting WEASEL 2.0: Reproduction, Sensitivity, and an Adaptive Ensemble-Size Rule

    Cian Higgins, Gerard Carrigan, Pinar Sungu Isiacik +1

    cs.LGarXiv:2608.18021v12026
  50. Towards Deep Learning Models Resistant to Adversarial Attacks

    Aleksander Madry, Aleksandar Makelov, Ludwig Schmidt +2

    stat.MLcs.LGcs.NEarXiv:1706.06083v42017
  51. Continuous control with deep reinforcement learning

    Timothy P. Lillicrap, Jonathan J. Hunt, Alexander Pritzel +5

    cs.LGstat.MLarXiv:1509.02971v62015
  52. Intriguing properties of neural networks

    Christian Szegedy, Wojciech Zaremba, Ilya Sutskever +4

    cs.CVcs.LGcs.NEarXiv:1312.6199v42013
  53. Cross-View Correspondence Is a Measurement Intervention: Two-Sided Validation for Agent Evaluation and Credit Assignment

    Zhen Zhang, Ahmad Hafez, Amr Alanwar

    cs.LGarXiv:2608.17713v12026
  54. Tight Bounds for Data-driven Multiple Hyper-parameter Tuning with Structured Loss Function

    Anh Tuan Nguyen, Viet Anh Nguyen

    cs.LGstat.MLarXiv:2608.17343v12026
  55. AppendiGrade: An XAI-Enhanced Deep Learning Framework for Grading Appendicitis in Ultrasound with Gaussian Blur and Grad-CAM

    Fahad Ahammed, Omar Faruq Shikdar, Navid Zaman +3

    cs.CVcs.LGarXiv:2608.17923v12026
  56. EfficientNet: Rethinking Model Scaling for Convolutional Neural Networks

    Mingxing Tan, Quoc V. Le

    cs.LGcs.CVstat.MLarXiv:1905.11946v52019
  57. Inductive Representation Learning on Large Graphs

    William L. Hamilton, Rex Ying, Jure Leskovec

    cs.SIcs.LGstat.MLarXiv:1706.02216v42017
  58. GUPO: Gradient Uncertainty-aware Policy Optimization for Post-Training Large Language Models

    Peizheng Guo, Jianqi Zhang, Xingyu Zhang +4

    cs.LGarXiv:2608.17411v12026
  59. Intent-Driven Dynamic Chunking: Segmenting Documents to Reflect Predicted Information Needs

    Christos Koutsiaris

    cs.IRcs.AIcs.CLarXiv:2602.14784v12026
  60. Delving Deep into Rectifiers: Surpassing Human-Level Performance on ImageNet Classification

    Kaiming He, Xiangyu Zhang, Shaoqing Ren +1

    cs.CVcs.AIcs.LGarXiv:1502.01852v12015