Machine Learning

Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

4,081 to 4,140 of 20,308

  1. Context-Aware Convolutional Neural Network for Grading of Colorectal Cancer Histology Images

    Muhammad Shaban, Ruqayya Awan, Muhammad Moazam Fraz +3

    eess.IVcs.LGstat.MLarXiv:1907.09478v12019
  2. RubRIX: Rubric-Driven Risk Mitigation in Caregiver-AI Interactions

    Drishti Goel, Jeongah Lee, Qiuyue Joy Zhong +5

    cs.HCcs.AIcs.CLarXiv:2601.13235v12026
  3. Parallelism and Generation Order in Masked Diffusion Language Models: Limits Today, Potential Tomorrow

    Yangyang Zhong, Yanmei Gu, Zhengqing Zang +14

    cs.CLcs.AIcs.LGarXiv:2601.15593v22026
  4. SpiralFormer: Looped Transformers Can Learn Hierarchical Dependencies via Multi-Resolution Recursion

    Chengting Yu, Xiaobo Shu, Yadao Wang +8

    cs.LGarXiv:2602.11698v22026
  5. Scaling Deep Contrastive Learning Batch Size under Memory Limited Setup

    Luyu Gao, Yunyi Zhang, Jiawei Han +1

    cs.LGcs.CLcs.IRarXiv:2101.06983v22021
  6. Covariance Matrix Adaptation for the Rapid Illumination of Behavior Space

    Matthew C. Fontaine, Julian Togelius, Stefanos Nikolaidis +1

    cs.LGstat.MLarXiv:1912.02400v22019
  7. If Influence Functions are the Answer, Then What is the Question?

    Juhan Bae, Nathan Ng, Alston Lo +2

    cs.LGstat.MLarXiv:2209.05364v12022
  8. MMErroR: A Benchmark for Erroneous Reasoning in Vision-Language Models

    Yang Shi, Yifeng Xie, Minzhe Guo +6

    cs.CVcs.AIcs.LGarXiv:2601.03331v22026
  9. Anomaly Detection in Dynamic Graphs via Transformer

    Yixin Liu, Shirui Pan, Yu Guang Wang +4

    cs.LGarXiv:2106.09876v22021
  10. Probability-Entropy Calibration: An Elastic Indicator for Adaptive Fine-tuning

    Wenhao Yu, Shaohang Wei, Jiahong Liu +5

    cs.LGcs.AIarXiv:2602.01745v22026
  11. The Robust Manifold Defense: Adversarial Training using Generative Models

    Ajil Jalal, Andrew Ilyas, Constantinos Daskalakis +1

    cs.CVcs.CRcs.LGarXiv:1712.09196v52017
  12. Match-SRNN: Modeling the Recursive Matching Structure with Spatial RNN

    Shengxian Wan, Yanyan Lan, Jun Xu +3

    cs.CLcs.AIcs.LGarXiv:1604.04378v12016
  13. SceneAlign: Aligning Multimodal Reasoning to Scene Graphs in Complex Visual Scenes

    Chuhan Wang, Xintong Li, Jennifer Yuntong Zhang +5

    cs.CVcs.CLcs.LGarXiv:2601.05600v12026
  14. Beyond Precision: Training-Inference Mismatch is an Optimization Problem and Simple LR Scheduling Fixes It

    Yaxiang Zhang, Yingru Li, Jiacai Liu +4

    cs.LGcs.AIarXiv:2602.01826v12026
  15. SpinalNet: Deep Neural Network with Gradual Input

    H M Dipu Kabir, Moloud Abdar, Seyed Mohammad Jafar Jalali +4

    cs.CVcs.LGcs.NEarXiv:2007.03347v32020
  16. Unsupervised Anomaly Localization using Variational Auto-Encoders

    David Zimmerer, Fabian Isensee, Jens Petersen +2

    cs.LGeess.IVstat.MLarXiv:1907.02796v22019
  17. NextMem: Towards Latent Factual Memory for LLM-based Agents

    Zeyu Zhang, Rui Li, Xiaoyan Zhao +4

    cs.AIcs.IRcs.LGarXiv:2603.15634v12026
  18. Channel-Aware Adversarial Attacks Against Deep Learning-Based Wireless Signal Classifiers

    Brian Kim, Yalin E. Sagduyu, Kemal Davaslioglu +2

    eess.SPcs.LGcs.NIarXiv:2005.05321v32020
  19. Small Language Models: Survey, Measurements, and Insights

    Zhenyan Lu, Xiang Li, Dongqi Cai +5

    cs.CLcs.AIcs.LGarXiv:2409.15790v32024
  20. Nested Slice Sampling: Vectorized Nested Sampling for GPU-Accelerated Inference

    David Yallup, Namu Kroupa, Will Handley

    stat.COcs.LGstat.MLarXiv:2601.23252v22026
  21. Do Latent-CoT Models Think Step-by-Step? A Mechanistic Study on Sequential Reasoning Tasks

    Jia Liang, Liangming Pan

    cs.AIcs.LGarXiv:2602.00449v12026
  22. Inductive Biases and Variable Creation in Self-Attention Mechanisms

    Benjamin L. Edelman, Surbhi Goel, Sham Kakade +1

    cs.LGstat.MLarXiv:2110.10090v22021
  23. Penalizing Gradient Norm for Efficiently Improving Generalization in Deep Learning

    Yang Zhao, Hao Zhang, Xiuyuan Hu

    cs.LGcs.AIarXiv:2202.03599v32022
  24. Between Pure and Approximate Differential Privacy

    Thomas Steinke, Jonathan Ullman

    cs.DScs.CRcs.LGarXiv:1501.06095v12015
  25. NeRS: Neural Reflectance Surfaces for Sparse-view 3D Reconstruction in the Wild

    Jason Y. Zhang, Gengshan Yang, Shubham Tulsiani +1

    cs.CVcs.LGarXiv:2110.07604v32021
  26. Training verified learners with learned verifiers

    Krishnamurthy Dvijotham, Sven Gowal, Robert Stanforth +4

    cs.LGstat.MLarXiv:1805.10265v22018
  27. Online 3D Bin Packing with Constrained Deep Reinforcement Learning

    Hang Zhao, Qijin She, Chenyang Zhu +2

    cs.LGstat.MLarXiv:2006.14978v52020
  28. Towards minimax policies for online linear optimization with bandit feedback

    Sébastien Bubeck, Nicolò Cesa-Bianchi, Sham M. Kakade

    cs.LGstat.MLarXiv:1202.3079v12012
  29. From LLMs to LRMs: Rethinking Pruning for Reasoning-Centric Models

    Longwei Ding, Anhao Zhao, Fanghua Ye +2

    cs.LGarXiv:2601.18091v12026
  30. BalDRO: A Distributionally Robust Optimization based Framework for Large Language Model Unlearning

    Pengyang Shao, Naixin Zhai, Lei Chen +4

    cs.LGarXiv:2601.09172v32026
  31. Temporal Multimodal Fusion for Video Emotion Classification in the Wild

    Valentin Vielzeuf, Stéphane Pateux, Frédéric Jurie

    cs.CVcs.LGcs.MMarXiv:1709.07200v12017
  32. A Signal Propagation Perspective for Pruning Neural Networks at Initialization

    Namhoon Lee, Thalaiyasingam Ajanthan, Stephen Gould +1

    cs.LGcs.CVstat.MLarXiv:1906.06307v22019
  33. Reasoning in Trees: Improving Retrieval-Augmented Generation for Multi-Hop Question Answering

    Yuling Shi, Maolin Sun, Zijun Liu +4

    cs.CLcs.LGarXiv:2601.11255v12026
  34. A Framework for Evaluating Gradient Leakage Attacks in Federated Learning

    Wenqi Wei, Ling Liu, Margaret Loper +4

    cs.LGcs.CRstat.MLarXiv:2004.10397v22020
  35. Image Generators with Conditionally-Independent Pixel Synthesis

    Ivan Anokhin, Kirill Demochkin, Taras Khakhulin +3

    cs.CVcs.AIcs.LGarXiv:2011.13775v12020
  36. Hadamard Response: Estimating Distributions Privately, Efficiently, and with Little Communication

    Jayadev Acharya, Ziteng Sun, Huanyu Zhang

    cs.LGcs.DScs.ITarXiv:1802.04705v22018
  37. Gaussian Process Prior Variational Autoencoders

    Francesco Paolo Casale, Adrian V Dalca, Luca Saglietti +2

    cs.LGstat.MLarXiv:1810.11738v22018
  38. Signed Graph Attention Networks

    Junjie Huang, Huawei Shen, Liang Hou +1

    cs.SIcs.LGphysics.soc-pharXiv:1906.10958v32019
  39. Quantization-Aware Collaborative Inference for Large Embodied AI Models

    Zhonghao Lyu, Ming Xiao, Mikael Skoglund +2

    cs.LGeess.SParXiv:2602.13052v12026
  40. Cast-R1: Learning Tool-Augmented Sequential Decision Policies for Time Series Forecasting

    Xiaoyu Tao, Mingyue Cheng, Chuang Jiang +3

    cs.LGarXiv:2602.13802v12026
  41. Hierarchical Decomposition of Prompt-Based Continual Learning: Rethinking Obscured Sub-optimality

    Liyuan Wang, Jingyi Xie, Xingxing Zhang +3

    cs.LGarXiv:2310.07234v12023
  42. Semi-Supervised Learning of Visual Features by Non-Parametrically Predicting View Assignments with Support Samples

    Mahmoud Assran, Mathilde Caron, Ishan Misra +4

    cs.CVcs.AIcs.LGarXiv:2104.13963v32021
  43. Just-In-Time Reinforcement Learning: Continual Learning in LLM Agents Without Gradient Updates

    Yibo Li, Zijie Lin, Ailin Deng +5

    cs.LGcs.AIarXiv:2601.18510v32026
  44. MNL-Bandit: A Dynamic Learning Approach to Assortment Selection

    Shipra Agrawal, Vashist Avadhanula, Vineet Goyal +1

    cs.LGarXiv:1706.03880v22017
  45. HiPER: Hierarchical Reinforcement Learning with Explicit Credit Assignment for Large Language Model Agents

    Jiangweizhi Peng, Yuanxin Liu, Ruida Zhou +4

    cs.LGcs.AIarXiv:2602.16165v22026
  46. Low-Dimensional and Transversely Curved Optimization Dynamics in Grokking

    Yongzhong Xu

    cs.LGcs.AIarXiv:2602.16746v32026
  47. Co-RedTeam: Orchestrated Security Discovery and Exploitation with LLM Agents

    Pengfei He, Ash Fox, Lesly Miculicich +7

    cs.LGcs.CRarXiv:2602.02164v22026
  48. Weakly-Supervised Video Moment Retrieval via Semantic Completion Network

    Zhijie Lin, Zhou Zhao, Zhu Zhang +2

    cs.CVcs.LGcs.MMarXiv:1911.08199v32019
  49. From Brute Force to Semantic Insight: Performance-Guided Data Transformation Design with LLMs

    Usha Shrestha, Dmitry Ignatov, Radu Timofte

    cs.CVcs.LGarXiv:2601.03808v22026
  50. Privately Learning High-Dimensional Distributions

    Gautam Kamath, Jerry Li, Vikrant Singhal +1

    cs.DScs.CRcs.LGarXiv:1805.00216v32018
  51. Tight Analyses for Non-Smooth Stochastic Gradient Descent

    Nicholas J. A. Harvey, Christopher Liaw, Yaniv Plan +1

    cs.LGmath.OCstat.MLarXiv:1812.05217v12018
  52. Neural Arabic Question Answering

    Hussein Mozannar, Karl El Hajal, Elie Maamary +1

    cs.CLcs.LGarXiv:1906.05394v12019
  53. Graph-Structured Deep Learning Framework for Multi-task Contention Identification with High-dimensional Metrics

    Xiao Yang, Yinan Ni, Yuqi Tang +3

    cs.DCcs.LGarXiv:2601.20389v12026
  54. Improving aircraft performance using machine learning: a review

    Soledad Le Clainche, Esteban Ferrer, Sam Gibson +3

    cs.LGphysics.data-anphysics.flu-dynarXiv:2210.11481v12022
  55. Evaluating explainable artificial intelligence methods for multi-label deep learning classification tasks in remote sensing

    Ioannis Kakogeorgiou, Konstantinos Karantzalos

    cs.LGcs.CVarXiv:2104.01375v22021
  56. Tackling Data Heterogeneity in Federated Learning with Class Prototypes

    Yutong Dai, Zeyuan Chen, Junnan Li +3

    cs.LGcs.AIarXiv:2212.02758v22022
  57. Does a Technique for Building Multimodal Representation Matter? -- Comparative Analysis

    Maciej Pawłowski, Anna Wróblewska, Sylwia Sysko-Romańczuk

    cs.LGarXiv:2206.06367v12022
  58. Understanding Neural Networks via Feature Visualization: A survey

    Anh Nguyen, Jason Yosinski, Jeff Clune

    cs.LGcs.AIcs.CVarXiv:1904.08939v12019
  59. HFedMoE: Resource-aware Heterogeneous Federated Learning with Mixture-of-Experts

    Zihan Fang, Zheng Lin, Senkang Hu +5

    cs.LGcs.AIcs.NIarXiv:2601.00583v12026
  60. Escaping Saddles with Stochastic Gradients

    Hadi Daneshmand, Jonas Kohler, Aurelien Lucchi +1

    cs.LGmath.OCstat.MLarXiv:1803.05999v22018