Machine Learning
Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
4,081 to 4,140 of 20,308
Context-Aware Convolutional Neural Network for Grading of Colorectal Cancer Histology Images
Muhammad Shaban, Ruqayya Awan, Muhammad Moazam Fraz +3
eess.IVcs.LGstat.MLarXiv:1907.09478v12019RubRIX: Rubric-Driven Risk Mitigation in Caregiver-AI Interactions
Drishti Goel, Jeongah Lee, Qiuyue Joy Zhong +5
cs.HCcs.AIcs.CLarXiv:2601.13235v12026Parallelism and Generation Order in Masked Diffusion Language Models: Limits Today, Potential Tomorrow
Yangyang Zhong, Yanmei Gu, Zhengqing Zang +14
cs.CLcs.AIcs.LGarXiv:2601.15593v22026SpiralFormer: Looped Transformers Can Learn Hierarchical Dependencies via Multi-Resolution Recursion
Chengting Yu, Xiaobo Shu, Yadao Wang +8
cs.LGarXiv:2602.11698v22026Scaling Deep Contrastive Learning Batch Size under Memory Limited Setup
Luyu Gao, Yunyi Zhang, Jiawei Han +1
cs.LGcs.CLcs.IRarXiv:2101.06983v22021Covariance Matrix Adaptation for the Rapid Illumination of Behavior Space
Matthew C. Fontaine, Julian Togelius, Stefanos Nikolaidis +1
cs.LGstat.MLarXiv:1912.02400v22019If Influence Functions are the Answer, Then What is the Question?
Juhan Bae, Nathan Ng, Alston Lo +2
cs.LGstat.MLarXiv:2209.05364v12022MMErroR: A Benchmark for Erroneous Reasoning in Vision-Language Models
Yang Shi, Yifeng Xie, Minzhe Guo +6
cs.CVcs.AIcs.LGarXiv:2601.03331v22026Anomaly Detection in Dynamic Graphs via Transformer
Yixin Liu, Shirui Pan, Yu Guang Wang +4
cs.LGarXiv:2106.09876v22021Probability-Entropy Calibration: An Elastic Indicator for Adaptive Fine-tuning
Wenhao Yu, Shaohang Wei, Jiahong Liu +5
cs.LGcs.AIarXiv:2602.01745v22026The Robust Manifold Defense: Adversarial Training using Generative Models
Ajil Jalal, Andrew Ilyas, Constantinos Daskalakis +1
cs.CVcs.CRcs.LGarXiv:1712.09196v52017Match-SRNN: Modeling the Recursive Matching Structure with Spatial RNN
Shengxian Wan, Yanyan Lan, Jun Xu +3
cs.CLcs.AIcs.LGarXiv:1604.04378v12016SceneAlign: Aligning Multimodal Reasoning to Scene Graphs in Complex Visual Scenes
Chuhan Wang, Xintong Li, Jennifer Yuntong Zhang +5
cs.CVcs.CLcs.LGarXiv:2601.05600v12026Beyond Precision: Training-Inference Mismatch is an Optimization Problem and Simple LR Scheduling Fixes It
Yaxiang Zhang, Yingru Li, Jiacai Liu +4
cs.LGcs.AIarXiv:2602.01826v12026SpinalNet: Deep Neural Network with Gradual Input
H M Dipu Kabir, Moloud Abdar, Seyed Mohammad Jafar Jalali +4
cs.CVcs.LGcs.NEarXiv:2007.03347v32020Unsupervised Anomaly Localization using Variational Auto-Encoders
David Zimmerer, Fabian Isensee, Jens Petersen +2
cs.LGeess.IVstat.MLarXiv:1907.02796v22019NextMem: Towards Latent Factual Memory for LLM-based Agents
Zeyu Zhang, Rui Li, Xiaoyan Zhao +4
cs.AIcs.IRcs.LGarXiv:2603.15634v12026Channel-Aware Adversarial Attacks Against Deep Learning-Based Wireless Signal Classifiers
Brian Kim, Yalin E. Sagduyu, Kemal Davaslioglu +2
eess.SPcs.LGcs.NIarXiv:2005.05321v32020Small Language Models: Survey, Measurements, and Insights
Zhenyan Lu, Xiang Li, Dongqi Cai +5
cs.CLcs.AIcs.LGarXiv:2409.15790v32024Nested Slice Sampling: Vectorized Nested Sampling for GPU-Accelerated Inference
David Yallup, Namu Kroupa, Will Handley
stat.COcs.LGstat.MLarXiv:2601.23252v22026Do Latent-CoT Models Think Step-by-Step? A Mechanistic Study on Sequential Reasoning Tasks
Jia Liang, Liangming Pan
cs.AIcs.LGarXiv:2602.00449v12026Inductive Biases and Variable Creation in Self-Attention Mechanisms
Benjamin L. Edelman, Surbhi Goel, Sham Kakade +1
cs.LGstat.MLarXiv:2110.10090v22021Penalizing Gradient Norm for Efficiently Improving Generalization in Deep Learning
Yang Zhao, Hao Zhang, Xiuyuan Hu
cs.LGcs.AIarXiv:2202.03599v32022Between Pure and Approximate Differential Privacy
Thomas Steinke, Jonathan Ullman
cs.DScs.CRcs.LGarXiv:1501.06095v12015NeRS: Neural Reflectance Surfaces for Sparse-view 3D Reconstruction in the Wild
Jason Y. Zhang, Gengshan Yang, Shubham Tulsiani +1
cs.CVcs.LGarXiv:2110.07604v32021Training verified learners with learned verifiers
Krishnamurthy Dvijotham, Sven Gowal, Robert Stanforth +4
cs.LGstat.MLarXiv:1805.10265v22018Online 3D Bin Packing with Constrained Deep Reinforcement Learning
Hang Zhao, Qijin She, Chenyang Zhu +2
cs.LGstat.MLarXiv:2006.14978v52020Towards minimax policies for online linear optimization with bandit feedback
Sébastien Bubeck, Nicolò Cesa-Bianchi, Sham M. Kakade
cs.LGstat.MLarXiv:1202.3079v12012From LLMs to LRMs: Rethinking Pruning for Reasoning-Centric Models
Longwei Ding, Anhao Zhao, Fanghua Ye +2
cs.LGarXiv:2601.18091v12026BalDRO: A Distributionally Robust Optimization based Framework for Large Language Model Unlearning
Pengyang Shao, Naixin Zhai, Lei Chen +4
cs.LGarXiv:2601.09172v32026Temporal Multimodal Fusion for Video Emotion Classification in the Wild
Valentin Vielzeuf, Stéphane Pateux, Frédéric Jurie
cs.CVcs.LGcs.MMarXiv:1709.07200v12017A Signal Propagation Perspective for Pruning Neural Networks at Initialization
Namhoon Lee, Thalaiyasingam Ajanthan, Stephen Gould +1
cs.LGcs.CVstat.MLarXiv:1906.06307v22019Reasoning in Trees: Improving Retrieval-Augmented Generation for Multi-Hop Question Answering
Yuling Shi, Maolin Sun, Zijun Liu +4
cs.CLcs.LGarXiv:2601.11255v12026A Framework for Evaluating Gradient Leakage Attacks in Federated Learning
Wenqi Wei, Ling Liu, Margaret Loper +4
cs.LGcs.CRstat.MLarXiv:2004.10397v22020Image Generators with Conditionally-Independent Pixel Synthesis
Ivan Anokhin, Kirill Demochkin, Taras Khakhulin +3
cs.CVcs.AIcs.LGarXiv:2011.13775v12020Hadamard Response: Estimating Distributions Privately, Efficiently, and with Little Communication
Jayadev Acharya, Ziteng Sun, Huanyu Zhang
cs.LGcs.DScs.ITarXiv:1802.04705v22018Gaussian Process Prior Variational Autoencoders
Francesco Paolo Casale, Adrian V Dalca, Luca Saglietti +2
cs.LGstat.MLarXiv:1810.11738v22018Signed Graph Attention Networks
Junjie Huang, Huawei Shen, Liang Hou +1
cs.SIcs.LGphysics.soc-pharXiv:1906.10958v32019Quantization-Aware Collaborative Inference for Large Embodied AI Models
Zhonghao Lyu, Ming Xiao, Mikael Skoglund +2
cs.LGeess.SParXiv:2602.13052v12026Cast-R1: Learning Tool-Augmented Sequential Decision Policies for Time Series Forecasting
Xiaoyu Tao, Mingyue Cheng, Chuang Jiang +3
cs.LGarXiv:2602.13802v12026Hierarchical Decomposition of Prompt-Based Continual Learning: Rethinking Obscured Sub-optimality
Liyuan Wang, Jingyi Xie, Xingxing Zhang +3
cs.LGarXiv:2310.07234v12023Semi-Supervised Learning of Visual Features by Non-Parametrically Predicting View Assignments with Support Samples
Mahmoud Assran, Mathilde Caron, Ishan Misra +4
cs.CVcs.AIcs.LGarXiv:2104.13963v32021Just-In-Time Reinforcement Learning: Continual Learning in LLM Agents Without Gradient Updates
Yibo Li, Zijie Lin, Ailin Deng +5
cs.LGcs.AIarXiv:2601.18510v32026MNL-Bandit: A Dynamic Learning Approach to Assortment Selection
Shipra Agrawal, Vashist Avadhanula, Vineet Goyal +1
cs.LGarXiv:1706.03880v22017HiPER: Hierarchical Reinforcement Learning with Explicit Credit Assignment for Large Language Model Agents
Jiangweizhi Peng, Yuanxin Liu, Ruida Zhou +4
cs.LGcs.AIarXiv:2602.16165v22026Low-Dimensional and Transversely Curved Optimization Dynamics in Grokking
Yongzhong Xu
cs.LGcs.AIarXiv:2602.16746v32026Co-RedTeam: Orchestrated Security Discovery and Exploitation with LLM Agents
Pengfei He, Ash Fox, Lesly Miculicich +7
cs.LGcs.CRarXiv:2602.02164v22026Weakly-Supervised Video Moment Retrieval via Semantic Completion Network
Zhijie Lin, Zhou Zhao, Zhu Zhang +2
cs.CVcs.LGcs.MMarXiv:1911.08199v32019From Brute Force to Semantic Insight: Performance-Guided Data Transformation Design with LLMs
Usha Shrestha, Dmitry Ignatov, Radu Timofte
cs.CVcs.LGarXiv:2601.03808v22026Privately Learning High-Dimensional Distributions
Gautam Kamath, Jerry Li, Vikrant Singhal +1
cs.DScs.CRcs.LGarXiv:1805.00216v32018Tight Analyses for Non-Smooth Stochastic Gradient Descent
Nicholas J. A. Harvey, Christopher Liaw, Yaniv Plan +1
cs.LGmath.OCstat.MLarXiv:1812.05217v12018Neural Arabic Question Answering
Hussein Mozannar, Karl El Hajal, Elie Maamary +1
cs.CLcs.LGarXiv:1906.05394v12019Graph-Structured Deep Learning Framework for Multi-task Contention Identification with High-dimensional Metrics
Xiao Yang, Yinan Ni, Yuqi Tang +3
cs.DCcs.LGarXiv:2601.20389v12026Improving aircraft performance using machine learning: a review
Soledad Le Clainche, Esteban Ferrer, Sam Gibson +3
cs.LGphysics.data-anphysics.flu-dynarXiv:2210.11481v12022Evaluating explainable artificial intelligence methods for multi-label deep learning classification tasks in remote sensing
Ioannis Kakogeorgiou, Konstantinos Karantzalos
cs.LGcs.CVarXiv:2104.01375v22021Tackling Data Heterogeneity in Federated Learning with Class Prototypes
Yutong Dai, Zeyuan Chen, Junnan Li +3
cs.LGcs.AIarXiv:2212.02758v22022Does a Technique for Building Multimodal Representation Matter? -- Comparative Analysis
Maciej Pawłowski, Anna Wróblewska, Sylwia Sysko-Romańczuk
cs.LGarXiv:2206.06367v12022Understanding Neural Networks via Feature Visualization: A survey
Anh Nguyen, Jason Yosinski, Jeff Clune
cs.LGcs.AIcs.CVarXiv:1904.08939v12019HFedMoE: Resource-aware Heterogeneous Federated Learning with Mixture-of-Experts
Zihan Fang, Zheng Lin, Senkang Hu +5
cs.LGcs.AIcs.NIarXiv:2601.00583v12026Escaping Saddles with Stochastic Gradients
Hadi Daneshmand, Jonas Kohler, Aurelien Lucchi +1
cs.LGmath.OCstat.MLarXiv:1803.05999v22018