Machine Learning
Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
17,341 to 17,400 of 20,454
S4L: Self-Supervised Semi-Supervised Learning
Xiaohua Zhai, Avital Oliver, Alexander Kolesnikov +1
cs.CVcs.LGarXiv:1905.03670v22019A Safety Report on GPT-5.2, Gemini 3 Pro, Qwen3-VL, Grok 4.1 Fast, Nano Banana Pro, and Seedream 4.5
Xingjun Ma, Yixu Wang, Hengyuan Xu +18
cs.AIcs.CLcs.CVarXiv:2601.10527v22026Fixed Point Quantization of Deep Convolutional Networks
Darryl D. Lin, Sachin S. Talathi, V. Sreekanth Annapureddy
cs.LGarXiv:1511.06393v32015CUA-Suite: Massive Human-annotated Video Demonstrations for Computer-Use Agents
Xiangru Jian, Shravan Nayak, Kevin Qinghong Lin +5
cs.LGcs.AIcs.CVarXiv:2603.24440v12026Clipper: A Low-Latency Online Prediction Serving System
Daniel Crankshaw, Xin Wang, Giulio Zhou +3
cs.DCcs.LGarXiv:1612.03079v22016Perceiver-Actor: A Multi-Task Transformer for Robotic Manipulation
Mohit Shridhar, Lucas Manuelli, Dieter Fox
cs.ROcs.AIcs.CLarXiv:2209.05451v22022Neural Relational Inference for Interacting Systems
Thomas Kipf, Ethan Fetaya, Kuan-Chieh Wang +2
stat.MLcs.LGarXiv:1802.04687v22018SE-DiCoW: Self-Enrolled Diarization-Conditioned Whisper
Alexander Polok, Dominik Klement, Samuele Cornell +4
eess.AScs.LGarXiv:2601.19194v12026Stable Architectures for Deep Neural Networks
Eldad Haber, Lars Ruthotto
cs.LGmath.NAmath.OCarXiv:1705.03341v32017Eternal Sunshine of the Spotless Net: Selective Forgetting in Deep Networks
Aditya Golatkar, Alessandro Achille, Stefano Soatto
cs.LGstat.MLarXiv:1911.04933v52019How to train your ViT? Data, Augmentation, and Regularization in Vision Transformers
Andreas Steiner, Alexander Kolesnikov, Xiaohua Zhai +3
cs.CVcs.AIcs.LGarXiv:2106.10270v22021Towards a Medical AI Scientist
Hongtao Wu, Boyun Zheng, Dingjie Song +5
cs.AIcs.LGarXiv:2603.28589v12026The Neuro-Symbolic Concept Learner: Interpreting Scenes, Words, and Sentences From Natural Supervision
Jiayuan Mao, Chuang Gan, Pushmeet Kohli +2
cs.CVcs.AIcs.CLarXiv:1904.12584v12019Recommendations as Treatments: Debiasing Learning and Evaluation
Tobias Schnabel, Adith Swaminathan, Ashudeep Singh +2
cs.LGcs.AIcs.IRarXiv:1602.05352v22016Adversarially Robust Generalization Requires More Data
Ludwig Schmidt, Shibani Santurkar, Dimitris Tsipras +2
cs.LGcs.NEstat.MLarXiv:1804.11285v22018MeKi: Memory-based Expert Knowledge Injection for Efficient LLM Scaling
Ning Ding, Fangcheng Liu, Kyungrae Kim +4
cs.LGcs.AIcs.CLarXiv:2602.03359v12026Canzona: A Unified, Asynchronous, and Load-Balanced Framework for Distributed Matrix-based Optimizers
Liangyu Wang, Siqi Zhang, Junjie Wang +7
cs.DCcs.LGarXiv:2602.06079v12026On the Scaling of PEFT: Towards Million Personal Models of Trillion Parameters
Mind Lab, :, Vin Bo +64
cs.LGcs.CLarXiv:2606.02437v22026Maxout Networks
Ian J. Goodfellow, David Warde-Farley, Mehdi Mirza +2
stat.MLcs.LGarXiv:1302.4389v42013Summaries:한국어Enriching ImageNet with Human Similarity Judgments and Psychological Embeddings
Brett D. Roads, Bradley C. Love
cs.CVcs.LGarXiv:2011.11015v12020Summaries:한국어Geometric Deep Learning: Grids, Groups, Graphs, Geodesics, and Gauges
Michael M. Bronstein, Joan Bruna, Taco Cohen +1
cs.LGcs.AIcs.CGarXiv:2104.13478v22021FAAST: Forward-Only Associative Learning via Closed-Form Fast Weights for Test-Time Supervised Adaptation
Guangsheng Bao, Hongbo Zhang, Han Cui +4
cs.LGcs.CLarXiv:2605.04651v22026A Neural Representation of Sketch Drawings
David Ha, Douglas Eck
cs.NEcs.LGstat.MLarXiv:1704.03477v42017#Exploration: A Study of Count-Based Exploration for Deep Reinforcement Learning
Haoran Tang, Rein Houthooft, Davis Foote +6
cs.AIcs.LGarXiv:1611.04717v32016Natural Language Processing (almost) from Scratch
Ronan Collobert, Jason Weston, Leon Bottou +3
cs.LGcs.CLarXiv:1103.0398v12011Summaries:한국어Medusa: Simple LLM Inference Acceleration Framework with Multiple Decoding Heads
Tianle Cai, Yuhong Li, Zhengyang Geng +4
cs.LGcs.CLarXiv:2401.10774v32024Post-LayerNorm Is Back: Stable, ExpressivE, and Deep
Chen Chen, Lai Wei
cs.LGcs.CLarXiv:2601.19895v22026Behavior Regularized Offline Reinforcement Learning
Yifan Wu, George Tucker, Ofir Nachum
cs.LGcs.AIstat.MLarXiv:1911.11361v12019StealthRL: Reinforcement Learning Paraphrase Attacks for Multi-Detector Evasion of AI-Text Detectors
Suraj Ranganath, Atharv Ramesh
cs.LGcs.AIcs.CRarXiv:2602.08934v22026Adaptive Graph Convolutional Neural Networks
Ruoyu Li, Sheng Wang, Feiyun Zhu +1
cs.LGstat.MLarXiv:1801.03226v12018Benchmarks Saturate When The Model Gets Smarter Than The Judge
Marthe Ballon, Andres Algaba, Brecht Verbeken +1
cs.AIcs.CLcs.LGarXiv:2601.19532v12026Time-Series Representation Learning via Temporal and Contextual Contrasting
Emadeldeen Eldele, Mohamed Ragab, Zhenghua Chen +4
cs.LGcs.AIarXiv:2106.14112v12021Continual GUI Agents
Ziwei Liu, Borui Kang, Hangjie Yuan +4
cs.LGcs.CVarXiv:2601.20732v42026Nature-Inspired Optimization Algorithms: Challenges and Open Problems
Xin-She Yang
cs.NEcs.LGmath.OCarXiv:2003.03776v12020Explainability in Graph Neural Networks: A Taxonomic Survey
Hao Yuan, Haiyang Yu, Shurui Gui +1
cs.LGcs.AIarXiv:2012.15445v32020Learning Robust Rewards with Adversarial Inverse Reinforcement Learning
Justin Fu, Katie Luo, Sergey Levine
cs.LGarXiv:1710.11248v22017Black-box Generation of Adversarial Text Sequences to Evade Deep Learning Classifiers
Ji Gao, Jack Lanchantin, Mary Lou Soffa +1
cs.CLcs.CRcs.IRarXiv:1801.04354v52018One-Step Evolution for Long-Time Extrapolation: An Error-Bound-Informed and Prior-Guided Neural Residual Framework for Autonomous PDEs
Maqun Zhang, Feng Gao, Wankun Chen +3
cs.AIcs.LGarXiv:2608.22026v12026Exposing the Systematic Vulnerability of Open-Weight Models to Prefill Attacks
Lukas Struppek, Adam Gleave, Kellin Pelrine
cs.CRcs.AIcs.CLarXiv:2602.14689v12026Deep Learning for Sensor-based Human Activity Recognition: Overview, Challenges and Opportunities
Kaixuan Chen, Dalin Zhang, Lina Yao +3
cs.HCcs.LGarXiv:2001.07416v22020AI4SLT: Empirical Processes in Lean 4 for Formal Statistical Learning Theory
Yuanhe Zhang, Jason D. Lee, Fanghui Liu
cs.LGcs.CLmath.STarXiv:2602.02285v22026FILIP: Fine-grained Interactive Language-Image Pre-Training
Lewei Yao, Runhui Huang, Lu Hou +7
cs.CVcs.LGarXiv:2111.07783v12021Conditional Neural Processes
Marta Garnelo, Dan Rosenbaum, Chris J. Maddison +6
cs.LGstat.MLarXiv:1807.01613v12018On Randomness in Agentic Evals
Bjarni Haukur Bjarnason, André Silva, Martin Monperrus
cs.LGcs.AIcs.SEarXiv:2602.07150v32026H$_2$O: Heavy-Hitter Oracle for Efficient Generative Inference of Large Language Models
Zhenyu Zhang, Ying Sheng, Tianyi Zhou +9
cs.LGarXiv:2306.14048v32023Explainable Machine Learning for Scientific Insights and Discoveries
Ribana Roscher, Bastian Bohn, Marco F. Duarte +1
cs.LGstat.MLarXiv:1905.08883v32019Learning a Generative Meta-Model of LLM Activations
Grace Luo, Jiahai Feng, Trevor Darrell +2
cs.LGcs.AIcs.CLarXiv:2602.06964v12026Personalized Cross-Silo Federated Learning on Non-IID Data
Yutao Huang, Lingyang Chu, Zirui Zhou +4
cs.LGcs.DCstat.MLarXiv:2007.03797v52020Masked Feature Prediction for Self-Supervised Visual Pre-Training
Chen Wei, Haoqi Fan, Saining Xie +3
cs.CVcs.LGarXiv:2112.09133v22021Steer2Adapt: Dynamically Composing Steering Vectors Elicits Efficient Adaptation of LLMs
Pengrui Han, Xueqiang Xu, Keyang Xuan +12
cs.AIcs.CLcs.LGarXiv:2602.07276v12026FrugalGPT: How to Use Large Language Models While Reducing Cost and Improving Performance
Lingjiao Chen, Matei Zaharia, James Zou
cs.LGcs.AIcs.CLarXiv:2305.05176v12023Deep Learning for Classical Japanese Literature
Tarin Clanuwat, Mikel Bober-Irizar, Asanobu Kitamoto +3
cs.CVcs.LGstat.MLarXiv:1812.01718v12018Are LLM Decisions Faithful to Verbal Confidence?
Jiawei Wang, Yanfei Zhou, Siddartha Devic +1
cs.LGcs.CLarXiv:2601.07767v12026Domain Adaptation: Learning Bounds and Algorithms
Yishay Mansour, Mehryar Mohri, Afshin Rostamizadeh
cs.LGcs.AIarXiv:0902.3430v32009Attend-and-Excite: Attention-Based Semantic Guidance for Text-to-Image Diffusion Models
Hila Chefer, Yuval Alaluf, Yael Vinker +2
cs.CVcs.CLcs.GRarXiv:2301.13826v22023Whose Opinions Do Language Models Reflect?
Shibani Santurkar, Esin Durmus, Faisal Ladhak +3
cs.CLcs.AIcs.CYarXiv:2303.17548v12023Hints, Critics, and Teachers: Prior Injection for Sparse-Reward RL in Vision-Language Math Reasoning
Qiqian Fu
cs.AIcs.LGarXiv:2608.21811v12026When Gaussian Process Meets Big Data: A Review of Scalable GPs
Haitao Liu, Yew-Soon Ong, Xiaobo Shen +1
stat.MLcs.LGarXiv:1807.01065v22018Semantic Uncertainty: Linguistic Invariances for Uncertainty Estimation in Natural Language Generation
Lorenz Kuhn, Yarin Gal, Sebastian Farquhar
cs.CLcs.AIcs.LGarXiv:2302.09664v32023A Comprehensive Survey of Neural Architecture Search: Challenges and Solutions
Pengzhen Ren, Yun Xiao, Xiaojun Chang +4
cs.LGstat.MLarXiv:2006.02903v32020