Machine Learning

Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

17,341 to 17,400 of 20,454

  1. S4L: Self-Supervised Semi-Supervised Learning

    Xiaohua Zhai, Avital Oliver, Alexander Kolesnikov +1

    cs.CVcs.LGarXiv:1905.03670v22019
  2. A Safety Report on GPT-5.2, Gemini 3 Pro, Qwen3-VL, Grok 4.1 Fast, Nano Banana Pro, and Seedream 4.5

    Xingjun Ma, Yixu Wang, Hengyuan Xu +18

    cs.AIcs.CLcs.CVarXiv:2601.10527v22026
  3. Fixed Point Quantization of Deep Convolutional Networks

    Darryl D. Lin, Sachin S. Talathi, V. Sreekanth Annapureddy

    cs.LGarXiv:1511.06393v32015
  4. CUA-Suite: Massive Human-annotated Video Demonstrations for Computer-Use Agents

    Xiangru Jian, Shravan Nayak, Kevin Qinghong Lin +5

    cs.LGcs.AIcs.CVarXiv:2603.24440v12026
  5. Clipper: A Low-Latency Online Prediction Serving System

    Daniel Crankshaw, Xin Wang, Giulio Zhou +3

    cs.DCcs.LGarXiv:1612.03079v22016
  6. Perceiver-Actor: A Multi-Task Transformer for Robotic Manipulation

    Mohit Shridhar, Lucas Manuelli, Dieter Fox

    cs.ROcs.AIcs.CLarXiv:2209.05451v22022
  7. Neural Relational Inference for Interacting Systems

    Thomas Kipf, Ethan Fetaya, Kuan-Chieh Wang +2

    stat.MLcs.LGarXiv:1802.04687v22018
  8. SE-DiCoW: Self-Enrolled Diarization-Conditioned Whisper

    Alexander Polok, Dominik Klement, Samuele Cornell +4

    eess.AScs.LGarXiv:2601.19194v12026
  9. Stable Architectures for Deep Neural Networks

    Eldad Haber, Lars Ruthotto

    cs.LGmath.NAmath.OCarXiv:1705.03341v32017
  10. Eternal Sunshine of the Spotless Net: Selective Forgetting in Deep Networks

    Aditya Golatkar, Alessandro Achille, Stefano Soatto

    cs.LGstat.MLarXiv:1911.04933v52019
  11. How to train your ViT? Data, Augmentation, and Regularization in Vision Transformers

    Andreas Steiner, Alexander Kolesnikov, Xiaohua Zhai +3

    cs.CVcs.AIcs.LGarXiv:2106.10270v22021
  12. Towards a Medical AI Scientist

    Hongtao Wu, Boyun Zheng, Dingjie Song +5

    cs.AIcs.LGarXiv:2603.28589v12026
  13. The Neuro-Symbolic Concept Learner: Interpreting Scenes, Words, and Sentences From Natural Supervision

    Jiayuan Mao, Chuang Gan, Pushmeet Kohli +2

    cs.CVcs.AIcs.CLarXiv:1904.12584v12019
  14. Recommendations as Treatments: Debiasing Learning and Evaluation

    Tobias Schnabel, Adith Swaminathan, Ashudeep Singh +2

    cs.LGcs.AIcs.IRarXiv:1602.05352v22016
  15. Adversarially Robust Generalization Requires More Data

    Ludwig Schmidt, Shibani Santurkar, Dimitris Tsipras +2

    cs.LGcs.NEstat.MLarXiv:1804.11285v22018
  16. MeKi: Memory-based Expert Knowledge Injection for Efficient LLM Scaling

    Ning Ding, Fangcheng Liu, Kyungrae Kim +4

    cs.LGcs.AIcs.CLarXiv:2602.03359v12026
  17. Canzona: A Unified, Asynchronous, and Load-Balanced Framework for Distributed Matrix-based Optimizers

    Liangyu Wang, Siqi Zhang, Junjie Wang +7

    cs.DCcs.LGarXiv:2602.06079v12026
  18. On the Scaling of PEFT: Towards Million Personal Models of Trillion Parameters

    Mind Lab, :, Vin Bo +64

    cs.LGcs.CLarXiv:2606.02437v22026
  19. Maxout Networks

    Ian J. Goodfellow, David Warde-Farley, Mehdi Mirza +2

    stat.MLcs.LGarXiv:1302.4389v42013
    Summaries:한국어
  20. Enriching ImageNet with Human Similarity Judgments and Psychological Embeddings

    Brett D. Roads, Bradley C. Love

    cs.CVcs.LGarXiv:2011.11015v12020
    Summaries:한국어
  21. Geometric Deep Learning: Grids, Groups, Graphs, Geodesics, and Gauges

    Michael M. Bronstein, Joan Bruna, Taco Cohen +1

    cs.LGcs.AIcs.CGarXiv:2104.13478v22021
  22. FAAST: Forward-Only Associative Learning via Closed-Form Fast Weights for Test-Time Supervised Adaptation

    Guangsheng Bao, Hongbo Zhang, Han Cui +4

    cs.LGcs.CLarXiv:2605.04651v22026
  23. A Neural Representation of Sketch Drawings

    David Ha, Douglas Eck

    cs.NEcs.LGstat.MLarXiv:1704.03477v42017
  24. #Exploration: A Study of Count-Based Exploration for Deep Reinforcement Learning

    Haoran Tang, Rein Houthooft, Davis Foote +6

    cs.AIcs.LGarXiv:1611.04717v32016
  25. Natural Language Processing (almost) from Scratch

    Ronan Collobert, Jason Weston, Leon Bottou +3

    cs.LGcs.CLarXiv:1103.0398v12011
    Summaries:한국어
  26. Medusa: Simple LLM Inference Acceleration Framework with Multiple Decoding Heads

    Tianle Cai, Yuhong Li, Zhengyang Geng +4

    cs.LGcs.CLarXiv:2401.10774v32024
  27. Post-LayerNorm Is Back: Stable, ExpressivE, and Deep

    Chen Chen, Lai Wei

    cs.LGcs.CLarXiv:2601.19895v22026
  28. Behavior Regularized Offline Reinforcement Learning

    Yifan Wu, George Tucker, Ofir Nachum

    cs.LGcs.AIstat.MLarXiv:1911.11361v12019
  29. StealthRL: Reinforcement Learning Paraphrase Attacks for Multi-Detector Evasion of AI-Text Detectors

    Suraj Ranganath, Atharv Ramesh

    cs.LGcs.AIcs.CRarXiv:2602.08934v22026
  30. Adaptive Graph Convolutional Neural Networks

    Ruoyu Li, Sheng Wang, Feiyun Zhu +1

    cs.LGstat.MLarXiv:1801.03226v12018
  31. Benchmarks Saturate When The Model Gets Smarter Than The Judge

    Marthe Ballon, Andres Algaba, Brecht Verbeken +1

    cs.AIcs.CLcs.LGarXiv:2601.19532v12026
  32. Time-Series Representation Learning via Temporal and Contextual Contrasting

    Emadeldeen Eldele, Mohamed Ragab, Zhenghua Chen +4

    cs.LGcs.AIarXiv:2106.14112v12021
  33. Continual GUI Agents

    Ziwei Liu, Borui Kang, Hangjie Yuan +4

    cs.LGcs.CVarXiv:2601.20732v42026
  34. Nature-Inspired Optimization Algorithms: Challenges and Open Problems

    Xin-She Yang

    cs.NEcs.LGmath.OCarXiv:2003.03776v12020
  35. Explainability in Graph Neural Networks: A Taxonomic Survey

    Hao Yuan, Haiyang Yu, Shurui Gui +1

    cs.LGcs.AIarXiv:2012.15445v32020
  36. Learning Robust Rewards with Adversarial Inverse Reinforcement Learning

    Justin Fu, Katie Luo, Sergey Levine

    cs.LGarXiv:1710.11248v22017
  37. Black-box Generation of Adversarial Text Sequences to Evade Deep Learning Classifiers

    Ji Gao, Jack Lanchantin, Mary Lou Soffa +1

    cs.CLcs.CRcs.IRarXiv:1801.04354v52018
  38. One-Step Evolution for Long-Time Extrapolation: An Error-Bound-Informed and Prior-Guided Neural Residual Framework for Autonomous PDEs

    Maqun Zhang, Feng Gao, Wankun Chen +3

    cs.AIcs.LGarXiv:2608.22026v12026
  39. Exposing the Systematic Vulnerability of Open-Weight Models to Prefill Attacks

    Lukas Struppek, Adam Gleave, Kellin Pelrine

    cs.CRcs.AIcs.CLarXiv:2602.14689v12026
  40. Deep Learning for Sensor-based Human Activity Recognition: Overview, Challenges and Opportunities

    Kaixuan Chen, Dalin Zhang, Lina Yao +3

    cs.HCcs.LGarXiv:2001.07416v22020
  41. AI4SLT: Empirical Processes in Lean 4 for Formal Statistical Learning Theory

    Yuanhe Zhang, Jason D. Lee, Fanghui Liu

    cs.LGcs.CLmath.STarXiv:2602.02285v22026
  42. FILIP: Fine-grained Interactive Language-Image Pre-Training

    Lewei Yao, Runhui Huang, Lu Hou +7

    cs.CVcs.LGarXiv:2111.07783v12021
  43. Conditional Neural Processes

    Marta Garnelo, Dan Rosenbaum, Chris J. Maddison +6

    cs.LGstat.MLarXiv:1807.01613v12018
  44. On Randomness in Agentic Evals

    Bjarni Haukur Bjarnason, André Silva, Martin Monperrus

    cs.LGcs.AIcs.SEarXiv:2602.07150v32026
  45. H$_2$O: Heavy-Hitter Oracle for Efficient Generative Inference of Large Language Models

    Zhenyu Zhang, Ying Sheng, Tianyi Zhou +9

    cs.LGarXiv:2306.14048v32023
  46. Explainable Machine Learning for Scientific Insights and Discoveries

    Ribana Roscher, Bastian Bohn, Marco F. Duarte +1

    cs.LGstat.MLarXiv:1905.08883v32019
  47. Learning a Generative Meta-Model of LLM Activations

    Grace Luo, Jiahai Feng, Trevor Darrell +2

    cs.LGcs.AIcs.CLarXiv:2602.06964v12026
  48. Personalized Cross-Silo Federated Learning on Non-IID Data

    Yutao Huang, Lingyang Chu, Zirui Zhou +4

    cs.LGcs.DCstat.MLarXiv:2007.03797v52020
  49. Masked Feature Prediction for Self-Supervised Visual Pre-Training

    Chen Wei, Haoqi Fan, Saining Xie +3

    cs.CVcs.LGarXiv:2112.09133v22021
  50. Steer2Adapt: Dynamically Composing Steering Vectors Elicits Efficient Adaptation of LLMs

    Pengrui Han, Xueqiang Xu, Keyang Xuan +12

    cs.AIcs.CLcs.LGarXiv:2602.07276v12026
  51. FrugalGPT: How to Use Large Language Models While Reducing Cost and Improving Performance

    Lingjiao Chen, Matei Zaharia, James Zou

    cs.LGcs.AIcs.CLarXiv:2305.05176v12023
  52. Deep Learning for Classical Japanese Literature

    Tarin Clanuwat, Mikel Bober-Irizar, Asanobu Kitamoto +3

    cs.CVcs.LGstat.MLarXiv:1812.01718v12018
  53. Are LLM Decisions Faithful to Verbal Confidence?

    Jiawei Wang, Yanfei Zhou, Siddartha Devic +1

    cs.LGcs.CLarXiv:2601.07767v12026
  54. Domain Adaptation: Learning Bounds and Algorithms

    Yishay Mansour, Mehryar Mohri, Afshin Rostamizadeh

    cs.LGcs.AIarXiv:0902.3430v32009
  55. Attend-and-Excite: Attention-Based Semantic Guidance for Text-to-Image Diffusion Models

    Hila Chefer, Yuval Alaluf, Yael Vinker +2

    cs.CVcs.CLcs.GRarXiv:2301.13826v22023
  56. Whose Opinions Do Language Models Reflect?

    Shibani Santurkar, Esin Durmus, Faisal Ladhak +3

    cs.CLcs.AIcs.CYarXiv:2303.17548v12023
  57. Hints, Critics, and Teachers: Prior Injection for Sparse-Reward RL in Vision-Language Math Reasoning

    Qiqian Fu

    cs.AIcs.LGarXiv:2608.21811v12026
  58. When Gaussian Process Meets Big Data: A Review of Scalable GPs

    Haitao Liu, Yew-Soon Ong, Xiaobo Shen +1

    stat.MLcs.LGarXiv:1807.01065v22018
  59. Semantic Uncertainty: Linguistic Invariances for Uncertainty Estimation in Natural Language Generation

    Lorenz Kuhn, Yarin Gal, Sebastian Farquhar

    cs.CLcs.AIcs.LGarXiv:2302.09664v32023
  60. A Comprehensive Survey of Neural Architecture Search: Challenges and Solutions

    Pengzhen Ren, Yun Xiao, Xiaojun Chang +4

    cs.LGstat.MLarXiv:2006.02903v32020