Machine Learning

Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

5,461 to 5,520 of 19,972

  1. Measuring consistency via ensemble margin and local prediction variability: Auditing decision systems in the presence of predictive multiplicity

    Sinjini Banerjee, Tim Marrinan, Anand D. Sarwate

    stat.MLcs.AIcs.LGarXiv:2609.01397v12026
  2. One-Prompt-One-Story: Free-Lunch Consistent Text-to-Image Generation Using a Single Prompt

    Tao Liu, Kai Wang, Senmao Li +6

    cs.CVcs.AIcs.LGarXiv:2501.13554v32025
  3. Spatial Broadcast Decoder: A Simple Architecture for Learning Disentangled Representations in VAEs

    Nicholas Watters, Loic Matthey, Christopher P. Burgess +1

    cs.LGcs.CVstat.MLarXiv:1901.07017v22019
  4. DualStake: Dual-Path Confidence Calibration in Deep Research Agents

    Yinuo Xu, Yuwei Liang, Jianjie Cheng +4

    cs.CLcs.AIcs.LGarXiv:2609.00935v12026
  5. A Wholistic View of Continual Learning with Deep Neural Networks: Forgotten Lessons and the Bridge to Active and Open World Learning

    Martin Mundt, Yongwon Hong, Iuliia Pliushch +1

    cs.LGstat.MLarXiv:2009.01797v32020
  6. Confess What You Know: Forget-Set Misalignment with Model Knowledge in LLM Unlearning

    Miso Kim, Georu Lee, Seungwon Jeong +1

    cs.LGcs.AIcs.CLarXiv:2609.00605v12026
  7. A Study of Hidden-State Optimization Order in Predictive Coding Networks

    Xueyuan Li, Danilo Vasconcellos Vargas

    cs.LGcs.AIarXiv:2609.00686v12026
  8. Fast, Exact and Multi-Scale Inference for Semantic Image Segmentation with Deep Gaussian CRFs

    Siddhartha Chandra, Iasonas Kokkinos

    cs.CVcs.LGarXiv:1603.08358v42016
  9. Universal Model Routing for Efficient LLM Inference

    Wittawat Jitkrittum, Harikrishna Narasimhan, Ankit Singh Rawat +9

    cs.CLcs.LGarXiv:2502.08773v22025
  10. ViTAMINS: An Empirical Study of Training Self-Supervised Vision Transformers with Synthetic Hard Negatives

    Nikos Giakoumoglou, Andreas Floros, Kleanthis-Marios Papadopoulos +1

    cs.CVcs.AIcs.LGarXiv:2609.01041v12026
  11. Sparse Autoencoders Do Not Find Canonical Units of Analysis

    Patrick Leask, Bart Bussmann, Michael Pearce +5

    cs.LGcs.AIarXiv:2502.04878v12025
  12. Momentum Contrastive Learning for Few-Shot COVID-19 Diagnosis from Chest CT Images

    Xiaocong Chen, Lina Yao, Tao Zhou +2

    eess.IVcs.CVcs.LGarXiv:2006.13276v12020
  13. SMART: Self-Aware Agent for Tool Overuse Mitigation

    Cheng Qian, Emre Can Acikgoz, Hongru Wang +5

    cs.AIcs.CLcs.LGarXiv:2502.11435v22025
  14. Learning Generalizable Robotic Reward Functions from "In-The-Wild" Human Videos

    Annie S. Chen, Suraj Nair, Chelsea Finn

    cs.ROcs.AIcs.CVarXiv:2103.16817v12021
  15. On the Doubt about Margin Explanation of Boosting

    Wei Gao, Zhi-Hua Zhou

    cs.LGarXiv:1009.3613v52010
  16. Embedded Conditional Independence Tests for Large Language Model Generated Text with an Application to German Parliament Speeches

    Marco Simnacher, Georg Keilbar, Benjamin König +2

    stat.MLcs.AIcs.LGarXiv:2609.00946v12026
  17. TimeNet: Pre-trained deep recurrent neural network for time series classification

    Pankaj Malhotra, Vishnu TV, Lovekesh Vig +2

    cs.LGarXiv:1706.08838v12017
  18. Learning the Travelling Salesperson Problem Requires Rethinking Generalization

    Chaitanya K. Joshi, Quentin Cappart, Louis-Martin Rousseau +1

    cs.LGstat.MLarXiv:2006.07054v62020
  19. GPTAQ: Efficient Finetuning-Free Quantization for Asymmetric Calibration

    Yuhang Li, Ruokai Yin, Donghyun Lee +2

    cs.LGarXiv:2504.02692v32025
  20. FlexiViT: One Model for All Patch Sizes

    Lucas Beyer, Pavel Izmailov, Alexander Kolesnikov +7

    cs.CVcs.AIcs.LGarXiv:2212.08013v22022
  21. Inference-Time Alignment in Diffusion Models with Reward-Guided Generation: Tutorial and Review

    Masatoshi Uehara, Yulai Zhao, Chenyu Wang +4

    cs.AIcs.LGq-bio.QMarXiv:2501.09685v22025
  22. Learning-Theoretic Foundation for General Coded Computing: The Straggler Setting

    Parsa Moradi, Behrooz Tahmasebi, Mohammad Ali Maddah-Ali

    cs.LGarXiv:2608.28910v12026
  23. FaSNet: Low-latency Adaptive Beamforming for Multi-microphone Audio Processing

    Yi Luo, Enea Ceolini, Cong Han +2

    eess.AScs.LGcs.SDarXiv:1909.13387v22019
  24. Reconstructing Training Data from Trained Neural Networks

    Niv Haim, Gal Vardi, Gilad Yehudai +2

    cs.LGcs.CRcs.CVarXiv:2206.07758v32022
  25. Not All Language Model Features Are One-Dimensionally Linear

    Joshua Engels, Eric J. Michaud, Isaac Liao +2

    cs.LGarXiv:2405.14860v32024
  26. One Token to Fool LLM-as-a-Judge

    Yulai Zhao, Haolin Liu, Dian Yu +4

    cs.LGcs.CLarXiv:2507.08794v32025
  27. HelpSteer3-Preference: Open Human-Annotated Preference Data across Diverse Tasks and Languages

    Zhilin Wang, Jiaqi Zeng, Olivier Delalleau +6

    cs.CLcs.AIcs.LGarXiv:2505.11475v22025
  28. Convergence issues in Relational Concept Analysis based on AOC-posets

    Xavier Dolques, Agnès Braud, Alain Gutierrez +2

    cs.LGarXiv:2609.00054v12026
  29. AdaComp : Adaptive Residual Gradient Compression for Data-Parallel Distributed Training

    Chia-Yu Chen, Jungwook Choi, Daniel Brand +3

    cs.LGstat.MLarXiv:1712.02679v12017
  30. GSPMD: General and Scalable Parallelization for ML Computation Graphs

    Yuanzhong Xu, HyoukJoong Lee, Dehao Chen +13

    cs.DCcs.LGarXiv:2105.04663v22021
  31. Sample Complexity Bounds for Stochastic Shortest Path with a Generative Model

    Jean Tarbouriech, Matteo Pirotta, Michal Valko +1

    cs.LGstat.MLarXiv:2604.16111v12026
  32. Dion: Distributed Orthonormalized Updates

    Kwangjun Ahn, Byron Xu, Natalie Abreu +5

    cs.LGcs.AImath.OCarXiv:2504.05295v32025
  33. A Study of Reinforcement Learning for Neural Machine Translation

    Lijun Wu, Fei Tian, Tao Qin +2

    cs.LGcs.AIstat.MLarXiv:1808.08866v12018
  34. An Attention Free Transformer

    Shuangfei Zhai, Walter Talbott, Nitish Srivastava +4

    cs.LGcs.CLcs.CVarXiv:2105.14103v22021
  35. Multi-Scale Contrastive Siamese Networks for Self-Supervised Graph Representation Learning

    Ming Jin, Yizhen Zheng, Yuan-Fang Li +3

    cs.LGcs.SIarXiv:2105.05682v22021
  36. DNA-inspired online behavioral modeling and its application to spambot detection

    Stefano Cresci, Roberto Di Pietro, Marinella Petrocchi +2

    cs.SIcs.CRcs.LGarXiv:1602.00110v12016
  37. Self-Reports Are Not Verification: Environment-Grounded Auditing of LLM Operators in Evolutionary Search

    Enrong Pan, Ryan Zhou, Ting Hu

    cs.AIcs.LGcs.NEarXiv:2609.00652v12026
  38. The challenge of realistic music generation: modelling raw audio at scale

    Sander Dieleman, Aäron van den Oord, Karen Simonyan

    cs.SDcs.LGeess.ASarXiv:1806.10474v12018
  39. VTool-R1: VLMs Learn to Think with Images via Reinforcement Learning on Multimodal Tool Use

    Mingyuan Wu, Jingcheng Yang, Jize Jiang +6

    cs.LGcs.AIarXiv:2505.19255v42025
  40. HDPO: Hybrid Distillation Policy Optimization via Privileged Self-Distillation

    Ken Ding

    cs.LGcs.AIarXiv:2603.23871v12026
  41. Iterative Amortized Inference

    Joseph Marino, Yisong Yue, Stephan Mandt

    cs.LGstat.MLarXiv:1807.09356v12018
  42. Effective Diversity in Population Based Reinforcement Learning

    Jack Parker-Holder, Aldo Pacchiano, Krzysztof Choromanski +1

    cs.LGstat.MLarXiv:2002.00632v32020
  43. DynaNDE: Dynamic Near-Data Expert Scheduling for Batched MoE Inference

    Xiaoyang Lu, Belthangady Akash Vi Narayana Pai, Xian-He Sun

    cs.ARcs.LGarXiv:2609.00407v12026
  44. Complete & Label: A Domain Adaptation Approach to Semantic Segmentation of LiDAR Point Clouds

    Li Yi, Boqing Gong, Thomas Funkhouser

    cs.CVcs.LGarXiv:2007.08488v22020
  45. Local Reference Geometry Residual Augmentation for Imbalanced Time Series Classification

    Chuanhang Qiu, Yanran Xu, Yue Wang +1

    cs.LGarXiv:2609.00093v12026
  46. POSEIDON: Privacy-Preserving Federated Neural Network Learning

    Sinem Sav, Apostolos Pyrgelis, Juan R. Troncoso-Pastoriza +4

    cs.CRcs.LGarXiv:2009.00349v32020
  47. Cartridges: Lightweight and general-purpose long context representations via self-study

    Sabri Eyuboglu, Ryan Ehrlich, Simran Arora +8

    cs.CLcs.AIcs.LGarXiv:2506.06266v32025
  48. The Invisible Leash: Why RLVR May or May Not Escape Its Origin

    Fang Wu, Weihao Xuan, Ximing Lu +4

    cs.LGcs.AIcs.CLarXiv:2507.14843v42025
  49. Linear Convergence in Federated Learning: Tackling Client Heterogeneity and Sparse Gradients

    Aritra Mitra, Rayana Jaafar, George J. Pappas +1

    cs.LGcs.DCeess.SYarXiv:2102.07053v22021
  50. Design Patterns for Securing LLM Agents against Prompt Injections

    Luca Beurer-Kellner, Beat Buesser, Ana-Maria Creţu +11

    cs.LGcs.CRarXiv:2506.08837v32025
  51. Towards Understanding Camera Motions in Any Video

    Zhiqiu Lin, Siyuan Cen, Daniel Jiang +12

    cs.CVcs.AIcs.CLarXiv:2504.15376v22025
  52. Urban Driver: Learning to Drive from Real-world Demonstrations Using Policy Gradients

    Oliver Scheel, Luca Bergamini, Maciej Wołczyk +2

    cs.ROcs.AIcs.CVarXiv:2109.13333v12021
  53. Efficient Online Reinforcement Learning for Diffusion Policy

    Haitong Ma, Tianyi Chen, Kai Wang +2

    cs.LGarXiv:2502.00361v42025
  54. Text2Reward: Reward Shaping with Language Models for Reinforcement Learning

    Tianbao Xie, Siheng Zhao, Chen Henry Wu +5

    cs.LGcs.AIcs.CLarXiv:2309.11489v32023
  55. A feature agnostic approach for glaucoma detection in OCT volumes

    Stefan Maetschke, Bhavna Antony, Hiroshi Ishikawa +3

    cs.CVcs.LGstat.MLarXiv:1807.04855v42018
  56. STEm-Seg: Spatio-temporal Embeddings for Instance Segmentation in Videos

    Ali Athar, Sabarinath Mahadevan, Aljoša Ošep +2

    cs.CVcs.LGeess.IVarXiv:2003.08429v42020
  57. Parametrized quantum policies for reinforcement learning

    Sofiene Jerbi, Casper Gyurik, Simon C. Marshall +2

    quant-phcs.AIcs.LGarXiv:2103.05577v22021
  58. MedRAX: Medical Reasoning Agent for Chest X-ray

    Adibvafa Fallahpour, Jun Ma, Alif Munim +2

    cs.LGcs.AIcs.MAarXiv:2502.02673v22025
  59. TimeFilter: Patch-Specific Spatial-Temporal Graph Filtration for Time Series Forecasting

    Yifan Hu, Guibin Zhang, Peiyuan Liu +6

    cs.LGarXiv:2501.13041v22025
  60. Character-Level Question Answering with Attention

    David Golub, Xiaodong He

    cs.CLcs.AIcs.LGarXiv:1604.00727v42016