Machine Learning
Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
5,701 to 5,760 of 20,214
Semantic-Guided Multimodal Preprocessing for Vision Transformer-Based Clear Cell Renal Cell Carcinoma Grading
Fatemeh Javadian, Zhu Chen, Zahra Aminparast +1
cs.CVcs.AIcs.LGarXiv:2609.01426v12026Measuring consistency via ensemble margin and local prediction variability: Auditing decision systems in the presence of predictive multiplicity
Sinjini Banerjee, Tim Marrinan, Anand D. Sarwate
stat.MLcs.AIcs.LGarXiv:2609.01397v12026One-Prompt-One-Story: Free-Lunch Consistent Text-to-Image Generation Using a Single Prompt
Tao Liu, Kai Wang, Senmao Li +6
cs.CVcs.AIcs.LGarXiv:2501.13554v32025Spatial Broadcast Decoder: A Simple Architecture for Learning Disentangled Representations in VAEs
Nicholas Watters, Loic Matthey, Christopher P. Burgess +1
cs.LGcs.CVstat.MLarXiv:1901.07017v22019DualStake: Dual-Path Confidence Calibration in Deep Research Agents
Yinuo Xu, Yuwei Liang, Jianjie Cheng +4
cs.CLcs.AIcs.LGarXiv:2609.00935v12026A Wholistic View of Continual Learning with Deep Neural Networks: Forgotten Lessons and the Bridge to Active and Open World Learning
Martin Mundt, Yongwon Hong, Iuliia Pliushch +1
cs.LGstat.MLarXiv:2009.01797v32020Confess What You Know: Forget-Set Misalignment with Model Knowledge in LLM Unlearning
Miso Kim, Georu Lee, Seungwon Jeong +1
cs.LGcs.AIcs.CLarXiv:2609.00605v12026A Study of Hidden-State Optimization Order in Predictive Coding Networks
Xueyuan Li, Danilo Vasconcellos Vargas
cs.LGcs.AIarXiv:2609.00686v12026Fast, Exact and Multi-Scale Inference for Semantic Image Segmentation with Deep Gaussian CRFs
Siddhartha Chandra, Iasonas Kokkinos
cs.CVcs.LGarXiv:1603.08358v42016Universal Model Routing for Efficient LLM Inference
Wittawat Jitkrittum, Harikrishna Narasimhan, Ankit Singh Rawat +9
cs.CLcs.LGarXiv:2502.08773v22025ViTAMINS: An Empirical Study of Training Self-Supervised Vision Transformers with Synthetic Hard Negatives
Nikos Giakoumoglou, Andreas Floros, Kleanthis-Marios Papadopoulos +1
cs.CVcs.AIcs.LGarXiv:2609.01041v12026Sparse Autoencoders Do Not Find Canonical Units of Analysis
Patrick Leask, Bart Bussmann, Michael Pearce +5
cs.LGcs.AIarXiv:2502.04878v12025Momentum Contrastive Learning for Few-Shot COVID-19 Diagnosis from Chest CT Images
Xiaocong Chen, Lina Yao, Tao Zhou +2
eess.IVcs.CVcs.LGarXiv:2006.13276v12020SMART: Self-Aware Agent for Tool Overuse Mitigation
Cheng Qian, Emre Can Acikgoz, Hongru Wang +5
cs.AIcs.CLcs.LGarXiv:2502.11435v22025Learning Generalizable Robotic Reward Functions from "In-The-Wild" Human Videos
Annie S. Chen, Suraj Nair, Chelsea Finn
cs.ROcs.AIcs.CVarXiv:2103.16817v12021On the Doubt about Margin Explanation of Boosting
Wei Gao, Zhi-Hua Zhou
cs.LGarXiv:1009.3613v52010Embedded Conditional Independence Tests for Large Language Model Generated Text with an Application to German Parliament Speeches
Marco Simnacher, Georg Keilbar, Benjamin König +2
stat.MLcs.AIcs.LGarXiv:2609.00946v12026TimeNet: Pre-trained deep recurrent neural network for time series classification
Pankaj Malhotra, Vishnu TV, Lovekesh Vig +2
cs.LGarXiv:1706.08838v12017Learning the Travelling Salesperson Problem Requires Rethinking Generalization
Chaitanya K. Joshi, Quentin Cappart, Louis-Martin Rousseau +1
cs.LGstat.MLarXiv:2006.07054v62020GPTAQ: Efficient Finetuning-Free Quantization for Asymmetric Calibration
Yuhang Li, Ruokai Yin, Donghyun Lee +2
cs.LGarXiv:2504.02692v32025FlexiViT: One Model for All Patch Sizes
Lucas Beyer, Pavel Izmailov, Alexander Kolesnikov +7
cs.CVcs.AIcs.LGarXiv:2212.08013v22022Inference-Time Alignment in Diffusion Models with Reward-Guided Generation: Tutorial and Review
Masatoshi Uehara, Yulai Zhao, Chenyu Wang +4
cs.AIcs.LGq-bio.QMarXiv:2501.09685v22025Learning-Theoretic Foundation for General Coded Computing: The Straggler Setting
Parsa Moradi, Behrooz Tahmasebi, Mohammad Ali Maddah-Ali
cs.LGarXiv:2608.28910v12026FaSNet: Low-latency Adaptive Beamforming for Multi-microphone Audio Processing
Yi Luo, Enea Ceolini, Cong Han +2
eess.AScs.LGcs.SDarXiv:1909.13387v22019Reconstructing Training Data from Trained Neural Networks
Niv Haim, Gal Vardi, Gilad Yehudai +2
cs.LGcs.CRcs.CVarXiv:2206.07758v32022Not All Language Model Features Are One-Dimensionally Linear
Joshua Engels, Eric J. Michaud, Isaac Liao +2
cs.LGarXiv:2405.14860v32024One Token to Fool LLM-as-a-Judge
Yulai Zhao, Haolin Liu, Dian Yu +4
cs.LGcs.CLarXiv:2507.08794v32025HelpSteer3-Preference: Open Human-Annotated Preference Data across Diverse Tasks and Languages
Zhilin Wang, Jiaqi Zeng, Olivier Delalleau +6
cs.CLcs.AIcs.LGarXiv:2505.11475v22025Convergence issues in Relational Concept Analysis based on AOC-posets
Xavier Dolques, Agnès Braud, Alain Gutierrez +2
cs.LGarXiv:2609.00054v12026AdaComp : Adaptive Residual Gradient Compression for Data-Parallel Distributed Training
Chia-Yu Chen, Jungwook Choi, Daniel Brand +3
cs.LGstat.MLarXiv:1712.02679v12017GSPMD: General and Scalable Parallelization for ML Computation Graphs
Yuanzhong Xu, HyoukJoong Lee, Dehao Chen +13
cs.DCcs.LGarXiv:2105.04663v22021Sample Complexity Bounds for Stochastic Shortest Path with a Generative Model
Jean Tarbouriech, Matteo Pirotta, Michal Valko +1
cs.LGstat.MLarXiv:2604.16111v12026Dion: Distributed Orthonormalized Updates
Kwangjun Ahn, Byron Xu, Natalie Abreu +5
cs.LGcs.AImath.OCarXiv:2504.05295v32025A Study of Reinforcement Learning for Neural Machine Translation
Lijun Wu, Fei Tian, Tao Qin +2
cs.LGcs.AIstat.MLarXiv:1808.08866v12018An Attention Free Transformer
Shuangfei Zhai, Walter Talbott, Nitish Srivastava +4
cs.LGcs.CLcs.CVarXiv:2105.14103v22021Multi-Scale Contrastive Siamese Networks for Self-Supervised Graph Representation Learning
Ming Jin, Yizhen Zheng, Yuan-Fang Li +3
cs.LGcs.SIarXiv:2105.05682v22021DNA-inspired online behavioral modeling and its application to spambot detection
Stefano Cresci, Roberto Di Pietro, Marinella Petrocchi +2
cs.SIcs.CRcs.LGarXiv:1602.00110v12016Self-Reports Are Not Verification: Environment-Grounded Auditing of LLM Operators in Evolutionary Search
Enrong Pan, Ryan Zhou, Ting Hu
cs.AIcs.LGcs.NEarXiv:2609.00652v12026The challenge of realistic music generation: modelling raw audio at scale
Sander Dieleman, Aäron van den Oord, Karen Simonyan
cs.SDcs.LGeess.ASarXiv:1806.10474v12018VTool-R1: VLMs Learn to Think with Images via Reinforcement Learning on Multimodal Tool Use
Mingyuan Wu, Jingcheng Yang, Jize Jiang +6
cs.LGcs.AIarXiv:2505.19255v42025HDPO: Hybrid Distillation Policy Optimization via Privileged Self-Distillation
Ken Ding
cs.LGcs.AIarXiv:2603.23871v12026Iterative Amortized Inference
Joseph Marino, Yisong Yue, Stephan Mandt
cs.LGstat.MLarXiv:1807.09356v12018Effective Diversity in Population Based Reinforcement Learning
Jack Parker-Holder, Aldo Pacchiano, Krzysztof Choromanski +1
cs.LGstat.MLarXiv:2002.00632v32020DynaNDE: Dynamic Near-Data Expert Scheduling for Batched MoE Inference
Xiaoyang Lu, Belthangady Akash Vi Narayana Pai, Xian-He Sun
cs.ARcs.LGarXiv:2609.00407v12026Complete & Label: A Domain Adaptation Approach to Semantic Segmentation of LiDAR Point Clouds
Li Yi, Boqing Gong, Thomas Funkhouser
cs.CVcs.LGarXiv:2007.08488v22020Local Reference Geometry Residual Augmentation for Imbalanced Time Series Classification
Chuanhang Qiu, Yanran Xu, Yue Wang +1
cs.LGarXiv:2609.00093v12026POSEIDON: Privacy-Preserving Federated Neural Network Learning
Sinem Sav, Apostolos Pyrgelis, Juan R. Troncoso-Pastoriza +4
cs.CRcs.LGarXiv:2009.00349v32020Cartridges: Lightweight and general-purpose long context representations via self-study
Sabri Eyuboglu, Ryan Ehrlich, Simran Arora +8
cs.CLcs.AIcs.LGarXiv:2506.06266v32025The Invisible Leash: Why RLVR May or May Not Escape Its Origin
Fang Wu, Weihao Xuan, Ximing Lu +4
cs.LGcs.AIcs.CLarXiv:2507.14843v42025Linear Convergence in Federated Learning: Tackling Client Heterogeneity and Sparse Gradients
Aritra Mitra, Rayana Jaafar, George J. Pappas +1
cs.LGcs.DCeess.SYarXiv:2102.07053v22021Design Patterns for Securing LLM Agents against Prompt Injections
Luca Beurer-Kellner, Beat Buesser, Ana-Maria Creţu +11
cs.LGcs.CRarXiv:2506.08837v32025Towards Understanding Camera Motions in Any Video
Zhiqiu Lin, Siyuan Cen, Daniel Jiang +12
cs.CVcs.AIcs.CLarXiv:2504.15376v22025Urban Driver: Learning to Drive from Real-world Demonstrations Using Policy Gradients
Oliver Scheel, Luca Bergamini, Maciej Wołczyk +2
cs.ROcs.AIcs.CVarXiv:2109.13333v12021Efficient Online Reinforcement Learning for Diffusion Policy
Haitong Ma, Tianyi Chen, Kai Wang +2
cs.LGarXiv:2502.00361v42025Text2Reward: Reward Shaping with Language Models for Reinforcement Learning
Tianbao Xie, Siheng Zhao, Chen Henry Wu +5
cs.LGcs.AIcs.CLarXiv:2309.11489v32023A feature agnostic approach for glaucoma detection in OCT volumes
Stefan Maetschke, Bhavna Antony, Hiroshi Ishikawa +3
cs.CVcs.LGstat.MLarXiv:1807.04855v42018STEm-Seg: Spatio-temporal Embeddings for Instance Segmentation in Videos
Ali Athar, Sabarinath Mahadevan, Aljoša Ošep +2
cs.CVcs.LGeess.IVarXiv:2003.08429v42020Parametrized quantum policies for reinforcement learning
Sofiene Jerbi, Casper Gyurik, Simon C. Marshall +2
quant-phcs.AIcs.LGarXiv:2103.05577v22021MedRAX: Medical Reasoning Agent for Chest X-ray
Adibvafa Fallahpour, Jun Ma, Alif Munim +2
cs.LGcs.AIcs.MAarXiv:2502.02673v22025TimeFilter: Patch-Specific Spatial-Temporal Graph Filtration for Time Series Forecasting
Yifan Hu, Guibin Zhang, Peiyuan Liu +6
cs.LGarXiv:2501.13041v22025