Machine Learning
Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
16,501 to 16,560 of 20,199
Shared Nature, Unique Nurture: PRISM for Pluralistic Reasoning via In-context Structure Modeling
Guancheng Tu, Shiyang Zhang, Tianyu Zhang +2
cs.LGarXiv:2602.21317v12026Easy to Learn, Yet Hard to Forget: Towards Robust Unlearning Under Bias
JuneHyoung Kwon, MiHyeon Kim, Eunju Lee +3
cs.LGcs.CVarXiv:2602.21773v12026Token-Level Likelihood-Array Regression for Membership Inference and AI-Generated Text Detection
Jiajun Sun, Zhanrui Cai
stat.MLcs.LGstat.MEarXiv:2608.22179v12026Mobile ALOHA: Learning Bimanual Mobile Manipulation with Low-Cost Whole-Body Teleoperation
Zipeng Fu, Tony Z. Zhao, Chelsea Finn
cs.ROcs.AIcs.CVarXiv:2401.02117v12024Multi-Head Low-Rank Attention
Songtao Liu, Hongwu Peng, Zhiwei Zhang +2
cs.LGarXiv:2603.02188v12026Reward Constrained Policy Optimization
Chen Tessler, Daniel J. Mankowitz, Shie Mannor
cs.LGcs.AIstat.MLarXiv:1805.11074v32018Feature Learning based Deep Supervised Hashing with Pairwise Labels
Wu-Jun Li, Sheng Wang, Wang-Cheng Kang
cs.LGcs.CVarXiv:1511.03855v22015Modeling Temporal Dependencies in High-Dimensional Sequences: Application to Polyphonic Music Generation and Transcription
Nicolas Boulanger-Lewandowski, Yoshua Bengio, Pascal Vincent
cs.LGcs.SDstat.MLarXiv:1206.6392v12012Reducing Transformer Depth on Demand with Structured Dropout
Angela Fan, Edouard Grave, Armand Joulin
cs.LGcs.CLstat.MLarXiv:1909.11556v12019SANE: State Anomaly Neutralization for Stable Extreme-Context Delta-Rule Models
Qingwen Lin, Boyan Xu, Xiao Liu +2
cs.LGcs.AIarXiv:2608.22354v22026Recovered in Translation: Efficient Pipeline for Automated Translation of Benchmarks and Datasets
Hanna Yukhymenko, Anton Alexandrov, Martin Vechev
cs.CLcs.AIcs.LGarXiv:2602.22207v12026Learning from the Test: Self-Referential Differential Testing for Deep RL Agents
Junda He, Jieke Shi, Zhou Yang +2
cs.SEcs.AIcs.LGarXiv:2608.22284v12026veScale-FSDP: Flexible and High-Performance FSDP at Scale
Zezhou Wang, Youjie Li, Zhiqi Lin +9
cs.DCcs.AIcs.LGarXiv:2602.22437v32026Text-to-SQL Empowered by Large Language Models: A Benchmark Evaluation
Dawei Gao, Haibin Wang, Yaliang Li +4
cs.DBcs.CLcs.LGarXiv:2308.15363v42023Mollified Value Learning
Hrishikesh Viswanath, Juanwu Lu, S. Talha Bukhari +4
cs.LGcs.ROarXiv:2602.23280v22026DySCo: Dynamically consistent data-driven downscaling of extremes in climate projections
S. Stamatelopoulos, M. Wang, I. Lopez-Gomez +5
cs.LGmath.NAphysics.ao-pharXiv:2608.21998v12026Scalable Global Optimization via Local Bayesian Optimization
David Eriksson, Michael Pearce, Jacob R Gardner +2
cs.LGstat.MLarXiv:1910.01739v42019DynaMoE: Dynamic Token-Level Expert Activation with Layer-Wise Adaptive Capacity for Mixture-of-Experts Neural Networks
Gökdeniz Gülmez
cs.LGcs.AIarXiv:2603.01697v12026Don't Just Assume; Look and Answer: Overcoming Priors for Visual Question Answering
Aishwarya Agrawal, Dhruv Batra, Devi Parikh +1
cs.CVcs.AIcs.CLarXiv:1712.00377v22017Order-independent constraint-based causal structure learning
Diego Colombo, Marloes H. Maathuis
stat.MLcs.LGarXiv:1211.3295v22012Overcoming Catastrophic Forgetting by Incremental Moment Matching
Sang-Woo Lee, Jin-Hwa Kim, Jaehyun Jun +2
cs.LGcs.AIarXiv:1703.08475v32017Are Emergent Abilities of Large Language Models a Mirage?
Rylan Schaeffer, Brando Miranda, Sanmi Koyejo
cs.AIcs.LGarXiv:2304.15004v22023Scaling Agentic Capabilities, Not Context: Efficient Reinforcement Finetuning for Large Toolspaces
Karan Gupta, Pranav Vajreshwari, Yash Pandya +3
cs.LGcs.AIarXiv:2603.06713v12026A Probabilistic U-Net for Segmentation of Ambiguous Images
Simon A. A. Kohl, Bernardino Romera-Paredes, Clemens Meyer +6
cs.CVcs.LGcs.NEarXiv:1806.05034v42018Deep Learning for Computational Chemistry
Garrett B. Goh, Nathan O. Hodas, Abhinav Vishnu
stat.MLcs.AIcs.CEarXiv:1701.04503v12017SAHOO: Safeguarded Alignment for High-Order Optimization Objectives in Recursive Self-Improvement
Subramanyam Sahoo, Aman Chadha, Vinija Jain +1
cs.AIcs.CLcs.LGarXiv:2603.06333v12026Invertible Residual Networks
Jens Behrmann, Will Grathwohl, Ricky T. Q. Chen +2
cs.LGcs.AIcs.CVarXiv:1811.00995v32018Beyond Fresh Starts: Stateful Inference for Streaming ASR in Conversational Voice Agents
Sameep Chattopadhyay, Alexander Erdmann, Mari Ostendorf
cs.LGarXiv:2608.22101v12026ByteFlow: Language Modeling through Adaptive Byte Compression without a Tokenizer
Chunyuan Deng, Sanket Lokegaonkar, Colin Lockard +3
cs.CLcs.LGarXiv:2603.03583v12026Probabilistic Model-Agnostic Meta-Learning
Chelsea Finn, Kelvin Xu, Sergey Levine
cs.LGcs.AIstat.MLarXiv:1806.02817v22018Risk-Sensitive Reinforcement Learning with Smoothed Quantile Objectives
Mohammad Alipour-Vaezi, Huaiyang Zhong, Sajad Khodadadian
cs.LGmath.OCarXiv:2608.22227v12026A Little Is Enough: Circumventing Defenses For Distributed Learning
Moran Baruch, Gilad Baruch, Yoav Goldberg
cs.LGcs.CRcs.DCarXiv:1902.06156v12019Layer by layer, module by module: Choose both for optimal OOD probing of ViT
Ambroise Odonnat, Vasilii Feofanov, Laetitia Chapel +2
cs.CVcs.LGstat.MLarXiv:2603.05280v12026TensorFlow Lite Micro: Embedded Machine Learning on TinyML Systems
Robert David, Jared Duke, Advait Jain +10
cs.LGcs.AIarXiv:2010.08678v32020PixARMesh: Autoregressive Mesh-Native Single-View Scene Reconstruction
Xiang Zhang, Sohyun Yoo, Hongrui Wu +3
cs.CVcs.GRcs.LGarXiv:2603.05888v12026Deep Unsupervised Clustering with Gaussian Mixture Variational Autoencoders
Nat Dilokthanakul, Pedro A. M. Mediano, Marta Garnelo +4
cs.LGcs.NEstat.MLarXiv:1611.02648v22016DC-DiT: Adaptive Compute and Elastic Inference for Visual Generation via Dynamic Chunking
Akash Haridas, Utkarsh Saxena, Parsa Ashrafi Fashi +3
cs.CVcs.AIcs.LGarXiv:2603.06351v22026Concept-Guided Fine-Tuning: Steering ViTs away from Spurious Correlations to Improve Robustness
Yehonatan Elisha, Oren Barkan, Noam Koenigstein
cs.CVcs.AIcs.LGarXiv:2603.08309v22026Kernel-based Conditional Independence Test and Application in Causal Discovery
Kun Zhang, Jonas Peters, Dominik Janzing +1
cs.LGstat.MLarXiv:1202.3775v12012Diffusion Models for Adversarial Purification
Weili Nie, Brandon Guo, Yujia Huang +3
cs.LGcs.CRcs.CVarXiv:2205.07460v12022GTA-RAG: Graph-Trajectory-Augmented Reinforcement Learning for Multi-Turn Retrieval-Augmented Reasoning
Jun Chen, Yongchao Liu, Pengyu Qiu +6
cs.CLcs.LGarXiv:2608.22479v12026Compiler-First State Space Duality and Portable $O(1)$ Autoregressive Caching for Inference
Cosmo Santoni, Anmol Thapar
cs.LGcs.AIcs.DCarXiv:2603.09555v22026Recursive Language Models Meet Uncertainty: The Surprising Effectiveness of Self-Reflective Program Search for Long Context
Keivan Alizadeh, Parshin Shojaee, Minsik Cho +1
cs.CLcs.AIcs.LGarXiv:2603.15653v12026Neural Text Generation with Unlikelihood Training
Sean Welleck, Ilia Kulikov, Stephen Roller +3
cs.LGcs.CLstat.MLarXiv:1908.04319v22019Unlocking Data Value in Finance: A Study on Distillation and Difficulty-Aware Training
Chuxue Cao, Honglin Lin, Zhanping Zhong +5
cs.LGarXiv:2603.07223v12026A Comparative analysis of Layer-wise Representational Capacity in AR and Diffusion LLMs
Raghavv Goel, Risheek Garrepalli, Sudhanshu Agrawal +3
cs.CLcs.LGarXiv:2603.07475v42026Gaussian process learning with flow map refinement for parameter estimation in dynamical systems
Yue Hao, Dongwei Ye
cs.LGcs.CEmath.NAarXiv:2608.22324v12026TasNet: time-domain audio separation network for real-time, single-channel speech separation
Yi Luo, Nima Mesgarani
cs.SDcs.LGcs.MMarXiv:1711.00541v22017MusicLM: Generating Music From Text
Andrea Agostinelli, Timo I. Denk, Zalán Borsos +10
cs.SDcs.LGeess.ASarXiv:2301.11325v12023Doubly Robust Off-policy Value Evaluation for Reinforcement Learning
Nan Jiang, Lihong Li
cs.LGcs.AIeess.SYarXiv:1511.03722v32015A Survey and Critique of Multiagent Deep Reinforcement Learning
Pablo Hernandez-Leal, Bilal Kartal, Matthew E. Taylor
cs.MAcs.AIcs.LGarXiv:1810.05587v32018BiCLIP: Domain Canonicalization via Structured Geometric Transformation
Pranav Mantini, Shishir K. Shah
cs.CVcs.AIcs.CLarXiv:2603.08942v22026The Reasoning Trap -- Logical Reasoning as a Mechanistic Pathway to Situational Awareness
Subramanyam Sahoo, Aman Chadha, Vinija Jain +1
cs.AIcs.CLcs.CYarXiv:2603.09200v12026Beyond Test-Time Memory: State-Space Optimal Control for LLM Reasoning
Peihao Wang, Shan Yang, Xijun Wang +8
cs.LGarXiv:2603.09221v22026Argmax Flows and Multinomial Diffusion: Learning Categorical Distributions
Emiel Hoogeboom, Didrik Nielsen, Priyank Jaini +2
stat.MLcs.CLcs.LGarXiv:2102.05379v32021Progressive Differentiable Architecture Search: Bridging the Depth Gap between Search and Evaluation
Xin Chen, Lingxi Xie, Jun Wu +1
cs.CVcs.LGarXiv:1904.12760v12019Towards a Neural Debugger for Python
Maximilian Beck, Jonas Gehring, Jannik Kossen +1
cs.LGcs.AIcs.SEarXiv:2603.09951v12026MLPerf Inference Benchmark
Vijay Janapa Reddi, Christine Cheng, David Kanter +44
cs.LGcs.PFstat.MLarXiv:1911.02549v22019Self-Execution Simulation Improves Coding Models
Gallil Maimon, Ori Yoran, Felix Kreuk +4
cs.CLcs.LGarXiv:2604.03253v12026Social-BiGAT: Multimodal Trajectory Forecasting using Bicycle-GAN and Graph Attention Networks
Vineet Kosaraju, Amir Sadeghian, Roberto Martín-Martín +3
cs.CVcs.LGarXiv:1907.03395v22019