Machine Learning
Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
17,041 to 17,100 of 20,215
Don't Repeat Yourself: Stopping Verbatim Loops at Sampling Time
Philipp Emanuel Weidmann, Allen Roush, Judah Goldfeder +2
cs.CLcs.AIcs.LGarXiv:2608.22761v12026A Commutator Framework for Selective Spectral Alignment in Deep Neural Networks
Kaj Nyström
stat.MLcs.LGarXiv:2608.22910v12026Parameterized Explainer for Graph Neural Network
Dongsheng Luo, Wei Cheng, Dongkuan Xu +4
cs.LGcs.AIarXiv:2011.04573v12020Beyond chlorophyll: machine learning estimates of diagnostic phytoplankton pigments from multispectral ocean colour data
David Moffat, Angus Laurenson, Victor Martinez-Vicente +4
q-bio.OTcs.LGphysics.opticsarXiv:2608.23348v12026Thinking at the Right Size: Amortized Distillation Across Post-Trained LLMs
Yan Zhou, Sara Kangaslahti, Jonathan Geuter +4
cs.LGarXiv:2608.22854v12026Least-Loaded Expert Parallelism: Load Balancing An Imbalanced Mixture-of-Experts
Xuan-Phi Nguyen, Shrey Pandit, Austin Xu +2
cs.LGcs.AIarXiv:2601.17111v12026Conditional Positional Encodings for Vision Transformers
Xiangxiang Chu, Zhi Tian, Bo Zhang +2
cs.CVcs.AIcs.LGarXiv:2102.10882v32021VLM-SubtleBench: How Far Are VLMs from Human-Level Subtle Comparative Reasoning?
Minkyu Kim, Sangheon Lee, Dongmin Park
cs.CVcs.AIcs.LGarXiv:2603.07888v12026Meta-Learning: A Survey
Joaquin Vanschoren
cs.LGstat.MLarXiv:1810.03548v12018XTC: Head-Aware Sampling by Excluding Top Choices
Philipp Emanuel Weidmann, Allen Roush, Judah Goldfeder +2
cs.CLcs.AIcs.LGarXiv:2608.22758v12026On Distillation of Guided Diffusion Models
Chenlin Meng, Robin Rombach, Ruiqi Gao +4
cs.CVcs.AIcs.LGarXiv:2210.03142v32022TERMINATOR: Learning Optimal Exit Points for Early Stopping in Chain-of-Thought Reasoning
Alliot Nagle, Jakhongir Saydaliev, Dhia Garbaya +3
cs.LGcs.AIcs.CLarXiv:2603.12529v22026The Web as a Knowledge-base for Answering Complex Questions
Alon Talmor, Jonathan Berant
cs.CLcs.AIcs.LGarXiv:1803.06643v12018PolyChirp: Multi-Species Birdsong Classification Using TinyML on Low-Power Acoustic Sensors
Nathan Duboisset, Zhaolan Huang, Felix Bießmann +3
cs.LGcs.AIarXiv:2608.23101v12026Algorithms for nonnegative matrix factorization with the beta-divergence
Cédric Févotte, Jérôme Idier
cs.LGarXiv:1010.1763v32010COVID-19 Image Data Collection: Prospective Predictions Are the Future
Joseph Paul Cohen, Paul Morrison, Lan Dao +3
q-bio.QMcs.CVcs.LGarXiv:2006.11988v32020Accelerating Very Deep Convolutional Networks for Classification and Detection
Xiangyu Zhang, Jianhua Zou, Kaiming He +1
cs.CVcs.LGcs.NEarXiv:1505.06798v22015Spatial-TTT: Streaming Visual-based Spatial Intelligence with Test-Time Training
Fangfu Liu, Diankun Wu, Jiawei Chi +7
cs.CVcs.LGarXiv:2603.12255v12026iDLG: Improved Deep Leakage from Gradients
Bo Zhao, Konda Reddy Mopuri, Hakan Bilen
cs.LGstat.MLarXiv:2001.02610v12020MultiPath: Multiple Probabilistic Anchor Trajectory Hypotheses for Behavior Prediction
Yuning Chai, Benjamin Sapp, Mayank Bansal +1
cs.LGcs.CVcs.ROarXiv:1910.05449v12019Why Steering Works: Toward a Unified View of Language Model Parameter Dynamics
Ziwen Xu, Chenyan Wu, Hengyu Sun +9
cs.CLcs.AIcs.CVarXiv:2602.02343v32026Time Series Data Augmentation for Deep Learning: A Survey
Qingsong Wen, Liang Sun, Fan Yang +4
cs.LGeess.SPstat.MLarXiv:2002.12478v42020hp-VPINNs: Variational Physics-Informed Neural Networks With Domain Decomposition
Ehsan Kharazmi, Zhongqiang Zhang, George Em Karniadakis
cs.NEcs.LGmath.NAarXiv:2003.05385v12020Revisiting Diffusion Model Predictions Through Dimensionality
Qing Jin, Chaoyang Wang
cs.LGcs.CVarXiv:2601.21419v22026Revisiting Distributed Synchronous SGD
Jianmin Chen, Xinghao Pan, Rajat Monga +2
cs.LGcs.DCcs.NEarXiv:1604.00981v32016CFG-Ctrl: Control-Based Classifier-Free Diffusion Guidance
Hanyang Wang, Yiyang Liu, Jiawei Chi +3
cs.CVcs.LGarXiv:2603.03281v22026Learning Representations for Counterfactual Inference
Fredrik D. Johansson, Uri Shalit, David Sontag
stat.MLcs.AIcs.LGarXiv:1605.03661v32016Rethinking LLM-as-a-Judge: Representation-as-a-Judge with Small Language Models via Semantic Capacity Asymmetry
Zhuochun Li, Yong Zhang, Ming Li +8
cs.CLcs.AIcs.LGarXiv:2601.22588v22026Improved Recurrent Neural Networks for Session-based Recommendations
Yong Kiam Tan, Xinxing Xu, Yong Liu
cs.LGarXiv:1606.08117v22016Conformal Risk Minimization for Semi-Supervised Domain Adaptation via Optimal Transport
Manos Giannopoulos, Yi Shen, Michael M. Zavlanos
cs.LGarXiv:2608.23153v12026SafePred: A Predictive Guardrail for Computer-Using Agents via World Models
Yurun Chen, Zeyi Liao, Ping Yin +3
cs.CLcs.AIcs.LGarXiv:2602.01725v12026HorizonMath: Measuring AI Progress Toward Mathematical Discovery with Automatic Verification
Erik Y. Wang, Sumeet Motwani, James V. Roggeveen +7
cs.LGarXiv:2603.15617v12026Variational Autoencoder for Deep Learning of Images, Labels and Captions
Yunchen Pu, Zhe Gan, Ricardo Henao +4
stat.MLcs.LGarXiv:1609.08976v12016Fast Transformer Decoding: One Write-Head is All You Need
Noam Shazeer
cs.NEcs.CLcs.LGarXiv:1911.02150v12019LatentChem: From Textual CoT to Latent Thinking in Chemical Reasoning
Xinwu Ye, Yicheng Mao, Yuxuan Liao +16
physics.chem-phcs.AIcs.CLarXiv:2602.07075v62026FlexGen: High-Throughput Generative Inference of Large Language Models with a Single GPU
Ying Sheng, Lianmin Zheng, Binhang Yuan +11
cs.LGcs.AIcs.PFarXiv:2303.06865v22023The Open Catalyst 2020 (OC20) Dataset and Community Challenges
Lowik Chanussot, Abhishek Das, Siddharth Goyal +14
cond-mat.mtrl-scics.LGarXiv:2010.09990v52020Characterizing Adversarial Subspaces Using Local Intrinsic Dimensionality
Xingjun Ma, Bo Li, Yisen Wang +6
cs.LGcs.CRcs.CVarXiv:1801.02613v32018Mitigating Bias in Algorithmic Hiring: Evaluating Claims and Practices
Manish Raghavan, Solon Barocas, Jon Kleinberg +1
cs.CYcs.AIcs.LGarXiv:1906.09208v32019Unveiling Implicit Advantage Symmetry: Why GRPO Struggles with Exploration and Difficulty Adaptation
Zhiqi Yu, Zhangquan Chen, Mengting Liu +2
cs.LGcs.AIarXiv:2602.05548v32026Role-Specialized Mixture-of-Agents with Open-Weight LLMs for Clinical Prediction
Jun Hou, Yi Fang, Xuan Wang
cs.AIcs.LGarXiv:2608.22176v12026Code-Space Response Oracles: Generating Interpretable Multi-Agent Policies with Large Language Models
Daniel Hennes, Zun Li, John Schultz +1
cs.GTcs.AIcs.LGarXiv:2603.10098v12026MCP-Universe RL: A Framework for Training MCP Tool-Use Agents via Reinforcement Learning
Ziyang Luo, Yan Yang, Xiangru Jian +5
cs.AIcs.LGarXiv:2608.22167v12026SPARKLING: Balancing Signal Preservation and Symmetry Breaking for Width-Progressive Learning
Qifan Yu, Xinyu Ma, Zhijian Zhuo +7
cs.LGcs.CLarXiv:2602.02472v22026Pruning neural networks without any data by iteratively conserving synaptic flow
Hidenori Tanaka, Daniel Kunin, Daniel L. K. Yamins +1
cs.LGcond-mat.dis-nncs.CVarXiv:2006.05467v32020Flow Equivariant World Models: Memory for Partially Observed Dynamic Environments
Hansen Jin Lillemark, Benhao Huang, Fangneng Zhan +2
cs.LGcs.AIcs.CVarXiv:2601.01075v22026Demystifying Reinforcement Learning for Long-Horizon Tool-Using Agents: A Comprehensive Recipe
Xixi Wu, Qianguo Sun, Ruiyang Zhang +4
cs.LGcs.CLarXiv:2603.21972v12026Complementary RL: Towards Efficient Experience-Driven Agent Learning
Dilxat Muhtar, Jiashun Liu, Wei Gao +8
cs.LGcs.CLarXiv:2603.17621v22026SimplE Embedding for Link Prediction in Knowledge Graphs
Seyed Mehran Kazemi, David Poole
stat.MLcs.LGarXiv:1802.04868v22018Noise2Self: Blind Denoising by Self-Supervision
Joshua Batson, Loic Royer
cs.CVcs.LGstat.MLarXiv:1901.11365v22019Mixture-of-Experts with Expert Choice Routing
Yanqi Zhou, Tao Lei, Hanxiao Liu +7
cs.LGcs.AIarXiv:2202.09368v22022NExT-GPT: Any-to-Any Multimodal LLM
Shengqiong Wu, Hao Fei, Leigang Qu +2
cs.AIcs.CLcs.LGarXiv:2309.05519v32023The many Shapley values for model explanation
Mukund Sundararajan, Amir Najmi
cs.AIcs.LGecon.THarXiv:1908.08474v22019The Flexibility Trap: Rethinking the Value of Arbitrary Order in Diffusion Language Models
Zanlin Ni, Shenzhi Wang, Yang Yue +8
cs.CLcs.AIcs.LGarXiv:2601.15165v42026Unbiased Scene Graph Generation from Biased Training
Kaihua Tang, Yulei Niu, Jianqiang Huang +2
cs.CVcs.LGarXiv:2002.11949v42020Solving math word problems with process- and outcome-based feedback
Jonathan Uesato, Nate Kushman, Ramana Kumar +6
cs.LGcs.AIcs.CLarXiv:2211.14275v12022Zero-Shot Relation Extraction via Reading Comprehension
Omer Levy, Minjoon Seo, Eunsol Choi +1
cs.CLcs.AIcs.LGarXiv:1706.04115v12017Counterfactual Reasoning and Learning Systems
Léon Bottou, Jonas Peters, Joaquin Quiñonero-Candela +6
cs.LGcs.AIcs.IRarXiv:1209.2355v52012The Design Space of Tri-Modal Masked Diffusion Models
Louis Bethune, Victor Turrisi, Bruno Kacper Mlodozeniec +21
cs.LGarXiv:2602.21472v12026S4L: Self-Supervised Semi-Supervised Learning
Xiaohua Zhai, Avital Oliver, Alexander Kolesnikov +1
cs.CVcs.LGarXiv:1905.03670v22019