Machine Learning
Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
16,741 to 16,800 of 20,217
A Critical Look at Targeted Instruction Selection: Disentangling What Matters (and What Doesn't)
Nihal V. Nayak, Paula Rodriguez-Diaz, Neha Hulkund +2
cs.LGarXiv:2602.14696v22026VLMo: Unified Vision-Language Pre-Training with Mixture-of-Modality-Experts
Hangbo Bao, Wenhui Wang, Li Dong +5
cs.CVcs.CLcs.LGarXiv:2111.02358v22021Meta Pseudo Labels
Hieu Pham, Zihang Dai, Qizhe Xie +2
cs.LGstat.MLarXiv:2003.10580v42020Tackling the Generative Learning Trilemma with Denoising Diffusion GANs
Zhisheng Xiao, Karsten Kreis, Arash Vahdat
cs.LGstat.MLarXiv:2112.07804v22021Assessing Generative Models via Precision and Recall
Mehdi S. M. Sajjadi, Olivier Bachem, Mario Lucic +2
stat.MLcs.LGarXiv:1806.00035v22018Uncertainty-Aware Vision-Language Segmentation for Medical Imaging
Aryan Das, Tanishq Rachamalla, Koushik Biswas +2
cs.CVcs.LGarXiv:2602.14498v22026Decoding as Optimisation on the Probability Simplex: From Top-K to Top-P (Nucleus) to Best-of-K Samplers
Xiaotong Ji, Rasul Tutunov, Matthieu Zimmer +1
cs.LGcs.AIarXiv:2602.18292v22026Unsupervised and Semi-supervised Learning with Categorical Generative Adversarial Networks
Jost Tobias Springenberg
stat.MLcs.LGarXiv:1511.06390v22015A Downsampled Variant of ImageNet as an Alternative to the CIFAR datasets
Patryk Chrabaszcz, Ilya Loshchilov, Frank Hutter
cs.CVcs.LGarXiv:1707.08819v32017Communication-Efficient On-Device Machine Learning: Federated Distillation and Augmentation under Non-IID Private Data
Eunjeong Jeong, Seungeun Oh, Hyesung Kim +3
cs.LGcs.NIstat.MLarXiv:1811.11479v22018Grid Search, Random Search, Genetic Algorithm: A Big Comparison for NAS
Petro Liashchynskyi, Pavlo Liashchynskyi
cs.LGcs.NEstat.MLarXiv:1912.06059v12019Crop Yield Prediction Using Deep Neural Networks
Saeed Khaki, Lizhi Wang
cs.LGstat.APstat.MLarXiv:1902.02860v32019Picking Winning Tickets Before Training by Preserving Gradient Flow
Chaoqi Wang, Guodong Zhang, Roger Grosse
cs.LGcs.CVstat.MLarXiv:2002.07376v22020Large-Scale Study of Curiosity-Driven Learning
Yuri Burda, Harri Edwards, Deepak Pathak +3
cs.LGcs.AIcs.CVarXiv:1808.04355v12018GRAM: Graph-based Attention Model for Healthcare Representation Learning
Edward Choi, Mohammad Taha Bahadori, Le Song +2
cs.LGstat.MLarXiv:1611.07012v32016Efficient Continual Learning in Language Models via Thalamically Routed Cortical Columns
Afshin Khadangi
cs.LGarXiv:2602.22479v62026BenthicDINO: Physics-Informed Self-Distillation for View-Invariant Side-Scan Sonar Representations
Taqi Hamoda, Hayat Rajani, Nuno Gracias
cs.CVcs.AIcs.LGarXiv:2608.23215v12026Whisper-RIR-Mega: A Paired Clean-Reverberant Speech Benchmark for ASR Robustness to Room Acoustics
Mandip Goswami
eess.AScs.AIcs.LGarXiv:2603.02252v32026Operator Learning Using Weak Supervision from Walk-on-Spheres
Hrishikesh Viswanath, Hong Chul Nam, Xi Deng +3
cs.LGarXiv:2603.01193v22026ZeroQuant: Efficient and Affordable Post-Training Quantization for Large-Scale Transformers
Zhewei Yao, Reza Yazdani Aminabadi, Minjia Zhang +3
cs.CLcs.LGarXiv:2206.01861v12022Simple and Controllable Music Generation
Jade Copet, Felix Kreuk, Itai Gat +5
cs.SDcs.AIcs.LGarXiv:2306.05284v32023Self-Sovereign Agent
Wenjie Qu, Xuandong Zhao, Jiaheng Zhang +1
cs.CRcs.CYcs.LGarXiv:2604.08551v12026Distribution-Conditioned Transport
Nic Fishman, Gokul Gowri, Paolo L. B. Fischer +3
cs.LGarXiv:2603.04736v12026TailSieve: Partial-Rollout-Guided Tail Routing for LLM Rollouts
Tianqi Xu, Lu Lv, Haoyang Huang +15
cs.AIcs.LGarXiv:2608.22788v12026A comprehensive study of non-adaptive and residual-based adaptive sampling for physics-informed neural networks
Chenxi Wu, Min Zhu, Qinyang Tan +2
physics.comp-phcs.LGarXiv:2207.10289v12022KARL: Knowledge Agents via Reinforcement Learning
Jonathan D. Chang, Andrew Drozdov, Shubham Toshniwal +23
cs.AIcs.LGarXiv:2603.05218v12026A Theoretical Analysis of Deep Q-Learning
Jianqing Fan, Zhaoran Wang, Yuchen Xie +1
cs.LGmath.OCstat.MLarXiv:1901.00137v32019Multi-talker Speech Separation with Utterance-level Permutation Invariant Training of Deep Recurrent Neural Networks
Morten Kolbæk, Dong Yu, Zheng-Hua Tan +1
cs.SDcs.LGeess.ASarXiv:1703.06284v22017Reasoning as Compression: Unifying Budget Forcing via the Conditional Information Bottleneck
Fabio Valerio Massoli, Andrey Kuzmin, Arash Behboodi
cs.LGarXiv:2603.08462v22026Federated learning with hierarchical clustering of local updates to improve training on non-IID data
Christopher Briggs, Zhong Fan, Peter Andras
cs.LGstat.MLarXiv:2004.11791v22020Hard Negative Mixing for Contrastive Learning
Yannis Kalantidis, Mert Bulent Sariyildiz, Noe Pion +2
cs.CVcs.LGarXiv:2010.01028v22020OSMDA: OpenStreetMap-based Domain Adaptation for Remote Sensing VLMs
Stefan Maria Ailuro, Mario Markov, Mohammad Mahdi +3
cs.CVcs.LGarXiv:2603.11804v32026Deep Gaussian Embedding of Graphs: Unsupervised Inductive Learning via Ranking
Aleksandar Bojchevski, Stephan Günnemann
stat.MLcs.LGcs.SIarXiv:1707.03815v42017Censored LLMs as a Natural Testbed for Secret Knowledge Elicitation
Helena Casademunt, Bartosz Cywiński, Khoi Tran +3
cs.LGcs.AIcs.CLarXiv:2603.05494v22026Variational Flow Maps: Make Some Noise for One-Step Conditional Generation
Abbas Mammadov, So Takao, Bohan Chen +4
cs.CVcs.LGstat.MLarXiv:2603.07276v12026Deep Ensembles: A Loss Landscape Perspective
Stanislav Fort, Huiyi Hu, Balaji Lakshminarayanan
stat.MLcs.LGarXiv:1912.02757v22019Learning to Optimize Via Posterior Sampling
Daniel Russo, Benjamin Van Roy
cs.LGarXiv:1301.2609v52013Low Data Drug Discovery with One-shot Learning
Han Altae-Tran, Bharath Ramsundar, Aneesh S. Pappu +1
cs.LGstat.MLarXiv:1611.03199v12016NaviDriveVLM: Decoupling High-Level Reasoning and Motion Planning for Autonomous Driving
Ximeng Tao, Pardis Taghavi, Dimitar Filev +2
cs.ROcs.LGarXiv:2603.07901v12026ConvergeFlow: Language Flow with Provable Convergence to Token Embeddings
Na Li, Yuchen Jiao, Changxiao Cai +1
cs.CLcs.AIcs.LGarXiv:2608.23551v12026$V_{0.5}$: Generalist Value Model as a Prior for Sparse RL Rollouts
Yi-Kai Zhang, Yueqing Sun, Hongyan Hao +4
cs.LGcs.AIcs.CLarXiv:2603.10848v12026LUKE: Deep Contextualized Entity Representations with Entity-aware Self-attention
Ikuya Yamada, Akari Asai, Hiroyuki Shindo +2
cs.CLcs.LGarXiv:2010.01057v12020VideoGPT: Video Generation using VQ-VAE and Transformers
Wilson Yan, Yunzhi Zhang, Pieter Abbeel +1
cs.CVcs.LGarXiv:2104.10157v22021Asynchronous Federated Optimization
Cong Xie, Sanmi Koyejo, Indranil Gupta
cs.DCcs.LGarXiv:1903.03934v52019A Comprehensive Survey on Graph Anomaly Detection with Deep Learning
Xiaoxiao Ma, Jia Wu, Shan Xue +5
cs.LGarXiv:2106.07178v52021Quantifying Generalization in Reinforcement Learning
Karl Cobbe, Oleg Klimov, Chris Hesse +2
cs.LGstat.MLarXiv:1812.02341v32018GCA: Global Centroid Alignment in Federated Learning
Jong-Ik Park, Harry Jiang, Logan Blakely +3
cs.LGcs.DCarXiv:2608.22593v12026Video Summarization with Long Short-term Memory
Ke Zhang, Wei-Lun Chao, Fei Sha +1
cs.CVcs.LGarXiv:1605.08110v22016Power-Performance Characterization of TinyML Systems
Yujie Zhang, Dhananjaya Wijerathne, Zhaoying Li +1
cs.LGcs.AIarXiv:2608.21646v12026Efficient Attention: Attention with Linear Complexities
Zhuoran Shen, Mingyuan Zhang, Haiyu Zhao +2
cs.CVcs.AIcs.LGarXiv:1812.01243v102018Using cognitive psychology to understand GPT-3
Marcel Binz, Eric Schulz
cs.CLcs.AIcs.LGarXiv:2206.14576v12022ESPIRE: A Diagnostic Benchmark for Embodied Spatial Reasoning of Vision-Language Models
Yanpeng Zhao, Wentao Ding, Hongtao Li +2
cs.CVcs.LGcs.ROarXiv:2603.13033v12026Harmonic Networks: Deep Translation and Rotation Equivariance
Daniel E. Worrall, Stephan J. Garbin, Daniyar Turmukhambetov +1
cs.CVcs.LGstat.MLarXiv:1612.04642v22016A General and Adaptive Robust Loss Function
Jonathan T. Barron
cs.CVcs.LGstat.MLarXiv:1701.03077v102017Open-Vocabulary Semantic Segmentation with Mask-adapted CLIP
Feng Liang, Bichen Wu, Xiaoliang Dai +6
cs.CVcs.LGarXiv:2210.04150v32022EESEN: End-to-End Speech Recognition using Deep RNN Models and WFST-based Decoding
Yajie Miao, Mohammad Gowayyed, Florian Metze
cs.CLcs.LGarXiv:1507.08240v32015FlashSampling: Fast and Memory-Efficient Exact Sampling
Tomas Ruiz, Zhen Qin, Yifan Zhang +3
cs.LGcs.AIcs.CLarXiv:2603.15854v22026Residual Stream Duality in Modern Transformer Architectures
Yifan Zhang
cs.LGcs.AIcs.CLarXiv:2603.16039v22026What AstroPT knows about galaxies, and what that can teach us about LLMs
UniverseTBD, :, Kshitij Duraphe +3
cs.LGastro-ph.IMarXiv:2608.22614v12026Generalized Discrete Diffusion from Snapshots
Oussama Zekri, Théo Uscidda, Nicolas Boullé +1
stat.MLcs.AIcs.CLarXiv:2603.21342v12026