Machine Learning

Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

16,741 to 16,800 of 20,217

  1. A Critical Look at Targeted Instruction Selection: Disentangling What Matters (and What Doesn't)

    Nihal V. Nayak, Paula Rodriguez-Diaz, Neha Hulkund +2

    cs.LGarXiv:2602.14696v22026
  2. VLMo: Unified Vision-Language Pre-Training with Mixture-of-Modality-Experts

    Hangbo Bao, Wenhui Wang, Li Dong +5

    cs.CVcs.CLcs.LGarXiv:2111.02358v22021
  3. Meta Pseudo Labels

    Hieu Pham, Zihang Dai, Qizhe Xie +2

    cs.LGstat.MLarXiv:2003.10580v42020
  4. Tackling the Generative Learning Trilemma with Denoising Diffusion GANs

    Zhisheng Xiao, Karsten Kreis, Arash Vahdat

    cs.LGstat.MLarXiv:2112.07804v22021
  5. Assessing Generative Models via Precision and Recall

    Mehdi S. M. Sajjadi, Olivier Bachem, Mario Lucic +2

    stat.MLcs.LGarXiv:1806.00035v22018
  6. Uncertainty-Aware Vision-Language Segmentation for Medical Imaging

    Aryan Das, Tanishq Rachamalla, Koushik Biswas +2

    cs.CVcs.LGarXiv:2602.14498v22026
  7. Decoding as Optimisation on the Probability Simplex: From Top-K to Top-P (Nucleus) to Best-of-K Samplers

    Xiaotong Ji, Rasul Tutunov, Matthieu Zimmer +1

    cs.LGcs.AIarXiv:2602.18292v22026
  8. Unsupervised and Semi-supervised Learning with Categorical Generative Adversarial Networks

    Jost Tobias Springenberg

    stat.MLcs.LGarXiv:1511.06390v22015
  9. A Downsampled Variant of ImageNet as an Alternative to the CIFAR datasets

    Patryk Chrabaszcz, Ilya Loshchilov, Frank Hutter

    cs.CVcs.LGarXiv:1707.08819v32017
  10. Communication-Efficient On-Device Machine Learning: Federated Distillation and Augmentation under Non-IID Private Data

    Eunjeong Jeong, Seungeun Oh, Hyesung Kim +3

    cs.LGcs.NIstat.MLarXiv:1811.11479v22018
  11. Grid Search, Random Search, Genetic Algorithm: A Big Comparison for NAS

    Petro Liashchynskyi, Pavlo Liashchynskyi

    cs.LGcs.NEstat.MLarXiv:1912.06059v12019
  12. Crop Yield Prediction Using Deep Neural Networks

    Saeed Khaki, Lizhi Wang

    cs.LGstat.APstat.MLarXiv:1902.02860v32019
  13. Picking Winning Tickets Before Training by Preserving Gradient Flow

    Chaoqi Wang, Guodong Zhang, Roger Grosse

    cs.LGcs.CVstat.MLarXiv:2002.07376v22020
  14. Large-Scale Study of Curiosity-Driven Learning

    Yuri Burda, Harri Edwards, Deepak Pathak +3

    cs.LGcs.AIcs.CVarXiv:1808.04355v12018
  15. GRAM: Graph-based Attention Model for Healthcare Representation Learning

    Edward Choi, Mohammad Taha Bahadori, Le Song +2

    cs.LGstat.MLarXiv:1611.07012v32016
  16. Efficient Continual Learning in Language Models via Thalamically Routed Cortical Columns

    Afshin Khadangi

    cs.LGarXiv:2602.22479v62026
  17. BenthicDINO: Physics-Informed Self-Distillation for View-Invariant Side-Scan Sonar Representations

    Taqi Hamoda, Hayat Rajani, Nuno Gracias

    cs.CVcs.AIcs.LGarXiv:2608.23215v12026
  18. Whisper-RIR-Mega: A Paired Clean-Reverberant Speech Benchmark for ASR Robustness to Room Acoustics

    Mandip Goswami

    eess.AScs.AIcs.LGarXiv:2603.02252v32026
  19. Operator Learning Using Weak Supervision from Walk-on-Spheres

    Hrishikesh Viswanath, Hong Chul Nam, Xi Deng +3

    cs.LGarXiv:2603.01193v22026
  20. ZeroQuant: Efficient and Affordable Post-Training Quantization for Large-Scale Transformers

    Zhewei Yao, Reza Yazdani Aminabadi, Minjia Zhang +3

    cs.CLcs.LGarXiv:2206.01861v12022
  21. Simple and Controllable Music Generation

    Jade Copet, Felix Kreuk, Itai Gat +5

    cs.SDcs.AIcs.LGarXiv:2306.05284v32023
  22. Self-Sovereign Agent

    Wenjie Qu, Xuandong Zhao, Jiaheng Zhang +1

    cs.CRcs.CYcs.LGarXiv:2604.08551v12026
  23. Distribution-Conditioned Transport

    Nic Fishman, Gokul Gowri, Paolo L. B. Fischer +3

    cs.LGarXiv:2603.04736v12026
  24. TailSieve: Partial-Rollout-Guided Tail Routing for LLM Rollouts

    Tianqi Xu, Lu Lv, Haoyang Huang +15

    cs.AIcs.LGarXiv:2608.22788v12026
  25. A comprehensive study of non-adaptive and residual-based adaptive sampling for physics-informed neural networks

    Chenxi Wu, Min Zhu, Qinyang Tan +2

    physics.comp-phcs.LGarXiv:2207.10289v12022
  26. KARL: Knowledge Agents via Reinforcement Learning

    Jonathan D. Chang, Andrew Drozdov, Shubham Toshniwal +23

    cs.AIcs.LGarXiv:2603.05218v12026
  27. A Theoretical Analysis of Deep Q-Learning

    Jianqing Fan, Zhaoran Wang, Yuchen Xie +1

    cs.LGmath.OCstat.MLarXiv:1901.00137v32019
  28. Multi-talker Speech Separation with Utterance-level Permutation Invariant Training of Deep Recurrent Neural Networks

    Morten Kolbæk, Dong Yu, Zheng-Hua Tan +1

    cs.SDcs.LGeess.ASarXiv:1703.06284v22017
  29. Reasoning as Compression: Unifying Budget Forcing via the Conditional Information Bottleneck

    Fabio Valerio Massoli, Andrey Kuzmin, Arash Behboodi

    cs.LGarXiv:2603.08462v22026
  30. Federated learning with hierarchical clustering of local updates to improve training on non-IID data

    Christopher Briggs, Zhong Fan, Peter Andras

    cs.LGstat.MLarXiv:2004.11791v22020
  31. Hard Negative Mixing for Contrastive Learning

    Yannis Kalantidis, Mert Bulent Sariyildiz, Noe Pion +2

    cs.CVcs.LGarXiv:2010.01028v22020
  32. OSMDA: OpenStreetMap-based Domain Adaptation for Remote Sensing VLMs

    Stefan Maria Ailuro, Mario Markov, Mohammad Mahdi +3

    cs.CVcs.LGarXiv:2603.11804v32026
  33. Deep Gaussian Embedding of Graphs: Unsupervised Inductive Learning via Ranking

    Aleksandar Bojchevski, Stephan Günnemann

    stat.MLcs.LGcs.SIarXiv:1707.03815v42017
  34. Censored LLMs as a Natural Testbed for Secret Knowledge Elicitation

    Helena Casademunt, Bartosz Cywiński, Khoi Tran +3

    cs.LGcs.AIcs.CLarXiv:2603.05494v22026
  35. Variational Flow Maps: Make Some Noise for One-Step Conditional Generation

    Abbas Mammadov, So Takao, Bohan Chen +4

    cs.CVcs.LGstat.MLarXiv:2603.07276v12026
  36. Deep Ensembles: A Loss Landscape Perspective

    Stanislav Fort, Huiyi Hu, Balaji Lakshminarayanan

    stat.MLcs.LGarXiv:1912.02757v22019
  37. Learning to Optimize Via Posterior Sampling

    Daniel Russo, Benjamin Van Roy

    cs.LGarXiv:1301.2609v52013
  38. Low Data Drug Discovery with One-shot Learning

    Han Altae-Tran, Bharath Ramsundar, Aneesh S. Pappu +1

    cs.LGstat.MLarXiv:1611.03199v12016
  39. NaviDriveVLM: Decoupling High-Level Reasoning and Motion Planning for Autonomous Driving

    Ximeng Tao, Pardis Taghavi, Dimitar Filev +2

    cs.ROcs.LGarXiv:2603.07901v12026
  40. ConvergeFlow: Language Flow with Provable Convergence to Token Embeddings

    Na Li, Yuchen Jiao, Changxiao Cai +1

    cs.CLcs.AIcs.LGarXiv:2608.23551v12026
  41. $V_{0.5}$: Generalist Value Model as a Prior for Sparse RL Rollouts

    Yi-Kai Zhang, Yueqing Sun, Hongyan Hao +4

    cs.LGcs.AIcs.CLarXiv:2603.10848v12026
  42. LUKE: Deep Contextualized Entity Representations with Entity-aware Self-attention

    Ikuya Yamada, Akari Asai, Hiroyuki Shindo +2

    cs.CLcs.LGarXiv:2010.01057v12020
  43. VideoGPT: Video Generation using VQ-VAE and Transformers

    Wilson Yan, Yunzhi Zhang, Pieter Abbeel +1

    cs.CVcs.LGarXiv:2104.10157v22021
  44. Asynchronous Federated Optimization

    Cong Xie, Sanmi Koyejo, Indranil Gupta

    cs.DCcs.LGarXiv:1903.03934v52019
  45. A Comprehensive Survey on Graph Anomaly Detection with Deep Learning

    Xiaoxiao Ma, Jia Wu, Shan Xue +5

    cs.LGarXiv:2106.07178v52021
  46. Quantifying Generalization in Reinforcement Learning

    Karl Cobbe, Oleg Klimov, Chris Hesse +2

    cs.LGstat.MLarXiv:1812.02341v32018
  47. GCA: Global Centroid Alignment in Federated Learning

    Jong-Ik Park, Harry Jiang, Logan Blakely +3

    cs.LGcs.DCarXiv:2608.22593v12026
  48. Video Summarization with Long Short-term Memory

    Ke Zhang, Wei-Lun Chao, Fei Sha +1

    cs.CVcs.LGarXiv:1605.08110v22016
  49. Power-Performance Characterization of TinyML Systems

    Yujie Zhang, Dhananjaya Wijerathne, Zhaoying Li +1

    cs.LGcs.AIarXiv:2608.21646v12026
  50. Efficient Attention: Attention with Linear Complexities

    Zhuoran Shen, Mingyuan Zhang, Haiyu Zhao +2

    cs.CVcs.AIcs.LGarXiv:1812.01243v102018
  51. Using cognitive psychology to understand GPT-3

    Marcel Binz, Eric Schulz

    cs.CLcs.AIcs.LGarXiv:2206.14576v12022
  52. ESPIRE: A Diagnostic Benchmark for Embodied Spatial Reasoning of Vision-Language Models

    Yanpeng Zhao, Wentao Ding, Hongtao Li +2

    cs.CVcs.LGcs.ROarXiv:2603.13033v12026
  53. Harmonic Networks: Deep Translation and Rotation Equivariance

    Daniel E. Worrall, Stephan J. Garbin, Daniyar Turmukhambetov +1

    cs.CVcs.LGstat.MLarXiv:1612.04642v22016
  54. A General and Adaptive Robust Loss Function

    Jonathan T. Barron

    cs.CVcs.LGstat.MLarXiv:1701.03077v102017
  55. Open-Vocabulary Semantic Segmentation with Mask-adapted CLIP

    Feng Liang, Bichen Wu, Xiaoliang Dai +6

    cs.CVcs.LGarXiv:2210.04150v32022
  56. EESEN: End-to-End Speech Recognition using Deep RNN Models and WFST-based Decoding

    Yajie Miao, Mohammad Gowayyed, Florian Metze

    cs.CLcs.LGarXiv:1507.08240v32015
  57. FlashSampling: Fast and Memory-Efficient Exact Sampling

    Tomas Ruiz, Zhen Qin, Yifan Zhang +3

    cs.LGcs.AIcs.CLarXiv:2603.15854v22026
  58. Residual Stream Duality in Modern Transformer Architectures

    Yifan Zhang

    cs.LGcs.AIcs.CLarXiv:2603.16039v22026
  59. What AstroPT knows about galaxies, and what that can teach us about LLMs

    UniverseTBD, :, Kshitij Duraphe +3

    cs.LGastro-ph.IMarXiv:2608.22614v12026
  60. Generalized Discrete Diffusion from Snapshots

    Oussama Zekri, Théo Uscidda, Nicolas Boullé +1

    stat.MLcs.AIcs.CLarXiv:2603.21342v12026