Machine Learning

Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

16,801 to 16,860 of 20,308

  1. Learning Generalizable Behaviors for Terminal Agents

    Yihang Yao, Bo Pang, Xuan Phi Nguyen +3

    cs.LGarXiv:2608.22631v12026
  2. Revisiting Small Batch Training for Deep Neural Networks

    Dominic Masters, Carlo Luschi

    cs.LGcs.CVstat.MLarXiv:1804.07612v12018
  3. Enhancing Graph Neural Network-based Fraud Detectors against Camouflaged Fraudsters

    Yingtong Dou, Zhiwei Liu, Li Sun +3

    cs.SIcs.CRcs.LGarXiv:2008.08692v12020
  4. PhysicsAgentABM: Physics-Guided Generative Agent-Based Modeling

    Kavana Venkatesh, Yinhan He, Jundong Li +1

    cs.MAcs.LGarXiv:2602.06030v22026
  5. Interpretable AI with Local Distillation

    Erin Craig, Yiling Huang, Snigdha Panigrahi

    stat.MEcs.LGstat.MLarXiv:2608.23538v12026
  6. Stochastic Gradient Descent as Approximate Bayesian Inference

    Stephan Mandt, Matthew D. Hoffman, David M. Blei

    stat.MLcs.LGarXiv:1704.04289v22017
  7. SCALE: Self-uncertainty Conditioned Adaptive Looking and Execution for Vision-Language-Action Models

    Hyeonbeom Choi, Daechul Ahn, Youhan Lee +3

    cs.ROcs.AIcs.LGarXiv:2602.04208v22026
  8. Cognitive Mapping and Planning for Visual Navigation

    Saurabh Gupta, Varun Tolani, James Davidson +3

    cs.CVcs.AIcs.LGarXiv:1702.03920v32017
  9. When More Modalities Hurt: Modality Dropout for Heavy-Duty Vehicle Engine Diagnostics

    Adeel Zafar, Slawomir Nowaczyk, Hamid Sarmadi +1

    cs.LGarXiv:2608.23161v12026
  10. Rigging the Lottery: Making All Tickets Winners

    Utku Evci, Trevor Gale, Jacob Menick +2

    cs.LGcs.CVstat.MLarXiv:1911.11134v32019
  11. FNet: Mixing Tokens with Fourier Transforms

    James Lee-Thorp, Joshua Ainslie, Ilya Eckstein +1

    cs.CLcs.LGarXiv:2105.03824v42021
  12. MemFly: On-the-Fly Memory Optimization via Information Bottleneck

    Zhenyuan Zhang, Xianzhang Jia, Zhiqin Yang +4

    cs.AIcs.LGarXiv:2602.07885v22026
  13. Quantized Evolution Strategies: High-precision Fine-tuning of Quantized LLMs at Low-precision Cost

    Yinggan Xu, Kajetan Schweighofer, Risto Miikkulainen +1

    cs.LGcs.AIarXiv:2602.03120v22026
  14. Learning Representations from EEG with Deep Recurrent-Convolutional Neural Networks

    Pouya Bashivan, Irina Rish, Mohammed Yeasin +1

    cs.LGcs.CVarXiv:1511.06448v32015
  15. Uncovering Cross-Objective Interference in Multi-Objective Alignment

    Yining Lu, Meng Jiang

    cs.CLcs.LGarXiv:2602.06869v22026
  16. Object detection via a multi-region & semantic segmentation-aware CNN model

    Spyros Gidaris, Nikos Komodakis

    cs.CVcs.LGcs.NEarXiv:1505.01749v32015
  17. pixelSplat: 3D Gaussian Splats from Image Pairs for Scalable Generalizable 3D Reconstruction

    David Charatan, Sizhe Li, Andrea Tagliasacchi +1

    cs.CVcs.LGarXiv:2312.12337v42023
  18. ChatGPT: Jack of all trades, master of none

    Jan Kocoń, Igor Cichecki, Oliwier Kaszyca +17

    cs.CLcs.AIcs.CYarXiv:2302.10724v42023
  19. DIME: Query-Efficient Framework for Membership Inference on Diffusion Models

    Tue Do, Daniel Alabi

    cs.LGcs.CRarXiv:2608.22824v12026
  20. Exploring Models and Data for Image Question Answering

    Mengye Ren, Ryan Kiros, Richard Zemel

    cs.LGcs.AIcs.CLarXiv:1505.02074v42015
  21. LLM-Planner: Few-Shot Grounded Planning for Embodied Agents with Large Language Models

    Chan Hee Song, Jiaman Wu, Clayton Washington +3

    cs.AIcs.CLcs.CVarXiv:2212.04088v32022
  22. Efficient Regression Models for Scan Statistics

    Gazi Abdur Rakib, Tristan Ashton, Ryan A. Loomis +4

    stat.MEcs.LGarXiv:2608.22201v12026
  23. Three dimensional Deep Learning approach for remote sensing image classification

    Amina Ben Hamida, A Benoit, Patrick Lambert +1

    cs.CVcs.LGstat.MLarXiv:1806.05824v12018
  24. SenTSR-Bench: Thinking with Injected Knowledge for Time-Series Reasoning

    Zelin He, Boran Han, Xiyuan Zhang +10

    cs.LGcs.AIcs.CLarXiv:2602.19455v12026
  25. Optoelectronic Reservoir Computing

    Yvan Paquot, François Duport, Anteo Smerieri +4

    cs.ETcs.LGcs.NEarXiv:1111.7219v12011
  26. Curriculum Learning for Reinforcement Learning Domains: A Framework and Survey

    Sanmit Narvekar, Bei Peng, Matteo Leonetti +3

    cs.LGcs.AIstat.MLarXiv:2003.04960v22020
  27. Turing Test on Screen: A Benchmark for Mobile GUI Agent Humanization

    Jiachen Zhu, Lingyu Yang, Rong Shan +6

    cs.AIcs.LGarXiv:2604.09574v12026
  28. Reasoning about Entailment with Neural Attention

    Tim Rocktäschel, Edward Grefenstette, Karl Moritz Hermann +2

    cs.CLcs.AIcs.LGarXiv:1509.06664v42015
  29. Deep Learning on a Data Diet: Finding Important Examples Early in Training

    Mansheej Paul, Surya Ganguli, Gintare Karolina Dziugaite

    cs.LGarXiv:2107.07075v22021
  30. Maximum-distance nonnegative matrix factorization for unmixing highly mixed grain-size distribution data: A generalization of AnalySize

    Qianqian Qi, Zhongming Chen, Peter G. M. van der Heijden

    cs.LGarXiv:2608.22681v12026
  31. A Critical Look at Targeted Instruction Selection: Disentangling What Matters (and What Doesn't)

    Nihal V. Nayak, Paula Rodriguez-Diaz, Neha Hulkund +2

    cs.LGarXiv:2602.14696v22026
  32. VLMo: Unified Vision-Language Pre-Training with Mixture-of-Modality-Experts

    Hangbo Bao, Wenhui Wang, Li Dong +5

    cs.CVcs.CLcs.LGarXiv:2111.02358v22021
  33. Meta Pseudo Labels

    Hieu Pham, Zihang Dai, Qizhe Xie +2

    cs.LGstat.MLarXiv:2003.10580v42020
  34. Tackling the Generative Learning Trilemma with Denoising Diffusion GANs

    Zhisheng Xiao, Karsten Kreis, Arash Vahdat

    cs.LGstat.MLarXiv:2112.07804v22021
  35. Assessing Generative Models via Precision and Recall

    Mehdi S. M. Sajjadi, Olivier Bachem, Mario Lucic +2

    stat.MLcs.LGarXiv:1806.00035v22018
  36. Uncertainty-Aware Vision-Language Segmentation for Medical Imaging

    Aryan Das, Tanishq Rachamalla, Koushik Biswas +2

    cs.CVcs.LGarXiv:2602.14498v22026
  37. Decoding as Optimisation on the Probability Simplex: From Top-K to Top-P (Nucleus) to Best-of-K Samplers

    Xiaotong Ji, Rasul Tutunov, Matthieu Zimmer +1

    cs.LGcs.AIarXiv:2602.18292v22026
  38. Unsupervised and Semi-supervised Learning with Categorical Generative Adversarial Networks

    Jost Tobias Springenberg

    stat.MLcs.LGarXiv:1511.06390v22015
  39. A Downsampled Variant of ImageNet as an Alternative to the CIFAR datasets

    Patryk Chrabaszcz, Ilya Loshchilov, Frank Hutter

    cs.CVcs.LGarXiv:1707.08819v32017
  40. Communication-Efficient On-Device Machine Learning: Federated Distillation and Augmentation under Non-IID Private Data

    Eunjeong Jeong, Seungeun Oh, Hyesung Kim +3

    cs.LGcs.NIstat.MLarXiv:1811.11479v22018
  41. Grid Search, Random Search, Genetic Algorithm: A Big Comparison for NAS

    Petro Liashchynskyi, Pavlo Liashchynskyi

    cs.LGcs.NEstat.MLarXiv:1912.06059v12019
  42. Crop Yield Prediction Using Deep Neural Networks

    Saeed Khaki, Lizhi Wang

    cs.LGstat.APstat.MLarXiv:1902.02860v32019
  43. Picking Winning Tickets Before Training by Preserving Gradient Flow

    Chaoqi Wang, Guodong Zhang, Roger Grosse

    cs.LGcs.CVstat.MLarXiv:2002.07376v22020
  44. Large-Scale Study of Curiosity-Driven Learning

    Yuri Burda, Harri Edwards, Deepak Pathak +3

    cs.LGcs.AIcs.CVarXiv:1808.04355v12018
  45. GRAM: Graph-based Attention Model for Healthcare Representation Learning

    Edward Choi, Mohammad Taha Bahadori, Le Song +2

    cs.LGstat.MLarXiv:1611.07012v32016
  46. Efficient Continual Learning in Language Models via Thalamically Routed Cortical Columns

    Afshin Khadangi

    cs.LGarXiv:2602.22479v62026
  47. BenthicDINO: Physics-Informed Self-Distillation for View-Invariant Side-Scan Sonar Representations

    Taqi Hamoda, Hayat Rajani, Nuno Gracias

    cs.CVcs.AIcs.LGarXiv:2608.23215v12026
  48. Whisper-RIR-Mega: A Paired Clean-Reverberant Speech Benchmark for ASR Robustness to Room Acoustics

    Mandip Goswami

    eess.AScs.AIcs.LGarXiv:2603.02252v32026
  49. Operator Learning Using Weak Supervision from Walk-on-Spheres

    Hrishikesh Viswanath, Hong Chul Nam, Xi Deng +3

    cs.LGarXiv:2603.01193v22026
  50. ZeroQuant: Efficient and Affordable Post-Training Quantization for Large-Scale Transformers

    Zhewei Yao, Reza Yazdani Aminabadi, Minjia Zhang +3

    cs.CLcs.LGarXiv:2206.01861v12022
  51. Simple and Controllable Music Generation

    Jade Copet, Felix Kreuk, Itai Gat +5

    cs.SDcs.AIcs.LGarXiv:2306.05284v32023
  52. Self-Sovereign Agent

    Wenjie Qu, Xuandong Zhao, Jiaheng Zhang +1

    cs.CRcs.CYcs.LGarXiv:2604.08551v12026
  53. Distribution-Conditioned Transport

    Nic Fishman, Gokul Gowri, Paolo L. B. Fischer +3

    cs.LGarXiv:2603.04736v12026
  54. TailSieve: Partial-Rollout-Guided Tail Routing for LLM Rollouts

    Tianqi Xu, Lu Lv, Haoyang Huang +15

    cs.AIcs.LGarXiv:2608.22788v12026
  55. A comprehensive study of non-adaptive and residual-based adaptive sampling for physics-informed neural networks

    Chenxi Wu, Min Zhu, Qinyang Tan +2

    physics.comp-phcs.LGarXiv:2207.10289v12022
  56. KARL: Knowledge Agents via Reinforcement Learning

    Jonathan D. Chang, Andrew Drozdov, Shubham Toshniwal +23

    cs.AIcs.LGarXiv:2603.05218v12026
  57. A Theoretical Analysis of Deep Q-Learning

    Jianqing Fan, Zhaoran Wang, Yuchen Xie +1

    cs.LGmath.OCstat.MLarXiv:1901.00137v32019
  58. Multi-talker Speech Separation with Utterance-level Permutation Invariant Training of Deep Recurrent Neural Networks

    Morten Kolbæk, Dong Yu, Zheng-Hua Tan +1

    cs.SDcs.LGeess.ASarXiv:1703.06284v22017
  59. Reasoning as Compression: Unifying Budget Forcing via the Conditional Information Bottleneck

    Fabio Valerio Massoli, Andrey Kuzmin, Arash Behboodi

    cs.LGarXiv:2603.08462v22026
  60. Federated learning with hierarchical clustering of local updates to improve training on non-IID data

    Christopher Briggs, Zhong Fan, Peter Andras

    cs.LGstat.MLarXiv:2004.11791v22020