Machine Learning
Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
16,801 to 16,860 of 20,308
Learning Generalizable Behaviors for Terminal Agents
Yihang Yao, Bo Pang, Xuan Phi Nguyen +3
cs.LGarXiv:2608.22631v12026Revisiting Small Batch Training for Deep Neural Networks
Dominic Masters, Carlo Luschi
cs.LGcs.CVstat.MLarXiv:1804.07612v12018Enhancing Graph Neural Network-based Fraud Detectors against Camouflaged Fraudsters
Yingtong Dou, Zhiwei Liu, Li Sun +3
cs.SIcs.CRcs.LGarXiv:2008.08692v12020PhysicsAgentABM: Physics-Guided Generative Agent-Based Modeling
Kavana Venkatesh, Yinhan He, Jundong Li +1
cs.MAcs.LGarXiv:2602.06030v22026Interpretable AI with Local Distillation
Erin Craig, Yiling Huang, Snigdha Panigrahi
stat.MEcs.LGstat.MLarXiv:2608.23538v12026Stochastic Gradient Descent as Approximate Bayesian Inference
Stephan Mandt, Matthew D. Hoffman, David M. Blei
stat.MLcs.LGarXiv:1704.04289v22017SCALE: Self-uncertainty Conditioned Adaptive Looking and Execution for Vision-Language-Action Models
Hyeonbeom Choi, Daechul Ahn, Youhan Lee +3
cs.ROcs.AIcs.LGarXiv:2602.04208v22026Cognitive Mapping and Planning for Visual Navigation
Saurabh Gupta, Varun Tolani, James Davidson +3
cs.CVcs.AIcs.LGarXiv:1702.03920v32017When More Modalities Hurt: Modality Dropout for Heavy-Duty Vehicle Engine Diagnostics
Adeel Zafar, Slawomir Nowaczyk, Hamid Sarmadi +1
cs.LGarXiv:2608.23161v12026Rigging the Lottery: Making All Tickets Winners
Utku Evci, Trevor Gale, Jacob Menick +2
cs.LGcs.CVstat.MLarXiv:1911.11134v32019FNet: Mixing Tokens with Fourier Transforms
James Lee-Thorp, Joshua Ainslie, Ilya Eckstein +1
cs.CLcs.LGarXiv:2105.03824v42021MemFly: On-the-Fly Memory Optimization via Information Bottleneck
Zhenyuan Zhang, Xianzhang Jia, Zhiqin Yang +4
cs.AIcs.LGarXiv:2602.07885v22026Quantized Evolution Strategies: High-precision Fine-tuning of Quantized LLMs at Low-precision Cost
Yinggan Xu, Kajetan Schweighofer, Risto Miikkulainen +1
cs.LGcs.AIarXiv:2602.03120v22026Learning Representations from EEG with Deep Recurrent-Convolutional Neural Networks
Pouya Bashivan, Irina Rish, Mohammed Yeasin +1
cs.LGcs.CVarXiv:1511.06448v32015Uncovering Cross-Objective Interference in Multi-Objective Alignment
Yining Lu, Meng Jiang
cs.CLcs.LGarXiv:2602.06869v22026Object detection via a multi-region & semantic segmentation-aware CNN model
Spyros Gidaris, Nikos Komodakis
cs.CVcs.LGcs.NEarXiv:1505.01749v32015pixelSplat: 3D Gaussian Splats from Image Pairs for Scalable Generalizable 3D Reconstruction
David Charatan, Sizhe Li, Andrea Tagliasacchi +1
cs.CVcs.LGarXiv:2312.12337v42023ChatGPT: Jack of all trades, master of none
Jan Kocoń, Igor Cichecki, Oliwier Kaszyca +17
cs.CLcs.AIcs.CYarXiv:2302.10724v42023DIME: Query-Efficient Framework for Membership Inference on Diffusion Models
Tue Do, Daniel Alabi
cs.LGcs.CRarXiv:2608.22824v12026Exploring Models and Data for Image Question Answering
Mengye Ren, Ryan Kiros, Richard Zemel
cs.LGcs.AIcs.CLarXiv:1505.02074v42015LLM-Planner: Few-Shot Grounded Planning for Embodied Agents with Large Language Models
Chan Hee Song, Jiaman Wu, Clayton Washington +3
cs.AIcs.CLcs.CVarXiv:2212.04088v32022Efficient Regression Models for Scan Statistics
Gazi Abdur Rakib, Tristan Ashton, Ryan A. Loomis +4
stat.MEcs.LGarXiv:2608.22201v12026Three dimensional Deep Learning approach for remote sensing image classification
Amina Ben Hamida, A Benoit, Patrick Lambert +1
cs.CVcs.LGstat.MLarXiv:1806.05824v12018SenTSR-Bench: Thinking with Injected Knowledge for Time-Series Reasoning
Zelin He, Boran Han, Xiyuan Zhang +10
cs.LGcs.AIcs.CLarXiv:2602.19455v12026Optoelectronic Reservoir Computing
Yvan Paquot, François Duport, Anteo Smerieri +4
cs.ETcs.LGcs.NEarXiv:1111.7219v12011Curriculum Learning for Reinforcement Learning Domains: A Framework and Survey
Sanmit Narvekar, Bei Peng, Matteo Leonetti +3
cs.LGcs.AIstat.MLarXiv:2003.04960v22020Turing Test on Screen: A Benchmark for Mobile GUI Agent Humanization
Jiachen Zhu, Lingyu Yang, Rong Shan +6
cs.AIcs.LGarXiv:2604.09574v12026Reasoning about Entailment with Neural Attention
Tim Rocktäschel, Edward Grefenstette, Karl Moritz Hermann +2
cs.CLcs.AIcs.LGarXiv:1509.06664v42015Deep Learning on a Data Diet: Finding Important Examples Early in Training
Mansheej Paul, Surya Ganguli, Gintare Karolina Dziugaite
cs.LGarXiv:2107.07075v22021Maximum-distance nonnegative matrix factorization for unmixing highly mixed grain-size distribution data: A generalization of AnalySize
Qianqian Qi, Zhongming Chen, Peter G. M. van der Heijden
cs.LGarXiv:2608.22681v12026A Critical Look at Targeted Instruction Selection: Disentangling What Matters (and What Doesn't)
Nihal V. Nayak, Paula Rodriguez-Diaz, Neha Hulkund +2
cs.LGarXiv:2602.14696v22026VLMo: Unified Vision-Language Pre-Training with Mixture-of-Modality-Experts
Hangbo Bao, Wenhui Wang, Li Dong +5
cs.CVcs.CLcs.LGarXiv:2111.02358v22021Meta Pseudo Labels
Hieu Pham, Zihang Dai, Qizhe Xie +2
cs.LGstat.MLarXiv:2003.10580v42020Tackling the Generative Learning Trilemma with Denoising Diffusion GANs
Zhisheng Xiao, Karsten Kreis, Arash Vahdat
cs.LGstat.MLarXiv:2112.07804v22021Assessing Generative Models via Precision and Recall
Mehdi S. M. Sajjadi, Olivier Bachem, Mario Lucic +2
stat.MLcs.LGarXiv:1806.00035v22018Uncertainty-Aware Vision-Language Segmentation for Medical Imaging
Aryan Das, Tanishq Rachamalla, Koushik Biswas +2
cs.CVcs.LGarXiv:2602.14498v22026Decoding as Optimisation on the Probability Simplex: From Top-K to Top-P (Nucleus) to Best-of-K Samplers
Xiaotong Ji, Rasul Tutunov, Matthieu Zimmer +1
cs.LGcs.AIarXiv:2602.18292v22026Unsupervised and Semi-supervised Learning with Categorical Generative Adversarial Networks
Jost Tobias Springenberg
stat.MLcs.LGarXiv:1511.06390v22015A Downsampled Variant of ImageNet as an Alternative to the CIFAR datasets
Patryk Chrabaszcz, Ilya Loshchilov, Frank Hutter
cs.CVcs.LGarXiv:1707.08819v32017Communication-Efficient On-Device Machine Learning: Federated Distillation and Augmentation under Non-IID Private Data
Eunjeong Jeong, Seungeun Oh, Hyesung Kim +3
cs.LGcs.NIstat.MLarXiv:1811.11479v22018Grid Search, Random Search, Genetic Algorithm: A Big Comparison for NAS
Petro Liashchynskyi, Pavlo Liashchynskyi
cs.LGcs.NEstat.MLarXiv:1912.06059v12019Crop Yield Prediction Using Deep Neural Networks
Saeed Khaki, Lizhi Wang
cs.LGstat.APstat.MLarXiv:1902.02860v32019Picking Winning Tickets Before Training by Preserving Gradient Flow
Chaoqi Wang, Guodong Zhang, Roger Grosse
cs.LGcs.CVstat.MLarXiv:2002.07376v22020Large-Scale Study of Curiosity-Driven Learning
Yuri Burda, Harri Edwards, Deepak Pathak +3
cs.LGcs.AIcs.CVarXiv:1808.04355v12018GRAM: Graph-based Attention Model for Healthcare Representation Learning
Edward Choi, Mohammad Taha Bahadori, Le Song +2
cs.LGstat.MLarXiv:1611.07012v32016Efficient Continual Learning in Language Models via Thalamically Routed Cortical Columns
Afshin Khadangi
cs.LGarXiv:2602.22479v62026BenthicDINO: Physics-Informed Self-Distillation for View-Invariant Side-Scan Sonar Representations
Taqi Hamoda, Hayat Rajani, Nuno Gracias
cs.CVcs.AIcs.LGarXiv:2608.23215v12026Whisper-RIR-Mega: A Paired Clean-Reverberant Speech Benchmark for ASR Robustness to Room Acoustics
Mandip Goswami
eess.AScs.AIcs.LGarXiv:2603.02252v32026Operator Learning Using Weak Supervision from Walk-on-Spheres
Hrishikesh Viswanath, Hong Chul Nam, Xi Deng +3
cs.LGarXiv:2603.01193v22026ZeroQuant: Efficient and Affordable Post-Training Quantization for Large-Scale Transformers
Zhewei Yao, Reza Yazdani Aminabadi, Minjia Zhang +3
cs.CLcs.LGarXiv:2206.01861v12022Simple and Controllable Music Generation
Jade Copet, Felix Kreuk, Itai Gat +5
cs.SDcs.AIcs.LGarXiv:2306.05284v32023Self-Sovereign Agent
Wenjie Qu, Xuandong Zhao, Jiaheng Zhang +1
cs.CRcs.CYcs.LGarXiv:2604.08551v12026Distribution-Conditioned Transport
Nic Fishman, Gokul Gowri, Paolo L. B. Fischer +3
cs.LGarXiv:2603.04736v12026TailSieve: Partial-Rollout-Guided Tail Routing for LLM Rollouts
Tianqi Xu, Lu Lv, Haoyang Huang +15
cs.AIcs.LGarXiv:2608.22788v12026A comprehensive study of non-adaptive and residual-based adaptive sampling for physics-informed neural networks
Chenxi Wu, Min Zhu, Qinyang Tan +2
physics.comp-phcs.LGarXiv:2207.10289v12022KARL: Knowledge Agents via Reinforcement Learning
Jonathan D. Chang, Andrew Drozdov, Shubham Toshniwal +23
cs.AIcs.LGarXiv:2603.05218v12026A Theoretical Analysis of Deep Q-Learning
Jianqing Fan, Zhaoran Wang, Yuchen Xie +1
cs.LGmath.OCstat.MLarXiv:1901.00137v32019Multi-talker Speech Separation with Utterance-level Permutation Invariant Training of Deep Recurrent Neural Networks
Morten Kolbæk, Dong Yu, Zheng-Hua Tan +1
cs.SDcs.LGeess.ASarXiv:1703.06284v22017Reasoning as Compression: Unifying Budget Forcing via the Conditional Information Bottleneck
Fabio Valerio Massoli, Andrey Kuzmin, Arash Behboodi
cs.LGarXiv:2603.08462v22026Federated learning with hierarchical clustering of local updates to improve training on non-IID data
Christopher Briggs, Zhong Fan, Peter Andras
cs.LGstat.MLarXiv:2004.11791v22020