Machine Learning
Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
3,721 to 3,780 of 20,454
Probabilistic Recursive Reasoning for Multi-Agent Reinforcement Learning
Ying Wen, Yaodong Yang, Rui Luo +2
cs.LGcs.AIstat.MLarXiv:1901.09207v22019Parallel Predictive Entropy Search for Batch Global Optimization of Expensive Objective Functions
Amar Shah, Zoubin Ghahramani
cs.LGstat.MLarXiv:1511.07130v12015Lorentz Group Equivariant Neural Network for Particle Physics
Alexander Bogatskiy, Brandon Anderson, Jan T. Offermann +3
hep-phcs.LGhep-exarXiv:2006.04780v12020ARC-Bench: Closed-Loop Replanning Masks Broken Action Ranking in Frozen JEPA World Models
Zhengshu Zhang, Zhiyuan Li
cs.AIcs.LGcs.ROarXiv:2609.05461v12026GEP-PG: Decoupling Exploration and Exploitation in Deep Reinforcement Learning Algorithms
Cédric Colas, Olivier Sigaud, Pierre-Yves Oudeyer
cs.LGarXiv:1802.05054v52018RecoGym: A Reinforcement Learning Environment for the problem of Product Recommendation in Online Advertising
David Rohde, Stephen Bonner, Travis Dunlop +2
cs.IRcs.LGarXiv:1808.00720v22018Learning From Noisy Singly-labeled Data
Ashish Khetan, Zachary C. Lipton, Anima Anandkumar
cs.LGarXiv:1712.04577v22017Generalization Error Bounds of Gradient Descent for Learning Over-parameterized Deep ReLU Networks
Yuan Cao, Quanquan Gu
cs.LGmath.OCstat.MLarXiv:1902.01384v42019Self-supervised Feature Learning for 3D Medical Images by Playing a Rubik's Cube
Xinrui Zhuang, Yuexiang Li, Yifan Hu +3
cs.CVcs.LGeess.IVarXiv:1910.02241v12019Neural Architecture Transfer
Zhichao Lu, Gautam Sreekumar, Erik Goodman +3
cs.CVcs.LGcs.NEarXiv:2005.05859v22020Woulda, Coulda, Shoulda: Counterfactually-Guided Policy Search
Lars Buesing, Theophane Weber, Yori Zwols +4
cs.LGstat.MLarXiv:1811.06272v12018Reason Through the Latent! Making Latent Visual Reasoning Necessary
Suhyeong Park, Junha Jung, Jaewoo Kang
cs.AIcs.CLcs.CVarXiv:2609.06746v12026Eliciting Weak-to-Strong Generalization with On-Policy Reverse Distillation
Youngrok Park, Sangmin Bae, Hojung Jung +6
cs.LGcs.CLarXiv:2609.08798v12026One Loss for All: Deep Hashing with a Single Cosine Similarity based Learning Objective
Jiun Tian Hoe, Kam Woh Ng, Tianyu Zhang +3
cs.CVcs.LGarXiv:2109.14449v12021Deep Signature Transforms
Patric Bonnier, Patrick Kidger, Imanol Perez Arribas +2
cs.LGstat.MLarXiv:1905.08494v22019Navigation with Large Language Models: Semantic Guesswork as a Heuristic for Planning
Dhruv Shah, Michael Equi, Blazej Osinski +3
cs.ROcs.AIcs.CLarXiv:2310.10103v12023Damage-Aware Bandit Pruning for Vision and Language Transformers
Salem Ameen, Sunil Vadera
cs.AIcs.LGarXiv:2609.05448v12026A geometric alternative to Nesterov's accelerated gradient descent
Sébastien Bubeck, Yin Tat Lee, Mohit Singh
math.OCcs.DScs.LGarXiv:1506.08187v12015Boosting Adversarial Training with Hypersphere Embedding
Tianyu Pang, Xiao Yang, Yinpeng Dong +3
cs.LGcs.CRcs.CVarXiv:2002.08619v32020General-purpose Tagging of Freesound Audio with AudioSet Labels: Task Description, Dataset, and Baseline
Eduardo Fonseca, Manoj Plakal, Frederic Font +4
cs.SDcs.LGeess.ASarXiv:1807.09902v32018Minimizing Energy Consumption Leads to the Emergence of Gaits in Legged Robots
Zipeng Fu, Ashish Kumar, Jitendra Malik +1
cs.ROcs.AIcs.CVarXiv:2111.01674v12021A Topology Layer for Machine Learning
Rickard Brüel-Gabrielsson, Bradley J. Nelson, Anjan Dwaraknath +3
cs.LGmath.ATstat.MLarXiv:1905.12200v22019Sampling-based sublinear low-rank matrix arithmetic framework for dequantizing quantum machine learning
Nai-Hui Chia, András Gilyén, Tongyang Li +3
cs.DScs.LGquant-pharXiv:1910.06151v42019Deep Metric Learning for Practical Person Re-Identification
Dong Yi, Zhen Lei, Stan Z. Li
cs.CVcs.LGcs.NEarXiv:1407.4979v12014High-dimensional Asymptotics of Feature Learning: How One Gradient Step Improves the Representation
Jimmy Ba, Murat A. Erdogdu, Taiji Suzuki +3
stat.MLcs.LGmath.STarXiv:2205.01445v12022Generalized Shape Metrics on Neural Representations
Alex H. Williams, Erin Kunz, Simon Kornblith +1
stat.MLcs.LGarXiv:2110.14739v22021Steering Geometry: Validating Human Value Geometry in LLM Steering Space
Mohammad Mahdi Abootorabi, Armin Saghafian, Ali Bazshoushtari +5
cs.CLcs.AIcs.LGarXiv:2609.06289v12026CLIP-Dissect: Automatic Description of Neuron Representations in Deep Vision Networks
Tuomas Oikarinen, Tsui-Wei Weng
cs.CVcs.AIcs.LGarXiv:2204.10965v52022Multi-Agent Adversarial Inverse Reinforcement Learning
Lantao Yu, Jiaming Song, Stefano Ermon
cs.LGstat.MLarXiv:1907.13220v12019CODA: A Real-World Road Corner Case Dataset for Object Detection in Autonomous Driving
Kaican Li, Kai Chen, Haoyu Wang +10
cs.CVcs.LGcs.ROarXiv:2203.07724v32022Cryptanalytic Extraction of Neural Network Models
Nicholas Carlini, Matthew Jagielski, Ilya Mironov
cs.LGcs.CRarXiv:2003.04884v22020SIRNN: A Math Library for Secure RNN Inference
Deevashwer Rathee, Mayank Rathee, Rahul Kranti Kiran Goli +4
cs.CRcs.LGcs.MSarXiv:2105.04236v12021DistGNN: Scalable Distributed Training for Large-Scale Graph Neural Networks
Vasimuddin Md, Sanchit Misra, Guixiang Ma +6
cs.LGcs.DCarXiv:2104.06700v32021Miles v0.1: Production-Level Post-Training
RadixArk, :, Tom Chen +11
cs.LGcs.CLarXiv:2609.08368v12026Interpretation and Generalization of Score Matching
Siwei Lyu
cs.LGstat.MLarXiv:1205.2629v12012TimeSHAP: Explaining Recurrent Models through Sequence Perturbations
João Bento, Pedro Saleiro, André F. Cruz +2
cs.LGcs.AIarXiv:2012.00073v22020Distributionally Robust Federated Averaging
Yuyang Deng, Mohammad Mahdi Kamani, Mehrdad Mahdavi
cs.LGcs.DCstat.MLarXiv:2102.12660v12021Black-box Explanation of Object Detectors via Saliency Maps
Vitali Petsiuk, Rajiv Jain, Varun Manjunatha +4
cs.CVcs.AIcs.LGarXiv:2006.03204v22020Robust Regression via Hard Thresholding
Kush Bhatia, Prateek Jain, Purushottam Kar
cs.LGstat.MLarXiv:1506.02428v12015TREC CAsT 2019: The Conversational Assistance Track Overview
Jeffrey Dalton, Chenyan Xiong, Jamie Callan
cs.IRcs.CLcs.LGarXiv:2003.13624v12020Learning and Evaluating Graph Neural Network Explanations based on Counterfactual and Factual Reasoning
Juntao Tan, Shijie Geng, Zuohui Fu +4
cs.IRcs.LGarXiv:2202.08816v32022Online Draft Co-Training for Speculative Decoding in Large-Scale, Long-Context RL Post-Training
Zili Wang, Zhaopeng Qiu, Yuekai Zhang +2
cs.LGcs.DCarXiv:2609.07108v12026On the Transfer of Inductive Bias from Simulation to the Real World: a New Disentanglement Dataset
Muhammad Waleed Gondal, Manuel Wüthrich, Đorđe Miladinović +7
stat.MLcs.LGarXiv:1906.03292v32019Environments as Scaffold: Enriching Feedback to Bootstrap Self-Evolving Agents in Long-Horizon Tasks
Hongbang Yuan, Zhuoran Jin, Yixin Cao
cs.LGcs.AIcs.CLarXiv:2609.08404v12026How to Exploit Hyperspherical Embeddings for Out-of-Distribution Detection?
Yifei Ming, Yiyou Sun, Ousmane Dia +1
cs.CVcs.LGarXiv:2203.04450v32022Biomedical Entity Representations with Synonym Marginalization
Mujeen Sung, Hwisang Jeon, Jinhyuk Lee +1
cs.CLcs.LGarXiv:2005.00239v12020Geometric Understanding of Deep Learning
Na Lei, Zhongxuan Luo, Shing-Tung Yau +1
cs.LGstat.MLarXiv:1805.10451v22018Triformer: Triangular, Variable-Specific Attentions for Long Sequence Multivariate Time Series Forecasting--Full Version
Razvan-Gabriel Cirstea, Chenjuan Guo, Bin Yang +3
cs.LGarXiv:2204.13767v12022Anti-DreamBooth: Protecting users from personalized text-to-image synthesis
Thanh Van Le, Hao Phung, Thuan Hoang Nguyen +3
cs.CVcs.CRcs.LGarXiv:2303.15433v22023Unsupervised Depth Completion from Visual Inertial Odometry
Alex Wong, Xiaohan Fei, Stephanie Tsuei +1
cs.CVcs.AIcs.LGarXiv:1905.08616v42019Improving Diffusion Inverse Problem Solving with Decoupled Noise Annealing
Bingliang Zhang, Wenda Chu, Julius Berner +3
cs.LGcs.AIcs.CVarXiv:2407.01521v32024Guaranteed Non-convex Optimization: Submodular Maximization over Continuous Domains
Andrew An Bian, Baharan Mirzasoleiman, Joachim M. Buhmann +1
cs.LGcs.DSarXiv:1606.05615v52016Low-Power Neuromorphic Hardware for Signal Processing Applications
Bipin Rajendran, Abu Sebastian, Michael Schmuker +2
cs.ETcs.LGcs.NEarXiv:1901.03690v32019Federated Learning: A Signal Processing Perspective
Tomer Gafni, Nir Shlezinger, Kobi Cohen +2
eess.SPcs.LGarXiv:2103.17150v22021Trainability of Dissipative Perceptron-Based Quantum Neural Networks
Kunal Sharma, M. Cerezo, Lukasz Cincio +1
quant-phcs.LGarXiv:2005.12458v22020Learning State Representations for Query Optimization with Deep Reinforcement Learning
Jennifer Ortiz, Magdalena Balazinska, Johannes Gehrke +1
cs.DBcs.AIcs.LGarXiv:1803.08604v12018SafeDrug: Dual Molecular Graph Encoders for Recommending Effective and Safe Drug Combinations
Chaoqi Yang, Cao Xiao, Fenglong Ma +2
cs.LGarXiv:2105.02711v22021Retro*: Learning Retrosynthetic Planning with Neural Guided A* Search
Binghong Chen, Chengtao Li, Hanjun Dai +1
cs.LGcs.AIstat.MLarXiv:2006.15820v12020Kalman Delta Networks: Uncertainty-aware Associative Memory
Ngoc Bui, Tinglin Huang, Rex Ying
cs.LGcs.AIarXiv:2609.07816v12026MOLE: Detecting Insider Threats in AI Agents
Aashiq Muhamed, Virginia Smith
cs.LGcs.CLcs.CRarXiv:2609.06966v12026