Machine Learning

Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

3,721 to 3,780 of 20,454

  1. Probabilistic Recursive Reasoning for Multi-Agent Reinforcement Learning

    Ying Wen, Yaodong Yang, Rui Luo +2

    cs.LGcs.AIstat.MLarXiv:1901.09207v22019
  2. Parallel Predictive Entropy Search for Batch Global Optimization of Expensive Objective Functions

    Amar Shah, Zoubin Ghahramani

    cs.LGstat.MLarXiv:1511.07130v12015
  3. Lorentz Group Equivariant Neural Network for Particle Physics

    Alexander Bogatskiy, Brandon Anderson, Jan T. Offermann +3

    hep-phcs.LGhep-exarXiv:2006.04780v12020
  4. ARC-Bench: Closed-Loop Replanning Masks Broken Action Ranking in Frozen JEPA World Models

    Zhengshu Zhang, Zhiyuan Li

    cs.AIcs.LGcs.ROarXiv:2609.05461v12026
  5. GEP-PG: Decoupling Exploration and Exploitation in Deep Reinforcement Learning Algorithms

    Cédric Colas, Olivier Sigaud, Pierre-Yves Oudeyer

    cs.LGarXiv:1802.05054v52018
  6. RecoGym: A Reinforcement Learning Environment for the problem of Product Recommendation in Online Advertising

    David Rohde, Stephen Bonner, Travis Dunlop +2

    cs.IRcs.LGarXiv:1808.00720v22018
  7. Learning From Noisy Singly-labeled Data

    Ashish Khetan, Zachary C. Lipton, Anima Anandkumar

    cs.LGarXiv:1712.04577v22017
  8. Generalization Error Bounds of Gradient Descent for Learning Over-parameterized Deep ReLU Networks

    Yuan Cao, Quanquan Gu

    cs.LGmath.OCstat.MLarXiv:1902.01384v42019
  9. Self-supervised Feature Learning for 3D Medical Images by Playing a Rubik's Cube

    Xinrui Zhuang, Yuexiang Li, Yifan Hu +3

    cs.CVcs.LGeess.IVarXiv:1910.02241v12019
  10. Neural Architecture Transfer

    Zhichao Lu, Gautam Sreekumar, Erik Goodman +3

    cs.CVcs.LGcs.NEarXiv:2005.05859v22020
  11. Woulda, Coulda, Shoulda: Counterfactually-Guided Policy Search

    Lars Buesing, Theophane Weber, Yori Zwols +4

    cs.LGstat.MLarXiv:1811.06272v12018
  12. Reason Through the Latent! Making Latent Visual Reasoning Necessary

    Suhyeong Park, Junha Jung, Jaewoo Kang

    cs.AIcs.CLcs.CVarXiv:2609.06746v12026
  13. Eliciting Weak-to-Strong Generalization with On-Policy Reverse Distillation

    Youngrok Park, Sangmin Bae, Hojung Jung +6

    cs.LGcs.CLarXiv:2609.08798v12026
  14. One Loss for All: Deep Hashing with a Single Cosine Similarity based Learning Objective

    Jiun Tian Hoe, Kam Woh Ng, Tianyu Zhang +3

    cs.CVcs.LGarXiv:2109.14449v12021
  15. Deep Signature Transforms

    Patric Bonnier, Patrick Kidger, Imanol Perez Arribas +2

    cs.LGstat.MLarXiv:1905.08494v22019
  16. Navigation with Large Language Models: Semantic Guesswork as a Heuristic for Planning

    Dhruv Shah, Michael Equi, Blazej Osinski +3

    cs.ROcs.AIcs.CLarXiv:2310.10103v12023
  17. Damage-Aware Bandit Pruning for Vision and Language Transformers

    Salem Ameen, Sunil Vadera

    cs.AIcs.LGarXiv:2609.05448v12026
  18. A geometric alternative to Nesterov's accelerated gradient descent

    Sébastien Bubeck, Yin Tat Lee, Mohit Singh

    math.OCcs.DScs.LGarXiv:1506.08187v12015
  19. Boosting Adversarial Training with Hypersphere Embedding

    Tianyu Pang, Xiao Yang, Yinpeng Dong +3

    cs.LGcs.CRcs.CVarXiv:2002.08619v32020
  20. General-purpose Tagging of Freesound Audio with AudioSet Labels: Task Description, Dataset, and Baseline

    Eduardo Fonseca, Manoj Plakal, Frederic Font +4

    cs.SDcs.LGeess.ASarXiv:1807.09902v32018
  21. Minimizing Energy Consumption Leads to the Emergence of Gaits in Legged Robots

    Zipeng Fu, Ashish Kumar, Jitendra Malik +1

    cs.ROcs.AIcs.CVarXiv:2111.01674v12021
  22. A Topology Layer for Machine Learning

    Rickard Brüel-Gabrielsson, Bradley J. Nelson, Anjan Dwaraknath +3

    cs.LGmath.ATstat.MLarXiv:1905.12200v22019
  23. Sampling-based sublinear low-rank matrix arithmetic framework for dequantizing quantum machine learning

    Nai-Hui Chia, András Gilyén, Tongyang Li +3

    cs.DScs.LGquant-pharXiv:1910.06151v42019
  24. Deep Metric Learning for Practical Person Re-Identification

    Dong Yi, Zhen Lei, Stan Z. Li

    cs.CVcs.LGcs.NEarXiv:1407.4979v12014
  25. High-dimensional Asymptotics of Feature Learning: How One Gradient Step Improves the Representation

    Jimmy Ba, Murat A. Erdogdu, Taiji Suzuki +3

    stat.MLcs.LGmath.STarXiv:2205.01445v12022
  26. Generalized Shape Metrics on Neural Representations

    Alex H. Williams, Erin Kunz, Simon Kornblith +1

    stat.MLcs.LGarXiv:2110.14739v22021
  27. Steering Geometry: Validating Human Value Geometry in LLM Steering Space

    Mohammad Mahdi Abootorabi, Armin Saghafian, Ali Bazshoushtari +5

    cs.CLcs.AIcs.LGarXiv:2609.06289v12026
  28. CLIP-Dissect: Automatic Description of Neuron Representations in Deep Vision Networks

    Tuomas Oikarinen, Tsui-Wei Weng

    cs.CVcs.AIcs.LGarXiv:2204.10965v52022
  29. Multi-Agent Adversarial Inverse Reinforcement Learning

    Lantao Yu, Jiaming Song, Stefano Ermon

    cs.LGstat.MLarXiv:1907.13220v12019
  30. CODA: A Real-World Road Corner Case Dataset for Object Detection in Autonomous Driving

    Kaican Li, Kai Chen, Haoyu Wang +10

    cs.CVcs.LGcs.ROarXiv:2203.07724v32022
  31. Cryptanalytic Extraction of Neural Network Models

    Nicholas Carlini, Matthew Jagielski, Ilya Mironov

    cs.LGcs.CRarXiv:2003.04884v22020
  32. SIRNN: A Math Library for Secure RNN Inference

    Deevashwer Rathee, Mayank Rathee, Rahul Kranti Kiran Goli +4

    cs.CRcs.LGcs.MSarXiv:2105.04236v12021
  33. DistGNN: Scalable Distributed Training for Large-Scale Graph Neural Networks

    Vasimuddin Md, Sanchit Misra, Guixiang Ma +6

    cs.LGcs.DCarXiv:2104.06700v32021
  34. Miles v0.1: Production-Level Post-Training

    RadixArk, :, Tom Chen +11

    cs.LGcs.CLarXiv:2609.08368v12026
  35. Interpretation and Generalization of Score Matching

    Siwei Lyu

    cs.LGstat.MLarXiv:1205.2629v12012
  36. TimeSHAP: Explaining Recurrent Models through Sequence Perturbations

    João Bento, Pedro Saleiro, André F. Cruz +2

    cs.LGcs.AIarXiv:2012.00073v22020
  37. Distributionally Robust Federated Averaging

    Yuyang Deng, Mohammad Mahdi Kamani, Mehrdad Mahdavi

    cs.LGcs.DCstat.MLarXiv:2102.12660v12021
  38. Black-box Explanation of Object Detectors via Saliency Maps

    Vitali Petsiuk, Rajiv Jain, Varun Manjunatha +4

    cs.CVcs.AIcs.LGarXiv:2006.03204v22020
  39. Robust Regression via Hard Thresholding

    Kush Bhatia, Prateek Jain, Purushottam Kar

    cs.LGstat.MLarXiv:1506.02428v12015
  40. TREC CAsT 2019: The Conversational Assistance Track Overview

    Jeffrey Dalton, Chenyan Xiong, Jamie Callan

    cs.IRcs.CLcs.LGarXiv:2003.13624v12020
  41. Learning and Evaluating Graph Neural Network Explanations based on Counterfactual and Factual Reasoning

    Juntao Tan, Shijie Geng, Zuohui Fu +4

    cs.IRcs.LGarXiv:2202.08816v32022
  42. Online Draft Co-Training for Speculative Decoding in Large-Scale, Long-Context RL Post-Training

    Zili Wang, Zhaopeng Qiu, Yuekai Zhang +2

    cs.LGcs.DCarXiv:2609.07108v12026
  43. On the Transfer of Inductive Bias from Simulation to the Real World: a New Disentanglement Dataset

    Muhammad Waleed Gondal, Manuel Wüthrich, Đorđe Miladinović +7

    stat.MLcs.LGarXiv:1906.03292v32019
  44. Environments as Scaffold: Enriching Feedback to Bootstrap Self-Evolving Agents in Long-Horizon Tasks

    Hongbang Yuan, Zhuoran Jin, Yixin Cao

    cs.LGcs.AIcs.CLarXiv:2609.08404v12026
  45. How to Exploit Hyperspherical Embeddings for Out-of-Distribution Detection?

    Yifei Ming, Yiyou Sun, Ousmane Dia +1

    cs.CVcs.LGarXiv:2203.04450v32022
  46. Biomedical Entity Representations with Synonym Marginalization

    Mujeen Sung, Hwisang Jeon, Jinhyuk Lee +1

    cs.CLcs.LGarXiv:2005.00239v12020
  47. Geometric Understanding of Deep Learning

    Na Lei, Zhongxuan Luo, Shing-Tung Yau +1

    cs.LGstat.MLarXiv:1805.10451v22018
  48. Triformer: Triangular, Variable-Specific Attentions for Long Sequence Multivariate Time Series Forecasting--Full Version

    Razvan-Gabriel Cirstea, Chenjuan Guo, Bin Yang +3

    cs.LGarXiv:2204.13767v12022
  49. Anti-DreamBooth: Protecting users from personalized text-to-image synthesis

    Thanh Van Le, Hao Phung, Thuan Hoang Nguyen +3

    cs.CVcs.CRcs.LGarXiv:2303.15433v22023
  50. Unsupervised Depth Completion from Visual Inertial Odometry

    Alex Wong, Xiaohan Fei, Stephanie Tsuei +1

    cs.CVcs.AIcs.LGarXiv:1905.08616v42019
  51. Improving Diffusion Inverse Problem Solving with Decoupled Noise Annealing

    Bingliang Zhang, Wenda Chu, Julius Berner +3

    cs.LGcs.AIcs.CVarXiv:2407.01521v32024
  52. Guaranteed Non-convex Optimization: Submodular Maximization over Continuous Domains

    Andrew An Bian, Baharan Mirzasoleiman, Joachim M. Buhmann +1

    cs.LGcs.DSarXiv:1606.05615v52016
  53. Low-Power Neuromorphic Hardware for Signal Processing Applications

    Bipin Rajendran, Abu Sebastian, Michael Schmuker +2

    cs.ETcs.LGcs.NEarXiv:1901.03690v32019
  54. Federated Learning: A Signal Processing Perspective

    Tomer Gafni, Nir Shlezinger, Kobi Cohen +2

    eess.SPcs.LGarXiv:2103.17150v22021
  55. Trainability of Dissipative Perceptron-Based Quantum Neural Networks

    Kunal Sharma, M. Cerezo, Lukasz Cincio +1

    quant-phcs.LGarXiv:2005.12458v22020
  56. Learning State Representations for Query Optimization with Deep Reinforcement Learning

    Jennifer Ortiz, Magdalena Balazinska, Johannes Gehrke +1

    cs.DBcs.AIcs.LGarXiv:1803.08604v12018
  57. SafeDrug: Dual Molecular Graph Encoders for Recommending Effective and Safe Drug Combinations

    Chaoqi Yang, Cao Xiao, Fenglong Ma +2

    cs.LGarXiv:2105.02711v22021
  58. Retro*: Learning Retrosynthetic Planning with Neural Guided A* Search

    Binghong Chen, Chengtao Li, Hanjun Dai +1

    cs.LGcs.AIstat.MLarXiv:2006.15820v12020
  59. Kalman Delta Networks: Uncertainty-aware Associative Memory

    Ngoc Bui, Tinglin Huang, Rex Ying

    cs.LGcs.AIarXiv:2609.07816v12026
  60. MOLE: Detecting Insider Threats in AI Agents

    Aashiq Muhamed, Virginia Smith

    cs.LGcs.CLcs.CRarXiv:2609.06966v12026