Machine Learning

Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

3,421 to 3,480 of 20,193

  1. Retrieval-Augmented Multimodal Language Modeling

    Michihiro Yasunaga, Armen Aghajanyan, Weijia Shi +6

    cs.CVcs.CLcs.LGarXiv:2211.12561v22022
  2. Proximal Algorithms in Statistics and Machine Learning

    Nicholas G. Polson, James G. Scott, Brandon T. Willard

    stat.MLcs.LGstat.MEarXiv:1502.03175v32015
  3. Deep belief networks are exact

    Gleb Smirnov

    cs.AIcs.LGmath.PRarXiv:2609.05572v12026
  4. Adversarial and Clean Data Are Not Twins

    Zhitao Gong, Wenlu Wang, Wei-Shinn Ku

    cs.LGcs.NEarXiv:1704.04960v12017
  5. Sales Demand Forecast in E-commerce using a Long Short-Term Memory Neural Network Methodology

    Kasun Bandara, Peibei Shi, Christoph Bergmeir +3

    cs.LGstat.MLarXiv:1901.04028v22019
  6. Large Language Models as Urban Residents: An LLM Agent Framework for Personal Mobility Generation

    Jiawei Wang, Renhe Jiang, Chuang Yang +5

    cs.AIcs.CLcs.CYarXiv:2402.14744v32024
  7. The Proper Care and Feeding of CAMELS: How Limited Training Data Affects Streamflow Prediction

    Martin Gauch, Juliane Mai, Jimmy Lin

    cs.LGstat.MLarXiv:1911.07249v32019
  8. When Benchmarks are Targets: Revealing the Sensitivity of Large Language Model Leaderboards

    Norah Alzahrani, Hisham Abdullah Alyahya, Yazeed Alnumay +9

    cs.CLcs.AIcs.LGarXiv:2402.01781v22024
  9. Red-Teaming for Generative AI: Silver Bullet or Security Theater?

    Michael Feffer, Anusha Sinha, Wesley Hanwen Deng +2

    cs.CYcs.HCcs.LGarXiv:2401.15897v32024
  10. Continual Learning with Node-Importance based Adaptive Group Sparse Regularization

    Sangwon Jung, Hongjoon Ahn, Sungmin Cha +1

    cs.LGstat.MLarXiv:2003.13726v42020
  11. CapsuleGAN: Generative Adversarial Capsule Network

    Ayush Jaiswal, Wael AbdAlmageed, Yue Wu +1

    stat.MLcs.LGarXiv:1802.06167v72018
  12. Real-time Faulted Line Localization and PMU Placement in Power Systems through Convolutional Neural Networks

    Wenting Li, Deepjyoti Deka, Michael Chertkov +1

    eess.SYcs.LGstat.MLarXiv:1810.05247v22018
  13. Dynamic Pricing with Limited Supply

    Moshe Babaioff, Shaddin Dughmi, Robert Kleinberg +1

    cs.GTcs.DScs.LGarXiv:1108.4142v32011
  14. Sparse DNNs with Improved Adversarial Robustness

    Yiwen Guo, Chao Zhang, Changshui Zhang +1

    cs.LGcs.CRcs.CVarXiv:1810.09619v22018
  15. Omni Interaction Agent Technical Report

    Orantqing, Shengpeng Ji, Junlong Tong +20

    eess.AScs.AIcs.LGarXiv:2609.08977v12026
  16. One-Shot Learning of Manipulation Skills with Online Dynamics Adaptation and Neural Network Priors

    Justin Fu, Sergey Levine, Pieter Abbeel

    cs.LGcs.ROarXiv:1509.06841v32015
  17. Strategies and Principles of Distributed Machine Learning on Big Data

    Eric P. Xing, Qirong Ho, Pengtao Xie +1

    stat.MLcs.DCcs.LGarXiv:1512.09295v12015
  18. When and What to Teach: Budget-Aware Online Adaptation for Web Agents

    Jianwei Zhang, Sihan Cao, Pengcheng Zheng +7

    cs.AIcs.CVcs.LGarXiv:2609.05513v12026
  19. Toward Understanding the Feature Learning Process of Self-supervised Contrastive Learning

    Zixin Wen, Yuanzhi Li

    cs.LGcs.CVstat.MLarXiv:2105.15134v32021
  20. Word2Vec applied to Recommendation: Hyperparameters Matter

    Hugo Caselles-Dupré, Florian Lesaint, Jimena Royo-Letelier

    cs.IRcs.CLcs.LGarXiv:1804.04212v32018
  21. On Robustness of Neural Ordinary Differential Equations

    Hanshu Yan, Jiawei Du, Vincent Y. F. Tan +1

    cs.LGstat.MLarXiv:1910.05513v42019
  22. Fighting Offensive Language on Social Media with Unsupervised Text Style Transfer

    Cicero Nogueira dos Santos, Igor Melnyk, Inkit Padhi

    cs.CLcs.LGarXiv:1805.07685v12018
  23. FedAT: A High-Performance and Communication-Efficient Federated Learning System with Asynchronous Tiers

    Zheng Chai, Yujing Chen, Ali Anwar +3

    cs.DCcs.LGcs.NIarXiv:2010.05958v22020
  24. Probabilistic Recursive Reasoning for Multi-Agent Reinforcement Learning

    Ying Wen, Yaodong Yang, Rui Luo +2

    cs.LGcs.AIstat.MLarXiv:1901.09207v22019
  25. Parallel Predictive Entropy Search for Batch Global Optimization of Expensive Objective Functions

    Amar Shah, Zoubin Ghahramani

    cs.LGstat.MLarXiv:1511.07130v12015
  26. Lorentz Group Equivariant Neural Network for Particle Physics

    Alexander Bogatskiy, Brandon Anderson, Jan T. Offermann +3

    hep-phcs.LGhep-exarXiv:2006.04780v12020
  27. ARC-Bench: Closed-Loop Replanning Masks Broken Action Ranking in Frozen JEPA World Models

    Zhengshu Zhang, Zhiyuan Li

    cs.AIcs.LGcs.ROarXiv:2609.05461v12026
  28. GEP-PG: Decoupling Exploration and Exploitation in Deep Reinforcement Learning Algorithms

    Cédric Colas, Olivier Sigaud, Pierre-Yves Oudeyer

    cs.LGarXiv:1802.05054v52018
  29. RecoGym: A Reinforcement Learning Environment for the problem of Product Recommendation in Online Advertising

    David Rohde, Stephen Bonner, Travis Dunlop +2

    cs.IRcs.LGarXiv:1808.00720v22018
  30. Learning From Noisy Singly-labeled Data

    Ashish Khetan, Zachary C. Lipton, Anima Anandkumar

    cs.LGarXiv:1712.04577v22017
  31. Generalization Error Bounds of Gradient Descent for Learning Over-parameterized Deep ReLU Networks

    Yuan Cao, Quanquan Gu

    cs.LGmath.OCstat.MLarXiv:1902.01384v42019
  32. Self-supervised Feature Learning for 3D Medical Images by Playing a Rubik's Cube

    Xinrui Zhuang, Yuexiang Li, Yifan Hu +3

    cs.CVcs.LGeess.IVarXiv:1910.02241v12019
  33. Neural Architecture Transfer

    Zhichao Lu, Gautam Sreekumar, Erik Goodman +3

    cs.CVcs.LGcs.NEarXiv:2005.05859v22020
  34. Woulda, Coulda, Shoulda: Counterfactually-Guided Policy Search

    Lars Buesing, Theophane Weber, Yori Zwols +4

    cs.LGstat.MLarXiv:1811.06272v12018
  35. Reason Through the Latent! Making Latent Visual Reasoning Necessary

    Suhyeong Park, Junha Jung, Jaewoo Kang

    cs.AIcs.CLcs.CVarXiv:2609.06746v12026
  36. Eliciting Weak-to-Strong Generalization with On-Policy Reverse Distillation

    Youngrok Park, Sangmin Bae, Hojung Jung +6

    cs.LGcs.CLarXiv:2609.08798v12026
  37. One Loss for All: Deep Hashing with a Single Cosine Similarity based Learning Objective

    Jiun Tian Hoe, Kam Woh Ng, Tianyu Zhang +3

    cs.CVcs.LGarXiv:2109.14449v12021
  38. Deep Signature Transforms

    Patric Bonnier, Patrick Kidger, Imanol Perez Arribas +2

    cs.LGstat.MLarXiv:1905.08494v22019
  39. Navigation with Large Language Models: Semantic Guesswork as a Heuristic for Planning

    Dhruv Shah, Michael Equi, Blazej Osinski +3

    cs.ROcs.AIcs.CLarXiv:2310.10103v12023
  40. Damage-Aware Bandit Pruning for Vision and Language Transformers

    Salem Ameen, Sunil Vadera

    cs.AIcs.LGarXiv:2609.05448v12026
  41. A geometric alternative to Nesterov's accelerated gradient descent

    Sébastien Bubeck, Yin Tat Lee, Mohit Singh

    math.OCcs.DScs.LGarXiv:1506.08187v12015
  42. Boosting Adversarial Training with Hypersphere Embedding

    Tianyu Pang, Xiao Yang, Yinpeng Dong +3

    cs.LGcs.CRcs.CVarXiv:2002.08619v32020
  43. General-purpose Tagging of Freesound Audio with AudioSet Labels: Task Description, Dataset, and Baseline

    Eduardo Fonseca, Manoj Plakal, Frederic Font +4

    cs.SDcs.LGeess.ASarXiv:1807.09902v32018
  44. Minimizing Energy Consumption Leads to the Emergence of Gaits in Legged Robots

    Zipeng Fu, Ashish Kumar, Jitendra Malik +1

    cs.ROcs.AIcs.CVarXiv:2111.01674v12021
  45. A Topology Layer for Machine Learning

    Rickard Brüel-Gabrielsson, Bradley J. Nelson, Anjan Dwaraknath +3

    cs.LGmath.ATstat.MLarXiv:1905.12200v22019
  46. Sampling-based sublinear low-rank matrix arithmetic framework for dequantizing quantum machine learning

    Nai-Hui Chia, András Gilyén, Tongyang Li +3

    cs.DScs.LGquant-pharXiv:1910.06151v42019
  47. Deep Metric Learning for Practical Person Re-Identification

    Dong Yi, Zhen Lei, Stan Z. Li

    cs.CVcs.LGcs.NEarXiv:1407.4979v12014
  48. High-dimensional Asymptotics of Feature Learning: How One Gradient Step Improves the Representation

    Jimmy Ba, Murat A. Erdogdu, Taiji Suzuki +3

    stat.MLcs.LGmath.STarXiv:2205.01445v12022
  49. Generalized Shape Metrics on Neural Representations

    Alex H. Williams, Erin Kunz, Simon Kornblith +1

    stat.MLcs.LGarXiv:2110.14739v22021
  50. Steering Geometry: Validating Human Value Geometry in LLM Steering Space

    Mohammad Mahdi Abootorabi, Armin Saghafian, Ali Bazshoushtari +5

    cs.CLcs.AIcs.LGarXiv:2609.06289v12026
  51. CLIP-Dissect: Automatic Description of Neuron Representations in Deep Vision Networks

    Tuomas Oikarinen, Tsui-Wei Weng

    cs.CVcs.AIcs.LGarXiv:2204.10965v52022
  52. Multi-Agent Adversarial Inverse Reinforcement Learning

    Lantao Yu, Jiaming Song, Stefano Ermon

    cs.LGstat.MLarXiv:1907.13220v12019
  53. CODA: A Real-World Road Corner Case Dataset for Object Detection in Autonomous Driving

    Kaican Li, Kai Chen, Haoyu Wang +10

    cs.CVcs.LGcs.ROarXiv:2203.07724v32022
  54. Cryptanalytic Extraction of Neural Network Models

    Nicholas Carlini, Matthew Jagielski, Ilya Mironov

    cs.LGcs.CRarXiv:2003.04884v22020
  55. SIRNN: A Math Library for Secure RNN Inference

    Deevashwer Rathee, Mayank Rathee, Rahul Kranti Kiran Goli +4

    cs.CRcs.LGcs.MSarXiv:2105.04236v12021
  56. DistGNN: Scalable Distributed Training for Large-Scale Graph Neural Networks

    Vasimuddin Md, Sanchit Misra, Guixiang Ma +6

    cs.LGcs.DCarXiv:2104.06700v32021
  57. Miles v0.1: Production-Level Post-Training

    RadixArk, :, Tom Chen +11

    cs.LGcs.CLarXiv:2609.08368v12026
  58. Interpretation and Generalization of Score Matching

    Siwei Lyu

    cs.LGstat.MLarXiv:1205.2629v12012
  59. TimeSHAP: Explaining Recurrent Models through Sequence Perturbations

    João Bento, Pedro Saleiro, André F. Cruz +2

    cs.LGcs.AIarXiv:2012.00073v22020
  60. Distributionally Robust Federated Averaging

    Yuyang Deng, Mohammad Mahdi Kamani, Mehrdad Mahdavi

    cs.LGcs.DCstat.MLarXiv:2102.12660v12021