Machine Learning
Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
3,421 to 3,480 of 20,193
Retrieval-Augmented Multimodal Language Modeling
Michihiro Yasunaga, Armen Aghajanyan, Weijia Shi +6
cs.CVcs.CLcs.LGarXiv:2211.12561v22022Proximal Algorithms in Statistics and Machine Learning
Nicholas G. Polson, James G. Scott, Brandon T. Willard
stat.MLcs.LGstat.MEarXiv:1502.03175v32015Deep belief networks are exact
Gleb Smirnov
cs.AIcs.LGmath.PRarXiv:2609.05572v12026Adversarial and Clean Data Are Not Twins
Zhitao Gong, Wenlu Wang, Wei-Shinn Ku
cs.LGcs.NEarXiv:1704.04960v12017Sales Demand Forecast in E-commerce using a Long Short-Term Memory Neural Network Methodology
Kasun Bandara, Peibei Shi, Christoph Bergmeir +3
cs.LGstat.MLarXiv:1901.04028v22019Large Language Models as Urban Residents: An LLM Agent Framework for Personal Mobility Generation
Jiawei Wang, Renhe Jiang, Chuang Yang +5
cs.AIcs.CLcs.CYarXiv:2402.14744v32024The Proper Care and Feeding of CAMELS: How Limited Training Data Affects Streamflow Prediction
Martin Gauch, Juliane Mai, Jimmy Lin
cs.LGstat.MLarXiv:1911.07249v32019When Benchmarks are Targets: Revealing the Sensitivity of Large Language Model Leaderboards
Norah Alzahrani, Hisham Abdullah Alyahya, Yazeed Alnumay +9
cs.CLcs.AIcs.LGarXiv:2402.01781v22024Red-Teaming for Generative AI: Silver Bullet or Security Theater?
Michael Feffer, Anusha Sinha, Wesley Hanwen Deng +2
cs.CYcs.HCcs.LGarXiv:2401.15897v32024Continual Learning with Node-Importance based Adaptive Group Sparse Regularization
Sangwon Jung, Hongjoon Ahn, Sungmin Cha +1
cs.LGstat.MLarXiv:2003.13726v42020CapsuleGAN: Generative Adversarial Capsule Network
Ayush Jaiswal, Wael AbdAlmageed, Yue Wu +1
stat.MLcs.LGarXiv:1802.06167v72018Real-time Faulted Line Localization and PMU Placement in Power Systems through Convolutional Neural Networks
Wenting Li, Deepjyoti Deka, Michael Chertkov +1
eess.SYcs.LGstat.MLarXiv:1810.05247v22018Dynamic Pricing with Limited Supply
Moshe Babaioff, Shaddin Dughmi, Robert Kleinberg +1
cs.GTcs.DScs.LGarXiv:1108.4142v32011Sparse DNNs with Improved Adversarial Robustness
Yiwen Guo, Chao Zhang, Changshui Zhang +1
cs.LGcs.CRcs.CVarXiv:1810.09619v22018Omni Interaction Agent Technical Report
Orantqing, Shengpeng Ji, Junlong Tong +20
eess.AScs.AIcs.LGarXiv:2609.08977v12026One-Shot Learning of Manipulation Skills with Online Dynamics Adaptation and Neural Network Priors
Justin Fu, Sergey Levine, Pieter Abbeel
cs.LGcs.ROarXiv:1509.06841v32015Strategies and Principles of Distributed Machine Learning on Big Data
Eric P. Xing, Qirong Ho, Pengtao Xie +1
stat.MLcs.DCcs.LGarXiv:1512.09295v12015When and What to Teach: Budget-Aware Online Adaptation for Web Agents
Jianwei Zhang, Sihan Cao, Pengcheng Zheng +7
cs.AIcs.CVcs.LGarXiv:2609.05513v12026Toward Understanding the Feature Learning Process of Self-supervised Contrastive Learning
Zixin Wen, Yuanzhi Li
cs.LGcs.CVstat.MLarXiv:2105.15134v32021Word2Vec applied to Recommendation: Hyperparameters Matter
Hugo Caselles-Dupré, Florian Lesaint, Jimena Royo-Letelier
cs.IRcs.CLcs.LGarXiv:1804.04212v32018On Robustness of Neural Ordinary Differential Equations
Hanshu Yan, Jiawei Du, Vincent Y. F. Tan +1
cs.LGstat.MLarXiv:1910.05513v42019Fighting Offensive Language on Social Media with Unsupervised Text Style Transfer
Cicero Nogueira dos Santos, Igor Melnyk, Inkit Padhi
cs.CLcs.LGarXiv:1805.07685v12018FedAT: A High-Performance and Communication-Efficient Federated Learning System with Asynchronous Tiers
Zheng Chai, Yujing Chen, Ali Anwar +3
cs.DCcs.LGcs.NIarXiv:2010.05958v22020Probabilistic Recursive Reasoning for Multi-Agent Reinforcement Learning
Ying Wen, Yaodong Yang, Rui Luo +2
cs.LGcs.AIstat.MLarXiv:1901.09207v22019Parallel Predictive Entropy Search for Batch Global Optimization of Expensive Objective Functions
Amar Shah, Zoubin Ghahramani
cs.LGstat.MLarXiv:1511.07130v12015Lorentz Group Equivariant Neural Network for Particle Physics
Alexander Bogatskiy, Brandon Anderson, Jan T. Offermann +3
hep-phcs.LGhep-exarXiv:2006.04780v12020ARC-Bench: Closed-Loop Replanning Masks Broken Action Ranking in Frozen JEPA World Models
Zhengshu Zhang, Zhiyuan Li
cs.AIcs.LGcs.ROarXiv:2609.05461v12026GEP-PG: Decoupling Exploration and Exploitation in Deep Reinforcement Learning Algorithms
Cédric Colas, Olivier Sigaud, Pierre-Yves Oudeyer
cs.LGarXiv:1802.05054v52018RecoGym: A Reinforcement Learning Environment for the problem of Product Recommendation in Online Advertising
David Rohde, Stephen Bonner, Travis Dunlop +2
cs.IRcs.LGarXiv:1808.00720v22018Learning From Noisy Singly-labeled Data
Ashish Khetan, Zachary C. Lipton, Anima Anandkumar
cs.LGarXiv:1712.04577v22017Generalization Error Bounds of Gradient Descent for Learning Over-parameterized Deep ReLU Networks
Yuan Cao, Quanquan Gu
cs.LGmath.OCstat.MLarXiv:1902.01384v42019Self-supervised Feature Learning for 3D Medical Images by Playing a Rubik's Cube
Xinrui Zhuang, Yuexiang Li, Yifan Hu +3
cs.CVcs.LGeess.IVarXiv:1910.02241v12019Neural Architecture Transfer
Zhichao Lu, Gautam Sreekumar, Erik Goodman +3
cs.CVcs.LGcs.NEarXiv:2005.05859v22020Woulda, Coulda, Shoulda: Counterfactually-Guided Policy Search
Lars Buesing, Theophane Weber, Yori Zwols +4
cs.LGstat.MLarXiv:1811.06272v12018Reason Through the Latent! Making Latent Visual Reasoning Necessary
Suhyeong Park, Junha Jung, Jaewoo Kang
cs.AIcs.CLcs.CVarXiv:2609.06746v12026Eliciting Weak-to-Strong Generalization with On-Policy Reverse Distillation
Youngrok Park, Sangmin Bae, Hojung Jung +6
cs.LGcs.CLarXiv:2609.08798v12026One Loss for All: Deep Hashing with a Single Cosine Similarity based Learning Objective
Jiun Tian Hoe, Kam Woh Ng, Tianyu Zhang +3
cs.CVcs.LGarXiv:2109.14449v12021Deep Signature Transforms
Patric Bonnier, Patrick Kidger, Imanol Perez Arribas +2
cs.LGstat.MLarXiv:1905.08494v22019Navigation with Large Language Models: Semantic Guesswork as a Heuristic for Planning
Dhruv Shah, Michael Equi, Blazej Osinski +3
cs.ROcs.AIcs.CLarXiv:2310.10103v12023Damage-Aware Bandit Pruning for Vision and Language Transformers
Salem Ameen, Sunil Vadera
cs.AIcs.LGarXiv:2609.05448v12026A geometric alternative to Nesterov's accelerated gradient descent
Sébastien Bubeck, Yin Tat Lee, Mohit Singh
math.OCcs.DScs.LGarXiv:1506.08187v12015Boosting Adversarial Training with Hypersphere Embedding
Tianyu Pang, Xiao Yang, Yinpeng Dong +3
cs.LGcs.CRcs.CVarXiv:2002.08619v32020General-purpose Tagging of Freesound Audio with AudioSet Labels: Task Description, Dataset, and Baseline
Eduardo Fonseca, Manoj Plakal, Frederic Font +4
cs.SDcs.LGeess.ASarXiv:1807.09902v32018Minimizing Energy Consumption Leads to the Emergence of Gaits in Legged Robots
Zipeng Fu, Ashish Kumar, Jitendra Malik +1
cs.ROcs.AIcs.CVarXiv:2111.01674v12021A Topology Layer for Machine Learning
Rickard Brüel-Gabrielsson, Bradley J. Nelson, Anjan Dwaraknath +3
cs.LGmath.ATstat.MLarXiv:1905.12200v22019Sampling-based sublinear low-rank matrix arithmetic framework for dequantizing quantum machine learning
Nai-Hui Chia, András Gilyén, Tongyang Li +3
cs.DScs.LGquant-pharXiv:1910.06151v42019Deep Metric Learning for Practical Person Re-Identification
Dong Yi, Zhen Lei, Stan Z. Li
cs.CVcs.LGcs.NEarXiv:1407.4979v12014High-dimensional Asymptotics of Feature Learning: How One Gradient Step Improves the Representation
Jimmy Ba, Murat A. Erdogdu, Taiji Suzuki +3
stat.MLcs.LGmath.STarXiv:2205.01445v12022Generalized Shape Metrics on Neural Representations
Alex H. Williams, Erin Kunz, Simon Kornblith +1
stat.MLcs.LGarXiv:2110.14739v22021Steering Geometry: Validating Human Value Geometry in LLM Steering Space
Mohammad Mahdi Abootorabi, Armin Saghafian, Ali Bazshoushtari +5
cs.CLcs.AIcs.LGarXiv:2609.06289v12026CLIP-Dissect: Automatic Description of Neuron Representations in Deep Vision Networks
Tuomas Oikarinen, Tsui-Wei Weng
cs.CVcs.AIcs.LGarXiv:2204.10965v52022Multi-Agent Adversarial Inverse Reinforcement Learning
Lantao Yu, Jiaming Song, Stefano Ermon
cs.LGstat.MLarXiv:1907.13220v12019CODA: A Real-World Road Corner Case Dataset for Object Detection in Autonomous Driving
Kaican Li, Kai Chen, Haoyu Wang +10
cs.CVcs.LGcs.ROarXiv:2203.07724v32022Cryptanalytic Extraction of Neural Network Models
Nicholas Carlini, Matthew Jagielski, Ilya Mironov
cs.LGcs.CRarXiv:2003.04884v22020SIRNN: A Math Library for Secure RNN Inference
Deevashwer Rathee, Mayank Rathee, Rahul Kranti Kiran Goli +4
cs.CRcs.LGcs.MSarXiv:2105.04236v12021DistGNN: Scalable Distributed Training for Large-Scale Graph Neural Networks
Vasimuddin Md, Sanchit Misra, Guixiang Ma +6
cs.LGcs.DCarXiv:2104.06700v32021Miles v0.1: Production-Level Post-Training
RadixArk, :, Tom Chen +11
cs.LGcs.CLarXiv:2609.08368v12026Interpretation and Generalization of Score Matching
Siwei Lyu
cs.LGstat.MLarXiv:1205.2629v12012TimeSHAP: Explaining Recurrent Models through Sequence Perturbations
João Bento, Pedro Saleiro, André F. Cruz +2
cs.LGcs.AIarXiv:2012.00073v22020Distributionally Robust Federated Averaging
Yuyang Deng, Mohammad Mahdi Kamani, Mehrdad Mahdavi
cs.LGcs.DCstat.MLarXiv:2102.12660v12021