Machine Learning
Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
11,941 to 12,000 of 20,192
NEAT: Neural Attention Fields for End-to-End Autonomous Driving
Kashyap Chitta, Aditya Prakash, Andreas Geiger
cs.CVcs.AIcs.LGarXiv:2109.04456v12021The Role of ImageNet Classes in Fréchet Inception Distance
Tuomas Kynkäänniemi, Tero Karras, Miika Aittala +2
cs.CVcs.AIcs.LGarXiv:2203.06026v32022Mobile Sensor Data Anonymization
Mohammad Malekzadeh, Richard G. Clegg, Andrea Cavallaro +1
cs.LGstat.MLarXiv:1810.11546v32018An Emphatic Approach to the Problem of Off-policy Temporal-Difference Learning
Richard S. Sutton, A. Rupam Mahmood, Martha White
cs.LGarXiv:1503.04269v22015Trajectory balance: Improved credit assignment in GFlowNets
Nikolay Malkin, Moksh Jain, Emmanuel Bengio +2
cs.LGstat.MLarXiv:2201.13259v32022Reinforcement Learning for Robust Parameterized Locomotion Control of Bipedal Robots
Zhongyu Li, Xuxin Cheng, Xue Bin Peng +4
cs.ROcs.AIcs.LGarXiv:2103.14295v12021ETA Prediction with Graph Neural Networks in Google Maps
Austin Derrow-Pinion, Jennifer She, David Wong +14
cs.LGcs.AIcs.SIarXiv:2108.11482v12021GANs for Medical Image Synthesis: An Empirical Study
Youssef Skandarani, Pierre-Marc Jodoin, Alain Lalande
eess.IVcs.CVcs.LGarXiv:2105.05318v22021Parallel Multi-Dimensional LSTM, With Application to Fast Biomedical Volumetric Image Segmentation
Marijn F. Stollenga, Wonmin Byeon, Marcus Liwicki +1
cs.CVcs.LGarXiv:1506.07452v12015Learning to Fly by Crashing
Dhiraj Gandhi, Lerrel Pinto, Abhinav Gupta
cs.ROcs.CVcs.LGarXiv:1704.05588v22017Deep Denoising Neural Network Assisted Compressive Channel Estimation for mmWave Intelligent Reflecting Surfaces
Shicong Liu, Zhen Gao, Jun Zhang +2
cs.ITcs.LGeess.SParXiv:2006.02201v22020Iteratively Pruned Deep Learning Ensembles for COVID-19 Detection in Chest X-rays
Sivaramakrishnan Rajaraman, Jen Siegelman, Philip O. Alderson +3
eess.IVcs.CVcs.LGarXiv:2004.08379v32020Few-Shot Learning via Learning the Representation, Provably
Simon S. Du, Wei Hu, Sham M. Kakade +2
cs.LGmath.OCstat.MLarXiv:2002.09434v22020VisionZip: Longer is Better but Not Necessary in Vision Language Models
Senqiao Yang, Yukang Chen, Zhuotao Tian +4
cs.CVcs.AIcs.CLarXiv:2412.04467v22024From Variational to Deterministic Autoencoders
Partha Ghosh, Mehdi S. M. Sajjadi, Antonio Vergari +2
cs.LGstat.MLarXiv:1903.12436v42019Small-Object Detection in Remote Sensing Images with End-to-End Edge-Enhanced GAN and Object Detector Network
Jakaria Rabbi, Nilanjan Ray, Matthias Schubert +2
cs.CVcs.LGarXiv:2003.09085v52020U-NO: U-shaped Neural Operators
Md Ashiqur Rahman, Zachary E. Ross, Kamyar Azizzadenesheli
cs.LGarXiv:2204.11127v32022Self-Evaluation Guided Beam Search for Reasoning
Yuxi Xie, Kenji Kawaguchi, Yiran Zhao +4
cs.CLcs.AIcs.LGarXiv:2305.00633v32023On the Practical Computational Power of Finite Precision RNNs for Language Recognition
Gail Weiss, Yoav Goldberg, Eran Yahav
cs.LGcs.CLstat.MLarXiv:1805.04908v12018Combinatorial Network Optimization with Unknown Variables: Multi-Armed Bandits with Linear Rewards
Yi Gai, Bhaskar Krishnamachari, Rahul Jain
math.OCcs.LGcs.NIarXiv:1011.4748v12010Muppet: Massive Multi-task Representations with Pre-Finetuning
Armen Aghajanyan, Anchit Gupta, Akshat Shrivastava +3
cs.CLcs.LGarXiv:2101.11038v12021ROLAND: Graph Learning Framework for Dynamic Graphs
Jiaxuan You, Tianyu Du, Jure Leskovec
cs.LGcs.AIcs.SIarXiv:2208.07239v12022Zero-Shot Video Question Answering via Frozen Bidirectional Language Models
Antoine Yang, Antoine Miech, Josef Sivic +2
cs.CVcs.CLcs.LGarXiv:2206.08155v22022Deep Learning of Subsurface Flow via Theory-guided Neural Network
Nanzhe Wang, Dongxiao Zhang, Haibin Chang +1
cs.LGstat.MLarXiv:1911.00103v12019Physics-informed deep learning for incompressible laminar flows
Chengping Rao, Hao Sun, Yang Liu
physics.flu-dyncs.LGphysics.comp-pharXiv:2002.10558v22020Federated Learning from Pre-Trained Models: A Contrastive Learning Approach
Yue Tan, Guodong Long, Jie Ma +3
cs.CRcs.AIcs.LGarXiv:2209.10083v12022Offline-to-Online Reinforcement Learning via Balanced Replay and Pessimistic Q-Ensemble
Seunghyun Lee, Younggyo Seo, Kimin Lee +2
cs.ROcs.LGarXiv:2107.00591v22021An elementary introduction to information geometry
Frank Nielsen
cs.LGcs.ITstat.MLarXiv:1808.08271v22018Learning in a Large Function Space: Privacy-Preserving Mechanisms for SVM Learning
Benjamin I. P. Rubinstein, Peter L. Bartlett, Ling Huang +1
cs.LGcs.CRcs.DBarXiv:0911.5708v12009AI Safety Gridworlds
Jan Leike, Miljan Martic, Victoria Krakovna +5
cs.LGcs.AIarXiv:1711.09883v22017Conditional Diffusion Probabilistic Model for Speech Enhancement
Yen-Ju Lu, Zhong-Qiu Wang, Shinji Watanabe +3
eess.AScs.LGcs.SDarXiv:2202.05256v12022Partial success in closing the gap between human and machine vision
Robert Geirhos, Kantharaju Narayanappa, Benjamin Mitzkus +4
cs.CVcs.AIcs.LGarXiv:2106.07411v22021Non-convex Robust PCA
Praneeth Netrapalli, U N Niranjan, Sujay Sanghavi +2
cs.ITcs.LGstat.MLarXiv:1410.7660v12014Hierarchical Open-Vocabulary 3D Scene Graphs for Language-Grounded Robot Navigation
Abdelrhman Werby, Chenguang Huang, Martin Büchner +2
cs.ROcs.AIcs.CLarXiv:2403.17846v22024Compositional Fairness Constraints for Graph Embeddings
Avishek Joey Bose, William L. Hamilton
cs.LGcs.AIstat.MLarXiv:1905.10674v42019On Mean Absolute Error for Deep Neural Network Based Vector-to-Vector Regression
Jun Qi, Jun Du, Sabato Marco Siniscalchi +2
eess.AScs.LGcs.SDarXiv:2008.07281v12020ReDMark: Framework for Residual Diffusion Watermarking on Deep Networks
Mahdi Ahmadi, Alireza Norouzi, S. M. Reza Soroushmehr +4
cs.MMcs.CRcs.LGarXiv:1810.07248v32018Signal Recovery on Graphs: Variation Minimization
Siheng Chen, Aliaksei Sandryhaila, José M. F. Moura +1
cs.SIcs.LGstat.MLarXiv:1411.7414v32014MobileLLM: Optimizing Sub-billion Parameter Language Models for On-Device Use Cases
Zechun Liu, Changsheng Zhao, Forrest Iandola +9
cs.LGcs.AIcs.CLarXiv:2402.14905v22024FedProc: Prototypical Contrastive Federated Learning on Non-IID data
Xutong Mu, Yulong Shen, Ke Cheng +4
cs.LGcs.DCarXiv:2109.12273v12021Relational Neural Expectation Maximization: Unsupervised Discovery of Objects and their Interactions
Sjoerd van Steenkiste, Michael Chang, Klaus Greff +1
cs.LGcs.AIcs.NEarXiv:1802.10353v12018Investigating Bi-Level Optimization for Learning and Vision from a Unified Perspective: A Survey and Beyond
Risheng Liu, Jiaxin Gao, Jin Zhang +2
cs.LGcs.CVmath.DSarXiv:2101.11517v32021Editing Conditional Radiance Fields
Steven Liu, Xiuming Zhang, Zhoutong Zhang +3
cs.CVcs.GRcs.LGarXiv:2105.06466v22021Lyapunov-based Safe Policy Optimization for Continuous Control
Yinlam Chow, Ofir Nachum, Aleksandra Faust +2
cs.LGcs.AIstat.MLarXiv:1901.10031v22019SpeechT5: Unified-Modal Encoder-Decoder Pre-Training for Spoken Language Processing
Junyi Ao, Rui Wang, Long Zhou +11
eess.AScs.CLcs.LGarXiv:2110.07205v32021Learning to Self-Train for Semi-Supervised Few-Shot Classification
Xinzhe Li, Qianru Sun, Yaoyao Liu +4
cs.CVcs.LGstat.MLarXiv:1906.00562v22019Portuguese Named Entity Recognition using BERT-CRF
Fábio Souza, Rodrigo Nogueira, Roberto Lotufo
cs.CLcs.IRcs.LGarXiv:1909.10649v22019ToolkenGPT: Augmenting Frozen Language Models with Massive Tools via Tool Embeddings
Shibo Hao, Tianyang Liu, Zhen Wang +1
cs.CLcs.LGarXiv:2305.11554v42023MVTN: Multi-View Transformation Network for 3D Shape Recognition
Abdullah Hamdi, Silvio Giancola, Bernard Ghanem
cs.CVcs.LGarXiv:2011.13244v32020Video Object Segmentation with Episodic Graph Memory Networks
Xiankai Lu, Wenguan Wang, Martin Danelljan +3
cs.CVcs.LGarXiv:2007.07020v42020Tensor Programs V: Tuning Large Neural Networks via Zero-Shot Hyperparameter Transfer
Greg Yang, Edward J. Hu, Igor Babuschkin +7
cs.LGcond-mat.dis-nncs.NEarXiv:2203.03466v22022Distributed Autonomous Online Learning: Regrets and Intrinsic Privacy-Preserving Properties
Feng Yan, Shreyas Sundaram, S. V. N. Vishwanathan +1
cs.LGcs.AIarXiv:1006.4039v32010Neural Architecture Search on ImageNet in Four GPU Hours: A Theoretically Inspired Perspective
Wuyang Chen, Xinyu Gong, Zhangyang Wang
cs.CVcs.LGarXiv:2102.11535v42021Delayed Optimizer-State Transport Shapes Short-Horizon Training Decisions
Jinhui Guo
cs.LGphysics.comp-pharXiv:2608.24593v12026Renormalization Group Flow Matching for Scalable Local Generative Modeling
Kanta Masuki, Yuto Ashida
cs.LGcond-mat.stat-mecharXiv:2608.23696v12026Probability-Preserving Transformer for the Time-Dependent Schrödinger Equation
Mushtaq Ali, Muzamil Tariq, Niaz Ali Khan
cs.LGmath-phquant-pharXiv:2608.15112v12026Pairton: Iterative Reconstruction of Short-Lived Particles
Andreas Hermansen, Chris Scheulen, Tobias Golling
hep-phcs.LGhep-exarXiv:2608.14278v12026Contrastive Learning for Interpretable Anomaly Detection at Collider Experiments
Haoyi Jia, Sagar Addepalli, Julia Gonski
cs.LGhep-exhep-pharXiv:2608.13652v12026The Geometry of Semantic Space: A Continuous Geometric Framework for the Transformer Architecture
Zhihua Liang
cond-mat.dis-nncs.CLcs.LGarXiv:2607.17146v12026Challenges and Opportunities in Quantum Machine Learning
M. Cerezo, Guillaume Verdon, Hsin-Yuan Huang +2
quant-phcs.LGstat.MLarXiv:2303.09491v12023