Machine Learning
Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
3,301 to 3,360 of 20,454
Grounding Language with Visual Affordances over Unstructured Data
Oier Mees, Jessica Borja-Diaz, Wolfram Burgard
cs.ROcs.AIcs.CLarXiv:2210.01911v32022Investigating the Limitations of Transformers with Simple Arithmetic Tasks
Rodrigo Nogueira, Zhiying Jiang, Jimmy Lin
cs.CLcs.AIcs.LGarXiv:2102.13019v32021All for 1-Bit: Towards Genuine 1-Bit Post-Training Quantization for LLMs
Zhixiong Zhao, Zukang Xu, Guangyu Sun +2
cs.LGcs.AIarXiv:2609.06161v12026An Attentive Inductive Bias for Sequential Recommendation beyond the Self-Attention
Yehjin Shin, Jeongwhan Choi, Hyowon Wi +1
cs.LGcs.AIcs.IRarXiv:2312.10325v22023Subdivision-Based Mesh Convolution Networks
Shi-Min Hu, Zheng-Ning Liu, Meng-Hao Guo +4
cs.CVcs.GRcs.LGarXiv:2106.02285v22021Topology-Aware Correlations Between Relations for Inductive Link Prediction in Knowledge Graphs
Jiajun Chen, Huarui He, Feng Wu +1
cs.LGcs.CLarXiv:2103.03642v12021Learning Algorithms for Active Learning
Philip Bachman, Alessandro Sordoni, Adam Trischler
cs.LGarXiv:1708.00088v12017Decision-Aware Suffix Prediction and Reasoning of Business Processes
Henryk Mustroph, Stefanie Rinderle-Ma
cs.LGcs.AIarXiv:2609.06169v12026Instance-Conditioned GAN
Arantxa Casanova, Marlène Careil, Jakob Verbeek +2
cs.CVcs.LGarXiv:2109.05070v22021Reconstruction of Markov Random Fields from Samples: Some Easy Observations and Algorithms
Guy Bresler, Elchanan Mossel, Allan Sly
cs.CCcs.LGarXiv:0712.1402v22007Contrastive Predictive Coding for Human Activity Recognition
Harish Haresamudram, Irfan Essa, Thomas Ploetz
cs.LGarXiv:2012.05333v12020Permutation invariant graph-to-sequence model for template-free retrosynthesis and reaction prediction
Zhengkai Tu, Connor W. Coley
cs.LGarXiv:2110.09681v12021Hyperspherical Prototype Networks
Pascal Mettes, Elise van der Pol, Cees G. M. Snoek
cs.LGstat.MLarXiv:1901.10514v32019Matérn Gaussian processes on Riemannian manifolds
Viacheslav Borovitskiy, Alexander Terenin, Peter Mostowsky +1
stat.MLcs.LGarXiv:2006.10160v62020Adversarial Attacks on Deep Neural Networks for Time Series Classification
Hassan Ismail Fawaz, Germain Forestier, Jonathan Weber +2
cs.LGcs.CRstat.MLarXiv:1903.07054v22019Fairness-aware Agnostic Federated Learning
Wei Du, Depeng Xu, Xintao Wu +1
cs.LGcs.CYarXiv:2010.05057v12020Stable Learning via Sample Reweighting
Zheyan Shen, Peng Cui, Tong Zhang +1
cs.LGstat.MLarXiv:1911.12580v12019Machine learning of hierarchical clustering to segment 2D and 3D images
Juan Nunez-Iglesias, Ryan Kennedy, Toufiq Parag +2
cs.CVcs.LGarXiv:1303.6163v32013Non-Autoregressive Machine Translation with Auxiliary Regularization
Yiren Wang, Fei Tian, Di He +3
cs.CLcs.LGstat.MLarXiv:1902.10245v12019What the Window Does Not Contain: Auditing Provenance in a Document-Grounded Instability Benchmark
Seyed Mosayeb Alam
cs.CLcs.AIcs.LGarXiv:2609.06147v12026Online Learning for Time Series Prediction
Oren Anava, Elad Hazan, Shie Mannor +1
cs.LGarXiv:1302.6927v12013POCO: Point Convolution for Surface Reconstruction
Alexandre Boulch, Renaud Marlet
cs.CVcs.CGcs.LGarXiv:2201.01831v22022Bilinear Factor Matrix Norm Minimization for Robust PCA: Algorithms and Applications
Fanhua Shang, James Cheng, Yuanyuan Liu +2
cs.LGcs.CVmath.OCarXiv:1810.05186v12018Summaries:한국어High-Throughput CNN Inference on Embedded ARM big.LITTLE Multi-Core Processors
Siqi Wang, Gayathri Ananthanarayanan, Yifan Zeng +3
cs.LGcs.DCcs.PFarXiv:1903.05898v32019Summaries:한국어An Information-Theoretic Approach to Transferability in Task Transfer Learning
Yajie Bao, Yang Li, Shao-Lun Huang +4
cs.LGcs.CVarXiv:2212.10082v12022Simple and Effective Text Matching with Richer Alignment Features
Runqi Yang, Jianhai Zhang, Xing Gao +2
cs.CLcs.LGarXiv:1908.00300v12019Message Passing for Hyper-Relational Knowledge Graphs
Mikhail Galkin, Priyansh Trivedi, Gaurav Maheshwari +2
cs.LGcs.AIcs.CLarXiv:2009.10847v12020Wind speed prediction using a hybrid model of the multi-layer perceptron and whale optimization algorithm
Saeed Samadianfard, Sajjad Hashemi, Katayoun Kargar +5
cs.LGstat.MLarXiv:2002.06226v12020Understanding Machine-learned Density Functionals
Li Li, John C. Snyder, Isabelle M. Pelaschier +6
physics.chem-phcs.LGphysics.comp-pharXiv:1404.1333v22014A Generalization of Amari's Bayesian Duality
Mohammad Emtiyaz Khan, Thomas Möllenhoff
cs.AIcs.LGstat.MLarXiv:2609.09126v12026Toxicity Prediction using Deep Learning
Thomas Unterthiner, Andreas Mayr, Günter Klambauer +1
stat.MLcs.LGcs.NEarXiv:1503.01445v12015AGSA-Net: Abundance-Guided Self-Attention Network for Spectral Unmixing-Aware Hyperspectral Remote Sensing Image Classification
Nafisa Anjum, Satavisa Dey Borno, Ananna Saha +4
cs.CVcs.AIcs.LGarXiv:2609.06359v12026Dynamical Regimes of Diffusion Models
Giulio Biroli, Tony Bonnaire, Valentin de Bortoli +1
cs.LGcond-mat.stat-mecharXiv:2402.18491v12024A Stochastic PCA and SVD Algorithm with an Exponential Convergence Rate
Ohad Shamir
cs.LGmath.NAmath.OCarXiv:1409.2848v52014Linear Algebra Foundations of Efficient Attention: A Phase Reversal in Rank Collapse Under SVD Compression
Anjaneya Teja Sarma Kalvakolanu
cs.LGcs.AIcs.NEarXiv:2609.06341v12026L4GM: Large 4D Gaussian Reconstruction Model
Jiawei Ren, Kevin Xie, Ashkan Mirzaei +8
cs.CVcs.LGarXiv:2406.10324v12024Zipformer: A faster and better encoder for automatic speech recognition
Zengwei Yao, Liyong Guo, Xiaoyu Yang +6
eess.AScs.LGcs.SDarXiv:2310.11230v42023On Multiplicative Integration with Recurrent Neural Networks
Yuhuai Wu, Saizheng Zhang, Ying Zhang +2
cs.LGarXiv:1606.06630v22016Reconfigurable Intelligent Surface Enabled Federated Learning: A Unified Communication-Learning Design Approach
Hang Liu, Xiaojun Yuan, Ying-Jun Angela Zhang
cs.ITcs.LGcs.NIarXiv:2011.10282v42020Pre-training via Paraphrasing
Mike Lewis, Marjan Ghazvininejad, Gargi Ghosh +3
cs.CLcs.LGstat.MLarXiv:2006.15020v12020Delay and Cooperation in Nonstochastic Bandits
Nicolo' Cesa-Bianchi, Claudio Gentile, Yishay Mansour +1
cs.LGarXiv:1602.04741v22016Towards Adversarially Robust Object Detection
Haichao Zhang, Jianyu Wang
cs.CVcs.LGeess.IVarXiv:1907.10310v12019Debiasing Graph Neural Networks via Learning Disentangled Causal Substructure
Shaohua Fan, Xiao Wang, Yanhu Mo +2
cs.LGcs.AIarXiv:2209.14107v12022TransEdge: Translating Relation-contextualized Embeddings for Knowledge Graphs
Zequn Sun, Jiacheng Huang, Wei Hu +3
cs.AIcs.CLcs.LGarXiv:2004.13579v12020Deep Unknown Intent Detection with Margin Loss
Ting-En Lin, Hua Xu
cs.CLcs.LGarXiv:1906.00434v12019No Metrics Are Perfect: Adversarial Reward Learning for Visual Storytelling
Xin Wang, Wenhu Chen, Yuan-Fang Wang +1
cs.CLcs.AIcs.CVarXiv:1804.09160v22018Active Adversarial Domain Adaptation
Jong-Chyi Su, Yi-Hsuan Tsai, Kihyuk Sohn +3
cs.CVcs.LGarXiv:1904.07848v22019Deep Gaussian Mixture Models
Cinzia Viroli, Geoffrey J. McLachlan
stat.MLcs.LGarXiv:1711.06929v12017Robotic Telekinesis: Learning a Robotic Hand Imitator by Watching Humans on Youtube
Aravind Sivakumar, Kenneth Shaw, Deepak Pathak
cs.ROcs.AIcs.CVarXiv:2202.10448v22022A review of Generative Adversarial Networks (GANs) and its applications in a wide variety of disciplines -- From Medical to Remote Sensing
Ankan Dash, Junyi Ye, Guiling Wang
cs.LGcs.AIcs.CVarXiv:2110.01442v12021Towards Foundation Models for Scientific Machine Learning: Characterizing Scaling and Transfer Behavior
Shashank Subramanian, Peter Harrington, Kurt Keutzer +4
cs.LGmath.NAarXiv:2306.00258v12023Random vector functional link neural network based ensemble deep learning for short-term load forecasting
Ruobin Gao, Liang Du, P. N. Suganthan +2
cs.LGcs.AIeess.SParXiv:2107.14385v12021RL-RRT: Kinodynamic Motion Planning via Learning Reachability Estimators from RL Policies
Hao-Tien Lewis Chiang, Jasmine Hsu, Marek Fiser +2
cs.ROcs.AIcs.LGarXiv:1907.04799v22019SAEScientist-Bench: Can AI Agents Conduct Autonomous SAE Interpretability Research?
Yuqiao Tan, Shizhu He, Jun Zhao +1
cs.AIcs.CLcs.LGarXiv:2609.09113v12026Protein sequence design with deep generative models
Zachary Wu, Kadina E. Johnston, Frances H. Arnold +1
q-bio.QMcs.LGq-bio.BMarXiv:2104.04457v12021LLM Critics Help Catch LLM Bugs
Nat McAleese, Rai Michael Pokorny, Juan Felipe Ceron Uribe +3
cs.SEcs.LGarXiv:2407.00215v12024Learning to Control Self-Assembling Morphologies: A Study of Generalization via Modularity
Deepak Pathak, Chris Lu, Trevor Darrell +2
cs.LGcs.AIcs.CVarXiv:1902.05546v22019Keyformer: KV Cache Reduction through Key Tokens Selection for Efficient Generative Inference
Muhammad Adnan, Akhil Arunkumar, Gaurav Jain +3
cs.LGcs.AIcs.ARarXiv:2403.09054v22024Machine Learning (ML)-Centric Resource Management in Cloud Computing: A Review and Future Directions
Tahseen Khan, Wenhong Tian, Rajkumar Buyya
cs.DCcs.LGarXiv:2105.05079v12021Overview frequency principle/spectral bias in deep learning
Zhi-Qin John Xu, Yaoyu Zhang, Tao Luo
cs.LGarXiv:2201.07395v42022