Machine Learning

Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

3,301 to 3,360 of 20,454

  1. Grounding Language with Visual Affordances over Unstructured Data

    Oier Mees, Jessica Borja-Diaz, Wolfram Burgard

    cs.ROcs.AIcs.CLarXiv:2210.01911v32022
  2. Investigating the Limitations of Transformers with Simple Arithmetic Tasks

    Rodrigo Nogueira, Zhiying Jiang, Jimmy Lin

    cs.CLcs.AIcs.LGarXiv:2102.13019v32021
  3. All for 1-Bit: Towards Genuine 1-Bit Post-Training Quantization for LLMs

    Zhixiong Zhao, Zukang Xu, Guangyu Sun +2

    cs.LGcs.AIarXiv:2609.06161v12026
  4. An Attentive Inductive Bias for Sequential Recommendation beyond the Self-Attention

    Yehjin Shin, Jeongwhan Choi, Hyowon Wi +1

    cs.LGcs.AIcs.IRarXiv:2312.10325v22023
  5. Subdivision-Based Mesh Convolution Networks

    Shi-Min Hu, Zheng-Ning Liu, Meng-Hao Guo +4

    cs.CVcs.GRcs.LGarXiv:2106.02285v22021
  6. Topology-Aware Correlations Between Relations for Inductive Link Prediction in Knowledge Graphs

    Jiajun Chen, Huarui He, Feng Wu +1

    cs.LGcs.CLarXiv:2103.03642v12021
  7. Learning Algorithms for Active Learning

    Philip Bachman, Alessandro Sordoni, Adam Trischler

    cs.LGarXiv:1708.00088v12017
  8. Decision-Aware Suffix Prediction and Reasoning of Business Processes

    Henryk Mustroph, Stefanie Rinderle-Ma

    cs.LGcs.AIarXiv:2609.06169v12026
  9. Instance-Conditioned GAN

    Arantxa Casanova, Marlène Careil, Jakob Verbeek +2

    cs.CVcs.LGarXiv:2109.05070v22021
  10. Reconstruction of Markov Random Fields from Samples: Some Easy Observations and Algorithms

    Guy Bresler, Elchanan Mossel, Allan Sly

    cs.CCcs.LGarXiv:0712.1402v22007
  11. Contrastive Predictive Coding for Human Activity Recognition

    Harish Haresamudram, Irfan Essa, Thomas Ploetz

    cs.LGarXiv:2012.05333v12020
  12. Permutation invariant graph-to-sequence model for template-free retrosynthesis and reaction prediction

    Zhengkai Tu, Connor W. Coley

    cs.LGarXiv:2110.09681v12021
  13. Hyperspherical Prototype Networks

    Pascal Mettes, Elise van der Pol, Cees G. M. Snoek

    cs.LGstat.MLarXiv:1901.10514v32019
  14. Matérn Gaussian processes on Riemannian manifolds

    Viacheslav Borovitskiy, Alexander Terenin, Peter Mostowsky +1

    stat.MLcs.LGarXiv:2006.10160v62020
  15. Adversarial Attacks on Deep Neural Networks for Time Series Classification

    Hassan Ismail Fawaz, Germain Forestier, Jonathan Weber +2

    cs.LGcs.CRstat.MLarXiv:1903.07054v22019
  16. Fairness-aware Agnostic Federated Learning

    Wei Du, Depeng Xu, Xintao Wu +1

    cs.LGcs.CYarXiv:2010.05057v12020
  17. Stable Learning via Sample Reweighting

    Zheyan Shen, Peng Cui, Tong Zhang +1

    cs.LGstat.MLarXiv:1911.12580v12019
  18. Machine learning of hierarchical clustering to segment 2D and 3D images

    Juan Nunez-Iglesias, Ryan Kennedy, Toufiq Parag +2

    cs.CVcs.LGarXiv:1303.6163v32013
  19. Non-Autoregressive Machine Translation with Auxiliary Regularization

    Yiren Wang, Fei Tian, Di He +3

    cs.CLcs.LGstat.MLarXiv:1902.10245v12019
  20. What the Window Does Not Contain: Auditing Provenance in a Document-Grounded Instability Benchmark

    Seyed Mosayeb Alam

    cs.CLcs.AIcs.LGarXiv:2609.06147v12026
  21. Online Learning for Time Series Prediction

    Oren Anava, Elad Hazan, Shie Mannor +1

    cs.LGarXiv:1302.6927v12013
  22. POCO: Point Convolution for Surface Reconstruction

    Alexandre Boulch, Renaud Marlet

    cs.CVcs.CGcs.LGarXiv:2201.01831v22022
  23. Bilinear Factor Matrix Norm Minimization for Robust PCA: Algorithms and Applications

    Fanhua Shang, James Cheng, Yuanyuan Liu +2

    cs.LGcs.CVmath.OCarXiv:1810.05186v12018
    Summaries:한국어
  24. High-Throughput CNN Inference on Embedded ARM big.LITTLE Multi-Core Processors

    Siqi Wang, Gayathri Ananthanarayanan, Yifan Zeng +3

    cs.LGcs.DCcs.PFarXiv:1903.05898v32019
    Summaries:한국어
  25. An Information-Theoretic Approach to Transferability in Task Transfer Learning

    Yajie Bao, Yang Li, Shao-Lun Huang +4

    cs.LGcs.CVarXiv:2212.10082v12022
  26. Simple and Effective Text Matching with Richer Alignment Features

    Runqi Yang, Jianhai Zhang, Xing Gao +2

    cs.CLcs.LGarXiv:1908.00300v12019
  27. Message Passing for Hyper-Relational Knowledge Graphs

    Mikhail Galkin, Priyansh Trivedi, Gaurav Maheshwari +2

    cs.LGcs.AIcs.CLarXiv:2009.10847v12020
  28. Wind speed prediction using a hybrid model of the multi-layer perceptron and whale optimization algorithm

    Saeed Samadianfard, Sajjad Hashemi, Katayoun Kargar +5

    cs.LGstat.MLarXiv:2002.06226v12020
  29. Understanding Machine-learned Density Functionals

    Li Li, John C. Snyder, Isabelle M. Pelaschier +6

    physics.chem-phcs.LGphysics.comp-pharXiv:1404.1333v22014
  30. A Generalization of Amari's Bayesian Duality

    Mohammad Emtiyaz Khan, Thomas Möllenhoff

    cs.AIcs.LGstat.MLarXiv:2609.09126v12026
  31. Toxicity Prediction using Deep Learning

    Thomas Unterthiner, Andreas Mayr, Günter Klambauer +1

    stat.MLcs.LGcs.NEarXiv:1503.01445v12015
  32. AGSA-Net: Abundance-Guided Self-Attention Network for Spectral Unmixing-Aware Hyperspectral Remote Sensing Image Classification

    Nafisa Anjum, Satavisa Dey Borno, Ananna Saha +4

    cs.CVcs.AIcs.LGarXiv:2609.06359v12026
  33. Dynamical Regimes of Diffusion Models

    Giulio Biroli, Tony Bonnaire, Valentin de Bortoli +1

    cs.LGcond-mat.stat-mecharXiv:2402.18491v12024
  34. A Stochastic PCA and SVD Algorithm with an Exponential Convergence Rate

    Ohad Shamir

    cs.LGmath.NAmath.OCarXiv:1409.2848v52014
  35. Linear Algebra Foundations of Efficient Attention: A Phase Reversal in Rank Collapse Under SVD Compression

    Anjaneya Teja Sarma Kalvakolanu

    cs.LGcs.AIcs.NEarXiv:2609.06341v12026
  36. L4GM: Large 4D Gaussian Reconstruction Model

    Jiawei Ren, Kevin Xie, Ashkan Mirzaei +8

    cs.CVcs.LGarXiv:2406.10324v12024
  37. Zipformer: A faster and better encoder for automatic speech recognition

    Zengwei Yao, Liyong Guo, Xiaoyu Yang +6

    eess.AScs.LGcs.SDarXiv:2310.11230v42023
  38. On Multiplicative Integration with Recurrent Neural Networks

    Yuhuai Wu, Saizheng Zhang, Ying Zhang +2

    cs.LGarXiv:1606.06630v22016
  39. Reconfigurable Intelligent Surface Enabled Federated Learning: A Unified Communication-Learning Design Approach

    Hang Liu, Xiaojun Yuan, Ying-Jun Angela Zhang

    cs.ITcs.LGcs.NIarXiv:2011.10282v42020
  40. Pre-training via Paraphrasing

    Mike Lewis, Marjan Ghazvininejad, Gargi Ghosh +3

    cs.CLcs.LGstat.MLarXiv:2006.15020v12020
  41. Delay and Cooperation in Nonstochastic Bandits

    Nicolo' Cesa-Bianchi, Claudio Gentile, Yishay Mansour +1

    cs.LGarXiv:1602.04741v22016
  42. Towards Adversarially Robust Object Detection

    Haichao Zhang, Jianyu Wang

    cs.CVcs.LGeess.IVarXiv:1907.10310v12019
  43. Debiasing Graph Neural Networks via Learning Disentangled Causal Substructure

    Shaohua Fan, Xiao Wang, Yanhu Mo +2

    cs.LGcs.AIarXiv:2209.14107v12022
  44. TransEdge: Translating Relation-contextualized Embeddings for Knowledge Graphs

    Zequn Sun, Jiacheng Huang, Wei Hu +3

    cs.AIcs.CLcs.LGarXiv:2004.13579v12020
  45. Deep Unknown Intent Detection with Margin Loss

    Ting-En Lin, Hua Xu

    cs.CLcs.LGarXiv:1906.00434v12019
  46. No Metrics Are Perfect: Adversarial Reward Learning for Visual Storytelling

    Xin Wang, Wenhu Chen, Yuan-Fang Wang +1

    cs.CLcs.AIcs.CVarXiv:1804.09160v22018
  47. Active Adversarial Domain Adaptation

    Jong-Chyi Su, Yi-Hsuan Tsai, Kihyuk Sohn +3

    cs.CVcs.LGarXiv:1904.07848v22019
  48. Deep Gaussian Mixture Models

    Cinzia Viroli, Geoffrey J. McLachlan

    stat.MLcs.LGarXiv:1711.06929v12017
  49. Robotic Telekinesis: Learning a Robotic Hand Imitator by Watching Humans on Youtube

    Aravind Sivakumar, Kenneth Shaw, Deepak Pathak

    cs.ROcs.AIcs.CVarXiv:2202.10448v22022
  50. A review of Generative Adversarial Networks (GANs) and its applications in a wide variety of disciplines -- From Medical to Remote Sensing

    Ankan Dash, Junyi Ye, Guiling Wang

    cs.LGcs.AIcs.CVarXiv:2110.01442v12021
  51. Towards Foundation Models for Scientific Machine Learning: Characterizing Scaling and Transfer Behavior

    Shashank Subramanian, Peter Harrington, Kurt Keutzer +4

    cs.LGmath.NAarXiv:2306.00258v12023
  52. Random vector functional link neural network based ensemble deep learning for short-term load forecasting

    Ruobin Gao, Liang Du, P. N. Suganthan +2

    cs.LGcs.AIeess.SParXiv:2107.14385v12021
  53. RL-RRT: Kinodynamic Motion Planning via Learning Reachability Estimators from RL Policies

    Hao-Tien Lewis Chiang, Jasmine Hsu, Marek Fiser +2

    cs.ROcs.AIcs.LGarXiv:1907.04799v22019
  54. SAEScientist-Bench: Can AI Agents Conduct Autonomous SAE Interpretability Research?

    Yuqiao Tan, Shizhu He, Jun Zhao +1

    cs.AIcs.CLcs.LGarXiv:2609.09113v12026
  55. Protein sequence design with deep generative models

    Zachary Wu, Kadina E. Johnston, Frances H. Arnold +1

    q-bio.QMcs.LGq-bio.BMarXiv:2104.04457v12021
  56. LLM Critics Help Catch LLM Bugs

    Nat McAleese, Rai Michael Pokorny, Juan Felipe Ceron Uribe +3

    cs.SEcs.LGarXiv:2407.00215v12024
  57. Learning to Control Self-Assembling Morphologies: A Study of Generalization via Modularity

    Deepak Pathak, Chris Lu, Trevor Darrell +2

    cs.LGcs.AIcs.CVarXiv:1902.05546v22019
  58. Keyformer: KV Cache Reduction through Key Tokens Selection for Efficient Generative Inference

    Muhammad Adnan, Akhil Arunkumar, Gaurav Jain +3

    cs.LGcs.AIcs.ARarXiv:2403.09054v22024
  59. Machine Learning (ML)-Centric Resource Management in Cloud Computing: A Review and Future Directions

    Tahseen Khan, Wenhong Tian, Rajkumar Buyya

    cs.DCcs.LGarXiv:2105.05079v12021
  60. Overview frequency principle/spectral bias in deep learning

    Zhi-Qin John Xu, Yaoyu Zhang, Tao Luo

    cs.LGarXiv:2201.07395v42022