Machine Learning

Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

9,001 to 9,060 of 20,201

  1. AutoGCL: Automated Graph Contrastive Learning via Learnable View Generators

    Yihang Yin, Qingzhong Wang, Siyu Huang +2

    cs.LGarXiv:2109.10259v22021
  2. Node-Based Learning of Multiple Gaussian Graphical Models

    Karthik Mohan, Palma London, Maryam Fazel +2

    stat.MLcs.LGmath.OCarXiv:1303.5145v42013
  3. Scaling description of generalization with number of parameters in deep learning

    Mario Geiger, Arthur Jacot, Stefano Spigler +6

    cond-mat.dis-nncs.LGarXiv:1901.01608v52019
  4. Adversarial Sample Detection for Deep Neural Network through Model Mutation Testing

    Jingyi Wang, Guoliang Dong, Jun Sun +2

    cs.LGcs.SEstat.MLarXiv:1812.05793v22018
  5. Improving model calibration with accuracy versus uncertainty optimization

    Ranganath Krishnan, Omesh Tickoo

    cs.LGarXiv:2012.07923v12020
  6. SE(3)-DiffusionFields: Learning smooth cost functions for joint grasp and motion optimization through diffusion

    Julen Urain, Niklas Funk, Jan Peters +1

    cs.ROcs.LGarXiv:2209.03855v42022
  7. HiPoly: a hierarchical polymer-native AI framework for property prediction and generative design

    Ge Sun, Gervasio Zaldivar, Yuan Tian +5

    physics.chem-phcond-mat.mtrl-scics.AIarXiv:2609.02746v12026
  8. A Bi-model based RNN Semantic Frame Parsing Model for Intent Detection and Slot Filling

    Yu Wang, Yilin Shen, Hongxia Jin

    cs.CLcs.AIcs.LGarXiv:1812.10235v12018
  9. S4NN: temporal backpropagation for spiking neural networks with one spike per neuron

    Saeed Reza Kheradpisheh, Timothée Masquelier

    cs.NEcs.CVcs.LGarXiv:1910.09495v42019
  10. Patch Slimming for Efficient Vision Transformers

    Yehui Tang, Kai Han, Yunhe Wang +4

    cs.CVcs.LGarXiv:2106.02852v22021
  11. Posterior Tempering Explains Variance Inflation in Linear and Generalized Linear Thompson Sampling

    Prateek Jaiswal, Debdeep Pati, Anirban Bhattacharya +1

    stat.MLcs.ITcs.LGarXiv:2609.01999v12026
  12. Semantic Hierarchy Emerges in Deep Generative Representations for Scene Synthesis

    Ceyuan Yang, Yujun Shen, Bolei Zhou

    cs.CVcs.GRcs.LGarXiv:1911.09267v32019
  13. Improving CUR Matrix Decomposition and the Nyström Approximation via Adaptive Sampling

    Shusen Wang, Zhihua Zhang

    cs.LGmath.NAarXiv:1303.4207v72013
  14. Batch Active Learning at Scale

    Gui Citovsky, Giulia DeSalvo, Claudio Gentile +4

    cs.LGcs.AIarXiv:2107.14263v12021
  15. Dropout Inference in Bayesian Neural Networks with Alpha-divergences

    Yingzhen Li, Yarin Gal

    cs.LGstat.MLarXiv:1703.02914v12017
  16. Should You Mask 15% in Masked Language Modeling?

    Alexander Wettig, Tianyu Gao, Zexuan Zhong +1

    cs.CLcs.LGarXiv:2202.08005v32022
  17. Percolation Dynamics in Optimization : Variance Cascades and Discrete Scale Invariance

    Sai Niranjan Ramachandran, Suvrit Sra

    cs.LGcond-mat.dis-nncond-mat.stat-mecharXiv:2609.02373v12026
  18. Humanoid Safe Stop via Learned Stoppability Value

    Junfeng Long, Pieter Abbeel, Koushil Sreenath +3

    cs.ROcs.LGeess.SYarXiv:2609.02358v12026
  19. Logical Neural Networks

    Ryan Riegel, Alexander Gray, Francois Luus +12

    cs.AIcs.LGcs.LOarXiv:2006.13155v12020
  20. Multi-Agent Retrieval-Augmented Generation for Efficient Cloud Knowledge Base Search in Telecom SNOC Environment

    Harish Saragadam, Sudhanshu Sharma, Ipsha Routray

    cs.IRcs.LGarXiv:2609.01618v12026
  21. Is a Good Representation Sufficient for Sample Efficient Reinforcement Learning?

    Simon S. Du, Sham M. Kakade, Ruosong Wang +1

    cs.LGcs.AImath.OCarXiv:1910.03016v42019
  22. Prompt-Space Meta-Learning Does Not Transfer Across Users: A Frozen-LLM Negative Result

    Liam Byrne, David Dylan, Orla Fitzgerald +4

    cs.LGarXiv:2609.01615v12026
  23. The Uncertainty Bellman Equation and Exploration

    Brendan O'Donoghue, Ian Osband, Remi Munos +1

    cs.AIcs.LGmath.OCarXiv:1709.05380v42017
  24. Exponential concentration in quantum kernel methods

    Supanut Thanasilp, Samson Wang, M. Cerezo +1

    quant-phcs.LGstat.MLarXiv:2208.11060v22022
  25. When Literature Data Mislead Artificial Intelligence in Materials Discovery

    Qian Wang, Ying Li, Ryuhei Sato +4

    cs.IRcond-mat.mtrl-scics.CEarXiv:2609.01621v12026
  26. Generating Useful Accident-Prone Driving Scenarios via a Learned Traffic Prior

    Davis Rempe, Jonah Philion, Leonidas J. Guibas +2

    cs.CVcs.LGcs.ROarXiv:2112.05077v22021
  27. Convolutional Conditional Neural Processes

    Jonathan Gordon, Wessel P. Bruinsma, Andrew Y. K. Foong +3

    stat.MLcs.LGarXiv:1910.13556v52019
  28. PRISM: An Agentic Multi-Model Architecture for Proactive Safety in Autonomous Transportation Systems

    Joyjit Roy, Samaresh Kumar Singh, Sushanta Das

    cs.MAcs.CVcs.ETarXiv:2609.01623v12026
  29. RecEvolve: A Knowledge-Driven Autonomous Agent System for Recommender Systems

    Weidi Pan, He Ma, Shuhao Ye +6

    cs.IRcs.AIcs.LGarXiv:2609.01622v12026
  30. Fast Approximate Natural Gradient Descent in a Kronecker-factored Eigenbasis

    Thomas George, César Laurent, Xavier Bouthillier +2

    cs.LGstat.MLarXiv:1806.03884v22018
  31. Bilinear Classes: A Structural Framework for Provable Generalization in RL

    Simon S. Du, Sham M. Kakade, Jason D. Lee +4

    cs.LGcs.AImath.OCarXiv:2103.10897v32021
  32. Hybrid Retrieval-Augmented Generation with Knowledge Graph Expansion, RRF Fusion, and Per-Chunk Grounded Evaluation for Enterprise Document Search

    Harish Saragadam, Sudhanshu Sharma, Meghana Pujari

    cs.IRcs.AIcs.LGarXiv:2609.01617v12026
  33. A Common Measure of Communication for Speech Brain-Computer Interfaces

    Dulhan Jayalath, Benjamin Ballyk, Oiwi Parker Jones

    cs.LGq-bio.NCarXiv:2609.02887v12026
  34. Large Associative Memory Problem in Neurobiology and Machine Learning

    Dmitry Krotov, John Hopfield

    q-bio.NCcond-mat.dis-nncs.CLarXiv:2008.06996v32020
  35. The Cost of Privacy: Optimal Rates of Convergence for Parameter Estimation with Differential Privacy

    T. Tony Cai, Yichen Wang, Linjun Zhang

    stat.MLcs.CRcs.DSarXiv:1902.04495v52019
  36. BottleNet: A Deep Learning Architecture for Intelligent Mobile Cloud Computing Services

    Amir Erfan Eshratifar, Amirhossein Esmaili, Massoud Pedram

    cs.DCcs.LGarXiv:1902.01000v12019
  37. Clustering on Multi-Layer Graphs via Subspace Analysis on Grassmann Manifolds

    Xiaowen Dong, Pascal Frossard, Pierre Vandergheynst +1

    cs.LGcs.CVcs.SIarXiv:1303.2221v12013
  38. RouterBench: A Benchmark for Multi-LLM Routing System

    Qitian Jason Hu, Jacob Bieker, Xiuyu Li +5

    cs.LGcs.AIarXiv:2403.12031v22024
  39. PAC-Bayesian Theory Meets Bayesian Inference

    Pascal Germain, Francis Bach, Alexandre Lacoste +1

    stat.MLcs.LGarXiv:1605.08636v42016
  40. Deep Learning in Finance

    J. B. Heaton, N. G. Polson, J. H. Witte

    cs.LGarXiv:1602.06561v32016
  41. Last Step Matters: Early Uncertainty Cannot Predict Failure in Long-Horizon Agents

    Zongyue Li, Chengyue Yu, Lei Zang +3

    cs.LGarXiv:2608.29685v12026
  42. A Complete Survey on Generative AI (AIGC): Is ChatGPT from GPT-4 to GPT-5 All You Need?

    Chaoning Zhang, Chenshuang Zhang, Sheng Zheng +14

    cs.AIcs.CVcs.LGarXiv:2303.11717v12023
  43. Land Cover Classification via Multi-temporal Spatial Data by Recurrent Neural Networks

    Dino Ienco, Raffaele Gaetano, Claire Dupaquier +1

    cs.CVcs.LGarXiv:1704.04055v12017
  44. Differential Privacy-enabled Federated Learning for Sensitive Health Data

    Olivia Choudhury, Aris Gkoulalas-Divanis, Theodoros Salonidis +4

    cs.LGcs.CRarXiv:1910.02578v32019
  45. Stochastic Variance-Reduced Policy Gradient

    Matteo Papini, Damiano Binaghi, Giuseppe Canonaco +2

    cs.LGstat.MLarXiv:1806.05618v12018
  46. Learning from Complementary Labels

    Takashi Ishida, Gang Niu, Weihua Hu +1

    stat.MLcs.LGarXiv:1705.07541v22017
  47. Graph Condensation for Graph Neural Networks

    Wei Jin, Lingxiao Zhao, Shichang Zhang +3

    cs.LGcs.AIarXiv:2110.07580v42021
  48. A Deep Latent Variable Framework for Jointly Modeling Missingness, Measurement Error, and Heterogeneity

    Yasin Khadem Charvadeh, Grace Y. Yi, Mithat Gönen +1

    stat.MLcs.LGarXiv:2608.30040v12026
  49. Remote Sensing Image Super-resolution and Object Detection: Benchmark and State of the Art

    Yi Wang, Syed Muhammad Arsalan Bashir, Mahrukh Khan +5

    cs.CVcs.AIcs.LGarXiv:2111.03260v12021
  50. Can Knowledge Graphs Reduce Hallucinations in LLMs? : A Survey

    Garima Agrawal, Tharindu Kumarage, Zeyad Alghamdi +1

    cs.CLcs.LGarXiv:2311.07914v22023
  51. A systematic comparison of supervised classifiers

    D. R. Amancio, C. H. Comin, D. Casanova +4

    cs.LGarXiv:1311.0202v12013
  52. Image Data Augmentation Approaches: A Comprehensive Survey and Future directions

    Teerath Kumar, Alessandra Mileo, Rob Brennan +1

    cs.CVcs.AIcs.LGarXiv:2301.02830v42023
  53. Forecasting day-ahead electricity prices in Europe: the importance of considering market integration

    Jesus Lago, Fjo De Ridder, Peter Vrancx +1

    q-fin.STcs.CEcs.LGarXiv:1708.07061v32017
  54. Partition-Aware Unlearning for Removing Spurious Correlations in Large Vision-Language Models

    Aditi Sarker, Nazreen Shah, Rafi Ibn Sultan +3

    cs.CVcs.AIcs.LGarXiv:2608.29996v12026
  55. The BOSARIS Toolkit: Theory, Algorithms and Code for Surviving the New DCF

    Niko Brümmer, Edward de Villiers

    stat.APcs.LGstat.MLarXiv:1304.2865v12013
  56. Automatic Semantic Augmentation of Language Model Prompts (for Code Summarization)

    Toufique Ahmed, Kunal Suresh Pai, Premkumar Devanbu +1

    cs.SEcs.LGarXiv:2304.06815v32023
  57. The Capacity and Robustness Trade-off: Revisiting the Channel Independent Strategy for Multivariate Time Series Forecasting

    Lu Han, Han-Jia Ye, De-Chuan Zhan

    cs.LGarXiv:2304.05206v12023
  58. Synthetic Data -- what, why and how?

    James Jordon, Lukasz Szpruch, Florimond Houssiau +5

    cs.LGarXiv:2205.03257v12022
  59. A Novel Ensemble Deep Learning Model for Stock Prediction Based on Stock Prices and News

    Yang Li, Yi Pan

    q-fin.STcs.LGarXiv:2007.12620v12020
  60. TicTac: Accelerating Distributed Deep Learning with Communication Scheduling

    Sayed Hadi Hashemi, Sangeetha Abdu Jyothi, Roy H. Campbell

    cs.DCcs.LGcs.PFarXiv:1803.03288v22018