Machine Learning

Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

14,701 to 14,760 of 20,199

  1. A Holistic Approach to Undesired Content Detection in the Real World

    Todor Markov, Chong Zhang, Sandhini Agarwal +5

    cs.CLcs.LGarXiv:2208.03274v22022
  2. On the Expressive Power of Deep Learning: A Tensor Analysis

    Nadav Cohen, Or Sharir, Amnon Shashua

    cs.NEcs.LGmath.NAarXiv:1509.05009v32015
  3. What actually runs: a measurement study of language model placement and decode speed on the Apple Neural Engine

    Shahir M A

    cs.LGcs.ARcs.PFarXiv:2608.22110v12026
  4. Multi-Instance Learning by Treating Instances As Non-I.I.D. Samples

    Zhi-Hua Zhou, Yu-Yin Sun, Yu-Feng Li

    cs.LGcs.AIarXiv:0807.1997v42008
  5. Focal Frequency Loss for Image Reconstruction and Synthesis

    Liming Jiang, Bo Dai, Wayne Wu +1

    cs.CVcs.LGeess.IVarXiv:2012.12821v32020
  6. A quantum-inspired classical algorithm for recommendation systems

    Ewin Tang

    cs.IRcs.DScs.LGarXiv:1807.04271v32018
  7. Stabilizing Transformers for Reinforcement Learning

    Emilio Parisotto, H. Francis Song, Jack W. Rae +10

    cs.LGcs.AIstat.MLarXiv:1910.06764v12019
  8. Cognitive Architectures for Language Agents

    Theodore R. Sumers, Shunyu Yao, Karthik Narasimhan +1

    cs.AIcs.CLcs.LGarXiv:2309.02427v32023
  9. Dataset Condensation with Distribution Matching

    Bo Zhao, Hakan Bilen

    cs.LGcs.CVarXiv:2110.04181v32021
  10. More Experts, Worse Dynamics: Inverse Scaling and Spectral Bias in Mixture-of-Experts State-Space Models

    Chandresh Pandey

    cs.LGarXiv:2608.21840v12026
  11. The Communication Map of a Transformer

    Richard Zhe Wang

    cs.LGcs.CLarXiv:2608.22007v12026
  12. Beyond Fixed Directions: Adaptive Representation Analysis of Reasoning and Memorization in LLMs

    Shaheen Nabi

    cs.LGarXiv:2608.21919v12026
  13. Layer-wise Analysis of a Self-supervised Speech Representation Model

    Ankita Pasad, Ju-Chieh Chou, Karen Livescu

    cs.CLcs.LGeess.ASarXiv:2107.04734v32021
  14. TextWorld: A Learning Environment for Text-based Games

    Marc-Alexandre Côté, Ákos Kádár, Xingdi Yuan +10

    cs.LGcs.CLstat.MLarXiv:1806.11532v22018
  15. Analog Bits: Generating Discrete Data using Diffusion Models with Self-Conditioning

    Ting Chen, Ruixiang Zhang, Geoffrey Hinton

    cs.CVcs.AIcs.CLarXiv:2208.04202v22022
  16. Learning Factored Representations in a Deep Mixture of Experts

    David Eigen, Marc'Aurelio Ranzato, Ilya Sutskever

    cs.LGarXiv:1312.4314v32013
  17. Polyglot: Distributed Word Representations for Multilingual NLP

    Rami Al-Rfou, Bryan Perozzi, Steven Skiena

    cs.CLcs.LGarXiv:1307.1662v22013
  18. Crafting Adversarial Input Sequences for Recurrent Neural Networks

    Nicolas Papernot, Patrick McDaniel, Ananthram Swami +1

    cs.CRcs.LGcs.NEarXiv:1604.08275v12016
  19. GraphSMOTE: Imbalanced Node Classification on Graphs with Graph Neural Networks

    Tianxiang Zhao, Xiang Zhang, Suhang Wang

    cs.LGarXiv:2103.08826v12021
  20. Hacking Smart Machines with Smarter Ones: How to Extract Meaningful Data from Machine Learning Classifiers

    Giuseppe Ateniese, Giovanni Felici, Luigi V. Mancini +3

    cs.CRcs.LGstat.MLarXiv:1306.4447v12013
  21. Device Placement Optimization with Reinforcement Learning

    Azalia Mirhoseini, Hieu Pham, Quoc V. Le +7

    cs.LGcs.AIarXiv:1706.04972v22017
  22. Failing Loudly: An Empirical Study of Methods for Detecting Dataset Shift

    Stephan Rabanser, Stephan Günnemann, Zachary C. Lipton

    stat.MLcs.LGarXiv:1810.11953v42018
  23. One Thousand and One Hours: Self-driving Motion Prediction Dataset

    John Houston, Guido Zuidhof, Luca Bergamini +6

    cs.CVcs.LGcs.ROarXiv:2006.14480v22020
  24. Discovering Dual-Origin Slow Wind from Solar Orbiter with Self-Supervised Contrastive Learning

    Henry Han, Jorge Yero Salazar

    astro-ph.SRcs.AIcs.LGarXiv:2608.22065v12026
  25. Quasi-Dense Similarity Learning for Multiple Object Tracking

    Jiangmiao Pang, Linlu Qiu, Xia Li +4

    cs.CVcs.LGarXiv:2006.06664v42020
  26. Multiplicative Normalizing Flows for Variational Bayesian Neural Networks

    Christos Louizos, Max Welling

    stat.MLcs.LGarXiv:1703.01961v22017
  27. GreenLeaf Law Embed Tiny: A Compact Embedding Model for Legal Domain Retrieval

    Surya Saka

    cs.LGcs.AIcs.CLarXiv:2608.24936v12026
  28. Generic Attention-model Explainability for Interpreting Bi-Modal and Encoder-Decoder Transformers

    Hila Chefer, Shir Gur, Lior Wolf

    cs.CVcs.LGarXiv:2103.15679v12021
  29. Optimizing Millions of Hyperparameters by Implicit Differentiation

    Jonathan Lorraine, Paul Vicol, David Duvenaud

    cs.LGstat.MLarXiv:1911.02590v12019
  30. Last Layer Re-Training is Sufficient for Robustness to Spurious Correlations

    Polina Kirichenko, Pavel Izmailov, Andrew Gordon Wilson

    cs.LGcs.CVstat.MLarXiv:2204.02937v22022
  31. A Survey on Graph Kernels

    Nils M. Kriege, Fredrik D. Johansson, Christopher Morris

    cs.LGstat.MLarXiv:1903.11835v22019
  32. Unnatural Instructions: Tuning Language Models with (Almost) No Human Labor

    Or Honovich, Thomas Scialom, Omer Levy +1

    cs.CLcs.AIcs.LGarXiv:2212.09689v12022
  33. COVID-ResNet: A Deep Learning Framework for Screening of COVID19 from Radiographs

    Muhammad Farooq, Abdul Hafeez

    eess.IVcs.CVcs.LGarXiv:2003.14395v12020
  34. No More Pesky Learning Rates

    Tom Schaul, Sixin Zhang, Yann LeCun

    stat.MLcs.LGarXiv:1206.1106v22012
  35. Two-branch Recurrent Network for Isolating Deepfakes in Videos

    Iacopo Masi, Aditya Killekar, Royston Marian Mascarenhas +2

    cs.CVcs.CYcs.LGarXiv:2008.03412v32020
  36. Mitigating Explanation Leakage in Financial Fraud Detection Systems

    Muhammad Waleed Gul, Elaheh Homayounvala

    cs.LGcs.CRarXiv:2608.22607v12026
  37. From Symmetry to Invariance: Learning Galois Equivalent Representations in Finite Fields

    Zheng Zhang, Na Zhang

    cs.LGarXiv:2608.22513v12026
  38. Rainbow Memory: Continual Learning with a Memory of Diverse Samples

    Jihwan Bang, Heesu Kim, YoungJoon Yoo +2

    cs.CVcs.LGarXiv:2103.17230v12021
  39. Adversarially Regularized Graph Autoencoder for Graph Embedding

    Shirui Pan, Ruiqi Hu, Guodong Long +3

    cs.LGstat.MLarXiv:1802.04407v22018
  40. When Test-Time Adaptation Helps, Harms, or Becomes Inactive: A Condition-Level Study on CIFAR-10-C

    Sreeja Guha Majumdar, Aratrika Saha

    cs.LGcs.CVarXiv:2608.22233v12026
  41. FreeMatch: Self-adaptive Thresholding for Semi-supervised Learning

    Yidong Wang, Hao Chen, Qiang Heng +9

    cs.LGcs.CVarXiv:2205.07246v32022
  42. Slicing Aided Hyper Inference and Fine-tuning for Small Object Detection

    Fatih Cagatay Akyon, Sinan Onur Altinuc, Alptekin Temizel

    cs.CVcs.LGarXiv:2202.06934v52022
  43. Shallow and Deep Convolutional Networks for Saliency Prediction

    Junting Pan, Kevin McGuinness, Elisa Sayrol +2

    cs.CVcs.LGarXiv:1603.00845v12016
  44. Structured Attention Networks

    Yoon Kim, Carl Denton, Luong Hoang +1

    cs.CLcs.LGcs.NEarXiv:1702.00887v32017
  45. Convolutional neural networks with low-rank regularization

    Cheng Tai, Tong Xiao, Yi Zhang +2

    cs.LGcs.CVstat.MLarXiv:1511.06067v32015
  46. Do Adversarially Robust ImageNet Models Transfer Better?

    Hadi Salman, Andrew Ilyas, Logan Engstrom +2

    cs.CVcs.LGstat.MLarXiv:2007.08489v22020
  47. Machine learning in cardiovascular flows modeling: Predicting arterial blood pressure from non-invasive 4D flow MRI data using physics-informed neural networks

    Georgios Kissas, Yibo Yang, Eileen Hwuang +3

    cs.LGstat.MLarXiv:1905.04817v22019
  48. Multi-Agent Cooperation and the Emergence of (Natural) Language

    Angeliki Lazaridou, Alexander Peysakhovich, Marco Baroni

    cs.CLcs.CVcs.GTarXiv:1612.07182v22016
  49. Learning feed-forward one-shot learners

    Luca Bertinetto, João F. Henriques, Jack Valmadre +2

    cs.CVcs.LGarXiv:1606.05233v12016
  50. Does a Modern-Handwriting Warm-Up Help Historical Arabic OCR? A Reproducible, Compute-Matched Evaluation on Muharaf and KHATT

    Sumaih Almarshad, Maram Alamri, Dona Aloraini +4

    cs.LGcs.CVarXiv:2608.22316v12026
  51. Deep Parametric Continuous Convolutional Neural Networks

    Shenlong Wang, Simon Suo, Wei-Chiu Ma +2

    cs.CVcs.AIcs.LGarXiv:2101.06742v12021
  52. No Fear of Heterogeneity: Classifier Calibration for Federated Learning with Non-IID Data

    Mi Luo, Fei Chen, Dapeng Hu +3

    cs.LGcs.CVcs.DCarXiv:2106.05001v22021
  53. Counterfactual Evaluation of Temporal Observation Protocols

    Xizhe Zhang

    cs.LGarXiv:2608.22221v12026
  54. KONTOGRAPH: Verified Point-in-Time Feature Consistency and Amortised Explanation for Real-Time Anti-Money Laundering under a 200 ms Decision Budget

    Ahmed Abolfadl

    cs.CRcs.AIcs.LGarXiv:2608.22389v12026
  55. NGBoost: Natural Gradient Boosting for Probabilistic Prediction

    Tony Duan, Anand Avati, Daisy Yi Ding +4

    cs.LGstat.MLarXiv:1910.03225v42019
  56. Privacy Amplification by Subsampling: Tight Analyses via Couplings and Divergences

    Borja Balle, Gilles Barthe, Marco Gaboardi

    cs.LGcs.CRstat.MLarXiv:1807.01647v22018
  57. Explaining Neural Scaling Laws

    Yasaman Bahri, Ethan Dyer, Jared Kaplan +2

    cs.LGcond-mat.dis-nnstat.MLarXiv:2102.06701v22021
  58. Power Hungry Processing: Watts Driving the Cost of AI Deployment?

    Alexandra Sasha Luccioni, Yacine Jernite, Emma Strubell

    cs.LGarXiv:2311.16863v32023
  59. Dynamics-Aware Unsupervised Discovery of Skills

    Archit Sharma, Shixiang Gu, Sergey Levine +2

    cs.LGcs.ROstat.MLarXiv:1907.01657v22019
  60. Disparities in Dermatology AI Performance on a Diverse, Curated Clinical Image Set

    Roxana Daneshjou, Kailas Vodrahalli, Roberto A Novoa +16

    eess.IVcs.AIcs.CVarXiv:2203.08807v12022