Machine Learning
Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
14,701 to 14,760 of 20,199
A Holistic Approach to Undesired Content Detection in the Real World
Todor Markov, Chong Zhang, Sandhini Agarwal +5
cs.CLcs.LGarXiv:2208.03274v22022On the Expressive Power of Deep Learning: A Tensor Analysis
Nadav Cohen, Or Sharir, Amnon Shashua
cs.NEcs.LGmath.NAarXiv:1509.05009v32015What actually runs: a measurement study of language model placement and decode speed on the Apple Neural Engine
Shahir M A
cs.LGcs.ARcs.PFarXiv:2608.22110v12026Multi-Instance Learning by Treating Instances As Non-I.I.D. Samples
Zhi-Hua Zhou, Yu-Yin Sun, Yu-Feng Li
cs.LGcs.AIarXiv:0807.1997v42008Focal Frequency Loss for Image Reconstruction and Synthesis
Liming Jiang, Bo Dai, Wayne Wu +1
cs.CVcs.LGeess.IVarXiv:2012.12821v32020A quantum-inspired classical algorithm for recommendation systems
Ewin Tang
cs.IRcs.DScs.LGarXiv:1807.04271v32018Stabilizing Transformers for Reinforcement Learning
Emilio Parisotto, H. Francis Song, Jack W. Rae +10
cs.LGcs.AIstat.MLarXiv:1910.06764v12019Cognitive Architectures for Language Agents
Theodore R. Sumers, Shunyu Yao, Karthik Narasimhan +1
cs.AIcs.CLcs.LGarXiv:2309.02427v32023Dataset Condensation with Distribution Matching
Bo Zhao, Hakan Bilen
cs.LGcs.CVarXiv:2110.04181v32021More Experts, Worse Dynamics: Inverse Scaling and Spectral Bias in Mixture-of-Experts State-Space Models
Chandresh Pandey
cs.LGarXiv:2608.21840v12026The Communication Map of a Transformer
Richard Zhe Wang
cs.LGcs.CLarXiv:2608.22007v12026Beyond Fixed Directions: Adaptive Representation Analysis of Reasoning and Memorization in LLMs
Shaheen Nabi
cs.LGarXiv:2608.21919v12026Layer-wise Analysis of a Self-supervised Speech Representation Model
Ankita Pasad, Ju-Chieh Chou, Karen Livescu
cs.CLcs.LGeess.ASarXiv:2107.04734v32021TextWorld: A Learning Environment for Text-based Games
Marc-Alexandre Côté, Ákos Kádár, Xingdi Yuan +10
cs.LGcs.CLstat.MLarXiv:1806.11532v22018Analog Bits: Generating Discrete Data using Diffusion Models with Self-Conditioning
Ting Chen, Ruixiang Zhang, Geoffrey Hinton
cs.CVcs.AIcs.CLarXiv:2208.04202v22022Learning Factored Representations in a Deep Mixture of Experts
David Eigen, Marc'Aurelio Ranzato, Ilya Sutskever
cs.LGarXiv:1312.4314v32013Polyglot: Distributed Word Representations for Multilingual NLP
Rami Al-Rfou, Bryan Perozzi, Steven Skiena
cs.CLcs.LGarXiv:1307.1662v22013Crafting Adversarial Input Sequences for Recurrent Neural Networks
Nicolas Papernot, Patrick McDaniel, Ananthram Swami +1
cs.CRcs.LGcs.NEarXiv:1604.08275v12016GraphSMOTE: Imbalanced Node Classification on Graphs with Graph Neural Networks
Tianxiang Zhao, Xiang Zhang, Suhang Wang
cs.LGarXiv:2103.08826v12021Hacking Smart Machines with Smarter Ones: How to Extract Meaningful Data from Machine Learning Classifiers
Giuseppe Ateniese, Giovanni Felici, Luigi V. Mancini +3
cs.CRcs.LGstat.MLarXiv:1306.4447v12013Device Placement Optimization with Reinforcement Learning
Azalia Mirhoseini, Hieu Pham, Quoc V. Le +7
cs.LGcs.AIarXiv:1706.04972v22017Failing Loudly: An Empirical Study of Methods for Detecting Dataset Shift
Stephan Rabanser, Stephan Günnemann, Zachary C. Lipton
stat.MLcs.LGarXiv:1810.11953v42018One Thousand and One Hours: Self-driving Motion Prediction Dataset
John Houston, Guido Zuidhof, Luca Bergamini +6
cs.CVcs.LGcs.ROarXiv:2006.14480v22020Discovering Dual-Origin Slow Wind from Solar Orbiter with Self-Supervised Contrastive Learning
Henry Han, Jorge Yero Salazar
astro-ph.SRcs.AIcs.LGarXiv:2608.22065v12026Quasi-Dense Similarity Learning for Multiple Object Tracking
Jiangmiao Pang, Linlu Qiu, Xia Li +4
cs.CVcs.LGarXiv:2006.06664v42020Multiplicative Normalizing Flows for Variational Bayesian Neural Networks
Christos Louizos, Max Welling
stat.MLcs.LGarXiv:1703.01961v22017GreenLeaf Law Embed Tiny: A Compact Embedding Model for Legal Domain Retrieval
Surya Saka
cs.LGcs.AIcs.CLarXiv:2608.24936v12026Generic Attention-model Explainability for Interpreting Bi-Modal and Encoder-Decoder Transformers
Hila Chefer, Shir Gur, Lior Wolf
cs.CVcs.LGarXiv:2103.15679v12021Optimizing Millions of Hyperparameters by Implicit Differentiation
Jonathan Lorraine, Paul Vicol, David Duvenaud
cs.LGstat.MLarXiv:1911.02590v12019Last Layer Re-Training is Sufficient for Robustness to Spurious Correlations
Polina Kirichenko, Pavel Izmailov, Andrew Gordon Wilson
cs.LGcs.CVstat.MLarXiv:2204.02937v22022A Survey on Graph Kernels
Nils M. Kriege, Fredrik D. Johansson, Christopher Morris
cs.LGstat.MLarXiv:1903.11835v22019Unnatural Instructions: Tuning Language Models with (Almost) No Human Labor
Or Honovich, Thomas Scialom, Omer Levy +1
cs.CLcs.AIcs.LGarXiv:2212.09689v12022COVID-ResNet: A Deep Learning Framework for Screening of COVID19 from Radiographs
Muhammad Farooq, Abdul Hafeez
eess.IVcs.CVcs.LGarXiv:2003.14395v12020No More Pesky Learning Rates
Tom Schaul, Sixin Zhang, Yann LeCun
stat.MLcs.LGarXiv:1206.1106v22012Two-branch Recurrent Network for Isolating Deepfakes in Videos
Iacopo Masi, Aditya Killekar, Royston Marian Mascarenhas +2
cs.CVcs.CYcs.LGarXiv:2008.03412v32020Mitigating Explanation Leakage in Financial Fraud Detection Systems
Muhammad Waleed Gul, Elaheh Homayounvala
cs.LGcs.CRarXiv:2608.22607v12026From Symmetry to Invariance: Learning Galois Equivalent Representations in Finite Fields
Zheng Zhang, Na Zhang
cs.LGarXiv:2608.22513v12026Rainbow Memory: Continual Learning with a Memory of Diverse Samples
Jihwan Bang, Heesu Kim, YoungJoon Yoo +2
cs.CVcs.LGarXiv:2103.17230v12021Adversarially Regularized Graph Autoencoder for Graph Embedding
Shirui Pan, Ruiqi Hu, Guodong Long +3
cs.LGstat.MLarXiv:1802.04407v22018When Test-Time Adaptation Helps, Harms, or Becomes Inactive: A Condition-Level Study on CIFAR-10-C
Sreeja Guha Majumdar, Aratrika Saha
cs.LGcs.CVarXiv:2608.22233v12026FreeMatch: Self-adaptive Thresholding for Semi-supervised Learning
Yidong Wang, Hao Chen, Qiang Heng +9
cs.LGcs.CVarXiv:2205.07246v32022Slicing Aided Hyper Inference and Fine-tuning for Small Object Detection
Fatih Cagatay Akyon, Sinan Onur Altinuc, Alptekin Temizel
cs.CVcs.LGarXiv:2202.06934v52022Shallow and Deep Convolutional Networks for Saliency Prediction
Junting Pan, Kevin McGuinness, Elisa Sayrol +2
cs.CVcs.LGarXiv:1603.00845v12016Structured Attention Networks
Yoon Kim, Carl Denton, Luong Hoang +1
cs.CLcs.LGcs.NEarXiv:1702.00887v32017Convolutional neural networks with low-rank regularization
Cheng Tai, Tong Xiao, Yi Zhang +2
cs.LGcs.CVstat.MLarXiv:1511.06067v32015Do Adversarially Robust ImageNet Models Transfer Better?
Hadi Salman, Andrew Ilyas, Logan Engstrom +2
cs.CVcs.LGstat.MLarXiv:2007.08489v22020Machine learning in cardiovascular flows modeling: Predicting arterial blood pressure from non-invasive 4D flow MRI data using physics-informed neural networks
Georgios Kissas, Yibo Yang, Eileen Hwuang +3
cs.LGstat.MLarXiv:1905.04817v22019Multi-Agent Cooperation and the Emergence of (Natural) Language
Angeliki Lazaridou, Alexander Peysakhovich, Marco Baroni
cs.CLcs.CVcs.GTarXiv:1612.07182v22016Learning feed-forward one-shot learners
Luca Bertinetto, João F. Henriques, Jack Valmadre +2
cs.CVcs.LGarXiv:1606.05233v12016Does a Modern-Handwriting Warm-Up Help Historical Arabic OCR? A Reproducible, Compute-Matched Evaluation on Muharaf and KHATT
Sumaih Almarshad, Maram Alamri, Dona Aloraini +4
cs.LGcs.CVarXiv:2608.22316v12026Deep Parametric Continuous Convolutional Neural Networks
Shenlong Wang, Simon Suo, Wei-Chiu Ma +2
cs.CVcs.AIcs.LGarXiv:2101.06742v12021No Fear of Heterogeneity: Classifier Calibration for Federated Learning with Non-IID Data
Mi Luo, Fei Chen, Dapeng Hu +3
cs.LGcs.CVcs.DCarXiv:2106.05001v22021Counterfactual Evaluation of Temporal Observation Protocols
Xizhe Zhang
cs.LGarXiv:2608.22221v12026KONTOGRAPH: Verified Point-in-Time Feature Consistency and Amortised Explanation for Real-Time Anti-Money Laundering under a 200 ms Decision Budget
Ahmed Abolfadl
cs.CRcs.AIcs.LGarXiv:2608.22389v12026NGBoost: Natural Gradient Boosting for Probabilistic Prediction
Tony Duan, Anand Avati, Daisy Yi Ding +4
cs.LGstat.MLarXiv:1910.03225v42019Privacy Amplification by Subsampling: Tight Analyses via Couplings and Divergences
Borja Balle, Gilles Barthe, Marco Gaboardi
cs.LGcs.CRstat.MLarXiv:1807.01647v22018Explaining Neural Scaling Laws
Yasaman Bahri, Ethan Dyer, Jared Kaplan +2
cs.LGcond-mat.dis-nnstat.MLarXiv:2102.06701v22021Power Hungry Processing: Watts Driving the Cost of AI Deployment?
Alexandra Sasha Luccioni, Yacine Jernite, Emma Strubell
cs.LGarXiv:2311.16863v32023Dynamics-Aware Unsupervised Discovery of Skills
Archit Sharma, Shixiang Gu, Sergey Levine +2
cs.LGcs.ROstat.MLarXiv:1907.01657v22019Disparities in Dermatology AI Performance on a Diverse, Curated Clinical Image Set
Roxana Daneshjou, Kailas Vodrahalli, Roberto A Novoa +16
eess.IVcs.AIcs.CVarXiv:2203.08807v12022