Machine Learning
Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
16,381 to 16,440 of 20,192
MobileFaceNets: Efficient CNNs for Accurate Real-Time Face Verification on Mobile Devices
Sheng Chen, Yang Liu, Xiang Gao +1
cs.CVcs.LGarXiv:1804.07573v42018Self-Supervised Graph Representation Learning for In-The-Wild Wearable and Smartphone based Emotion Recognition
Ioannis N. Ziogas, Leontios J. Hadjileontiadis, Ahsan H. Khandoker +1
cs.LGcs.AIeess.SParXiv:2608.22387v12026Large Language Models Struggle to Learn Long-Tail Knowledge
Nikhil Kandpal, Haikang Deng, Adam Roberts +2
cs.CLcs.LGarXiv:2211.08411v22022Deep Learning of Representations: Looking Forward
Yoshua Bengio
cs.LGarXiv:1305.0445v22013Artificial Entanglement in the Fine-Tuning of Large Language Models
Min Chen, Zihan Wang, Canyu Chen +3
cs.LGcs.AIhep-tharXiv:2601.06788v12026DeepGauge: Multi-Granularity Testing Criteria for Deep Learning Systems
Lei Ma, Felix Juefei-Xu, Fuyuan Zhang +9
cs.SEcs.CRcs.LGarXiv:1803.07519v42018Variational Lossy Autoencoder
Xi Chen, Diederik P. Kingma, Tim Salimans +5
cs.LGstat.MLarXiv:1611.02731v22016What makes ImageNet good for transfer learning?
Minyoung Huh, Pulkit Agrawal, Alexei A. Efros
cs.CVcs.AIcs.LGarXiv:1608.08614v22016Membership Inference Attacks on Machine Learning: A Survey
Hongsheng Hu, Zoran Salcic, Lichao Sun +3
cs.LGcs.CRarXiv:2103.07853v42021Scalable quantum simulation of continuous-time generative models via tensor networks
Nathan X. Kodama, L. Andrew Wray, Sam Cochran +3
quant-phcs.AIcs.LGarXiv:2608.21700v12026robosuite: A Modular Simulation Framework and Benchmark for Robot Learning
Yuke Zhu, Josiah Wong, Ajay Mandlekar +6
cs.ROcs.AIcs.LGarXiv:2009.12293v32020Learning from positive and unlabeled data: a survey
Jessa Bekker, Jesse Davis
cs.LGstat.MLarXiv:1811.04820v32018Memory-V2V: Memory-Augmented Video-to-Video Diffusion for Consistent Multi-Turn Editing
Dohun Lee, Chun-Hao Paul Huang, Xuelin Chen +3
cs.CVcs.AIcs.LGarXiv:2601.16296v22026Towards Out-Of-Distribution Generalization: A Survey
Jiashuo Liu, Zheyan Shen, Yue He +4
cs.LGarXiv:2108.13624v22021Simple Black-box Adversarial Attacks
Chuan Guo, Jacob R. Gardner, Yurong You +2
cs.LGcs.CRstat.MLarXiv:1905.07121v22019Explainable AI (XAI): A Systematic Meta-Survey of Current Challenges and Future Opportunities
Waddah Saeed, Christian Omlin
cs.LGcs.AIarXiv:2111.06420v12021A Survey on Metric Learning for Feature Vectors and Structured Data
Aurélien Bellet, Amaury Habrard, Marc Sebban
cs.LGcs.AIstat.MLarXiv:1306.6709v42013Applied Federated Learning: Improving Google Keyboard Query Suggestions
Timothy Yang, Galen Andrew, Hubert Eichner +5
cs.LGstat.MLarXiv:1812.02903v12018RM -RF: Reward Model for Run-Free Unit Test Evaluation
Elena Bruches, Daniil Grebenkin, Mikhail Klementev +8
cs.SEcs.LGarXiv:2601.13097v12026Distilling a Neural Network Into a Soft Decision Tree
Nicholas Frosst, Geoffrey Hinton
cs.LGcs.AIstat.MLarXiv:1711.09784v12017Grokking: Generalization Beyond Overfitting on Small Algorithmic Datasets
Alethea Power, Yuri Burda, Harri Edwards +2
cs.LGarXiv:2201.02177v12022Segment Length Matters: A Study of Segment Lengths on Audio Fingerprinting Performance
Ziling Gong, Yunyan Ouyang, Iram Kamdar +5
cs.SDcs.AIcs.IRarXiv:2601.17690v12026Domain Generalization for Object Recognition with Multi-task Autoencoders
Muhammad Ghifary, W. Bastiaan Kleijn, Mengjie Zhang +1
cs.CVcs.AIcs.LGarXiv:1508.07680v12015TensorLens: End-to-End Transformer Analysis via High-Order Attention Tensors
Ido Andrew Atad, Itamar Zimerman, Shahar Katz +1
cs.LGarXiv:2601.17958v12026Estimating Training Data Influence by Tracing Gradient Descent
Garima Pruthi, Frederick Liu, Mukund Sundararajan +1
cs.LGstat.MLarXiv:2002.08484v32020Hopfield Networks is All You Need
Hubert Ramsauer, Bernhard Schäfl, Johannes Lehner +13
cs.NEcs.CLcs.LGarXiv:2008.02217v32020Artificial Intelligence in the Creative Industries: A Review
Nantheera Anantrasirichai, David Bull
cs.CVcs.AIcs.LGarXiv:2007.12391v62020Hierarchical Exponential-Gaussian Mixtures for Watch-Time Distribution Prediction
Sofia Gulevskaia, Mikhail Trapeznikov, Aleksandr Poslavsky +1
cs.IRcs.LGstat.MLarXiv:2608.23356v12026Inertial Manifold Neural Operator for Dissipative Time-Dependent Partial Differential Equations
Xiaoyang Xie, Clarence W. Rowley
math.NAcs.LGmath.DSarXiv:2608.23546v12026Neural Architecture Optimization
Renqian Luo, Fei Tian, Tao Qin +2
cs.LGstat.MLarXiv:1808.07233v52018Spicing up Genetic Netlist Generation with LLMs
Stefan Uhlich, Yağız Gençer, Andrea Bonetti +4
cs.NEcs.ARcs.LGarXiv:2608.23317v12026EGAMA-RC: Risk-Calibrated Evidence-Gated Adaptive Malware Analysis for Robust and Interpretable Memory-Forensic Triage
Isaac Kofi Nti
cs.CRcs.LGarXiv:2608.22721v12026Stochastic gradient descent on Riemannian manifolds
Silvere Bonnabel
math.OCcs.LGstat.MLarXiv:1111.5280v42011Making AI Forget You: Data Deletion in Machine Learning
Antonio Ginart, Melody Y. Guan, Gregory Valiant +1
cs.LGstat.MLarXiv:1907.05012v22019Bottom-Up Abstractive Summarization
Sebastian Gehrmann, Yuntian Deng, Alexander M. Rush
cs.CLcs.AIcs.LGarXiv:1808.10792v22018Rotation-invariant convolutional neural networks for galaxy morphology prediction
Sander Dieleman, Kyle W. Willett, Joni Dambre
astro-ph.IMastro-ph.GAcs.CVarXiv:1503.07077v12015Selective Steering: Norm-Preserving Control Through Discriminative Layer Selection
Quy-Anh Dang, Chris Ngo
cs.LGcs.AIarXiv:2601.19375v12026ConceptMoE: Adaptive Token-to-Concept Compression for Implicit Compute Allocation
Zihao Huang, Jundong Zhou, Xingwei Qu +2
cs.LGarXiv:2601.21420v12026SARAH: A Novel Method for Machine Learning Problems Using Stochastic Recursive Gradient
Lam M. Nguyen, Jie Liu, Katya Scheinberg +1
stat.MLcs.LGmath.OCarXiv:1703.00102v22017Deep Learning in Multimodal Remote Sensing Data Fusion: A Comprehensive Review
Jiaxin Li, Danfeng Hong, Lianru Gao +4
cs.CVcs.LGeess.SParXiv:2205.01380v12022Routing the Lottery: Adaptive Subnetworks for Heterogeneous Data
Grzegorz Stefanski, Alberto Presta, Michal Byra
cs.AIcs.CVcs.LGarXiv:2601.22141v12026Quantum Reservoir Computing with Physics-Informed Correction for Reduced-Order PDE Forecasting
Krishna Bhatia, Harsh, Shalini Devendrababu
quant-phcs.LGarXiv:2608.23119v12026Clipping-Free Policy Optimization for Large Language Models
Ömer Veysel Çağatan, Barış Akgün, Gözde Gül Şahin +1
cs.LGarXiv:2601.22801v12026How Far Ahead Do LLMs Plan? Uncovering the Latent Horizon in Chain-of-Thought Reasoning
Liyan Xu, Mo Yu, Fandong Meng +1
cs.LGcs.CLarXiv:2602.02103v22026Self-Rewarding Sequential Monte Carlo for Masked Diffusion Language Models
Ziwei Luo, Ziqi Jin, Lei Wang +2
cs.LGarXiv:2602.01849v12026DASH: Faster Shampoo via Batched Block Preconditioning and Efficient Inverse-Root Solvers
Ionut-Vlad Modoranu, Philip Zmushko, Erik Schultheis +2
cs.LGarXiv:2602.02016v22026Neural-Symbolic VQA: Disentangling Reasoning from Vision and Language Understanding
Kexin Yi, Jiajun Wu, Chuang Gan +3
cs.AIcs.CLcs.CVarXiv:1810.02338v22018ECHO-2: A Large-Scale Distributed Rollout Framework for Cost-Efficient Reinforcement Learning
Jingwei Song, Meng Chen, Jie Xiao +15
cs.LGcs.DCarXiv:2602.02192v52026Reliable and Responsible Foundation Models: A Comprehensive Survey
Xinyu Yang, Junlin Han, Rishi Bommasani +49
cs.LGcs.AIcs.CLarXiv:2602.08145v12026Agent-Omit: Adaptive Context Omission for Efficient LLM Agents
Yansong Ning, Jun Fang, Naiqiang Tan +1
cs.AIcs.LGarXiv:2602.04284v22026Making Expert Reasoning Learnable with Self-Distillation
Ethan Mendes, Jungsoo Park, Alan Ritter
cs.LGcs.AIarXiv:2602.02405v22026"I May Not Have Articulated Myself Clearly": Diagnosing Dynamic Instability in LLM Reasoning at Inference Time
Jinkun Chen, Fengxiang Cheng, Sijia Han +1
cs.AIcs.LGarXiv:2602.02863v12026Token Sparse Attention: Efficient Long-Context Inference with Interleaved Token Selection
Dongwon Jo, Beomseok Kang, Jiwon Song +1
cs.CLcs.LGarXiv:2602.03216v32026Training Data Efficiency in Multimodal Process Reward Models
Jinyuan Li, Chengsong Huang, Langlin Huang +4
cs.LGcs.CLcs.MMarXiv:2602.04145v22026ChatGPT for Robotics: Design Principles and Model Abilities
Sai Vemprala, Rogerio Bonatti, Arthur Bucker +1
cs.AIcs.CLcs.HCarXiv:2306.17582v22023In Search of the Real Inductive Bias: On the Role of Implicit Regularization in Deep Learning
Behnam Neyshabur, Ryota Tomioka, Nathan Srebro
cs.LGcs.AIcs.CVarXiv:1412.6614v42014Machine Learning in Python: Main developments and technology trends in data science, machine learning, and artificial intelligence
Sebastian Raschka, Joshua Patterson, Corey Nolet
cs.LGstat.MLarXiv:2002.04803v22020Multi-agent Reinforcement Learning in Sequential Social Dilemmas
Joel Z. Leibo, Vinicius Zambaldi, Marc Lanctot +2
cs.MAcs.AIcs.GTarXiv:1702.03037v12017SplitLite: Low-Rank Residual Compression for Split Learning
Tao Li, Yulin Tang, Qi Guo +1
cs.LGcs.AIarXiv:2608.23018v12026Efficiently Scaling Transformer Inference
Reiner Pope, Sholto Douglas, Aakanksha Chowdhery +7
cs.LGcs.CLarXiv:2211.05102v12022