Machine Learning

Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

16,381 to 16,440 of 20,192

  1. MobileFaceNets: Efficient CNNs for Accurate Real-Time Face Verification on Mobile Devices

    Sheng Chen, Yang Liu, Xiang Gao +1

    cs.CVcs.LGarXiv:1804.07573v42018
  2. Self-Supervised Graph Representation Learning for In-The-Wild Wearable and Smartphone based Emotion Recognition

    Ioannis N. Ziogas, Leontios J. Hadjileontiadis, Ahsan H. Khandoker +1

    cs.LGcs.AIeess.SParXiv:2608.22387v12026
  3. Large Language Models Struggle to Learn Long-Tail Knowledge

    Nikhil Kandpal, Haikang Deng, Adam Roberts +2

    cs.CLcs.LGarXiv:2211.08411v22022
  4. Deep Learning of Representations: Looking Forward

    Yoshua Bengio

    cs.LGarXiv:1305.0445v22013
  5. Artificial Entanglement in the Fine-Tuning of Large Language Models

    Min Chen, Zihan Wang, Canyu Chen +3

    cs.LGcs.AIhep-tharXiv:2601.06788v12026
  6. DeepGauge: Multi-Granularity Testing Criteria for Deep Learning Systems

    Lei Ma, Felix Juefei-Xu, Fuyuan Zhang +9

    cs.SEcs.CRcs.LGarXiv:1803.07519v42018
  7. Variational Lossy Autoencoder

    Xi Chen, Diederik P. Kingma, Tim Salimans +5

    cs.LGstat.MLarXiv:1611.02731v22016
  8. What makes ImageNet good for transfer learning?

    Minyoung Huh, Pulkit Agrawal, Alexei A. Efros

    cs.CVcs.AIcs.LGarXiv:1608.08614v22016
  9. Membership Inference Attacks on Machine Learning: A Survey

    Hongsheng Hu, Zoran Salcic, Lichao Sun +3

    cs.LGcs.CRarXiv:2103.07853v42021
  10. Scalable quantum simulation of continuous-time generative models via tensor networks

    Nathan X. Kodama, L. Andrew Wray, Sam Cochran +3

    quant-phcs.AIcs.LGarXiv:2608.21700v12026
  11. robosuite: A Modular Simulation Framework and Benchmark for Robot Learning

    Yuke Zhu, Josiah Wong, Ajay Mandlekar +6

    cs.ROcs.AIcs.LGarXiv:2009.12293v32020
  12. Learning from positive and unlabeled data: a survey

    Jessa Bekker, Jesse Davis

    cs.LGstat.MLarXiv:1811.04820v32018
  13. Memory-V2V: Memory-Augmented Video-to-Video Diffusion for Consistent Multi-Turn Editing

    Dohun Lee, Chun-Hao Paul Huang, Xuelin Chen +3

    cs.CVcs.AIcs.LGarXiv:2601.16296v22026
  14. Towards Out-Of-Distribution Generalization: A Survey

    Jiashuo Liu, Zheyan Shen, Yue He +4

    cs.LGarXiv:2108.13624v22021
  15. Simple Black-box Adversarial Attacks

    Chuan Guo, Jacob R. Gardner, Yurong You +2

    cs.LGcs.CRstat.MLarXiv:1905.07121v22019
  16. Explainable AI (XAI): A Systematic Meta-Survey of Current Challenges and Future Opportunities

    Waddah Saeed, Christian Omlin

    cs.LGcs.AIarXiv:2111.06420v12021
  17. A Survey on Metric Learning for Feature Vectors and Structured Data

    Aurélien Bellet, Amaury Habrard, Marc Sebban

    cs.LGcs.AIstat.MLarXiv:1306.6709v42013
  18. Applied Federated Learning: Improving Google Keyboard Query Suggestions

    Timothy Yang, Galen Andrew, Hubert Eichner +5

    cs.LGstat.MLarXiv:1812.02903v12018
  19. RM -RF: Reward Model for Run-Free Unit Test Evaluation

    Elena Bruches, Daniil Grebenkin, Mikhail Klementev +8

    cs.SEcs.LGarXiv:2601.13097v12026
  20. Distilling a Neural Network Into a Soft Decision Tree

    Nicholas Frosst, Geoffrey Hinton

    cs.LGcs.AIstat.MLarXiv:1711.09784v12017
  21. Grokking: Generalization Beyond Overfitting on Small Algorithmic Datasets

    Alethea Power, Yuri Burda, Harri Edwards +2

    cs.LGarXiv:2201.02177v12022
  22. Segment Length Matters: A Study of Segment Lengths on Audio Fingerprinting Performance

    Ziling Gong, Yunyan Ouyang, Iram Kamdar +5

    cs.SDcs.AIcs.IRarXiv:2601.17690v12026
  23. Domain Generalization for Object Recognition with Multi-task Autoencoders

    Muhammad Ghifary, W. Bastiaan Kleijn, Mengjie Zhang +1

    cs.CVcs.AIcs.LGarXiv:1508.07680v12015
  24. TensorLens: End-to-End Transformer Analysis via High-Order Attention Tensors

    Ido Andrew Atad, Itamar Zimerman, Shahar Katz +1

    cs.LGarXiv:2601.17958v12026
  25. Estimating Training Data Influence by Tracing Gradient Descent

    Garima Pruthi, Frederick Liu, Mukund Sundararajan +1

    cs.LGstat.MLarXiv:2002.08484v32020
  26. Hopfield Networks is All You Need

    Hubert Ramsauer, Bernhard Schäfl, Johannes Lehner +13

    cs.NEcs.CLcs.LGarXiv:2008.02217v32020
  27. Artificial Intelligence in the Creative Industries: A Review

    Nantheera Anantrasirichai, David Bull

    cs.CVcs.AIcs.LGarXiv:2007.12391v62020
  28. Hierarchical Exponential-Gaussian Mixtures for Watch-Time Distribution Prediction

    Sofia Gulevskaia, Mikhail Trapeznikov, Aleksandr Poslavsky +1

    cs.IRcs.LGstat.MLarXiv:2608.23356v12026
  29. Inertial Manifold Neural Operator for Dissipative Time-Dependent Partial Differential Equations

    Xiaoyang Xie, Clarence W. Rowley

    math.NAcs.LGmath.DSarXiv:2608.23546v12026
  30. Neural Architecture Optimization

    Renqian Luo, Fei Tian, Tao Qin +2

    cs.LGstat.MLarXiv:1808.07233v52018
  31. Spicing up Genetic Netlist Generation with LLMs

    Stefan Uhlich, Yağız Gençer, Andrea Bonetti +4

    cs.NEcs.ARcs.LGarXiv:2608.23317v12026
  32. EGAMA-RC: Risk-Calibrated Evidence-Gated Adaptive Malware Analysis for Robust and Interpretable Memory-Forensic Triage

    Isaac Kofi Nti

    cs.CRcs.LGarXiv:2608.22721v12026
  33. Stochastic gradient descent on Riemannian manifolds

    Silvere Bonnabel

    math.OCcs.LGstat.MLarXiv:1111.5280v42011
  34. Making AI Forget You: Data Deletion in Machine Learning

    Antonio Ginart, Melody Y. Guan, Gregory Valiant +1

    cs.LGstat.MLarXiv:1907.05012v22019
  35. Bottom-Up Abstractive Summarization

    Sebastian Gehrmann, Yuntian Deng, Alexander M. Rush

    cs.CLcs.AIcs.LGarXiv:1808.10792v22018
  36. Rotation-invariant convolutional neural networks for galaxy morphology prediction

    Sander Dieleman, Kyle W. Willett, Joni Dambre

    astro-ph.IMastro-ph.GAcs.CVarXiv:1503.07077v12015
  37. Selective Steering: Norm-Preserving Control Through Discriminative Layer Selection

    Quy-Anh Dang, Chris Ngo

    cs.LGcs.AIarXiv:2601.19375v12026
  38. ConceptMoE: Adaptive Token-to-Concept Compression for Implicit Compute Allocation

    Zihao Huang, Jundong Zhou, Xingwei Qu +2

    cs.LGarXiv:2601.21420v12026
  39. SARAH: A Novel Method for Machine Learning Problems Using Stochastic Recursive Gradient

    Lam M. Nguyen, Jie Liu, Katya Scheinberg +1

    stat.MLcs.LGmath.OCarXiv:1703.00102v22017
  40. Deep Learning in Multimodal Remote Sensing Data Fusion: A Comprehensive Review

    Jiaxin Li, Danfeng Hong, Lianru Gao +4

    cs.CVcs.LGeess.SParXiv:2205.01380v12022
  41. Routing the Lottery: Adaptive Subnetworks for Heterogeneous Data

    Grzegorz Stefanski, Alberto Presta, Michal Byra

    cs.AIcs.CVcs.LGarXiv:2601.22141v12026
  42. Quantum Reservoir Computing with Physics-Informed Correction for Reduced-Order PDE Forecasting

    Krishna Bhatia, Harsh, Shalini Devendrababu

    quant-phcs.LGarXiv:2608.23119v12026
  43. Clipping-Free Policy Optimization for Large Language Models

    Ömer Veysel Çağatan, Barış Akgün, Gözde Gül Şahin +1

    cs.LGarXiv:2601.22801v12026
  44. How Far Ahead Do LLMs Plan? Uncovering the Latent Horizon in Chain-of-Thought Reasoning

    Liyan Xu, Mo Yu, Fandong Meng +1

    cs.LGcs.CLarXiv:2602.02103v22026
  45. Self-Rewarding Sequential Monte Carlo for Masked Diffusion Language Models

    Ziwei Luo, Ziqi Jin, Lei Wang +2

    cs.LGarXiv:2602.01849v12026
  46. DASH: Faster Shampoo via Batched Block Preconditioning and Efficient Inverse-Root Solvers

    Ionut-Vlad Modoranu, Philip Zmushko, Erik Schultheis +2

    cs.LGarXiv:2602.02016v22026
  47. Neural-Symbolic VQA: Disentangling Reasoning from Vision and Language Understanding

    Kexin Yi, Jiajun Wu, Chuang Gan +3

    cs.AIcs.CLcs.CVarXiv:1810.02338v22018
  48. ECHO-2: A Large-Scale Distributed Rollout Framework for Cost-Efficient Reinforcement Learning

    Jingwei Song, Meng Chen, Jie Xiao +15

    cs.LGcs.DCarXiv:2602.02192v52026
  49. Reliable and Responsible Foundation Models: A Comprehensive Survey

    Xinyu Yang, Junlin Han, Rishi Bommasani +49

    cs.LGcs.AIcs.CLarXiv:2602.08145v12026
  50. Agent-Omit: Adaptive Context Omission for Efficient LLM Agents

    Yansong Ning, Jun Fang, Naiqiang Tan +1

    cs.AIcs.LGarXiv:2602.04284v22026
  51. Making Expert Reasoning Learnable with Self-Distillation

    Ethan Mendes, Jungsoo Park, Alan Ritter

    cs.LGcs.AIarXiv:2602.02405v22026
  52. "I May Not Have Articulated Myself Clearly": Diagnosing Dynamic Instability in LLM Reasoning at Inference Time

    Jinkun Chen, Fengxiang Cheng, Sijia Han +1

    cs.AIcs.LGarXiv:2602.02863v12026
  53. Token Sparse Attention: Efficient Long-Context Inference with Interleaved Token Selection

    Dongwon Jo, Beomseok Kang, Jiwon Song +1

    cs.CLcs.LGarXiv:2602.03216v32026
  54. Training Data Efficiency in Multimodal Process Reward Models

    Jinyuan Li, Chengsong Huang, Langlin Huang +4

    cs.LGcs.CLcs.MMarXiv:2602.04145v22026
  55. ChatGPT for Robotics: Design Principles and Model Abilities

    Sai Vemprala, Rogerio Bonatti, Arthur Bucker +1

    cs.AIcs.CLcs.HCarXiv:2306.17582v22023
  56. In Search of the Real Inductive Bias: On the Role of Implicit Regularization in Deep Learning

    Behnam Neyshabur, Ryota Tomioka, Nathan Srebro

    cs.LGcs.AIcs.CVarXiv:1412.6614v42014
  57. Machine Learning in Python: Main developments and technology trends in data science, machine learning, and artificial intelligence

    Sebastian Raschka, Joshua Patterson, Corey Nolet

    cs.LGstat.MLarXiv:2002.04803v22020
  58. Multi-agent Reinforcement Learning in Sequential Social Dilemmas

    Joel Z. Leibo, Vinicius Zambaldi, Marc Lanctot +2

    cs.MAcs.AIcs.GTarXiv:1702.03037v12017
  59. SplitLite: Low-Rank Residual Compression for Split Learning

    Tao Li, Yulin Tang, Qi Guo +1

    cs.LGcs.AIarXiv:2608.23018v12026
  60. Efficiently Scaling Transformer Inference

    Reiner Pope, Sholto Douglas, Aakanksha Chowdhery +7

    cs.LGcs.CLarXiv:2211.05102v12022