Machine Learning

Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

16,021 to 16,080 of 20,193

  1. SPINAL -- Scaling-law and Preference Integration in Neural Alignment Layers

    Arion Das, Partha Pratim Saha, Amit Dhanda +3

    cs.LGcs.AIcs.CLarXiv:2601.06238v12026
  2. Challenges in Deploying Machine Learning: a Survey of Case Studies

    Andrei Paleyes, Raoul-Gabriel Urma, Neil D. Lawrence

    cs.LGarXiv:2011.09926v32020
  3. Neural Operator: Learning Maps Between Function Spaces

    Nikola Kovachki, Zongyi Li, Burigede Liu +4

    cs.LGmath.NAarXiv:2108.08481v62021
  4. Disentangled Skill Representations for Predictive Human Modeling

    Mariah Schrum, Deepak Gopinath, Srijan Srivatsa +2

    cs.LGcs.AIarXiv:2608.23776v12026
  5. Gecko: An Efficient Neural Architecture Inherently Processing Sequences with Arbitrary Lengths

    Xuezhe Ma, Shicheng Wen, Linghao Jin +11

    cs.LGcs.CLarXiv:2601.06463v12026
  6. A Feature-Major Codebook for Memory-Efficient Sparse-Binary Self-Organizing Maps: Scaling a MEDLINE Atlas to 1.05 Million Neurons on a Single Consumer GPU

    Andrew James Amos

    cs.LGcs.DCarXiv:2608.24067v12026
  7. Short-Term Forecasting of Passenger Demand under On-Demand Ride Services: A Spatio-Temporal Deep Learning Approach

    Jintao Ke, Hongyu Zheng, Hai Yang +2

    cs.LGarXiv:1706.06279v12017
  8. The Natural Language Decathlon: Multitask Learning as Question Answering

    Bryan McCann, Nitish Shirish Keskar, Caiming Xiong +1

    cs.CLcs.AIcs.LGarXiv:1806.08730v12018
  9. FrugalSOT - Frugal Search Over the Models

    Pradheep P, Yuvanesh S, Harish KB +4

    cs.LGarXiv:2608.21621v12026
  10. Analysis Methods in Neural Language Processing: A Survey

    Yonatan Belinkov, James Glass

    cs.CLcs.LGcs.NEarXiv:1812.08951v22018
  11. A Framework for Understanding Sources of Harm throughout the Machine Learning Life Cycle

    Harini Suresh, John V. Guttag

    cs.LGstat.MLarXiv:1901.10002v52019
  12. Data-Efficient Off-Policy Policy Evaluation for Reinforcement Learning

    Philip S. Thomas, Emma Brunskill

    cs.LGcs.AIarXiv:1604.00923v12016
  13. RetrievalFormer: A Dual-Encoder Transformer for Efficient Approximate Nearest Neighbor Retrieval and Cold-Item Recommendation

    Theodore Rogers, Joe Standerfer, Dmitrii Timoshenko +3

    cs.IRcs.LGarXiv:2608.24079v12026
  14. TabDDPM: Modelling Tabular Data with Diffusion Models

    Akim Kotelnikov, Dmitry Baranchuk, Ivan Rubachev +1

    cs.LGarXiv:2209.15421v22022
  15. Learning Confidence for Out-of-Distribution Detection in Neural Networks

    Terrance DeVries, Graham W. Taylor

    stat.MLcs.LGarXiv:1802.04865v12018
  16. Adversarial Attacks on Graph Neural Networks via Meta Learning

    Daniel Zügner, Stephan Günnemann

    cs.LGcs.CRstat.MLarXiv:1902.08412v22019
  17. By-passing the Kohn-Sham equations with machine learning

    Felix Brockherde, Leslie Vogt, Li Li +3

    physics.comp-phcs.LGphysics.chem-pharXiv:1609.02815v32016
  18. KAGE-Bench: Fast Known-Axis Visual Generalization Evaluation for Reinforcement Learning

    Egor Cherepanov, Daniil Zelezetsky, Alexey K. Kovalev +1

    cs.LGcs.AIcs.CVarXiv:2601.14232v22026
  19. Fairness Without Demographics in Repeated Loss Minimization

    Tatsunori B. Hashimoto, Megha Srivastava, Hongseok Namkoong +1

    stat.MLcs.LGarXiv:1806.08010v22018
  20. CLEVRER: CoLlision Events for Video REpresentation and Reasoning

    Kexin Yi, Chuang Gan, Yunzhu Li +4

    cs.CVcs.AIcs.CLarXiv:1910.01442v22019
  21. PredRNN: A Recurrent Neural Network for Spatiotemporal Predictive Learning

    Yunbo Wang, Haixu Wu, Jianjin Zhang +4

    cs.LGcs.CVarXiv:2103.09504v42021
  22. Deep Voice: Real-time Neural Text-to-Speech

    Sercan O. Arik, Mike Chrzanowski, Adam Coates +9

    cs.CLcs.LGcs.NEarXiv:1702.07825v22017
  23. METIS: Mentoring Engine for Thoughtful Inquiry & Solutions

    Abhinav Rajeev Kumar, Dhruv Trehan, Paras Chopra

    cs.LGcs.AIarXiv:2601.13075v12026
  24. Learning with Pseudo-Ensembles

    Philip Bachman, Ouais Alsharif, Doina Precup

    stat.MLcs.LGcs.NEarXiv:1412.4864v12014
  25. An Evaluation Dataset for Intent Classification and Out-of-Scope Prediction

    Stefan Larson, Anish Mahendran, Joseph J. Peper +8

    cs.CLcs.AIcs.LGarXiv:1909.02027v12019
  26. Low-Rank Ternary Adaptation for Fine-Tuning Transformers

    Alexandru-Dragos Manolache, Yunqiang Li, Jan van Gemert

    cs.CVcs.LGarXiv:2608.24469v12026
  27. EvasionBench: A Large-Scale Benchmark for Detecting Managerial Evasion in Earnings Call Q&A

    Shijian Ma, Yan Lin, Yi Yang

    cs.LGcs.CLarXiv:2601.09142v22026
  28. Cluster Workload Allocation: Semantic Soft Affinity Using Natural Language Processing

    Leszek Sliwko, Jolanta Mizeria-Pietraszko

    cs.AIcs.DCcs.LGarXiv:2601.09282v22026
  29. Improving Reproducibility in Machine Learning Research (A Report from the NeurIPS 2019 Reproducibility Program)

    Joelle Pineau, Philippe Vincent-Lamarre, Koustuv Sinha +5

    cs.LGstat.MLarXiv:2003.12206v42020
  30. Adversarial Risk and the Dangers of Evaluating Against Weak Attacks

    Jonathan Uesato, Brendan O'Donoghue, Aaron van den Oord +1

    cs.LGcs.CRstat.MLarXiv:1802.05666v22018
  31. Beyond Cosine Similarity: Taming Semantic Drift and Antonym Intrusion in a 15-Million Node Turkish Synonym Graph

    Ebubekir Tosun, Mehmet Emin Buldur, Özay Ezerceli +1

    cs.CLcs.LGarXiv:2601.13251v12026
  32. A Hybrid Protocol for Large-Scale Semantic Dataset Generation in Low-Resource Languages: The Turkish Semantic Relations Corpus

    Ebubekir Tosun, Mehmet Emin Buldur, Özay Ezerceli +1

    cs.CLcs.LGarXiv:2601.13253v12026
  33. Knowledge-aware Graph Neural Networks with Label Smoothness Regularization for Recommender Systems

    Hongwei Wang, Fuzheng Zhang, Mengdi Zhang +4

    cs.LGcs.IRstat.MLarXiv:1905.04413v32019
  34. Uncertainty-Aware Gradient Signal-to-Noise Data Selection for Instruction Tuning

    Zhihang Yuan, Chengyu Yue, Long Huang +2

    cs.CLcs.AIcs.LGarXiv:2601.13697v12026
  35. Low-Latency Activation-Regularized Sparse Neural Operators with Distillation Assistance Towards Real-Time Edge-Deployable Virtual Sensing

    William Howes, Farid Ahmed, Syed Bahauddin Alam

    cs.LGarXiv:2608.23987v12026
  36. AdaBelief Optimizer: Adapting Stepsizes by the Belief in Observed Gradients

    Juntang Zhuang, Tommy Tang, Yifan Ding +4

    cs.LGcs.CVstat.MLarXiv:2010.07468v52020
  37. AR-Omni: A Unified Autoregressive Model for Any-to-Any Generation

    Dongjie Cheng, Ruifeng Yuan, Yongqi Li +5

    cs.LGcs.AIcs.CLarXiv:2601.17761v12026
  38. A Variational Perspective on Accelerated Methods in Optimization

    Andre Wibisono, Ashia C. Wilson, Michael I. Jordan

    math.OCcs.LGstat.MLarXiv:1603.04245v12016
  39. The KL-UCB Algorithm for Bounded Stochastic Bandits and Beyond

    Aurélien Garivier, Olivier Cappé

    math.STcs.LGeess.SYarXiv:1102.2490v52011
  40. Machine Learning Operations (MLOps): Overview, Definition, and Architecture

    Dominik Kreuzberger, Niklas Kühl, Sebastian Hirschl

    cs.LGarXiv:2205.02302v32022
  41. Learning to Adapt in Dynamic, Real-World Environments Through Meta-Reinforcement Learning

    Anusha Nagabandi, Ignasi Clavera, Simin Liu +4

    cs.LGcs.ROstat.MLarXiv:1803.11347v62018
  42. RSA: Byzantine-Robust Stochastic Aggregation Methods for Distributed Learning from Heterogeneous Datasets

    Liping Li, Wei Xu, Tianyi Chen +2

    cs.LGcs.CRcs.MAarXiv:1811.03761v22018
  43. InterpretML: A Unified Framework for Machine Learning Interpretability

    Harsha Nori, Samuel Jenkins, Paul Koch +1

    cs.LGstat.MLarXiv:1909.09223v12019
  44. GameTalk: Training LLMs for Strategic Conversation

    Victor Conchello Vendrell, Max Ruiz Luyten, Mihaela van der Schaar

    cs.CLcs.AIcs.GTarXiv:2601.16276v12026
  45. Context-Aware Attentive Knowledge Tracing

    Aritra Ghosh, Neil Heffernan, Andrew S. Lan

    cs.LGcs.AIarXiv:2007.12324v12020
  46. Three Approaches for Personalization with Applications to Federated Learning

    Yishay Mansour, Mehryar Mohri, Jae Ro +1

    cs.LGstat.MLarXiv:2002.10619v22020
  47. ChemRL-GEM: Geometry Enhanced Molecular Representation Learning for Property Prediction

    Xiaomin Fang, Lihang Liu, Jieqiong Lei +6

    cs.LGphysics.chem-phq-bio.MNarXiv:2106.06130v42021
  48. GPCR-Filter: a deep learning framework for efficient and precise GPCR modulator discovery

    Jingjie Ning, Xiangzhen Shen, Li Hou +8

    cs.LGq-bio.QMarXiv:2601.19149v22026
  49. Adversarial Logit Pairing

    Harini Kannan, Alexey Kurakin, Ian Goodfellow

    cs.LGstat.MLarXiv:1803.06373v12018
  50. Is Conditional Generative Modeling all you need for Decision-Making?

    Anurag Ajay, Yilun Du, Abhi Gupta +3

    cs.LGcs.AIarXiv:2211.15657v42022
  51. Physics-guided Neural Networks (PGNN): An Application in Lake Temperature Modeling

    Arka Daw, Anuj Karpatne, William Watkins +2

    cs.LGcs.AIcs.CVarXiv:1710.11431v32017
  52. Latent Adversarial Regularization for Offline Preference Optimization

    Enyi Jiang, Yibo Jacky Zhang, Yinglun Xu +3

    cs.LGcs.AIarXiv:2601.22083v22026
  53. The FLUXCOM ensemble of global land-atmosphere energy fluxes

    Martin Jung, Sujan Koirala, Ulrich Weber +7

    physics.ao-phcs.LGstat.MLarXiv:1812.04951v12018
  54. Learning What to Predict: Downstream-Guided Task Design for Continued Pretraining

    Shuqi Ke, Giulia Fanti

    cs.LGcs.AIarXiv:2601.22108v22026
  55. Enhancing Bayesian Optimization and Active Learning Through Kernel Diversity

    Heng Zhang, Haotian Xiang, Qin Lu +2

    cs.LGcs.AIarXiv:2608.24721v12026
  56. Mind the Student: Behavioral and Contextual Cues for Automated Engagement Prediction in Online Learning

    Alperen Kantarci, Visvanathan Ramesh, Gemma Roig

    cs.CVcs.AIcs.HCarXiv:2608.24340v12026
  57. Automata from Agent Traces: Failure and Next-Step Prediction

    Seonglae Cho, Franklin Cardenoso Fernandez, Umar Mohammed +4

    cs.AIcs.CLcs.LGarXiv:2608.23670v12026
  58. RAPTOR: Ridge-Adaptive Logistic Probes

    Ziqi Gao, Yaotian Zhu, Qingcheng Zeng +4

    cs.LGcs.AIarXiv:2602.00158v22026
  59. nocaps: novel object captioning at scale

    Harsh Agrawal, Karan Desai, Yufei Wang +7

    cs.CVcs.AIcs.CLarXiv:1812.08658v32018
  60. Not All Samples Are Created Equal: Deep Learning with Importance Sampling

    Angelos Katharopoulos, François Fleuret

    cs.LGarXiv:1803.00942v32018