Machine Learning
Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
16,021 to 16,080 of 20,193
SPINAL -- Scaling-law and Preference Integration in Neural Alignment Layers
Arion Das, Partha Pratim Saha, Amit Dhanda +3
cs.LGcs.AIcs.CLarXiv:2601.06238v12026Challenges in Deploying Machine Learning: a Survey of Case Studies
Andrei Paleyes, Raoul-Gabriel Urma, Neil D. Lawrence
cs.LGarXiv:2011.09926v32020Neural Operator: Learning Maps Between Function Spaces
Nikola Kovachki, Zongyi Li, Burigede Liu +4
cs.LGmath.NAarXiv:2108.08481v62021Disentangled Skill Representations for Predictive Human Modeling
Mariah Schrum, Deepak Gopinath, Srijan Srivatsa +2
cs.LGcs.AIarXiv:2608.23776v12026Gecko: An Efficient Neural Architecture Inherently Processing Sequences with Arbitrary Lengths
Xuezhe Ma, Shicheng Wen, Linghao Jin +11
cs.LGcs.CLarXiv:2601.06463v12026A Feature-Major Codebook for Memory-Efficient Sparse-Binary Self-Organizing Maps: Scaling a MEDLINE Atlas to 1.05 Million Neurons on a Single Consumer GPU
Andrew James Amos
cs.LGcs.DCarXiv:2608.24067v12026Short-Term Forecasting of Passenger Demand under On-Demand Ride Services: A Spatio-Temporal Deep Learning Approach
Jintao Ke, Hongyu Zheng, Hai Yang +2
cs.LGarXiv:1706.06279v12017The Natural Language Decathlon: Multitask Learning as Question Answering
Bryan McCann, Nitish Shirish Keskar, Caiming Xiong +1
cs.CLcs.AIcs.LGarXiv:1806.08730v12018FrugalSOT - Frugal Search Over the Models
Pradheep P, Yuvanesh S, Harish KB +4
cs.LGarXiv:2608.21621v12026Analysis Methods in Neural Language Processing: A Survey
Yonatan Belinkov, James Glass
cs.CLcs.LGcs.NEarXiv:1812.08951v22018A Framework for Understanding Sources of Harm throughout the Machine Learning Life Cycle
Harini Suresh, John V. Guttag
cs.LGstat.MLarXiv:1901.10002v52019Data-Efficient Off-Policy Policy Evaluation for Reinforcement Learning
Philip S. Thomas, Emma Brunskill
cs.LGcs.AIarXiv:1604.00923v12016RetrievalFormer: A Dual-Encoder Transformer for Efficient Approximate Nearest Neighbor Retrieval and Cold-Item Recommendation
Theodore Rogers, Joe Standerfer, Dmitrii Timoshenko +3
cs.IRcs.LGarXiv:2608.24079v12026TabDDPM: Modelling Tabular Data with Diffusion Models
Akim Kotelnikov, Dmitry Baranchuk, Ivan Rubachev +1
cs.LGarXiv:2209.15421v22022Learning Confidence for Out-of-Distribution Detection in Neural Networks
Terrance DeVries, Graham W. Taylor
stat.MLcs.LGarXiv:1802.04865v12018Adversarial Attacks on Graph Neural Networks via Meta Learning
Daniel Zügner, Stephan Günnemann
cs.LGcs.CRstat.MLarXiv:1902.08412v22019By-passing the Kohn-Sham equations with machine learning
Felix Brockherde, Leslie Vogt, Li Li +3
physics.comp-phcs.LGphysics.chem-pharXiv:1609.02815v32016KAGE-Bench: Fast Known-Axis Visual Generalization Evaluation for Reinforcement Learning
Egor Cherepanov, Daniil Zelezetsky, Alexey K. Kovalev +1
cs.LGcs.AIcs.CVarXiv:2601.14232v22026Fairness Without Demographics in Repeated Loss Minimization
Tatsunori B. Hashimoto, Megha Srivastava, Hongseok Namkoong +1
stat.MLcs.LGarXiv:1806.08010v22018CLEVRER: CoLlision Events for Video REpresentation and Reasoning
Kexin Yi, Chuang Gan, Yunzhu Li +4
cs.CVcs.AIcs.CLarXiv:1910.01442v22019PredRNN: A Recurrent Neural Network for Spatiotemporal Predictive Learning
Yunbo Wang, Haixu Wu, Jianjin Zhang +4
cs.LGcs.CVarXiv:2103.09504v42021Deep Voice: Real-time Neural Text-to-Speech
Sercan O. Arik, Mike Chrzanowski, Adam Coates +9
cs.CLcs.LGcs.NEarXiv:1702.07825v22017METIS: Mentoring Engine for Thoughtful Inquiry & Solutions
Abhinav Rajeev Kumar, Dhruv Trehan, Paras Chopra
cs.LGcs.AIarXiv:2601.13075v12026Learning with Pseudo-Ensembles
Philip Bachman, Ouais Alsharif, Doina Precup
stat.MLcs.LGcs.NEarXiv:1412.4864v12014An Evaluation Dataset for Intent Classification and Out-of-Scope Prediction
Stefan Larson, Anish Mahendran, Joseph J. Peper +8
cs.CLcs.AIcs.LGarXiv:1909.02027v12019Low-Rank Ternary Adaptation for Fine-Tuning Transformers
Alexandru-Dragos Manolache, Yunqiang Li, Jan van Gemert
cs.CVcs.LGarXiv:2608.24469v12026EvasionBench: A Large-Scale Benchmark for Detecting Managerial Evasion in Earnings Call Q&A
Shijian Ma, Yan Lin, Yi Yang
cs.LGcs.CLarXiv:2601.09142v22026Cluster Workload Allocation: Semantic Soft Affinity Using Natural Language Processing
Leszek Sliwko, Jolanta Mizeria-Pietraszko
cs.AIcs.DCcs.LGarXiv:2601.09282v22026Improving Reproducibility in Machine Learning Research (A Report from the NeurIPS 2019 Reproducibility Program)
Joelle Pineau, Philippe Vincent-Lamarre, Koustuv Sinha +5
cs.LGstat.MLarXiv:2003.12206v42020Adversarial Risk and the Dangers of Evaluating Against Weak Attacks
Jonathan Uesato, Brendan O'Donoghue, Aaron van den Oord +1
cs.LGcs.CRstat.MLarXiv:1802.05666v22018Beyond Cosine Similarity: Taming Semantic Drift and Antonym Intrusion in a 15-Million Node Turkish Synonym Graph
Ebubekir Tosun, Mehmet Emin Buldur, Özay Ezerceli +1
cs.CLcs.LGarXiv:2601.13251v12026A Hybrid Protocol for Large-Scale Semantic Dataset Generation in Low-Resource Languages: The Turkish Semantic Relations Corpus
Ebubekir Tosun, Mehmet Emin Buldur, Özay Ezerceli +1
cs.CLcs.LGarXiv:2601.13253v12026Knowledge-aware Graph Neural Networks with Label Smoothness Regularization for Recommender Systems
Hongwei Wang, Fuzheng Zhang, Mengdi Zhang +4
cs.LGcs.IRstat.MLarXiv:1905.04413v32019Uncertainty-Aware Gradient Signal-to-Noise Data Selection for Instruction Tuning
Zhihang Yuan, Chengyu Yue, Long Huang +2
cs.CLcs.AIcs.LGarXiv:2601.13697v12026Low-Latency Activation-Regularized Sparse Neural Operators with Distillation Assistance Towards Real-Time Edge-Deployable Virtual Sensing
William Howes, Farid Ahmed, Syed Bahauddin Alam
cs.LGarXiv:2608.23987v12026AdaBelief Optimizer: Adapting Stepsizes by the Belief in Observed Gradients
Juntang Zhuang, Tommy Tang, Yifan Ding +4
cs.LGcs.CVstat.MLarXiv:2010.07468v52020AR-Omni: A Unified Autoregressive Model for Any-to-Any Generation
Dongjie Cheng, Ruifeng Yuan, Yongqi Li +5
cs.LGcs.AIcs.CLarXiv:2601.17761v12026A Variational Perspective on Accelerated Methods in Optimization
Andre Wibisono, Ashia C. Wilson, Michael I. Jordan
math.OCcs.LGstat.MLarXiv:1603.04245v12016The KL-UCB Algorithm for Bounded Stochastic Bandits and Beyond
Aurélien Garivier, Olivier Cappé
math.STcs.LGeess.SYarXiv:1102.2490v52011Machine Learning Operations (MLOps): Overview, Definition, and Architecture
Dominik Kreuzberger, Niklas Kühl, Sebastian Hirschl
cs.LGarXiv:2205.02302v32022Learning to Adapt in Dynamic, Real-World Environments Through Meta-Reinforcement Learning
Anusha Nagabandi, Ignasi Clavera, Simin Liu +4
cs.LGcs.ROstat.MLarXiv:1803.11347v62018RSA: Byzantine-Robust Stochastic Aggregation Methods for Distributed Learning from Heterogeneous Datasets
Liping Li, Wei Xu, Tianyi Chen +2
cs.LGcs.CRcs.MAarXiv:1811.03761v22018InterpretML: A Unified Framework for Machine Learning Interpretability
Harsha Nori, Samuel Jenkins, Paul Koch +1
cs.LGstat.MLarXiv:1909.09223v12019GameTalk: Training LLMs for Strategic Conversation
Victor Conchello Vendrell, Max Ruiz Luyten, Mihaela van der Schaar
cs.CLcs.AIcs.GTarXiv:2601.16276v12026Context-Aware Attentive Knowledge Tracing
Aritra Ghosh, Neil Heffernan, Andrew S. Lan
cs.LGcs.AIarXiv:2007.12324v12020Three Approaches for Personalization with Applications to Federated Learning
Yishay Mansour, Mehryar Mohri, Jae Ro +1
cs.LGstat.MLarXiv:2002.10619v22020ChemRL-GEM: Geometry Enhanced Molecular Representation Learning for Property Prediction
Xiaomin Fang, Lihang Liu, Jieqiong Lei +6
cs.LGphysics.chem-phq-bio.MNarXiv:2106.06130v42021GPCR-Filter: a deep learning framework for efficient and precise GPCR modulator discovery
Jingjie Ning, Xiangzhen Shen, Li Hou +8
cs.LGq-bio.QMarXiv:2601.19149v22026Adversarial Logit Pairing
Harini Kannan, Alexey Kurakin, Ian Goodfellow
cs.LGstat.MLarXiv:1803.06373v12018Is Conditional Generative Modeling all you need for Decision-Making?
Anurag Ajay, Yilun Du, Abhi Gupta +3
cs.LGcs.AIarXiv:2211.15657v42022Physics-guided Neural Networks (PGNN): An Application in Lake Temperature Modeling
Arka Daw, Anuj Karpatne, William Watkins +2
cs.LGcs.AIcs.CVarXiv:1710.11431v32017Latent Adversarial Regularization for Offline Preference Optimization
Enyi Jiang, Yibo Jacky Zhang, Yinglun Xu +3
cs.LGcs.AIarXiv:2601.22083v22026The FLUXCOM ensemble of global land-atmosphere energy fluxes
Martin Jung, Sujan Koirala, Ulrich Weber +7
physics.ao-phcs.LGstat.MLarXiv:1812.04951v12018Learning What to Predict: Downstream-Guided Task Design for Continued Pretraining
Shuqi Ke, Giulia Fanti
cs.LGcs.AIarXiv:2601.22108v22026Enhancing Bayesian Optimization and Active Learning Through Kernel Diversity
Heng Zhang, Haotian Xiang, Qin Lu +2
cs.LGcs.AIarXiv:2608.24721v12026Mind the Student: Behavioral and Contextual Cues for Automated Engagement Prediction in Online Learning
Alperen Kantarci, Visvanathan Ramesh, Gemma Roig
cs.CVcs.AIcs.HCarXiv:2608.24340v12026Automata from Agent Traces: Failure and Next-Step Prediction
Seonglae Cho, Franklin Cardenoso Fernandez, Umar Mohammed +4
cs.AIcs.CLcs.LGarXiv:2608.23670v12026RAPTOR: Ridge-Adaptive Logistic Probes
Ziqi Gao, Yaotian Zhu, Qingcheng Zeng +4
cs.LGcs.AIarXiv:2602.00158v22026nocaps: novel object captioning at scale
Harsh Agrawal, Karan Desai, Yufei Wang +7
cs.CVcs.AIcs.CLarXiv:1812.08658v32018Not All Samples Are Created Equal: Deep Learning with Importance Sampling
Angelos Katharopoulos, François Fleuret
cs.LGarXiv:1803.00942v32018