Machine Learning
Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
18,181 to 18,240 of 20,454
Decision Tree and K-Means Analysis of Raman Spectra for Edible Oils: A Physics-Informed AI Approach
Amrita Shaw, Chandrasekar S. N., Sai Muthukumar V. +2
cs.LGcs.AIarXiv:2608.20440v12026Training Deep Neural Networks on Noisy Labels with Bootstrapping
Scott Reed, Honglak Lee, Dragomir Anguelov +3
cs.CVcs.LGcs.NEarXiv:1412.6596v32014Approximate Homomorphisms and Convergent Representations in Transducers
Santiago Cifuentes
cs.LGcs.AIarXiv:2608.20428v12026Simplified State Space Layers for Sequence Modeling
Jimmy T. H. Smith, Andrew Warrington, Scott W. Linderman
cs.LGarXiv:2208.04933v32022VA-DPO: Valence-Arousal Direct Preference Optimization for Controllable Emotion Generation in Language Models
Hyunwoo Kim
cs.CLcs.AIcs.LGarXiv:2608.20374v12026EviRank: Structured Relevance Evidence for Multimodal Image Re-ranking
Enjun Du, Siyi Liu, Zirong Chen +8
cs.CVcs.LGarXiv:2608.20886v12026Let's Scale Step by Step: Compute-Efficient Hyperparameter Transfer for Large-Scale Mixture-of-Experts
Nayeon Kim, Hojin Lee, Yunju Bak +2
cs.LGcs.AIcs.CLarXiv:2608.20061v12026Are GANs Created Equal? A Large-Scale Study
Mario Lucic, Karol Kurach, Marcin Michalski +2
stat.MLcs.LGarXiv:1711.10337v42017ReCurveflow: A Flow Matching Framework that Learns Curved Reaction Trajectories to Predict Transition State Geometries
Seungheun Baek, Mogan Gim, Jaewoo Kang
cs.AIcs.LGarXiv:2608.20869v12026Differentiable Volumetric Rendering: Learning Implicit 3D Representations without 3D Supervision
Michael Niemeyer, Lars Mescheder, Michael Oechsle +1
cs.CVcs.LGeess.IVarXiv:1912.07372v22019Summaries:한국어NeuroStrata: An Electroencephalographic Connectivity-Aware Deep Representation Learning Framework for Dynamic Brain Network Analysis of Mental Stress
Sayantan Acharya, Hamzeh Asgharnezhad, Abbas Khosravi +3
q-bio.NCcs.AIcs.LGarXiv:2608.20354v12026AutoInt: Automatic Feature Interaction Learning via Self-Attentive Neural Networks
Weiping Song, Chence Shi, Zhiping Xiao +4
cs.IRcs.AIcs.LGarXiv:1810.11921v22018Resnet in Resnet: Generalizing Residual Architectures
Sasha Targ, Diogo Almeida, Kevin Lyman
cs.LGcs.CVcs.NEarXiv:1603.08029v12016Unsupervised Learning for Physical Interaction through Video Prediction
Chelsea Finn, Ian Goodfellow, Sergey Levine
cs.LGcs.AIcs.CVarXiv:1605.07157v42016Personalized Privacy Control in LLMs via Attention Head Intervention
Junseok Kim, Nakyeong Yang, Kyomin Jung
cs.AIcs.CLcs.LGarXiv:2608.21209v12026ROCKET: Exceptionally fast and accurate time series classification using random convolutional kernels
Angus Dempster, François Petitjean, Geoffrey I. Webb
cs.LGstat.MLarXiv:1910.13051v12019TreeWY: Speculative Verification for Gated DeltaNet Hybrids
Sneha Murthy Ghantasala
cs.AIcs.CLcs.DCarXiv:2608.20961v12026Neuro-Geospatial Modelling of EEG Affective States Using Literature-Informed Environmental Context
Utsav Poudel, Jagannath Aryal, Subramaniyaswamy Vairavasundaram
cs.AIcs.HCcs.LGarXiv:2608.20807v12026NVAE: A Deep Hierarchical Variational Autoencoder
Arash Vahdat, Jan Kautz
stat.MLcs.CVcs.LGarXiv:2007.03898v32020PointPainting: Sequential Fusion for 3D Object Detection
Sourabh Vora, Alex H. Lang, Bassam Helou +1
cs.CVcs.LGeess.IVarXiv:1911.10150v22019Dual-Cache Latent Space Communication between Heterogeneous Language Models
Jiyao Liu, Qi Zhang, Yaoyi Jia +2
cs.AIcs.LGarXiv:2608.20617v12026Deep Graph Contrastive Representation Learning
Yanqiao Zhu, Yichen Xu, Feng Yu +3
cs.LGstat.MLarXiv:2006.04131v22020Multimodal Learning with Transformers: A Survey
Peng Xu, Xiatian Zhu, David A. Clifton
cs.CVcs.LGarXiv:2206.06488v22022Universal Adversarial Triggers for Attacking and Analyzing NLP
Eric Wallace, Shi Feng, Nikhil Kandpal +2
cs.CLcs.LGarXiv:1908.07125v32019Reasoning with Language Model is Planning with World Model
Shibo Hao, Yi Gu, Haodi Ma +4
cs.CLcs.AIcs.LGarXiv:2305.14992v22023Time-LLM: Time Series Forecasting by Reprogramming Large Language Models
Ming Jin, Shiyu Wang, Lintao Ma +8
cs.LGcs.AIarXiv:2310.01728v22023Long Short-Term Memory Based Recurrent Neural Network Architectures for Large Vocabulary Speech Recognition
Haşim Sak, Andrew Senior, Françoise Beaufays
cs.NEcs.CLcs.LGarXiv:1402.1128v12014The UCR Time Series Archive
Hoang Anh Dau, Anthony Bagnall, Kaveh Kamgar +5
cs.LGstat.MLarXiv:1810.07758v22018SimPO: Simple Preference Optimization with a Reference-Free Reward
Yu Meng, Mengzhou Xia, Danqi Chen
cs.CLcs.LGarXiv:2405.14734v32024Siren's Song in the AI Ocean: A Survey on Hallucination in Large Language Models
Yue Zhang, Yafu Li, Leyang Cui +13
cs.CLcs.AIcs.CYarXiv:2309.01219v32023It's Not Just Size That Matters: Small Language Models Are Also Few-Shot Learners
Timo Schick, Hinrich Schütze
cs.CLcs.AIcs.LGarXiv:2009.07118v22020Low-rank Matrix Completion using Alternating Minimization
Prateek Jain, Praneeth Netrapalli, Sujay Sanghavi
stat.MLcs.LGmath.OCarXiv:1212.0467v12012Argoverse 2: Next Generation Datasets for Self-Driving Perception and Forecasting
Benjamin Wilson, William Qi, Tanmay Agarwal +10
cs.CVcs.AIcs.LGarXiv:2301.00493v12023Contrastive Learning of Medical Visual Representations from Paired Images and Text
Yuhao Zhang, Hang Jiang, Yasuhide Miura +2
cs.CVcs.CLcs.LGarXiv:2010.00747v22020Joint Deep Modeling of Users and Items Using Reviews for Recommendation
Lei Zheng, Vahid Noroozi, Philip S. Yu
cs.LGcs.IRarXiv:1701.04783v12017QANet: Combining Local Convolution with Global Self-Attention for Reading Comprehension
Adams Wei Yu, David Dohan, Minh-Thang Luong +4
cs.CLcs.AIcs.LGarXiv:1804.09541v12018BERTweet: A pre-trained language model for English Tweets
Dat Quoc Nguyen, Thanh Vu, Anh Tuan Nguyen
cs.CLcs.LGarXiv:2005.10200v22020StateSMix: Online Lossless Compression via Mamba State Space Models and Sparse N-gram Context Mixing
Roberto Tacconelli
cs.LGcs.ITarXiv:2605.02904v12026A Generalist Agent
Scott Reed, Konrad Zolna, Emilio Parisotto +17
cs.AIcs.CLcs.LGarXiv:2205.06175v32022Be Your Own Teacher: Improve the Performance of Convolutional Neural Networks via Self Distillation
Linfeng Zhang, Jiebo Song, Anni Gao +3
cs.LGstat.MLarXiv:1905.08094v12019Multi-scale Attributed Node Embedding
Benedek Rozemberczki, Carl Allen, Rik Sarkar
cs.LGcs.NIcs.SIarXiv:1909.13021v32019When Retrieval Fails Before It Begins: Structurally Indirect Prerequisite Eviction as a Retention Failure in Agentic Memory
Minkyu Song
cs.AIcs.CLcs.LGarXiv:2608.20400v12026World models of environment, agent and joint agent-environment systems
Manuel Baltieri, Filippo Torresan, Yivan Zhang +2
cs.AIcs.LGarXiv:2608.20401v12026A Hybrid Approach to Privacy-Preserving Federated Learning
Stacey Truex, Nathalie Baracaldo, Ali Anwar +4
cs.LGstat.MLarXiv:1812.03224v22018Inf-Net: Automatic COVID-19 Lung Infection Segmentation from CT Images
Deng-Ping Fan, Tao Zhou, Ge-Peng Ji +5
eess.IVcs.CVcs.LGarXiv:2004.14133v42020A Review on Generative Adversarial Networks: Algorithms, Theory, and Applications
Jie Gui, Zhenan Sun, Yonggang Wen +2
cs.LGstat.MLarXiv:2001.06937v12020Bayesian Active Learning for Classification and Preference Learning
Neil Houlsby, Ferenc Huszár, Zoubin Ghahramani +1
stat.MLcs.LGarXiv:1112.5745v12011A Hierarchical Latent Variable Encoder-Decoder Model for Generating Dialogues
Iulian Vlad Serban, Alessandro Sordoni, Ryan Lowe +4
cs.CLcs.AIcs.LGarXiv:1605.06069v32016Deep Forest
Zhi-Hua Zhou, Ji Feng
cs.LGstat.MLarXiv:1702.08835v42017Agnostic Federated Learning
Mehryar Mohri, Gary Sivek, Ananda Theertha Suresh
cs.LGstat.MLarXiv:1902.00146v12019MIMIC-CXR-JPG, a large publicly available database of labeled chest radiographs
Alistair E. W. Johnson, Tom J. Pollard, Nathaniel R. Greenbaum +7
cs.CVcs.LGeess.IVarXiv:1901.07042v52019Florence: A New Foundation Model for Computer Vision
Lu Yuan, Dongdong Chen, Yi-Ling Chen +20
cs.CVcs.AIcs.LGarXiv:2111.11432v12021Evading Defenses to Transferable Adversarial Examples by Translation-Invariant Attacks
Yinpeng Dong, Tianyu Pang, Hang Su +1
cs.CVcs.CRcs.LGarXiv:1904.02884v12019Tune: A Research Platform for Distributed Model Selection and Training
Richard Liaw, Eric Liang, Robert Nishihara +3
cs.LGcs.DCstat.MLarXiv:1807.05118v12018Self-Supervised Learning from Images with a Joint-Embedding Predictive Architecture
Mahmoud Assran, Quentin Duval, Ishan Misra +5
cs.CVcs.AIcs.LGarXiv:2301.08243v32023An Explanation of In-context Learning as Implicit Bayesian Inference
Sang Michael Xie, Aditi Raghunathan, Percy Liang +1
cs.CLcs.LGarXiv:2111.02080v62021TUDataset: A collection of benchmark datasets for learning with graphs
Christopher Morris, Nils M. Kriege, Franka Bause +3
cs.LGcs.NEstat.MLarXiv:2007.08663v12020Recipe for a General, Powerful, Scalable Graph Transformer
Ladislav Rampášek, Mikhail Galkin, Vijay Prakash Dwivedi +3
cs.LGarXiv:2205.12454v42022Visualizing and Understanding Recurrent Networks
Andrej Karpathy, Justin Johnson, Li Fei-Fei
cs.LGcs.CLcs.NEarXiv:1506.02078v22015VectorNet: Encoding HD Maps and Agent Dynamics from Vectorized Representation
Jiyang Gao, Chen Sun, Hang Zhao +4
cs.CVcs.LGstat.MLarXiv:2005.04259v12020