Machine Learning
Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
5,941 to 6,000 of 20,198
Gossip Learning with Linear Models on Fully Distributed Data
Róbert Ormándi, István Hegedüs, Márk Jelasity
cs.LGcs.DCarXiv:1109.1396v32011AceReason-Nemotron 1.1: Advancing Math and Code Reasoning through SFT and RL Synergy
Zihan Liu, Zhuolin Yang, Yang Chen +4
cs.CLcs.AIcs.LGarXiv:2506.13284v12025When Models Manipulate Manifolds: The Geometry of a Counting Task
Wes Gurnee, Emmanuel Ameisen, Isaac Kauvar +4
cs.LGarXiv:2601.04480v12026Kernel-based Reconstruction of Graph Signals
Daniel Romero, Meng Ma, Georgios B. Giannakis
stat.MLcs.LGarXiv:1605.07174v12016When Decodability Is Not Enough: Logical Validity Representations, Behavioral Dissociation, and Causal Tests in Language Models
Smitha Muthya Sudheendra, Jaideep Srivastava
cs.CLcs.LGarXiv:2609.02438v12026Compressed Sensing with Deep Image Prior and Learned Regularization
Dave Van Veen, Ajil Jalal, Mahdi Soltanolkotabi +3
stat.MLcs.ITcs.LGarXiv:1806.06438v42018Residual LSTM: Design of a Deep Recurrent Architecture for Distant Speech Recognition
Jaeyoung Kim, Mostafa El-Khamy, Jungwon Lee
cs.LGcs.AIcs.SDarXiv:1701.03360v32017Do Multilingual LLMs Think In English?
Lisa Schut, Yarin Gal, Sebastian Farquhar
cs.CLcs.AIcs.LGarXiv:2502.15603v12025A Context-Aware Citation Recommendation Model with BERT and Graph Convolutional Networks
Chanwoo Jeong, Sion Jang, Hyuna Shin +2
cs.CLcs.IRcs.LGarXiv:1903.06464v12019SMart: A Multi-source Multi-phase Time Series Representation Transfer Framework
Fang He, Wang-chien Lee
cs.LGcs.AIarXiv:2609.02203v12026Schrödinger Bridges on Lie Group Manifolds for Probabilistic Intrinsic Generation
Shizhe Zhang, Mingyang Zhao, Lei Ma
stat.MLcs.AIcs.LGarXiv:2609.02196v12026Anonymous Walk Embeddings
Sergey Ivanov, Evgeny Burnaev
cs.LGstat.MLarXiv:1805.11921v32018A Biologically Plausible Supervised Learning Method for Spiking Neural Networks Using the Symmetric STDP Rule
Yunzhe Hao, Xuhui Huang, Meng Dong +1
cs.NEcs.AIcs.LGarXiv:1812.06574v32018Few-Shot Learning with Embedded Class Models and Shot-Free Meta Training
Avinash Ravichandran, Rahul Bhotika, Stefano Soatto
cs.LGcs.CVstat.MLarXiv:1905.04398v22019RSL-RL: A Learning Library for Robotics Research
Clemens Schwarke, Mayank Mittal, Nikita Rudin +2
cs.ROcs.LGarXiv:2509.10771v12025Visual Planning: Let's Think Only with Images
Yi Xu, Chengzu Li, Han Zhou +4
cs.LGcs.AIcs.CLarXiv:2505.11409v32025Implicit Latent Variable Model for Scene-Consistent Motion Forecasting
Sergio Casas, Cole Gulino, Simon Suo +3
cs.CVcs.LGcs.ROarXiv:2007.12036v12020Variational Bayesian Unlearning
Quoc Phong Nguyen, Bryan Kian Hsiang Low, Patrick Jaillet
cs.LGstat.MLarXiv:2010.12883v12020Tell me about yourself: LLMs are aware of their learned behaviors
Jan Betley, Xuchan Bao, Martín Soto +3
cs.CLcs.AIcs.CRarXiv:2501.11120v12025Subspace Learning and Imputation for Streaming Big Data Matrices and Tensors
Morteza Mardani, Gonzalo Mateos, Georgios B. Giannakis
stat.MLcs.ITcs.LGarXiv:1404.4667v12014Deep Learning for Time Series Forecasting: A Survey
Xiangjie Kong, Zhenghao Chen, Weiyao Liu +6
cs.LGcs.AIarXiv:2503.10198v12025Certified Robustness to Label-Flipping Attacks via Randomized Smoothing
Elan Rosenfeld, Ezra Winston, Pradeep Ravikumar +1
cs.LGcs.AIcs.CRarXiv:2002.03018v42020AMA-Bench: Evaluating Long-Horizon Memory for Agentic Applications
Yujie Zhao, Boqin Yuan, Junbo Huang +9
cs.AIcs.LGarXiv:2602.22769v42026Influence-Balanced Loss for Imbalanced Visual Classification
Seulki Park, Jongin Lim, Younghan Jeon +1
cs.CVcs.LGarXiv:2110.02444v12021Words or Vision: Do Vision-Language Models Have Blind Faith in Text?
Ailin Deng, Tri Cao, Zhirui Chen +1
cs.CVcs.AIcs.CLarXiv:2503.02199v12025BooookScore: A systematic exploration of book-length summarization in the era of LLMs
Yapei Chang, Kyle Lo, Tanya Goyal +1
cs.CLcs.AIcs.LGarXiv:2310.00785v42023SE(3)-Stochastic Flow Matching for Protein Backbone Generation
Avishek Joey Bose, Tara Akhound-Sadegh, Guillaume Huguet +7
cs.LGcs.AIarXiv:2310.02391v42023Multiresolution Recurrent Neural Networks: An Application to Dialogue Response Generation
Iulian Vlad Serban, Tim Klinger, Gerald Tesauro +4
cs.CLcs.AIcs.LGarXiv:1606.00776v22016Sparse Autoencoders Trained on the Same Data Learn Different Features
Gonçalo Paulo, Nora Belrose
cs.LGarXiv:2501.16615v22025To Talk or to Work: Flexible Communication Compression for Energy Efficient Federated Learning over Heterogeneous Mobile Edge Devices
Liang Li, Dian Shi, Ronghui Hou +3
cs.LGcs.AIarXiv:2012.11804v12020A Convergent Gradient Descent Algorithm for Rank Minimization and Semidefinite Programming from Random Linear Measurements
Qinqing Zheng, John Lafferty
stat.MLcs.LGarXiv:1506.06081v32015Imitation Learning from Imperfect Demonstration
Yueh-Hua Wu, Nontawat Charoenphakdee, Han Bao +2
cs.LGcs.AIstat.MLarXiv:1901.09387v32019Multi-Contrast Super-Resolution MRI Through a Progressive Network
Qing Lyu, Hongming Shan, Ge Wang
eess.IVcs.LGphysics.med-pharXiv:1908.01612v22019STP: Self-play LLM Theorem Provers with Iterative Conjecturing and Proving
Kefan Dong, Tengyu Ma
cs.LGcs.AIcs.LOarXiv:2502.00212v42025QTEA: Ternary LLMs with Sparse Residual Salient Weight and By-Column Optimization
Yipin Guo, Arun M George, Jie Fu +3
cs.LGcs.AIarXiv:2609.00224v22026Superintelligent Agents Pose Catastrophic Risks: Can Scientist AI Offer a Safer Path?
Yoshua Bengio, Michael Cohen, Damiano Fornasiere +10
cs.AIcs.LGarXiv:2502.15657v22025Neural Pruning via Growing Regularization
Huan Wang, Can Qin, Yulun Zhang +1
cs.CVcs.AIcs.LGarXiv:2012.09243v22020ArtGS: Building Interactable Replicas of Complex Articulated Objects via Gaussian Splatting
Yu Liu, Baoxiong Jia, Ruijie Lu +3
cs.CVcs.GRcs.LGarXiv:2502.19459v22025SymmCD: Symmetry-Preserving Crystal Generation with Diffusion Models
Daniel Levy, Siba Smarak Panigrahi, Sékou-Oumar Kaba +5
cond-mat.mtrl-scics.LGarXiv:2502.03638v32025Curriculum Reinforcement Learning from Easy to Hard Tasks Improves LLM Reasoning
Shubham Parashar, Shurui Gui, Xiner Li +8
cs.LGcs.AIcs.CLarXiv:2506.06632v32025LimiX: Unleashing Structured-Data Modeling Capability for Generalist Intelligence
Xingxuan Zhang, Gang Ren, Han Yu +35
cs.LGcs.AIcs.CLarXiv:2509.03505v22025Poisoning Attacks on LLMs Require a Near-constant Number of Poison Samples
Alexandra Souly, Javier Rando, Ed Chapman +10
cs.LGarXiv:2510.07192v12025WebSailor-V2: Bridging the Chasm to Proprietary Agents via Synthetic Data and Scalable Reinforcement Learning
Kuan Li, Zhongwang Zhang, Huifeng Yin +14
cs.LGcs.CLarXiv:2509.13305v12025On the Anatomy of MCMC-Based Maximum Likelihood Learning of Energy-Based Models
Erik Nijkamp, Mitch Hill, Tian Han +2
stat.MLcs.CVcs.LGarXiv:1903.12370v42019Graph Normalizing Flows
Jenny Liu, Aviral Kumar, Jimmy Ba +2
cs.LGstat.MLarXiv:1905.13177v12019Training High-Performance Low-Latency Spiking Neural Networks by Differentiation on Spike Representation
Qingyan Meng, Mingqing Xiao, Shen Yan +3
cs.NEcs.LGarXiv:2205.00459v22022Differentiable Learning-to-Normalize via Switchable Normalization
Ping Luo, Jiamin Ren, Zhanglin Peng +2
cs.CVcs.LGarXiv:1806.10779v52018WiSDoM: Wireless Sparse Decision Transformer with Mixture-of-Experts for Multi-Task Mobile Network Optimization
Fatih Temiz, Shavbo Salehi, Melike Erol-Kantarci
cs.NIcs.AIcs.LGarXiv:2609.00284v12026DeepAgent: A General Reasoning Agent with Scalable Toolsets
Xiaoxi Li, Wenxiang Jiao, Jiarui Jin +8
cs.AIcs.CLcs.IRarXiv:2510.21618v32025Faster Than Flash: Exploiting Attention Sparsity for Efficient Long-Context Decoding
Zhigeng Liu, Zhiyuan Ning, Ruixiao Li +5
cs.LGcs.AIarXiv:2609.00097v12026Data-Dependent Stability of Stochastic Gradient Descent
Ilja Kuzborskij, Christoph H. Lampert
cs.LGarXiv:1703.01678v42017Detecting Harmful Memes and Their Targets
Shraman Pramanick, Dimitar Dimitrov, Rituparna Mukherjee +4
cs.CLcs.LGcs.MMarXiv:2110.00413v12021Which Algorithmic Choices Matter at Which Batch Sizes? Insights From a Noisy Quadratic Model
Guodong Zhang, Lala Li, Zachary Nado +5
cs.LGstat.MLarXiv:1907.04164v22019Deep Learning Advancements in Anomaly Detection: A Comprehensive Survey
Haoqi Huang, Ping Wang, Jianhua Pei +3
cs.LGarXiv:2503.13195v12025Specifying Object Attributes and Relations in Interactive Scene Generation
Oron Ashual, Lior Wolf
cs.CVcs.LGarXiv:1909.05379v22019The Amazon Nova Family of Models: Technical Report and Model Card
Amazon AGI, Aaron Langford, Aayush Shah +783
cs.AIcs.CYcs.LGarXiv:2506.12103v12025The Illusion of State in State-Space Models
William Merrill, Jackson Petty, Ashish Sabharwal
cs.LGcs.CCcs.CLarXiv:2404.08819v32024Learning with Feature-Dependent Label Noise: A Progressive Approach
Yikai Zhang, Songzhu Zheng, Pengxiang Wu +2
cs.LGcs.CVstat.AParXiv:2103.07756v32021Temporal Convolution for Real-time Keyword Spotting on Mobile Devices
Seungwoo Choi, Seokjun Seo, Beomjun Shin +5
cs.SDcs.LGcs.NEarXiv:1904.03814v22019SWEET-RL: Training Multi-Turn LLM Agents on Collaborative Reasoning Tasks
Yifei Zhou, Song Jiang, Yuandong Tian +4
cs.LGarXiv:2503.15478v12025