Machine Learning
Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
16,981 to 17,040 of 20,199
A Physical Response-and-Memory Model for Muon Optimization
Yinze Hu, Hongjun Xiang, Xingao Gong +1
cs.LGcond-mat.dis-nncond-mat.stat-mecharXiv:2608.22994v12026MOReL : Model-Based Offline Reinforcement Learning
Rahul Kidambi, Aravind Rajeswaran, Praneeth Netrapalli +1
cs.LGcs.AIstat.MLarXiv:2005.05951v32020Revisiting Spatial-Temporal Similarity: A Deep Learning Framework for Traffic Prediction
Huaxiu Yao, Xianfeng Tang, Hua Wei +2
cs.LGarXiv:1803.01254v22018Knowing Isn't Understanding: Re-grounding Generative Proactivity with Epistemic and Behavioral Insight
Kirandeep Kaur, Xingda Lyu, Chirag Shah
cs.CYcs.AIcs.LGarXiv:2602.15259v22026Multi-Scale Progressive Fusion Network for Single Image Deraining
Kui Jiang, Zhongyuan Wang, Peng Yi +5
cs.CVcs.LGeess.IVarXiv:2003.10985v22020Few-Shot Learning via Embedding Adaptation with Set-to-Set Functions
Han-Jia Ye, Hexiang Hu, De-Chuan Zhan +1
cs.LGcs.CVarXiv:1812.03664v62018Neural Operator based Multi-Field Reconstruction of Inner Solar Boundary State
Vignesh Kumar Pandian Sathia, Reza Mansouri, Dustin J. Kempton +2
cs.LGastro-ph.IMastro-ph.SRarXiv:2608.22782v12026CCNet: Extracting High Quality Monolingual Datasets from Web Crawl Data
Guillaume Wenzek, Marie-Anne Lachaux, Alexis Conneau +4
cs.CLcs.IRcs.LGarXiv:1911.00359v22019MiniCPM-SALA: Hybridizing Sparse and Linear Attention for Efficient Long-Context Modeling
MiniCPM Team, Wenhao An, Yingfa Chen +44
cs.CLcs.AIcs.LGarXiv:2602.11761v22026Using Pre-Training Can Improve Model Robustness and Uncertainty
Dan Hendrycks, Kimin Lee, Mantas Mazeika
cs.LGcs.CVstat.MLarXiv:1901.09960v52019InnoEval: On Research Idea Evaluation as a Knowledge-Grounded, Multi-Perspective Reasoning Problem
Shuofei Qiao, Yunxiang Wei, Xuehai Wang +10
cs.CLcs.AIcs.IRarXiv:2602.14367v22026Physics-informed neural networks with hard constraints for inverse design
Lu Lu, Raphael Pestourie, Wenjie Yao +3
physics.comp-phcs.LGarXiv:2102.04626v12021Attack of the Tails: Yes, You Really Can Backdoor Federated Learning
Hongyi Wang, Kartik Sreenivasan, Shashank Rajput +5
cs.LGcs.CRcs.DCarXiv:2007.05084v12020DF-MoE: Generalizable Deepfake Detection via Multimodal Sparse Mixture-of-Experts
Vlad Hondru, Florinel Alin Croitoru, Iuliana Georgescu +2
cs.CVcs.AIcs.LGarXiv:2608.23363v12026LoST: Level of Semantics Tokenization for 3D Shapes
Niladri Shekhar Dutt, Zifan Shi, Paul Guerrero +4
cs.CVcs.GRcs.LGarXiv:2603.17995v12026HopSkipJumpAttack: A Query-Efficient Decision-Based Attack
Jianbo Chen, Michael I. Jordan, Martin J. Wainwright
cs.LGcs.CRmath.OCarXiv:1904.02144v52019Reservoir of Importance: Learning Semi-Structured Sparsity with Differentiable Subset Sampling
Ha Dinh, Xuan Duy Ta, Khoat Than +1
cs.LGarXiv:2608.23048v12026Spanning the Visual Analogy Space with a Weight Basis of LoRAs
Hila Manor, Rinon Gal, Haggai Maron +2
cs.CVcs.AIcs.GRarXiv:2602.15727v22026CMI-RewardBench: Evaluating Music Reward Models with Compositional Multimodal Instruction
Yinghao Ma, Haiwen Xia, Hewei Gao +9
cs.SDcs.AIcs.LGarXiv:2603.00610v32026Segment Anything Model for Medical Image Analysis: an Experimental Study
Maciej A. Mazurowski, Haoyu Dong, Hanxue Gu +3
cs.CVcs.AIcs.LGarXiv:2304.10517v32023Contrastive Clustering
Yunfan Li, Peng Hu, Zitao Liu +3
cs.LGcs.CVstat.MLarXiv:2009.09687v12020Spend Search Where It Pays: Value-Guided Structured Sampling and Optimization for Generative Recommendation
Jie Jiang, Yangru Huang, Zeyu Wang +4
cs.AIcs.LGarXiv:2602.10699v22026KILT: a Benchmark for Knowledge Intensive Language Tasks
Fabio Petroni, Aleksandra Piktus, Angela Fan +10
cs.CLcs.AIcs.IRarXiv:2009.02252v42020High-Fidelity Audio Compression with Improved RVQGAN
Rithesh Kumar, Prem Seetharaman, Alejandro Luebs +2
cs.SDcs.LGeess.ASarXiv:2306.06546v22023Transfer learning enhanced physics informed neural network for phase-field modeling of fracture
Somdatta Goswami, Cosmin Anitescu, Souvik Chakraborty +1
stat.MLcs.LGarXiv:1907.02531v12019An introduction to Topological Data Analysis: fundamental and practical aspects for data scientists
Frédéric Chazal, Bertrand Michel
math.STcs.LGmath.ATarXiv:1710.04019v22017Online Learning: A Comprehensive Survey
Steven C. H. Hoi, Doyen Sahoo, Jing Lu +1
cs.LGarXiv:1802.02871v22018Counterfactual Transition Graphs: Evaluating Cross-Class Transition Quality
Syed Muhammad Hamza Zaidi, Szymon Bobek, Grzegorz J. Nalepa +1
cs.LGcs.AIarXiv:2608.23164v12026PixelDefend: Leveraging Generative Models to Understand and Defend against Adversarial Examples
Yang Song, Taesup Kim, Sebastian Nowozin +2
cs.LGarXiv:1710.10766v32017Mirror descent algorithms with logarithmic barriers
Alberto De Marchi, Yura Malitsky, Adrien B. Taylor
math.OCcs.LGmath.NAarXiv:2608.22834v12026Transformers learn in-context by gradient descent
Johannes von Oswald, Eyvind Niklasson, Ettore Randazzo +4
cs.LGcs.AIcs.CLarXiv:2212.07677v22022Model compression via distillation and quantization
Antonio Polino, Razvan Pascanu, Dan Alistarh
cs.NEcs.LGarXiv:1802.05668v12018Likelihood Ratios for Out-of-Distribution Detection
Jie Ren, Peter J. Liu, Emily Fertig +5
stat.MLcs.LGarXiv:1906.02845v22019Next Embedding Prediction Makes World Models Stronger
George Bredis, Nikita Balagansky, Daniil Gavrilov +1
cs.LGcs.AIarXiv:2603.02765v12026Reward-Free Continual Adaptation for Resilient Space Robots
Andrej Orsula, Miguel Olivares-Mendez, Carol Martinez
cs.ROcs.AIcs.LGarXiv:2608.23452v12026Machine learning for molecular simulation
Frank Noé, Alexandre Tkatchenko, Klaus-Robert Müller +1
physics.chem-phcs.LGphysics.comp-pharXiv:1911.02792v12019ISO-Bench: Can Coding Agents Optimize Real-World Inference Workloads?
Ayush Nangia, Shikhar Mishra, Aman Gokrani +1
cs.LGarXiv:2602.19594v12026BeamPERL: Parameter-Efficient RL with Verifiable Rewards Specializes Compact LLMs for Structured Beam Mechanics Reasoning
Tarjei Paule Hage, Markus J. Buehler
cs.AIcond-mat.mtrl-scics.CLarXiv:2603.04124v12026Plug-and-Play Benchmarking of Reinforcement Learning Algorithms for Large-Scale Flow Control
Jannis Becktepe, Aleksandra Franz, Nils Thuerey +1
cs.LGarXiv:2601.15015v22026Document Ranking with a Pretrained Sequence-to-Sequence Model
Rodrigo Nogueira, Zhiying Jiang, Jimmy Lin
cs.IRcs.LGarXiv:2003.06713v12020Diffusion Model Alignment Using Direct Preference Optimization
Bram Wallace, Meihua Dang, Rafael Rafailov +7
cs.CVcs.AIcs.GRarXiv:2311.12908v12023DenseCLIP: Language-Guided Dense Prediction with Context-Aware Prompting
Yongming Rao, Wenliang Zhao, Guangyi Chen +5
cs.CVcs.AIcs.LGarXiv:2112.01518v22021What Can Transformers Learn In-Context? A Case Study of Simple Function Classes
Shivam Garg, Dimitris Tsipras, Percy Liang +1
cs.CLcs.LGarXiv:2208.01066v32022Bridging Theory and Algorithm for Domain Adaptation
Yuchen Zhang, Tianle Liu, Mingsheng Long +1
cs.LGstat.MLarXiv:1904.05801v22019Don't Repeat Yourself: Stopping Verbatim Loops at Sampling Time
Philipp Emanuel Weidmann, Allen Roush, Judah Goldfeder +2
cs.CLcs.AIcs.LGarXiv:2608.22761v12026A Commutator Framework for Selective Spectral Alignment in Deep Neural Networks
Kaj Nyström
stat.MLcs.LGarXiv:2608.22910v12026Parameterized Explainer for Graph Neural Network
Dongsheng Luo, Wei Cheng, Dongkuan Xu +4
cs.LGcs.AIarXiv:2011.04573v12020Beyond chlorophyll: machine learning estimates of diagnostic phytoplankton pigments from multispectral ocean colour data
David Moffat, Angus Laurenson, Victor Martinez-Vicente +4
q-bio.OTcs.LGphysics.opticsarXiv:2608.23348v12026Thinking at the Right Size: Amortized Distillation Across Post-Trained LLMs
Yan Zhou, Sara Kangaslahti, Jonathan Geuter +4
cs.LGarXiv:2608.22854v12026Least-Loaded Expert Parallelism: Load Balancing An Imbalanced Mixture-of-Experts
Xuan-Phi Nguyen, Shrey Pandit, Austin Xu +2
cs.LGcs.AIarXiv:2601.17111v12026Conditional Positional Encodings for Vision Transformers
Xiangxiang Chu, Zhi Tian, Bo Zhang +2
cs.CVcs.AIcs.LGarXiv:2102.10882v32021VLM-SubtleBench: How Far Are VLMs from Human-Level Subtle Comparative Reasoning?
Minkyu Kim, Sangheon Lee, Dongmin Park
cs.CVcs.AIcs.LGarXiv:2603.07888v12026Meta-Learning: A Survey
Joaquin Vanschoren
cs.LGstat.MLarXiv:1810.03548v12018XTC: Head-Aware Sampling by Excluding Top Choices
Philipp Emanuel Weidmann, Allen Roush, Judah Goldfeder +2
cs.CLcs.AIcs.LGarXiv:2608.22758v12026On Distillation of Guided Diffusion Models
Chenlin Meng, Robin Rombach, Ruiqi Gao +4
cs.CVcs.AIcs.LGarXiv:2210.03142v32022TERMINATOR: Learning Optimal Exit Points for Early Stopping in Chain-of-Thought Reasoning
Alliot Nagle, Jakhongir Saydaliev, Dhia Garbaya +3
cs.LGcs.AIcs.CLarXiv:2603.12529v22026The Web as a Knowledge-base for Answering Complex Questions
Alon Talmor, Jonathan Berant
cs.CLcs.AIcs.LGarXiv:1803.06643v12018PolyChirp: Multi-Species Birdsong Classification Using TinyML on Low-Power Acoustic Sensors
Nathan Duboisset, Zhaolan Huang, Felix Bießmann +3
cs.LGcs.AIarXiv:2608.23101v12026Algorithms for nonnegative matrix factorization with the beta-divergence
Cédric Févotte, Jérôme Idier
cs.LGarXiv:1010.1763v32010COVID-19 Image Data Collection: Prospective Predictions Are the Future
Joseph Paul Cohen, Paul Morrison, Lan Dao +3
q-bio.QMcs.CVcs.LGarXiv:2006.11988v32020