Machine Learning
Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
8,161 to 8,220 of 20,199
Optimal CUR Matrix Decompositions
Christos Boutsidis, David P. Woodruff
cs.DScs.LGmath.NAarXiv:1405.7910v22014ANODE: Unconditionally Accurate Memory-Efficient Gradients for Neural ODEs
Amir Gholami, Kurt Keutzer, George Biros
cs.LGarXiv:1902.10298v32019RAGEN: Understanding Self-Evolution in LLM Agents via Multi-Turn Reinforcement Learning
Zihan Wang, Kangrui Wang, Qineng Wang +15
cs.LGcs.AIcs.CLarXiv:2504.20073v22025TradingAgents: Multi-Agents LLM Financial Trading Framework
Yijia Xiao, Edward Sun, Di Luo +1
q-fin.TRcs.AIcs.CEarXiv:2412.20138v72024Deep learning-based synthetic-CT generation in radiotherapy and PET: a review
Maria Francesca Spadea, Matteo Maspero, Paolo Zaffino +1
physics.med-phcs.LGeess.IVarXiv:2102.02734v22021MemoryWalker: Stop Training Agents on Contexts They Never Saw
Zinco J, Xunjie Zhu, Shen Huang +3
cs.LGcs.CLarXiv:2609.00865v12026Time-Dependent Deep Image Prior for Dynamic MRI
Jaejun Yoo, Kyong Hwan Jin, Harshit Gupta +3
eess.IVcs.CVcs.LGarXiv:1910.01684v22019Summaries:한국어CATeye: Coupled Attribute-Topology Invariance Learning for Voucher Abuse Detection
Tian Tian, Shuaicheng Niu, Hao Kuang +3
cs.LGarXiv:2609.01425v12026Spatial-MLLM: Boosting MLLM Capabilities in Visual-based Spatial Intelligence
Diankun Wu, Fangfu Liu, Yi-Hsin Hung +1
cs.CVcs.AIcs.LGarXiv:2505.23747v22025Summaries:한국어Scalable Adaptive Computation for Iterative Generation
Allan Jabri, David Fleet, Ting Chen
cs.LGcs.CVcs.NEarXiv:2212.11972v22022A Generalizable and Accessible Approach to Machine Learning with Global Satellite Imagery
Esther Rolf, Jonathan Proctor, Tamma Carleton +5
cs.LGcs.CVarXiv:2010.08168v12020A Survey on Self-Improving Test-Time Intelligence: Feedback-Driven Adapting, Learning, and Scaling at Inference
Shuaicheng Niu, Guohao Chen, Yaofo Chen +14
cs.LGarXiv:2609.01679v12026FinRL: A Deep Reinforcement Learning Library for Automated Stock Trading in Quantitative Finance
Xiao-Yang Liu, Hongyang Yang, Qian Chen +4
q-fin.TRcs.LGarXiv:2011.09607v22020Learning to Rearrange Deformable Cables, Fabrics, and Bags with Goal-Conditioned Transporter Networks
Daniel Seita, Pete Florence, Jonathan Tompson +4
cs.ROcs.LGarXiv:2012.03385v42020FedGH: Heterogeneous Federated Learning with Generalized Global Header
Liping Yi, Gang Wang, Xiaoguang Liu +2
cs.LGcs.DCarXiv:2303.13137v22023KernelBench: Can LLMs Write Efficient GPU Kernels?
Anne Ouyang, Simon Guo, Simran Arora +4
cs.LGcs.AIcs.PFarXiv:2502.10517v12025Tuning Hyperparameters without Grad Students: Scalable and Robust Bayesian Optimisation with Dragonfly
Kirthevasan Kandasamy, Karun Raju Vysyaraju, Willie Neiswanger +5
stat.MLcs.AIcs.LGarXiv:1903.06694v22019TTRL: Test-Time Reinforcement Learning
Yuxin Zuo, Kaiyan Zhang, Li Sheng +13
cs.CLcs.LGarXiv:2504.16084v32025MisGAN: Learning from Incomplete Data with Generative Adversarial Networks
Steven Cheng-Xian Li, Bo Jiang, Benjamin Marlin
cs.LGstat.MLarXiv:1902.09599v12019When Metropolis and Hastings Meet Bradley and Terry: Exact MCMC From Preference Voting
Ariel Smogorghevski, Nir Rosenfeld, Yaniv Romano
cs.LGstat.COstat.MLarXiv:2609.00905v12026A Survey of Data Quality Measurement and Monitoring Tools
Lisa Ehrlinger, Elisa Rusz, Wolfram Wöß
cs.DBcs.LGarXiv:1907.08138v12019Chronos-2: From Univariate to Universal Forecasting
Abdul Fatir Ansari, Oleksandr Shchur, Jaris Küken +20
cs.LGcs.AIstat.MLarXiv:2510.15821v12025What Can ResNet Learn Efficiently, Going Beyond Kernels?
Zeyuan Allen-Zhu, Yuanzhi Li
cs.LGcs.DScs.NEarXiv:1905.10337v32019LLaDA 1.5: Variance-Reduced Preference Optimization for Large Language Diffusion Models
Fengqi Zhu, Rongzhen Wang, Shen Nie +8
cs.LGarXiv:2505.19223v22025SkeleMotion: A New Representation of Skeleton Joint Sequences Based on Motion Information for 3D Action Recognition
Carlos Caetano, Jessica Sena, François Brémond +2
cs.CVcs.LGeess.IVarXiv:1907.13025v12019Layer by Layer: Uncovering Hidden Representations in Language Models
Oscar Skean, Md Rifat Arefin, Dan Zhao +4
cs.LGcs.AIcs.CLarXiv:2502.02013v22025Loss minimization and parameter estimation with heavy tails
Daniel Hsu, Sivan Sabato
cs.LGstat.MLarXiv:1307.1827v72013The Offset Tree for Learning with Partial Labels
Alina Beygelzimer, John Langford
cs.LGcs.AIarXiv:0812.4044v32008Improving Video Generation with Human Feedback
Jie Liu, Gongye Liu, Jiajun Liang +14
cs.CVcs.AIcs.GRarXiv:2501.13918v22025Deep Neural Network Compression for Aircraft Collision Avoidance Systems
Kyle D. Julian, Mykel J. Kochenderfer, Michael P. Owen
cs.LGstat.MLarXiv:1810.04240v12018FoundationStereo: Zero-Shot Stereo Matching
Bowen Wen, Matthew Trepte, Joseph Aribido +3
cs.CVcs.LGcs.ROarXiv:2501.09898v42025GDPO: Group reward-Decoupled Normalization Policy Optimization for Multi-reward RL Optimization
Shih-Yang Liu, Xin Dong, Ximing Lu +10
cs.CLcs.AIcs.LGarXiv:2601.05242v12026FlashInfer: Efficient and Customizable Attention Engine for LLM Inference Serving
Zihao Ye, Lequn Chen, Ruihang Lai +8
cs.DCcs.AIcs.LGarXiv:2501.01005v22025LLaDA2.0: Scaling Up Diffusion Language Models to 100B
Tiwei Bie, Maosong Cao, Kun Chen +28
cs.LGcs.AIcs.CLarXiv:2512.15745v22025Attribution Patching Outperforms Automated Circuit Discovery
Aaquib Syed, Can Rager, Arthur Conmy
cs.LGcs.AIcs.CLarXiv:2310.10348v22023Meta-Learning without Memorization
Mingzhang Yin, George Tucker, Mingyuan Zhou +2
cs.LGcs.AIstat.MLarXiv:1912.03820v32019Event-based Asynchronous Sparse Convolutional Networks
Nico Messikommer, Daniel Gehrig, Antonio Loquercio +1
cs.CVcs.LGeess.SParXiv:2003.09148v22020On Feature Learning in the Presence of Spurious Correlations
Pavel Izmailov, Polina Kirichenko, Nate Gruver +1
cs.LGcs.CVstat.MLarXiv:2210.11369v12022Trial and Error: Exploration-Based Trajectory Optimization for LLM Agents
Yifan Song, Da Yin, Xiang Yue +3
cs.CLcs.AIcs.LGarXiv:2403.02502v22024AlphaEarth Foundations: An embedding field model for accurate and efficient global mapping from sparse label data
Christopher F. Brown, Michal R. Kazmierski, Valerie J. Pasquarella +16
cs.CVcs.LGarXiv:2507.22291v22025CAT-Flow: Curvature-Adaptive sTeps for Flow Matching
Qinchan Li, Pedro Cisneros-Velarde, Keru Fu +3
cs.LGarXiv:2609.01746v12026Tri-Band Channel Measurement-Enabled Multi-Layer Digital Twin for Terahertz Wireless Data Centers
Mingjie Zhu, Ziming Yu, Guangjian Wang +1
cs.LGcs.ITarXiv:2609.01699v12026Kimi-Audio Technical Report
KimiTeam, Ding Ding, Zeqian Ju +37
eess.AScs.AIcs.CLarXiv:2504.18425v12025The Structure of Quantization Damage in LLMs: Why the Next Bit Should Be Spent Globally
Jundong Hu, Shekar Ramachandran
cs.LGcs.CLarXiv:2609.01587v12026Improved Gradient Descent Lower Bounds Beyond Nesterov
Yuhan Ye, Kaizhao Liu
math.OCcs.LGstat.MLarXiv:2609.02855v22026Demystifying Long Chain-of-Thought Reasoning in LLMs
Edward Yeo, Yuxuan Tong, Morry Niu +2
cs.CLcs.LGarXiv:2502.03373v12025DiffusionNFT: Online Diffusion Reinforcement with Forward Process
Kaiwen Zheng, Huayu Chen, Haotian Ye +7
cs.LGcs.AIcs.CVarXiv:2509.16117v22025AReaL: A Large-Scale Asynchronous Reinforcement Learning System for Language Reasoning
Wei Fu, Jiaxuan Gao, Xujie Shen +10
cs.LGcs.AIarXiv:2505.24298v52025TabICL: A Tabular Foundation Model for In-Context Learning on Large Data
Jingang Qu, David Holzmüller, Gaël Varoquaux +1
cs.LGcs.AIarXiv:2502.05564v22025Learning the Pareto Front with Hypernetworks
Aviv Navon, Aviv Shamsian, Gal Chechik +1
cs.LGarXiv:2010.04104v22020UMA: A Family of Universal Models for Atoms
Brandon M. Wood, Misko Dzamba, Xiang Fu +15
cs.LGarXiv:2506.23971v22025Automated Directed Fairness Testing
Sakshi Udeshi, Pryanshu Arora, Sudipta Chattopadhyay
cs.LGcs.AIcs.SEarXiv:1807.00468v22018DENSE: Data-Free One-Shot Federated Learning
Jie Zhang, Chen Chen, Bo Li +5
cs.LGcs.CVarXiv:2112.12371v22021Strongly Adaptive Online Learning
Amit Daniely, Alon Gonen, Shai Shalev-Shwartz
cs.LGarXiv:1502.07073v32015Doc2EDAG: An End-to-End Document-level Framework for Chinese Financial Event Extraction
Shun Zheng, Wei Cao, Wei Xu +1
cs.CLcs.LGarXiv:1904.07535v22019Graph Neural Networks in IoT: A Survey
Guimin Dong, Mingyue Tang, Zhiyuan Wang +7
cs.LGarXiv:2203.15935v22022Cascaded V-Net using ROI masks for brain tumor segmentation
Adrià Casamitjana, Marcel Catà, Irina Sánchez +2
cs.CVcs.AIcs.CYarXiv:1812.11588v12018A Survey of Safety and Trustworthiness of Large Language Models through the Lens of Verification and Validation
Xiaowei Huang, Wenjie Ruan, Wei Huang +14
cs.AIcs.LGarXiv:2305.11391v22023When Vision Meets Graphs: A Survey on Graph Reasoning and Learning
Xinjian Zhao, Wei Pang, Zhixuan Yu +8
cs.SIcs.CVcs.LGarXiv:2609.03816v12026Most Ligand-Based Classification Benchmarks Reward Memorization Rather than Generalization
Izhar Wallach, Abraham Heifets
q-bio.QMcs.LGstat.MLarXiv:1706.06619v22017