Machine Learning

Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

8,161 to 8,220 of 20,199

  1. Optimal CUR Matrix Decompositions

    Christos Boutsidis, David P. Woodruff

    cs.DScs.LGmath.NAarXiv:1405.7910v22014
  2. ANODE: Unconditionally Accurate Memory-Efficient Gradients for Neural ODEs

    Amir Gholami, Kurt Keutzer, George Biros

    cs.LGarXiv:1902.10298v32019
  3. RAGEN: Understanding Self-Evolution in LLM Agents via Multi-Turn Reinforcement Learning

    Zihan Wang, Kangrui Wang, Qineng Wang +15

    cs.LGcs.AIcs.CLarXiv:2504.20073v22025
  4. TradingAgents: Multi-Agents LLM Financial Trading Framework

    Yijia Xiao, Edward Sun, Di Luo +1

    q-fin.TRcs.AIcs.CEarXiv:2412.20138v72024
  5. Deep learning-based synthetic-CT generation in radiotherapy and PET: a review

    Maria Francesca Spadea, Matteo Maspero, Paolo Zaffino +1

    physics.med-phcs.LGeess.IVarXiv:2102.02734v22021
  6. MemoryWalker: Stop Training Agents on Contexts They Never Saw

    Zinco J, Xunjie Zhu, Shen Huang +3

    cs.LGcs.CLarXiv:2609.00865v12026
  7. Time-Dependent Deep Image Prior for Dynamic MRI

    Jaejun Yoo, Kyong Hwan Jin, Harshit Gupta +3

    eess.IVcs.CVcs.LGarXiv:1910.01684v22019
    Summaries:한국어
  8. CATeye: Coupled Attribute-Topology Invariance Learning for Voucher Abuse Detection

    Tian Tian, Shuaicheng Niu, Hao Kuang +3

    cs.LGarXiv:2609.01425v12026
  9. Spatial-MLLM: Boosting MLLM Capabilities in Visual-based Spatial Intelligence

    Diankun Wu, Fangfu Liu, Yi-Hsin Hung +1

    cs.CVcs.AIcs.LGarXiv:2505.23747v22025
    Summaries:한국어
  10. Scalable Adaptive Computation for Iterative Generation

    Allan Jabri, David Fleet, Ting Chen

    cs.LGcs.CVcs.NEarXiv:2212.11972v22022
  11. A Generalizable and Accessible Approach to Machine Learning with Global Satellite Imagery

    Esther Rolf, Jonathan Proctor, Tamma Carleton +5

    cs.LGcs.CVarXiv:2010.08168v12020
  12. A Survey on Self-Improving Test-Time Intelligence: Feedback-Driven Adapting, Learning, and Scaling at Inference

    Shuaicheng Niu, Guohao Chen, Yaofo Chen +14

    cs.LGarXiv:2609.01679v12026
  13. FinRL: A Deep Reinforcement Learning Library for Automated Stock Trading in Quantitative Finance

    Xiao-Yang Liu, Hongyang Yang, Qian Chen +4

    q-fin.TRcs.LGarXiv:2011.09607v22020
  14. Learning to Rearrange Deformable Cables, Fabrics, and Bags with Goal-Conditioned Transporter Networks

    Daniel Seita, Pete Florence, Jonathan Tompson +4

    cs.ROcs.LGarXiv:2012.03385v42020
  15. FedGH: Heterogeneous Federated Learning with Generalized Global Header

    Liping Yi, Gang Wang, Xiaoguang Liu +2

    cs.LGcs.DCarXiv:2303.13137v22023
  16. KernelBench: Can LLMs Write Efficient GPU Kernels?

    Anne Ouyang, Simon Guo, Simran Arora +4

    cs.LGcs.AIcs.PFarXiv:2502.10517v12025
  17. Tuning Hyperparameters without Grad Students: Scalable and Robust Bayesian Optimisation with Dragonfly

    Kirthevasan Kandasamy, Karun Raju Vysyaraju, Willie Neiswanger +5

    stat.MLcs.AIcs.LGarXiv:1903.06694v22019
  18. TTRL: Test-Time Reinforcement Learning

    Yuxin Zuo, Kaiyan Zhang, Li Sheng +13

    cs.CLcs.LGarXiv:2504.16084v32025
  19. MisGAN: Learning from Incomplete Data with Generative Adversarial Networks

    Steven Cheng-Xian Li, Bo Jiang, Benjamin Marlin

    cs.LGstat.MLarXiv:1902.09599v12019
  20. When Metropolis and Hastings Meet Bradley and Terry: Exact MCMC From Preference Voting

    Ariel Smogorghevski, Nir Rosenfeld, Yaniv Romano

    cs.LGstat.COstat.MLarXiv:2609.00905v12026
  21. A Survey of Data Quality Measurement and Monitoring Tools

    Lisa Ehrlinger, Elisa Rusz, Wolfram Wöß

    cs.DBcs.LGarXiv:1907.08138v12019
  22. Chronos-2: From Univariate to Universal Forecasting

    Abdul Fatir Ansari, Oleksandr Shchur, Jaris Küken +20

    cs.LGcs.AIstat.MLarXiv:2510.15821v12025
  23. What Can ResNet Learn Efficiently, Going Beyond Kernels?

    Zeyuan Allen-Zhu, Yuanzhi Li

    cs.LGcs.DScs.NEarXiv:1905.10337v32019
  24. LLaDA 1.5: Variance-Reduced Preference Optimization for Large Language Diffusion Models

    Fengqi Zhu, Rongzhen Wang, Shen Nie +8

    cs.LGarXiv:2505.19223v22025
  25. SkeleMotion: A New Representation of Skeleton Joint Sequences Based on Motion Information for 3D Action Recognition

    Carlos Caetano, Jessica Sena, François Brémond +2

    cs.CVcs.LGeess.IVarXiv:1907.13025v12019
  26. Layer by Layer: Uncovering Hidden Representations in Language Models

    Oscar Skean, Md Rifat Arefin, Dan Zhao +4

    cs.LGcs.AIcs.CLarXiv:2502.02013v22025
  27. Loss minimization and parameter estimation with heavy tails

    Daniel Hsu, Sivan Sabato

    cs.LGstat.MLarXiv:1307.1827v72013
  28. The Offset Tree for Learning with Partial Labels

    Alina Beygelzimer, John Langford

    cs.LGcs.AIarXiv:0812.4044v32008
  29. Improving Video Generation with Human Feedback

    Jie Liu, Gongye Liu, Jiajun Liang +14

    cs.CVcs.AIcs.GRarXiv:2501.13918v22025
  30. Deep Neural Network Compression for Aircraft Collision Avoidance Systems

    Kyle D. Julian, Mykel J. Kochenderfer, Michael P. Owen

    cs.LGstat.MLarXiv:1810.04240v12018
  31. FoundationStereo: Zero-Shot Stereo Matching

    Bowen Wen, Matthew Trepte, Joseph Aribido +3

    cs.CVcs.LGcs.ROarXiv:2501.09898v42025
  32. GDPO: Group reward-Decoupled Normalization Policy Optimization for Multi-reward RL Optimization

    Shih-Yang Liu, Xin Dong, Ximing Lu +10

    cs.CLcs.AIcs.LGarXiv:2601.05242v12026
  33. FlashInfer: Efficient and Customizable Attention Engine for LLM Inference Serving

    Zihao Ye, Lequn Chen, Ruihang Lai +8

    cs.DCcs.AIcs.LGarXiv:2501.01005v22025
  34. LLaDA2.0: Scaling Up Diffusion Language Models to 100B

    Tiwei Bie, Maosong Cao, Kun Chen +28

    cs.LGcs.AIcs.CLarXiv:2512.15745v22025
  35. Attribution Patching Outperforms Automated Circuit Discovery

    Aaquib Syed, Can Rager, Arthur Conmy

    cs.LGcs.AIcs.CLarXiv:2310.10348v22023
  36. Meta-Learning without Memorization

    Mingzhang Yin, George Tucker, Mingyuan Zhou +2

    cs.LGcs.AIstat.MLarXiv:1912.03820v32019
  37. Event-based Asynchronous Sparse Convolutional Networks

    Nico Messikommer, Daniel Gehrig, Antonio Loquercio +1

    cs.CVcs.LGeess.SParXiv:2003.09148v22020
  38. On Feature Learning in the Presence of Spurious Correlations

    Pavel Izmailov, Polina Kirichenko, Nate Gruver +1

    cs.LGcs.CVstat.MLarXiv:2210.11369v12022
  39. Trial and Error: Exploration-Based Trajectory Optimization for LLM Agents

    Yifan Song, Da Yin, Xiang Yue +3

    cs.CLcs.AIcs.LGarXiv:2403.02502v22024
  40. AlphaEarth Foundations: An embedding field model for accurate and efficient global mapping from sparse label data

    Christopher F. Brown, Michal R. Kazmierski, Valerie J. Pasquarella +16

    cs.CVcs.LGarXiv:2507.22291v22025
  41. CAT-Flow: Curvature-Adaptive sTeps for Flow Matching

    Qinchan Li, Pedro Cisneros-Velarde, Keru Fu +3

    cs.LGarXiv:2609.01746v12026
  42. Tri-Band Channel Measurement-Enabled Multi-Layer Digital Twin for Terahertz Wireless Data Centers

    Mingjie Zhu, Ziming Yu, Guangjian Wang +1

    cs.LGcs.ITarXiv:2609.01699v12026
  43. Kimi-Audio Technical Report

    KimiTeam, Ding Ding, Zeqian Ju +37

    eess.AScs.AIcs.CLarXiv:2504.18425v12025
  44. The Structure of Quantization Damage in LLMs: Why the Next Bit Should Be Spent Globally

    Jundong Hu, Shekar Ramachandran

    cs.LGcs.CLarXiv:2609.01587v12026
  45. Improved Gradient Descent Lower Bounds Beyond Nesterov

    Yuhan Ye, Kaizhao Liu

    math.OCcs.LGstat.MLarXiv:2609.02855v22026
  46. Demystifying Long Chain-of-Thought Reasoning in LLMs

    Edward Yeo, Yuxuan Tong, Morry Niu +2

    cs.CLcs.LGarXiv:2502.03373v12025
  47. DiffusionNFT: Online Diffusion Reinforcement with Forward Process

    Kaiwen Zheng, Huayu Chen, Haotian Ye +7

    cs.LGcs.AIcs.CVarXiv:2509.16117v22025
  48. AReaL: A Large-Scale Asynchronous Reinforcement Learning System for Language Reasoning

    Wei Fu, Jiaxuan Gao, Xujie Shen +10

    cs.LGcs.AIarXiv:2505.24298v52025
  49. TabICL: A Tabular Foundation Model for In-Context Learning on Large Data

    Jingang Qu, David Holzmüller, Gaël Varoquaux +1

    cs.LGcs.AIarXiv:2502.05564v22025
  50. Learning the Pareto Front with Hypernetworks

    Aviv Navon, Aviv Shamsian, Gal Chechik +1

    cs.LGarXiv:2010.04104v22020
  51. UMA: A Family of Universal Models for Atoms

    Brandon M. Wood, Misko Dzamba, Xiang Fu +15

    cs.LGarXiv:2506.23971v22025
  52. Automated Directed Fairness Testing

    Sakshi Udeshi, Pryanshu Arora, Sudipta Chattopadhyay

    cs.LGcs.AIcs.SEarXiv:1807.00468v22018
  53. DENSE: Data-Free One-Shot Federated Learning

    Jie Zhang, Chen Chen, Bo Li +5

    cs.LGcs.CVarXiv:2112.12371v22021
  54. Strongly Adaptive Online Learning

    Amit Daniely, Alon Gonen, Shai Shalev-Shwartz

    cs.LGarXiv:1502.07073v32015
  55. Doc2EDAG: An End-to-End Document-level Framework for Chinese Financial Event Extraction

    Shun Zheng, Wei Cao, Wei Xu +1

    cs.CLcs.LGarXiv:1904.07535v22019
  56. Graph Neural Networks in IoT: A Survey

    Guimin Dong, Mingyue Tang, Zhiyuan Wang +7

    cs.LGarXiv:2203.15935v22022
  57. Cascaded V-Net using ROI masks for brain tumor segmentation

    Adrià Casamitjana, Marcel Catà, Irina Sánchez +2

    cs.CVcs.AIcs.CYarXiv:1812.11588v12018
  58. A Survey of Safety and Trustworthiness of Large Language Models through the Lens of Verification and Validation

    Xiaowei Huang, Wenjie Ruan, Wei Huang +14

    cs.AIcs.LGarXiv:2305.11391v22023
  59. When Vision Meets Graphs: A Survey on Graph Reasoning and Learning

    Xinjian Zhao, Wei Pang, Zhixuan Yu +8

    cs.SIcs.CVcs.LGarXiv:2609.03816v12026
  60. Most Ligand-Based Classification Benchmarks Reward Memorization Rather than Generalization

    Izhar Wallach, Abraham Heifets

    q-bio.QMcs.LGstat.MLarXiv:1706.06619v22017