Machine Learning

Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

8,221 to 8,280 of 20,199

  1. Survey of state-of-the-art mixed data clustering algorithms

    Amir Ahmad, Shehroz S. Khan

    cs.LGcs.AIstat.MLarXiv:1811.04364v62018
  2. Enhanced Ensemble Clustering via Fast Propagation of Cluster-wise Similarities

    Dong Huang, Chang-Dong Wang, Hongxing Peng +2

    cs.LGstat.MLarXiv:1810.12544v12018
  3. Solving In-Table Prediction Problems by Deep Neural Networks with Performance Evaluation Using Synthetic Data

    Xiao Zhao, Daniela Oelke

    cs.LGarXiv:2609.01262v12026
  4. Motus: A Unified Latent Action World Model

    Hongzhe Bi, Hengkai Tan, Shenghao Xie +13

    cs.CVcs.LGcs.ROarXiv:2512.13030v22025
  5. Smart Predict-and-Optimize for Hard Combinatorial Optimization Problems

    Jaynta Mandi, Emir Demirović, Peter. J Stuckey +1

    cs.LGcs.AImath.OCarXiv:1911.10092v12019
  6. Latent unified smooth Hamiltonians for excited state chemistry

    David Juergens, Martin Stöhr, Andreas E. Hillers-Bendtsen +2

    physics.chem-phcs.LGarXiv:2609.01871v12026
  7. VL-Rethinker: Incentivizing Self-Reflection of Vision-Language Models with Reinforcement Learning

    Haozhe Wang, Chao Qu, Zuming Huang +3

    cs.LGcs.AIarXiv:2504.08837v32025
  8. Classification-based Financial Markets Prediction using Deep Neural Networks

    Matthew Dixon, Diego Klabjan, Jin Hoon Bang

    cs.LGcs.CEarXiv:1603.08604v22016
  9. Learning to Reason under Off-Policy Guidance

    Jianhao Yan, Yafu Li, Zican Hu +5

    cs.LGcs.AIcs.CLarXiv:2504.14945v52025
  10. Frozen Cores Need Task Signal: Fisher-Whitened Cross-Covariance for Low-Resource LLM Adaptation

    Wentao Ye, Zhanming Shen, Zhiqing Xiao +3

    cs.LGarXiv:2609.00762v12026
  11. LeJEPA: Provable and Scalable Self-Supervised Learning Without the Heuristics

    Randall Balestriero, Yann LeCun

    cs.LGcs.AIcs.CVarXiv:2511.08544v32025
  12. Compute and Energy Consumption Trends in Deep Learning Inference

    Radosvet Desislavov, Fernando Martínez-Plumed, José Hernández-Orallo

    cs.LGcs.AIarXiv:2109.05472v22021
  13. DRACO: Fine-Grained Credit Assignment with Dynamic Rubrics for Long-Horizon Agent Training

    Shubham Gandhi, Saurabh Goyal, Kiran Kate +1

    cs.AIcs.LGcs.SEarXiv:2609.04094v12026
  14. Learning Mixtures of Submodular Shells with Application to Document Summarization

    Hui Lin, Jeff A. Bilmes

    cs.LGcs.CLcs.IRarXiv:1210.4871v12012
  15. Accelerating scientific discovery with Co-Scientist

    Juraj Gottweis, Wei-Hung Weng, Alexander Daryin +48

    cs.AIcs.CLcs.HCarXiv:2502.18864v22025
  16. DenseRaC: Joint 3D Pose and Shape Estimation by Dense Render-and-Compare

    Yuanlu Xu, Song-Chun Zhu, Tony Tung

    cs.CVcs.LGeess.IVarXiv:1910.00116v22019
  17. ToolRL: Reward is All Tool Learning Needs

    Cheng Qian, Emre Can Acikgoz, Qi He +5

    cs.LGcs.AIcs.CLarXiv:2504.13958v12025
  18. Early Detection of Breast Cancer using SVM Classifier Technique

    Y. Ireaneus Anna Rejani, S. Thamarai Selvi

    cs.LGarXiv:0912.2314v12009
  19. Humanity's Last Exam

    Long Phan, Alice Gatti, Ziwen Han +1155

    cs.LGcs.AIcs.CLarXiv:2501.14249v112025
  20. ICDAR 2019 Competition on Large-scale Street View Text with Partial Labeling -- RRC-LSVT

    Yipeng Sun, Zihan Ni, Chee-Kheng Chng +9

    cs.CVcs.LGcs.MMarXiv:1909.07741v12019
  21. Fast and Eager k-Medoids Clustering: O(k) Runtime Improvement of the PAM, CLARA, and CLARANS Algorithms

    Erich Schubert, Peter J. Rousseeuw

    cs.LGcs.AIstat.MLarXiv:2008.05171v22020
  22. The Illusion of Thinking: Understanding the Strengths and Limitations of Reasoning Models via the Lens of Problem Complexity

    Parshin Shojaee, Iman Mirzadeh, Keivan Alizadeh +3

    cs.AIcs.CLcs.LGarXiv:2506.06941v32025
  23. Tree-based Intelligent Intrusion Detection System in Internet of Vehicles

    Li Yang, Abdallah Moubayed, Ismail Hamieh +1

    cs.LGcs.CRstat.MLarXiv:1910.08635v22019
  24. On the Effect of Dropping Layers of Pre-trained Transformer Models

    Hassan Sajjad, Fahim Dalvi, Nadir Durrani +1

    cs.CLcs.LGarXiv:2004.03844v32020
  25. L1: Controlling How Long A Reasoning Model Thinks With Reinforcement Learning

    Pranjal Aggarwal, Sean Welleck

    cs.CLcs.AIcs.LGarXiv:2503.04697v22025
  26. Deep Fully-Connected Networks for Video Compressive Sensing

    Michael Iliadis, Leonidas Spinoulas, Aggelos K. Katsaggelos

    cs.CVcs.LGcs.MMarXiv:1603.04930v22016
  27. Agent Laboratory: Using LLM Agents as Research Assistants

    Samuel Schmidgall, Yusheng Su, Ze Wang +7

    cs.HCcs.AIcs.CLarXiv:2501.04227v22025
  28. GCN-GAN: A Non-linear Temporal Link Prediction Model for Weighted Dynamic Networks

    Kai Lei, Meng Qin, Bo Bai +2

    cs.SIcs.LGcs.NIarXiv:1901.09165v12019
  29. Early Detection of Combustion Instabilities using Deep Convolutional Selective Autoencoders on Hi-speed Flame Video

    Adedotun Akintayo, Kin Gwn Lore, Soumalya Sarkar +1

    cs.CVcs.LGcs.NEarXiv:1603.07839v12016
  30. Modality Competition: What Makes Joint Training of Multi-modal Network Fail in Deep Learning? (Provably)

    Yu Huang, Junyang Lin, Chang Zhou +2

    cs.LGarXiv:2203.12221v12022
  31. Accelerating DNN Training in Wireless Federated Edge Learning Systems

    Jinke Ren, Guanding Yu, Guangyao Ding

    cs.LGeess.SParXiv:1905.09712v32019
  32. Scaling up Test-Time Compute with Latent Reasoning: A Recurrent Depth Approach

    Jonas Geiping, Sean McLeish, Neel Jain +6

    cs.LGcs.CLarXiv:2502.05171v22025
  33. GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning

    GLM-V Team, :, Wenyi Hong +91

    cs.CVcs.AIcs.LGarXiv:2507.01006v62025
  34. Adaptive Feature Selection Guided Deep Forest for COVID-19 Classification with Chest CT

    Liang Sun, Zhanhao Mo, Fuhua Yan +14

    eess.IVcs.CVcs.LGarXiv:2005.03264v12020
  35. Practical Deep Reinforcement Learning Approach for Stock Trading

    Xiao-Yang Liu, Zhuoran Xiong, Shan Zhong +2

    cs.LGq-fin.TRstat.MLarXiv:1811.07522v32018
  36. Persona Vectors: Monitoring and Controlling Character Traits in Language Models

    Runjin Chen, Andy Arditi, Henry Sleight +2

    cs.CLcs.LGarXiv:2507.21509v32025
  37. Momentum in large-batch training: Polyak enlarges the critical batch size, Nesterov improves data efficiency

    Jia-Nan Wang, Zixun Huang, Kairui Li +1

    stat.MLcs.LGmath.OCarXiv:2609.02728v12026
  38. Large-Scale Screening of COVID-19 from Community Acquired Pneumonia using Infection Size-Aware Classification

    Feng Shi, Liming Xia, Fei Shan +7

    eess.IVcs.CVcs.LGarXiv:2003.09860v12020
  39. QCell: Recombining and Aligning Cell Queries for Overlapping Instance Segmentation

    Yaroslav Prytula, Anton Popov, Dmytro Fishman

    cs.CVcs.AIcs.LGarXiv:2608.29253v12026
  40. Muon is Scalable for LLM Training

    Jingyuan Liu, Jianlin Su, Xingcheng Yao +25

    cs.LGcs.AIcs.CLarXiv:2502.16982v12025
  41. The Lessons of Developing Process Reward Models in Mathematical Reasoning

    Zhenru Zhang, Chujie Zheng, Yangzhen Wu +6

    cs.CLcs.AIcs.LGarXiv:2501.07301v22025
  42. The Right Tool for the Job: Matching Model and Instance Complexities

    Roy Schwartz, Gabriel Stanovsky, Swabha Swayamdipta +2

    cs.CLcs.LGarXiv:2004.07453v22020
  43. A Unified Approach to Error Bounds for Structured Convex Optimization Problems

    Zirui Zhou, Anthony Man-Cho So

    math.OCcs.LGmath.NAarXiv:1512.03518v12015
  44. Deep Probabilistic Programming

    Dustin Tran, Matthew D. Hoffman, Rif A. Saurous +3

    stat.MLcs.AIcs.LGarXiv:1701.03757v22017
  45. Diffusion Transformers with Representation Autoencoders

    Boyang Zheng, Nanye Ma, Shengbang Tong +1

    cs.CVcs.LGarXiv:2510.11690v12025
  46. Differentiable plasticity: training plastic neural networks with backpropagation

    Thomas Miconi, Jeff Clune, Kenneth O. Stanley

    cs.NEcs.LGstat.MLarXiv:1804.02464v32018
  47. Contrastive Learning for Label-Efficient Semantic Segmentation

    Xiangyun Zhao, Raviteja Vemulapalli, Philip Mansfield +4

    cs.CVcs.AIcs.LGarXiv:2012.06985v42020
  48. Few-Shot Class-Incremental Learning by Sampling Multi-Phase Tasks

    Da-Wei Zhou, Han-Jia Ye, Liang Ma +3

    cs.CVcs.LGarXiv:2203.17030v22022
  49. Cognitive Behaviors that Enable Self-Improving Reasoners, or, Four Habits of Highly Effective STaRs

    Kanishk Gandhi, Ayush Chakravarthy, Anikait Singh +2

    cs.CLcs.LGarXiv:2503.01307v22025
  50. Overparameterized Nonlinear Learning: Gradient Descent Takes the Shortest Path?

    Samet Oymak, Mahdi Soltanolkotabi

    cs.LGmath.OCstat.MLarXiv:1812.10004v12018
  51. Agentic Context Engineering: Evolving Contexts for Self-Improving Language Models

    Qizheng Zhang, Changran Hu, Shubhangi Upasani +10

    cs.LGcs.AIcs.CLarXiv:2510.04618v32025
  52. Reconstruction vs. Generation: Taming Optimization Dilemma in Latent Diffusion Models

    Jingfeng Yao, Bin Yang, Xinggang Wang

    cs.CVcs.LGarXiv:2501.01423v32025
  53. Block Diffusion: Interpolating Between Autoregressive and Diffusion Language Models

    Marianne Arriola, Aaron Gokaslan, Justin T. Chiu +5

    cs.LGcs.AIarXiv:2503.09573v32025
  54. Who Should I Trust: AI or Myself? Leveraging Human and AI Correctness Likelihood to Promote Appropriate Trust in AI-Assisted Decision-Making

    Shuai Ma, Ying Lei, Xinru Wang +4

    cs.HCcs.AIcs.LGarXiv:2301.05809v12023
  55. Process Reinforcement through Implicit Rewards

    Ganqu Cui, Lifan Yuan, Zefan Wang +22

    cs.LGcs.AIcs.CLarXiv:2502.01456v22025
  56. TrajMind: Chaining Role-Specialized LoRAs for Fast-and-Slow Collective Trajectory Anomaly Diagnosis

    Jiahao Wu, Zhenqun Yang, Chen Jason Zhang +1

    cs.LGarXiv:2609.02540v12026
    Summaries:한국어
  57. From Reweighting to Rewriting: Unlocking the Intervention Effects of Influential Samples in Training Data Attribution

    Yuzhang Luo, Chenpeng Wang, Jianhui Chen +1

    cs.CLcs.AIcs.LGarXiv:2609.02771v12026
  58. Coverage, Not Targeting: A Structural Regime in Multi-Turn Agent Credit Assignment

    Chenyu Zhou, Qiliang Jiang, Shuning Wu +1

    cs.LGcs.AIarXiv:2609.02417v12026
  59. Node Feature Extraction by Self-Supervised Multi-scale Neighborhood Prediction

    Eli Chien, Wei-Cheng Chang, Cho-Jui Hsieh +4

    cs.LGarXiv:2111.00064v32021
  60. Machine-Learning-Based Diagnostics of EEG Pathology

    Lukas Alexander Wilhelm Gemein, Robin Tibor Schirrmeister, Patryk Chrabąszcz +5

    eess.IVcs.LGeess.SParXiv:2002.05115v12020