Machine Learning
Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
8,221 to 8,280 of 20,199
Survey of state-of-the-art mixed data clustering algorithms
Amir Ahmad, Shehroz S. Khan
cs.LGcs.AIstat.MLarXiv:1811.04364v62018Enhanced Ensemble Clustering via Fast Propagation of Cluster-wise Similarities
Dong Huang, Chang-Dong Wang, Hongxing Peng +2
cs.LGstat.MLarXiv:1810.12544v12018Solving In-Table Prediction Problems by Deep Neural Networks with Performance Evaluation Using Synthetic Data
Xiao Zhao, Daniela Oelke
cs.LGarXiv:2609.01262v12026Motus: A Unified Latent Action World Model
Hongzhe Bi, Hengkai Tan, Shenghao Xie +13
cs.CVcs.LGcs.ROarXiv:2512.13030v22025Smart Predict-and-Optimize for Hard Combinatorial Optimization Problems
Jaynta Mandi, Emir Demirović, Peter. J Stuckey +1
cs.LGcs.AImath.OCarXiv:1911.10092v12019Latent unified smooth Hamiltonians for excited state chemistry
David Juergens, Martin Stöhr, Andreas E. Hillers-Bendtsen +2
physics.chem-phcs.LGarXiv:2609.01871v12026VL-Rethinker: Incentivizing Self-Reflection of Vision-Language Models with Reinforcement Learning
Haozhe Wang, Chao Qu, Zuming Huang +3
cs.LGcs.AIarXiv:2504.08837v32025Classification-based Financial Markets Prediction using Deep Neural Networks
Matthew Dixon, Diego Klabjan, Jin Hoon Bang
cs.LGcs.CEarXiv:1603.08604v22016Learning to Reason under Off-Policy Guidance
Jianhao Yan, Yafu Li, Zican Hu +5
cs.LGcs.AIcs.CLarXiv:2504.14945v52025Frozen Cores Need Task Signal: Fisher-Whitened Cross-Covariance for Low-Resource LLM Adaptation
Wentao Ye, Zhanming Shen, Zhiqing Xiao +3
cs.LGarXiv:2609.00762v12026LeJEPA: Provable and Scalable Self-Supervised Learning Without the Heuristics
Randall Balestriero, Yann LeCun
cs.LGcs.AIcs.CVarXiv:2511.08544v32025Compute and Energy Consumption Trends in Deep Learning Inference
Radosvet Desislavov, Fernando Martínez-Plumed, José Hernández-Orallo
cs.LGcs.AIarXiv:2109.05472v22021DRACO: Fine-Grained Credit Assignment with Dynamic Rubrics for Long-Horizon Agent Training
Shubham Gandhi, Saurabh Goyal, Kiran Kate +1
cs.AIcs.LGcs.SEarXiv:2609.04094v12026Learning Mixtures of Submodular Shells with Application to Document Summarization
Hui Lin, Jeff A. Bilmes
cs.LGcs.CLcs.IRarXiv:1210.4871v12012Accelerating scientific discovery with Co-Scientist
Juraj Gottweis, Wei-Hung Weng, Alexander Daryin +48
cs.AIcs.CLcs.HCarXiv:2502.18864v22025DenseRaC: Joint 3D Pose and Shape Estimation by Dense Render-and-Compare
Yuanlu Xu, Song-Chun Zhu, Tony Tung
cs.CVcs.LGeess.IVarXiv:1910.00116v22019ToolRL: Reward is All Tool Learning Needs
Cheng Qian, Emre Can Acikgoz, Qi He +5
cs.LGcs.AIcs.CLarXiv:2504.13958v12025Early Detection of Breast Cancer using SVM Classifier Technique
Y. Ireaneus Anna Rejani, S. Thamarai Selvi
cs.LGarXiv:0912.2314v12009Humanity's Last Exam
Long Phan, Alice Gatti, Ziwen Han +1155
cs.LGcs.AIcs.CLarXiv:2501.14249v112025ICDAR 2019 Competition on Large-scale Street View Text with Partial Labeling -- RRC-LSVT
Yipeng Sun, Zihan Ni, Chee-Kheng Chng +9
cs.CVcs.LGcs.MMarXiv:1909.07741v12019Fast and Eager k-Medoids Clustering: O(k) Runtime Improvement of the PAM, CLARA, and CLARANS Algorithms
Erich Schubert, Peter J. Rousseeuw
cs.LGcs.AIstat.MLarXiv:2008.05171v22020The Illusion of Thinking: Understanding the Strengths and Limitations of Reasoning Models via the Lens of Problem Complexity
Parshin Shojaee, Iman Mirzadeh, Keivan Alizadeh +3
cs.AIcs.CLcs.LGarXiv:2506.06941v32025Tree-based Intelligent Intrusion Detection System in Internet of Vehicles
Li Yang, Abdallah Moubayed, Ismail Hamieh +1
cs.LGcs.CRstat.MLarXiv:1910.08635v22019On the Effect of Dropping Layers of Pre-trained Transformer Models
Hassan Sajjad, Fahim Dalvi, Nadir Durrani +1
cs.CLcs.LGarXiv:2004.03844v32020L1: Controlling How Long A Reasoning Model Thinks With Reinforcement Learning
Pranjal Aggarwal, Sean Welleck
cs.CLcs.AIcs.LGarXiv:2503.04697v22025Deep Fully-Connected Networks for Video Compressive Sensing
Michael Iliadis, Leonidas Spinoulas, Aggelos K. Katsaggelos
cs.CVcs.LGcs.MMarXiv:1603.04930v22016Agent Laboratory: Using LLM Agents as Research Assistants
Samuel Schmidgall, Yusheng Su, Ze Wang +7
cs.HCcs.AIcs.CLarXiv:2501.04227v22025GCN-GAN: A Non-linear Temporal Link Prediction Model for Weighted Dynamic Networks
Kai Lei, Meng Qin, Bo Bai +2
cs.SIcs.LGcs.NIarXiv:1901.09165v12019Early Detection of Combustion Instabilities using Deep Convolutional Selective Autoencoders on Hi-speed Flame Video
Adedotun Akintayo, Kin Gwn Lore, Soumalya Sarkar +1
cs.CVcs.LGcs.NEarXiv:1603.07839v12016Modality Competition: What Makes Joint Training of Multi-modal Network Fail in Deep Learning? (Provably)
Yu Huang, Junyang Lin, Chang Zhou +2
cs.LGarXiv:2203.12221v12022Accelerating DNN Training in Wireless Federated Edge Learning Systems
Jinke Ren, Guanding Yu, Guangyao Ding
cs.LGeess.SParXiv:1905.09712v32019Scaling up Test-Time Compute with Latent Reasoning: A Recurrent Depth Approach
Jonas Geiping, Sean McLeish, Neel Jain +6
cs.LGcs.CLarXiv:2502.05171v22025GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning
GLM-V Team, :, Wenyi Hong +91
cs.CVcs.AIcs.LGarXiv:2507.01006v62025Adaptive Feature Selection Guided Deep Forest for COVID-19 Classification with Chest CT
Liang Sun, Zhanhao Mo, Fuhua Yan +14
eess.IVcs.CVcs.LGarXiv:2005.03264v12020Practical Deep Reinforcement Learning Approach for Stock Trading
Xiao-Yang Liu, Zhuoran Xiong, Shan Zhong +2
cs.LGq-fin.TRstat.MLarXiv:1811.07522v32018Persona Vectors: Monitoring and Controlling Character Traits in Language Models
Runjin Chen, Andy Arditi, Henry Sleight +2
cs.CLcs.LGarXiv:2507.21509v32025Momentum in large-batch training: Polyak enlarges the critical batch size, Nesterov improves data efficiency
Jia-Nan Wang, Zixun Huang, Kairui Li +1
stat.MLcs.LGmath.OCarXiv:2609.02728v12026Large-Scale Screening of COVID-19 from Community Acquired Pneumonia using Infection Size-Aware Classification
Feng Shi, Liming Xia, Fei Shan +7
eess.IVcs.CVcs.LGarXiv:2003.09860v12020QCell: Recombining and Aligning Cell Queries for Overlapping Instance Segmentation
Yaroslav Prytula, Anton Popov, Dmytro Fishman
cs.CVcs.AIcs.LGarXiv:2608.29253v12026Muon is Scalable for LLM Training
Jingyuan Liu, Jianlin Su, Xingcheng Yao +25
cs.LGcs.AIcs.CLarXiv:2502.16982v12025The Lessons of Developing Process Reward Models in Mathematical Reasoning
Zhenru Zhang, Chujie Zheng, Yangzhen Wu +6
cs.CLcs.AIcs.LGarXiv:2501.07301v22025The Right Tool for the Job: Matching Model and Instance Complexities
Roy Schwartz, Gabriel Stanovsky, Swabha Swayamdipta +2
cs.CLcs.LGarXiv:2004.07453v22020A Unified Approach to Error Bounds for Structured Convex Optimization Problems
Zirui Zhou, Anthony Man-Cho So
math.OCcs.LGmath.NAarXiv:1512.03518v12015Deep Probabilistic Programming
Dustin Tran, Matthew D. Hoffman, Rif A. Saurous +3
stat.MLcs.AIcs.LGarXiv:1701.03757v22017Diffusion Transformers with Representation Autoencoders
Boyang Zheng, Nanye Ma, Shengbang Tong +1
cs.CVcs.LGarXiv:2510.11690v12025Differentiable plasticity: training plastic neural networks with backpropagation
Thomas Miconi, Jeff Clune, Kenneth O. Stanley
cs.NEcs.LGstat.MLarXiv:1804.02464v32018Contrastive Learning for Label-Efficient Semantic Segmentation
Xiangyun Zhao, Raviteja Vemulapalli, Philip Mansfield +4
cs.CVcs.AIcs.LGarXiv:2012.06985v42020Few-Shot Class-Incremental Learning by Sampling Multi-Phase Tasks
Da-Wei Zhou, Han-Jia Ye, Liang Ma +3
cs.CVcs.LGarXiv:2203.17030v22022Cognitive Behaviors that Enable Self-Improving Reasoners, or, Four Habits of Highly Effective STaRs
Kanishk Gandhi, Ayush Chakravarthy, Anikait Singh +2
cs.CLcs.LGarXiv:2503.01307v22025Overparameterized Nonlinear Learning: Gradient Descent Takes the Shortest Path?
Samet Oymak, Mahdi Soltanolkotabi
cs.LGmath.OCstat.MLarXiv:1812.10004v12018Agentic Context Engineering: Evolving Contexts for Self-Improving Language Models
Qizheng Zhang, Changran Hu, Shubhangi Upasani +10
cs.LGcs.AIcs.CLarXiv:2510.04618v32025Reconstruction vs. Generation: Taming Optimization Dilemma in Latent Diffusion Models
Jingfeng Yao, Bin Yang, Xinggang Wang
cs.CVcs.LGarXiv:2501.01423v32025Block Diffusion: Interpolating Between Autoregressive and Diffusion Language Models
Marianne Arriola, Aaron Gokaslan, Justin T. Chiu +5
cs.LGcs.AIarXiv:2503.09573v32025Who Should I Trust: AI or Myself? Leveraging Human and AI Correctness Likelihood to Promote Appropriate Trust in AI-Assisted Decision-Making
Shuai Ma, Ying Lei, Xinru Wang +4
cs.HCcs.AIcs.LGarXiv:2301.05809v12023Process Reinforcement through Implicit Rewards
Ganqu Cui, Lifan Yuan, Zefan Wang +22
cs.LGcs.AIcs.CLarXiv:2502.01456v22025TrajMind: Chaining Role-Specialized LoRAs for Fast-and-Slow Collective Trajectory Anomaly Diagnosis
Jiahao Wu, Zhenqun Yang, Chen Jason Zhang +1
cs.LGarXiv:2609.02540v12026Summaries:한국어From Reweighting to Rewriting: Unlocking the Intervention Effects of Influential Samples in Training Data Attribution
Yuzhang Luo, Chenpeng Wang, Jianhui Chen +1
cs.CLcs.AIcs.LGarXiv:2609.02771v12026Coverage, Not Targeting: A Structural Regime in Multi-Turn Agent Credit Assignment
Chenyu Zhou, Qiliang Jiang, Shuning Wu +1
cs.LGcs.AIarXiv:2609.02417v12026Node Feature Extraction by Self-Supervised Multi-scale Neighborhood Prediction
Eli Chien, Wei-Cheng Chang, Cho-Jui Hsieh +4
cs.LGarXiv:2111.00064v32021Machine-Learning-Based Diagnostics of EEG Pathology
Lukas Alexander Wilhelm Gemein, Robin Tibor Schirrmeister, Patryk Chrabąszcz +5
eess.IVcs.LGeess.SParXiv:2002.05115v12020