Machine Learning
Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
8,341 to 8,400 of 20,219
CoT-VLA: Visual Chain-of-Thought Reasoning for Vision-Language-Action Models
Qingqing Zhao, Yao Lu, Moo Jin Kim +12
cs.CVcs.AIcs.LGarXiv:2503.22020v12025Machine Learning for Combinatorial Optimization: a Methodological Tour d'Horizon
Yoshua Bengio, Andrea Lodi, Antoine Prouvost
cs.LGstat.MLarXiv:1811.06128v22018Emergence of Invariance and Disentanglement in Deep Representations
Alessandro Achille, Stefano Soatto
cs.LGcs.AIstat.MLarXiv:1706.01350v32017Uncertainty Quantification in Machine Learning for Engineering Design and Health Prognostics: A Tutorial
Venkat Nemani, Luca Biggio, Xun Huan +6
cs.LGcs.AIarXiv:2305.04933v22023A Systematic Survey on Deep Generative Models for Graph Generation
Xiaojie Guo, Liang Zhao
cs.LGstat.MLarXiv:2007.06686v32020MADGAN: unsupervised Medical Anomaly Detection GAN using multiple adjacent brain MRI slice reconstruction
Changhee Han, Leonardo Rundo, Kohei Murao +7
cs.CVcs.LGeess.IVarXiv:2007.13559v22020A Comprehensive Survey of Convolutions in Deep Learning: Applications, Challenges, and Future Trends
Abolfazl Younesi, Mohsen Ansari, MohammadAmin Fazli +3
cs.LGcs.NEarXiv:2402.15490v22024Deep Learning on Chest X-ray Images to Detect and Evaluate Pneumonia Cases at the Era of COVID-19
Karim Hammoudi, Halim Benhabiles, Mahmoud Melkemi +4
eess.IVcs.CVcs.LGarXiv:2004.03399v12020F*: An Interpretable Transformation of the F-measure
David J. Hand, Peter Christen, Nishadi Kirielle
cs.LGcs.AIcs.CVarXiv:2008.00103v320208-Bit Approximations for Parallelism in Deep Learning
Tim Dettmers
cs.NEcs.LGarXiv:1511.04561v42015Baldur: Whole-Proof Generation and Repair with Large Language Models
Emily First, Markus N. Rabe, Talia Ringer +1
cs.LGcs.LOcs.SEarXiv:2303.04910v22023Signal Processing on Higher-Order Networks: Livin' on the Edge ... and Beyond
Michael T. Schaub, Yu Zhu, Jean-Baptiste Seby +2
cs.SIcs.LGphysics.soc-pharXiv:2101.05510v42021TableNet: Deep Learning model for end-to-end Table detection and Tabular data extraction from Scanned Document Images
Shubham Paliwal, Vishwanath D, Rohit Rahul +2
cs.CVcs.LGeess.IVarXiv:2001.01469v12020Continual Learning with Pre-Trained Models: A Survey
Da-Wei Zhou, Hai-Long Sun, Jingyi Ning +2
cs.LGcs.CVarXiv:2401.16386v22024HookNet: multi-resolution convolutional neural networks for semantic segmentation in histopathology whole-slide images
Mart van Rijthoven, Maschenka Balkenhol, Karina Siliņa +2
eess.IVcs.CVcs.LGarXiv:2006.12230v12020TGANet: Text-guided attention for improved polyp segmentation
Nikhil Kumar Tomar, Debesh Jha, Ulas Bagci +1
eess.IVcs.CVcs.LGarXiv:2205.04280v12022A Survey on Explainable Anomaly Detection
Zhong Li, Yuxuan Zhu, Matthijs van Leeuwen
cs.LGarXiv:2210.06959v22022Speculative Decoding: Exploiting Speculative Execution for Accelerating Seq2seq Generation
Heming Xia, Tao Ge, Peiyi Wang +3
cs.CLcs.LGarXiv:2203.16487v62022ResRep: Lossless CNN Pruning via Decoupling Remembering and Forgetting
Xiaohan Ding, Tianxiang Hao, Jianchao Tan +4
cs.LGcs.CVeess.IVarXiv:2007.03260v42020SAINT+: Integrating Temporal Features for EdNet Correctness Prediction
Dongmin Shin, Yugeun Shim, Hangyeol Yu +3
cs.CYcs.AIcs.LGarXiv:2010.12042v22020Evaluating Prerequisite Qualities for Learning End-to-End Dialog Systems
Jesse Dodge, Andreea Gane, Xiang Zhang +5
cs.CLcs.LGarXiv:1511.06931v62015Knowledge Distillation via Route Constrained Optimization
Xiao Jin, Baoyun Peng, Yichao Wu +5
cs.LGcs.CVarXiv:1904.09149v12019TensorFlow.js: Machine Learning for the Web and Beyond
Daniel Smilkov, Nikhil Thorat, Yannick Assogba +17
cs.LGarXiv:1901.05350v22019Scaling Laws, Tabular Data and Actuarial Ratemaking Models
Ronald Richman
cs.LGq-fin.RMarXiv:2609.03106v12026Learning 2-opt Heuristics for the Traveling Salesman Problem via Deep Reinforcement Learning
Paulo R. de O. da Costa, Jason Rhuggenaath, Yingqian Zhang +1
cs.LGcs.AIstat.MLarXiv:2004.01608v32020In-Hand Object Rotation via Rapid Motor Adaptation
Haozhi Qi, Ashish Kumar, Roberto Calandra +2
cs.ROcs.AIcs.CVarXiv:2210.04887v12022Learning Program Embeddings to Propagate Feedback on Student Code
Chris Piech, Jonathan Huang, Andy Nguyen +3
cs.LGcs.NEcs.SEarXiv:1505.05969v12015SenseBERT: Driving Some Sense into BERT
Yoav Levine, Barak Lenz, Or Dagan +6
cs.CLcs.LGarXiv:1908.05646v22019A First Look at Deep Learning Apps on Smartphones
Mengwei Xu, Jiawei Liu, Yuanqiang Liu +3
cs.LGcs.CYarXiv:1812.05448v42018BharatGather: A Culturally-Informed Benchmark Dataset for Misinformation and Fake News Detection in Indian Public Events
Parth Bramhecha, Smit Deshmukh, Sairaj Bodhale +2
cs.CLcs.LGarXiv:2609.02895v12026Training Data Influence Analysis and Estimation: A Survey
Zayd Hammoudeh, Daniel Lowd
cs.LGarXiv:2212.04612v32022Inversion by Direct Iteration: An Alternative to Denoising Diffusion for Image Restoration
Mauricio Delbracio, Peyman Milanfar
eess.IVcs.CVcs.LGarXiv:2303.11435v52023SCOP: Scientific Control for Reliable Neural Network Pruning
Yehui Tang, Yunhe Wang, Yixing Xu +4
cs.CVcs.LGarXiv:2010.10732v22020A Simple Neural Attentive Meta-Learner
Nikhil Mishra, Mostafa Rohaninejad, Xi Chen +1
cs.AIcs.LGcs.NEarXiv:1707.03141v32017Use HiResCAM instead of Grad-CAM for faithful explanations of convolutional neural networks
Rachel Lea Draelos, Lawrence Carin
eess.IVcs.CVcs.LGarXiv:2011.08891v42020UniVLA: Learning to Act Anywhere with Task-centric Latent Actions
Qingwen Bu, Yanting Yang, Jisong Cai +5
cs.ROcs.AIcs.LGarXiv:2505.06111v32025SimpleRL-Zoo: Investigating and Taming Zero Reinforcement Learning for Open Base Models in the Wild
Weihao Zeng, Yuzhen Huang, Qian Liu +4
cs.LGcs.AIcs.CLarXiv:2503.18892v32025Scaling Near-Optimal SFT-RL Annotation Budget Allocation from Small to Large LLMs
Jingtan Wang, Arun Verma, Xiaoqiang Lin +4
cs.CLcs.AIcs.LGarXiv:2609.01573v12026Intriguing Properties of Contrastive Losses
Ting Chen, Calvin Luo, Lala Li
cs.LGcs.AIcs.CVarXiv:2011.02803v32020GEPA: Reflective Prompt Evolution Can Outperform Reinforcement Learning
Lakshya A Agrawal, Shangyin Tan, Dilara Soylu +14
cs.CLcs.AIcs.LGarXiv:2507.19457v22025A Study of Conditional Diffusion Models for Open-Loop Control under Dry Friction and Stiction
Eric Aislan Antonelo
cs.LGarXiv:2609.01756v12026Sequence-to-Sequence Knowledge Graph Completion and Question Answering
Apoorv Saxena, Adrian Kochsiek, Rainer Gemulla
cs.CLcs.LGarXiv:2203.10321v12022LeWorldModel: Stable End-to-End Joint-Embedding Predictive Architecture from Pixels
Lucas Maes, Quentin Le Lidec, Damien Scieur +2
cs.LGcs.AIarXiv:2603.19312v32026Graph Convolution for Multimodal Information Extraction from Visually Rich Documents
Xiaojing Liu, Feiyu Gao, Qiong Zhang +1
cs.IRcs.CVcs.LGarXiv:1903.11279v12019TSLANet: Rethinking Transformers for Time Series Representation Learning
Emadeldeen Eldele, Mohamed Ragab, Zhenghua Chen +2
cs.LGstat.MLarXiv:2404.08472v22024Global Self-Attention as a Replacement for Graph Convolution
Md Shamim Hussain, Mohammed J. Zaki, Dharmashankar Subramanian
cs.LGarXiv:2108.03348v32021ChronoNet: A Deep Recurrent Neural Network for Abnormal EEG Identification
Subhrajit Roy, Isabell Kiral-Kornek, Stefan Harrer
eess.SPcs.LGarXiv:1802.00308v22018The Entropy Mechanism of Reinforcement Learning for Reasoning Language Models
Ganqu Cui, Yuchen Zhang, Jiacheng Chen +14
cs.LGcs.AIcs.CLarXiv:2505.22617v12025Vision-R1: Incentivizing Reasoning Capability in Multimodal Large Language Models
Wenxuan Huang, Bohan Jia, Zijie Zhai +7
cs.CVcs.AIcs.CLarXiv:2503.06749v42025Simple linear attention language models balance the recall-throughput tradeoff
Simran Arora, Sabri Eyuboglu, Michael Zhang +6
cs.CLcs.LGarXiv:2402.18668v22024FAST: Efficient Action Tokenization for Vision-Language-Action Models
Karl Pertsch, Kyle Stachowicz, Brian Ichter +6
cs.ROcs.LGarXiv:2501.09747v12025Beyond the 80/20 Rule: High-Entropy Minority Tokens Drive Effective Reinforcement Learning for LLM Reasoning
Shenzhi Wang, Le Yu, Chang Gao +15
cs.CLcs.AIcs.LGarXiv:2506.01939v22025Optimus: Organizing Sentences via Pre-trained Modeling of a Latent Space
Chunyuan Li, Xiang Gao, Yuan Li +4
cs.CLcs.LGstat.MLarXiv:2004.04092v42020The State of the Art in Enhancing Trust in Machine Learning Models with the Use of Visualizations
A. Chatzimparmpas, R. Martins, I. Jusufi +3
cs.LGcs.HCstat.MLarXiv:2212.11737v22022CrossFit: A Few-shot Learning Challenge for Cross-task Generalization in NLP
Qinyuan Ye, Bill Yuchen Lin, Xiang Ren
cs.CLcs.LGarXiv:2104.08835v22021Phi-4-Mini Technical Report: Compact yet Powerful Multimodal Language Models via Mixture-of-LoRAs
Microsoft, :, Abdelrahman Abouelenin +73
cs.CLcs.AIcs.LGarXiv:2503.01743v22025Online Non-Monotone DR-Submodular Maximization Matching the Offline $0.401$ Factor
Vaneet Aggarwal, Yiyang Lu
cs.LGcs.AIcs.CCarXiv:2609.02145v12026Open-ended Learning in Symmetric Zero-sum Games
David Balduzzi, Marta Garnelo, Yoram Bachrach +4
cs.LGcs.GTcs.MAarXiv:1901.08106v22019Self-Distilled Reasoner: On-Policy Self-Distillation for Large Language Models
Siyan Zhao, Zhihui Xie, Mengchen Liu +4
cs.LGcs.CLarXiv:2601.18734v32026Codebook Agent: Amortized Topology Design for LLM Multi-Agent Systems
Jinxi Yu, Yubei Li, Eric Hanchen Jiang +6
cs.AIcs.LGcs.MAarXiv:2609.02264v12026