Machine Learning

Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

8,341 to 8,400 of 20,219

  1. CoT-VLA: Visual Chain-of-Thought Reasoning for Vision-Language-Action Models

    Qingqing Zhao, Yao Lu, Moo Jin Kim +12

    cs.CVcs.AIcs.LGarXiv:2503.22020v12025
  2. Machine Learning for Combinatorial Optimization: a Methodological Tour d'Horizon

    Yoshua Bengio, Andrea Lodi, Antoine Prouvost

    cs.LGstat.MLarXiv:1811.06128v22018
  3. Emergence of Invariance and Disentanglement in Deep Representations

    Alessandro Achille, Stefano Soatto

    cs.LGcs.AIstat.MLarXiv:1706.01350v32017
  4. Uncertainty Quantification in Machine Learning for Engineering Design and Health Prognostics: A Tutorial

    Venkat Nemani, Luca Biggio, Xun Huan +6

    cs.LGcs.AIarXiv:2305.04933v22023
  5. A Systematic Survey on Deep Generative Models for Graph Generation

    Xiaojie Guo, Liang Zhao

    cs.LGstat.MLarXiv:2007.06686v32020
  6. MADGAN: unsupervised Medical Anomaly Detection GAN using multiple adjacent brain MRI slice reconstruction

    Changhee Han, Leonardo Rundo, Kohei Murao +7

    cs.CVcs.LGeess.IVarXiv:2007.13559v22020
  7. A Comprehensive Survey of Convolutions in Deep Learning: Applications, Challenges, and Future Trends

    Abolfazl Younesi, Mohsen Ansari, MohammadAmin Fazli +3

    cs.LGcs.NEarXiv:2402.15490v22024
  8. Deep Learning on Chest X-ray Images to Detect and Evaluate Pneumonia Cases at the Era of COVID-19

    Karim Hammoudi, Halim Benhabiles, Mahmoud Melkemi +4

    eess.IVcs.CVcs.LGarXiv:2004.03399v12020
  9. F*: An Interpretable Transformation of the F-measure

    David J. Hand, Peter Christen, Nishadi Kirielle

    cs.LGcs.AIcs.CVarXiv:2008.00103v32020
  10. 8-Bit Approximations for Parallelism in Deep Learning

    Tim Dettmers

    cs.NEcs.LGarXiv:1511.04561v42015
  11. Baldur: Whole-Proof Generation and Repair with Large Language Models

    Emily First, Markus N. Rabe, Talia Ringer +1

    cs.LGcs.LOcs.SEarXiv:2303.04910v22023
  12. Signal Processing on Higher-Order Networks: Livin' on the Edge ... and Beyond

    Michael T. Schaub, Yu Zhu, Jean-Baptiste Seby +2

    cs.SIcs.LGphysics.soc-pharXiv:2101.05510v42021
  13. TableNet: Deep Learning model for end-to-end Table detection and Tabular data extraction from Scanned Document Images

    Shubham Paliwal, Vishwanath D, Rohit Rahul +2

    cs.CVcs.LGeess.IVarXiv:2001.01469v12020
  14. Continual Learning with Pre-Trained Models: A Survey

    Da-Wei Zhou, Hai-Long Sun, Jingyi Ning +2

    cs.LGcs.CVarXiv:2401.16386v22024
  15. HookNet: multi-resolution convolutional neural networks for semantic segmentation in histopathology whole-slide images

    Mart van Rijthoven, Maschenka Balkenhol, Karina Siliņa +2

    eess.IVcs.CVcs.LGarXiv:2006.12230v12020
  16. TGANet: Text-guided attention for improved polyp segmentation

    Nikhil Kumar Tomar, Debesh Jha, Ulas Bagci +1

    eess.IVcs.CVcs.LGarXiv:2205.04280v12022
  17. A Survey on Explainable Anomaly Detection

    Zhong Li, Yuxuan Zhu, Matthijs van Leeuwen

    cs.LGarXiv:2210.06959v22022
  18. Speculative Decoding: Exploiting Speculative Execution for Accelerating Seq2seq Generation

    Heming Xia, Tao Ge, Peiyi Wang +3

    cs.CLcs.LGarXiv:2203.16487v62022
  19. ResRep: Lossless CNN Pruning via Decoupling Remembering and Forgetting

    Xiaohan Ding, Tianxiang Hao, Jianchao Tan +4

    cs.LGcs.CVeess.IVarXiv:2007.03260v42020
  20. SAINT+: Integrating Temporal Features for EdNet Correctness Prediction

    Dongmin Shin, Yugeun Shim, Hangyeol Yu +3

    cs.CYcs.AIcs.LGarXiv:2010.12042v22020
  21. Evaluating Prerequisite Qualities for Learning End-to-End Dialog Systems

    Jesse Dodge, Andreea Gane, Xiang Zhang +5

    cs.CLcs.LGarXiv:1511.06931v62015
  22. Knowledge Distillation via Route Constrained Optimization

    Xiao Jin, Baoyun Peng, Yichao Wu +5

    cs.LGcs.CVarXiv:1904.09149v12019
  23. TensorFlow.js: Machine Learning for the Web and Beyond

    Daniel Smilkov, Nikhil Thorat, Yannick Assogba +17

    cs.LGarXiv:1901.05350v22019
  24. Scaling Laws, Tabular Data and Actuarial Ratemaking Models

    Ronald Richman

    cs.LGq-fin.RMarXiv:2609.03106v12026
  25. Learning 2-opt Heuristics for the Traveling Salesman Problem via Deep Reinforcement Learning

    Paulo R. de O. da Costa, Jason Rhuggenaath, Yingqian Zhang +1

    cs.LGcs.AIstat.MLarXiv:2004.01608v32020
  26. In-Hand Object Rotation via Rapid Motor Adaptation

    Haozhi Qi, Ashish Kumar, Roberto Calandra +2

    cs.ROcs.AIcs.CVarXiv:2210.04887v12022
  27. Learning Program Embeddings to Propagate Feedback on Student Code

    Chris Piech, Jonathan Huang, Andy Nguyen +3

    cs.LGcs.NEcs.SEarXiv:1505.05969v12015
  28. SenseBERT: Driving Some Sense into BERT

    Yoav Levine, Barak Lenz, Or Dagan +6

    cs.CLcs.LGarXiv:1908.05646v22019
  29. A First Look at Deep Learning Apps on Smartphones

    Mengwei Xu, Jiawei Liu, Yuanqiang Liu +3

    cs.LGcs.CYarXiv:1812.05448v42018
  30. BharatGather: A Culturally-Informed Benchmark Dataset for Misinformation and Fake News Detection in Indian Public Events

    Parth Bramhecha, Smit Deshmukh, Sairaj Bodhale +2

    cs.CLcs.LGarXiv:2609.02895v12026
  31. Training Data Influence Analysis and Estimation: A Survey

    Zayd Hammoudeh, Daniel Lowd

    cs.LGarXiv:2212.04612v32022
  32. Inversion by Direct Iteration: An Alternative to Denoising Diffusion for Image Restoration

    Mauricio Delbracio, Peyman Milanfar

    eess.IVcs.CVcs.LGarXiv:2303.11435v52023
  33. SCOP: Scientific Control for Reliable Neural Network Pruning

    Yehui Tang, Yunhe Wang, Yixing Xu +4

    cs.CVcs.LGarXiv:2010.10732v22020
  34. A Simple Neural Attentive Meta-Learner

    Nikhil Mishra, Mostafa Rohaninejad, Xi Chen +1

    cs.AIcs.LGcs.NEarXiv:1707.03141v32017
  35. Use HiResCAM instead of Grad-CAM for faithful explanations of convolutional neural networks

    Rachel Lea Draelos, Lawrence Carin

    eess.IVcs.CVcs.LGarXiv:2011.08891v42020
  36. UniVLA: Learning to Act Anywhere with Task-centric Latent Actions

    Qingwen Bu, Yanting Yang, Jisong Cai +5

    cs.ROcs.AIcs.LGarXiv:2505.06111v32025
  37. SimpleRL-Zoo: Investigating and Taming Zero Reinforcement Learning for Open Base Models in the Wild

    Weihao Zeng, Yuzhen Huang, Qian Liu +4

    cs.LGcs.AIcs.CLarXiv:2503.18892v32025
  38. Scaling Near-Optimal SFT-RL Annotation Budget Allocation from Small to Large LLMs

    Jingtan Wang, Arun Verma, Xiaoqiang Lin +4

    cs.CLcs.AIcs.LGarXiv:2609.01573v12026
  39. Intriguing Properties of Contrastive Losses

    Ting Chen, Calvin Luo, Lala Li

    cs.LGcs.AIcs.CVarXiv:2011.02803v32020
  40. GEPA: Reflective Prompt Evolution Can Outperform Reinforcement Learning

    Lakshya A Agrawal, Shangyin Tan, Dilara Soylu +14

    cs.CLcs.AIcs.LGarXiv:2507.19457v22025
  41. A Study of Conditional Diffusion Models for Open-Loop Control under Dry Friction and Stiction

    Eric Aislan Antonelo

    cs.LGarXiv:2609.01756v12026
  42. Sequence-to-Sequence Knowledge Graph Completion and Question Answering

    Apoorv Saxena, Adrian Kochsiek, Rainer Gemulla

    cs.CLcs.LGarXiv:2203.10321v12022
  43. LeWorldModel: Stable End-to-End Joint-Embedding Predictive Architecture from Pixels

    Lucas Maes, Quentin Le Lidec, Damien Scieur +2

    cs.LGcs.AIarXiv:2603.19312v32026
  44. Graph Convolution for Multimodal Information Extraction from Visually Rich Documents

    Xiaojing Liu, Feiyu Gao, Qiong Zhang +1

    cs.IRcs.CVcs.LGarXiv:1903.11279v12019
  45. TSLANet: Rethinking Transformers for Time Series Representation Learning

    Emadeldeen Eldele, Mohamed Ragab, Zhenghua Chen +2

    cs.LGstat.MLarXiv:2404.08472v22024
  46. Global Self-Attention as a Replacement for Graph Convolution

    Md Shamim Hussain, Mohammed J. Zaki, Dharmashankar Subramanian

    cs.LGarXiv:2108.03348v32021
  47. ChronoNet: A Deep Recurrent Neural Network for Abnormal EEG Identification

    Subhrajit Roy, Isabell Kiral-Kornek, Stefan Harrer

    eess.SPcs.LGarXiv:1802.00308v22018
  48. The Entropy Mechanism of Reinforcement Learning for Reasoning Language Models

    Ganqu Cui, Yuchen Zhang, Jiacheng Chen +14

    cs.LGcs.AIcs.CLarXiv:2505.22617v12025
  49. Vision-R1: Incentivizing Reasoning Capability in Multimodal Large Language Models

    Wenxuan Huang, Bohan Jia, Zijie Zhai +7

    cs.CVcs.AIcs.CLarXiv:2503.06749v42025
  50. Simple linear attention language models balance the recall-throughput tradeoff

    Simran Arora, Sabri Eyuboglu, Michael Zhang +6

    cs.CLcs.LGarXiv:2402.18668v22024
  51. FAST: Efficient Action Tokenization for Vision-Language-Action Models

    Karl Pertsch, Kyle Stachowicz, Brian Ichter +6

    cs.ROcs.LGarXiv:2501.09747v12025
  52. Beyond the 80/20 Rule: High-Entropy Minority Tokens Drive Effective Reinforcement Learning for LLM Reasoning

    Shenzhi Wang, Le Yu, Chang Gao +15

    cs.CLcs.AIcs.LGarXiv:2506.01939v22025
  53. Optimus: Organizing Sentences via Pre-trained Modeling of a Latent Space

    Chunyuan Li, Xiang Gao, Yuan Li +4

    cs.CLcs.LGstat.MLarXiv:2004.04092v42020
  54. The State of the Art in Enhancing Trust in Machine Learning Models with the Use of Visualizations

    A. Chatzimparmpas, R. Martins, I. Jusufi +3

    cs.LGcs.HCstat.MLarXiv:2212.11737v22022
  55. CrossFit: A Few-shot Learning Challenge for Cross-task Generalization in NLP

    Qinyuan Ye, Bill Yuchen Lin, Xiang Ren

    cs.CLcs.LGarXiv:2104.08835v22021
  56. Phi-4-Mini Technical Report: Compact yet Powerful Multimodal Language Models via Mixture-of-LoRAs

    Microsoft, :, Abdelrahman Abouelenin +73

    cs.CLcs.AIcs.LGarXiv:2503.01743v22025
  57. Online Non-Monotone DR-Submodular Maximization Matching the Offline $0.401$ Factor

    Vaneet Aggarwal, Yiyang Lu

    cs.LGcs.AIcs.CCarXiv:2609.02145v12026
  58. Open-ended Learning in Symmetric Zero-sum Games

    David Balduzzi, Marta Garnelo, Yoram Bachrach +4

    cs.LGcs.GTcs.MAarXiv:1901.08106v22019
  59. Self-Distilled Reasoner: On-Policy Self-Distillation for Large Language Models

    Siyan Zhao, Zhihui Xie, Mengchen Liu +4

    cs.LGcs.CLarXiv:2601.18734v32026
  60. Codebook Agent: Amortized Topology Design for LLM Multi-Agent Systems

    Jinxi Yu, Yubei Li, Eric Hanchen Jiang +6

    cs.AIcs.LGcs.MAarXiv:2609.02264v12026