Machine Learning

Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

6,661 to 6,720 of 20,217

  1. Tabular LLMs for Interpretable Few-Shot Alzheimer's Disease Prediction with Multimodal Biomedical Data

    Sophie Kearney, Shu Yang, Zixuan Wen +8

    cs.CLcs.LGq-bio.QMarXiv:2603.17191v12026
  2. Vision-Language-Action Models for Robotics: A Review Towards Real-World Applications

    Kento Kawaharazuka, Jihoon Oh, Jun Yamada +2

    cs.ROcs.AIcs.CVarXiv:2510.07077v12025
  3. Training-free Detection of Generated Videos via Spatial-Temporal Likelihoods

    Omer Ben Hayun, Roy Betser, Meir Yossef Levi +2

    cs.CVcs.LGarXiv:2603.15026v22026
  4. Designing ECG Monitoring Healthcare System with Federated Transfer Learning and Explainable AI

    Ali Raza, Kim Phuc Tran, Ludovic Koehl +1

    cs.LGcs.AIeess.SParXiv:2105.12497v22021
  5. Unified Vision-Language Modeling via Concept Space Alignment

    Yifu Qiu, Paul-Ambroise Duquenne, Holger Schwenk

    cs.CVcs.AIcs.CLarXiv:2603.01096v12026
  6. Sim2Signal: Sim-to-Real Benchmarks for Traffic Signal Control

    Ferdous Al Rafi, Susrik Mukherjee, Latika Liladhar Dekate +6

    cs.LGarXiv:2609.01676v12026
  7. References Improve LLM Alignment in Non-Verifiable Domains

    Kejian Shi, Yixin Liu, Peifeng Wang +3

    cs.CLcs.AIcs.LGarXiv:2602.16802v12026
  8. MemoryLLM: Plug-n-Play Interpretable Feed-Forward Memory for Transformers

    Ajay Jaiswal, Lauren Hannah, Han-Byul Kim +4

    cs.LGarXiv:2602.00398v12026
  9. s1: Simple test-time scaling

    Niklas Muennighoff, Zitong Yang, Weijia Shi +7

    cs.CLcs.AIcs.LGarXiv:2501.19393v32025
  10. Contact-Anchored Policies: Contact Conditioning Creates Strong Robot Utility Models

    Zichen Jeff Cui, Omar Rayyan, Haritheja Etukuru +16

    cs.ROcs.LGarXiv:2602.09017v12026
  11. TTCS: Test-Time Curriculum Synthesis for Self-Evolving

    Chengyi Yang, Zhishang Xiang, Yunbo Tang +5

    cs.LGcs.AIcs.CLarXiv:2601.22628v12026
  12. REASONING GYM: Reasoning Environments for Reinforcement Learning with Verifiable Rewards

    Zafir Stojanovski, Oliver Stanley, Joe Sharratt +4

    cs.LGcs.AIcs.CLarXiv:2505.24760v22025
  13. Multiplex Thinking: Reasoning via Token-wise Branch-and-Merge

    Yao Tang, Li Dong, Yaru Hao +3

    cs.CLcs.AIcs.LGarXiv:2601.08808v12026
  14. Reinforcement Learning with Action Chunking

    Qiyang Li, Zhiyuan Zhou, Sergey Levine

    cs.LGcs.AIcs.ROarXiv:2507.07969v42025
  15. VIBE: Visual Instruction Based Editor

    Grigorii Alekseenko, Aleksandr Gordeev, Irina Tolstykh +7

    cs.CVcs.AIcs.LGarXiv:2601.02242v12026
  16. SampoNLP: A Self-Referential Toolkit for Morphological Analysis of Subword Tokenizers

    Iaroslav Chelombitko, Ekaterina Chelombitko, Aleksey Komissarov

    cs.CLcs.IRcs.LGarXiv:2601.04469v12026
  17. OmniRetarget: Interaction-Preserving Data Generation for Humanoid Whole-Body Loco-Manipulation and Scene Interaction

    Lujie Yang, Xiaoyu Huang, Zhen Wu +6

    cs.ROcs.AIcs.LGarXiv:2509.26633v32025
  18. Learning Smooth and Expressive Interatomic Potentials for Physical Property Prediction

    Xiang Fu, Brandon M. Wood, Luis Barroso-Luque +4

    physics.comp-phcs.LGarXiv:2502.12147v22025
  19. Learning to Act from Actionless Videos through Dense Correspondences

    Po-Chen Ko, Jiayuan Mao, Yilun Du +2

    cs.ROcs.CVcs.LGarXiv:2310.08576v12023
  20. Poseidon: Efficient Foundation Models for PDEs

    Maximilian Herde, Bogdan Raonić, Tobias Rohner +4

    cs.LGarXiv:2405.19101v22024
  21. PaliGemma: A versatile 3B VLM for transfer

    Lucas Beyer, Andreas Steiner, André Susano Pinto +32

    cs.CVcs.AIcs.CLarXiv:2407.07726v22024
  22. DataComp-LM: In search of the next generation of training sets for language models

    Jeffrey Li, Alex Fang, Georgios Smyrnis +56

    cs.LGcs.CLarXiv:2406.11794v42024
  23. Simplified and Generalized Masked Diffusion for Discrete Data

    Jiaxin Shi, Kehang Han, Zhe Wang +2

    cs.LGstat.MLarXiv:2406.04329v42024
  24. ReFT: Representation Finetuning for Language Models

    Zhengxuan Wu, Aryaman Arora, Zheng Wang +4

    cs.CLcs.AIcs.LGarXiv:2404.03592v32024
  25. V-STaR: Training Verifiers for Self-Taught Reasoners

    Arian Hosseini, Xingdi Yuan, Nikolay Malkin +3

    cs.LGcs.AIcs.CLarXiv:2402.06457v22024
  26. Unexpected Improvements to Expected Improvement for Bayesian Optimization

    Sebastian Ament, Samuel Daulton, David Eriksson +2

    cs.LGmath.NAstat.MLarXiv:2310.20708v32023
    Summaries:한국어
  27. Advancements in Generative AI: A Comprehensive Review of GANs, GPT, Autoencoders, Diffusion Model, and Transformers

    Staphord Bengesi, Hoda El-Sayed, Md Kamruzzaman Sarker +3

    cs.LGcs.AIarXiv:2311.10242v22023
  28. Convolutional Recurrent Neural Networks for Small-Footprint Keyword Spotting

    Sercan O. Arik, Markus Kliegl, Rewon Child +5

    cs.CLcs.AIcs.LGarXiv:1703.05390v32017
  29. Being-H0.7: A Latent World-Action Model from Egocentric Videos

    Hao Luo, Wanpeng Zhang, Yicheng Feng +6

    cs.ROcs.CVcs.LGarXiv:2605.00078v12026
  30. iTransformer: Inverted Transformers Are Effective for Time Series Forecasting

    Yong Liu, Tengge Hu, Haoran Zhang +4

    cs.LGarXiv:2310.06625v42023
  31. Large Language Models as Optimizers

    Chengrun Yang, Xuezhi Wang, Yifeng Lu +4

    cs.LGcs.AIcs.CLarXiv:2309.03409v32023
  32. LIMR: Less is More for RL Scaling

    Xuefeng Li, Haoyang Zou, Pengfei Liu

    cs.LGcs.AIcs.CLarXiv:2502.11886v12025
  33. Negative Momentum for Improved Game Dynamics

    Gauthier Gidel, Reyhane Askari Hemmat, Mohammad Pezeshki +4

    cs.LGstat.MLarXiv:1807.04740v52018
  34. A Comprehensive Survey of Deep Transfer Learning for Anomaly Detection in Industrial Time Series: Methods, Applications, and Directions

    Peng Yan, Ahmed Abdulkadir, Paul-Philipp Luley +4

    cs.LGcs.AIarXiv:2307.05638v22023
  35. Language is All a Graph Needs

    Ruosong Ye, Caiqi Zhang, Runhui Wang +2

    cs.CLcs.AIcs.IRarXiv:2308.07134v52023
  36. DragDiffusion: Harnessing Diffusion Models for Interactive Point-based Image Editing

    Yujun Shi, Chuhui Xue, Jun Hao Liew +5

    cs.CVcs.LGarXiv:2306.14435v62023
  37. mPLUG-Owl: Modularization Empowers Large Language Models with Multimodality

    Qinghao Ye, Haiyang Xu, Guohai Xu +15

    cs.CLcs.CVcs.LGarXiv:2304.14178v32023
  38. Interactive and Explainable Region-guided Radiology Report Generation

    Tim Tanida, Philip Müller, Georgios Kaissis +1

    cs.CVcs.CLcs.LGarXiv:2304.08295v12023
  39. GraphPrompt: Unifying Pre-Training and Downstream Tasks for Graph Neural Networks

    Zemin Liu, Xingtong Yu, Yuan Fang +1

    cs.LGcs.CLarXiv:2302.08043v32023
  40. Grounding Large Language Models in Interactive Environments with Online Reinforcement Learning

    Thomas Carta, Clément Romac, Thomas Wolf +3

    cs.LGarXiv:2302.02662v52023
  41. Experimentally realized in situ backpropagation for deep learning in nanophotonic neural networks

    Sunil Pai, Zhanghao Sun, Tyler W. Hughes +11

    cs.ETcs.LGphysics.opticsarXiv:2205.08501v12022
  42. Navigating to Objects in the Real World

    Theophile Gervet, Soumith Chintala, Dhruv Batra +2

    cs.ROcs.CVcs.LGarXiv:2212.00922v12022
  43. Data Augmentation techniques in time series domain: A survey and taxonomy

    Guillermo Iglesias, Edgar Talavera, Ángel González-Prieto +2

    cs.LGcs.AIarXiv:2206.13508v42022
  44. From parcel to continental scale -- A first European crop type map based on Sentinel-1 and LUCAS Copernicus in-situ observations

    Raphaël d'Andrimont, Astrid Verhegghen, Guido Lemoine +3

    stat.MLcs.LGstat.AParXiv:2105.09261v22021
  45. Equivariant message passing for the prediction of tensorial properties and molecular spectra

    Kristof T. Schütt, Oliver T. Unke, Michael Gastegger

    cs.LGphysics.chem-pharXiv:2102.03150v42021
  46. Anomaly Detection in Univariate Time-series: A Survey on the State-of-the-Art

    Mohammad Braei, Sebastian Wagner

    cs.LGstat.MLarXiv:2004.00433v12020
  47. Deep ROC Analysis and AUC as Balanced Average Accuracy to Improve Model Selection, Understanding and Interpretation

    André M. Carrington, Douglas G. Manuel, Paul W. Fieguth +9

    stat.MEcs.AIcs.LGarXiv:2103.11357v12021
  48. Trident: Efficient 4PC Framework for Privacy Preserving Machine Learning

    Harsh Chaudhari, Rahul Rachuri, Ajith Suresh

    cs.LGcs.CRstat.MLarXiv:1912.02631v22019
  49. DL-Droid: Deep learning based android malware detection using real devices

    Mohammed K. Alzaylaee, Suleiman Y. Yerima, Sakir Sezer

    cs.CRcs.LGcs.NEarXiv:1911.10113v12019
  50. PyTorch: An Imperative Style, High-Performance Deep Learning Library

    Adam Paszke, Sam Gross, Francisco Massa +18

    cs.LGcs.MSstat.MLarXiv:1912.01703v12019
    Summaries:한국어
  51. Importance of spatial predictor variable selection in machine learning applications -- Moving from data reproduction to spatial prediction

    Hanna Meyer, Christoph Reudenbach, Stephan Wöllauer +1

    stat.APcs.LGstat.MLarXiv:1908.07805v12019
  52. ALEX: An Updatable Adaptive Learned Index

    Jialin Ding, Umar Farooq Minhas, Jia Yu +9

    cs.DBcs.DScs.LGarXiv:1905.08898v22019
  53. Class-incremental learning: survey and performance evaluation on image classification

    Marc Masana, Xialei Liu, Bartlomiej Twardowski +3

    cs.LGcs.CVarXiv:2010.15277v32020
  54. Regularization Matters: Generalization and Optimization of Neural Nets v.s. their Induced Kernel

    Colin Wei, Jason D. Lee, Qiang Liu +1

    stat.MLcs.LGarXiv:1810.05369v42018
  55. TESSERACT: Eliminating Experimental Bias in Malware Classification across Space and Time

    Feargus Pendlebury, Fabio Pierazzi, Roberto Jordaney +2

    cs.CRcs.LGarXiv:1807.07838v42018
  56. Backdoor Learning: A Survey

    Yiming Li, Yong Jiang, Zhifeng Li +1

    cs.CRcs.CVcs.LGarXiv:2007.08745v52020
  57. Blockchain-Federated-Learning and Deep Learning Models for COVID-19 detection using CT Imaging

    Rajesh Kumar, Abdullah Aman Khan, Sinmin Zhang +7

    eess.IVcs.CVcs.LGarXiv:2007.06537v22020
  58. CoCoA: A General Framework for Communication-Efficient Distributed Optimization

    Virginia Smith, Simone Forte, Chenxin Ma +3

    cs.LGarXiv:1611.02189v22016
  59. ProtTrans: Towards Cracking the Language of Life's Code Through Self-Supervised Deep Learning and High Performance Computing

    Ahmed Elnaggar, Michael Heinzinger, Christian Dallago +9

    cs.LGcs.CLcs.DCarXiv:2007.06225v32020
  60. A Blockchain-based Decentralized Federated Learning Framework with Committee Consensus

    Yuzheng Li, Chuan Chen, Nan Liu +3

    cs.DCcs.LGarXiv:2004.00773v12020