Machine Learning

Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

6,241 to 6,300 of 20,199

  1. Learning to Prove Theorems via Interacting with Proof Assistants

    Kaiyu Yang, Jia Deng

    cs.LOcs.AIcs.LGarXiv:1905.09381v12019
  2. Elite-Weighted Supervised Fine-tuning for Goal-Directed Molecular Optimization

    Shiyun Wa, Yifei Wang, Anna G. Green +2

    cs.LGarXiv:2609.00189v12026
  3. On The Reasons Behind Decisions

    Adnan Darwiche, Auguste Hirth

    cs.AIcs.LGarXiv:2002.09284v12020
  4. OpenMM 8: Molecular Dynamics Simulation with Machine Learning Potentials

    Peter Eastman, Raimondas Galvelis, Raúl P. Peláez +22

    physics.chem-phcs.LGarXiv:2310.03121v22023
  5. Neural 3D Morphable Models: Spiral Convolutional Networks for 3D Shape Representation Learning and Generation

    Giorgos Bouritsas, Sergiy Bokhnyak, Stylianos Ploumpis +2

    cs.CVcs.AIcs.GRarXiv:1905.02876v32019
  6. Compressing DMA Engine: Leveraging Activation Sparsity for Training Deep Neural Networks

    Minsoo Rhu, Mike O'Connor, Niladrish Chatterjee +2

    cs.LGcs.ARarXiv:1705.01626v12017
  7. Hilbert: Recursively Building Formal Proofs with Informal Reasoning

    Sumanth Varambally, Thomas Voice, Yanchao Sun +3

    cs.AIcs.FLcs.LGarXiv:2509.22819v22025
  8. Automated Conjecture Resolution with Formal Verification

    Haocheng Ju, Guoxiong Gao, Jiedong Jiang +13

    cs.LGcs.AIarXiv:2604.03789v22026
  9. Variational Bayesian Optimal Experimental Design

    Adam Foster, Martin Jankowiak, Eli Bingham +4

    stat.MLcs.LGstat.COarXiv:1903.05480v32019
  10. XingGAN for Person Image Generation

    Hao Tang, Song Bai, Li Zhang +2

    cs.CVcs.LGeess.IVarXiv:2007.09278v12020
  11. CHIP: CHannel Independence-based Pruning for Compact Neural Networks

    Yang Sui, Miao Yin, Yi Xie +3

    cs.CVcs.AIcs.LGarXiv:2110.13981v32021
  12. KVzip: Query-Agnostic KV Cache Compression with Context Reconstruction

    Jang-Hyun Kim, Jinuk Kim, Sangwoo Kwon +3

    cs.DBcs.LGarXiv:2505.23416v22025
  13. Revisiting Reinforcement Learning for LLM Reasoning from A Cross-Domain Perspective

    Zhoujun Cheng, Shibo Hao, Tianyang Liu +21

    cs.LGcs.AIcs.CLarXiv:2506.14965v12025
  14. AgentFold: Long-Horizon Web Agents with Proactive Context Management

    Rui Ye, Zhongwang Zhang, Kuan Li +12

    cs.CLcs.AIcs.LGarXiv:2510.24699v12025
  15. How Do Language Models Choose Between Context and Memory?

    Benjamin Shih, John Winnicki, Arianna Cao

    cs.LGcs.CLarXiv:2609.00753v12026
  16. Latent Replay for Real-Time Continual Learning

    Lorenzo Pellegrini, Gabriele Graffieti, Vincenzo Lomonaco +1

    cs.LGcs.CVstat.MLarXiv:1912.01100v22019
  17. TreeRL: LLM Reinforcement Learning with On-Policy Tree Search

    Zhenyu Hou, Ziniu Hu, Yujiang Li +3

    cs.LGcs.CLarXiv:2506.11902v12025
  18. Echo Chamber: RL Post-training Amplifies Behaviors Learned in Pretraining

    Rosie Zhao, Alexandru Meterez, Sham Kakade +3

    cs.LGarXiv:2504.07912v22025
  19. VALOR: Vision-Audio-Language Omni-Perception Pretraining Model and Dataset

    Jing Liu, Sihan Chen, Xingjian He +4

    cs.LGcs.CLcs.CVarXiv:2304.08345v22023
  20. From AI for Science to Agentic Science: A Survey on Autonomous Scientific Discovery

    Jiaqi Wei, Yuejin Yang, Xiang Zhang +24

    cs.LGarXiv:2508.14111v22025
  21. Tail-Likelihood Reinforcement Learning

    Shrinivas Ramasubramanian, Daman Arora, Fahim Tajwar +11

    cs.LGstat.MLarXiv:2609.02987v12026
  22. Rewarding the Unlikely: Lifting GRPO Beyond Distribution Sharpening

    Andre He, Daniel Fried, Sean Welleck

    cs.LGarXiv:2506.02355v22025
  23. Does Fault Localization Beat a Fresh Attempt? A Placebo-Controlled Study of Test-Guided Code Repair

    Anik Jha

    cs.SEcs.AIcs.LGarXiv:2609.00854v12026
  24. Spawn Freely, Act Sparingly: Progressive Risk Vesting for Recursive LLM-Agent Trees

    Molly Wang

    cs.AIcs.LGmath.PRarXiv:2609.01035v12026
  25. Patch Diffusion: Faster and More Data-Efficient Training of Diffusion Models

    Zhendong Wang, Yifan Jiang, Huangjie Zheng +5

    cs.CVcs.LGarXiv:2304.12526v22023
  26. Skillful joint probabilistic weather forecasting from marginals

    Ferran Alet, Ilan Price, Andrew El-Kadi +8

    cs.LGphysics.ao-pharXiv:2506.10772v12025
  27. Automated data processing and feature engineering for deep learning and big data applications: a survey

    Alhassan Mumuni, Fuseini Mumuni

    cs.LGcs.AIcs.DBarXiv:2403.11395v22024
  28. SkillRouter: Skill Routing for LLM Agents at Scale

    YanZhao Zheng, ZhenTao Zhang, Chao Ma +8

    cs.LGarXiv:2603.22455v52026
  29. Adaptive Attacks Break Defenses Against Indirect Prompt Injection Attacks on LLM Agents

    Qiusi Zhan, Richard Fang, Henil Shalin Panchal +1

    cs.CRcs.LGarXiv:2503.00061v22025
  30. Competitive Programming with Large Reasoning Models

    OpenAI, :, Ahmed El-Kishky +23

    cs.LGcs.AIcs.CLarXiv:2502.06807v22025
  31. Complex spectrogram enhancement by convolutional neural network with multi-metrics learning

    Szu-Wei Fu, Ting-yao Hu, Yu Tsao +1

    stat.MLcs.LGcs.SDarXiv:1704.08504v22017
  32. Who Speaks for the Pruned? Visual Token Pruning as Coverage Optimization

    Qingchan Zhu, Weihang You, Hanqi Jiang +3

    cs.CVcs.CLcs.LGarXiv:2609.03158v12026
  33. Accelerating Diffusion LLMs via Adaptive Parallel Decoding

    Daniel Israel, Guy Van den Broeck, Aditya Grover

    cs.CLcs.AIcs.LGarXiv:2506.00413v22025
  34. Challenges in Training PINNs: A Loss Landscape Perspective

    Pratik Rathore, Weimu Lei, Zachary Frangella +2

    cs.LGmath.OCstat.MLarXiv:2402.01868v22024
  35. Latent Matters: Learning Deep State-Space Models

    Alexej Klushyn, Richard Kurle, Maximilian Soelch +2

    cs.LGarXiv:2602.23050v22026
  36. Preference Leakage: A Contamination Problem in LLM-as-a-judge

    Dawei Li, Renliang Sun, Yue Huang +6

    cs.LGcs.AIcs.CLarXiv:2502.01534v32025
  37. Graph Few-shot Learning via Knowledge Transfer

    Huaxiu Yao, Chuxu Zhang, Ying Wei +5

    cs.LGstat.MLarXiv:1910.03053v32019
  38. Inductive Moment Matching

    Linqi Zhou, Stefano Ermon, Jiaming Song

    cs.LGcs.AIstat.MLarXiv:2503.07565v72025
  39. CRAD: Class-wise Reliability-Aware Distillation for Decentralized Heterogeneous Federated Learning

    Baraa Bilbeisi, Mengchen Fan, Baocheng Geng +1

    cs.LGcs.CVarXiv:2609.00446v12026
  40. Cluster-guided Contrastive Graph Clustering Network

    Xihong Yang, Yue Liu, Sihang Zhou +6

    cs.LGarXiv:2301.01098v12023
  41. CWEval: Outcome-driven Evaluation on Functionality and Security of LLM Code Generation

    Jinjun Peng, Leyi Cui, Kele Huang +2

    cs.SEcs.CLcs.LGarXiv:2501.08200v12025
  42. Audio Word2Vec: Unsupervised Learning of Audio Segment Representations using Sequence-to-sequence Autoencoder

    Yu-An Chung, Chao-Chung Wu, Chia-Hao Shen +2

    cs.SDcs.LGarXiv:1603.00982v42016
  43. Accelerated Sampling from Masked Diffusion Models via Entropy Bounded Unmasking

    Heli Ben-Hamu, Itai Gat, Daniel Severo +2

    cs.LGarXiv:2505.24857v12025
  44. NVIDIA Nemotron Nano 2: An Accurate and Efficient Hybrid Mamba-Transformer Reasoning Model

    NVIDIA, :, Aarti Basant +214

    cs.CLcs.AIcs.LGarXiv:2508.14444v42025
  45. PyG 2.0: Scalable Learning on Real World Graphs

    Matthias Fey, Jinu Sunil, Akihiro Nitta +10

    cs.LGcs.AIarXiv:2507.16991v22025
  46. I Know What You Trained Last Summer: A Survey on Stealing Machine Learning Models and Defences

    Daryna Oliynyk, Rudolf Mayer, Andreas Rauber

    cs.LGcs.AIcs.CRarXiv:2206.08451v22022
  47. DLIME: A Deterministic Local Interpretable Model-Agnostic Explanations Approach for Computer-Aided Diagnosis Systems

    Muhammad Rehman Zafar, Naimul Mefraz Khan

    cs.LGcs.AIstat.MLarXiv:1906.10263v12019
  48. Align Your Flow: Scaling Continuous-Time Flow Map Distillation

    Amirmojtaba Sabour, Sanja Fidler, Karsten Kreis

    cs.CVcs.LGarXiv:2506.14603v12025
  49. GraKeL: A Graph Kernel Library in Python

    Giannis Siglidis, Giannis Nikolentzos, Stratis Limnios +3

    stat.MLcs.LGarXiv:1806.02193v22018
  50. Challenges and Countermeasures for Adversarial Attacks on Deep Reinforcement Learning

    Inaam Ilahi, Muhammad Usama, Junaid Qadir +4

    cs.LGcs.AIcs.CRarXiv:2001.09684v22020
  51. Point3R: Streaming 3D Reconstruction with Explicit Spatial Pointer Memory

    Yuqi Wu, Wenzhao Zheng, Jie Zhou +1

    cs.CVcs.AIcs.LGarXiv:2507.02863v22025
  52. Unveiling Causal Reasoning in Large Language Models: Reality or Mirage?

    Haoang Chi, He Li, Wenjing Yang +5

    cs.AIcs.CLcs.LGarXiv:2506.21215v12025
  53. DeepReview: Improving LLM-based Paper Review with Human-like Deep Thinking Process

    Minjun Zhu, Yixuan Weng, Linyi Yang +1

    cs.CLcs.LGarXiv:2503.08569v12025
  54. Norm matters: efficient and accurate normalization schemes in deep networks

    Elad Hoffer, Ron Banner, Itay Golan +1

    stat.MLcs.LGarXiv:1803.01814v32018
  55. A Comprehensive Survey in LLM(-Agent) Full Stack Safety: Data, Training and Deployment

    Kun Wang, Guibin Zhang, Zhenhong Zhou +100

    cs.CRcs.AIcs.CLarXiv:2504.15585v42025
  56. AceReason-Nemotron: Advancing Math and Code Reasoning through Reinforcement Learning

    Yang Chen, Zhuolin Yang, Zihan Liu +5

    cs.LGcs.AIcs.CLarXiv:2505.16400v32025
  57. Probably Approximately Correct MDP Learning and Control With Temporal Logic Constraints

    Jie Fu, Ufuk Topcu

    eess.SYcs.LGcs.LOarXiv:1404.7073v22014
  58. MEt3R: Measuring Multi-View Consistency in Generated Images

    Mohammad Asim, Christopher Wewer, Thomas Wimmer +2

    cs.CVcs.LGeess.IVarXiv:2501.06336v22025
  59. Machine Generated Text: A Comprehensive Survey of Threat Models and Detection Methods

    Evan Crothers, Nathalie Japkowicz, Herna Viktor

    cs.CLcs.CRcs.CYarXiv:2210.07321v42022
  60. Transparency and Explanation in Deep Reinforcement Learning Neural Networks

    Rahul Iyer, Yuezhang Li, Huao Li +3

    cs.LGstat.MLarXiv:1809.06061v12018