Machine Learning
Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
6,241 to 6,300 of 20,199
Learning to Prove Theorems via Interacting with Proof Assistants
Kaiyu Yang, Jia Deng
cs.LOcs.AIcs.LGarXiv:1905.09381v12019Elite-Weighted Supervised Fine-tuning for Goal-Directed Molecular Optimization
Shiyun Wa, Yifei Wang, Anna G. Green +2
cs.LGarXiv:2609.00189v12026On The Reasons Behind Decisions
Adnan Darwiche, Auguste Hirth
cs.AIcs.LGarXiv:2002.09284v12020OpenMM 8: Molecular Dynamics Simulation with Machine Learning Potentials
Peter Eastman, Raimondas Galvelis, Raúl P. Peláez +22
physics.chem-phcs.LGarXiv:2310.03121v22023Neural 3D Morphable Models: Spiral Convolutional Networks for 3D Shape Representation Learning and Generation
Giorgos Bouritsas, Sergiy Bokhnyak, Stylianos Ploumpis +2
cs.CVcs.AIcs.GRarXiv:1905.02876v32019Compressing DMA Engine: Leveraging Activation Sparsity for Training Deep Neural Networks
Minsoo Rhu, Mike O'Connor, Niladrish Chatterjee +2
cs.LGcs.ARarXiv:1705.01626v12017Hilbert: Recursively Building Formal Proofs with Informal Reasoning
Sumanth Varambally, Thomas Voice, Yanchao Sun +3
cs.AIcs.FLcs.LGarXiv:2509.22819v22025Automated Conjecture Resolution with Formal Verification
Haocheng Ju, Guoxiong Gao, Jiedong Jiang +13
cs.LGcs.AIarXiv:2604.03789v22026Variational Bayesian Optimal Experimental Design
Adam Foster, Martin Jankowiak, Eli Bingham +4
stat.MLcs.LGstat.COarXiv:1903.05480v32019XingGAN for Person Image Generation
Hao Tang, Song Bai, Li Zhang +2
cs.CVcs.LGeess.IVarXiv:2007.09278v12020CHIP: CHannel Independence-based Pruning for Compact Neural Networks
Yang Sui, Miao Yin, Yi Xie +3
cs.CVcs.AIcs.LGarXiv:2110.13981v32021KVzip: Query-Agnostic KV Cache Compression with Context Reconstruction
Jang-Hyun Kim, Jinuk Kim, Sangwoo Kwon +3
cs.DBcs.LGarXiv:2505.23416v22025Revisiting Reinforcement Learning for LLM Reasoning from A Cross-Domain Perspective
Zhoujun Cheng, Shibo Hao, Tianyang Liu +21
cs.LGcs.AIcs.CLarXiv:2506.14965v12025AgentFold: Long-Horizon Web Agents with Proactive Context Management
Rui Ye, Zhongwang Zhang, Kuan Li +12
cs.CLcs.AIcs.LGarXiv:2510.24699v12025How Do Language Models Choose Between Context and Memory?
Benjamin Shih, John Winnicki, Arianna Cao
cs.LGcs.CLarXiv:2609.00753v12026Latent Replay for Real-Time Continual Learning
Lorenzo Pellegrini, Gabriele Graffieti, Vincenzo Lomonaco +1
cs.LGcs.CVstat.MLarXiv:1912.01100v22019TreeRL: LLM Reinforcement Learning with On-Policy Tree Search
Zhenyu Hou, Ziniu Hu, Yujiang Li +3
cs.LGcs.CLarXiv:2506.11902v12025Echo Chamber: RL Post-training Amplifies Behaviors Learned in Pretraining
Rosie Zhao, Alexandru Meterez, Sham Kakade +3
cs.LGarXiv:2504.07912v22025VALOR: Vision-Audio-Language Omni-Perception Pretraining Model and Dataset
Jing Liu, Sihan Chen, Xingjian He +4
cs.LGcs.CLcs.CVarXiv:2304.08345v22023From AI for Science to Agentic Science: A Survey on Autonomous Scientific Discovery
Jiaqi Wei, Yuejin Yang, Xiang Zhang +24
cs.LGarXiv:2508.14111v22025Tail-Likelihood Reinforcement Learning
Shrinivas Ramasubramanian, Daman Arora, Fahim Tajwar +11
cs.LGstat.MLarXiv:2609.02987v12026Rewarding the Unlikely: Lifting GRPO Beyond Distribution Sharpening
Andre He, Daniel Fried, Sean Welleck
cs.LGarXiv:2506.02355v22025Does Fault Localization Beat a Fresh Attempt? A Placebo-Controlled Study of Test-Guided Code Repair
Anik Jha
cs.SEcs.AIcs.LGarXiv:2609.00854v12026Spawn Freely, Act Sparingly: Progressive Risk Vesting for Recursive LLM-Agent Trees
Molly Wang
cs.AIcs.LGmath.PRarXiv:2609.01035v12026Patch Diffusion: Faster and More Data-Efficient Training of Diffusion Models
Zhendong Wang, Yifan Jiang, Huangjie Zheng +5
cs.CVcs.LGarXiv:2304.12526v22023Skillful joint probabilistic weather forecasting from marginals
Ferran Alet, Ilan Price, Andrew El-Kadi +8
cs.LGphysics.ao-pharXiv:2506.10772v12025Automated data processing and feature engineering for deep learning and big data applications: a survey
Alhassan Mumuni, Fuseini Mumuni
cs.LGcs.AIcs.DBarXiv:2403.11395v22024SkillRouter: Skill Routing for LLM Agents at Scale
YanZhao Zheng, ZhenTao Zhang, Chao Ma +8
cs.LGarXiv:2603.22455v52026Adaptive Attacks Break Defenses Against Indirect Prompt Injection Attacks on LLM Agents
Qiusi Zhan, Richard Fang, Henil Shalin Panchal +1
cs.CRcs.LGarXiv:2503.00061v22025Competitive Programming with Large Reasoning Models
OpenAI, :, Ahmed El-Kishky +23
cs.LGcs.AIcs.CLarXiv:2502.06807v22025Complex spectrogram enhancement by convolutional neural network with multi-metrics learning
Szu-Wei Fu, Ting-yao Hu, Yu Tsao +1
stat.MLcs.LGcs.SDarXiv:1704.08504v22017Who Speaks for the Pruned? Visual Token Pruning as Coverage Optimization
Qingchan Zhu, Weihang You, Hanqi Jiang +3
cs.CVcs.CLcs.LGarXiv:2609.03158v12026Accelerating Diffusion LLMs via Adaptive Parallel Decoding
Daniel Israel, Guy Van den Broeck, Aditya Grover
cs.CLcs.AIcs.LGarXiv:2506.00413v22025Challenges in Training PINNs: A Loss Landscape Perspective
Pratik Rathore, Weimu Lei, Zachary Frangella +2
cs.LGmath.OCstat.MLarXiv:2402.01868v22024Latent Matters: Learning Deep State-Space Models
Alexej Klushyn, Richard Kurle, Maximilian Soelch +2
cs.LGarXiv:2602.23050v22026Preference Leakage: A Contamination Problem in LLM-as-a-judge
Dawei Li, Renliang Sun, Yue Huang +6
cs.LGcs.AIcs.CLarXiv:2502.01534v32025Graph Few-shot Learning via Knowledge Transfer
Huaxiu Yao, Chuxu Zhang, Ying Wei +5
cs.LGstat.MLarXiv:1910.03053v32019Inductive Moment Matching
Linqi Zhou, Stefano Ermon, Jiaming Song
cs.LGcs.AIstat.MLarXiv:2503.07565v72025CRAD: Class-wise Reliability-Aware Distillation for Decentralized Heterogeneous Federated Learning
Baraa Bilbeisi, Mengchen Fan, Baocheng Geng +1
cs.LGcs.CVarXiv:2609.00446v12026Cluster-guided Contrastive Graph Clustering Network
Xihong Yang, Yue Liu, Sihang Zhou +6
cs.LGarXiv:2301.01098v12023CWEval: Outcome-driven Evaluation on Functionality and Security of LLM Code Generation
Jinjun Peng, Leyi Cui, Kele Huang +2
cs.SEcs.CLcs.LGarXiv:2501.08200v12025Audio Word2Vec: Unsupervised Learning of Audio Segment Representations using Sequence-to-sequence Autoencoder
Yu-An Chung, Chao-Chung Wu, Chia-Hao Shen +2
cs.SDcs.LGarXiv:1603.00982v42016Accelerated Sampling from Masked Diffusion Models via Entropy Bounded Unmasking
Heli Ben-Hamu, Itai Gat, Daniel Severo +2
cs.LGarXiv:2505.24857v12025NVIDIA Nemotron Nano 2: An Accurate and Efficient Hybrid Mamba-Transformer Reasoning Model
NVIDIA, :, Aarti Basant +214
cs.CLcs.AIcs.LGarXiv:2508.14444v42025PyG 2.0: Scalable Learning on Real World Graphs
Matthias Fey, Jinu Sunil, Akihiro Nitta +10
cs.LGcs.AIarXiv:2507.16991v22025I Know What You Trained Last Summer: A Survey on Stealing Machine Learning Models and Defences
Daryna Oliynyk, Rudolf Mayer, Andreas Rauber
cs.LGcs.AIcs.CRarXiv:2206.08451v22022DLIME: A Deterministic Local Interpretable Model-Agnostic Explanations Approach for Computer-Aided Diagnosis Systems
Muhammad Rehman Zafar, Naimul Mefraz Khan
cs.LGcs.AIstat.MLarXiv:1906.10263v12019Align Your Flow: Scaling Continuous-Time Flow Map Distillation
Amirmojtaba Sabour, Sanja Fidler, Karsten Kreis
cs.CVcs.LGarXiv:2506.14603v12025GraKeL: A Graph Kernel Library in Python
Giannis Siglidis, Giannis Nikolentzos, Stratis Limnios +3
stat.MLcs.LGarXiv:1806.02193v22018Challenges and Countermeasures for Adversarial Attacks on Deep Reinforcement Learning
Inaam Ilahi, Muhammad Usama, Junaid Qadir +4
cs.LGcs.AIcs.CRarXiv:2001.09684v22020Point3R: Streaming 3D Reconstruction with Explicit Spatial Pointer Memory
Yuqi Wu, Wenzhao Zheng, Jie Zhou +1
cs.CVcs.AIcs.LGarXiv:2507.02863v22025Unveiling Causal Reasoning in Large Language Models: Reality or Mirage?
Haoang Chi, He Li, Wenjing Yang +5
cs.AIcs.CLcs.LGarXiv:2506.21215v12025DeepReview: Improving LLM-based Paper Review with Human-like Deep Thinking Process
Minjun Zhu, Yixuan Weng, Linyi Yang +1
cs.CLcs.LGarXiv:2503.08569v12025Norm matters: efficient and accurate normalization schemes in deep networks
Elad Hoffer, Ron Banner, Itay Golan +1
stat.MLcs.LGarXiv:1803.01814v32018A Comprehensive Survey in LLM(-Agent) Full Stack Safety: Data, Training and Deployment
Kun Wang, Guibin Zhang, Zhenhong Zhou +100
cs.CRcs.AIcs.CLarXiv:2504.15585v42025AceReason-Nemotron: Advancing Math and Code Reasoning through Reinforcement Learning
Yang Chen, Zhuolin Yang, Zihan Liu +5
cs.LGcs.AIcs.CLarXiv:2505.16400v32025Probably Approximately Correct MDP Learning and Control With Temporal Logic Constraints
Jie Fu, Ufuk Topcu
eess.SYcs.LGcs.LOarXiv:1404.7073v22014MEt3R: Measuring Multi-View Consistency in Generated Images
Mohammad Asim, Christopher Wewer, Thomas Wimmer +2
cs.CVcs.LGeess.IVarXiv:2501.06336v22025Machine Generated Text: A Comprehensive Survey of Threat Models and Detection Methods
Evan Crothers, Nathalie Japkowicz, Herna Viktor
cs.CLcs.CRcs.CYarXiv:2210.07321v42022Transparency and Explanation in Deep Reinforcement Learning Neural Networks
Rahul Iyer, Yuezhang Li, Huao Li +3
cs.LGstat.MLarXiv:1809.06061v12018