Machine Learning

Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

6,781 to 6,840 of 20,199

  1. xLSTM: Extended Long Short-Term Memory

    Maximilian Beck, Korbinian Pöppel, Markus Spanring +6

    cs.LGcs.AIstat.MLarXiv:2405.04517v22024
  2. LION: Latent Point Diffusion Models for 3D Shape Generation

    Xiaohui Zeng, Arash Vahdat, Francis Williams +4

    cs.CVcs.LGstat.MLarXiv:2210.06978v12022
  3. AnyGPT: Unified Multimodal LLM with Discrete Sequence Modeling

    Jun Zhan, Junqi Dai, Jiasheng Ye +13

    cs.CLcs.AIcs.CVarXiv:2402.12226v52024
  4. Towards Automated Circuit Discovery for Mechanistic Interpretability

    Arthur Conmy, Augustine N. Mavor-Parker, Aengus Lynch +2

    cs.LGarXiv:2304.14997v42023
  5. AlpacaFarm: A Simulation Framework for Methods that Learn from Human Feedback

    Yann Dubois, Xuechen Li, Rohan Taori +6

    cs.LGcs.AIcs.CLarXiv:2305.14387v42023
  6. Deep Semi-Supervised Anomaly Detection

    Lukas Ruff, Robert A. Vandermeulen, Nico Görnitz +4

    cs.LGstat.MLarXiv:1906.02694v22019
  7. Learning Local Equivariant Representations for Large-Scale Atomistic Dynamics

    Albert Musaelian, Simon Batzner, Anders Johansson +4

    physics.comp-phcond-mat.mtrl-scics.LGarXiv:2204.05249v12022
  8. Virtual Adversarial Training: A Regularization Method for Supervised and Semi-Supervised Learning

    Takeru Miyato, Shin-ichi Maeda, Masanori Koyama +1

    stat.MLcs.LGarXiv:1704.03976v22017
  9. Deep Models Under the GAN: Information Leakage from Collaborative Deep Learning

    Briland Hitaj, Giuseppe Ateniese, Fernando Perez-Cruz

    cs.CRcs.LGstat.MLarXiv:1702.07464v32017
  10. Barren plateaus in quantum neural network training landscapes

    Jarrod R. McClean, Sergio Boixo, Vadim N. Smelyanskiy +2

    quant-phcs.LGphysics.chem-pharXiv:1803.11173v12018
  11. Do Deep Nets Really Need to be Deep?

    Lei Jimmy Ba, Rich Caruana

    cs.LGcs.NEarXiv:1312.6184v72013
  12. Loss is its own Reward: Self-Supervision for Reinforcement Learning

    Evan Shelhamer, Parsa Mahmoudieh, Max Argus +1

    cs.LGarXiv:1612.07307v22016
  13. Do LLMs Recognize Your Preferences? Evaluating Personalized Preference Following in LLMs

    Siyan Zhao, Mingyi Hong, Yang Liu +2

    cs.LGcs.CLarXiv:2502.09597v12025
  14. Locked at the Entrance, Open Inside: Where RLVR Narrows the Solution Space

    Qiancheng Zhou, Ruizhe Li

    cs.LGcs.AIcs.CLarXiv:2608.29188v12026
  15. Discovering Hidden Factors of Variation in Deep Networks

    Brian Cheung, Jesse A. Livezey, Arjun K. Bansal +1

    cs.LGcs.CVcs.NEarXiv:1412.6583v42014
  16. Gradients Know What Outcomes Don't: Unlocking Reinforcement Learning for LLM Reasoning with Gradient-Aligned Rewards

    Leqi Zheng, Jinbo Su, Fang Niu +8

    cs.LGarXiv:2609.03342v12026
  17. Machine Teaching: A New Paradigm for Building Machine Learning Systems

    Patrice Y. Simard, Saleema Amershi, David M. Chickering +8

    cs.LGcs.AIcs.HCarXiv:1707.06742v32017
  18. Magma: A Foundation Model for Multimodal AI Agents

    Jianwei Yang, Reuben Tan, Qianhui Wu +10

    cs.CVcs.AIcs.HCarXiv:2502.13130v12025
  19. Adversarial Generation of Continuous Images

    Ivan Skorokhodov, Savva Ignatyev, Mohamed Elhoseiny

    cs.CVcs.AIcs.LGarXiv:2011.12026v22020
  20. MMBERT: Multimodal BERT Pretraining for Improved Medical VQA

    Yash Khare, Viraj Bagal, Minesh Mathew +3

    cs.CVcs.CLcs.LGarXiv:2104.01394v12021
  21. A Peer-Relative Representation Learning Framework for Energy Inefficiency Identification in Mobile Network Sites

    Eliud Nyakweba Koto, Jaco du Toit, Adham Stoltz +1

    cs.LGarXiv:2609.03809v12026
  22. Mixture-of-Recursions: Learning Dynamic Recursive Depths for Adaptive Token-Level Computation

    Sangmin Bae, Yujin Kim, Reza Bayat +8

    cs.CLcs.LGarXiv:2507.10524v32025
  23. Learning Plannable Representations with Causal InfoGAN

    Thanard Kurutach, Aviv Tamar, Ge Yang +2

    cs.LGcs.AIcs.CVarXiv:1807.09341v12018
  24. R2E-Gym: Procedural Environments and Hybrid Verifiers for Scaling Open-Weights SWE Agents

    Naman Jain, Jaskirat Singh, Manish Shetty +3

    cs.SEcs.CLcs.LGarXiv:2504.07164v12025
  25. Agentic Reinforced Policy Optimization

    Guanting Dong, Hangyu Mao, Kai Ma +11

    cs.LGcs.AIcs.CLarXiv:2507.19849v12025
  26. AdaptThink: Reasoning Models Can Learn When to Think

    Jiajie Zhang, Nianyi Lin, Lei Hou +2

    cs.CLcs.AIcs.LGarXiv:2505.13417v12025
  27. The Attacker Moves Second: Stronger Adaptive Attacks Bypass Defenses Against Llm Jailbreaks and Prompt Injections

    Milad Nasr, Nicholas Carlini, Chawin Sitawarin +11

    cs.LGcs.CRarXiv:2510.09023v12025
  28. Country-wide high-resolution vegetation height mapping with Sentinel-2

    Nico Lang, Konrad Schindler, Jan Dirk Wegner

    eess.IVcs.CVcs.LGarXiv:1904.13270v22019
  29. Self-Distilled RLVR

    Chenxu Yang, Chuanyu Qin, Qingyi Si +7

    cs.LGcs.CLarXiv:2604.03128v22026
  30. Deep learning with convolutional neural networks for decoding and visualization of EEG pathology

    Robin Tibor Schirrmeister, Lukas Gemein, Katharina Eggensperger +2

    cs.LGcs.NEstat.MLarXiv:1708.08012v32017
  31. PLAS: Latent Action Space for Offline Reinforcement Learning

    Wenxuan Zhou, Sujay Bajracharya, David Held

    cs.ROcs.AIcs.LGarXiv:2011.07213v12020
  32. Two-Stage Reinforcement Learning for Sound and Adversarial Test Generation in Code LLMs

    Jiacheng Xu, Wentao Zhang, Zhiyi Lyu +4

    cs.CLcs.LGarXiv:2609.03955v12026
  33. Do GANs actually learn the distribution? An empirical study

    Sanjeev Arora, Yi Zhang

    cs.LGarXiv:1706.08224v22017
  34. Deep Feature Space Trojan Attack of Neural Networks by Controlled Detoxification

    Siyuan Cheng, Yingqi Liu, Shiqing Ma +1

    cs.LGcs.CVarXiv:2012.11212v22020
  35. Benchmarking Cognitive Biases in Large Language Models as Evaluators

    Ryan Koo, Minhwa Lee, Vipul Raheja +3

    cs.CLcs.AIcs.LGarXiv:2309.17012v32023
  36. PhysBench: Benchmarking and Enhancing Vision-Language Models for Physical World Understanding

    Wei Chow, Jiageng Mao, Boyi Li +3

    cs.CVcs.AIcs.CLarXiv:2501.16411v22025
  37. Contrastive Behavioral Similarity Embeddings for Generalization in Reinforcement Learning

    Rishabh Agarwal, Marlos C. Machado, Pablo Samuel Castro +1

    cs.LGcs.AIstat.MLarXiv:2101.05265v22021
  38. ICON Decomposition: Multivariate Concept-Level Explanations of Deep Representations for Model Auditing

    Roshan Prakash Rane, Marco Simnacher, Manuel Pfeuffer +7

    cs.LGcs.AIcs.CVarXiv:2608.26083v12026
  39. Lost but not erased: Finding traces of a forgotten language in neural speech models

    Peter Plantinga, Charlotte Moore, Peter W. Donhauser +2

    cs.CLcs.LGarXiv:2608.25976v12026
  40. Which Economic Tasks are Performed with AI? Evidence from Millions of Claude Conversations

    Kunal Handa, Alex Tamkin, Miles McCain +12

    cs.CYcs.AIcs.CLarXiv:2503.04761v12025
  41. Extracting Forgotten Prompts from Targeted Unlearned Models

    Au Ashley Hoi-Ting, Meghdad Kurmanji, William F. Shen +2

    cs.LGarXiv:2609.03662v12026
  42. DE-Venus: A Data-Efficient RLVR Framework for Large Language Models

    Shenzhi Yang, Guangcheng Zhu, Kai Tang +11

    cs.LGarXiv:2609.03324v12026
  43. AI Control: Improving Safety Despite Intentional Subversion

    Ryan Greenblatt, Buck Shlegeris, Kshitij Sachan +1

    cs.LGarXiv:2312.06942v52023
  44. FlexPrefill: A Context-Aware Sparse Attention Mechanism for Efficient Long-Sequence Inference

    Xunhao Lai, Jianqiao Lu, Yao Luo +2

    cs.LGcs.CLarXiv:2502.20766v12025
  45. Speck: A Smart event-based Vision Sensor with a low latency 327K Neuron Convolutional Neuronal Network Processing Pipeline

    Ole Richter, Yannan Xing, Michele De Marchi +9

    cs.NEcs.LGeess.IVarXiv:2304.06793v22023
  46. Knowledge Base Completion: Baselines Strike Back

    Rudolf Kadlec, Ondrej Bajgar, Jan Kleindienst

    cs.LGcs.AIarXiv:1705.10744v12017
  47. Multi-Agent Risks from Advanced AI

    Lewis Hammond, Alan Chan, Jesse Clifton +41

    cs.MAcs.AIcs.CYarXiv:2502.14143v12025
  48. DAST: Difficulty-Adaptive Slow-Thinking for Large Reasoning Models

    Yi Shen, Jian Zhang, Jieyun Huang +7

    cs.LGcs.AIarXiv:2503.04472v32025
  49. Functional Adversarial Attacks

    Cassidy Laidlaw, Soheil Feizi

    cs.LGcs.CVarXiv:1906.00001v22019
  50. Goedel-Prover: A Frontier Model for Open-Source Automated Theorem Proving

    Yong Lin, Shange Tang, Bohan Lyu +8

    cs.LGcs.AIarXiv:2502.07640v32025
  51. Deep Neural Network for Respiratory Sound Classification in Wearable Devices Enabled by Patient Specific Model Tuning

    Jyotibdha Acharya, Arindam Basu

    eess.AScs.LGcs.SDarXiv:2004.08287v12020
  52. ShieldAgent: Shielding Agents via Verifiable Safety Policy Reasoning

    Zhaorun Chen, Mintong Kang, Bo Li

    cs.LGcs.CRarXiv:2503.22738v22025
  53. Knowledge Insulating Vision-Language-Action Models: Train Fast, Run Fast, Generalize Better

    Danny Driess, Jost Tobias Springenberg, Brian Ichter +8

    cs.LGcs.ROarXiv:2505.23705v12025
  54. Multi-Source Deep Domain Adaptation with Weak Supervision for Time-Series Sensor Data

    Garrett Wilson, Janardhan Rao Doppa, Diane J. Cook

    cs.LGstat.MLarXiv:2005.10996v12020
  55. Re-IQA: Unsupervised Learning for Image Quality Assessment in the Wild

    Avinab Saha, Sandeep Mishra, Alan C. Bovik

    cs.CVcs.LGcs.MMarXiv:2304.00451v22023
  56. Skywork Open Reasoner 1 Technical Report

    Jujie He, Jiacai Liu, Chris Yuhao Liu +14

    cs.LGcs.AIcs.CLarXiv:2505.22312v22025
  57. Score Approximation, Estimation and Distribution Recovery of Diffusion Models on Low-Dimensional Data

    Minshuo Chen, Kaixuan Huang, Tuo Zhao +1

    cs.LGstat.MLarXiv:2302.07194v12023
  58. Certified Defenses for Adversarial Patches

    Ping-Yeh Chiang, Renkun Ni, Ahmed Abdelkader +3

    cs.CRcs.LGstat.MLarXiv:2003.06693v22020
  59. A mathematical perspective on Transformers

    Borjan Geshkovski, Cyril Letrouit, Yury Polyanskiy +1

    cs.LGmath.APmath.DSarXiv:2312.10794v52023
  60. Implicit Gradient Regularization

    David G. T. Barrett, Benoit Dherin

    cs.LGstat.MLarXiv:2009.11162v32020