Machine Learning
Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
6,781 to 6,840 of 20,199
xLSTM: Extended Long Short-Term Memory
Maximilian Beck, Korbinian Pöppel, Markus Spanring +6
cs.LGcs.AIstat.MLarXiv:2405.04517v22024LION: Latent Point Diffusion Models for 3D Shape Generation
Xiaohui Zeng, Arash Vahdat, Francis Williams +4
cs.CVcs.LGstat.MLarXiv:2210.06978v12022AnyGPT: Unified Multimodal LLM with Discrete Sequence Modeling
Jun Zhan, Junqi Dai, Jiasheng Ye +13
cs.CLcs.AIcs.CVarXiv:2402.12226v52024Towards Automated Circuit Discovery for Mechanistic Interpretability
Arthur Conmy, Augustine N. Mavor-Parker, Aengus Lynch +2
cs.LGarXiv:2304.14997v42023AlpacaFarm: A Simulation Framework for Methods that Learn from Human Feedback
Yann Dubois, Xuechen Li, Rohan Taori +6
cs.LGcs.AIcs.CLarXiv:2305.14387v42023Deep Semi-Supervised Anomaly Detection
Lukas Ruff, Robert A. Vandermeulen, Nico Görnitz +4
cs.LGstat.MLarXiv:1906.02694v22019Learning Local Equivariant Representations for Large-Scale Atomistic Dynamics
Albert Musaelian, Simon Batzner, Anders Johansson +4
physics.comp-phcond-mat.mtrl-scics.LGarXiv:2204.05249v12022Virtual Adversarial Training: A Regularization Method for Supervised and Semi-Supervised Learning
Takeru Miyato, Shin-ichi Maeda, Masanori Koyama +1
stat.MLcs.LGarXiv:1704.03976v22017Deep Models Under the GAN: Information Leakage from Collaborative Deep Learning
Briland Hitaj, Giuseppe Ateniese, Fernando Perez-Cruz
cs.CRcs.LGstat.MLarXiv:1702.07464v32017Barren plateaus in quantum neural network training landscapes
Jarrod R. McClean, Sergio Boixo, Vadim N. Smelyanskiy +2
quant-phcs.LGphysics.chem-pharXiv:1803.11173v12018Do Deep Nets Really Need to be Deep?
Lei Jimmy Ba, Rich Caruana
cs.LGcs.NEarXiv:1312.6184v72013Loss is its own Reward: Self-Supervision for Reinforcement Learning
Evan Shelhamer, Parsa Mahmoudieh, Max Argus +1
cs.LGarXiv:1612.07307v22016Do LLMs Recognize Your Preferences? Evaluating Personalized Preference Following in LLMs
Siyan Zhao, Mingyi Hong, Yang Liu +2
cs.LGcs.CLarXiv:2502.09597v12025Locked at the Entrance, Open Inside: Where RLVR Narrows the Solution Space
Qiancheng Zhou, Ruizhe Li
cs.LGcs.AIcs.CLarXiv:2608.29188v12026Discovering Hidden Factors of Variation in Deep Networks
Brian Cheung, Jesse A. Livezey, Arjun K. Bansal +1
cs.LGcs.CVcs.NEarXiv:1412.6583v42014Gradients Know What Outcomes Don't: Unlocking Reinforcement Learning for LLM Reasoning with Gradient-Aligned Rewards
Leqi Zheng, Jinbo Su, Fang Niu +8
cs.LGarXiv:2609.03342v12026Machine Teaching: A New Paradigm for Building Machine Learning Systems
Patrice Y. Simard, Saleema Amershi, David M. Chickering +8
cs.LGcs.AIcs.HCarXiv:1707.06742v32017Magma: A Foundation Model for Multimodal AI Agents
Jianwei Yang, Reuben Tan, Qianhui Wu +10
cs.CVcs.AIcs.HCarXiv:2502.13130v12025Adversarial Generation of Continuous Images
Ivan Skorokhodov, Savva Ignatyev, Mohamed Elhoseiny
cs.CVcs.AIcs.LGarXiv:2011.12026v22020MMBERT: Multimodal BERT Pretraining for Improved Medical VQA
Yash Khare, Viraj Bagal, Minesh Mathew +3
cs.CVcs.CLcs.LGarXiv:2104.01394v12021A Peer-Relative Representation Learning Framework for Energy Inefficiency Identification in Mobile Network Sites
Eliud Nyakweba Koto, Jaco du Toit, Adham Stoltz +1
cs.LGarXiv:2609.03809v12026Mixture-of-Recursions: Learning Dynamic Recursive Depths for Adaptive Token-Level Computation
Sangmin Bae, Yujin Kim, Reza Bayat +8
cs.CLcs.LGarXiv:2507.10524v32025Learning Plannable Representations with Causal InfoGAN
Thanard Kurutach, Aviv Tamar, Ge Yang +2
cs.LGcs.AIcs.CVarXiv:1807.09341v12018R2E-Gym: Procedural Environments and Hybrid Verifiers for Scaling Open-Weights SWE Agents
Naman Jain, Jaskirat Singh, Manish Shetty +3
cs.SEcs.CLcs.LGarXiv:2504.07164v12025Agentic Reinforced Policy Optimization
Guanting Dong, Hangyu Mao, Kai Ma +11
cs.LGcs.AIcs.CLarXiv:2507.19849v12025AdaptThink: Reasoning Models Can Learn When to Think
Jiajie Zhang, Nianyi Lin, Lei Hou +2
cs.CLcs.AIcs.LGarXiv:2505.13417v12025The Attacker Moves Second: Stronger Adaptive Attacks Bypass Defenses Against Llm Jailbreaks and Prompt Injections
Milad Nasr, Nicholas Carlini, Chawin Sitawarin +11
cs.LGcs.CRarXiv:2510.09023v12025Country-wide high-resolution vegetation height mapping with Sentinel-2
Nico Lang, Konrad Schindler, Jan Dirk Wegner
eess.IVcs.CVcs.LGarXiv:1904.13270v22019Self-Distilled RLVR
Chenxu Yang, Chuanyu Qin, Qingyi Si +7
cs.LGcs.CLarXiv:2604.03128v22026Deep learning with convolutional neural networks for decoding and visualization of EEG pathology
Robin Tibor Schirrmeister, Lukas Gemein, Katharina Eggensperger +2
cs.LGcs.NEstat.MLarXiv:1708.08012v32017PLAS: Latent Action Space for Offline Reinforcement Learning
Wenxuan Zhou, Sujay Bajracharya, David Held
cs.ROcs.AIcs.LGarXiv:2011.07213v12020Two-Stage Reinforcement Learning for Sound and Adversarial Test Generation in Code LLMs
Jiacheng Xu, Wentao Zhang, Zhiyi Lyu +4
cs.CLcs.LGarXiv:2609.03955v12026Do GANs actually learn the distribution? An empirical study
Sanjeev Arora, Yi Zhang
cs.LGarXiv:1706.08224v22017Deep Feature Space Trojan Attack of Neural Networks by Controlled Detoxification
Siyuan Cheng, Yingqi Liu, Shiqing Ma +1
cs.LGcs.CVarXiv:2012.11212v22020Benchmarking Cognitive Biases in Large Language Models as Evaluators
Ryan Koo, Minhwa Lee, Vipul Raheja +3
cs.CLcs.AIcs.LGarXiv:2309.17012v32023PhysBench: Benchmarking and Enhancing Vision-Language Models for Physical World Understanding
Wei Chow, Jiageng Mao, Boyi Li +3
cs.CVcs.AIcs.CLarXiv:2501.16411v22025Contrastive Behavioral Similarity Embeddings for Generalization in Reinforcement Learning
Rishabh Agarwal, Marlos C. Machado, Pablo Samuel Castro +1
cs.LGcs.AIstat.MLarXiv:2101.05265v22021ICON Decomposition: Multivariate Concept-Level Explanations of Deep Representations for Model Auditing
Roshan Prakash Rane, Marco Simnacher, Manuel Pfeuffer +7
cs.LGcs.AIcs.CVarXiv:2608.26083v12026Lost but not erased: Finding traces of a forgotten language in neural speech models
Peter Plantinga, Charlotte Moore, Peter W. Donhauser +2
cs.CLcs.LGarXiv:2608.25976v12026Which Economic Tasks are Performed with AI? Evidence from Millions of Claude Conversations
Kunal Handa, Alex Tamkin, Miles McCain +12
cs.CYcs.AIcs.CLarXiv:2503.04761v12025Extracting Forgotten Prompts from Targeted Unlearned Models
Au Ashley Hoi-Ting, Meghdad Kurmanji, William F. Shen +2
cs.LGarXiv:2609.03662v12026DE-Venus: A Data-Efficient RLVR Framework for Large Language Models
Shenzhi Yang, Guangcheng Zhu, Kai Tang +11
cs.LGarXiv:2609.03324v12026AI Control: Improving Safety Despite Intentional Subversion
Ryan Greenblatt, Buck Shlegeris, Kshitij Sachan +1
cs.LGarXiv:2312.06942v52023FlexPrefill: A Context-Aware Sparse Attention Mechanism for Efficient Long-Sequence Inference
Xunhao Lai, Jianqiao Lu, Yao Luo +2
cs.LGcs.CLarXiv:2502.20766v12025Speck: A Smart event-based Vision Sensor with a low latency 327K Neuron Convolutional Neuronal Network Processing Pipeline
Ole Richter, Yannan Xing, Michele De Marchi +9
cs.NEcs.LGeess.IVarXiv:2304.06793v22023Knowledge Base Completion: Baselines Strike Back
Rudolf Kadlec, Ondrej Bajgar, Jan Kleindienst
cs.LGcs.AIarXiv:1705.10744v12017Multi-Agent Risks from Advanced AI
Lewis Hammond, Alan Chan, Jesse Clifton +41
cs.MAcs.AIcs.CYarXiv:2502.14143v12025DAST: Difficulty-Adaptive Slow-Thinking for Large Reasoning Models
Yi Shen, Jian Zhang, Jieyun Huang +7
cs.LGcs.AIarXiv:2503.04472v32025Functional Adversarial Attacks
Cassidy Laidlaw, Soheil Feizi
cs.LGcs.CVarXiv:1906.00001v22019Goedel-Prover: A Frontier Model for Open-Source Automated Theorem Proving
Yong Lin, Shange Tang, Bohan Lyu +8
cs.LGcs.AIarXiv:2502.07640v32025Deep Neural Network for Respiratory Sound Classification in Wearable Devices Enabled by Patient Specific Model Tuning
Jyotibdha Acharya, Arindam Basu
eess.AScs.LGcs.SDarXiv:2004.08287v12020ShieldAgent: Shielding Agents via Verifiable Safety Policy Reasoning
Zhaorun Chen, Mintong Kang, Bo Li
cs.LGcs.CRarXiv:2503.22738v22025Knowledge Insulating Vision-Language-Action Models: Train Fast, Run Fast, Generalize Better
Danny Driess, Jost Tobias Springenberg, Brian Ichter +8
cs.LGcs.ROarXiv:2505.23705v12025Multi-Source Deep Domain Adaptation with Weak Supervision for Time-Series Sensor Data
Garrett Wilson, Janardhan Rao Doppa, Diane J. Cook
cs.LGstat.MLarXiv:2005.10996v12020Re-IQA: Unsupervised Learning for Image Quality Assessment in the Wild
Avinab Saha, Sandeep Mishra, Alan C. Bovik
cs.CVcs.LGcs.MMarXiv:2304.00451v22023Skywork Open Reasoner 1 Technical Report
Jujie He, Jiacai Liu, Chris Yuhao Liu +14
cs.LGcs.AIcs.CLarXiv:2505.22312v22025Score Approximation, Estimation and Distribution Recovery of Diffusion Models on Low-Dimensional Data
Minshuo Chen, Kaixuan Huang, Tuo Zhao +1
cs.LGstat.MLarXiv:2302.07194v12023Certified Defenses for Adversarial Patches
Ping-Yeh Chiang, Renkun Ni, Ahmed Abdelkader +3
cs.CRcs.LGstat.MLarXiv:2003.06693v22020A mathematical perspective on Transformers
Borjan Geshkovski, Cyril Letrouit, Yury Polyanskiy +1
cs.LGmath.APmath.DSarXiv:2312.10794v52023Implicit Gradient Regularization
David G. T. Barrett, Benoit Dherin
cs.LGstat.MLarXiv:2009.11162v32020