Machine Learning
Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
841 to 900 of 20,178
Review of Swarm Intelligence-based Feature Selection Methods
Mehrdad Rostami, Kamal Berahmand, Saman Forouzandeh
cs.LGcs.NEstat.MLarXiv:2008.04103v12020Robust Adversarial Reinforcement Learning
Lerrel Pinto, James Davidson, Rahul Sukthankar +1
cs.LGcs.AIcs.MAarXiv:1703.02702v12017Aligning AI With Shared Human Values
Dan Hendrycks, Collin Burns, Steven Basart +4
cs.CYcs.AIcs.CLarXiv:2008.02275v62020An Analytical Formula of Population Gradient for two-layered ReLU network and its Applications in Convergence and Critical Point Analysis
Yuandong Tian
cs.LGarXiv:1703.00560v22017Segmentation of optic disc, fovea and retinal vasculature using a single convolutional neural network
Jen Hong Tan, U. Rajendra Acharya, Sulatha V. Bhandary +2
cs.CVcs.LGarXiv:1702.00509v12017Learn&Fuzz: Machine Learning for Input Fuzzing
Patrice Godefroid, Hila Peleg, Rishabh Singh
cs.AIcs.CRcs.LGarXiv:1701.07232v12017GenMol: A Drug Discovery Generalist with Discrete Diffusion
Seul Lee, Karsten Kreis, Srimukh Prasad Veccham +6
cs.LGarXiv:2501.06158v32025Test-time Alignment of Diffusion Models without Reward Over-optimization
Sunwoo Kim, Minkyu Kim, Dongmin Park
cs.LGcs.AIcs.CVarXiv:2501.05803v32025An Explainable Machine Learning Framework for Predicting Blood-Brain Barrier Permeability Using Molecular Descriptors
Fatemeh Mahmoudi
cs.LGcond-mat.mtrl-sciarXiv:2609.10012v12026Infra-Bench CLS: A Global, Open-Source Benchmark for Critical Infrastructure Classification with Earth Observation Foundation Models
Justin Guthrie, Edward Oughton, Konrad Wessels +2
cs.CVcs.LGarXiv:2609.09482v12026A Survey on Large Language Models with some Insights on their Capabilities and Limitations
Andrea Matarazzo, Riccardo Torlone
cs.CLcs.AIcs.LGarXiv:2501.04040v22025Beyond Skip Connections: Top-Down Modulation for Object Detection
Abhinav Shrivastava, Rahul Sukthankar, Jitendra Malik +1
cs.CVcs.LGarXiv:1612.06851v22016Predicting Process Behaviour using Deep Learning
Joerg Evermann, Jana-Rebecca Rehse, Peter Fettke
cs.LGstat.MLarXiv:1612.04600v22016Training-Free Task Vectors for LLM Behavioral Control
Gabriel J. Perin, Lucas Boscaini, André Araujo +1
cs.LGcs.AIarXiv:2609.09054v12026Backdoor Attacks Against Deep Learning Systems in the Physical World
Emily Wenger, Josephine Passananti, Arjun Bhagoji +3
cs.CVcs.CRcs.LGarXiv:2006.14580v42020The BrowserGym Ecosystem for Web Agent Research
Thibault Le Sellier De Chezelles, Maxime Gasse, Alexandre Drouin +17
cs.LGcs.AIcs.SEarXiv:2412.05467v42024Pyramidal Convolution: Rethinking Convolutional Neural Networks for Visual Recognition
Ionut Cosmin Duta, Li Liu, Fan Zhu +1
cs.CVcs.LGeess.IVarXiv:2006.11538v12020A causal framework for discovering and removing direct and indirect discrimination
Lu Zhang, Yongkai Wu, Xintao Wu
cs.LGarXiv:1611.07509v12016wav2vec 2.0: A Framework for Self-Supervised Learning of Speech Representations
Alexei Baevski, Henry Zhou, Abdelrahman Mohamed +1
cs.CLcs.LGcs.SDarXiv:2006.11477v32020AWAC: Accelerating Online Reinforcement Learning with Offline Datasets
Ashvin Nair, Abhishek Gupta, Murtaza Dalal +1
cs.LGcs.ROstat.MLarXiv:2006.09359v62020SUN: Reaching for Novelty in Reinforcement Learning
Wenyan Yang, Arsenii Mustafin, Dominik Baumann +2
cs.LGcs.AIcs.ROarXiv:2609.08642v12026Foundations of Structural Causal Models with Cycles and Latent Variables
Stephan Bongers, Patrick Forré, Jonas Peters +1
stat.MEcs.AIcs.LGarXiv:1611.06221v62016Robust Learning Through Cross-Task Consistency
Amir Zamir, Alexander Sax, Teresa Yeo +6
cs.CVcs.GRcs.LGarXiv:2006.04096v12020SaliencyMix: A Saliency Guided Data Augmentation Strategy for Better Regularization
A. F. M. Shahab Uddin, Mst. Sirazam Monira, Wheemyung Shin +2
cs.LGstat.MLarXiv:2006.01791v22020Embodied Agent Interface: Benchmarking LLMs for Embodied Decision Making
Manling Li, Shiyu Zhao, Qineng Wang +12
cs.CLcs.AIcs.LGarXiv:2410.07166v32024SWIFT: Super-fast and Robust Privacy-Preserving Machine Learning
Nishat Koti, Mahak Pancholi, Arpita Patra +1
cs.CRcs.LGarXiv:2005.10296v32020A Better Use of Audio-Visual Cues: Dense Video Captioning with Bi-modal Transformer
Vladimir Iashin, Esa Rahtu
cs.CVcs.CLcs.LGarXiv:2005.08271v22020SoundNet: Learning Sound Representations from Unlabeled Video
Yusuf Aytar, Carl Vondrick, Antonio Torralba
cs.CVcs.LGcs.SDarXiv:1610.09001v12016Bit-pragmatic Deep Neural Network Computing
J. Albericio, P. Judd, A. Delmás +2
cs.LGcs.AIcs.ARarXiv:1610.06920v12016Explainable Reinforcement Learning: A Survey
Erika Puiutta, Eric MSP Veith
cs.LGstat.MLarXiv:2005.06247v12020VILA-U: a Unified Foundation Model Integrating Visual Understanding and Generation
Yecheng Wu, Zhuoyang Zhang, Junyu Chen +9
cs.CVcs.LGarXiv:2409.04429v32024We-Math: Does Your Large Multimodal Model Achieve Human-like Mathematical Reasoning?
Runqi Qiao, Qiuna Tan, Guanting Dong +15
cs.AIcs.CLcs.CVarXiv:2407.01284v12024Genetic Algorithms for Tractable Bayesian Network Fusion via Pre-Fusion Edge Pruning
Pablo Torrijos, José A. Gámez, José M. Puerta +1
cs.NEcs.LGarXiv:2609.03724v12026Dude: A Dual-Detection Multi-Agent System for Paper-Code Discrepancy Detection
Weijie Liu, Running Zhao, Wenhao Yuan +4
cs.AIcs.LGarXiv:2609.03416v12026Unique3D: High-Quality and Efficient 3D Mesh Generation from a Single Image
Kailu Wu, Fangfu Liu, Zhihan Cai +5
cs.CVcs.GRcs.LGarXiv:2405.20343v32024Reinforcement learning
Sarod Yatawatta
astro-ph.IMcs.AIcs.LGarXiv:2405.10369v12024Adversarial Machine Learning in Network Intrusion Detection Systems
Elie Alhajjar, Paul Maxwell, Nathaniel D. Bastian
cs.CRcs.LGcs.NEarXiv:2004.11898v12020Improving Dictionary Learning with Gated Sparse Autoencoders
Senthooran Rajamanoharan, Arthur Conmy, Lewis Smith +5
cs.LGcs.AIarXiv:2404.16014v22024WorkArena: How Capable Are Web Agents at Solving Common Knowledge Work Tasks?
Alexandre Drouin, Maxime Gasse, Massimo Caccia +9
cs.LGcs.AIarXiv:2403.07718v52024Mining Fashion Outfit Composition Using An End-to-End Deep Learning Approach on Set Data
Yuncheng Li, LiangLiang Cao, Jiang Zhu +1
cs.MMcs.LGarXiv:1608.03016v22016Assessing the Brittleness of Safety Alignment via Pruning and Low-Rank Modifications
Boyi Wei, Kaixuan Huang, Yangsibo Huang +6
cs.LGcs.AIcs.CLarXiv:2402.05162v42024Subliminal Learning as Trait-Direction Drift: A Mechanism and Targeted Control under SFT Distillation
Zhixuan Liu, Zhichen Dong, Yuyu Fan +2
cs.LGarXiv:2609.01091v22026HarmBench: A Standardized Evaluation Framework for Automated Red Teaming and Robust Refusal
Mantas Mazeika, Long Phan, Xuwang Yin +9
cs.LGcs.AIcs.CLarXiv:2402.04249v22024EEG-AS: Instance-Level Foundation Model Selection for EEG Foundation Models via Behavior Reconstruction
Yunzhen Zhang, Ruoxi Piao, Hasan Onur Keles +1
cs.LGcs.AIarXiv:2609.00653v12026Machine learning-based network intrusion detection for big and imbalanced data using oversampling, stacking feature embedding and feature extraction
Md. Alamin Talukder, Md. Manowarul Islam, Md Ashraf Uddin +4
cs.CRcs.LGarXiv:2401.12262v12024Towards unsupervised representation learning for quantum data: quantum models with inference and generation
Robin Lorenz, Eric Brunner, Marcello Benedetti
quant-phcs.LGarXiv:2609.00372v12026Self-Play Fine-Tuning Converts Weak Language Models to Strong Language Models
Zixiang Chen, Yihe Deng, Huizhuo Yuan +2
cs.LGcs.AIcs.CLarXiv:2401.01335v32024Synthetic Worlds for Temporal Evaluation and Knowledge Updating in LLMs
Jonathan Zheng, Zirui Shao, Alan Ritter +1
cs.CLcs.LGarXiv:2609.00184v12026Backprop KF: Learning Discriminative Deterministic State Estimators
Tuomas Haarnoja, Anurag Ajay, Sergey Levine +1
cs.LGcs.AIarXiv:1605.07148v42016Different representation learning objectives recover distinct latent structures from the same psychometric data
Cong Cao, Tassos C. Kyriakides, Pambos Vrasidas
cs.AIcs.LGstat.MEarXiv:2609.00100v12026Align Your Gaussians: Text-to-4D with Dynamic 3D Gaussians and Composed Diffusion Models
Huan Ling, Seung Wook Kim, Antonio Torralba +2
cs.CVcs.LGarXiv:2312.13763v22023Learning Dynamics of Logits Debiasing for Long-Tailed Semi-Supervised Learning
Yue Cheng, Jiajun Zhang, Xiaohui Gao +2
cs.LGcs.AIarXiv:2608.30699v12026Steering Llama 2 via Contrastive Activation Addition
Nina Panickssery, Nick Gabrieli, Julian Schulz +3
cs.CLcs.AIcs.LGarXiv:2312.06681v42023Baseline Defenses for Adversarial Attacks Against Aligned Language Models
Neel Jain, Avi Schwarzschild, Yuxin Wen +7
cs.LGcs.CLcs.CRarXiv:2309.00614v22023D4: Improving LLM Pretraining via Document De-Duplication and Diversification
Kushal Tirumala, Daniel Simig, Armen Aghajanyan +1
cs.CLcs.AIcs.LGarXiv:2308.12284v12023PMET: Precise Model Editing in a Transformer
Xiaopeng Li, Shasha Li, Shezheng Song +3
cs.CLcs.AIcs.LGarXiv:2308.08742v62023On the use of deep learning for phase recovery
Kaiqiang Wang, Li Song, Chutian Wang +8
physics.opticscs.LGeess.IVarXiv:2308.00942v12023Aligning Multi-Trajectory Supervision with Policy Optimization for VLA Driving
Tian Zhang, Zhuo Huang, Hongrui Ye +3
cs.CVcs.AIcs.LGarXiv:2608.30122v12026A Survey of Techniques for Optimizing Transformer Inference
Krishna Teja Chitty-Venkata, Sparsh Mittal, Murali Emani +2
cs.LGcs.ARcs.CLarXiv:2307.07982v12023Generating images with recurrent adversarial networks
Daniel Jiwoong Im, Chris Dongjoo Kim, Hui Jiang +1
cs.LGcs.CVarXiv:1602.05110v52016