Machine Learning

Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

841 to 900 of 20,178

  1. Review of Swarm Intelligence-based Feature Selection Methods

    Mehrdad Rostami, Kamal Berahmand, Saman Forouzandeh

    cs.LGcs.NEstat.MLarXiv:2008.04103v12020
  2. Robust Adversarial Reinforcement Learning

    Lerrel Pinto, James Davidson, Rahul Sukthankar +1

    cs.LGcs.AIcs.MAarXiv:1703.02702v12017
  3. Aligning AI With Shared Human Values

    Dan Hendrycks, Collin Burns, Steven Basart +4

    cs.CYcs.AIcs.CLarXiv:2008.02275v62020
  4. An Analytical Formula of Population Gradient for two-layered ReLU network and its Applications in Convergence and Critical Point Analysis

    Yuandong Tian

    cs.LGarXiv:1703.00560v22017
  5. Segmentation of optic disc, fovea and retinal vasculature using a single convolutional neural network

    Jen Hong Tan, U. Rajendra Acharya, Sulatha V. Bhandary +2

    cs.CVcs.LGarXiv:1702.00509v12017
  6. Learn&Fuzz: Machine Learning for Input Fuzzing

    Patrice Godefroid, Hila Peleg, Rishabh Singh

    cs.AIcs.CRcs.LGarXiv:1701.07232v12017
  7. GenMol: A Drug Discovery Generalist with Discrete Diffusion

    Seul Lee, Karsten Kreis, Srimukh Prasad Veccham +6

    cs.LGarXiv:2501.06158v32025
  8. Test-time Alignment of Diffusion Models without Reward Over-optimization

    Sunwoo Kim, Minkyu Kim, Dongmin Park

    cs.LGcs.AIcs.CVarXiv:2501.05803v32025
  9. An Explainable Machine Learning Framework for Predicting Blood-Brain Barrier Permeability Using Molecular Descriptors

    Fatemeh Mahmoudi

    cs.LGcond-mat.mtrl-sciarXiv:2609.10012v12026
  10. Infra-Bench CLS: A Global, Open-Source Benchmark for Critical Infrastructure Classification with Earth Observation Foundation Models

    Justin Guthrie, Edward Oughton, Konrad Wessels +2

    cs.CVcs.LGarXiv:2609.09482v12026
  11. A Survey on Large Language Models with some Insights on their Capabilities and Limitations

    Andrea Matarazzo, Riccardo Torlone

    cs.CLcs.AIcs.LGarXiv:2501.04040v22025
  12. Beyond Skip Connections: Top-Down Modulation for Object Detection

    Abhinav Shrivastava, Rahul Sukthankar, Jitendra Malik +1

    cs.CVcs.LGarXiv:1612.06851v22016
  13. Predicting Process Behaviour using Deep Learning

    Joerg Evermann, Jana-Rebecca Rehse, Peter Fettke

    cs.LGstat.MLarXiv:1612.04600v22016
  14. Training-Free Task Vectors for LLM Behavioral Control

    Gabriel J. Perin, Lucas Boscaini, André Araujo +1

    cs.LGcs.AIarXiv:2609.09054v12026
  15. Backdoor Attacks Against Deep Learning Systems in the Physical World

    Emily Wenger, Josephine Passananti, Arjun Bhagoji +3

    cs.CVcs.CRcs.LGarXiv:2006.14580v42020
  16. The BrowserGym Ecosystem for Web Agent Research

    Thibault Le Sellier De Chezelles, Maxime Gasse, Alexandre Drouin +17

    cs.LGcs.AIcs.SEarXiv:2412.05467v42024
  17. Pyramidal Convolution: Rethinking Convolutional Neural Networks for Visual Recognition

    Ionut Cosmin Duta, Li Liu, Fan Zhu +1

    cs.CVcs.LGeess.IVarXiv:2006.11538v12020
  18. A causal framework for discovering and removing direct and indirect discrimination

    Lu Zhang, Yongkai Wu, Xintao Wu

    cs.LGarXiv:1611.07509v12016
  19. wav2vec 2.0: A Framework for Self-Supervised Learning of Speech Representations

    Alexei Baevski, Henry Zhou, Abdelrahman Mohamed +1

    cs.CLcs.LGcs.SDarXiv:2006.11477v32020
  20. AWAC: Accelerating Online Reinforcement Learning with Offline Datasets

    Ashvin Nair, Abhishek Gupta, Murtaza Dalal +1

    cs.LGcs.ROstat.MLarXiv:2006.09359v62020
  21. SUN: Reaching for Novelty in Reinforcement Learning

    Wenyan Yang, Arsenii Mustafin, Dominik Baumann +2

    cs.LGcs.AIcs.ROarXiv:2609.08642v12026
  22. Foundations of Structural Causal Models with Cycles and Latent Variables

    Stephan Bongers, Patrick Forré, Jonas Peters +1

    stat.MEcs.AIcs.LGarXiv:1611.06221v62016
  23. Robust Learning Through Cross-Task Consistency

    Amir Zamir, Alexander Sax, Teresa Yeo +6

    cs.CVcs.GRcs.LGarXiv:2006.04096v12020
  24. SaliencyMix: A Saliency Guided Data Augmentation Strategy for Better Regularization

    A. F. M. Shahab Uddin, Mst. Sirazam Monira, Wheemyung Shin +2

    cs.LGstat.MLarXiv:2006.01791v22020
  25. Embodied Agent Interface: Benchmarking LLMs for Embodied Decision Making

    Manling Li, Shiyu Zhao, Qineng Wang +12

    cs.CLcs.AIcs.LGarXiv:2410.07166v32024
  26. SWIFT: Super-fast and Robust Privacy-Preserving Machine Learning

    Nishat Koti, Mahak Pancholi, Arpita Patra +1

    cs.CRcs.LGarXiv:2005.10296v32020
  27. A Better Use of Audio-Visual Cues: Dense Video Captioning with Bi-modal Transformer

    Vladimir Iashin, Esa Rahtu

    cs.CVcs.CLcs.LGarXiv:2005.08271v22020
  28. SoundNet: Learning Sound Representations from Unlabeled Video

    Yusuf Aytar, Carl Vondrick, Antonio Torralba

    cs.CVcs.LGcs.SDarXiv:1610.09001v12016
  29. Bit-pragmatic Deep Neural Network Computing

    J. Albericio, P. Judd, A. Delmás +2

    cs.LGcs.AIcs.ARarXiv:1610.06920v12016
  30. Explainable Reinforcement Learning: A Survey

    Erika Puiutta, Eric MSP Veith

    cs.LGstat.MLarXiv:2005.06247v12020
  31. VILA-U: a Unified Foundation Model Integrating Visual Understanding and Generation

    Yecheng Wu, Zhuoyang Zhang, Junyu Chen +9

    cs.CVcs.LGarXiv:2409.04429v32024
  32. We-Math: Does Your Large Multimodal Model Achieve Human-like Mathematical Reasoning?

    Runqi Qiao, Qiuna Tan, Guanting Dong +15

    cs.AIcs.CLcs.CVarXiv:2407.01284v12024
  33. Genetic Algorithms for Tractable Bayesian Network Fusion via Pre-Fusion Edge Pruning

    Pablo Torrijos, José A. Gámez, José M. Puerta +1

    cs.NEcs.LGarXiv:2609.03724v12026
  34. Dude: A Dual-Detection Multi-Agent System for Paper-Code Discrepancy Detection

    Weijie Liu, Running Zhao, Wenhao Yuan +4

    cs.AIcs.LGarXiv:2609.03416v12026
  35. Unique3D: High-Quality and Efficient 3D Mesh Generation from a Single Image

    Kailu Wu, Fangfu Liu, Zhihan Cai +5

    cs.CVcs.GRcs.LGarXiv:2405.20343v32024
  36. Reinforcement learning

    Sarod Yatawatta

    astro-ph.IMcs.AIcs.LGarXiv:2405.10369v12024
  37. Adversarial Machine Learning in Network Intrusion Detection Systems

    Elie Alhajjar, Paul Maxwell, Nathaniel D. Bastian

    cs.CRcs.LGcs.NEarXiv:2004.11898v12020
  38. Improving Dictionary Learning with Gated Sparse Autoencoders

    Senthooran Rajamanoharan, Arthur Conmy, Lewis Smith +5

    cs.LGcs.AIarXiv:2404.16014v22024
  39. WorkArena: How Capable Are Web Agents at Solving Common Knowledge Work Tasks?

    Alexandre Drouin, Maxime Gasse, Massimo Caccia +9

    cs.LGcs.AIarXiv:2403.07718v52024
  40. Mining Fashion Outfit Composition Using An End-to-End Deep Learning Approach on Set Data

    Yuncheng Li, LiangLiang Cao, Jiang Zhu +1

    cs.MMcs.LGarXiv:1608.03016v22016
  41. Assessing the Brittleness of Safety Alignment via Pruning and Low-Rank Modifications

    Boyi Wei, Kaixuan Huang, Yangsibo Huang +6

    cs.LGcs.AIcs.CLarXiv:2402.05162v42024
  42. Subliminal Learning as Trait-Direction Drift: A Mechanism and Targeted Control under SFT Distillation

    Zhixuan Liu, Zhichen Dong, Yuyu Fan +2

    cs.LGarXiv:2609.01091v22026
  43. HarmBench: A Standardized Evaluation Framework for Automated Red Teaming and Robust Refusal

    Mantas Mazeika, Long Phan, Xuwang Yin +9

    cs.LGcs.AIcs.CLarXiv:2402.04249v22024
  44. EEG-AS: Instance-Level Foundation Model Selection for EEG Foundation Models via Behavior Reconstruction

    Yunzhen Zhang, Ruoxi Piao, Hasan Onur Keles +1

    cs.LGcs.AIarXiv:2609.00653v12026
  45. Machine learning-based network intrusion detection for big and imbalanced data using oversampling, stacking feature embedding and feature extraction

    Md. Alamin Talukder, Md. Manowarul Islam, Md Ashraf Uddin +4

    cs.CRcs.LGarXiv:2401.12262v12024
  46. Towards unsupervised representation learning for quantum data: quantum models with inference and generation

    Robin Lorenz, Eric Brunner, Marcello Benedetti

    quant-phcs.LGarXiv:2609.00372v12026
  47. Self-Play Fine-Tuning Converts Weak Language Models to Strong Language Models

    Zixiang Chen, Yihe Deng, Huizhuo Yuan +2

    cs.LGcs.AIcs.CLarXiv:2401.01335v32024
  48. Synthetic Worlds for Temporal Evaluation and Knowledge Updating in LLMs

    Jonathan Zheng, Zirui Shao, Alan Ritter +1

    cs.CLcs.LGarXiv:2609.00184v12026
  49. Backprop KF: Learning Discriminative Deterministic State Estimators

    Tuomas Haarnoja, Anurag Ajay, Sergey Levine +1

    cs.LGcs.AIarXiv:1605.07148v42016
  50. Different representation learning objectives recover distinct latent structures from the same psychometric data

    Cong Cao, Tassos C. Kyriakides, Pambos Vrasidas

    cs.AIcs.LGstat.MEarXiv:2609.00100v12026
  51. Align Your Gaussians: Text-to-4D with Dynamic 3D Gaussians and Composed Diffusion Models

    Huan Ling, Seung Wook Kim, Antonio Torralba +2

    cs.CVcs.LGarXiv:2312.13763v22023
  52. Learning Dynamics of Logits Debiasing for Long-Tailed Semi-Supervised Learning

    Yue Cheng, Jiajun Zhang, Xiaohui Gao +2

    cs.LGcs.AIarXiv:2608.30699v12026
  53. Steering Llama 2 via Contrastive Activation Addition

    Nina Panickssery, Nick Gabrieli, Julian Schulz +3

    cs.CLcs.AIcs.LGarXiv:2312.06681v42023
  54. Baseline Defenses for Adversarial Attacks Against Aligned Language Models

    Neel Jain, Avi Schwarzschild, Yuxin Wen +7

    cs.LGcs.CLcs.CRarXiv:2309.00614v22023
  55. D4: Improving LLM Pretraining via Document De-Duplication and Diversification

    Kushal Tirumala, Daniel Simig, Armen Aghajanyan +1

    cs.CLcs.AIcs.LGarXiv:2308.12284v12023
  56. PMET: Precise Model Editing in a Transformer

    Xiaopeng Li, Shasha Li, Shezheng Song +3

    cs.CLcs.AIcs.LGarXiv:2308.08742v62023
  57. On the use of deep learning for phase recovery

    Kaiqiang Wang, Li Song, Chutian Wang +8

    physics.opticscs.LGeess.IVarXiv:2308.00942v12023
  58. Aligning Multi-Trajectory Supervision with Policy Optimization for VLA Driving

    Tian Zhang, Zhuo Huang, Hongrui Ye +3

    cs.CVcs.AIcs.LGarXiv:2608.30122v12026
  59. A Survey of Techniques for Optimizing Transformer Inference

    Krishna Teja Chitty-Venkata, Sparsh Mittal, Murali Emani +2

    cs.LGcs.ARcs.CLarXiv:2307.07982v12023
  60. Generating images with recurrent adversarial networks

    Daniel Jiwoong Im, Chris Dongjoo Kim, Hui Jiang +1

    cs.LGcs.CVarXiv:1602.05110v52016