Machine Learning
Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
15,121 to 15,180 of 20,192
Group-Shared Low-Rank Approximation for Mobile-Efficient Pointwise Convolutions in Large-Kernel CNNs
Hao Luo, Yiting Yang, Wenyi Zhao +10
cs.LGarXiv:2608.26069v12026Attend and Diagnose: Clinical Time Series Analysis using Attention Models
Huan Song, Deepta Rajan, Jayaraman J. Thiagarajan +1
stat.MLcs.LGarXiv:1711.03905v22017Robust CurveMoE: Multi-Norm Adversarial Defense for Mixture-of-Experts Models via Mode Connectivity
Xu Zhang, Ren Wang
cs.LGarXiv:2608.26043v12026Spectral Allocation: Why Muon Outperforms Adam, and How to Improve Muon
Xiaodong Wu, Wenyi Yu, Chao Zhang +1
cs.LGarXiv:2608.25990v12026Forecasting Multiple Observables with SCROLL: Score-Trained Uncertainty for Stochastic Dynamics
Pavel Prochazka
cs.LGarXiv:2608.25898v12026Learning Continuous Regional Temperature Fields with Lead-Time and Resolution Queries
Chunlei Shi, Jiong Wang, Yi-Lin Wei +4
cs.LGcs.MMarXiv:2608.25823v12026Deep neural networks for the evaluation and design of photonic devices
Jiaqi Jiang, Mingkun Chen, Jonathan A. Fan
eess.IVcs.LGphysics.app-pharXiv:2007.00084v12020CEDAR: Controlled and Event-Driven Demand Forecasting via Residual Decomposition
Junjie Meng, Ranxu Zhang, Zi-an Zhang +6
cs.LGarXiv:2608.25871v12026Massively Parallel Methods for Deep Reinforcement Learning
Arun Nair, Praveen Srinivasan, Sam Blackwell +11
cs.LGcs.AIcs.DCarXiv:1507.04296v22015How Edge of Stability Hinders SCAFFOLD in Federated Optimization
Anant Khandelwal, Michael Crawshaw, Mingrui Liu
cs.LGarXiv:2608.25873v12026EXAONE Tabular 1.0 : Technical Report
Moonjung Eo, Min-Kook Suh, Hye-Seung Cho +4
cs.LGarXiv:2608.25774v12026Deep Learning for Image and Point Cloud Fusion in Autonomous Driving: A Review
Yaodong Cui, Ren Chen, Wenbo Chu +4
cs.CVcs.LGcs.ROarXiv:2004.05224v22020Large Scale Fine-Grained Categorization and Domain-Specific Transfer Learning
Yin Cui, Yang Song, Chen Sun +2
cs.CVcs.LGarXiv:1806.06193v12018Towards Optimally Decentralized Multi-Robot Collision Avoidance via Deep Reinforcement Learning
Pinxin Long, Tingxiang Fan, Xinyi Liao +3
cs.ROcs.AIcs.LGarXiv:1709.10082v32017$R^3$: Training Robots to Reason in Natural Language via Reinforcement Learning
Lehong Wu, Yuxiao Qu, Zheyuan Hu +4
cs.ROcs.AIcs.CLarXiv:2608.26053v12026Git Re-Basin: Merging Models modulo Permutation Symmetries
Samuel K. Ainsworth, Jonathan Hayase, Siddhartha Srinivasa
cs.LGcs.AIarXiv:2209.04836v62022BlockDrop: Dynamic Inference Paths in Residual Networks
Zuxuan Wu, Tushar Nagarajan, Abhishek Kumar +4
cs.CVcs.LGarXiv:1711.08393v42017One-shot Learning with Memory-Augmented Neural Networks
Adam Santoro, Sergey Bartunov, Matthew Botvinick +2
cs.LGarXiv:1605.06065v12016Soft-to-Hard Vector Quantization for End-to-End Learning Compressible Representations
Eirikur Agustsson, Fabian Mentzer, Michael Tschannen +4
cs.LGcs.CVarXiv:1704.00648v22017Large Language Models Can Be Strong Differentially Private Learners
Xuechen Li, Florian Tramèr, Percy Liang +1
cs.LGcs.CLarXiv:2110.05679v62021HyenaDNA: Long-Range Genomic Sequence Modeling at Single Nucleotide Resolution
Eric Nguyen, Michael Poli, Marjan Faizi +10
cs.LGq-bio.GNarXiv:2306.15794v22023It's a matter of timescale: non-linear utility in successor features and multi-objective planning and learning
Liam P. H. Mertens, Lucas N. Alegre, Florent Delgrange +3
cs.LGcs.AIarXiv:2608.25723v12026Fairness-Aware Test-Time Prompt Tuning
Yoann Launay, Parameswaran Kamalaruban, Tom Kempton +2
cs.LGarXiv:2608.25707v12026True Few-Shot Learning with Language Models
Ethan Perez, Douwe Kiela, Kyunghyun Cho
cs.CLcs.LGstat.MLarXiv:2105.11447v12021LDP-Fed: Federated Learning with Local Differential Privacy
Stacey Truex, Ling Liu, Ka-Ho Chow +2
cs.LGcs.CRstat.MLarXiv:2006.03637v12020Development and evaluation of a deep learning model for protein-ligand binding affinity prediction
Marta M. Stepniewska-Dziubinska, Piotr Zielenkiewicz, Pawel Siedlecki
stat.MLcs.LGq-bio.BMarXiv:1712.07042v22017Delayed Impact of Fair Machine Learning
Lydia T. Liu, Sarah Dean, Esther Rolf +2
cs.LGstat.MLarXiv:1803.04383v22018RLPrompt: Optimizing Discrete Text Prompts with Reinforcement Learning
Mingkai Deng, Jianyu Wang, Cheng-Ping Hsieh +6
cs.CLcs.LGarXiv:2205.12548v32022Sionna: An Open-Source Library for Next-Generation Physical Layer Research
Jakob Hoydis, Sebastian Cammerer, Fayçal Ait Aoudia +4
cs.ITcs.AIcs.LGarXiv:2203.11854v22022Adversarial Training of Linear Models under Stealthy Attacks
Lovisa Eriksson, Dave Zachariah, André M. H. Teixeira
cs.LGcs.CReess.SYarXiv:2608.25681v12026Machine learning based disease diagnosis: A comprehensive review
Md Manjurul Ahsan, Zahed Siddique
cs.LGarXiv:2112.15538v12021Learning Features by Watching Objects Move
Deepak Pathak, Ross Girshick, Piotr Dollár +2
cs.CVcs.AIcs.LGarXiv:1612.06370v22016How Much Rank Does LoRA Need? Rank-Error Bounds for Transformer Attention
Gerard Conangla Planes
cs.LGcs.AIcs.CLarXiv:2608.26052v12026On the Tractability of SHAP Explanations
Guy Van den Broeck, Anton Lykov, Maximilian Schleich +1
cs.AIcs.CCcs.LGarXiv:2009.08634v22020Balanced Distribution Adaptation for Transfer Learning
Jindong Wang, Yiqiang Chen, Shuji Hao +2
cs.LGstat.MLarXiv:1807.00516v12018COMBO: Conservative Offline Model-Based Policy Optimization
Tianhe Yu, Aviral Kumar, Rafael Rafailov +3
cs.LGcs.AIcs.ROarXiv:2102.08363v22021A survey on modern trainable activation functions
Andrea Apicella, Francesco Donnarumma, Francesco Isgrò +1
cs.LGcs.NEstat.MLarXiv:2005.00817v42020Focal Self-attention for Local-Global Interactions in Vision Transformers
Jianwei Yang, Chunyuan Li, Pengchuan Zhang +4
cs.CVcs.AIcs.LGarXiv:2107.00641v12021One Symptom, Three Levers: A Critical Review of On-Policy Self-Distillation
Justin Robert, Raheel Qader
cs.LGcs.AIcs.CLarXiv:2608.25936v12026Learning Task Grouping and Overlap in Multi-task Learning
Abhishek Kumar, Hal Daume
cs.LGstat.MLarXiv:1206.6417v12012Deep Learning over Multi-field Categorical Data: A Case Study on User Response Prediction
Weinan Zhang, Tianming Du, Jun Wang
cs.LGcs.IRarXiv:1601.02376v12016Task-Agnostic Meta-Learning for Few-shot Learning
Muhammad Abdullah Jamal, Guo-Jun Qi, Mubarak Shah
cs.LGstat.MLarXiv:1805.07722v12018Convolutional Recurrent Neural Networks for Music Classification
Keunwoo Choi, George Fazekas, Mark Sandler +1
cs.NEcs.LGcs.MMarXiv:1609.04243v32016StyleSpace Analysis: Disentangled Controls for StyleGAN Image Generation
Zongze Wu, Dani Lischinski, Eli Shechtman
cs.CVcs.GRcs.LGarXiv:2011.12799v22020Why Does Graph Learning Fail to Fully Benefit from a Text Teacher?
Fumiaki Kimino, Ryoma Sato
cs.LGcs.CLarXiv:2608.25741v12026Prefix Sliding for efficient test-time scaling
Niklas Muennighoff, Zhengyang Wang, Zeyi Chen +15
cs.CLcs.AIcs.LGarXiv:2608.26070v12026Bayesian Optimization with Unknown Constraints
Michael A. Gelbart, Jasper Snoek, Ryan P. Adams
stat.MLcs.LGarXiv:1403.5607v12014Multi-layer Representation Learning for Medical Concepts
Edward Choi, Mohammad Taha Bahadori, Elizabeth Searles +2
cs.LGarXiv:1602.05568v12016The Secret Revealer: Generative Model-Inversion Attacks Against Deep Neural Networks
Yuheng Zhang, Ruoxi Jia, Hengzhi Pei +3
cs.LGstat.MLarXiv:1911.07135v22019Statistical guarantees for the EM algorithm: From population to sample-based analysis
Sivaraman Balakrishnan, Martin J. Wainwright, Bin Yu
math.STcs.LGstat.MLarXiv:1408.2156v12014Structured Inference Networks for Nonlinear State Space Models
Rahul G. Krishnan, Uri Shalit, David Sontag
stat.MLcs.AIcs.LGarXiv:1609.09869v22016AI4COVID-19: AI Enabled Preliminary Diagnosis for COVID-19 from Cough Samples via an App
Ali Imran, Iryna Posokhova, Haneya N. Qureshi +6
eess.AScs.LGcs.SDarXiv:2004.01275v62020BERT: A Review of Applications in Natural Language Processing and Understanding
M. V. Koroteev
cs.CLcs.AIcs.LGarXiv:2103.11943v12021Conditional Probability Models for Deep Image Compression
Fabian Mentzer, Eirikur Agustsson, Michael Tschannen +2
cs.CVcs.LGarXiv:1801.04260v42018Learning to Hash for Indexing Big Data - A Survey
Jun Wang, Wei Liu, Sanjiv Kumar +1
cs.LGarXiv:1509.05472v12015An Empirical Study of Spatial Attention Mechanisms in Deep Networks
Xizhou Zhu, Dazhi Cheng, Zheng Zhang +2
cs.CVcs.CLcs.LGarXiv:1904.05873v12019Vote3Deep: Fast Object Detection in 3D Point Clouds Using Efficient Convolutional Neural Networks
Martin Engelcke, Dushyant Rao, Dominic Zeng Wang +2
cs.ROcs.AIcs.CVarXiv:1609.06666v22016Summaries:한국어Technical Report on the CleverHans v2.1.0 Adversarial Examples Library
Nicolas Papernot, Fartash Faghri, Nicholas Carlini +23
cs.LGcs.CRstat.MLarXiv:1610.00768v62016Iterative Deep Graph Learning for Graph Neural Networks: Better and Robust Node Embeddings
Yu Chen, Lingfei Wu, Mohammed J. Zaki
cs.LGstat.MLarXiv:2006.13009v22020Topology Attack and Defense for Graph Neural Networks: An Optimization Perspective
Kaidi Xu, Hongge Chen, Sijia Liu +4
cs.LGcs.CRcs.SIarXiv:1906.04214v32019