Machine Learning

Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

15,121 to 15,180 of 20,192

  1. Group-Shared Low-Rank Approximation for Mobile-Efficient Pointwise Convolutions in Large-Kernel CNNs

    Hao Luo, Yiting Yang, Wenyi Zhao +10

    cs.LGarXiv:2608.26069v12026
  2. Attend and Diagnose: Clinical Time Series Analysis using Attention Models

    Huan Song, Deepta Rajan, Jayaraman J. Thiagarajan +1

    stat.MLcs.LGarXiv:1711.03905v22017
  3. Robust CurveMoE: Multi-Norm Adversarial Defense for Mixture-of-Experts Models via Mode Connectivity

    Xu Zhang, Ren Wang

    cs.LGarXiv:2608.26043v12026
  4. Spectral Allocation: Why Muon Outperforms Adam, and How to Improve Muon

    Xiaodong Wu, Wenyi Yu, Chao Zhang +1

    cs.LGarXiv:2608.25990v12026
  5. Forecasting Multiple Observables with SCROLL: Score-Trained Uncertainty for Stochastic Dynamics

    Pavel Prochazka

    cs.LGarXiv:2608.25898v12026
  6. Learning Continuous Regional Temperature Fields with Lead-Time and Resolution Queries

    Chunlei Shi, Jiong Wang, Yi-Lin Wei +4

    cs.LGcs.MMarXiv:2608.25823v12026
  7. Deep neural networks for the evaluation and design of photonic devices

    Jiaqi Jiang, Mingkun Chen, Jonathan A. Fan

    eess.IVcs.LGphysics.app-pharXiv:2007.00084v12020
  8. CEDAR: Controlled and Event-Driven Demand Forecasting via Residual Decomposition

    Junjie Meng, Ranxu Zhang, Zi-an Zhang +6

    cs.LGarXiv:2608.25871v12026
  9. Massively Parallel Methods for Deep Reinforcement Learning

    Arun Nair, Praveen Srinivasan, Sam Blackwell +11

    cs.LGcs.AIcs.DCarXiv:1507.04296v22015
  10. How Edge of Stability Hinders SCAFFOLD in Federated Optimization

    Anant Khandelwal, Michael Crawshaw, Mingrui Liu

    cs.LGarXiv:2608.25873v12026
  11. EXAONE Tabular 1.0 : Technical Report

    Moonjung Eo, Min-Kook Suh, Hye-Seung Cho +4

    cs.LGarXiv:2608.25774v12026
  12. Deep Learning for Image and Point Cloud Fusion in Autonomous Driving: A Review

    Yaodong Cui, Ren Chen, Wenbo Chu +4

    cs.CVcs.LGcs.ROarXiv:2004.05224v22020
  13. Large Scale Fine-Grained Categorization and Domain-Specific Transfer Learning

    Yin Cui, Yang Song, Chen Sun +2

    cs.CVcs.LGarXiv:1806.06193v12018
  14. Towards Optimally Decentralized Multi-Robot Collision Avoidance via Deep Reinforcement Learning

    Pinxin Long, Tingxiang Fan, Xinyi Liao +3

    cs.ROcs.AIcs.LGarXiv:1709.10082v32017
  15. $R^3$: Training Robots to Reason in Natural Language via Reinforcement Learning

    Lehong Wu, Yuxiao Qu, Zheyuan Hu +4

    cs.ROcs.AIcs.CLarXiv:2608.26053v12026
  16. Git Re-Basin: Merging Models modulo Permutation Symmetries

    Samuel K. Ainsworth, Jonathan Hayase, Siddhartha Srinivasa

    cs.LGcs.AIarXiv:2209.04836v62022
  17. BlockDrop: Dynamic Inference Paths in Residual Networks

    Zuxuan Wu, Tushar Nagarajan, Abhishek Kumar +4

    cs.CVcs.LGarXiv:1711.08393v42017
  18. One-shot Learning with Memory-Augmented Neural Networks

    Adam Santoro, Sergey Bartunov, Matthew Botvinick +2

    cs.LGarXiv:1605.06065v12016
  19. Soft-to-Hard Vector Quantization for End-to-End Learning Compressible Representations

    Eirikur Agustsson, Fabian Mentzer, Michael Tschannen +4

    cs.LGcs.CVarXiv:1704.00648v22017
  20. Large Language Models Can Be Strong Differentially Private Learners

    Xuechen Li, Florian Tramèr, Percy Liang +1

    cs.LGcs.CLarXiv:2110.05679v62021
  21. HyenaDNA: Long-Range Genomic Sequence Modeling at Single Nucleotide Resolution

    Eric Nguyen, Michael Poli, Marjan Faizi +10

    cs.LGq-bio.GNarXiv:2306.15794v22023
  22. It's a matter of timescale: non-linear utility in successor features and multi-objective planning and learning

    Liam P. H. Mertens, Lucas N. Alegre, Florent Delgrange +3

    cs.LGcs.AIarXiv:2608.25723v12026
  23. Fairness-Aware Test-Time Prompt Tuning

    Yoann Launay, Parameswaran Kamalaruban, Tom Kempton +2

    cs.LGarXiv:2608.25707v12026
  24. True Few-Shot Learning with Language Models

    Ethan Perez, Douwe Kiela, Kyunghyun Cho

    cs.CLcs.LGstat.MLarXiv:2105.11447v12021
  25. LDP-Fed: Federated Learning with Local Differential Privacy

    Stacey Truex, Ling Liu, Ka-Ho Chow +2

    cs.LGcs.CRstat.MLarXiv:2006.03637v12020
  26. Development and evaluation of a deep learning model for protein-ligand binding affinity prediction

    Marta M. Stepniewska-Dziubinska, Piotr Zielenkiewicz, Pawel Siedlecki

    stat.MLcs.LGq-bio.BMarXiv:1712.07042v22017
  27. Delayed Impact of Fair Machine Learning

    Lydia T. Liu, Sarah Dean, Esther Rolf +2

    cs.LGstat.MLarXiv:1803.04383v22018
  28. RLPrompt: Optimizing Discrete Text Prompts with Reinforcement Learning

    Mingkai Deng, Jianyu Wang, Cheng-Ping Hsieh +6

    cs.CLcs.LGarXiv:2205.12548v32022
  29. Sionna: An Open-Source Library for Next-Generation Physical Layer Research

    Jakob Hoydis, Sebastian Cammerer, Fayçal Ait Aoudia +4

    cs.ITcs.AIcs.LGarXiv:2203.11854v22022
  30. Adversarial Training of Linear Models under Stealthy Attacks

    Lovisa Eriksson, Dave Zachariah, André M. H. Teixeira

    cs.LGcs.CReess.SYarXiv:2608.25681v12026
  31. Machine learning based disease diagnosis: A comprehensive review

    Md Manjurul Ahsan, Zahed Siddique

    cs.LGarXiv:2112.15538v12021
  32. Learning Features by Watching Objects Move

    Deepak Pathak, Ross Girshick, Piotr Dollár +2

    cs.CVcs.AIcs.LGarXiv:1612.06370v22016
  33. How Much Rank Does LoRA Need? Rank-Error Bounds for Transformer Attention

    Gerard Conangla Planes

    cs.LGcs.AIcs.CLarXiv:2608.26052v12026
  34. On the Tractability of SHAP Explanations

    Guy Van den Broeck, Anton Lykov, Maximilian Schleich +1

    cs.AIcs.CCcs.LGarXiv:2009.08634v22020
  35. Balanced Distribution Adaptation for Transfer Learning

    Jindong Wang, Yiqiang Chen, Shuji Hao +2

    cs.LGstat.MLarXiv:1807.00516v12018
  36. COMBO: Conservative Offline Model-Based Policy Optimization

    Tianhe Yu, Aviral Kumar, Rafael Rafailov +3

    cs.LGcs.AIcs.ROarXiv:2102.08363v22021
  37. A survey on modern trainable activation functions

    Andrea Apicella, Francesco Donnarumma, Francesco Isgrò +1

    cs.LGcs.NEstat.MLarXiv:2005.00817v42020
  38. Focal Self-attention for Local-Global Interactions in Vision Transformers

    Jianwei Yang, Chunyuan Li, Pengchuan Zhang +4

    cs.CVcs.AIcs.LGarXiv:2107.00641v12021
  39. One Symptom, Three Levers: A Critical Review of On-Policy Self-Distillation

    Justin Robert, Raheel Qader

    cs.LGcs.AIcs.CLarXiv:2608.25936v12026
  40. Learning Task Grouping and Overlap in Multi-task Learning

    Abhishek Kumar, Hal Daume

    cs.LGstat.MLarXiv:1206.6417v12012
  41. Deep Learning over Multi-field Categorical Data: A Case Study on User Response Prediction

    Weinan Zhang, Tianming Du, Jun Wang

    cs.LGcs.IRarXiv:1601.02376v12016
  42. Task-Agnostic Meta-Learning for Few-shot Learning

    Muhammad Abdullah Jamal, Guo-Jun Qi, Mubarak Shah

    cs.LGstat.MLarXiv:1805.07722v12018
  43. Convolutional Recurrent Neural Networks for Music Classification

    Keunwoo Choi, George Fazekas, Mark Sandler +1

    cs.NEcs.LGcs.MMarXiv:1609.04243v32016
  44. StyleSpace Analysis: Disentangled Controls for StyleGAN Image Generation

    Zongze Wu, Dani Lischinski, Eli Shechtman

    cs.CVcs.GRcs.LGarXiv:2011.12799v22020
  45. Why Does Graph Learning Fail to Fully Benefit from a Text Teacher?

    Fumiaki Kimino, Ryoma Sato

    cs.LGcs.CLarXiv:2608.25741v12026
  46. Prefix Sliding for efficient test-time scaling

    Niklas Muennighoff, Zhengyang Wang, Zeyi Chen +15

    cs.CLcs.AIcs.LGarXiv:2608.26070v12026
  47. Bayesian Optimization with Unknown Constraints

    Michael A. Gelbart, Jasper Snoek, Ryan P. Adams

    stat.MLcs.LGarXiv:1403.5607v12014
  48. Multi-layer Representation Learning for Medical Concepts

    Edward Choi, Mohammad Taha Bahadori, Elizabeth Searles +2

    cs.LGarXiv:1602.05568v12016
  49. The Secret Revealer: Generative Model-Inversion Attacks Against Deep Neural Networks

    Yuheng Zhang, Ruoxi Jia, Hengzhi Pei +3

    cs.LGstat.MLarXiv:1911.07135v22019
  50. Statistical guarantees for the EM algorithm: From population to sample-based analysis

    Sivaraman Balakrishnan, Martin J. Wainwright, Bin Yu

    math.STcs.LGstat.MLarXiv:1408.2156v12014
  51. Structured Inference Networks for Nonlinear State Space Models

    Rahul G. Krishnan, Uri Shalit, David Sontag

    stat.MLcs.AIcs.LGarXiv:1609.09869v22016
  52. AI4COVID-19: AI Enabled Preliminary Diagnosis for COVID-19 from Cough Samples via an App

    Ali Imran, Iryna Posokhova, Haneya N. Qureshi +6

    eess.AScs.LGcs.SDarXiv:2004.01275v62020
  53. BERT: A Review of Applications in Natural Language Processing and Understanding

    M. V. Koroteev

    cs.CLcs.AIcs.LGarXiv:2103.11943v12021
  54. Conditional Probability Models for Deep Image Compression

    Fabian Mentzer, Eirikur Agustsson, Michael Tschannen +2

    cs.CVcs.LGarXiv:1801.04260v42018
  55. Learning to Hash for Indexing Big Data - A Survey

    Jun Wang, Wei Liu, Sanjiv Kumar +1

    cs.LGarXiv:1509.05472v12015
  56. An Empirical Study of Spatial Attention Mechanisms in Deep Networks

    Xizhou Zhu, Dazhi Cheng, Zheng Zhang +2

    cs.CVcs.CLcs.LGarXiv:1904.05873v12019
  57. Vote3Deep: Fast Object Detection in 3D Point Clouds Using Efficient Convolutional Neural Networks

    Martin Engelcke, Dushyant Rao, Dominic Zeng Wang +2

    cs.ROcs.AIcs.CVarXiv:1609.06666v22016
    Summaries:한국어
  58. Technical Report on the CleverHans v2.1.0 Adversarial Examples Library

    Nicolas Papernot, Fartash Faghri, Nicholas Carlini +23

    cs.LGcs.CRstat.MLarXiv:1610.00768v62016
  59. Iterative Deep Graph Learning for Graph Neural Networks: Better and Robust Node Embeddings

    Yu Chen, Lingfei Wu, Mohammed J. Zaki

    cs.LGstat.MLarXiv:2006.13009v22020
  60. Topology Attack and Defense for Graph Neural Networks: An Optimization Perspective

    Kaidi Xu, Hongge Chen, Sijia Liu +4

    cs.LGcs.CRcs.SIarXiv:1906.04214v32019