Machine Learning

Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

13,501 to 13,560 of 20,219

  1. Common Geodesics Do Not Guarantee Fisher Consistency of the Structured SVM: Minimal Counterexamples and a Tree-Metric Classification

    Jintao Fei, Jiangying Luo

    cs.LGarXiv:2608.27203v12026
  2. Towards better decoding and language model integration in sequence to sequence models

    Jan Chorowski, Navdeep Jaitly

    cs.NEcs.CLcs.LGarXiv:1612.02695v12016
  3. On the Apparent Conflict Between Individual and Group Fairness

    Reuben Binns

    cs.LGcs.CYstat.MLarXiv:1912.06883v12019
  4. Domain-Specific Self-Supervised Representation Learning for Retinal Fundus Classification

    Bekzat Nurlanbekova, Fung Fung Ting

    cs.CVcs.LGarXiv:2608.26686v12026
  5. FairGAN: Fairness-aware Generative Adversarial Networks

    Depeng Xu, Shuhan Yuan, Lu Zhang +1

    cs.LGcs.CYstat.MLarXiv:1805.11202v12018
  6. pix2code: Generating Code from a Graphical User Interface Screenshot

    Tony Beltramelli

    cs.LGcs.AIcs.CLarXiv:1705.07962v22017
  7. SAGE: Variate-Wise Semantic Augmentation for Vision-Language Time Series Forecasting

    Haizhao Fan, Xinyi Le

    cs.LGcs.CVarXiv:2608.26829v12026
  8. Towards End-to-End Speech Recognition with Deep Convolutional Neural Networks

    Ying Zhang, Mohammad Pezeshki, Philemon Brakel +3

    cs.CLcs.LGstat.MLarXiv:1701.02720v12017
  9. Information Leakage in Embedding Models

    Congzheng Song, Ananth Raghunathan

    cs.LGcs.CLcs.CRarXiv:2004.00053v22020
  10. The Why and How of Nonnegative Matrix Factorization

    Nicolas Gillis

    stat.MLcs.IRcs.LGarXiv:1401.5226v22014
  11. Tabular Deep Learning for Algorithmic Trading: Cross-Regime Bayesian Optimisation for Equity Signal Generation

    Joshua Le Grice

    cs.LGq-fin.CPq-fin.TRarXiv:2608.27076v12026
  12. CrossFormer: A Versatile Vision Transformer Hinging on Cross-scale Attention

    Wenxiao Wang, Lu Yao, Long Chen +4

    cs.CVcs.LGarXiv:2108.00154v22021
  13. ARCH: Animatable Reconstruction of Clothed Humans

    Zeng Huang, Yuanlu Xu, Christoph Lassner +2

    cs.GRcs.CVcs.LGarXiv:2004.04572v22020
  14. Disentangling Optimization Scale from Preference Scale in DPO

    Ivan Kruzhilov

    cs.LGarXiv:2608.27032v12026
  15. Online Structured Laplace Approximations For Overcoming Catastrophic Forgetting

    Hippolyt Ritter, Aleksandar Botev, David Barber

    stat.MLcs.LGarXiv:1805.07810v12018
  16. Multiple Futures Prediction

    Yichuan Charlie Tang, Ruslan Salakhutdinov

    cs.LGcs.CVcs.MAarXiv:1911.00997v22019
  17. Spatially Adaptive Computation Time for Residual Networks

    Michael Figurnov, Maxwell D. Collins, Yukun Zhu +4

    cs.CVcs.LGarXiv:1612.02297v22016
  18. Aligning Domain-specific Distribution and Classifier for Cross-domain Classification from Multiple Sources

    Yongchun Zhu, Fuzhen Zhuang, Deqing Wang

    cs.LGcs.AIcs.CVarXiv:2201.01003v12022
  19. ClusterAttention: A training-free speedup of bidirectional attention

    Kasper Nordenram, Amelie Dittmann

    cs.LGcs.CVarXiv:2608.26965v12026
  20. Permutation Invariant Graph Generation via Score-Based Generative Modeling

    Chenhao Niu, Yang Song, Jiaming Song +3

    cs.LGstat.MLarXiv:2003.00638v12020
  21. S-Prompts Learning with Pre-trained Transformers: An Occam's Razor for Domain Incremental Learning

    Yabin Wang, Zhiwu Huang, Xiaopeng Hong

    cs.CVcs.LGarXiv:2207.12819v22022
  22. A Layer Importance Metric for Quantization Accounting for the Speed-Quality Trade-off in Autoregressive Models

    Artem Safronov

    cs.LGarXiv:2608.26926v12026
  23. Gradient Matching for Domain Generalization

    Yuge Shi, Jeffrey Seely, Philip H. S. Torr +4

    cs.LGstat.MLarXiv:2104.09937v32021
  24. On the Convergence of A Class of Adam-Type Algorithms for Non-Convex Optimization

    Xiangyi Chen, Sijia Liu, Ruoyu Sun +1

    cs.LGmath.OCstat.MLarXiv:1808.02941v22018
  25. Beyond Client Averaging: A Client-Independent Second-Order Stationary-Bias Component in Stochastic SCAFFOLD

    Yi-Ping Tang, Guan-Ju Peng

    cs.LGmath.STarXiv:2608.26765v12026
  26. A Unified Analysis of Extra-gradient and Optimistic Gradient Methods for Saddle Point Problems: Proximal Point Approach

    Aryan Mokhtari, Asuman Ozdaglar, Sarath Pattathil

    math.OCcs.LGstat.MLarXiv:1901.08511v42019
  27. Exploring the Landscape of Spatial Robustness

    Logan Engstrom, Brandon Tran, Dimitris Tsipras +2

    cs.LGcs.CVcs.NEarXiv:1712.02779v42017
  28. Flexibly Fair Representation Learning by Disentanglement

    Elliot Creager, David Madras, Jörn-Henrik Jacobsen +4

    cs.LGcs.AIstat.MLarXiv:1906.02589v12019
  29. Distillation-Based Semi-Supervised Federated Learning for Communication-Efficient Collaborative Training with Non-IID Private Data

    Sohei Itahara, Takayuki Nishio, Yusuke Koda +2

    cs.DCcs.LGarXiv:2008.06180v22020
  30. Discrete Graph Structure Learning for Forecasting Multiple Time Series

    Chao Shang, Jie Chen, Jinbo Bi

    cs.LGstat.MLarXiv:2101.06861v32021
  31. Attention-based Graph Neural Network for Semi-supervised Learning

    Kiran K. Thekumparampil, Chong Wang, Sewoong Oh +1

    stat.MLcs.AIcs.LGarXiv:1803.03735v12018
  32. Block-Coordinate Frank-Wolfe Optimization for Structural SVMs

    Simon Lacoste-Julien, Martin Jaggi, Mark Schmidt +1

    cs.LGmath.OCstat.MLarXiv:1207.4747v42012
  33. Why ResNet Works? Residuals Generalize

    Fengxiang He, Tongliang Liu, Dacheng Tao

    stat.MLcs.LGarXiv:1904.01367v12019
  34. Enhanced Membership Inference Attacks against Machine Learning Models

    Jiayuan Ye, Aadyaa Maddi, Sasi Kumar Murakonda +2

    cs.LGcs.CRstat.MLarXiv:2111.09679v42021
  35. Axiom-based Grad-CAM: Towards Accurate Visualization and Explanation of CNNs

    Ruigang Fu, Qingyong Hu, Xiaohu Dong +3

    cs.CVcs.AIcs.LGarXiv:2008.02312v42020
  36. Identifying Mislabeled Data using the Area Under the Margin Ranking

    Geoff Pleiss, Tianyi Zhang, Ethan R. Elenberg +1

    cs.LGcs.CVstat.MLarXiv:2001.10528v42020
  37. ClusterGAN : Latent Space Clustering in Generative Adversarial Networks

    Sudipto Mukherjee, Himanshu Asnani, Eugene Lin +1

    cs.LGstat.MLarXiv:1809.03627v22018
  38. Learning to Remember Rare Events

    Łukasz Kaiser, Ofir Nachum, Aurko Roy +1

    cs.LGarXiv:1703.03129v12017
  39. Generalization Properties of Learning with Random Features

    Alessandro Rudi, Lorenzo Rosasco

    stat.MLcs.LGarXiv:1602.04474v52016
  40. How much data is needed to train a medical image deep learning system to achieve necessary high accuracy?

    Junghwan Cho, Kyewook Lee, Ellie Shin +2

    cs.LGcs.CVcs.NEarXiv:1511.06348v22015
  41. Rethinking Transformer-based Set Prediction for Object Detection

    Zhiqing Sun, Shengcao Cao, Yiming Yang +1

    cs.CVcs.LGarXiv:2011.10881v22020
  42. Implicit Bias of Gradient Descent for Wide Two-layer Neural Networks Trained with the Logistic Loss

    Lenaic Chizat, Francis Bach

    math.OCcs.LGstat.MLarXiv:2002.04486v42020
  43. An approach to reachability analysis for feed-forward ReLU neural networks

    Alessio Lomuscio, Lalit Maganti

    cs.AIcs.LGcs.LOarXiv:1706.07351v12017
  44. Recursive Neural Conditional Random Fields for Aspect-based Sentiment Analysis

    Wenya Wang, Sinno Jialin Pan, Daniel Dahlmeier +1

    cs.CLcs.IRcs.LGarXiv:1603.06679v32016
  45. Discrimination in the Age of Algorithms

    Jon Kleinberg, Jens Ludwig, Sendhil Mullainathan +1

    cs.CYcs.AIcs.LGarXiv:1902.03731v12019
  46. Nonlinear Transform Source-Channel Coding for Semantic Communications

    Jincheng Dai, Sixian Wang, Kailin Tan +4

    cs.ITcs.CVcs.LGarXiv:2112.10961v32021
  47. GRAS: Guided Reduced-Variance Proposals and Adaptive Selection for Training-Free Reward Alignment in Discrete Diffusion

    Kwanyoung Kim

    cs.LGcs.CEq-bio.QMarXiv:2608.26585v12026
  48. catch22: CAnonical Time-series CHaracteristics

    Carl H Lubba, Sarab S Sethi, Philip Knaute +3

    cs.IRcs.LGstat.MLarXiv:1901.10200v22019
  49. Neural Topological SLAM for Visual Navigation

    Devendra Singh Chaplot, Ruslan Salakhutdinov, Abhinav Gupta +1

    cs.CVcs.AIcs.LGarXiv:2005.12256v22020
  50. Autoencoders for Unsupervised Anomaly Segmentation in Brain MR Images: A Comparative Study

    Christoph Baur, Stefan Denner, Benedikt Wiestler +2

    eess.IVcs.CVcs.LGarXiv:2004.03271v22020
  51. FoldPipe: Bounded Remote Streaming of Native Molecular Shards with Asynchronous Prefetch

    Dhiren Mukesh Khatri

    cs.PFcs.LGarXiv:2608.27029v12026
  52. Cross-Lingual Ability of Multilingual BERT: An Empirical Study

    Karthikeyan K, Zihan Wang, Stephen Mayhew +1

    cs.CLcs.AIcs.LGarXiv:1912.07840v22019
  53. When Is the Sharp Covariance Envelope Tight? Feature-Only Geometry for Volume-Sampled Least Squares

    Kihun Rhee

    cs.LGstat.MLarXiv:2608.26877v12026
  54. CheXclusion: Fairness gaps in deep chest X-ray classifiers

    Laleh Seyyed-Kalantari, Guanxiong Liu, Matthew McDermott +2

    cs.CVcs.AIcs.LGarXiv:2003.00827v22020
  55. How good is my GAN?

    Konstantin Shmelkov, Cordelia Schmid, Karteek Alahari

    cs.CVcs.LGarXiv:1807.09499v12018
  56. Cross-Layer Distillation with Semantic Calibration

    Defang Chen, Jian-Ping Mei, Yuan Zhang +3

    cs.CVcs.AIcs.LGarXiv:2012.03236v22020
  57. A Unified Descriptive-Complexity Framework for Model Selection under Correlated Designs

    Yanhang Zhang, Wei Liu, Yuhong Yang

    stat.MLcs.LGarXiv:2608.26618v12026
  58. Active Bias: Training More Accurate Neural Networks by Emphasizing High Variance Samples

    Haw-Shiuan Chang, Erik Learned-Miller, Andrew McCallum

    stat.MLcs.LGarXiv:1704.07433v42017
  59. LaserNet: An Efficient Probabilistic 3D Object Detector for Autonomous Driving

    Gregory P. Meyer, Ankit Laddha, Eric Kee +2

    cs.CVcs.LGcs.ROarXiv:1903.08701v12019
  60. On the Convergence and Robustness of Adversarial Training

    Yisen Wang, Xingjun Ma, James Bailey +3

    cs.LGarXiv:2112.08304v22021