Machine Learning
Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
13,501 to 13,560 of 20,219
Common Geodesics Do Not Guarantee Fisher Consistency of the Structured SVM: Minimal Counterexamples and a Tree-Metric Classification
Jintao Fei, Jiangying Luo
cs.LGarXiv:2608.27203v12026Towards better decoding and language model integration in sequence to sequence models
Jan Chorowski, Navdeep Jaitly
cs.NEcs.CLcs.LGarXiv:1612.02695v12016On the Apparent Conflict Between Individual and Group Fairness
Reuben Binns
cs.LGcs.CYstat.MLarXiv:1912.06883v12019Domain-Specific Self-Supervised Representation Learning for Retinal Fundus Classification
Bekzat Nurlanbekova, Fung Fung Ting
cs.CVcs.LGarXiv:2608.26686v12026FairGAN: Fairness-aware Generative Adversarial Networks
Depeng Xu, Shuhan Yuan, Lu Zhang +1
cs.LGcs.CYstat.MLarXiv:1805.11202v12018pix2code: Generating Code from a Graphical User Interface Screenshot
Tony Beltramelli
cs.LGcs.AIcs.CLarXiv:1705.07962v22017SAGE: Variate-Wise Semantic Augmentation for Vision-Language Time Series Forecasting
Haizhao Fan, Xinyi Le
cs.LGcs.CVarXiv:2608.26829v12026Towards End-to-End Speech Recognition with Deep Convolutional Neural Networks
Ying Zhang, Mohammad Pezeshki, Philemon Brakel +3
cs.CLcs.LGstat.MLarXiv:1701.02720v12017Information Leakage in Embedding Models
Congzheng Song, Ananth Raghunathan
cs.LGcs.CLcs.CRarXiv:2004.00053v22020The Why and How of Nonnegative Matrix Factorization
Nicolas Gillis
stat.MLcs.IRcs.LGarXiv:1401.5226v22014Tabular Deep Learning for Algorithmic Trading: Cross-Regime Bayesian Optimisation for Equity Signal Generation
Joshua Le Grice
cs.LGq-fin.CPq-fin.TRarXiv:2608.27076v12026CrossFormer: A Versatile Vision Transformer Hinging on Cross-scale Attention
Wenxiao Wang, Lu Yao, Long Chen +4
cs.CVcs.LGarXiv:2108.00154v22021ARCH: Animatable Reconstruction of Clothed Humans
Zeng Huang, Yuanlu Xu, Christoph Lassner +2
cs.GRcs.CVcs.LGarXiv:2004.04572v22020Disentangling Optimization Scale from Preference Scale in DPO
Ivan Kruzhilov
cs.LGarXiv:2608.27032v12026Online Structured Laplace Approximations For Overcoming Catastrophic Forgetting
Hippolyt Ritter, Aleksandar Botev, David Barber
stat.MLcs.LGarXiv:1805.07810v12018Multiple Futures Prediction
Yichuan Charlie Tang, Ruslan Salakhutdinov
cs.LGcs.CVcs.MAarXiv:1911.00997v22019Spatially Adaptive Computation Time for Residual Networks
Michael Figurnov, Maxwell D. Collins, Yukun Zhu +4
cs.CVcs.LGarXiv:1612.02297v22016Aligning Domain-specific Distribution and Classifier for Cross-domain Classification from Multiple Sources
Yongchun Zhu, Fuzhen Zhuang, Deqing Wang
cs.LGcs.AIcs.CVarXiv:2201.01003v12022ClusterAttention: A training-free speedup of bidirectional attention
Kasper Nordenram, Amelie Dittmann
cs.LGcs.CVarXiv:2608.26965v12026Permutation Invariant Graph Generation via Score-Based Generative Modeling
Chenhao Niu, Yang Song, Jiaming Song +3
cs.LGstat.MLarXiv:2003.00638v12020S-Prompts Learning with Pre-trained Transformers: An Occam's Razor for Domain Incremental Learning
Yabin Wang, Zhiwu Huang, Xiaopeng Hong
cs.CVcs.LGarXiv:2207.12819v22022A Layer Importance Metric for Quantization Accounting for the Speed-Quality Trade-off in Autoregressive Models
Artem Safronov
cs.LGarXiv:2608.26926v12026Gradient Matching for Domain Generalization
Yuge Shi, Jeffrey Seely, Philip H. S. Torr +4
cs.LGstat.MLarXiv:2104.09937v32021On the Convergence of A Class of Adam-Type Algorithms for Non-Convex Optimization
Xiangyi Chen, Sijia Liu, Ruoyu Sun +1
cs.LGmath.OCstat.MLarXiv:1808.02941v22018Beyond Client Averaging: A Client-Independent Second-Order Stationary-Bias Component in Stochastic SCAFFOLD
Yi-Ping Tang, Guan-Ju Peng
cs.LGmath.STarXiv:2608.26765v12026A Unified Analysis of Extra-gradient and Optimistic Gradient Methods for Saddle Point Problems: Proximal Point Approach
Aryan Mokhtari, Asuman Ozdaglar, Sarath Pattathil
math.OCcs.LGstat.MLarXiv:1901.08511v42019Exploring the Landscape of Spatial Robustness
Logan Engstrom, Brandon Tran, Dimitris Tsipras +2
cs.LGcs.CVcs.NEarXiv:1712.02779v42017Flexibly Fair Representation Learning by Disentanglement
Elliot Creager, David Madras, Jörn-Henrik Jacobsen +4
cs.LGcs.AIstat.MLarXiv:1906.02589v12019Distillation-Based Semi-Supervised Federated Learning for Communication-Efficient Collaborative Training with Non-IID Private Data
Sohei Itahara, Takayuki Nishio, Yusuke Koda +2
cs.DCcs.LGarXiv:2008.06180v22020Discrete Graph Structure Learning for Forecasting Multiple Time Series
Chao Shang, Jie Chen, Jinbo Bi
cs.LGstat.MLarXiv:2101.06861v32021Attention-based Graph Neural Network for Semi-supervised Learning
Kiran K. Thekumparampil, Chong Wang, Sewoong Oh +1
stat.MLcs.AIcs.LGarXiv:1803.03735v12018Block-Coordinate Frank-Wolfe Optimization for Structural SVMs
Simon Lacoste-Julien, Martin Jaggi, Mark Schmidt +1
cs.LGmath.OCstat.MLarXiv:1207.4747v42012Why ResNet Works? Residuals Generalize
Fengxiang He, Tongliang Liu, Dacheng Tao
stat.MLcs.LGarXiv:1904.01367v12019Enhanced Membership Inference Attacks against Machine Learning Models
Jiayuan Ye, Aadyaa Maddi, Sasi Kumar Murakonda +2
cs.LGcs.CRstat.MLarXiv:2111.09679v42021Axiom-based Grad-CAM: Towards Accurate Visualization and Explanation of CNNs
Ruigang Fu, Qingyong Hu, Xiaohu Dong +3
cs.CVcs.AIcs.LGarXiv:2008.02312v42020Identifying Mislabeled Data using the Area Under the Margin Ranking
Geoff Pleiss, Tianyi Zhang, Ethan R. Elenberg +1
cs.LGcs.CVstat.MLarXiv:2001.10528v42020ClusterGAN : Latent Space Clustering in Generative Adversarial Networks
Sudipto Mukherjee, Himanshu Asnani, Eugene Lin +1
cs.LGstat.MLarXiv:1809.03627v22018Learning to Remember Rare Events
Łukasz Kaiser, Ofir Nachum, Aurko Roy +1
cs.LGarXiv:1703.03129v12017Generalization Properties of Learning with Random Features
Alessandro Rudi, Lorenzo Rosasco
stat.MLcs.LGarXiv:1602.04474v52016How much data is needed to train a medical image deep learning system to achieve necessary high accuracy?
Junghwan Cho, Kyewook Lee, Ellie Shin +2
cs.LGcs.CVcs.NEarXiv:1511.06348v22015Rethinking Transformer-based Set Prediction for Object Detection
Zhiqing Sun, Shengcao Cao, Yiming Yang +1
cs.CVcs.LGarXiv:2011.10881v22020Implicit Bias of Gradient Descent for Wide Two-layer Neural Networks Trained with the Logistic Loss
Lenaic Chizat, Francis Bach
math.OCcs.LGstat.MLarXiv:2002.04486v42020An approach to reachability analysis for feed-forward ReLU neural networks
Alessio Lomuscio, Lalit Maganti
cs.AIcs.LGcs.LOarXiv:1706.07351v12017Recursive Neural Conditional Random Fields for Aspect-based Sentiment Analysis
Wenya Wang, Sinno Jialin Pan, Daniel Dahlmeier +1
cs.CLcs.IRcs.LGarXiv:1603.06679v32016Discrimination in the Age of Algorithms
Jon Kleinberg, Jens Ludwig, Sendhil Mullainathan +1
cs.CYcs.AIcs.LGarXiv:1902.03731v12019Nonlinear Transform Source-Channel Coding for Semantic Communications
Jincheng Dai, Sixian Wang, Kailin Tan +4
cs.ITcs.CVcs.LGarXiv:2112.10961v32021GRAS: Guided Reduced-Variance Proposals and Adaptive Selection for Training-Free Reward Alignment in Discrete Diffusion
Kwanyoung Kim
cs.LGcs.CEq-bio.QMarXiv:2608.26585v12026catch22: CAnonical Time-series CHaracteristics
Carl H Lubba, Sarab S Sethi, Philip Knaute +3
cs.IRcs.LGstat.MLarXiv:1901.10200v22019Neural Topological SLAM for Visual Navigation
Devendra Singh Chaplot, Ruslan Salakhutdinov, Abhinav Gupta +1
cs.CVcs.AIcs.LGarXiv:2005.12256v22020Autoencoders for Unsupervised Anomaly Segmentation in Brain MR Images: A Comparative Study
Christoph Baur, Stefan Denner, Benedikt Wiestler +2
eess.IVcs.CVcs.LGarXiv:2004.03271v22020FoldPipe: Bounded Remote Streaming of Native Molecular Shards with Asynchronous Prefetch
Dhiren Mukesh Khatri
cs.PFcs.LGarXiv:2608.27029v12026Cross-Lingual Ability of Multilingual BERT: An Empirical Study
Karthikeyan K, Zihan Wang, Stephen Mayhew +1
cs.CLcs.AIcs.LGarXiv:1912.07840v22019When Is the Sharp Covariance Envelope Tight? Feature-Only Geometry for Volume-Sampled Least Squares
Kihun Rhee
cs.LGstat.MLarXiv:2608.26877v12026CheXclusion: Fairness gaps in deep chest X-ray classifiers
Laleh Seyyed-Kalantari, Guanxiong Liu, Matthew McDermott +2
cs.CVcs.AIcs.LGarXiv:2003.00827v22020How good is my GAN?
Konstantin Shmelkov, Cordelia Schmid, Karteek Alahari
cs.CVcs.LGarXiv:1807.09499v12018Cross-Layer Distillation with Semantic Calibration
Defang Chen, Jian-Ping Mei, Yuan Zhang +3
cs.CVcs.AIcs.LGarXiv:2012.03236v22020A Unified Descriptive-Complexity Framework for Model Selection under Correlated Designs
Yanhang Zhang, Wei Liu, Yuhong Yang
stat.MLcs.LGarXiv:2608.26618v12026Active Bias: Training More Accurate Neural Networks by Emphasizing High Variance Samples
Haw-Shiuan Chang, Erik Learned-Miller, Andrew McCallum
stat.MLcs.LGarXiv:1704.07433v42017LaserNet: An Efficient Probabilistic 3D Object Detector for Autonomous Driving
Gregory P. Meyer, Ankit Laddha, Eric Kee +2
cs.CVcs.LGcs.ROarXiv:1903.08701v12019On the Convergence and Robustness of Adversarial Training
Yisen Wang, Xingjun Ma, James Bailey +3
cs.LGarXiv:2112.08304v22021