Machine Learning

Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

11,101 to 11,160 of 20,063

  1. Understanding Alternating Minimization for Matrix Completion

    Moritz Hardt

    cs.LGcs.DSstat.MLarXiv:1312.0925v42013
  2. Language Model Tokenizers Introduce Unfairness Between Languages

    Aleksandar Petrov, Emanuele La Malfa, Philip H. S. Torr +1

    cs.CLcs.LGarXiv:2305.15425v22023
  3. Joint Distribution Matters: Deep Brownian Distance Covariance for Few-Shot Classification

    Jiangtao Xie, Fei Long, Jiaming Lv +2

    cs.CVcs.LGstat.MLarXiv:2204.04567v12022
  4. PHiSeg: Capturing Uncertainty in Medical Image Segmentation

    Christian F. Baumgartner, Kerem C. Tezcan, Krishna Chaitanya +6

    eess.IVcs.LGstat.MLarXiv:1906.04045v22019
  5. BanglaMed-QA: A Question Answering System for Healthcare Support in Bangla

    Rowzatul Zannat, Abdullah Al Shafi, K. M. Azharul Hasan +1

    cs.CLcs.AIcs.LGarXiv:2608.28329v12026
  6. A Survey on Data Selection for Language Models

    Alon Albalak, Yanai Elazar, Sang Michael Xie +11

    cs.CLcs.LGarXiv:2402.16827v32024
  7. Knowledge Tracing with Sequential Key-Value Memory Networks

    Ghodai Abdelrahman, Qing Wang

    cs.LGcs.AIcs.IRarXiv:1910.13197v12019
  8. DeepReach: A Deep Learning Approach to High-Dimensional Reachability

    Somil Bansal, Claire Tomlin

    cs.ROcs.AIcs.LGarXiv:2011.02082v12020
  9. Tensor Completion Algorithms in Big Data Analytics

    Qingquan Song, Hancheng Ge, James Caverlee +1

    stat.MLcs.AIcs.LGarXiv:1711.10105v22017
  10. Explaining Deep Neural Networks with a Polynomial Time Algorithm for Shapley Values Approximation

    Marco Ancona, Cengiz Öztireli, Markus Gross

    cs.LGstat.MLarXiv:1903.10992v42019
  11. DeepNet: Scaling Transformers to 1,000 Layers

    Hongyu Wang, Shuming Ma, Li Dong +3

    cs.CLcs.LGarXiv:2203.00555v12022
  12. NATTACK: Learning the Distributions of Adversarial Examples for an Improved Black-Box Attack on Deep Neural Networks

    Yandong Li, Lijun Li, Liqiang Wang +2

    cs.LGcs.CRcs.CVarXiv:1905.00441v32019
  13. RecVAE: a New Variational Autoencoder for Top-N Recommendations with Implicit Feedback

    Ilya Shenbin, Anton Alekseev, Elena Tutubalina +2

    cs.IRcs.LGarXiv:1912.11160v12019
  14. DKM: Dense Kernelized Feature Matching for Geometry Estimation

    Johan Edstedt, Ioannis Athanasiadis, Mårten Wadenbäck +1

    cs.CVcs.LGarXiv:2202.00667v32022
  15. Deriving Scaling Laws for OpenEuroLLM Models: Learning Rate, Batch Size and Loss

    Niccolò Ajroldi, Diana Alexandra Onutu, Haider Al-Tahan +4

    cs.LGcs.AIarXiv:2608.28308v12026
  16. BIGPATENT: A Large-Scale Dataset for Abstractive and Coherent Summarization

    Eva Sharma, Chen Li, Lu Wang

    cs.CLcs.LGarXiv:1906.03741v12019
  17. Machine Learning Techniques for Biomedical Image Segmentation: An Overview of Technical Aspects and Introduction to State-of-Art Applications

    Hyunseok Seo, Masoud Badiei Khuzani, Varun Vasudevan +5

    eess.IVcs.CVcs.LGarXiv:1911.02521v12019
  18. Federated Learning for Internet of Things: Applications, Challenges, and Opportunities

    Tuo Zhang, Lei Gao, Chaoyang He +3

    cs.LGarXiv:2111.07494v42021
  19. Learning from All Vehicles

    Dian Chen, Philipp Krähenbühl

    cs.ROcs.CVcs.LGarXiv:2203.11934v32022
  20. RLHF Workflow: From Reward Modeling to Online RLHF

    Hanze Dong, Wei Xiong, Bo Pang +7

    cs.LGcs.AIcs.CLarXiv:2405.07863v32024
  21. Deep Deterministic Uncertainty: A Simple Baseline

    Jishnu Mukhoti, Andreas Kirsch, Joost van Amersfoort +2

    cs.LGstat.MLarXiv:2102.11582v32021
  22. VISTA: Verifier-Informed Student-to-Teacher Adaptation for On-Policy Self-Distillation

    Zewen Ding, Zezhong Wu, Zhou Tao +5

    cs.LGcs.AIcs.CLarXiv:2608.28306v12026
  23. Private Stochastic Convex Optimization with Optimal Rates

    Raef Bassily, Vitaly Feldman, Kunal Talwar +1

    cs.LGcs.CRcs.DSarXiv:1908.09970v12019
  24. An Introductory Survey on Attention Mechanisms in NLP Problems

    Dichao Hu

    cs.CLcs.LGstat.MLarXiv:1811.05544v12018
  25. Lifelong Federated Reinforcement Learning: A Learning Architecture for Navigation in Cloud Robotic Systems

    Boyi Liu, Lujia Wang, Ming Liu

    cs.ROcs.AIcs.DCarXiv:1901.06455v32019
  26. Multi-objective Evolutionary Federated Learning

    Hangyu Zhu, Yaochu Jin

    cs.LGcs.AIstat.MLarXiv:1812.07478v22018
  27. Conformal Risk-Averse Decision Making with Optimized Certainty Equivalent Risk Control

    Amirmohammad Farzaneh, Osvaldo Simeone

    stat.MLcs.AIcs.ITarXiv:2608.28179v12026
  28. Leveraging BERT for Extractive Text Summarization on Lectures

    Derek Miller

    cs.CLcs.LGcs.SDarXiv:1906.04165v12019
  29. DRACO: Byzantine-resilient Distributed Training via Redundant Gradients

    Lingjiao Chen, Hongyi Wang, Zachary Charles +1

    stat.MLcs.DCcs.ITarXiv:1803.09877v42018
  30. OneLLM: One Framework to Align All Modalities with Language

    Jiaming Han, Kaixiong Gong, Yiyuan Zhang +6

    cs.CVcs.AIcs.CLarXiv:2312.03700v22023
  31. Generalize a Small Pre-trained Model to Arbitrarily Large TSP Instances

    Zhang-Hua Fu, Kai-Bin Qiu, Hongyuan Zha

    cs.LGarXiv:2012.10658v22020
  32. Adversarially Robust Distillation

    Micah Goldblum, Liam Fowl, Soheil Feizi +1

    cs.LGcs.CVstat.MLarXiv:1905.09747v22019
  33. On the Universality of Invariant Networks

    Haggai Maron, Ethan Fetaya, Nimrod Segol +1

    cs.LGstat.MLarXiv:1901.09342v42019
  34. VICT: Verifier-Instrumented Credit Tracing for Long-Horizon LLM Agent Reinforcement Learning

    Pengcheng Li, Zhengyang Zhang, Dongxu Zhang +2

    cs.LGcs.AIarXiv:2608.28128v12026
  35. Performative Privacy: When Differential Privacy Maximizes Utility

    Uddalak Mukherjee, Edwige Cyffers, Yann Chevaleyre

    cs.LGcs.AIstat.MLarXiv:2608.28198v12026
  36. COVID-19 Cough Classification using Machine Learning and Global Smartphone Recordings

    Madhurananda Pahar, Marisa Klopper, Robin Warren +1

    cs.SDcs.LGeess.ASarXiv:2012.01926v22020
  37. Beyond Flat Netlist: Hierarchical Graph Representation Learning for Scalable Analysis of Sequential Circuits

    Jingyi Zhou, Zhengyuan Shi, Jiaying Zhu +2

    cs.LGcs.AIcs.ARarXiv:2608.28188v12026
  38. Effectively Unbiased FID and Inception Score and where to find them

    Min Jin Chong, David Forsyth

    cs.CVcs.LGarXiv:1911.07023v32019
  39. Improve Unsupervised Domain Adaptation with Mixup Training

    Shen Yan, Huan Song, Nanxiang Li +2

    stat.MLcs.CVcs.LGarXiv:2001.00677v12020
  40. Safety Verification and Robustness Analysis of Neural Networks via Quadratic Constraints and Semidefinite Programming

    Mahyar Fazlyab, Manfred Morari, George J. Pappas

    math.OCcs.LGarXiv:1903.01287v32019
  41. OccWorld: Learning a 3D Occupancy World Model for Autonomous Driving

    Wenzhao Zheng, Weiliang Chen, Yuanhui Huang +3

    cs.CVcs.AIcs.LGarXiv:2311.16038v12023
  42. The Approximation Rank of Softmax Attention: Sharp Geometric Laws and Robust Interaction Dimension

    Yuhe Sui, Jianing Zhang

    cs.LGcs.AIarXiv:2608.28150v12026
  43. Understanding Self-Training for Gradual Domain Adaptation

    Ananya Kumar, Tengyu Ma, Percy Liang

    cs.LGstat.MLarXiv:2002.11361v12020
  44. Functional Variational Bayesian Neural Networks

    Shengyang Sun, Guodong Zhang, Jiaxin Shi +1

    cs.LGstat.MLarXiv:1903.05779v12019
  45. An Exponential Learning Rate Schedule for Deep Learning

    Zhiyuan Li, Sanjeev Arora

    cs.LGstat.MLarXiv:1910.07454v32019
  46. SoftMatch: Addressing the Quantity-Quality Trade-off in Semi-supervised Learning

    Hao Chen, Ran Tao, Yue Fan +6

    cs.LGcs.AIcs.CVarXiv:2301.10921v22023
  47. Cross-Batch Memory for Embedding Learning

    Xun Wang, Haozhi Zhang, Weilin Huang +1

    cs.LGcs.CVarXiv:1912.06798v32019
  48. CheXtriev: Anatomy-Centered Representation for Case-Based Retrieval of Chest Radiographs

    Naren Akash, Arihanth Tadanki, Jayanthi Sivaswamy

    eess.IVcs.AIcs.CVarXiv:2608.28137v12026
  49. Style Aligned Image Generation via Shared Attention

    Amir Hertz, Andrey Voynov, Shlomi Fruchter +1

    cs.CVcs.GRcs.LGarXiv:2312.02133v22023
  50. DeltaGrad: Rapid retraining of machine learning models

    Yinjun Wu, Edgar Dobriban, Susan B. Davidson

    cs.LGstat.MLarXiv:2006.14755v22020
  51. Car Detection using Unmanned Aerial Vehicles: Comparison between Faster R-CNN and YOLOv3

    Bilel Benjdira, Taha Khursheed, Anis Koubaa +2

    cs.ROcs.CVcs.LGarXiv:1812.10968v12018
  52. MiME: Multilevel Medical Embedding of Electronic Health Records for Predictive Healthcare

    Edward Choi, Cao Xiao, Walter F. Stewart +1

    cs.LGcs.CLstat.MLarXiv:1810.09593v12018
  53. Towards End-to-End Learning for Dialog State Tracking and Management using Deep Reinforcement Learning

    Tiancheng Zhao, Maxine Eskenazi

    cs.AIcs.CLcs.LGarXiv:1606.02560v22016
  54. Community detection in node-attributed social networks: a survey

    Petr Chunaev

    cs.SIcs.LGcs.PFarXiv:1912.09816v22019
  55. Self-supervised learning methods and applications in medical imaging analysis: A survey

    Saeed Shurrab, Rehab Duwairi

    eess.IVcs.CVcs.LGarXiv:2109.08685v32021
  56. Distilling Reasoning Capabilities into Smaller Language Models

    Kumar Shridhar, Alessandro Stolfo, Mrinmaya Sachan

    cs.LGcs.CLarXiv:2212.00193v22022
  57. Re-imagining Algorithmic Fairness in India and Beyond

    Nithya Sambasivan, Erin Arnesen, Ben Hutchinson +2

    cs.CYcs.AIcs.CLarXiv:2101.09995v22021
  58. Extremely Simple Activation Shaping for Out-of-Distribution Detection

    Andrija Djurisic, Nebojsa Bozanic, Arjun Ashok +1

    cs.LGcs.CVarXiv:2209.09858v22022
  59. Do Medical Vision Models Reason About Anatomy? Probing the Spatial Inductive Biases of Learned Visual Representations

    Naren Akash, Neeraja Ramanan

    eess.IVcs.AIcs.CVarXiv:2608.28092v12026
  60. Deep Learning for Community Detection: Progress, Challenges and Opportunities

    Fanzhen Liu, Shan Xue, Jia Wu +6

    cs.SIcs.AIcs.LGarXiv:2005.08225v22020