Machine Learning

Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

15,721 to 15,780 of 20,186

  1. Sliced Wasserstein Discrepancy for Unsupervised Domain Adaptation

    Chen-Yu Lee, Tanmay Batra, Mohammad Haris Baig +1

    cs.CVcs.LGstat.MLarXiv:1903.04064v12019
  2. IAPO: Influence-Aware Policy Optimization for Credit Assignment in Multi-Turn Service Agents

    Bo Ren, Yirong Mao, Yi Yang +1

    cs.LGarXiv:2608.24588v12026
  3. BOLD: Dataset and Metrics for Measuring Biases in Open-Ended Language Generation

    Jwala Dhamala, Tony Sun, Varun Kumar +4

    cs.CLcs.AIcs.LGarXiv:2101.11718v12021
  4. Efficient GAN-Based Anomaly Detection

    Houssam Zenati, Chuan Sheng Foo, Bruno Lecouat +2

    cs.LGstat.MLarXiv:1802.06222v22018
  5. Connecting the Dots in Trustworthy Artificial Intelligence: From AI Principles, Ethics, and Key Requirements to Responsible AI Systems and Regulation

    Natalia Díaz-Rodríguez, Javier Del Ser, Mark Coeckelbergh +3

    cs.CYcs.AIcs.LGarXiv:2305.02231v22023
  6. A Comprehensive Survey on Test-Time Adaptation under Distribution Shifts

    Jian Liang, Ran He, Tieniu Tan

    cs.LGcs.AIcs.CVarXiv:2303.15361v22023
  7. Gradient Sparsification for Communication-Efficient Distributed Optimization

    Jianqiao Wangni, Jialei Wang, Ji Liu +1

    cs.LGmath.NAstat.MLarXiv:1710.09854v12017
  8. Mechanistic Circuit Identification for Controllable Data Generation

    Nakyung Lee, Sangwoo Hong, Jungwoo Lee

    cs.LGcs.AIcs.CLarXiv:2608.24065v12026
  9. JailbreakBench: An Open Robustness Benchmark for Jailbreaking Large Language Models

    Patrick Chao, Edoardo Debenedetti, Alexander Robey +9

    cs.CRcs.LGarXiv:2404.01318v52024
  10. Do Transformers Really Perform Bad for Graph Representation?

    Chengxuan Ying, Tianle Cai, Shengjie Luo +5

    cs.LGcs.AIarXiv:2106.05234v52021
  11. Tight Majorizations and Convergence Rates of Nuclear Norm Minimization IRLS

    Christian Kümmerle, Tomas Masak, Dominik Stöger

    cs.LGmath.NAmath.OCarXiv:2608.23765v12026
  12. A Practical Guide to Multi-Objective Reinforcement Learning and Planning

    Conor F. Hayes, Roxana Rădulescu, Eugenio Bargiacchi +15

    cs.AIcs.LGarXiv:2103.09568v12021
  13. Enhancing Computational Fluid Dynamics with Machine Learning

    Ricardo Vinuesa, Steven L. Brunton

    physics.flu-dyncs.LGphysics.comp-pharXiv:2110.02085v22021
  14. Survey of Deep Reinforcement Learning for Motion Planning of Autonomous Vehicles

    Szilárd Aradi

    cs.LGeess.SYstat.MLarXiv:2001.11231v12020
  15. SWAD: Domain Generalization by Seeking Flat Minima

    Junbum Cha, Sanghyuk Chun, Kyungjae Lee +4

    cs.LGcs.CVarXiv:2102.08604v42021
  16. Calibration-Preserving Pruning: Compression as a Reliability Contract

    Ibne Farabi Shihab, Adria Binte Habib, Anuj Sharma

    cs.LGcs.CLarXiv:2608.23744v12026
  17. Q-Align: Teaching LMMs for Visual Scoring via Discrete Text-Defined Levels

    Haoning Wu, Zicheng Zhang, Weixia Zhang +11

    cs.CVcs.CLcs.LGarXiv:2312.17090v12023
  18. Enabling Factorized Piano Music Modeling and Generation with the MAESTRO Dataset

    Curtis Hawthorne, Andriy Stasyuk, Adam Roberts +6

    cs.SDcs.LGeess.ASarXiv:1810.12247v52018
  19. Positive-Unlabeled Learning with Non-Negative Risk Estimator

    Ryuichi Kiryo, Gang Niu, Marthinus C. du Plessis +1

    cs.LGstat.MLarXiv:1703.00593v22017
  20. Mitigating Exploration Bias in RL for Multi-Instruction Following

    Mian Zhang, Yueqin Yin, Kaiyu He +4

    cs.CLcs.LGarXiv:2608.23830v12026
  21. RouteLLM: Learning to Route LLMs with Preference Data

    Isaac Ong, Amjad Almahairi, Vincent Wu +5

    cs.LGcs.AIcs.CLarXiv:2406.18665v42024
  22. AASIST: Audio Anti-Spoofing using Integrated Spectro-Temporal Graph Attention Networks

    Jee-weon Jung, Hee-Soo Heo, Hemlata Tak +5

    eess.AScs.AIcs.LGarXiv:2110.01200v12021
  23. FedKD: Communication Efficient Federated Learning via Knowledge Distillation

    Chuhan Wu, Fangzhao Wu, Lingjuan Lyu +2

    cs.LGcs.CLarXiv:2108.13323v22021
  24. Fast Abstractive Summarization with Reinforce-Selected Sentence Rewriting

    Yen-Chun Chen, Mohit Bansal

    cs.CLcs.AIcs.LGarXiv:1805.11080v12018
  25. Conditional GraphGANFed: Optimizing Graph-Structured Molecule Generation in Federated Generative Adversarial Networks

    Daniel Manu, Abee Alazzwi

    cs.LGcs.DCarXiv:2608.24610v12026
  26. PPINN: Parareal Physics-Informed Neural Network for time-dependent PDEs

    Xuhui Meng, Zhen Li, Dongkun Zhang +1

    physics.comp-phcs.LGstat.MLarXiv:1909.10145v12019
  27. Learning Data Augmentation Strategies for Object Detection

    Barret Zoph, Ekin D. Cubuk, Golnaz Ghiasi +3

    cs.CVcs.LGarXiv:1906.11172v12019
  28. Deep Global Registration

    Christopher Choy, Wei Dong, Vladlen Koltun

    cs.CVcs.CGcs.LGarXiv:2004.11540v22020
  29. An LSTM Network for Highway Trajectory Prediction

    Florent Altché, Arnaud de La Fortelle

    cs.ROcs.LGarXiv:1801.07962v12018
  30. On Loss Functions for Deep Neural Networks in Classification

    Katarzyna Janocha, Wojciech Marian Czarnecki

    cs.LGarXiv:1702.05659v12017
  31. YaRN: Efficient Context Window Extension of Large Language Models

    Bowen Peng, Jeffrey Quesnelle, Honglu Fan +1

    cs.CLcs.AIcs.LGarXiv:2309.00071v32023
  32. Large-scale Multi-view Subspace Clustering in Linear Time

    Zhao Kang, Wangtao Zhou, Zhitong Zhao +3

    cs.LGcs.CVstat.MLarXiv:1911.09290v12019
  33. On Smoothing and Inference for Topic Models

    Arthur Asuncion, Max Welling, Padhraic Smyth +1

    cs.LGstat.MLarXiv:1205.2662v12012
  34. FlowNeg: GFlowNet-Guided Diverse Hard Negative Sampling for Knowledge Graph Embedding

    Ibne Farabi Shihab, Naoshin Anzum Hridi, Joyanta Jyoti Mondal

    cs.LGarXiv:2608.23849v12026
  35. Meta R-CNN : Towards General Solver for Instance-level Few-shot Learning

    Xiaopeng Yan, Ziliang Chen, Anni Xu +3

    cs.CVcs.LGarXiv:1909.13032v22019
  36. Towards End-to-End Prosody Transfer for Expressive Speech Synthesis with Tacotron

    RJ Skerry-Ryan, Eric Battenberg, Ying Xiao +6

    cs.CLcs.LGcs.SDarXiv:1803.09047v12018
  37. Invariance Matters: Exemplar Memory for Domain Adaptive Person Re-identification

    Zhun Zhong, Liang Zheng, Zhiming Luo +2

    cs.CVcs.LGarXiv:1904.01990v12019
  38. A review of domain adaptation without target labels

    Wouter M. Kouw, Marco Loog

    cs.LGstat.MLarXiv:1901.05335v22019
  39. SAITS: Self-Attention-based Imputation for Time Series

    Wenjie Du, David Cote, Yan Liu

    cs.LGarXiv:2202.08516v52022
  40. A study of the effect of JPG compression on adversarial images

    Gintare Karolina Dziugaite, Zoubin Ghahramani, Daniel M. Roy

    cs.CVcs.LGarXiv:1608.00853v12016
  41. Differential Learning for Robust Prediction of Thermal Stability with Application to Energetic Materials

    Megan C. Davis, R. Seaton Ullberg, Jeremy N. Schroeder +5

    physics.chem-phcond-mat.mtrl-scics.LGarXiv:2608.23874v12026
  42. MirrorGAN: Learning Text-to-image Generation by Redescription

    Tingting Qiao, Jing Zhang, Duanqing Xu +1

    cs.CLcs.CVcs.LGarXiv:1903.05854v12019
  43. SMART: Robust and Efficient Fine-Tuning for Pre-trained Natural Language Models through Principled Regularized Optimization

    Haoming Jiang, Pengcheng He, Weizhu Chen +3

    cs.CLcs.LGmath.OCarXiv:1911.03437v52019
  44. On Gradient Descent Ascent for Nonconvex-Concave Minimax Problems

    Tianyi Lin, Chi Jin, Michael I. Jordan

    cs.LGmath.OCstat.MLarXiv:1906.00331v102019
  45. Mapping the Concept Landscape: Structural Perception of Global Distributions for Transparent Data Pruning

    Dongyue Wu, Tao Ma

    cs.LGcs.CVarXiv:2608.22858v12026
  46. End-to-End Text-Dependent Speaker Verification

    Georg Heigold, Ignacio Moreno, Samy Bengio +1

    cs.LGcs.SDarXiv:1509.08062v12015
  47. Clotho: An Audio Captioning Dataset

    Konstantinos Drossos, Samuel Lipping, Tuomas Virtanen

    cs.SDcs.CLcs.LGarXiv:1910.09387v12019
  48. DeMixPert: Decomposed Response Modeling with Gaussian Mixtures for OOD Single-Cell Perturbation Prediction

    Jiawen Liu, Xuechenxiao Cao, Yutong Li +5

    cs.LGcs.AIarXiv:2608.23114v12026
  49. CatchBench: When Can an Agent Failure Be Caught?

    Yue Zhao

    cs.LGcs.MAcs.PFarXiv:2608.22808v12026
  50. GSM-Symbolic: Understanding the Limitations of Mathematical Reasoning in Large Language Models

    Iman Mirzadeh, Keivan Alizadeh, Hooman Shahrokhi +3

    cs.LGcs.AIarXiv:2410.05229v22024
  51. Efficient Machine Learning for Big Data: A Review

    O. Y. Al-Jarrah, P. D. Yoo, S Muhaidat +2

    cs.LGcs.AIarXiv:1503.05296v12015
  52. MotionCtrl: A Unified and Flexible Motion Controller for Video Generation

    Zhouxia Wang, Ziyang Yuan, Xintao Wang +4

    cs.CVcs.AIcs.LGarXiv:2312.03641v22023
  53. On the (In)fidelity and Sensitivity for Explanations

    Chih-Kuan Yeh, Cheng-Yu Hsieh, Arun Sai Suggala +2

    cs.LGstat.MLarXiv:1901.09392v42019
  54. 3D Steerable CNNs: Learning Rotationally Equivariant Features in Volumetric Data

    Maurice Weiler, Mario Geiger, Max Welling +2

    cs.LGstat.MLarXiv:1807.02547v22018
  55. Exploration in Deep Reinforcement Learning: A Survey

    Pawel Ladosz, Lilian Weng, Minwoo Kim +1

    cs.LGarXiv:2205.00824v12022
  56. Prometheus: Inducing Fine-grained Evaluation Capability in Language Models

    Seungone Kim, Jamin Shin, Yejin Cho +8

    cs.CLcs.LGarXiv:2310.08491v22023
  57. Using Trusted Data to Train Deep Networks on Labels Corrupted by Severe Noise

    Dan Hendrycks, Mantas Mazeika, Duncan Wilson +1

    cs.LGcs.CLcs.CVarXiv:1802.05300v42018
  58. A Formal Methodological Framework for Auditing Robustness and Fidelity in Explainable AI: From Application to Trust Certification

    Rosa Elysabeth Ralinirina, Jean Christian Ralaivao, Niaiko Michaël Ralaivao +2

    cs.AIcs.CYcs.LGarXiv:2608.23817v12026
  59. MolEmb: Multimodal Large Language Models Can Be Strong Molecular Embedding Models

    Xinjian Zhao, Xiangru Jian, Yaoyao Xu +4

    cs.AIcs.LGarXiv:2608.23646v12026
  60. Generative Neural Networks for Sinkhorn Distributionally Robust Hypothesis Testing

    Fenglin Zhang, Teyan Liu, Jie Wang

    stat.MLcs.LGmath.OCarXiv:2608.22746v12026