Machine Learning
Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
15,721 to 15,780 of 20,186
Sliced Wasserstein Discrepancy for Unsupervised Domain Adaptation
Chen-Yu Lee, Tanmay Batra, Mohammad Haris Baig +1
cs.CVcs.LGstat.MLarXiv:1903.04064v12019IAPO: Influence-Aware Policy Optimization for Credit Assignment in Multi-Turn Service Agents
Bo Ren, Yirong Mao, Yi Yang +1
cs.LGarXiv:2608.24588v12026BOLD: Dataset and Metrics for Measuring Biases in Open-Ended Language Generation
Jwala Dhamala, Tony Sun, Varun Kumar +4
cs.CLcs.AIcs.LGarXiv:2101.11718v12021Efficient GAN-Based Anomaly Detection
Houssam Zenati, Chuan Sheng Foo, Bruno Lecouat +2
cs.LGstat.MLarXiv:1802.06222v22018Connecting the Dots in Trustworthy Artificial Intelligence: From AI Principles, Ethics, and Key Requirements to Responsible AI Systems and Regulation
Natalia Díaz-Rodríguez, Javier Del Ser, Mark Coeckelbergh +3
cs.CYcs.AIcs.LGarXiv:2305.02231v22023A Comprehensive Survey on Test-Time Adaptation under Distribution Shifts
Jian Liang, Ran He, Tieniu Tan
cs.LGcs.AIcs.CVarXiv:2303.15361v22023Gradient Sparsification for Communication-Efficient Distributed Optimization
Jianqiao Wangni, Jialei Wang, Ji Liu +1
cs.LGmath.NAstat.MLarXiv:1710.09854v12017Mechanistic Circuit Identification for Controllable Data Generation
Nakyung Lee, Sangwoo Hong, Jungwoo Lee
cs.LGcs.AIcs.CLarXiv:2608.24065v12026JailbreakBench: An Open Robustness Benchmark for Jailbreaking Large Language Models
Patrick Chao, Edoardo Debenedetti, Alexander Robey +9
cs.CRcs.LGarXiv:2404.01318v52024Do Transformers Really Perform Bad for Graph Representation?
Chengxuan Ying, Tianle Cai, Shengjie Luo +5
cs.LGcs.AIarXiv:2106.05234v52021Tight Majorizations and Convergence Rates of Nuclear Norm Minimization IRLS
Christian Kümmerle, Tomas Masak, Dominik Stöger
cs.LGmath.NAmath.OCarXiv:2608.23765v12026A Practical Guide to Multi-Objective Reinforcement Learning and Planning
Conor F. Hayes, Roxana Rădulescu, Eugenio Bargiacchi +15
cs.AIcs.LGarXiv:2103.09568v12021Enhancing Computational Fluid Dynamics with Machine Learning
Ricardo Vinuesa, Steven L. Brunton
physics.flu-dyncs.LGphysics.comp-pharXiv:2110.02085v22021Survey of Deep Reinforcement Learning for Motion Planning of Autonomous Vehicles
Szilárd Aradi
cs.LGeess.SYstat.MLarXiv:2001.11231v12020SWAD: Domain Generalization by Seeking Flat Minima
Junbum Cha, Sanghyuk Chun, Kyungjae Lee +4
cs.LGcs.CVarXiv:2102.08604v42021Calibration-Preserving Pruning: Compression as a Reliability Contract
Ibne Farabi Shihab, Adria Binte Habib, Anuj Sharma
cs.LGcs.CLarXiv:2608.23744v12026Q-Align: Teaching LMMs for Visual Scoring via Discrete Text-Defined Levels
Haoning Wu, Zicheng Zhang, Weixia Zhang +11
cs.CVcs.CLcs.LGarXiv:2312.17090v12023Enabling Factorized Piano Music Modeling and Generation with the MAESTRO Dataset
Curtis Hawthorne, Andriy Stasyuk, Adam Roberts +6
cs.SDcs.LGeess.ASarXiv:1810.12247v52018Positive-Unlabeled Learning with Non-Negative Risk Estimator
Ryuichi Kiryo, Gang Niu, Marthinus C. du Plessis +1
cs.LGstat.MLarXiv:1703.00593v22017Mitigating Exploration Bias in RL for Multi-Instruction Following
Mian Zhang, Yueqin Yin, Kaiyu He +4
cs.CLcs.LGarXiv:2608.23830v12026RouteLLM: Learning to Route LLMs with Preference Data
Isaac Ong, Amjad Almahairi, Vincent Wu +5
cs.LGcs.AIcs.CLarXiv:2406.18665v42024AASIST: Audio Anti-Spoofing using Integrated Spectro-Temporal Graph Attention Networks
Jee-weon Jung, Hee-Soo Heo, Hemlata Tak +5
eess.AScs.AIcs.LGarXiv:2110.01200v12021FedKD: Communication Efficient Federated Learning via Knowledge Distillation
Chuhan Wu, Fangzhao Wu, Lingjuan Lyu +2
cs.LGcs.CLarXiv:2108.13323v22021Fast Abstractive Summarization with Reinforce-Selected Sentence Rewriting
Yen-Chun Chen, Mohit Bansal
cs.CLcs.AIcs.LGarXiv:1805.11080v12018Conditional GraphGANFed: Optimizing Graph-Structured Molecule Generation in Federated Generative Adversarial Networks
Daniel Manu, Abee Alazzwi
cs.LGcs.DCarXiv:2608.24610v12026PPINN: Parareal Physics-Informed Neural Network for time-dependent PDEs
Xuhui Meng, Zhen Li, Dongkun Zhang +1
physics.comp-phcs.LGstat.MLarXiv:1909.10145v12019Learning Data Augmentation Strategies for Object Detection
Barret Zoph, Ekin D. Cubuk, Golnaz Ghiasi +3
cs.CVcs.LGarXiv:1906.11172v12019Deep Global Registration
Christopher Choy, Wei Dong, Vladlen Koltun
cs.CVcs.CGcs.LGarXiv:2004.11540v22020An LSTM Network for Highway Trajectory Prediction
Florent Altché, Arnaud de La Fortelle
cs.ROcs.LGarXiv:1801.07962v12018On Loss Functions for Deep Neural Networks in Classification
Katarzyna Janocha, Wojciech Marian Czarnecki
cs.LGarXiv:1702.05659v12017YaRN: Efficient Context Window Extension of Large Language Models
Bowen Peng, Jeffrey Quesnelle, Honglu Fan +1
cs.CLcs.AIcs.LGarXiv:2309.00071v32023Large-scale Multi-view Subspace Clustering in Linear Time
Zhao Kang, Wangtao Zhou, Zhitong Zhao +3
cs.LGcs.CVstat.MLarXiv:1911.09290v12019On Smoothing and Inference for Topic Models
Arthur Asuncion, Max Welling, Padhraic Smyth +1
cs.LGstat.MLarXiv:1205.2662v12012FlowNeg: GFlowNet-Guided Diverse Hard Negative Sampling for Knowledge Graph Embedding
Ibne Farabi Shihab, Naoshin Anzum Hridi, Joyanta Jyoti Mondal
cs.LGarXiv:2608.23849v12026Meta R-CNN : Towards General Solver for Instance-level Few-shot Learning
Xiaopeng Yan, Ziliang Chen, Anni Xu +3
cs.CVcs.LGarXiv:1909.13032v22019Towards End-to-End Prosody Transfer for Expressive Speech Synthesis with Tacotron
RJ Skerry-Ryan, Eric Battenberg, Ying Xiao +6
cs.CLcs.LGcs.SDarXiv:1803.09047v12018Invariance Matters: Exemplar Memory for Domain Adaptive Person Re-identification
Zhun Zhong, Liang Zheng, Zhiming Luo +2
cs.CVcs.LGarXiv:1904.01990v12019A review of domain adaptation without target labels
Wouter M. Kouw, Marco Loog
cs.LGstat.MLarXiv:1901.05335v22019SAITS: Self-Attention-based Imputation for Time Series
Wenjie Du, David Cote, Yan Liu
cs.LGarXiv:2202.08516v52022A study of the effect of JPG compression on adversarial images
Gintare Karolina Dziugaite, Zoubin Ghahramani, Daniel M. Roy
cs.CVcs.LGarXiv:1608.00853v12016Differential Learning for Robust Prediction of Thermal Stability with Application to Energetic Materials
Megan C. Davis, R. Seaton Ullberg, Jeremy N. Schroeder +5
physics.chem-phcond-mat.mtrl-scics.LGarXiv:2608.23874v12026MirrorGAN: Learning Text-to-image Generation by Redescription
Tingting Qiao, Jing Zhang, Duanqing Xu +1
cs.CLcs.CVcs.LGarXiv:1903.05854v12019SMART: Robust and Efficient Fine-Tuning for Pre-trained Natural Language Models through Principled Regularized Optimization
Haoming Jiang, Pengcheng He, Weizhu Chen +3
cs.CLcs.LGmath.OCarXiv:1911.03437v52019On Gradient Descent Ascent for Nonconvex-Concave Minimax Problems
Tianyi Lin, Chi Jin, Michael I. Jordan
cs.LGmath.OCstat.MLarXiv:1906.00331v102019Mapping the Concept Landscape: Structural Perception of Global Distributions for Transparent Data Pruning
Dongyue Wu, Tao Ma
cs.LGcs.CVarXiv:2608.22858v12026End-to-End Text-Dependent Speaker Verification
Georg Heigold, Ignacio Moreno, Samy Bengio +1
cs.LGcs.SDarXiv:1509.08062v12015Clotho: An Audio Captioning Dataset
Konstantinos Drossos, Samuel Lipping, Tuomas Virtanen
cs.SDcs.CLcs.LGarXiv:1910.09387v12019DeMixPert: Decomposed Response Modeling with Gaussian Mixtures for OOD Single-Cell Perturbation Prediction
Jiawen Liu, Xuechenxiao Cao, Yutong Li +5
cs.LGcs.AIarXiv:2608.23114v12026CatchBench: When Can an Agent Failure Be Caught?
Yue Zhao
cs.LGcs.MAcs.PFarXiv:2608.22808v12026GSM-Symbolic: Understanding the Limitations of Mathematical Reasoning in Large Language Models
Iman Mirzadeh, Keivan Alizadeh, Hooman Shahrokhi +3
cs.LGcs.AIarXiv:2410.05229v22024Efficient Machine Learning for Big Data: A Review
O. Y. Al-Jarrah, P. D. Yoo, S Muhaidat +2
cs.LGcs.AIarXiv:1503.05296v12015MotionCtrl: A Unified and Flexible Motion Controller for Video Generation
Zhouxia Wang, Ziyang Yuan, Xintao Wang +4
cs.CVcs.AIcs.LGarXiv:2312.03641v22023On the (In)fidelity and Sensitivity for Explanations
Chih-Kuan Yeh, Cheng-Yu Hsieh, Arun Sai Suggala +2
cs.LGstat.MLarXiv:1901.09392v420193D Steerable CNNs: Learning Rotationally Equivariant Features in Volumetric Data
Maurice Weiler, Mario Geiger, Max Welling +2
cs.LGstat.MLarXiv:1807.02547v22018Exploration in Deep Reinforcement Learning: A Survey
Pawel Ladosz, Lilian Weng, Minwoo Kim +1
cs.LGarXiv:2205.00824v12022Prometheus: Inducing Fine-grained Evaluation Capability in Language Models
Seungone Kim, Jamin Shin, Yejin Cho +8
cs.CLcs.LGarXiv:2310.08491v22023Using Trusted Data to Train Deep Networks on Labels Corrupted by Severe Noise
Dan Hendrycks, Mantas Mazeika, Duncan Wilson +1
cs.LGcs.CLcs.CVarXiv:1802.05300v42018A Formal Methodological Framework for Auditing Robustness and Fidelity in Explainable AI: From Application to Trust Certification
Rosa Elysabeth Ralinirina, Jean Christian Ralaivao, Niaiko Michaël Ralaivao +2
cs.AIcs.CYcs.LGarXiv:2608.23817v12026MolEmb: Multimodal Large Language Models Can Be Strong Molecular Embedding Models
Xinjian Zhao, Xiangru Jian, Yaoyao Xu +4
cs.AIcs.LGarXiv:2608.23646v12026Generative Neural Networks for Sinkhorn Distributionally Robust Hypothesis Testing
Fenglin Zhang, Teyan Liu, Jie Wang
stat.MLcs.LGmath.OCarXiv:2608.22746v12026