Machine Learning

Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

8,761 to 8,820 of 20,193

  1. DeepRMSA: A Deep Reinforcement Learning Framework for Routing, Modulation and Spectrum Assignment in Elastic Optical Networks

    Xiaoliang Chen, Baojia Li, Roberto Proietti +3

    cs.NIcs.LGeess.SParXiv:1905.02248v22019
  2. CACTUS: Mask-Guided Semantic Clean-Label Backdoors in Decentralized Federated Learning

    Chao Feng, Burkhard Stiller

    cs.LGcs.DCarXiv:2609.02450v12026
  3. Kymatio: Scattering Transforms in Python

    Mathieu Andreux, Tomás Angles, Georgios Exarchakis +15

    cs.LGcs.CVcs.SDarXiv:1812.11214v32018
  4. DMRL: Document-Mediated Reinforcement Learning for Skill Optimization in Advertising Recommendation

    Wei Zhang, Hongji Li, Song Sun +4

    cs.LGarXiv:2609.02170v12026
  5. Compositional Spectral Prompts for LLM-based Online Time Series Forecasting

    Seungyoon Choi, Hyunchul Kim, Jae-Gil Lee +1

    cs.LGarXiv:2609.02093v12026
  6. Cronus: Robust and Heterogeneous Collaborative Learning with Black-Box Knowledge Transfer

    Hongyan Chang, Virat Shejwalkar, Reza Shokri +1

    stat.MLcs.CRcs.LGarXiv:1912.11279v12019
  7. On Constrained Spectral Clustering and Its Applications

    Xiang Wang, Buyue Qian, Ian Davidson

    cs.LGstat.MLarXiv:1201.5338v22012
  8. Mining gold from implicit models to improve likelihood-free inference

    Johann Brehmer, Gilles Louppe, Juan Pavez +1

    stat.MLcs.LGhep-pharXiv:1805.12244v42018
  9. Model Reconstruction from Model Explanations

    Smitha Milli, Ludwig Schmidt, Anca D. Dragan +1

    stat.MLcs.LGarXiv:1807.05185v12018
  10. Convolutional Neural Networks on Graphs with Chebyshev Approximation, Revisited

    Mingguo He, Zhewei Wei, Ji-Rong Wen

    cs.LGcs.AIarXiv:2202.03580v52022
  11. SPECTRE: Defending Against Backdoor Attacks Using Robust Statistics

    Jonathan Hayase, Weihao Kong, Raghav Somani +1

    cs.LGcs.AIstat.MLarXiv:2104.11315v12021
  12. Gradient Inversion with Generative Image Prior

    Jinwoo Jeon, Jaechang Kim, Kangwook Lee +2

    cs.LGarXiv:2110.14962v12021
  13. Median-of-Means as an Extremal Convex Estimator and a Nonconvex Route to the Trimmed Oracle

    Angshul Majumdar

    cs.LGarXiv:2609.01689v12026
  14. ParSeNet: A Parametric Surface Fitting Network for 3D Point Clouds

    Gopal Sharma, Difan Liu, Subhransu Maji +3

    cs.CVcs.LGarXiv:2003.12181v52020
  15. Dutch Books for Language Models

    Isaiah Andrews, Suproteem Sarkar

    econ.GNcs.AIcs.CLarXiv:2609.02797v12026
  16. Attracting and Dispersing: A Simple Approach for Source-free Domain Adaptation

    Shiqi Yang, Yaxing Wang, Kai Wang +2

    cs.CVcs.LGarXiv:2205.04183v32022
  17. More Adaptive Algorithms for Adversarial Bandits

    Chen-Yu Wei, Haipeng Luo

    cs.LGstat.MLarXiv:1801.03265v32018
  18. Transferability and Hardness of Supervised Classification Tasks

    Anh T. Tran, Cuong V. Nguyen, Tal Hassner

    cs.LGcs.CVstat.MLarXiv:1908.08142v12019
  19. Adversarial Deep Learning for Robust Detection of Binary Encoded Malware

    Abdullah Al-Dujaili, Alex Huang, Erik Hemberg +1

    cs.CRcs.LGstat.MLarXiv:1801.02950v32018
  20. Evaluating and Calibrating Uncertainty Prediction in Regression Tasks

    Dan Levi, Liran Gispan, Niv Giladi +1

    cs.LGstat.MLarXiv:1905.11659v32019
  21. PubTables-1M: Towards comprehensive table extraction from unstructured documents

    Brandon Smock, Rohith Pesala, Robin Abraham

    cs.LGcs.CVarXiv:2110.00061v32021
  22. BAFFLE : Blockchain Based Aggregator Free Federated Learning

    Paritosh Ramanan, Kiyoshi Nakayama

    cs.LGcs.CRcs.DCarXiv:1909.07452v32019
  23. The importance of stain normalization in colorectal tissue classification with convolutional networks

    Francesco Ciompi, Oscar Geessink, Babak Ehteshami Bejnordi +6

    cs.CVcs.LGarXiv:1702.05931v22017
  24. Benchmarking the Performance of Bayesian Optimization across Multiple Experimental Materials Science Domains

    Qiaohao Liang, Aldair E. Gongora, Zekun Ren +12

    cond-mat.mtrl-scics.LGphysics.data-anarXiv:2106.01309v12021
  25. Meta-learning via Language Model In-context Tuning

    Yanda Chen, Ruiqi Zhong, Sheng Zha +2

    cs.CLcs.LGarXiv:2110.07814v22021
  26. Training seeds and model-selection stability in recommender-system evaluation

    Juan Manuel Rodriguez, Oleg Lesota, Antonela Tommasel

    cs.IRcs.LGarXiv:2609.02499v12026
  27. Neither hype nor gloom do DNNs justice

    Felix A. Wichmann, Simon Kornblith, Robert Geirhos

    cs.LGcs.CVq-bio.NCarXiv:2312.05355v12023
  28. Extracting Automata from Recurrent Neural Networks Using Queries and Counterexamples

    Gail Weiss, Yoav Goldberg, Eran Yahav

    cs.LGcs.FLarXiv:1711.09576v42017
  29. OGBench: Benchmarking Offline Goal-Conditioned RL

    Seohong Park, Kevin Frans, Benjamin Eysenbach +1

    cs.LGcs.AIarXiv:2410.20092v22024
  30. HuatuoGPT-Vision, Towards Injecting Medical Visual Knowledge into Multimodal LLMs at Scale

    Junying Chen, Chi Gui, Ruyi Ouyang +10

    cs.CVcs.AIcs.CLarXiv:2406.19280v42024
  31. Reasoning About Generalization via Conditional Mutual Information

    Thomas Steinke, Lydia Zakynthinou

    cs.LGcs.CRcs.DSarXiv:2001.09122v32020
  32. Deep Learning for Survival Analysis: A Review

    Simon Wiegrebe, Philipp Kopper, Raphael Sonabend +2

    stat.MLcs.LGarXiv:2305.14961v42023
  33. FORGE: Forward-Only Test-Time Adaptation for Integer-Only Vision Models on Microcontrollers

    Muhammad Rehan, Haider Ali, Muhammad Ali Munir +1

    cs.CVcs.ARcs.LGarXiv:2609.01683v12026
  34. IFW-BLS: Dual-Robust Broad Learning System with Intuitionistic Fuzzy Wave Loss

    Mushir Akhtar, M. Tanveer

    cs.LGarXiv:2609.02422v12026
  35. An invertible crystallographic representation for general inverse design of inorganic crystals with targeted properties

    Zekun Ren, Siyu Isaac Parker Tian, Juhwan Noh +14

    physics.comp-phcond-mat.mtrl-scics.LGarXiv:2005.07609v32020
  36. Reinforcement Learning from Imperfect Demonstrations

    Yang Gao, Huazhe Xu, Ji Lin +3

    cs.AIcs.LGstat.MLarXiv:1802.05313v22018
  37. Raw Waveform-based Speech Enhancement by Fully Convolutional Networks

    Szu-Wei Fu, Yu Tsao, Xugang Lu +1

    stat.MLcs.LGcs.SDarXiv:1703.02205v32017
  38. Generative Models for Effective ML on Private, Decentralized Datasets

    Sean Augenstein, H. Brendan McMahan, Daniel Ramage +5

    cs.LGstat.MLarXiv:1911.06679v22019
  39. ArCHer: Training Language Model Agents via Hierarchical Multi-Turn RL

    Yifei Zhou, Andrea Zanette, Jiayi Pan +2

    cs.LGcs.AIcs.CLarXiv:2402.19446v12024
  40. Exact Limits of Random Projections for Preserving Geometry: Distance Recovery, Nearest-Neighbor Rankings, and Covariance Shape in Gaussian Models

    Piyush Sao

    cs.LGcs.ITmath.NAarXiv:2609.02155v12026
  41. Multi-Reward Reinforced Summarization with Saliency and Entailment

    Ramakanth Pasunuru, Mohit Bansal

    cs.CLcs.AIcs.LGarXiv:1804.06451v22018
  42. RINSE: Robust Target-Time Normality Estimation for Zero-Shot Graph Anomaly Detection

    Taufikur Rahman Fuad, Md Abrar Jahin, Amir Hussain

    cs.LGcs.AIarXiv:2609.02497v12026
  43. Having Beer after Prayer? Measuring Cultural Bias in Large Language Models

    Tarek Naous, Michael J. Ryan, Alan Ritter +1

    cs.CLcs.AIcs.LGarXiv:2305.14456v42023
  44. Multi-modal Dense Video Captioning

    Vladimir Iashin, Esa Rahtu

    cs.CVcs.CLcs.LGarXiv:2003.07758v22020
  45. Quantum machine learning for image classification

    Arsenii Senokosov, Alexandr Sedykh, Asel Sagingalieva +2

    quant-phcs.CVcs.LGarXiv:2304.09224v22023
  46. Reward-rational (implicit) choice: A unifying formalism for reward learning

    Hong Jun Jeon, Smitha Milli, Anca D. Dragan

    cs.LGcs.AIcs.HCarXiv:2002.04833v42020
  47. What Algorithms can Transformers Learn? A Study in Length Generalization

    Hattie Zhou, Arwen Bradley, Etai Littwin +5

    cs.LGcs.AIcs.CLarXiv:2310.16028v12023
  48. COLD Decoding: Energy-based Constrained Text Generation with Langevin Dynamics

    Lianhui Qin, Sean Welleck, Daniel Khashabi +1

    cs.CLcs.AIcs.LGarXiv:2202.11705v32022
  49. Entity Abstraction in Visual Model-Based Reinforcement Learning

    Rishi Veerapaneni, John D. Co-Reyes, Michael Chang +5

    cs.LGcs.CVcs.NEarXiv:1910.12827v52019
  50. On the Representation Collapse of Sparse Mixture of Experts

    Zewen Chi, Li Dong, Shaohan Huang +9

    cs.CLcs.LGarXiv:2204.09179v32022
  51. Convergence of score-based generative modeling for general data distributions

    Holden Lee, Jianfeng Lu, Yixin Tan

    cs.LGmath.PRmath.STarXiv:2209.12381v22022
  52. Act More, Decide Less: Skill-Guided Adaptive Action Chunking for Long-Horizon LLM Agents

    Yanting Yang, Can Jin, Jinman Zhao +6

    cs.LGarXiv:2609.02042v12026
  53. Graph Optimal Transport for Cross-Domain Alignment

    Liqun Chen, Zhe Gan, Yu Cheng +3

    cs.CLcs.CVcs.LGarXiv:2006.14744v32020
  54. Sparse Coding and Dictionary Learning for Symmetric Positive Definite Matrices: A Kernel Approach

    Mehrtash T. Harandi, Conrad Sanderson, Richard Hartley +1

    cs.LGcs.CVstat.MLarXiv:1304.4344v12013
  55. Celebrating Diversity in Shared Multi-Agent Reinforcement Learning

    Chenghao Li, Tonghan Wang, Chengjie Wu +3

    cs.LGarXiv:2106.02195v22021
  56. DiDrive: A Risk-Aware Hierarchical Diffusion Framework for Safe Offline Reinforcement Learning in Autonomous Driving

    Qisong Guo, Jingtang Chen, Zhilin Chen +4

    cs.LGcs.ROarXiv:2609.01609v12026
  57. Bayesian Structure Learning with Generative Flow Networks

    Tristan Deleu, António Góis, Chris Emezue +4

    cs.LGstat.MLarXiv:2202.13903v22022
  58. ZeroSCROLLS: A Zero-Shot Benchmark for Long Text Understanding

    Uri Shaham, Maor Ivgi, Avia Efrat +2

    cs.CLcs.AIcs.LGarXiv:2305.14196v32023
  59. Motion-Attentive Transition for Zero-Shot Video Object Segmentation

    Tianfei Zhou, Shunzhou Wang, Yi Zhou +3

    cs.CVcs.LGeess.IVarXiv:2003.04253v32020
  60. Learning-Based Reconstruction Attacks on Coordinate-Obfuscated Point Clouds

    Mohammad Waquas Usmani, Susmit Shannigrahi, Michael Zink

    cs.CRcs.LGarXiv:2609.02568v12026