Machine Learning
Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
8,761 to 8,820 of 20,193
DeepRMSA: A Deep Reinforcement Learning Framework for Routing, Modulation and Spectrum Assignment in Elastic Optical Networks
Xiaoliang Chen, Baojia Li, Roberto Proietti +3
cs.NIcs.LGeess.SParXiv:1905.02248v22019CACTUS: Mask-Guided Semantic Clean-Label Backdoors in Decentralized Federated Learning
Chao Feng, Burkhard Stiller
cs.LGcs.DCarXiv:2609.02450v12026Kymatio: Scattering Transforms in Python
Mathieu Andreux, Tomás Angles, Georgios Exarchakis +15
cs.LGcs.CVcs.SDarXiv:1812.11214v32018DMRL: Document-Mediated Reinforcement Learning for Skill Optimization in Advertising Recommendation
Wei Zhang, Hongji Li, Song Sun +4
cs.LGarXiv:2609.02170v12026Compositional Spectral Prompts for LLM-based Online Time Series Forecasting
Seungyoon Choi, Hyunchul Kim, Jae-Gil Lee +1
cs.LGarXiv:2609.02093v12026Cronus: Robust and Heterogeneous Collaborative Learning with Black-Box Knowledge Transfer
Hongyan Chang, Virat Shejwalkar, Reza Shokri +1
stat.MLcs.CRcs.LGarXiv:1912.11279v12019On Constrained Spectral Clustering and Its Applications
Xiang Wang, Buyue Qian, Ian Davidson
cs.LGstat.MLarXiv:1201.5338v22012Mining gold from implicit models to improve likelihood-free inference
Johann Brehmer, Gilles Louppe, Juan Pavez +1
stat.MLcs.LGhep-pharXiv:1805.12244v42018Model Reconstruction from Model Explanations
Smitha Milli, Ludwig Schmidt, Anca D. Dragan +1
stat.MLcs.LGarXiv:1807.05185v12018Convolutional Neural Networks on Graphs with Chebyshev Approximation, Revisited
Mingguo He, Zhewei Wei, Ji-Rong Wen
cs.LGcs.AIarXiv:2202.03580v52022SPECTRE: Defending Against Backdoor Attacks Using Robust Statistics
Jonathan Hayase, Weihao Kong, Raghav Somani +1
cs.LGcs.AIstat.MLarXiv:2104.11315v12021Gradient Inversion with Generative Image Prior
Jinwoo Jeon, Jaechang Kim, Kangwook Lee +2
cs.LGarXiv:2110.14962v12021Median-of-Means as an Extremal Convex Estimator and a Nonconvex Route to the Trimmed Oracle
Angshul Majumdar
cs.LGarXiv:2609.01689v12026ParSeNet: A Parametric Surface Fitting Network for 3D Point Clouds
Gopal Sharma, Difan Liu, Subhransu Maji +3
cs.CVcs.LGarXiv:2003.12181v52020Dutch Books for Language Models
Isaiah Andrews, Suproteem Sarkar
econ.GNcs.AIcs.CLarXiv:2609.02797v12026Attracting and Dispersing: A Simple Approach for Source-free Domain Adaptation
Shiqi Yang, Yaxing Wang, Kai Wang +2
cs.CVcs.LGarXiv:2205.04183v32022More Adaptive Algorithms for Adversarial Bandits
Chen-Yu Wei, Haipeng Luo
cs.LGstat.MLarXiv:1801.03265v32018Transferability and Hardness of Supervised Classification Tasks
Anh T. Tran, Cuong V. Nguyen, Tal Hassner
cs.LGcs.CVstat.MLarXiv:1908.08142v12019Adversarial Deep Learning for Robust Detection of Binary Encoded Malware
Abdullah Al-Dujaili, Alex Huang, Erik Hemberg +1
cs.CRcs.LGstat.MLarXiv:1801.02950v32018Evaluating and Calibrating Uncertainty Prediction in Regression Tasks
Dan Levi, Liran Gispan, Niv Giladi +1
cs.LGstat.MLarXiv:1905.11659v32019PubTables-1M: Towards comprehensive table extraction from unstructured documents
Brandon Smock, Rohith Pesala, Robin Abraham
cs.LGcs.CVarXiv:2110.00061v32021BAFFLE : Blockchain Based Aggregator Free Federated Learning
Paritosh Ramanan, Kiyoshi Nakayama
cs.LGcs.CRcs.DCarXiv:1909.07452v32019The importance of stain normalization in colorectal tissue classification with convolutional networks
Francesco Ciompi, Oscar Geessink, Babak Ehteshami Bejnordi +6
cs.CVcs.LGarXiv:1702.05931v22017Benchmarking the Performance of Bayesian Optimization across Multiple Experimental Materials Science Domains
Qiaohao Liang, Aldair E. Gongora, Zekun Ren +12
cond-mat.mtrl-scics.LGphysics.data-anarXiv:2106.01309v12021Meta-learning via Language Model In-context Tuning
Yanda Chen, Ruiqi Zhong, Sheng Zha +2
cs.CLcs.LGarXiv:2110.07814v22021Training seeds and model-selection stability in recommender-system evaluation
Juan Manuel Rodriguez, Oleg Lesota, Antonela Tommasel
cs.IRcs.LGarXiv:2609.02499v12026Neither hype nor gloom do DNNs justice
Felix A. Wichmann, Simon Kornblith, Robert Geirhos
cs.LGcs.CVq-bio.NCarXiv:2312.05355v12023Extracting Automata from Recurrent Neural Networks Using Queries and Counterexamples
Gail Weiss, Yoav Goldberg, Eran Yahav
cs.LGcs.FLarXiv:1711.09576v42017OGBench: Benchmarking Offline Goal-Conditioned RL
Seohong Park, Kevin Frans, Benjamin Eysenbach +1
cs.LGcs.AIarXiv:2410.20092v22024HuatuoGPT-Vision, Towards Injecting Medical Visual Knowledge into Multimodal LLMs at Scale
Junying Chen, Chi Gui, Ruyi Ouyang +10
cs.CVcs.AIcs.CLarXiv:2406.19280v42024Reasoning About Generalization via Conditional Mutual Information
Thomas Steinke, Lydia Zakynthinou
cs.LGcs.CRcs.DSarXiv:2001.09122v32020Deep Learning for Survival Analysis: A Review
Simon Wiegrebe, Philipp Kopper, Raphael Sonabend +2
stat.MLcs.LGarXiv:2305.14961v42023FORGE: Forward-Only Test-Time Adaptation for Integer-Only Vision Models on Microcontrollers
Muhammad Rehan, Haider Ali, Muhammad Ali Munir +1
cs.CVcs.ARcs.LGarXiv:2609.01683v12026IFW-BLS: Dual-Robust Broad Learning System with Intuitionistic Fuzzy Wave Loss
Mushir Akhtar, M. Tanveer
cs.LGarXiv:2609.02422v12026An invertible crystallographic representation for general inverse design of inorganic crystals with targeted properties
Zekun Ren, Siyu Isaac Parker Tian, Juhwan Noh +14
physics.comp-phcond-mat.mtrl-scics.LGarXiv:2005.07609v32020Reinforcement Learning from Imperfect Demonstrations
Yang Gao, Huazhe Xu, Ji Lin +3
cs.AIcs.LGstat.MLarXiv:1802.05313v22018Raw Waveform-based Speech Enhancement by Fully Convolutional Networks
Szu-Wei Fu, Yu Tsao, Xugang Lu +1
stat.MLcs.LGcs.SDarXiv:1703.02205v32017Generative Models for Effective ML on Private, Decentralized Datasets
Sean Augenstein, H. Brendan McMahan, Daniel Ramage +5
cs.LGstat.MLarXiv:1911.06679v22019ArCHer: Training Language Model Agents via Hierarchical Multi-Turn RL
Yifei Zhou, Andrea Zanette, Jiayi Pan +2
cs.LGcs.AIcs.CLarXiv:2402.19446v12024Exact Limits of Random Projections for Preserving Geometry: Distance Recovery, Nearest-Neighbor Rankings, and Covariance Shape in Gaussian Models
Piyush Sao
cs.LGcs.ITmath.NAarXiv:2609.02155v12026Multi-Reward Reinforced Summarization with Saliency and Entailment
Ramakanth Pasunuru, Mohit Bansal
cs.CLcs.AIcs.LGarXiv:1804.06451v22018RINSE: Robust Target-Time Normality Estimation for Zero-Shot Graph Anomaly Detection
Taufikur Rahman Fuad, Md Abrar Jahin, Amir Hussain
cs.LGcs.AIarXiv:2609.02497v12026Having Beer after Prayer? Measuring Cultural Bias in Large Language Models
Tarek Naous, Michael J. Ryan, Alan Ritter +1
cs.CLcs.AIcs.LGarXiv:2305.14456v42023Multi-modal Dense Video Captioning
Vladimir Iashin, Esa Rahtu
cs.CVcs.CLcs.LGarXiv:2003.07758v22020Quantum machine learning for image classification
Arsenii Senokosov, Alexandr Sedykh, Asel Sagingalieva +2
quant-phcs.CVcs.LGarXiv:2304.09224v22023Reward-rational (implicit) choice: A unifying formalism for reward learning
Hong Jun Jeon, Smitha Milli, Anca D. Dragan
cs.LGcs.AIcs.HCarXiv:2002.04833v42020What Algorithms can Transformers Learn? A Study in Length Generalization
Hattie Zhou, Arwen Bradley, Etai Littwin +5
cs.LGcs.AIcs.CLarXiv:2310.16028v12023COLD Decoding: Energy-based Constrained Text Generation with Langevin Dynamics
Lianhui Qin, Sean Welleck, Daniel Khashabi +1
cs.CLcs.AIcs.LGarXiv:2202.11705v32022Entity Abstraction in Visual Model-Based Reinforcement Learning
Rishi Veerapaneni, John D. Co-Reyes, Michael Chang +5
cs.LGcs.CVcs.NEarXiv:1910.12827v52019On the Representation Collapse of Sparse Mixture of Experts
Zewen Chi, Li Dong, Shaohan Huang +9
cs.CLcs.LGarXiv:2204.09179v32022Convergence of score-based generative modeling for general data distributions
Holden Lee, Jianfeng Lu, Yixin Tan
cs.LGmath.PRmath.STarXiv:2209.12381v22022Act More, Decide Less: Skill-Guided Adaptive Action Chunking for Long-Horizon LLM Agents
Yanting Yang, Can Jin, Jinman Zhao +6
cs.LGarXiv:2609.02042v12026Graph Optimal Transport for Cross-Domain Alignment
Liqun Chen, Zhe Gan, Yu Cheng +3
cs.CLcs.CVcs.LGarXiv:2006.14744v32020Sparse Coding and Dictionary Learning for Symmetric Positive Definite Matrices: A Kernel Approach
Mehrtash T. Harandi, Conrad Sanderson, Richard Hartley +1
cs.LGcs.CVstat.MLarXiv:1304.4344v12013Celebrating Diversity in Shared Multi-Agent Reinforcement Learning
Chenghao Li, Tonghan Wang, Chengjie Wu +3
cs.LGarXiv:2106.02195v22021DiDrive: A Risk-Aware Hierarchical Diffusion Framework for Safe Offline Reinforcement Learning in Autonomous Driving
Qisong Guo, Jingtang Chen, Zhilin Chen +4
cs.LGcs.ROarXiv:2609.01609v12026Bayesian Structure Learning with Generative Flow Networks
Tristan Deleu, António Góis, Chris Emezue +4
cs.LGstat.MLarXiv:2202.13903v22022ZeroSCROLLS: A Zero-Shot Benchmark for Long Text Understanding
Uri Shaham, Maor Ivgi, Avia Efrat +2
cs.CLcs.AIcs.LGarXiv:2305.14196v32023Motion-Attentive Transition for Zero-Shot Video Object Segmentation
Tianfei Zhou, Shunzhou Wang, Yi Zhou +3
cs.CVcs.LGeess.IVarXiv:2003.04253v32020Learning-Based Reconstruction Attacks on Coordinate-Obfuscated Point Clouds
Mohammad Waquas Usmani, Susmit Shannigrahi, Michael Zink
cs.CRcs.LGarXiv:2609.02568v12026