Machine Learning
Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
2,521 to 2,580 of 20,069
Risk-Conditioned Fine-Tuning of Large Language Models
Zixuan Liu, Fangzheng Wu, Brian Summa +1
cs.LGarXiv:2609.08064v12026Robust Subspace Clustering via Thresholding
Reinhard Heckel, Helmut Bölcskei
stat.MLcs.ITcs.LGarXiv:1307.4891v42013Hyperbolic Image-Text Representations
Karan Desai, Maximilian Nickel, Tanmay Rajpurohit +2
cs.CVcs.LGarXiv:2304.09172v32023Secure Evaluation of Quantized Neural Networks
Anders Dalskov, Daniel Escudero, Marcel Keller
cs.CRcs.LGarXiv:1910.12435v22019Adversarially Trained Actor Critic for Offline Reinforcement Learning
Ching-An Cheng, Tengyang Xie, Nan Jiang +1
cs.LGarXiv:2202.02446v22022Learning Length-Extrapolatable Recurrent Models
Hanwen Jiang
cs.LGcs.CLarXiv:2609.09157v12026Submodular Functions: from Discrete to Continous Domains
Francis Bach
cs.LGmath.OCarXiv:1511.00394v22015The Well: a Large-Scale Collection of Diverse Physics Simulations for Machine Learning
Ruben Ohana, Michael McCabe, Lucas Meyer +24
cs.LGphysics.flu-dynarXiv:2412.00568v22024Symmetry and Group in Attribute-Object Compositions
Yong-Lu Li, Yue Xu, Xiaohan Mao +1
cs.CVcs.LGarXiv:2004.00587v12020Thompson Sampling for Combinatorial Semi-Bandits
Siwei Wang, Wei Chen
cs.LGarXiv:1803.04623v52018Rates of Convergence for Nearest Neighbor Classification
Kamalika Chaudhuri, Sanjoy Dasgupta
cs.LGmath.STstat.MLarXiv:1407.0067v22014Unsupervised Multi-source Domain Adaptation Without Access to Source Data
Sk Miraj Ahmed, Dripta S. Raychaudhuri, Sujoy Paul +2
cs.LGcs.CVarXiv:2104.01845v12021Entropy-Regularized Rank-Masked Policy Optimization for Test-Time Reinforcement Learning in Code Generation
Jiacheng Xu, Feng Chen, Xiuneng Xu +1
cs.LGcs.CLarXiv:2609.09135v12026A Disease Diagnosis and Treatment Recommendation System Based on Big Data Mining and Cloud Computing
Jianguo Chen, Kenli Li, Huigui Rong +3
cs.LGstat.MLarXiv:1810.07762v12018Beyond One-Step-Ahead Forecasting: Evaluation of Alternative Multi-Step-Ahead Forecasting Models for Crude Oil Prices
Tao Xiong, Yukun Bao, Zhongyi Hu
cs.LGcs.AIarXiv:1401.1560v12014Pedestrian Attribute Recognition: A Survey
Xiao Wang, Shaofei Zheng, Rui Yang +4
cs.CVcs.AIcs.LGarXiv:1901.07474v22019A Closed-Form Estimator and Diagnostic Battery for Anchor-Judge Error Correlation, Under a Single-Common-Factor Model
Veerendra Kumar Sunkavalli
stat.MEcs.CLcs.LGarXiv:2609.08826v12026A Survey on Deep Active Learning: Recent Advances and New Frontiers
Dongyuan Li, Zhen Wang, Yankai Chen +3
cs.LGarXiv:2405.00334v22024Consensus Multi-Agent Reinforcement Learning for Volt-VAR Control in Power Distribution Networks
Yuanqi Gao, Wei Wang, Nanpeng Yu
eess.SYcs.LGarXiv:2007.02991v12020Accurate Genomic Prediction Of Human Height
Louis Lello, Steven G. Avery, Laurent Tellier +3
q-bio.GNcs.LGq-bio.QMarXiv:1709.06489v12017A Closer Look at Classification Evaluation Metrics and a Critical Reflection of Common Evaluation Practice
Juri Opitz
cs.LGcs.CLarXiv:2404.16958v22024Improving Chemical Autoencoder Latent Space and Molecular De novo Generation Diversity with Heteroencoders
Esben Jannik Bjerrum, Boris Sattarov
cs.LGstat.MLarXiv:1806.09300v22018Opportunities and Challenges in Deep Learning Adversarial Robustness: A Survey
Samuel Henrique Silva, Peyman Najafirad
cs.LGcs.AIstat.MLarXiv:2007.00753v22020Data-driven polynomial chaos expansion for machine learning regression
E. Torre, S. Marelli, P. Embrechts +1
stat.MLcs.LGstat.COarXiv:1808.03216v22018Rotating without Seeing: Towards In-hand Dexterity through Touch
Zhao-Heng Yin, Binghao Huang, Yuzhe Qin +2
cs.ROcs.AIcs.LGarXiv:2303.10880v42023TontaubeV1: Streaming Text-to-Speech with Hierarchical Codec Modeling and Bounded Context
Fritz Cremer, Jonathan Cremer
cs.SDcs.CLcs.LGarXiv:2609.08703v12026Bias and Generalization in Deep Generative Models: An Empirical Study
Shengjia Zhao, Hongyu Ren, Arianna Yuan +3
cs.LGstat.MLarXiv:1811.03259v12018Physics of Language Models: Part 3.2, Knowledge Manipulation
Zeyuan Allen-Zhu, Yuanzhi Li
cs.CLcs.AIcs.LGarXiv:2309.14402v22023Distillation as Probability Transport: Routed On-Policy Distillation
Tianle Xia, Lingxiang Hu, Yiding Sun +6
cs.LGcs.CLarXiv:2609.08337v12026Solving for high dimensional committor functions using artificial neural networks
Yuehaw Khoo, Jianfeng Lu, Lexing Ying
cs.LGmath.NAstat.MLarXiv:1802.10275v12018HoneyRoute: Honeypot-Model Routing for Adversarial LLM Serving
Han Jin
cs.CRcs.CLcs.LGarXiv:2609.08306v22026The mixed deep energy method for resolving concentration features in finite strain hyperelasticity
Jan N. Fuhg, Nikolaos Bouklas
cs.CEcs.LGarXiv:2104.09623v12021To Believe or Not to Believe Your LLM
Yasin Abbasi Yadkori, Ilja Kuzborskij, András György +1
cs.LGcs.AIcs.CLarXiv:2406.02543v22024Pathways: Asynchronous Distributed Dataflow for ML
Paul Barham, Aakanksha Chowdhery, Jeff Dean +13
cs.DCcs.LGarXiv:2203.12533v12022GHRS: Graph-based Hybrid Recommendation System with Application to Movie Recommendation
Zahra Zamanzadeh Darban, Mohammad Hadi Valipour
cs.IRcs.AIcs.LGarXiv:2111.11293v22021A Comprehensive Review of Digital Twin -- Part 2: Roles of Uncertainty Quantification and Optimization, a Battery Digital Twin, and Perspectives
Adam Thelen, Xiaoge Zhang, Olga Fink +7
cs.LGmath.OCarXiv:2208.12904v12022Evaluation of Contextual Understanding in Large Language Models
Subavarshana Arumugam, Mamta Nallaretnam, Kithuni Wickramasinghe +4
cs.CLcs.LGarXiv:2609.09004v12026Collaborative Machine Learning with Incentive-Aware Model Rewards
Rachael Hwee Ling Sim, Yehong Zhang, Mun Choon Chan +1
cs.LGcs.GTcs.MAarXiv:2010.12797v12020Second-Order Optimization for Non-Convex Machine Learning: An Empirical Study
Peng Xu, Farbod Roosta-Khorasani, Michael W. Mahoney
math.OCcs.LGmath.NAarXiv:1708.07827v22017Intention-aware Long Horizon Trajectory Prediction of Surrounding Vehicles using Dual LSTM Networks
Long Xin, Pin Wang, Ching-Yao Chan +3
cs.LGcs.ROstat.MLarXiv:1906.02815v12019Fast Multi-language LSTM-based Online Handwriting Recognition
Victor Carbune, Pedro Gonnet, Thomas Deselaers +7
cs.CLcs.LGstat.MLarXiv:1902.10525v22019MahNMF: Manhattan Non-negative Matrix Factorization
Naiyang Guan, Dacheng Tao, Zhigang Luo +1
stat.MLcs.LGmath.NAarXiv:1207.3438v12012Eigenvalue and Generalized Eigenvalue Problems: Tutorial
Benyamin Ghojogh, Fakhri Karray, Mark Crowley
stat.MLcs.LGarXiv:1903.11240v32019AgentDrift: A Step-Labeled Benchmark of Injection-Hijacked LLM Agent Trajectories
Asif Pinjari, Mithun Paul Saint-Germain
cs.CRcs.AIcs.LGarXiv:2609.06972v12026TextScanner: Reading Characters in Order for Robust Scene Text Recognition
Zhaoyi Wan, Minghang He, Haoran Chen +2
cs.CVcs.CLcs.LGarXiv:1912.12422v22019Federated learning with class imbalance reduction
Miao Yang, Akitanoshou Wong, Hongbin Zhu +2
cs.LGcs.AIcs.DCarXiv:2011.11266v12020Steering Interference Reflects the Model's Defaults, Not the Behavior Directions
Srikanth Malla, Chiho Choi, Joon Hee Choi
cs.LGcs.AIarXiv:2609.06951v12026Model Extraction Warning in MLaaS Paradigm
Manish Kesarwani, Bhaskar Mukhoty, Vijay Arya +1
cs.LGcs.CRcs.DCarXiv:1711.07221v12017The Geometry of Refusal: Why Post-Hoc Safety Is Fragile and Pretraining-Time Safety Persists
Srikanth Malla, Chiho Choi, Joon Hee Choi
cs.LGcs.AIarXiv:2609.06934v12026Learning to Infer and Execute 3D Shape Programs
Yonglong Tian, Andrew Luo, Xingyuan Sun +4
cs.CVcs.AIcs.GRarXiv:1901.02875v32019Analogies Explained: Towards Understanding Word Embeddings
Carl Allen, Timothy Hospedales
cs.CLcs.LGstat.MLarXiv:1901.09813v22019Pavement Image Datasets: A New Benchmark Dataset to Classify and Densify Pavement Distresses
Hamed Majidifard, Peng Jin, Yaw Adu-Gyamfi +1
cs.CVcs.LGstat.MLarXiv:1910.11123v22019Few-Shot Text Generation with Pattern-Exploiting Training
Timo Schick, Hinrich Schütze
cs.CLcs.LGarXiv:2012.11926v22020Constrained Online Learning with Noisy Constraint Values
Vaneet Aggarwal
cs.LGcs.AImath.OCarXiv:2609.06921v12026From Synthetic Priors to Model Behavior: Structural Coverage in Tabular Foundation Models
He Zhao, Ryan Thompson, Daniel M. Steinberg +3
cs.LGcs.AIarXiv:2609.06912v12026PFNN: A Penalty-Free Neural Network Method for Solving a Class of Second-Order Boundary-Value Problems on Complex Geometries
Hailong Sheng, Chao Yang
math.NAcs.LGarXiv:2004.06490v22020Noisy-Space Policy Gradient for Diffusion Policies in Offline Reinforcement Learning
Mahmoud Selim, Cristina Cipriani, Karl H. Johansson
cs.LGcs.AIcs.ROarXiv:2609.06882v12026Personalized Language Modeling from Personalized Human Feedback
Xinyu Li, Ruiyang Zhou, Zachary C. Lipton +1
cs.CLcs.AIcs.LGarXiv:2402.05133v32024Towards Explainable NLP: A Generative Explanation Framework for Text Classification
Hui Liu, Qingyu Yin, William Yang Wang
cs.CLcs.AIcs.LGarXiv:1811.00196v22018iVideoGPT: Interactive VideoGPTs are Scalable World Models
Jialong Wu, Shaofeng Yin, Ningya Feng +4
cs.CVcs.LGcs.ROarXiv:2405.15223v32024