Machine Learning

Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

2,521 to 2,580 of 20,069

  1. Risk-Conditioned Fine-Tuning of Large Language Models

    Zixuan Liu, Fangzheng Wu, Brian Summa +1

    cs.LGarXiv:2609.08064v12026
  2. Robust Subspace Clustering via Thresholding

    Reinhard Heckel, Helmut Bölcskei

    stat.MLcs.ITcs.LGarXiv:1307.4891v42013
  3. Hyperbolic Image-Text Representations

    Karan Desai, Maximilian Nickel, Tanmay Rajpurohit +2

    cs.CVcs.LGarXiv:2304.09172v32023
  4. Secure Evaluation of Quantized Neural Networks

    Anders Dalskov, Daniel Escudero, Marcel Keller

    cs.CRcs.LGarXiv:1910.12435v22019
  5. Adversarially Trained Actor Critic for Offline Reinforcement Learning

    Ching-An Cheng, Tengyang Xie, Nan Jiang +1

    cs.LGarXiv:2202.02446v22022
  6. Learning Length-Extrapolatable Recurrent Models

    Hanwen Jiang

    cs.LGcs.CLarXiv:2609.09157v12026
  7. Submodular Functions: from Discrete to Continous Domains

    Francis Bach

    cs.LGmath.OCarXiv:1511.00394v22015
  8. The Well: a Large-Scale Collection of Diverse Physics Simulations for Machine Learning

    Ruben Ohana, Michael McCabe, Lucas Meyer +24

    cs.LGphysics.flu-dynarXiv:2412.00568v22024
  9. Symmetry and Group in Attribute-Object Compositions

    Yong-Lu Li, Yue Xu, Xiaohan Mao +1

    cs.CVcs.LGarXiv:2004.00587v12020
  10. Thompson Sampling for Combinatorial Semi-Bandits

    Siwei Wang, Wei Chen

    cs.LGarXiv:1803.04623v52018
  11. Rates of Convergence for Nearest Neighbor Classification

    Kamalika Chaudhuri, Sanjoy Dasgupta

    cs.LGmath.STstat.MLarXiv:1407.0067v22014
  12. Unsupervised Multi-source Domain Adaptation Without Access to Source Data

    Sk Miraj Ahmed, Dripta S. Raychaudhuri, Sujoy Paul +2

    cs.LGcs.CVarXiv:2104.01845v12021
  13. Entropy-Regularized Rank-Masked Policy Optimization for Test-Time Reinforcement Learning in Code Generation

    Jiacheng Xu, Feng Chen, Xiuneng Xu +1

    cs.LGcs.CLarXiv:2609.09135v12026
  14. A Disease Diagnosis and Treatment Recommendation System Based on Big Data Mining and Cloud Computing

    Jianguo Chen, Kenli Li, Huigui Rong +3

    cs.LGstat.MLarXiv:1810.07762v12018
  15. Beyond One-Step-Ahead Forecasting: Evaluation of Alternative Multi-Step-Ahead Forecasting Models for Crude Oil Prices

    Tao Xiong, Yukun Bao, Zhongyi Hu

    cs.LGcs.AIarXiv:1401.1560v12014
  16. Pedestrian Attribute Recognition: A Survey

    Xiao Wang, Shaofei Zheng, Rui Yang +4

    cs.CVcs.AIcs.LGarXiv:1901.07474v22019
  17. A Closed-Form Estimator and Diagnostic Battery for Anchor-Judge Error Correlation, Under a Single-Common-Factor Model

    Veerendra Kumar Sunkavalli

    stat.MEcs.CLcs.LGarXiv:2609.08826v12026
  18. A Survey on Deep Active Learning: Recent Advances and New Frontiers

    Dongyuan Li, Zhen Wang, Yankai Chen +3

    cs.LGarXiv:2405.00334v22024
  19. Consensus Multi-Agent Reinforcement Learning for Volt-VAR Control in Power Distribution Networks

    Yuanqi Gao, Wei Wang, Nanpeng Yu

    eess.SYcs.LGarXiv:2007.02991v12020
  20. Accurate Genomic Prediction Of Human Height

    Louis Lello, Steven G. Avery, Laurent Tellier +3

    q-bio.GNcs.LGq-bio.QMarXiv:1709.06489v12017
  21. A Closer Look at Classification Evaluation Metrics and a Critical Reflection of Common Evaluation Practice

    Juri Opitz

    cs.LGcs.CLarXiv:2404.16958v22024
  22. Improving Chemical Autoencoder Latent Space and Molecular De novo Generation Diversity with Heteroencoders

    Esben Jannik Bjerrum, Boris Sattarov

    cs.LGstat.MLarXiv:1806.09300v22018
  23. Opportunities and Challenges in Deep Learning Adversarial Robustness: A Survey

    Samuel Henrique Silva, Peyman Najafirad

    cs.LGcs.AIstat.MLarXiv:2007.00753v22020
  24. Data-driven polynomial chaos expansion for machine learning regression

    E. Torre, S. Marelli, P. Embrechts +1

    stat.MLcs.LGstat.COarXiv:1808.03216v22018
  25. Rotating without Seeing: Towards In-hand Dexterity through Touch

    Zhao-Heng Yin, Binghao Huang, Yuzhe Qin +2

    cs.ROcs.AIcs.LGarXiv:2303.10880v42023
  26. TontaubeV1: Streaming Text-to-Speech with Hierarchical Codec Modeling and Bounded Context

    Fritz Cremer, Jonathan Cremer

    cs.SDcs.CLcs.LGarXiv:2609.08703v12026
  27. Bias and Generalization in Deep Generative Models: An Empirical Study

    Shengjia Zhao, Hongyu Ren, Arianna Yuan +3

    cs.LGstat.MLarXiv:1811.03259v12018
  28. Physics of Language Models: Part 3.2, Knowledge Manipulation

    Zeyuan Allen-Zhu, Yuanzhi Li

    cs.CLcs.AIcs.LGarXiv:2309.14402v22023
  29. Distillation as Probability Transport: Routed On-Policy Distillation

    Tianle Xia, Lingxiang Hu, Yiding Sun +6

    cs.LGcs.CLarXiv:2609.08337v12026
  30. Solving for high dimensional committor functions using artificial neural networks

    Yuehaw Khoo, Jianfeng Lu, Lexing Ying

    cs.LGmath.NAstat.MLarXiv:1802.10275v12018
  31. HoneyRoute: Honeypot-Model Routing for Adversarial LLM Serving

    Han Jin

    cs.CRcs.CLcs.LGarXiv:2609.08306v22026
  32. The mixed deep energy method for resolving concentration features in finite strain hyperelasticity

    Jan N. Fuhg, Nikolaos Bouklas

    cs.CEcs.LGarXiv:2104.09623v12021
  33. To Believe or Not to Believe Your LLM

    Yasin Abbasi Yadkori, Ilja Kuzborskij, András György +1

    cs.LGcs.AIcs.CLarXiv:2406.02543v22024
  34. Pathways: Asynchronous Distributed Dataflow for ML

    Paul Barham, Aakanksha Chowdhery, Jeff Dean +13

    cs.DCcs.LGarXiv:2203.12533v12022
  35. GHRS: Graph-based Hybrid Recommendation System with Application to Movie Recommendation

    Zahra Zamanzadeh Darban, Mohammad Hadi Valipour

    cs.IRcs.AIcs.LGarXiv:2111.11293v22021
  36. A Comprehensive Review of Digital Twin -- Part 2: Roles of Uncertainty Quantification and Optimization, a Battery Digital Twin, and Perspectives

    Adam Thelen, Xiaoge Zhang, Olga Fink +7

    cs.LGmath.OCarXiv:2208.12904v12022
  37. Evaluation of Contextual Understanding in Large Language Models

    Subavarshana Arumugam, Mamta Nallaretnam, Kithuni Wickramasinghe +4

    cs.CLcs.LGarXiv:2609.09004v12026
  38. Collaborative Machine Learning with Incentive-Aware Model Rewards

    Rachael Hwee Ling Sim, Yehong Zhang, Mun Choon Chan +1

    cs.LGcs.GTcs.MAarXiv:2010.12797v12020
  39. Second-Order Optimization for Non-Convex Machine Learning: An Empirical Study

    Peng Xu, Farbod Roosta-Khorasani, Michael W. Mahoney

    math.OCcs.LGmath.NAarXiv:1708.07827v22017
  40. Intention-aware Long Horizon Trajectory Prediction of Surrounding Vehicles using Dual LSTM Networks

    Long Xin, Pin Wang, Ching-Yao Chan +3

    cs.LGcs.ROstat.MLarXiv:1906.02815v12019
  41. Fast Multi-language LSTM-based Online Handwriting Recognition

    Victor Carbune, Pedro Gonnet, Thomas Deselaers +7

    cs.CLcs.LGstat.MLarXiv:1902.10525v22019
  42. MahNMF: Manhattan Non-negative Matrix Factorization

    Naiyang Guan, Dacheng Tao, Zhigang Luo +1

    stat.MLcs.LGmath.NAarXiv:1207.3438v12012
  43. Eigenvalue and Generalized Eigenvalue Problems: Tutorial

    Benyamin Ghojogh, Fakhri Karray, Mark Crowley

    stat.MLcs.LGarXiv:1903.11240v32019
  44. AgentDrift: A Step-Labeled Benchmark of Injection-Hijacked LLM Agent Trajectories

    Asif Pinjari, Mithun Paul Saint-Germain

    cs.CRcs.AIcs.LGarXiv:2609.06972v12026
  45. TextScanner: Reading Characters in Order for Robust Scene Text Recognition

    Zhaoyi Wan, Minghang He, Haoran Chen +2

    cs.CVcs.CLcs.LGarXiv:1912.12422v22019
  46. Federated learning with class imbalance reduction

    Miao Yang, Akitanoshou Wong, Hongbin Zhu +2

    cs.LGcs.AIcs.DCarXiv:2011.11266v12020
  47. Steering Interference Reflects the Model's Defaults, Not the Behavior Directions

    Srikanth Malla, Chiho Choi, Joon Hee Choi

    cs.LGcs.AIarXiv:2609.06951v12026
  48. Model Extraction Warning in MLaaS Paradigm

    Manish Kesarwani, Bhaskar Mukhoty, Vijay Arya +1

    cs.LGcs.CRcs.DCarXiv:1711.07221v12017
  49. The Geometry of Refusal: Why Post-Hoc Safety Is Fragile and Pretraining-Time Safety Persists

    Srikanth Malla, Chiho Choi, Joon Hee Choi

    cs.LGcs.AIarXiv:2609.06934v12026
  50. Learning to Infer and Execute 3D Shape Programs

    Yonglong Tian, Andrew Luo, Xingyuan Sun +4

    cs.CVcs.AIcs.GRarXiv:1901.02875v32019
  51. Analogies Explained: Towards Understanding Word Embeddings

    Carl Allen, Timothy Hospedales

    cs.CLcs.LGstat.MLarXiv:1901.09813v22019
  52. Pavement Image Datasets: A New Benchmark Dataset to Classify and Densify Pavement Distresses

    Hamed Majidifard, Peng Jin, Yaw Adu-Gyamfi +1

    cs.CVcs.LGstat.MLarXiv:1910.11123v22019
  53. Few-Shot Text Generation with Pattern-Exploiting Training

    Timo Schick, Hinrich Schütze

    cs.CLcs.LGarXiv:2012.11926v22020
  54. Constrained Online Learning with Noisy Constraint Values

    Vaneet Aggarwal

    cs.LGcs.AImath.OCarXiv:2609.06921v12026
  55. From Synthetic Priors to Model Behavior: Structural Coverage in Tabular Foundation Models

    He Zhao, Ryan Thompson, Daniel M. Steinberg +3

    cs.LGcs.AIarXiv:2609.06912v12026
  56. PFNN: A Penalty-Free Neural Network Method for Solving a Class of Second-Order Boundary-Value Problems on Complex Geometries

    Hailong Sheng, Chao Yang

    math.NAcs.LGarXiv:2004.06490v22020
  57. Noisy-Space Policy Gradient for Diffusion Policies in Offline Reinforcement Learning

    Mahmoud Selim, Cristina Cipriani, Karl H. Johansson

    cs.LGcs.AIcs.ROarXiv:2609.06882v12026
  58. Personalized Language Modeling from Personalized Human Feedback

    Xinyu Li, Ruiyang Zhou, Zachary C. Lipton +1

    cs.CLcs.AIcs.LGarXiv:2402.05133v32024
  59. Towards Explainable NLP: A Generative Explanation Framework for Text Classification

    Hui Liu, Qingyu Yin, William Yang Wang

    cs.CLcs.AIcs.LGarXiv:1811.00196v22018
  60. iVideoGPT: Interactive VideoGPTs are Scalable World Models

    Jialong Wu, Shaofeng Yin, Ningya Feng +4

    cs.CVcs.LGcs.ROarXiv:2405.15223v32024