Machine Learning

Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

1,561 to 1,620 of 20,199

  1. Universal Jailbreak Backdoors from Poisoned Human Feedback

    Javier Rando, Florian Tramèr

    cs.AIcs.CLcs.CRarXiv:2311.14455v42023
  2. Zero-shot rib design: merging training-free generative prior with topology optimization

    Yongmin Kwon, Namwoo Kang

    cs.LGphysics.app-pharXiv:2609.10643v12026
  3. The People's Speech: A Large-Scale Diverse English Speech Recognition Dataset for Commercial Usage

    Daniel Galvez, Greg Diamos, Juan Ciro +7

    cs.LGstat.MLarXiv:2111.09344v12021
  4. CALF: Aligning LLMs for Time Series Forecasting via Cross-modal Fine-Tuning

    Peiyuan Liu, Hang Guo, Tao Dai +5

    cs.LGcs.CLarXiv:2403.07300v32024
  5. Halo: Improving forecast accuracy through heteroscedastic estimation

    Adam Cataldo

    cs.LGarXiv:2609.10589v12026
  6. M3-Former: Multimodal Transformer with Mixture-of-Experts for Long-Term Vessel Trajectory Prediction

    Wenzhe Jin, Haina Tang

    cs.LGcs.CVarXiv:2609.10559v12026
  7. Diverse Trajectory Forecasting with Determinantal Point Processes

    Ye Yuan, Kris Kitani

    cs.CVcs.LGcs.ROarXiv:1907.04967v22019
  8. CausalWorld: A Robotic Manipulation Benchmark for Causal Structure and Transfer Learning

    Ossama Ahmed, Frederik Träuble, Anirudh Goyal +5

    cs.ROcs.LGstat.MLarXiv:2010.04296v22020
  9. Data Scarcity and Model Sparsity: Mixtures-of-Experts Overfit More to Repeated Data

    Atindra Jha, Margaret Li, Jure Leskovec +2

    cs.LGcs.CLarXiv:2609.11917v12026
  10. HOI Analysis: Integrating and Decomposing Human-Object Interaction

    Yong-Lu Li, Xinpeng Liu, Xiaoqian Wu +2

    cs.CVcs.LGeess.IVarXiv:2010.16219v22020
  11. Automated Data Slicing for Model Validation:A Big data - AI Integration Approach

    Yeounoh Chung, Tim Kraska, Neoklis Polyzotis +2

    cs.DBcs.LGarXiv:1807.06068v32018
  12. Denoising Diffusion Samplers

    Francisco Vargas, Will Grathwohl, Arnaud Doucet

    cs.LGstat.MLarXiv:2302.13834v22023
  13. Epistemic Neural Networks

    Ian Osband, Zheng Wen, Seyed Mohammad Asghari +4

    cs.LGcs.AIstat.MLarXiv:2107.08924v82021
  14. Defending Against Model Stealing Attacks with Adaptive Misinformation

    Sanjay Kariyappa, Moinuddin K Qureshi

    stat.MLcs.CRcs.LGarXiv:1911.07100v12019
  15. A Unified Per-Token Gating Family for On-Policy Distillation: FKL/RKL Mixing with Multi-Channel and Bias Coefficients

    Suwan Wu, Yumeng Lin, Pengcheng Yuan +1

    cs.AIcs.CLcs.LGarXiv:2609.11768v12026
  16. SIRF: A Spec-Internalized Risk Foundation Model for Industrial Content Risk Control

    Suwan Wu, Yumeng Lin, Pengcheng Yuan +1

    cs.AIcs.CLcs.LGarXiv:2609.11752v12026
  17. A Systematic Survey and Critical Review on Evaluating Large Language Models: Challenges, Limitations, and Recommendations

    Md Tahmid Rahman Laskar, Sawsan Alqahtani, M Saiful Bari +10

    cs.CLcs.AIcs.LGarXiv:2407.04069v22024
  18. The Last AI Built by Humans: Toward Genuine Recursive Self-Improvement

    Yi Duan, Ying Liu, Zirui Tang +30

    cs.LGcs.AIcs.CLarXiv:2609.11873v12026
  19. Deep Generalized Method of Moments for Instrumental Variable Analysis

    Andrew Bennett, Nathan Kallus, Tobias Schnabel

    stat.MLcs.LGecon.EMarXiv:1905.12495v22019
  20. Hidden Biases of End-to-End Driving Models

    Bernhard Jaeger, Kashyap Chitta, Andreas Geiger

    cs.CVcs.AIcs.LGarXiv:2306.07957v22023
  21. Stochastic Segmentation Networks: Modelling Spatially Correlated Aleatoric Uncertainty

    Miguel Monteiro, Loïc Le Folgoc, Daniel Coelho de Castro +5

    cs.CVcs.LGarXiv:2006.06015v22020
  22. IPGuard: Protecting Intellectual Property of Deep Neural Networks via Fingerprinting the Classification Boundary

    Xiaoyu Cao, Jinyuan Jia, Neil Zhenqiang Gong

    cs.CRcs.AIcs.LGarXiv:1910.12903v52019
  23. Why Does Post-Training Quantization Work?

    Yuxiang Chen, Michael Beyer, Jun Zhu +1

    cs.LGcs.CLarXiv:2609.11716v12026
  24. Survey on reinforcement learning for language processing

    Victor Uc-Cetina, Nicolas Navarro-Guerrero, Anabel Martin-Gonzalez +2

    cs.CLcs.AIcs.LGarXiv:2104.05565v32021
  25. VikingRAG: Accurate and Token-efficient Retrieval-augmented Generation over Structured Documents

    Peiyuan Gao, Gaoyuan Zhang, Haojie Qin +5

    cs.IRcs.AIcs.CLarXiv:2609.11390v12026
  26. EnergyStar++: Towards more accurate and explanatory building energy benchmarking

    Pandarasamy Arjunan, Kameshwar Poolla, Clayton Miller

    stat.APcs.LGeess.SYarXiv:1910.14563v22019
  27. Analysis of Bayesian Classification based Approaches for Android Malware Detection

    Suleiman Y. Yerima, Sakir Sezer, Gavin McWilliams

    cs.CRcs.LGarXiv:1608.05812v12016
  28. LILA: Calibration-Free Structured Pruning of Large Language Models via Latent Spectral Geometry

    Sankar Behera, Dhruv Singh, Anshika Agnihotri +3

    cs.LGcs.CLarXiv:2609.11163v12026
  29. Learning a bidirectional mapping between human whole-body motion and natural language using deep recurrent neural networks

    Matthias Plappert, Christian Mandery, Tamim Asfour

    cs.LGcs.CLcs.ROarXiv:1705.06400v22017
  30. Using machine learning to correct model error in data assimilation and forecast applications

    Alban Farchi, Patrick Laloyaux, Massimo Bonavita +1

    stat.MLcs.LGphysics.data-anarXiv:2010.12605v22020
  31. ChemBO: Bayesian Optimization of Small Organic Molecules with Synthesizable Recommendations

    Ksenia Korovina, Sailun Xu, Kirthevasan Kandasamy +4

    cs.LGphysics.chem-phstat.MLarXiv:1908.01425v22019
  32. Generative chemistry: drug discovery with deep learning generative models

    Yuemin Bian, Xiang-Qun Xie

    q-bio.BMcs.LGq-bio.QMarXiv:2008.09000v12020
  33. A Federated Learning Aggregation Algorithm for Pervasive Computing: Evaluation and Comparison

    Sannara Ek, François Portet, Philippe Lalanda +1

    cs.LGcs.AIcs.DCarXiv:2110.10223v12021
  34. MUtE: A Dual Framework for Concept Erasure and Counterfactual Interventions

    Antoine Saillenfest

    cs.LGcs.CLarXiv:2609.11253v12026
  35. A Review of Machine Learning Methods Applied to Structural Dynamics and Vibroacoustic

    Barbara Cunha, Christophe Droz, Abdelmalek Zine +2

    cs.LGcs.SDeess.ASarXiv:2204.06362v22022
  36. Toward General-Purpose Robots via Foundation Models: A Survey and Meta-Analysis

    Yafei Hu, Quanting Xie, Vidhi Jain +20

    cs.ROcs.AIcs.CVarXiv:2312.08782v32023
  37. Deep linear neural networks with arbitrary loss: All local minima are global

    Thomas Laurent, James von Brecht

    cs.LGstat.MLarXiv:1712.01473v22017
  38. REVA: Reusable Evidence View Aggregation for Context-Efficient RAG Serving

    Tuan Nguyen, Qiran Hu, Banruo Liu +3

    cs.LGcs.CLcs.IRarXiv:2609.11209v12026
  39. Learning Visuotactile Skills with Two Multifingered Hands

    Toru Lin, Yu Zhang, Qiyang Li +4

    cs.ROcs.AIcs.CVarXiv:2404.16823v22024
  40. FedBiOT: LLM Local Fine-tuning in Federated Learning without Full Model

    Feijie Wu, Zitao Li, Yaliang Li +2

    cs.LGcs.CLcs.DCarXiv:2406.17706v12024
  41. Iterative Instance Segmentation

    Ke Li, Bharath Hariharan, Jitendra Malik

    cs.CVcs.LGarXiv:1511.08498v32015
  42. The Oligarch Barely Steers Model Collapse in Multi-Model Ecosystems

    Yangze Liu, Zhongyi Han

    cs.AIcs.CLcs.LGarXiv:2609.11146v12026
  43. Bayesian Optimization with Machine Learning Algorithms Towards Anomaly Detection

    MohammadNoor Injadat, Fadi Salo, Ali Bou Nassif +2

    cs.LGcs.NIstat.MLarXiv:2008.02327v12020
  44. Transferability in Deep Learning: A Survey

    Junguang Jiang, Yang Shu, Jianmin Wang +1

    cs.LGarXiv:2201.05867v12022
  45. The information geometry of large language models is shared, learned, and controllable

    Dario Picozzi

    cs.LGcs.CLarXiv:2609.11063v12026
  46. Principal Graphs and Manifolds

    A. N. Gorban, A. Y. Zinovyev

    cs.LGcs.NEstat.MLarXiv:0809.0490v22008
  47. Spectrum Sensing Based on Deep Learning Classification for Cognitive Radios

    Shilian Zheng, Shichuan Chen, Peihan Qi +2

    eess.SPcs.LGarXiv:1909.06020v12019
  48. Goal Misgeneralization: Why Correct Specifications Aren't Enough For Correct Goals

    Rohin Shah, Vikrant Varma, Ramana Kumar +4

    cs.LGarXiv:2210.01790v22022
  49. ChatGPT Needs SPADE (Sustainability, PrivAcy, Digital divide, and Ethics) Evaluation: A Review

    Sunder Ali Khowaja, Parus Khuwaja, Kapal Dev +2

    cs.CYcs.AIcs.CLarXiv:2305.03123v42023
  50. A Comprehensive Survey of Foundation Models in Medicine

    Wasif Khan, Seowung Leem, Kyle B. See +3

    cs.LGcs.AIcs.CVarXiv:2406.10729v32024
  51. Story Imprinting: AI Assistants Absorb Traits from Human Characters They Resemble

    Jorio Cocola, Lev McKinney, Harry Mayne +2

    cs.LGcs.AIcs.CLarXiv:2609.10883v12026
  52. Model-Driven Deep Learning Based Channel Estimation and Feedback for Millimeter-Wave Massive Hybrid MIMO Systems

    Xisuo Ma, Zhen Gao, Feifei Gao +1

    cs.ITcs.AIcs.LGarXiv:2104.11052v32021
  53. Beyond Solver Verdicts: Generative Reward Models for Autoformalization

    Vikash Singh, Debargha Ganguly, Aman Goel +5

    cs.LGcs.CLarXiv:2609.11085v12026
  54. Unbiased split variable selection for random survival forests using maximally selected rank statistics

    Marvin N. Wright, Theresa Dankowski, Andreas Ziegler

    stat.MLcs.LGarXiv:1605.03391v22016
  55. New Evidence, Same Choice: Testing Physical Experiment Selection in Vision Language Models

    Sourajit Saha, Shubhashis Roy Dipta, Nobin Sarwar +4

    cs.CVcs.AIcs.CLarXiv:2609.11022v12026
  56. A Comprehensive Survey on Deep Music Generation: Multi-level Representations, Algorithms, Evaluations, and Future Directions

    Shulei Ji, Jing Luo, Xinyu Yang

    cs.SDcs.LGeess.ASarXiv:2011.06801v12020
  57. Empirical Evaluation of Membership Inference Attacks on NLP Text Classifiers: A Baseline Study on SST-2

    William Novak, Muhammad Abusaqer

    cs.CRcs.CLcs.LGarXiv:2609.10935v12026
  58. Non-convex Min-Max Optimization: Applications, Challenges, and Recent Theoretical Advances

    Meisam Razaviyayn, Tianjian Huang, Songtao Lu +3

    math.OCcs.LGstat.MLarXiv:2006.08141v22020
  59. Hardware Approximate Techniques for Deep Neural Network Accelerators: A Survey

    Giorgos Armeniakos, Georgios Zervakis, Dimitrios Soudris +1

    cs.ARcs.LGarXiv:2203.08737v12022
  60. CATCH: Channel-Aware multivariate Time Series Anomaly Detection via Frequency Patching

    Xingjian Wu, Xiangfei Qiu, Zhengyu Li +5

    cs.LGcs.AIarXiv:2410.12261v52024