Machine Learning
Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
1,561 to 1,620 of 20,199
Universal Jailbreak Backdoors from Poisoned Human Feedback
Javier Rando, Florian Tramèr
cs.AIcs.CLcs.CRarXiv:2311.14455v42023Zero-shot rib design: merging training-free generative prior with topology optimization
Yongmin Kwon, Namwoo Kang
cs.LGphysics.app-pharXiv:2609.10643v12026The People's Speech: A Large-Scale Diverse English Speech Recognition Dataset for Commercial Usage
Daniel Galvez, Greg Diamos, Juan Ciro +7
cs.LGstat.MLarXiv:2111.09344v12021CALF: Aligning LLMs for Time Series Forecasting via Cross-modal Fine-Tuning
Peiyuan Liu, Hang Guo, Tao Dai +5
cs.LGcs.CLarXiv:2403.07300v32024Halo: Improving forecast accuracy through heteroscedastic estimation
Adam Cataldo
cs.LGarXiv:2609.10589v12026M3-Former: Multimodal Transformer with Mixture-of-Experts for Long-Term Vessel Trajectory Prediction
Wenzhe Jin, Haina Tang
cs.LGcs.CVarXiv:2609.10559v12026Diverse Trajectory Forecasting with Determinantal Point Processes
Ye Yuan, Kris Kitani
cs.CVcs.LGcs.ROarXiv:1907.04967v22019CausalWorld: A Robotic Manipulation Benchmark for Causal Structure and Transfer Learning
Ossama Ahmed, Frederik Träuble, Anirudh Goyal +5
cs.ROcs.LGstat.MLarXiv:2010.04296v22020Data Scarcity and Model Sparsity: Mixtures-of-Experts Overfit More to Repeated Data
Atindra Jha, Margaret Li, Jure Leskovec +2
cs.LGcs.CLarXiv:2609.11917v12026HOI Analysis: Integrating and Decomposing Human-Object Interaction
Yong-Lu Li, Xinpeng Liu, Xiaoqian Wu +2
cs.CVcs.LGeess.IVarXiv:2010.16219v22020Automated Data Slicing for Model Validation:A Big data - AI Integration Approach
Yeounoh Chung, Tim Kraska, Neoklis Polyzotis +2
cs.DBcs.LGarXiv:1807.06068v32018Denoising Diffusion Samplers
Francisco Vargas, Will Grathwohl, Arnaud Doucet
cs.LGstat.MLarXiv:2302.13834v22023Epistemic Neural Networks
Ian Osband, Zheng Wen, Seyed Mohammad Asghari +4
cs.LGcs.AIstat.MLarXiv:2107.08924v82021Defending Against Model Stealing Attacks with Adaptive Misinformation
Sanjay Kariyappa, Moinuddin K Qureshi
stat.MLcs.CRcs.LGarXiv:1911.07100v12019A Unified Per-Token Gating Family for On-Policy Distillation: FKL/RKL Mixing with Multi-Channel and Bias Coefficients
Suwan Wu, Yumeng Lin, Pengcheng Yuan +1
cs.AIcs.CLcs.LGarXiv:2609.11768v12026SIRF: A Spec-Internalized Risk Foundation Model for Industrial Content Risk Control
Suwan Wu, Yumeng Lin, Pengcheng Yuan +1
cs.AIcs.CLcs.LGarXiv:2609.11752v12026A Systematic Survey and Critical Review on Evaluating Large Language Models: Challenges, Limitations, and Recommendations
Md Tahmid Rahman Laskar, Sawsan Alqahtani, M Saiful Bari +10
cs.CLcs.AIcs.LGarXiv:2407.04069v22024The Last AI Built by Humans: Toward Genuine Recursive Self-Improvement
Yi Duan, Ying Liu, Zirui Tang +30
cs.LGcs.AIcs.CLarXiv:2609.11873v12026Deep Generalized Method of Moments for Instrumental Variable Analysis
Andrew Bennett, Nathan Kallus, Tobias Schnabel
stat.MLcs.LGecon.EMarXiv:1905.12495v22019Hidden Biases of End-to-End Driving Models
Bernhard Jaeger, Kashyap Chitta, Andreas Geiger
cs.CVcs.AIcs.LGarXiv:2306.07957v22023Stochastic Segmentation Networks: Modelling Spatially Correlated Aleatoric Uncertainty
Miguel Monteiro, Loïc Le Folgoc, Daniel Coelho de Castro +5
cs.CVcs.LGarXiv:2006.06015v22020IPGuard: Protecting Intellectual Property of Deep Neural Networks via Fingerprinting the Classification Boundary
Xiaoyu Cao, Jinyuan Jia, Neil Zhenqiang Gong
cs.CRcs.AIcs.LGarXiv:1910.12903v52019Why Does Post-Training Quantization Work?
Yuxiang Chen, Michael Beyer, Jun Zhu +1
cs.LGcs.CLarXiv:2609.11716v12026Survey on reinforcement learning for language processing
Victor Uc-Cetina, Nicolas Navarro-Guerrero, Anabel Martin-Gonzalez +2
cs.CLcs.AIcs.LGarXiv:2104.05565v32021VikingRAG: Accurate and Token-efficient Retrieval-augmented Generation over Structured Documents
Peiyuan Gao, Gaoyuan Zhang, Haojie Qin +5
cs.IRcs.AIcs.CLarXiv:2609.11390v12026EnergyStar++: Towards more accurate and explanatory building energy benchmarking
Pandarasamy Arjunan, Kameshwar Poolla, Clayton Miller
stat.APcs.LGeess.SYarXiv:1910.14563v22019Analysis of Bayesian Classification based Approaches for Android Malware Detection
Suleiman Y. Yerima, Sakir Sezer, Gavin McWilliams
cs.CRcs.LGarXiv:1608.05812v12016LILA: Calibration-Free Structured Pruning of Large Language Models via Latent Spectral Geometry
Sankar Behera, Dhruv Singh, Anshika Agnihotri +3
cs.LGcs.CLarXiv:2609.11163v12026Learning a bidirectional mapping between human whole-body motion and natural language using deep recurrent neural networks
Matthias Plappert, Christian Mandery, Tamim Asfour
cs.LGcs.CLcs.ROarXiv:1705.06400v22017Using machine learning to correct model error in data assimilation and forecast applications
Alban Farchi, Patrick Laloyaux, Massimo Bonavita +1
stat.MLcs.LGphysics.data-anarXiv:2010.12605v22020ChemBO: Bayesian Optimization of Small Organic Molecules with Synthesizable Recommendations
Ksenia Korovina, Sailun Xu, Kirthevasan Kandasamy +4
cs.LGphysics.chem-phstat.MLarXiv:1908.01425v22019Generative chemistry: drug discovery with deep learning generative models
Yuemin Bian, Xiang-Qun Xie
q-bio.BMcs.LGq-bio.QMarXiv:2008.09000v12020A Federated Learning Aggregation Algorithm for Pervasive Computing: Evaluation and Comparison
Sannara Ek, François Portet, Philippe Lalanda +1
cs.LGcs.AIcs.DCarXiv:2110.10223v12021MUtE: A Dual Framework for Concept Erasure and Counterfactual Interventions
Antoine Saillenfest
cs.LGcs.CLarXiv:2609.11253v12026A Review of Machine Learning Methods Applied to Structural Dynamics and Vibroacoustic
Barbara Cunha, Christophe Droz, Abdelmalek Zine +2
cs.LGcs.SDeess.ASarXiv:2204.06362v22022Toward General-Purpose Robots via Foundation Models: A Survey and Meta-Analysis
Yafei Hu, Quanting Xie, Vidhi Jain +20
cs.ROcs.AIcs.CVarXiv:2312.08782v32023Deep linear neural networks with arbitrary loss: All local minima are global
Thomas Laurent, James von Brecht
cs.LGstat.MLarXiv:1712.01473v22017REVA: Reusable Evidence View Aggregation for Context-Efficient RAG Serving
Tuan Nguyen, Qiran Hu, Banruo Liu +3
cs.LGcs.CLcs.IRarXiv:2609.11209v12026Learning Visuotactile Skills with Two Multifingered Hands
Toru Lin, Yu Zhang, Qiyang Li +4
cs.ROcs.AIcs.CVarXiv:2404.16823v22024FedBiOT: LLM Local Fine-tuning in Federated Learning without Full Model
Feijie Wu, Zitao Li, Yaliang Li +2
cs.LGcs.CLcs.DCarXiv:2406.17706v12024Iterative Instance Segmentation
Ke Li, Bharath Hariharan, Jitendra Malik
cs.CVcs.LGarXiv:1511.08498v32015The Oligarch Barely Steers Model Collapse in Multi-Model Ecosystems
Yangze Liu, Zhongyi Han
cs.AIcs.CLcs.LGarXiv:2609.11146v12026Bayesian Optimization with Machine Learning Algorithms Towards Anomaly Detection
MohammadNoor Injadat, Fadi Salo, Ali Bou Nassif +2
cs.LGcs.NIstat.MLarXiv:2008.02327v12020Transferability in Deep Learning: A Survey
Junguang Jiang, Yang Shu, Jianmin Wang +1
cs.LGarXiv:2201.05867v12022The information geometry of large language models is shared, learned, and controllable
Dario Picozzi
cs.LGcs.CLarXiv:2609.11063v12026Principal Graphs and Manifolds
A. N. Gorban, A. Y. Zinovyev
cs.LGcs.NEstat.MLarXiv:0809.0490v22008Spectrum Sensing Based on Deep Learning Classification for Cognitive Radios
Shilian Zheng, Shichuan Chen, Peihan Qi +2
eess.SPcs.LGarXiv:1909.06020v12019Goal Misgeneralization: Why Correct Specifications Aren't Enough For Correct Goals
Rohin Shah, Vikrant Varma, Ramana Kumar +4
cs.LGarXiv:2210.01790v22022ChatGPT Needs SPADE (Sustainability, PrivAcy, Digital divide, and Ethics) Evaluation: A Review
Sunder Ali Khowaja, Parus Khuwaja, Kapal Dev +2
cs.CYcs.AIcs.CLarXiv:2305.03123v42023A Comprehensive Survey of Foundation Models in Medicine
Wasif Khan, Seowung Leem, Kyle B. See +3
cs.LGcs.AIcs.CVarXiv:2406.10729v32024Story Imprinting: AI Assistants Absorb Traits from Human Characters They Resemble
Jorio Cocola, Lev McKinney, Harry Mayne +2
cs.LGcs.AIcs.CLarXiv:2609.10883v12026Model-Driven Deep Learning Based Channel Estimation and Feedback for Millimeter-Wave Massive Hybrid MIMO Systems
Xisuo Ma, Zhen Gao, Feifei Gao +1
cs.ITcs.AIcs.LGarXiv:2104.11052v32021Beyond Solver Verdicts: Generative Reward Models for Autoformalization
Vikash Singh, Debargha Ganguly, Aman Goel +5
cs.LGcs.CLarXiv:2609.11085v12026Unbiased split variable selection for random survival forests using maximally selected rank statistics
Marvin N. Wright, Theresa Dankowski, Andreas Ziegler
stat.MLcs.LGarXiv:1605.03391v22016New Evidence, Same Choice: Testing Physical Experiment Selection in Vision Language Models
Sourajit Saha, Shubhashis Roy Dipta, Nobin Sarwar +4
cs.CVcs.AIcs.CLarXiv:2609.11022v12026A Comprehensive Survey on Deep Music Generation: Multi-level Representations, Algorithms, Evaluations, and Future Directions
Shulei Ji, Jing Luo, Xinyu Yang
cs.SDcs.LGeess.ASarXiv:2011.06801v12020Empirical Evaluation of Membership Inference Attacks on NLP Text Classifiers: A Baseline Study on SST-2
William Novak, Muhammad Abusaqer
cs.CRcs.CLcs.LGarXiv:2609.10935v12026Non-convex Min-Max Optimization: Applications, Challenges, and Recent Theoretical Advances
Meisam Razaviyayn, Tianjian Huang, Songtao Lu +3
math.OCcs.LGstat.MLarXiv:2006.08141v22020Hardware Approximate Techniques for Deep Neural Network Accelerators: A Survey
Giorgos Armeniakos, Georgios Zervakis, Dimitrios Soudris +1
cs.ARcs.LGarXiv:2203.08737v12022CATCH: Channel-Aware multivariate Time Series Anomaly Detection via Frequency Patching
Xingjian Wu, Xiangfei Qiu, Zhengyu Li +5
cs.LGcs.AIarXiv:2410.12261v52024