Machine Learning
Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
10,201 to 10,260 of 19,908
Semantic Entropy Probes: Robust and Cheap Hallucination Detection in LLMs
Jannik Kossen, Jiatong Han, Muhammed Razzak +3
cs.CLcs.AIcs.LGarXiv:2406.15927v12024Dual Cross-Attention Learning for Fine-Grained Visual Categorization and Object Re-Identification
Haowei Zhu, Wenjing Ke, Dong Li +3
cs.CVcs.AIcs.LGarXiv:2205.02151v12022Lies We Can See: Joint Verbal and Non-Verbal Deception by VLM Agents in Embodied Social Interactions
Jaewoo Ahn, Junseo Kim, Hyunseo Kim +4
cs.CLcs.AIcs.CVarXiv:2608.30428v12026Branch and Bound for Piecewise Linear Neural Network Verification
Rudy Bunel, Jingyue Lu, Ilker Turkaslan +3
cs.LGcs.LOstat.MLarXiv:1909.06588v52019How to Build a Graph-Based Deep Learning Architecture in Traffic Domain: A Survey
Jiexia Ye, Juanjuan Zhao, Kejiang Ye +1
eess.SPcs.LGarXiv:2005.11691v62020Orthogonal Random Features
Felix X. Yu, Ananda Theertha Suresh, Krzysztof Choromanski +2
cs.LGstat.MLarXiv:1610.09072v12016Few-Shot Knowledge Graph Completion
Chuxu Zhang, Huaxiu Yao, Chao Huang +3
cs.CLcs.AIcs.LGarXiv:1911.11298v12019The Illusion of Replacement: Rethinking Specialized Machine Learning Models in the Foundation Model Era
Kiyan Rezaee
cs.CLcs.AIcs.LGarXiv:2608.28980v12026The Hallucination Signal Is a Mean Shift: Why Simple Probes Suffice
Jungseob Lee, Jaehyung Seo, Heuiseok Lim
cs.CLcs.AIcs.LGarXiv:2608.28930v12026Scalable End-to-End Autonomous Vehicle Testing via Rare-event Simulation
Matthew O'Kelly, Aman Sinha, Hongseok Namkoong +2
cs.LGcs.ROstat.MLarXiv:1811.00145v32018Time Will Tell: New Outlooks and A Baseline for Temporal Multi-View 3D Object Detection
Jinhyung Park, Chenfeng Xu, Shijia Yang +4
cs.CVcs.AIcs.LGarXiv:2210.02443v12022Labeling Trick: A Theory of Using Graph Neural Networks for Multi-Node Representation Learning
Muhan Zhang, Pan Li, Yinglong Xia +2
cs.LGarXiv:2010.16103v52020Understanding and Mitigating Copying in Diffusion Models
Gowthami Somepalli, Vasu Singla, Micah Goldblum +2
cs.LGcs.CRcs.CVarXiv:2305.20086v12023Exact and Stable Covariance Estimation from Quadratic Sampling via Convex Programming
Yuxin Chen, Yuejie Chi, Andrea Goldsmith
cs.ITcs.LGmath.NAarXiv:1310.0807v52013Spectral Graph Convolutions for Population-based Disease Prediction
Sarah Parisot, Sofia Ira Ktena, Enzo Ferrante +4
stat.MLcs.LGarXiv:1703.03020v32017Improved Analysis of Score-based Generative Modeling: User-Friendly Bounds under Minimal Smoothness Assumptions
Hongrui Chen, Holden Lee, Jianfeng Lu
cs.LGarXiv:2211.01916v22022Estimating Counterfactual Treatment Outcomes over Time Through Adversarially Balanced Representations
Ioana Bica, Ahmed M. Alaa, James Jordon +1
cs.LGstat.MLarXiv:2002.04083v12020Data-Free Model Extraction
Jean-Baptiste Truong, Pratyush Maini, Robert J. Walls +1
cs.LGarXiv:2011.14779v22020Progressive Transformers for End-to-End Sign Language Production
Ben Saunders, Necati Cihan Camgoz, Richard Bowden
cs.CVcs.CLcs.LGarXiv:2004.14874v22020Deep Embedded Multi-view Clustering with Collaborative Training
Jie Xu, Yazhou Ren, Guofeng Li +3
cs.LGstat.MLarXiv:2007.13067v12020Leveraging Turn-taking Dynamics for Intent Recognition in Multi-party Conversations
Galo Castillo-López, Alexis Lombard, Gaël de Chalendar +1
cs.CLcs.LGarXiv:2608.28926v12026Will we run out of data? Limits of LLM scaling based on human-generated data
Pablo Villalobos, Anson Ho, Jaime Sevilla +3
cs.LGcs.AIcs.CLarXiv:2211.04325v22022Modality to Modality Translation: An Adversarial Representation Learning and Graph Fusion Network for Multimodal Fusion
Sijie Mai, Haifeng Hu, Songlong Xing
cs.CVcs.LGcs.MMarXiv:1911.07848v42019Topology and Geometry of Half-Rectified Network Optimization
C. Daniel Freeman, Joan Bruna
stat.MLcs.LGarXiv:1611.01540v42016Optimal Client Sampling for Federated Learning
Wenlin Chen, Samuel Horvath, Peter Richtarik
cs.LGcs.DCarXiv:2010.13723v32020Wave-ViT: Unifying Wavelet and Transformers for Visual Representation Learning
Ting Yao, Yingwei Pan, Yehao Li +2
cs.CVcs.LGarXiv:2207.04978v12022Moving the Mean Toward the Known Good, Not Beyond It: What Inference-Time Interventions and Weight Consolidation Buy in Open-Ended Generation
Roberto I. Ono Filho
cs.CLcs.AIcs.LGarXiv:2608.28886v12026Dynamic stochastic blockmodels for time-evolving social networks
Kevin S. Xu, Alfred O. Hero
cs.SIcs.LGphysics.soc-pharXiv:1403.0921v12014Text-To-4D Dynamic Scene Generation
Uriel Singer, Shelly Sheynin, Adam Polyak +8
cs.CVcs.AIcs.LGarXiv:2301.11280v12023TimeVAE: A Variational Auto-Encoder for Multivariate Time Series Generation
Abhyuday Desai, Cynthia Freeman, Zuhui Wang +1
cs.LGarXiv:2111.08095v32021The Curious Case of Hallucinations in Neural Machine Translation
Vikas Raunak, Arul Menezes, Marcin Junczys-Dowmunt
cs.CLcs.AIcs.LGarXiv:2104.06683v12021Equinox: neural networks in JAX via callable PyTrees and filtered transformations
Patrick Kidger, Cristian Garcia
cs.LGcs.PLarXiv:2111.00254v12021Tool Zero: Training Tool-Augmented LLMs via Pure RL from Scratch
Yirong Zeng, Xiao Ding, Yutai Hou +9
cs.LGcs.AIarXiv:2511.01934v22025Group-in-Group Policy Optimization for LLM Agent Training
Lang Feng, Zhenghai Xue, Tingcong Liu +1
cs.LGcs.AIarXiv:2505.10978v32025Degenerate Feedback Loops in Recommender Systems
Ray Jiang, Silvia Chiappa, Tor Lattimore +2
stat.MLcs.LGarXiv:1902.10730v32019The Feeling of Success: Does Touch Sensing Help Predict Grasp Outcomes?
Roberto Calandra, Andrew Owens, Manu Upadhyaya +4
cs.ROcs.CVcs.LGarXiv:1710.05512v22017Global canopy height regression and uncertainty estimation from GEDI LIDAR waveforms with deep ensembles
Nico Lang, Nikolai Kalischek, John Armston +3
cs.LGcs.CVarXiv:2103.03975v22021A rigor-matched audit of periodic-step layer skipping for efficient llm inference: conflayers versus swift, with a supplemental analysis of trained routing alternatives
Prateek Kumar Sikdar, Arpan Ghosh
cs.CLcs.AIcs.LGarXiv:2608.28846v12026Semi-supervised Multitask Learning for Sequence Labeling
Marek Rei
cs.CLcs.LGcs.NEarXiv:1704.07156v12017Generative replay with feedback connections as a general strategy for continual learning
Gido M. van de Ven, Andreas S. Tolias
cs.LGcs.AIcs.CVarXiv:1809.10635v22018Towards a Science of Human-AI Decision Making: A Survey of Empirical Studies
Vivian Lai, Chacha Chen, Q. Vera Liao +2
cs.AIcs.CLcs.CYarXiv:2112.11471v12021CXPlain: Causal Explanations for Model Interpretation under Uncertainty
Patrick Schwab, Walter Karlen
cs.LGstat.MLarXiv:1910.12336v12019giotto-tda: A Topological Data Analysis Toolkit for Machine Learning and Data Exploration
Guillaume Tauzin, Umberto Lupo, Lewis Tunstall +6
cs.LGmath.ATstat.MLarXiv:2004.02551v22020On learning to localize objects with minimal supervision
Hyun Oh Song, Ross Girshick, Stefanie Jegelka +3
cs.CVcs.LGarXiv:1403.1024v42014TopicRNN: A Recurrent Neural Network with Long-Range Semantic Dependency
Adji B. Dieng, Chong Wang, Jianfeng Gao +1
cs.CLcs.AIcs.LGarXiv:1611.01702v22016Test-Time Scaling for Scientific Equation Discovery
Haowei Lin, Hubert Lim, Xiangyu Wang +2
cs.CLcs.AIcs.LGarXiv:2608.28660v12026Lasso Screening Rules via Dual Polytope Projection
Jie Wang, Peter Wonka, Jieping Ye
cs.LGstat.MLarXiv:1211.3966v32012Measuring Sample Quality with Stein's Method
Jackson Gorham, Lester Mackey
stat.MLcs.LGmath.PRarXiv:1506.03039v62015MADLAD-400: A Multilingual And Document-Level Large Audited Dataset
Sneha Kudugunta, Isaac Caswell, Biao Zhang +8
cs.CLcs.LGarXiv:2309.04662v12023Timer: Generative Pre-trained Transformers Are Large Time Series Models
Yong Liu, Haoran Zhang, Chenyu Li +3
cs.LGstat.MLarXiv:2402.02368v32024Invariant Representations without Adversarial Training
Daniel Moyer, Shuyang Gao, Rob Brekelmans +2
cs.LGstat.MLarXiv:1805.09458v42018Perturbed Iterate Analysis for Asynchronous Stochastic Optimization
Horia Mania, Xinghao Pan, Dimitris Papailiopoulos +3
stat.MLcs.DCcs.DSarXiv:1507.06970v22015Wind Power Forecasting Considering Data Privacy Protection: A Federated Deep Reinforcement Learning Approach
Yang Li, Ruinong Wang, Yuanzheng Li +2
cs.LGeess.SYarXiv:2211.02674v12022Polymer Informatics: Current Status and Critical Next Steps
Lihua Chen, Ghanshyam Pilania, Rohit Batra +4
cond-mat.softcs.LGarXiv:2011.00508v12020Explainable Medical Imaging AI Needs Human-Centered Design: Guidelines and Evidence from a Systematic Review
Haomin Chen, Catalina Gomez, Chien-Ming Huang +1
cs.HCcs.CVcs.LGarXiv:2112.12596v42021Decentralized Collaborative Learning of Personalized Models over Networks
Paul Vanhaesebrouck, Aurélien Bellet, Marc Tommasi
cs.LGcs.AIcs.DCarXiv:1610.05202v22016Human Perceptions of Fairness in Algorithmic Decision Making: A Case Study of Criminal Risk Prediction
Nina Grgić-Hlača, Elissa M. Redmiles, Krishna P. Gummadi +1
stat.MLcs.CYcs.LGarXiv:1802.09548v12018Counterfactuals and Causability in Explainable Artificial Intelligence: Theory, Algorithms, and Applications
Yu-Liang Chou, Catarina Moreira, Peter Bruza +2
cs.AIcs.LGarXiv:2103.04244v22021SPICE, A Dataset of Drug-like Molecules and Peptides for Training Machine Learning Potentials
Peter Eastman, Pavan Kumar Behara, David L. Dotson +9
physics.chem-phcs.LGq-bio.BMarXiv:2209.10702v22022CBLUE: A Chinese Biomedical Language Understanding Evaluation Benchmark
Ningyu Zhang, Mosha Chen, Zhen Bi +20
cs.CLcs.AIcs.IRarXiv:2106.08087v62021