Machine Learning

Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

10,201 to 10,260 of 19,908

  1. Semantic Entropy Probes: Robust and Cheap Hallucination Detection in LLMs

    Jannik Kossen, Jiatong Han, Muhammed Razzak +3

    cs.CLcs.AIcs.LGarXiv:2406.15927v12024
  2. Dual Cross-Attention Learning for Fine-Grained Visual Categorization and Object Re-Identification

    Haowei Zhu, Wenjing Ke, Dong Li +3

    cs.CVcs.AIcs.LGarXiv:2205.02151v12022
  3. Lies We Can See: Joint Verbal and Non-Verbal Deception by VLM Agents in Embodied Social Interactions

    Jaewoo Ahn, Junseo Kim, Hyunseo Kim +4

    cs.CLcs.AIcs.CVarXiv:2608.30428v12026
  4. Branch and Bound for Piecewise Linear Neural Network Verification

    Rudy Bunel, Jingyue Lu, Ilker Turkaslan +3

    cs.LGcs.LOstat.MLarXiv:1909.06588v52019
  5. How to Build a Graph-Based Deep Learning Architecture in Traffic Domain: A Survey

    Jiexia Ye, Juanjuan Zhao, Kejiang Ye +1

    eess.SPcs.LGarXiv:2005.11691v62020
  6. Orthogonal Random Features

    Felix X. Yu, Ananda Theertha Suresh, Krzysztof Choromanski +2

    cs.LGstat.MLarXiv:1610.09072v12016
  7. Few-Shot Knowledge Graph Completion

    Chuxu Zhang, Huaxiu Yao, Chao Huang +3

    cs.CLcs.AIcs.LGarXiv:1911.11298v12019
  8. The Illusion of Replacement: Rethinking Specialized Machine Learning Models in the Foundation Model Era

    Kiyan Rezaee

    cs.CLcs.AIcs.LGarXiv:2608.28980v12026
  9. The Hallucination Signal Is a Mean Shift: Why Simple Probes Suffice

    Jungseob Lee, Jaehyung Seo, Heuiseok Lim

    cs.CLcs.AIcs.LGarXiv:2608.28930v12026
  10. Scalable End-to-End Autonomous Vehicle Testing via Rare-event Simulation

    Matthew O'Kelly, Aman Sinha, Hongseok Namkoong +2

    cs.LGcs.ROstat.MLarXiv:1811.00145v32018
  11. Time Will Tell: New Outlooks and A Baseline for Temporal Multi-View 3D Object Detection

    Jinhyung Park, Chenfeng Xu, Shijia Yang +4

    cs.CVcs.AIcs.LGarXiv:2210.02443v12022
  12. Labeling Trick: A Theory of Using Graph Neural Networks for Multi-Node Representation Learning

    Muhan Zhang, Pan Li, Yinglong Xia +2

    cs.LGarXiv:2010.16103v52020
  13. Understanding and Mitigating Copying in Diffusion Models

    Gowthami Somepalli, Vasu Singla, Micah Goldblum +2

    cs.LGcs.CRcs.CVarXiv:2305.20086v12023
  14. Exact and Stable Covariance Estimation from Quadratic Sampling via Convex Programming

    Yuxin Chen, Yuejie Chi, Andrea Goldsmith

    cs.ITcs.LGmath.NAarXiv:1310.0807v52013
  15. Spectral Graph Convolutions for Population-based Disease Prediction

    Sarah Parisot, Sofia Ira Ktena, Enzo Ferrante +4

    stat.MLcs.LGarXiv:1703.03020v32017
  16. Improved Analysis of Score-based Generative Modeling: User-Friendly Bounds under Minimal Smoothness Assumptions

    Hongrui Chen, Holden Lee, Jianfeng Lu

    cs.LGarXiv:2211.01916v22022
  17. Estimating Counterfactual Treatment Outcomes over Time Through Adversarially Balanced Representations

    Ioana Bica, Ahmed M. Alaa, James Jordon +1

    cs.LGstat.MLarXiv:2002.04083v12020
  18. Data-Free Model Extraction

    Jean-Baptiste Truong, Pratyush Maini, Robert J. Walls +1

    cs.LGarXiv:2011.14779v22020
  19. Progressive Transformers for End-to-End Sign Language Production

    Ben Saunders, Necati Cihan Camgoz, Richard Bowden

    cs.CVcs.CLcs.LGarXiv:2004.14874v22020
  20. Deep Embedded Multi-view Clustering with Collaborative Training

    Jie Xu, Yazhou Ren, Guofeng Li +3

    cs.LGstat.MLarXiv:2007.13067v12020
  21. Leveraging Turn-taking Dynamics for Intent Recognition in Multi-party Conversations

    Galo Castillo-López, Alexis Lombard, Gaël de Chalendar +1

    cs.CLcs.LGarXiv:2608.28926v12026
  22. Will we run out of data? Limits of LLM scaling based on human-generated data

    Pablo Villalobos, Anson Ho, Jaime Sevilla +3

    cs.LGcs.AIcs.CLarXiv:2211.04325v22022
  23. Modality to Modality Translation: An Adversarial Representation Learning and Graph Fusion Network for Multimodal Fusion

    Sijie Mai, Haifeng Hu, Songlong Xing

    cs.CVcs.LGcs.MMarXiv:1911.07848v42019
  24. Topology and Geometry of Half-Rectified Network Optimization

    C. Daniel Freeman, Joan Bruna

    stat.MLcs.LGarXiv:1611.01540v42016
  25. Optimal Client Sampling for Federated Learning

    Wenlin Chen, Samuel Horvath, Peter Richtarik

    cs.LGcs.DCarXiv:2010.13723v32020
  26. Wave-ViT: Unifying Wavelet and Transformers for Visual Representation Learning

    Ting Yao, Yingwei Pan, Yehao Li +2

    cs.CVcs.LGarXiv:2207.04978v12022
  27. Moving the Mean Toward the Known Good, Not Beyond It: What Inference-Time Interventions and Weight Consolidation Buy in Open-Ended Generation

    Roberto I. Ono Filho

    cs.CLcs.AIcs.LGarXiv:2608.28886v12026
  28. Dynamic stochastic blockmodels for time-evolving social networks

    Kevin S. Xu, Alfred O. Hero

    cs.SIcs.LGphysics.soc-pharXiv:1403.0921v12014
  29. Text-To-4D Dynamic Scene Generation

    Uriel Singer, Shelly Sheynin, Adam Polyak +8

    cs.CVcs.AIcs.LGarXiv:2301.11280v12023
  30. TimeVAE: A Variational Auto-Encoder for Multivariate Time Series Generation

    Abhyuday Desai, Cynthia Freeman, Zuhui Wang +1

    cs.LGarXiv:2111.08095v32021
  31. The Curious Case of Hallucinations in Neural Machine Translation

    Vikas Raunak, Arul Menezes, Marcin Junczys-Dowmunt

    cs.CLcs.AIcs.LGarXiv:2104.06683v12021
  32. Equinox: neural networks in JAX via callable PyTrees and filtered transformations

    Patrick Kidger, Cristian Garcia

    cs.LGcs.PLarXiv:2111.00254v12021
  33. Tool Zero: Training Tool-Augmented LLMs via Pure RL from Scratch

    Yirong Zeng, Xiao Ding, Yutai Hou +9

    cs.LGcs.AIarXiv:2511.01934v22025
  34. Group-in-Group Policy Optimization for LLM Agent Training

    Lang Feng, Zhenghai Xue, Tingcong Liu +1

    cs.LGcs.AIarXiv:2505.10978v32025
  35. Degenerate Feedback Loops in Recommender Systems

    Ray Jiang, Silvia Chiappa, Tor Lattimore +2

    stat.MLcs.LGarXiv:1902.10730v32019
  36. The Feeling of Success: Does Touch Sensing Help Predict Grasp Outcomes?

    Roberto Calandra, Andrew Owens, Manu Upadhyaya +4

    cs.ROcs.CVcs.LGarXiv:1710.05512v22017
  37. Global canopy height regression and uncertainty estimation from GEDI LIDAR waveforms with deep ensembles

    Nico Lang, Nikolai Kalischek, John Armston +3

    cs.LGcs.CVarXiv:2103.03975v22021
  38. A rigor-matched audit of periodic-step layer skipping for efficient llm inference: conflayers versus swift, with a supplemental analysis of trained routing alternatives

    Prateek Kumar Sikdar, Arpan Ghosh

    cs.CLcs.AIcs.LGarXiv:2608.28846v12026
  39. Semi-supervised Multitask Learning for Sequence Labeling

    Marek Rei

    cs.CLcs.LGcs.NEarXiv:1704.07156v12017
  40. Generative replay with feedback connections as a general strategy for continual learning

    Gido M. van de Ven, Andreas S. Tolias

    cs.LGcs.AIcs.CVarXiv:1809.10635v22018
  41. Towards a Science of Human-AI Decision Making: A Survey of Empirical Studies

    Vivian Lai, Chacha Chen, Q. Vera Liao +2

    cs.AIcs.CLcs.CYarXiv:2112.11471v12021
  42. CXPlain: Causal Explanations for Model Interpretation under Uncertainty

    Patrick Schwab, Walter Karlen

    cs.LGstat.MLarXiv:1910.12336v12019
  43. giotto-tda: A Topological Data Analysis Toolkit for Machine Learning and Data Exploration

    Guillaume Tauzin, Umberto Lupo, Lewis Tunstall +6

    cs.LGmath.ATstat.MLarXiv:2004.02551v22020
  44. On learning to localize objects with minimal supervision

    Hyun Oh Song, Ross Girshick, Stefanie Jegelka +3

    cs.CVcs.LGarXiv:1403.1024v42014
  45. TopicRNN: A Recurrent Neural Network with Long-Range Semantic Dependency

    Adji B. Dieng, Chong Wang, Jianfeng Gao +1

    cs.CLcs.AIcs.LGarXiv:1611.01702v22016
  46. Test-Time Scaling for Scientific Equation Discovery

    Haowei Lin, Hubert Lim, Xiangyu Wang +2

    cs.CLcs.AIcs.LGarXiv:2608.28660v12026
  47. Lasso Screening Rules via Dual Polytope Projection

    Jie Wang, Peter Wonka, Jieping Ye

    cs.LGstat.MLarXiv:1211.3966v32012
  48. Measuring Sample Quality with Stein's Method

    Jackson Gorham, Lester Mackey

    stat.MLcs.LGmath.PRarXiv:1506.03039v62015
  49. MADLAD-400: A Multilingual And Document-Level Large Audited Dataset

    Sneha Kudugunta, Isaac Caswell, Biao Zhang +8

    cs.CLcs.LGarXiv:2309.04662v12023
  50. Timer: Generative Pre-trained Transformers Are Large Time Series Models

    Yong Liu, Haoran Zhang, Chenyu Li +3

    cs.LGstat.MLarXiv:2402.02368v32024
  51. Invariant Representations without Adversarial Training

    Daniel Moyer, Shuyang Gao, Rob Brekelmans +2

    cs.LGstat.MLarXiv:1805.09458v42018
  52. Perturbed Iterate Analysis for Asynchronous Stochastic Optimization

    Horia Mania, Xinghao Pan, Dimitris Papailiopoulos +3

    stat.MLcs.DCcs.DSarXiv:1507.06970v22015
  53. Wind Power Forecasting Considering Data Privacy Protection: A Federated Deep Reinforcement Learning Approach

    Yang Li, Ruinong Wang, Yuanzheng Li +2

    cs.LGeess.SYarXiv:2211.02674v12022
  54. Polymer Informatics: Current Status and Critical Next Steps

    Lihua Chen, Ghanshyam Pilania, Rohit Batra +4

    cond-mat.softcs.LGarXiv:2011.00508v12020
  55. Explainable Medical Imaging AI Needs Human-Centered Design: Guidelines and Evidence from a Systematic Review

    Haomin Chen, Catalina Gomez, Chien-Ming Huang +1

    cs.HCcs.CVcs.LGarXiv:2112.12596v42021
  56. Decentralized Collaborative Learning of Personalized Models over Networks

    Paul Vanhaesebrouck, Aurélien Bellet, Marc Tommasi

    cs.LGcs.AIcs.DCarXiv:1610.05202v22016
  57. Human Perceptions of Fairness in Algorithmic Decision Making: A Case Study of Criminal Risk Prediction

    Nina Grgić-Hlača, Elissa M. Redmiles, Krishna P. Gummadi +1

    stat.MLcs.CYcs.LGarXiv:1802.09548v12018
  58. Counterfactuals and Causability in Explainable Artificial Intelligence: Theory, Algorithms, and Applications

    Yu-Liang Chou, Catarina Moreira, Peter Bruza +2

    cs.AIcs.LGarXiv:2103.04244v22021
  59. SPICE, A Dataset of Drug-like Molecules and Peptides for Training Machine Learning Potentials

    Peter Eastman, Pavan Kumar Behara, David L. Dotson +9

    physics.chem-phcs.LGq-bio.BMarXiv:2209.10702v22022
  60. CBLUE: A Chinese Biomedical Language Understanding Evaluation Benchmark

    Ningyu Zhang, Mosha Chen, Zhen Bi +20

    cs.CLcs.AIcs.IRarXiv:2106.08087v62021