Machine Learning

Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

10,501 to 10,560 of 20,199

  1. Grammar-Constrained Decoding for Structured NLP Tasks without Finetuning

    Saibo Geng, Martin Josifoski, Maxime Peyrard +1

    cs.CLcs.AIcs.LGarXiv:2305.13971v62023
  2. Adaptive Machine Unlearning

    Varun Gupta, Christopher Jung, Seth Neel +3

    cs.LGstat.MLarXiv:2106.04378v12021
  3. Who Said What: Modeling Individual Labelers Improves Classification

    Melody Y. Guan, Varun Gulshan, Andrew M. Dai +1

    cs.LGcs.CVarXiv:1703.08774v22017
  4. Apprenticeship Learning using Inverse Reinforcement Learning and Gradient Methods

    Gergely Neu, Csaba Szepesvari

    cs.LGstat.MLarXiv:1206.5264v12012
  5. Semantic Entropy Probes: Robust and Cheap Hallucination Detection in LLMs

    Jannik Kossen, Jiatong Han, Muhammed Razzak +3

    cs.CLcs.AIcs.LGarXiv:2406.15927v12024
  6. Dual Cross-Attention Learning for Fine-Grained Visual Categorization and Object Re-Identification

    Haowei Zhu, Wenjing Ke, Dong Li +3

    cs.CVcs.AIcs.LGarXiv:2205.02151v12022
  7. Lies We Can See: Joint Verbal and Non-Verbal Deception by VLM Agents in Embodied Social Interactions

    Jaewoo Ahn, Junseo Kim, Hyunseo Kim +4

    cs.CLcs.AIcs.CVarXiv:2608.30428v12026
  8. Branch and Bound for Piecewise Linear Neural Network Verification

    Rudy Bunel, Jingyue Lu, Ilker Turkaslan +3

    cs.LGcs.LOstat.MLarXiv:1909.06588v52019
  9. How to Build a Graph-Based Deep Learning Architecture in Traffic Domain: A Survey

    Jiexia Ye, Juanjuan Zhao, Kejiang Ye +1

    eess.SPcs.LGarXiv:2005.11691v62020
  10. Orthogonal Random Features

    Felix X. Yu, Ananda Theertha Suresh, Krzysztof Choromanski +2

    cs.LGstat.MLarXiv:1610.09072v12016
  11. Few-Shot Knowledge Graph Completion

    Chuxu Zhang, Huaxiu Yao, Chao Huang +3

    cs.CLcs.AIcs.LGarXiv:1911.11298v12019
  12. The Illusion of Replacement: Rethinking Specialized Machine Learning Models in the Foundation Model Era

    Kiyan Rezaee

    cs.CLcs.AIcs.LGarXiv:2608.28980v12026
  13. The Hallucination Signal Is a Mean Shift: Why Simple Probes Suffice

    Jungseob Lee, Jaehyung Seo, Heuiseok Lim

    cs.CLcs.AIcs.LGarXiv:2608.28930v12026
  14. Scalable End-to-End Autonomous Vehicle Testing via Rare-event Simulation

    Matthew O'Kelly, Aman Sinha, Hongseok Namkoong +2

    cs.LGcs.ROstat.MLarXiv:1811.00145v32018
  15. Time Will Tell: New Outlooks and A Baseline for Temporal Multi-View 3D Object Detection

    Jinhyung Park, Chenfeng Xu, Shijia Yang +4

    cs.CVcs.AIcs.LGarXiv:2210.02443v12022
  16. Labeling Trick: A Theory of Using Graph Neural Networks for Multi-Node Representation Learning

    Muhan Zhang, Pan Li, Yinglong Xia +2

    cs.LGarXiv:2010.16103v52020
  17. Understanding and Mitigating Copying in Diffusion Models

    Gowthami Somepalli, Vasu Singla, Micah Goldblum +2

    cs.LGcs.CRcs.CVarXiv:2305.20086v12023
  18. Exact and Stable Covariance Estimation from Quadratic Sampling via Convex Programming

    Yuxin Chen, Yuejie Chi, Andrea Goldsmith

    cs.ITcs.LGmath.NAarXiv:1310.0807v52013
  19. Spectral Graph Convolutions for Population-based Disease Prediction

    Sarah Parisot, Sofia Ira Ktena, Enzo Ferrante +4

    stat.MLcs.LGarXiv:1703.03020v32017
  20. Improved Analysis of Score-based Generative Modeling: User-Friendly Bounds under Minimal Smoothness Assumptions

    Hongrui Chen, Holden Lee, Jianfeng Lu

    cs.LGarXiv:2211.01916v22022
  21. Estimating Counterfactual Treatment Outcomes over Time Through Adversarially Balanced Representations

    Ioana Bica, Ahmed M. Alaa, James Jordon +1

    cs.LGstat.MLarXiv:2002.04083v12020
  22. Data-Free Model Extraction

    Jean-Baptiste Truong, Pratyush Maini, Robert J. Walls +1

    cs.LGarXiv:2011.14779v22020
  23. Progressive Transformers for End-to-End Sign Language Production

    Ben Saunders, Necati Cihan Camgoz, Richard Bowden

    cs.CVcs.CLcs.LGarXiv:2004.14874v22020
  24. Deep Embedded Multi-view Clustering with Collaborative Training

    Jie Xu, Yazhou Ren, Guofeng Li +3

    cs.LGstat.MLarXiv:2007.13067v12020
  25. Leveraging Turn-taking Dynamics for Intent Recognition in Multi-party Conversations

    Galo Castillo-López, Alexis Lombard, Gaël de Chalendar +1

    cs.CLcs.LGarXiv:2608.28926v12026
  26. Will we run out of data? Limits of LLM scaling based on human-generated data

    Pablo Villalobos, Anson Ho, Jaime Sevilla +3

    cs.LGcs.AIcs.CLarXiv:2211.04325v22022
  27. Modality to Modality Translation: An Adversarial Representation Learning and Graph Fusion Network for Multimodal Fusion

    Sijie Mai, Haifeng Hu, Songlong Xing

    cs.CVcs.LGcs.MMarXiv:1911.07848v42019
  28. Topology and Geometry of Half-Rectified Network Optimization

    C. Daniel Freeman, Joan Bruna

    stat.MLcs.LGarXiv:1611.01540v42016
  29. Optimal Client Sampling for Federated Learning

    Wenlin Chen, Samuel Horvath, Peter Richtarik

    cs.LGcs.DCarXiv:2010.13723v32020
  30. Wave-ViT: Unifying Wavelet and Transformers for Visual Representation Learning

    Ting Yao, Yingwei Pan, Yehao Li +2

    cs.CVcs.LGarXiv:2207.04978v12022
  31. Moving the Mean Toward the Known Good, Not Beyond It: What Inference-Time Interventions and Weight Consolidation Buy in Open-Ended Generation

    Roberto I. Ono Filho

    cs.CLcs.AIcs.LGarXiv:2608.28886v12026
  32. Dynamic stochastic blockmodels for time-evolving social networks

    Kevin S. Xu, Alfred O. Hero

    cs.SIcs.LGphysics.soc-pharXiv:1403.0921v12014
  33. Text-To-4D Dynamic Scene Generation

    Uriel Singer, Shelly Sheynin, Adam Polyak +8

    cs.CVcs.AIcs.LGarXiv:2301.11280v12023
  34. TimeVAE: A Variational Auto-Encoder for Multivariate Time Series Generation

    Abhyuday Desai, Cynthia Freeman, Zuhui Wang +1

    cs.LGarXiv:2111.08095v32021
  35. The Curious Case of Hallucinations in Neural Machine Translation

    Vikas Raunak, Arul Menezes, Marcin Junczys-Dowmunt

    cs.CLcs.AIcs.LGarXiv:2104.06683v12021
  36. Equinox: neural networks in JAX via callable PyTrees and filtered transformations

    Patrick Kidger, Cristian Garcia

    cs.LGcs.PLarXiv:2111.00254v12021
  37. Tool Zero: Training Tool-Augmented LLMs via Pure RL from Scratch

    Yirong Zeng, Xiao Ding, Yutai Hou +9

    cs.LGcs.AIarXiv:2511.01934v22025
  38. Group-in-Group Policy Optimization for LLM Agent Training

    Lang Feng, Zhenghai Xue, Tingcong Liu +1

    cs.LGcs.AIarXiv:2505.10978v32025
  39. Degenerate Feedback Loops in Recommender Systems

    Ray Jiang, Silvia Chiappa, Tor Lattimore +2

    stat.MLcs.LGarXiv:1902.10730v32019
  40. The Feeling of Success: Does Touch Sensing Help Predict Grasp Outcomes?

    Roberto Calandra, Andrew Owens, Manu Upadhyaya +4

    cs.ROcs.CVcs.LGarXiv:1710.05512v22017
  41. Global canopy height regression and uncertainty estimation from GEDI LIDAR waveforms with deep ensembles

    Nico Lang, Nikolai Kalischek, John Armston +3

    cs.LGcs.CVarXiv:2103.03975v22021
  42. A rigor-matched audit of periodic-step layer skipping for efficient llm inference: conflayers versus swift, with a supplemental analysis of trained routing alternatives

    Prateek Kumar Sikdar, Arpan Ghosh

    cs.CLcs.AIcs.LGarXiv:2608.28846v12026
  43. Semi-supervised Multitask Learning for Sequence Labeling

    Marek Rei

    cs.CLcs.LGcs.NEarXiv:1704.07156v12017
  44. Generative replay with feedback connections as a general strategy for continual learning

    Gido M. van de Ven, Andreas S. Tolias

    cs.LGcs.AIcs.CVarXiv:1809.10635v22018
  45. Towards a Science of Human-AI Decision Making: A Survey of Empirical Studies

    Vivian Lai, Chacha Chen, Q. Vera Liao +2

    cs.AIcs.CLcs.CYarXiv:2112.11471v12021
  46. CXPlain: Causal Explanations for Model Interpretation under Uncertainty

    Patrick Schwab, Walter Karlen

    cs.LGstat.MLarXiv:1910.12336v12019
  47. giotto-tda: A Topological Data Analysis Toolkit for Machine Learning and Data Exploration

    Guillaume Tauzin, Umberto Lupo, Lewis Tunstall +6

    cs.LGmath.ATstat.MLarXiv:2004.02551v22020
  48. On learning to localize objects with minimal supervision

    Hyun Oh Song, Ross Girshick, Stefanie Jegelka +3

    cs.CVcs.LGarXiv:1403.1024v42014
  49. TopicRNN: A Recurrent Neural Network with Long-Range Semantic Dependency

    Adji B. Dieng, Chong Wang, Jianfeng Gao +1

    cs.CLcs.AIcs.LGarXiv:1611.01702v22016
  50. Test-Time Scaling for Scientific Equation Discovery

    Haowei Lin, Hubert Lim, Xiangyu Wang +2

    cs.CLcs.AIcs.LGarXiv:2608.28660v12026
  51. Lasso Screening Rules via Dual Polytope Projection

    Jie Wang, Peter Wonka, Jieping Ye

    cs.LGstat.MLarXiv:1211.3966v32012
  52. Measuring Sample Quality with Stein's Method

    Jackson Gorham, Lester Mackey

    stat.MLcs.LGmath.PRarXiv:1506.03039v62015
  53. MADLAD-400: A Multilingual And Document-Level Large Audited Dataset

    Sneha Kudugunta, Isaac Caswell, Biao Zhang +8

    cs.CLcs.LGarXiv:2309.04662v12023
  54. Timer: Generative Pre-trained Transformers Are Large Time Series Models

    Yong Liu, Haoran Zhang, Chenyu Li +3

    cs.LGstat.MLarXiv:2402.02368v32024
  55. Invariant Representations without Adversarial Training

    Daniel Moyer, Shuyang Gao, Rob Brekelmans +2

    cs.LGstat.MLarXiv:1805.09458v42018
  56. Perturbed Iterate Analysis for Asynchronous Stochastic Optimization

    Horia Mania, Xinghao Pan, Dimitris Papailiopoulos +3

    stat.MLcs.DCcs.DSarXiv:1507.06970v22015
  57. Wind Power Forecasting Considering Data Privacy Protection: A Federated Deep Reinforcement Learning Approach

    Yang Li, Ruinong Wang, Yuanzheng Li +2

    cs.LGeess.SYarXiv:2211.02674v12022
  58. Polymer Informatics: Current Status and Critical Next Steps

    Lihua Chen, Ghanshyam Pilania, Rohit Batra +4

    cond-mat.softcs.LGarXiv:2011.00508v12020
  59. Explainable Medical Imaging AI Needs Human-Centered Design: Guidelines and Evidence from a Systematic Review

    Haomin Chen, Catalina Gomez, Chien-Ming Huang +1

    cs.HCcs.CVcs.LGarXiv:2112.12596v42021
  60. Decentralized Collaborative Learning of Personalized Models over Networks

    Paul Vanhaesebrouck, Aurélien Bellet, Marc Tommasi

    cs.LGcs.AIcs.DCarXiv:1610.05202v22016