Machine Learning

Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

19,201 to 19,260 of 20,201

  1. Graph Convolutional Neural Networks for Web-Scale Recommender Systems

    Rex Ying, Ruining He, Kaifeng Chen +3

    cs.IRcs.LGstat.MLarXiv:1806.01973v12018
  2. Beyond Teacher Likelihood: Group-Calibrated On-Policy Distillation for Long-Context Reasoning

    Zhu Zhang, Jixun Wang, Xiaoang Xu +6

    cs.LGcs.AIcs.CLarXiv:2608.19181v12026
  3. Wide & Deep Learning for Recommender Systems

    Heng-Tze Cheng, Levent Koc, Jeremiah Harmsen +13

    cs.LGcs.IRstat.MLarXiv:1606.07792v12016
  4. DeepONet: Learning nonlinear operators for identifying differential equations based on the universal approximation theorem of operators

    Lu Lu, Pengzhan Jin, George Em Karniadakis

    cs.LGstat.MLarXiv:1910.03193v32019
  5. Bernstein-Vazirani Networks: Quantum Machine Learning by Interference

    Natacha Kuete Meli, Tolga Birdal, Prayag Tiwari +2

    quant-phcs.AIcs.CVarXiv:2608.19043v12026
  6. The Lottery Ticket Hypothesis: Finding Sparse, Trainable Neural Networks

    Jonathan Frankle, Michael Carbin

    cs.LGcs.AIcs.NEarXiv:1803.03635v52018
  7. Masked-attention Mask Transformer for Universal Image Segmentation

    Bowen Cheng, Ishan Misra, Alexander G. Schwing +2

    cs.CVcs.AIcs.LGarXiv:2112.01527v32021
  8. Leaf Values as Coordinates: Exact Contrastive Explanation for Gradient-Boosted Ensembles

    Emanuele Luzio

    cs.LGcs.AIcs.CYarXiv:2608.19127v12026
  9. Scalability in Perception for Autonomous Driving: Waymo Open Dataset

    Pei Sun, Henrik Kretzschmar, Xerxes Dotiwalla +22

    cs.CVcs.LGstat.MLarXiv:1912.04838v72019
  10. Discretizing Continuous Time Series for Imputation with Masked Diffusion Training

    Dongbin Kim, Seungyun Lee, Geonwoo Shin +1

    cs.LGcs.AIarXiv:2608.19119v12026
  11. Learning to Prompt for Vision-Language Models

    Kaiyang Zhou, Jingkang Yang, Chen Change Loy +1

    cs.CVcs.AIcs.LGarXiv:2109.01134v62021
  12. Open-MOPD: Diagnosing and Fixing Capability Imbalance in Multi-Teacher On-Policy Distillation

    Huan-ang Gao, Haohan Chi, Yong Yan +7

    cs.LGcs.AIcs.CLarXiv:2608.19098v12026
  13. A Reduction of Imitation Learning and Structured Prediction to No-Regret Online Learning

    Stephane Ross, Geoffrey J. Gordon, J. Andrew Bagnell

    cs.LGcs.AIstat.MLarXiv:1011.0686v32010
  14. Harness Continual Learning: Continual Adaptation Beyond Model Parameters

    Borui Kang, Jinrui Gu, Junhan Lv +3

    cs.LGcs.AIarXiv:2608.19013v12026
  15. Making the V in VQA Matter: Elevating the Role of Image Understanding in Visual Question Answering

    Yash Goyal, Tejas Khot, Douglas Summers-Stay +2

    cs.CVcs.AIcs.CLarXiv:1612.00837v32016
  16. High-Resolution Image Synthesis and Semantic Manipulation with Conditional GANs

    Ting-Chun Wang, Ming-Yu Liu, Jun-Yan Zhu +3

    cs.CVcs.GRcs.LGarXiv:1711.11585v22017
  17. Fourier Neural Operator for Parametric Partial Differential Equations

    Zongyi Li, Nikola Kovachki, Kamyar Azizzadenesheli +4

    cs.LGmath.NAarXiv:2010.08895v32020
  18. Tree of Thoughts: Deliberate Problem Solving with Large Language Models

    Shunyu Yao, Dian Yu, Jeffrey Zhao +4

    cs.CLcs.AIcs.LGarXiv:2305.10601v22023
  19. A Critical Synthesis of Uncertainty Quantification and Foundation Models for Semantic Segmentation

    Steven Landgraf, Joceline Hinz, Markus Ulrich

    cs.CVcs.AIcs.LGarXiv:2608.18709v12026
  20. Transformer-XL: Attentive Language Models Beyond a Fixed-Length Context

    Zihang Dai, Zhilin Yang, Yiming Yang +3

    cs.LGcs.CLstat.MLarXiv:1901.02860v32019
  21. DreamBooth: Fine Tuning Text-to-Image Diffusion Models for Subject-Driven Generation

    Nataniel Ruiz, Yuanzhen Li, Varun Jampani +3

    cs.CVcs.GRcs.LGarXiv:2208.12242v22022
  22. Graphical Design of Interpretable Architectures

    Pietro Barbiero

    cs.LGcs.AIcs.NEarXiv:2608.18936v12026
  23. Knowledge Distillation: A Survey

    Jianping Gou, Baosheng Yu, Stephen John Maybank +1

    cs.LGstat.MLarXiv:2006.05525v72020
  24. Group Normalization

    Yuxin Wu, Kaiming He

    cs.CVcs.LGarXiv:1803.08494v32018
  25. Dueling Network Architectures for Deep Reinforcement Learning

    Ziyu Wang, Tom Schaul, Matteo Hessel +3

    cs.LGarXiv:1511.06581v32015
  26. Density estimation using Real NVP

    Laurent Dinh, Jascha Sohl-Dickstein, Samy Bengio

    cs.LGcs.AIcs.NEarXiv:1605.08803v32016
  27. A Baseline for Detecting Misclassified and Out-of-Distribution Examples in Neural Networks

    Dan Hendrycks, Kevin Gimpel

    cs.NEcs.CVcs.LGarXiv:1610.02136v32016
  28. InfoGAN: Interpretable Representation Learning by Information Maximizing Generative Adversarial Nets

    Xi Chen, Yan Duan, Rein Houthooft +3

    cs.LGstat.MLarXiv:1606.03657v12016
  29. Forgetting, plasticity, and co-observation: a third facet of continual learning

    Timm Hess, Abhishek Jha, Gido M. van de Ven +1

    cs.LGcs.AIarXiv:2608.18803v12026
  30. The Mythos of Model Interpretability

    Zachary C. Lipton

    cs.LGcs.AIcs.CVarXiv:1606.03490v32016
  31. EEGNet: A Compact Convolutional Network for EEG-based Brain-Computer Interfaces

    Vernon J. Lawhern, Amelia J. Solon, Nicholas R. Waytowich +3

    cs.LGq-bio.NCstat.MLarXiv:1611.08024v42016
  32. Beyond Predictive Fairness: Quantifying Attribution Consistency Across Demographic Groups in Diabetic Retinopathy Screening

    Kerol Djoumessi, Philipp Berens

    cs.LGcs.AIarXiv:2608.18759v12026
  33. Benchmarking Neural Network Robustness to Common Corruptions and Perturbations

    Dan Hendrycks, Thomas Dietterich

    cs.LGcs.CVstat.MLarXiv:1903.12261v12019
  34. FlowNet: Learning Optical Flow with Convolutional Networks

    Philipp Fischer, Alexey Dosovitskiy, Eddy Ilg +6

    cs.CVcs.LGarXiv:1504.06852v22015
  35. End to End Learning for Self-Driving Cars

    Mariusz Bojarski, Davide Del Testa, Daniel Dworakowski +10

    cs.CVcs.LGcs.NEarXiv:1604.07316v12016
  36. PointPillars: Fast Encoders for Object Detection from Point Clouds

    Alex H. Lang, Sourabh Vora, Holger Caesar +3

    cs.LGcs.CVstat.MLarXiv:1812.05784v22018
  37. SimCSE: Simple Contrastive Learning of Sentence Embeddings

    Tianyu Gao, Xingcheng Yao, Danqi Chen

    cs.CLcs.LGarXiv:2104.08821v42021
  38. The Impact of CutMix on Reliability and Robustness in Semantic Segmentation

    Steven Landgraf, Markus Ulrich

    cs.CVcs.AIcs.LGarXiv:2608.18715v12026
  39. Evaluating and Explaining Prompt Sensitivity of LLMs Using Interactions

    Ruiyang Qin, Qingzhuo Wang, Tian Wang +2

    cs.LGcs.AIcs.CLarXiv:2608.18539v12026
  40. OPT: Open Pre-trained Transformer Language Models

    Susan Zhang, Stephen Roller, Naman Goyal +16

    cs.CLcs.LGarXiv:2205.01068v42022
  41. MorphoGP: A Nonparametric Framework for Predicting Equilibrium Beach Profiles Under Tidal Influence

    Xi Wu, Yanqing Wei, Hang Yin +3

    cs.LGcs.AIphysics.geo-pharXiv:2608.18558v12026
  42. HuBERT: Self-Supervised Speech Representation Learning by Masked Prediction of Hidden Units

    Wei-Ning Hsu, Benjamin Bolte, Yao-Hung Hubert Tsai +3

    cs.CLcs.AIcs.LGarXiv:2106.07447v12021
  43. GPT-4o System Card

    OpenAI, :, Aaron Hurst +417

    cs.CLcs.AIcs.CVarXiv:2410.21276v12024
  44. DART-SD: Diamond-topology Aware Retrieval and Tuning for Self-Distillation of Multi-Turn Tool-Calling Agents

    Hangrui Xu, Jiarui Wang, Yang Yang +5

    cs.CLcs.AIcs.LGarXiv:2608.18524v12026
  45. FitNets: Hints for Thin Deep Nets

    Adriana Romero, Nicolas Ballas, Samira Ebrahimi Kahou +3

    cs.LGcs.NEarXiv:1412.6550v42014
  46. FixMatch: Simplifying Semi-Supervised Learning with Consistency and Confidence

    Kihyuk Sohn, David Berthelot, Chun-Liang Li +6

    cs.LGcs.CVstat.MLarXiv:2001.07685v22020
  47. High-Dimensional Continuous Control Using Generalized Advantage Estimation

    John Schulman, Philipp Moritz, Sergey Levine +2

    cs.LGcs.ROeess.SYarXiv:1506.02438v62015
  48. Learning Important Features Through Propagating Activation Differences

    Avanti Shrikumar, Peyton Greenside, Anshul Kundaje

    cs.CVcs.LGcs.NEarXiv:1704.02685v22017
  49. Europe's Climate Ambition Under Scrutiny: Evidence from Deep Learning Emission Projections

    Jacopo Ghirri, Carlos Rodriguez-Pardo, Lara Aleluia Reis +1

    cs.LGcs.AIecon.GNarXiv:2608.18690v12026
  50. GLIDE: Towards Photorealistic Image Generation and Editing with Text-Guided Diffusion Models

    Alex Nichol, Prafulla Dhariwal, Aditya Ramesh +5

    cs.CVcs.GRcs.LGarXiv:2112.10741v32021
  51. Reflexion: Language Agents with Verbal Reinforcement Learning

    Noah Shinn, Federico Cassano, Edward Berman +3

    cs.AIcs.CLcs.LGarXiv:2303.11366v42023
  52. iCaRL: Incremental Classifier and Representation Learning

    Sylvestre-Alvise Rebuffi, Alexander Kolesnikov, Georg Sperl +1

    cs.CVcs.LGstat.MLarXiv:1611.07725v22016
  53. Variational Inference with Normalizing Flows

    Danilo Jimenez Rezende, Shakir Mohamed

    stat.MLcs.AIcs.LGarXiv:1505.05770v62015
  54. Spectral Normalization for Generative Adversarial Networks

    Takeru Miyato, Toshiki Kataoka, Masanori Koyama +1

    cs.LGcs.CVstat.MLarXiv:1802.05957v12018
  55. FedCoRe: Target-Adaptive Completion for Missing Modalities in Healthcare Federated Learning

    Holger R. Roth, Ziyue Xu, Peter Cnudde

    cs.CVcs.AIcs.LGarXiv:2608.18311v12026
  56. A Survey Of Methods For Explaining Black Box Models

    Riccardo Guidotti, Anna Monreale, Salvatore Ruggieri +3

    cs.CYcs.AIcs.LGarXiv:1802.01933v32018
  57. EEVEE: Towards Test-time Prompt Learning in the Real World for Self-Improving Agents

    Weixian Xu, Shilong Liu, Mengdi Wang

    cs.LGcs.AIarXiv:2606.11182v12026
  58. Task-Conditioned Least-Privilege Learning for Executable Terminal and MCP Agents

    Alexander Tu, Michael Tu

    cs.CRcs.AIcs.LGarXiv:2608.18351v12026
  59. ERASE: EaRly bAckpropagation SchEdule for Faster Training of Modern Recommendation Systems

    Ergan Shang, Flavio Sales Truzzi

    cs.LGcs.AIarXiv:2608.18469v12026
  60. Flash-GMM: A Memory-Efficient Kernel for Scalable Soft Clustering

    Gal Bloch, Ariel Gera, Matan Orbach +2

    cs.LGcs.DBcs.IRarXiv:2606.10896v12026