Machine Learning
Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
19,201 to 19,260 of 20,201
Graph Convolutional Neural Networks for Web-Scale Recommender Systems
Rex Ying, Ruining He, Kaifeng Chen +3
cs.IRcs.LGstat.MLarXiv:1806.01973v12018Beyond Teacher Likelihood: Group-Calibrated On-Policy Distillation for Long-Context Reasoning
Zhu Zhang, Jixun Wang, Xiaoang Xu +6
cs.LGcs.AIcs.CLarXiv:2608.19181v12026Wide & Deep Learning for Recommender Systems
Heng-Tze Cheng, Levent Koc, Jeremiah Harmsen +13
cs.LGcs.IRstat.MLarXiv:1606.07792v12016DeepONet: Learning nonlinear operators for identifying differential equations based on the universal approximation theorem of operators
Lu Lu, Pengzhan Jin, George Em Karniadakis
cs.LGstat.MLarXiv:1910.03193v32019Bernstein-Vazirani Networks: Quantum Machine Learning by Interference
Natacha Kuete Meli, Tolga Birdal, Prayag Tiwari +2
quant-phcs.AIcs.CVarXiv:2608.19043v12026The Lottery Ticket Hypothesis: Finding Sparse, Trainable Neural Networks
Jonathan Frankle, Michael Carbin
cs.LGcs.AIcs.NEarXiv:1803.03635v52018Masked-attention Mask Transformer for Universal Image Segmentation
Bowen Cheng, Ishan Misra, Alexander G. Schwing +2
cs.CVcs.AIcs.LGarXiv:2112.01527v32021Leaf Values as Coordinates: Exact Contrastive Explanation for Gradient-Boosted Ensembles
Emanuele Luzio
cs.LGcs.AIcs.CYarXiv:2608.19127v12026Scalability in Perception for Autonomous Driving: Waymo Open Dataset
Pei Sun, Henrik Kretzschmar, Xerxes Dotiwalla +22
cs.CVcs.LGstat.MLarXiv:1912.04838v72019Discretizing Continuous Time Series for Imputation with Masked Diffusion Training
Dongbin Kim, Seungyun Lee, Geonwoo Shin +1
cs.LGcs.AIarXiv:2608.19119v12026Learning to Prompt for Vision-Language Models
Kaiyang Zhou, Jingkang Yang, Chen Change Loy +1
cs.CVcs.AIcs.LGarXiv:2109.01134v62021Open-MOPD: Diagnosing and Fixing Capability Imbalance in Multi-Teacher On-Policy Distillation
Huan-ang Gao, Haohan Chi, Yong Yan +7
cs.LGcs.AIcs.CLarXiv:2608.19098v12026A Reduction of Imitation Learning and Structured Prediction to No-Regret Online Learning
Stephane Ross, Geoffrey J. Gordon, J. Andrew Bagnell
cs.LGcs.AIstat.MLarXiv:1011.0686v32010Harness Continual Learning: Continual Adaptation Beyond Model Parameters
Borui Kang, Jinrui Gu, Junhan Lv +3
cs.LGcs.AIarXiv:2608.19013v12026Making the V in VQA Matter: Elevating the Role of Image Understanding in Visual Question Answering
Yash Goyal, Tejas Khot, Douglas Summers-Stay +2
cs.CVcs.AIcs.CLarXiv:1612.00837v32016High-Resolution Image Synthesis and Semantic Manipulation with Conditional GANs
Ting-Chun Wang, Ming-Yu Liu, Jun-Yan Zhu +3
cs.CVcs.GRcs.LGarXiv:1711.11585v22017Fourier Neural Operator for Parametric Partial Differential Equations
Zongyi Li, Nikola Kovachki, Kamyar Azizzadenesheli +4
cs.LGmath.NAarXiv:2010.08895v32020Tree of Thoughts: Deliberate Problem Solving with Large Language Models
Shunyu Yao, Dian Yu, Jeffrey Zhao +4
cs.CLcs.AIcs.LGarXiv:2305.10601v22023A Critical Synthesis of Uncertainty Quantification and Foundation Models for Semantic Segmentation
Steven Landgraf, Joceline Hinz, Markus Ulrich
cs.CVcs.AIcs.LGarXiv:2608.18709v12026Transformer-XL: Attentive Language Models Beyond a Fixed-Length Context
Zihang Dai, Zhilin Yang, Yiming Yang +3
cs.LGcs.CLstat.MLarXiv:1901.02860v32019DreamBooth: Fine Tuning Text-to-Image Diffusion Models for Subject-Driven Generation
Nataniel Ruiz, Yuanzhen Li, Varun Jampani +3
cs.CVcs.GRcs.LGarXiv:2208.12242v22022Graphical Design of Interpretable Architectures
Pietro Barbiero
cs.LGcs.AIcs.NEarXiv:2608.18936v12026Knowledge Distillation: A Survey
Jianping Gou, Baosheng Yu, Stephen John Maybank +1
cs.LGstat.MLarXiv:2006.05525v72020Group Normalization
Yuxin Wu, Kaiming He
cs.CVcs.LGarXiv:1803.08494v32018Dueling Network Architectures for Deep Reinforcement Learning
Ziyu Wang, Tom Schaul, Matteo Hessel +3
cs.LGarXiv:1511.06581v32015Density estimation using Real NVP
Laurent Dinh, Jascha Sohl-Dickstein, Samy Bengio
cs.LGcs.AIcs.NEarXiv:1605.08803v32016A Baseline for Detecting Misclassified and Out-of-Distribution Examples in Neural Networks
Dan Hendrycks, Kevin Gimpel
cs.NEcs.CVcs.LGarXiv:1610.02136v32016InfoGAN: Interpretable Representation Learning by Information Maximizing Generative Adversarial Nets
Xi Chen, Yan Duan, Rein Houthooft +3
cs.LGstat.MLarXiv:1606.03657v12016Forgetting, plasticity, and co-observation: a third facet of continual learning
Timm Hess, Abhishek Jha, Gido M. van de Ven +1
cs.LGcs.AIarXiv:2608.18803v12026The Mythos of Model Interpretability
Zachary C. Lipton
cs.LGcs.AIcs.CVarXiv:1606.03490v32016EEGNet: A Compact Convolutional Network for EEG-based Brain-Computer Interfaces
Vernon J. Lawhern, Amelia J. Solon, Nicholas R. Waytowich +3
cs.LGq-bio.NCstat.MLarXiv:1611.08024v42016Beyond Predictive Fairness: Quantifying Attribution Consistency Across Demographic Groups in Diabetic Retinopathy Screening
Kerol Djoumessi, Philipp Berens
cs.LGcs.AIarXiv:2608.18759v12026Benchmarking Neural Network Robustness to Common Corruptions and Perturbations
Dan Hendrycks, Thomas Dietterich
cs.LGcs.CVstat.MLarXiv:1903.12261v12019FlowNet: Learning Optical Flow with Convolutional Networks
Philipp Fischer, Alexey Dosovitskiy, Eddy Ilg +6
cs.CVcs.LGarXiv:1504.06852v22015End to End Learning for Self-Driving Cars
Mariusz Bojarski, Davide Del Testa, Daniel Dworakowski +10
cs.CVcs.LGcs.NEarXiv:1604.07316v12016PointPillars: Fast Encoders for Object Detection from Point Clouds
Alex H. Lang, Sourabh Vora, Holger Caesar +3
cs.LGcs.CVstat.MLarXiv:1812.05784v22018SimCSE: Simple Contrastive Learning of Sentence Embeddings
Tianyu Gao, Xingcheng Yao, Danqi Chen
cs.CLcs.LGarXiv:2104.08821v42021The Impact of CutMix on Reliability and Robustness in Semantic Segmentation
Steven Landgraf, Markus Ulrich
cs.CVcs.AIcs.LGarXiv:2608.18715v12026Evaluating and Explaining Prompt Sensitivity of LLMs Using Interactions
Ruiyang Qin, Qingzhuo Wang, Tian Wang +2
cs.LGcs.AIcs.CLarXiv:2608.18539v12026OPT: Open Pre-trained Transformer Language Models
Susan Zhang, Stephen Roller, Naman Goyal +16
cs.CLcs.LGarXiv:2205.01068v42022MorphoGP: A Nonparametric Framework for Predicting Equilibrium Beach Profiles Under Tidal Influence
Xi Wu, Yanqing Wei, Hang Yin +3
cs.LGcs.AIphysics.geo-pharXiv:2608.18558v12026HuBERT: Self-Supervised Speech Representation Learning by Masked Prediction of Hidden Units
Wei-Ning Hsu, Benjamin Bolte, Yao-Hung Hubert Tsai +3
cs.CLcs.AIcs.LGarXiv:2106.07447v12021GPT-4o System Card
OpenAI, :, Aaron Hurst +417
cs.CLcs.AIcs.CVarXiv:2410.21276v12024DART-SD: Diamond-topology Aware Retrieval and Tuning for Self-Distillation of Multi-Turn Tool-Calling Agents
Hangrui Xu, Jiarui Wang, Yang Yang +5
cs.CLcs.AIcs.LGarXiv:2608.18524v12026FitNets: Hints for Thin Deep Nets
Adriana Romero, Nicolas Ballas, Samira Ebrahimi Kahou +3
cs.LGcs.NEarXiv:1412.6550v42014FixMatch: Simplifying Semi-Supervised Learning with Consistency and Confidence
Kihyuk Sohn, David Berthelot, Chun-Liang Li +6
cs.LGcs.CVstat.MLarXiv:2001.07685v22020High-Dimensional Continuous Control Using Generalized Advantage Estimation
John Schulman, Philipp Moritz, Sergey Levine +2
cs.LGcs.ROeess.SYarXiv:1506.02438v62015Learning Important Features Through Propagating Activation Differences
Avanti Shrikumar, Peyton Greenside, Anshul Kundaje
cs.CVcs.LGcs.NEarXiv:1704.02685v22017Europe's Climate Ambition Under Scrutiny: Evidence from Deep Learning Emission Projections
Jacopo Ghirri, Carlos Rodriguez-Pardo, Lara Aleluia Reis +1
cs.LGcs.AIecon.GNarXiv:2608.18690v12026GLIDE: Towards Photorealistic Image Generation and Editing with Text-Guided Diffusion Models
Alex Nichol, Prafulla Dhariwal, Aditya Ramesh +5
cs.CVcs.GRcs.LGarXiv:2112.10741v32021Reflexion: Language Agents with Verbal Reinforcement Learning
Noah Shinn, Federico Cassano, Edward Berman +3
cs.AIcs.CLcs.LGarXiv:2303.11366v42023iCaRL: Incremental Classifier and Representation Learning
Sylvestre-Alvise Rebuffi, Alexander Kolesnikov, Georg Sperl +1
cs.CVcs.LGstat.MLarXiv:1611.07725v22016Variational Inference with Normalizing Flows
Danilo Jimenez Rezende, Shakir Mohamed
stat.MLcs.AIcs.LGarXiv:1505.05770v62015Spectral Normalization for Generative Adversarial Networks
Takeru Miyato, Toshiki Kataoka, Masanori Koyama +1
cs.LGcs.CVstat.MLarXiv:1802.05957v12018FedCoRe: Target-Adaptive Completion for Missing Modalities in Healthcare Federated Learning
Holger R. Roth, Ziyue Xu, Peter Cnudde
cs.CVcs.AIcs.LGarXiv:2608.18311v12026A Survey Of Methods For Explaining Black Box Models
Riccardo Guidotti, Anna Monreale, Salvatore Ruggieri +3
cs.CYcs.AIcs.LGarXiv:1802.01933v32018EEVEE: Towards Test-time Prompt Learning in the Real World for Self-Improving Agents
Weixian Xu, Shilong Liu, Mengdi Wang
cs.LGcs.AIarXiv:2606.11182v12026Task-Conditioned Least-Privilege Learning for Executable Terminal and MCP Agents
Alexander Tu, Michael Tu
cs.CRcs.AIcs.LGarXiv:2608.18351v12026ERASE: EaRly bAckpropagation SchEdule for Faster Training of Modern Recommendation Systems
Ergan Shang, Flavio Sales Truzzi
cs.LGcs.AIarXiv:2608.18469v12026Flash-GMM: A Memory-Efficient Kernel for Scalable Soft Clustering
Gal Bloch, Ariel Gera, Matan Orbach +2
cs.LGcs.DBcs.IRarXiv:2606.10896v12026