Machine Learning
Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
18,901 to 18,960 of 20,199
Graph Contrastive Learning with Augmentations
Yuning You, Tianlong Chen, Yongduo Sui +3
cs.LGcs.AIarXiv:2010.13902v32020Semi-Supervised Learning with Deep Generative Models
Diederik P. Kingma, Danilo J. Rezende, Shakir Mohamed +1
cs.LGstat.MLarXiv:1406.5298v22014RotatE: Knowledge Graph Embedding by Relational Rotation in Complex Space
Zhiqing Sun, Zhi-Hong Deng, Jian-Yun Nie +1
cs.LGcs.CLstat.MLarXiv:1902.10197v12019Prompt-to-Prompt Image Editing with Cross Attention Control
Amir Hertz, Ron Mokady, Jay Tenenbaum +3
cs.CVcs.CLcs.GRarXiv:2208.01626v12022Temporal Ensembling for Semi-Supervised Learning
Samuli Laine, Timo Aila
cs.NEcs.LGarXiv:1610.02242v32016A Survey on Deep Transfer Learning
Chuanqi Tan, Fuchun Sun, Tao Kong +3
cs.LGstat.MLarXiv:1808.01974v12018UNETR: Transformers for 3D Medical Image Segmentation
Ali Hatamizadeh, Yucheng Tang, Vishwesh Nath +5
eess.IVcs.CVcs.LGarXiv:2103.10504v32021CLEVR: A Diagnostic Dataset for Compositional Language and Elementary Visual Reasoning
Justin Johnson, Bharath Hariharan, Laurens van der Maaten +3
cs.CVcs.CLcs.LGarXiv:1612.06890v12016MatConvNet - Convolutional Neural Networks for MATLAB
Andrea Vedaldi, Karel Lenc
cs.CVcs.LGcs.MSarXiv:1412.4564v32014LXMERT: Learning Cross-Modality Encoder Representations from Transformers
Hao Tan, Mohit Bansal
cs.CLcs.CVcs.LGarXiv:1908.07490v32019T-GCN: A Temporal Graph ConvolutionalNetwork for Traffic Prediction
Ling Zhao, Yujiao Song, Chao Zhang +5
cs.LGstat.MLarXiv:1811.05320v32018Reformer: The Efficient Transformer
Nikita Kitaev, Łukasz Kaiser, Anselm Levskaya
cs.LGcs.CLstat.MLarXiv:2001.04451v22020CNN Architectures for Large-Scale Audio Classification
Shawn Hershey, Sourish Chaudhuri, Daniel P. W. Ellis +10
cs.SDcs.LGstat.MLarXiv:1609.09430v22016Efficient Neural Architecture Search via Parameter Sharing
Hieu Pham, Melody Y. Guan, Barret Zoph +2
cs.LGcs.CLcs.CVarXiv:1802.03268v22018Transformers are RNNs: Fast Autoregressive Transformers with Linear Attention
Angelos Katharopoulos, Apoorv Vyas, Nikolaos Pappas +1
cs.LGstat.MLarXiv:2006.16236v32020OpenVLA: An Open-Source Vision-Language-Action Model
Moo Jin Kim, Karl Pertsch, Siddharth Karamcheti +15
cs.ROcs.LGarXiv:2406.09246v32024Progressive Neural Networks
Andrei A. Rusu, Neil C. Rabinowitz, Guillaume Desjardins +5
cs.LGarXiv:1606.04671v42016SPK: Eliciting Structured Prior Knowledge for Interpretable Out-of-Distribution Detection in Real-Time Object Detection
Changshun Wu, Weicheng He, Xiaowei Huang +1
cs.CVcs.LGarXiv:2608.19080v12026Cyclical Learning Rates for Training Neural Networks
Leslie N. Smith
cs.CVcs.LGcs.NEarXiv:1506.01186v62015PolicyGuide: From Guarding One Action to Guiding the Whole Workflow for Policy-Compliant LLM Agents
Seongjae Kang, Taehyung Yu, Sung Ju Hwang
cs.AIcs.CLcs.LGarXiv:2608.19861v12026Session-based Recommendations with Recurrent Neural Networks
Balázs Hidasi, Alexandros Karatzoglou, Linas Baltrunas +1
cs.LGcs.IRcs.NEarXiv:1511.06939v42015PaLM-E: An Embodied Multimodal Language Model
Danny Driess, Fei Xia, Mehdi S. M. Sajjadi +19
cs.LGcs.AIcs.ROarXiv:2303.03378v12023Big Bird: Transformers for Longer Sequences
Manzil Zaheer, Guru Guruganesh, Avinava Dubey +8
cs.LGcs.CLstat.MLarXiv:2007.14062v22020A Large Dataset to Train Convolutional Networks for Disparity, Optical Flow, and Scene Flow Estimation
Nikolaus Mayer, Eddy Ilg, Philip Häusser +4
cs.CVcs.LGstat.MLarXiv:1512.02134v12015A Survey on Multi-Task Learning
Yu Zhang, Qiang Yang
cs.LGcs.AIarXiv:1707.08114v32017Locating and Editing Factual Associations in GPT
Kevin Meng, David Bau, Alex Andonian +1
cs.CLcs.LGarXiv:2202.05262v52022Self-Evolving Visual Questioner
Yijun Liang, Hengguang Zhou, Ming Li +3
cs.CVcs.LGarXiv:2606.13929v12026Privacy-Preserving Dataset Curation for Kuala Lumpur Urban Traffic: Grounded Vision-Language Detection with Spatial Vehicle-Context Filtering
Mohammed Abdul Al Arafat Tanzin, Rudzidatul Akmam Dziyauddin
cs.CVcs.AIcs.LGarXiv:2608.14724v12026Do As I Can, Not As I Say: Grounding Language in Robotic Affordances
Michael Ahn, Anthony Brohan, Noah Brown +42
cs.ROcs.CLcs.LGarXiv:2204.01691v22022What Makes Software Issue Resolution Tasks Difficult for Agents?
Ebtesam Al-Haque, Brittany Johnson
cs.SEcs.AIcs.CLarXiv:2608.18280v12026A Pre-Specified Construction-Confirmation Test of Operation-Level Causal Transfer Across Finite Isomorphic Symbolic Domains
Xinyi Shan
cs.LGarXiv:2608.15809v12026Towards a Physics Foundation Model
Florian Wiesner, Zoë J. Gray, Matthias Wessling +1
cs.LGcs.AIstat.MLarXiv:2509.13805v42025LUNG-KGMM: Knowledge-Guided Multimodal Learning for Lung Cancer Incidence Prediction
Chunlei Yang, Shuyan Li, Zhong Cao
cs.LGcs.CVarXiv:2608.14657v12026Summaries:한국어When Single-Dataset Conclusions Fail: A 45-Task Study of Threshold Tuning and Resampling for Imbalanced Classification
Diyorbek Musaev
cs.AIcs.LGarXiv:2608.16147v12026Summaries:한국어Mind the Gap: Understanding the Modality Gap in Multi-modal Contrastive Representation Learning
Weixin Liang, Yuhui Zhang, Yongchan Kwon +2
cs.CLcs.AIcs.CVarXiv:2203.02053v22022Summaries:한국어Proactive Road Safety Intervention in Australia: Predicting Risky Driving Hotspots from Connected Vehicle Data
Adriana-Simona Mihăiţă, Clarence Cheung, Artur Grigorev +2
cs.LGcs.CYarXiv:2608.16913v12026Deep learning with convolutional neural networks for EEG decoding and visualization
Robin Tibor Schirrmeister, Jost Tobias Springenberg, Lukas Dominique Josef Fiederer +6
cs.LGcs.NEarXiv:1703.05051v52017MirrorWorld: Taming Video Diffusion Models for Mirror Reflection Generation
Youjun Zhao, Alex Warren, Gary K. L. Tam +1
cs.CVcs.LGarXiv:2608.07463v12026The Physics of Multi-Turn Long-Horizon Planning: From Pre-training to Post-training via Single- and Multi-Teacher On-Policy Agentic Distillation
Tianyi Men, Zhuoran Jin, Kang Liu +1
cs.CLcs.AIcs.LGarXiv:2607.24720v12026DiFA: Inference-Time Forward-Process Alignment for Diffusion Models
Shigui Li, Delu Zeng
cs.LGarXiv:2607.17972v12026Masked Diffusion Language Models are Strong and Steerable Text-Based World Models for Agentic RL
Darshan Deshpande
cs.AIcs.LGarXiv:2607.16204v12026Do Thinking Tokens Help with Safety?
Narutatsu Ri, Abhishek Panigrahi, Sanjeev Arora
cs.LGcs.AIcs.CLarXiv:2606.25013v12026Summaries:한국어Seeing Before Reasoning: Decoupling Perception and Reasoning for Shortcut-Resilient Multimodal On-Policy Self-Distillation
Sihan Wang, Xiyao Liu, Lianqing Liu +1
cs.LGcs.CVarXiv:2606.19120v22026MaxProof: Scaling Mathematical Proof with Generative-Verifier RL and Population-Level Test-Time Scaling
Jiacheng Chen, Xinyu Zhang, Shunkai Zhang +20
cs.LGcs.AIcs.CLarXiv:2606.13473v12026Agents' Last Exam
Yiyou Sun, Xinyang Han, Weichen Zhang +307
cs.AIcs.CLcs.LGarXiv:2606.05405v22026BenchEvolver: Frontier Task Synthesis via Solution-Centric Evolution
Yangzhen Wu, Aaron J. Li, Wenjie Ma +10
cs.SEcs.AIcs.CLarXiv:2606.01286v12026Multi-Stream LLMs: Unblocking Language Models with Parallel Streams of Thoughts, Inputs and Outputs
Guinan Su, Yanwu Yang, Xueyan Li +1
cs.LGcs.CLarXiv:2605.12460v12026Multimodal Deep Learning
Cem Akkus, Luyang Chu, Vladana Djakovic +14
cs.CLcs.LGstat.MLarXiv:2301.04856v12023Summaries:한국어DarwinX: Evolving Agent Harnesses Through Natural Selection
Yifan Zhang, Yutong Dai, Juntao Tan +9
cs.NEcs.AIcs.LGarXiv:2608.07545v12026CoM$^3$eT: A foundation model for medical image analysis through federated, multidimensional context integration
J. Raphael Schäfer, Kai Geissler, Till Nicke +27
cs.CVcs.LGarXiv:2608.16268v12026Domain-Specific Text Embedding Models for Entity Resolution
Khajesh Sapram, Srivardhani Raju, Kishore Konda
cs.IRcs.AIcs.LGarXiv:2608.16161v12026A Scalable Pipeline for LLM-Teacher Distillation Labeling: Work-Stealing Job Scheduling and Memory-Aware GPU Concurrency
Ravi Satya Durga Prasad Yenugula
cs.DCcs.AIcs.CLarXiv:2608.15975v12026Efficient Knowledge Distillation for LLMs: Offline Top-K Logits and a Fused Chunked KL Loss
Bakbergen Ryskulov, Iker García-Ferrero, David Montero +5
cs.CLcs.AIcs.LGarXiv:2608.03796v12026Enhanced Privacy and Communication Efficiency in Non-IID Federated Learning with Adaptive Quantization and Differential Privacy
Emre Ardıç, Yakup Genç
cs.CVcs.LGarXiv:2604.23426v12026On the Opportunities and Risks of Foundation Models
Rishi Bommasani, Drew A. Hudson, Ehsan Adeli +111
cs.LGcs.AIcs.CYarXiv:2108.07258v32021Deep Learning with Differential Privacy
Martín Abadi, Andy Chu, Ian Goodfellow +4
stat.MLcs.CRcs.LGarXiv:1607.00133v22016Look Before You Lift: Visual and Quantitative Diagnostics for Topological Deep Learning
Mathilde Papillon, Guillermo Bernárdez, Álvaro Ballón Barreiro +4
cs.LGarXiv:2608.15388v12026Robust Risk Under Evolving Uncertainty: A Wasserstein Counterpart of the Entropic Value-at-Risk
Deep Kumar Ganguly, Jan Křetínský
cs.AIcs.LGstat.MLarXiv:2608.19073v12026Training Chemical Plausibility-Aware Large Language Models for Single-Step Retrosynthesis
Bogdan Zagribelnyy, Ivan Ilin, Nikita Bondarev +5
cs.LGcs.AIcs.CEarXiv:2608.18940v12026Understanding Multilingual Medical ASR Adaptation Through Layer-Wise Analysis
Souranil Kahali, Rituparna Bose, Abner Hernandez +4
cs.CLcs.AIcs.LGarXiv:2608.18825v12026