Machine Learning

Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

18,901 to 18,960 of 20,199

  1. Graph Contrastive Learning with Augmentations

    Yuning You, Tianlong Chen, Yongduo Sui +3

    cs.LGcs.AIarXiv:2010.13902v32020
  2. Semi-Supervised Learning with Deep Generative Models

    Diederik P. Kingma, Danilo J. Rezende, Shakir Mohamed +1

    cs.LGstat.MLarXiv:1406.5298v22014
  3. RotatE: Knowledge Graph Embedding by Relational Rotation in Complex Space

    Zhiqing Sun, Zhi-Hong Deng, Jian-Yun Nie +1

    cs.LGcs.CLstat.MLarXiv:1902.10197v12019
  4. Prompt-to-Prompt Image Editing with Cross Attention Control

    Amir Hertz, Ron Mokady, Jay Tenenbaum +3

    cs.CVcs.CLcs.GRarXiv:2208.01626v12022
  5. Temporal Ensembling for Semi-Supervised Learning

    Samuli Laine, Timo Aila

    cs.NEcs.LGarXiv:1610.02242v32016
  6. A Survey on Deep Transfer Learning

    Chuanqi Tan, Fuchun Sun, Tao Kong +3

    cs.LGstat.MLarXiv:1808.01974v12018
  7. UNETR: Transformers for 3D Medical Image Segmentation

    Ali Hatamizadeh, Yucheng Tang, Vishwesh Nath +5

    eess.IVcs.CVcs.LGarXiv:2103.10504v32021
  8. CLEVR: A Diagnostic Dataset for Compositional Language and Elementary Visual Reasoning

    Justin Johnson, Bharath Hariharan, Laurens van der Maaten +3

    cs.CVcs.CLcs.LGarXiv:1612.06890v12016
  9. MatConvNet - Convolutional Neural Networks for MATLAB

    Andrea Vedaldi, Karel Lenc

    cs.CVcs.LGcs.MSarXiv:1412.4564v32014
  10. LXMERT: Learning Cross-Modality Encoder Representations from Transformers

    Hao Tan, Mohit Bansal

    cs.CLcs.CVcs.LGarXiv:1908.07490v32019
  11. T-GCN: A Temporal Graph ConvolutionalNetwork for Traffic Prediction

    Ling Zhao, Yujiao Song, Chao Zhang +5

    cs.LGstat.MLarXiv:1811.05320v32018
  12. Reformer: The Efficient Transformer

    Nikita Kitaev, Łukasz Kaiser, Anselm Levskaya

    cs.LGcs.CLstat.MLarXiv:2001.04451v22020
  13. CNN Architectures for Large-Scale Audio Classification

    Shawn Hershey, Sourish Chaudhuri, Daniel P. W. Ellis +10

    cs.SDcs.LGstat.MLarXiv:1609.09430v22016
  14. Efficient Neural Architecture Search via Parameter Sharing

    Hieu Pham, Melody Y. Guan, Barret Zoph +2

    cs.LGcs.CLcs.CVarXiv:1802.03268v22018
  15. Transformers are RNNs: Fast Autoregressive Transformers with Linear Attention

    Angelos Katharopoulos, Apoorv Vyas, Nikolaos Pappas +1

    cs.LGstat.MLarXiv:2006.16236v32020
  16. OpenVLA: An Open-Source Vision-Language-Action Model

    Moo Jin Kim, Karl Pertsch, Siddharth Karamcheti +15

    cs.ROcs.LGarXiv:2406.09246v32024
  17. Progressive Neural Networks

    Andrei A. Rusu, Neil C. Rabinowitz, Guillaume Desjardins +5

    cs.LGarXiv:1606.04671v42016
  18. SPK: Eliciting Structured Prior Knowledge for Interpretable Out-of-Distribution Detection in Real-Time Object Detection

    Changshun Wu, Weicheng He, Xiaowei Huang +1

    cs.CVcs.LGarXiv:2608.19080v12026
  19. Cyclical Learning Rates for Training Neural Networks

    Leslie N. Smith

    cs.CVcs.LGcs.NEarXiv:1506.01186v62015
  20. PolicyGuide: From Guarding One Action to Guiding the Whole Workflow for Policy-Compliant LLM Agents

    Seongjae Kang, Taehyung Yu, Sung Ju Hwang

    cs.AIcs.CLcs.LGarXiv:2608.19861v12026
  21. Session-based Recommendations with Recurrent Neural Networks

    Balázs Hidasi, Alexandros Karatzoglou, Linas Baltrunas +1

    cs.LGcs.IRcs.NEarXiv:1511.06939v42015
  22. PaLM-E: An Embodied Multimodal Language Model

    Danny Driess, Fei Xia, Mehdi S. M. Sajjadi +19

    cs.LGcs.AIcs.ROarXiv:2303.03378v12023
  23. Big Bird: Transformers for Longer Sequences

    Manzil Zaheer, Guru Guruganesh, Avinava Dubey +8

    cs.LGcs.CLstat.MLarXiv:2007.14062v22020
  24. A Large Dataset to Train Convolutional Networks for Disparity, Optical Flow, and Scene Flow Estimation

    Nikolaus Mayer, Eddy Ilg, Philip Häusser +4

    cs.CVcs.LGstat.MLarXiv:1512.02134v12015
  25. A Survey on Multi-Task Learning

    Yu Zhang, Qiang Yang

    cs.LGcs.AIarXiv:1707.08114v32017
  26. Locating and Editing Factual Associations in GPT

    Kevin Meng, David Bau, Alex Andonian +1

    cs.CLcs.LGarXiv:2202.05262v52022
  27. Self-Evolving Visual Questioner

    Yijun Liang, Hengguang Zhou, Ming Li +3

    cs.CVcs.LGarXiv:2606.13929v12026
  28. Privacy-Preserving Dataset Curation for Kuala Lumpur Urban Traffic: Grounded Vision-Language Detection with Spatial Vehicle-Context Filtering

    Mohammed Abdul Al Arafat Tanzin, Rudzidatul Akmam Dziyauddin

    cs.CVcs.AIcs.LGarXiv:2608.14724v12026
  29. Do As I Can, Not As I Say: Grounding Language in Robotic Affordances

    Michael Ahn, Anthony Brohan, Noah Brown +42

    cs.ROcs.CLcs.LGarXiv:2204.01691v22022
  30. What Makes Software Issue Resolution Tasks Difficult for Agents?

    Ebtesam Al-Haque, Brittany Johnson

    cs.SEcs.AIcs.CLarXiv:2608.18280v12026
  31. A Pre-Specified Construction-Confirmation Test of Operation-Level Causal Transfer Across Finite Isomorphic Symbolic Domains

    Xinyi Shan

    cs.LGarXiv:2608.15809v12026
  32. Towards a Physics Foundation Model

    Florian Wiesner, Zoë J. Gray, Matthias Wessling +1

    cs.LGcs.AIstat.MLarXiv:2509.13805v42025
  33. LUNG-KGMM: Knowledge-Guided Multimodal Learning for Lung Cancer Incidence Prediction

    Chunlei Yang, Shuyan Li, Zhong Cao

    cs.LGcs.CVarXiv:2608.14657v12026
    Summaries:한국어
  34. When Single-Dataset Conclusions Fail: A 45-Task Study of Threshold Tuning and Resampling for Imbalanced Classification

    Diyorbek Musaev

    cs.AIcs.LGarXiv:2608.16147v12026
    Summaries:한국어
  35. Mind the Gap: Understanding the Modality Gap in Multi-modal Contrastive Representation Learning

    Weixin Liang, Yuhui Zhang, Yongchan Kwon +2

    cs.CLcs.AIcs.CVarXiv:2203.02053v22022
    Summaries:한국어
  36. Proactive Road Safety Intervention in Australia: Predicting Risky Driving Hotspots from Connected Vehicle Data

    Adriana-Simona Mihăiţă, Clarence Cheung, Artur Grigorev +2

    cs.LGcs.CYarXiv:2608.16913v12026
  37. Deep learning with convolutional neural networks for EEG decoding and visualization

    Robin Tibor Schirrmeister, Jost Tobias Springenberg, Lukas Dominique Josef Fiederer +6

    cs.LGcs.NEarXiv:1703.05051v52017
  38. MirrorWorld: Taming Video Diffusion Models for Mirror Reflection Generation

    Youjun Zhao, Alex Warren, Gary K. L. Tam +1

    cs.CVcs.LGarXiv:2608.07463v12026
  39. The Physics of Multi-Turn Long-Horizon Planning: From Pre-training to Post-training via Single- and Multi-Teacher On-Policy Agentic Distillation

    Tianyi Men, Zhuoran Jin, Kang Liu +1

    cs.CLcs.AIcs.LGarXiv:2607.24720v12026
  40. DiFA: Inference-Time Forward-Process Alignment for Diffusion Models

    Shigui Li, Delu Zeng

    cs.LGarXiv:2607.17972v12026
  41. Masked Diffusion Language Models are Strong and Steerable Text-Based World Models for Agentic RL

    Darshan Deshpande

    cs.AIcs.LGarXiv:2607.16204v12026
  42. Do Thinking Tokens Help with Safety?

    Narutatsu Ri, Abhishek Panigrahi, Sanjeev Arora

    cs.LGcs.AIcs.CLarXiv:2606.25013v12026
    Summaries:한국어
  43. Seeing Before Reasoning: Decoupling Perception and Reasoning for Shortcut-Resilient Multimodal On-Policy Self-Distillation

    Sihan Wang, Xiyao Liu, Lianqing Liu +1

    cs.LGcs.CVarXiv:2606.19120v22026
  44. MaxProof: Scaling Mathematical Proof with Generative-Verifier RL and Population-Level Test-Time Scaling

    Jiacheng Chen, Xinyu Zhang, Shunkai Zhang +20

    cs.LGcs.AIcs.CLarXiv:2606.13473v12026
  45. Agents' Last Exam

    Yiyou Sun, Xinyang Han, Weichen Zhang +307

    cs.AIcs.CLcs.LGarXiv:2606.05405v22026
  46. BenchEvolver: Frontier Task Synthesis via Solution-Centric Evolution

    Yangzhen Wu, Aaron J. Li, Wenjie Ma +10

    cs.SEcs.AIcs.CLarXiv:2606.01286v12026
  47. Multi-Stream LLMs: Unblocking Language Models with Parallel Streams of Thoughts, Inputs and Outputs

    Guinan Su, Yanwu Yang, Xueyan Li +1

    cs.LGcs.CLarXiv:2605.12460v12026
  48. Multimodal Deep Learning

    Cem Akkus, Luyang Chu, Vladana Djakovic +14

    cs.CLcs.LGstat.MLarXiv:2301.04856v12023
    Summaries:한국어
  49. DarwinX: Evolving Agent Harnesses Through Natural Selection

    Yifan Zhang, Yutong Dai, Juntao Tan +9

    cs.NEcs.AIcs.LGarXiv:2608.07545v12026
  50. CoM$^3$eT: A foundation model for medical image analysis through federated, multidimensional context integration

    J. Raphael Schäfer, Kai Geissler, Till Nicke +27

    cs.CVcs.LGarXiv:2608.16268v12026
  51. Domain-Specific Text Embedding Models for Entity Resolution

    Khajesh Sapram, Srivardhani Raju, Kishore Konda

    cs.IRcs.AIcs.LGarXiv:2608.16161v12026
  52. A Scalable Pipeline for LLM-Teacher Distillation Labeling: Work-Stealing Job Scheduling and Memory-Aware GPU Concurrency

    Ravi Satya Durga Prasad Yenugula

    cs.DCcs.AIcs.CLarXiv:2608.15975v12026
  53. Efficient Knowledge Distillation for LLMs: Offline Top-K Logits and a Fused Chunked KL Loss

    Bakbergen Ryskulov, Iker García-Ferrero, David Montero +5

    cs.CLcs.AIcs.LGarXiv:2608.03796v12026
  54. Enhanced Privacy and Communication Efficiency in Non-IID Federated Learning with Adaptive Quantization and Differential Privacy

    Emre Ardıç, Yakup Genç

    cs.CVcs.LGarXiv:2604.23426v12026
  55. On the Opportunities and Risks of Foundation Models

    Rishi Bommasani, Drew A. Hudson, Ehsan Adeli +111

    cs.LGcs.AIcs.CYarXiv:2108.07258v32021
  56. Deep Learning with Differential Privacy

    Martín Abadi, Andy Chu, Ian Goodfellow +4

    stat.MLcs.CRcs.LGarXiv:1607.00133v22016
  57. Look Before You Lift: Visual and Quantitative Diagnostics for Topological Deep Learning

    Mathilde Papillon, Guillermo Bernárdez, Álvaro Ballón Barreiro +4

    cs.LGarXiv:2608.15388v12026
  58. Robust Risk Under Evolving Uncertainty: A Wasserstein Counterpart of the Entropic Value-at-Risk

    Deep Kumar Ganguly, Jan Křetínský

    cs.AIcs.LGstat.MLarXiv:2608.19073v12026
  59. Training Chemical Plausibility-Aware Large Language Models for Single-Step Retrosynthesis

    Bogdan Zagribelnyy, Ivan Ilin, Nikita Bondarev +5

    cs.LGcs.AIcs.CEarXiv:2608.18940v12026
  60. Understanding Multilingual Medical ASR Adaptation Through Layer-Wise Analysis

    Souranil Kahali, Rituparna Bose, Abner Hernandez +4

    cs.CLcs.AIcs.LGarXiv:2608.18825v12026