Machine Learning

Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

18,661 to 18,720 of 20,219

  1. Multitask Prompted Training Enables Zero-Shot Task Generalization

    Victor Sanh, Albert Webson, Colin Raffel +38

    cs.LGcs.CLarXiv:2110.08207v32021
  2. Send a SCOUT First: Pre-hoc Reasoning for Adaptive Detector Allocation in Prompt-Injection Defense

    Shuhao Zhang, Jiarui Li, Qi Cao +2

    cs.CRcs.LGarXiv:2605.30837v22026
  3. Natural Adversarial Examples

    Dan Hendrycks, Kevin Zhao, Steven Basart +2

    cs.LGcs.CVstat.MLarXiv:1907.07174v42019
  4. Semi-Supervised Noise Adaptation: Transferring Knowledge from Noise Domain

    Yuan Yao, Jin Song, Huixia Li +3

    cs.LGarXiv:2606.00558v22026
  5. Covid-19: Automatic detection from X-Ray images utilizing Transfer Learning with Convolutional Neural Networks

    Ioannis D. Apostolopoulos, Tzani Bessiana

    eess.IVcs.CVcs.LGarXiv:2003.11617v12020
  6. Representation over Routing: Diagnosing Temporal Routing Pathologies in Multi-Timescale PPO

    Jing Sun

    cs.LGcs.AIarXiv:2604.13517v42026
  7. Confidence-Adaptive SwiGLU for Mixture-of-Experts

    Shaohua Li, Xiuchao Sui, Xiaobing Sun +4

    cs.LGcs.CLarXiv:2606.00761v12026
  8. GMAN: A Graph Multi-Attention Network for Traffic Prediction

    Chuanpan Zheng, Xiaoliang Fan, Cheng Wang +1

    eess.SPcs.LGarXiv:1911.08415v22019
  9. Ensemble deep learning: A review

    M. A. Ganaie, Minghui Hu, A. K. Malik +2

    cs.LGcs.AIcs.CVarXiv:2104.02395v32021
  10. The Shape of Addition: Geometric Structures of Arithmetic in Large Language Models

    Liuyuan Wen, Xun Zhu, Lihao Huang +2

    cs.LGcs.AIarXiv:2606.03645v12026
  11. Functional Attention: From Pairwise Affinities to Functional Correspondences

    Jiefang Xiao, Maolin Gao, Simon Weber +2

    cs.LGarXiv:2605.31559v12026
  12. Honest Lying: Understanding Memory Confabulation in Reflexive Agents

    Prakhar Dixit, Sadia Kamal, Tim Oates

    cs.LGcs.AIarXiv:2605.29463v22026
  13. PaintBench: Deterministic Evaluation of Precise Visual Editing

    Kai Xu, Ellis Brown, Shrikar Madhu +3

    cs.GRcs.CVcs.LGarXiv:2606.00188v12026
  14. Relational Knowledge Distillation

    Wonpyo Park, Dongju Kim, Yan Lu +1

    cs.CVcs.LGarXiv:1904.05068v22019
  15. Stacked Attention Networks for Image Question Answering

    Zichao Yang, Xiaodong He, Jianfeng Gao +2

    cs.LGcs.CLcs.CVarXiv:1511.02274v22015
  16. Enhancing the Locality and Breaking the Memory Bottleneck of Transformer on Time Series Forecasting

    Shiyang Li, Xiaoyong Jin, Yao Xuan +4

    cs.LGstat.MLarXiv:1907.00235v32019
  17. DRAW: A Recurrent Neural Network For Image Generation

    Karol Gregor, Ivo Danihelka, Alex Graves +2

    cs.CVcs.LGcs.NEarXiv:1502.04623v22015
  18. Simple and Deep Graph Convolutional Networks

    Ming Chen, Zhewei Wei, Zengfeng Huang +2

    cs.LGstat.MLarXiv:2007.02133v12020
  19. Tutorial on Variational Autoencoders

    Carl Doersch

    stat.MLcs.LGarXiv:1606.05908v32016
  20. Learning Latent Dynamics for Planning from Pixels

    Danijar Hafner, Timothy Lillicrap, Ian Fischer +4

    cs.LGcs.AIstat.MLarXiv:1811.04551v52018
  21. Improving Variational Inference with Inverse Autoregressive Flow

    Diederik P. Kingma, Tim Salimans, Rafal Jozefowicz +3

    cs.LGstat.MLarXiv:1606.04934v22016
  22. KAN: Kolmogorov-Arnold Networks

    Ziming Liu, Yixuan Wang, Sachin Vaidya +5

    cs.LGcond-mat.dis-nncs.AIarXiv:2404.19756v52024
  23. On the Limits of LLM Adaptability: Impact of Model-Internalized Priors on Annotation Task Performance

    Etienne Casanova, Rafal Kocielnik, R. Michael Alvarez

    cs.CLcs.AIcs.LGarXiv:2606.00467v12026
  24. DiffWave: A Versatile Diffusion Model for Audio Synthesis

    Zhifeng Kong, Wei Ping, Jiaji Huang +2

    eess.AScs.CLcs.LGarXiv:2009.09761v32020
  25. Deep Bayesian Active Learning with Image Data

    Yarin Gal, Riashat Islam, Zoubin Ghahramani

    cs.LGcs.CVstat.MLarXiv:1703.02910v12017
  26. Identifying the Best Machine Learning Algorithms for Brain Tumor Segmentation, Progression Assessment, and Overall Survival Prediction in the BRATS Challenge

    Spyridon Bakas, Mauricio Reyes, Andras Jakab +424

    cs.CVcs.AIcs.LGarXiv:1811.02629v32018
  27. Measuring the Symmetry--Data Exchange Rate

    Ahmed M. Adly

    stat.MEcs.LGarXiv:2606.01090v12026
  28. Challenges in Representation Learning: A report on three machine learning contests

    Ian J. Goodfellow, Dumitru Erhan, Pierre Luc Carrier +25

    stat.MLcs.LGarXiv:1307.0414v12013
  29. Sharpness-Aware Minimization for Efficiently Improving Generalization

    Pierre Foret, Ariel Kleiner, Hossein Mobahi +1

    cs.LGstat.MLarXiv:2010.01412v32020
  30. Habitat: A Platform for Embodied AI Research

    Manolis Savva, Abhishek Kadian, Oleksandr Maksymets +9

    cs.CVcs.AIcs.CLarXiv:1904.01201v22019
  31. DOT-MoE: Differentiable Optimal Transport for MoEfication

    Udbhav Bamba, Arnav Chavan, Aryamaan Thakur +2

    cs.LGcs.AIarXiv:2606.01666v12026
  32. Alias-Free Generative Adversarial Networks

    Tero Karras, Miika Aittala, Samuli Laine +4

    cs.CVcs.AIcs.LGarXiv:2106.12423v42021
    Summaries:한국어
  33. A Local Perturbation Theory for Cross-Domain Interference and Recovery in Multi-Domain RL

    Lei Yang, Siyu Ding, Deyi Xiong

    cs.LGcs.CLarXiv:2606.02398v12026
  34. Scaling LLM Test-Time Compute Optimally can be More Effective than Scaling Model Parameters

    Charlie Snell, Jaehoon Lee, Kelvin Xu +1

    cs.LGcs.CLarXiv:2408.03314v12024
  35. An Enigma of Artificial Reason: Investigating the Production-Evaluation Gap in Large Reasoning Models

    Mingzhong Sun, Teresa Yeo, Armando Solar-Lezama +1

    cs.AIcs.CLcs.LGarXiv:2606.01462v12026
  36. Gradient Surgery for Multi-Task Learning

    Tianhe Yu, Saurabh Kumar, Abhishek Gupta +3

    cs.LGcs.CVcs.ROarXiv:2001.06782v42020
  37. Scalable Inference-Time Annealing with Surrogate Likelihood Estimators

    Daniel Peñaherrera, Rishal Aggarwal, David Ryan Koes

    cs.LGq-bio.BMarXiv:2605.31498v32026
  38. Free-Form Image Inpainting with Gated Convolution

    Jiahui Yu, Zhe Lin, Jimei Yang +3

    cs.CVcs.GRcs.LGarXiv:1806.03589v22018
  39. ToolLLM: Facilitating Large Language Models to Master 16000+ Real-world APIs

    Yujia Qin, Shihao Liang, Yining Ye +16

    cs.AIcs.CLcs.LGarXiv:2307.16789v22023
  40. Trust Functions: Near-Lossless Weak-to-Strong Generalization by Learning When to Trust the Weak Teacher

    Arda Uzunoglu, Alvin Zhang, Daniel Khashabi

    cs.LGcs.CLarXiv:2606.01000v12026
  41. Is Your Code Generated by ChatGPT Really Correct? Rigorous Evaluation of Large Language Models for Code Generation

    Jiawei Liu, Chunqiu Steven Xia, Yuyao Wang +1

    cs.SEcs.CLcs.LGarXiv:2305.01210v32023
  42. Mixtral of Experts

    Albert Q. Jiang, Alexandre Sablayrolles, Antoine Roux +23

    cs.LGcs.CLarXiv:2401.04088v12024
  43. DeBERTaV3: Improving DeBERTa using ELECTRA-Style Pre-Training with Gradient-Disentangled Embedding Sharing

    Pengcheng He, Jianfeng Gao, Weizhu Chen

    cs.CLcs.LGarXiv:2111.09543v42021
  44. Tacotron: Towards End-to-End Speech Synthesis

    Yuxuan Wang, RJ Skerry-Ryan, Daisy Stanton +11

    cs.CLcs.LGcs.SDarXiv:1703.10135v22017
  45. Vision Mamba: Efficient Visual Representation Learning with Bidirectional State Space Model

    Lianghui Zhu, Bencheng Liao, Qian Zhang +3

    cs.CVcs.LGarXiv:2401.09417v32024
  46. Solving high-dimensional partial differential equations using deep learning

    Jiequn Han, Arnulf Jentzen, Weinan E

    math.NAcs.LGmath.OCarXiv:1707.02568v32017
  47. RandLA-Net: Efficient Semantic Segmentation of Large-Scale Point Clouds

    Qingyong Hu, Bo Yang, Linhai Xie +5

    cs.CVcs.LGeess.IVarXiv:1911.11236v32019
  48. Jailbroken: How Does LLM Safety Training Fail?

    Alexander Wei, Nika Haghtalab, Jacob Steinhardt

    cs.LGcs.CRarXiv:2307.02483v12023
    Summaries:한국어
  49. LayerRoute: Input-Conditioned Adaptive Layer Skipping via LoRA Fine-Tuning for Agentic Language Models

    Prateek Kumar Sikdar

    cs.CLcs.AIcs.LGarXiv:2606.01838v12026
  50. Return of Frustratingly Easy Domain Adaptation

    Baochen Sun, Jiashi Feng, Kate Saenko

    cs.CVcs.AIcs.LGarXiv:1511.05547v22015
  51. $Ψ$-Bench: Evaluating Persona-Sensitive Influencing in Persuasive Dialogues

    Peixuan Han, Hongyi Du, Jiayu Liu +3

    cs.LGarXiv:2606.02754v12026
  52. MMG2Skill: Can Agents Distill In-the-Wild Guides into Self-Evolving Skills?

    Xinyu Che, Junqi Xiong, Yunfei Ge +10

    cs.CLcs.AIcs.LGarXiv:2606.01993v12026
  53. Adaptive Graph Convolutional Recurrent Network for Traffic Forecasting

    Lei Bai, Lina Yao, Can Li +2

    cs.LGstat.MLarXiv:2007.02842v22020
  54. Training-Free Multi-Concept LoRA Composition with Prompt-Aware Weighting

    Georgios Tsoumplekas, Stella Bounareli, Vasileios Argyriou

    cs.CVcs.LGarXiv:2606.03792v12026
  55. Large Language Models Hack Rewards, and Society

    Wei Liu, Xinyi Mou, Hanqi Yan +2

    cs.LGcs.AIcs.CLarXiv:2606.04075v22026
  56. Noise2Noise: Learning Image Restoration without Clean Data

    Jaakko Lehtinen, Jacob Munkberg, Jon Hasselgren +4

    cs.CVcs.LGstat.MLarXiv:1803.04189v32018
  57. StarGAN v2: Diverse Image Synthesis for Multiple Domains

    Yunjey Choi, Youngjung Uh, Jaejun Yoo +1

    cs.CVcs.LGarXiv:1912.01865v22019
  58. Imagen Video: High Definition Video Generation with Diffusion Models

    Jonathan Ho, William Chan, Chitwan Saharia +8

    cs.CVcs.LGarXiv:2210.02303v12022
  59. Improving Factuality and Reasoning in Language Models through Multiagent Debate

    Yilun Du, Shuang Li, Antonio Torralba +2

    cs.CLcs.AIcs.CVarXiv:2305.14325v12023
  60. Hyperparameters and Tuning Strategies for Random Forest

    Philipp Probst, Marvin Wright, Anne-Laure Boulesteix

    stat.MLcs.LGarXiv:1804.03515v22018