Machine Learning

Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

7,921 to 7,980 of 20,177

  1. LangFlow: Continuous Diffusion Rivals Discrete in Language Modeling

    Yuxin Chen, Chumeng Liang, Hangke Sui +4

    cs.CLcs.LGarXiv:2604.11748v32026
  2. Efficient and Principled Scientific Discovery through Bayesian Optimization: A Tutorial

    Zhongwei Yu, Rasul Tutunov, Alexandre Max Maraval +13

    cs.LGarXiv:2604.01328v32026
  3. How Far Can Unsupervised RLVR Scale LLM Training?

    Bingxiang He, Yuxin Zuo, Zeyuan Liu +18

    cs.LGcs.CLarXiv:2603.08660v12026
  4. Initialisation Determines the Basin: Efficient Codebook Optimisation for Extreme LLM Quantization

    Ian W. Kennedy, Nafise Sadat Moosavi

    cs.CLcs.LGarXiv:2604.08118v12026
  5. DreamPartGen: Semantically Grounded Part-Level 3D Generation via Collaborative Latent Denoising

    Tianjiao Yu, Xinzhuo Li, Muntasir Wahed +4

    cs.CVcs.AIcs.LGarXiv:2603.19216v22026
  6. Breaking the Capability Ceiling of LLM Post-Training by Reintroducing Markov States

    Yurun Yuan, Tengyang Xie

    cs.LGcs.AIcs.CLarXiv:2603.19987v12026
  7. POLCA: Stochastic Generative Optimization with LLM

    Xuanfei Ren, Allen Nie, Tengyang Xie +1

    cs.LGcs.AIarXiv:2603.14769v12026
  8. AgilePruner: An Empirical Study of Attention and Diversity for Adaptive Visual Token Pruning in Large Vision-Language Models

    Changwoo Baek, Jouwon Song, Sohyeon Kim +1

    cs.CVcs.LGarXiv:2603.01236v12026
  9. Efficient Reasoning with Balanced Thinking

    Yulin Li, Tengyao Tu, Li Ding +5

    cs.AIcs.CLcs.LGarXiv:2603.12372v32026
  10. SuperLocalMemory V3: Information-Geometric Foundations for Zero-LLM Enterprise Agent Memory

    Varun Pratap Bhardwaj

    cs.AIcs.IRcs.LGarXiv:2603.14588v12026
  11. The Curse and Blessing of Mean Bias in FP4-Quantized LLM Training

    Hengjie Cao, Zhendong Huang, Mengyi Chen +15

    cs.LGcs.AIarXiv:2603.10444v22026
  12. Spectral Condition for $μ$P under Width-Depth Scaling

    Chenyu Zheng, Rongzhen Wang, Xinyu Zhang +1

    cs.LGstat.MLarXiv:2603.00541v22026
  13. Prescriptive Scaling Reveals the Evolution of Language Model Capabilities

    Hanlin Zhang, Jikai Jin, Vasilis Syrgkanis +1

    cs.LGcs.AIcs.CLarXiv:2602.15327v22026
    Summaries:한국어
  14. Learning from Trials and Errors: Reflective Test-Time Planning for Embodied LLMs

    Yining Hong, Huang Huang, Manling Li +4

    cs.LGcs.AIcs.CLarXiv:2602.21198v32026
  15. FLAC: Maximum Entropy RL via Kinetic Energy Regularized Bridge Matching

    Lei Lv, Yunfei Li, Yu Luo +2

    cs.LGcs.AIarXiv:2602.12829v12026
  16. TADA! Tuning Audio Diffusion Models through Activation Steering

    Łukasz Staniszewski, Katarzyna Zaleska, Mateusz Modrzejewski +1

    cs.SDcs.LGarXiv:2602.11910v22026
  17. Vision Transformer Finetuning Benefits from Non-Smooth Components

    Ambroise Odonnat, Laetitia Chapel, Romain Tavenard +1

    cs.LGcs.CVstat.MLarXiv:2602.06883v32026
  18. Dr. MAS: Stable Reinforcement Learning for Multi-Agent LLM Systems

    Lang Feng, Longtao Zheng, Shuo He +2

    cs.LGcs.AIarXiv:2602.08847v12026
  19. Approximation of Log-Partition Function in Policy Mirror Descent Induces Implicit Regularization for LLM Post-Training

    Zhenghao Xu, Qin Lu, Changlong Yu +1

    cs.LGarXiv:2602.05933v12026
  20. Learning Rate Matters: Vanilla LoRA May Suffice for LLM Fine-tuning

    Yu-Ang Lee, Ching-Yun Ko, Pin-Yu Chen +1

    cs.LGcs.AIcs.CLarXiv:2602.04998v22026
  21. Demystifying the Slash Pattern in Attention: The Role of RoPE

    Yuan Cheng, Fengzhuo Zhang, Yunlong Hou +5

    cs.LGcs.AIcs.CLarXiv:2601.08297v22026
  22. Neural Predictor-Corrector: Solving Homotopy Problems with Reinforcement Learning

    Jiayao Mai, Bangyan Liao, Zhenjun Zhao +6

    cs.LGcs.CVarXiv:2602.03086v12026
  23. On the Relationship Between Representation Geometry and Generalization in Deep Neural Networks

    Sumit Yadav

    cs.LGarXiv:2602.00130v22026
  24. Grounding and Enhancing Informativeness and Utility in Dataset Distillation

    Shaobo Wang, Yantai Yang, Guo Chen +5

    cs.LGcs.AIarXiv:2601.21296v12026
  25. UniversalRAG: Retrieval-Augmented Generation over Corpora of Diverse Modalities and Granularities

    Woongyeong Yeo, Kangsan Kim, Soyeong Jeong +2

    cs.CLcs.AIcs.CVarXiv:2504.20734v52025
    Summaries:한국어
  26. HalluGuard: Demystifying Data-Driven and Reasoning-Driven Hallucinations in LLMs

    Xinyue Zeng, Junhong Lin, Yujun Yan +4

    cs.LGcs.AIarXiv:2601.18753v22026
  27. FP8-RL: A Practical and Stable Low-Precision Stack for LLM Reinforcement Learning

    Zhaopeng Qiu, Shuang Yu, Jingqi Zhang +4

    cs.LGcs.CLarXiv:2601.18150v22026
  28. Diffusion Policy Policy Optimization

    Allen Z. Ren, Justin Lidard, Lars L. Ankile +6

    cs.ROcs.LGarXiv:2409.00588v32024
  29. Diffusion for World Modeling: Visual Details Matter in Atari

    Eloi Alonso, Adam Jelley, Vincent Micheli +4

    cs.LGcs.AIcs.CVarXiv:2405.12399v22024
  30. The Faiss library

    Matthijs Douze, Alexandr Guzhva, Chengqi Deng +6

    cs.LGcs.CVcs.SEarXiv:2401.08281v42024
  31. Latent Consistency Models: Synthesizing High-Resolution Images with Few-Step Inference

    Simian Luo, Yiqin Tan, Longbo Huang +2

    cs.CVcs.LGarXiv:2310.04378v12023
  32. Nougat: Neural Optical Understanding for Academic Documents

    Lukas Blecher, Guillem Cucurull, Thomas Scialom +1

    cs.LGcs.CVarXiv:2308.13418v12023
  33. A Survey on Graph Neural Networks for Time Series: Forecasting, Classification, Imputation, and Anomaly Detection

    Ming Jin, Huan Yee Koh, Qingsong Wen +5

    cs.LGcs.AIarXiv:2307.03759v32023
  34. LLM4TS: Aligning Pre-Trained LLMs as Data-Efficient Time-Series Forecasters

    Ching Chang, Wei-Yao Wang, Wen-Chih Peng +1

    cs.LGarXiv:2308.08469v62023
  35. Direct Preference Optimization: Your Language Model is Secretly a Reward Model

    Rafael Rafailov, Archit Sharma, Eric Mitchell +3

    cs.LGcs.AIcs.CLarXiv:2305.18290v32023
  36. Principle-Driven Self-Alignment of Language Models from Scratch with Minimal Human Supervision

    Zhiqing Sun, Yikang Shen, Qinhong Zhou +5

    cs.LGcs.AIcs.CLarXiv:2305.03047v22023
    Summaries:한국어
  37. Align your Latents: High-Resolution Video Synthesis with Latent Diffusion Models

    Andreas Blattmann, Robin Rombach, Huan Ling +4

    cs.CVcs.LGarXiv:2304.08818v22023
  38. Analyzing Leakage of Personally Identifiable Information in Language Models

    Nils Lukas, Ahmed Salem, Robert Sim +3

    cs.LGarXiv:2302.00539v42023
  39. A Comprehensive Survey of Continual Learning: Theory, Method and Application

    Liyuan Wang, Xingxing Zhang, Hang Su +1

    cs.LGcs.AIcs.CVarXiv:2302.00487v32023
  40. PromptCast: A New Prompt-based Learning Paradigm for Time Series Forecasting

    Hao Xue, Flora D. Salim

    stat.MEcs.AIcs.CLarXiv:2210.08964v52022
  41. Diffusion Models in Vision: A Survey

    Florinel-Alin Croitoru, Vlad Hondru, Radu Tudor Ionescu +1

    cs.CVcs.AIcs.LGarXiv:2209.04747v62022
  42. ProgPrompt: Generating Situated Robot Task Plans using Large Language Models

    Ishika Singh, Valts Blukis, Arsalan Mousavian +6

    cs.ROcs.AIcs.CLarXiv:2209.11302v12022
  43. Beyond Transmitting Bits: Context, Semantics, and Task-Oriented Communications

    Deniz Gunduz, Zhijin Qin, Inaki Estella Aguerri +5

    cs.ITcs.AIcs.LGarXiv:2207.09353v22022
  44. MACE: Higher Order Equivariant Message Passing Neural Networks for Fast and Accurate Force Fields

    Ilyes Batatia, Dávid Péter Kovács, Gregor N. C. Simm +2

    stat.MLcond-mat.mtrl-scics.LGarXiv:2206.07697v22022
  45. Photorealistic Text-to-Image Diffusion Models with Deep Language Understanding

    Chitwan Saharia, William Chan, Saurabh Saxena +11

    cs.CVcs.LGarXiv:2205.11487v12022
  46. FLEURS: Few-shot Learning Evaluation of Universal Representations of Speech

    Alexis Conneau, Min Ma, Simran Khanuja +6

    cs.CLcs.LGcs.SDarXiv:2205.12446v12022
  47. Fast Sampling of Diffusion Models with Exponential Integrator

    Qinsheng Zhang, Yongxin Chen

    cs.LGarXiv:2204.13902v42022
  48. Training language models to follow instructions with human feedback

    Long Ouyang, Jeff Wu, Xu Jiang +17

    cs.CLcs.AIcs.LGarXiv:2203.02155v12022
    Summaries:한국어
  49. Data Collection and Quality Challenges in Deep Learning: A Data-Centric AI Perspective

    Steven Euijong Whang, Yuji Roh, Hwanjun Song +1

    cs.LGarXiv:2112.06409v32021
  50. Temporal Context Mining for Learned Video Compression

    Xihua Sheng, Jiahao Li, Bin Li +3

    cs.CVcs.LGeess.IVarXiv:2111.13850v22021
  51. Combining Recurrent, Convolutional, and Continuous-time Models with Linear State-Space Layers

    Albert Gu, Isys Johnson, Karan Goel +4

    cs.LGcs.AIarXiv:2110.13985v12021
  52. An Evaluation of Anomaly Detection and Diagnosis in Multivariate Time Series

    Astha Garg, Wenyu Zhang, Jules Samaran +2

    cs.LGcs.AIstat.MLarXiv:2109.11428v12021
  53. Pre-train, Prompt, and Predict: A Systematic Survey of Prompting Methods in Natural Language Processing

    Pengfei Liu, Weizhe Yuan, Jinlan Fu +3

    cs.CLcs.AIcs.LGarXiv:2107.13586v12021
  54. Spatio-temporal graph neural networks for multi-site PV power forecasting

    Jelena Simeunović, Baptiste Schubnel, Pierre-Jean Alet +1

    cs.LGeess.SParXiv:2107.13875v22021
  55. A Survey of Uncertainty in Deep Neural Networks

    Jakob Gawlikowski, Cedrique Rovile Njieutcheu Tassi, Mohsin Ali +11

    cs.LGstat.MLarXiv:2107.03342v32021
  56. SoundStream: An End-to-End Neural Audio Codec

    Neil Zeghidour, Alejandro Luebs, Ahmed Omran +2

    cs.SDcs.LGeess.ASarXiv:2107.03312v12021
  57. Structured Denoising Diffusion Models in Discrete State-Spaces

    Jacob Austin, Daniel D. Johnson, Jonathan Ho +2

    cs.LGcs.AIcs.CLarXiv:2107.03006v32021
  58. SpeechBrain: A General-Purpose Speech Toolkit

    Mirco Ravanelli, Titouan Parcollet, Peter Plantinga +18

    eess.AScs.AIcs.LGarXiv:2106.04624v12021
  59. How Attentive are Graph Attention Networks?

    Shaked Brody, Uri Alon, Eran Yahav

    cs.LGarXiv:2105.14491v32021
  60. Numerical Composition of Differential Privacy

    Sivakanth Gopi, Yin Tat Lee, Lukas Wutschitz

    cs.DScs.CRcs.LGarXiv:2106.02848v32021