Machine Learning
Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
7,921 to 7,980 of 20,177
LangFlow: Continuous Diffusion Rivals Discrete in Language Modeling
Yuxin Chen, Chumeng Liang, Hangke Sui +4
cs.CLcs.LGarXiv:2604.11748v32026Efficient and Principled Scientific Discovery through Bayesian Optimization: A Tutorial
Zhongwei Yu, Rasul Tutunov, Alexandre Max Maraval +13
cs.LGarXiv:2604.01328v32026How Far Can Unsupervised RLVR Scale LLM Training?
Bingxiang He, Yuxin Zuo, Zeyuan Liu +18
cs.LGcs.CLarXiv:2603.08660v12026Initialisation Determines the Basin: Efficient Codebook Optimisation for Extreme LLM Quantization
Ian W. Kennedy, Nafise Sadat Moosavi
cs.CLcs.LGarXiv:2604.08118v12026DreamPartGen: Semantically Grounded Part-Level 3D Generation via Collaborative Latent Denoising
Tianjiao Yu, Xinzhuo Li, Muntasir Wahed +4
cs.CVcs.AIcs.LGarXiv:2603.19216v22026Breaking the Capability Ceiling of LLM Post-Training by Reintroducing Markov States
Yurun Yuan, Tengyang Xie
cs.LGcs.AIcs.CLarXiv:2603.19987v12026POLCA: Stochastic Generative Optimization with LLM
Xuanfei Ren, Allen Nie, Tengyang Xie +1
cs.LGcs.AIarXiv:2603.14769v12026AgilePruner: An Empirical Study of Attention and Diversity for Adaptive Visual Token Pruning in Large Vision-Language Models
Changwoo Baek, Jouwon Song, Sohyeon Kim +1
cs.CVcs.LGarXiv:2603.01236v12026Efficient Reasoning with Balanced Thinking
Yulin Li, Tengyao Tu, Li Ding +5
cs.AIcs.CLcs.LGarXiv:2603.12372v32026SuperLocalMemory V3: Information-Geometric Foundations for Zero-LLM Enterprise Agent Memory
Varun Pratap Bhardwaj
cs.AIcs.IRcs.LGarXiv:2603.14588v12026The Curse and Blessing of Mean Bias in FP4-Quantized LLM Training
Hengjie Cao, Zhendong Huang, Mengyi Chen +15
cs.LGcs.AIarXiv:2603.10444v22026Spectral Condition for $μ$P under Width-Depth Scaling
Chenyu Zheng, Rongzhen Wang, Xinyu Zhang +1
cs.LGstat.MLarXiv:2603.00541v22026Prescriptive Scaling Reveals the Evolution of Language Model Capabilities
Hanlin Zhang, Jikai Jin, Vasilis Syrgkanis +1
cs.LGcs.AIcs.CLarXiv:2602.15327v22026Summaries:한국어Learning from Trials and Errors: Reflective Test-Time Planning for Embodied LLMs
Yining Hong, Huang Huang, Manling Li +4
cs.LGcs.AIcs.CLarXiv:2602.21198v32026FLAC: Maximum Entropy RL via Kinetic Energy Regularized Bridge Matching
Lei Lv, Yunfei Li, Yu Luo +2
cs.LGcs.AIarXiv:2602.12829v12026TADA! Tuning Audio Diffusion Models through Activation Steering
Łukasz Staniszewski, Katarzyna Zaleska, Mateusz Modrzejewski +1
cs.SDcs.LGarXiv:2602.11910v22026Vision Transformer Finetuning Benefits from Non-Smooth Components
Ambroise Odonnat, Laetitia Chapel, Romain Tavenard +1
cs.LGcs.CVstat.MLarXiv:2602.06883v32026Dr. MAS: Stable Reinforcement Learning for Multi-Agent LLM Systems
Lang Feng, Longtao Zheng, Shuo He +2
cs.LGcs.AIarXiv:2602.08847v12026Approximation of Log-Partition Function in Policy Mirror Descent Induces Implicit Regularization for LLM Post-Training
Zhenghao Xu, Qin Lu, Changlong Yu +1
cs.LGarXiv:2602.05933v12026Learning Rate Matters: Vanilla LoRA May Suffice for LLM Fine-tuning
Yu-Ang Lee, Ching-Yun Ko, Pin-Yu Chen +1
cs.LGcs.AIcs.CLarXiv:2602.04998v22026Demystifying the Slash Pattern in Attention: The Role of RoPE
Yuan Cheng, Fengzhuo Zhang, Yunlong Hou +5
cs.LGcs.AIcs.CLarXiv:2601.08297v22026Neural Predictor-Corrector: Solving Homotopy Problems with Reinforcement Learning
Jiayao Mai, Bangyan Liao, Zhenjun Zhao +6
cs.LGcs.CVarXiv:2602.03086v12026On the Relationship Between Representation Geometry and Generalization in Deep Neural Networks
Sumit Yadav
cs.LGarXiv:2602.00130v22026Grounding and Enhancing Informativeness and Utility in Dataset Distillation
Shaobo Wang, Yantai Yang, Guo Chen +5
cs.LGcs.AIarXiv:2601.21296v12026UniversalRAG: Retrieval-Augmented Generation over Corpora of Diverse Modalities and Granularities
Woongyeong Yeo, Kangsan Kim, Soyeong Jeong +2
cs.CLcs.AIcs.CVarXiv:2504.20734v52025Summaries:한국어HalluGuard: Demystifying Data-Driven and Reasoning-Driven Hallucinations in LLMs
Xinyue Zeng, Junhong Lin, Yujun Yan +4
cs.LGcs.AIarXiv:2601.18753v22026FP8-RL: A Practical and Stable Low-Precision Stack for LLM Reinforcement Learning
Zhaopeng Qiu, Shuang Yu, Jingqi Zhang +4
cs.LGcs.CLarXiv:2601.18150v22026Diffusion Policy Policy Optimization
Allen Z. Ren, Justin Lidard, Lars L. Ankile +6
cs.ROcs.LGarXiv:2409.00588v32024Diffusion for World Modeling: Visual Details Matter in Atari
Eloi Alonso, Adam Jelley, Vincent Micheli +4
cs.LGcs.AIcs.CVarXiv:2405.12399v22024The Faiss library
Matthijs Douze, Alexandr Guzhva, Chengqi Deng +6
cs.LGcs.CVcs.SEarXiv:2401.08281v42024Latent Consistency Models: Synthesizing High-Resolution Images with Few-Step Inference
Simian Luo, Yiqin Tan, Longbo Huang +2
cs.CVcs.LGarXiv:2310.04378v12023Nougat: Neural Optical Understanding for Academic Documents
Lukas Blecher, Guillem Cucurull, Thomas Scialom +1
cs.LGcs.CVarXiv:2308.13418v12023A Survey on Graph Neural Networks for Time Series: Forecasting, Classification, Imputation, and Anomaly Detection
Ming Jin, Huan Yee Koh, Qingsong Wen +5
cs.LGcs.AIarXiv:2307.03759v32023LLM4TS: Aligning Pre-Trained LLMs as Data-Efficient Time-Series Forecasters
Ching Chang, Wei-Yao Wang, Wen-Chih Peng +1
cs.LGarXiv:2308.08469v62023Direct Preference Optimization: Your Language Model is Secretly a Reward Model
Rafael Rafailov, Archit Sharma, Eric Mitchell +3
cs.LGcs.AIcs.CLarXiv:2305.18290v32023Principle-Driven Self-Alignment of Language Models from Scratch with Minimal Human Supervision
Zhiqing Sun, Yikang Shen, Qinhong Zhou +5
cs.LGcs.AIcs.CLarXiv:2305.03047v22023Summaries:한국어Align your Latents: High-Resolution Video Synthesis with Latent Diffusion Models
Andreas Blattmann, Robin Rombach, Huan Ling +4
cs.CVcs.LGarXiv:2304.08818v22023Analyzing Leakage of Personally Identifiable Information in Language Models
Nils Lukas, Ahmed Salem, Robert Sim +3
cs.LGarXiv:2302.00539v42023A Comprehensive Survey of Continual Learning: Theory, Method and Application
Liyuan Wang, Xingxing Zhang, Hang Su +1
cs.LGcs.AIcs.CVarXiv:2302.00487v32023PromptCast: A New Prompt-based Learning Paradigm for Time Series Forecasting
Hao Xue, Flora D. Salim
stat.MEcs.AIcs.CLarXiv:2210.08964v52022Diffusion Models in Vision: A Survey
Florinel-Alin Croitoru, Vlad Hondru, Radu Tudor Ionescu +1
cs.CVcs.AIcs.LGarXiv:2209.04747v62022ProgPrompt: Generating Situated Robot Task Plans using Large Language Models
Ishika Singh, Valts Blukis, Arsalan Mousavian +6
cs.ROcs.AIcs.CLarXiv:2209.11302v12022Beyond Transmitting Bits: Context, Semantics, and Task-Oriented Communications
Deniz Gunduz, Zhijin Qin, Inaki Estella Aguerri +5
cs.ITcs.AIcs.LGarXiv:2207.09353v22022MACE: Higher Order Equivariant Message Passing Neural Networks for Fast and Accurate Force Fields
Ilyes Batatia, Dávid Péter Kovács, Gregor N. C. Simm +2
stat.MLcond-mat.mtrl-scics.LGarXiv:2206.07697v22022Photorealistic Text-to-Image Diffusion Models with Deep Language Understanding
Chitwan Saharia, William Chan, Saurabh Saxena +11
cs.CVcs.LGarXiv:2205.11487v12022FLEURS: Few-shot Learning Evaluation of Universal Representations of Speech
Alexis Conneau, Min Ma, Simran Khanuja +6
cs.CLcs.LGcs.SDarXiv:2205.12446v12022Fast Sampling of Diffusion Models with Exponential Integrator
Qinsheng Zhang, Yongxin Chen
cs.LGarXiv:2204.13902v42022Training language models to follow instructions with human feedback
Long Ouyang, Jeff Wu, Xu Jiang +17
cs.CLcs.AIcs.LGarXiv:2203.02155v12022Summaries:한국어Data Collection and Quality Challenges in Deep Learning: A Data-Centric AI Perspective
Steven Euijong Whang, Yuji Roh, Hwanjun Song +1
cs.LGarXiv:2112.06409v32021Temporal Context Mining for Learned Video Compression
Xihua Sheng, Jiahao Li, Bin Li +3
cs.CVcs.LGeess.IVarXiv:2111.13850v22021Combining Recurrent, Convolutional, and Continuous-time Models with Linear State-Space Layers
Albert Gu, Isys Johnson, Karan Goel +4
cs.LGcs.AIarXiv:2110.13985v12021An Evaluation of Anomaly Detection and Diagnosis in Multivariate Time Series
Astha Garg, Wenyu Zhang, Jules Samaran +2
cs.LGcs.AIstat.MLarXiv:2109.11428v12021Pre-train, Prompt, and Predict: A Systematic Survey of Prompting Methods in Natural Language Processing
Pengfei Liu, Weizhe Yuan, Jinlan Fu +3
cs.CLcs.AIcs.LGarXiv:2107.13586v12021Spatio-temporal graph neural networks for multi-site PV power forecasting
Jelena Simeunović, Baptiste Schubnel, Pierre-Jean Alet +1
cs.LGeess.SParXiv:2107.13875v22021A Survey of Uncertainty in Deep Neural Networks
Jakob Gawlikowski, Cedrique Rovile Njieutcheu Tassi, Mohsin Ali +11
cs.LGstat.MLarXiv:2107.03342v32021SoundStream: An End-to-End Neural Audio Codec
Neil Zeghidour, Alejandro Luebs, Ahmed Omran +2
cs.SDcs.LGeess.ASarXiv:2107.03312v12021Structured Denoising Diffusion Models in Discrete State-Spaces
Jacob Austin, Daniel D. Johnson, Jonathan Ho +2
cs.LGcs.AIcs.CLarXiv:2107.03006v32021SpeechBrain: A General-Purpose Speech Toolkit
Mirco Ravanelli, Titouan Parcollet, Peter Plantinga +18
eess.AScs.AIcs.LGarXiv:2106.04624v12021How Attentive are Graph Attention Networks?
Shaked Brody, Uri Alon, Eran Yahav
cs.LGarXiv:2105.14491v32021Numerical Composition of Differential Privacy
Sivakanth Gopi, Yin Tat Lee, Lukas Wutschitz
cs.DScs.CRcs.LGarXiv:2106.02848v32021