Machine Learning

Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

421 to 480 of 20,454

  1. MuProp: Unbiased Backpropagation for Stochastic Neural Networks

    Shixiang Gu, Sergey Levine, Ilya Sutskever +1

    cs.LGarXiv:1511.05176v32015
  2. Gaussian Process Upper Confidence Bound Achieves Nearly-Optimal Regret in Noise-Free Gaussian Process Bandits

    Shogo Iwazaki

    cs.LGarXiv:2502.19006v22025
  3. Can AutoML outperform humans? An evaluation on popular OpenML datasets using AutoML Benchmark

    Marc Hanussek, Matthias Blohm, Maximilien Kintz

    cs.LGstat.MLarXiv:2009.01564v22020
  4. Learning diverse attacks on large language models for robust red-teaming and safety tuning

    Seanie Lee, Minsu Kim, Lynn Cherif +8

    cs.CLcs.CRcs.LGarXiv:2405.18540v32024
  5. Directed Diffusion: Direct Control of Object Placement through Attention Guidance

    Wan-Duo Kurt Ma, J. P. Lewis, Avisek Lahiri +2

    cs.CVcs.GRcs.LGarXiv:2302.13153v32023
  6. Aggregated Momentum: Stability Through Passive Damping

    James Lucas, Shengyang Sun, Richard Zemel +1

    cs.LGcs.AImath.OCarXiv:1804.00325v32018
  7. Attacking the Trusted Imagination: Oracle-Level Integrity Attacks on Imagine-then-Act World Models

    Linghan Chen, Kaiyan Ji, Minyu Guo

    cs.LGcs.AIcs.CRarXiv:2606.22966v12026
  8. Self-Supervised Graph Representation Learning via Global Context Prediction

    Zhen Peng, Yixiang Dong, Minnan Luo +2

    cs.LGstat.MLarXiv:2003.01604v12020
  9. Tomayto, Tomahto. Beyond Token-level Answer Equivalence for Question Answering Evaluation

    Jannis Bulian, Christian Buck, Wojciech Gajewski +2

    cs.CLcs.LGarXiv:2202.07654v22022
  10. A theory of continuous generative flow networks

    Salem Lahlou, Tristan Deleu, Pablo Lemos +6

    cs.LGstat.MLarXiv:2301.12594v22023
  11. Reparameterizable Subset Sampling via Continuous Relaxations

    Sang Michael Xie, Stefano Ermon

    cs.LGstat.MLarXiv:1901.10517v52019
  12. MaskLLM: Learnable Semi-Structured Sparsity for Large Language Models

    Gongfan Fang, Hongxu Yin, Saurav Muralidharan +5

    cs.AIcs.CLcs.LGarXiv:2409.17481v22024
  13. Optimal and Adaptive Algorithms for Online Boosting

    Alina Beygelzimer, Satyen Kale, Haipeng Luo

    cs.LGarXiv:1502.02651v12015
  14. Measuring In-Context Computation Complexity via Hidden State Prediction

    Vincent Herrmann, Róbert Csordás, Jürgen Schmidhuber

    cs.LGarXiv:2503.13431v12025
  15. Image Generation Via Minimizing Fr\'echet Distance in Discriminator Feature Space

    Khoa D. Doan, Saurav Manchanda, Fengjiao Wang +3

    cs.CVcs.LGeess.IVarXiv:2003.11774v22020
  16. A Comprehensive Benchmark of Machine and Deep Learning Across Diverse Tabular Datasets

    Assaf Shmuel, Oren Glickman, Teddy Lazebnik

    cs.LGcs.AIarXiv:2408.14817v12024
  17. TabReD: Analyzing Pitfalls and Filling the Gaps in Tabular Deep Learning Benchmarks

    Ivan Rubachev, Nikolay Kartashev, Yury Gorishniy +1

    cs.LGarXiv:2406.19380v42024
  18. fev-bench: A Realistic Benchmark for Time Series Forecasting

    Oleksandr Shchur, Abdul Fatir Ansari, Caner Turkmen +5

    cs.LGarXiv:2509.26468v42025
  19. A Closer Look at Deep Learning Methods on Tabular Datasets

    Han-Jia Ye, Si-Yang Liu, Hao-Run Cai +2

    cs.LGarXiv:2407.00956v42024
  20. Unreflected Use of Tabular Data Repositories Can Undermine Research Quality

    Andrej Tschalzev, Lennart Purucker, Stefan Lüdtke +3

    cs.LGarXiv:2503.09159v12025
  21. TetraJet-v2: Accurate NVFP4 Training for Large Language Models with Oscillation Suppression and Outlier Control

    Yuxiang Chen, Yifan Liu, Xiaoming Xu +5

    cs.LGcs.AIarXiv:2510.27527v32025
  22. Distinguishing the Knowable from the Unknowable with Language Models

    Gustaf Ahdritz, Tian Qin, Nikhil Vyas +2

    cs.LGcs.AIcs.CLarXiv:2402.03563v22024
  23. TruthRL: Incentivizing Truthful LLMs via Reinforcement Learning

    Zhepei Wei, Xiao Yang, Kai Sun +12

    cs.CLcs.AIcs.LGarXiv:2509.25760v22025
  24. blob loss: instance imbalance aware loss functions for semantic segmentation

    Florian Kofler, Suprosanna Shit, Ivan Ezhov +14

    cs.CVcs.LGeess.IVarXiv:2205.08209v32022
  25. Representation Learning for Grounded Spatial Reasoning

    Michael Janner, Karthik Narasimhan, Regina Barzilay

    cs.CLcs.AIcs.LGarXiv:1707.03938v22017
  26. Agentic Reinforcement Learning with Observation-Calibrated Self-Distillation

    Yi Yang, Cong Qin, Xiaodan Liu +8

    cs.LGcs.AIcs.CLarXiv:2608.04788v12026
  27. Scalable Second Order Optimization for Deep Learning

    Rohan Anil, Vineet Gupta, Tomer Koren +2

    cs.LGmath.OCstat.MLarXiv:2002.09018v22020
  28. Syntactic and Semantic Control of Large Language Models via Sequential Monte Carlo

    João Loula, Benjamin LeBrun, Li Du +12

    cs.CLcs.AIcs.LGarXiv:2504.13139v22025
  29. Is Private Learning Possible with Instance Encoding?

    Nicholas Carlini, Samuel Deng, Sanjam Garg +6

    cs.CRcs.CVcs.LGarXiv:2011.05315v22020
  30. The Alignment Problem in Constrained Code Generation

    Matteo Biagiola, Jahrim Gabriele Cesario, Luca Di Grazia +2

    cs.SEcs.LGcs.PLarXiv:2606.21619v12026
  31. CRUST-Bench: A Comprehensive Benchmark for C-to-safe-Rust Transpilation

    Anirudh Khatry, Robert Zhang, Jia Pan +4

    cs.SEcs.CLcs.LGarXiv:2504.15254v32025
  32. Guiding LLMs The Right Way: Fast, Non-Invasive Constrained Generation

    Luca Beurer-Kellner, Marc Fischer, Martin Vechev

    cs.LGcs.CLarXiv:2403.06988v12024
  33. Explaining Question Answering Models through Text Generation

    Veronica Latcinnik, Jonathan Berant

    cs.CLcs.AIcs.LGarXiv:2004.05569v12020
  34. Knowledge-Grounded Self-Rationalization via Extractive and Natural Language Explanations

    Bodhisattwa Prasad Majumder, Oana-Maria Camburu, Thomas Lukasiewicz +1

    cs.CLcs.AIcs.LGarXiv:2106.13876v42021
  35. Goal Driven Discovery of Distributional Differences via Language Descriptions

    Ruiqi Zhong, Peter Zhang, Steve Li +3

    cs.CLcs.AIcs.LGarXiv:2302.14233v22023
  36. Sum-max Submodular Bandits

    Stephen Pasteris, Alberto Rumi, Fabio Vitale +1

    cs.LGarXiv:2311.05975v12023
  37. A Meta-Learning Approach for Graph Representation Learning in Multi-Task Settings

    Davide Buffelli, Fabio Vandin

    cs.LGcs.AIarXiv:2012.06755v12020
  38. Lossless Acceleration for Seq2seq Generation with Aggressive Decoding

    Tao Ge, Heming Xia, Xin Sun +2

    cs.CLcs.LGarXiv:2205.10350v12022
  39. On the insufficiency of existing momentum schemes for Stochastic Optimization

    Rahul Kidambi, Praneeth Netrapalli, Prateek Jain +1

    cs.LGmath.OCstat.MLarXiv:1803.05591v22018
  40. Evaluating Machine Unlearning via Epistemic Uncertainty

    Alexander Becker, Thomas Liebig

    cs.LGarXiv:2208.10836v22022
  41. TimeDP: Learning to Generate Multi-Domain Time Series with Domain Prompts

    Yu-Hao Huang, Chang Xu, Yueying Wu +2

    cs.LGcs.AIarXiv:2501.05403v12025
  42. Reliable Graph Neural Networks via Robust Aggregation

    Simon Geisler, Daniel Zügner, Stephan Günnemann

    cs.LGstat.MLarXiv:2010.15651v12020
  43. Efficient Online Data Mixing For Language Model Pre-Training

    Alon Albalak, Liangming Pan, Colin Raffel +1

    cs.CLcs.LGarXiv:2312.02406v22023
  44. Language models scale reliably with over-training and on downstream tasks

    Samir Yitzhak Gadre, Georgios Smyrnis, Vaishaal Shankar +22

    cs.CLcs.LGarXiv:2403.08540v22024
  45. GES: Generalized Exponential Splatting for Efficient Radiance Field Rendering

    Abdullah Hamdi, Luke Melas-Kyriazi, Jinjie Mai +5

    cs.CVcs.GRcs.LGarXiv:2402.10128v22024
  46. Pathways on the Image Manifold: Image Editing via Video Generation

    Noam Rotstein, Gal Yona, Daniel Silver +3

    cs.CVcs.AIcs.LGarXiv:2411.16819v42024
  47. Model-Preserving Adaptive Rounding

    Albert Tseng, Zhaofeng Sun, Christopher De Sa

    cs.LGcs.AIarXiv:2505.22988v32025
  48. AgentOhana: Design Unified Data and Training Pipeline for Effective Agent Learning

    Jianguo Zhang, Tian Lan, Rithesh Murthy +15

    cs.AIcs.CLcs.LGarXiv:2402.15506v42024
  49. Collaborative Unsupervised Visual Representation Learning from Decentralized Data

    Weiming Zhuang, Xin Gan, Yonggang Wen +2

    cs.DCcs.AIcs.CVarXiv:2108.06492v12021
  50. FAHT: An Adaptive Fairness-aware Decision Tree Classifier

    Wenbin Zhang, Eirini Ntoutsi

    cs.LGcs.AIstat.MLarXiv:1907.07237v12019
  51. DiscoGen: Procedural Generation of Algorithm Discovery Tasks in Machine Learning

    Alexander D. Goldie, Zilin Wang, Adrian Hayler +17

    cs.LGcs.AIarXiv:2603.17863v22026
  52. All SMILES Variational Autoencoder

    Zaccary Alperstein, Artem Cherkasov, Jason Tyler Rolfe

    cs.LGstat.MLarXiv:1905.13343v22019
  53. Temporal Logic Specification-Conditioned Decision Transformer for Offline Safe Reinforcement Learning

    Zijian Guo, Weichao Zhou, Wenchao Li

    cs.LGcs.AIarXiv:2402.17217v22024
  54. DenseFormer: Enhancing Information Flow in Transformers via Depth Weighted Averaging

    Matteo Pagliardini, Amirkeivan Mohtashami, Francois Fleuret +1

    cs.CLcs.LGarXiv:2402.02622v22024
  55. Training a Generally Curious Agent

    Fahim Tajwar, Yiding Jiang, Abitha Thankaraj +4

    cs.LGcs.AIcs.CLarXiv:2502.17543v42025
  56. Evaluating and Enhancing Large Language Models for Novelty Assessment in Scholarly Publications

    Ethan Lin, Zhiyuan Peng, Yi Fang

    cs.CLcs.AIcs.IRarXiv:2409.16605v12024
  57. CryoBench: Diverse and challenging datasets for the heterogeneity problem in cryo-EM

    Minkyu Jeon, Rishwanth Raghu, Miro Astore +6

    cs.CVcs.AIcs.CEarXiv:2408.05526v22024
  58. Mixture of neural fields for heterogeneous reconstruction in cryo-EM

    Axel Levy, Rishwanth Raghu, David Shustin +5

    cs.LGarXiv:2412.09420v12024
  59. RAP: Retrieval-Augmented Personalization for Multimodal Large Language Models

    Haoran Hao, Jiaming Han, Changsheng Li +2

    cs.CVcs.AIcs.CLarXiv:2410.13360v32024
  60. POLY-HOOT: Monte-Carlo Planning in Continuous Space MDPs with Non-Asymptotic Analysis

    Weichao Mao, Kaiqing Zhang, Qiaomin Xie +1

    cs.AIcs.LGarXiv:2006.04672v22020