Machine Learning

Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

15,841 to 15,900 of 20,199

  1. The AI Scientist: Towards Fully Automated Open-Ended Scientific Discovery

    Chris Lu, Cong Lu, Robert Tjarko Lange +3

    cs.AIcs.CLcs.LGarXiv:2408.06292v32024
    Summaries:한국어
  2. Channel Estimation for RIS-Empowered Multi-User MISO Wireless Communications

    Li Wei, Chongwen Huang, George C. Alexandropoulos +3

    cs.ITcs.LGeess.SParXiv:2008.01459v22020
  3. Can AI-Generated Text be Reliably Detected?

    Vinu Sankar Sadasivan, Aounon Kumar, Sriram Balasubramanian +2

    cs.CLcs.AIcs.LGarXiv:2303.11156v42023
  4. The Loss Floor of Denoising Score Matching: Fisher Geometry from Schrödinger Bridges

    Avinash Raju, Kai Zhang

    cs.LGcond-mat.stat-mecharXiv:2608.23916v12026
  5. Towards a Guideline for Evaluation Metrics in Medical Image Segmentation

    Dominik Müller, Iñaki Soto-Rey, Frank Kramer

    eess.IVcs.CVcs.LGarXiv:2202.05273v12022
  6. Revisiting the Arcade Learning Environment: Evaluation Protocols and Open Problems for General Agents

    Marlos C. Machado, Marc G. Bellemare, Erik Talvitie +3

    cs.LGarXiv:1709.06009v22017
  7. End-to-end Learning of Action Detection from Frame Glimpses in Videos

    Serena Yeung, Olga Russakovsky, Greg Mori +1

    cs.CVcs.LGarXiv:1511.06984v22015
  8. DeepCoder: Learning to Write Programs

    Matej Balog, Alexander L. Gaunt, Marc Brockschmidt +2

    cs.LGarXiv:1611.01989v22016
  9. Dipole: Diagnosis Prediction in Healthcare via Attention-based Bidirectional Recurrent Neural Networks

    Fenglong Ma, Radha Chitta, Jing Zhou +3

    cs.LGarXiv:1706.05764v12017
  10. GATNextHop: A GAT for Shortest Path Routing with Cross-Topology Generalization

    Chia-Hong Chou, Katerina Potika

    cs.LGarXiv:2608.23917v12026
  11. A robust self-learning method for fully unsupervised cross-lingual mappings of word embeddings

    Mikel Artetxe, Gorka Labaka, Eneko Agirre

    cs.CLcs.AIcs.LGarXiv:1805.06297v22018
  12. Generative and Discriminative Voxel Modeling with Convolutional Neural Networks

    Andrew Brock, Theodore Lim, J. M. Ritchie +1

    cs.CVcs.HCcs.LGarXiv:1608.04236v22016
  13. Katyusha: The First Direct Acceleration of Stochastic Gradient Methods

    Zeyuan Allen-Zhu

    math.OCcs.DScs.LGarXiv:1603.05953v62016
  14. CODA-Prompt: COntinual Decomposed Attention-based Prompting for Rehearsal-Free Continual Learning

    James Seale Smith, Leonid Karlinsky, Vyshnavi Gutta +6

    cs.CVcs.AIcs.LGarXiv:2211.13218v22022
  15. Edge Artificial Intelligence for 6G: Vision, Enabling Technologies, and Applications

    Khaled B. Letaief, Yuanming Shi, Jianmin Lu +1

    cs.ITcs.LGcs.NIarXiv:2111.12444v12021
  16. Scalable agent alignment via reward modeling: a research direction

    Jan Leike, David Krueger, Tom Everitt +3

    cs.LGcs.AIcs.NEarXiv:1811.07871v12018
  17. Knowledge Transfer via Distillation of Activation Boundaries Formed by Hidden Neurons

    Byeongho Heo, Minsik Lee, Sangdoo Yun +1

    cs.LGcs.CVstat.MLarXiv:1811.03233v22018
  18. Fairness-Aware Mixture-of-Experts via Subgroup Reweighting and Gate Regularization

    Sunhee Hwang

    cs.LGcs.AIarXiv:2608.22820v12026
  19. QuaRot: Outlier-Free 4-Bit Inference in Rotated LLMs

    Saleh Ashkboos, Amirkeivan Mohtashami, Maximilian L. Croci +6

    cs.LGarXiv:2404.00456v22024
  20. A Convolutional Attention Network for Extreme Summarization of Source Code

    Miltiadis Allamanis, Hao Peng, Charles Sutton

    cs.LGcs.CLcs.SEarXiv:1602.03001v22016
  21. Convergence Rates of Inexact Proximal-Gradient Methods for Convex Optimization

    Mark Schmidt, Nicolas Le Roux, Francis Bach

    cs.LGmath.OCarXiv:1109.2415v22011
  22. Towards Open World Object Detection

    K J Joseph, Salman Khan, Fahad Shahbaz Khan +1

    cs.CVcs.AIcs.LGarXiv:2103.02603v22021
  23. Reliable Fidelity and Diversity Metrics for Generative Models

    Muhammad Ferjad Naeem, Seong Joon Oh, Youngjung Uh +2

    cs.CVcs.LGstat.MLarXiv:2002.09797v22020
  24. Lagrangian Neural Networks

    Miles Cranmer, Sam Greydanus, Stephan Hoyer +3

    cs.LGmath.DSphysics.comp-pharXiv:2003.04630v22020
  25. Sequence-to-Sequence Learning as Beam-Search Optimization

    Sam Wiseman, Alexander M. Rush

    cs.CLcs.LGcs.NEarXiv:1606.02960v22016
  26. When Similarity Is Interaction-Driven: Quantum Kernels for Regime-Sensitive Learning

    Hanqiu Peng, Jianlong Lu, Ying Chen

    quant-phcs.LGarXiv:2608.24631v12026
  27. Deep Learning-based Computational Pathology Predicts Origins for Cancers of Unknown Primary

    Ming Y. Lu, Melissa Zhao, Maha Shady +4

    q-bio.TOcs.LGq-bio.QMarXiv:2006.13932v22020
  28. DeepInf: Social Influence Prediction with Deep Learning

    Jiezhong Qiu, Jian Tang, Hao Ma +3

    cs.SIcs.LGarXiv:1807.05560v12018
  29. DSOD: Learning Deeply Supervised Object Detectors from Scratch

    Zhiqiang Shen, Zhuang Liu, Jianguo Li +3

    cs.CVcs.LGarXiv:1708.01241v22017
  30. Zephyr: Direct Distillation of LM Alignment

    Lewis Tunstall, Edward Beeching, Nathan Lambert +11

    cs.LGcs.CLarXiv:2310.16944v12023
  31. DeeperForensics-1.0: A Large-Scale Dataset for Real-World Face Forgery Detection

    Liming Jiang, Ren Li, Wayne Wu +2

    cs.CVcs.LGarXiv:2001.03024v22020
  32. Ab-Initio Solution of the Many-Electron Schrödinger Equation with Deep Neural Networks

    David Pfau, James S. Spencer, Alexander G. de G. Matthews +1

    physics.chem-phcs.LGphysics.comp-pharXiv:1909.02487v32019
  33. COVID-CAPS: A Capsule Network-based Framework for Identification of COVID-19 cases from X-ray Images

    Parnian Afshar, Shahin Heidarian, Farnoosh Naderkhani +3

    cs.CVcs.LGeess.IVarXiv:2004.02696v22020
  34. The UEA multivariate time series classification archive, 2018

    Anthony Bagnall, Hoang Anh Dau, Jason Lines +5

    cs.LGstat.MLarXiv:1811.00075v12018
  35. Deep Semantic Ranking Based Hashing for Multi-Label Image Retrieval

    Fang Zhao, Yongzhen Huang, Liang Wang +1

    cs.CVcs.LGarXiv:1501.06272v22015
  36. Sequential operator learning under dependent data

    Rafael Oliveira

    stat.MLcs.LGarXiv:2608.24426v12026
  37. Movement Pruning: Adaptive Sparsity by Fine-Tuning

    Victor Sanh, Thomas Wolf, Alexander M. Rush

    cs.CLcs.LGarXiv:2005.07683v22020
  38. Open-Set Recognition: a Good Closed-Set Classifier is All You Need?

    Sagar Vaze, Kai Han, Andrea Vedaldi +1

    cs.CVcs.LGarXiv:2110.06207v22021
  39. Neural Additive Models: Interpretable Machine Learning with Neural Nets

    Rishabh Agarwal, Levi Melnick, Nicholas Frosst +4

    cs.LGcs.AIstat.MLarXiv:2004.13912v22020
  40. Motion Planning Among Dynamic, Decision-Making Agents with Deep Reinforcement Learning

    Michael Everett, Yu Fan Chen, Jonathan P. How

    cs.ROcs.AIcs.LGarXiv:1805.01956v12018
  41. Learning to Optimize

    Ke Li, Jitendra Malik

    cs.LGcs.AImath.OCarXiv:1606.01885v12016
  42. Federated Learning over Wireless Fading Channels

    Mohammad Mohammadi Amiri, Deniz Gunduz

    cs.ITcs.DCcs.LGarXiv:1907.09769v22019
  43. XP-JEPA: Cross-Predictive Physics Grounding for Forecastable Latent Dynamics

    Kehan Wen, Ziming Li, Siyuan Luo +1

    cs.LGarXiv:2608.24044v12026
  44. GPT detectors are biased against non-native English writers

    Weixin Liang, Mert Yuksekgonul, Yining Mao +2

    cs.CLcs.AIcs.HCarXiv:2304.02819v32023
  45. Residual Gated Graph ConvNets

    Xavier Bresson, Thomas Laurent

    cs.LGstat.MLarXiv:1711.07553v22017
  46. Playing FPS Games with Deep Reinforcement Learning

    Guillaume Lample, Devendra Singh Chaplot

    cs.AIcs.LGarXiv:1609.05521v22016
  47. LLaVA-NeXT-Interleave: Tackling Multi-image, Video, and 3D in Large Multimodal Models

    Feng Li, Renrui Zhang, Hao Zhang +5

    cs.CVcs.CLcs.LGarXiv:2407.07895v22024
  48. On Mixup Training: Improved Calibration and Predictive Uncertainty for Deep Neural Networks

    Sunil Thulasidasan, Gopinath Chennupati, Jeff Bilmes +2

    stat.MLcs.LGarXiv:1905.11001v52019
  49. Why do tree-based models still outperform deep learning on tabular data?

    Léo Grinsztajn, Edouard Oyallon, Gaël Varoquaux

    cs.LGcs.AIstat.MEarXiv:2207.08815v12022
  50. Distral: Robust Multitask Reinforcement Learning

    Yee Whye Teh, Victor Bapst, Wojciech Marian Czarnecki +5

    cs.LGstat.MLarXiv:1707.04175v12017
  51. Evaluation of a Tree-based Pipeline Optimization Tool for Automating Data Science

    Randal S. Olson, Nathan Bartley, Ryan J. Urbanowicz +1

    cs.NEcs.AIcs.LGarXiv:1603.06212v12016
  52. Improving Cross-Problem Vehicle Routing with Locally Augmented Preferences and Representation Disentanglement

    Arthur Corrêa, Paulo Nascimento, Samuel Moniz

    cs.LGarXiv:2608.24859v12026
  53. Where Entropy Is Measured Matters: Policy Geometry in Bounded Continuous-Control PPO

    Yiyang He, Zhichun Zhou, Ziwei Wang +2

    cs.LGarXiv:2608.24488v12026
  54. Why ReLU networks yield high-confidence predictions far away from the training data and how to mitigate the problem

    Matthias Hein, Maksym Andriushchenko, Julian Bitterwolf

    cs.LGcs.CVstat.MLarXiv:1812.05720v22018
  55. FlashAttention-3: Fast and Accurate Attention with Asynchrony and Low-precision

    Jay Shah, Ganesh Bikshandi, Ying Zhang +3

    cs.LGcs.AIarXiv:2407.08608v22024
  56. ZeRO-Offload: Democratizing Billion-Scale Model Training

    Jie Ren, Samyam Rajbhandari, Reza Yazdani Aminabadi +5

    cs.DCcs.LGarXiv:2101.06840v12021
  57. Memory Is Not Always Needed: Characterizing Conditional Memory in Scientific Reasoning

    Zhen Bi, Xueshu Chen, Yan Wang +6

    cs.AIcs.CLcs.LGarXiv:2608.23982v12026
  58. Deep Reinforcement Learning for Intelligent Transportation Systems: A Survey

    Ammar Haydari, Yasin Yilmaz

    cs.LGcs.MAeess.SParXiv:2005.00935v12020
  59. K-Adapter: Infusing Knowledge into Pre-Trained Models with Adapters

    Ruize Wang, Duyu Tang, Nan Duan +6

    cs.CLcs.LGarXiv:2002.01808v52020
  60. Diffusion Forcing: Next-token Prediction Meets Full-Sequence Diffusion

    Boyuan Chen, Diego Marti Monso, Yilun Du +3

    cs.LGcs.CVcs.ROarXiv:2407.01392v42024