Machine Learning
Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
15,841 to 15,900 of 20,199
The AI Scientist: Towards Fully Automated Open-Ended Scientific Discovery
Chris Lu, Cong Lu, Robert Tjarko Lange +3
cs.AIcs.CLcs.LGarXiv:2408.06292v32024Summaries:한국어Channel Estimation for RIS-Empowered Multi-User MISO Wireless Communications
Li Wei, Chongwen Huang, George C. Alexandropoulos +3
cs.ITcs.LGeess.SParXiv:2008.01459v22020Can AI-Generated Text be Reliably Detected?
Vinu Sankar Sadasivan, Aounon Kumar, Sriram Balasubramanian +2
cs.CLcs.AIcs.LGarXiv:2303.11156v42023The Loss Floor of Denoising Score Matching: Fisher Geometry from Schrödinger Bridges
Avinash Raju, Kai Zhang
cs.LGcond-mat.stat-mecharXiv:2608.23916v12026Towards a Guideline for Evaluation Metrics in Medical Image Segmentation
Dominik Müller, Iñaki Soto-Rey, Frank Kramer
eess.IVcs.CVcs.LGarXiv:2202.05273v12022Revisiting the Arcade Learning Environment: Evaluation Protocols and Open Problems for General Agents
Marlos C. Machado, Marc G. Bellemare, Erik Talvitie +3
cs.LGarXiv:1709.06009v22017End-to-end Learning of Action Detection from Frame Glimpses in Videos
Serena Yeung, Olga Russakovsky, Greg Mori +1
cs.CVcs.LGarXiv:1511.06984v22015DeepCoder: Learning to Write Programs
Matej Balog, Alexander L. Gaunt, Marc Brockschmidt +2
cs.LGarXiv:1611.01989v22016Dipole: Diagnosis Prediction in Healthcare via Attention-based Bidirectional Recurrent Neural Networks
Fenglong Ma, Radha Chitta, Jing Zhou +3
cs.LGarXiv:1706.05764v12017GATNextHop: A GAT for Shortest Path Routing with Cross-Topology Generalization
Chia-Hong Chou, Katerina Potika
cs.LGarXiv:2608.23917v12026A robust self-learning method for fully unsupervised cross-lingual mappings of word embeddings
Mikel Artetxe, Gorka Labaka, Eneko Agirre
cs.CLcs.AIcs.LGarXiv:1805.06297v22018Generative and Discriminative Voxel Modeling with Convolutional Neural Networks
Andrew Brock, Theodore Lim, J. M. Ritchie +1
cs.CVcs.HCcs.LGarXiv:1608.04236v22016Katyusha: The First Direct Acceleration of Stochastic Gradient Methods
Zeyuan Allen-Zhu
math.OCcs.DScs.LGarXiv:1603.05953v62016CODA-Prompt: COntinual Decomposed Attention-based Prompting for Rehearsal-Free Continual Learning
James Seale Smith, Leonid Karlinsky, Vyshnavi Gutta +6
cs.CVcs.AIcs.LGarXiv:2211.13218v22022Edge Artificial Intelligence for 6G: Vision, Enabling Technologies, and Applications
Khaled B. Letaief, Yuanming Shi, Jianmin Lu +1
cs.ITcs.LGcs.NIarXiv:2111.12444v12021Scalable agent alignment via reward modeling: a research direction
Jan Leike, David Krueger, Tom Everitt +3
cs.LGcs.AIcs.NEarXiv:1811.07871v12018Knowledge Transfer via Distillation of Activation Boundaries Formed by Hidden Neurons
Byeongho Heo, Minsik Lee, Sangdoo Yun +1
cs.LGcs.CVstat.MLarXiv:1811.03233v22018Fairness-Aware Mixture-of-Experts via Subgroup Reweighting and Gate Regularization
Sunhee Hwang
cs.LGcs.AIarXiv:2608.22820v12026QuaRot: Outlier-Free 4-Bit Inference in Rotated LLMs
Saleh Ashkboos, Amirkeivan Mohtashami, Maximilian L. Croci +6
cs.LGarXiv:2404.00456v22024A Convolutional Attention Network for Extreme Summarization of Source Code
Miltiadis Allamanis, Hao Peng, Charles Sutton
cs.LGcs.CLcs.SEarXiv:1602.03001v22016Convergence Rates of Inexact Proximal-Gradient Methods for Convex Optimization
Mark Schmidt, Nicolas Le Roux, Francis Bach
cs.LGmath.OCarXiv:1109.2415v22011Towards Open World Object Detection
K J Joseph, Salman Khan, Fahad Shahbaz Khan +1
cs.CVcs.AIcs.LGarXiv:2103.02603v22021Reliable Fidelity and Diversity Metrics for Generative Models
Muhammad Ferjad Naeem, Seong Joon Oh, Youngjung Uh +2
cs.CVcs.LGstat.MLarXiv:2002.09797v22020Lagrangian Neural Networks
Miles Cranmer, Sam Greydanus, Stephan Hoyer +3
cs.LGmath.DSphysics.comp-pharXiv:2003.04630v22020Sequence-to-Sequence Learning as Beam-Search Optimization
Sam Wiseman, Alexander M. Rush
cs.CLcs.LGcs.NEarXiv:1606.02960v22016When Similarity Is Interaction-Driven: Quantum Kernels for Regime-Sensitive Learning
Hanqiu Peng, Jianlong Lu, Ying Chen
quant-phcs.LGarXiv:2608.24631v12026Deep Learning-based Computational Pathology Predicts Origins for Cancers of Unknown Primary
Ming Y. Lu, Melissa Zhao, Maha Shady +4
q-bio.TOcs.LGq-bio.QMarXiv:2006.13932v22020DeepInf: Social Influence Prediction with Deep Learning
Jiezhong Qiu, Jian Tang, Hao Ma +3
cs.SIcs.LGarXiv:1807.05560v12018DSOD: Learning Deeply Supervised Object Detectors from Scratch
Zhiqiang Shen, Zhuang Liu, Jianguo Li +3
cs.CVcs.LGarXiv:1708.01241v22017Zephyr: Direct Distillation of LM Alignment
Lewis Tunstall, Edward Beeching, Nathan Lambert +11
cs.LGcs.CLarXiv:2310.16944v12023DeeperForensics-1.0: A Large-Scale Dataset for Real-World Face Forgery Detection
Liming Jiang, Ren Li, Wayne Wu +2
cs.CVcs.LGarXiv:2001.03024v22020Ab-Initio Solution of the Many-Electron Schrödinger Equation with Deep Neural Networks
David Pfau, James S. Spencer, Alexander G. de G. Matthews +1
physics.chem-phcs.LGphysics.comp-pharXiv:1909.02487v32019COVID-CAPS: A Capsule Network-based Framework for Identification of COVID-19 cases from X-ray Images
Parnian Afshar, Shahin Heidarian, Farnoosh Naderkhani +3
cs.CVcs.LGeess.IVarXiv:2004.02696v22020The UEA multivariate time series classification archive, 2018
Anthony Bagnall, Hoang Anh Dau, Jason Lines +5
cs.LGstat.MLarXiv:1811.00075v12018Deep Semantic Ranking Based Hashing for Multi-Label Image Retrieval
Fang Zhao, Yongzhen Huang, Liang Wang +1
cs.CVcs.LGarXiv:1501.06272v22015Sequential operator learning under dependent data
Rafael Oliveira
stat.MLcs.LGarXiv:2608.24426v12026Movement Pruning: Adaptive Sparsity by Fine-Tuning
Victor Sanh, Thomas Wolf, Alexander M. Rush
cs.CLcs.LGarXiv:2005.07683v22020Open-Set Recognition: a Good Closed-Set Classifier is All You Need?
Sagar Vaze, Kai Han, Andrea Vedaldi +1
cs.CVcs.LGarXiv:2110.06207v22021Neural Additive Models: Interpretable Machine Learning with Neural Nets
Rishabh Agarwal, Levi Melnick, Nicholas Frosst +4
cs.LGcs.AIstat.MLarXiv:2004.13912v22020Motion Planning Among Dynamic, Decision-Making Agents with Deep Reinforcement Learning
Michael Everett, Yu Fan Chen, Jonathan P. How
cs.ROcs.AIcs.LGarXiv:1805.01956v12018Learning to Optimize
Ke Li, Jitendra Malik
cs.LGcs.AImath.OCarXiv:1606.01885v12016Federated Learning over Wireless Fading Channels
Mohammad Mohammadi Amiri, Deniz Gunduz
cs.ITcs.DCcs.LGarXiv:1907.09769v22019XP-JEPA: Cross-Predictive Physics Grounding for Forecastable Latent Dynamics
Kehan Wen, Ziming Li, Siyuan Luo +1
cs.LGarXiv:2608.24044v12026GPT detectors are biased against non-native English writers
Weixin Liang, Mert Yuksekgonul, Yining Mao +2
cs.CLcs.AIcs.HCarXiv:2304.02819v32023Residual Gated Graph ConvNets
Xavier Bresson, Thomas Laurent
cs.LGstat.MLarXiv:1711.07553v22017Playing FPS Games with Deep Reinforcement Learning
Guillaume Lample, Devendra Singh Chaplot
cs.AIcs.LGarXiv:1609.05521v22016LLaVA-NeXT-Interleave: Tackling Multi-image, Video, and 3D in Large Multimodal Models
Feng Li, Renrui Zhang, Hao Zhang +5
cs.CVcs.CLcs.LGarXiv:2407.07895v22024On Mixup Training: Improved Calibration and Predictive Uncertainty for Deep Neural Networks
Sunil Thulasidasan, Gopinath Chennupati, Jeff Bilmes +2
stat.MLcs.LGarXiv:1905.11001v52019Why do tree-based models still outperform deep learning on tabular data?
Léo Grinsztajn, Edouard Oyallon, Gaël Varoquaux
cs.LGcs.AIstat.MEarXiv:2207.08815v12022Distral: Robust Multitask Reinforcement Learning
Yee Whye Teh, Victor Bapst, Wojciech Marian Czarnecki +5
cs.LGstat.MLarXiv:1707.04175v12017Evaluation of a Tree-based Pipeline Optimization Tool for Automating Data Science
Randal S. Olson, Nathan Bartley, Ryan J. Urbanowicz +1
cs.NEcs.AIcs.LGarXiv:1603.06212v12016Improving Cross-Problem Vehicle Routing with Locally Augmented Preferences and Representation Disentanglement
Arthur Corrêa, Paulo Nascimento, Samuel Moniz
cs.LGarXiv:2608.24859v12026Where Entropy Is Measured Matters: Policy Geometry in Bounded Continuous-Control PPO
Yiyang He, Zhichun Zhou, Ziwei Wang +2
cs.LGarXiv:2608.24488v12026Why ReLU networks yield high-confidence predictions far away from the training data and how to mitigate the problem
Matthias Hein, Maksym Andriushchenko, Julian Bitterwolf
cs.LGcs.CVstat.MLarXiv:1812.05720v22018FlashAttention-3: Fast and Accurate Attention with Asynchrony and Low-precision
Jay Shah, Ganesh Bikshandi, Ying Zhang +3
cs.LGcs.AIarXiv:2407.08608v22024ZeRO-Offload: Democratizing Billion-Scale Model Training
Jie Ren, Samyam Rajbhandari, Reza Yazdani Aminabadi +5
cs.DCcs.LGarXiv:2101.06840v12021Memory Is Not Always Needed: Characterizing Conditional Memory in Scientific Reasoning
Zhen Bi, Xueshu Chen, Yan Wang +6
cs.AIcs.CLcs.LGarXiv:2608.23982v12026Deep Reinforcement Learning for Intelligent Transportation Systems: A Survey
Ammar Haydari, Yasin Yilmaz
cs.LGcs.MAeess.SParXiv:2005.00935v12020K-Adapter: Infusing Knowledge into Pre-Trained Models with Adapters
Ruize Wang, Duyu Tang, Nan Duan +6
cs.CLcs.LGarXiv:2002.01808v52020Diffusion Forcing: Next-token Prediction Meets Full-Sequence Diffusion
Boyuan Chen, Diego Marti Monso, Yilun Du +3
cs.LGcs.CVcs.ROarXiv:2407.01392v42024