Machine Learning
Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
17,101 to 17,160 of 20,193
#Exploration: A Study of Count-Based Exploration for Deep Reinforcement Learning
Haoran Tang, Rein Houthooft, Davis Foote +6
cs.AIcs.LGarXiv:1611.04717v32016Natural Language Processing (almost) from Scratch
Ronan Collobert, Jason Weston, Leon Bottou +3
cs.LGcs.CLarXiv:1103.0398v12011Summaries:한국어Medusa: Simple LLM Inference Acceleration Framework with Multiple Decoding Heads
Tianle Cai, Yuhong Li, Zhengyang Geng +4
cs.LGcs.CLarXiv:2401.10774v32024Post-LayerNorm Is Back: Stable, ExpressivE, and Deep
Chen Chen, Lai Wei
cs.LGcs.CLarXiv:2601.19895v22026Behavior Regularized Offline Reinforcement Learning
Yifan Wu, George Tucker, Ofir Nachum
cs.LGcs.AIstat.MLarXiv:1911.11361v12019StealthRL: Reinforcement Learning Paraphrase Attacks for Multi-Detector Evasion of AI-Text Detectors
Suraj Ranganath, Atharv Ramesh
cs.LGcs.AIcs.CRarXiv:2602.08934v22026Adaptive Graph Convolutional Neural Networks
Ruoyu Li, Sheng Wang, Feiyun Zhu +1
cs.LGstat.MLarXiv:1801.03226v12018Benchmarks Saturate When The Model Gets Smarter Than The Judge
Marthe Ballon, Andres Algaba, Brecht Verbeken +1
cs.AIcs.CLcs.LGarXiv:2601.19532v12026Time-Series Representation Learning via Temporal and Contextual Contrasting
Emadeldeen Eldele, Mohamed Ragab, Zhenghua Chen +4
cs.LGcs.AIarXiv:2106.14112v12021Continual GUI Agents
Ziwei Liu, Borui Kang, Hangjie Yuan +4
cs.LGcs.CVarXiv:2601.20732v42026Nature-Inspired Optimization Algorithms: Challenges and Open Problems
Xin-She Yang
cs.NEcs.LGmath.OCarXiv:2003.03776v12020Explainability in Graph Neural Networks: A Taxonomic Survey
Hao Yuan, Haiyang Yu, Shurui Gui +1
cs.LGcs.AIarXiv:2012.15445v32020Learning Robust Rewards with Adversarial Inverse Reinforcement Learning
Justin Fu, Katie Luo, Sergey Levine
cs.LGarXiv:1710.11248v22017Black-box Generation of Adversarial Text Sequences to Evade Deep Learning Classifiers
Ji Gao, Jack Lanchantin, Mary Lou Soffa +1
cs.CLcs.CRcs.IRarXiv:1801.04354v52018One-Step Evolution for Long-Time Extrapolation: An Error-Bound-Informed and Prior-Guided Neural Residual Framework for Autonomous PDEs
Maqun Zhang, Feng Gao, Wankun Chen +3
cs.AIcs.LGarXiv:2608.22026v12026Exposing the Systematic Vulnerability of Open-Weight Models to Prefill Attacks
Lukas Struppek, Adam Gleave, Kellin Pelrine
cs.CRcs.AIcs.CLarXiv:2602.14689v12026Deep Learning for Sensor-based Human Activity Recognition: Overview, Challenges and Opportunities
Kaixuan Chen, Dalin Zhang, Lina Yao +3
cs.HCcs.LGarXiv:2001.07416v22020AI4SLT: Empirical Processes in Lean 4 for Formal Statistical Learning Theory
Yuanhe Zhang, Jason D. Lee, Fanghui Liu
cs.LGcs.CLmath.STarXiv:2602.02285v22026FILIP: Fine-grained Interactive Language-Image Pre-Training
Lewei Yao, Runhui Huang, Lu Hou +7
cs.CVcs.LGarXiv:2111.07783v12021Conditional Neural Processes
Marta Garnelo, Dan Rosenbaum, Chris J. Maddison +6
cs.LGstat.MLarXiv:1807.01613v12018On Randomness in Agentic Evals
Bjarni Haukur Bjarnason, André Silva, Martin Monperrus
cs.LGcs.AIcs.SEarXiv:2602.07150v32026H$_2$O: Heavy-Hitter Oracle for Efficient Generative Inference of Large Language Models
Zhenyu Zhang, Ying Sheng, Tianyi Zhou +9
cs.LGarXiv:2306.14048v32023Explainable Machine Learning for Scientific Insights and Discoveries
Ribana Roscher, Bastian Bohn, Marco F. Duarte +1
cs.LGstat.MLarXiv:1905.08883v32019Learning a Generative Meta-Model of LLM Activations
Grace Luo, Jiahai Feng, Trevor Darrell +2
cs.LGcs.AIcs.CLarXiv:2602.06964v12026Personalized Cross-Silo Federated Learning on Non-IID Data
Yutao Huang, Lingyang Chu, Zirui Zhou +4
cs.LGcs.DCstat.MLarXiv:2007.03797v52020Masked Feature Prediction for Self-Supervised Visual Pre-Training
Chen Wei, Haoqi Fan, Saining Xie +3
cs.CVcs.LGarXiv:2112.09133v22021Steer2Adapt: Dynamically Composing Steering Vectors Elicits Efficient Adaptation of LLMs
Pengrui Han, Xueqiang Xu, Keyang Xuan +12
cs.AIcs.CLcs.LGarXiv:2602.07276v12026FrugalGPT: How to Use Large Language Models While Reducing Cost and Improving Performance
Lingjiao Chen, Matei Zaharia, James Zou
cs.LGcs.AIcs.CLarXiv:2305.05176v12023Deep Learning for Classical Japanese Literature
Tarin Clanuwat, Mikel Bober-Irizar, Asanobu Kitamoto +3
cs.CVcs.LGstat.MLarXiv:1812.01718v12018Are LLM Decisions Faithful to Verbal Confidence?
Jiawei Wang, Yanfei Zhou, Siddartha Devic +1
cs.LGcs.CLarXiv:2601.07767v12026Domain Adaptation: Learning Bounds and Algorithms
Yishay Mansour, Mehryar Mohri, Afshin Rostamizadeh
cs.LGcs.AIarXiv:0902.3430v32009Attend-and-Excite: Attention-Based Semantic Guidance for Text-to-Image Diffusion Models
Hila Chefer, Yuval Alaluf, Yael Vinker +2
cs.CVcs.CLcs.GRarXiv:2301.13826v22023Whose Opinions Do Language Models Reflect?
Shibani Santurkar, Esin Durmus, Faisal Ladhak +3
cs.CLcs.AIcs.CYarXiv:2303.17548v12023Hints, Critics, and Teachers: Prior Injection for Sparse-Reward RL in Vision-Language Math Reasoning
Qiqian Fu
cs.AIcs.LGarXiv:2608.21811v12026When Gaussian Process Meets Big Data: A Review of Scalable GPs
Haitao Liu, Yew-Soon Ong, Xiaobo Shen +1
stat.MLcs.LGarXiv:1807.01065v22018Semantic Uncertainty: Linguistic Invariances for Uncertainty Estimation in Natural Language Generation
Lorenz Kuhn, Yarin Gal, Sebastian Farquhar
cs.CLcs.AIcs.LGarXiv:2302.09664v32023A Comprehensive Survey of Neural Architecture Search: Challenges and Solutions
Pengzhen Ren, Yun Xiao, Xiaojun Chang +4
cs.LGstat.MLarXiv:2006.02903v32020VisAdj: Learning Adjacency Matrices from Node-Link Images
Jiahao Xie, Guangmo Tong
cs.AIcs.CVcs.LGarXiv:2608.21825v12026MOOSE-Star: Unlocking Tractable Training for Scientific Discovery by Breaking the Complexity Barrier
Zonglin Yang, Lidong Bing
cs.LGcs.CEcs.CLarXiv:2603.03756v42026ChauffeurNet: Learning to Drive by Imitating the Best and Synthesizing the Worst
Mayank Bansal, Alex Krizhevsky, Abhijit Ogale
cs.ROcs.CVcs.LGarXiv:1812.03079v12018HIRA: A Human-in-the-Loop Retrieval-Augmented Cascade for Document Classification in Regulated Industries
Shangxuan Tian, Yanhui Chen, Carlos Queiroz
cs.AIcs.CVcs.IRarXiv:2608.21792v12026QA-GNN: Reasoning with Language Models and Knowledge Graphs for Question Answering
Michihiro Yasunaga, Hongyu Ren, Antoine Bosselut +2
cs.CLcs.LGarXiv:2104.06378v52021GLTR: Statistical Detection and Visualization of Generated Text
Sebastian Gehrmann, Hendrik Strobelt, Alexander M. Rush
cs.CLcs.AIcs.HCarXiv:1906.04043v12019ReflexiCoder: Teaching Large Language Models to Self-Reflect on Generated Code and Self-Correct It via Reinforcement Learning
Juyong Jiang, Jiasi Shen, Sunghun Kim +3
cs.CLcs.LGcs.SEarXiv:2603.05863v22026k-Nearest Neighbour Classifiers: 2nd Edition (with Python examples)
Padraig Cunningham, Sarah Jane Delany
cs.LGstat.MLarXiv:2004.04523v22020Zip-NeRF: Anti-Aliased Grid-Based Neural Radiance Fields
Jonathan T. Barron, Ben Mildenhall, Dor Verbin +2
cs.CVcs.GRcs.LGarXiv:2304.06706v32023ERASER: A Benchmark to Evaluate Rationalized NLP Models
Jay DeYoung, Sarthak Jain, Nazneen Fatema Rajani +4
cs.CLcs.AIcs.LGarXiv:1911.03429v22019PaperSearchQA: Learning to Search and Reason over Scientific Papers with RLVR
James Burgess, Jan N. Hansen, Duo Peng +5
cs.LGcs.AIcs.CLarXiv:2601.18207v12026CondConv: Conditionally Parameterized Convolutions for Efficient Inference
Brandon Yang, Gabriel Bender, Quoc V. Le +1
cs.CVcs.AIcs.LGarXiv:1904.04971v32019SCINet: Time Series Modeling and Forecasting with Sample Convolution and Interaction
Minhao Liu, Ailing Zeng, Muxi Chen +4
cs.LGcs.AIarXiv:2106.09305v32021A Reproducible, License-Aware Distillation Recipe for CPUDeployable Safety Classification
Edson Rodrigues da Cruz Filho, Paulo Ricardo Ferreira Neves, Paulo Henrique Eleuterio Falsetti +7
cs.AIcs.LGarXiv:2608.21570v12026KAT-Coder-V2 Technical Report
Fengxiang Li, Han Zhang, Haoyang Huang +43
cs.CLcs.LGarXiv:2603.27703v12026Dynamic Key-Value Memory Networks for Knowledge Tracing
Jiani Zhang, Xingjian Shi, Irwin King +1
cs.AIcs.LGarXiv:1611.08108v22016Manipulating Machine Learning: Poisoning Attacks and Countermeasures for Regression Learning
Matthew Jagielski, Alina Oprea, Battista Biggio +3
cs.CRcs.GTcs.LGarXiv:1804.00308v32018ACNet: Strengthening the Kernel Skeletons for Powerful CNN via Asymmetric Convolution Blocks
Xiaohan Ding, Yuchen Guo, Guiguang Ding +1
cs.CVcs.LGcs.NEarXiv:1908.03930v32019Encoding Sentences with Graph Convolutional Networks for Semantic Role Labeling
Diego Marcheggiani, Ivan Titov
cs.CLcs.LGarXiv:1703.04826v42017Test-Time Training with KV Binding Is Secretly Linear Attention
Junchen Liu, Sven Elflein, Or Litany +2
cs.LGcs.AIcs.CVarXiv:2602.21204v42026Aletheia tackles FirstProof autonomously
Tony Feng, Junehyuk Jung, Sang-hyun Kim +14
cs.AIcs.CLcs.LGarXiv:2602.21201v32026Not Just a Black Box: Learning Important Features Through Propagating Activation Differences
Avanti Shrikumar, Peyton Greenside, Anna Shcherbina +1
cs.LGcs.CVcs.NEarXiv:1605.01713v32016Speaker Recognition from Raw Waveform with SincNet
Mirco Ravanelli, Yoshua Bengio
eess.AScs.LGcs.SDarXiv:1808.00158v32018