Machine Learning

Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

17,101 to 17,160 of 20,193

  1. #Exploration: A Study of Count-Based Exploration for Deep Reinforcement Learning

    Haoran Tang, Rein Houthooft, Davis Foote +6

    cs.AIcs.LGarXiv:1611.04717v32016
  2. Natural Language Processing (almost) from Scratch

    Ronan Collobert, Jason Weston, Leon Bottou +3

    cs.LGcs.CLarXiv:1103.0398v12011
    Summaries:한국어
  3. Medusa: Simple LLM Inference Acceleration Framework with Multiple Decoding Heads

    Tianle Cai, Yuhong Li, Zhengyang Geng +4

    cs.LGcs.CLarXiv:2401.10774v32024
  4. Post-LayerNorm Is Back: Stable, ExpressivE, and Deep

    Chen Chen, Lai Wei

    cs.LGcs.CLarXiv:2601.19895v22026
  5. Behavior Regularized Offline Reinforcement Learning

    Yifan Wu, George Tucker, Ofir Nachum

    cs.LGcs.AIstat.MLarXiv:1911.11361v12019
  6. StealthRL: Reinforcement Learning Paraphrase Attacks for Multi-Detector Evasion of AI-Text Detectors

    Suraj Ranganath, Atharv Ramesh

    cs.LGcs.AIcs.CRarXiv:2602.08934v22026
  7. Adaptive Graph Convolutional Neural Networks

    Ruoyu Li, Sheng Wang, Feiyun Zhu +1

    cs.LGstat.MLarXiv:1801.03226v12018
  8. Benchmarks Saturate When The Model Gets Smarter Than The Judge

    Marthe Ballon, Andres Algaba, Brecht Verbeken +1

    cs.AIcs.CLcs.LGarXiv:2601.19532v12026
  9. Time-Series Representation Learning via Temporal and Contextual Contrasting

    Emadeldeen Eldele, Mohamed Ragab, Zhenghua Chen +4

    cs.LGcs.AIarXiv:2106.14112v12021
  10. Continual GUI Agents

    Ziwei Liu, Borui Kang, Hangjie Yuan +4

    cs.LGcs.CVarXiv:2601.20732v42026
  11. Nature-Inspired Optimization Algorithms: Challenges and Open Problems

    Xin-She Yang

    cs.NEcs.LGmath.OCarXiv:2003.03776v12020
  12. Explainability in Graph Neural Networks: A Taxonomic Survey

    Hao Yuan, Haiyang Yu, Shurui Gui +1

    cs.LGcs.AIarXiv:2012.15445v32020
  13. Learning Robust Rewards with Adversarial Inverse Reinforcement Learning

    Justin Fu, Katie Luo, Sergey Levine

    cs.LGarXiv:1710.11248v22017
  14. Black-box Generation of Adversarial Text Sequences to Evade Deep Learning Classifiers

    Ji Gao, Jack Lanchantin, Mary Lou Soffa +1

    cs.CLcs.CRcs.IRarXiv:1801.04354v52018
  15. One-Step Evolution for Long-Time Extrapolation: An Error-Bound-Informed and Prior-Guided Neural Residual Framework for Autonomous PDEs

    Maqun Zhang, Feng Gao, Wankun Chen +3

    cs.AIcs.LGarXiv:2608.22026v12026
  16. Exposing the Systematic Vulnerability of Open-Weight Models to Prefill Attacks

    Lukas Struppek, Adam Gleave, Kellin Pelrine

    cs.CRcs.AIcs.CLarXiv:2602.14689v12026
  17. Deep Learning for Sensor-based Human Activity Recognition: Overview, Challenges and Opportunities

    Kaixuan Chen, Dalin Zhang, Lina Yao +3

    cs.HCcs.LGarXiv:2001.07416v22020
  18. AI4SLT: Empirical Processes in Lean 4 for Formal Statistical Learning Theory

    Yuanhe Zhang, Jason D. Lee, Fanghui Liu

    cs.LGcs.CLmath.STarXiv:2602.02285v22026
  19. FILIP: Fine-grained Interactive Language-Image Pre-Training

    Lewei Yao, Runhui Huang, Lu Hou +7

    cs.CVcs.LGarXiv:2111.07783v12021
  20. Conditional Neural Processes

    Marta Garnelo, Dan Rosenbaum, Chris J. Maddison +6

    cs.LGstat.MLarXiv:1807.01613v12018
  21. On Randomness in Agentic Evals

    Bjarni Haukur Bjarnason, André Silva, Martin Monperrus

    cs.LGcs.AIcs.SEarXiv:2602.07150v32026
  22. H$_2$O: Heavy-Hitter Oracle for Efficient Generative Inference of Large Language Models

    Zhenyu Zhang, Ying Sheng, Tianyi Zhou +9

    cs.LGarXiv:2306.14048v32023
  23. Explainable Machine Learning for Scientific Insights and Discoveries

    Ribana Roscher, Bastian Bohn, Marco F. Duarte +1

    cs.LGstat.MLarXiv:1905.08883v32019
  24. Learning a Generative Meta-Model of LLM Activations

    Grace Luo, Jiahai Feng, Trevor Darrell +2

    cs.LGcs.AIcs.CLarXiv:2602.06964v12026
  25. Personalized Cross-Silo Federated Learning on Non-IID Data

    Yutao Huang, Lingyang Chu, Zirui Zhou +4

    cs.LGcs.DCstat.MLarXiv:2007.03797v52020
  26. Masked Feature Prediction for Self-Supervised Visual Pre-Training

    Chen Wei, Haoqi Fan, Saining Xie +3

    cs.CVcs.LGarXiv:2112.09133v22021
  27. Steer2Adapt: Dynamically Composing Steering Vectors Elicits Efficient Adaptation of LLMs

    Pengrui Han, Xueqiang Xu, Keyang Xuan +12

    cs.AIcs.CLcs.LGarXiv:2602.07276v12026
  28. FrugalGPT: How to Use Large Language Models While Reducing Cost and Improving Performance

    Lingjiao Chen, Matei Zaharia, James Zou

    cs.LGcs.AIcs.CLarXiv:2305.05176v12023
  29. Deep Learning for Classical Japanese Literature

    Tarin Clanuwat, Mikel Bober-Irizar, Asanobu Kitamoto +3

    cs.CVcs.LGstat.MLarXiv:1812.01718v12018
  30. Are LLM Decisions Faithful to Verbal Confidence?

    Jiawei Wang, Yanfei Zhou, Siddartha Devic +1

    cs.LGcs.CLarXiv:2601.07767v12026
  31. Domain Adaptation: Learning Bounds and Algorithms

    Yishay Mansour, Mehryar Mohri, Afshin Rostamizadeh

    cs.LGcs.AIarXiv:0902.3430v32009
  32. Attend-and-Excite: Attention-Based Semantic Guidance for Text-to-Image Diffusion Models

    Hila Chefer, Yuval Alaluf, Yael Vinker +2

    cs.CVcs.CLcs.GRarXiv:2301.13826v22023
  33. Whose Opinions Do Language Models Reflect?

    Shibani Santurkar, Esin Durmus, Faisal Ladhak +3

    cs.CLcs.AIcs.CYarXiv:2303.17548v12023
  34. Hints, Critics, and Teachers: Prior Injection for Sparse-Reward RL in Vision-Language Math Reasoning

    Qiqian Fu

    cs.AIcs.LGarXiv:2608.21811v12026
  35. When Gaussian Process Meets Big Data: A Review of Scalable GPs

    Haitao Liu, Yew-Soon Ong, Xiaobo Shen +1

    stat.MLcs.LGarXiv:1807.01065v22018
  36. Semantic Uncertainty: Linguistic Invariances for Uncertainty Estimation in Natural Language Generation

    Lorenz Kuhn, Yarin Gal, Sebastian Farquhar

    cs.CLcs.AIcs.LGarXiv:2302.09664v32023
  37. A Comprehensive Survey of Neural Architecture Search: Challenges and Solutions

    Pengzhen Ren, Yun Xiao, Xiaojun Chang +4

    cs.LGstat.MLarXiv:2006.02903v32020
  38. VisAdj: Learning Adjacency Matrices from Node-Link Images

    Jiahao Xie, Guangmo Tong

    cs.AIcs.CVcs.LGarXiv:2608.21825v12026
  39. MOOSE-Star: Unlocking Tractable Training for Scientific Discovery by Breaking the Complexity Barrier

    Zonglin Yang, Lidong Bing

    cs.LGcs.CEcs.CLarXiv:2603.03756v42026
  40. ChauffeurNet: Learning to Drive by Imitating the Best and Synthesizing the Worst

    Mayank Bansal, Alex Krizhevsky, Abhijit Ogale

    cs.ROcs.CVcs.LGarXiv:1812.03079v12018
  41. HIRA: A Human-in-the-Loop Retrieval-Augmented Cascade for Document Classification in Regulated Industries

    Shangxuan Tian, Yanhui Chen, Carlos Queiroz

    cs.AIcs.CVcs.IRarXiv:2608.21792v12026
  42. QA-GNN: Reasoning with Language Models and Knowledge Graphs for Question Answering

    Michihiro Yasunaga, Hongyu Ren, Antoine Bosselut +2

    cs.CLcs.LGarXiv:2104.06378v52021
  43. GLTR: Statistical Detection and Visualization of Generated Text

    Sebastian Gehrmann, Hendrik Strobelt, Alexander M. Rush

    cs.CLcs.AIcs.HCarXiv:1906.04043v12019
  44. ReflexiCoder: Teaching Large Language Models to Self-Reflect on Generated Code and Self-Correct It via Reinforcement Learning

    Juyong Jiang, Jiasi Shen, Sunghun Kim +3

    cs.CLcs.LGcs.SEarXiv:2603.05863v22026
  45. k-Nearest Neighbour Classifiers: 2nd Edition (with Python examples)

    Padraig Cunningham, Sarah Jane Delany

    cs.LGstat.MLarXiv:2004.04523v22020
  46. Zip-NeRF: Anti-Aliased Grid-Based Neural Radiance Fields

    Jonathan T. Barron, Ben Mildenhall, Dor Verbin +2

    cs.CVcs.GRcs.LGarXiv:2304.06706v32023
  47. ERASER: A Benchmark to Evaluate Rationalized NLP Models

    Jay DeYoung, Sarthak Jain, Nazneen Fatema Rajani +4

    cs.CLcs.AIcs.LGarXiv:1911.03429v22019
  48. PaperSearchQA: Learning to Search and Reason over Scientific Papers with RLVR

    James Burgess, Jan N. Hansen, Duo Peng +5

    cs.LGcs.AIcs.CLarXiv:2601.18207v12026
  49. CondConv: Conditionally Parameterized Convolutions for Efficient Inference

    Brandon Yang, Gabriel Bender, Quoc V. Le +1

    cs.CVcs.AIcs.LGarXiv:1904.04971v32019
  50. SCINet: Time Series Modeling and Forecasting with Sample Convolution and Interaction

    Minhao Liu, Ailing Zeng, Muxi Chen +4

    cs.LGcs.AIarXiv:2106.09305v32021
  51. A Reproducible, License-Aware Distillation Recipe for CPUDeployable Safety Classification

    Edson Rodrigues da Cruz Filho, Paulo Ricardo Ferreira Neves, Paulo Henrique Eleuterio Falsetti +7

    cs.AIcs.LGarXiv:2608.21570v12026
  52. KAT-Coder-V2 Technical Report

    Fengxiang Li, Han Zhang, Haoyang Huang +43

    cs.CLcs.LGarXiv:2603.27703v12026
  53. Dynamic Key-Value Memory Networks for Knowledge Tracing

    Jiani Zhang, Xingjian Shi, Irwin King +1

    cs.AIcs.LGarXiv:1611.08108v22016
  54. Manipulating Machine Learning: Poisoning Attacks and Countermeasures for Regression Learning

    Matthew Jagielski, Alina Oprea, Battista Biggio +3

    cs.CRcs.GTcs.LGarXiv:1804.00308v32018
  55. ACNet: Strengthening the Kernel Skeletons for Powerful CNN via Asymmetric Convolution Blocks

    Xiaohan Ding, Yuchen Guo, Guiguang Ding +1

    cs.CVcs.LGcs.NEarXiv:1908.03930v32019
  56. Encoding Sentences with Graph Convolutional Networks for Semantic Role Labeling

    Diego Marcheggiani, Ivan Titov

    cs.CLcs.LGarXiv:1703.04826v42017
  57. Test-Time Training with KV Binding Is Secretly Linear Attention

    Junchen Liu, Sven Elflein, Or Litany +2

    cs.LGcs.AIcs.CVarXiv:2602.21204v42026
  58. Aletheia tackles FirstProof autonomously

    Tony Feng, Junehyuk Jung, Sang-hyun Kim +14

    cs.AIcs.CLcs.LGarXiv:2602.21201v32026
  59. Not Just a Black Box: Learning Important Features Through Propagating Activation Differences

    Avanti Shrikumar, Peyton Greenside, Anna Shcherbina +1

    cs.LGcs.CVcs.NEarXiv:1605.01713v32016
  60. Speaker Recognition from Raw Waveform with SincNet

    Mirco Ravanelli, Yoshua Bengio

    eess.AScs.LGcs.SDarXiv:1808.00158v32018