Machine Learning

Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

5,041 to 5,100 of 20,175

  1. Forecasting directional movements of stock prices for intraday trading using LSTM and random forests

    Pushpendu Ghosh, Ariel Neufeld, Jajati Keshari Sahoo

    cs.LGq-fin.STstat.MLarXiv:2004.10178v22020
  2. Polis: 3D Self-Supervision at City Scale

    Alexander Rusnak, Sophia Kovalenko, Jingru Wang +3

    cs.CVcs.AIcs.LGarXiv:2608.29426v12026
  3. Optimizing the Optimizer for Physics-Informed Neural Networks and Kolmogorov-Arnold Networks

    Elham Kiyani, Khemraj Shukla, Jorge F. Urbán +2

    cs.LGcs.AImath.OCarXiv:2501.16371v62025
  4. UBnormal: New Benchmark for Supervised Open-Set Video Anomaly Detection

    Andra Acsintoae, Andrei Florescu, Mariana-Iuliana Georgescu +5

    cs.CVcs.LGarXiv:2111.08644v32021
  5. On Symmetric and Asymmetric LSHs for Inner Product Search

    Behnam Neyshabur, Nathan Srebro

    stat.MLcs.DScs.IRarXiv:1410.5518v32014
  6. Circulant Binary Embedding

    Felix X. Yu, Sanjiv Kumar, Yunchao Gong +1

    stat.MLcs.LGarXiv:1405.3162v12014
  7. A Finite Time Analysis of Two Time-Scale Actor Critic Methods

    Yue Wu, Weitong Zhang, Pan Xu +1

    cs.LGmath.OCstat.MLarXiv:2005.01350v32020
  8. The Impact of Feature Scaling In Machine Learning: Effects on Regression and Classification Tasks

    João Manoel Herrera Pinheiro, Suzana Vilas Boas de Oliveira, Thiago Henrique Segreto Silva +5

    cs.LGstat.MLarXiv:2506.08274v52025
  9. Stabilizing Deep Q-Learning with ConvNets and Vision Transformers under Data Augmentation

    Nicklas Hansen, Hao Su, Xiaolong Wang

    cs.LGcs.CVcs.ROarXiv:2107.00644v22021
  10. Hierarchical Planning with Latent World Models

    Wancong Zhang, Basile Terver, Artem Zholus +8

    cs.LGarXiv:2604.03208v22026
  11. nPINNs: nonlocal Physics-Informed Neural Networks for a parametrized nonlocal universal Laplacian operator. Algorithms and Applications

    Guofei Pang, Marta D'Elia, Michael Parks +1

    math.APcs.LGmath.OCarXiv:2004.04276v12020
  12. DeepSeek vs. ChatGPT vs. Claude: A Comparative Study for Scientific Computing and Scientific Machine Learning Tasks

    Qile Jiang, Zhiwei Gao, George Em Karniadakis

    cs.LGcs.AIarXiv:2502.17764v22025
  13. Spurious Forgetting in Continual Learning of Language Models

    Junhao Zheng, Xidi Cai, Shengjie Qiu +1

    cs.LGarXiv:2501.13453v12025
  14. Graph2Seq: Graph to Sequence Learning with Attention-based Neural Networks

    Kun Xu, Lingfei Wu, Zhiguo Wang +3

    cs.AIcs.CLcs.LGarXiv:1804.00823v42018
  15. REAL-Q: E2E LLM Quantization via Dynamic Gradient Descent

    Qian Zhang, Yaoming Li, Zhewen Tan +9

    cs.LGcs.AIarXiv:2609.00049v12026
  16. Meta Flow Maps enable scalable reward alignment

    Peter Potaptchik, Adhi Saravanan, Abbas Mammadov +3

    stat.MLcs.LGarXiv:2601.14430v22026
  17. A Functional Taxonomy of Music Generation Systems

    Dorien Herremans, Ching-Hua Chuan, Elaine Chew

    cs.SDcs.LGeess.ASarXiv:1812.04186v12018
  18. DeepSWE: Measuring Frontier Coding Agents on Original, Long-Horizon Engineering Tasks

    Wenqi Huang, Charley Lee, Leonard Tng +1

    cs.SEcs.LGarXiv:2607.07946v12026
  19. TransDeepLab: Convolution-Free Transformer-based DeepLab v3+ for Medical Image Segmentation

    Reza Azad, Moein Heidari, Moein Shariatnia +4

    eess.IVcs.CVcs.LGarXiv:2208.00713v12022
  20. Context-Alignment: Activating and Enhancing LLM Capabilities in Time Series

    Yuxiao Hu, Qian Li, Dongxiao Zhang +2

    cs.LGcs.CLstat.AParXiv:2501.03747v32025
  21. GCR: Gradient Coreset Based Replay Buffer Selection For Continual Learning

    Rishabh Tiwari, Krishnateja Killamsetty, Rishabh Iyer +1

    cs.LGcs.AIarXiv:2111.11210v32021
  22. BEHAVIOR Robot Suite: Streamlining Real-World Whole-Body Manipulation for Everyday Household Activities

    Yunfan Jiang, Ruohan Zhang, Josiah Wong +7

    cs.ROcs.AIcs.CVarXiv:2503.05652v22025
  23. Multifidelity deep neural operators for efficient learning of partial differential equations with application to fast inverse design of nanoscale heat transport

    Lu Lu, Raphael Pestourie, Steven G. Johnson +1

    physics.comp-phcs.LGarXiv:2204.06684v12022
  24. Pomegranate: fast and flexible probabilistic modeling in python

    Jacob Schreiber

    cs.AIcs.LGstat.MLarXiv:1711.00137v22017
  25. MolecularRNN: Generating realistic molecular graphs with optimized properties

    Mariya Popova, Mykhailo Shvets, Junier Oliva +1

    cs.LGcs.AIq-bio.MNarXiv:1905.13372v12019
  26. Decentralized Computation Offloading for Multi-User Mobile Edge Computing: A Deep Reinforcement Learning Approach

    Zhao Chen, Xiaodong Wang

    cs.LGeess.SPmath.OCarXiv:1812.07394v12018
  27. All Bark and No Bite: Rogue Dimensions in Transformer Language Models Obscure Representational Quality

    William Timkey, Marten van Schijndel

    cs.CLcs.LGarXiv:2109.04404v12021
  28. Do Sparse Autoencoders Capture Concept Manifolds?

    Usha Bhalla, Thomas Fel, Can Rager +9

    cs.LGcs.AIarXiv:2604.28119v12026
  29. Convolutional-Recurrent Neural Networks for Speech Enhancement

    Han Zhao, Shuayb Zarar, Ivan Tashev +1

    cs.SDcs.CLcs.LGarXiv:1805.00579v12018
  30. Deep Learning for Procedural Content Generation

    Jialin Liu, Sam Snodgrass, Ahmed Khalifa +3

    cs.AIcs.LGarXiv:2010.04548v12020
  31. Not too little, not too much: a theoretical analysis of graph (over)smoothing

    Nicolas Keriven

    stat.MLcs.LGarXiv:2205.12156v22022
  32. When Quantization Breaks Memory: Recurrent-State Write-Back in Low-Precision Temporal Inference

    Ismail Erbas, Xavier Intes, Vikas Pandey

    cs.AIcs.LGphysics.opticsarXiv:2609.04490v12026
  33. Aligning with Human Judgement: The Role of Pairwise Preference in Large Language Model Evaluators

    Yinhong Liu, Han Zhou, Zhijiang Guo +4

    cs.CLcs.AIcs.LGarXiv:2403.16950v52024
  34. A Framework for Evaluating Approximation Methods for Gaussian Process Regression

    Krzysztof Chalupka, Christopher K. I. Williams, Iain Murray

    stat.MLcs.LGstat.COarXiv:1205.6326v22012
  35. PUe: Biased Positive-Unlabeled Learning Enhancement by Causal Inference

    Xutao Wang, Hanting Chen, Tianyu Guo +1

    cs.LGarXiv:2607.13428v12026
  36. Mitigating Over-Optimization in PRM-Guided Search in Mathematical Reasoning by Optimizing the Guide

    Taejong Joo, Diego Klabjan

    cs.AIcs.LGarXiv:2608.30051v12026
  37. POPE: Learning to Reason on Hard Problems via Privileged On-Policy Exploration

    Yuxiao Qu, Amrith Setlur, Virginia Smith +2

    cs.LGcs.AIcs.CLarXiv:2601.18779v12026
  38. MEGABYTE: Predicting Million-byte Sequences with Multiscale Transformers

    Lili Yu, Dániel Simig, Colin Flaherty +3

    cs.LGarXiv:2305.07185v22023
  39. Humanoid Manipulation Interface: Humanoid Whole-Body Manipulation from Robot-Free Demonstrations

    Ruiqian Nai, Boyuan Zheng, Junming Zhao +8

    cs.ROcs.AIcs.LGarXiv:2602.06643v22026
  40. Estimating Node Importance in Knowledge Graphs Using Graph Neural Networks

    Namyong Park, Andrey Kan, Xin Luna Dong +2

    cs.LGcs.IRstat.MLarXiv:1905.08865v22019
  41. POLYGLOT-NER: Massive Multilingual Named Entity Recognition

    Rami Al-Rfou, Vivek Kulkarni, Bryan Perozzi +1

    cs.CLcs.LGarXiv:1410.3791v12014
  42. A Modern Take on the Bias-Variance Tradeoff in Neural Networks

    Brady Neal, Sarthak Mittal, Aristide Baratin +4

    cs.LGstat.MLarXiv:1810.08591v42018
  43. Frequency-Aligned Knowledge Distillation for Lightweight Spatiotemporal Forecasting

    Yuqi Li, Chuanguang Yang, Hansheng Zeng +5

    cs.LGcs.AIcs.CVarXiv:2507.02939v22025
  44. LoongServe: Efficiently Serving Long-Context Large Language Models with Elastic Sequence Parallelism

    Bingyang Wu, Shengyu Liu, Yinmin Zhong +3

    cs.DCcs.LGarXiv:2404.09526v22024
  45. Stop Summation: Min-Form Credit Assignment Is All Process Reward Model Needs for Reasoning

    Jie Cheng, Gang Xiong, Ruixi Qiao +5

    cs.AIcs.LGarXiv:2504.15275v32025
  46. Focused Transformer: Contrastive Training for Context Scaling

    Szymon Tworkowski, Konrad Staniszewski, Mikołaj Pacek +3

    cs.CLcs.AIcs.LGarXiv:2307.03170v22023
  47. Chemception: A Deep Neural Network with Minimal Chemistry Knowledge Matches the Performance of Expert-developed QSAR/QSPR Models

    Garrett B. Goh, Charles Siegel, Abhinav Vishnu +2

    stat.MLcs.AIcs.CEarXiv:1706.06689v12017
  48. Do Large Language Model Benchmarks Test Reliability?

    Joshua Vendrow, Edward Vendrow, Sara Beery +1

    cs.LGcs.CLarXiv:2502.03461v12025
  49. Vchitect-2.0: Parallel Transformer for Scaling Up Video Diffusion Models

    Weichen Fan, Chenyang Si, Junhao Song +16

    cs.CVcs.LGarXiv:2501.08453v12025
  50. Error Detection for PET/CT Radiology Reports: Domain-Specific vs Large Language Models

    Hermione Warr, Harry Anthony, Lilli J Freischem +3

    cs.LGcs.AIarXiv:2608.30021v12026
  51. On the Instance Hardness as a Decision Criterion in TinyML Systems

    Tobiasz Puslecki, Krzysztof Walkowiak

    cs.AIcs.LGarXiv:2608.29913v12026
  52. Walk the Talk? Measuring the Faithfulness of Large Language Model Explanations

    Katie Matton, Robert Osazuwa Ness, John Guttag +1

    cs.CLcs.AIcs.LGarXiv:2504.14150v22025
  53. Adversarial Dropout for Supervised and Semi-supervised Learning

    Sungrae Park, Jun-Keon Park, Su-Jin Shin +1

    cs.LGcs.CVarXiv:1707.03631v22017
  54. On Vanishing Gradients, Over-Smoothing, and Over-Squashing in GNNs: Bridging Recurrent and Graph Learning

    Álvaro Arroyo, Alessio Gravina, Benjamin Gutteridge +5

    cs.LGcs.AIarXiv:2502.10818v22025
  55. Text-to-Image Diffusion Models are Zero-Shot Classifiers

    Kevin Clark, Priyank Jaini

    cs.CVcs.AIcs.LGarXiv:2303.15233v22023
  56. Dataset Pruning: Reducing Training Data by Examining Generalization Influence

    Shuo Yang, Zeke Xie, Hanyu Peng +3

    cs.LGarXiv:2205.09329v22022
  57. The Intervention Gap in Latent World Models

    Donna Vakalis

    cs.LGarXiv:2608.29998v12026
  58. Sparse-Interest Network for Sequential Recommendation

    Qiaoyu Tan, Jianwei Zhang, Jiangchao Yao +4

    cs.IRcs.LGarXiv:2102.09267v12021
  59. Adversarial Attacks on Machine Learning Cybersecurity Defences in Industrial Control Systems

    Eirini Anthi, Lowri Williams, Matilda Rhode +2

    cs.LGcs.CReess.SParXiv:2004.05005v12020
  60. Joint Spatiotemporal Spectral Neural Operators for Learning PDEs on Irregular Domains

    Abdolmehdi Behroozi, Chaopeng Shen

    cs.LGarXiv:2608.29892v12026