Machine Learning
Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
5,041 to 5,100 of 20,175
Forecasting directional movements of stock prices for intraday trading using LSTM and random forests
Pushpendu Ghosh, Ariel Neufeld, Jajati Keshari Sahoo
cs.LGq-fin.STstat.MLarXiv:2004.10178v22020Polis: 3D Self-Supervision at City Scale
Alexander Rusnak, Sophia Kovalenko, Jingru Wang +3
cs.CVcs.AIcs.LGarXiv:2608.29426v12026Optimizing the Optimizer for Physics-Informed Neural Networks and Kolmogorov-Arnold Networks
Elham Kiyani, Khemraj Shukla, Jorge F. Urbán +2
cs.LGcs.AImath.OCarXiv:2501.16371v62025UBnormal: New Benchmark for Supervised Open-Set Video Anomaly Detection
Andra Acsintoae, Andrei Florescu, Mariana-Iuliana Georgescu +5
cs.CVcs.LGarXiv:2111.08644v32021On Symmetric and Asymmetric LSHs for Inner Product Search
Behnam Neyshabur, Nathan Srebro
stat.MLcs.DScs.IRarXiv:1410.5518v32014Circulant Binary Embedding
Felix X. Yu, Sanjiv Kumar, Yunchao Gong +1
stat.MLcs.LGarXiv:1405.3162v12014A Finite Time Analysis of Two Time-Scale Actor Critic Methods
Yue Wu, Weitong Zhang, Pan Xu +1
cs.LGmath.OCstat.MLarXiv:2005.01350v32020The Impact of Feature Scaling In Machine Learning: Effects on Regression and Classification Tasks
João Manoel Herrera Pinheiro, Suzana Vilas Boas de Oliveira, Thiago Henrique Segreto Silva +5
cs.LGstat.MLarXiv:2506.08274v52025Stabilizing Deep Q-Learning with ConvNets and Vision Transformers under Data Augmentation
Nicklas Hansen, Hao Su, Xiaolong Wang
cs.LGcs.CVcs.ROarXiv:2107.00644v22021Hierarchical Planning with Latent World Models
Wancong Zhang, Basile Terver, Artem Zholus +8
cs.LGarXiv:2604.03208v22026nPINNs: nonlocal Physics-Informed Neural Networks for a parametrized nonlocal universal Laplacian operator. Algorithms and Applications
Guofei Pang, Marta D'Elia, Michael Parks +1
math.APcs.LGmath.OCarXiv:2004.04276v12020DeepSeek vs. ChatGPT vs. Claude: A Comparative Study for Scientific Computing and Scientific Machine Learning Tasks
Qile Jiang, Zhiwei Gao, George Em Karniadakis
cs.LGcs.AIarXiv:2502.17764v22025Spurious Forgetting in Continual Learning of Language Models
Junhao Zheng, Xidi Cai, Shengjie Qiu +1
cs.LGarXiv:2501.13453v12025Graph2Seq: Graph to Sequence Learning with Attention-based Neural Networks
Kun Xu, Lingfei Wu, Zhiguo Wang +3
cs.AIcs.CLcs.LGarXiv:1804.00823v42018REAL-Q: E2E LLM Quantization via Dynamic Gradient Descent
Qian Zhang, Yaoming Li, Zhewen Tan +9
cs.LGcs.AIarXiv:2609.00049v12026Meta Flow Maps enable scalable reward alignment
Peter Potaptchik, Adhi Saravanan, Abbas Mammadov +3
stat.MLcs.LGarXiv:2601.14430v22026A Functional Taxonomy of Music Generation Systems
Dorien Herremans, Ching-Hua Chuan, Elaine Chew
cs.SDcs.LGeess.ASarXiv:1812.04186v12018DeepSWE: Measuring Frontier Coding Agents on Original, Long-Horizon Engineering Tasks
Wenqi Huang, Charley Lee, Leonard Tng +1
cs.SEcs.LGarXiv:2607.07946v12026TransDeepLab: Convolution-Free Transformer-based DeepLab v3+ for Medical Image Segmentation
Reza Azad, Moein Heidari, Moein Shariatnia +4
eess.IVcs.CVcs.LGarXiv:2208.00713v12022Context-Alignment: Activating and Enhancing LLM Capabilities in Time Series
Yuxiao Hu, Qian Li, Dongxiao Zhang +2
cs.LGcs.CLstat.AParXiv:2501.03747v32025GCR: Gradient Coreset Based Replay Buffer Selection For Continual Learning
Rishabh Tiwari, Krishnateja Killamsetty, Rishabh Iyer +1
cs.LGcs.AIarXiv:2111.11210v32021BEHAVIOR Robot Suite: Streamlining Real-World Whole-Body Manipulation for Everyday Household Activities
Yunfan Jiang, Ruohan Zhang, Josiah Wong +7
cs.ROcs.AIcs.CVarXiv:2503.05652v22025Multifidelity deep neural operators for efficient learning of partial differential equations with application to fast inverse design of nanoscale heat transport
Lu Lu, Raphael Pestourie, Steven G. Johnson +1
physics.comp-phcs.LGarXiv:2204.06684v12022Pomegranate: fast and flexible probabilistic modeling in python
Jacob Schreiber
cs.AIcs.LGstat.MLarXiv:1711.00137v22017MolecularRNN: Generating realistic molecular graphs with optimized properties
Mariya Popova, Mykhailo Shvets, Junier Oliva +1
cs.LGcs.AIq-bio.MNarXiv:1905.13372v12019Decentralized Computation Offloading for Multi-User Mobile Edge Computing: A Deep Reinforcement Learning Approach
Zhao Chen, Xiaodong Wang
cs.LGeess.SPmath.OCarXiv:1812.07394v12018All Bark and No Bite: Rogue Dimensions in Transformer Language Models Obscure Representational Quality
William Timkey, Marten van Schijndel
cs.CLcs.LGarXiv:2109.04404v12021Do Sparse Autoencoders Capture Concept Manifolds?
Usha Bhalla, Thomas Fel, Can Rager +9
cs.LGcs.AIarXiv:2604.28119v12026Convolutional-Recurrent Neural Networks for Speech Enhancement
Han Zhao, Shuayb Zarar, Ivan Tashev +1
cs.SDcs.CLcs.LGarXiv:1805.00579v12018Deep Learning for Procedural Content Generation
Jialin Liu, Sam Snodgrass, Ahmed Khalifa +3
cs.AIcs.LGarXiv:2010.04548v12020Not too little, not too much: a theoretical analysis of graph (over)smoothing
Nicolas Keriven
stat.MLcs.LGarXiv:2205.12156v22022When Quantization Breaks Memory: Recurrent-State Write-Back in Low-Precision Temporal Inference
Ismail Erbas, Xavier Intes, Vikas Pandey
cs.AIcs.LGphysics.opticsarXiv:2609.04490v12026Aligning with Human Judgement: The Role of Pairwise Preference in Large Language Model Evaluators
Yinhong Liu, Han Zhou, Zhijiang Guo +4
cs.CLcs.AIcs.LGarXiv:2403.16950v52024A Framework for Evaluating Approximation Methods for Gaussian Process Regression
Krzysztof Chalupka, Christopher K. I. Williams, Iain Murray
stat.MLcs.LGstat.COarXiv:1205.6326v22012PUe: Biased Positive-Unlabeled Learning Enhancement by Causal Inference
Xutao Wang, Hanting Chen, Tianyu Guo +1
cs.LGarXiv:2607.13428v12026Mitigating Over-Optimization in PRM-Guided Search in Mathematical Reasoning by Optimizing the Guide
Taejong Joo, Diego Klabjan
cs.AIcs.LGarXiv:2608.30051v12026POPE: Learning to Reason on Hard Problems via Privileged On-Policy Exploration
Yuxiao Qu, Amrith Setlur, Virginia Smith +2
cs.LGcs.AIcs.CLarXiv:2601.18779v12026MEGABYTE: Predicting Million-byte Sequences with Multiscale Transformers
Lili Yu, Dániel Simig, Colin Flaherty +3
cs.LGarXiv:2305.07185v22023Humanoid Manipulation Interface: Humanoid Whole-Body Manipulation from Robot-Free Demonstrations
Ruiqian Nai, Boyuan Zheng, Junming Zhao +8
cs.ROcs.AIcs.LGarXiv:2602.06643v22026Estimating Node Importance in Knowledge Graphs Using Graph Neural Networks
Namyong Park, Andrey Kan, Xin Luna Dong +2
cs.LGcs.IRstat.MLarXiv:1905.08865v22019POLYGLOT-NER: Massive Multilingual Named Entity Recognition
Rami Al-Rfou, Vivek Kulkarni, Bryan Perozzi +1
cs.CLcs.LGarXiv:1410.3791v12014A Modern Take on the Bias-Variance Tradeoff in Neural Networks
Brady Neal, Sarthak Mittal, Aristide Baratin +4
cs.LGstat.MLarXiv:1810.08591v42018Frequency-Aligned Knowledge Distillation for Lightweight Spatiotemporal Forecasting
Yuqi Li, Chuanguang Yang, Hansheng Zeng +5
cs.LGcs.AIcs.CVarXiv:2507.02939v22025LoongServe: Efficiently Serving Long-Context Large Language Models with Elastic Sequence Parallelism
Bingyang Wu, Shengyu Liu, Yinmin Zhong +3
cs.DCcs.LGarXiv:2404.09526v22024Stop Summation: Min-Form Credit Assignment Is All Process Reward Model Needs for Reasoning
Jie Cheng, Gang Xiong, Ruixi Qiao +5
cs.AIcs.LGarXiv:2504.15275v32025Focused Transformer: Contrastive Training for Context Scaling
Szymon Tworkowski, Konrad Staniszewski, Mikołaj Pacek +3
cs.CLcs.AIcs.LGarXiv:2307.03170v22023Chemception: A Deep Neural Network with Minimal Chemistry Knowledge Matches the Performance of Expert-developed QSAR/QSPR Models
Garrett B. Goh, Charles Siegel, Abhinav Vishnu +2
stat.MLcs.AIcs.CEarXiv:1706.06689v12017Do Large Language Model Benchmarks Test Reliability?
Joshua Vendrow, Edward Vendrow, Sara Beery +1
cs.LGcs.CLarXiv:2502.03461v12025Vchitect-2.0: Parallel Transformer for Scaling Up Video Diffusion Models
Weichen Fan, Chenyang Si, Junhao Song +16
cs.CVcs.LGarXiv:2501.08453v12025Error Detection for PET/CT Radiology Reports: Domain-Specific vs Large Language Models
Hermione Warr, Harry Anthony, Lilli J Freischem +3
cs.LGcs.AIarXiv:2608.30021v12026On the Instance Hardness as a Decision Criterion in TinyML Systems
Tobiasz Puslecki, Krzysztof Walkowiak
cs.AIcs.LGarXiv:2608.29913v12026Walk the Talk? Measuring the Faithfulness of Large Language Model Explanations
Katie Matton, Robert Osazuwa Ness, John Guttag +1
cs.CLcs.AIcs.LGarXiv:2504.14150v22025Adversarial Dropout for Supervised and Semi-supervised Learning
Sungrae Park, Jun-Keon Park, Su-Jin Shin +1
cs.LGcs.CVarXiv:1707.03631v22017On Vanishing Gradients, Over-Smoothing, and Over-Squashing in GNNs: Bridging Recurrent and Graph Learning
Álvaro Arroyo, Alessio Gravina, Benjamin Gutteridge +5
cs.LGcs.AIarXiv:2502.10818v22025Text-to-Image Diffusion Models are Zero-Shot Classifiers
Kevin Clark, Priyank Jaini
cs.CVcs.AIcs.LGarXiv:2303.15233v22023Dataset Pruning: Reducing Training Data by Examining Generalization Influence
Shuo Yang, Zeke Xie, Hanyu Peng +3
cs.LGarXiv:2205.09329v22022The Intervention Gap in Latent World Models
Donna Vakalis
cs.LGarXiv:2608.29998v12026Sparse-Interest Network for Sequential Recommendation
Qiaoyu Tan, Jianwei Zhang, Jiangchao Yao +4
cs.IRcs.LGarXiv:2102.09267v12021Adversarial Attacks on Machine Learning Cybersecurity Defences in Industrial Control Systems
Eirini Anthi, Lowri Williams, Matilda Rhode +2
cs.LGcs.CReess.SParXiv:2004.05005v12020Joint Spatiotemporal Spectral Neural Operators for Learning PDEs on Irregular Domains
Abdolmehdi Behroozi, Chaopeng Shen
cs.LGarXiv:2608.29892v12026