Machine Learning
Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
721 to 780 of 20,221
WebQA: Multihop and Multimodal QA
Yingshan Chang, Mridu Narang, Hisami Suzuki +3
cs.CLcs.AIcs.CVarXiv:2109.00590v42021R2VC: Modular Fact-Checking with Retrieval, Verification, and Confidence Calibration
Dhruv Dixit, Paritosh Pandey
cs.CLcs.LGarXiv:2609.11955v12026InRTL: Effective Intra-Inter Interaction Learning for Relational Tables
Weichen Li, Ken Zhong, Zheng Wang +2
cs.LGcs.AIarXiv:2609.12712v12026Look Before You Leap: Pre-Action Verification for LLM Agents
Asaad Althoubi
cs.LGcs.MAarXiv:2609.11957v12026EHRSHOT: An EHR Benchmark for Few-Shot Evaluation of Foundation Models
Michael Wornow, Rahul Thapa, Ethan Steinberg +2
cs.LGcs.AIcs.CLarXiv:2307.02028v32023SIMS: Scale-Invariant Merit-Function-Based Scalarization for Multi-Task Learning
Zebin Chen, Fei Xing, Yang Chen +4
cs.LGcs.AImath.OCarXiv:2609.12599v12026MPCFormer: fast, performant and private Transformer inference with MPC
Dacheng Li, Rulin Shao, Hongyi Wang +3
cs.LGcs.CRarXiv:2211.01452v22022MyoFlow: Anchor-Tied Rectified Flow for HD-sEMG Gesture Recognition Across Sessions and Subjects
Chenhao Wu, Dingjie Peng, Satoshi Funabashi +5
cs.LGarXiv:2609.17194v12026From Foundation Embeddings to Cropland Maps: Label Efficiency, Temporal Transferability and Independent Human Validation
Mohammad Ammar Mughees, Giovanni Montefoschi, Zhongxin Chen +1
cs.CVcs.LGeess.IVarXiv:2609.17138v12026On the disintegration of the stochastic majority vote: From PAC-Bayesian bounds to a self-bounding algorithm
Julien Bastian, Benjamin Leblanc, Pascal Germain +4
stat.MLcs.LGarXiv:2609.16803v12026Space as an Interventional Invariant: Cross-Modal Predictive Geometry for Stratified Cities and Em-Spaced Intelligence
Tao Yang, Xuhui Lin, Kunyao Li +1
cs.LGcs.CLarXiv:2609.11959v12026Decoding Mixture Perception through Computational Modeling of Component Interactions
Fei Wang, Xiaoya Xie, Junfei Liu +6
cs.LGarXiv:2609.11958v12026Diagnosing Faults in Reinforcement Learning Simulators and World Models with Canonical Polynomial Invariants
Tesfay Zemuy Gebrekidan, Hadush Hailu Gebrerufael
cs.LGcs.AIarXiv:2609.13194v12026Exact ReLU realization of binary affine refinement iterates via reflection folding and cone switching
Boldsaikhan Bolorkhuu, Tsogtgerel Gantumur
math.OCcs.LGarXiv:2609.11962v12026A Comprehensive Review of Deep Learning-based Single Image Super-resolution
Syed Muhammad Arsalan Bashir, Yi Wang, Mahrukh Khan +1
cs.CVcs.LGeess.IVarXiv:2102.09351v32021Towards Spatio-Temporal Aware Traffic Time Series Forecasting--Full Version
Razvan-Gabriel Cirstea, Bin Yang, Chenjuan Guo +2
cs.LGarXiv:2203.15737v32022On the Importance of Gating: Memorization vs. In-Context Learning in State Space Models
William L. Tong, Aryo Lotfi, Emmanuel Abbe +6
cs.LGcs.AIarXiv:2609.16540v12026Sampling using $SU(N)$ gauge equivariant flows
Denis Boyda, Gurtej Kanwar, Sébastien Racanière +5
hep-latcs.LGstat.MLarXiv:2008.05456v22020Enable Deep Learning on Mobile Devices: Methods, Systems, and Applications
Han Cai, Ji Lin, Yujun Lin +5
cs.LGcs.CLcs.CVarXiv:2204.11786v12022Unified Heterogeneous Graph Neural Network solver for Power Flow, Optimal Power Flow and State Estimation
Ferran Bohigas-Daranas, Hamid Latif-Martínez, Eduardo Prieto-Araujo +2
eess.SYcs.LGarXiv:2609.16738v12026Continual Learning for Traversability Prediction with Uncertainty-Aware Adaptation
Hojin Lee, Yunho Lee, Daniel A Duecker +1
cs.ROcs.AIcs.LGarXiv:2609.17141v12026A Multi-Vehicle Dataset with Camera, LiDAR, and Radar Sensors and Scanned 3D Models for Custom Auto-Annotation using RTK-GNSS
Philipp Berthold, Bianca Forkel, Mirko Maehlisch
cs.ROcs.AIcs.CVarXiv:2609.12871v12026The Router Within: Eliciting Native Skill Routing from a Frozen LLM
Ruishuo Chen, Xun Wang, Yu Chen +2
cs.LGcs.AIcs.CLarXiv:2609.15982v12026RLP: Reinforcement as a Pretraining Objective
Ali Hatamizadeh, Syeda Nahida Akter, Shrimai Prabhumoye +5
cs.LGcs.AIcs.CLarXiv:2510.01265v22025ResLRP: The Role of Residual Cancellation in Attribution Instability in Vision Transformers
Jim Berend, Reduan Achtibat, Daniel Schäffer +4
cs.CVcs.AIcs.LGarXiv:2609.17152v12026LLMs as Master Forgers: Generating Synthetic Time Series Data for Manufacturing
Mantek Singh, Jeshwanth Challagundla, Prateek Karnal +3
cs.LGcs.AIarXiv:2609.16155v12026Decoder Design Matters for ECG Delineation
Joseph Scharpf, William Han, Chaojing Duan +3
cs.LGcs.AIarXiv:2609.16489v12026High-Fidelity Digital Twin Data Models by Randomized Dynamic Mode Decomposition and Deep Learning with Applications in Fluid Dynamics
Diana A. Bistrian
cs.LGmath.NAarXiv:2609.17101v12026Federated stochastic bilevel optimization with fully first-order gradients
Yihan Zhang, Rohit Dhaipule, Chiu C Tan +2
cs.LGarXiv:2609.16350v12026Information Geometric Self-Organization at the Edge of Stability in High-Capacity Kernel Associative Memories
Akira Tamamori
cs.LGcs.NEarXiv:2609.16827v12026GraphText: Graph Reasoning in Text Space
Jianan Zhao, Le Zhuo, Yikang Shen +5
cs.CLcs.LGarXiv:2310.01089v12023gradSim: Differentiable simulation for system identification and visuomotor control
Krishna Murthy Jatavallabhula, Miles Macklin, Florian Golemo +11
cs.CVcs.AIcs.LGarXiv:2104.02646v12021Neuro-Symbolic Hierarchical Intention Anticipation in Human Behavior
Farnaz Soleimani, Abdelghani Chibani, Yacine Amirat +1
cs.AIcs.CVcs.HCarXiv:2609.17064v12026Contrastive Triple Extraction with Generative Transformer
Hongbin Ye, Ningyu Zhang, Shumin Deng +4
cs.CLcs.AIcs.DBarXiv:2009.06207v82020TEMPO: Learning Temporal Context for Dynamic Robot Manipulation
Zhenyang Feng, Jimin Heo, Erik B. Sudderth +1
cs.ROcs.CVcs.LGarXiv:2609.16864v12026Near-Optimal Nonconvex Matrix Completion
Jian-Feng Cai, Xiliang Lu, Juntao You
math.NAcs.LGarXiv:2609.17048v12026How to estimate carbon footprint when training deep learning models? A guide and review
Lucia Bouza Heguerte, Aurélie Bugeau, Loïc Lannelongue
cs.LGcs.AIcs.CYarXiv:2306.08323v22023Latent Undertow: How Ordinary Typos Break Probes
Elad David, Max Fomin, Amit LeVi
cs.CLcs.LGarXiv:2609.15994v12026The METRIC-framework for assessing data quality for trustworthy AI in medicine: a systematic review
Daniel Schwabe, Katinka Becker, Martin Seyferth +2
cs.LGcs.AIarXiv:2402.13635v12024Detecting Adversarial Samples Using Influence Functions and Nearest Neighbors
Gilad Cohen, Guillermo Sapiro, Raja Giryes
cs.LGstat.MLarXiv:1909.06872v22019UI-S1: Advancing GUI Automation via Semi-online Reinforcement Learning
Zhengxi Lu, Jiabo Ye, Fei Tang +8
cs.LGcs.AIarXiv:2509.11543v22025Self-Improving LLM Agents at Test-Time
Emre Can Acikgoz, Cheng Qian, Heng Ji +2
cs.LGcs.AIcs.CLarXiv:2510.07841v12025An Empirical Study of Language CNN for Image Captioning
Jiuxiang Gu, Gang Wang, Jianfei Cai +1
cs.CVcs.LGarXiv:1612.07086v32016InfLLM-V2: Dense-Sparse Switchable Attention for Seamless Short-to-Long Adaptation
Weilin Zhao, Zihan Zhou, Zhou Su +10
cs.CLcs.AIcs.LGarXiv:2509.24663v12025FrontierCS: Evolving Challenges for Evolving Intelligence
Qiuyang Mang, Wenhao Chai, Zhifei Li +48
cs.LGcs.SEarXiv:2512.15699v12025Repurposing Deep Limit Order Book Forecasting for Scenario-Conditioned Market Impact Modeling
Eljas Linna, Kestutis Baltakys, Derrick Manoharan +2
cs.LGcs.AIarXiv:2609.16930v12026A panoramic aerodynamic performance prediction method for turbomachinery cascades using transformer-enhanced neural operator
Qineng Wang, Zhendong Guo, Liming Song +1
cs.LGphysics.comp-phphysics.flu-dynarXiv:2609.16066v12026Type-IV Code Clone Detection via Layer-Wise Non-Contrastive Representation Learning
Luciano Marchezan, Kevin Delcourt, Eugene Syriani +1
cs.SEcs.LGarXiv:2609.17338v12026Convolutional Neural Networks for Global Human Settlements Mapping from Sentinel-2 Satellite Imagery
Christina Corbane, Vasileios Syrris, Filip Sabo +5
eess.IVcs.CVcs.LGarXiv:2006.03267v22020Generalization Can Emerge in Tabular Foundation Models From a Single Table
Junwei Ma, Nour Shaheen, Alex Labach +4
cs.LGcs.AIarXiv:2511.09665v12025A Comparative Analysis of Forecasting Financial Time Series Using ARIMA, LSTM, and BiLSTM
Sima Siami-Namini, Neda Tavakoli, Akbar Siami Namin
cs.LGcs.CEcs.PFarXiv:1911.09512v12019Word meaning in minds and machines
Brenden M. Lake, Gregory L. Murphy
cs.CLcs.AIcs.LGarXiv:2008.01766v32020A Systematic Evaluation of Machine Learning Methods for Fault Detection and Line Identification in Electrical Power Grids
Julian Oelhaf, Georg Kordowich, Paula Andrea Pérez-Toro +4
cs.LGeess.SParXiv:2609.16744v12026Towards Optimizing SQL Generation via LLM Routing
Mohammadhossein Malekpour, Nour Shaheen, Foutse Khomh +1
cs.DBcs.AIcs.LGarXiv:2411.04319v12024Representing Numbers in NLP: a Survey and a Vision
Avijit Thawani, Jay Pujara, Pedro A. Szekely +1
cs.CLcs.AIcs.LGarXiv:2103.13136v12021Seeing What Matters: Visual Cue Guided Video Planning for Generalizable Robot Navigation
Hojin Lee, Sizhe Lester Li, Maximilian Hilger +4
cs.ROcs.AIcs.CVarXiv:2609.16737v12026Conflict-Aware Client Selection for Multi-Server Federated Learning
Mingwei Hong, Zheng Lin, Zehang Lin +7
cs.LGcs.NIarXiv:2602.02458v12026SkipGNN: Predicting Molecular Interactions with Skip-Graph Networks
Kexin Huang, Cao Xiao, Lucas Glass +2
q-bio.MNcs.LGarXiv:2004.14949v22020Low-shot Learning via Covariance-Preserving Adversarial Augmentation Networks
Hang Gao, Zheng Shou, Alireza Zareian +2
cs.LGstat.MLarXiv:1810.11730v32018Towards better understanding of gradient-based attribution methods for Deep Neural Networks
Marco Ancona, Enea Ceolini, Cengiz Öztireli +1
cs.LGstat.MLarXiv:1711.06104v42017