Machine Learning
Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
6,721 to 6,780 of 20,219
ProtTrans: Towards Cracking the Language of Life's Code Through Self-Supervised Deep Learning and High Performance Computing
Ahmed Elnaggar, Michael Heinzinger, Christian Dallago +9
cs.LGcs.CLcs.DCarXiv:2007.06225v32020A Blockchain-based Decentralized Federated Learning Framework with Committee Consensus
Yuzheng Li, Chuan Chen, Nan Liu +3
cs.DCcs.LGarXiv:2004.00773v12020Distributional Soft Actor-Critic: Off-Policy Reinforcement Learning for Addressing Value Estimation Errors
Jingliang Duan, Yang Guan, Shengbo Eben Li +2
cs.LGcs.AIeess.SYarXiv:2001.02811v32020Federated Learning for Wireless Communications: Motivation, Opportunities and Challenges
Solmaz Niknam, Harpreet S. Dhillon, Jeffery H. Reed
eess.SPcs.LGstat.MLarXiv:1908.06847v42019Attributed Graph Clustering via Adaptive Graph Convolution
Xiaotong Zhang, Han Liu, Qimai Li +1
cs.LGcs.AIstat.MLarXiv:1906.01210v12019Utilizing Deep Learning Towards Multi-modal Bio-sensing and Vision-based Affective Computing
Siddharth Siddharth, Tzyy-Ping Jung, Terrence J. Sejnowski
cs.LGcs.HCeess.SParXiv:1905.07039v12019BCIJelly: An integrated ecosystem for brain-computer interface research
Liyuan Han, Xinrui Yang, Tianyu Zheng +12
cs.HCcs.LGq-bio.NCarXiv:2608.13576v12026Deep Learning in Alzheimer's disease: Diagnostic Classification and Prognostic Prediction using Neuroimaging Data
Taeho Jo, Kwangsik Nho, Andrew J. Saykin
eess.IVcs.LGstat.MLarXiv:1905.00931v42019Model Evaluation, Model Selection, and Algorithm Selection in Machine Learning
Sebastian Raschka
cs.LGstat.MLarXiv:1811.12808v32018Deep Reinforcement Learning for Resource Management in Network Slicing
Rongpeng Li, Zhifeng Zhao, Qi Sun +5
cs.NIcs.LGarXiv:1805.06591v32018Beyond Binary Rewards: Training LMs to Reason About Their Uncertainty
Mehul Damani, Isha Puri, Stewart Slocum +4
cs.LGcs.AIcs.CLarXiv:2507.16806v22025Tensor Decomposition for Signal Processing and Machine Learning
Nicholas D. Sidiropoulos, Lieven De Lathauwer, Xiao Fu +3
stat.MLcs.LGmath.NAarXiv:1607.01668v22016RoboCasa365: A Large-Scale Simulation Framework for Training and Benchmarking Generalist Robots
Soroush Nasiriany, Sepehr Nasiriany, Abhiram Maddukuri +1
cs.ROcs.AIcs.LGarXiv:2603.04356v12026TWIST2: Scalable, Portable, and Holistic Humanoid Data Collection System
Yanjie Ze, Siheng Zhao, Weizhuo Wang +6
cs.ROcs.CVcs.LGarXiv:2511.02832v12025Subliminal Learning: Language models transmit behavioral traits via hidden signals in data
Alex Cloud, Minh Le, James Chua +5
cs.LGcs.AIarXiv:2507.14805v12025Deep Unfolding: Model-Based Inspiration of Novel Deep Architectures
John R. Hershey, Jonathan Le Roux, Felix Weninger
cs.LGcs.NEstat.MLarXiv:1409.2574v42014Do generative video models understand physical principles?
Saman Motamed, Laura Culp, Kevin Swersky +2
cs.CVcs.AIcs.GRarXiv:2501.09038v32025Language-encoded network topology enables large language models to reason about complex networks
Ucchwas Talukder Utsha, Sakib Mostafa, James Zou +1
cs.LGarXiv:2609.03229v12026Safe Reinforcement Learning with Model Uncertainty Estimates
Björn Lütjens, Michael Everett, Jonathan P. How
cs.ROcs.AIcs.LGarXiv:1810.08700v22018Restricted Eigenvalues Beyond Gaussian Width: Threshold Occupancy under Heavy Tails
Shi Fu, Huibo Xu, Qixin Zhang +1
cs.LGarXiv:2609.03504v12026A Primer on Motion Capture with Deep Learning: Principles, Pitfalls and Perspectives
Alexander Mathis, Steffen Schneider, Jessy Lauer +1
cs.CVcs.LGq-bio.NCarXiv:2009.00564v22020A Brief Review of Hypernetworks in Deep Learning
Vinod Kumar Chauhan, Jiandong Zhou, Ping Lu +2
cs.LGarXiv:2306.06955v32023A Two-Stage Forecasting System for CPU Workload Prediction in Private Clouds
Ashir Javeed, Anton Borg, Håkan Grahn +3
cs.LGarXiv:2609.03457v12026Verify Before You Distill: Prompt-Level Teacher Gating for On-Policy Distillation
Zhiwei Zhang, Zechen Sun, Fei Zhao +6
cs.LGcs.AIcs.CLarXiv:2609.02998v12026Investigating EEG-Based Functional Connectivity Patterns for Multimodal Emotion Recognition
Xun Wu, Wei-Long Zheng, Bao-Liang Lu
cs.HCcs.LGarXiv:2004.01973v12020xbench: Tracking Agents Productivity Scaling with Profession-Aligned Real-World Evaluations
Kaiyuan Chen, Yixin Ren, Yang Liu +30
cs.LGarXiv:2506.13651v12025Agentic Empirical Asset Pricing: Methodological Foundations
Yingjian Pan, Xiaowei Ding, Kay Giesecke
cs.AIcs.LGq-fin.STarXiv:2609.00731v12026Summaries:한국어A Comparative Study of Label-free Representation Quality Metrics in Deep Learning
Daniel Richards Arputharaj, Daniel Jönsson, Gabriel Eilertsen
cs.LGcs.CVarXiv:2608.23182v12026Training Alignment Auditors via Reinforcement Learning
Paul Rosu, Rowan Wang
cs.AIcs.LGarXiv:2608.25460v12026Bellman Calibration for Marginalized Importance Weighting in Offline Reinforcement Learning
Lars van der Laan, Nathan Kallus
cs.LGstat.MLarXiv:2608.24858v12026Functional compatibility as a determinant of persistent neural learning
Hossein Javidnia
cs.LGcs.AIarXiv:2608.22462v22026RIBOSPAN: A Long-Context RNA Foundation Model for Versatile RNA Modeling
Ziyuan Wang, Bohao Tang, Fei Zhang +2
cs.LGq-bio.GNarXiv:2608.22849v12026Self-Evolving Curriculum for LLM Reasoning
Xiaoyin Chen, Jiarui Lu, Minsu Kim +6
cs.AIcs.LGarXiv:2505.14970v42025The Laws of Context Allocation: Causal Measurement and Closed-Loop Orchestration in Generative Search
Peiyang Liu, Xi Wang, Di Liang +1
cs.LGcs.CLcs.IRarXiv:2608.23252v12026Length-Adaptive Decoding for Masked Diffusion Machine Translation
Yan Zhan, Mengkai Hou, Wanting Zhang +1
cs.CLcs.AIcs.LGarXiv:2608.22274v12026FlavourBench: Ranking Frontier Language Models with Executable Culinary Ground Truth
Josef Chen, Erim Hayretci
cs.AIcs.CYcs.LGarXiv:2608.20574v12026SparseGS: Sparse View Synthesis using 3D Gaussian Splatting
Haolin Xiong, Sairisheek Muttukuru, Hanyuan Xiao +4
cs.CVcs.LGeess.IVarXiv:2312.00206v42023Denoising-Aware Inversion: Revealing Privacy Risks in Noise-Protected Text Embeddings
Yubo Wang, Shujie Cui, James Bailey +5
cs.LGcs.AIcs.CRarXiv:2608.18610v12026Physics-Unrolled Neural Operator for Wireless Field Modeling
Rafid Umayer Murshed, Saif Ur Rahman, Mingyue Tang +1
cs.LGcs.AIarXiv:2608.18495v12026Harnessing Magnitude-Only and Complex Measurements for Improved Dynamic MRI Reconstruction with Learned Priors
Mahdi Saberi, Yaşar Utku Alçalar, Merve Gülle +2
eess.IVcs.AIcs.CVarXiv:2608.18036v12026CaliBench: Are the Stochastic Dynamics of Video World Models Physically Calibrated?
Jonathan Sadeghi, Jenny Seidenschwarz, Jesse Allardice +3
cs.LGcs.AIarXiv:2608.16829v12026Reference-free logged energy-oracle recovery for neural approximations of symmetric coercive variational problems: conforming Riesz reconstruction and archive-level selection
Karim Bounja, Lahcen Laayouni, Boujemaa Achchab +1
cs.LGmath.NAarXiv:2608.16473v12026Asymptotics-guided learning and symbolic regression for dispersive resonances
Konstantinos Alexopoulos, Josselin Garnier
math.NAcs.LGmath-pharXiv:2608.16152v12026QuantumPhaseNet: A Gauge-Covariant Geometric and Quantum-Spectral Theory of Semantic Concept Hierarchies with Prototype Validation of a Classical Quantum-Inspired Model
Kiyotaka Kasubuchi, Kazuo Fukiya
cs.CLcs.LGarXiv:2608.15820v12026Data-driven techniques for translational neuroscience and personalized neuro-health
Vishal Subedi, Shashipraba N. K. Rajakaruna, Pratyusha Sarkar +10
q-bio.NCcs.AIcs.LGarXiv:2608.13749v12026A Reproducibility Study of Partial Residual Ablations in Pre-LN Transformers
Pratikkumar Babariya
cs.LGarXiv:2608.14689v12026iFuzz-Meta: An Interpretable Fuzzy Learning Framework Bridging Top-Down and Bottom-Up Knowledge Integration
Xiaowei Jiang, Daniel Leong, Beining Cao +5
cs.LGcs.AIcs.HCarXiv:2608.14646v12026GRPO Beyond English: A Large-Scale Study of GRPO in Non-English and Multilingual Settings
Konstantin Dobler, Federico Scozzafava, Jonathan Janke +2
cs.CLcs.LGarXiv:2608.13698v12026Resume Means Resume: A Machine-Checked Conformance Contract for Checkpoint, Interrupt, and Resume Semantics in Workflow Persistence Layers
Sajjad Khan
cs.LGcs.DCcs.LOarXiv:2608.03836v32026Retrieval-Augmented Generation for Code Summarization via Hybrid GNN
Shangqing Liu, Yu Chen, Xiaofei Xie +2
cs.LGcs.AIarXiv:2006.05405v52020SAF-OPD: Stable Advantage Fusion for On-Policy Distillation
Yifan Ding, Xincheng Wei, Yoshua Y. Li +7
cs.LGcs.AIarXiv:2607.29209v12026Constitutional Midtraining: Content Presence Drives Alignment Gains
Desiree Cho, Cameron Tice, Bernie Hogan +4
cs.CLcs.AIcs.CYarXiv:2607.26654v22026Discrete Diffusion Models: A Unified Framework from Tokenization to Generation
Ye Yuan, Weien Li, Rui Song +19
cs.LGcs.AIcs.CLarXiv:2607.13431v12026Uncovering Latent Reasoning Strategies in Language Models
Awni Altabaa, John Lafferty
cs.LGcs.AIarXiv:2607.17674v12026Three-Body Scattering for Generative Modeling
Peng Sun, Zhenglin Cheng, Deyuan Liu +3
cs.LGcs.CVarXiv:2607.18198v12026KVpop -- Key-Value Cache Compression with Predictive Online Pruning
Lukas Hauzenberger, Niklas Schmidinger, Anamaria-Roberta Hartl +5
cs.LGcs.AIarXiv:2607.05061v22026Hierarchical Experimentalist Agents
Abhranil Chandra, Sankaran Vaidyanathan, Utsav Dhanuka +2
cs.AIcs.LGarXiv:2606.29315v12026Scaling Laws for Grid-Based Approximate Nearest Neighbor Search in High Dimensions
Matthew J Liu, Wei Hang Zheng, Vidhan Purohit +4
cs.LGcs.AIarXiv:2607.01283v12026Proxy OPD: On-Policy Distillation with Transferable Relative Proxy Update
Daocheng Fu, Rong Wu, Yu Yang +7
cs.LGcs.AIarXiv:2607.11505v22026DataClaw0: Agentic Tailoring Multimodal Data from Raw Streams
Cong Wan, Zeyu Guo, Zijian Cai +6
cs.LGcs.AIarXiv:2606.21337v22026