Machine Learning
Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
4,441 to 4,500 of 20,192
An Attention-based Collaboration Framework for Multi-View Network Representation Learning
Meng Qu, Jian Tang, Jingbo Shang +3
cs.SIcs.LGstat.MLarXiv:1709.06636v12017Reuse your FLOPs: Scaling RL on Hard Problems by Conditioning on Very Off-Policy Prefixes
Amrith Setlur, Zijian Wang, Andrew Cohen +2
cs.LGcs.AIcs.CLarXiv:2601.18795v22026Stochastic Variance Reduced Ensemble Adversarial Attack for Boosting the Adversarial Transferability
Yifeng Xiong, Jiadong Lin, Min Zhang +2
cs.LGcs.CRcs.CVarXiv:2111.10752v22021SpArSe: Sparse Architecture Search for CNNs on Resource-Constrained Microcontrollers
Igor Fedorov, Ryan P. Adams, Matthew Mattina +1
cs.LGcs.CVarXiv:1905.12107v12019Hyperbolic Graph Attention Network
Yiding Zhang, Xiao Wang, Xunqiang Jiang +2
cs.LGstat.MLarXiv:1912.03046v12019Interpolated Policy Gradient: Merging On-Policy and Off-Policy Gradient Estimation for Deep Reinforcement Learning
Shixiang Gu, Timothy Lillicrap, Zoubin Ghahramani +3
cs.LGcs.AIcs.ROarXiv:1706.00387v12017Memori: A Persistent Memory Layer for Efficient, Context-Aware LLM Agents
Luiz C. Borro, Luiz A. B. Macarini, Gordon Tindall +2
cs.LGarXiv:2603.19935v12026Shifting Attention to Relevance: Towards the Predictive Uncertainty Quantification of Free-Form Large Language Models
Jinhao Duan, Hao Cheng, Shiqi Wang +5
cs.CLcs.AIcs.LGarXiv:2307.01379v32023Stop Treating Collisions Equally: Qualification-Aware Semantic ID Learning for Recommendation at Industrial Scale
Zheng Hu, Yuxin Chen, Yongsen Pan +13
cs.IRcs.LGarXiv:2603.00632v12026How Language Models Choose Sides: Internal Representations of Instruction Hierarchy
Enrique Balp-Straffon, Chih-Hao Hsu, Rushiraj Gadhvi +3
cs.AIcs.CLcs.LGarXiv:2608.28648v12026Tweet2Vec: Character-Based Distributed Representations for Social Media
Bhuwan Dhingra, Zhong Zhou, Dylan Fitzpatrick +2
cs.LGcs.CLarXiv:1605.03481v22016PLeak: Prompt Leaking Attacks against Large Language Model Applications
Bo Hui, Haolin Yuan, Neil Gong +2
cs.CRcs.AIcs.LGarXiv:2405.06823v32024PlasticineLab: A Soft-Body Manipulation Benchmark with Differentiable Physics
Zhiao Huang, Yuanming Hu, Tao Du +4
cs.LGcs.AIcs.CVarXiv:2104.03311v12021The Halt Vector: Internalizing a Causal Steering Intervention for Efficient Reasoning
Dylan Jayabahu, Tinuade Adeleke
cs.LGcs.AIcs.CLarXiv:2608.28859v12026Improving Named Entity Recognition by External Context Retrieving and Cooperative Learning
Xinyu Wang, Yong Jiang, Nguyen Bach +4
cs.CLcs.AIcs.LGarXiv:2105.03654v32021Improved Conditional VRNNs for Video Prediction
Lluis Castrejon, Nicolas Ballas, Aaron Courville
cs.CVcs.LGarXiv:1904.12165v12019Auditing Reasoning-Trace Memorization Claims after Unlearning with Head-Conditioned Canaries
Yanhang Li, Zhichao Fan, Zexin Zhuang
cs.LGcs.AIarXiv:2605.18891v12026Survey and cross-benchmark comparison of remaining time prediction methods in business process monitoring
Ilya Verenich, Marlon Dumas, Marcello La Rosa +2
cs.AIcs.LGarXiv:1805.02896v22018Self-Attentive Classification-Based Anomaly Detection in Unstructured Logs
Sasho Nedelkoski, Jasmin Bogatinovski, Alexander Acker +2
cs.LGcs.IRstat.MLarXiv:2008.09340v12020MoE Lens -- An Expert Is All You Need
Marmik Chaudhari, Idhant Gulati, Nishkal Hundia +2
cs.LGarXiv:2603.05806v12026Independent SE(3)-Equivariant Models for End-to-End Rigid Protein Docking
Octavian-Eugen Ganea, Xinyuan Huang, Charlotte Bunne +4
cs.AIcs.LGarXiv:2111.07786v22021A Survey on Graph Neural Networks and Graph Transformers in Computer Vision: A Task-Oriented Perspective
Chaoqi Chen, Yushuang Wu, Qiyuan Dai +5
cs.CVcs.AIcs.LGarXiv:2209.13232v42022SortedRL: Accelerating RL Training for LLMs through Online Length-Aware Scheduling
Yiqi Zhang, Huiqiang Jiang, Xufang Luo +7
cs.LGcs.AIarXiv:2603.23414v12026CommonsenseQA 2.0: Exposing the Limits of AI through Gamification
Alon Talmor, Ori Yoran, Ronan Le Bras +4
cs.CLcs.AIcs.LGarXiv:2201.05320v12022Purple Llama CyberSecEval: A Secure Coding Benchmark for Language Models
Manish Bhatt, Sahana Chennabasappa, Cyrus Nikolaidis +18
cs.CRcs.LGarXiv:2312.04724v12023Noise Contrastive Estimation and Negative Sampling for Conditional Models: Consistency and Statistical Efficiency
Zhuang Ma, Michael Collins
cs.CLcs.LGstat.MEarXiv:1809.01812v12018KernelSkill: A Multi-Agent Framework for GPU Kernel Optimization
Qitong Sun, Jun Han, Tianlin Li +6
cs.LGcs.AIcs.MAarXiv:2603.10085v12026Reconstruction or Semantics? What Makes a Latent Space Useful for Robotic World Models
Nilaksh, Saurav Jha, Artem Zholus +1
cs.CVcs.LGcs.ROarXiv:2605.06388v12026On the linearity of large non-linear models: when and why the tangent kernel is constant
Chaoyue Liu, Libin Zhu, Mikhail Belkin
cs.LGstat.MLarXiv:2010.01092v32020Don't Jump Through Hoops and Remove Those Loops: SVRG and Katyusha are Better Without the Outer Loop
Dmitry Kovalev, Samuel Horvath, Peter Richtarik
cs.LGmath.OCstat.MLarXiv:1901.08689v22019Large Language Models for Code Generation: A Comprehensive Survey of Challenges, Techniques, Evaluation, and Applications
Nam Huynh, Beiyu Lin
cs.SEcs.LGarXiv:2503.01245v22025Noisy but Valid: Robust Statistical Evaluation of LLMs with Imperfect Judges
Chen Feng, Minghe Shen, Ananth Balashankar +2
cs.LGcs.AIcs.CVarXiv:2601.20913v12026CORE: Context-Robust Remasking for Diffusion Language Models
Kevin Zhai, Sabbir Mollah, Zhenyi Wang +1
cs.LGarXiv:2602.04096v32026Label Efficient Semi-Supervised Learning via Graph Filtering
Qimai Li, Xiao-Ming Wu, Han Liu +2
cs.LGcs.AIstat.MLarXiv:1901.09993v32019Sharp Restricted Isometry Thresholds for Global Minima of Rank-Restricted Matrix LASSO
Richard Y. Zhang
stat.MLcs.ITcs.LGarXiv:2608.29018v12026Scaling Beyond Masked Diffusion Language Models
Subham Sekhar Sahoo, Jean-Marie Lemercier, Zhihan Yang +4
cs.LGcs.CLarXiv:2602.15014v12026A Systematic Review for Transformer-based Long-term Series Forecasting
Liyilei Su, Xumin Zuo, Rui Li +3
cs.LGcs.AIarXiv:2310.20218v12023Variance-Reduced and Projection-Free Stochastic Optimization
Elad Hazan, Haipeng Luo
cs.LGarXiv:1602.02101v22016Rethinking the Value of Multi-Agent Workflow: A Strong Single Agent Baseline
Jiawei Xu, Arief Koesdwiady, Sisong Bei +8
cs.MAcs.CLcs.LGarXiv:2601.12307v12026Why Reasoning Fails to Plan: A Planning-Centric Analysis of Long-Horizon Decision Making in LLM Agents
Zehong Wang, Fang Wu, Hongru Wang +8
cs.AIcs.CLcs.LGarXiv:2601.22311v12026GUI-GENESIS: Automated Synthesis of Efficient Environments with Verifiable Rewards for GUI Agent Post-Training
Yuan Cao, Dezhi Ran, Mengzhou Wu +9
cs.AIcs.LGarXiv:2602.14093v12026MEL: Coordinate-Preserving EEG Tokenization for fMRI Translation
Xiangyu Liu, Zeting Yan, Zhitong Yin +2
cs.LGarXiv:2608.29304v12026LightLDA: Big Topic Models on Modest Compute Clusters
Jinhui Yuan, Fei Gao, Qirong Ho +6
stat.MLcs.DCcs.IRarXiv:1412.1576v12014Tropical Geometry of Deep Neural Networks
Liwen Zhang, Gregory Naitzat, Lek-Heng Lim
cs.LGmath.AGstat.MLarXiv:1805.07091v12018UCNN: Exploiting Computational Reuse in Deep Neural Networks via Weight Repetition
Kartik Hegde, Jiyong Yu, Rohit Agrawal +3
cs.NEcs.LGarXiv:1804.06508v12018Commercial LLM Agents Are Already Vulnerable to Simple Yet Dangerous Attacks
Ang Li, Yin Zhou, Vethavikashini Chithrra Raghuram +2
cs.LGcs.AIarXiv:2502.08586v12025It's TIME: Towards the Next Generation of Time Series Forecasting Benchmarks
Zhongzheng Qiao, Sheng Pan, Anni Wang +7
cs.LGarXiv:2602.12147v42026Microstructure Representation and Reconstruction of Heterogeneous Materials via Deep Belief Network for Computational Material Design
Ruijin Cang, Yaopengxiao Xu, Shaohua Chen +3
cond-mat.mtrl-scics.LGstat.MLarXiv:1612.07401v32016Jigsaw-CRL: Recovering Global Latent Causal Order from Fragmented Multi-Client Interventions
Haijie Xu, Chen Zhang
stat.MLcs.LGarXiv:2608.28991v12026Extended depth-of-field in holographic image reconstruction using deep learning based auto-focusing and phase-recovery
Yichen Wu, Yair Rivenson, Yibo Zhang +4
cs.CVcs.LGphysics.opticsarXiv:1803.08138v12018A Spectral Identifiability Threshold for Dissipative Rate Recovery from Truncated Liouvillian Spectra
Yujun Ji, Somyajit Chakraborty
cs.LGphysics.comp-phquant-pharXiv:2608.29302v12026Limits of trust in medical AI
Joshua Hatherley
cs.LGcs.AIcs.CYarXiv:2503.16692v22025Development of an Autonomous AI Coding Agent using Monte Carlo Tree Search (MCTS) and Gemini LLM Frameworks
Pravin Game, Vipin Ramakrishnan, Prathamesh Wagh
cs.LGcs.AIarXiv:2608.29096v12026An effective algorithm for hyperparameter optimization of neural networks
Gonzalo Diaz, Achille Fokoue, Giacomo Nannicini +1
cs.AIcs.LGcs.NEarXiv:1705.08520v12017Topic Matching in the Wild: Benchmark and Lessons from Real-World ASR Transcripts
Saman Rahbar, Xiliang Zhu, Irvin Cardoza +1
cs.CLcs.AIcs.LGarXiv:2609.00330v12026Large-Scale Optimization Model Auto-Formulation: Harnessing LLM Flexibility via Structured Workflow
Kuo Liang, Yuhang Lu, Jianming Mao +7
cs.AIcs.LGarXiv:2601.09635v32026MASPO: Unifying Gradient Utilization, Probability Mass, and Signal Reliability for Robust and Sample-Efficient LLM Reasoning
Xiaoliang Fu, Jiaye Lin, Yangyi Fang +7
cs.LGcs.AIarXiv:2602.17550v32026Action Transformer: A Self-Attention Model for Short-Time Pose-Based Human Action Recognition
Vittorio Mazzia, Simone Angarano, Francesco Salvetti +2
cs.CVcs.LGarXiv:2107.00606v62021KernelFoundry: Hardware-aware evolutionary GPU kernel optimization
Nina Wiedemann, Quentin Leboutet, Michael Paulitsch +2
cs.DCcs.LGarXiv:2603.12440v22026Temperature-Adaptive Transformed Teacher Matching
Hiroaki Aizawa, Yoshikazu Hayashi
cs.LGcs.CVarXiv:2608.29099v12026