Machine Learning

Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

4,441 to 4,500 of 20,192

  1. An Attention-based Collaboration Framework for Multi-View Network Representation Learning

    Meng Qu, Jian Tang, Jingbo Shang +3

    cs.SIcs.LGstat.MLarXiv:1709.06636v12017
  2. Reuse your FLOPs: Scaling RL on Hard Problems by Conditioning on Very Off-Policy Prefixes

    Amrith Setlur, Zijian Wang, Andrew Cohen +2

    cs.LGcs.AIcs.CLarXiv:2601.18795v22026
  3. Stochastic Variance Reduced Ensemble Adversarial Attack for Boosting the Adversarial Transferability

    Yifeng Xiong, Jiadong Lin, Min Zhang +2

    cs.LGcs.CRcs.CVarXiv:2111.10752v22021
  4. SpArSe: Sparse Architecture Search for CNNs on Resource-Constrained Microcontrollers

    Igor Fedorov, Ryan P. Adams, Matthew Mattina +1

    cs.LGcs.CVarXiv:1905.12107v12019
  5. Hyperbolic Graph Attention Network

    Yiding Zhang, Xiao Wang, Xunqiang Jiang +2

    cs.LGstat.MLarXiv:1912.03046v12019
  6. Interpolated Policy Gradient: Merging On-Policy and Off-Policy Gradient Estimation for Deep Reinforcement Learning

    Shixiang Gu, Timothy Lillicrap, Zoubin Ghahramani +3

    cs.LGcs.AIcs.ROarXiv:1706.00387v12017
  7. Memori: A Persistent Memory Layer for Efficient, Context-Aware LLM Agents

    Luiz C. Borro, Luiz A. B. Macarini, Gordon Tindall +2

    cs.LGarXiv:2603.19935v12026
  8. Shifting Attention to Relevance: Towards the Predictive Uncertainty Quantification of Free-Form Large Language Models

    Jinhao Duan, Hao Cheng, Shiqi Wang +5

    cs.CLcs.AIcs.LGarXiv:2307.01379v32023
  9. Stop Treating Collisions Equally: Qualification-Aware Semantic ID Learning for Recommendation at Industrial Scale

    Zheng Hu, Yuxin Chen, Yongsen Pan +13

    cs.IRcs.LGarXiv:2603.00632v12026
  10. How Language Models Choose Sides: Internal Representations of Instruction Hierarchy

    Enrique Balp-Straffon, Chih-Hao Hsu, Rushiraj Gadhvi +3

    cs.AIcs.CLcs.LGarXiv:2608.28648v12026
  11. Tweet2Vec: Character-Based Distributed Representations for Social Media

    Bhuwan Dhingra, Zhong Zhou, Dylan Fitzpatrick +2

    cs.LGcs.CLarXiv:1605.03481v22016
  12. PLeak: Prompt Leaking Attacks against Large Language Model Applications

    Bo Hui, Haolin Yuan, Neil Gong +2

    cs.CRcs.AIcs.LGarXiv:2405.06823v32024
  13. PlasticineLab: A Soft-Body Manipulation Benchmark with Differentiable Physics

    Zhiao Huang, Yuanming Hu, Tao Du +4

    cs.LGcs.AIcs.CVarXiv:2104.03311v12021
  14. The Halt Vector: Internalizing a Causal Steering Intervention for Efficient Reasoning

    Dylan Jayabahu, Tinuade Adeleke

    cs.LGcs.AIcs.CLarXiv:2608.28859v12026
  15. Improving Named Entity Recognition by External Context Retrieving and Cooperative Learning

    Xinyu Wang, Yong Jiang, Nguyen Bach +4

    cs.CLcs.AIcs.LGarXiv:2105.03654v32021
  16. Improved Conditional VRNNs for Video Prediction

    Lluis Castrejon, Nicolas Ballas, Aaron Courville

    cs.CVcs.LGarXiv:1904.12165v12019
  17. Auditing Reasoning-Trace Memorization Claims after Unlearning with Head-Conditioned Canaries

    Yanhang Li, Zhichao Fan, Zexin Zhuang

    cs.LGcs.AIarXiv:2605.18891v12026
  18. Survey and cross-benchmark comparison of remaining time prediction methods in business process monitoring

    Ilya Verenich, Marlon Dumas, Marcello La Rosa +2

    cs.AIcs.LGarXiv:1805.02896v22018
  19. Self-Attentive Classification-Based Anomaly Detection in Unstructured Logs

    Sasho Nedelkoski, Jasmin Bogatinovski, Alexander Acker +2

    cs.LGcs.IRstat.MLarXiv:2008.09340v12020
  20. MoE Lens -- An Expert Is All You Need

    Marmik Chaudhari, Idhant Gulati, Nishkal Hundia +2

    cs.LGarXiv:2603.05806v12026
  21. Independent SE(3)-Equivariant Models for End-to-End Rigid Protein Docking

    Octavian-Eugen Ganea, Xinyuan Huang, Charlotte Bunne +4

    cs.AIcs.LGarXiv:2111.07786v22021
  22. A Survey on Graph Neural Networks and Graph Transformers in Computer Vision: A Task-Oriented Perspective

    Chaoqi Chen, Yushuang Wu, Qiyuan Dai +5

    cs.CVcs.AIcs.LGarXiv:2209.13232v42022
  23. SortedRL: Accelerating RL Training for LLMs through Online Length-Aware Scheduling

    Yiqi Zhang, Huiqiang Jiang, Xufang Luo +7

    cs.LGcs.AIarXiv:2603.23414v12026
  24. CommonsenseQA 2.0: Exposing the Limits of AI through Gamification

    Alon Talmor, Ori Yoran, Ronan Le Bras +4

    cs.CLcs.AIcs.LGarXiv:2201.05320v12022
  25. Purple Llama CyberSecEval: A Secure Coding Benchmark for Language Models

    Manish Bhatt, Sahana Chennabasappa, Cyrus Nikolaidis +18

    cs.CRcs.LGarXiv:2312.04724v12023
  26. Noise Contrastive Estimation and Negative Sampling for Conditional Models: Consistency and Statistical Efficiency

    Zhuang Ma, Michael Collins

    cs.CLcs.LGstat.MEarXiv:1809.01812v12018
  27. KernelSkill: A Multi-Agent Framework for GPU Kernel Optimization

    Qitong Sun, Jun Han, Tianlin Li +6

    cs.LGcs.AIcs.MAarXiv:2603.10085v12026
  28. Reconstruction or Semantics? What Makes a Latent Space Useful for Robotic World Models

    Nilaksh, Saurav Jha, Artem Zholus +1

    cs.CVcs.LGcs.ROarXiv:2605.06388v12026
  29. On the linearity of large non-linear models: when and why the tangent kernel is constant

    Chaoyue Liu, Libin Zhu, Mikhail Belkin

    cs.LGstat.MLarXiv:2010.01092v32020
  30. Don't Jump Through Hoops and Remove Those Loops: SVRG and Katyusha are Better Without the Outer Loop

    Dmitry Kovalev, Samuel Horvath, Peter Richtarik

    cs.LGmath.OCstat.MLarXiv:1901.08689v22019
  31. Large Language Models for Code Generation: A Comprehensive Survey of Challenges, Techniques, Evaluation, and Applications

    Nam Huynh, Beiyu Lin

    cs.SEcs.LGarXiv:2503.01245v22025
  32. Noisy but Valid: Robust Statistical Evaluation of LLMs with Imperfect Judges

    Chen Feng, Minghe Shen, Ananth Balashankar +2

    cs.LGcs.AIcs.CVarXiv:2601.20913v12026
  33. CORE: Context-Robust Remasking for Diffusion Language Models

    Kevin Zhai, Sabbir Mollah, Zhenyi Wang +1

    cs.LGarXiv:2602.04096v32026
  34. Label Efficient Semi-Supervised Learning via Graph Filtering

    Qimai Li, Xiao-Ming Wu, Han Liu +2

    cs.LGcs.AIstat.MLarXiv:1901.09993v32019
  35. Sharp Restricted Isometry Thresholds for Global Minima of Rank-Restricted Matrix LASSO

    Richard Y. Zhang

    stat.MLcs.ITcs.LGarXiv:2608.29018v12026
  36. Scaling Beyond Masked Diffusion Language Models

    Subham Sekhar Sahoo, Jean-Marie Lemercier, Zhihan Yang +4

    cs.LGcs.CLarXiv:2602.15014v12026
  37. A Systematic Review for Transformer-based Long-term Series Forecasting

    Liyilei Su, Xumin Zuo, Rui Li +3

    cs.LGcs.AIarXiv:2310.20218v12023
  38. Variance-Reduced and Projection-Free Stochastic Optimization

    Elad Hazan, Haipeng Luo

    cs.LGarXiv:1602.02101v22016
  39. Rethinking the Value of Multi-Agent Workflow: A Strong Single Agent Baseline

    Jiawei Xu, Arief Koesdwiady, Sisong Bei +8

    cs.MAcs.CLcs.LGarXiv:2601.12307v12026
  40. Why Reasoning Fails to Plan: A Planning-Centric Analysis of Long-Horizon Decision Making in LLM Agents

    Zehong Wang, Fang Wu, Hongru Wang +8

    cs.AIcs.CLcs.LGarXiv:2601.22311v12026
  41. GUI-GENESIS: Automated Synthesis of Efficient Environments with Verifiable Rewards for GUI Agent Post-Training

    Yuan Cao, Dezhi Ran, Mengzhou Wu +9

    cs.AIcs.LGarXiv:2602.14093v12026
  42. MEL: Coordinate-Preserving EEG Tokenization for fMRI Translation

    Xiangyu Liu, Zeting Yan, Zhitong Yin +2

    cs.LGarXiv:2608.29304v12026
  43. LightLDA: Big Topic Models on Modest Compute Clusters

    Jinhui Yuan, Fei Gao, Qirong Ho +6

    stat.MLcs.DCcs.IRarXiv:1412.1576v12014
  44. Tropical Geometry of Deep Neural Networks

    Liwen Zhang, Gregory Naitzat, Lek-Heng Lim

    cs.LGmath.AGstat.MLarXiv:1805.07091v12018
  45. UCNN: Exploiting Computational Reuse in Deep Neural Networks via Weight Repetition

    Kartik Hegde, Jiyong Yu, Rohit Agrawal +3

    cs.NEcs.LGarXiv:1804.06508v12018
  46. Commercial LLM Agents Are Already Vulnerable to Simple Yet Dangerous Attacks

    Ang Li, Yin Zhou, Vethavikashini Chithrra Raghuram +2

    cs.LGcs.AIarXiv:2502.08586v12025
  47. It's TIME: Towards the Next Generation of Time Series Forecasting Benchmarks

    Zhongzheng Qiao, Sheng Pan, Anni Wang +7

    cs.LGarXiv:2602.12147v42026
  48. Microstructure Representation and Reconstruction of Heterogeneous Materials via Deep Belief Network for Computational Material Design

    Ruijin Cang, Yaopengxiao Xu, Shaohua Chen +3

    cond-mat.mtrl-scics.LGstat.MLarXiv:1612.07401v32016
  49. Jigsaw-CRL: Recovering Global Latent Causal Order from Fragmented Multi-Client Interventions

    Haijie Xu, Chen Zhang

    stat.MLcs.LGarXiv:2608.28991v12026
  50. Extended depth-of-field in holographic image reconstruction using deep learning based auto-focusing and phase-recovery

    Yichen Wu, Yair Rivenson, Yibo Zhang +4

    cs.CVcs.LGphysics.opticsarXiv:1803.08138v12018
  51. A Spectral Identifiability Threshold for Dissipative Rate Recovery from Truncated Liouvillian Spectra

    Yujun Ji, Somyajit Chakraborty

    cs.LGphysics.comp-phquant-pharXiv:2608.29302v12026
  52. Limits of trust in medical AI

    Joshua Hatherley

    cs.LGcs.AIcs.CYarXiv:2503.16692v22025
  53. Development of an Autonomous AI Coding Agent using Monte Carlo Tree Search (MCTS) and Gemini LLM Frameworks

    Pravin Game, Vipin Ramakrishnan, Prathamesh Wagh

    cs.LGcs.AIarXiv:2608.29096v12026
  54. An effective algorithm for hyperparameter optimization of neural networks

    Gonzalo Diaz, Achille Fokoue, Giacomo Nannicini +1

    cs.AIcs.LGcs.NEarXiv:1705.08520v12017
  55. Topic Matching in the Wild: Benchmark and Lessons from Real-World ASR Transcripts

    Saman Rahbar, Xiliang Zhu, Irvin Cardoza +1

    cs.CLcs.AIcs.LGarXiv:2609.00330v12026
  56. Large-Scale Optimization Model Auto-Formulation: Harnessing LLM Flexibility via Structured Workflow

    Kuo Liang, Yuhang Lu, Jianming Mao +7

    cs.AIcs.LGarXiv:2601.09635v32026
  57. MASPO: Unifying Gradient Utilization, Probability Mass, and Signal Reliability for Robust and Sample-Efficient LLM Reasoning

    Xiaoliang Fu, Jiaye Lin, Yangyi Fang +7

    cs.LGcs.AIarXiv:2602.17550v32026
  58. Action Transformer: A Self-Attention Model for Short-Time Pose-Based Human Action Recognition

    Vittorio Mazzia, Simone Angarano, Francesco Salvetti +2

    cs.CVcs.LGarXiv:2107.00606v62021
  59. KernelFoundry: Hardware-aware evolutionary GPU kernel optimization

    Nina Wiedemann, Quentin Leboutet, Michael Paulitsch +2

    cs.DCcs.LGarXiv:2603.12440v22026
  60. Temperature-Adaptive Transformed Teacher Matching

    Hiroaki Aizawa, Yoshikazu Hayashi

    cs.LGcs.CVarXiv:2608.29099v12026