Artificial Intelligence

Papers filed under cs.AI on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

6,361 to 6,420 of 15,502

  1. STAPO: Stabilizing Reinforcement Learning for LLMs by Silencing Rare Spurious Tokens

    Shiqi Liu, Zeyu He, Guojian Zhan +10

    cs.CLcs.AIarXiv:2602.15620v52026
  2. CrispEdit: Low-Curvature Projections for Scalable Non-Destructive LLM Editing

    Zarif Ikram, Arad Firouzkouhi, Stephen Tu +2

    cs.LGcs.AIarXiv:2602.15823v22026
  3. On Surprising Effectiveness of Masking Updates in Adaptive Optimizers

    Taejong Joo, Wenhan Xia, Cheolmin Kim +2

    cs.LGcs.AIarXiv:2602.15322v12026
  4. DivPrune: Diversity-based Visual Token Pruning for Large Multimodal Models

    Saeed Ranjbar Alvar, Gursimran Singh, Mohammad Akbari +1

    cs.CVcs.AIcs.LGarXiv:2503.02175v22025
  5. Neural Additive Experts: Context-Gated Experts for Controllable Model Additivity

    Guangzhi Xiong, Sanchit Sinha, Aidong Zhang

    cs.LGcs.AIarXiv:2602.10585v12026
  6. RISE-Video: Can Video Generators Decode Implicit World Rules?

    Mingxin Liu, Shuran Ma, Shibei Meng +9

    cs.CVcs.AIarXiv:2602.05986v12026
  7. MIND: Benchmarking Memory Consistency and Action Control in World Models

    Yixuan Ye, Xuanyu Lu, Yuxin Jiang +7

    cs.CVcs.AIarXiv:2602.08025v22026
  8. Concept-Aware Privacy Mechanisms for Defending Embedding Inversion Attacks

    Yu-Che Tsai, Hsiang Hsiao, Kuan-Yu Chen +1

    cs.CRcs.AIarXiv:2602.07090v12026
  9. ParalESN: Enabling parallel information processing in Reservoir Computing

    Matteo Pinna, Giacomo Lagomarsini, Andrea Ceni +1

    cs.LGcs.AIarXiv:2601.22296v22026
  10. Movie Gen: A Cast of Media Foundation Models

    Adam Polyak, Amit Zohar, Andrew Brown +85

    cs.CVcs.AIcs.LGarXiv:2410.13720v22024
  11. Avoiding Premature Collapse: Adaptive Annealing for Entropy-Regularized Structural Inference

    Yizhi Liu

    cs.LGcs.AIarXiv:2601.23039v32026
  12. Group Distributionally Robust Optimization-Driven Reinforcement Learning for LLM Reasoning

    Kishan Panaganti, Zhenwen Liang, Wenhao Yu +2

    cs.LGcs.AIcs.CLarXiv:2601.19280v12026
  13. CooperBench: Why Coding Agents Cannot be Your Teammates Yet

    Arpandeep Khatua, Hao Zhu, Peter Tran +8

    cs.LGcs.AIcs.CLarXiv:2601.13295v22026
  14. Harder Is Better: Boosting Mathematical Reasoning via Difficulty-Aware GRPO and Multi-Aspect Question Reformulation

    Yanqi Dai, Yuxiang Ji, Xiao Zhang +3

    cs.AIcs.CLarXiv:2601.20614v12026
  15. ProFit: Leveraging High-Value Signals in SFT via Probability-Guided Token Selection

    Tao Liu, Taiqiang Wu, Runming Yang +3

    cs.CLcs.AIarXiv:2601.09195v32026
  16. Imagine-then-Plan: Agent Learning from Adaptive Lookahead with World Models

    Youwei Liu, Jian Wang, Hanlin Wang +2

    cs.CLcs.AIcs.LGarXiv:2601.08955v22026
  17. $τ^2$-Bench: Evaluating Conversational Agents in a Dual-Control Environment

    Victor Barres, Honghua Dong, Soham Ray +2

    cs.AIcs.CLarXiv:2506.07982v12025
  18. Windows Agent Arena: Evaluating Multi-Modal OS Agents at Scale

    Rogerio Bonatti, Dan Zhao, Francesco Bonacci +9

    cs.AIarXiv:2409.08264v22024
  19. Qwen2.5-Math Technical Report: Toward Mathematical Expert Model via Self-Improvement

    An Yang, Beichen Zhang, Binyuan Hui +13

    cs.CLcs.AIcs.LGarXiv:2409.12122v12024
  20. Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context

    Gemini Team, Petko Georgiev, Ving Ian Lei +1134

    cs.CLcs.AIarXiv:2403.05530v52024
  21. RankRAG: Unifying Context Ranking with Retrieval-Augmented Generation in LLMs

    Yue Yu, Wei Ping, Zihan Liu +5

    cs.CLcs.AIcs.IRarXiv:2407.02485v12024
  22. OSWorld: Benchmarking Multimodal Agents for Open-Ended Tasks in Real Computer Environments

    Tianbao Xie, Danyang Zhang, Jixuan Chen +14

    cs.AIcs.CLarXiv:2404.07972v22024
  23. Scaling and evaluating sparse autoencoders

    Leo Gao, Tom Dupré la Tour, Henk Tillman +6

    cs.LGcs.AIarXiv:2406.04093v12024
  24. Kimina-Prover Preview: Towards Large Formal Reasoning Models with Reinforcement Learning

    Haiming Wang, Mert Unsal, Xiaohan Lin +37

    cs.AIarXiv:2504.11354v12025
  25. WildBench: Benchmarking LLMs with Challenging Tasks from Real Users in the Wild

    Bill Yuchen Lin, Yuntian Deng, Khyathi Chandu +6

    cs.CLcs.AIarXiv:2406.04770v22024
  26. A Comprehensive Survey on Process-Oriented Automatic Text Summarization with Exploration of LLM-Based Methods

    Yang Zhang, Hanlei Jin, Dan Meng +2

    cs.AIarXiv:2403.02901v32024
  27. Learning Longer-term Dependencies in RNNs with Auxiliary Losses

    Trieu H. Trinh, Andrew M. Dai, Minh-Thang Luong +1

    cs.LGcs.AIstat.MLarXiv:1803.00144v32018
  28. The opportunities and risks of large language models in mental health

    Hannah R. Lawrence, Renee A. Schneider, Susan B. Rubin +3

    cs.CLcs.AIcs.CYarXiv:2403.14814v32024
  29. LIBERO: Benchmarking Knowledge Transfer for Lifelong Robot Learning

    Bo Liu, Yifeng Zhu, Chongkai Gao +4

    cs.AIarXiv:2306.03310v22023
  30. NodeFormer: A Scalable Graph Structure Learning Transformer for Node Classification

    Qitian Wu, Wentao Zhao, Zenan Li +2

    cs.LGcs.AIarXiv:2306.08385v12023
  31. Extending Context Window of Large Language Models via Positional Interpolation

    Shouyuan Chen, Sherman Wong, Liangjian Chen +1

    cs.CLcs.AIcs.LGarXiv:2306.15595v22023
  32. Chameleon: Plug-and-Play Compositional Reasoning with Large Language Models

    Pan Lu, Baolin Peng, Hao Cheng +5

    cs.CLcs.AIcs.CVarXiv:2304.09842v32023
  33. One Fits All:Power General Time Series Analysis by Pretrained LM

    Tian Zhou, PeiSong Niu, Xue Wang +2

    cs.LGcs.AIarXiv:2302.11939v62023
  34. ADBench: Anomaly Detection Benchmark

    Songqiao Han, Xiyang Hu, Hailiang Huang +2

    cs.LGcs.AIarXiv:2206.09426v22022
  35. MedXpertQA: Benchmarking Expert-Level Medical Reasoning and Understanding

    Yuxin Zuo, Shang Qu, Yifei Li +6

    cs.AIcs.CLcs.CVarXiv:2501.18362v32025
  36. TabLLM: Few-shot Classification of Tabular Data with Large Language Models

    Stefan Hegselmann, Alejandro Buendia, Hunter Lang +3

    cs.CLcs.AIarXiv:2210.10723v22022
  37. Pure Transformers are Powerful Graph Learners

    Jinwoo Kim, Tien Dat Nguyen, Seonwoo Min +4

    cs.LGcs.AIarXiv:2207.02505v22022
  38. UI-TARS-2 Technical Report: Advancing GUI Agent with Multi-Turn Reinforcement Learning

    Haoming Wang, Haoyang Zou, Huatong Song +109

    cs.AIcs.CLcs.CVarXiv:2509.02544v22025
  39. Solving Quantitative Reasoning Problems with Language Models

    Aitor Lewkowycz, Anders Andreassen, David Dohan +11

    cs.CLcs.AIcs.LGarXiv:2206.14858v22022
  40. EmbeddingGemma: Powerful and Lightweight Text Representations

    Henrique Schechter Vera, Sahil Dua, Biao Zhang +86

    cs.CLcs.AIarXiv:2509.20354v32025
  41. Improving Graph Collaborative Filtering with Neighborhood-enriched Contrastive Learning

    Zihan Lin, Changxin Tian, Yupeng Hou +1

    cs.IRcs.AIarXiv:2202.06200v22022
  42. TransMorph: Transformer for unsupervised medical image registration

    Junyu Chen, Eric C. Frey, Yufan He +3

    eess.IVcs.AIcs.CVarXiv:2111.10480v62021
  43. MultiBench: Multiscale Benchmarks for Multimodal Representation Learning

    Paul Pu Liang, Yiwei Lyu, Xiang Fan +9

    cs.LGcs.AIcs.CLarXiv:2107.07502v22021
  44. Preservation of the Global Knowledge by Not-True Distillation in Federated Learning

    Gihun Lee, Minchan Jeong, Yongjin Shin +2

    cs.LGcs.AIcs.CVarXiv:2106.03097v52021
  45. MITRE at SemEval-2016 Task 6: Transfer Learning for Stance Detection

    Guido Zarrella, Amy Marsh

    cs.AIcs.CLarXiv:1606.03784v12016
  46. STAN: Spatio-Temporal Attention Network for Next Location Recommendation

    Yingtao Luo, Qiang Liu, Zhaocheng Liu

    cs.IRcs.AIarXiv:2102.04095v12021
  47. On the eigenvector bias of Fourier feature networks: From regression to solving multi-scale PDEs with physics-informed neural networks

    Sifan Wang, Hanwen Wang, Paris Perdikaris

    cs.LGcs.AIstat.MLarXiv:2012.10047v12020
  48. QPLEX: Duplex Dueling Multi-Agent Q-Learning

    Jianhao Wang, Zhizhou Ren, Terry Liu +2

    cs.LGcs.AIcs.MAarXiv:2008.01062v32020
  49. Decision-Making with Auto-Encoding Variational Bayes

    Romain Lopez, Pierre Boyeau, Nir Yosef +2

    stat.MLcs.AIcs.LGarXiv:2002.07217v32020
  50. Meta-World: A Benchmark and Evaluation for Multi-Task and Meta Reinforcement Learning

    Tianhe Yu, Deirdre Quillen, Zhanpeng He +7

    cs.LGcs.AIcs.ROarXiv:1910.10897v22019
  51. Graph Neural Tangent Kernel: Fusing Graph Neural Networks with Graph Kernels

    Simon S. Du, Kangcheng Hou, Barnabás Póczos +3

    cs.LGcs.AIcs.CVarXiv:1905.13192v22019
  52. Text Classification Algorithms: A Survey

    Kamran Kowsari, Kiana Jafari Meimandi, Mojtaba Heidarysafa +3

    cs.LGcs.AIcs.CLarXiv:1904.08067v52019
  53. Deep Reinforcement Learning for Sepsis Treatment

    Aniruddh Raghu, Matthieu Komorowski, Imran Ahmed +3

    cs.AIcs.LGarXiv:1711.09602v12017
  54. Mercury: Ultra-Fast Language Models Based on Diffusion

    Inception Labs, Samar Khanna, Siddhant Kharbanda +10

    cs.CLcs.AIcs.LGarXiv:2506.17298v12025
  55. A Survey on Traffic Signal Control Methods

    Hua Wei, Guanjie Zheng, Vikash Gayah +1

    cs.LGcs.AIstat.MLarXiv:1904.08117v32019
  56. Three scenarios for continual learning

    Gido M. van de Ven, Andreas S. Tolias

    cs.LGcs.AIcs.CVarXiv:1904.07734v12019
  57. Large Batch Optimization for Deep Learning: Training BERT in 76 minutes

    Yang You, Jing Li, Sashank Reddi +7

    cs.LGcs.AIcs.CLarXiv:1904.00962v52019
  58. Soft Actor-Critic Algorithms and Applications

    Tuomas Haarnoja, Aurick Zhou, Kristian Hartikainen +8

    cs.LGcs.AIcs.ROarXiv:1812.05905v22018
  59. The relativistic discriminator: a key element missing from standard GAN

    Alexia Jolicoeur-Martineau

    cs.LGcs.AIcs.CRarXiv:1807.00734v32018
  60. Understanding Batch Normalization

    Johan Bjorck, Carla Gomes, Bart Selman +1

    cs.LGcs.AIstat.MLarXiv:1806.02375v42018