Computation and Language
Papers filed under cs.CL on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
2,641 to 2,700 of 11,245
SETS: Leveraging Self-Verification and Self-Correction for Improved Test-Time Scaling
Jiefeng Chen, Jie Ren, Xinyun Chen +4
cs.AIcs.CLarXiv:2501.19306v52025PlotMachines: Outline-Conditioned Generation with Dynamic Plot State Tracking
Hannah Rashkin, Asli Celikyilmaz, Yejin Choi +1
cs.CLarXiv:2004.14967v22020EmoStance: Response-Side Affective-Orientation Control for Empathetic Response Generation via Emoji Weak Supervision
Ziyuan Jin, Yuxuan Ge, Zheng Tian
cs.AIcs.CLarXiv:2609.02133v12026Vision-Language Models for Edge Networks: A Comprehensive Survey
Ahmed Sharshar, Latif U. Khan, Waseem Ullah +1
cs.CVcs.AIcs.CLarXiv:2502.07855v22025SelfDoc: Self-Supervised Document Representation Learning
Peizhao Li, Jiuxiang Gu, Jason Kuen +5
cs.CVcs.CLarXiv:2106.03331v12021Towards Trustworthy Retrieval Augmented Generation for Large Language Models: A Survey
Bo Ni, Zheyuan Liu, Leyao Wang +17
cs.CLcs.AIarXiv:2502.06872v12025Content-Based Citation Recommendation
Chandra Bhagavatula, Sergey Feldman, Russell Power +1
cs.CLcs.DLcs.IRarXiv:1802.08301v12018Think Deep, Not Just Long: Measuring LLM Reasoning Effort via Deep-Thinking Tokens
Wei-Lin Chen, Liqian Peng, Tian Tan +5
cs.CLarXiv:2602.13517v22026DIET: Lightweight Language Understanding for Dialogue Systems
Tanja Bunk, Daksh Varshneya, Vladimir Vlasov +1
cs.CLarXiv:2004.09936v32020GME: Improving Universal Multimodal Retrieval by Multimodal LLMs
Xin Zhang, Yanzhao Zhang, Wen Xie +7
cs.CLcs.IRarXiv:2412.16855v22024Prompting4Debugging: Red-Teaming Text-to-Image Diffusion Models by Finding Problematic Prompts
Zhi-Yi Chin, Chieh-Ming Jiang, Ching-Chun Huang +2
cs.CLcs.CVarXiv:2309.06135v32023Large language models for automated scholarly paper review: A survey
Zhenzhen Zhuang, Jiandong Chen, Hongfeng Xu +2
cs.AIcs.CLcs.DLarXiv:2501.10326v22025The simulation of judgment in LLMs
Edoardo Loru, Jacopo Nudo, Niccolò Di Marco +6
cs.CLcs.AIcs.CYarXiv:2502.04426v32025Graph2Seq: Graph to Sequence Learning with Attention-based Neural Networks
Kun Xu, Lingfei Wu, Zhiguo Wang +3
cs.AIcs.CLcs.LGarXiv:1804.00823v42018A Survey of Agent Memory in the Second Half: Towards Self-Evolving and Long-Horizon Agents
Wei-Chieh Huang, Weizhi Zhang, Yueqing Liang +57
cs.CLcs.AIarXiv:2602.06052v42026Massive Values in Self-Attention Modules are the Key to Contextual Knowledge Understanding
Mingyu Jin, Kai Mei, Wujiang Xu +5
cs.CLarXiv:2502.01563v42025Federated Learning Of Out-Of-Vocabulary Words
Mingqing Chen, Rajiv Mathews, Tom Ouyang +1
cs.CLarXiv:1903.10635v12019Context-Alignment: Activating and Enhancing LLM Capabilities in Time Series
Yuxiao Hu, Qian Li, Dongxiao Zhang +2
cs.LGcs.CLstat.AParXiv:2501.03747v32025MMDocIR: Benchmarking Multimodal Retrieval for Long Documents
Kuicai Dong, Yujing Chang, Xin Deik Goh +3
cs.IRcs.AIcs.CLarXiv:2501.08828v32025Evaluating LLM-based Agents for Multi-Turn Conversations: A Survey
Shengyue Guan, Jindong Wang, Jiang Bian +3
cs.CLcs.AIarXiv:2503.22458v22025All Bark and No Bite: Rogue Dimensions in Transformer Language Models Obscure Representational Quality
William Timkey, Marten van Schijndel
cs.CLcs.LGarXiv:2109.04404v12021SemEval-2026 Task 3: Dimensional Aspect-Based Sentiment Analysis (DimABSA)
Liang-Chih Yu, Jonas Becker, Shamsuddeen Hassan Muhammad +14
cs.CLarXiv:2604.07066v12026Convolutional-Recurrent Neural Networks for Speech Enhancement
Han Zhao, Shuayb Zarar, Ivan Tashev +1
cs.SDcs.CLcs.LGarXiv:1805.00579v12018Efficient Inference for Large Reasoning Models: A Survey
Yue Liu, Jiaying Wu, Yufei He +11
cs.CLarXiv:2503.23077v32025Can Large Language Models Detect Errors in Long Chain-of-Thought Reasoning?
Yancheng He, Shilong Li, Jiaheng Liu +8
cs.CLarXiv:2502.19361v32025Agentic Reward Modeling: Integrating Human Preferences with Verifiable Correctness Signals for Reliable Reward Systems
Hao Peng, Yunjia Qi, Xiaozhi Wang +4
cs.CLcs.AIarXiv:2502.19328v12025Aligning with Human Judgement: The Role of Pairwise Preference in Large Language Model Evaluators
Yinhong Liu, Han Zhou, Zhijiang Guo +4
cs.CLcs.AIcs.LGarXiv:2403.16950v52024When Models Edit Too Much: On the Fidelity of Minimal Code Edits
Tongyao Zhu, Wei Hern Lim, Min-Yen Kan
cs.SEcs.AIcs.CLarXiv:2609.04061v12026Skywork R1V: Pioneering Multimodal Reasoning with Chain-of-Thought
Yi Peng, Peiyu Wang, Xiaokun Wang +12
cs.CVcs.CLarXiv:2504.05599v22025Multilingual Machine Translation with Open Large Language Models at Practical Scale: An Empirical Study
Menglong Cui, Pengzhi Gao, Wei Liu +2
cs.CLarXiv:2502.02481v42025Combining Recurrent and Convolutional Neural Networks for Relation Classification
Ngoc Thang Vu, Heike Adel, Pankaj Gupta +1
cs.CLarXiv:1605.07333v12016GLM-OCR Technical Report
Shuaiqi Duan, Yadong Xue, Weihan Wang +20
cs.CLarXiv:2603.10910v22026DEMix Layers: Disentangling Domains for Modular Language Modeling
Suchin Gururangan, Mike Lewis, Ari Holtzman +2
cs.CLcs.AIarXiv:2108.05036v22021FlashDLM: Accelerating Diffusion Language Model Inference via Efficient KV Caching and Guided Diffusion
Zhanqiu Hu, Jian Meng, Yash Akhauri +4
cs.CLarXiv:2505.21467v22025POPE: Learning to Reason on Hard Problems via Privileged On-Policy Exploration
Yuxiao Qu, Amrith Setlur, Virginia Smith +2
cs.LGcs.AIcs.CLarXiv:2601.18779v12026POLYGLOT-NER: Massive Multilingual Named Entity Recognition
Rami Al-Rfou, Vivek Kulkarni, Bryan Perozzi +1
cs.CLcs.LGarXiv:1410.3791v12014Reasoning Beyond Language: A Comprehensive Survey on Latent Chain-of-Thought Reasoning
Xinghao Chen, Anhao Zhao, Heming Xia +7
cs.CLarXiv:2505.16782v32025ReWOO: Decoupling Reasoning from Observations for Efficient Augmented Language Models
Binfeng Xu, Zhiyuan Peng, Bowen Lei +3
cs.CLcs.AIarXiv:2305.18323v12023Focused Transformer: Contrastive Training for Context Scaling
Szymon Tworkowski, Konrad Staniszewski, Mikołaj Pacek +3
cs.CLcs.AIcs.LGarXiv:2307.03170v22023Satori: Reinforcement Learning with Chain-of-Action-Thought Enhances LLM Reasoning via Autoregressive Search
Maohao Shen, Guangtao Zeng, Zhenting Qi +7
cs.CLcs.AIarXiv:2502.02508v32025Chain-of-Knowledge: Grounding Large Language Models via Dynamic Knowledge Adapting over Heterogeneous Sources
Xingxuan Li, Ruochen Zhao, Yew Ken Chia +4
cs.CLarXiv:2305.13269v42023Do Large Language Model Benchmarks Test Reliability?
Joshua Vendrow, Edward Vendrow, Sara Beery +1
cs.LGcs.CLarXiv:2502.03461v12025Walk the Talk? Measuring the Faithfulness of Large Language Model Explanations
Katie Matton, Robert Osazuwa Ness, John Guttag +1
cs.CLcs.AIcs.LGarXiv:2504.14150v22025ACUTE-EVAL: Improved Dialogue Evaluation with Optimized Questions and Multi-turn Comparisons
Margaret Li, Jason Weston, Stephen Roller
cs.CLarXiv:1909.03087v12019Hierarchical Memory for High-Efficiency Long-Term Reasoning in LLM Agents
Haoran Sun, Shaoning Zeng
cs.CLcs.AIarXiv:2507.22925v12025Scalable Vision Language Model Training via High Quality Data Curation
Hongyuan Dong, Zijian Kang, Weijie Yin +3
cs.CVcs.CLarXiv:2501.05952v32025The MASK Benchmark: Disentangling Honesty From Accuracy in AI Systems
Richard Ren, Arunim Agarwal, Mantas Mazeika +13
cs.LGcs.AIcs.CLarXiv:2503.03750v32025SocioVerse: A World Model for Social Simulation Powered by LLM Agents and A Pool of 10 Million Real-World Users
Xinnong Zhang, Jiayu Lin, Xinyi Mou +18
cs.CLcs.CYarXiv:2504.10157v32025GNN-RAG: Graph Neural Retrieval for Large Language Model Reasoning
Costas Mavromatis, George Karypis
cs.CLcs.AIcs.LGarXiv:2405.20139v12024The Landscape of Prompt Injection Threats in LLM Agents: From Taxonomy to Analysis
Peiran Wang, Xinfeng Li, Chong Xiang +5
cs.CRcs.CLarXiv:2602.10453v12026MathDial: A Dialogue Tutoring Dataset with Rich Pedagogical Properties Grounded in Math Reasoning Problems
Jakub Macina, Nico Daheim, Sankalan Pal Chowdhury +4
cs.CLarXiv:2305.14536v22023OPSDL: On-Policy Self-Distillation for Long-Context Language Models
Xinsen Zhang, Zhenkai Ding, Tianjun Pan +4
cs.CLcs.AIarXiv:2604.17535v12026VnCoreNLP: A Vietnamese Natural Language Processing Toolkit
Thanh Vu, Dat Quoc Nguyen, Dai Quoc Nguyen +2
cs.CLarXiv:1801.01331v22018HM-RAG: Hierarchical Multi-Agent Multimodal Retrieval Augmented Generation
Pei Liu, Xin Liu, Ruoyu Yao +4
cs.CLcs.AIarXiv:2504.12330v12025Large Language Models Share Representations of Latent Grammatical Concepts Across Typologically Diverse Languages
Jannik Brinkmann, Chris Wendler, Christian Bartelt +1
cs.CLarXiv:2501.06346v22025Talking Turns: Benchmarking Audio Foundation Models on Turn-Taking Dynamics
Siddhant Arora, Zhiyun Lu, Chung-Cheng Chiu +2
cs.CLcs.SDeess.ASarXiv:2503.01174v12025Pushing Mixture of Experts to the Limit: Extremely Parameter Efficient MoE for Instruction Tuning
Ted Zadouri, Ahmet Üstün, Arash Ahmadian +3
cs.CLcs.LGarXiv:2309.05444v12023In-depth Analysis of Graph-based RAG in a Unified Framework
Yingli Zhou, Yaodong Su, Youran Sun +8
cs.IRcs.CLcs.DBarXiv:2503.04338v22025Word Network Topic Model: A Simple but General Solution for Short and Imbalanced Texts
Yuan Zuo, Jichang Zhao, Ke Xu
cs.CLcs.IRarXiv:1412.5404v12014Captioning Images with Diverse Objects
Subhashini Venugopalan, Lisa Anne Hendricks, Marcus Rohrbach +3
cs.CVcs.CLarXiv:1606.07770v32016