Computation and Language
Papers filed under cs.CL on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
5,581 to 5,640 of 11,242
PrivBench: A Holistic and Modular Benchmarking Platform for Evaluating Text-to-Text Privatization
Stephen Meisenbacher, Andreea-Elena Bodea, Ahmet Bilal Akın +3
cs.CLarXiv:2608.29624v12026Memory-First Fact-Checking: A Knowledge-Graph-Grounded Multi-Agent System for Misinformation Detection
Amelia Petrenciuc, Alexandru Lecu, Adrian Groza
cs.CLcs.AIarXiv:2608.29617v12026Head-to-Tail: How Knowledgeable are Large Language Models (LLMs)? A.K.A. Will LLMs Replace Knowledge Graphs?
Kai Sun, Yifan Ethan Xu, Hanwen Zha +2
cs.CLarXiv:2308.10168v22023Agent Zero Memory: Provenance-Aware Long-Term Memory for LLM Agents
Ming Wu, Pengyuan Zhu
cs.CLarXiv:2608.29606v12026Hindsight Memory-PRM: Supervising Memory Management with Auditable Hindsight Credit
Haoxuan Jia, Yang Liu, Yingguang Yang +14
cs.CLarXiv:2608.29605v12026On Extractive and Abstractive Neural Document Summarization with Transformer Language Models
Sandeep Subramanian, Raymond Li, Jonathan Pilault +1
cs.CLarXiv:1909.03186v22019SemTrace: Source-Grounded Semantic Signatures for Tracing LLM Exposure to Protected Documents
Junyan Zhang, Yudong Zeng, Yongwei Huang +3
cs.CLarXiv:2608.29575v12026Adversarial Filters of Dataset Biases
Ronan Le Bras, Swabha Swayamdipta, Chandra Bhagavatula +4
cs.LGcs.AIcs.CLarXiv:2002.04108v32020KLUE: Korean Language Understanding Evaluation
Sungjoon Park, Jihyung Moon, Sungdong Kim +28
cs.CLarXiv:2105.09680v42021The Emergent Symbolic Structure of Artificial Neural Networks
R. Thomas McCoy, Paul Soulos, Tal Linzen +1
cs.CLcs.AIarXiv:2608.29530v120262 OLMo 2 Furious
Team OLMo, Pete Walsh, Luca Soldaini +40
cs.CLcs.LGarXiv:2501.00656v32024The Unreasonable Ineffectiveness of the Deeper Layers
Andrey Gromov, Kushal Tirumala, Hassan Shapourian +2
cs.CLcs.LGstat.MLarXiv:2403.17887v22024Progressive Prompts: Continual Learning for Language Models
Anastasia Razdaibiedina, Yuning Mao, Rui Hou +3
cs.CLcs.AIcs.LGarXiv:2301.12314v12023Defending Against Indirect Prompt Injection Attacks With Spotlighting
Keegan Hines, Gary Lopez, Matthew Hall +3
cs.CRcs.CLcs.LGarXiv:2403.14720v12024SUP-MIMIC: A Multi-Task Clinical Diagnosis Benchmark for Evaluating LLMs' Robustness to Contradictory Evidence
Yi Yu, Bo Wang, Chong Feng +4
cs.CLcs.AIarXiv:2608.29582v12026Large Language Models Understand and Can be Enhanced by Emotional Stimuli
Cheng Li, Jindong Wang, Yixuan Zhang +6
cs.CLcs.AIcs.HCarXiv:2307.11760v72023Combined Scaling for Zero-shot Transfer Learning
Hieu Pham, Zihang Dai, Golnaz Ghiasi +9
cs.LGcs.CLcs.CVarXiv:2111.10050v32021Which one is banana man? Evaluating vision-language models in multi-turn pragmatic interpretation
Alvin Wei Ming Tan, Ben Prystawski, Veronica Boyce
cs.CLarXiv:2608.29571v12026TACS: Trajectory-Aware Candidate Selection for LLM Jailbreak Suffix Optimization
Shiliang Xiao
cs.CLcs.LGarXiv:2608.29564v12026Evaluating LLMs on Conversational Text-to-SQL under Chain Ambiguity and Intent Drift
Yujia Liu, Jiayan Lin, Zijin Hong +6
cs.CLcs.AIcs.DBarXiv:2608.29543v12026Structured Pruning Learns Compact and Accurate Models
Mengzhou Xia, Zexuan Zhong, Danqi Chen
cs.CLcs.LGarXiv:2204.00408v32022LoGo: Token-Level Dynamic Local-Global Attention
Yuqi Pan, Zheng Li, Bohao Tang +2
cs.CLcs.LGarXiv:2608.29539v12026Recurrent Convolutional Neural Networks for Discourse Compositionality
Nal Kalchbrenner, Phil Blunsom
cs.CLarXiv:1306.3584v12013QServe: W4A8KV4 Quantization and System Co-design for Efficient LLM Serving
Yujun Lin, Haotian Tang, Shang Yang +4
cs.CLcs.AIcs.LGarXiv:2405.04532v32024Super Library Agent: Joint Generation and Maintenance of Multiple Applications Beyond the Single Codebase
Daegyu Sung, Yukyeong Lee, Geon Park +2
cs.SEcs.AIcs.CLarXiv:2608.29310v12026A Dataset for Document Grounded Conversations
Kangyan Zhou, Shrimai Prabhumoye, Alan W Black
cs.CLarXiv:1809.07358v12018A Survey on Recent Advances in LLM-Based Multi-turn Dialogue Systems
Zihao Yi, Jiarui Ouyang, Zhe Xu +4
cs.CLcs.AIarXiv:2402.18013v22024CoCoA: Context-Conditional Cultural Alignment for Large Language Models
Kyungdon Lee, Wei Xu, Alan Ritter +2
cs.CLarXiv:2608.29492v12026Ontology-Guided Multi-Agent Extraction of Evaluation Objects from Academic Review Texts: Evidence from Chinese Library and Information Science
Haolin Chen, Hongyi Dong, Yu Zhu +3
cs.CLarXiv:2608.29526v12026A Pretrainer's Guide to Training Data: Measuring the Effects of Data Age, Domain Coverage, Quality, & Toxicity
Shayne Longpre, Gregory Yauney, Emily Reif +8
cs.CLcs.LGarXiv:2305.13169v22023LLM Judges as Raters: A Pre-Registered Audit of Severity, Halo, Reliability, and Version Instability in LLM Essay Scoring on Public Corpora
Veerendra Kumar Sunkavalli
cs.CLarXiv:2608.29517v12026Item-Mean Surrogates: Why Richer Persona Data Fail to Improve LLMs as Human Surrogates
Daehwan Ahn, Chengfeng Mao, Dokyun Lee
cs.CLcs.CYcs.LGarXiv:2608.29455v12026Chain-of-Thought Faithfulness of Reasoning Models Varies with Where and How Preference Cues Are Delivered
Aryo Pradipta Gema, Neel Rajani, Rohit Saxena +2
cs.CLcs.AIarXiv:2608.29464v12026Graph Neural Networks Meet Neural-Symbolic Computing: A Survey and Perspective
Luis C. Lamb, Artur Garcez, Marco Gori +3
cs.AIcs.CLcs.LGarXiv:2003.00330v72020Argument-Aware Semantic Alignment of Normative Texts: A Toulmin-Based Neuro-Symbolic Approach
William Schroeder
cs.CLcs.AIarXiv:2608.29529v12026Emotion-Cause Pair Extraction: A New Task to Emotion Analysis in Texts
Rui Xia, Zixiang Ding
cs.CLarXiv:1906.01267v12019SIC-Agents: Benchmarking and Building an Adaptive Simulator for Pediatric Serious Illness Communication Training
Zihan Wang, Anita Marie Slominska, Rennie Bimman +10
cs.CLarXiv:2608.29481v12026The Landscape of Emerging AI Agent Architectures for Reasoning, Planning, and Tool Calling: A Survey
Tula Masterman, Sandi Besen, Mason Sawtell +1
cs.AIcs.CLarXiv:2404.11584v12024MUDDLE: Measuring Understanding of Documents under Distractor and Length Effects
Jason Luo, Saibilila Abudukelimu, Judy Song +4
cs.CLcs.AIarXiv:2608.29477v12026Repairing the Cracked Foundation: A Survey of Obstacles in Evaluation Practices for Generated Text
Sebastian Gehrmann, Elizabeth Clark, Thibault Sellam
cs.CLcs.AIcs.LGarXiv:2202.06935v12022AI Can Be Easily Persuaded in Clinical Decision Making
Jiayuan Zhu, Jiazhen Pan, Fenglin Liu +2
cs.CLcs.AIarXiv:2608.29453v12026Language as an Abstraction for Hierarchical Deep Reinforcement Learning
Yiding Jiang, Shixiang Gu, Kevin Murphy +1
cs.LGcs.AIcs.CLarXiv:1906.07343v22019Learning by Abstraction: The Neural State Machine
Drew A. Hudson, Christopher D. Manning
cs.AIcs.CLcs.CVarXiv:1907.03950v42019FireAct: Toward Language Agent Fine-tuning
Baian Chen, Chang Shu, Ehsan Shareghi +3
cs.CLcs.AIcs.LGarXiv:2310.05915v12023Evaluating the Semantic Specificity of Representation Steering in Language Models
Zhangdie Yuan, Andreas Vlachos
cs.CLarXiv:2608.29431v12026LoftQ: LoRA-Fine-Tuning-Aware Quantization for Large Language Models
Yixiao Li, Yifan Yu, Chen Liang +4
cs.CLcs.AIcs.LGarXiv:2310.08659v42023Arabic Safety Alignment as Selective Refusal: An Empirical Study of SFT, DPO, and Guard Calibration
Mohamad Zbib, Ammar Mohanna
cs.CLcs.AIarXiv:2608.29378v12026Neural Networks and the Chomsky Hierarchy
Grégoire Delétang, Anian Ruoss, Jordi Grau-Moya +8
cs.LGcs.AIcs.CLarXiv:2207.02098v32022When to Adapt: Conditional Memory Adapters for Retention-Preserving Domain Specialization
Jiayu Hou, Lei Wang
cs.CLarXiv:2608.29327v12026StageWell: A Process-Aligned Chinese Corpus for Positive-Psychology Support Dialogue
Yuxiong Wang, Ziwei Lin, Bo Wang +2
cs.CLcs.AIarXiv:2608.29326v12026WebWorld: The Browser as a World Model for Self-Improving Web Code
Jiajun Wu, Jian Yang, Yaxin Du +7
cs.CLcs.SEarXiv:2608.30530v12026All You Need Is Non-Commutative Words
Carla M. Quispe Flores, Stanley Salvatierra, Renan Cabrera
cs.CLarXiv:2608.29314v12026Towards Conversational Recommendation over Multi-Type Dialogs
Zeming Liu, Haifeng Wang, Zheng-Yu Niu +3
cs.CLcs.AIarXiv:2005.03954v32020A Hierarchical Multi-task Approach for Learning Embeddings from Semantic Tasks
Victor Sanh, Thomas Wolf, Sebastian Ruder
cs.CLarXiv:1811.06031v22018Whose Assessment of Distress? Community Perspectives and LLM Alignment on Well-Being Posts
Andrew Aquilina, Xiang Lorraine Li, Yu-Ru Li
cs.CLarXiv:2608.29446v12026Learning latent representations for style control and transfer in end-to-end speech synthesis
Ya-Jie Zhang, Shifeng Pan, Lei He +1
cs.CLcs.SDeess.ASarXiv:1812.04342v22018AlgoWorlds: Benchmarking Tool Use for Global Optimization in Algorithmic Worlds
Zixiang Xu, Jiaan Wang, Fandong Meng
cs.CLarXiv:2608.29397v12026BPEmb: Tokenization-free Pre-trained Subword Embeddings in 275 Languages
Benjamin Heinzerling, Michael Strube
cs.CLarXiv:1710.02187v12017Evaluating the State-of-the-Art of End-to-End Natural Language Generation: The E2E NLG Challenge
Ondřej Dušek, Jekaterina Novikova, Verena Rieser
cs.CLarXiv:1901.07931v32019Padārtha: Ontology-Grounded Fine-Grained NER Benchmark for Classical Sanskrit
Sujoy Sarkar, Pretam Ray, Paramhans Shah +4
cs.CLarXiv:2608.29324v12026