Computation and Language
Papers filed under cs.CL on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
961 to 1,020 of 11,247
Generative AI for Programming Education: Benchmarking ChatGPT, GPT-4, and Human Tutors
Tung Phung, Victor-Alexandru Pădurean, José Cambronero +5
cs.CYcs.AIcs.CLarXiv:2306.17156v32023K/V-Cache Interventions Dissociate Representation Alignment from Persona Expression in Decoder-Only Language Models
Yu Sun, Mengyin Lu, Cong Feng +2
cs.CLarXiv:2609.11020v12026Ground-Truth Labels Matter: A Deeper Look into Input-Label Demonstrations
Kang Min Yoo, Junyeob Kim, Hyuhng Joon Kim +5
cs.CLcs.AIcs.LGarXiv:2205.12685v22022Walking Down the Memory Maze: Beyond Context Limit through Interactive Reading
Howard Chen, Ramakanth Pasunuru, Jason Weston +1
cs.CLarXiv:2310.05029v12023LLMCarbon: Modeling the end-to-end Carbon Footprint of Large Language Models
Ahmad Faiz, Sotaro Kaneda, Ruhan Wang +4
cs.CLcs.AIcs.CYarXiv:2309.14393v22023Robust Multimodal Sentiment Analysis with Incomplete Modalities via Semantic-aware Completeness based Reconstruction
Han-Jun Choi, Byunggill Joe, Saim Shin +1
cs.CLcs.AIcs.LGarXiv:2609.10950v12026MultiVis-Agent: A Multi-Agent Framework with Logic Rules for Reliable and Comprehensive Cross-Modal Data Visualization
Jinwei Lu, Yuanfeng Song, Chen Zhang +1
cs.CLcs.AIcs.DBarXiv:2601.18320v12026Structurally Speaking: Motif-Oriented Graph Captioning through Bidirectional Graph-Text Translation
Hsiao-Ying Lu, Dongyu Liu, Kwan-Liu Ma
cs.CLcs.LGarXiv:2609.10923v12026Auto-RecSys: Harnessing Autonomous Research Agents for Industry-Scale Recommender System
Ming Li, Dai Li, Xuying Ning +11
cs.CLarXiv:2609.10922v12026Can ChatGPT Replace Traditional KBQA Models? An In-depth Analysis of the Question Answering Performance of the GPT LLM Family
Yiming Tan, Dehai Min, Yu Li +4
cs.CLarXiv:2303.07992v32023SearchAtlas: Analyzing Agentic Search Strategies via Evidential Query Graphs
Jiacheng Sang, Mengyuan Li, Sanxing Chen +3
cs.CLarXiv:2609.10901v12026Retrieving Multimodal Information for Augmented Generation: A Survey
Ruochen Zhao, Hailin Chen, Weishi Wang +8
cs.CLarXiv:2303.10868v32023EmoLLMs: A Series of Emotional Large Language Models and Annotation Tools for Comprehensive Affective Analysis
Zhiwei Liu, Kailai Yang, Tianlin Zhang +2
cs.CLarXiv:2401.08508v22024NumGLUE: A Suite of Fundamental yet Challenging Mathematical Reasoning Tasks
Swaroop Mishra, Arindam Mitra, Neeraj Varshney +4
cs.CLcs.AIcs.LGarXiv:2204.05660v12022SCOTT: Self-Consistent Chain-of-Thought Distillation
Peifeng Wang, Zhengyang Wang, Zheng Li +3
cs.CLarXiv:2305.01879v42023Revisiting Zeroth-Order Optimization for Memory-Efficient LLM Fine-Tuning: A Benchmark
Yihua Zhang, Pingzhi Li, Junyuan Hong +10
cs.LGcs.CLarXiv:2402.11592v32024Instructions as Backdoors: Backdoor Vulnerabilities of Instruction Tuning for Large Language Models
Jiashu Xu, Mingyu Derek Ma, Fei Wang +2
cs.CLcs.AIcs.CRarXiv:2305.14710v22023CoDAR: Continuous Diffusion Language Models are More Powerful Than You Think
Junzhe Shen, Jieru Zhao, Ziwei He +1
cs.CLcs.AIcs.LGarXiv:2603.02547v12026LLM-Anchored Paralinguistic Enrichment for Alzheimer's Disease Detection
Xiao Wei, Yuqin Lin, Yaru Cao +6
cs.CLcs.SDarXiv:2609.10896v12026Let's Sample Step by Step: Adaptive-Consistency for Efficient Reasoning and Coding with LLMs
Pranjal Aggarwal, Aman Madaan, Yiming Yang +1
cs.CLarXiv:2305.11860v22023Does Linguistic Structure Enrichment Enhance Coherence Assessment? Not With Current Architectures
Victor Mazzotti, Luiz Pereira, Marina Bitencourt dos Santos +4
cs.CLcs.AIarXiv:2609.10893v12026Detectable Only Where It Is Confounded: What Verified Duplication Counts Say About Membership Evidence in Language Models
Arman Nik Khah
cs.CLcs.CRcs.LGarXiv:2609.10830v12026MMedAgent: Learning to Use Medical Tools with Multi-modal Agent
Binxu Li, Tiankai Yan, Yuanting Pan +8
cs.CLcs.AIarXiv:2407.02483v22024Larger Context Window, Fewer Overcorrections: Optimizing Prompts and Batching for Minimal-Edit Grammatical Error Correction
Kateryna Karpo, Artem Chernodub
cs.CLarXiv:2609.10810v12026Analyzing Traditional and Neural Approaches to Multilingual Readability Assessment
Joshua Wong, Chris Tanner
cs.CLarXiv:2609.10792v12026Generating Benchmarks for Factuality Evaluation of Language Models
Dor Muhlgay, Ori Ram, Inbal Magar +7
cs.CLcs.AIarXiv:2307.06908v22023Multilingual in Name Only? Cultural and Linguistic Weaknesses of LLMs in Urdu
Farah Adeeba, Abdul Rafae Khan, Rajesh Bhatt +1
cs.CLcs.AIcs.LGarXiv:2609.10758v12026ReviewerGPT? An Exploratory Study on Using Large Language Models for Paper Reviewing
Ryan Liu, Nihar B. Shah
cs.CLcs.AIcs.DLarXiv:2306.00622v12023Think Before You Link: Rarity, Reasoning, and Retrieval in Multilingual Entity Linking
Parinthapat Pengpun, Simran Khanuja, Graham Neubig
cs.CLarXiv:2609.10745v12026Summaries:한국어Neural Networks for Entity Matching: A Survey
Nils Barlaug, Jon Atle Gulla
cs.DBcs.CLcs.LGarXiv:2010.11075v22020CMNIE: An Information Extraction Benchmark for Chinese Military News
Yan Yu, Mengna Zhu, Zhenyu Song +3
cs.CLarXiv:2609.10722v12026On the Creativity of Large Language Models
Giorgio Franceschelli, Mirco Musolesi
cs.AIcs.CLcs.CYarXiv:2304.00008v52023Better to Ask in English: Cross-Lingual Evaluation of Large Language Models for Healthcare Queries
Yiqiao Jin, Mohit Chandra, Gaurav Verma +3
cs.CLcs.AIarXiv:2310.13132v22023Colosseum: Auditing Collusion in Cooperative Multi-Agent Systems
Mason Nakamura, Abhinav Kumar, Saswat Das +5
cs.MAcs.AIcs.CLarXiv:2602.15198v22026Data-Efficient Language Modeling: From Frontier Advancement to Principle-Guided Model Improvement
Shuxing Yang, Kaihao Zhu, Junjie Yang +13
cs.CLcs.AIarXiv:2609.10702v12026Not All Experts are Equal: Efficient Expert Pruning and Skipping for Mixture-of-Experts Large Language Models
Xudong Lu, Qi Liu, Yuhui Xu +5
cs.CLcs.AIcs.LGarXiv:2402.14800v22024Generative agent-based modeling with actions grounded in physical, social, or digital space using Concordia
Alexander Sasha Vezhnevets, John P. Agapiou, Avia Aharon +7
cs.AIcs.CLarXiv:2312.03664v22023Lost in Multilinguality: Dissecting Cross-lingual Factual Inconsistency in Transformer Language Models
Mingyang Wang, Heike Adel, Lukas Lange +4
cs.CLarXiv:2504.04264v12025Conformal Prediction with Large Language Models for Multi-Choice Question Answering
Bhawesh Kumar, Charlie Lu, Gauri Gupta +4
cs.CLcs.LGstat.MLarXiv:2305.18404v32023AssistantBench: Can Web Agents Solve Realistic and Time-Consuming Tasks?
Ori Yoran, Samuel Joseph Amouyal, Chaitanya Malaviya +3
cs.CLarXiv:2407.15711v22024Translation Artifacts in Cross-lingual Transfer Learning
Mikel Artetxe, Gorka Labaka, Eneko Agirre
cs.CLcs.LGarXiv:2004.04721v42020Decoupling KL and Trajectories: A Unified Perspective for SFT, DAgger, Offline RL, and OPD in LLM Distillation
Anhao Zhao, Haoran Xin, Yingqi Fan +3
cs.LGcs.AIcs.CLarXiv:2605.16826v12026Large Language Models and Knowledge Graphs: Opportunities and Challenges
Jeff Z. Pan, Simon Razniewski, Jan-Christoph Kalo +13
cs.AIcs.CLarXiv:2308.06374v12023CycleResearcher: Improving Automated Research via Automated Review
Yixuan Weng, Minjun Zhu, Guangsheng Bao +4
cs.CLcs.AIcs.CYarXiv:2411.00816v32024We're Different, We're the Same: Creative Homogeneity Across LLMs
Emily Wenger, Yoed Kenett
cs.CYcs.AIcs.CLarXiv:2501.19361v12025The AI Scientist-v2: Workshop-Level Automated Scientific Discovery via Agentic Tree Search
Yutaro Yamada, Robert Tjarko Lange, Cong Lu +5
cs.AIcs.CLcs.LGarXiv:2504.08066v12025Summaries:한국어UNICORN on RAINBOW: A Universal Commonsense Reasoning Model on a New Multitask Benchmark
Nicholas Lourie, Ronan Le Bras, Chandra Bhagavatula +1
cs.CLarXiv:2103.13009v12021BioT5: Enriching Cross-modal Integration in Biology with Chemical Knowledge and Natural Language Associations
Qizhi Pei, Wei Zhang, Jinhua Zhu +5
cs.CLcs.AIcs.LGarXiv:2310.07276v32023NCP-ArchPreview Technical Report: Moving towards Latent Space Language Models through Next Concept Prediction
The Intern-NCP Team, :, Jiaqi Cao +26
cs.CLarXiv:2609.10715v12026TopiOCQA: Open-domain Conversational Question Answering with Topic Switching
Vaibhav Adlakha, Shehzaad Dhuliawala, Kaheer Suleman +2
cs.CLarXiv:2110.00768v32021Images are Achilles' Heel of Alignment: Exploiting Visual Vulnerabilities for Jailbreaking Multimodal Large Language Models
Yifan Li, Hangyu Guo, Kun Zhou +2
cs.CVcs.CLarXiv:2403.09792v32024Synthetic Text Generation with Differential Privacy: A Simple and Practical Recipe
Xiang Yue, Huseyin A. Inan, Xuechen Li +6
cs.CLcs.CRarXiv:2210.14348v32022X-OPD: Cross-Modal On-Policy Distillation for Capability Alignment in Speech LLMs
Di Cao, Dongjie Fu, Hai Yu +3
eess.AScs.AIcs.CLarXiv:2603.24596v32026Few-shot In-context Learning for Knowledge Base Question Answering
Tianle Li, Xueguang Ma, Alex Zhuang +3
cs.CLcs.AIarXiv:2305.01750v22023DeCap: Decoding CLIP Latents for Zero-Shot Captioning via Text-Only Training
Wei Li, Linchao Zhu, Longyin Wen +1
cs.CVcs.AIcs.CLarXiv:2303.03032v12023Large Language Models and the Reverse Turing Test
Terrence Sejnowski
cs.CLcs.AIcs.LGarXiv:2207.14382v92022IPIGuard: A Novel Tool Dependency Graph-Based Defense Against Indirect Prompt Injection in LLM Agents
Hengyu An, Jinghuai Zhang, Tianyu Du +4
cs.CRcs.AIcs.CLarXiv:2508.15310v12025A spelling correction model for end-to-end speech recognition
Jinxi Guo, Tara N. Sainath, Ron J. Weiss
eess.AScs.AIcs.CLarXiv:1902.07178v12019DeepAnalyze: Agentic Large Language Models for Autonomous Data Science
Shaolei Zhang, Ju Fan, Meihao Fan +2
cs.AIcs.CLcs.DBarXiv:2510.16872v12025DeepScientist: Advancing Frontier-Pushing Scientific Findings Progressively
Yixuan Weng, Minjun Zhu, Qiujie Xie +4
cs.CLcs.LGarXiv:2509.26603v12025