Computation and Language

Papers filed under cs.CL on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

1,021 to 1,080 of 11,310

  1. Weakly Supervised Cross-Lingual Named Entity Recognition via Effective Annotation and Representation Projection

    Jian Ni, Georgiana Dinu, Radu Florian

    cs.CLcs.IRarXiv:1707.02483v12017
  2. When Noise Fabricates Bias: The Fragility of LLM-as-a-Judge Bias Measurement under Noisy Text

    DongHyun Ryu, Jaehyeok Lee, YeongJun Hwang +1

    cs.CLcs.LGarXiv:2609.11067v12026
  3. Cultural Alignment in Large Language Models: An Explanatory Analysis Based on Hofstede's Cultural Dimensions

    Reem I. Masoud, Ziquan Liu, Martin Ferianc +2

    cs.CYcs.CLcs.LGarXiv:2309.12342v22023
  4. Generative AI for Programming Education: Benchmarking ChatGPT, GPT-4, and Human Tutors

    Tung Phung, Victor-Alexandru Pădurean, José Cambronero +5

    cs.CYcs.AIcs.CLarXiv:2306.17156v32023
  5. K/V-Cache Interventions Dissociate Representation Alignment from Persona Expression in Decoder-Only Language Models

    Yu Sun, Mengyin Lu, Cong Feng +2

    cs.CLarXiv:2609.11020v12026
  6. Ground-Truth Labels Matter: A Deeper Look into Input-Label Demonstrations

    Kang Min Yoo, Junyeob Kim, Hyuhng Joon Kim +5

    cs.CLcs.AIcs.LGarXiv:2205.12685v22022
  7. Walking Down the Memory Maze: Beyond Context Limit through Interactive Reading

    Howard Chen, Ramakanth Pasunuru, Jason Weston +1

    cs.CLarXiv:2310.05029v12023
  8. LLMCarbon: Modeling the end-to-end Carbon Footprint of Large Language Models

    Ahmad Faiz, Sotaro Kaneda, Ruhan Wang +4

    cs.CLcs.AIcs.CYarXiv:2309.14393v22023
  9. Robust Multimodal Sentiment Analysis with Incomplete Modalities via Semantic-aware Completeness based Reconstruction

    Han-Jun Choi, Byunggill Joe, Saim Shin +1

    cs.CLcs.AIcs.LGarXiv:2609.10950v12026
  10. MultiVis-Agent: A Multi-Agent Framework with Logic Rules for Reliable and Comprehensive Cross-Modal Data Visualization

    Jinwei Lu, Yuanfeng Song, Chen Zhang +1

    cs.CLcs.AIcs.DBarXiv:2601.18320v12026
  11. Structurally Speaking: Motif-Oriented Graph Captioning through Bidirectional Graph-Text Translation

    Hsiao-Ying Lu, Dongyu Liu, Kwan-Liu Ma

    cs.CLcs.LGarXiv:2609.10923v12026
  12. Auto-RecSys: Harnessing Autonomous Research Agents for Industry-Scale Recommender System

    Ming Li, Dai Li, Xuying Ning +11

    cs.CLarXiv:2609.10922v12026
  13. Can ChatGPT Replace Traditional KBQA Models? An In-depth Analysis of the Question Answering Performance of the GPT LLM Family

    Yiming Tan, Dehai Min, Yu Li +4

    cs.CLarXiv:2303.07992v32023
  14. SearchAtlas: Analyzing Agentic Search Strategies via Evidential Query Graphs

    Jiacheng Sang, Mengyuan Li, Sanxing Chen +3

    cs.CLarXiv:2609.10901v12026
  15. Retrieving Multimodal Information for Augmented Generation: A Survey

    Ruochen Zhao, Hailin Chen, Weishi Wang +8

    cs.CLarXiv:2303.10868v32023
  16. EmoLLMs: A Series of Emotional Large Language Models and Annotation Tools for Comprehensive Affective Analysis

    Zhiwei Liu, Kailai Yang, Tianlin Zhang +2

    cs.CLarXiv:2401.08508v22024
  17. NumGLUE: A Suite of Fundamental yet Challenging Mathematical Reasoning Tasks

    Swaroop Mishra, Arindam Mitra, Neeraj Varshney +4

    cs.CLcs.AIcs.LGarXiv:2204.05660v12022
  18. SCOTT: Self-Consistent Chain-of-Thought Distillation

    Peifeng Wang, Zhengyang Wang, Zheng Li +3

    cs.CLarXiv:2305.01879v42023
  19. Revisiting Zeroth-Order Optimization for Memory-Efficient LLM Fine-Tuning: A Benchmark

    Yihua Zhang, Pingzhi Li, Junyuan Hong +10

    cs.LGcs.CLarXiv:2402.11592v32024
  20. Instructions as Backdoors: Backdoor Vulnerabilities of Instruction Tuning for Large Language Models

    Jiashu Xu, Mingyu Derek Ma, Fei Wang +2

    cs.CLcs.AIcs.CRarXiv:2305.14710v22023
  21. CoDAR: Continuous Diffusion Language Models are More Powerful Than You Think

    Junzhe Shen, Jieru Zhao, Ziwei He +1

    cs.CLcs.AIcs.LGarXiv:2603.02547v12026
  22. LLM-Anchored Paralinguistic Enrichment for Alzheimer's Disease Detection

    Xiao Wei, Yuqin Lin, Yaru Cao +6

    cs.CLcs.SDarXiv:2609.10896v12026
  23. Let's Sample Step by Step: Adaptive-Consistency for Efficient Reasoning and Coding with LLMs

    Pranjal Aggarwal, Aman Madaan, Yiming Yang +1

    cs.CLarXiv:2305.11860v22023
  24. Does Linguistic Structure Enrichment Enhance Coherence Assessment? Not With Current Architectures

    Victor Mazzotti, Luiz Pereira, Marina Bitencourt dos Santos +4

    cs.CLcs.AIarXiv:2609.10893v12026
  25. Detectable Only Where It Is Confounded: What Verified Duplication Counts Say About Membership Evidence in Language Models

    Arman Nik Khah

    cs.CLcs.CRcs.LGarXiv:2609.10830v12026
  26. MMedAgent: Learning to Use Medical Tools with Multi-modal Agent

    Binxu Li, Tiankai Yan, Yuanting Pan +8

    cs.CLcs.AIarXiv:2407.02483v22024
  27. Larger Context Window, Fewer Overcorrections: Optimizing Prompts and Batching for Minimal-Edit Grammatical Error Correction

    Kateryna Karpo, Artem Chernodub

    cs.CLarXiv:2609.10810v12026
  28. Analyzing Traditional and Neural Approaches to Multilingual Readability Assessment

    Joshua Wong, Chris Tanner

    cs.CLarXiv:2609.10792v12026
  29. Generating Benchmarks for Factuality Evaluation of Language Models

    Dor Muhlgay, Ori Ram, Inbal Magar +7

    cs.CLcs.AIarXiv:2307.06908v22023
  30. Multilingual in Name Only? Cultural and Linguistic Weaknesses of LLMs in Urdu

    Farah Adeeba, Abdul Rafae Khan, Rajesh Bhatt +1

    cs.CLcs.AIcs.LGarXiv:2609.10758v12026
  31. ReviewerGPT? An Exploratory Study on Using Large Language Models for Paper Reviewing

    Ryan Liu, Nihar B. Shah

    cs.CLcs.AIcs.DLarXiv:2306.00622v12023
  32. Think Before You Link: Rarity, Reasoning, and Retrieval in Multilingual Entity Linking

    Parinthapat Pengpun, Simran Khanuja, Graham Neubig

    cs.CLarXiv:2609.10745v12026
    Summaries:한국어
  33. Neural Networks for Entity Matching: A Survey

    Nils Barlaug, Jon Atle Gulla

    cs.DBcs.CLcs.LGarXiv:2010.11075v22020
  34. CMNIE: An Information Extraction Benchmark for Chinese Military News

    Yan Yu, Mengna Zhu, Zhenyu Song +3

    cs.CLarXiv:2609.10722v12026
  35. On the Creativity of Large Language Models

    Giorgio Franceschelli, Mirco Musolesi

    cs.AIcs.CLcs.CYarXiv:2304.00008v52023
  36. Better to Ask in English: Cross-Lingual Evaluation of Large Language Models for Healthcare Queries

    Yiqiao Jin, Mohit Chandra, Gaurav Verma +3

    cs.CLcs.AIarXiv:2310.13132v22023
  37. Colosseum: Auditing Collusion in Cooperative Multi-Agent Systems

    Mason Nakamura, Abhinav Kumar, Saswat Das +5

    cs.MAcs.AIcs.CLarXiv:2602.15198v22026
  38. Data-Efficient Language Modeling: From Frontier Advancement to Principle-Guided Model Improvement

    Shuxing Yang, Kaihao Zhu, Junjie Yang +13

    cs.CLcs.AIarXiv:2609.10702v12026
  39. Not All Experts are Equal: Efficient Expert Pruning and Skipping for Mixture-of-Experts Large Language Models

    Xudong Lu, Qi Liu, Yuhui Xu +5

    cs.CLcs.AIcs.LGarXiv:2402.14800v22024
  40. Generative agent-based modeling with actions grounded in physical, social, or digital space using Concordia

    Alexander Sasha Vezhnevets, John P. Agapiou, Avia Aharon +7

    cs.AIcs.CLarXiv:2312.03664v22023
  41. Lost in Multilinguality: Dissecting Cross-lingual Factual Inconsistency in Transformer Language Models

    Mingyang Wang, Heike Adel, Lukas Lange +4

    cs.CLarXiv:2504.04264v12025
  42. Conformal Prediction with Large Language Models for Multi-Choice Question Answering

    Bhawesh Kumar, Charlie Lu, Gauri Gupta +4

    cs.CLcs.LGstat.MLarXiv:2305.18404v32023
  43. AssistantBench: Can Web Agents Solve Realistic and Time-Consuming Tasks?

    Ori Yoran, Samuel Joseph Amouyal, Chaitanya Malaviya +3

    cs.CLarXiv:2407.15711v22024
  44. Translation Artifacts in Cross-lingual Transfer Learning

    Mikel Artetxe, Gorka Labaka, Eneko Agirre

    cs.CLcs.LGarXiv:2004.04721v42020
  45. Decoupling KL and Trajectories: A Unified Perspective for SFT, DAgger, Offline RL, and OPD in LLM Distillation

    Anhao Zhao, Haoran Xin, Yingqi Fan +3

    cs.LGcs.AIcs.CLarXiv:2605.16826v12026
  46. Large Language Models and Knowledge Graphs: Opportunities and Challenges

    Jeff Z. Pan, Simon Razniewski, Jan-Christoph Kalo +13

    cs.AIcs.CLarXiv:2308.06374v12023
  47. CycleResearcher: Improving Automated Research via Automated Review

    Yixuan Weng, Minjun Zhu, Guangsheng Bao +4

    cs.CLcs.AIcs.CYarXiv:2411.00816v32024
  48. We're Different, We're the Same: Creative Homogeneity Across LLMs

    Emily Wenger, Yoed Kenett

    cs.CYcs.AIcs.CLarXiv:2501.19361v12025
  49. The AI Scientist-v2: Workshop-Level Automated Scientific Discovery via Agentic Tree Search

    Yutaro Yamada, Robert Tjarko Lange, Cong Lu +5

    cs.AIcs.CLcs.LGarXiv:2504.08066v12025
    Summaries:한국어
  50. UNICORN on RAINBOW: A Universal Commonsense Reasoning Model on a New Multitask Benchmark

    Nicholas Lourie, Ronan Le Bras, Chandra Bhagavatula +1

    cs.CLarXiv:2103.13009v12021
  51. BioT5: Enriching Cross-modal Integration in Biology with Chemical Knowledge and Natural Language Associations

    Qizhi Pei, Wei Zhang, Jinhua Zhu +5

    cs.CLcs.AIcs.LGarXiv:2310.07276v32023
  52. NCP-ArchPreview Technical Report: Moving towards Latent Space Language Models through Next Concept Prediction

    The Intern-NCP Team, :, Jiaqi Cao +26

    cs.CLarXiv:2609.10715v12026
  53. TopiOCQA: Open-domain Conversational Question Answering with Topic Switching

    Vaibhav Adlakha, Shehzaad Dhuliawala, Kaheer Suleman +2

    cs.CLarXiv:2110.00768v32021
  54. Images are Achilles' Heel of Alignment: Exploiting Visual Vulnerabilities for Jailbreaking Multimodal Large Language Models

    Yifan Li, Hangyu Guo, Kun Zhou +2

    cs.CVcs.CLarXiv:2403.09792v32024
  55. Synthetic Text Generation with Differential Privacy: A Simple and Practical Recipe

    Xiang Yue, Huseyin A. Inan, Xuechen Li +6

    cs.CLcs.CRarXiv:2210.14348v32022
  56. X-OPD: Cross-Modal On-Policy Distillation for Capability Alignment in Speech LLMs

    Di Cao, Dongjie Fu, Hai Yu +3

    eess.AScs.AIcs.CLarXiv:2603.24596v32026
  57. Few-shot In-context Learning for Knowledge Base Question Answering

    Tianle Li, Xueguang Ma, Alex Zhuang +3

    cs.CLcs.AIarXiv:2305.01750v22023
  58. DeCap: Decoding CLIP Latents for Zero-Shot Captioning via Text-Only Training

    Wei Li, Linchao Zhu, Longyin Wen +1

    cs.CVcs.AIcs.CLarXiv:2303.03032v12023
  59. Large Language Models and the Reverse Turing Test

    Terrence Sejnowski

    cs.CLcs.AIcs.LGarXiv:2207.14382v92022
  60. IPIGuard: A Novel Tool Dependency Graph-Based Defense Against Indirect Prompt Injection in LLM Agents

    Hengyu An, Jinghuai Zhang, Tianyu Du +4

    cs.CRcs.AIcs.CLarXiv:2508.15310v12025