Computation and Language

Papers filed under cs.CL on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

901 to 960 of 11,220

  1. Natural SQL: Making SQL Easier to Infer from Natural Language Specifications

    Yujian Gan, Xinyun Chen, Jinxia Xie +4

    cs.CLarXiv:2109.05153v12021
  2. Performance-Efficiency Trade-offs in Unsupervised Pre-training for Speech Recognition

    Felix Wu, Kwangyoun Kim, Jing Pan +3

    cs.CLcs.LGcs.SDarXiv:2109.06870v12021
  3. ReGround: Grounding Reviewer Comments in Multimodal Evidence

    Serwar Basch, Lizhen Qu, Iryna Gurevych

    cs.CLcs.IRarXiv:2609.11460v12026
  4. FaithDial: A Faithful Benchmark for Information-Seeking Dialogue

    Nouha Dziri, Ehsan Kamalloo, Sivan Milton +4

    cs.CLarXiv:2204.10757v32022
  5. Cross-Lingual Clinical Annotation Projection as Constrained Text Generation: A Six-Language Study

    Álvaro Rey-Blanes, Francisco J. Moreno-Barea, Francisco J. Veredas

    cs.CLcs.AIarXiv:2609.11450v12026
  6. On the Impact of Anonymization on the Performance of Large Language Models

    Tobias Deußer, Max Hahnbück, Lorenz Sparrenberg +3

    cs.CLcs.AIarXiv:2609.11335v12026
  7. MultiHuSE: A Multimodal Dataset for Humour Styles and Emotions

    Mary Ogbuka Kenneth, Foaad Khosmood, Abbas Edalat

    cs.CLcs.CVcs.MMarXiv:2609.11322v12026
  8. Automatic Lyric Transcription for Greek Songs: Scaling and Task Composition Effects in Whisper Adaptation

    Maria Frangiadaki, Dimitrios Damianos, Kosmas Kritsis +1

    cs.CLcs.SDarXiv:2609.11302v12026
  9. The Illusion of Balanced Multimodal Sentiment Analysis: Beyond the Limits of Optimization-Based Methods

    Ioanna Kaffeza, Efthymios Georgiou, Alexandros Potamianos

    cs.CLarXiv:2609.11247v12026
  10. Mixture-of-Transformers: A Sparse and Scalable Architecture for Multi-Modal Foundation Models

    Weixin Liang, Lili Yu, Liang Luo +8

    cs.CLarXiv:2411.04996v22024
  11. Can LLMs Normalize Databases? A Benchmark and Multi-Agent Framework for Schema Normalization

    Dong-Jae Koh, Huisu Kim, SeongHwan Yoon +4

    cs.CLarXiv:2609.11141v12026
  12. Truncation Sampling as Language Model Desmoothing

    John Hewitt, Christopher D. Manning, Percy Liang

    cs.CLarXiv:2210.15191v12022
  13. A Fragility Spectrum for Recursive Language-Model Training

    Yangze Liu, Zhongyi Han

    cs.CLcs.AIcs.LGarXiv:2609.11149v12026
  14. Overview of the NLPCC 2026 Shared Task 11: Agent-Based Experiment Reproduction from Scientific Papers

    Hanhua Hong, Yizhi Li, Luu Gia Huy +3

    cs.CLarXiv:2609.11117v12026
  15. Assessing the Reusability of Public Speech Resources for Low-Resource Languages: A Central Kurdish Case Study

    Hiwa Asadpour

    cs.CLarXiv:2609.11246v12026
  16. HittER: Hierarchical Transformers for Knowledge Graph Embeddings

    Sanxing Chen, Xiaodong Liu, Jianfeng Gao +3

    cs.CLcs.LGarXiv:2008.12813v22020
  17. OmniHallu: Unified Hallucination Detection for Cross-Modal Comprehension and Generation in Multimodal Large Language Models

    Jianjiang Yang, Peihang Li, Shanqing Xu +3

    cs.CLcs.CVarXiv:2609.11244v12026
  18. Does Neural Machine Translation Benefit from Larger Context?

    Sebastien Jean, Stanislas Lauly, Orhan Firat +1

    stat.MLcs.CLcs.LGarXiv:1704.05135v12017
  19. Automated Identification of Competing Narratives in Political Discourse on Social Media

    Sergej Wildemann, Erick Elejalde

    cs.CLcs.SIarXiv:2609.11202v12026
  20. FlexComp: One Model for Every Ratio in Context Compression

    Kaiyan Zhao, Zhongtao Miao, Akiko Aizawa +1

    cs.CLarXiv:2609.11192v12026
  21. Have Faith in Faithfulness: Going Beyond Circuit Overlap When Finding Model Mechanisms

    Michael Hanna, Sandro Pezzelle, Yonatan Belinkov

    cs.LGcs.CLarXiv:2403.17806v22024
  22. Rubric-Aligned Disentangled Evaluation of Human Simultaneous Interpreting

    Ziyu Zhang, Satoshi Nakamura

    cs.CLarXiv:2609.11131v12026
  23. From Repetition to Recognition: Inductive Discovery of Disinformation Narratives

    Max Upravitelev, Veronika Solopova, Jing Yang +4

    cs.CLarXiv:2609.11128v12026
  24. SentiBERT: A Transferable Transformer-Based Architecture for Compositional Sentiment Semantics

    Da Yin, Tao Meng, Kai-Wei Chang

    cs.CLarXiv:2005.04114v42020
  25. ProMediConv: Benchmarking Proactive Conversational Agents in Legal Dispute Mediation

    Zesheng Wei, Mengfan Li, Wenhao Liu +3

    cs.CLarXiv:2609.11101v12026
  26. Distribution-aware Language Neuron Identification in Multilingual Large Language Models

    Minjun Kim, Inho Won, Junghun Yuk +3

    cs.CLarXiv:2609.10993v12026
  27. Rebalancing Token Importance in Language Models with TF-IDF Weighted Cross-Entropy Loss

    Zhijian Li, Stefan Larson, Kevin Leach

    cs.CLcs.LGarXiv:2609.11029v12026
  28. Beyond Consensus: Perspectivist Modeling and Evaluation of Annotator Disagreement in NLP

    Yinuo Xu, David Jurgens

    cs.CLarXiv:2601.09065v22026
  29. Rethinking Verbalized Confidence for LLM-as-a-Judge: A Compatibility Shift on Post-2025 Proprietary Models

    Yu-Chung Hsiao

    cs.CLarXiv:2609.10996v12026
  30. Using Semantic Uncertainty to Estimate Transition Relevance in Turn-taking

    Muhammad Umair, Jan P. de Ruiter

    cs.CLarXiv:2609.10934v12026
  31. Weakly Supervised Cross-Lingual Named Entity Recognition via Effective Annotation and Representation Projection

    Jian Ni, Georgiana Dinu, Radu Florian

    cs.CLcs.IRarXiv:1707.02483v12017
  32. When Noise Fabricates Bias: The Fragility of LLM-as-a-Judge Bias Measurement under Noisy Text

    DongHyun Ryu, Jaehyeok Lee, YeongJun Hwang +1

    cs.CLcs.LGarXiv:2609.11067v12026
  33. Cultural Alignment in Large Language Models: An Explanatory Analysis Based on Hofstede's Cultural Dimensions

    Reem I. Masoud, Ziquan Liu, Martin Ferianc +2

    cs.CYcs.CLcs.LGarXiv:2309.12342v22023
  34. Generative AI for Programming Education: Benchmarking ChatGPT, GPT-4, and Human Tutors

    Tung Phung, Victor-Alexandru Pădurean, José Cambronero +5

    cs.CYcs.AIcs.CLarXiv:2306.17156v32023
  35. K/V-Cache Interventions Dissociate Representation Alignment from Persona Expression in Decoder-Only Language Models

    Yu Sun, Mengyin Lu, Cong Feng +2

    cs.CLarXiv:2609.11020v12026
  36. Ground-Truth Labels Matter: A Deeper Look into Input-Label Demonstrations

    Kang Min Yoo, Junyeob Kim, Hyuhng Joon Kim +5

    cs.CLcs.AIcs.LGarXiv:2205.12685v22022
  37. Walking Down the Memory Maze: Beyond Context Limit through Interactive Reading

    Howard Chen, Ramakanth Pasunuru, Jason Weston +1

    cs.CLarXiv:2310.05029v12023
  38. LLMCarbon: Modeling the end-to-end Carbon Footprint of Large Language Models

    Ahmad Faiz, Sotaro Kaneda, Ruhan Wang +4

    cs.CLcs.AIcs.CYarXiv:2309.14393v22023
  39. Robust Multimodal Sentiment Analysis with Incomplete Modalities via Semantic-aware Completeness based Reconstruction

    Han-Jun Choi, Byunggill Joe, Saim Shin +1

    cs.CLcs.AIcs.LGarXiv:2609.10950v12026
  40. MultiVis-Agent: A Multi-Agent Framework with Logic Rules for Reliable and Comprehensive Cross-Modal Data Visualization

    Jinwei Lu, Yuanfeng Song, Chen Zhang +1

    cs.CLcs.AIcs.DBarXiv:2601.18320v12026
  41. Structurally Speaking: Motif-Oriented Graph Captioning through Bidirectional Graph-Text Translation

    Hsiao-Ying Lu, Dongyu Liu, Kwan-Liu Ma

    cs.CLcs.LGarXiv:2609.10923v12026
  42. Auto-RecSys: Harnessing Autonomous Research Agents for Industry-Scale Recommender System

    Ming Li, Dai Li, Xuying Ning +11

    cs.CLarXiv:2609.10922v12026
  43. Can ChatGPT Replace Traditional KBQA Models? An In-depth Analysis of the Question Answering Performance of the GPT LLM Family

    Yiming Tan, Dehai Min, Yu Li +4

    cs.CLarXiv:2303.07992v32023
  44. SearchAtlas: Analyzing Agentic Search Strategies via Evidential Query Graphs

    Jiacheng Sang, Mengyuan Li, Sanxing Chen +3

    cs.CLarXiv:2609.10901v12026
  45. Retrieving Multimodal Information for Augmented Generation: A Survey

    Ruochen Zhao, Hailin Chen, Weishi Wang +8

    cs.CLarXiv:2303.10868v32023
  46. EmoLLMs: A Series of Emotional Large Language Models and Annotation Tools for Comprehensive Affective Analysis

    Zhiwei Liu, Kailai Yang, Tianlin Zhang +2

    cs.CLarXiv:2401.08508v22024
  47. NumGLUE: A Suite of Fundamental yet Challenging Mathematical Reasoning Tasks

    Swaroop Mishra, Arindam Mitra, Neeraj Varshney +4

    cs.CLcs.AIcs.LGarXiv:2204.05660v12022
  48. SCOTT: Self-Consistent Chain-of-Thought Distillation

    Peifeng Wang, Zhengyang Wang, Zheng Li +3

    cs.CLarXiv:2305.01879v42023
  49. Revisiting Zeroth-Order Optimization for Memory-Efficient LLM Fine-Tuning: A Benchmark

    Yihua Zhang, Pingzhi Li, Junyuan Hong +10

    cs.LGcs.CLarXiv:2402.11592v32024
  50. Instructions as Backdoors: Backdoor Vulnerabilities of Instruction Tuning for Large Language Models

    Jiashu Xu, Mingyu Derek Ma, Fei Wang +2

    cs.CLcs.AIcs.CRarXiv:2305.14710v22023
  51. CoDAR: Continuous Diffusion Language Models are More Powerful Than You Think

    Junzhe Shen, Jieru Zhao, Ziwei He +1

    cs.CLcs.AIcs.LGarXiv:2603.02547v12026
  52. LLM-Anchored Paralinguistic Enrichment for Alzheimer's Disease Detection

    Xiao Wei, Yuqin Lin, Yaru Cao +6

    cs.CLcs.SDarXiv:2609.10896v12026
  53. Let's Sample Step by Step: Adaptive-Consistency for Efficient Reasoning and Coding with LLMs

    Pranjal Aggarwal, Aman Madaan, Yiming Yang +1

    cs.CLarXiv:2305.11860v22023
  54. Does Linguistic Structure Enrichment Enhance Coherence Assessment? Not With Current Architectures

    Victor Mazzotti, Luiz Pereira, Marina Bitencourt dos Santos +4

    cs.CLcs.AIarXiv:2609.10893v12026
  55. Detectable Only Where It Is Confounded: What Verified Duplication Counts Say About Membership Evidence in Language Models

    Arman Nik Khah

    cs.CLcs.CRcs.LGarXiv:2609.10830v12026
  56. MMedAgent: Learning to Use Medical Tools with Multi-modal Agent

    Binxu Li, Tiankai Yan, Yuanting Pan +8

    cs.CLcs.AIarXiv:2407.02483v22024
  57. Larger Context Window, Fewer Overcorrections: Optimizing Prompts and Batching for Minimal-Edit Grammatical Error Correction

    Kateryna Karpo, Artem Chernodub

    cs.CLarXiv:2609.10810v12026
  58. Analyzing Traditional and Neural Approaches to Multilingual Readability Assessment

    Joshua Wong, Chris Tanner

    cs.CLarXiv:2609.10792v12026
  59. Generating Benchmarks for Factuality Evaluation of Language Models

    Dor Muhlgay, Ori Ram, Inbal Magar +7

    cs.CLcs.AIarXiv:2307.06908v22023
  60. Multilingual in Name Only? Cultural and Linguistic Weaknesses of LLMs in Urdu

    Farah Adeeba, Abdul Rafae Khan, Rajesh Bhatt +1

    cs.CLcs.AIcs.LGarXiv:2609.10758v12026