Computation and Language
Papers filed under cs.CL on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
901 to 960 of 11,220
Natural SQL: Making SQL Easier to Infer from Natural Language Specifications
Yujian Gan, Xinyun Chen, Jinxia Xie +4
cs.CLarXiv:2109.05153v12021Performance-Efficiency Trade-offs in Unsupervised Pre-training for Speech Recognition
Felix Wu, Kwangyoun Kim, Jing Pan +3
cs.CLcs.LGcs.SDarXiv:2109.06870v12021ReGround: Grounding Reviewer Comments in Multimodal Evidence
Serwar Basch, Lizhen Qu, Iryna Gurevych
cs.CLcs.IRarXiv:2609.11460v12026FaithDial: A Faithful Benchmark for Information-Seeking Dialogue
Nouha Dziri, Ehsan Kamalloo, Sivan Milton +4
cs.CLarXiv:2204.10757v32022Cross-Lingual Clinical Annotation Projection as Constrained Text Generation: A Six-Language Study
Álvaro Rey-Blanes, Francisco J. Moreno-Barea, Francisco J. Veredas
cs.CLcs.AIarXiv:2609.11450v12026On the Impact of Anonymization on the Performance of Large Language Models
Tobias Deußer, Max Hahnbück, Lorenz Sparrenberg +3
cs.CLcs.AIarXiv:2609.11335v12026MultiHuSE: A Multimodal Dataset for Humour Styles and Emotions
Mary Ogbuka Kenneth, Foaad Khosmood, Abbas Edalat
cs.CLcs.CVcs.MMarXiv:2609.11322v12026Automatic Lyric Transcription for Greek Songs: Scaling and Task Composition Effects in Whisper Adaptation
Maria Frangiadaki, Dimitrios Damianos, Kosmas Kritsis +1
cs.CLcs.SDarXiv:2609.11302v12026The Illusion of Balanced Multimodal Sentiment Analysis: Beyond the Limits of Optimization-Based Methods
Ioanna Kaffeza, Efthymios Georgiou, Alexandros Potamianos
cs.CLarXiv:2609.11247v12026Mixture-of-Transformers: A Sparse and Scalable Architecture for Multi-Modal Foundation Models
Weixin Liang, Lili Yu, Liang Luo +8
cs.CLarXiv:2411.04996v22024Can LLMs Normalize Databases? A Benchmark and Multi-Agent Framework for Schema Normalization
Dong-Jae Koh, Huisu Kim, SeongHwan Yoon +4
cs.CLarXiv:2609.11141v12026Truncation Sampling as Language Model Desmoothing
John Hewitt, Christopher D. Manning, Percy Liang
cs.CLarXiv:2210.15191v12022A Fragility Spectrum for Recursive Language-Model Training
Yangze Liu, Zhongyi Han
cs.CLcs.AIcs.LGarXiv:2609.11149v12026Overview of the NLPCC 2026 Shared Task 11: Agent-Based Experiment Reproduction from Scientific Papers
Hanhua Hong, Yizhi Li, Luu Gia Huy +3
cs.CLarXiv:2609.11117v12026Assessing the Reusability of Public Speech Resources for Low-Resource Languages: A Central Kurdish Case Study
Hiwa Asadpour
cs.CLarXiv:2609.11246v12026HittER: Hierarchical Transformers for Knowledge Graph Embeddings
Sanxing Chen, Xiaodong Liu, Jianfeng Gao +3
cs.CLcs.LGarXiv:2008.12813v22020OmniHallu: Unified Hallucination Detection for Cross-Modal Comprehension and Generation in Multimodal Large Language Models
Jianjiang Yang, Peihang Li, Shanqing Xu +3
cs.CLcs.CVarXiv:2609.11244v12026Does Neural Machine Translation Benefit from Larger Context?
Sebastien Jean, Stanislas Lauly, Orhan Firat +1
stat.MLcs.CLcs.LGarXiv:1704.05135v12017Automated Identification of Competing Narratives in Political Discourse on Social Media
Sergej Wildemann, Erick Elejalde
cs.CLcs.SIarXiv:2609.11202v12026FlexComp: One Model for Every Ratio in Context Compression
Kaiyan Zhao, Zhongtao Miao, Akiko Aizawa +1
cs.CLarXiv:2609.11192v12026Have Faith in Faithfulness: Going Beyond Circuit Overlap When Finding Model Mechanisms
Michael Hanna, Sandro Pezzelle, Yonatan Belinkov
cs.LGcs.CLarXiv:2403.17806v22024Rubric-Aligned Disentangled Evaluation of Human Simultaneous Interpreting
Ziyu Zhang, Satoshi Nakamura
cs.CLarXiv:2609.11131v12026From Repetition to Recognition: Inductive Discovery of Disinformation Narratives
Max Upravitelev, Veronika Solopova, Jing Yang +4
cs.CLarXiv:2609.11128v12026SentiBERT: A Transferable Transformer-Based Architecture for Compositional Sentiment Semantics
Da Yin, Tao Meng, Kai-Wei Chang
cs.CLarXiv:2005.04114v42020ProMediConv: Benchmarking Proactive Conversational Agents in Legal Dispute Mediation
Zesheng Wei, Mengfan Li, Wenhao Liu +3
cs.CLarXiv:2609.11101v12026Distribution-aware Language Neuron Identification in Multilingual Large Language Models
Minjun Kim, Inho Won, Junghun Yuk +3
cs.CLarXiv:2609.10993v12026Rebalancing Token Importance in Language Models with TF-IDF Weighted Cross-Entropy Loss
Zhijian Li, Stefan Larson, Kevin Leach
cs.CLcs.LGarXiv:2609.11029v12026Beyond Consensus: Perspectivist Modeling and Evaluation of Annotator Disagreement in NLP
Yinuo Xu, David Jurgens
cs.CLarXiv:2601.09065v22026Rethinking Verbalized Confidence for LLM-as-a-Judge: A Compatibility Shift on Post-2025 Proprietary Models
Yu-Chung Hsiao
cs.CLarXiv:2609.10996v12026Using Semantic Uncertainty to Estimate Transition Relevance in Turn-taking
Muhammad Umair, Jan P. de Ruiter
cs.CLarXiv:2609.10934v12026Weakly Supervised Cross-Lingual Named Entity Recognition via Effective Annotation and Representation Projection
Jian Ni, Georgiana Dinu, Radu Florian
cs.CLcs.IRarXiv:1707.02483v12017When Noise Fabricates Bias: The Fragility of LLM-as-a-Judge Bias Measurement under Noisy Text
DongHyun Ryu, Jaehyeok Lee, YeongJun Hwang +1
cs.CLcs.LGarXiv:2609.11067v12026Cultural Alignment in Large Language Models: An Explanatory Analysis Based on Hofstede's Cultural Dimensions
Reem I. Masoud, Ziquan Liu, Martin Ferianc +2
cs.CYcs.CLcs.LGarXiv:2309.12342v22023Generative AI for Programming Education: Benchmarking ChatGPT, GPT-4, and Human Tutors
Tung Phung, Victor-Alexandru Pădurean, José Cambronero +5
cs.CYcs.AIcs.CLarXiv:2306.17156v32023K/V-Cache Interventions Dissociate Representation Alignment from Persona Expression in Decoder-Only Language Models
Yu Sun, Mengyin Lu, Cong Feng +2
cs.CLarXiv:2609.11020v12026Ground-Truth Labels Matter: A Deeper Look into Input-Label Demonstrations
Kang Min Yoo, Junyeob Kim, Hyuhng Joon Kim +5
cs.CLcs.AIcs.LGarXiv:2205.12685v22022Walking Down the Memory Maze: Beyond Context Limit through Interactive Reading
Howard Chen, Ramakanth Pasunuru, Jason Weston +1
cs.CLarXiv:2310.05029v12023LLMCarbon: Modeling the end-to-end Carbon Footprint of Large Language Models
Ahmad Faiz, Sotaro Kaneda, Ruhan Wang +4
cs.CLcs.AIcs.CYarXiv:2309.14393v22023Robust Multimodal Sentiment Analysis with Incomplete Modalities via Semantic-aware Completeness based Reconstruction
Han-Jun Choi, Byunggill Joe, Saim Shin +1
cs.CLcs.AIcs.LGarXiv:2609.10950v12026MultiVis-Agent: A Multi-Agent Framework with Logic Rules for Reliable and Comprehensive Cross-Modal Data Visualization
Jinwei Lu, Yuanfeng Song, Chen Zhang +1
cs.CLcs.AIcs.DBarXiv:2601.18320v12026Structurally Speaking: Motif-Oriented Graph Captioning through Bidirectional Graph-Text Translation
Hsiao-Ying Lu, Dongyu Liu, Kwan-Liu Ma
cs.CLcs.LGarXiv:2609.10923v12026Auto-RecSys: Harnessing Autonomous Research Agents for Industry-Scale Recommender System
Ming Li, Dai Li, Xuying Ning +11
cs.CLarXiv:2609.10922v12026Can ChatGPT Replace Traditional KBQA Models? An In-depth Analysis of the Question Answering Performance of the GPT LLM Family
Yiming Tan, Dehai Min, Yu Li +4
cs.CLarXiv:2303.07992v32023SearchAtlas: Analyzing Agentic Search Strategies via Evidential Query Graphs
Jiacheng Sang, Mengyuan Li, Sanxing Chen +3
cs.CLarXiv:2609.10901v12026Retrieving Multimodal Information for Augmented Generation: A Survey
Ruochen Zhao, Hailin Chen, Weishi Wang +8
cs.CLarXiv:2303.10868v32023EmoLLMs: A Series of Emotional Large Language Models and Annotation Tools for Comprehensive Affective Analysis
Zhiwei Liu, Kailai Yang, Tianlin Zhang +2
cs.CLarXiv:2401.08508v22024NumGLUE: A Suite of Fundamental yet Challenging Mathematical Reasoning Tasks
Swaroop Mishra, Arindam Mitra, Neeraj Varshney +4
cs.CLcs.AIcs.LGarXiv:2204.05660v12022SCOTT: Self-Consistent Chain-of-Thought Distillation
Peifeng Wang, Zhengyang Wang, Zheng Li +3
cs.CLarXiv:2305.01879v42023Revisiting Zeroth-Order Optimization for Memory-Efficient LLM Fine-Tuning: A Benchmark
Yihua Zhang, Pingzhi Li, Junyuan Hong +10
cs.LGcs.CLarXiv:2402.11592v32024Instructions as Backdoors: Backdoor Vulnerabilities of Instruction Tuning for Large Language Models
Jiashu Xu, Mingyu Derek Ma, Fei Wang +2
cs.CLcs.AIcs.CRarXiv:2305.14710v22023CoDAR: Continuous Diffusion Language Models are More Powerful Than You Think
Junzhe Shen, Jieru Zhao, Ziwei He +1
cs.CLcs.AIcs.LGarXiv:2603.02547v12026LLM-Anchored Paralinguistic Enrichment for Alzheimer's Disease Detection
Xiao Wei, Yuqin Lin, Yaru Cao +6
cs.CLcs.SDarXiv:2609.10896v12026Let's Sample Step by Step: Adaptive-Consistency for Efficient Reasoning and Coding with LLMs
Pranjal Aggarwal, Aman Madaan, Yiming Yang +1
cs.CLarXiv:2305.11860v22023Does Linguistic Structure Enrichment Enhance Coherence Assessment? Not With Current Architectures
Victor Mazzotti, Luiz Pereira, Marina Bitencourt dos Santos +4
cs.CLcs.AIarXiv:2609.10893v12026Detectable Only Where It Is Confounded: What Verified Duplication Counts Say About Membership Evidence in Language Models
Arman Nik Khah
cs.CLcs.CRcs.LGarXiv:2609.10830v12026MMedAgent: Learning to Use Medical Tools with Multi-modal Agent
Binxu Li, Tiankai Yan, Yuanting Pan +8
cs.CLcs.AIarXiv:2407.02483v22024Larger Context Window, Fewer Overcorrections: Optimizing Prompts and Batching for Minimal-Edit Grammatical Error Correction
Kateryna Karpo, Artem Chernodub
cs.CLarXiv:2609.10810v12026Analyzing Traditional and Neural Approaches to Multilingual Readability Assessment
Joshua Wong, Chris Tanner
cs.CLarXiv:2609.10792v12026Generating Benchmarks for Factuality Evaluation of Language Models
Dor Muhlgay, Ori Ram, Inbal Magar +7
cs.CLcs.AIarXiv:2307.06908v22023Multilingual in Name Only? Cultural and Linguistic Weaknesses of LLMs in Urdu
Farah Adeeba, Abdul Rafae Khan, Rajesh Bhatt +1
cs.CLcs.AIcs.LGarXiv:2609.10758v12026