Computation and Language
Papers filed under cs.CL on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
8,401 to 8,460 of 11,242
Evaluating Language Models on Cross-Language Code Functional Equivalence
Hui Sun, Anderson Uchôa, Rohit Gheyi +1
cs.SEcs.AIcs.CLarXiv:2608.23961v12026Investigating Knowledge Transfer Across Interactive Dialogue Games
Filippo Momentè, Mir Nafis Sharear Shopnil, Andrea de Varda +5
cs.CLarXiv:2608.23969v12026Preference Data Selection for Mitigating the Alignment Tax in Large Language Models
Minsu Kim, Jianxun Lian, Xing Xie +1
cs.AIcs.CLarXiv:2608.24192v12026Density-aware Soft Context Compression with Semi-Dynamic Compression Ratio
Yijiong Yu, Shuai Yuan, Jie Zheng +2
cs.CLarXiv:2603.25926v12026TAPS: Task Aware Proposal Distributions for Speculative Sampling
Mohamad Zbib, Mohamad Bazzi, Ammar Mohanna +2
cs.CLcs.AIarXiv:2603.27027v12026Explainable Prediction of Medical Codes from Clinical Text
James Mullenbach, Sarah Wiegreffe, Jon Duke +2
cs.CLcs.LGstat.MLarXiv:1802.05695v22018RLAIF vs. RLHF: Scaling Reinforcement Learning from Human Feedback with AI Feedback
Harrison Lee, Samrat Phatale, Hassan Mansoor +8
cs.CLcs.AIcs.LGarXiv:2309.00267v32023DamageScope: Vision-Language Retrieval at Scale for Disaster Damage Assessment from Satellite Imagery
Ravi K. Rajendran, Biplob Debnath, Murugan Sankaradas +1
cs.CVcs.CLcs.IRarXiv:2608.21529v12026LëtzCross: A Cross-Lingual Page-Level Benchmark for Multimodal Retrieval over Luxembourgish Documents
Omar El Bachyr, Fred Philippy, Laura Maria Bernardy +3
cs.CLarXiv:2608.21714v12026Can MLLMs Read Students' Minds? Unpacking Multimodal Error Analysis in Handwritten Math
Dingjie Song, Tianlong Xu, Yi-Fan Zhang +6
cs.AIcs.CLcs.CVarXiv:2603.24961v12026Good Debt or Bad Debt: Detecting Semantic Orientations in Economic Texts
Pekka Malo, Ankur Sinha, Pyry Takala +2
cs.CLcs.IRq-fin.CParXiv:1307.5336v22013Working Notes on Late Interaction Dynamics: Analyzing Targeted Behaviors of Late Interaction Models
Antoine Edy, Max Conti, Quentin Macé
cs.IRcs.AIcs.CLarXiv:2603.26259v22026HiDiffTIR: Hierarchical Difficulty-Aware Policy Optimization for Multi-Turn Tool-Integrated Reasoning
Yucan Guo, Xiaohan Wang, Miao Su +8
cs.CLcs.AIarXiv:2608.21863v12026Text Data Integration
Md Ataur Rahman, Dimitris Sacharidis, Oscar Romero +1
cs.CLcs.IRarXiv:2603.27055v12026Story2Proposal: A Scaffold for Structured Scientific Paper Writing
Zhuoyang Qian, Wei Shi, Xu Lin +19
cs.CLarXiv:2603.27065v12026How is ChatGPT's behavior changing over time?
Lingjiao Chen, Matei Zaharia, James Zou
cs.CLcs.AIcs.LGarXiv:2307.09009v32023daVinci-LLM:Towards the Science of Pretraining
Yiwei Qin, Yixiu Liu, Tiantian Mi +12
cs.AIcs.CLarXiv:2603.27164v12026A Unified MRC Framework for Named Entity Recognition
Xiaoya Li, Jingrong Feng, Yuxian Meng +4
cs.CLarXiv:1910.11476v92019Taming Visual Neglect: A Variational Information Bottleneck Framework for Adaptive Attention in Multimodal In-Context Learning
Kaito Tanaka, Yuji Nishimura, Keisuke Matsuda +1
cs.CLarXiv:2608.23570v12026When and why vision-language models behave like bags-of-words, and what to do about it?
Mert Yuksekgonul, Federico Bianchi, Pratyusha Kalluri +2
cs.CVcs.AIcs.CLarXiv:2210.01936v32022LongVideoBench: A Benchmark for Long-context Interleaved Video-Language Understanding
Haoning Wu, Dongxu Li, Bei Chen +1
cs.CVcs.CLcs.LGarXiv:2407.15754v12024Distilling Human-Aligned Privacy Sensitivity Assessment from Large Language Models
Gabriel Loiseau, Damien Sileo, Damien Riquet +2
cs.CLarXiv:2603.29497v12026CALVIN: A Benchmark for Language-Conditioned Policy Learning for Long-Horizon Robot Manipulation Tasks
Oier Mees, Lukas Hermann, Erick Rosete-Beas +1
cs.ROcs.AIcs.CLarXiv:2112.03227v42021KSE-Web: An Analysis of Hybrid Retrieval and LLM-Assisted Query Expansion for Low-Resource Khmer Semantic Search
Nimol Thuon
cs.CLcs.AIcs.IRarXiv:2608.21365v12026An Empirical Evaluation of doc2vec with Practical Insights into Document Embedding Generation
Jey Han Lau, Timothy Baldwin
cs.CLarXiv:1607.05368v12016Semantics or Structure? Auditing Text Sensitivity in Multimodal Time-Series Forecasting
Karthik Sridhar, Atharva Gupta, Nishant Pradhan +3
cs.CLarXiv:2608.22321v12026Joint Extraction of Entities and Relations Based on a Novel Tagging Scheme
Suncong Zheng, Feng Wang, Hongyun Bao +3
cs.CLcs.AIcs.LGarXiv:1706.05075v12017Text-Anchored Semantic Perturbations for Transferable Jailbreak Attacks on Multimodal Large Language Models
Wenyun Li, Guiping Cao, Xiangyuan Lan +1
cs.CLarXiv:2608.22312v12026Distinguishing Revision and Delayed Elaboration in Incremental Narrative Interpretation
Yi-Chun Chen
cs.CLcs.AIcs.MMarXiv:2608.21364v12026Contrastive Decoding: Open-ended Text Generation as Optimization
Xiang Lisa Li, Ari Holtzman, Daniel Fried +5
cs.CLcs.AIcs.LGarXiv:2210.15097v22022GUI-Primitives: Diagnosing Spatial Reasoning Failures in Vision-Language GUI Grounding
Md Abrar Jahin, Md Rizwan Parvez
cs.CLarXiv:2608.21832v12026Wazobia Eval: A Benchmark for Nigerian Pidgin Emotion Understanding, Sarcasm Detection, and Cultural Reasoning
Stephanie Okoye
cs.CLcs.AIarXiv:2608.21369v12026ARBERT & MARBERT: Deep Bidirectional Transformers for Arabic
Muhammad Abdul-Mageed, AbdelRahim Elmadany, El Moatez Billah Nagoudi
cs.CLarXiv:2101.01785v32020A Multiscale Visualization of Attention in the Transformer Model
Jesse Vig
cs.HCcs.CLcs.LGarXiv:1906.05714v12019The Collaboration Tax: How Much LLM Multi-Agent Systems Pay to Coordinate
Weixiang Sun, Zehong Wang, Hong Huang +2
cs.CLarXiv:2608.22152v12026GeoRisk-RAG: A Hierarchy-Aware Risk Framework for Improving RAG Reliability through Selective Answering
Meenu Ravi, Shailik Sarkar, Lulwah AlKulaib +2
cs.CLcs.AIarXiv:2608.22634v12026Agentic Scaffolding Amplifies Sycophantic Behavior in Large Language Models
Thantham Jittham
cs.CLcs.AIcs.LGarXiv:2608.21377v12026PersonaMem-v3: Toward Omni-Platform Personal Intelligence for Holistic User Understanding, Recommendation, and Agentic Tasks
Bowen Jiang, Yuan Yuan, Zhuoqun Hao +11
cs.CYcs.CLarXiv:2608.21381v12026A Social Media Analysis of Discourse on the Israel--Palestine Conflict on Telegram
Michail Zafeiropoulos, Despoina Antonakaki, Sotiris Ioannidis
cs.CLcs.AIcs.CYarXiv:2608.21385v12026Beyond Two Bytes per Letter: Tokenization Overhead in Cyrillic AI Systems
Ivan Dobrovolskyi
cs.CLcs.AIarXiv:2608.21384v12026SpeechGPT: Empowering Large Language Models with Intrinsic Cross-Modal Conversational Abilities
Dong Zhang, Shimin Li, Xin Zhang +4
cs.CLarXiv:2305.11000v22023Phrase-Based & Neural Unsupervised Machine Translation
Guillaume Lample, Myle Ott, Alexis Conneau +2
cs.CLarXiv:1804.07755v22018A Rising Tide Lifts All Boats: MTQE Rewards for Idioms Improve General Translation Quality
Ishika Agarwal, Zhenlin He, Dhruva Patil +1
cs.CLarXiv:2601.06307v12026EpiCaR: Knowing What You Don't Know Matters for Better Reasoning in LLMs
Jewon Yeom, Jaewon Sok, Seonghyeon Park +2
cs.CLarXiv:2601.06786v12026Architecture as Capability Equalizer for Coding Agents
Arquimedes Canedo
cs.SEcs.AIcs.CLarXiv:2608.21747v12026Large Language Models Struggle to Learn Long-Tail Knowledge
Nikhil Kandpal, Haikang Deng, Adam Roberts +2
cs.CLcs.LGarXiv:2211.08411v22022Aligning Text, Code, and Vision: A Multi-Objective Reinforcement Learning Framework for Text-to-Visualization
Mizanur Rahman, Mohammed Saidul Islam, Md Tahmid Rahman Laskar +2
cs.CLarXiv:2601.04582v12026User-Oriented Multi-Turn Dialogue Generation with Tool Use at scale
Jungho Cho, Minbyul Jeong, Sungrae Park
cs.CLarXiv:2601.08225v12026sui-1: Grounded and Verifiable Long-Form Summarization
Benedikt Droste, Jan Philipp Harries, Maximilian Idahl +1
cs.CLcs.AIarXiv:2601.08472v12026mPLUG-Owl2: Revolutionizing Multi-modal Large Language Model with Modality Collaboration
Qinghao Ye, Haiyang Xu, Jiabo Ye +7
cs.CLcs.CVarXiv:2311.04257v22023Bulbul: A Dataset for Dialectal Arabic Speech Recognition
Ahmed Ashraf, Aisha Alansari, Fadel Al Abbas +30
cs.CLcs.AIarXiv:2608.21950v12026TabFact: A Large-scale Dataset for Table-based Fact Verification
Wenhu Chen, Hongmin Wang, Jianshu Chen +5
cs.CLcs.AIarXiv:1909.02164v52019Sequential LLM Release Facilitates Manipulation in Regulated Markets
Eilam Shapira, Moshe Tennenholtz, Roi Reichart
cs.GTcs.AIcs.CLarXiv:2601.11496v32026PingPong: A Natural Benchmark for Multi-Turn Code-Switching Dialogues
Mohammad Rifqi Farhansyah, Hanif Muhammad Zhafran, Farid Adilazuarda +6
cs.CLarXiv:2601.17277v12026RIR-Mega-Speech: A Reverberant Speech Corpus with Comprehensive Acoustic Metadata and Reproducible Evaluation
Mandip Goswami
eess.AScs.CLcs.SDarXiv:2601.19949v12026Figurative Justice: Detecting metaphors in Hindi judgements with qualitative assessment and transformers
Bhumika Bhattacharyya, Shouvik Kumar Guha, Indranil Dutta
cs.CLarXiv:2608.22446v12026LIBERTy: A Causal Framework for Benchmarking Concept-Based Explanations of LLMs with Structural Counterfactuals
Gilat Toker, Nitay Calderon, Ohad Amosy +1
cs.CLcs.AIarXiv:2601.10700v22026Register Shifts Break LLM Safety: A Bengali Benchmark with Culturally Grounded Harms
Naymul Islam, Nusrat Jahan Lia, Shubhashis Roy Dipta +2
cs.CLcs.AIarXiv:2608.22335v12026AstroReason-Bench: Evaluating Unified Agentic Planning across Heterogeneous Space Planning Problems
Weiyi Wang, Xinchi Chen, Jingjing Gong +2
cs.AIcs.CLarXiv:2601.11354v12026Scaling Speech Technology to 1,000+ Languages
Vineel Pratap, Andros Tjandra, Bowen Shi +13
cs.CLcs.SDeess.ASarXiv:2305.13516v12023