Computation and Language
Papers filed under cs.CL on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
10,561 to 10,620 of 11,240
Computational Orientalism: Measuring Structural Discourse Bias in Large Language Models Using the Middle East Cultural Sensitivity Score (MECSS)
Maha Shahid
cs.CLcs.AIcs.CYarXiv:2608.18100v12026Fractional Decay KV-Cache: Ownership-Aware Memory Management for Improved Inference Relevancy in Dialog Systems
Sukanta Ganguly
cs.CLcs.AIarXiv:2608.18098v12026Longformer: The Long-Document Transformer
Iz Beltagy, Matthew E. Peters, Arman Cohan
cs.CLarXiv:2004.05150v22020NE-BERT: A Multilingual Language Model for Nine Northeast Indian Languages
Badal Nyalang
cs.CLcs.AIarXiv:2608.18094v12026Abliteration Mitigation via Refusal Aliases
Nathan Truong
cs.CLcs.AIcs.CRarXiv:2608.18093v12026Nine Emotion Centroids: A Label-Free Valence Axis That Transfers Across Four Modalities
Yousef Radwan
cs.CLcs.AIarXiv:2608.18090v12026SuTRA : Structurally-Unified Tokenization with Root Awareness
Vaibhav Rathore, Siddhant Gole, Dadhichi Telwadkar +4
cs.CLcs.AIarXiv:2608.18087v12026Adaptive Memory and Reflection Multi-Agent System for Medical Question Answering
Pradeep Murugesan, Luoxiao Yang, Xueli Chen +1
cs.AIcs.CLcs.MAarXiv:2608.19029v12026Metrics That Write Themselves: Evolving an Evaluator from Its Own Blind Spots
Xing Zhang, Yanwei Cui, Guanghui Wang +2
cs.AIcs.CLcs.SEarXiv:2608.18744v12026Can a Lightweight Multimodal Model Estimate LLM Reasoning Performance? A Study for Compute-Optimal Document Inference
Zishan Ahmad, Vishal Vaddina
cs.AIcs.CLarXiv:2608.18591v12026Character-level Convolutional Networks for Text Classification
Xiang Zhang, Junbo Zhao, Yann LeCun
cs.LGcs.CLarXiv:1509.01626v32015Google's Neural Machine Translation System: Bridging the Gap between Human and Machine Translation
Yonghui Wu, Mike Schuster, Zhifeng Chen +28
cs.CLcs.AIcs.LGarXiv:1609.08144v22016Self-Consistency Improves Chain of Thought Reasoning in Language Models
Xuezhi Wang, Jason Wei, Dale Schuurmans +5
cs.CLcs.AIarXiv:2203.11171v42022On the Properties of Neural Machine Translation: Encoder-Decoder Approaches
Kyunghyun Cho, Bart van Merrienboer, Dzmitry Bahdanau +1
cs.CLstat.MLarXiv:1409.1259v22014ALBERT: A Lite BERT for Self-supervised Learning of Language Representations
Zhenzhong Lan, Mingda Chen, Sebastian Goodman +3
cs.CLcs.AIarXiv:1909.11942v62019Efficient Adaptation of LLMs for Hate Speech Detection in Low-Resource Languages: A Comparative Study on Roman Urdu
Toneema Zubair, Muhammad Junaid Asif, Faisal Kamiran +2
cs.AIcs.CLarXiv:2608.18142v12026Neural Machine Translation of Rare Words with Subword Units
Rico Sennrich, Barry Haddow, Alexandra Birch
cs.CLarXiv:1508.07909v52015Speech Recognition with Deep Recurrent Neural Networks
Alex Graves, Abdel-rahman Mohamed, Geoffrey Hinton
cs.NEcs.CLarXiv:1303.5778v12013PaLM: Scaling Language Modeling with Pathways
Aakanksha Chowdhery, Sharan Narang, Jacob Devlin +64
cs.CLarXiv:2204.02311v52022Safety Alignment Illusion: The Cross-Lingual Safety Gap in LLMs
Namya Bhatnagar
cs.AIcs.CLcs.CYarXiv:2608.18131v12026Effective Approaches to Attention-based Neural Machine Translation
Minh-Thang Luong, Hieu Pham, Christopher D. Manning
cs.CLarXiv:1508.04025v52015GLUE: A Multi-Task Benchmark and Analysis Platform for Natural Language Understanding
Alex Wang, Amanpreet Singh, Julian Michael +3
cs.CLarXiv:1804.07461v32018Unsupervised Cross-lingual Representation Learning at Scale
Alexis Conneau, Kartikay Khandelwal, Naman Goyal +7
cs.CLarXiv:1911.02116v22019Measuring Massive Multitask Language Understanding
Dan Hendrycks, Collin Burns, Steven Basart +4
cs.CYcs.AIcs.CLarXiv:2009.03300v32020BERTScore: Evaluating Text Generation with BERT
Tianyi Zhang, Varsha Kishore, Felix Wu +2
cs.CLarXiv:1904.09675v32019VIA-SD: Verification via Intra-Model Routing for Speculative Decoding
Yuchen Xian, Yang He, Yunqiu Xu +1
cs.CLcs.AIarXiv:2606.12243v12026Fine-tuning Multi-modal LLMs with ART: Art-based Reinforcement Training
Michal Chudoba, Sergey Alyaev, Petra Galuscakova +1
cs.LGcs.AIcs.CLarXiv:2606.11854v12026XLNet: Generalized Autoregressive Pretraining for Language Understanding
Zhilin Yang, Zihang Dai, Yiming Yang +3
cs.CLcs.LGarXiv:1906.08237v22019When is Your LLM Steerable?
Chenrui Fan, Yize Cheng, Ming Li +2
cs.CLcs.LGarXiv:2606.11599v12026Unstable Features, Reproducible Subspaces: Understanding Seed Dependence in Sparse Autoencoders
Gleb Gerasimov, Timofei Rusalev, Nikita Balagansky +3
cs.LGcs.AIcs.CLarXiv:2606.12138v12026Distributed Representations of Sentences and Documents
Quoc V. Le, Tomas Mikolov
cs.CLcs.AIcs.LGarXiv:1405.4053v22014Getting Better at Working With You: Compiling User Corrections into Runtime Enforcement for Coding Agents
Yujun Zhou, Kehan Guo, Haomin Zhuang +8
cs.LGcs.CLarXiv:2606.13174v12026Training Verifiers to Solve Math Word Problems
Karl Cobbe, Vineet Kosaraju, Mohammad Bavarian +9
cs.LGcs.CLarXiv:2110.14168v22021DistilBERT, a distilled version of BERT: smaller, faster, cheaper and lighter
Victor Sanh, Lysandre Debut, Julien Chaumond +1
cs.CLarXiv:1910.01108v42019Judging LLM-as-a-Judge with MT-Bench and Chatbot Arena
Lianmin Zheng, Wei-Lin Chiang, Ying Sheng +10
cs.CLcs.AIarXiv:2306.05685v42023Visual Instruction Tuning
Haotian Liu, Chunyuan Li, Qingyang Wu +1
cs.CVcs.AIcs.CLarXiv:2304.08485v22023Enriching Word Vectors with Subword Information
Piotr Bojanowski, Edouard Grave, Armand Joulin +1
cs.CLcs.LGarXiv:1607.04606v22016BART: Denoising Sequence-to-Sequence Pre-training for Natural Language Generation, Translation, and Comprehension
Mike Lewis, Yinhan Liu, Naman Goyal +5
cs.CLcs.LGstat.MLarXiv:1910.13461v12019Summaries:한국어Deep contextualized word representations
Matthew E. Peters, Mark Neumann, Mohit Iyyer +4
cs.CLarXiv:1802.05365v22018Convolutional Neural Networks for Sentence Classification
Yoon Kim
cs.CLcs.NEarXiv:1408.5882v22014The Llama 3 Herd of Models
Aaron Grattafiori, Abhimanyu Dubey, Abhinav Jauhri +558
cs.AIcs.CLcs.CVarXiv:2407.21783v32024Sentence-BERT: Sentence Embeddings using Siamese BERT-Networks
Nils Reimers, Iryna Gurevych
cs.CLarXiv:1908.10084v12019Intent-Driven Dynamic Chunking: Segmenting Documents to Reflect Predicted Information Needs
Christos Koutsiaris
cs.IRcs.AIcs.CLarXiv:2602.14784v12026Auxiliary uncertainty signals for LLM-assisted systematic review screening: a benchmark across eight Cohen drug-class reviews
Arya Rahgozar, Pouria Mortezaagha
cs.CLcs.DLcs.IRarXiv:2608.14551v12026Learning Phrase Representations using RNN Encoder-Decoder for Statistical Machine Translation
Kyunghyun Cho, Bart van Merrienboer, Caglar Gulcehre +4
cs.CLcs.LGcs.NEarXiv:1406.1078v32014Notes2Skills: From Lab Notebooks to Certainty-Aware Scientific Agent Skills
Shi Liu, Jiayao Chen, Chengwei Qin +3
cs.CLarXiv:2606.11897v12026Which Source Wins? Task-Dependent Reliance in Vision-Language Models
Rodela Ghosh, Aviral Gupta, Guangjing Wang
cs.CLcs.CVarXiv:2608.17205v12026BengaliMCQ: Automatic Generation and Answer Prediction of Academic Multiple-Choice Questions in a Low-Resource Language
Abu Tarabin Surzo, A. K. M. Nihalul Kabir, Sm Azmain Faysal +3
cs.CLarXiv:2608.15547v12026Distributed Representations of Words and Phrases and their Compositionality
Tomas Mikolov, Ilya Sutskever, Kai Chen +2
cs.CLcs.LGstat.MLarXiv:1310.4546v12013Thinking in a Low-Resource Language: What SFT Builds, What RL Fixes, What Accuracy Cannot See
Ayoub Kirouane, Christos Petrocheilos
cs.CLcs.LGcs.ROarXiv:2608.17744v12026HarmProfile: Characterizing Harmful Distributions in Frontier LLMs
Zhouyuan Ma, Yutao Wu, Hanxun Huang +6
cs.CLcs.AIarXiv:2608.14577v12026Harness the Memory: A Holistic Evaluation of Memory Substrates in Memory Agents
Wei-Chieh Huang, Weizhi Zhang, Yuchen Wu +12
cs.CLarXiv:2608.15008v12026Building Social World Models with Large Language Models
Haofei Yu, Yining Zhao, Guanyu Lin +1
cs.SIcs.CLarXiv:2606.11482v12026Beyond Single Object: Learning 3D Relations with Large Language Models
Kohsuke Ide, Ryousuke Yamada, Yue Qiu +4
cs.CVcs.AIcs.CLarXiv:2608.15710v12026Do Language Models Consistently Encode the Current Year?
Suze van Adrichem, Aditi Bhaskar, Diyi Yang +2
cs.CLcs.LGarXiv:2608.15507v12026More Context, Larger Models, or Moral Knowledge? A Systematic Study of Schwartz Value Detection in Political Texts
Víctor Yeste, Paolo Rosso
cs.CLcs.AIcs.LGarXiv:2605.22641v32026Large Language Models as Implicit Sociological Models: Reconstructing Voting Behaviour from Sociodemographic Profiles
Roman Neruda, Martin Bakoš, Josef Šlerka +3
cs.CYcs.CLcs.LGarXiv:2608.15871v12026Polaris: Learning to Generate Table Descriptions from Retrieval Feedback
Ting Cai, Tuan Minh Phan, AnHai Doan
cs.CLcs.DBarXiv:2608.17171v12026Connect the Dots: Training LLMs for Long-Lifecycle Agents with Cross-Domain Generalization Via Reinforcement Learning
Yanxi Chen, Weijie Shi, Yuexiang Xie +4
cs.LGcs.AIcs.CLarXiv:2606.20002v12026The Commercial Tax: Rent-vs-Own Blind Spots in Multi-Hop Retrieval Benchmarks
Luis M. Sanchez, Kosrow Dehnad
cs.IRcs.CLarXiv:2608.16096v12026