Computation and Language

Papers filed under cs.CL on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

10,561 to 10,620 of 11,240

  1. Computational Orientalism: Measuring Structural Discourse Bias in Large Language Models Using the Middle East Cultural Sensitivity Score (MECSS)

    Maha Shahid

    cs.CLcs.AIcs.CYarXiv:2608.18100v12026
  2. Fractional Decay KV-Cache: Ownership-Aware Memory Management for Improved Inference Relevancy in Dialog Systems

    Sukanta Ganguly

    cs.CLcs.AIarXiv:2608.18098v12026
  3. Longformer: The Long-Document Transformer

    Iz Beltagy, Matthew E. Peters, Arman Cohan

    cs.CLarXiv:2004.05150v22020
  4. NE-BERT: A Multilingual Language Model for Nine Northeast Indian Languages

    Badal Nyalang

    cs.CLcs.AIarXiv:2608.18094v12026
  5. Abliteration Mitigation via Refusal Aliases

    Nathan Truong

    cs.CLcs.AIcs.CRarXiv:2608.18093v12026
  6. Nine Emotion Centroids: A Label-Free Valence Axis That Transfers Across Four Modalities

    Yousef Radwan

    cs.CLcs.AIarXiv:2608.18090v12026
  7. SuTRA : Structurally-Unified Tokenization with Root Awareness

    Vaibhav Rathore, Siddhant Gole, Dadhichi Telwadkar +4

    cs.CLcs.AIarXiv:2608.18087v12026
  8. Adaptive Memory and Reflection Multi-Agent System for Medical Question Answering

    Pradeep Murugesan, Luoxiao Yang, Xueli Chen +1

    cs.AIcs.CLcs.MAarXiv:2608.19029v12026
  9. Metrics That Write Themselves: Evolving an Evaluator from Its Own Blind Spots

    Xing Zhang, Yanwei Cui, Guanghui Wang +2

    cs.AIcs.CLcs.SEarXiv:2608.18744v12026
  10. Can a Lightweight Multimodal Model Estimate LLM Reasoning Performance? A Study for Compute-Optimal Document Inference

    Zishan Ahmad, Vishal Vaddina

    cs.AIcs.CLarXiv:2608.18591v12026
  11. Character-level Convolutional Networks for Text Classification

    Xiang Zhang, Junbo Zhao, Yann LeCun

    cs.LGcs.CLarXiv:1509.01626v32015
  12. Google's Neural Machine Translation System: Bridging the Gap between Human and Machine Translation

    Yonghui Wu, Mike Schuster, Zhifeng Chen +28

    cs.CLcs.AIcs.LGarXiv:1609.08144v22016
  13. Self-Consistency Improves Chain of Thought Reasoning in Language Models

    Xuezhi Wang, Jason Wei, Dale Schuurmans +5

    cs.CLcs.AIarXiv:2203.11171v42022
  14. On the Properties of Neural Machine Translation: Encoder-Decoder Approaches

    Kyunghyun Cho, Bart van Merrienboer, Dzmitry Bahdanau +1

    cs.CLstat.MLarXiv:1409.1259v22014
  15. ALBERT: A Lite BERT for Self-supervised Learning of Language Representations

    Zhenzhong Lan, Mingda Chen, Sebastian Goodman +3

    cs.CLcs.AIarXiv:1909.11942v62019
  16. Efficient Adaptation of LLMs for Hate Speech Detection in Low-Resource Languages: A Comparative Study on Roman Urdu

    Toneema Zubair, Muhammad Junaid Asif, Faisal Kamiran +2

    cs.AIcs.CLarXiv:2608.18142v12026
  17. Neural Machine Translation of Rare Words with Subword Units

    Rico Sennrich, Barry Haddow, Alexandra Birch

    cs.CLarXiv:1508.07909v52015
  18. Speech Recognition with Deep Recurrent Neural Networks

    Alex Graves, Abdel-rahman Mohamed, Geoffrey Hinton

    cs.NEcs.CLarXiv:1303.5778v12013
  19. PaLM: Scaling Language Modeling with Pathways

    Aakanksha Chowdhery, Sharan Narang, Jacob Devlin +64

    cs.CLarXiv:2204.02311v52022
  20. Safety Alignment Illusion: The Cross-Lingual Safety Gap in LLMs

    Namya Bhatnagar

    cs.AIcs.CLcs.CYarXiv:2608.18131v12026
  21. Effective Approaches to Attention-based Neural Machine Translation

    Minh-Thang Luong, Hieu Pham, Christopher D. Manning

    cs.CLarXiv:1508.04025v52015
  22. GLUE: A Multi-Task Benchmark and Analysis Platform for Natural Language Understanding

    Alex Wang, Amanpreet Singh, Julian Michael +3

    cs.CLarXiv:1804.07461v32018
  23. Unsupervised Cross-lingual Representation Learning at Scale

    Alexis Conneau, Kartikay Khandelwal, Naman Goyal +7

    cs.CLarXiv:1911.02116v22019
  24. Measuring Massive Multitask Language Understanding

    Dan Hendrycks, Collin Burns, Steven Basart +4

    cs.CYcs.AIcs.CLarXiv:2009.03300v32020
  25. BERTScore: Evaluating Text Generation with BERT

    Tianyi Zhang, Varsha Kishore, Felix Wu +2

    cs.CLarXiv:1904.09675v32019
  26. VIA-SD: Verification via Intra-Model Routing for Speculative Decoding

    Yuchen Xian, Yang He, Yunqiu Xu +1

    cs.CLcs.AIarXiv:2606.12243v12026
  27. Fine-tuning Multi-modal LLMs with ART: Art-based Reinforcement Training

    Michal Chudoba, Sergey Alyaev, Petra Galuscakova +1

    cs.LGcs.AIcs.CLarXiv:2606.11854v12026
  28. XLNet: Generalized Autoregressive Pretraining for Language Understanding

    Zhilin Yang, Zihang Dai, Yiming Yang +3

    cs.CLcs.LGarXiv:1906.08237v22019
  29. When is Your LLM Steerable?

    Chenrui Fan, Yize Cheng, Ming Li +2

    cs.CLcs.LGarXiv:2606.11599v12026
  30. Unstable Features, Reproducible Subspaces: Understanding Seed Dependence in Sparse Autoencoders

    Gleb Gerasimov, Timofei Rusalev, Nikita Balagansky +3

    cs.LGcs.AIcs.CLarXiv:2606.12138v12026
  31. Distributed Representations of Sentences and Documents

    Quoc V. Le, Tomas Mikolov

    cs.CLcs.AIcs.LGarXiv:1405.4053v22014
  32. Getting Better at Working With You: Compiling User Corrections into Runtime Enforcement for Coding Agents

    Yujun Zhou, Kehan Guo, Haomin Zhuang +8

    cs.LGcs.CLarXiv:2606.13174v12026
  33. Training Verifiers to Solve Math Word Problems

    Karl Cobbe, Vineet Kosaraju, Mohammad Bavarian +9

    cs.LGcs.CLarXiv:2110.14168v22021
  34. DistilBERT, a distilled version of BERT: smaller, faster, cheaper and lighter

    Victor Sanh, Lysandre Debut, Julien Chaumond +1

    cs.CLarXiv:1910.01108v42019
  35. Judging LLM-as-a-Judge with MT-Bench and Chatbot Arena

    Lianmin Zheng, Wei-Lin Chiang, Ying Sheng +10

    cs.CLcs.AIarXiv:2306.05685v42023
  36. Visual Instruction Tuning

    Haotian Liu, Chunyuan Li, Qingyang Wu +1

    cs.CVcs.AIcs.CLarXiv:2304.08485v22023
  37. Enriching Word Vectors with Subword Information

    Piotr Bojanowski, Edouard Grave, Armand Joulin +1

    cs.CLcs.LGarXiv:1607.04606v22016
  38. BART: Denoising Sequence-to-Sequence Pre-training for Natural Language Generation, Translation, and Comprehension

    Mike Lewis, Yinhan Liu, Naman Goyal +5

    cs.CLcs.LGstat.MLarXiv:1910.13461v12019
    Summaries:한국어
  39. Deep contextualized word representations

    Matthew E. Peters, Mark Neumann, Mohit Iyyer +4

    cs.CLarXiv:1802.05365v22018
  40. Convolutional Neural Networks for Sentence Classification

    Yoon Kim

    cs.CLcs.NEarXiv:1408.5882v22014
  41. The Llama 3 Herd of Models

    Aaron Grattafiori, Abhimanyu Dubey, Abhinav Jauhri +558

    cs.AIcs.CLcs.CVarXiv:2407.21783v32024
  42. Sentence-BERT: Sentence Embeddings using Siamese BERT-Networks

    Nils Reimers, Iryna Gurevych

    cs.CLarXiv:1908.10084v12019
  43. Intent-Driven Dynamic Chunking: Segmenting Documents to Reflect Predicted Information Needs

    Christos Koutsiaris

    cs.IRcs.AIcs.CLarXiv:2602.14784v12026
  44. Auxiliary uncertainty signals for LLM-assisted systematic review screening: a benchmark across eight Cohen drug-class reviews

    Arya Rahgozar, Pouria Mortezaagha

    cs.CLcs.DLcs.IRarXiv:2608.14551v12026
  45. Learning Phrase Representations using RNN Encoder-Decoder for Statistical Machine Translation

    Kyunghyun Cho, Bart van Merrienboer, Caglar Gulcehre +4

    cs.CLcs.LGcs.NEarXiv:1406.1078v32014
  46. Notes2Skills: From Lab Notebooks to Certainty-Aware Scientific Agent Skills

    Shi Liu, Jiayao Chen, Chengwei Qin +3

    cs.CLarXiv:2606.11897v12026
  47. Which Source Wins? Task-Dependent Reliance in Vision-Language Models

    Rodela Ghosh, Aviral Gupta, Guangjing Wang

    cs.CLcs.CVarXiv:2608.17205v12026
  48. BengaliMCQ: Automatic Generation and Answer Prediction of Academic Multiple-Choice Questions in a Low-Resource Language

    Abu Tarabin Surzo, A. K. M. Nihalul Kabir, Sm Azmain Faysal +3

    cs.CLarXiv:2608.15547v12026
  49. Distributed Representations of Words and Phrases and their Compositionality

    Tomas Mikolov, Ilya Sutskever, Kai Chen +2

    cs.CLcs.LGstat.MLarXiv:1310.4546v12013
  50. Thinking in a Low-Resource Language: What SFT Builds, What RL Fixes, What Accuracy Cannot See

    Ayoub Kirouane, Christos Petrocheilos

    cs.CLcs.LGcs.ROarXiv:2608.17744v12026
  51. HarmProfile: Characterizing Harmful Distributions in Frontier LLMs

    Zhouyuan Ma, Yutao Wu, Hanxun Huang +6

    cs.CLcs.AIarXiv:2608.14577v12026
  52. Harness the Memory: A Holistic Evaluation of Memory Substrates in Memory Agents

    Wei-Chieh Huang, Weizhi Zhang, Yuchen Wu +12

    cs.CLarXiv:2608.15008v12026
  53. Building Social World Models with Large Language Models

    Haofei Yu, Yining Zhao, Guanyu Lin +1

    cs.SIcs.CLarXiv:2606.11482v12026
  54. Beyond Single Object: Learning 3D Relations with Large Language Models

    Kohsuke Ide, Ryousuke Yamada, Yue Qiu +4

    cs.CVcs.AIcs.CLarXiv:2608.15710v12026
  55. Do Language Models Consistently Encode the Current Year?

    Suze van Adrichem, Aditi Bhaskar, Diyi Yang +2

    cs.CLcs.LGarXiv:2608.15507v12026
  56. More Context, Larger Models, or Moral Knowledge? A Systematic Study of Schwartz Value Detection in Political Texts

    Víctor Yeste, Paolo Rosso

    cs.CLcs.AIcs.LGarXiv:2605.22641v32026
  57. Large Language Models as Implicit Sociological Models: Reconstructing Voting Behaviour from Sociodemographic Profiles

    Roman Neruda, Martin Bakoš, Josef Šlerka +3

    cs.CYcs.CLcs.LGarXiv:2608.15871v12026
  58. Polaris: Learning to Generate Table Descriptions from Retrieval Feedback

    Ting Cai, Tuan Minh Phan, AnHai Doan

    cs.CLcs.DBarXiv:2608.17171v12026
  59. Connect the Dots: Training LLMs for Long-Lifecycle Agents with Cross-Domain Generalization Via Reinforcement Learning

    Yanxi Chen, Weijie Shi, Yuexiang Xie +4

    cs.LGcs.AIcs.CLarXiv:2606.20002v12026
  60. The Commercial Tax: Rent-vs-Own Blind Spots in Multi-Hop Retrieval Benchmarks

    Luis M. Sanchez, Kosrow Dehnad

    cs.IRcs.CLarXiv:2608.16096v12026