Computation and Language

Papers filed under cs.CL on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

4,501 to 4,560 of 11,199

  1. Time Travel in LLMs: Tracing Data Contamination in Large Language Models

    Shahriar Golchin, Mihai Surdeanu

    cs.CLcs.AIcs.CRarXiv:2308.08493v32023
  2. Adaptive Critical Token-Aware Retrieval for Repository-Level Code Generation

    Kefeng Duan, Dewu Zheng, Yanlin Wang +8

    cs.SEcs.AIcs.CLarXiv:2609.01601v12026
  3. MIDR: Enrichment-Augmented Indexing for Multimodal Document Retrieval

    Debanjan Mahata, Atharva Tendle, Daniel Preotiuc-Pietro +2

    cs.IRcs.AIcs.CLarXiv:2609.01316v12026
  4. Rubrics as Rewards: Reinforcement Learning Beyond Verifiable Domains

    Anisha Gunjal, Anthony Wang, Elaine Lau +4

    cs.LGcs.AIcs.CLarXiv:2507.17746v22025
  5. DAPO: An Open-Source LLM Reinforcement Learning System at Scale

    Qiying Yu, Zheng Zhang, Ruofei Zhu +32

    cs.LGcs.CLarXiv:2503.14476v22025
  6. Qwen2.5-VL Technical Report

    Shuai Bai, Keqin Chen, Xuejing Liu +24

    cs.CVcs.CLarXiv:2502.13923v12025
  7. A Dataset for Modeling Iterative Problem-Solving

    Fagun Patel, Sang T. Truong, Duc Q. Nguyen +4

    cs.CLarXiv:2609.00940v12026
  8. StoryBuddy: A Human-AI Collaborative Chatbot for Parent-Child Interactive Storytelling with Flexible Parental Involvement

    Zheng Zhang, Ying Xu, Yanhao Wang +6

    cs.HCcs.AIcs.CLarXiv:2202.06205v22022
  9. Learning to Attend via Word-Aspect Associative Fusion for Aspect-based Sentiment Analysis

    Yi Tay, Anh Tuan Luu, Siu Cheung Hui

    cs.CLcs.AIcs.IRarXiv:1712.05403v12017
  10. NeuroLogic A*esque Decoding: Constrained Text Generation with Lookahead Heuristics

    Ximing Lu, Sean Welleck, Peter West +9

    cs.CLarXiv:2112.08726v12021
  11. Graphologue: Exploring Large Language Model Responses with Interactive Diagrams

    Peiling Jiang, Jude Rayan, Steven P. Dow +1

    cs.HCcs.AIcs.CLarXiv:2305.11473v22023
  12. A Language Agent for Autonomous Driving

    Jiageng Mao, Junjie Ye, Yuxi Qian +2

    cs.CVcs.AIcs.CLarXiv:2311.10813v42023
  13. Cite or Decline: A Strict Course-Grounded Chatbot for STEM Lecture Videos

    S M Masrur Ahmed, Jaspal Subhlok

    cs.CLcs.CYcs.IRarXiv:2609.01846v12026
  14. Linguistic Structure Guided Context Modeling for Referring Image Segmentation

    Tianrui Hui, Si Liu, Shaofei Huang +4

    cs.CVcs.CLarXiv:2010.00515v32020
  15. Compositional Exemplars for In-context Learning

    Jiacheng Ye, Zhiyong Wu, Jiangtao Feng +2

    cs.CLcs.AIcs.LGarXiv:2302.05698v32023
  16. Enhancing Person-Job Fit for Talent Recruitment: An Ability-aware Neural Network Approach

    Chuan Qin, Hengshu Zhu, Tong Xu +4

    cs.AIcs.CLcs.LGarXiv:1812.08947v12018
  17. Knowledge-Grounded Dialogue Generation with Pre-trained Language Models

    Xueliang Zhao, Wei Wu, Can Xu +3

    cs.CLarXiv:2010.08824v12020
  18. Predicting Program Exit Code with LLMs and Programming Language Semantics

    Lara Marinov, Aditya Thimmaiah, Jayanth Srinivasa +2

    cs.PLcs.AIcs.CLarXiv:2609.00579v12026
  19. Verifiable Disaster Storylines and Causal Knowledge Graphs: A Citation-Grounded Pipeline from Heterogeneous Humanitarian Sources

    Ivan Decostanzi, Michele Ronco, Sergio Consoli +8

    cs.AIcs.CLarXiv:2609.00858v12026
  20. GLUECoS : An Evaluation Benchmark for Code-Switched NLP

    Simran Khanuja, Sandipan Dandapat, Anirudh Srinivasan +2

    cs.CLarXiv:2004.12376v22020
  21. Compile by Training: Turning Natural-Language Specifications into Local Neural Functions

    Yuntian Deng, Pengyu Nie, Stuart Shieber

    cs.CLcs.AIcs.LGarXiv:2609.04199v12026
  22. Talk2Car: Taking Control of Your Self-Driving Car

    Thierry Deruyttere, Simon Vandenhende, Dusan Grujicic +2

    cs.AIcs.CLcs.ROarXiv:1909.10838v22019
  23. SemEval-2016 Task 3: Community Question Answering

    Preslav Nakov, Lluís Màrquez, Alessandro Moschitti +5

    cs.CLcs.IRarXiv:1912.01972v12019
  24. Compositional generalization through meta sequence-to-sequence learning

    Brenden M. Lake

    cs.CLcs.AIcs.LGarXiv:1906.05381v22019
  25. SciREX: A Challenge Dataset for Document-Level Information Extraction

    Sarthak Jain, Madeleine van Zuylen, Hannaneh Hajishirzi +1

    cs.CLcs.IRcs.LGarXiv:2005.00512v12020
  26. Linearity of Relation Decoding in Transformer Language Models

    Evan Hernandez, Arnab Sen Sharma, Tal Haklay +5

    cs.CLarXiv:2308.09124v22023
  27. From Confusion to Clarity: Confusion-Aware Retrieval and Knowledge Injection for Text Classification

    Manish Gupta, Chaitanya Giri, Jayasimha Talur

    cs.CLcs.AIarXiv:2609.01564v12026
  28. DuoRC: Towards Complex Language Understanding with Paraphrased Reading Comprehension

    Amrita Saha, Rahul Aralikatte, Mitesh M. Khapra +1

    cs.CLarXiv:1804.07927v42018
  29. MRKL Systems: A modular, neuro-symbolic architecture that combines large language models, external knowledge sources and discrete reasoning

    Ehud Karpas, Omri Abend, Yonatan Belinkov +14

    cs.CLcs.AIarXiv:2205.00445v12022
  30. PACE: Towards Surfacing Hidden Conflicts in User Requests

    Yoojin Kim, Jihyoung Jang, Hyounghun Kim

    cs.CLarXiv:2609.03293v12026
  31. Whose Judgments Count? Representation Gaps in Crowdsourced Content Moderation Produce Unequal Protection from Perceived Toxicity

    Zhaodi Chen, Byungkyu Lee

    cs.SIcs.CLcs.CYarXiv:2609.01625v12026
  32. Visualizing and Understanding the Effectiveness of BERT

    Yaru Hao, Li Dong, Furu Wei +1

    cs.CLcs.LGarXiv:1908.05620v12019
  33. InternLM-XComposer-2.5: A Versatile Large Vision Language Model Supporting Long-Contextual Input and Output

    Pan Zhang, Xiaoyi Dong, Yuhang Zang +24

    cs.CVcs.CLarXiv:2407.03320v12024
  34. Make Your LLM Fully Utilize the Context

    Shengnan An, Zexiong Ma, Zeqi Lin +2

    cs.CLcs.AIarXiv:2404.16811v22024
  35. Editable Visual Design

    Junyan Ye, Wei Liu, Dongzhi Jiang +9

    cs.CVcs.CLarXiv:2609.04034v12026
  36. Masked Diffusion Models are Secretly Time-Agnostic Masked Models and Exploit Inaccurate Categorical Sampling

    Kaiwen Zheng, Yongxin Chen, Hanzi Mao +3

    cs.LGcs.AIcs.CLarXiv:2409.02908v62024
  37. Fast Domain Adaptation for Neural Machine Translation

    Markus Freitag, Yaser Al-Onaizan

    cs.CLarXiv:1612.06897v12016
  38. CoNLL-SIGMORPHON 2017 Shared Task: Universal Morphological Reinflection in 52 Languages

    Ryan Cotterell, Christo Kirov, John Sylak-Glassman +8

    cs.CLarXiv:1706.09031v22017
  39. Clinical Concept Embeddings Learned from Massive Sources of Multimodal Medical Data

    Andrew L. Beam, Benjamin Kompa, Allen Schmaltz +6

    cs.CLcs.AIstat.MLarXiv:1804.01486v32018
  40. Recursive Introspection: Teaching Language Model Agents How to Self-Improve

    Yuxiao Qu, Tianjun Zhang, Naman Garg +1

    cs.LGcs.AIcs.CLarXiv:2407.18219v22024
  41. CubeMLP: An MLP-based Model for Multimodal Sentiment Analysis and Depression Estimation

    Hao Sun, Hongyi Wang, Jiaqing Liu +2

    cs.MMcs.CLcs.CVarXiv:2207.14087v32022
  42. ShallowStream: Index Shallow then Answer Deep for Streaming Video Understanding

    Jitai Hao, Ke Yang, Qiang Huang +1

    cs.CVcs.CLarXiv:2609.02780v12026
  43. Knowledge Distillation from Internal Representations

    Gustavo Aguilar, Yuan Ling, Yu Zhang +3

    cs.CLarXiv:1910.03723v22019
  44. When Persona Attributes Improve Population Alignment in Large Language Models

    Leon Fröhling, Jens Rupprecht, Markus Strohmaier +1

    cs.CLcs.CYarXiv:2609.02526v12026
  45. A Layered Taxonomy for Chinese Learner Grammatical Error Annotation

    Mengyang Qiu, Jungyeul Park

    cs.CLarXiv:2609.02153v12026
  46. A Simple Recipe for Multilingual Grammatical Error Correction

    Sascha Rothe, Jonathan Mallinson, Eric Malmi +2

    cs.CLarXiv:2106.03830v22021
  47. Sentence Similarity Learning by Lexical Decomposition and Composition

    Zhiguo Wang, Haitao Mi, Abraham Ittycheriah

    cs.CLarXiv:1602.07019v22016
  48. NS-Copilot: An LLM-Driven Agent System for Autonomous Neuroscience Analysis

    Wuche Liu, Yiran Qiao, Linlin Hou +4

    cs.CLarXiv:2609.01971v12026
  49. Multimodal Named Entity Recognition for Short Social Media Posts

    Seungwhan Moon, Leonardo Neves, Vitor Carvalho

    cs.CLarXiv:1802.07862v12018
  50. TaRA: Training-Aware Low-Rank Adaptation Initialization

    Taehyeon Kim, Eunhyeok Park

    cs.CLcs.AIcs.LGarXiv:2609.02639v12026
  51. RQ-RAG: Learning to Refine Queries for Retrieval Augmented Generation

    Chi-Min Chan, Chunpu Xu, Ruibin Yuan +4

    cs.CLarXiv:2404.00610v12024
  52. Investigating the Factual Knowledge Boundary of Large Language Models with Retrieval Augmentation

    Ruiyang Ren, Yuhao Wang, Yingqi Qu +6

    cs.CLcs.IRarXiv:2307.11019v32023
  53. Ferret-UI: Grounded Mobile UI Understanding with Multimodal LLMs

    Keen You, Haotian Zhang, Eldon Schoop +5

    cs.CVcs.CLcs.HCarXiv:2404.05719v12024
  54. SALA: Semantic-Aware Logical Alignment for Complex Reasoning in In-Context Learning

    Zhao Ji, Wenqing Chen, Zhixuan Chu +4

    cs.AIcs.CLarXiv:2609.02336v12026
  55. CoMerge: Conflict-Driven Preference Optimization for Multi-Task Model Merging

    Mingjie Zheng, Zihao Chen, Wenqing Chen +4

    cs.AIcs.CLarXiv:2609.02273v12026
  56. Selective Agent Guidance via Entropy: Learning Autonomous Policies from Imperfect VLM Teachers

    Giovanni Bonetta, Matteo Merler, Davide Zago +2

    cs.AIcs.CLcs.LGarXiv:2609.01567v22026
  57. Efficient Dialogue State Tracking by Selectively Overwriting Memory

    Sungdong Kim, Sohee Yang, Gyuwan Kim +1

    cs.CLarXiv:1911.03906v22019
  58. UKP-Athene: Multi-Sentence Textual Entailment for Claim Verification

    Andreas Hanselowski, Hao Zhang, Zile Li +4

    cs.IRcs.AIcs.CLarXiv:1809.01479v52018
  59. SciRepEval: A Multi-Format Benchmark for Scientific Document Representations

    Amanpreet Singh, Mike D'Arcy, Arman Cohan +2

    cs.CLcs.AIcs.IRarXiv:2211.13308v42022
  60. Simulating Classroom Education with LLM-Empowered Agents

    Zheyuan Zhang, Daniel Zhang-Li, Jifan Yu +9

    cs.CLcs.HCarXiv:2406.19226v22024