Computation and Language

Papers filed under cs.CL on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

4,681 to 4,740 of 11,332

  1. CoNLL-SIGMORPHON 2017 Shared Task: Universal Morphological Reinflection in 52 Languages

    Ryan Cotterell, Christo Kirov, John Sylak-Glassman +8

    cs.CLarXiv:1706.09031v22017
  2. Clinical Concept Embeddings Learned from Massive Sources of Multimodal Medical Data

    Andrew L. Beam, Benjamin Kompa, Allen Schmaltz +6

    cs.CLcs.AIstat.MLarXiv:1804.01486v32018
  3. Recursive Introspection: Teaching Language Model Agents How to Self-Improve

    Yuxiao Qu, Tianjun Zhang, Naman Garg +1

    cs.LGcs.AIcs.CLarXiv:2407.18219v22024
  4. CubeMLP: An MLP-based Model for Multimodal Sentiment Analysis and Depression Estimation

    Hao Sun, Hongyi Wang, Jiaqing Liu +2

    cs.MMcs.CLcs.CVarXiv:2207.14087v32022
  5. ShallowStream: Index Shallow then Answer Deep for Streaming Video Understanding

    Jitai Hao, Ke Yang, Qiang Huang +1

    cs.CVcs.CLarXiv:2609.02780v12026
  6. Knowledge Distillation from Internal Representations

    Gustavo Aguilar, Yuan Ling, Yu Zhang +3

    cs.CLarXiv:1910.03723v22019
  7. When Persona Attributes Improve Population Alignment in Large Language Models

    Leon Fröhling, Jens Rupprecht, Markus Strohmaier +1

    cs.CLcs.CYarXiv:2609.02526v12026
  8. A Layered Taxonomy for Chinese Learner Grammatical Error Annotation

    Mengyang Qiu, Jungyeul Park

    cs.CLarXiv:2609.02153v12026
  9. A Simple Recipe for Multilingual Grammatical Error Correction

    Sascha Rothe, Jonathan Mallinson, Eric Malmi +2

    cs.CLarXiv:2106.03830v22021
  10. Sentence Similarity Learning by Lexical Decomposition and Composition

    Zhiguo Wang, Haitao Mi, Abraham Ittycheriah

    cs.CLarXiv:1602.07019v22016
  11. NS-Copilot: An LLM-Driven Agent System for Autonomous Neuroscience Analysis

    Wuche Liu, Yiran Qiao, Linlin Hou +4

    cs.CLarXiv:2609.01971v12026
  12. Multimodal Named Entity Recognition for Short Social Media Posts

    Seungwhan Moon, Leonardo Neves, Vitor Carvalho

    cs.CLarXiv:1802.07862v12018
  13. TaRA: Training-Aware Low-Rank Adaptation Initialization

    Taehyeon Kim, Eunhyeok Park

    cs.CLcs.AIcs.LGarXiv:2609.02639v12026
  14. RQ-RAG: Learning to Refine Queries for Retrieval Augmented Generation

    Chi-Min Chan, Chunpu Xu, Ruibin Yuan +4

    cs.CLarXiv:2404.00610v12024
  15. Investigating the Factual Knowledge Boundary of Large Language Models with Retrieval Augmentation

    Ruiyang Ren, Yuhao Wang, Yingqi Qu +6

    cs.CLcs.IRarXiv:2307.11019v32023
  16. Ferret-UI: Grounded Mobile UI Understanding with Multimodal LLMs

    Keen You, Haotian Zhang, Eldon Schoop +5

    cs.CVcs.CLcs.HCarXiv:2404.05719v12024
  17. SALA: Semantic-Aware Logical Alignment for Complex Reasoning in In-Context Learning

    Zhao Ji, Wenqing Chen, Zhixuan Chu +4

    cs.AIcs.CLarXiv:2609.02336v12026
  18. CoMerge: Conflict-Driven Preference Optimization for Multi-Task Model Merging

    Mingjie Zheng, Zihao Chen, Wenqing Chen +4

    cs.AIcs.CLarXiv:2609.02273v12026
  19. Selective Agent Guidance via Entropy: Learning Autonomous Policies from Imperfect VLM Teachers

    Giovanni Bonetta, Matteo Merler, Davide Zago +2

    cs.AIcs.CLcs.LGarXiv:2609.01567v22026
  20. Efficient Dialogue State Tracking by Selectively Overwriting Memory

    Sungdong Kim, Sohee Yang, Gyuwan Kim +1

    cs.CLarXiv:1911.03906v22019
  21. UKP-Athene: Multi-Sentence Textual Entailment for Claim Verification

    Andreas Hanselowski, Hao Zhang, Zile Li +4

    cs.IRcs.AIcs.CLarXiv:1809.01479v52018
  22. SciRepEval: A Multi-Format Benchmark for Scientific Document Representations

    Amanpreet Singh, Mike D'Arcy, Arman Cohan +2

    cs.CLcs.AIcs.IRarXiv:2211.13308v42022
  23. Simulating Classroom Education with LLM-Empowered Agents

    Zheyuan Zhang, Daniel Zhang-Li, Jifan Yu +9

    cs.CLcs.HCarXiv:2406.19226v22024
  24. A Survey of Deep Learning for Mathematical Reasoning

    Pan Lu, Liang Qiu, Wenhao Yu +2

    cs.AIcs.CLcs.CVarXiv:2212.10535v22022
  25. Could a Large Language Model be Conscious?

    David J. Chalmers

    cs.AIcs.CLcs.LGarXiv:2303.07103v32023
  26. Learning Neural Templates for Text Generation

    Sam Wiseman, Stuart M. Shieber, Alexander M. Rush

    cs.CLcs.LGarXiv:1808.10122v32018
  27. An Empirical Study on Robustness to Spurious Correlations using Pre-trained Language Models

    Lifu Tu, Garima Lalwani, Spandana Gella +1

    cs.CLcs.LGarXiv:2007.06778v32020
  28. Counterfactual Memorization in Neural Language Models

    Chiyuan Zhang, Daphne Ippolito, Katherine Lee +3

    cs.CLcs.AIcs.LGarXiv:2112.12938v22021
  29. Why Does Surprisal From Larger Transformer-Based Language Models Provide a Poorer Fit to Human Reading Times?

    Byung-Doh Oh, William Schuler

    cs.CLarXiv:2212.12131v12022
  30. Embodied Question Answering in Photorealistic Environments with Point Cloud Perception

    Erik Wijmans, Samyak Datta, Oleksandr Maksymets +6

    cs.CVcs.AIcs.CLarXiv:1904.03461v12019
  31. Efficient Guided Generation for Large Language Models

    Brandon T. Willard, Rémi Louf

    cs.CLcs.LGarXiv:2307.09702v42023
  32. Hurtful Words: Quantifying Biases in Clinical Contextual Word Embeddings

    Haoran Zhang, Amy X. Lu, Mohamed Abdalla +2

    cs.CLcs.CYcs.LGarXiv:2003.11515v12020
  33. Adversarial Training for Large Neural Language Models

    Xiaodong Liu, Hao Cheng, Pengcheng He +4

    cs.CLarXiv:2004.08994v22020
  34. Interpretable Adversarial Perturbation in Input Embedding Space for Text

    Motoki Sato, Jun Suzuki, Hiroyuki Shindo +1

    cs.LGcs.CLstat.MLarXiv:1805.02917v12018
  35. VerTox: Verifiable Reward-Guided Corpus Poisoning Against Neural Ranking Models

    Zhiqi Huang, Vivek Datla, Zhichao Xu +3

    cs.CLcs.IRarXiv:2609.01325v12026
  36. SciTrue: Reliable Scientific Claim Validation with Frontier and Open Language Models at the NTCIR SciClaimEval Task

    Qiming Bao, Neşet Özkan Tan, Siyuan Wang +1

    cs.AIcs.CLarXiv:2609.00654v12026
  37. Abstractive Summarization of Reddit Posts with Multi-level Memory Networks

    Byeongchang Kim, Hyunwoo Kim, Gunhee Kim

    cs.CLarXiv:1811.00783v22018
  38. A comprehensive evaluation of ChatGPT's zero-shot Text-to-SQL capability

    Aiwei Liu, Xuming Hu, Lijie Wen +1

    cs.CLcs.AIarXiv:2303.13547v12023
  39. A Mechanistic Understanding of Alignment Algorithms: A Case Study on DPO and Toxicity

    Andrew Lee, Xiaoyan Bai, Itamar Pres +3

    cs.CLcs.AIarXiv:2401.01967v12024
  40. End-to-End Automatic Speech Translation of Audiobooks

    Alexandre Bérard, Laurent Besacier, Ali Can Kocabiyikoglu +1

    cs.CLarXiv:1802.04200v12018
  41. FANToM: A Benchmark for Stress-testing Machine Theory of Mind in Interactions

    Hyunwoo Kim, Melanie Sclar, Xuhui Zhou +4

    cs.CLcs.AIarXiv:2310.15421v32023
  42. CHASE-SQL: Multi-Path Reasoning and Preference Optimized Candidate Selection in Text-to-SQL

    Mohammadreza Pourreza, Hailong Li, Ruoxi Sun +7

    cs.LGcs.AIcs.CLarXiv:2410.01943v12024
  43. From Matching to Generation: A Survey on Generative Information Retrieval

    Xiaoxi Li, Jiajie Jin, Yujia Zhou +4

    cs.IRcs.AIcs.CLarXiv:2404.14851v42024
  44. PaperQA: Retrieval-Augmented Generative Agent for Scientific Research

    Jakub Lála, Odhran O'Donoghue, Aleksandar Shtedritski +3

    cs.CLcs.AIcs.LGarXiv:2312.07559v22023
  45. Reducing hallucination in structured outputs via Retrieval-Augmented Generation

    Patrice Béchard, Orlando Marquez Ayala

    cs.LGcs.AIcs.CLarXiv:2404.08189v12024
  46. Terminal-Universe: Turning Agent Trajectories into Scalable Terminal Environments

    Jie Wu, Zhenru Zhang, Beichen Zhang +11

    cs.AIcs.CLarXiv:2609.04148v12026
  47. Aggression-annotated Corpus of Hindi-English Code-mixed Data

    Ritesh Kumar, Aishwarya N. Reganti, Akshit Bhatia +1

    cs.CLarXiv:1803.09402v12018
  48. A Dataset for Answering Time-Sensitive Questions

    Wenhu Chen, Xinyi Wang, William Yang Wang

    cs.CLcs.AIarXiv:2108.06314v52021
  49. Who Wrote this Code? Watermarking for Code Generation

    Taehyun Lee, Seokhee Hong, Jaewoo Ahn +5

    cs.CLarXiv:2305.15060v42023
  50. Few-shot Slot Tagging with Collapsed Dependency Transfer and Label-enhanced Task-adaptive Projection Network

    Yutai Hou, Wanxiang Che, Yongkui Lai +4

    cs.CLcs.LGarXiv:2006.05702v12020
  51. Cross-Attention is All You Need: Adapting Pretrained Transformers for Machine Translation

    Mozhdeh Gheini, Xiang Ren, Jonathan May

    cs.CLarXiv:2104.08771v22021
  52. CORE: Improving Compositional Reasoning in MLLM Embedding via Reranker Distillation

    Tingyu Song, Mingxin Li, Yanzhao Zhang +5

    cs.CVcs.AIcs.CLarXiv:2609.04083v12026
  53. Rethinking On-Policy Distillation of Large Language Models II: One Training Example

    Zixuan Fu, Bingxiang He, Yuxin Zuo +10

    cs.AIcs.CLarXiv:2609.04172v12026
  54. Random Attention: Rethinking KV Cache Eviction for Efficient Reasoning

    Heng Wang, Jielin Qiu, Wenting Zhao +7

    cs.CLarXiv:2609.03430v12026
  55. Streaming automatic speech recognition with the transformer model

    Niko Moritz, Takaaki Hori, Jonathan Le Roux

    cs.SDcs.CLcs.LGarXiv:2001.02674v52020
  56. Emformer: Efficient Memory Transformer Based Acoustic Model For Low Latency Streaming Speech Recognition

    Yangyang Shi, Yongqiang Wang, Chunyang Wu +5

    cs.SDcs.CLcs.LGarXiv:2010.10759v42020
  57. MojiTalk: Generating Emotional Responses at Scale

    Xianda Zhou, William Yang Wang

    cs.CLcs.AIarXiv:1711.04090v22017
  58. APIGen: Automated Pipeline for Generating Verifiable and Diverse Function-Calling Datasets

    Zuxin Liu, Thai Hoang, Jianguo Zhang +14

    cs.CLcs.AIcs.LGarXiv:2406.18518v12024
  59. User Feedback Provides a Unique Signal that LLMs Can not Detect

    Shachar Don-Yehiya, Leshem Choshen, Omri Abend

    cs.CLarXiv:2609.02859v12026
  60. MathCoder: Seamless Code Integration in LLMs for Enhanced Mathematical Reasoning

    Ke Wang, Houxing Ren, Aojun Zhou +7

    cs.CLcs.AIcs.CVarXiv:2310.03731v12023