Computation and Language
Papers filed under cs.CL on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
4,681 to 4,740 of 11,332
CoNLL-SIGMORPHON 2017 Shared Task: Universal Morphological Reinflection in 52 Languages
Ryan Cotterell, Christo Kirov, John Sylak-Glassman +8
cs.CLarXiv:1706.09031v22017Clinical Concept Embeddings Learned from Massive Sources of Multimodal Medical Data
Andrew L. Beam, Benjamin Kompa, Allen Schmaltz +6
cs.CLcs.AIstat.MLarXiv:1804.01486v32018Recursive Introspection: Teaching Language Model Agents How to Self-Improve
Yuxiao Qu, Tianjun Zhang, Naman Garg +1
cs.LGcs.AIcs.CLarXiv:2407.18219v22024CubeMLP: An MLP-based Model for Multimodal Sentiment Analysis and Depression Estimation
Hao Sun, Hongyi Wang, Jiaqing Liu +2
cs.MMcs.CLcs.CVarXiv:2207.14087v32022ShallowStream: Index Shallow then Answer Deep for Streaming Video Understanding
Jitai Hao, Ke Yang, Qiang Huang +1
cs.CVcs.CLarXiv:2609.02780v12026Knowledge Distillation from Internal Representations
Gustavo Aguilar, Yuan Ling, Yu Zhang +3
cs.CLarXiv:1910.03723v22019When Persona Attributes Improve Population Alignment in Large Language Models
Leon Fröhling, Jens Rupprecht, Markus Strohmaier +1
cs.CLcs.CYarXiv:2609.02526v12026A Layered Taxonomy for Chinese Learner Grammatical Error Annotation
Mengyang Qiu, Jungyeul Park
cs.CLarXiv:2609.02153v12026A Simple Recipe for Multilingual Grammatical Error Correction
Sascha Rothe, Jonathan Mallinson, Eric Malmi +2
cs.CLarXiv:2106.03830v22021Sentence Similarity Learning by Lexical Decomposition and Composition
Zhiguo Wang, Haitao Mi, Abraham Ittycheriah
cs.CLarXiv:1602.07019v22016NS-Copilot: An LLM-Driven Agent System for Autonomous Neuroscience Analysis
Wuche Liu, Yiran Qiao, Linlin Hou +4
cs.CLarXiv:2609.01971v12026Multimodal Named Entity Recognition for Short Social Media Posts
Seungwhan Moon, Leonardo Neves, Vitor Carvalho
cs.CLarXiv:1802.07862v12018TaRA: Training-Aware Low-Rank Adaptation Initialization
Taehyeon Kim, Eunhyeok Park
cs.CLcs.AIcs.LGarXiv:2609.02639v12026RQ-RAG: Learning to Refine Queries for Retrieval Augmented Generation
Chi-Min Chan, Chunpu Xu, Ruibin Yuan +4
cs.CLarXiv:2404.00610v12024Investigating the Factual Knowledge Boundary of Large Language Models with Retrieval Augmentation
Ruiyang Ren, Yuhao Wang, Yingqi Qu +6
cs.CLcs.IRarXiv:2307.11019v32023Ferret-UI: Grounded Mobile UI Understanding with Multimodal LLMs
Keen You, Haotian Zhang, Eldon Schoop +5
cs.CVcs.CLcs.HCarXiv:2404.05719v12024SALA: Semantic-Aware Logical Alignment for Complex Reasoning in In-Context Learning
Zhao Ji, Wenqing Chen, Zhixuan Chu +4
cs.AIcs.CLarXiv:2609.02336v12026CoMerge: Conflict-Driven Preference Optimization for Multi-Task Model Merging
Mingjie Zheng, Zihao Chen, Wenqing Chen +4
cs.AIcs.CLarXiv:2609.02273v12026Selective Agent Guidance via Entropy: Learning Autonomous Policies from Imperfect VLM Teachers
Giovanni Bonetta, Matteo Merler, Davide Zago +2
cs.AIcs.CLcs.LGarXiv:2609.01567v22026Efficient Dialogue State Tracking by Selectively Overwriting Memory
Sungdong Kim, Sohee Yang, Gyuwan Kim +1
cs.CLarXiv:1911.03906v22019UKP-Athene: Multi-Sentence Textual Entailment for Claim Verification
Andreas Hanselowski, Hao Zhang, Zile Li +4
cs.IRcs.AIcs.CLarXiv:1809.01479v52018SciRepEval: A Multi-Format Benchmark for Scientific Document Representations
Amanpreet Singh, Mike D'Arcy, Arman Cohan +2
cs.CLcs.AIcs.IRarXiv:2211.13308v42022Simulating Classroom Education with LLM-Empowered Agents
Zheyuan Zhang, Daniel Zhang-Li, Jifan Yu +9
cs.CLcs.HCarXiv:2406.19226v22024A Survey of Deep Learning for Mathematical Reasoning
Pan Lu, Liang Qiu, Wenhao Yu +2
cs.AIcs.CLcs.CVarXiv:2212.10535v22022Could a Large Language Model be Conscious?
David J. Chalmers
cs.AIcs.CLcs.LGarXiv:2303.07103v32023Learning Neural Templates for Text Generation
Sam Wiseman, Stuart M. Shieber, Alexander M. Rush
cs.CLcs.LGarXiv:1808.10122v32018An Empirical Study on Robustness to Spurious Correlations using Pre-trained Language Models
Lifu Tu, Garima Lalwani, Spandana Gella +1
cs.CLcs.LGarXiv:2007.06778v32020Counterfactual Memorization in Neural Language Models
Chiyuan Zhang, Daphne Ippolito, Katherine Lee +3
cs.CLcs.AIcs.LGarXiv:2112.12938v22021Why Does Surprisal From Larger Transformer-Based Language Models Provide a Poorer Fit to Human Reading Times?
Byung-Doh Oh, William Schuler
cs.CLarXiv:2212.12131v12022Embodied Question Answering in Photorealistic Environments with Point Cloud Perception
Erik Wijmans, Samyak Datta, Oleksandr Maksymets +6
cs.CVcs.AIcs.CLarXiv:1904.03461v12019Efficient Guided Generation for Large Language Models
Brandon T. Willard, Rémi Louf
cs.CLcs.LGarXiv:2307.09702v42023Hurtful Words: Quantifying Biases in Clinical Contextual Word Embeddings
Haoran Zhang, Amy X. Lu, Mohamed Abdalla +2
cs.CLcs.CYcs.LGarXiv:2003.11515v12020Adversarial Training for Large Neural Language Models
Xiaodong Liu, Hao Cheng, Pengcheng He +4
cs.CLarXiv:2004.08994v22020Interpretable Adversarial Perturbation in Input Embedding Space for Text
Motoki Sato, Jun Suzuki, Hiroyuki Shindo +1
cs.LGcs.CLstat.MLarXiv:1805.02917v12018VerTox: Verifiable Reward-Guided Corpus Poisoning Against Neural Ranking Models
Zhiqi Huang, Vivek Datla, Zhichao Xu +3
cs.CLcs.IRarXiv:2609.01325v12026SciTrue: Reliable Scientific Claim Validation with Frontier and Open Language Models at the NTCIR SciClaimEval Task
Qiming Bao, Neşet Özkan Tan, Siyuan Wang +1
cs.AIcs.CLarXiv:2609.00654v12026Abstractive Summarization of Reddit Posts with Multi-level Memory Networks
Byeongchang Kim, Hyunwoo Kim, Gunhee Kim
cs.CLarXiv:1811.00783v22018A comprehensive evaluation of ChatGPT's zero-shot Text-to-SQL capability
Aiwei Liu, Xuming Hu, Lijie Wen +1
cs.CLcs.AIarXiv:2303.13547v12023A Mechanistic Understanding of Alignment Algorithms: A Case Study on DPO and Toxicity
Andrew Lee, Xiaoyan Bai, Itamar Pres +3
cs.CLcs.AIarXiv:2401.01967v12024End-to-End Automatic Speech Translation of Audiobooks
Alexandre Bérard, Laurent Besacier, Ali Can Kocabiyikoglu +1
cs.CLarXiv:1802.04200v12018FANToM: A Benchmark for Stress-testing Machine Theory of Mind in Interactions
Hyunwoo Kim, Melanie Sclar, Xuhui Zhou +4
cs.CLcs.AIarXiv:2310.15421v32023CHASE-SQL: Multi-Path Reasoning and Preference Optimized Candidate Selection in Text-to-SQL
Mohammadreza Pourreza, Hailong Li, Ruoxi Sun +7
cs.LGcs.AIcs.CLarXiv:2410.01943v12024From Matching to Generation: A Survey on Generative Information Retrieval
Xiaoxi Li, Jiajie Jin, Yujia Zhou +4
cs.IRcs.AIcs.CLarXiv:2404.14851v42024PaperQA: Retrieval-Augmented Generative Agent for Scientific Research
Jakub Lála, Odhran O'Donoghue, Aleksandar Shtedritski +3
cs.CLcs.AIcs.LGarXiv:2312.07559v22023Reducing hallucination in structured outputs via Retrieval-Augmented Generation
Patrice Béchard, Orlando Marquez Ayala
cs.LGcs.AIcs.CLarXiv:2404.08189v12024Terminal-Universe: Turning Agent Trajectories into Scalable Terminal Environments
Jie Wu, Zhenru Zhang, Beichen Zhang +11
cs.AIcs.CLarXiv:2609.04148v12026Aggression-annotated Corpus of Hindi-English Code-mixed Data
Ritesh Kumar, Aishwarya N. Reganti, Akshit Bhatia +1
cs.CLarXiv:1803.09402v12018A Dataset for Answering Time-Sensitive Questions
Wenhu Chen, Xinyi Wang, William Yang Wang
cs.CLcs.AIarXiv:2108.06314v52021Who Wrote this Code? Watermarking for Code Generation
Taehyun Lee, Seokhee Hong, Jaewoo Ahn +5
cs.CLarXiv:2305.15060v42023Few-shot Slot Tagging with Collapsed Dependency Transfer and Label-enhanced Task-adaptive Projection Network
Yutai Hou, Wanxiang Che, Yongkui Lai +4
cs.CLcs.LGarXiv:2006.05702v12020Cross-Attention is All You Need: Adapting Pretrained Transformers for Machine Translation
Mozhdeh Gheini, Xiang Ren, Jonathan May
cs.CLarXiv:2104.08771v22021CORE: Improving Compositional Reasoning in MLLM Embedding via Reranker Distillation
Tingyu Song, Mingxin Li, Yanzhao Zhang +5
cs.CVcs.AIcs.CLarXiv:2609.04083v12026Rethinking On-Policy Distillation of Large Language Models II: One Training Example
Zixuan Fu, Bingxiang He, Yuxin Zuo +10
cs.AIcs.CLarXiv:2609.04172v12026Random Attention: Rethinking KV Cache Eviction for Efficient Reasoning
Heng Wang, Jielin Qiu, Wenting Zhao +7
cs.CLarXiv:2609.03430v12026Streaming automatic speech recognition with the transformer model
Niko Moritz, Takaaki Hori, Jonathan Le Roux
cs.SDcs.CLcs.LGarXiv:2001.02674v52020Emformer: Efficient Memory Transformer Based Acoustic Model For Low Latency Streaming Speech Recognition
Yangyang Shi, Yongqiang Wang, Chunyang Wu +5
cs.SDcs.CLcs.LGarXiv:2010.10759v42020MojiTalk: Generating Emotional Responses at Scale
Xianda Zhou, William Yang Wang
cs.CLcs.AIarXiv:1711.04090v22017APIGen: Automated Pipeline for Generating Verifiable and Diverse Function-Calling Datasets
Zuxin Liu, Thai Hoang, Jianguo Zhang +14
cs.CLcs.AIcs.LGarXiv:2406.18518v12024User Feedback Provides a Unique Signal that LLMs Can not Detect
Shachar Don-Yehiya, Leshem Choshen, Omri Abend
cs.CLarXiv:2609.02859v12026MathCoder: Seamless Code Integration in LLMs for Enhanced Mathematical Reasoning
Ke Wang, Houxing Ren, Aojun Zhou +7
cs.CLcs.AIcs.CVarXiv:2310.03731v12023