Computation and Language
Papers filed under cs.CL on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
10,081 to 10,140 of 11,333
Approximate Nearest Neighbor Negative Contrastive Learning for Dense Text Retrieval
Lee Xiong, Chenyan Xiong, Ye Li +5
cs.IRcs.CLcs.LGarXiv:2007.00808v22020GQA: Training Generalized Multi-Query Transformer Models from Multi-Head Checkpoints
Joshua Ainslie, James Lee-Thorp, Michiel de Jong +3
cs.CLcs.LGarXiv:2305.13245v32023Model soups: averaging weights of multiple fine-tuned models improves accuracy without increasing inference time
Mitchell Wortsman, Gabriel Ilharco, Samir Yitzhak Gadre +8
cs.LGcs.CLcs.CVarXiv:2203.05482v32022ChatGPT Outperforms Crowd-Workers for Text-Annotation Tasks
Fabrizio Gilardi, Meysam Alizadeh, Maël Kubli
cs.CLcs.CYarXiv:2303.15056v22023MELD: A Multimodal Multi-Party Dataset for Emotion Recognition in Conversations
Soujanya Poria, Devamanyu Hazarika, Navonil Majumder +3
cs.CLarXiv:1810.02508v62018M3-Embedding: Multi-Linguality, Multi-Functionality, Multi-Granularity Text Embeddings Through Self-Knowledge Distillation
Jianlv Chen, Shitao Xiao, Peitian Zhang +3
cs.CLcs.AIcs.LGarXiv:2402.03216v52024Pre-trained Models for Natural Language Processing: A Survey
Xipeng Qiu, Tianxiang Sun, Yige Xu +3
cs.CLcs.LGarXiv:2003.08271v42020A simple neural network module for relational reasoning
Adam Santoro, David Raposo, David G. T. Barrett +4
cs.CLcs.LGarXiv:1706.01427v12017word2vec Explained: deriving Mikolov et al.'s negative-sampling word-embedding method
Yoav Goldberg, Omer Levy
cs.CLcs.LGstat.MLarXiv:1402.3722v12014ChatGLM: A Family of Large Language Models from GLM-130B to GLM-4 All Tools
Team GLM, :, Aohan Zeng +56
cs.CLarXiv:2406.12793v22024"Liar, Liar Pants on Fire": A New Benchmark Dataset for Fake News Detection
William Yang Wang
cs.CLcs.CYarXiv:1705.00648v12017Character-Aware Neural Language Models
Yoon Kim, Yacine Jernite, David Sontag +1
cs.CLcs.NEstat.MLarXiv:1508.06615v42015ChartQA: A Benchmark for Question Answering about Charts with Visual and Logical Reasoning
Ahmed Masry, Do Xuan Long, Jia Qing Tan +2
cs.CLarXiv:2203.10244v12022ESC-Skills: Discovering and Self-Evolving Skills for Emotional Support Conversations
Jie Zhu, Huaixia Dou, Shuo Jiang +5
cs.CLcs.AIarXiv:2605.27908v12026OmniInteract: Benchmarking Real-World Streaming Interaction for Real-Time Omnimodal Assistants
Xudong Lu, Xueying Li, Annan Wang +8
cs.CVcs.CLarXiv:2605.26485v12026Adversarial Examples for Evaluating Reading Comprehension Systems
Robin Jia, Percy Liang
cs.CLcs.LGarXiv:1707.07328v12017Alignment Tampering: How Reinforcement Learning from Human Feedback Is Exploited to Optimize Misaligned Biases
Dongyoon Hahm, Dylan Hadfield-Menell, Kimin Lee
cs.AIcs.CLcs.LGarXiv:2605.27355v22026Grounded Language-Image Pre-training
Liunian Harold Li, Pengchuan Zhang, Haotian Zhang +9
cs.CVcs.AIcs.CLarXiv:2112.03857v22021ESPnet: End-to-End Speech Processing Toolkit
Shinji Watanabe, Takaaki Hori, Shigeki Karita +9
cs.CLarXiv:1804.00015v12018Revealing Algorithmic Deductive Circuits for Logical Reasoning
Phuong Minh Nguyen, Tien Huu Dang, Naoya Inoue
cs.AIcs.CLarXiv:2605.27824v12026Models That Know How Evaluations Are Designed Score Safer
Katharina Deckenbach, Haritz Puerto, Jonas Geiping +1
cs.CLcs.AIarXiv:2605.28591v32026DialoGPT: Large-Scale Generative Pre-training for Conversational Response Generation
Yizhe Zhang, Siqi Sun, Michel Galley +6
cs.CLcs.LGarXiv:1911.00536v32019How multilingual is Multilingual BERT?
Telmo Pires, Eva Schlinger, Dan Garrette
cs.CLcs.AIcs.LGarXiv:1906.01502v12019PEFT-Arena: Understanding Parameter-Efficient Finetuning from a Stability-Plasticity Perspective
Yangyi Huang, Ruotian Peng, Zeju Qiu +4
cs.LGcs.CLarXiv:2605.28819v12026Hierarchical Question-Image Co-Attention for Visual Question Answering
Jiasen Lu, Jianwei Yang, Dhruv Batra +1
cs.CVcs.CLarXiv:1606.00061v52016Experience Makes Skillful: Enabling Generalizable Medical Agent Reasoning via Self-Evolving Skill Memory
Haoran Sun, Wenjie Li, Yujie Zhang +8
cs.AIcs.CLarXiv:2606.09365v32026Summaries:한국어LLaVA-Med: Training a Large Language-and-Vision Assistant for Biomedicine in One Day
Chunyuan Li, Cliff Wong, Sheng Zhang +6
cs.CVcs.CLarXiv:2306.00890v12023Flickr30k Entities: Collecting Region-to-Phrase Correspondences for Richer Image-to-Sentence Models
Bryan A. Plummer, Liwei Wang, Chris M. Cervantes +3
cs.CVcs.CLarXiv:1505.04870v42015Teaching Machines to Read and Comprehend
Karl Moritz Hermann, Tomáš Kočiský, Edward Grefenstette +4
cs.CLcs.AIcs.NEarXiv:1506.03340v32015CAMEL: Communicative Agents for "Mind" Exploration of Large Language Model Society
Guohao Li, Hasan Abed Al Kader Hammoud, Hani Itani +2
cs.AIcs.CLcs.CYarXiv:2303.17760v22023Learn from Weaknesses: Automated Domain Specialization for Small Computer-Use Agents
Suji Kim, Kangsan Kim, Sung Ju Hwang
cs.LGcs.AIcs.CLarXiv:2605.28775v12026VibeSearchBench: Benchmarking Long-horizon Proactive Search in the Wild
Xiaohongshu Inc
cs.CLcs.AIarXiv:2605.27882v22026Summaries:한국어Generation and Comprehension of Unambiguous Object Descriptions
Junhua Mao, Jonathan Huang, Alexander Toshev +3
cs.CVcs.CLcs.LGarXiv:1511.02283v32015GUI-CIDER: Mid-training GUI Agents via Causal Internalization and Density-aware Exemplar Reselection
Zheng Wu, Chengcheng Han, Zhengxi Lu +5
cs.CLarXiv:2605.28534v12026Rethinking Memory as Continuously Evolving Connectivity
Jizhan Fang, Buqiang Xu, Zhixian Wang +12
cs.CLcs.AIcs.LGarXiv:2605.28773v12026When Confidence Misleads: Suffix Anchoring and Anchor-Proximity Confidence Modulation for Diffusion Language Models
Jungwon Park, Jimyeong Kim, Jungmin Ko +2
cs.CLarXiv:2605.28181v22026Show, Don't TELL: Explainable AI-Generated Text Detection
Aldan Creo, Suraj Ranganath
cs.AIcs.CLcs.CYarXiv:2605.27921v12026Not what you've signed up for: Compromising Real-World LLM-Integrated Applications with Indirect Prompt Injection
Kai Greshake, Sahar Abdelnabi, Shailesh Mishra +3
cs.CRcs.AIcs.CLarXiv:2302.12173v22023AI, Take the Wheel: What Drives Delegation and Trust in Human-Computer Cooperative Question Answering?
Maharshi Gor, Yoo Yeon Sung, Yu Hou +4
cs.AIcs.CLcs.HCarXiv:2605.28255v12026Same Question, Different Source, Different Answer: Auditing Source-Dependence in Medical Multi-Source RAG
Yubo Li, Rema Padman, Ramayya Krishnan
cs.CLcs.AIcs.IRarXiv:2605.29084v12026Deep Anomaly Detection with Outlier Exposure
Dan Hendrycks, Mantas Mazeika, Thomas Dietterich
cs.LGcs.CLcs.CVarXiv:1812.04606v32018Reducing Political Manipulation with Consistency Training
Long Phan, Devin Kim, Alexander Pan +3
cs.CLcs.AIarXiv:2605.22771v22026Sequence Level Training with Recurrent Neural Networks
Marc'Aurelio Ranzato, Sumit Chopra, Michael Auli +1
cs.LGcs.CLarXiv:1511.06732v72015RealToxicityPrompts: Evaluating Neural Toxic Degeneration in Language Models
Samuel Gehman, Suchin Gururangan, Maarten Sap +2
cs.CLarXiv:2009.11462v22020Building End-To-End Dialogue Systems Using Generative Hierarchical Neural Network Models
Iulian V. Serban, Alessandro Sordoni, Yoshua Bengio +2
cs.CLcs.AIcs.LGarXiv:1507.04808v32015SmoothQuant: Accurate and Efficient Post-Training Quantization for Large Language Models
Guangxuan Xiao, Ji Lin, Mickael Seznec +3
cs.CLcs.AIcs.LGarXiv:2211.10438v72022Towards Verifiable Multimodal Deep Research: A Multi-Agent Harness for Interleaved Report Generation
Chenghao Zhang, Guanting Dong, Yufan Liu +3
cs.CLcs.AIarXiv:2605.29861v22026LoMo: Local Modality Substitution for Deeper Vision-Language Fusion
Feng Han, Zhixiong Zhang, Zheming Liang +2
cs.CVcs.CLarXiv:2605.30265v12026OmniRetrieval: Unified Retrieval across Heterogeneous Knowledge Sources
Jinheon Baek, Soyeong Jeong, Sangwoo Park +5
cs.CLcs.AIcs.IRarXiv:2605.29250v12026A Neural Conversational Model
Oriol Vinyals, Quoc Le
cs.CLarXiv:1506.05869v32015Modeling Context in Referring Expressions
Licheng Yu, Patrick Poirson, Shan Yang +2
cs.CVcs.CLarXiv:1608.00272v32016A Multitask, Multilingual, Multimodal Evaluation of ChatGPT on Reasoning, Hallucination, and Interactivity
Yejin Bang, Samuel Cahyawijaya, Nayeon Lee +10
cs.CLcs.AIarXiv:2302.04023v42023Language (Technology) is Power: A Critical Survey of "Bias" in NLP
Su Lin Blodgett, Solon Barocas, Hal Daumé +1
cs.CLcs.CYarXiv:2005.14050v22020REPOT: Recoverable Program-of-Thought via Checkpoint Repair
Parsa Mazaheri
cs.SEcs.AIcs.CLarXiv:2605.30052v12026Word Translation Without Parallel Data
Alexis Conneau, Guillaume Lample, Marc'Aurelio Ranzato +2
cs.CLarXiv:1710.04087v32017WorldMemArena: Evaluating Multimodal Agent Memory Through Action-World Interaction
Chengzhi Liu, Yuzhe Yang, Sophia Xiao Pu +14
cs.CVcs.CLarXiv:2605.29341v22026UniSteer: Text-Guided Flow Matching in Activation Space for Versatile LLM Steering
Yingdong Shi, Ruiming Zhang, Changming Li +4
cs.CLarXiv:2605.30076v12026LongDS-Bench: On the Failure of Long-Horizon Agentic Data Analysis
Kewei Xu, Xiaoben Lu, Shuofei Qiao +4
cs.LGcs.AIcs.CLarXiv:2605.30434v12026MAAT: Multi-phase Adapter-Aware Targeted Unlearning
Suryash Yagnik, Shubham Gaur, Saksham Thakur +3
cs.LGcs.CLarXiv:2605.30514v12026PubMedQA: A Dataset for Biomedical Research Question Answering
Qiao Jin, Bhuwan Dhingra, Zhengping Liu +2
cs.CLcs.LGq-bio.QMarXiv:1909.06146v12019