Computation and Language
Papers filed under cs.CL on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
1,861 to 1,920 of 11,226
Leveraging Weakly Supervised Data to Improve End-to-End Speech-to-Text Translation
Ye Jia, Melvin Johnson, Wolfgang Macherey +6
cs.CLcs.LGcs.SDarXiv:1811.02050v22018TaskBench: Benchmarking Large Language Models for Task Automation
Yongliang Shen, Kaitao Song, Xu Tan +6
cs.CLcs.AIarXiv:2311.18760v42023Pattern Over-Generalization of Knowledge Graph Embedding
Junsik Kim, Kangil Kim
cs.CLcs.AIarXiv:2609.03487v12026CMB: A Comprehensive Medical Benchmark in Chinese
Xidong Wang, Guiming Hardy Chen, Dingjie Song +8
cs.CLcs.AIarXiv:2308.08833v22023LLM-in-the-loop: Leveraging Large Language Model for Thematic Analysis
Shih-Chieh Dai, Aiping Xiong, Lun-Wei Ku
cs.CLarXiv:2310.15100v12023Building and Evaluating Fixed-Voice Thai TTS from Synthetic Speech
Kunat Pipatanakul, Potsawee Manakul, Warit Sirichotedumrong +3
cs.CLcs.AIarXiv:2609.03502v12026DISC-LawLLM: Fine-tuning Large Language Models for Intelligent Legal Services
Shengbin Yue, Wei Chen, Siyuan Wang +8
cs.CLarXiv:2309.11325v22023When Users Don't Ask: Benchmarking Context-Driven Memory Retrieval in Conversational Agents
Wen-Yu Chang, Yun-Nung Chen
cs.CLcs.AIarXiv:2609.03467v12026Language Models Meet World Models: Embodied Experiences Enhance Language Models
Jiannan Xiang, Tianhua Tao, Yi Gu +4
cs.CLcs.AIcs.LGarXiv:2305.10626v32023"Why is 'Chicago' deceptive?" Towards Building Model-Driven Tutorials for Humans
Vivian Lai, Han Liu, Chenhao Tan
cs.HCcs.AIcs.CLarXiv:2001.05871v12020Everything at Once -- Multi-modal Fusion Transformer for Video Retrieval
Nina Shvetsova, Brian Chen, Andrew Rouditchenko +6
cs.CVcs.CLcs.SDarXiv:2112.04446v22021Plan Pointers and Record-Directive Form in Budgeted Verification of Inherited Agent Memory
Kazuki Nakayashiki
cs.IRcs.AIcs.CLarXiv:2609.03450v12026It's the Problem, Not the Path: Budget and Difficulty Confounds in LLM Reasoning Trajectories
Yigit Utku Bulut
cs.LGcs.AIcs.CLarXiv:2609.03436v12026VALSE: A Task-Independent Benchmark for Vision and Language Models Centered on Linguistic Phenomena
Letitia Parcalabescu, Michele Cafagna, Lilitta Muradjan +3
cs.CLcs.CVarXiv:2112.07566v22021GrIPS: Gradient-free, Edit-based Instruction Search for Prompting Large Language Models
Archiki Prasad, Peter Hase, Xiang Zhou +1
cs.CLcs.AIcs.LGarXiv:2203.07281v22022TabScope: Question-Adaptive Scope Selection for Table Question Answering
Yuxiang Wang, Junhao Gan, Jianzhong Qi
cs.CLcs.AIarXiv:2609.03395v12026End-to-End Speech Translation with Knowledge Distillation
Yuchen Liu, Hao Xiong, Zhongjun He +4
cs.CLarXiv:1904.08075v12019Android in the Zoo: Chain-of-Action-Thought for GUI Agents
Jiwen Zhang, Jihao Wu, Yihua Teng +5
cs.CLcs.CVcs.HCarXiv:2403.02713v22024ESimCSE: Enhanced Sample Building Method for Contrastive Learning of Unsupervised Sentence Embedding
Xing Wu, Chaochen Gao, Liangjun Zang +3
cs.CLcs.AIarXiv:2109.04380v22021PutnamBench: Evaluating Neural Theorem-Provers on the Putnam Mathematical Competition
George Tsoukalas, Jasper Lee, John Jennings +5
cs.AIcs.CLcs.LGarXiv:2407.11214v22024Learning-Based Single-Document Summarization with Compression and Anaphoricity Constraints
Greg Durrett, Taylor Berg-Kirkpatrick, Dan Klein
cs.CLarXiv:1603.08887v22016Mimicking Word Embeddings using Subword RNNs
Yuval Pinter, Robert Guthrie, Jacob Eisenstein
cs.CLarXiv:1707.06961v12017Measuring Compositionality in Representation Learning
Jacob Andreas
cs.LGcs.CLstat.MLarXiv:1902.07181v22019Instruction Duplication as an Inference-Time Control Primitive
Victor Lavrenko
cs.AIcs.CLarXiv:2609.04024v12026An Incremental Parser for Abstract Meaning Representation
Marco Damonte, Shay B. Cohen, Giorgio Satta
cs.CLarXiv:1608.06111v52016SuS-X: Training-Free Name-Only Transfer of Vision-Language Models
Vishaal Udandarao, Ankush Gupta, Samuel Albanie
cs.CVcs.CLcs.MMarXiv:2211.16198v42022Semi-supervised User Geolocation via Graph Convolutional Networks
Afshin Rahimi, Trevor Cohn, Timothy Baldwin
cs.CLarXiv:1804.08049v42018QuRating: Selecting High-Quality Data for Training Language Models
Alexander Wettig, Aatmik Gupta, Saumya Malik +1
cs.CLcs.LGarXiv:2402.09739v32024Large Language Models Meet NLP: A Survey
Libo Qin, Qiguang Chen, Xiachong Feng +6
cs.CLcs.AIarXiv:2405.12819v22024FiMI Banking: A Sovereign Model for Indian Retail Banking
NPCI AI Research Team, Aman Kumar, Asit Desai +15
cs.AIcs.CLarXiv:2609.03960v12026SpellGCN: Incorporating Phonological and Visual Similarities into Language Models for Chinese Spelling Check
Xingyi Cheng, Weidi Xu, Kunlong Chen +5
cs.CLarXiv:2004.14166v22020More Criticism Does Not Make a Better Review: EquiReview-R
Zexing Zhang, Jichao Li, Tianyang Lei +2
cs.AIcs.CLarXiv:2609.03943v12026A Comprehensive Survey of Hallucination in Large Language, Image, Video and Audio Foundation Models
Pranab Sahoo, Prabhash Meharia, Akash Ghosh +3
cs.LGcs.AIcs.CLarXiv:2405.09589v42024Speak for Me: Giving LLMs the Situational Awareness to Participate in a Meeting
Muneeb Khan, Frederic Kirstein, Terry Ruas +1
cs.AIcs.CLarXiv:2609.03923v12026Dive into Deep Learning
Aston Zhang, Zachary C. Lipton, Mu Li +1
cs.LGcs.AIcs.CLarXiv:2106.11342v52021Summaries:한국어INTENT-AS-A-TOOL Makes it Easy to Track Agentic Misalignment
Yutong Zhang, Jianshuo Dong, Peng Xu +5
cs.CLarXiv:2608.27348v12026GPTEval: A Survey on Assessments of ChatGPT and GPT-4
Rui Mao, Guanyi Chen, Xulang Zhang +2
cs.AIcs.CLarXiv:2308.12488v22023The Wisdom of Polarized Crowds
Feng Shi, Misha Teplitskiy, Eamon Duede +1
cs.SIcs.CLcs.CYarXiv:1712.06414v12017Very Deep Self-Attention Networks for End-to-End Speech Recognition
Ngoc-Quan Pham, Thai-Son Nguyen, Jan Niehues +3
cs.CLcs.LGcs.SDarXiv:1904.13377v22019Retrieval Augmented Generation or Long-Context LLMs? A Comprehensive Study and Hybrid Approach
Zhuowan Li, Cheng Li, Mingyang Zhang +2
cs.CLcs.AIcs.LGarXiv:2407.16833v22024Unlimiformer: Long-Range Transformers with Unlimited Length Input
Amanda Bertsch, Uri Alon, Graham Neubig +1
cs.CLarXiv:2305.01625v32023HalluPeer: A Taxonomy-driven Benchmark for Detecting Hallucinations in Scientific Peer Reviews
Tzu-Ling Lin, Dong-Ting Yao, Teng-Fang Hsiao +2
cs.AIcs.CLarXiv:2609.03580v12026Knowledge Graph-Augmented Abstractive Summarization with Semantic-Driven Cloze Reward
Luyang Huang, Lingfei Wu, Lu Wang
cs.CLarXiv:2005.01159v12020Interventional Video Grounding with Dual Contrastive Learning
Guoshun Nan, Rui Qiao, Yao Xiao +4
cs.CVcs.CLarXiv:2106.11013v22021Calibrating Sequence likelihood Improves Conditional Language Generation
Yao Zhao, Misha Khalman, Rishabh Joshi +3
cs.CLarXiv:2210.00045v12022A Novel Graph-based Multi-modal Fusion Encoder for Neural Machine Translation
Yongjing Yin, Fandong Meng, Jinsong Su +4
cs.CLarXiv:2007.08742v12020ReCode: Robustness Evaluation of Code Generation Models
Shiqi Wang, Zheng Li, Haifeng Qian +11
cs.LGcs.CLcs.SEarXiv:2212.10264v12022Distilled Feature Fields Enable Few-Shot Language-Guided Manipulation
William Shen, Ge Yang, Alan Yu +3
cs.CVcs.AIcs.CLarXiv:2308.07931v22023Introducing the VoicePrivacy Initiative
Natalia Tomashenko, Brij Mohan Lal Srivastava, Xin Wang +8
cs.CLarXiv:2005.01387v32020Is ChatGPT a Highly Fluent Grammatical Error Correction System? A Comprehensive Evaluation
Tao Fang, Shu Yang, Kaixin Lan +4
cs.CLarXiv:2304.01746v12023PyramidInfer: Pyramid KV Cache Compression for High-throughput LLM Inference
Dongjie Yang, XiaoDong Han, Yan Gao +3
cs.CLarXiv:2405.12532v22024Self-training Improves Pre-training for Natural Language Understanding
Jingfei Du, Edouard Grave, Beliz Gunel +5
cs.CLarXiv:2010.02194v12020With Little Power Comes Great Responsibility
Dallas Card, Peter Henderson, Urvashi Khandelwal +3
cs.CLcs.AIcs.LGarXiv:2010.06595v12020Topic Modelling Meets Deep Neural Networks: A Survey
He Zhao, Dinh Phung, Viet Huynh +3
cs.LGcs.CLcs.IRarXiv:2103.00498v12021LaMini-LM: A Diverse Herd of Distilled Models from Large-Scale Instructions
Minghao Wu, Abdul Waheed, Chiyu Zhang +2
cs.CLarXiv:2304.14402v32023Multi-step Retriever-Reader Interaction for Scalable Open-domain Question Answering
Rajarshi Das, Shehzaad Dhuliawala, Manzil Zaheer +1
cs.CLcs.LGarXiv:1905.05733v12019CommonLID: Re-evaluating State-of-the-Art Language Identification Performance on Web Data
Pedro Ortiz Suarez, Laurie Burchell, Catherine Arnett +94
cs.CLarXiv:2601.18026v22026DOBF: A Deobfuscation Pre-Training Objective for Programming Languages
Baptiste Roziere, Marie-Anne Lachaux, Marc Szafraniec +1
cs.CLarXiv:2102.07492v32021TableBench: A Comprehensive and Complex Benchmark for Table Question Answering
Xianjie Wu, Jian Yang, Linzheng Chai +10
cs.CLarXiv:2408.09174v22024Global-to-local Memory Pointer Networks for Task-Oriented Dialogue
Chien-Sheng Wu, Richard Socher, Caiming Xiong
cs.CLcs.AIarXiv:1901.04713v22019