Computation and Language

Papers filed under cs.CL on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

1,861 to 1,920 of 11,226

  1. Leveraging Weakly Supervised Data to Improve End-to-End Speech-to-Text Translation

    Ye Jia, Melvin Johnson, Wolfgang Macherey +6

    cs.CLcs.LGcs.SDarXiv:1811.02050v22018
  2. TaskBench: Benchmarking Large Language Models for Task Automation

    Yongliang Shen, Kaitao Song, Xu Tan +6

    cs.CLcs.AIarXiv:2311.18760v42023
  3. Pattern Over-Generalization of Knowledge Graph Embedding

    Junsik Kim, Kangil Kim

    cs.CLcs.AIarXiv:2609.03487v12026
  4. CMB: A Comprehensive Medical Benchmark in Chinese

    Xidong Wang, Guiming Hardy Chen, Dingjie Song +8

    cs.CLcs.AIarXiv:2308.08833v22023
  5. LLM-in-the-loop: Leveraging Large Language Model for Thematic Analysis

    Shih-Chieh Dai, Aiping Xiong, Lun-Wei Ku

    cs.CLarXiv:2310.15100v12023
  6. Building and Evaluating Fixed-Voice Thai TTS from Synthetic Speech

    Kunat Pipatanakul, Potsawee Manakul, Warit Sirichotedumrong +3

    cs.CLcs.AIarXiv:2609.03502v12026
  7. DISC-LawLLM: Fine-tuning Large Language Models for Intelligent Legal Services

    Shengbin Yue, Wei Chen, Siyuan Wang +8

    cs.CLarXiv:2309.11325v22023
  8. When Users Don't Ask: Benchmarking Context-Driven Memory Retrieval in Conversational Agents

    Wen-Yu Chang, Yun-Nung Chen

    cs.CLcs.AIarXiv:2609.03467v12026
  9. Language Models Meet World Models: Embodied Experiences Enhance Language Models

    Jiannan Xiang, Tianhua Tao, Yi Gu +4

    cs.CLcs.AIcs.LGarXiv:2305.10626v32023
  10. "Why is 'Chicago' deceptive?" Towards Building Model-Driven Tutorials for Humans

    Vivian Lai, Han Liu, Chenhao Tan

    cs.HCcs.AIcs.CLarXiv:2001.05871v12020
  11. Everything at Once -- Multi-modal Fusion Transformer for Video Retrieval

    Nina Shvetsova, Brian Chen, Andrew Rouditchenko +6

    cs.CVcs.CLcs.SDarXiv:2112.04446v22021
  12. Plan Pointers and Record-Directive Form in Budgeted Verification of Inherited Agent Memory

    Kazuki Nakayashiki

    cs.IRcs.AIcs.CLarXiv:2609.03450v12026
  13. It's the Problem, Not the Path: Budget and Difficulty Confounds in LLM Reasoning Trajectories

    Yigit Utku Bulut

    cs.LGcs.AIcs.CLarXiv:2609.03436v12026
  14. VALSE: A Task-Independent Benchmark for Vision and Language Models Centered on Linguistic Phenomena

    Letitia Parcalabescu, Michele Cafagna, Lilitta Muradjan +3

    cs.CLcs.CVarXiv:2112.07566v22021
  15. GrIPS: Gradient-free, Edit-based Instruction Search for Prompting Large Language Models

    Archiki Prasad, Peter Hase, Xiang Zhou +1

    cs.CLcs.AIcs.LGarXiv:2203.07281v22022
  16. TabScope: Question-Adaptive Scope Selection for Table Question Answering

    Yuxiang Wang, Junhao Gan, Jianzhong Qi

    cs.CLcs.AIarXiv:2609.03395v12026
  17. End-to-End Speech Translation with Knowledge Distillation

    Yuchen Liu, Hao Xiong, Zhongjun He +4

    cs.CLarXiv:1904.08075v12019
  18. Android in the Zoo: Chain-of-Action-Thought for GUI Agents

    Jiwen Zhang, Jihao Wu, Yihua Teng +5

    cs.CLcs.CVcs.HCarXiv:2403.02713v22024
  19. ESimCSE: Enhanced Sample Building Method for Contrastive Learning of Unsupervised Sentence Embedding

    Xing Wu, Chaochen Gao, Liangjun Zang +3

    cs.CLcs.AIarXiv:2109.04380v22021
  20. PutnamBench: Evaluating Neural Theorem-Provers on the Putnam Mathematical Competition

    George Tsoukalas, Jasper Lee, John Jennings +5

    cs.AIcs.CLcs.LGarXiv:2407.11214v22024
  21. Learning-Based Single-Document Summarization with Compression and Anaphoricity Constraints

    Greg Durrett, Taylor Berg-Kirkpatrick, Dan Klein

    cs.CLarXiv:1603.08887v22016
  22. Mimicking Word Embeddings using Subword RNNs

    Yuval Pinter, Robert Guthrie, Jacob Eisenstein

    cs.CLarXiv:1707.06961v12017
  23. Measuring Compositionality in Representation Learning

    Jacob Andreas

    cs.LGcs.CLstat.MLarXiv:1902.07181v22019
  24. Instruction Duplication as an Inference-Time Control Primitive

    Victor Lavrenko

    cs.AIcs.CLarXiv:2609.04024v12026
  25. An Incremental Parser for Abstract Meaning Representation

    Marco Damonte, Shay B. Cohen, Giorgio Satta

    cs.CLarXiv:1608.06111v52016
  26. SuS-X: Training-Free Name-Only Transfer of Vision-Language Models

    Vishaal Udandarao, Ankush Gupta, Samuel Albanie

    cs.CVcs.CLcs.MMarXiv:2211.16198v42022
  27. Semi-supervised User Geolocation via Graph Convolutional Networks

    Afshin Rahimi, Trevor Cohn, Timothy Baldwin

    cs.CLarXiv:1804.08049v42018
  28. QuRating: Selecting High-Quality Data for Training Language Models

    Alexander Wettig, Aatmik Gupta, Saumya Malik +1

    cs.CLcs.LGarXiv:2402.09739v32024
  29. Large Language Models Meet NLP: A Survey

    Libo Qin, Qiguang Chen, Xiachong Feng +6

    cs.CLcs.AIarXiv:2405.12819v22024
  30. FiMI Banking: A Sovereign Model for Indian Retail Banking

    NPCI AI Research Team, Aman Kumar, Asit Desai +15

    cs.AIcs.CLarXiv:2609.03960v12026
  31. SpellGCN: Incorporating Phonological and Visual Similarities into Language Models for Chinese Spelling Check

    Xingyi Cheng, Weidi Xu, Kunlong Chen +5

    cs.CLarXiv:2004.14166v22020
  32. More Criticism Does Not Make a Better Review: EquiReview-R

    Zexing Zhang, Jichao Li, Tianyang Lei +2

    cs.AIcs.CLarXiv:2609.03943v12026
  33. A Comprehensive Survey of Hallucination in Large Language, Image, Video and Audio Foundation Models

    Pranab Sahoo, Prabhash Meharia, Akash Ghosh +3

    cs.LGcs.AIcs.CLarXiv:2405.09589v42024
  34. Speak for Me: Giving LLMs the Situational Awareness to Participate in a Meeting

    Muneeb Khan, Frederic Kirstein, Terry Ruas +1

    cs.AIcs.CLarXiv:2609.03923v12026
  35. Dive into Deep Learning

    Aston Zhang, Zachary C. Lipton, Mu Li +1

    cs.LGcs.AIcs.CLarXiv:2106.11342v52021
    Summaries:한국어
  36. INTENT-AS-A-TOOL Makes it Easy to Track Agentic Misalignment

    Yutong Zhang, Jianshuo Dong, Peng Xu +5

    cs.CLarXiv:2608.27348v12026
  37. GPTEval: A Survey on Assessments of ChatGPT and GPT-4

    Rui Mao, Guanyi Chen, Xulang Zhang +2

    cs.AIcs.CLarXiv:2308.12488v22023
  38. The Wisdom of Polarized Crowds

    Feng Shi, Misha Teplitskiy, Eamon Duede +1

    cs.SIcs.CLcs.CYarXiv:1712.06414v12017
  39. Very Deep Self-Attention Networks for End-to-End Speech Recognition

    Ngoc-Quan Pham, Thai-Son Nguyen, Jan Niehues +3

    cs.CLcs.LGcs.SDarXiv:1904.13377v22019
  40. Retrieval Augmented Generation or Long-Context LLMs? A Comprehensive Study and Hybrid Approach

    Zhuowan Li, Cheng Li, Mingyang Zhang +2

    cs.CLcs.AIcs.LGarXiv:2407.16833v22024
  41. Unlimiformer: Long-Range Transformers with Unlimited Length Input

    Amanda Bertsch, Uri Alon, Graham Neubig +1

    cs.CLarXiv:2305.01625v32023
  42. HalluPeer: A Taxonomy-driven Benchmark for Detecting Hallucinations in Scientific Peer Reviews

    Tzu-Ling Lin, Dong-Ting Yao, Teng-Fang Hsiao +2

    cs.AIcs.CLarXiv:2609.03580v12026
  43. Knowledge Graph-Augmented Abstractive Summarization with Semantic-Driven Cloze Reward

    Luyang Huang, Lingfei Wu, Lu Wang

    cs.CLarXiv:2005.01159v12020
  44. Interventional Video Grounding with Dual Contrastive Learning

    Guoshun Nan, Rui Qiao, Yao Xiao +4

    cs.CVcs.CLarXiv:2106.11013v22021
  45. Calibrating Sequence likelihood Improves Conditional Language Generation

    Yao Zhao, Misha Khalman, Rishabh Joshi +3

    cs.CLarXiv:2210.00045v12022
  46. A Novel Graph-based Multi-modal Fusion Encoder for Neural Machine Translation

    Yongjing Yin, Fandong Meng, Jinsong Su +4

    cs.CLarXiv:2007.08742v12020
  47. ReCode: Robustness Evaluation of Code Generation Models

    Shiqi Wang, Zheng Li, Haifeng Qian +11

    cs.LGcs.CLcs.SEarXiv:2212.10264v12022
  48. Distilled Feature Fields Enable Few-Shot Language-Guided Manipulation

    William Shen, Ge Yang, Alan Yu +3

    cs.CVcs.AIcs.CLarXiv:2308.07931v22023
  49. Introducing the VoicePrivacy Initiative

    Natalia Tomashenko, Brij Mohan Lal Srivastava, Xin Wang +8

    cs.CLarXiv:2005.01387v32020
  50. Is ChatGPT a Highly Fluent Grammatical Error Correction System? A Comprehensive Evaluation

    Tao Fang, Shu Yang, Kaixin Lan +4

    cs.CLarXiv:2304.01746v12023
  51. PyramidInfer: Pyramid KV Cache Compression for High-throughput LLM Inference

    Dongjie Yang, XiaoDong Han, Yan Gao +3

    cs.CLarXiv:2405.12532v22024
  52. Self-training Improves Pre-training for Natural Language Understanding

    Jingfei Du, Edouard Grave, Beliz Gunel +5

    cs.CLarXiv:2010.02194v12020
  53. With Little Power Comes Great Responsibility

    Dallas Card, Peter Henderson, Urvashi Khandelwal +3

    cs.CLcs.AIcs.LGarXiv:2010.06595v12020
  54. Topic Modelling Meets Deep Neural Networks: A Survey

    He Zhao, Dinh Phung, Viet Huynh +3

    cs.LGcs.CLcs.IRarXiv:2103.00498v12021
  55. LaMini-LM: A Diverse Herd of Distilled Models from Large-Scale Instructions

    Minghao Wu, Abdul Waheed, Chiyu Zhang +2

    cs.CLarXiv:2304.14402v32023
  56. Multi-step Retriever-Reader Interaction for Scalable Open-domain Question Answering

    Rajarshi Das, Shehzaad Dhuliawala, Manzil Zaheer +1

    cs.CLcs.LGarXiv:1905.05733v12019
  57. CommonLID: Re-evaluating State-of-the-Art Language Identification Performance on Web Data

    Pedro Ortiz Suarez, Laurie Burchell, Catherine Arnett +94

    cs.CLarXiv:2601.18026v22026
  58. DOBF: A Deobfuscation Pre-Training Objective for Programming Languages

    Baptiste Roziere, Marie-Anne Lachaux, Marc Szafraniec +1

    cs.CLarXiv:2102.07492v32021
  59. TableBench: A Comprehensive and Complex Benchmark for Table Question Answering

    Xianjie Wu, Jian Yang, Linzheng Chai +10

    cs.CLarXiv:2408.09174v22024
  60. Global-to-local Memory Pointer Networks for Task-Oriented Dialogue

    Chien-Sheng Wu, Richard Socher, Caiming Xiong

    cs.CLcs.AIarXiv:1901.04713v22019