Computation and Language

Papers filed under cs.CL on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

1,921 to 1,980 of 11,259

  1. QuRating: Selecting High-Quality Data for Training Language Models

    Alexander Wettig, Aatmik Gupta, Saumya Malik +1

    cs.CLcs.LGarXiv:2402.09739v32024
  2. Large Language Models Meet NLP: A Survey

    Libo Qin, Qiguang Chen, Xiachong Feng +6

    cs.CLcs.AIarXiv:2405.12819v22024
  3. FiMI Banking: A Sovereign Model for Indian Retail Banking

    NPCI AI Research Team, Aman Kumar, Asit Desai +15

    cs.AIcs.CLarXiv:2609.03960v12026
  4. SpellGCN: Incorporating Phonological and Visual Similarities into Language Models for Chinese Spelling Check

    Xingyi Cheng, Weidi Xu, Kunlong Chen +5

    cs.CLarXiv:2004.14166v22020
  5. More Criticism Does Not Make a Better Review: EquiReview-R

    Zexing Zhang, Jichao Li, Tianyang Lei +2

    cs.AIcs.CLarXiv:2609.03943v12026
  6. A Comprehensive Survey of Hallucination in Large Language, Image, Video and Audio Foundation Models

    Pranab Sahoo, Prabhash Meharia, Akash Ghosh +3

    cs.LGcs.AIcs.CLarXiv:2405.09589v42024
  7. Speak for Me: Giving LLMs the Situational Awareness to Participate in a Meeting

    Muneeb Khan, Frederic Kirstein, Terry Ruas +1

    cs.AIcs.CLarXiv:2609.03923v12026
  8. Dive into Deep Learning

    Aston Zhang, Zachary C. Lipton, Mu Li +1

    cs.LGcs.AIcs.CLarXiv:2106.11342v52021
    Summaries:한국어
  9. INTENT-AS-A-TOOL Makes it Easy to Track Agentic Misalignment

    Yutong Zhang, Jianshuo Dong, Peng Xu +5

    cs.CLarXiv:2608.27348v12026
  10. GPTEval: A Survey on Assessments of ChatGPT and GPT-4

    Rui Mao, Guanyi Chen, Xulang Zhang +2

    cs.AIcs.CLarXiv:2308.12488v22023
  11. The Wisdom of Polarized Crowds

    Feng Shi, Misha Teplitskiy, Eamon Duede +1

    cs.SIcs.CLcs.CYarXiv:1712.06414v12017
  12. Very Deep Self-Attention Networks for End-to-End Speech Recognition

    Ngoc-Quan Pham, Thai-Son Nguyen, Jan Niehues +3

    cs.CLcs.LGcs.SDarXiv:1904.13377v22019
  13. Retrieval Augmented Generation or Long-Context LLMs? A Comprehensive Study and Hybrid Approach

    Zhuowan Li, Cheng Li, Mingyang Zhang +2

    cs.CLcs.AIcs.LGarXiv:2407.16833v22024
  14. Unlimiformer: Long-Range Transformers with Unlimited Length Input

    Amanda Bertsch, Uri Alon, Graham Neubig +1

    cs.CLarXiv:2305.01625v32023
  15. HalluPeer: A Taxonomy-driven Benchmark for Detecting Hallucinations in Scientific Peer Reviews

    Tzu-Ling Lin, Dong-Ting Yao, Teng-Fang Hsiao +2

    cs.AIcs.CLarXiv:2609.03580v12026
  16. Knowledge Graph-Augmented Abstractive Summarization with Semantic-Driven Cloze Reward

    Luyang Huang, Lingfei Wu, Lu Wang

    cs.CLarXiv:2005.01159v12020
  17. Interventional Video Grounding with Dual Contrastive Learning

    Guoshun Nan, Rui Qiao, Yao Xiao +4

    cs.CVcs.CLarXiv:2106.11013v22021
  18. Calibrating Sequence likelihood Improves Conditional Language Generation

    Yao Zhao, Misha Khalman, Rishabh Joshi +3

    cs.CLarXiv:2210.00045v12022
  19. A Novel Graph-based Multi-modal Fusion Encoder for Neural Machine Translation

    Yongjing Yin, Fandong Meng, Jinsong Su +4

    cs.CLarXiv:2007.08742v12020
  20. ReCode: Robustness Evaluation of Code Generation Models

    Shiqi Wang, Zheng Li, Haifeng Qian +11

    cs.LGcs.CLcs.SEarXiv:2212.10264v12022
  21. Distilled Feature Fields Enable Few-Shot Language-Guided Manipulation

    William Shen, Ge Yang, Alan Yu +3

    cs.CVcs.AIcs.CLarXiv:2308.07931v22023
  22. Introducing the VoicePrivacy Initiative

    Natalia Tomashenko, Brij Mohan Lal Srivastava, Xin Wang +8

    cs.CLarXiv:2005.01387v32020
  23. Is ChatGPT a Highly Fluent Grammatical Error Correction System? A Comprehensive Evaluation

    Tao Fang, Shu Yang, Kaixin Lan +4

    cs.CLarXiv:2304.01746v12023
  24. PyramidInfer: Pyramid KV Cache Compression for High-throughput LLM Inference

    Dongjie Yang, XiaoDong Han, Yan Gao +3

    cs.CLarXiv:2405.12532v22024
  25. Self-training Improves Pre-training for Natural Language Understanding

    Jingfei Du, Edouard Grave, Beliz Gunel +5

    cs.CLarXiv:2010.02194v12020
  26. With Little Power Comes Great Responsibility

    Dallas Card, Peter Henderson, Urvashi Khandelwal +3

    cs.CLcs.AIcs.LGarXiv:2010.06595v12020
  27. Topic Modelling Meets Deep Neural Networks: A Survey

    He Zhao, Dinh Phung, Viet Huynh +3

    cs.LGcs.CLcs.IRarXiv:2103.00498v12021
  28. LaMini-LM: A Diverse Herd of Distilled Models from Large-Scale Instructions

    Minghao Wu, Abdul Waheed, Chiyu Zhang +2

    cs.CLarXiv:2304.14402v32023
  29. Multi-step Retriever-Reader Interaction for Scalable Open-domain Question Answering

    Rajarshi Das, Shehzaad Dhuliawala, Manzil Zaheer +1

    cs.CLcs.LGarXiv:1905.05733v12019
  30. CommonLID: Re-evaluating State-of-the-Art Language Identification Performance on Web Data

    Pedro Ortiz Suarez, Laurie Burchell, Catherine Arnett +94

    cs.CLarXiv:2601.18026v22026
  31. DOBF: A Deobfuscation Pre-Training Objective for Programming Languages

    Baptiste Roziere, Marie-Anne Lachaux, Marc Szafraniec +1

    cs.CLarXiv:2102.07492v32021
  32. TableBench: A Comprehensive and Complex Benchmark for Table Question Answering

    Xianjie Wu, Jian Yang, Linzheng Chai +10

    cs.CLarXiv:2408.09174v22024
  33. Global-to-local Memory Pointer Networks for Task-Oriented Dialogue

    Chien-Sheng Wu, Richard Socher, Caiming Xiong

    cs.CLcs.AIarXiv:1901.04713v22019
  34. Panoptic Scene Graph Generation

    Jingkang Yang, Yi Zhe Ang, Zujin Guo +3

    cs.CVcs.AIcs.CLarXiv:2207.11247v12022
  35. MAmmoTH2: Scaling Instructions from the Web

    Xiang Yue, Tuney Zheng, Ge Zhang +1

    cs.CLarXiv:2405.03548v42024
  36. RegMix: Data Mixture as Regression for Language Model Pre-training

    Qian Liu, Xiaosen Zheng, Niklas Muennighoff +5

    cs.CLcs.AIarXiv:2407.01492v22024
  37. Visual Pivoting for (Unsupervised) Entity Alignment

    Fangyu Liu, Muhao Chen, Dan Roth +1

    cs.CLcs.AIarXiv:2009.13603v22020
  38. The Illusion of Insight in Reasoning Models

    Liv G. d'Aliberti, Manoel Horta Ribeiro

    cs.AIcs.CLarXiv:2601.00514v22026
  39. SentiBench - a benchmark comparison of state-of-the-practice sentiment analysis methods

    Filipe Nunes Ribeiro, Matheus Araújo, Pollyanna Gonçalves +2

    cs.CLcs.SIarXiv:1512.01818v52015
  40. Lift Yourself Up: Retrieval-augmented Text Generation with Self Memory

    Xin Cheng, Di Luo, Xiuying Chen +3

    cs.CLcs.AIarXiv:2305.02437v32023
  41. EvoPrompting: Language Models for Code-Level Neural Architecture Search

    Angelica Chen, David M. Dohan, David R. So

    cs.NEcs.AIcs.CLarXiv:2302.14838v32023
  42. Transformers with convolutional context for ASR

    Abdelrahman Mohamed, Dmytro Okhonko, Luke Zettlemoyer

    cs.CLarXiv:1904.11660v22019
  43. ETHOS: an Online Hate Speech Detection Dataset

    Ioannis Mollas, Zoe Chrysopoulou, Stamatis Karlos +1

    cs.CLcs.LGstat.MLarXiv:2006.08328v22020
  44. Saturated Transformers are Constant-Depth Threshold Circuits

    William Merrill, Ashish Sabharwal, Noah A. Smith

    cs.CLcs.CCcs.LGarXiv:2106.16213v32021
  45. EdgeBERT: Sentence-Level Energy Optimizations for Latency-Aware Multi-Task NLP Inference

    Thierry Tambe, Coleman Hooper, Lillian Pentecost +8

    cs.ARcs.CLarXiv:2011.14203v52020
  46. Beyond Chinchilla-Optimal: Accounting for Inference in Language Model Scaling Laws

    Nikhil Sardana, Jacob Portes, Sasha Doubov +1

    cs.LGcs.CLarXiv:2401.00448v32023
  47. Emergent Communication through Negotiation

    Kris Cao, Angeliki Lazaridou, Marc Lanctot +3

    cs.AIcs.CLcs.LGarXiv:1804.03980v12018
  48. Cross-lingual Prompting: Improving Zero-shot Chain-of-Thought Reasoning across Languages

    Libo Qin, Qiguang Chen, Fuxuan Wei +2

    cs.CLcs.AIarXiv:2310.14799v12023
  49. LongCoder: A Long-Range Pre-trained Language Model for Code Completion

    Daya Guo, Canwen Xu, Nan Duan +2

    cs.SEcs.AIcs.CLarXiv:2306.14893v12023
  50. Liquid Structural State-Space Models

    Ramin Hasani, Mathias Lechner, Tsun-Hsuan Wang +3

    cs.LGcs.AIcs.CLarXiv:2209.12951v12022
  51. Improving Neural Machine Translation with Conditional Sequence Generative Adversarial Nets

    Zhen Yang, Wei Chen, Feng Wang +1

    cs.CLarXiv:1703.04887v42017
  52. MuirBench: A Comprehensive Benchmark for Robust Multi-image Understanding

    Fei Wang, Xingyu Fu, James Y. Huang +18

    cs.CVcs.AIcs.CLarXiv:2406.09411v22024
  53. Visual Coreference Resolution in Visual Dialog using Neural Module Networks

    Satwik Kottur, José M. F. Moura, Devi Parikh +2

    cs.CVcs.AIcs.CLarXiv:1809.01816v12018
  54. AGQA: A Benchmark for Compositional Spatio-Temporal Reasoning

    Madeleine Grunde-McLaughlin, Ranjay Krishna, Maneesh Agrawala

    cs.CVcs.CLarXiv:2103.16002v12021
  55. Stack-Pointer Networks for Dependency Parsing

    Xuezhe Ma, Zecong Hu, Jingzhou Liu +3

    cs.CLcs.LGarXiv:1805.01087v12018
  56. Joint Reasoning for Temporal and Causal Relations

    Qiang Ning, Zhili Feng, Hao Wu +1

    cs.CLcs.AIcs.IRarXiv:1906.04941v12019
  57. Towards Exploiting Background Knowledge for Building Conversation Systems

    Nikita Moghe, Siddhartha Arora, Suman Banerjee +1

    cs.CLarXiv:1809.08205v12018
  58. Revolutionizing Finance with LLMs: An Overview of Applications and Insights

    Huaqin Zhao, Zhengliang Liu, Zihao Wu +16

    cs.CLarXiv:2401.11641v52024
  59. A Survey on Retrieval-Augmented Text Generation for Large Language Models

    Yizheng Huang, Jimmy Huang

    cs.IRcs.AIcs.CLarXiv:2404.10981v22024
  60. CRISPR-GPT for Agentic Automation of Gene-editing Experiments

    Yuanhao Qu, Kaixuan Huang, Ming Yin +11

    cs.AIcs.CLcs.HCarXiv:2404.18021v22024