Computation and Language

Papers filed under cs.CL on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

8,221 to 8,280 of 11,365

  1. Calibration-Preserving Pruning: Compression as a Reliability Contract

    Ibne Farabi Shihab, Adria Binte Habib, Anuj Sharma

    cs.LGcs.CLarXiv:2608.23744v12026
  2. Q-Align: Teaching LMMs for Visual Scoring via Discrete Text-Defined Levels

    Haoning Wu, Zicheng Zhang, Weixia Zhang +11

    cs.CVcs.CLcs.LGarXiv:2312.17090v12023
  3. A Hierarchical Neural Autoencoder for Paragraphs and Documents

    Jiwei Li, Minh-Thang Luong, Dan Jurafsky

    cs.CLarXiv:1506.01057v22015
  4. Mitigating Exploration Bias in RL for Multi-Instruction Following

    Mian Zhang, Yueqin Yin, Kaiyu He +4

    cs.CLcs.LGarXiv:2608.23830v12026
  5. RouteLLM: Learning to Route LLMs with Preference Data

    Isaac Ong, Amjad Almahairi, Vincent Wu +5

    cs.LGcs.AIcs.CLarXiv:2406.18665v42024
  6. PatchWrite: One Line, Not One Section -- Compile-Gated, Validity-Preserving Editing for AI-Drafted Manuscripts

    Weiwei Yang

    cs.AIcs.CLcs.SEarXiv:2608.23001v12026
  7. FedKD: Communication Efficient Federated Learning via Knowledge Distillation

    Chuhan Wu, Fangzhao Wu, Lingjuan Lyu +2

    cs.LGcs.CLarXiv:2108.13323v22021
  8. From Triage to Discharge: A Survey of NLP Tasks, Methods, and Open Challenges in the Emergency Department

    Dipankar Srirag, Aditya Joshi, Salil Kanhere +1

    cs.CLarXiv:2608.23627v12026
  9. Fast Abstractive Summarization with Reinforce-Selected Sentence Rewriting

    Yen-Chun Chen, Mohit Bansal

    cs.CLcs.AIcs.LGarXiv:1805.11080v12018
  10. Inter-dimension Dependence for Multi-Dimensional Evaluation of Open-Ended Text

    Haoyuan Li, Snigdha Chaturvedi

    cs.CLarXiv:2608.23783v12026
  11. YaRN: Efficient Context Window Extension of Large Language Models

    Bowen Peng, Jeffrey Quesnelle, Honglu Fan +1

    cs.CLcs.AIcs.LGarXiv:2309.00071v32023
  12. Towards End-to-End Prosody Transfer for Expressive Speech Synthesis with Tacotron

    RJ Skerry-Ryan, Eric Battenberg, Ying Xiao +6

    cs.CLcs.LGcs.SDarXiv:1803.09047v12018
  13. PAWS: Paraphrase Adversaries from Word Scrambling

    Yuan Zhang, Jason Baldridge, Luheng He

    cs.CLarXiv:1904.01130v12019
  14. MirrorGAN: Learning Text-to-image Generation by Redescription

    Tingting Qiao, Jing Zhang, Duanqing Xu +1

    cs.CLcs.CVcs.LGarXiv:1903.05854v12019
  15. Markets, Not Planners: Decentralized Orchestration of LLM Agents with Private Information

    Xiao Liu, Haoyang Li, Songwei Li +4

    cs.MAcs.CLarXiv:2608.23867v12026
  16. Asking and Answering Questions to Evaluate the Factual Consistency of Summaries

    Alex Wang, Kyunghyun Cho, Mike Lewis

    cs.CLarXiv:2004.04228v12020
  17. SMART: Robust and Efficient Fine-Tuning for Pre-trained Natural Language Models through Principled Regularized Optimization

    Haoming Jiang, Pengcheng He, Weizhu Chen +3

    cs.CLcs.LGmath.OCarXiv:1911.03437v52019
  18. ERNIE 3.0: Large-scale Knowledge Enhanced Pre-training for Language Understanding and Generation

    Yu Sun, Shuohuan Wang, Shikun Feng +19

    cs.CLarXiv:2107.02137v12021
  19. A Survey of Hallucination in Large Foundation Models

    Vipula Rawte, Amit Sheth, Amitava Das

    cs.AIcs.CLcs.IRarXiv:2309.05922v12023
  20. Dataset Cartography: Mapping and Diagnosing Datasets with Training Dynamics

    Swabha Swayamdipta, Roy Schwartz, Nicholas Lourie +4

    cs.CLarXiv:2009.10795v22020
  21. Clotho: An Audio Captioning Dataset

    Konstantinos Drossos, Samuel Lipping, Tuomas Virtanen

    cs.SDcs.CLcs.LGarXiv:1910.09387v12019
  22. Prometheus: Inducing Fine-grained Evaluation Capability in Language Models

    Seungone Kim, Jamin Shin, Yejin Cho +8

    cs.CLcs.LGarXiv:2310.08491v22023
  23. Using Trusted Data to Train Deep Networks on Labels Corrupted by Severe Noise

    Dan Hendrycks, Mantas Mazeika, Duncan Wilson +1

    cs.LGcs.CLcs.CVarXiv:1802.05300v42018
  24. Beyond the Nav-Graph: Vision-and-Language Navigation in Continuous Environments

    Jacob Krantz, Erik Wijmans, Arjun Majumdar +2

    cs.CVcs.CLcs.ROarXiv:2004.02857v22020
  25. Adaptive-RAG: Learning to Adapt Retrieval-Augmented Large Language Models through Question Complexity

    Soyeong Jeong, Jinheon Baek, Sukmin Cho +2

    cs.CLcs.AIarXiv:2403.14403v22024
  26. Weight Poisoning Attacks on Pre-trained Models

    Keita Kurita, Paul Michel, Graham Neubig

    cs.LGcs.CLcs.CRarXiv:2004.06660v12020
  27. Supervised Contrastive Learning for Pre-trained Language Model Fine-tuning

    Beliz Gunel, Jingfei Du, Alexis Conneau +1

    cs.CLcs.LGarXiv:2011.01403v32020
  28. How Can We Know When Language Models Know? On the Calibration of Language Models for Question Answering

    Zhengbao Jiang, Jun Araki, Haibo Ding +1

    cs.CLarXiv:2012.00955v22020
  29. Text and Code Embeddings by Contrastive Pre-Training

    Arvind Neelakantan, Tao Xu, Raul Puri +22

    cs.CLcs.LGarXiv:2201.10005v12022
  30. MemoryBank: Enhancing Large Language Models with Long-Term Memory

    Wanjun Zhong, Lianghong Guo, Qiqi Gao +2

    cs.CLcs.AIarXiv:2305.10250v32023
  31. MLQA: Evaluating Cross-lingual Extractive Question Answering

    Patrick Lewis, Barlas Oğuz, Ruty Rinott +2

    cs.CLcs.AIcs.LGarXiv:1910.07475v32019
  32. SkillForge: Evolving Verifiable Skills for Reinforcement Learning Agents

    Shidong Yang, Ziyu Ma, Tongwen Huang +5

    cs.CLarXiv:2608.24747v12026
  33. Arbitrary Polygon Oscillator: Generalizing Polygonal Synthesis to Arbitrary Shapes, Morphing, and Three-Dimensional Polyhedra

    Antonio Argentieri, Francesco Scagliola

    cs.CLarXiv:2608.24726v12026
  34. One Timeline, Many Renderings: A Wolfram Language Paclet for heterogeneous musical output

    Francesco Vitucci, Michele Lorusso, Francesco Scagliola

    cs.CLarXiv:2608.24683v12026
  35. From local kernels to global form: modeling the emergence of musical content

    Francesco Vitucci, Michele Lorusso, Francesco Scagliola

    cs.CLarXiv:2608.24660v12026
  36. Paritok-4B: Intent-Conditioned Context Compression for Coding Agents

    Jiayu Shi, Luzhuo Chen

    cs.AIcs.CLcs.LGarXiv:2608.24188v12026
  37. Sentiment Analysis Based on Deep Learning: A Comparative Study

    Nhan Cach Dang, María N. Moreno-García, Fernando De la Prieta

    cs.CLcs.IRcs.LGarXiv:2006.03541v12020
  38. GigaSpeech: An Evolving, Multi-domain ASR Corpus with 10,000 Hours of Transcribed Audio

    Guoguo Chen, Shuzhou Chai, Guanbo Wang +18

    cs.SDcs.CLeess.ASarXiv:2106.06909v12021
  39. HelaBERT: Enhancing Sinhala Language Understanding with Dual Pooling Classification Head

    Thisen Ekanayake, Nisansa de Silva

    cs.CLarXiv:2608.22922v12026
  40. Nuanced Metrics for Measuring Unintended Bias with Real Data for Text Classification

    Daniel Borkan, Lucas Dixon, Jeffrey Sorensen +2

    cs.LGcs.CLstat.MLarXiv:1903.04561v22019
  41. MM-REACT: Prompting ChatGPT for Multimodal Reasoning and Action

    Zhengyuan Yang, Linjie Li, Jianfeng Wang +7

    cs.CVcs.CLcs.LGarXiv:2303.11381v12023
  42. The Limits of Automatic Evaluation of Creativity in Large Language Models

    Alessandro Tutone, Giorgio Franceschelli, Mirco Musolesi

    cs.CLcs.AIcs.CYarXiv:2608.23705v12026
  43. The AI Scientist: Towards Fully Automated Open-Ended Scientific Discovery

    Chris Lu, Cong Lu, Robert Tjarko Lange +3

    cs.AIcs.CLcs.LGarXiv:2408.06292v32024
    Summaries:한국어
  44. Can AI-Generated Text be Reliably Detected?

    Vinu Sankar Sadasivan, Aounon Kumar, Sriram Balasubramanian +2

    cs.CLcs.AIcs.LGarXiv:2303.11156v42023
  45. Beyond Static and Linear: What Attention Constraints Best Fit Human Reading Times?

    Lanni Bu, Xiulin Yang, Christian Clark +2

    cs.CLarXiv:2608.23818v12026
  46. A robust self-learning method for fully unsupervised cross-lingual mappings of word embeddings

    Mikel Artetxe, Gorka Labaka, Eneko Agirre

    cs.CLcs.AIcs.LGarXiv:1805.06297v22018
  47. A Source-Grounded Framework for Constructing and Evaluating Progressive Multimodal Diagnostic Dialogues from Clinical Case Reports

    Yufan Wang, Rui Yang, Yi Liu +2

    cs.CLarXiv:2608.22713v12026
  48. Logic-LM: Empowering Large Language Models with Symbolic Solvers for Faithful Logical Reasoning

    Liangming Pan, Alon Albalak, Xinyi Wang +1

    cs.CLcs.AIarXiv:2305.12295v22023
  49. Go for a Walk and Arrive at the Answer: Reasoning Over Paths in Knowledge Bases using Reinforcement Learning

    Rajarshi Das, Shehzaad Dhuliawala, Manzil Zaheer +5

    cs.CLcs.AIarXiv:1711.05851v22017
  50. A Convolutional Attention Network for Extreme Summarization of Source Code

    Miltiadis Allamanis, Hao Peng, Charles Sutton

    cs.LGcs.CLcs.SEarXiv:1602.03001v22016
  51. Sequence-to-Sequence Learning as Beam-Search Optimization

    Sam Wiseman, Alexander M. Rush

    cs.CLcs.LGcs.NEarXiv:1606.02960v22016
  52. Mathematical Foundations for a Compositional Distributional Model of Meaning

    Bob Coecke, Mehrnoosh Sadrzadeh, Stephen Clark

    cs.CLcs.LOmath.CTarXiv:1003.4394v12010
  53. Zephyr: Direct Distillation of LM Alignment

    Lewis Tunstall, Edward Beeching, Nathan Lambert +11

    cs.LGcs.CLarXiv:2310.16944v12023
  54. Movement Pruning: Adaptive Sparsity by Fine-Tuning

    Victor Sanh, Thomas Wolf, Alexander M. Rush

    cs.CLcs.LGarXiv:2005.07683v22020
  55. Linear models and linear mixed effects models in R with linguistic applications

    Bodo Winter

    cs.CLarXiv:1308.5499v12013
  56. Newsroom: A Dataset of 1.3 Million Summaries with Diverse Extractive Strategies

    Max Grusky, Mor Naaman, Yoav Artzi

    cs.CLarXiv:1804.11283v22018
  57. SeMoCo: A Semantic-First Motion Codec for Motion Language Modeling

    Tianlv Huang, Hetian Guo, Ziyi Cai +6

    cs.CVcs.CLcs.GRarXiv:2608.24334v12026
  58. Measuring Digital Labour Market Transitions with a Digital Semantic Score: An AI-Based Methodology Applied to the Dutch Labour Market

    Sadegh Shahmohammadi, Xavier Pinho, Mairi Bowdler +2

    cs.CLarXiv:2608.24222v12026
  59. Improving Multimodal Fusion with Hierarchical Mutual Information Maximization for Multimodal Sentiment Analysis

    Wei Han, Hui Chen, Soujanya Poria

    cs.CLcs.AIarXiv:2109.00412v22021
  60. GPT detectors are biased against non-native English writers

    Weixin Liang, Mert Yuksekgonul, Yining Mao +2

    cs.CLcs.AIcs.HCarXiv:2304.02819v32023