Computation and Language

Papers filed under cs.CL on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

8,101 to 8,160 of 11,259

  1. SteerCheck: Attribution Specificity and Alignment Leakage in Activation-Steering Audits

    Daming Luo, Christy Liang, Junyu Xuan

    cs.CLarXiv:2608.24335v12026
  2. WinCLIP: Zero-/Few-Shot Anomaly Classification and Segmentation

    Jongheon Jeong, Yang Zou, Taewan Kim +3

    cs.CVcs.AIcs.CLarXiv:2303.14814v12023
  3. Curved Inference II: Sleeper Agent Geometry - Extending Interpretability Beyond Probes

    Rob Manson

    cs.CLarXiv:2608.24037v12026
  4. DiffuSeq: Sequence to Sequence Text Generation with Diffusion Models

    Shansan Gong, Mukai Li, Jiangtao Feng +2

    cs.CLcs.LGarXiv:2210.08933v32022
  5. WiC: the Word-in-Context Dataset for Evaluating Context-Sensitive Meaning Representations

    Mohammad Taher Pilehvar, Jose Camacho-Collados

    cs.CLarXiv:1808.09121v32018
  6. MineDojo: Building Open-Ended Embodied Agents with Internet-Scale Knowledge

    Linxi Fan, Guanzhi Wang, Yunfan Jiang +7

    cs.LGcs.AIcs.CLarXiv:2206.08853v22022
  7. A Survey Of Cross-lingual Word Embedding Models

    Sebastian Ruder, Ivan Vulić, Anders Søgaard

    cs.CLcs.LGarXiv:1706.04902v42017
  8. Language Agent Tree Search Unifies Reasoning Acting and Planning in Language Models

    Andy Zhou, Kai Yan, Michal Shlapentokh-Rothman +2

    cs.AIcs.CLcs.CVarXiv:2310.04406v32023
  9. BOLD: Dataset and Metrics for Measuring Biases in Open-Ended Language Generation

    Jwala Dhamala, Tony Sun, Varun Kumar +4

    cs.CLcs.AIcs.LGarXiv:2101.11718v12021
  10. Mechanistic Circuit Identification for Controllable Data Generation

    Nakyung Lee, Sangwoo Hong, Jungwoo Lee

    cs.LGcs.AIcs.CLarXiv:2608.24065v12026
  11. Large Language Models are Zero-Shot Rankers for Recommender Systems

    Yupeng Hou, Junjie Zhang, Zihan Lin +4

    cs.IRcs.CLarXiv:2305.08845v22023
  12. Aspect Based Sentiment Analysis with Gated Convolutional Networks

    Wei Xue, Tao Li

    cs.CLarXiv:1805.07043v12018
  13. GPT-4V(ision) is a Generalist Web Agent, if Grounded

    Boyuan Zheng, Boyu Gou, Jihyung Kil +2

    cs.IRcs.AIcs.CLarXiv:2401.01614v22024
  14. When Youth Enter The Chat: An Epistemic Shift in the Validation of LLM-Based Measures of Student Talk

    Liliana Santos-Deonizio, James Malamut, Ramón Martínez +1

    cs.CLcs.AIcs.HCarXiv:2608.23780v12026
  15. Calibration-Preserving Pruning: Compression as a Reliability Contract

    Ibne Farabi Shihab, Adria Binte Habib, Anuj Sharma

    cs.LGcs.CLarXiv:2608.23744v12026
  16. Q-Align: Teaching LMMs for Visual Scoring via Discrete Text-Defined Levels

    Haoning Wu, Zicheng Zhang, Weixia Zhang +11

    cs.CVcs.CLcs.LGarXiv:2312.17090v12023
  17. A Hierarchical Neural Autoencoder for Paragraphs and Documents

    Jiwei Li, Minh-Thang Luong, Dan Jurafsky

    cs.CLarXiv:1506.01057v22015
  18. Mitigating Exploration Bias in RL for Multi-Instruction Following

    Mian Zhang, Yueqin Yin, Kaiyu He +4

    cs.CLcs.LGarXiv:2608.23830v12026
  19. RouteLLM: Learning to Route LLMs with Preference Data

    Isaac Ong, Amjad Almahairi, Vincent Wu +5

    cs.LGcs.AIcs.CLarXiv:2406.18665v42024
  20. PatchWrite: One Line, Not One Section -- Compile-Gated, Validity-Preserving Editing for AI-Drafted Manuscripts

    Weiwei Yang

    cs.AIcs.CLcs.SEarXiv:2608.23001v12026
  21. FedKD: Communication Efficient Federated Learning via Knowledge Distillation

    Chuhan Wu, Fangzhao Wu, Lingjuan Lyu +2

    cs.LGcs.CLarXiv:2108.13323v22021
  22. From Triage to Discharge: A Survey of NLP Tasks, Methods, and Open Challenges in the Emergency Department

    Dipankar Srirag, Aditya Joshi, Salil Kanhere +1

    cs.CLarXiv:2608.23627v12026
  23. Fast Abstractive Summarization with Reinforce-Selected Sentence Rewriting

    Yen-Chun Chen, Mohit Bansal

    cs.CLcs.AIcs.LGarXiv:1805.11080v12018
  24. Inter-dimension Dependence for Multi-Dimensional Evaluation of Open-Ended Text

    Haoyuan Li, Snigdha Chaturvedi

    cs.CLarXiv:2608.23783v12026
  25. YaRN: Efficient Context Window Extension of Large Language Models

    Bowen Peng, Jeffrey Quesnelle, Honglu Fan +1

    cs.CLcs.AIcs.LGarXiv:2309.00071v32023
  26. Towards End-to-End Prosody Transfer for Expressive Speech Synthesis with Tacotron

    RJ Skerry-Ryan, Eric Battenberg, Ying Xiao +6

    cs.CLcs.LGcs.SDarXiv:1803.09047v12018
  27. PAWS: Paraphrase Adversaries from Word Scrambling

    Yuan Zhang, Jason Baldridge, Luheng He

    cs.CLarXiv:1904.01130v12019
  28. MirrorGAN: Learning Text-to-image Generation by Redescription

    Tingting Qiao, Jing Zhang, Duanqing Xu +1

    cs.CLcs.CVcs.LGarXiv:1903.05854v12019
  29. Markets, Not Planners: Decentralized Orchestration of LLM Agents with Private Information

    Xiao Liu, Haoyang Li, Songwei Li +4

    cs.MAcs.CLarXiv:2608.23867v12026
  30. Asking and Answering Questions to Evaluate the Factual Consistency of Summaries

    Alex Wang, Kyunghyun Cho, Mike Lewis

    cs.CLarXiv:2004.04228v12020
  31. SMART: Robust and Efficient Fine-Tuning for Pre-trained Natural Language Models through Principled Regularized Optimization

    Haoming Jiang, Pengcheng He, Weizhu Chen +3

    cs.CLcs.LGmath.OCarXiv:1911.03437v52019
  32. ERNIE 3.0: Large-scale Knowledge Enhanced Pre-training for Language Understanding and Generation

    Yu Sun, Shuohuan Wang, Shikun Feng +19

    cs.CLarXiv:2107.02137v12021
  33. A Survey of Hallucination in Large Foundation Models

    Vipula Rawte, Amit Sheth, Amitava Das

    cs.AIcs.CLcs.IRarXiv:2309.05922v12023
  34. Dataset Cartography: Mapping and Diagnosing Datasets with Training Dynamics

    Swabha Swayamdipta, Roy Schwartz, Nicholas Lourie +4

    cs.CLarXiv:2009.10795v22020
  35. Clotho: An Audio Captioning Dataset

    Konstantinos Drossos, Samuel Lipping, Tuomas Virtanen

    cs.SDcs.CLcs.LGarXiv:1910.09387v12019
  36. Prometheus: Inducing Fine-grained Evaluation Capability in Language Models

    Seungone Kim, Jamin Shin, Yejin Cho +8

    cs.CLcs.LGarXiv:2310.08491v22023
  37. Using Trusted Data to Train Deep Networks on Labels Corrupted by Severe Noise

    Dan Hendrycks, Mantas Mazeika, Duncan Wilson +1

    cs.LGcs.CLcs.CVarXiv:1802.05300v42018
  38. Beyond the Nav-Graph: Vision-and-Language Navigation in Continuous Environments

    Jacob Krantz, Erik Wijmans, Arjun Majumdar +2

    cs.CVcs.CLcs.ROarXiv:2004.02857v22020
  39. Adaptive-RAG: Learning to Adapt Retrieval-Augmented Large Language Models through Question Complexity

    Soyeong Jeong, Jinheon Baek, Sukmin Cho +2

    cs.CLcs.AIarXiv:2403.14403v22024
  40. Weight Poisoning Attacks on Pre-trained Models

    Keita Kurita, Paul Michel, Graham Neubig

    cs.LGcs.CLcs.CRarXiv:2004.06660v12020
  41. Supervised Contrastive Learning for Pre-trained Language Model Fine-tuning

    Beliz Gunel, Jingfei Du, Alexis Conneau +1

    cs.CLcs.LGarXiv:2011.01403v32020
  42. How Can We Know When Language Models Know? On the Calibration of Language Models for Question Answering

    Zhengbao Jiang, Jun Araki, Haibo Ding +1

    cs.CLarXiv:2012.00955v22020
  43. Text and Code Embeddings by Contrastive Pre-Training

    Arvind Neelakantan, Tao Xu, Raul Puri +22

    cs.CLcs.LGarXiv:2201.10005v12022
  44. MemoryBank: Enhancing Large Language Models with Long-Term Memory

    Wanjun Zhong, Lianghong Guo, Qiqi Gao +2

    cs.CLcs.AIarXiv:2305.10250v32023
  45. MLQA: Evaluating Cross-lingual Extractive Question Answering

    Patrick Lewis, Barlas Oğuz, Ruty Rinott +2

    cs.CLcs.AIcs.LGarXiv:1910.07475v32019
  46. SkillForge: Evolving Verifiable Skills for Reinforcement Learning Agents

    Shidong Yang, Ziyu Ma, Tongwen Huang +5

    cs.CLarXiv:2608.24747v12026
  47. Arbitrary Polygon Oscillator: Generalizing Polygonal Synthesis to Arbitrary Shapes, Morphing, and Three-Dimensional Polyhedra

    Antonio Argentieri, Francesco Scagliola

    cs.CLarXiv:2608.24726v12026
  48. One Timeline, Many Renderings: A Wolfram Language Paclet for heterogeneous musical output

    Francesco Vitucci, Michele Lorusso, Francesco Scagliola

    cs.CLarXiv:2608.24683v12026
  49. From local kernels to global form: modeling the emergence of musical content

    Francesco Vitucci, Michele Lorusso, Francesco Scagliola

    cs.CLarXiv:2608.24660v12026
  50. Paritok-4B: Intent-Conditioned Context Compression for Coding Agents

    Jiayu Shi, Luzhuo Chen

    cs.AIcs.CLcs.LGarXiv:2608.24188v12026
  51. Sentiment Analysis Based on Deep Learning: A Comparative Study

    Nhan Cach Dang, María N. Moreno-García, Fernando De la Prieta

    cs.CLcs.IRcs.LGarXiv:2006.03541v12020
  52. GigaSpeech: An Evolving, Multi-domain ASR Corpus with 10,000 Hours of Transcribed Audio

    Guoguo Chen, Shuzhou Chai, Guanbo Wang +18

    cs.SDcs.CLeess.ASarXiv:2106.06909v12021
  53. HelaBERT: Enhancing Sinhala Language Understanding with Dual Pooling Classification Head

    Thisen Ekanayake, Nisansa de Silva

    cs.CLarXiv:2608.22922v12026
  54. Nuanced Metrics for Measuring Unintended Bias with Real Data for Text Classification

    Daniel Borkan, Lucas Dixon, Jeffrey Sorensen +2

    cs.LGcs.CLstat.MLarXiv:1903.04561v22019
  55. MM-REACT: Prompting ChatGPT for Multimodal Reasoning and Action

    Zhengyuan Yang, Linjie Li, Jianfeng Wang +7

    cs.CVcs.CLcs.LGarXiv:2303.11381v12023
  56. The Limits of Automatic Evaluation of Creativity in Large Language Models

    Alessandro Tutone, Giorgio Franceschelli, Mirco Musolesi

    cs.CLcs.AIcs.CYarXiv:2608.23705v12026
  57. The AI Scientist: Towards Fully Automated Open-Ended Scientific Discovery

    Chris Lu, Cong Lu, Robert Tjarko Lange +3

    cs.AIcs.CLcs.LGarXiv:2408.06292v32024
    Summaries:한국어
  58. Can AI-Generated Text be Reliably Detected?

    Vinu Sankar Sadasivan, Aounon Kumar, Sriram Balasubramanian +2

    cs.CLcs.AIcs.LGarXiv:2303.11156v42023
  59. Beyond Static and Linear: What Attention Constraints Best Fit Human Reading Times?

    Lanni Bu, Xiulin Yang, Christian Clark +2

    cs.CLarXiv:2608.23818v12026
  60. A robust self-learning method for fully unsupervised cross-lingual mappings of word embeddings

    Mikel Artetxe, Gorka Labaka, Eneko Agirre

    cs.CLcs.AIcs.LGarXiv:1805.06297v22018