Computation and Language

Papers filed under cs.CL on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

8,581 to 8,640 of 11,245

  1. Mending the Holes: Mitigating Reward Hacking in Reinforcement Learning for Multilingual Translation

    Yifeng Liu, Siqi Ouyang, Yatish Hosmane Revanasiddappa +1

    cs.CLarXiv:2603.13045v12026
  2. VoXtream2: Full-stream TTS with dynamic speaking rate control

    Nikita Torgashov, Gustav Eje Henter, Gabriel Skantze

    eess.AScs.CLcs.HCarXiv:2603.13518v12026
  3. ECG-Reasoning-Benchmark: A Benchmark for Evaluating Clinical Reasoning Capabilities in ECG Interpretation

    Jungwoo Oh, Hyunseung Chung, Junhee Lee +6

    cs.LGcs.AIcs.CLarXiv:2603.14326v12026
  4. PARSA-Bench: A Comprehensive Persian Audio-Language Model Benchmark

    Mohammad Javad Ranjbar Kalahroodi, Mohammad Amini, Parmis Bathayan +2

    cs.CLcs.SDarXiv:2603.14456v12026
  5. Towards Scalable Multi-domain Conversational Agents: The Schema-Guided Dialogue Dataset

    Abhinav Rastogi, Xiaoxue Zang, Srinivas Sunkara +2

    cs.CLarXiv:1909.05855v22019
  6. DataComp: In search of the next generation of multimodal datasets

    Samir Yitzhak Gadre, Gabriel Ilharco, Alex Fang +31

    cs.CVcs.CLcs.LGarXiv:2304.14108v52023
  7. A Multi-World Approach to Question Answering about Real-World Scenes based on Uncertain Input

    Mateusz Malinowski, Mario Fritz

    cs.AIcs.CLcs.CVarXiv:1410.0210v42014
  8. Learning to Ask: Neural Question Generation for Reading Comprehension

    Xinya Du, Junru Shao, Claire Cardie

    cs.CLcs.AIarXiv:1705.00106v12017
  9. Loc3R-VLM: Language-based Localization and 3D Reasoning with Vision-Language Models

    Kevin Qu, Haozhe Qi, Mihai Dusmanu +3

    cs.CVcs.AIcs.CLarXiv:2603.18002v12026
  10. Socratic Models: Composing Zero-Shot Multimodal Reasoning with Language

    Andy Zeng, Maria Attarian, Brian Ichter +10

    cs.CVcs.AIcs.CLarXiv:2204.00598v22022
  11. ByT5: Towards a token-free future with pre-trained byte-to-byte models

    Linting Xue, Aditya Barua, Noah Constant +5

    cs.CLarXiv:2105.13626v32021
  12. Efficient Document Parsing via Parallel Token Prediction

    Lei Li, Ze Zhao, Meng Li +6

    cs.CLcs.CVarXiv:2603.15206v12026
  13. A Comprehensive Analysis of Arabic Natural Language Processing Research: Trends, Topic Evolution, and Research Gaps -- A Bibliometric and Topic-Based Study

    Mullosharaf K. Arabov

    cs.CLarXiv:2608.23421v22026
  14. FormuEvo: LLM-Guided Evolution for Discovering Solver-Efficient Mixed-Integer Programming Formulations

    Haofeng Yuan, Jianing Peng, Jieyi Bi +3

    cs.CLcs.NEarXiv:2608.23353v12026
  15. What's the Catch? Evaluating Temporal Consistency in Vision-Language Models

    Marek Hradil, Danae Sánchez Villegas

    cs.CLcs.AIcs.CVarXiv:2608.23474v22026
  16. Multi-User Large Language Model Agents

    Shu Yang, Shenzhe Zhu, Hao Zhu +5

    cs.CLcs.MAarXiv:2604.08567v22026
  17. DIAG: Diagnostic Iterative Alignment and Generation for Data-Efficient Mathematical Preference Distillation

    Guhan Chen, Songtao Tian, Bohan Li +3

    cs.CLarXiv:2608.22806v12026
  18. SPOC-SQL: Stage-wise Preference Optimization for Controllable Text-to-SQL

    Yingnan Chen, Chun Ding, Tianshi Xu +2

    cs.CLarXiv:2608.22772v12026
  19. Retentive Network: A Successor to Transformer for Large Language Models

    Yutao Sun, Li Dong, Shaohan Huang +5

    cs.CLcs.LGarXiv:2307.08621v42023
  20. Tree of Attacks: Jailbreaking Black-Box LLMs Automatically

    Anay Mehrotra, Manolis Zampetakis, Paul Kassianik +4

    cs.LGcs.AIcs.CLarXiv:2312.02119v32023
  21. The Emergence of Relevance Through Axiomatic Attention Patterns During LoRA Fine-Tuning

    Matthew Perlman, Atharva Nijasure, James Allan

    cs.CLcs.AIcs.IRarXiv:2608.23338v12026
  22. WorldCache: Content-Aware Caching for Accelerated Video World Models

    Umair Nawaz, Ahmed Heakl, Ufaq Khan +3

    cs.CVcs.AIcs.CLarXiv:2603.22286v12026
  23. SpecEyes: Accelerating Agentic Multimodal LLMs via Speculative Perception and Planning

    Haoyu Huang, Jinfa Huang, Zhongwei Wan +3

    cs.CVcs.CLarXiv:2603.23483v22026
  24. Robust Reasoning Benchmark

    Pavel Golikov, Evgenii Opryshko, Gennady Pekhimenko +1

    cs.LGcs.AIcs.CLarXiv:2604.08571v32026
  25. Xpertbench: Expert Level Tasks with Rubrics-Based Evaluation

    Xue Liu, Xin Ma, Yuxin Ma +36

    cs.AIcs.CLarXiv:2604.02368v42026
  26. VoxPopuli: A Large-Scale Multilingual Speech Corpus for Representation Learning, Semi-Supervised Learning and Interpretation

    Changhan Wang, Morgane Rivière, Ann Lee +6

    cs.CLeess.ASarXiv:2101.00390v22021
  27. Neural Approaches to Conversational AI

    Jianfeng Gao, Michel Galley, Lihong Li

    cs.CLarXiv:1809.08267v32018
  28. ExecRubrics: Executable Tool-Augmented Rubrics for Verifiable and Efficient Long-Form Evaluation

    Kaustubh D. Dhole, Charles L. A. Clarke, Eugene Y. Agichtein

    cs.AIcs.CLcs.IRarXiv:2608.22559v12026
  29. Fine-Tuning Pretrained Language Models: Weight Initializations, Data Orders, and Early Stopping

    Jesse Dodge, Gabriel Ilharco, Roy Schwartz +3

    cs.CLcs.LGarXiv:2002.06305v12020
  30. The Design and Implementation of XiaoIce, an Empathetic Social Chatbot

    Li Zhou, Jianfeng Gao, Di Li +1

    cs.HCcs.AIcs.CLarXiv:1812.08989v22018
  31. PerceptionComp: A Video Benchmark for Complex Perception-Centric Reasoning

    Shaoxuan Li, Zhixuan Zhao, Hanze Deng +9

    cs.CVcs.AIcs.CLarXiv:2603.26653v12026
  32. Emergent Social Intelligence Risks in Generative Multi-Agent Systems

    Yue Huang, Yu Jiang, Wenjie Wang +12

    cs.MAcs.CLcs.CYarXiv:2603.27771v22026
  33. WeMM-Embedding: WeChat Multi-Modal Embedding Technical Report

    Junjie Zhou, Ke Mei, Lei Li +3

    cs.CVcs.CLcs.IRarXiv:2608.24053v12026
    Summaries:한국어
  34. Friends and Grandmothers in Silico: Localizing Entity Cells in Language Models

    Itay Yona, Dan Barzilay, Michael Karasik +1

    cs.CLcs.AIarXiv:2604.01404v22026
  35. MemRerank: Preference Memory for Personalized Product Reranking

    Zhiyuan Peng, Xuyang Wu, Huaixiao Tou +2

    cs.CLcs.AIcs.LGarXiv:2603.29247v32026
  36. Ebisu: Benchmarking Large Language Models in Japanese Finance

    Xueqing Peng, Ruoyu Xiang, Fan Zhang +9

    cs.CLarXiv:2602.01479v12026
  37. WizardMath: Empowering Mathematical Reasoning for Large Language Models via Reinforced Evol-Instruct

    Haipeng Luo, Qingfeng Sun, Can Xu +8

    cs.CLcs.AIcs.LGarXiv:2308.09583v32023
  38. Hybrid Panels: Toward Human-AI Collaboration in Survey Research

    Julia Romberg, Tobias Gummer, Gabriella Lapesa +2

    cs.CLcs.AIcs.CYarXiv:2608.22582v12026
  39. ProBel: Propaganda Detection with Techniques, Spans, and Explanations

    Mohamed Bayan Kmainasi, Ali Ezzat Shahroor, Elisa Sartori +2

    cs.CLcs.AIcs.LGarXiv:2608.22388v12026
  40. Evaluating Very Long-Term Conversational Memory of LLM Agents

    Adyasha Maharana, Dong-Ho Lee, Sergey Tulyakov +3

    cs.CLcs.AIcs.LGarXiv:2402.17753v12024
  41. Beto, Bentz, Becas: The Surprising Cross-Lingual Effectiveness of BERT

    Shijie Wu, Mark Dredze

    cs.CLarXiv:1904.09077v22019
  42. Gender Bias in Coreference Resolution

    Rachel Rudinger, Jason Naradowsky, Brian Leonard +1

    cs.CLarXiv:1804.09301v12018
  43. ROCKET: Rapid Optimization via Calibration-guided Knapsack Enhanced Truncation for Efficient Model Compression

    Ammar Ali, Baher Mohammad, Denis Makhov +3

    cs.LGcs.AIcs.CLarXiv:2602.11008v12026
  44. Rubrics as an Attack Surface: Stealthy Preference Drift in LLM Judges

    Ruomeng Ding, Yifei Pang, He Sun +3

    cs.CRcs.AIcs.CLarXiv:2602.13576v12026
  45. AdapterHub: A Framework for Adapting Transformers

    Jonas Pfeiffer, Andreas Rücklé, Clifton Poth +5

    cs.CLarXiv:2007.07779v32020
  46. The Internal State of an LLM Knows When It's Lying

    Amos Azaria, Tom Mitchell

    cs.CLcs.AIcs.LGarXiv:2304.13734v22023
  47. Phi-4 Technical Report

    Marah Abdin, Jyoti Aneja, Harkirat Behl +24

    cs.CLcs.AIarXiv:2412.08905v12024
  48. Summary of ChatGPT-Related Research and Perspective Towards the Future of Large Language Models

    Yiheng Liu, Tianle Han, Siyuan Ma +15

    cs.CLarXiv:2304.01852v42023
  49. Structured Episodic Event Memory

    Zhengxuan Lu, Dongfang Li, Yukun Shi +3

    cs.CLarXiv:2601.06411v22026
  50. MiniCPM: Unveiling the Potential of Small Language Models with Scalable Training Strategies

    Shengding Hu, Yuge Tu, Xu Han +22

    cs.CLcs.LGarXiv:2404.06395v32024
  51. From Diagnosis to Redesign: Using Quantitative Ethnography to Improve Multi-Agent LLM Reasoning

    Vedant Khatri, Anthony Cusimano, Zachari Swiecki +3

    cs.CLarXiv:2608.22566v12026
  52. When Personalization Misleads: Understanding and Mitigating Hallucinations in Personalized LLMs

    Zhongxiang Sun, Yi Zhan, Chenglei Shen +4

    cs.CLcs.AIarXiv:2601.11000v12026
  53. Solar Open Technical Report

    Sungrae Park, Sanghoon Kim, Jungho Cho +34

    cs.CLarXiv:2601.07022v12026
  54. MAD-X: An Adapter-Based Framework for Multi-Task Cross-Lingual Transfer

    Jonas Pfeiffer, Ivan Vulić, Iryna Gurevych +1

    cs.CLarXiv:2005.00052v32020
  55. AudioNoisePrints: Model-free audio watermarking using spatial correlation in flow matching TTS

    Timothy Tin-Long, Jian Zhu, Aidan Pine +1

    cs.SDcs.CLarXiv:2608.22186v12026
  56. GutenOCR: A Grounded Vision-Language Front-End for Documents

    Hunter Heidenreich, Ben Elliott, Olivia Dinica +1

    cs.CVcs.AIcs.CLarXiv:2601.14490v22026
  57. Sparking Scientific Creativity via LLM-Driven Interdisciplinary Inspiration

    Priyanka Kargupta, Shuhaib Mehri, Dilek Hakkani-Tur +1

    cs.CLcs.AIarXiv:2603.12226v12026
  58. DARC: Decoupled Asymmetric Reasoning Curriculum for LLM Evolution

    Shengda Fan, Xuyan Ye, Yankai Lin

    cs.AIcs.CLarXiv:2601.13761v22026
  59. CodeT5+: Open Code Large Language Models for Code Understanding and Generation

    Yue Wang, Hung Le, Akhilesh Deepak Gotmare +3

    cs.CLcs.LGcs.PLarXiv:2305.07922v22023
  60. An Empirical Study of Catastrophic Forgetting in Large Language Models During Continual Fine-tuning

    Yun Luo, Zhen Yang, Fandong Meng +3

    cs.CLarXiv:2308.08747v52023