Computation and Language
Papers filed under cs.CL on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
8,581 to 8,640 of 11,245
Mending the Holes: Mitigating Reward Hacking in Reinforcement Learning for Multilingual Translation
Yifeng Liu, Siqi Ouyang, Yatish Hosmane Revanasiddappa +1
cs.CLarXiv:2603.13045v12026VoXtream2: Full-stream TTS with dynamic speaking rate control
Nikita Torgashov, Gustav Eje Henter, Gabriel Skantze
eess.AScs.CLcs.HCarXiv:2603.13518v12026ECG-Reasoning-Benchmark: A Benchmark for Evaluating Clinical Reasoning Capabilities in ECG Interpretation
Jungwoo Oh, Hyunseung Chung, Junhee Lee +6
cs.LGcs.AIcs.CLarXiv:2603.14326v12026PARSA-Bench: A Comprehensive Persian Audio-Language Model Benchmark
Mohammad Javad Ranjbar Kalahroodi, Mohammad Amini, Parmis Bathayan +2
cs.CLcs.SDarXiv:2603.14456v12026Towards Scalable Multi-domain Conversational Agents: The Schema-Guided Dialogue Dataset
Abhinav Rastogi, Xiaoxue Zang, Srinivas Sunkara +2
cs.CLarXiv:1909.05855v22019DataComp: In search of the next generation of multimodal datasets
Samir Yitzhak Gadre, Gabriel Ilharco, Alex Fang +31
cs.CVcs.CLcs.LGarXiv:2304.14108v52023A Multi-World Approach to Question Answering about Real-World Scenes based on Uncertain Input
Mateusz Malinowski, Mario Fritz
cs.AIcs.CLcs.CVarXiv:1410.0210v42014Learning to Ask: Neural Question Generation for Reading Comprehension
Xinya Du, Junru Shao, Claire Cardie
cs.CLcs.AIarXiv:1705.00106v12017Loc3R-VLM: Language-based Localization and 3D Reasoning with Vision-Language Models
Kevin Qu, Haozhe Qi, Mihai Dusmanu +3
cs.CVcs.AIcs.CLarXiv:2603.18002v12026Socratic Models: Composing Zero-Shot Multimodal Reasoning with Language
Andy Zeng, Maria Attarian, Brian Ichter +10
cs.CVcs.AIcs.CLarXiv:2204.00598v22022ByT5: Towards a token-free future with pre-trained byte-to-byte models
Linting Xue, Aditya Barua, Noah Constant +5
cs.CLarXiv:2105.13626v32021Efficient Document Parsing via Parallel Token Prediction
Lei Li, Ze Zhao, Meng Li +6
cs.CLcs.CVarXiv:2603.15206v12026A Comprehensive Analysis of Arabic Natural Language Processing Research: Trends, Topic Evolution, and Research Gaps -- A Bibliometric and Topic-Based Study
Mullosharaf K. Arabov
cs.CLarXiv:2608.23421v22026FormuEvo: LLM-Guided Evolution for Discovering Solver-Efficient Mixed-Integer Programming Formulations
Haofeng Yuan, Jianing Peng, Jieyi Bi +3
cs.CLcs.NEarXiv:2608.23353v12026What's the Catch? Evaluating Temporal Consistency in Vision-Language Models
Marek Hradil, Danae Sánchez Villegas
cs.CLcs.AIcs.CVarXiv:2608.23474v22026Multi-User Large Language Model Agents
Shu Yang, Shenzhe Zhu, Hao Zhu +5
cs.CLcs.MAarXiv:2604.08567v22026DIAG: Diagnostic Iterative Alignment and Generation for Data-Efficient Mathematical Preference Distillation
Guhan Chen, Songtao Tian, Bohan Li +3
cs.CLarXiv:2608.22806v12026SPOC-SQL: Stage-wise Preference Optimization for Controllable Text-to-SQL
Yingnan Chen, Chun Ding, Tianshi Xu +2
cs.CLarXiv:2608.22772v12026Retentive Network: A Successor to Transformer for Large Language Models
Yutao Sun, Li Dong, Shaohan Huang +5
cs.CLcs.LGarXiv:2307.08621v42023Tree of Attacks: Jailbreaking Black-Box LLMs Automatically
Anay Mehrotra, Manolis Zampetakis, Paul Kassianik +4
cs.LGcs.AIcs.CLarXiv:2312.02119v32023The Emergence of Relevance Through Axiomatic Attention Patterns During LoRA Fine-Tuning
Matthew Perlman, Atharva Nijasure, James Allan
cs.CLcs.AIcs.IRarXiv:2608.23338v12026WorldCache: Content-Aware Caching for Accelerated Video World Models
Umair Nawaz, Ahmed Heakl, Ufaq Khan +3
cs.CVcs.AIcs.CLarXiv:2603.22286v12026SpecEyes: Accelerating Agentic Multimodal LLMs via Speculative Perception and Planning
Haoyu Huang, Jinfa Huang, Zhongwei Wan +3
cs.CVcs.CLarXiv:2603.23483v22026Robust Reasoning Benchmark
Pavel Golikov, Evgenii Opryshko, Gennady Pekhimenko +1
cs.LGcs.AIcs.CLarXiv:2604.08571v32026Xpertbench: Expert Level Tasks with Rubrics-Based Evaluation
Xue Liu, Xin Ma, Yuxin Ma +36
cs.AIcs.CLarXiv:2604.02368v42026VoxPopuli: A Large-Scale Multilingual Speech Corpus for Representation Learning, Semi-Supervised Learning and Interpretation
Changhan Wang, Morgane Rivière, Ann Lee +6
cs.CLeess.ASarXiv:2101.00390v22021Neural Approaches to Conversational AI
Jianfeng Gao, Michel Galley, Lihong Li
cs.CLarXiv:1809.08267v32018ExecRubrics: Executable Tool-Augmented Rubrics for Verifiable and Efficient Long-Form Evaluation
Kaustubh D. Dhole, Charles L. A. Clarke, Eugene Y. Agichtein
cs.AIcs.CLcs.IRarXiv:2608.22559v12026Fine-Tuning Pretrained Language Models: Weight Initializations, Data Orders, and Early Stopping
Jesse Dodge, Gabriel Ilharco, Roy Schwartz +3
cs.CLcs.LGarXiv:2002.06305v12020The Design and Implementation of XiaoIce, an Empathetic Social Chatbot
Li Zhou, Jianfeng Gao, Di Li +1
cs.HCcs.AIcs.CLarXiv:1812.08989v22018PerceptionComp: A Video Benchmark for Complex Perception-Centric Reasoning
Shaoxuan Li, Zhixuan Zhao, Hanze Deng +9
cs.CVcs.AIcs.CLarXiv:2603.26653v12026Emergent Social Intelligence Risks in Generative Multi-Agent Systems
Yue Huang, Yu Jiang, Wenjie Wang +12
cs.MAcs.CLcs.CYarXiv:2603.27771v22026WeMM-Embedding: WeChat Multi-Modal Embedding Technical Report
Junjie Zhou, Ke Mei, Lei Li +3
cs.CVcs.CLcs.IRarXiv:2608.24053v12026Summaries:한국어Friends and Grandmothers in Silico: Localizing Entity Cells in Language Models
Itay Yona, Dan Barzilay, Michael Karasik +1
cs.CLcs.AIarXiv:2604.01404v22026MemRerank: Preference Memory for Personalized Product Reranking
Zhiyuan Peng, Xuyang Wu, Huaixiao Tou +2
cs.CLcs.AIcs.LGarXiv:2603.29247v32026Ebisu: Benchmarking Large Language Models in Japanese Finance
Xueqing Peng, Ruoyu Xiang, Fan Zhang +9
cs.CLarXiv:2602.01479v12026WizardMath: Empowering Mathematical Reasoning for Large Language Models via Reinforced Evol-Instruct
Haipeng Luo, Qingfeng Sun, Can Xu +8
cs.CLcs.AIcs.LGarXiv:2308.09583v32023Hybrid Panels: Toward Human-AI Collaboration in Survey Research
Julia Romberg, Tobias Gummer, Gabriella Lapesa +2
cs.CLcs.AIcs.CYarXiv:2608.22582v12026ProBel: Propaganda Detection with Techniques, Spans, and Explanations
Mohamed Bayan Kmainasi, Ali Ezzat Shahroor, Elisa Sartori +2
cs.CLcs.AIcs.LGarXiv:2608.22388v12026Evaluating Very Long-Term Conversational Memory of LLM Agents
Adyasha Maharana, Dong-Ho Lee, Sergey Tulyakov +3
cs.CLcs.AIcs.LGarXiv:2402.17753v12024Beto, Bentz, Becas: The Surprising Cross-Lingual Effectiveness of BERT
Shijie Wu, Mark Dredze
cs.CLarXiv:1904.09077v22019Gender Bias in Coreference Resolution
Rachel Rudinger, Jason Naradowsky, Brian Leonard +1
cs.CLarXiv:1804.09301v12018ROCKET: Rapid Optimization via Calibration-guided Knapsack Enhanced Truncation for Efficient Model Compression
Ammar Ali, Baher Mohammad, Denis Makhov +3
cs.LGcs.AIcs.CLarXiv:2602.11008v12026Rubrics as an Attack Surface: Stealthy Preference Drift in LLM Judges
Ruomeng Ding, Yifei Pang, He Sun +3
cs.CRcs.AIcs.CLarXiv:2602.13576v12026AdapterHub: A Framework for Adapting Transformers
Jonas Pfeiffer, Andreas Rücklé, Clifton Poth +5
cs.CLarXiv:2007.07779v32020The Internal State of an LLM Knows When It's Lying
Amos Azaria, Tom Mitchell
cs.CLcs.AIcs.LGarXiv:2304.13734v22023Phi-4 Technical Report
Marah Abdin, Jyoti Aneja, Harkirat Behl +24
cs.CLcs.AIarXiv:2412.08905v12024Summary of ChatGPT-Related Research and Perspective Towards the Future of Large Language Models
Yiheng Liu, Tianle Han, Siyuan Ma +15
cs.CLarXiv:2304.01852v42023Structured Episodic Event Memory
Zhengxuan Lu, Dongfang Li, Yukun Shi +3
cs.CLarXiv:2601.06411v22026MiniCPM: Unveiling the Potential of Small Language Models with Scalable Training Strategies
Shengding Hu, Yuge Tu, Xu Han +22
cs.CLcs.LGarXiv:2404.06395v32024From Diagnosis to Redesign: Using Quantitative Ethnography to Improve Multi-Agent LLM Reasoning
Vedant Khatri, Anthony Cusimano, Zachari Swiecki +3
cs.CLarXiv:2608.22566v12026When Personalization Misleads: Understanding and Mitigating Hallucinations in Personalized LLMs
Zhongxiang Sun, Yi Zhan, Chenglei Shen +4
cs.CLcs.AIarXiv:2601.11000v12026Solar Open Technical Report
Sungrae Park, Sanghoon Kim, Jungho Cho +34
cs.CLarXiv:2601.07022v12026MAD-X: An Adapter-Based Framework for Multi-Task Cross-Lingual Transfer
Jonas Pfeiffer, Ivan Vulić, Iryna Gurevych +1
cs.CLarXiv:2005.00052v32020AudioNoisePrints: Model-free audio watermarking using spatial correlation in flow matching TTS
Timothy Tin-Long, Jian Zhu, Aidan Pine +1
cs.SDcs.CLarXiv:2608.22186v12026GutenOCR: A Grounded Vision-Language Front-End for Documents
Hunter Heidenreich, Ben Elliott, Olivia Dinica +1
cs.CVcs.AIcs.CLarXiv:2601.14490v22026Sparking Scientific Creativity via LLM-Driven Interdisciplinary Inspiration
Priyanka Kargupta, Shuhaib Mehri, Dilek Hakkani-Tur +1
cs.CLcs.AIarXiv:2603.12226v12026DARC: Decoupled Asymmetric Reasoning Curriculum for LLM Evolution
Shengda Fan, Xuyan Ye, Yankai Lin
cs.AIcs.CLarXiv:2601.13761v22026CodeT5+: Open Code Large Language Models for Code Understanding and Generation
Yue Wang, Hung Le, Akhilesh Deepak Gotmare +3
cs.CLcs.LGcs.PLarXiv:2305.07922v22023An Empirical Study of Catastrophic Forgetting in Large Language Models During Continual Fine-tuning
Yun Luo, Zhen Yang, Fandong Meng +3
cs.CLarXiv:2308.08747v52023