Computation and Language
Papers filed under cs.CL on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
5,281 to 5,340 of 11,248
XTREME-R: Towards More Challenging and Nuanced Multilingual Evaluation
Sebastian Ruder, Noah Constant, Jan Botha +8
cs.CLcs.AIarXiv:2104.07412v22021Understanding Knowledge Distillation in Non-autoregressive Machine Translation
Chunting Zhou, Graham Neubig, Jiatao Gu
cs.CLarXiv:1911.02727v32019Interpretable Rumor Detection in Microblogs by Attending to User Interactions
Ling Min Serena Khoo, Hai Leong Chieu, Zhong Qian +1
cs.CLcs.IRcs.SIarXiv:2001.10667v12020Fine-grained Sentiment Classification using BERT
Manish Munikar, Sushil Shakya, Aakash Shrestha
cs.CLcs.LGstat.MLarXiv:1910.03474v12019SwiftSage: A Generative Agent with Fast and Slow Thinking for Complex Interactive Tasks
Bill Yuchen Lin, Yicheng Fu, Karina Yang +6
cs.CLcs.AIcs.LGarXiv:2305.17390v22023Large Language Models are few(1)-shot Table Reasoners
Wenhu Chen
cs.CLarXiv:2210.06710v22022Data Augmentation Approaches in Natural Language Processing: A Survey
Bohan Li, Yutai Hou, Wanxiang Che
cs.CLcs.AIcs.LGarXiv:2110.01852v32021Bias and Unfairness in Information Retrieval Systems: New Challenges in the LLM Era
Sunhao Dai, Chen Xu, Shicheng Xu +3
cs.IRcs.AIcs.CLarXiv:2404.11457v22024End-to-End Speaker Diarization for an Unknown Number of Speakers with Encoder-Decoder Based Attractors
Shota Horiguchi, Yusuke Fujita, Shinji Watanabe +2
eess.AScs.CLcs.SDarXiv:2005.09921v32020An Analysis of Simple Data Augmentation for Named Entity Recognition
Xiang Dai, Heike Adel
cs.CLarXiv:2010.11683v12020Stay on the Path: Instruction Fidelity in Vision-and-Language Navigation
Vihan Jain, Gabriel Magalhaes, Alexander Ku +3
cs.AIcs.CLarXiv:1905.12255v32019Iterative Answer Prediction with Pointer-Augmented Multimodal Transformers for TextVQA
Ronghang Hu, Amanpreet Singh, Trevor Darrell +1
cs.CVcs.CLarXiv:1911.06258v32019The Cost of Training NLP Models: A Concise Overview
Or Sharir, Barak Peleg, Yoav Shoham
cs.CLcs.LGcs.NEarXiv:2004.08900v12020Token-Efficient Data Reasoning Agents via Adaptive Structuring of Unstructured Data
Milad Rezaei Hajidehi, Qitong Wang, Stratos Idreos
cs.AIcs.CLcs.DBarXiv:2608.31082v12026Benchmarking Foundation Models with Language-Model-as-an-Examiner
Yushi Bai, Jiahao Ying, Yixin Cao +10
cs.CLcs.LGarXiv:2306.04181v22023Hurdles to Progress in Long-form Question Answering
Kalpesh Krishna, Aurko Roy, Mohit Iyyer
cs.CLcs.LGarXiv:2103.06332v22021Controlling Refusal Behavior of LLMs via Stiefel-Constrained Rotation Steering
Kirill Bunin, Dmitry Bylinkin, Vladimir Aletov +3
cs.LGcs.CLarXiv:2608.30986v12026Arcee's MergeKit: A Toolkit for Merging Large Language Models
Charles Goddard, Shamane Siriwardhana, Malikeh Ehghaghi +5
cs.CLcs.AIcs.LGarXiv:2403.13257v32024The Hermon Moment: AI Self-Transcendence and Its Human Narration
Alexei Grinbaum
cs.CYcs.CLarXiv:2608.30971v12026In-Context Impersonation Reveals Large Language Models' Strengths and Biases
Leonard Salewski, Stephan Alaniz, Isabel Rio-Torto +2
cs.AIcs.CLcs.LGarXiv:2305.14930v22023Applying Large Language Models and Chain-of-Thought for Automatic Scoring
Gyeong-Geon Lee, Ehsan Latif, Xuansheng Wu +2
cs.CLcs.AIarXiv:2312.03748v22023BLOOM-WILT: Logit Tilting for Behaviour Elicitation in Automated LLM Auditing
Adrians Skapars, Edoardo Manino
cs.AIcs.CLarXiv:2608.31105v12026A Model with No Head and Many Thoughts
Nikita Koriagin, Yaroslav Aksenov, George Bredis +3
cs.LGcs.CLarXiv:2608.31069v12026When Can We Work in Embedding Space? What Text Embeddings Preserve
Simon Freyaldenhoven
econ.EMcs.CLstat.MLarXiv:2608.31059v12026Augmenting Interviewer Judgments of Patient Experience with Automatic Language Analysis
Aowen Shi, Michal Balazia, Danilo Postin +4
cs.HCcs.CLarXiv:2608.31007v12026E-Commerce Bench: Evaluating LLM Agents on Long-Horizon Autonomous Business Operation
Wei Fan, Xinjie Shen, Xudong Guo +8
cs.LGcs.CLarXiv:2608.30730v12026PLC-DPO: Posterior Label Correction in Noisy and Ambiguous Preference Optimization
Boryeong Cho, Sumyeong Ahn, Se-Young Yun
cs.LGcs.CLarXiv:2608.30597v12026ScienceArena: Benchmarking LLMs on Latest Scientific Olympiad Competitions
Guangxiang Zhao, Qilong Shi, Xusen Xiao +13
cs.AIcs.CLarXiv:2608.30517v12026One Policy Is Enough: Single-Agent Reinforcement Learning Outperforms Tree Search for Chemistry Tool Learning
Armin Dariani, Sifan Wu, Bang Liu +1
cs.LGcs.CLarXiv:2608.30952v12026The NetHack Learning Environment
Heinrich Küttler, Nantas Nardelli, Alexander H. Miller +4
cs.LGcs.AIcs.CLarXiv:2006.13760v22020Responsible Integration of AI in Cancer Genomics: Barriers, Risks, and Pathways to Trustworthy Clinical Translation
Bahar İlgen, Yiannos Tolias, Denise Kühnert +5
cs.AIcs.CLarXiv:2608.30912v12026ResearchAgent: Iterative Research Idea Generation over Scientific Literature with Large Language Models
Jinheon Baek, Sujay Kumar Jauhar, Silviu Cucerzan +1
cs.CLcs.AIcs.LGarXiv:2404.07738v22024S3C-LLM: Skill-Code Guided Agentic Language Models for Spectrum-to-Structure Elucidation
Xuanle Zhao, Xinyuan Cai, Xiang Cheng +1
cs.LGcs.CLarXiv:2608.30910v12026Invariant Rationalization
Shiyu Chang, Yang Zhang, Mo Yu +1
cs.LGcs.AIcs.CLarXiv:2003.09772v12020Vocal Music under Phoneme-Conditional Analysis
Hayoon Kim, Kyogu Lee
cs.SDcs.CLarXiv:2608.30823v12026The Fragility of Jailbreak Robustness Across Operational States
Yuna Park, Hwang Youn Kim, Yujin Kim +3
cs.CRcs.CLarXiv:2608.30748v12026SingProbe Technical Report
Sing Team
cs.CRcs.AIcs.CLarXiv:2608.30703v12026Beyond the Payload: How User Invocation Shapes Coding Agent Vulnerability to Repository Poisoning
Fukang Zhu, Binbin Zhao, Ruixiao Lin +3
cs.CRcs.CLarXiv:2608.30686v12026Context Staircase: Signature-Aligned Dynamics of Token Embeddings under Small Initialization
Junjie Yao, Liangkai Hang, Zhi-Qin John Xu
cs.LGcs.CLarXiv:2608.30315v12026Learning Where Outcomes Change:Credit-Addressable Reasoning for Multimodal Geometry
Jiani Guo, Junjie Wang, Jie Wu +5
cs.LGcs.CLarXiv:2608.30457v12026Judging the Judges: Evaluating Alignment and Vulnerabilities in LLMs-as-Judges
Aman Singh Thakur, Kartik Choudhary, Venkat Srinik Ramayapally +2
cs.CLcs.AIarXiv:2406.12624v62024Full Parameter Fine-tuning for Large Language Models with Limited Resources
Kai Lv, Yuqing Yang, Tengxiao Liu +3
cs.CLarXiv:2306.09782v22023Towards Faithful Model Explanation in NLP: A Survey
Qing Lyu, Marianna Apidianaki, Chris Callison-Burch
cs.CLarXiv:2209.11326v42022RoleLLM: Benchmarking, Eliciting, and Enhancing Role-Playing Abilities of Large Language Models
Zekun Moore Wang, Zhongyuan Peng, Haoran Que +14
cs.CLcs.AIarXiv:2310.00746v32023A.X K2 Technical Report
Cheolseung Baek, Dhammiko Arya, Eunki Kim +40
cs.AIcs.CLarXiv:2608.30181v12026Towards a Joint Khmer Text Recognition and Word Segmentation
Marry Kong, Rina Buoy, Sovisal Chenda +3
cs.CVcs.CLarXiv:2608.30213v12026VIBE: Video Instruction-aligned Background music gEneration
Aryan Vijay Bhosale, Vaibhavi Lokegaonkar, Vishnu Raj +5
cs.SDcs.AIcs.CLarXiv:2608.30125v12026Lot Machine: Multimodal Lot Extraction from Auction Catalogs
Mathias Zinnen, Alisha Mund, Sabine Lang +3
cs.CVcs.AIcs.CLarXiv:2608.30510v12026Prompting PaLM for Translation: Assessing Strategies and Performance
David Vilar, Markus Freitag, Colin Cherry +3
cs.CLarXiv:2211.09102v32022EvoSkill Injection: Red-Teaming Autonomous Skill Generation and Evolution in Self-Evolving Agents
Doyun Kim, Chanwoo Kim, Sugyeong Eo +2
cs.AIcs.CLarXiv:2608.30429v12026Will the User Ever Know? Covert Indirect Prompt Injection on Tool-Using LLM Agents
Yunseok Lee, Yunji Kim, Woojin Lee
cs.AIcs.CLarXiv:2608.30362v12026Ignorance or Incompetence? Constructing Knowledge-Gated, Verifiable Tasks for LLM Agents
Hanlin Tian, Minhao Li, Yu Mi +5
cs.AIcs.CLarXiv:2608.30322v12026Syntactic Structure from Deep Learning
Tal Linzen, Marco Baroni
cs.CLarXiv:2004.10827v12020Paraformer: Fast and Accurate Parallel Transformer for Non-autoregressive End-to-End Speech Recognition
Zhifu Gao, Shiliang Zhang, Ian McLoughlin +1
cs.SDcs.CLeess.ASarXiv:2206.08317v32022How Much Reading Does Reading Comprehension Require? A Critical Investigation of Popular Benchmarks
Divyansh Kaushik, Zachary C. Lipton
cs.CLcs.AIcs.LGarXiv:1808.04926v22018LaMoC: Loss-Aware Modular Compression for LLMs
Mohanad Odema, Jacob Song
cs.AIcs.CLcs.PFarXiv:2608.30226v12026MultiVerS: Improving scientific claim verification with weak supervision and full-document context
David Wadden, Kyle Lo, Lucy Lu Wang +3
cs.CLcs.AIarXiv:2112.01640v22021Noisy Channel Language Model Prompting for Few-Shot Text Classification
Sewon Min, Mike Lewis, Hannaneh Hajishirzi +1
cs.CLcs.AIarXiv:2108.04106v32021Better Fine-Tuning by Reducing Representational Collapse
Armen Aghajanyan, Akshat Shrivastava, Anchit Gupta +3
cs.LGcs.CLstat.MLarXiv:2008.03156v12020Towards Debiasing Fact Verification Models
Tal Schuster, Darsh J Shah, Yun Jie Serene Yeo +3
cs.CLarXiv:1908.05267v22019