Computation and Language
Papers filed under cs.CL on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
10,201 to 10,260 of 11,332
Combinatorial Synthesis: Scaling Code RLVR via Atomic Decomposition and Recombination
Jiasheng Zheng, Boxi Cao, Boxi Yu +6
cs.CLcs.SEarXiv:2605.31058v12026Stacked Attention Networks for Image Question Answering
Zichao Yang, Xiaodong He, Jianfeng Gao +2
cs.LGcs.CLcs.CVarXiv:1511.02274v22015OCC-RAG: Optimal Cognitive Core for Faithful Question Answering
Maksim Savkin, Mikhail Goncharov, Alexander Gambashidze +7
cs.CLarXiv:2606.00683v12026Do Text Edits Generalize to Visual Generation? Benchmarking Cross-Modal Knowledge Editing in UMMs
Xin Gao, Cheng Yang, Chufan Shi +1
cs.CLcs.CVarXiv:2606.00477v12026On the Limits of LLM Adaptability: Impact of Model-Internalized Priors on Annotation Task Performance
Etienne Casanova, Rafal Kocielnik, R. Michael Alvarez
cs.CLcs.AIcs.LGarXiv:2606.00467v12026DiffWave: A Versatile Diffusion Model for Audio Synthesis
Zhifeng Kong, Wei Ping, Jiaji Huang +2
eess.AScs.CLcs.LGarXiv:2009.09761v32020LongAttnComp: Cross-Family Context Compression for Long-Context Reasoning
Mengmeng Ji, Ravi Shanker Raju, Jonathan Lingjie Li +1
cs.CLarXiv:2606.01336v22026Habitat: A Platform for Embodied AI Research
Manolis Savva, Abhishek Kadian, Oleksandr Maksymets +9
cs.CVcs.AIcs.CLarXiv:1904.01201v22019Who Annotates in NLP? A Large-scale Assessment of Human Annotation Reporting between 2018 and 2025
Maria Kunilovskaya, Gagan Bhatia, Lisa Sophie Albertelli +10
cs.CLcs.AIarXiv:2606.02255v12026Summaries:한국어Semantic Motion Anchors: Bridging Motion and Meaning in Co-Speech Gestures
Varsha Suresh, Mohammad Mahdi Abootorabi, Mohamed Salman +5
cs.CLarXiv:2605.30608v32026A Local Perturbation Theory for Cross-Domain Interference and Recovery in Multi-Domain RL
Lei Yang, Siyu Ding, Deyi Xiong
cs.LGcs.CLarXiv:2606.02398v12026Scaling LLM Test-Time Compute Optimally can be More Effective than Scaling Model Parameters
Charlie Snell, Jaehoon Lee, Kelvin Xu +1
cs.LGcs.CLarXiv:2408.03314v12024Spider: A Large-Scale Human-Labeled Dataset for Complex and Cross-Domain Semantic Parsing and Text-to-SQL Task
Tao Yu, Rui Zhang, Kai Yang +9
cs.CLcs.AIarXiv:1809.08887v52018An Enigma of Artificial Reason: Investigating the Production-Evaluation Gap in Large Reasoning Models
Mingzhong Sun, Teresa Yeo, Armando Solar-Lezama +1
cs.AIcs.CLcs.LGarXiv:2606.01462v12026Speech Commands: A Dataset for Limited-Vocabulary Speech Recognition
Pete Warden
cs.CLcs.HCarXiv:1804.03209v12018Hierarchical Neural Story Generation
Angela Fan, Mike Lewis, Yann Dauphin
cs.CLarXiv:1805.04833v12018ToolLLM: Facilitating Large Language Models to Master 16000+ Real-world APIs
Yujia Qin, Shihao Liang, Yining Ye +16
cs.AIcs.CLcs.LGarXiv:2307.16789v22023Trust Functions: Near-Lossless Weak-to-Strong Generalization by Learning When to Trust the Weak Teacher
Arda Uzunoglu, Alvin Zhang, Daniel Khashabi
cs.LGcs.CLarXiv:2606.01000v12026Pythia: A Suite for Analyzing Large Language Models Across Training and Scaling
Stella Biderman, Hailey Schoelkopf, Quentin Anthony +10
cs.CLarXiv:2304.01373v22023Is Your Code Generated by ChatGPT Really Correct? Rigorous Evaluation of Large Language Models for Code Generation
Jiawei Liu, Chunqiu Steven Xia, Yuyao Wang +1
cs.SEcs.CLcs.LGarXiv:2305.01210v32023Geometric Latent Reasoning Induces Shorter Generations in LLMs
Shashi Kumar, Yacouba Kaloga, Petr Motlicek +2
cs.CLarXiv:2606.02248v12026Mixtral of Experts
Albert Q. Jiang, Alexandre Sablayrolles, Antoine Roux +23
cs.LGcs.CLarXiv:2401.04088v12024DeBERTaV3: Improving DeBERTa using ELECTRA-Style Pre-Training with Gradient-Disentangled Embedding Sharing
Pengcheng He, Jianfeng Gao, Weizhu Chen
cs.CLcs.LGarXiv:2111.09543v42021Tacotron: Towards End-to-End Speech Synthesis
Yuxuan Wang, RJ Skerry-Ryan, Daisy Stanton +11
cs.CLcs.LGcs.SDarXiv:1703.10135v22017LLM Anonymization Against Agentic Re-Identification
Ziwen Li, Jianing Wen, Tianshi Li
cs.CRcs.CLarXiv:2605.30848v22026LayerRoute: Input-Conditioned Adaptive Layer Skipping via LoRA Fine-Tuning for Agentic Language Models
Prateek Kumar Sikdar
cs.CLcs.AIcs.LGarXiv:2606.01838v12026Graph Convolutional Networks for Text Classification
Liang Yao, Chengsheng Mao, Yuan Luo
cs.CLcs.AIarXiv:1809.05679v32018MMG2Skill: Can Agents Distill In-the-Wild Guides into Self-Evolving Skills?
Xinyu Che, Junqi Xiong, Yunfei Ge +10
cs.CLcs.AIcs.LGarXiv:2606.01993v12026AgentCL: Toward Rigorous Evaluation of Continual Learning in Language Agents
Yiheng Shu, Bernal Jiménez Gutiérrez, Saisri Padmaja Jonnalagedda +3
cs.AIcs.CLarXiv:2606.02461v22026Agentic Chain-of-Thought Steering for Efficient and Controllable LLM Reasoning
Yu Xia, Zhouhang Xie, Xin Xu +4
cs.CLcs.AIarXiv:2606.03965v12026Large Language Models Hack Rewards, and Society
Wei Liu, Xinyi Mou, Hanqi Yan +2
cs.LGcs.AIcs.CLarXiv:2606.04075v22026Improving Factuality and Reasoning in Language Models through Multiagent Debate
Yilun Du, Shuang Li, Antonio Torralba +2
cs.CLcs.AIcs.CVarXiv:2305.14325v12023Universal Sentence Encoder
Daniel Cer, Yinfei Yang, Sheng-yi Kong +10
cs.CLarXiv:1803.11175v22018Challenging BIG-Bench Tasks and Whether Chain-of-Thought Can Solve Them
Mirac Suzgun, Nathan Scales, Nathanael Schärli +8
cs.CLcs.AIarXiv:2210.09261v12022What Does BERT Look At? An Analysis of BERT's Attention
Kevin Clark, Urvashi Khandelwal, Omer Levy +1
cs.CLarXiv:1906.04341v12019MemTrain: Self-Supervised Context Memory Training
Ziheng Li, Xingrun Xing, Haoqing Wang +2
cs.CLarXiv:2606.03197v12026STRIDE: Training Data Attribution via Sparse Recovery from Subset Perturbations
Rishit Dagli, Abir Harrasse, Luke Zhang +4
cs.LGcs.CLarXiv:2606.05165v12026WebRISE: Requirement-Induced State Evaluation for MLLM-Generated Web Artifacts
Yuxin Meng, Yuhan Suo, Junjie Wang +9
cs.CLcs.AIarXiv:2606.03220v12026SemEval-2017 Task 1: Semantic Textual Similarity - Multilingual and Cross-lingual Focused Evaluation
Daniel Cer, Mona Diab, Eneko Agirre +2
cs.CLarXiv:1708.00055v12017M$^3$Eval: Multi-Modal Memory Evaluation through Cognitively-Grounded Video Tasks
Jie Huang, Ruixun Liu, Sirui Sun +4
cs.CVcs.AIcs.CLarXiv:2606.05008v12026Bidirectional Attention Flow for Machine Comprehension
Minjoon Seo, Aniruddha Kembhavi, Ali Farhadi +1
cs.CLarXiv:1611.01603v62016Rethinking the Role of Demonstrations: What Makes In-Context Learning Work?
Sewon Min, Xinxi Lyu, Ari Holtzman +4
cs.CLcs.AIarXiv:2202.12837v22022Small RL Controller, Large Language Model: RL-Guided Adaptive Sampling for Test-Time Scaling
Runpeng Dai, Tong Zheng, Rui Liu +2
cs.CLarXiv:2606.03102v12026World Models Meet Language Models: On the Complementarity of Concrete and Abstract Reasoning
Yucheng Zhou, Wei Tao, Yiwen Guo +1
cs.CVcs.CLarXiv:2606.03603v12026Multilingual Denoising Pre-training for Neural Machine Translation
Yinhan Liu, Jiatao Gu, Naman Goyal +5
cs.CLarXiv:2001.08210v22020KletterMix: Climbing Toward High-Quality German Pretraining Data - The Full Report
Maurice Kraus, Ruben Härle, Sebastian Sztwiertnia +5
cs.CLarXiv:2606.03773v22026Stateful Visual Encoders for Vision-Language Models
Zirui Wang, Junwei Yu, Adam Yala +3
cs.CVcs.CLcs.LGarXiv:2606.04433v12026Probing Outcome-Level Resemblance and Mechanism-Level Alignment in LLM Risk Decisions: Evidence from the St. Petersburg Game
Chensong Huang, Changyu Chen, Chenwei Lin +3
cs.CLcs.CYecon.GNarXiv:2606.04978v12026Stanza: A Python Natural Language Processing Toolkit for Many Human Languages
Peng Qi, Yuhao Zhang, Yuhui Zhang +2
cs.CLarXiv:2003.07082v22020DAR: Deontic Reasoning with Agentic Harnesses
Guangyao Dou, William Jurayj, Nils Holzenberger +1
cs.CLcs.AIarXiv:2606.05009v12026Streaming Communication in Multi-Agent Reasoning
Zhen Yang, Xiaogang Xu, Wen Wang +3
cs.CLcs.AIcs.MAarXiv:2606.05158v22026Audio Interaction Model
Zhifei Xie, Zihang Liu, Ze An +8
cs.SDcs.AIcs.CLarXiv:2606.05121v12026TIDE: Proactive Multi-Problem Discovery via Template-Guided Iteration
Soyeong Jeong, Jinheon Baek, Minki Kang +1
cs.CLcs.AIcs.LGarXiv:2606.04743v22026Revising Context, Shifting Simulated Stance: Auditing LLM-Based Stance Simulation in Online Discussions
Xinnong Zhang, Wanting Shan, Hanjia Lyu +2
cs.CLcs.MMcs.SIarXiv:2606.06443v22026Towards Truly Multilingual ASR: Generalizing Code-Switching ASR to Unseen Language Pairs
Gio Paik, Hyunseo Shin, Soungmin Lee
cs.CLeess.ASarXiv:2606.05846v22026SePO: Self-Evolving Prompt Agent for System Prompt Optimization
Wangcheng Tao, Han Wu, Weng-Fai Wong
cs.CLcs.AIarXiv:2606.04465v12026GENEB: Why Genomic Models Are Hard to Compare
Daria Ledneva, Mikhail Nuridinov, Denis Kuznetsov
cs.CLcs.LGq-bio.GNarXiv:2606.04525v42026SpanBERT: Improving Pre-training by Representing and Predicting Spans
Mandar Joshi, Danqi Chen, Yinhan Liu +3
cs.CLcs.LGarXiv:1907.10529v32019Statistically Reliable LLM-Based Ranking Evaluation via Prediction-Powered Inference
Abhishek Divekar
cs.LGcs.AIcs.CLarXiv:2606.05308v12026AURA: Intent-Directed Probing for Implicit-Need Surfacing in Situated LLM Agents
Yang Li, Jiaxiang Liu, Jiang Cai +1
cs.CLarXiv:2606.05557v12026