Computation and Language
Papers filed under cs.CL on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
7,861 to 7,920 of 11,278
Skill Issue: Are Skills Language-Invariant in LLMs?
Bobby Cheng, Adam Gaber, Zhengyuan Liu +4
cs.CLcs.AIcs.GTarXiv:2608.25832v12026BioMistral: A Collection of Open-Source Pretrained Large Language Models for Medical Domains
Yanis Labrak, Adrien Bazoge, Emmanuel Morin +3
cs.CLcs.AIcs.LGarXiv:2402.10373v32024Unfolding Scientific Papers into Multi-Turn Generation Trajectories for Continued Pre-Training
Qiankai Xu, Qiguang Chen, Zixin Su +4
cs.CLcs.AIcs.LGarXiv:2608.25826v12026LongMemEval: Benchmarking Chat Assistants on Long-Term Interactive Memory
Di Wu, Hongwei Wang, Wenhao Yu +3
cs.CLarXiv:2410.10813v22024TOFU: A Task of Fictitious Unlearning for LLMs
Pratyush Maini, Zhili Feng, Avi Schwarzschild +2
cs.LGcs.CLarXiv:2401.06121v12024Learning New Facts with QLoRA: An Acquisition-Retention Frontier
Estelle Zheng, Sébastien Warichet, Emmanuel Helbert +1
cs.CLcs.AIcs.LGarXiv:2608.25677v12026Think-Probe-Respond: Improving Large Language Models as Judges of Research Idea Novelty
Tim Schopf, Tobias Schreieder, Akiko Aizawa
cs.CLcs.AIarXiv:2608.25660v12026Style Transfer in Text: Exploration and Evaluation
Zhenxin Fu, Xiaoye Tan, Nanyun Peng +2
cs.CLarXiv:1711.06861v22017Localize-Then-Decide Guarantees for LLM Judgments
Xinyu Li, Yi Zhou, Guanqun Cao +3
cs.CLarXiv:2608.25824v12026MoganBert-TR: A Turkish Encoder Foundation Model Trained from Scratch with a CLM-to-MLM Curriculum
Furkan Yilmaz, Habibe Aleyna Tasdemir, Muhammed Faruk Gozay
cs.CLcs.AIarXiv:2608.25768v12026Beam Search, Self-Consistency, and the Limits of Inference-Time Scaling for Grammar-Constrained Text-to-SQL in Small Language Models
Ty Chermsirivatana, John MacCormick
cs.CLcs.AIarXiv:2608.25761v12026When RAG Fails to Equalize: Geo-bias in Factual Question Answering over Public Companies
Abhinav Havaldar, Enrico Santus
cs.CLcs.AIarXiv:2608.25717v12026Overview of SHROOM-Visions 2026: A Shared Task on Hallucination Detection in Large Vision-Language Models
Raúl Vázquez, Aman Sinha, Chuyuan Li +10
cs.CLarXiv:2608.25662v12026Reconstructing the Right Episode: Evaluating Interleaved Conversational Memory Beyond Long Context
Zhexi Feng, Ruiyi Zhang, Yongbo Yang +1
cs.CLcs.AIarXiv:2608.25655v12026The Chase Is the Curriculum, the Capture Anchors the Credit: Pursuit-Evasion Self-Play for Zero-Data LLM Reasoning
Jing Yu, Shengchao Chen, Yiyun Tan
cs.CLcs.LGarXiv:2608.21871v12026Easily Accessible Text-to-Image Generation Amplifies Demographic Stereotypes at Large Scale
Federico Bianchi, Pratyusha Kalluri, Esin Durmus +7
cs.CLcs.CVarXiv:2211.03759v22022Large Language Models Are Reasoning Teachers
Namgyu Ho, Laura Schmid, Se-Young Yun
cs.CLcs.AIcs.LGarXiv:2212.10071v22022Causal Reasoning and Large Language Models: Opening a New Frontier for Causality
Emre Kıcıman, Robert Ness, Amit Sharma +1
cs.AIcs.CLcs.CYarXiv:2305.00050v32023RepoCoder: Repository-Level Code Completion Through Iterative Retrieval and Generation
Fengji Zhang, Bei Chen, Yue Zhang +6
cs.CLcs.AIcs.PLarXiv:2303.12570v32023RRHF: Rank Responses to Align Language Models with Human Feedback without tears
Zheng Yuan, Hongyi Yuan, Chuanqi Tan +3
cs.CLarXiv:2304.05302v32023LexGLUE: A Benchmark Dataset for Legal Language Understanding in English
Ilias Chalkidis, Abhik Jana, Dirk Hartung +4
cs.CLarXiv:2110.00976v42021DExperts: Decoding-Time Controlled Text Generation with Experts and Anti-Experts
Alisa Liu, Maarten Sap, Ximing Lu +4
cs.CLarXiv:2105.03023v22021DeeBERT: Dynamic Early Exiting for Accelerating BERT Inference
Ji Xin, Raphael Tang, Jaejun Lee +2
cs.CLcs.LGarXiv:2004.12993v12020Learning to Generate Reviews and Discovering Sentiment
Alec Radford, Rafal Jozefowicz, Ilya Sutskever
cs.LGcs.CLcs.NEarXiv:1704.01444v22017Learning Mixtures of Plackett-Luce Models for Multi-Objective Alignment
Dongyue Li, Ziniu Zhang, Lu Wang +1
cs.LGcs.AIcs.CLarXiv:2608.25200v12026Reasoning on Graphs: Faithful and Interpretable Large Language Model Reasoning
Linhao Luo, Yuan-Fang Li, Gholamreza Haffari +1
cs.CLcs.AIarXiv:2310.01061v22023Recurrent Neural Network Grammars
Chris Dyer, Adhiguna Kuncoro, Miguel Ballesteros +1
cs.CLcs.NEarXiv:1602.07776v42016Conditional Total Correlation and the Serial Depth of Adaptive Parallel Sampling
Chuling Wen, Weijie Liang, Jian Lu
cs.ITcs.CLarXiv:2608.25505v12026DCGC: Draft-Conditioned Global Correction for Complex Reasoning with Masked Diffusion Models
Minhae Oh, Nakyung Lee, Jungwoo Lee
cs.CLcs.AIarXiv:2608.25428v12026GGSS: Geodesic-Gated Spherical Steering for Inference-Time Debiasing of Generative Vision-Language Models
Yiqun Sun, Junyu Chen, Pengfei Wei +1
cs.CYcs.CLcs.CVarXiv:2608.25375v12026Hyena Hierarchy: Towards Larger Convolutional Language Models
Michael Poli, Stefano Massaroli, Eric Nguyen +6
cs.LGcs.CLarXiv:2302.10866v32023GRIP: Granular Reward-Guided Parameter Interpolation for Efficient Reasoning
Lam So, Canhui Wu, Han Lin
cs.CLarXiv:2608.25583v12026Adaptive Triggering for Bias Correction in LLM Reasoning
Nayoung Kim, Mickey Mancenido, Huan Liu
cs.CLcs.AIarXiv:2608.25379v12026Abductive Commonsense Reasoning
Chandra Bhagavatula, Ronan Le Bras, Chaitanya Malaviya +6
cs.CLarXiv:1908.05739v22019Demystifying Reinforcement Learning Post-Training of Language Models
Donovan Clay, Saket Gollapudi, Sankar Harilal +4
cs.LGcs.AIcs.CLarXiv:2608.24949v12026Retrieved But Not Reliable: A Survey on Attacks, and Defenses in Retrieval-Augmented Generation
Minh Tran, Cuong Dang, Tuc Nguyen +9
cs.CRcs.CLcs.LGarXiv:2608.24977v12026Modulating early visual processing by language
Harm de Vries, Florian Strub, Jérémie Mary +3
cs.CVcs.CLcs.LGarXiv:1707.00683v32017Chain-of-Verification Reduces Hallucination in Large Language Models
Shehzaad Dhuliawala, Mojtaba Komeili, Jing Xu +4
cs.CLcs.AIarXiv:2309.11495v22023FinRiskAtlas: Decision-Aligned Evaluation of Large Language Models for Financial Risk Review
Suyang Zhong, Jingzhe Zhu, Qi Xu +5
cs.AIcs.CLarXiv:2608.25325v12026Understanding and Improving Layer Normalization
Jingjing Xu, Xu Sun, Zhiyuan Zhang +2
cs.LGcs.CLstat.MLarXiv:1911.07013v12019MathAdv: What Theorem Provers Know, Reason, Formalize, and Generalize
Jiaxin Yuan, Connor Martinez Lockhart, Xiaoyu Liu +11
cs.CLcs.AIcs.LOarXiv:2608.25449v12026The Changing Geometry of Grammar: Dimensionality and Neighborhood Reorganization across Transformer Layers
Samuele Vallisa, Federico Ravenda, Claudio Palominos +5
cs.CLarXiv:2608.25166v12026Belief Cascades Drive Persuasion in LLM Agent Networks
Haoyi Qiu, Genglin Liu, Pranav Narayanan Venkit +4
cs.CLcs.AIarXiv:2608.25152v12026MC-CXR: A Multi-Context Chest X-ray Benchmark for Context-Induced Disruption in Vision-Language Models
Junhyeok Lee, Songsoo Kim, Kyu Sung Choi
cs.CLarXiv:2608.24118v12026Experts, Errors, and Context: A Large-Scale Study of Human Evaluation for Machine Translation
Markus Freitag, George Foster, David Grangier +3
cs.CLcs.AIcs.LGarXiv:2104.14478v12021LibriBrain100: One Hundred Hours of Broad and Deep MEG Data for Neural Speech Decoding at Scale
Francesco Mantegna, Dulhan Jayalath, Gereon Elvers +12
cs.LGcs.CLarXiv:2608.25204v12026Apples to Apples? Towards Comparable Crosslingual Language Model Evaluation
Xiulin Yang, Ethan Gotlieb Wilcox, Catherine Arnett
cs.CLarXiv:2608.25089v12026Tensor2Tensor for Neural Machine Translation
Ashish Vaswani, Samy Bengio, Eugene Brevdo +10
cs.LGcs.CLstat.MLarXiv:1803.07416v12018Massively Multilingual Neural Machine Translation
Roee Aharoni, Melvin Johnson, Orhan Firat
cs.CLarXiv:1903.00089v32019Retrieve, Match, Escalate: Accurate and Scalable Product Linking with VLM-Distilled Cross-Encoders and Agentic VLMs
Jian Wang, Steven Xu, Sanjyot Thete +5
cs.AIcs.CLcs.DBarXiv:2608.25037v12026RefLAM: A Reference-Grounded Line Annotation Pipeline for Historical Arabic Manuscripts
Mohamed Guechaoui, Mohamed Diaa Zellagui, Souleyman Chaib +1
cs.CVcs.CLarXiv:2608.25140v12026Plans You Can Check: Verifier-Grounded Learning of an Open-Weight Planner for Executable Video-Editing
Haoyu Wang, Cheng Feng, Liuyang Bian +5
cs.CVcs.CLarXiv:2608.25622v12026Learning What to Share and What to Personalize: Hierarchical Strategy Co-Evolution for Agent Memory
Yupeng Han, Shuochen Liu, Kai Zhang +3
cs.AIcs.CLarXiv:2608.25329v12026ClueWeaver: Reward-Guided Dual-Agent Evidence Reasoning for Compact LLMs on Literary Long Narratives
Jihao Zhu, Zhiwei Yang, Wenxiao Zhang +7
cs.CLarXiv:2608.25531v12026SelfGraphRAG: Bridging the Supervision Gap in Graph-Based RAG with Synthetic QA Generation
Ben Lagnese, Manas Gaur
cs.CLcs.AIarXiv:2608.25123v12026A Survey on Text Classification: From Shallow to Deep Learning
Qian Li, Hao Peng, Jianxin Li +5
cs.CLarXiv:2008.00364v62020EgoArgus: Benchmarking VLMs as Situational Assistants for Modality-Grounded User Supports
Yu-Chien Tang, Yu-Hsiang Liu, An-Zi Yen
cs.CLarXiv:2608.25561v12026OmniPhys: A Unified Multimodal Benchmark for Physics Understanding and Generation from Chinese Educational Corpora
Hao Chen, Yumin Lin, Nadila Yushanjiang +2
cs.CLarXiv:2608.25398v12026Leveraging Speech Acts for Low-Data and Cross-Domain Conversation Derailment Forecasting
Angela Yifei Yuan, Christine De Kock, Christopher Leckie
cs.CLarXiv:2608.25359v12026Can We Read the Mind of an Audio LLM? A Verbalizable, Multilingual Middle-Layer Workspace
Jiajun Fan, Jingyuan Li, Prashanth Gurunath Shivakumar +8
cs.SDcs.AIcs.CLarXiv:2608.24958v12026