Computation and Language
Papers filed under cs.CL on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
7,801 to 7,860 of 11,237
VISA: Agentic Self-Evolving Data Synthesis for Multimodal Instruction Following
Min Zeng, Guanxin Tan, Libin Cen +5
cs.CLarXiv:2608.26013v12026When Personality Meets Quantization: A Layer-wise MBTI Analysis of Quantized LLMs
Yao Fu, Lijia Huang, Xiaomin Li +3
cs.CLarXiv:2608.25977v12026LLM2Vec: Large Language Models Are Secretly Powerful Text Encoders
Parishad BehnamGhader, Vaibhav Adlakha, Marius Mosbach +3
cs.CLcs.AIarXiv:2404.05961v22024Unveiling Spectral Mechanisms in Training-Free LLM Text Detection
Haitong Luo, Xuying Meng, Weiyao Zhang +5
cs.CLarXiv:2608.25944v12026BERT: A Review of Applications in Natural Language Processing and Understanding
M. V. Koroteev
cs.CLcs.AIcs.LGarXiv:2103.11943v12021Visual Storytelling
Ting-Hao, Huang, Francis Ferraro +13
cs.CLcs.AIcs.CVarXiv:1604.03968v12016Query-Side Attacks on GNN-Based KGQA: Tracing Failures from Entity Linking to Answer Generation
Pankaj Kumar, Subhankar Mishra
cs.CLcs.AIcs.IRarXiv:2608.25922v12026Graph Retrieval-Augmented Generation: A Survey
Boci Peng, Yun Zhu, Yongchao Liu +5
cs.AIcs.CLcs.IRarXiv:2408.08921v22024XL-Sum: Large-Scale Multilingual Abstractive Summarization for 44 Languages
Tahmid Hasan, Abhik Bhattacharjee, Md Saiful Islam +5
cs.CLarXiv:2106.13822v12021An Empirical Study of Spatial Attention Mechanisms in Deep Networks
Xizhou Zhu, Dazhi Cheng, Zheng Zhang +2
cs.CVcs.CLcs.LGarXiv:1904.05873v12019Loss-Based Active Learning for Neural Abstractive Summarization
Michail Ioannou, Tatiana Passali, George Michalopoulos +1
cs.CLarXiv:2608.25881v12026Anchoring Bias in LLM-as-a-Judge Systems: Prior Scores Compromise Evaluation Independence
Ante Kapetanovic, Kemal Altwlkany, Andro Mercep +2
cs.CLarXiv:2608.25869v12026Massive Exploration of Neural Machine Translation Architectures
Denny Britz, Anna Goldie, Minh-Thang Luong +1
cs.CLarXiv:1703.03906v22017Transformer Transducer: A Streamable Speech Recognition Model with Transformer Encoders and RNN-T Loss
Qian Zhang, Han Lu, Hasim Sak +4
eess.AScs.CLcs.SDarXiv:2002.02562v22020SAMpLE: A SystemC-AMS Machine LEarning-based Framework for Virtual Prototyping
Andrei Mihai Albu, Sara Vinco
cs.CLcs.LGarXiv:2608.25910v12026DeCLUTR: Deep Contrastive Learning for Unsupervised Textual Representations
John Giorgi, Osvald Nitski, Bo Wang +1
cs.CLcs.LGarXiv:2006.03659v42020One Form to Transfer Them All: Pretraining Multilingual Language Models Beyond Native Orthography
Muge Zhang, Aaron Jencks, Krishna Badikela +2
cs.CLarXiv:2608.25904v12026From Passive Response to Proactive Correction: Enhancing LLM Robustness Against Input Fact Perturbations
Ping Wang, Xiangguo Sun, Bingbing Xu +2
cs.CLarXiv:2608.25894v12026Key Point Analysis Needs Structure Recovery: Task Definition, Dataset Diagnosis, and a Structure-Aware Benchmark
Zhiqiang Shi, Oana Cocarascu
cs.CLcs.LGarXiv:2608.25854v12026Skill Issue: Are Skills Language-Invariant in LLMs?
Bobby Cheng, Adam Gaber, Zhengyuan Liu +4
cs.CLcs.AIcs.GTarXiv:2608.25832v12026BioMistral: A Collection of Open-Source Pretrained Large Language Models for Medical Domains
Yanis Labrak, Adrien Bazoge, Emmanuel Morin +3
cs.CLcs.AIcs.LGarXiv:2402.10373v32024Unfolding Scientific Papers into Multi-Turn Generation Trajectories for Continued Pre-Training
Qiankai Xu, Qiguang Chen, Zixin Su +4
cs.CLcs.AIcs.LGarXiv:2608.25826v12026LongMemEval: Benchmarking Chat Assistants on Long-Term Interactive Memory
Di Wu, Hongwei Wang, Wenhao Yu +3
cs.CLarXiv:2410.10813v22024TOFU: A Task of Fictitious Unlearning for LLMs
Pratyush Maini, Zhili Feng, Avi Schwarzschild +2
cs.LGcs.CLarXiv:2401.06121v12024Learning New Facts with QLoRA: An Acquisition-Retention Frontier
Estelle Zheng, Sébastien Warichet, Emmanuel Helbert +1
cs.CLcs.AIcs.LGarXiv:2608.25677v12026Think-Probe-Respond: Improving Large Language Models as Judges of Research Idea Novelty
Tim Schopf, Tobias Schreieder, Akiko Aizawa
cs.CLcs.AIarXiv:2608.25660v12026Style Transfer in Text: Exploration and Evaluation
Zhenxin Fu, Xiaoye Tan, Nanyun Peng +2
cs.CLarXiv:1711.06861v22017Localize-Then-Decide Guarantees for LLM Judgments
Xinyu Li, Yi Zhou, Guanqun Cao +3
cs.CLarXiv:2608.25824v12026MoganBert-TR: A Turkish Encoder Foundation Model Trained from Scratch with a CLM-to-MLM Curriculum
Furkan Yilmaz, Habibe Aleyna Tasdemir, Muhammed Faruk Gozay
cs.CLcs.AIarXiv:2608.25768v12026Beam Search, Self-Consistency, and the Limits of Inference-Time Scaling for Grammar-Constrained Text-to-SQL in Small Language Models
Ty Chermsirivatana, John MacCormick
cs.CLcs.AIarXiv:2608.25761v12026When RAG Fails to Equalize: Geo-bias in Factual Question Answering over Public Companies
Abhinav Havaldar, Enrico Santus
cs.CLcs.AIarXiv:2608.25717v12026Overview of SHROOM-Visions 2026: A Shared Task on Hallucination Detection in Large Vision-Language Models
Raúl Vázquez, Aman Sinha, Chuyuan Li +10
cs.CLarXiv:2608.25662v12026Reconstructing the Right Episode: Evaluating Interleaved Conversational Memory Beyond Long Context
Zhexi Feng, Ruiyi Zhang, Yongbo Yang +1
cs.CLcs.AIarXiv:2608.25655v12026The Chase Is the Curriculum, the Capture Anchors the Credit: Pursuit-Evasion Self-Play for Zero-Data LLM Reasoning
Jing Yu, Shengchao Chen, Yiyun Tan
cs.CLcs.LGarXiv:2608.21871v12026Easily Accessible Text-to-Image Generation Amplifies Demographic Stereotypes at Large Scale
Federico Bianchi, Pratyusha Kalluri, Esin Durmus +7
cs.CLcs.CVarXiv:2211.03759v22022Large Language Models Are Reasoning Teachers
Namgyu Ho, Laura Schmid, Se-Young Yun
cs.CLcs.AIcs.LGarXiv:2212.10071v22022Causal Reasoning and Large Language Models: Opening a New Frontier for Causality
Emre Kıcıman, Robert Ness, Amit Sharma +1
cs.AIcs.CLcs.CYarXiv:2305.00050v32023RepoCoder: Repository-Level Code Completion Through Iterative Retrieval and Generation
Fengji Zhang, Bei Chen, Yue Zhang +6
cs.CLcs.AIcs.PLarXiv:2303.12570v32023RRHF: Rank Responses to Align Language Models with Human Feedback without tears
Zheng Yuan, Hongyi Yuan, Chuanqi Tan +3
cs.CLarXiv:2304.05302v32023LexGLUE: A Benchmark Dataset for Legal Language Understanding in English
Ilias Chalkidis, Abhik Jana, Dirk Hartung +4
cs.CLarXiv:2110.00976v42021DExperts: Decoding-Time Controlled Text Generation with Experts and Anti-Experts
Alisa Liu, Maarten Sap, Ximing Lu +4
cs.CLarXiv:2105.03023v22021DeeBERT: Dynamic Early Exiting for Accelerating BERT Inference
Ji Xin, Raphael Tang, Jaejun Lee +2
cs.CLcs.LGarXiv:2004.12993v12020Learning to Generate Reviews and Discovering Sentiment
Alec Radford, Rafal Jozefowicz, Ilya Sutskever
cs.LGcs.CLcs.NEarXiv:1704.01444v22017Learning Mixtures of Plackett-Luce Models for Multi-Objective Alignment
Dongyue Li, Ziniu Zhang, Lu Wang +1
cs.LGcs.AIcs.CLarXiv:2608.25200v12026Reasoning on Graphs: Faithful and Interpretable Large Language Model Reasoning
Linhao Luo, Yuan-Fang Li, Gholamreza Haffari +1
cs.CLcs.AIarXiv:2310.01061v22023Recurrent Neural Network Grammars
Chris Dyer, Adhiguna Kuncoro, Miguel Ballesteros +1
cs.CLcs.NEarXiv:1602.07776v42016Conditional Total Correlation and the Serial Depth of Adaptive Parallel Sampling
Chuling Wen, Weijie Liang, Jian Lu
cs.ITcs.CLarXiv:2608.25505v12026DCGC: Draft-Conditioned Global Correction for Complex Reasoning with Masked Diffusion Models
Minhae Oh, Nakyung Lee, Jungwoo Lee
cs.CLcs.AIarXiv:2608.25428v12026GGSS: Geodesic-Gated Spherical Steering for Inference-Time Debiasing of Generative Vision-Language Models
Yiqun Sun, Junyu Chen, Pengfei Wei +1
cs.CYcs.CLcs.CVarXiv:2608.25375v12026Hyena Hierarchy: Towards Larger Convolutional Language Models
Michael Poli, Stefano Massaroli, Eric Nguyen +6
cs.LGcs.CLarXiv:2302.10866v32023GRIP: Granular Reward-Guided Parameter Interpolation for Efficient Reasoning
Lam So, Canhui Wu, Han Lin
cs.CLarXiv:2608.25583v12026Adaptive Triggering for Bias Correction in LLM Reasoning
Nayoung Kim, Mickey Mancenido, Huan Liu
cs.CLcs.AIarXiv:2608.25379v12026Abductive Commonsense Reasoning
Chandra Bhagavatula, Ronan Le Bras, Chaitanya Malaviya +6
cs.CLarXiv:1908.05739v22019Demystifying Reinforcement Learning Post-Training of Language Models
Donovan Clay, Saket Gollapudi, Sankar Harilal +4
cs.LGcs.AIcs.CLarXiv:2608.24949v12026Retrieved But Not Reliable: A Survey on Attacks, and Defenses in Retrieval-Augmented Generation
Minh Tran, Cuong Dang, Tuc Nguyen +9
cs.CRcs.CLcs.LGarXiv:2608.24977v12026Modulating early visual processing by language
Harm de Vries, Florian Strub, Jérémie Mary +3
cs.CVcs.CLcs.LGarXiv:1707.00683v32017Chain-of-Verification Reduces Hallucination in Large Language Models
Shehzaad Dhuliawala, Mojtaba Komeili, Jing Xu +4
cs.CLcs.AIarXiv:2309.11495v22023FinRiskAtlas: Decision-Aligned Evaluation of Large Language Models for Financial Risk Review
Suyang Zhong, Jingzhe Zhu, Qi Xu +5
cs.AIcs.CLarXiv:2608.25325v12026Understanding and Improving Layer Normalization
Jingjing Xu, Xu Sun, Zhiyuan Zhang +2
cs.LGcs.CLstat.MLarXiv:1911.07013v12019MathAdv: What Theorem Provers Know, Reason, Formalize, and Generalize
Jiaxin Yuan, Connor Martinez Lockhart, Xiaoyu Liu +11
cs.CLcs.AIcs.LOarXiv:2608.25449v12026