Computation and Language
Papers filed under cs.CL on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
7,741 to 7,800 of 11,199
Fine-tune BERT for Extractive Summarization
Yang Liu
cs.CLarXiv:1903.10318v22019LAVT: Language-Aware Vision Transformer for Referring Image Segmentation
Zhao Yang, Jiaqi Wang, Yansong Tang +3
cs.CVcs.CLarXiv:2112.02244v22021$R^3$: Training Robots to Reason in Natural Language via Reinforcement Learning
Lehong Wu, Yuxiao Qu, Zheyuan Hu +4
cs.ROcs.AIcs.CLarXiv:2608.26053v12026Large Language Models Can Be Strong Differentially Private Learners
Xuechen Li, Florian Tramèr, Percy Liang +1
cs.LGcs.CLarXiv:2110.05679v62021Trace Integrity for LLM Data Agents: A Vision for Auditable Structured Reasoning in Real-World Systems
Srimonti Dutta, Akshata Kishore Moharir
cs.AIcs.CLarXiv:2608.26036v12026True Few-Shot Learning with Language Models
Ethan Perez, Douwe Kiela, Kyunghyun Cho
cs.CLcs.LGstat.MLarXiv:2105.11447v12021SemEval-2020 Task 12: Multilingual Offensive Language Identification in Social Media (OffensEval 2020)
Marcos Zampieri, Preslav Nakov, Sara Rosenthal +6
cs.CLarXiv:2006.07235v22020RLPrompt: Optimizing Discrete Text Prompts with Reinforcement Learning
Mingkai Deng, Jianyu Wang, Cheng-Ping Hsieh +6
cs.CLcs.LGarXiv:2205.12548v32022SwarmWorld: Stigmergic technological evolution in societies of language-model agents
Subhadeep Pal, Fiona Y. Wang, Markus J. Buehler
cs.AIcond-mat.mtrl-scics.CLarXiv:2608.26081v12026How Much Rank Does LoRA Need? Rank-Error Bounds for Transformer Attention
Gerard Conangla Planes
cs.LGcs.AIcs.CLarXiv:2608.26052v12026One Symptom, Three Levers: A Critical Review of On-Policy Self-Distillation
Justin Robert, Raheel Qader
cs.LGcs.AIcs.CLarXiv:2608.25936v12026Cross-Sentence N-ary Relation Extraction with Graph LSTMs
Nanyun Peng, Hoifung Poon, Chris Quirk +2
cs.CLarXiv:1708.03743v12017Beyond Local Surprise: Grounded Dialogue as Selective Belief Revision under Referential Uncertainty
Ziming Liu, Bhanu Chaitanya Jasti, Ziyang Xu +3
cs.CLarXiv:2608.26035v12026Fine-Tuning Whisper for Automatic Speech Recognition in Baniwa: A Preliminary Study
Leonardo Duart, Tiago Fonseca, Thiago Chacón
cs.CLstat.MLarXiv:2608.26060v12026Inter-GPS: Interpretable Geometry Problem Solving with Formal Language and Symbolic Reasoning
Pan Lu, Ran Gong, Shibiao Jiang +4
cs.CLcs.AIcs.CVarXiv:2105.04165v32021Distinct dynamics of conceptual and referential disruptions in human reading and large language model processing
Rui He, Nihal Altay, Wolfram Hinzen
cs.CLarXiv:2608.25999v12026A Self-Evolving Multi-Agent Framework Defense against LLM Jailbreak Attacks
Tongyan Hu, Bryan Hooi
cs.CRcs.CLarXiv:2608.26008v12026AsymSpec: Context-Asymmetric Speculative Decoding for Agentic LLMs
Sheng Liang, Yongyue Zhang, Nathanael Brian +4
cs.AIcs.CLarXiv:2608.26004v12026Check Your Facts and Try Again: Improving Large Language Models with External Knowledge and Automated Feedback
Baolin Peng, Michel Galley, Pengcheng He +8
cs.CLcs.AIarXiv:2302.12813v32023Formal, Executable and Explainable Runtime Monitoring of Spoken Air Traffic Control Operational Procedures
Roberto Luvini, Giacomo Longo, Alessandro Armando +1
cs.AIcs.CLeess.ASarXiv:2608.25926v12026Why Does Graph Learning Fail to Fully Benefit from a Text Teacher?
Fumiaki Kimino, Ryoma Sato
cs.LGcs.CLarXiv:2608.25741v12026Prefix Sliding for efficient test-time scaling
Niklas Muennighoff, Zhengyang Wang, Zeyi Chen +15
cs.CLcs.AIcs.LGarXiv:2608.26070v12026VISA: Agentic Self-Evolving Data Synthesis for Multimodal Instruction Following
Min Zeng, Guanxin Tan, Libin Cen +5
cs.CLarXiv:2608.26013v12026When Personality Meets Quantization: A Layer-wise MBTI Analysis of Quantized LLMs
Yao Fu, Lijia Huang, Xiaomin Li +3
cs.CLarXiv:2608.25977v12026LLM2Vec: Large Language Models Are Secretly Powerful Text Encoders
Parishad BehnamGhader, Vaibhav Adlakha, Marius Mosbach +3
cs.CLcs.AIarXiv:2404.05961v22024Unveiling Spectral Mechanisms in Training-Free LLM Text Detection
Haitong Luo, Xuying Meng, Weiyao Zhang +5
cs.CLarXiv:2608.25944v12026BERT: A Review of Applications in Natural Language Processing and Understanding
M. V. Koroteev
cs.CLcs.AIcs.LGarXiv:2103.11943v12021Visual Storytelling
Ting-Hao, Huang, Francis Ferraro +13
cs.CLcs.AIcs.CVarXiv:1604.03968v12016Query-Side Attacks on GNN-Based KGQA: Tracing Failures from Entity Linking to Answer Generation
Pankaj Kumar, Subhankar Mishra
cs.CLcs.AIcs.IRarXiv:2608.25922v12026Graph Retrieval-Augmented Generation: A Survey
Boci Peng, Yun Zhu, Yongchao Liu +5
cs.AIcs.CLcs.IRarXiv:2408.08921v22024XL-Sum: Large-Scale Multilingual Abstractive Summarization for 44 Languages
Tahmid Hasan, Abhik Bhattacharjee, Md Saiful Islam +5
cs.CLarXiv:2106.13822v12021An Empirical Study of Spatial Attention Mechanisms in Deep Networks
Xizhou Zhu, Dazhi Cheng, Zheng Zhang +2
cs.CVcs.CLcs.LGarXiv:1904.05873v12019Loss-Based Active Learning for Neural Abstractive Summarization
Michail Ioannou, Tatiana Passali, George Michalopoulos +1
cs.CLarXiv:2608.25881v12026Anchoring Bias in LLM-as-a-Judge Systems: Prior Scores Compromise Evaluation Independence
Ante Kapetanovic, Kemal Altwlkany, Andro Mercep +2
cs.CLarXiv:2608.25869v12026Massive Exploration of Neural Machine Translation Architectures
Denny Britz, Anna Goldie, Minh-Thang Luong +1
cs.CLarXiv:1703.03906v22017Transformer Transducer: A Streamable Speech Recognition Model with Transformer Encoders and RNN-T Loss
Qian Zhang, Han Lu, Hasim Sak +4
eess.AScs.CLcs.SDarXiv:2002.02562v22020SAMpLE: A SystemC-AMS Machine LEarning-based Framework for Virtual Prototyping
Andrei Mihai Albu, Sara Vinco
cs.CLcs.LGarXiv:2608.25910v12026DeCLUTR: Deep Contrastive Learning for Unsupervised Textual Representations
John Giorgi, Osvald Nitski, Bo Wang +1
cs.CLcs.LGarXiv:2006.03659v42020One Form to Transfer Them All: Pretraining Multilingual Language Models Beyond Native Orthography
Muge Zhang, Aaron Jencks, Krishna Badikela +2
cs.CLarXiv:2608.25904v12026From Passive Response to Proactive Correction: Enhancing LLM Robustness Against Input Fact Perturbations
Ping Wang, Xiangguo Sun, Bingbing Xu +2
cs.CLarXiv:2608.25894v12026Key Point Analysis Needs Structure Recovery: Task Definition, Dataset Diagnosis, and a Structure-Aware Benchmark
Zhiqiang Shi, Oana Cocarascu
cs.CLcs.LGarXiv:2608.25854v12026Skill Issue: Are Skills Language-Invariant in LLMs?
Bobby Cheng, Adam Gaber, Zhengyuan Liu +4
cs.CLcs.AIcs.GTarXiv:2608.25832v12026BioMistral: A Collection of Open-Source Pretrained Large Language Models for Medical Domains
Yanis Labrak, Adrien Bazoge, Emmanuel Morin +3
cs.CLcs.AIcs.LGarXiv:2402.10373v32024Unfolding Scientific Papers into Multi-Turn Generation Trajectories for Continued Pre-Training
Qiankai Xu, Qiguang Chen, Zixin Su +4
cs.CLcs.AIcs.LGarXiv:2608.25826v12026LongMemEval: Benchmarking Chat Assistants on Long-Term Interactive Memory
Di Wu, Hongwei Wang, Wenhao Yu +3
cs.CLarXiv:2410.10813v22024TOFU: A Task of Fictitious Unlearning for LLMs
Pratyush Maini, Zhili Feng, Avi Schwarzschild +2
cs.LGcs.CLarXiv:2401.06121v12024Learning New Facts with QLoRA: An Acquisition-Retention Frontier
Estelle Zheng, Sébastien Warichet, Emmanuel Helbert +1
cs.CLcs.AIcs.LGarXiv:2608.25677v12026Think-Probe-Respond: Improving Large Language Models as Judges of Research Idea Novelty
Tim Schopf, Tobias Schreieder, Akiko Aizawa
cs.CLcs.AIarXiv:2608.25660v12026Style Transfer in Text: Exploration and Evaluation
Zhenxin Fu, Xiaoye Tan, Nanyun Peng +2
cs.CLarXiv:1711.06861v22017Localize-Then-Decide Guarantees for LLM Judgments
Xinyu Li, Yi Zhou, Guanqun Cao +3
cs.CLarXiv:2608.25824v12026MoganBert-TR: A Turkish Encoder Foundation Model Trained from Scratch with a CLM-to-MLM Curriculum
Furkan Yilmaz, Habibe Aleyna Tasdemir, Muhammed Faruk Gozay
cs.CLcs.AIarXiv:2608.25768v12026Beam Search, Self-Consistency, and the Limits of Inference-Time Scaling for Grammar-Constrained Text-to-SQL in Small Language Models
Ty Chermsirivatana, John MacCormick
cs.CLcs.AIarXiv:2608.25761v12026When RAG Fails to Equalize: Geo-bias in Factual Question Answering over Public Companies
Abhinav Havaldar, Enrico Santus
cs.CLcs.AIarXiv:2608.25717v12026Overview of SHROOM-Visions 2026: A Shared Task on Hallucination Detection in Large Vision-Language Models
Raúl Vázquez, Aman Sinha, Chuyuan Li +10
cs.CLarXiv:2608.25662v12026Reconstructing the Right Episode: Evaluating Interleaved Conversational Memory Beyond Long Context
Zhexi Feng, Ruiyi Zhang, Yongbo Yang +1
cs.CLcs.AIarXiv:2608.25655v12026The Chase Is the Curriculum, the Capture Anchors the Credit: Pursuit-Evasion Self-Play for Zero-Data LLM Reasoning
Jing Yu, Shengchao Chen, Yiyun Tan
cs.CLcs.LGarXiv:2608.21871v12026Easily Accessible Text-to-Image Generation Amplifies Demographic Stereotypes at Large Scale
Federico Bianchi, Pratyusha Kalluri, Esin Durmus +7
cs.CLcs.CVarXiv:2211.03759v22022Large Language Models Are Reasoning Teachers
Namgyu Ho, Laura Schmid, Se-Young Yun
cs.CLcs.AIcs.LGarXiv:2212.10071v22022Causal Reasoning and Large Language Models: Opening a New Frontier for Causality
Emre Kıcıman, Robert Ness, Amit Sharma +1
cs.AIcs.CLcs.CYarXiv:2305.00050v32023RepoCoder: Repository-Level Code Completion Through Iterative Retrieval and Generation
Fengji Zhang, Bei Chen, Yue Zhang +6
cs.CLcs.AIcs.PLarXiv:2303.12570v32023