Computation and Language
Papers filed under cs.CL on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
1,741 to 1,800 of 11,242
Neural Latent Extractive Document Summarization
Xingxing Zhang, Mirella Lapata, Furu Wei +1
cs.CLcs.AIcs.LGarXiv:1808.07187v22018LLM-Assisted Content Analysis: Using Large Language Models to Support Deductive Coding
Robert Chew, John Bollenbacher, Michael Wenger +2
cs.CLcs.AIcs.LGarXiv:2306.14924v12023A Reparameterized Discrete Diffusion Model for Text Generation
Lin Zheng, Jianbo Yuan, Lei Yu +1
cs.CLcs.LGarXiv:2302.05737v32023Efficient and Robust Question Answering from Minimal Context over Documents
Sewon Min, Victor Zhong, Richard Socher +1
cs.CLarXiv:1805.08092v12018Black-Box Prompt Optimization: Aligning Large Language Models without Model Training
Jiale Cheng, Xiao Liu, Kehan Zheng +5
cs.CLarXiv:2311.04155v32023The Parallel Meaning Bank: Towards a Multilingual Corpus of Translations Annotated with Compositional Meaning Representations
Lasha Abzianidze, Johannes Bjerva, Kilian Evang +5
cs.CLarXiv:1702.03964v12017Coarse-to-Fine Vision-Language Pre-training with Fusion in the Backbone
Zi-Yi Dou, Aishwarya Kamath, Zhe Gan +9
cs.CVcs.CLcs.LGarXiv:2206.07643v22022X-FACTR: Multilingual Factual Knowledge Retrieval from Pretrained Language Models
Zhengbao Jiang, Antonios Anastasopoulos, Jun Araki +2
cs.CLarXiv:2010.06189v32020Extractive Summarization of Long Documents by Combining Global and Local Context
Wen Xiao, Giuseppe Carenini
cs.CLarXiv:1909.08089v12019Vidur: A Large-Scale Simulation Framework For LLM Inference
Amey Agrawal, Nitin Kedia, Jayashree Mohan +5
cs.LGcs.AIcs.CLarXiv:2405.05465v22024LinCE: A Centralized Benchmark for Linguistic Code-switching Evaluation
Gustavo Aguilar, Sudipta Kar, Thamar Solorio
cs.CLarXiv:2005.04322v12020Topic Modeling Based Multi-modal Depression Detection
Yuan Gong, Christian Poellabauer
cs.CLcs.IRcs.LGarXiv:1803.10384v12018Chart-to-Text: Generating Natural Language Descriptions for Charts by Adapting the Transformer Model
Jason Obeid, Enamul Hoque
cs.CLcs.AIarXiv:2010.09142v22020Likelihood-Based Diffusion Language Models
Ishaan Gulrajani, Tatsunori B. Hashimoto
cs.CLcs.LGarXiv:2305.18619v12023Neural Open Information Extraction
Lei Cui, Furu Wei, Ming Zhou
cs.CLarXiv:1805.04270v12018AuK Technical Report: An Open-Source Foundational Model for Speech Generation and Editing
Ziyang Ma, Zhikang Niu, Wenming Tu +30
cs.SDcs.CLcs.MMarXiv:2609.08936v12026Turning Up the Heat: Min-p Sampling for Creative and Coherent LLM Outputs
Minh Nhat Nguyen, Andrew Baker, Clement Neo +3
cs.CLarXiv:2407.01082v82024Lookback Lens: Detecting and Mitigating Contextual Hallucinations in Large Language Models Using Only Attention Maps
Yung-Sung Chuang, Linlu Qiu, Cheng-Yu Hsieh +3
cs.CLcs.AIcs.LGarXiv:2407.07071v22024MixLoRA: Enhancing Large Language Models Fine-Tuning with LoRA-based Mixture of Experts
Dengchun Li, Yingzi Ma, Naizheng Wang +8
cs.CLcs.AIarXiv:2404.15159v32024Dreaddit: A Reddit Dataset for Stress Analysis in Social Media
Elsbeth Turcan, Kathleen McKeown
cs.CLarXiv:1911.00133v12019Dialog State Tracking: A Neural Reading Comprehension Approach
Shuyang Gao, Abhishek Sethi, Sanchit Agarwal +2
cs.CLcs.LGarXiv:1908.01946v32019Deep Multimodal Semantic Embeddings for Speech and Images
David Harwath, James Glass
cs.CVcs.AIcs.CLarXiv:1511.03690v12015Can neural machine translation do simultaneous translation?
Kyunghyun Cho, Masha Esipova
cs.CLarXiv:1606.02012v12016XuanYuan 2.0: A Large Chinese Financial Chat Model with Hundreds of Billions Parameters
Xuanyu Zhang, Qing Yang, Dongliang Xu
cs.CLarXiv:2305.12002v12023End-to-End Slot Alignment and Recognition for Cross-Lingual NLU
Weijia Xu, Batool Haider, Saab Mansour
cs.CLcs.LGarXiv:2004.14353v22020EnvCraft: Synthesizing Executable Environments in Agentic RL for Claw-like Agent
Yirong Zeng, Shen You, Jinhang Feng +8
cs.AIcs.CLarXiv:2609.05576v12026meProp: Sparsified Back Propagation for Accelerated Deep Learning with Reduced Overfitting
Xu Sun, Xuancheng Ren, Shuming Ma +1
cs.LGcs.AIcs.CLarXiv:1706.06197v52017Overview of the CLEF-2018 CheckThat! Lab on Automatic Identification and Verification of Political Claims. Task 1: Check-Worthiness
Pepa Atanasova, Alberto Barron-Cedeno, Tamer Elsayed +5
cs.CLarXiv:1808.05542v12018Retrieval-Augmented Multimodal Language Modeling
Michihiro Yasunaga, Armen Aghajanyan, Weijia Shi +6
cs.CVcs.CLcs.LGarXiv:2211.12561v22022A Survey on Gender Bias in Natural Language Processing
Karolina Stanczak, Isabelle Augenstein
cs.CLcs.CYarXiv:2112.14168v12021RLEF: Grounding Code LLMs in Execution Feedback with Reinforcement Learning
Jonas Gehring, Kunhao Zheng, Jade Copet +4
cs.CLcs.AIarXiv:2410.02089v22024Large Generative AI Models for Telecom: The Next Big Thing?
Lina Bariah, Qiyang Zhao, Hang Zou +3
cs.CLcs.AIarXiv:2306.10249v22023Large Language Models as Urban Residents: An LLM Agent Framework for Personal Mobility Generation
Jiawei Wang, Renhe Jiang, Chuang Yang +5
cs.AIcs.CLcs.CYarXiv:2402.14744v32024When Benchmarks are Targets: Revealing the Sensitivity of Large Language Model Leaderboards
Norah Alzahrani, Hisham Abdullah Alyahya, Yazeed Alnumay +9
cs.CLcs.AIcs.LGarXiv:2402.01781v22024Reasoning-Aware Compression: Identifying and Protecting Vulnerable Reasoning Circuits for Energy-Efficient LLM Deployment
Leonard Twagirayezu, Prasenjit Mitra
cs.AIcs.CLcs.PFarXiv:2609.05512v12026AutoWebGLM: A Large Language Model-based Web Navigating Agent
Hanyu Lai, Xiao Liu, Iat Long Iong +8
cs.CLarXiv:2404.03648v22024SCAFFOLD: Self-Improving Web Agents via Recursive Parametric Skill Abstraction
Bowei He, Xiaokun Zhang, Meng Ding +1
cs.AIcs.CLarXiv:2609.05511v12026Word2Vec applied to Recommendation: Hyperparameters Matter
Hugo Caselles-Dupré, Florian Lesaint, Jimena Royo-Letelier
cs.IRcs.CLcs.LGarXiv:1804.04212v32018Fighting Offensive Language on Social Media with Unsupervised Text Style Transfer
Cicero Nogueira dos Santos, Igor Melnyk, Inkit Padhi
cs.CLcs.LGarXiv:1805.07685v12018Fake News in Sheep's Clothing: Robust Fake News Detection Against LLM-Empowered Style Attacks
Jiaying Wu, Jiafeng Guo, Bryan Hooi
cs.CLarXiv:2310.10830v22023How Grammatical is Character-level Neural Machine Translation? Assessing MT Quality with Contrastive Translation Pairs
Rico Sennrich
cs.CLarXiv:1612.04629v32016Rich Knowledge Sources Bring Complex Knowledge Conflicts: Recalibrating Models to Reflect Conflicting Evidence
Hung-Ting Chen, Michael J. Q. Zhang, Eunsol Choi
cs.CLarXiv:2210.13701v12022Multi-hop Reading Comprehension across Multiple Documents by Reasoning over Heterogeneous Graphs
Ming Tu, Guangtao Wang, Jing Huang +3
cs.CLarXiv:1905.07374v22019Procedural Graphs: Self-Evolving Execution Structures for LLM Agents
Yuxing Lu, Yicheng Chen, Shanchan Wu +1
cs.AIcs.CLcs.MAarXiv:2609.09153v12026Racial Disparity in Natural Language Processing: A Case Study of Social Media African-American English
Su Lin Blodgett, Brendan O'Connor
cs.CYcs.CLarXiv:1707.00061v12017Reason Through the Latent! Making Latent Visual Reasoning Necessary
Suhyeong Park, Junha Jung, Jaewoo Kang
cs.AIcs.CLcs.CVarXiv:2609.06746v12026Microsoft Translator at WMT 2019: Towards Large-Scale Document-Level Neural Machine Translation
Marcin Junczys-Dowmunt
cs.CLarXiv:1907.06170v12019Eliciting Weak-to-Strong Generalization with On-Policy Reverse Distillation
Youngrok Park, Sangmin Bae, Hojung Jung +6
cs.LGcs.CLarXiv:2609.08798v12026Navigation with Large Language Models: Semantic Guesswork as a Heuristic for Planning
Dhruv Shah, Michael Equi, Blazej Osinski +3
cs.ROcs.AIcs.CLarXiv:2310.10103v12023Hierarchical Question Answering for Long Documents
Eunsol Choi, Daniel Hewlett, Alexandre Lacoste +3
cs.CLarXiv:1611.01839v22016Design Challenges and Misconceptions in Neural Sequence Labeling
Jie Yang, Shuailong Liang, Yue Zhang
cs.CLarXiv:1806.04470v22018Steering Geometry: Validating Human Value Geometry in LLM Steering Space
Mohammad Mahdi Abootorabi, Armin Saghafian, Ali Bazshoushtari +5
cs.CLcs.AIcs.LGarXiv:2609.06289v12026Sequence-to-Sequence Neural Net Models for Grapheme-to-Phoneme Conversion
Kaisheng Yao, Geoffrey Zweig
cs.CLarXiv:1506.00196v32015On NMT Search Errors and Model Errors: Cat Got Your Tongue?
Felix Stahlberg, Bill Byrne
cs.CLarXiv:1908.10090v12019The Impact of Synthetic Data Augmentation on Discourse-Pragmatic Function Classification
Sara Sorahi, Kevin Tang, Reza Kazemian
cs.CLarXiv:2609.03652v12026EditNTS: An Neural Programmer-Interpreter Model for Sentence Simplification through Explicit Editing
Yue Dong, Zichao Li, Mehdi Rezagholizadeh +1
cs.CLarXiv:1906.08104v12019Rephrasing the Web: A Recipe for Compute and Data-Efficient Language Modeling
Pratyush Maini, Skyler Seto, He Bai +3
cs.CLarXiv:2401.16380v12024Miles v0.1: Production-Level Post-Training
RadixArk, :, Tom Chen +11
cs.LGcs.CLarXiv:2609.08368v12026A Circuit for Plural Reference: How LLMs Represent and Retrieve Singular and Plural Entities
Anh Danh, Rick Nouwen, Massimo Poesio
cs.CLarXiv:2609.03687v12026Opening mind by opening architecture: analysis strategies
Francesco Vitucci, Giuseppe Silvi, Daniele Giuseppe Annese +2
cs.CLarXiv:2609.03719v12026