Computation and Language
Papers filed under cs.CL on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
7,261 to 7,320 of 11,332
Assessing Gender Bias in Machine Translation -- A Case Study with Google Translate
Marcelo O. R. Prates, Pedro H. C. Avelar, Luis Lamb
cs.CYcs.CLarXiv:1809.02208v42018Neural Text Summarization: A Critical Evaluation
Wojciech Kryściński, Nitish Shirish Keskar, Bryan McCann +2
cs.CLarXiv:1908.08960v12019Trusting Your Evidence: Hallucinate Less with Context-aware Decoding
Weijia Shi, Xiaochuang Han, Mike Lewis +3
cs.CLarXiv:2305.14739v12023Towards Multimodal Sarcasm Detection (An _Obviously_ Perfect Paper)
Santiago Castro, Devamanyu Hazarika, Verónica Pérez-Rosas +3
cs.CLcs.CVarXiv:1906.01815v12019Span-based Localizing Network for Natural Language Video Localization
Hao Zhang, Aixin Sun, Wei Jing +1
cs.CLcs.CVarXiv:2004.13931v22020Speech Model Pre-training for End-to-End Spoken Language Understanding
Loren Lugosch, Mirco Ravanelli, Patrick Ignoto +2
eess.AScs.CLcs.LGarXiv:1904.03670v22019From Pretraining Data to Language Models to Downstream Tasks: Tracking the Trails of Political Biases Leading to Unfair NLP Models
Shangbin Feng, Chan Young Park, Yuhan Liu +1
cs.CLarXiv:2305.08283v32023The Thousand-Graph Hypothesis: A Testable Hypothesis of Task-Conditioned Relation Materialization in Repository-Level Code Reasoning
Fei Ding
cs.SEcs.CLarXiv:2608.26602v12026Research Design Tracking and Assessment for the Social Sciences
Marco Rovera, Sergiu Burlacu, Dominique Cappelletti +5
cs.CLcs.CYarXiv:2608.27049v12026Instruction Quality Matters: Refining Instructions for Effective Preference Learning
Seohyeong Lee, Hwaran Lee, Buru Chang
cs.CLarXiv:2608.26779v12026NLP Evaluation in trouble: On the Need to Measure LLM Data Contamination for each Benchmark
Oscar Sainz, Jon Ander Campos, Iker García-Ferrero +3
cs.CLarXiv:2310.18018v12023STAR : Sentence Translation Alignment Rate for Document-to-Document Machine Translation
Yichen Dong, Hao Wang, Junhui Li +3
cs.CLarXiv:2608.27161v12026Understanding Factuality in Abstractive Summarization with FRANK: A Benchmark for Factuality Metrics
Artidoro Pagnoni, Vidhisha Balachandran, Yulia Tsvetkov
cs.CLarXiv:2104.13346v22021SemEval-2013 Task 2: Sentiment Analysis in Twitter
Preslav Nakov, Zornitsa Kozareva, Alan Ritter +3
cs.CLcs.IRcs.LGarXiv:1912.06806v12019Multi-Expert Conformal Risk Control for Pairwise LLM Judging in Open-Ended Dialogue
Ming Cheng, Yusheng Dai, Qiuhong Ke +2
cs.CLarXiv:2608.26529v12026SWIFT:A Scalable lightWeight Infrastructure for Fine-Tuning
Yuze Zhao, Jintao Huang, Jinghan Hu +10
cs.CLarXiv:2408.05517v42024FOCUS & RePAIR: Mitigating Text Degeneration via Token-Level Guidance for Pruned Large Language Models
Junyoung Lee, Sehyeon Park, Shinhyoung Jang +5
cs.CLcs.AIcs.LGarXiv:2608.26676v12026Synthesizer: Rethinking Self-Attention in Transformer Models
Yi Tay, Dara Bahri, Donald Metzler +3
cs.CLcs.IRcs.LGarXiv:2005.00743v32020MEDITRON-70B: Scaling Medical Pretraining for Large Language Models
Zeming Chen, Alejandro Hernández Cano, Angelika Romanou +17
cs.CLcs.AIcs.LGarXiv:2311.16079v12023Multi-Hop Knowledge Graph Reasoning with Reward Shaping
Xi Victoria Lin, Richard Socher, Caiming Xiong
cs.AIcs.CLcs.LGarXiv:1808.10568v22018UniVL: A Unified Video and Language Pre-Training Model for Multimodal Understanding and Generation
Huaishao Luo, Lei Ji, Botian Shi +6
cs.CVcs.CLcs.LGarXiv:2002.06353v32020Cross-lingual Representation Learning via Centroid Intervention Fusion
Wei Sun, Marie-Francine Moens
cs.CLarXiv:2608.26357v12026Coarse-to-Fine Decoding for Neural Semantic Parsing
Li Dong, Mirella Lapata
cs.CLarXiv:1805.04793v12018Are All Languages Created Equal in Multilingual BERT?
Shijie Wu, Mark Dredze
cs.CLarXiv:2005.09093v22020CHiME-6 Challenge:Tackling Multispeaker Speech Recognition for Unsegmented Recordings
Shinji Watanabe, Michael Mandel, Jon Barker +18
cs.SDcs.CLeess.ASarXiv:2004.09249v22020Not Enough Data? Deep Learning to the Rescue!
Ateret Anaby-Tavor, Boaz Carmeli, Esther Goldbraich +5
cs.CLcs.LGarXiv:1911.03118v22019Good Friends, Bad News - Affect and Virality in Twitter
Lars Kai Hansen, Adam Arvidsson, Finn Årup Nielsen +2
cs.SIcs.CLphysics.soc-pharXiv:1101.0510v12011SpeechGym: An Audio-Native Gym for Training Voice Agents via Reinforcement Learning
Jiajun Fan, Jingyuan Li, Prashanth Gurunath Shivakumar +6
cs.SDcs.AIcs.CLarXiv:2608.26432v12026LowRankArena: A Standardized Evaluation Platform for SVD-Based LLM Compression
Zishan Shao, Lixun Zhang, Kangning Cui +10
cs.CLcs.LGarXiv:2608.26389v12026LLM-QAT: Data-Free Quantization Aware Training for Large Language Models
Zechun Liu, Barlas Oguz, Changsheng Zhao +6
cs.CLarXiv:2305.17888v12023A Hierarchical Approach for Generating Descriptive Image Paragraphs
Jonathan Krause, Justin Johnson, Ranjay Krishna +1
cs.CVcs.CLarXiv:1611.06607v22016Is ChatGPT A Good Translator? Yes With GPT-4 As The Engine
Wenxiang Jiao, Wenxuan Wang, Jen-tse Huang +3
cs.CLarXiv:2301.08745v42023Gender Bias in Neural Natural Language Processing
Kaiji Lu, Piotr Mardziel, Fangjing Wu +2
cs.CLarXiv:1807.11714v22018Data Augmentation using Pre-trained Transformer Models
Varun Kumar, Ashutosh Choudhary, Eunah Cho
cs.CLcs.LGarXiv:2003.02245v22020Style Transfer Through Back-Translation
Shrimai Prabhumoye, Yulia Tsvetkov, Ruslan Salakhutdinov +1
cs.CLarXiv:1804.09000v32018QASC: A Dataset for Question Answering via Sentence Composition
Tushar Khot, Peter Clark, Michal Guerquin +2
cs.CLarXiv:1910.11473v22019Compositionality decomposed: how do neural networks generalise?
Dieuwke Hupkes, Verna Dankers, Mathijs Mul +1
cs.CLcs.AIcs.LGarXiv:1908.08351v22019DoReMi: Optimizing Data Mixtures Speeds Up Language Model Pretraining
Sang Michael Xie, Hieu Pham, Xuanyi Dong +7
cs.CLcs.LGarXiv:2305.10429v42023Interpretable Preferences via Multi-Objective Reward Modeling and Mixture-of-Experts
Haoxiang Wang, Wei Xiong, Tengyang Xie +2
cs.LGcs.CLarXiv:2406.12845v12024D2C-Routing: Dimension-to-Composition Evidence Routing for Mixed-Origin AI-Generated Text Detection
Xin Chen, Fuwei Zhang, Yiqi Tong +3
cs.CLarXiv:2608.27380v12026Vid2Seq: Large-Scale Pretraining of a Visual Language Model for Dense Video Captioning
Antoine Yang, Arsha Nagrani, Paul Hongsuck Seo +5
cs.CVcs.AIcs.CLarXiv:2302.14115v22023Representing and Parsing Korean Constituency Structure at Different Levels of Granularity
Jungyeul Park, KyungTae Lim, Zihao Huang +3
cs.CLarXiv:2608.27035v12026PragAlign: Evidence-Sensitive Reply Assistance Across Chinese and Japanese Appropriateness Judgments
Xin Zhong, Satori Hachisuka
cs.CLarXiv:2608.26700v12026Beyond Reflection: Affirmation as a Promising Behavioral Marker Associated with Quality in Text-Based Counseling
Michimasa Inaba
cs.CLarXiv:2608.26689v12026AudioGPT: Understanding and Generating Speech, Music, Sound, and Talking Head
Rongjie Huang, Mingze Li, Dongchao Yang +10
cs.CLcs.AIcs.SDarXiv:2304.12995v12023Generated Knowledge Prompting for Commonsense Reasoning
Jiacheng Liu, Alisa Liu, Ximing Lu +5
cs.CLarXiv:2110.08387v32021TERA: Self-Supervised Learning of Transformer Encoder Representation for Speech
Andy T. Liu, Shang-Wen Li, Hung-yi Lee
eess.AScs.CLcs.LGarXiv:2007.06028v32020MixText: Linguistically-Informed Interpolation of Hidden Space for Semi-Supervised Text Classification
Jiaao Chen, Zichao Yang, Diyi Yang
cs.CLcs.LGarXiv:2004.12239v12020Learning to Navigate Unseen Environments: Back Translation with Environmental Dropout
Hao Tan, Licheng Yu, Mohit Bansal
cs.CLcs.CVcs.LGarXiv:1904.04195v12019Polisis: Automated Analysis and Presentation of Privacy Policies Using Deep Learning
Hamza Harkous, Kassem Fawaz, Rémi Lebret +3
cs.CLcs.CRcs.HCarXiv:1802.02561v22018Theoretical Limitations of Self-Attention in Neural Sequence Models
Michael Hahn
cs.CLcs.FLcs.LGarXiv:1906.06755v22019PandaLM: An Automatic Evaluation Benchmark for LLM Instruction Tuning Optimization
Yidong Wang, Zhuohao Yu, Zhengran Zeng +10
cs.CLcs.AIarXiv:2306.05087v22023FastPitch: Parallel Text-to-speech with Pitch Prediction
Adrian Łańcucki
eess.AScs.CLcs.LGarXiv:2006.06873v22020COVID-Twitter-BERT: A Natural Language Processing Model to Analyse COVID-19 Content on Twitter
Martin Müller, Marcel Salathé, Per E Kummervold
cs.CLcs.LGcs.SIarXiv:2005.07503v12020Vision-and-Dialog Navigation
Jesse Thomason, Michael Murray, Maya Cakmak +1
cs.CLcs.AIcs.CVarXiv:1907.04957v32019RepoBench: Benchmarking Repository-Level Code Auto-Completion Systems
Tianyang Liu, Canwen Xu, Julian McAuley
cs.CLcs.AIcs.SEarXiv:2306.03091v22023NeMo: a toolkit for building AI applications using Neural Modules
Oleksii Kuchaiev, Jason Li, Huyen Nguyen +11
cs.LGcs.CLcs.SDarXiv:1909.09577v12019NaturalSpeech 3: Zero-Shot Speech Synthesis with Factorized Codec and Diffusion Models
Zeqian Ju, Yuancheng Wang, Kai Shen +16
eess.AScs.AIcs.CLarXiv:2403.03100v32024ReZero is All You Need: Fast Convergence at Large Depth
Thomas Bachlechner, Bodhisattwa Prasad Majumder, Huanru Henry Mao +2
cs.LGcs.CLstat.MLarXiv:2003.04887v22020Pre-training is a Hot Topic: Contextualized Document Embeddings Improve Topic Coherence
Federico Bianchi, Silvia Terragni, Dirk Hovy
cs.CLarXiv:2004.03974v22020