Computation and Language

Papers filed under cs.CL on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

1,321 to 1,380 of 11,226

  1. Language agents achieve superhuman synthesis of scientific knowledge

    Michael D. Skarlinski, Sam Cox, Jon M. Laurent +6

    cs.CLcs.AIcs.IRarXiv:2409.13740v22024
  2. On the Exploitability of Instruction Tuning

    Manli Shu, Jiongxiao Wang, Chen Zhu +3

    cs.CRcs.CLcs.LGarXiv:2306.17194v22023
  3. Effectiveness of self-supervised pre-training for speech recognition

    Alexei Baevski, Michael Auli, Abdelrahman Mohamed

    cs.CLcs.LGarXiv:1911.03912v32019
  4. Reducing Gender Bias in Neural Machine Translation as a Domain Adaptation Problem

    Danielle Saunders, Bill Byrne

    cs.CLarXiv:2004.04498v32020
  5. Byte Latent Transformer: Patches Scale Better Than Tokens

    Artidoro Pagnoni, Ram Pasunuru, Pedro Rodriguez +11

    cs.CLarXiv:2412.09871v12024
  6. Federated Fine-tuning of Large Language Models under Heterogeneous Tasks and Client Resources

    Jiamu Bai, Daoyuan Chen, Bingchen Qian +2

    cs.CLcs.AIarXiv:2402.11505v22024
  7. A Confederacy of Models: a Comprehensive Evaluation of LLMs on Creative Writing

    Carlos Gómez-Rodríguez, Paul Williams

    cs.CLcs.CYarXiv:2310.08433v12023
  8. Quantifying Logical Consistency in Transformers via Query-Key Alignment

    Eduard Tulchinskii, Anastasia Voznyuk, Laida Kushnareva +4

    cs.CLcs.AIcs.ITarXiv:2502.17017v12025
  9. GREEN: Generative Radiology Report Evaluation and Error Notation

    Sophie Ostmeier, Justin Xu, Zhihong Chen +8

    cs.CLcs.AIarXiv:2405.03595v22024
  10. Certifiably Robust RAG against Retrieval Corruption

    Chong Xiang, Tong Wu, Zexuan Zhong +3

    cs.LGcs.CLcs.CRarXiv:2405.15556v22024
  11. Commonsense Knowledge Base Completion with Structural and Semantic Context

    Chaitanya Malaviya, Chandra Bhagavatula, Antoine Bosselut +1

    cs.CLcs.AIarXiv:1910.02915v22019
  12. UnitBoost: Managing Compound LLM Systems with a Merge Operator, Not a Model

    Xing Zhang, Guanghui Wang, Yanwei Cui +2

    cs.AIcs.CLcs.MAarXiv:2609.09815v12026
  13. From Plausible to Actionable: A Position on LLM Self-Explanations

    Elize Herrewijnen, Benedetta Muscato, Gizem Gezici +1

    cs.CLcs.AIarXiv:2607.15957v32026
  14. More Data, More Relations, More Context and More Openness: A Review and Outlook for Relation Extraction

    Xu Han, Tianyu Gao, Yankai Lin +7

    cs.CLarXiv:2004.03186v32020
  15. Can Artificial Intelligence Support Healthcare and Mental Health Through Early Cyberbullying Detection ? The Impact of Emotion-Aware AI on Proactive Online Safety

    Hamed Jelodar, Amir Firouzi, Yen-Wu Lo +2

    cs.AIcs.CLarXiv:2609.09735v12026
  16. Evaluating Large Language Models on a Highly-specialized Topic, Radiation Oncology Physics

    Jason Holmes, Zhengliang Liu, Lian Zhang +8

    physics.med-phcs.CLphysics.ed-pharXiv:2304.01938v12023
  17. In conversation with Artificial Intelligence: aligning language models with human values

    Atoosa Kasirzadeh, Iason Gabriel

    cs.CYcs.CLarXiv:2209.00731v22022
  18. CityPlanner: A Sandbox Agent for Executable Urban Planning

    Wentao Zhang, Jingyuan Wang, Zetong Zhou +2

    cs.AIcs.CLarXiv:2609.09578v12026
  19. rVAD: An Unsupervised Segment-Based Robust Voice Activity Detection Method

    Zheng-Hua Tan, Achintya kr. Sarkar, Najim Dehak

    cs.SDcs.CLcs.LGarXiv:1906.03588v22019
  20. SpotServe: Serving Generative Large Language Models on Preemptible Instances

    Xupeng Miao, Chunan Shi, Jiangfei Duan +4

    cs.DCcs.CLcs.LGarXiv:2311.15566v12023
  21. Every Activation Boosted: Scaling General Reasoner to 1 Trillion Open Language Foundation

    Ling Team, Ang Li, Ben Liu +139

    cs.CLcs.AIcs.LGarXiv:2510.22115v22025
  22. Learning like a Child: Fast Novel Visual Concept Learning from Sentence Descriptions of Images

    Junhua Mao, Wei Xu, Yi Yang +3

    cs.CVcs.CLcs.LGarXiv:1504.06692v22015
  23. EventKG: A Multilingual Event-Centric Temporal Knowledge Graph

    Simon Gottschalk, Elena Demidova

    cs.CLcs.DBarXiv:1804.04526v12018
  24. Towards Conversational Diagnostic AI

    Tao Tu, Anil Palepu, Mike Schaekermann +22

    cs.AIcs.CLcs.LGarXiv:2401.05654v12024
  25. DianShi-RxnDB: A Large-Scale, Fine-Grained Organic Reaction Data Platform Built via a Fully Automated Pipeline for Researchers and AI Agents

    Yubin Wang, Xingjian Wei, Jiang Wu +33

    cs.CLarXiv:2609.06703v12026
  26. Lifelong Learning for Sentiment Classification

    Zhiyuan Chen, Nianzu Ma, Bing Liu

    cs.CLcs.IRcs.LGarXiv:1801.02808v12018
  27. Question Answering for Privacy Policies: Combining Computational and Legal Perspectives

    Abhilasha Ravichander, Alan W Black, Shomir Wilson +2

    cs.CLarXiv:1911.00841v12019
  28. EvalLM: Interactive Evaluation of Large Language Model Prompts on User-Defined Criteria

    Tae Soo Kim, Yoonjoo Lee, Jamin Shin +2

    cs.HCcs.AIcs.CLarXiv:2309.13633v22023
  29. One Country, 700+ Languages: NLP Challenges for Underrepresented Languages and Dialects in Indonesia

    Alham Fikri Aji, Genta Indra Winata, Fajri Koto +9

    cs.CLarXiv:2203.13357v12022
  30. Subagents vs Agent Skills: Executing Reusable Knowledge for Long-Horizon Agentic Tasks

    Wasu Top Piriyakulkij, Rachel Lawrence, Alicia Curth +2

    cs.AIcs.CLcs.LGarXiv:2609.09233v12026
  31. Leveraging Graph to Improve Abstractive Multi-Document Summarization

    Wei Li, Xinyan Xiao, Jiachen Liu +3

    cs.CLarXiv:2005.10043v12020
  32. Generate & Rank: A Multi-task Framework for Math Word Problems

    Jianhao Shen, Yichun Yin, Lin Li +4

    cs.CLcs.AIarXiv:2109.03034v12021
  33. Multilingual Distributed Representations without Word Alignment

    Karl Moritz Hermann, Phil Blunsom

    cs.CLarXiv:1312.6173v42013
  34. Routing to the Expert: Efficient Reward-guided Ensemble of Large Language Models

    Keming Lu, Hongyi Yuan, Runji Lin +4

    cs.CLcs.LGarXiv:2311.08692v12023
  35. Truthful AI: Developing and governing AI that does not lie

    Owain Evans, Owen Cotton-Barratt, Lukas Finnveden +5

    cs.CYcs.AIcs.CLarXiv:2110.06674v12021
  36. Text-Free Prosody-Aware Generative Spoken Language Modeling

    Eugene Kharitonov, Ann Lee, Adam Polyak +8

    cs.CLcs.LGcs.SDarXiv:2109.03264v22021
  37. KronA: Parameter Efficient Tuning with Kronecker Adapter

    Ali Edalati, Marzieh Tahaei, Ivan Kobyzev +3

    cs.CLarXiv:2212.10650v12022
  38. A Recipe For Arbitrary Text Style Transfer with Large Language Models

    Emily Reif, Daphne Ippolito, Ann Yuan +3

    cs.CLarXiv:2109.03910v42021
  39. Pile of Law: Learning Responsible Data Filtering from the Law and a 256GB Open-Source Legal Dataset

    Peter Henderson, Mark S. Krass, Lucia Zheng +4

    cs.CLcs.CYarXiv:2207.00220v22022
  40. Ovis: Structural Embedding Alignment for Multimodal Large Language Model

    Shiyin Lu, Yang Li, Qing-Guo Chen +4

    cs.CVcs.AIcs.CLarXiv:2405.20797v22024
  41. M$^3$CoT: A Novel Benchmark for Multi-Domain Multi-step Multi-modal Chain-of-Thought

    Qiguang Chen, Libo Qin, Jin Zhang +3

    cs.CVcs.AIcs.CLarXiv:2405.16473v12024
  42. Rethinking Cooperative Rationalization: Introspective Extraction and Complement Control

    Mo Yu, Shiyu Chang, Yang Zhang +1

    cs.CLcs.LGarXiv:1910.13294v22019
  43. End-to-End Multimodal Fact-Checking and Explanation Generation: A Challenging Dataset and Models

    Barry Menglong Yao, Aditya Shah, Lichao Sun +2

    cs.CLarXiv:2205.12487v22022
  44. Turn the Combination Lock: Learnable Textual Backdoor Attacks via Word Substitution

    Fanchao Qi, Yuan Yao, Sophia Xu +2

    cs.CLcs.CRarXiv:2106.06361v12021
  45. Revisiting Complete Reasoning Traces for Post-Training

    Jaehui Hwang, Sangdoo Yun, Byeongho Heo +1

    cs.CLarXiv:2609.07103v12026
  46. $Φ$-Bench: Can Large Language Models Engineer the Infrastructure That Powers Them?

    Leilei Ding, Shumin Wang, Yuting Huang +10

    cs.CLarXiv:2609.10226v12026
  47. X-LLM: Bootstrapping Advanced Large Language Models by Treating Multi-Modalities as Foreign Languages

    Feilong Chen, Minglun Han, Haozhi Zhao +4

    cs.CLcs.AIcs.CVarXiv:2305.04160v32023
  48. EfficientQAT: Efficient Quantization-Aware Training for Large Language Models

    Mengzhao Chen, Wenqi Shao, Peng Xu +4

    cs.LGcs.AIcs.CLarXiv:2407.11062v32024
  49. Superfiltering: Weak-to-Strong Data Filtering for Fast Instruction-Tuning

    Ming Li, Yong Zhang, Shwai He +5

    cs.CLarXiv:2402.00530v22024
  50. Scruples: A Corpus of Community Ethical Judgments on 32,000 Real-Life Anecdotes

    Nicholas Lourie, Ronan Le Bras, Yejin Choi

    cs.CLarXiv:2008.09094v22020
  51. AgenticGen: Reward-Guided Agentic Video Generation for Advertising

    Xingyuan Bu, Chengru Song, Hao Zhou +9

    cs.CVcs.AIcs.CLarXiv:2609.09187v12026
  52. Continual Pre-Training for Cross-Lingual LLM Adaptation: Enhancing Japanese Language Capabilities

    Kazuki Fujii, Taishi Nakamura, Mengsay Loem +7

    cs.CLcs.AIarXiv:2404.17790v12024
  53. Parallel Iterative Edit Models for Local Sequence Transduction

    Abhijeet Awasthi, Sunita Sarawagi, Rasna Goyal +2

    cs.CLcs.LGarXiv:1910.02893v22019
  54. Attenuating Bias in Word Vectors

    Sunipa Dev, Jeff Phillips

    cs.CLarXiv:1901.07656v12019
  55. An Improved Baseline for Sentence-level Relation Extraction

    Wenxuan Zhou, Muhao Chen

    cs.CLarXiv:2102.01373v42021
  56. Efficient One-Pass End-to-End Entity Linking for Questions

    Belinda Z. Li, Sewon Min, Srinivasan Iyer +2

    cs.CLcs.AIarXiv:2010.02413v12020
  57. From Scores to Evidence: Auditable Decisions Can Improve Speech Deepfake Detection

    Mengzhe Geng, Yujia Lu, Patrick Littell +2

    cs.SDcs.CLeess.ASarXiv:2609.08899v22026
  58. Evaluation of sentence embeddings in downstream and linguistic probing tasks

    Christian S. Perone, Roberto Silveira, Thomas S. Paula

    cs.CLarXiv:1806.06259v12018
  59. The Rater Ising-Potts Model with LLM-Derived Weights: An Application to Multi-Category Scoring Reliability

    Matthias von Davier

    stat.APcs.CLarXiv:2609.08797v12026
  60. Learning Length-Extrapolatable Recurrent Models

    Hanwen Jiang

    cs.LGcs.CLarXiv:2609.09157v12026