Computation and Language

Papers filed under cs.CL on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

2,041 to 2,100 of 11,233

  1. Reasoning in Trees: Improving Retrieval-Augmented Generation for Multi-Hop Question Answering

    Yuling Shi, Maolin Sun, Zijun Liu +4

    cs.CLcs.LGarXiv:2601.11255v12026
  2. Dynamics Within Latent Chain-of-Thought: An Empirical Study of Causal Structure

    Zirui Li, Xuefeng Bai, Kehai Chen +4

    cs.AIcs.CLarXiv:2602.08783v32026
  3. Freeze-Omni: A Smart and Low Latency Speech-to-speech Dialogue Model with Frozen LLM

    Xiong Wang, Yangze Li, Chaoyou Fu +5

    cs.SDcs.AIcs.CLarXiv:2411.00774v52024
  4. CoSDA-ML: Multi-Lingual Code-Switching Data Augmentation for Zero-Shot Cross-Lingual NLP

    Libo Qin, Minheng Ni, Yue Zhang +1

    cs.CLarXiv:2006.06402v22020
  5. PrivGemo: Privacy-Preserving Dual-Tower Graph Retrieval for Empowering LLM Reasoning with Memory Augmentation

    Xingyu Tan, Xiaoyang Wang, Qing Liu +4

    cs.CLarXiv:2601.08739v12026
  6. Chaining the Evidence: Robust Reinforcement Learning for Deep Search Agents with Citation-Aware Rubric Rewards

    Jiajie Zhang, Xin Lv, Ling Feng +2

    cs.CLarXiv:2601.06021v12026
  7. Sculpting the Vector Space: Towards Efficient Multi-Vector Visual Document Retrieval via Prune-then-Merge Framework

    Yibo Yan, Mingdong Ou, Yi Cao +5

    cs.CLcs.CVcs.IRarXiv:2602.19549v22026
  8. ES-MemEval: Benchmarking Conversational Agents on Personalized Long-Term Emotional Support

    Tiantian Chen, Jiaqi Lu, Ying Shen +1

    cs.CLcs.AIarXiv:2602.01885v12026
  9. Simple Question Answering by Attentive Convolutional Neural Network

    Wenpeng Yin, Mo Yu, Bing Xiang +2

    cs.CLarXiv:1606.03391v22016
  10. MuTual: A Dataset for Multi-Turn Dialogue Reasoning

    Leyang Cui, Yu Wu, Shujie Liu +2

    cs.CLarXiv:2004.04494v12020
  11. Industrialized Deception: The Collateral Effects of LLM-Generated Misinformation on Digital Ecosystems

    Alexander Loth, Martin Kappes, Marc-Oliver Pahl

    cs.CYcs.AIcs.CLarXiv:2601.21963v22026
  12. Neural Arabic Question Answering

    Hussein Mozannar, Karl El Hajal, Elie Maamary +1

    cs.CLcs.LGarXiv:1906.05394v12019
  13. Read As Human: Compressing Context via Parallelizable Close Reading and Skimming

    Jiwei Tang, Shilei Liu, Zhicheng Zhang +9

    cs.CLarXiv:2602.01840v22026
  14. BiasScope: Towards Automated Detection of Bias in LLM-as-a-Judge Evaluation

    Peng Lai, Zhihao Ou, Yong Wang +4

    cs.CLcs.AIcs.SEarXiv:2602.09383v12026
  15. Improving Neural Question Generation using Answer Separation

    Yanghoon Kim, Hwanhee Lee, Joongbo Shin +1

    cs.CLcs.AIcs.NEarXiv:1809.02393v22018
  16. Step Potential Advantage Estimation: Harnessing Intermediate Confidence and Correctness for Efficient Mathematical Reasoning

    Fei Wu, Zhenrong Zhang, Qikai Chang +3

    cs.CLarXiv:2601.03823v12026
  17. Exploring the Use of Text Classification in the Legal Domain

    Octavia-Maria Sulea, Marcos Zampieri, Shervin Malmasi +3

    cs.CLarXiv:1710.09306v12017
  18. AQAScore: Evaluating Semantic Alignment in Text-to-Audio Generation via Audio Question Answering

    Chun-Yi Kuan, Kai-Wei Chang, Hung-yi Lee

    eess.AScs.AIcs.CLarXiv:2601.14728v12026
  19. Confidence Estimation for LLMs in Multi-turn Interactions

    Caiqi Zhang, Ruihan Yang, Xiaochen Zhu +5

    cs.CLarXiv:2601.02179v22026
  20. Language Model Behavior: A Comprehensive Survey

    Tyler A. Chang, Benjamin K. Bergen

    cs.CLarXiv:2303.11504v22023
  21. Addressing Overthinking in Large Vision-Language Models via Gated Perception-Reasoning Optimization

    Xingjian Diao, Zheyuan Liu, Chunhui Zhang +6

    cs.CVcs.CLarXiv:2601.04442v22026
  22. Dialogue Learning with Human Teaching and Feedback in End-to-End Trainable Task-Oriented Dialogue Systems

    Bing Liu, Gokhan Tur, Dilek Hakkani-Tur +2

    cs.CLarXiv:1804.06512v12018
  23. LLaVA-Phi: Efficient Multi-Modal Assistant with Small Language Model

    Yichen Zhu, Minjie Zhu, Ning Liu +3

    cs.CVcs.CLarXiv:2401.02330v42024
  24. Stop Uploading Test Data in Plain Text: Practical Strategies for Mitigating Data Contamination by Evaluation Benchmarks

    Alon Jacovi, Avi Caciularu, Omer Goldman +1

    cs.CLcs.AIarXiv:2305.10160v22023
  25. LLaMA-MoE: Building Mixture-of-Experts from LLaMA with Continual Pre-training

    Tong Zhu, Xiaoye Qu, Daize Dong +4

    cs.CLarXiv:2406.16554v12024
  26. Adaptive Information Control for Search-Augmented LLM Reasoning

    Siheng Xiong, Oguzhan Gungordu, James C. Kerce +1

    cs.CLarXiv:2602.01672v22026
  27. Overview of the HASOC track at FIRE 2020: Hate Speech and Offensive Content Identification in Indo-European Languages

    Thomas Mandla, Sandip Modha, Gautam Kishore Shahi +5

    cs.CLcs.CYarXiv:2108.05927v12021
  28. Towards Execution-Grounded Automated AI Research

    Chenglei Si, Zitong Yang, Yejin Choi +3

    cs.CLcs.AIcs.LGarXiv:2601.14525v12026
  29. Latent Multi-task Architecture Learning

    Sebastian Ruder, Joachim Bingel, Isabelle Augenstein +1

    stat.MLcs.AIcs.CLarXiv:1705.08142v32017
  30. Listening while Speaking: Speech Chain by Deep Learning

    Andros Tjandra, Sakriani Sakti, Satoshi Nakamura

    cs.CLcs.LGcs.SDarXiv:1707.04879v12017
  31. EvoRoute: Experience-Driven Self-Routing LLM Agent Systems

    Guibin Zhang, Haiyang Yu, Kaiming Yang +4

    cs.CLcs.MAarXiv:2601.02695v12026
  32. On the Opportunities and Challenges of Foundation Models for Geospatial Artificial Intelligence

    Gengchen Mai, Weiming Huang, Jin Sun +11

    cs.AIcs.CLcs.CVarXiv:2304.06798v12023
  33. Measure and Improve Robustness in NLP Models: A Survey

    Xuezhi Wang, Haohan Wang, Diyi Yang

    cs.CLcs.LGarXiv:2112.08313v22021
  34. Birth of a Transformer: A Memory Viewpoint

    Alberto Bietti, Vivien Cabannes, Diane Bouchacourt +2

    stat.MLcs.CLcs.LGarXiv:2306.00802v22023
  35. Minimum Word Error Rate Training for Attention-based Sequence-to-Sequence Models

    Rohit Prabhavalkar, Tara N. Sainath, Yonghui Wu +4

    cs.CLeess.ASstat.MLarXiv:1712.01818v12017
  36. SaulLM-7B: A pioneering Large Language Model for Law

    Pierre Colombo, Telmo Pessoa Pires, Malik Boudiaf +8

    cs.CLarXiv:2403.03883v22024
  37. ScienceAgentBench: Toward Rigorous Assessment of Language Agents for Data-Driven Scientific Discovery

    Ziru Chen, Shijie Chen, Yuting Ning +17

    cs.CLcs.AIcs.LGarXiv:2410.05080v32024
  38. Sign-Based Optimizers Are Effective Under Heavy-Tailed Noise

    Dingzhi Yu, Hongyi Tao, Yuanyu Wan +2

    cs.LGcs.CLmath.OCarXiv:2602.07425v22026
  39. ToolGate: Contract-Grounded and Verified Tool Execution for LLMs

    Yanming Liu, Xinyue Peng, Jiannan Cao +5

    cs.CLcs.AIcs.FLarXiv:2601.04688v12026
  40. Search-P1: Path-Centric Reward Shaping for Stable and Efficient Agentic RAG Training

    Tianle Xia, Ming Xu, Lingxiang Hu +7

    cs.CLcs.IRcs.LGarXiv:2602.22576v12026
  41. Value of Information: A Framework for Human-Agent Communication

    Yijiang River Dong, Tiancheng Hu, Zheng Hui +4

    cs.CLarXiv:2601.06407v12026
  42. AfriqueLLM: How Data Mixing and Model Architecture Impact Continued Pre-training for African Languages

    Hao Yu, Tianyi Xu, Michael A. Hedderich +3

    cs.CLarXiv:2601.06395v32026
  43. Is Temperature the Creativity Parameter of Large Language Models?

    Max Peeperkorn, Tom Kouwenhoven, Dan Brown +1

    cs.CLcs.AIarXiv:2405.00492v12024
  44. Pre-training LLM without Learning Rate Decay Enhances Supervised Fine-Tuning

    Kazuki Yano, Shun Kiyono, Sosuke Kobayashi +2

    cs.CLcs.LGarXiv:2603.16127v12026
  45. OpenNRE: An Open and Extensible Toolkit for Neural Relation Extraction

    Xu Han, Tianyu Gao, Yuan Yao +3

    cs.CLarXiv:1909.13078v12019
  46. Rewards-in-Context: Multi-objective Alignment of Foundation Models with Dynamic Preference Adjustment

    Rui Yang, Xiaoman Pan, Feng Luo +4

    cs.LGcs.AIcs.CLarXiv:2402.10207v62024
  47. Adaptive Milestone Reward for GUI Agents

    Congmin Zheng, Xiaoyun Mo, Xinbei Ma +10

    cs.LGcs.AIcs.CLarXiv:2602.11524v12026
  48. SpecForge: A Flexible and Efficient Open-Source Training Framework for Speculative Decoding

    Shenggui Li, Chao Wang, Yikai Zhu +14

    cs.LGcs.AIcs.CLarXiv:2603.18567v12026
  49. Speaker-Aware BERT for Multi-Turn Response Selection in Retrieval-Based Chatbots

    Jia-Chen Gu, Tianda Li, Quan Liu +4

    cs.CLarXiv:2004.03588v22020
  50. Compositionality and Generalization in Emergent Languages

    Rahma Chaabouni, Eugene Kharitonov, Diane Bouchacourt +2

    cs.CLcs.AIcs.LGarXiv:2004.09124v12020
  51. Countdown-Code: A Testbed for Studying The Emergence and Generalization of Reward Hacking in RLVR

    Muhammad Khalifa, Zohaib Khan, Omer Tafveez +2

    cs.LGcs.AIcs.CLarXiv:2603.07084v22026
  52. R2-Router: A New Paradigm for LLM Routing with Reasoning

    Jiaqi Xue, Qian Lou, Jiarong Xing +1

    cs.CLarXiv:2602.02823v22026
  53. SUMBT: Slot-Utterance Matching for Universal and Scalable Belief Tracking

    Hwaran Lee, Jinsik Lee, Tae-Yoon Kim

    cs.CLcs.LGarXiv:1907.07421v12019
  54. Deferred Commitment Decoding for Diffusion Language Models

    Yingte Shu, Yuchuan Tian, Chao Xu +2

    cs.CLcs.AIarXiv:2601.02076v22026
  55. Natural Language Processing for EHR-Based Computational Phenotyping

    Zexian Zeng, Yu Deng, Xiaoyu Li +2

    cs.CLarXiv:1806.04820v22018
  56. Same Trajectory, Contradictory Rewards (ROBORMBENCH): Paraphrase Fragility in Vision Language Reward Models

    Wonje Jeung, Sangyeon Yoon, Hyesoo Hong +6

    cs.ROcs.CLarXiv:2609.05401v12026
  57. Repeated Queries Exhaust an LLM's Brand Recommendations but Not Its Sources

    Dmitrij Żatuchin

    cs.IRcs.CLarXiv:2609.05059v12026
  58. Maximizing Local Entropy Where It Matters: Prefix-Aware Localized LLM Unlearning

    Naixin Zhai, Pengyang Shao, Binbin Zheng +4

    cs.CLarXiv:2601.03190v42026
  59. Current Agents Fail to Leverage World Model as Tool for Foresight

    Cheng Qian, Emre Can Acikgoz, Bingxuan Li +8

    cs.AIcs.CLcs.LGarXiv:2601.03905v22026
  60. Learn-to-Distance: Distance Learning for Detecting LLM-Generated Text

    Hongyi Zhou, Jin Zhu, Kai Ye +3

    cs.CLcs.AIstat.MLarXiv:2601.21895v22026