Computation and Language
Papers filed under cs.CL on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
2,041 to 2,100 of 11,233
Reasoning in Trees: Improving Retrieval-Augmented Generation for Multi-Hop Question Answering
Yuling Shi, Maolin Sun, Zijun Liu +4
cs.CLcs.LGarXiv:2601.11255v12026Dynamics Within Latent Chain-of-Thought: An Empirical Study of Causal Structure
Zirui Li, Xuefeng Bai, Kehai Chen +4
cs.AIcs.CLarXiv:2602.08783v32026Freeze-Omni: A Smart and Low Latency Speech-to-speech Dialogue Model with Frozen LLM
Xiong Wang, Yangze Li, Chaoyou Fu +5
cs.SDcs.AIcs.CLarXiv:2411.00774v52024CoSDA-ML: Multi-Lingual Code-Switching Data Augmentation for Zero-Shot Cross-Lingual NLP
Libo Qin, Minheng Ni, Yue Zhang +1
cs.CLarXiv:2006.06402v22020PrivGemo: Privacy-Preserving Dual-Tower Graph Retrieval for Empowering LLM Reasoning with Memory Augmentation
Xingyu Tan, Xiaoyang Wang, Qing Liu +4
cs.CLarXiv:2601.08739v12026Chaining the Evidence: Robust Reinforcement Learning for Deep Search Agents with Citation-Aware Rubric Rewards
Jiajie Zhang, Xin Lv, Ling Feng +2
cs.CLarXiv:2601.06021v12026Sculpting the Vector Space: Towards Efficient Multi-Vector Visual Document Retrieval via Prune-then-Merge Framework
Yibo Yan, Mingdong Ou, Yi Cao +5
cs.CLcs.CVcs.IRarXiv:2602.19549v22026ES-MemEval: Benchmarking Conversational Agents on Personalized Long-Term Emotional Support
Tiantian Chen, Jiaqi Lu, Ying Shen +1
cs.CLcs.AIarXiv:2602.01885v12026Simple Question Answering by Attentive Convolutional Neural Network
Wenpeng Yin, Mo Yu, Bing Xiang +2
cs.CLarXiv:1606.03391v22016MuTual: A Dataset for Multi-Turn Dialogue Reasoning
Leyang Cui, Yu Wu, Shujie Liu +2
cs.CLarXiv:2004.04494v12020Industrialized Deception: The Collateral Effects of LLM-Generated Misinformation on Digital Ecosystems
Alexander Loth, Martin Kappes, Marc-Oliver Pahl
cs.CYcs.AIcs.CLarXiv:2601.21963v22026Neural Arabic Question Answering
Hussein Mozannar, Karl El Hajal, Elie Maamary +1
cs.CLcs.LGarXiv:1906.05394v12019Read As Human: Compressing Context via Parallelizable Close Reading and Skimming
Jiwei Tang, Shilei Liu, Zhicheng Zhang +9
cs.CLarXiv:2602.01840v22026BiasScope: Towards Automated Detection of Bias in LLM-as-a-Judge Evaluation
Peng Lai, Zhihao Ou, Yong Wang +4
cs.CLcs.AIcs.SEarXiv:2602.09383v12026Improving Neural Question Generation using Answer Separation
Yanghoon Kim, Hwanhee Lee, Joongbo Shin +1
cs.CLcs.AIcs.NEarXiv:1809.02393v22018Step Potential Advantage Estimation: Harnessing Intermediate Confidence and Correctness for Efficient Mathematical Reasoning
Fei Wu, Zhenrong Zhang, Qikai Chang +3
cs.CLarXiv:2601.03823v12026Exploring the Use of Text Classification in the Legal Domain
Octavia-Maria Sulea, Marcos Zampieri, Shervin Malmasi +3
cs.CLarXiv:1710.09306v12017AQAScore: Evaluating Semantic Alignment in Text-to-Audio Generation via Audio Question Answering
Chun-Yi Kuan, Kai-Wei Chang, Hung-yi Lee
eess.AScs.AIcs.CLarXiv:2601.14728v12026Confidence Estimation for LLMs in Multi-turn Interactions
Caiqi Zhang, Ruihan Yang, Xiaochen Zhu +5
cs.CLarXiv:2601.02179v22026Language Model Behavior: A Comprehensive Survey
Tyler A. Chang, Benjamin K. Bergen
cs.CLarXiv:2303.11504v22023Addressing Overthinking in Large Vision-Language Models via Gated Perception-Reasoning Optimization
Xingjian Diao, Zheyuan Liu, Chunhui Zhang +6
cs.CVcs.CLarXiv:2601.04442v22026Dialogue Learning with Human Teaching and Feedback in End-to-End Trainable Task-Oriented Dialogue Systems
Bing Liu, Gokhan Tur, Dilek Hakkani-Tur +2
cs.CLarXiv:1804.06512v12018LLaVA-Phi: Efficient Multi-Modal Assistant with Small Language Model
Yichen Zhu, Minjie Zhu, Ning Liu +3
cs.CVcs.CLarXiv:2401.02330v42024Stop Uploading Test Data in Plain Text: Practical Strategies for Mitigating Data Contamination by Evaluation Benchmarks
Alon Jacovi, Avi Caciularu, Omer Goldman +1
cs.CLcs.AIarXiv:2305.10160v22023LLaMA-MoE: Building Mixture-of-Experts from LLaMA with Continual Pre-training
Tong Zhu, Xiaoye Qu, Daize Dong +4
cs.CLarXiv:2406.16554v12024Adaptive Information Control for Search-Augmented LLM Reasoning
Siheng Xiong, Oguzhan Gungordu, James C. Kerce +1
cs.CLarXiv:2602.01672v22026Overview of the HASOC track at FIRE 2020: Hate Speech and Offensive Content Identification in Indo-European Languages
Thomas Mandla, Sandip Modha, Gautam Kishore Shahi +5
cs.CLcs.CYarXiv:2108.05927v12021Towards Execution-Grounded Automated AI Research
Chenglei Si, Zitong Yang, Yejin Choi +3
cs.CLcs.AIcs.LGarXiv:2601.14525v12026Latent Multi-task Architecture Learning
Sebastian Ruder, Joachim Bingel, Isabelle Augenstein +1
stat.MLcs.AIcs.CLarXiv:1705.08142v32017Listening while Speaking: Speech Chain by Deep Learning
Andros Tjandra, Sakriani Sakti, Satoshi Nakamura
cs.CLcs.LGcs.SDarXiv:1707.04879v12017EvoRoute: Experience-Driven Self-Routing LLM Agent Systems
Guibin Zhang, Haiyang Yu, Kaiming Yang +4
cs.CLcs.MAarXiv:2601.02695v12026On the Opportunities and Challenges of Foundation Models for Geospatial Artificial Intelligence
Gengchen Mai, Weiming Huang, Jin Sun +11
cs.AIcs.CLcs.CVarXiv:2304.06798v12023Measure and Improve Robustness in NLP Models: A Survey
Xuezhi Wang, Haohan Wang, Diyi Yang
cs.CLcs.LGarXiv:2112.08313v22021Birth of a Transformer: A Memory Viewpoint
Alberto Bietti, Vivien Cabannes, Diane Bouchacourt +2
stat.MLcs.CLcs.LGarXiv:2306.00802v22023Minimum Word Error Rate Training for Attention-based Sequence-to-Sequence Models
Rohit Prabhavalkar, Tara N. Sainath, Yonghui Wu +4
cs.CLeess.ASstat.MLarXiv:1712.01818v12017SaulLM-7B: A pioneering Large Language Model for Law
Pierre Colombo, Telmo Pessoa Pires, Malik Boudiaf +8
cs.CLarXiv:2403.03883v22024ScienceAgentBench: Toward Rigorous Assessment of Language Agents for Data-Driven Scientific Discovery
Ziru Chen, Shijie Chen, Yuting Ning +17
cs.CLcs.AIcs.LGarXiv:2410.05080v32024Sign-Based Optimizers Are Effective Under Heavy-Tailed Noise
Dingzhi Yu, Hongyi Tao, Yuanyu Wan +2
cs.LGcs.CLmath.OCarXiv:2602.07425v22026ToolGate: Contract-Grounded and Verified Tool Execution for LLMs
Yanming Liu, Xinyue Peng, Jiannan Cao +5
cs.CLcs.AIcs.FLarXiv:2601.04688v12026Search-P1: Path-Centric Reward Shaping for Stable and Efficient Agentic RAG Training
Tianle Xia, Ming Xu, Lingxiang Hu +7
cs.CLcs.IRcs.LGarXiv:2602.22576v12026Value of Information: A Framework for Human-Agent Communication
Yijiang River Dong, Tiancheng Hu, Zheng Hui +4
cs.CLarXiv:2601.06407v12026AfriqueLLM: How Data Mixing and Model Architecture Impact Continued Pre-training for African Languages
Hao Yu, Tianyi Xu, Michael A. Hedderich +3
cs.CLarXiv:2601.06395v32026Is Temperature the Creativity Parameter of Large Language Models?
Max Peeperkorn, Tom Kouwenhoven, Dan Brown +1
cs.CLcs.AIarXiv:2405.00492v12024Pre-training LLM without Learning Rate Decay Enhances Supervised Fine-Tuning
Kazuki Yano, Shun Kiyono, Sosuke Kobayashi +2
cs.CLcs.LGarXiv:2603.16127v12026OpenNRE: An Open and Extensible Toolkit for Neural Relation Extraction
Xu Han, Tianyu Gao, Yuan Yao +3
cs.CLarXiv:1909.13078v12019Rewards-in-Context: Multi-objective Alignment of Foundation Models with Dynamic Preference Adjustment
Rui Yang, Xiaoman Pan, Feng Luo +4
cs.LGcs.AIcs.CLarXiv:2402.10207v62024Adaptive Milestone Reward for GUI Agents
Congmin Zheng, Xiaoyun Mo, Xinbei Ma +10
cs.LGcs.AIcs.CLarXiv:2602.11524v12026SpecForge: A Flexible and Efficient Open-Source Training Framework for Speculative Decoding
Shenggui Li, Chao Wang, Yikai Zhu +14
cs.LGcs.AIcs.CLarXiv:2603.18567v12026Speaker-Aware BERT for Multi-Turn Response Selection in Retrieval-Based Chatbots
Jia-Chen Gu, Tianda Li, Quan Liu +4
cs.CLarXiv:2004.03588v22020Compositionality and Generalization in Emergent Languages
Rahma Chaabouni, Eugene Kharitonov, Diane Bouchacourt +2
cs.CLcs.AIcs.LGarXiv:2004.09124v12020Countdown-Code: A Testbed for Studying The Emergence and Generalization of Reward Hacking in RLVR
Muhammad Khalifa, Zohaib Khan, Omer Tafveez +2
cs.LGcs.AIcs.CLarXiv:2603.07084v22026R2-Router: A New Paradigm for LLM Routing with Reasoning
Jiaqi Xue, Qian Lou, Jiarong Xing +1
cs.CLarXiv:2602.02823v22026SUMBT: Slot-Utterance Matching for Universal and Scalable Belief Tracking
Hwaran Lee, Jinsik Lee, Tae-Yoon Kim
cs.CLcs.LGarXiv:1907.07421v12019Deferred Commitment Decoding for Diffusion Language Models
Yingte Shu, Yuchuan Tian, Chao Xu +2
cs.CLcs.AIarXiv:2601.02076v22026Natural Language Processing for EHR-Based Computational Phenotyping
Zexian Zeng, Yu Deng, Xiaoyu Li +2
cs.CLarXiv:1806.04820v22018Same Trajectory, Contradictory Rewards (ROBORMBENCH): Paraphrase Fragility in Vision Language Reward Models
Wonje Jeung, Sangyeon Yoon, Hyesoo Hong +6
cs.ROcs.CLarXiv:2609.05401v12026Repeated Queries Exhaust an LLM's Brand Recommendations but Not Its Sources
Dmitrij Żatuchin
cs.IRcs.CLarXiv:2609.05059v12026Maximizing Local Entropy Where It Matters: Prefix-Aware Localized LLM Unlearning
Naixin Zhai, Pengyang Shao, Binbin Zheng +4
cs.CLarXiv:2601.03190v42026Current Agents Fail to Leverage World Model as Tool for Foresight
Cheng Qian, Emre Can Acikgoz, Bingxuan Li +8
cs.AIcs.CLcs.LGarXiv:2601.03905v22026Learn-to-Distance: Distance Learning for Detecting LLM-Generated Text
Hongyi Zhou, Jin Zhu, Kai Ye +3
cs.CLcs.AIstat.MLarXiv:2601.21895v22026