Computation and Language
Papers filed under cs.CL on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
2,881 to 2,940 of 11,247
Multi-domain Dialog State Tracking using Recurrent Neural Networks
Nikola Mrkšić, Diarmuid Ó Séaghdha, Blaise Thomson +5
cs.CLcs.LGarXiv:1506.07190v12015IHEval: Evaluating Language Models on Following the Instruction Hierarchy
Zhihan Zhang, Shiyang Li, Zixuan Zhang +11
cs.CLarXiv:2502.08745v22025A Survey of LLM-based Deep Search Agents: Paradigm, Optimization, Evaluation, and Challenges
Yunjia Xi, Jianghao Lin, Yongzhao Xiao +7
cs.IRcs.AIcs.CLarXiv:2508.05668v32025Stabilizing MoE Reinforcement Learning by Aligning Training and Inference Routers
Wenhan Ma, Hailin Zhang, Liang Zhao +4
cs.CLcs.AIcs.LGarXiv:2510.11370v22025Sheared LLaMA: Accelerating Language Model Pre-training via Structured Pruning
Mengzhou Xia, Tianyu Gao, Zhiyuan Zeng +1
cs.CLcs.AIcs.LGarXiv:2310.06694v22023Human-AI Collaboration Enables More Empathic Conversations in Text-based Peer-to-Peer Mental Health Support
Ashish Sharma, Inna W. Lin, Adam S. Miner +2
cs.CLcs.HCcs.SIarXiv:2203.15144v12022Release Strategies and the Social Impacts of Language Models
Irene Solaiman, Miles Brundage, Jack Clark +12
cs.CLcs.AIcs.CYarXiv:1908.09203v22019R1-Reward: Training Multimodal Reward Model Through Stable Reinforcement Learning
Yi-Fan Zhang, Xingyu Lu, Xiao Hu +13
cs.CVcs.CLarXiv:2505.02835v22025Large Language Models and Games: A Survey and Roadmap
Roberto Gallotta, Graham Todd, Marvin Zammit +4
cs.CLcs.AIcs.HCarXiv:2402.18659v52024Speech2Vec: A Sequence-to-Sequence Framework for Learning Word Embeddings from Speech
Yu-An Chung, James Glass
cs.CLarXiv:1803.08976v22018InfiGUIAgent: A Multimodal Generalist GUI Agent with Native Reasoning and Reflection
Yuhang Liu, Pengxiang Li, Zishu Wei +7
cs.AIcs.CLcs.HCarXiv:2501.04575v12025Beyond BLEU: Training Neural Machine Translation with Semantic Similarity
John Wieting, Taylor Berg-Kirkpatrick, Kevin Gimpel +1
cs.CLarXiv:1909.06694v12019EvoFlint: An Evolutionary Atlas of Multi-Turn LLM Vulnerabilities
Feitong Qiao, Liren Peng, Shiming Ren +7
cs.CLcs.AIcs.CRarXiv:2609.00487v12026SCROLLS: Standardized CompaRison Over Long Language Sequences
Uri Shaham, Elad Segal, Maor Ivgi +8
cs.CLcs.AIcs.LGarXiv:2201.03533v22022Demystifying Reasoning Dynamics with Mutual Information: Thinking Tokens are Information Peaks in LLM Reasoning
Chen Qian, Dongrui Liu, Haochen Wen +3
cs.AIcs.CLarXiv:2506.02867v22025The CoT Collection: Improving Zero-shot and Few-shot Learning of Language Models via Chain-of-Thought Fine-Tuning
Seungone Kim, Se June Joo, Doyoung Kim +4
cs.CLcs.AIcs.LGarXiv:2305.14045v22023Voting or Consensus? Decision-Making in Multi-Agent Debate
Lars Benedikt Kaesberg, Jonas Becker, Jan Philip Wahle +2
cs.MAcs.AIcs.CLarXiv:2502.19130v42025Self-Edit: Fault-Aware Code Editor for Code Generation
Kechi Zhang, Zhuo Li, Jia Li +2
cs.SEcs.CLarXiv:2305.04087v52023SWE-Debate: Competitive Multi-Agent Debate for Software Issue Resolution
Han Li, Yuling Shi, Shaoxin Lin +6
cs.SEcs.CLcs.LGarXiv:2507.23348v12025DreamOn: Diffusion Language Models For Code Infilling Beyond Fixed-size Canvas
Zirui Wu, Lin Zheng, Zhihui Xie +8
cs.CLarXiv:2602.01326v12026UFT: Unifying Supervised and Reinforcement Fine-Tuning
Mingyang Liu, Gabriele Farina, Asuman Ozdaglar
cs.LGcs.CLarXiv:2505.16984v22025mPLUG-DocOwl: Modularized Multimodal Large Language Model for Document Understanding
Jiabo Ye, Anwen Hu, Haiyang Xu +10
cs.CLcs.AIarXiv:2307.02499v12023DeSTA2.5-Audio: Toward General-Purpose Large Audio Language Model with Self-Generated Cross-Modal Alignment
Ke-Han Lu, Zhehuai Chen, Szu-Wei Fu +25
eess.AScs.CLcs.SDarXiv:2507.02768v22025Direct Nash Optimization: Teaching Language Models to Self-Improve with General Preferences
Corby Rosset, Ching-An Cheng, Arindam Mitra +3
cs.LGcs.AIcs.CLarXiv:2404.03715v12024Rank-R1: Enhancing Reasoning in LLM-based Document Rerankers via Reinforcement Learning
Shengyao Zhuang, Xueguang Ma, Bevan Koopman +2
cs.IRcs.CLarXiv:2503.06034v12025MixLLM: Dynamic Routing in Mixed Large Language Models
Xinyuan Wang, Yanchi Liu, Wei Cheng +5
cs.CLcs.AIcs.DBarXiv:2502.18482v12025Defeating the Training-Inference Mismatch via FP16
Penghui Qi, Zichen Liu, Xiangxin Zhou +4
cs.LGcs.AIcs.CLarXiv:2510.26788v12025DuQuant: Distributing Outliers via Dual Transformation Makes Stronger Quantized LLMs
Haokun Lin, Haobo Xu, Yichen Wu +6
cs.CLarXiv:2406.01721v32024Unlocking Efficient Long-to-Short LLM Reasoning with Model Merging
Han Wu, Yuxuan Yao, Shuqi Liu +7
cs.CLarXiv:2503.20641v22025SimpleDeepSearcher: Deep Information Seeking via Web-Powered Reasoning Trajectory Synthesis
Shuang Sun, Huatong Song, Yuhao Wang +10
cs.CLcs.AIcs.IRarXiv:2505.16834v32025Learning When to Think: Shaping Adaptive Reasoning in R1-Style Models via Multi-Stage RL
Songjun Tu, Jiahao Lin, Qichao Zhang +4
cs.CLcs.AIarXiv:2505.10832v32025You Impress Me: Dialogue Generation via Mutual Persona Perception
Qian Liu, Yihong Chen, Bei Chen +4
cs.CLcs.AIarXiv:2004.05388v12020ReMA: Learning to Meta-think for LLMs with Multi-Agent Reinforcement Learning
Ziyu Wan, Yunxiang Li, Xiaoyu Wen +8
cs.AIcs.CLcs.LGarXiv:2503.09501v32025AutoHarness: improving LLM agents by automatically synthesizing a code harness
Xinghua Lou, Miguel Lázaro-Gredilla, Antoine Dedieu +3
cs.CLcs.AIarXiv:2603.03329v12026A Report on the Complex Word Identification Shared Task 2018
Seid Muhie Yimam, Chris Biemann, Shervin Malmasi +5
cs.CLarXiv:1804.09132v12018SWE-Exp: Experience-Driven Software Issue Resolution
Silin Chen, Shaoxin Lin, Yuling Shi +8
cs.SEcs.CLcs.LGarXiv:2507.23361v22025Logic Attention Based Neighborhood Aggregation for Inductive Knowledge Graph Embedding
Peifeng Wang, Jialong Han, Chenliang Li +1
cs.AIcs.CLarXiv:1811.01399v22018OpenCodeInstruct: A Large-scale Instruction Tuning Dataset for Code LLMs
Wasi Uddin Ahmad, Aleksander Ficek, Mehrzad Samadi +4
cs.SEcs.CLarXiv:2504.04030v22025Learning to Remember Translation History with a Continuous Cache
Zhaopeng Tu, Yang Liu, Shuming Shi +1
cs.CLarXiv:1711.09367v12017In-context Autoencoder for Context Compression in a Large Language Model
Tao Ge, Jing Hu, Lei Wang +3
cs.CLcs.AIcs.LGarXiv:2307.06945v42023Reasoning Over Semantic-Level Graph for Fact Checking
Wanjun Zhong, Jingjing Xu, Duyu Tang +5
cs.CLcs.AIarXiv:1909.03745v32019Segment Policy Optimization: Effective Segment-Level Credit Assignment in RL for Large Language Models
Yiran Guo, Lijie Xu, Jie Liu +2
cs.LGcs.AIcs.CLarXiv:2505.23564v22025Nemotron-CLIMB: CLustering-based Iterative Data Mixture Bootstrapping for Language Model Pre-training
Shizhe Diao, Yu Yang, Yonggan Fu +11
cs.CLarXiv:2504.13161v22025Safety in Large Reasoning Models: A Survey
Cheng Wang, Yue Liu, Baolong Bi +9
cs.CLarXiv:2504.17704v32025On the Interplay of Pre-Training, Mid-Training, and RL on Reasoning Language Models
Charlie Zhang, Graham Neubig, Xiang Yue
cs.CLarXiv:2512.07783v12025Do LLMs Change Their Minds Like Humans? Diagnosing Human--LLM Divergence in Single-Turn Persuasion Judgments
Lin Chen, Yitong Chen, Yong Li
cs.CYcs.CLarXiv:2608.29803v12026A Dual Reinforcement Learning Framework for Unsupervised Text Style Transfer
Fuli Luo, Peng Li, Jie Zhou +4
cs.CLarXiv:1905.10060v12019Scaling Test-Time Compute Without Verification or RL is Suboptimal
Amrith Setlur, Nived Rajaraman, Sergey Levine +1
cs.LGcs.CLarXiv:2502.12118v22025KVLink: Accelerating Large Language Models via Efficient KV Cache Reuse
Jingbo Yang, Bairu Hou, Wei Wei +2
cs.CLarXiv:2502.16002v42025VSE++: Improving Visual-Semantic Embeddings with Hard Negatives
Fartash Faghri, David J. Fleet, Jamie Ryan Kiros +1
cs.LGcs.CLcs.CVarXiv:1707.05612v42017GUI-G$^2$: Gaussian Reward Modeling for GUI Grounding
Fei Tang, Zhangxuan Gu, Zhengxi Lu +9
cs.LGcs.AIcs.CLarXiv:2507.15846v32025MiniCPM4: Ultra-Efficient LLMs on End Devices
MiniCPM Team, Chaojun Xiao, Yuxuan Li +80
cs.CLcs.AIarXiv:2506.07900v22025Classical Structured Prediction Losses for Sequence to Sequence Learning
Sergey Edunov, Myle Ott, Michael Auli +2
cs.CLarXiv:1711.04956v52017Prompt for Extraction? PAIE: Prompting Argument Interaction for Event Argument Extraction
Yubo Ma, Zehao Wang, Yixin Cao +4
cs.CLcs.AIarXiv:2202.12109v22022CoT-Kinetics: A Theoretical Modeling Assessing LRM Reasoning Process
Jinhe Bi, Danqi Yan, Yifan Wang +8
cs.AIcs.CLarXiv:2505.13408v12025Jointly Learning Entity and Relation Representations for Entity Alignment
Yuting Wu, Xiao Liu, Yansong Feng +2
cs.CLarXiv:1909.09317v12019RePro: Proof-Verified Benchmark Rewriting for Reliable Evaluation of LLM Mathematical Problem Solving
Xiyuan Zhou, Zhuoqi Li, Xinlei Wang +6
cs.CLcs.AIarXiv:2609.00062v12026Emergent autonomous scientific research capabilities of large language models
Daniil A. Boiko, Robert MacKnight, Gabe Gomes
physics.chem-phcs.CLarXiv:2304.05332v12023BEST-Route: Adaptive LLM Routing with Test-Time Optimal Compute
Dujian Ding, Ankur Mallick, Shaokun Zhang +7
cs.LGcs.AIcs.CLarXiv:2506.22716v12025START: Self-taught Reasoner with Tools
Chengpeng Li, Mingfeng Xue, Zhenru Zhang +7
cs.CLarXiv:2503.04625v22025