Computation and Language

Papers filed under cs.CL on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

2,881 to 2,940 of 11,247

  1. Multi-domain Dialog State Tracking using Recurrent Neural Networks

    Nikola Mrkšić, Diarmuid Ó Séaghdha, Blaise Thomson +5

    cs.CLcs.LGarXiv:1506.07190v12015
  2. IHEval: Evaluating Language Models on Following the Instruction Hierarchy

    Zhihan Zhang, Shiyang Li, Zixuan Zhang +11

    cs.CLarXiv:2502.08745v22025
  3. A Survey of LLM-based Deep Search Agents: Paradigm, Optimization, Evaluation, and Challenges

    Yunjia Xi, Jianghao Lin, Yongzhao Xiao +7

    cs.IRcs.AIcs.CLarXiv:2508.05668v32025
  4. Stabilizing MoE Reinforcement Learning by Aligning Training and Inference Routers

    Wenhan Ma, Hailin Zhang, Liang Zhao +4

    cs.CLcs.AIcs.LGarXiv:2510.11370v22025
  5. Sheared LLaMA: Accelerating Language Model Pre-training via Structured Pruning

    Mengzhou Xia, Tianyu Gao, Zhiyuan Zeng +1

    cs.CLcs.AIcs.LGarXiv:2310.06694v22023
  6. Human-AI Collaboration Enables More Empathic Conversations in Text-based Peer-to-Peer Mental Health Support

    Ashish Sharma, Inna W. Lin, Adam S. Miner +2

    cs.CLcs.HCcs.SIarXiv:2203.15144v12022
  7. Release Strategies and the Social Impacts of Language Models

    Irene Solaiman, Miles Brundage, Jack Clark +12

    cs.CLcs.AIcs.CYarXiv:1908.09203v22019
  8. R1-Reward: Training Multimodal Reward Model Through Stable Reinforcement Learning

    Yi-Fan Zhang, Xingyu Lu, Xiao Hu +13

    cs.CVcs.CLarXiv:2505.02835v22025
  9. Large Language Models and Games: A Survey and Roadmap

    Roberto Gallotta, Graham Todd, Marvin Zammit +4

    cs.CLcs.AIcs.HCarXiv:2402.18659v52024
  10. Speech2Vec: A Sequence-to-Sequence Framework for Learning Word Embeddings from Speech

    Yu-An Chung, James Glass

    cs.CLarXiv:1803.08976v22018
  11. InfiGUIAgent: A Multimodal Generalist GUI Agent with Native Reasoning and Reflection

    Yuhang Liu, Pengxiang Li, Zishu Wei +7

    cs.AIcs.CLcs.HCarXiv:2501.04575v12025
  12. Beyond BLEU: Training Neural Machine Translation with Semantic Similarity

    John Wieting, Taylor Berg-Kirkpatrick, Kevin Gimpel +1

    cs.CLarXiv:1909.06694v12019
  13. EvoFlint: An Evolutionary Atlas of Multi-Turn LLM Vulnerabilities

    Feitong Qiao, Liren Peng, Shiming Ren +7

    cs.CLcs.AIcs.CRarXiv:2609.00487v12026
  14. SCROLLS: Standardized CompaRison Over Long Language Sequences

    Uri Shaham, Elad Segal, Maor Ivgi +8

    cs.CLcs.AIcs.LGarXiv:2201.03533v22022
  15. Demystifying Reasoning Dynamics with Mutual Information: Thinking Tokens are Information Peaks in LLM Reasoning

    Chen Qian, Dongrui Liu, Haochen Wen +3

    cs.AIcs.CLarXiv:2506.02867v22025
  16. The CoT Collection: Improving Zero-shot and Few-shot Learning of Language Models via Chain-of-Thought Fine-Tuning

    Seungone Kim, Se June Joo, Doyoung Kim +4

    cs.CLcs.AIcs.LGarXiv:2305.14045v22023
  17. Voting or Consensus? Decision-Making in Multi-Agent Debate

    Lars Benedikt Kaesberg, Jonas Becker, Jan Philip Wahle +2

    cs.MAcs.AIcs.CLarXiv:2502.19130v42025
  18. Self-Edit: Fault-Aware Code Editor for Code Generation

    Kechi Zhang, Zhuo Li, Jia Li +2

    cs.SEcs.CLarXiv:2305.04087v52023
  19. SWE-Debate: Competitive Multi-Agent Debate for Software Issue Resolution

    Han Li, Yuling Shi, Shaoxin Lin +6

    cs.SEcs.CLcs.LGarXiv:2507.23348v12025
  20. DreamOn: Diffusion Language Models For Code Infilling Beyond Fixed-size Canvas

    Zirui Wu, Lin Zheng, Zhihui Xie +8

    cs.CLarXiv:2602.01326v12026
  21. UFT: Unifying Supervised and Reinforcement Fine-Tuning

    Mingyang Liu, Gabriele Farina, Asuman Ozdaglar

    cs.LGcs.CLarXiv:2505.16984v22025
  22. mPLUG-DocOwl: Modularized Multimodal Large Language Model for Document Understanding

    Jiabo Ye, Anwen Hu, Haiyang Xu +10

    cs.CLcs.AIarXiv:2307.02499v12023
  23. DeSTA2.5-Audio: Toward General-Purpose Large Audio Language Model with Self-Generated Cross-Modal Alignment

    Ke-Han Lu, Zhehuai Chen, Szu-Wei Fu +25

    eess.AScs.CLcs.SDarXiv:2507.02768v22025
  24. Direct Nash Optimization: Teaching Language Models to Self-Improve with General Preferences

    Corby Rosset, Ching-An Cheng, Arindam Mitra +3

    cs.LGcs.AIcs.CLarXiv:2404.03715v12024
  25. Rank-R1: Enhancing Reasoning in LLM-based Document Rerankers via Reinforcement Learning

    Shengyao Zhuang, Xueguang Ma, Bevan Koopman +2

    cs.IRcs.CLarXiv:2503.06034v12025
  26. MixLLM: Dynamic Routing in Mixed Large Language Models

    Xinyuan Wang, Yanchi Liu, Wei Cheng +5

    cs.CLcs.AIcs.DBarXiv:2502.18482v12025
  27. Defeating the Training-Inference Mismatch via FP16

    Penghui Qi, Zichen Liu, Xiangxin Zhou +4

    cs.LGcs.AIcs.CLarXiv:2510.26788v12025
  28. DuQuant: Distributing Outliers via Dual Transformation Makes Stronger Quantized LLMs

    Haokun Lin, Haobo Xu, Yichen Wu +6

    cs.CLarXiv:2406.01721v32024
  29. Unlocking Efficient Long-to-Short LLM Reasoning with Model Merging

    Han Wu, Yuxuan Yao, Shuqi Liu +7

    cs.CLarXiv:2503.20641v22025
  30. SimpleDeepSearcher: Deep Information Seeking via Web-Powered Reasoning Trajectory Synthesis

    Shuang Sun, Huatong Song, Yuhao Wang +10

    cs.CLcs.AIcs.IRarXiv:2505.16834v32025
  31. Learning When to Think: Shaping Adaptive Reasoning in R1-Style Models via Multi-Stage RL

    Songjun Tu, Jiahao Lin, Qichao Zhang +4

    cs.CLcs.AIarXiv:2505.10832v32025
  32. You Impress Me: Dialogue Generation via Mutual Persona Perception

    Qian Liu, Yihong Chen, Bei Chen +4

    cs.CLcs.AIarXiv:2004.05388v12020
  33. ReMA: Learning to Meta-think for LLMs with Multi-Agent Reinforcement Learning

    Ziyu Wan, Yunxiang Li, Xiaoyu Wen +8

    cs.AIcs.CLcs.LGarXiv:2503.09501v32025
  34. AutoHarness: improving LLM agents by automatically synthesizing a code harness

    Xinghua Lou, Miguel Lázaro-Gredilla, Antoine Dedieu +3

    cs.CLcs.AIarXiv:2603.03329v12026
  35. A Report on the Complex Word Identification Shared Task 2018

    Seid Muhie Yimam, Chris Biemann, Shervin Malmasi +5

    cs.CLarXiv:1804.09132v12018
  36. SWE-Exp: Experience-Driven Software Issue Resolution

    Silin Chen, Shaoxin Lin, Yuling Shi +8

    cs.SEcs.CLcs.LGarXiv:2507.23361v22025
  37. Logic Attention Based Neighborhood Aggregation for Inductive Knowledge Graph Embedding

    Peifeng Wang, Jialong Han, Chenliang Li +1

    cs.AIcs.CLarXiv:1811.01399v22018
  38. OpenCodeInstruct: A Large-scale Instruction Tuning Dataset for Code LLMs

    Wasi Uddin Ahmad, Aleksander Ficek, Mehrzad Samadi +4

    cs.SEcs.CLarXiv:2504.04030v22025
  39. Learning to Remember Translation History with a Continuous Cache

    Zhaopeng Tu, Yang Liu, Shuming Shi +1

    cs.CLarXiv:1711.09367v12017
  40. In-context Autoencoder for Context Compression in a Large Language Model

    Tao Ge, Jing Hu, Lei Wang +3

    cs.CLcs.AIcs.LGarXiv:2307.06945v42023
  41. Reasoning Over Semantic-Level Graph for Fact Checking

    Wanjun Zhong, Jingjing Xu, Duyu Tang +5

    cs.CLcs.AIarXiv:1909.03745v32019
  42. Segment Policy Optimization: Effective Segment-Level Credit Assignment in RL for Large Language Models

    Yiran Guo, Lijie Xu, Jie Liu +2

    cs.LGcs.AIcs.CLarXiv:2505.23564v22025
  43. Nemotron-CLIMB: CLustering-based Iterative Data Mixture Bootstrapping for Language Model Pre-training

    Shizhe Diao, Yu Yang, Yonggan Fu +11

    cs.CLarXiv:2504.13161v22025
  44. Safety in Large Reasoning Models: A Survey

    Cheng Wang, Yue Liu, Baolong Bi +9

    cs.CLarXiv:2504.17704v32025
  45. On the Interplay of Pre-Training, Mid-Training, and RL on Reasoning Language Models

    Charlie Zhang, Graham Neubig, Xiang Yue

    cs.CLarXiv:2512.07783v12025
  46. Do LLMs Change Their Minds Like Humans? Diagnosing Human--LLM Divergence in Single-Turn Persuasion Judgments

    Lin Chen, Yitong Chen, Yong Li

    cs.CYcs.CLarXiv:2608.29803v12026
  47. A Dual Reinforcement Learning Framework for Unsupervised Text Style Transfer

    Fuli Luo, Peng Li, Jie Zhou +4

    cs.CLarXiv:1905.10060v12019
  48. Scaling Test-Time Compute Without Verification or RL is Suboptimal

    Amrith Setlur, Nived Rajaraman, Sergey Levine +1

    cs.LGcs.CLarXiv:2502.12118v22025
  49. KVLink: Accelerating Large Language Models via Efficient KV Cache Reuse

    Jingbo Yang, Bairu Hou, Wei Wei +2

    cs.CLarXiv:2502.16002v42025
  50. VSE++: Improving Visual-Semantic Embeddings with Hard Negatives

    Fartash Faghri, David J. Fleet, Jamie Ryan Kiros +1

    cs.LGcs.CLcs.CVarXiv:1707.05612v42017
  51. GUI-G$^2$: Gaussian Reward Modeling for GUI Grounding

    Fei Tang, Zhangxuan Gu, Zhengxi Lu +9

    cs.LGcs.AIcs.CLarXiv:2507.15846v32025
  52. MiniCPM4: Ultra-Efficient LLMs on End Devices

    MiniCPM Team, Chaojun Xiao, Yuxuan Li +80

    cs.CLcs.AIarXiv:2506.07900v22025
  53. Classical Structured Prediction Losses for Sequence to Sequence Learning

    Sergey Edunov, Myle Ott, Michael Auli +2

    cs.CLarXiv:1711.04956v52017
  54. Prompt for Extraction? PAIE: Prompting Argument Interaction for Event Argument Extraction

    Yubo Ma, Zehao Wang, Yixin Cao +4

    cs.CLcs.AIarXiv:2202.12109v22022
  55. CoT-Kinetics: A Theoretical Modeling Assessing LRM Reasoning Process

    Jinhe Bi, Danqi Yan, Yifan Wang +8

    cs.AIcs.CLarXiv:2505.13408v12025
  56. Jointly Learning Entity and Relation Representations for Entity Alignment

    Yuting Wu, Xiao Liu, Yansong Feng +2

    cs.CLarXiv:1909.09317v12019
  57. RePro: Proof-Verified Benchmark Rewriting for Reliable Evaluation of LLM Mathematical Problem Solving

    Xiyuan Zhou, Zhuoqi Li, Xinlei Wang +6

    cs.CLcs.AIarXiv:2609.00062v12026
  58. Emergent autonomous scientific research capabilities of large language models

    Daniil A. Boiko, Robert MacKnight, Gabe Gomes

    physics.chem-phcs.CLarXiv:2304.05332v12023
  59. BEST-Route: Adaptive LLM Routing with Test-Time Optimal Compute

    Dujian Ding, Ankur Mallick, Shaokun Zhang +7

    cs.LGcs.AIcs.CLarXiv:2506.22716v12025
  60. START: Self-taught Reasoner with Tools

    Chengpeng Li, Mingfeng Xue, Zhenru Zhang +7

    cs.CLarXiv:2503.04625v22025