Computation and Language

Papers filed under cs.CL on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

9,241 to 9,300 of 11,310

  1. A C-LSTM Neural Network for Text Classification

    Chunting Zhou, Chonglin Sun, Zhiyuan Liu +1

    cs.CLarXiv:1511.08630v22015
  2. Lie to Me: How Faithful Is Chain-of-Thought Reasoning in Reasoning Models?

    Richard J. Young

    cs.CLcs.AIarXiv:2603.22582v12026
  3. Outcome Accuracy is Not Enough: Aligning the Reasoning Process of Reward Models

    Binghai Wang, Yantao Liu, Yuxuan Liu +13

    cs.CLarXiv:2602.04649v12026
  4. Grammar as a Foreign Language

    Oriol Vinyals, Lukasz Kaiser, Terry Koo +3

    cs.CLcs.LGstat.MLarXiv:1412.7449v32014
  5. K-BERT: Enabling Language Representation with Knowledge Graph

    Weijie Liu, Peng Zhou, Zhe Zhao +4

    cs.CLcs.LGarXiv:1909.07606v12019
  6. Can LLMs Clean Up Your Mess? A Survey of Application-Ready Data Preparation with LLMs

    Wei Zhou, Jun Zhou, Haoyu Wang +16

    cs.DBcs.AIcs.CLarXiv:2601.17058v12026
  7. From Directions to Regions: Decomposing Activations in Language Models via Local Geometry

    Or Shafran, Shaked Ronen, Omri Fahn +3

    cs.CLarXiv:2602.02464v12026
  8. LatentMem: Customizing Latent Memory for Multi-Agent Systems

    Muxin Fu, Xiangyuan Xue, Yafu Li +5

    cs.CLcs.LGcs.MAarXiv:2602.03036v22026
  9. ResAdapt: Adaptive Resolution for Efficient Multimodal Reasoning

    Huanxuan Liao, Zhongtao Jiang, Yupu Hao +6

    cs.CVcs.AIcs.CLarXiv:2603.28610v22026
  10. Learned in Translation: Contextualized Word Vectors

    Bryan McCann, James Bradbury, Caiming Xiong +1

    cs.CLcs.AIcs.LGarXiv:1708.00107v22017
  11. DSPy: Compiling Declarative Language Model Calls into Self-Improving Pipelines

    Omar Khattab, Arnav Singhvi, Paridhi Maheshwari +10

    cs.CLcs.AIcs.IRarXiv:2310.03714v12023
  12. Meta-Reinforcement Learning with Self-Reflection for Agentic Search

    Teng Xiao, Yige Yuan, Hamish Ivison +6

    cs.LGcs.CLarXiv:2603.11327v22026
  13. CreativeBench: Benchmarking and Enhancing Machine Creativity via Self-Evolving Challenges

    Zi-Han Wang, Lam Nguyen, Zhengyang Zhao +4

    cs.AIcs.CLarXiv:2603.11863v22026
  14. Towards Faithfully Interpretable NLP Systems: How should we define and evaluate faithfulness?

    Alon Jacovi, Yoav Goldberg

    cs.CLcs.LGarXiv:2004.03685v32020
  15. Controlled Self-Evolution for Algorithmic Code Optimization

    Tu Hu, Ronghao Chen, Shuo Zhang +9

    cs.CLcs.AIcs.NEarXiv:2601.07348v52026
  16. When Actions Go Off-Task: Detecting and Correcting Misaligned Actions in Computer-Use Agents

    Yuting Ning, Jaylen Jones, Zhehao Zhang +5

    cs.CLarXiv:2602.08995v22026
  17. Chameleon: Mixed-Modal Early-Fusion Foundation Models

    Chameleon Team

    cs.CLarXiv:2405.09818v22024
  18. Towards Reasoning in Large Language Models: A Survey

    Jie Huang, Kevin Chen-Chuan Chang

    cs.CLcs.AIarXiv:2212.10403v22022
  19. Ref-Adv: Exploring MLLM Visual Reasoning in Referring Expression Tasks

    Qihua Dong, Kuo Yang, Lin Ju +6

    cs.CVcs.AIcs.CLarXiv:2602.23898v12026
  20. When and How Much to Imagine: Adaptive Test-Time Scaling with World Models for Visual Spatial Reasoning

    Shoubin Yu, Yue Zhang, Zun Wang +4

    cs.CVcs.AIcs.CLarXiv:2602.08236v22026
  21. MMOU: A Massive Multi-Task Omni Understanding and Reasoning Benchmark for Long and Complex Real-World Videos

    Arushi Goel, Sreyan Ghosh, Vatsal Agarwal +16

    cs.CLcs.CVarXiv:2603.14145v22026
  22. Language-driven Semantic Segmentation

    Boyi Li, Kilian Q. Weinberger, Serge Belongie +2

    cs.CVcs.CLcs.LGarXiv:2201.03546v22022
  23. Fanar-Sadiq: A Multi-Agent Architecture for Grounded Islamic QA

    Ummar Abbas, Mourad Ouzzani, Mohamed Y. Eltabakh +7

    cs.CLarXiv:2603.08501v32026
  24. Automatic Detection of Fake News

    Verónica Pérez-Rosas, Bennett Kleinberg, Alexandra Lefevre +1

    cs.CLarXiv:1708.07104v12017
  25. Decoupling Reasoning and Confidence: Resurrecting Calibration in Reinforcement Learning from Verifiable Rewards

    Zhengzhao Ma, Xueru Wen, Boxi Cao +6

    cs.LGcs.AIcs.CLarXiv:2603.09117v32026
  26. DSDR: Dual-Scale Diversity Regularization for Exploration in LLM Reasoning

    Zhongwei Wan, Yun Shen, Zhihao Dou +9

    cs.LGcs.CLarXiv:2602.19895v12026
  27. OpenHands: An Open Platform for AI Software Developers as Generalist Agents

    Xingyao Wang, Boxuan Li, Yufan Song +21

    cs.SEcs.AIcs.CLarXiv:2407.16741v32024
  28. Spurious Rewards Paradox: Mechanistically Understanding How RLVR Activates Memorization Shortcuts in LLMs

    Lecheng Yan, Ruizhe Li, Guanhua Chen +5

    cs.LGcs.CLarXiv:2601.11061v22026
  29. dVoting: Fast Voting for dLLMs

    Sicheng Feng, Zigeng Chen, Xinyin Ma +2

    cs.CLcs.AIarXiv:2602.12153v12026
  30. NewsQA: A Machine Comprehension Dataset

    Adam Trischler, Tong Wang, Xingdi Yuan +4

    cs.CLcs.AIarXiv:1611.09830v32016
  31. ThinkJEPA: Empowering Latent World Models with Large Vision-Language Reasoning Model

    Haichao Zhang, Yijiang Li, Shwai He +5

    cs.CVcs.AIcs.CLarXiv:2603.22281v22026
  32. LongCat-Flash-Prover: Advancing Native Formal Reasoning via Agentic Tool-Integrated Reinforcement Learning

    Jianing Wang, Jianfei Zhang, Qi Guo +24

    cs.AIcs.CLarXiv:2603.21065v12026
  33. The Confidence Dichotomy: Analyzing and Mitigating Miscalibration in Tool-Use Agents

    Weihao Xuan, Qingcheng Zeng, Heli Qi +3

    cs.CLarXiv:2601.07264v12026
  34. PROGRESSLM: Towards Progress Reasoning in Vision-Language Models

    Jianshu Zhang, Chengxuan Qian, Haosen Sun +4

    cs.CVcs.CLarXiv:2601.15224v22026
  35. A Simple and Effective Pruning Approach for Large Language Models

    Mingjie Sun, Zhuang Liu, Anna Bair +1

    cs.CLcs.AIcs.LGarXiv:2306.11695v32023
  36. Unified Pre-training for Program Understanding and Generation

    Wasi Uddin Ahmad, Saikat Chakraborty, Baishakhi Ray +1

    cs.CLcs.PLarXiv:2103.06333v22021
  37. Ragas: Automated Evaluation of Retrieval Augmented Generation

    Shahul Es, Jithin James, Luis Espinosa-Anke +1

    cs.CLarXiv:2309.15217v22023
  38. Multi-Vector Index Compression in Any Modality

    Hanxiang Qin, Alexander Martin, Rohan Jha +3

    cs.IRcs.CLcs.CVarXiv:2602.21202v12026
  39. RealMem: Benchmarking LLMs in Real-World Memory-Driven Interaction

    Haonan Bian, Zhiyuan Yao, Sen Hu +7

    cs.CLcs.AIarXiv:2601.06966v12026
  40. Open Problems and Fundamental Limitations of Reinforcement Learning from Human Feedback

    Stephen Casper, Xander Davies, Claudia Shi +29

    cs.AIcs.CLcs.LGarXiv:2307.15217v22023
  41. Calibrate-Then-Act: Cost-Aware Exploration in LLM Agents

    Wenxuan Ding, Nicholas Tomlin, Greg Durrett

    cs.CLcs.AIarXiv:2602.16699v32026
  42. Deduplicating Training Data Makes Language Models Better

    Katherine Lee, Daphne Ippolito, Andrew Nystrom +4

    cs.CLcs.LGarXiv:2107.06499v22021
  43. Training Language Models via Neural Cellular Automata

    Dan Lee, Seungwook Han, Akarsh Kumar +1

    cs.LGcs.AIcs.CLarXiv:2603.10055v12026
  44. e5-omni: Explicit Cross-modal Alignment for Omni-modal Embeddings

    Haonan Chen, Sicheng Gao, Radu Timofte +2

    cs.CLcs.AIcs.CVarXiv:2601.03666v22026
  45. The Flan Collection: Designing Data and Methods for Effective Instruction Tuning

    Shayne Longpre, Le Hou, Tu Vu +8

    cs.AIcs.CLcs.LGarXiv:2301.13688v22023
  46. Enhancing Chat Language Models by Scaling High-quality Instructional Conversations

    Ning Ding, Yulin Chen, Bokai Xu +6

    cs.CLcs.AIarXiv:2305.14233v12023
  47. Kernel-Smith: A Unified Recipe for Evolutionary Kernel Optimization

    He Du, Qiming Ge, Jiakai Hu +18

    cs.CLcs.LGarXiv:2603.28342v22026
  48. Quantifying Memorization Across Neural Language Models

    Nicholas Carlini, Daphne Ippolito, Matthew Jagielski +3

    cs.LGcs.CLarXiv:2202.07646v32022
  49. Multi-Task GRPO: Reliable LLM Reasoning Across Tasks

    Shyam Sundhar Ramesh, Xiaotong Ji, Matthieu Zimmer +5

    cs.CLcs.AIcs.LGarXiv:2602.05547v32026
  50. Permutation Invariant Training of Deep Models for Speaker-Independent Multi-talker Speech Separation

    Dong Yu, Morten Kolbæk, Zheng-Hua Tan +1

    cs.CLcs.LGcs.SDarXiv:1607.00325v22016
  51. Learning To Retrieve Prompts for In-Context Learning

    Ohad Rubin, Jonathan Herzig, Jonathan Berant

    cs.CLcs.LGarXiv:2112.08633v22021
  52. Refusal in Language Models Is Mediated by a Single Direction

    Andy Arditi, Oscar Obeso, Aaquib Syed +4

    cs.LGcs.AIcs.CLarXiv:2406.11717v32024
  53. Beyond Static Tools: Test-Time Tool Evolution for Scientific Reasoning

    Jiaxuan Lu, Ziyu Kong, Yemin Wang +10

    cs.AIcs.CLcs.MAarXiv:2601.07641v12026
  54. Intrinsic Dimensionality Explains the Effectiveness of Language Model Fine-Tuning

    Armen Aghajanyan, Luke Zettlemoyer, Sonal Gupta

    cs.LGcs.CLarXiv:2012.13255v12020
  55. A Discourse-Aware Attention Model for Abstractive Summarization of Long Documents

    Arman Cohan, Franck Dernoncourt, Doo Soon Kim +4

    cs.CLarXiv:1804.05685v22018
  56. ChatEval: Towards Better LLM-based Evaluators through Multi-Agent Debate

    Chi-Min Chan, Weize Chen, Yusheng Su +5

    cs.CLarXiv:2308.07201v12023
  57. Lost in Stories: Consistency Bugs in Long Story Generation by LLMs

    Junjie Li, Xinrui Guo, Yuhao Wu +3

    cs.CLcs.AIarXiv:2603.05890v12026
  58. Transfer Learning from Speaker Verification to Multispeaker Text-To-Speech Synthesis

    Ye Jia, Yu Zhang, Ron J. Weiss +8

    cs.CLcs.LGcs.SDarXiv:1806.04558v42018
  59. Rewarding the Rare: Uniqueness-Aware RL for Creative Problem Solving in LLMs

    Zhiyuan Hu, Yucheng Wang, Yufei He +7

    cs.LGcs.CLarXiv:2601.08763v22026
  60. SimVLM: Simple Visual Language Model Pretraining with Weak Supervision

    Zirui Wang, Jiahui Yu, Adams Wei Yu +3

    cs.CVcs.CLcs.LGarXiv:2108.10904v32021