Computation and Language

Papers filed under cs.CL on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

2,821 to 2,880 of 11,247

  1. MMBench-GUI: Hierarchical Multi-Platform Evaluation Framework for GUI Agents

    Xuehui Wang, Zhenyu Wu, JingJing Xie +25

    cs.CVcs.CLarXiv:2507.19478v12025
  2. OS-ATLAS: A Foundation Action Model for Generalist GUI Agents

    Zhiyong Wu, Zhenyu Wu, Fangzhi Xu +8

    cs.CLcs.CVcs.HCarXiv:2410.23218v12024
  3. Look Inward to Explore Outward: Learning Temperature Policy from LLM Internal States via Hierarchical RL

    Yixiao Zhou, Yang Li, Dongzhou Cheng +2

    cs.LGcs.AIcs.CLarXiv:2602.13035v12026
  4. Seed-Coder: Let the Code Model Curate Data for Itself

    ByteDance Seed, Yuyu Zhang, Jing Su +24

    cs.CLcs.SEarXiv:2506.03524v22025
  5. AI chatbots versus human healthcare professionals: a systematic review and meta-analysis of empathy in patient care

    Alastair Howcroft, Amber Bennett-Weston, Ahmad Khan +3

    cs.HCcs.AIcs.CLarXiv:2602.05628v12026
  6. CodeI/O: Condensing Reasoning Patterns via Code Input-Output Prediction

    Junlong Li, Daya Guo, Dejian Yang +3

    cs.CLcs.AIarXiv:2502.07316v42025
  7. OmniVoice: Towards Omnilingual Zero-Shot Text-to-Speech with Diffusion Language Models

    Han Zhu, Lingxuan Ye, Wei Kang +7

    cs.CLeess.ASarXiv:2604.00688v32026
  8. Understanding and Improving Information Transfer in Multi-Task Learning

    Sen Wu, Hongyang R. Zhang, Christopher Ré

    cs.LGcs.CLarXiv:2005.00944v12020
  9. Large Language Model Routing with Benchmark Datasets

    Tal Shnitzer, Anthony Ou, Mírian Silva +5

    cs.CLcs.LGarXiv:2309.15789v12023
  10. Latent On-Policy Self-Distillation

    Guibin Zhang, Jiayang Lyu, Ran Sun +4

    cs.LGcs.CLarXiv:2608.13040v12026
  11. Can Sensitive Information Be Deleted From LLMs? Objectives for Defending Against Extraction Attacks

    Vaidehi Patil, Peter Hase, Mohit Bansal

    cs.CLcs.AIcs.LGarXiv:2309.17410v12023
  12. Covo-Audio Technical Report

    Wenfu Wang, Chenxing Li, Liqiang Zhang +23

    cs.SDcs.CLeess.ASarXiv:2602.09823v22026
  13. Specializing Large Language Models to Simulate Survey Response Distributions for Global Populations

    Yong Cao, Haijiang Liu, Arnav Arora +3

    cs.CLarXiv:2502.07068v22025
  14. A Systematic Literature Review of Retrieval-Augmented Generation: Techniques, Metrics, and Challenges

    Andrew Brown, Muhammad Roman, Barry Devereux

    cs.DLcs.AIcs.CLarXiv:2508.06401v32025
  15. Eagle and Finch: RWKV with Matrix-Valued States and Dynamic Recurrence

    Bo Peng, Daniel Goldstein, Quentin Anthony +27

    cs.CLcs.AIarXiv:2404.05892v42024
  16. Agent S: An Open Agentic Framework that Uses Computers Like a Human

    Saaket Agashe, Jiuzhou Han, Shuyu Gan +3

    cs.AIcs.CLcs.CVarXiv:2410.08164v12024
  17. Can LLM feedback enhance review quality? A randomized study of 20K reviews at ICLR 2025

    Nitya Thakkar, Mert Yuksekgonul, Jake Silberg +6

    cs.AIcs.CLcs.HCarXiv:2504.09737v12025
  18. Command A: An Enterprise-Ready Large Language Model

    Team Cohere, :, Aakanksha +227

    cs.CLcs.AIcs.LGarXiv:2504.00698v22025
  19. DistiLLM-2: A Contrastive Approach Boosts the Distillation of LLMs

    Jongwoo Ko, Tianyi Chen, Sungnyun Kim +4

    cs.CLcs.AIcs.LGarXiv:2503.07067v22025
  20. Offensive Language and Hate Speech Detection for Danish

    Gudbjartur Ingi Sigurbergsson, Leon Derczynski

    cs.CLarXiv:1908.04531v22019
  21. Sequence-to-Sequence Generation for Spoken Dialogue via Deep Syntax Trees and Strings

    Ondřej Dušek, Filip Jurčíček

    cs.CLarXiv:1606.05491v12016
  22. LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming?

    Zihan Zheng, Zerui Cheng, Zeyu Shen +16

    cs.SEcs.AIcs.CLarXiv:2506.11928v12025
  23. Improving Data Efficiency for LLM Reinforcement Fine-tuning Through Difficulty-targeted Online Data Selection and Rollout Replay

    Yifan Sun, Jingyan Shen, Yibin Wang +4

    cs.LGcs.AIcs.CLarXiv:2506.05316v42025
  24. Tactical Rewind: Self-Correction via Backtracking in Vision-and-Language Navigation

    Liyiming Ke, Xiujun Li, Yonatan Bisk +6

    cs.CLcs.CVcs.LGarXiv:1903.02547v22019
  25. What In-Context Learning "Learns" In-Context: Disentangling Task Recognition and Task Learning

    Jane Pan, Tianyu Gao, Howard Chen +1

    cs.CLcs.LGarXiv:2305.09731v12023
  26. Multilingual Topic Models for Unaligned Text

    Jordan Boyd-Graber, David Blei

    cs.CLcs.IRcs.LGarXiv:1205.2657v12012
  27. How Mental Health Self-Disclosure Becomes Visible: Evidence from Eight Conditions on Reddit

    Renkai Ma, Lingyao Li, Shanting Chen +3

    cs.HCcs.CLcs.CYarXiv:2608.29010v12026
  28. Speculative Thinking: Enhancing Small-Model Reasoning with Large Model Guidance at Inference Time

    Wang Yang, Xiang Yue, Vipin Chaudhary +1

    cs.CLcs.AIarXiv:2504.12329v22025
  29. Show, Describe and Conclude: On Exploiting the Structure Information of Chest X-Ray Reports

    Baoyu Jing, Zeya Wang, Eric Xing

    cs.CLcs.CVeess.IVarXiv:2004.12274v22020
  30. AgentRxiv: Towards Collaborative Autonomous Research

    Samuel Schmidgall, Michael Moor

    cs.AIcs.CLcs.LGarXiv:2503.18102v12025
  31. Multi-SpatialMLLM: Multi-Frame Spatial Understanding with Multi-Modal Large Language Models

    Runsen Xu, Weiyao Wang, Hao Tang +5

    cs.CVcs.CLarXiv:2505.17015v22025
  32. Avoiding Latent Variable Collapse With Generative Skip Models

    Adji B. Dieng, Yoon Kim, Alexander M. Rush +1

    stat.MLcs.CLcs.LGarXiv:1807.04863v22018
  33. InternLM-XComposer2-4KHD: A Pioneering Large Vision-Language Model Handling Resolutions from 336 Pixels to 4K HD

    Xiaoyi Dong, Pan Zhang, Yuhang Zang +21

    cs.CVcs.CLarXiv:2404.06512v12024
  34. Analyzing Polarization in Social Media: Method and Application to Tweets on 21 Mass Shootings

    Dorottya Demszky, Nikhil Garg, Rob Voigt +4

    cs.CLarXiv:1904.01596v22019
  35. ConCISE: Confidence-guided Compression in Step-by-step Efficient Reasoning

    Ziqing Qiao, Yongheng Deng, Jiali Zeng +7

    cs.LGcs.AIcs.CLarXiv:2505.04881v22025
  36. A Survey of Uncertainty Estimation Methods on Large Language Models

    Zhiqiu Xia, Jinxuan Xu, Yuqian Zhang +1

    cs.CLarXiv:2503.00172v22025
  37. EmoVoice: LLM-based Emotional Text-To-Speech Model with Freestyle Text Prompting

    Guanrou Yang, Chen Yang, Qian Chen +12

    eess.AScs.AIcs.CLarXiv:2504.12867v42025
  38. STransE: a novel embedding model of entities and relationships in knowledge bases

    Dat Quoc Nguyen, Kairit Sirts, Lizhen Qu +1

    cs.CLcs.AIarXiv:1606.08140v32016
  39. Critique Fine-Tuning: Learning to Critique is More Effective than Learning to Imitate

    Yubo Wang, Xiang Yue, Wenhu Chen

    cs.CLarXiv:2501.17703v42025
  40. Effectively Controlling Reasoning Models through Thinking Intervention

    Tong Wu, Chong Xiang, Jiachen T. Wang +2

    cs.LGcs.AIcs.CLarXiv:2503.24370v32025
  41. Reasoning Theater: Disentangling Model Beliefs from Chain-of-Thought

    Siddharth Boppana, Annabel Ma, Max Loeffler +5

    cs.CLcs.AIcs.LGarXiv:2603.05488v42026
  42. Comparative Analysis Based on DeepSeek, ChatGPT, and Google Gemini: Features, Techniques, Performance, Future Prospects

    Anichur Rahman, Shahariar Hossain Mahir, Md Tanjum An Tashrif +6

    cs.CLcs.CRarXiv:2503.04783v12025
  43. GPT Models in Construction Industry: Opportunities, Limitations, and a Use Case Validation

    Abdullahi Saka, Ridwan Taiwo, Nurudeen Saka +4

    cs.HCcs.AIcs.CLarXiv:2305.18997v12023
  44. The Homogenizing Effect of Large Language Models on Human Expression and Thought

    Zhivar Sourati, Alireza S. Ziabari, Morteza Dehghani

    cs.CLarXiv:2508.01491v22025
  45. LayoutXLM: Multimodal Pre-training for Multilingual Visually-rich Document Understanding

    Yiheng Xu, Tengchao Lv, Lei Cui +5

    cs.CLarXiv:2104.08836v32021
  46. Ross3D: Reconstructive Visual Instruction Tuning with 3D-Awareness

    Haochen Wang, Yucheng Zhao, Tiancai Wang +3

    cs.CVcs.AIcs.CLarXiv:2504.01901v12025
  47. Compact Language Models via Pruning and Knowledge Distillation

    Saurav Muralidharan, Sharath Turuvekere Sreenivas, Raviraj Joshi +6

    cs.CLcs.AIcs.LGarXiv:2407.14679v22024
  48. DeltaProduct: Improving State-Tracking in Linear RNNs via Householder Products

    Julien Siems, Timur Carstensen, Arber Zela +3

    cs.LGcs.CLcs.FLarXiv:2502.10297v72025
  49. Injecting Domain-Specific Knowledge into Large Language Models: A Comprehensive Survey

    Zirui Song, Bin Yan, Yuhan Liu +4

    cs.CLarXiv:2502.10708v22025
  50. LLM-Rec: Personalized Recommendation via Prompting Large Language Models

    Hanjia Lyu, Song Jiang, Hanqing Zeng +7

    cs.CLcs.AIcs.IRarXiv:2307.15780v32023
  51. Token Pruning in Multimodal Large Language Models: Are We Solving the Right Problem?

    Zichen Wen, Yifeng Gao, Weijia Li +2

    cs.CLcs.CVarXiv:2502.11501v22025
  52. Large Language Models for Time Series: A Survey

    Xiyuan Zhang, Ranak Roy Chowdhury, Rajesh K. Gupta +1

    cs.LGcs.AIcs.CLarXiv:2402.01801v32024
  53. ICLE++: Modeling Fine-Grained Traits for Holistic Essay Scoring

    Shengjie Li, Vincent Ng

    cs.CLarXiv:2607.27671v12026
  54. A Survey on Knowledge-Oriented Retrieval-Augmented Generation

    Mingyue Cheng, Yucong Luo, Jie Ouyang +9

    cs.CLcs.AIarXiv:2503.10677v32025
  55. LLM-SRBench: A New Benchmark for Scientific Equation Discovery with Large Language Models

    Parshin Shojaee, Ngoc-Hieu Nguyen, Kazem Meidani +3

    cs.CLcs.AIcs.LGarXiv:2504.10415v22025
  56. ATLAS: Learning to Optimally Memorize the Context at Test Time

    Ali Behrouz, Zeman Li, Praneeth Kacham +5

    cs.CLcs.AIarXiv:2505.23735v12025
  57. Learning Adaptive Parallel Reasoning with Language Models

    Jiayi Pan, Xiuyu Li, Long Lian +6

    cs.AIcs.CLarXiv:2504.15466v22025
  58. AuditBench: Evaluating Alignment Auditing Techniques on Models with Hidden Behaviors

    Abhay Sheshadri, Aidan Ewart, Kai Fronsdal +5

    cs.CLarXiv:2602.22755v32026
  59. AgentNet: Decentralized Evolutionary Coordination for LLM-based Multi-Agent Systems

    Yingxuan Yang, Huacan Chai, Shuai Shao +4

    cs.MAcs.CLarXiv:2504.00587v22025
  60. MTRAG-UN: A Benchmark for Open Challenges in Multi-Turn RAG Conversations

    Sara Rosenthal, Yannis Katsis, Vraj Shah +3

    cs.CLarXiv:2602.23184v12026