Computation and Language

Papers filed under cs.CL on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

4,321 to 4,380 of 11,252

  1. VideoChat: Chat-Centric Video Understanding

    KunChang Li, Yinan He, Yi Wang +6

    cs.CVcs.CLarXiv:2305.06355v22023
  2. Learning to Exploit Temporal Structure for Biomedical Vision-Language Processing

    Shruthi Bannur, Stephanie Hyland, Qianchu Liu +13

    cs.CVcs.CLarXiv:2301.04558v22023
  3. PromptCast: A New Prompt-based Learning Paradigm for Time Series Forecasting

    Hao Xue, Flora D. Salim

    stat.MEcs.AIcs.CLarXiv:2210.08964v52022
  4. ProgPrompt: Generating Situated Robot Task Plans using Large Language Models

    Ishika Singh, Valts Blukis, Arsalan Mousavian +6

    cs.ROcs.AIcs.CLarXiv:2209.11302v12022
  5. FLEURS: Few-shot Learning Evaluation of Universal Representations of Speech

    Alexis Conneau, Min Ma, Simran Khanuja +6

    cs.CLcs.LGcs.SDarXiv:2205.12446v12022
  6. Training language models to follow instructions with human feedback

    Long Ouyang, Jeff Wu, Xu Jiang +17

    cs.CLcs.AIcs.LGarXiv:2203.02155v12022
    Summaries:한국어
  7. Chain-of-Thought Prompting Elicits Reasoning in Large Language Models

    Jason Wei, Xuezhi Wang, Dale Schuurmans +6

    cs.CLcs.AIarXiv:2201.11903v62022
  8. LaMDA: Language Models for Dialog Applications

    Romal Thoppilan, Daniel De Freitas, Jamie Hall +57

    cs.CLcs.AIarXiv:2201.08239v32022
  9. WavLM: Large-Scale Self-Supervised Pre-Training for Full Stack Speech Processing

    Sanyuan Chen, Chengyi Wang, Zhengyang Chen +16

    cs.CLcs.SDeess.ASarXiv:2110.13900v52021
  10. Pre-train, Prompt, and Predict: A Systematic Survey of Prompting Methods in Natural Language Processing

    Pengfei Liu, Weizhe Yuan, Jinlan Fu +3

    cs.CLcs.AIcs.LGarXiv:2107.13586v12021
  11. From Show to Tell: A Survey on Deep Learning-based Image Captioning

    Matteo Stefanini, Marcella Cornia, Lorenzo Baraldi +3

    cs.CVcs.CLarXiv:2107.06912v32021
  12. Structured Denoising Diffusion Models in Discrete State-Spaces

    Jacob Austin, Daniel D. Johnson, Jonathan Ho +2

    cs.LGcs.AIcs.CLarXiv:2107.03006v32021
  13. InfographicVQA

    Minesh Mathew, Viraj Bagal, Rubèn Pérez Tito +3

    cs.CVcs.CLarXiv:2104.12756v22021
  14. Identifying and Controlling Important Neurons in Neural Machine Translation

    Anthony Bau, Yonatan Belinkov, Hassan Sajjad +3

    cs.CLarXiv:1811.01157v12018
  15. Inference-Time Scaling for Generalist Reward Modeling

    Zijun Liu, Peiyi Wang, Runxin Xu +5

    cs.CLcs.AIcs.LGarXiv:2504.02495v32025
  16. Counterfactual VQA: A Cause-Effect Look at Language Bias

    Yulei Niu, Kaihua Tang, Hanwang Zhang +3

    cs.CVcs.CLarXiv:2006.04315v42020
  17. Chain of Draft: Thinking Faster by Writing Less

    Silei Xu, Wenhao Xie, Lingxiao Zhao +1

    cs.CLarXiv:2502.18600v22025
  18. WebSailor: Navigating Super-human Reasoning for Web Agent

    Kuan Li, Zhongwang Zhang, Huifeng Yin +16

    cs.CLcs.AIarXiv:2507.02592v12025
  19. Artificial Hivemind: The Open-Ended Homogeneity of Language Models (and Beyond)

    Liwei Jiang, Yuanjun Chai, Margaret Li +7

    cs.CLarXiv:2510.22954v12025
  20. TextBugger: Generating Adversarial Text Against Real-world Applications

    Jinfeng Li, Shouling Ji, Tianyu Du +2

    cs.CRcs.CLcs.LGarXiv:1812.05271v12018
  21. Entropy-Aware On-Policy Distillation of Language Models

    Woogyeol Jin, Taywon Min, Yongjin Yang +5

    cs.LGcs.CLarXiv:2603.07079v32026
  22. From RAG to Memory: Non-Parametric Continual Learning for Large Language Models

    Bernal Jiménez Gutiérrez, Yiheng Shu, Weijian Qi +2

    cs.CLcs.AIarXiv:2502.14802v22025
  23. Understanding and Overcoming the Challenges of Efficient Transformer Quantization

    Yelysei Bondarenko, Markus Nagel, Tijmen Blankevoort

    cs.LGcs.AIcs.CLarXiv:2109.12948v12021
  24. Attaining the Unattainable? Reassessing Claims of Human Parity in Neural Machine Translation

    Antonio Toral, Sheila Castilho, Ke Hu +1

    cs.CLarXiv:1808.10432v12018
  25. IDEEA: training-free Input-Dependent stEEring via Activation cluster matching

    Zheng Wang, Muchen Li, Renjie Liao +1

    cs.CLcs.LGarXiv:2609.02089v12026
  26. MiniMax-M1: Scaling Test-Time Compute Efficiently with Lightning Attention

    MiniMax, :, Aili Chen +125

    cs.CLcs.LGarXiv:2506.13585v12025
  27. Recent Trends in Deep Learning Based Natural Language Processing

    Tom Young, Devamanyu Hazarika, Soujanya Poria +1

    cs.CLarXiv:1708.02709v82017
  28. Outrageously Large Neural Networks: The Sparsely-Gated Mixture-of-Experts Layer

    Noam Shazeer, Azalia Mirhoseini, Krzysztof Maziarz +4

    cs.LGcs.CLcs.NEarXiv:1701.06538v12017
  29. Sentiment Analysis of Twitter Data for Predicting Stock Market Movements

    Venkata Sasank Pagolu, Kamal Nayan Reddy Challa, Ganapati Panda +1

    cs.IRcs.CLcs.SIarXiv:1610.09225v12016
  30. SQuAD: 100,000+ Questions for Machine Comprehension of Text

    Pranav Rajpurkar, Jian Zhang, Konstantin Lopyrev +1

    cs.CLarXiv:1606.05250v32016
  31. Deep Sentence Embedding Using Long Short-Term Memory Networks: Analysis and Application to Information Retrieval

    Hamid Palangi, Li Deng, Yelong Shen +5

    cs.CLcs.IRcs.LGarXiv:1502.06922v32015
  32. Neural Machine Translation by Jointly Learning to Align and Translate

    Dzmitry Bahdanau, Kyunghyun Cho, Yoshua Bengio

    cs.CLcs.LGcs.NEarXiv:1409.0473v72014
  33. Home Location Identification of Twitter Users

    Jalal Mahmud, Jeffrey Nichols, Clemens Drews

    cs.SIcs.CLcs.CYarXiv:1403.2345v12014
  34. R1-VL: Learning to Reason with Multimodal Large Language Models via Step-wise Group Relative Policy Optimization

    Jingyi Zhang, Jiaxing Huang, Huanjin Yao +4

    cs.AIcs.CLcs.CVarXiv:2503.12937v22025
  35. RAGEN: Understanding Self-Evolution in LLM Agents via Multi-Turn Reinforcement Learning

    Zihan Wang, Kangrui Wang, Qineng Wang +15

    cs.LGcs.AIcs.CLarXiv:2504.20073v22025
  36. Lingshu: A Generalist Foundation Model for Unified Multimodal Medical Understanding and Reasoning

    LASA Team, Weiwen Xu, Hou Pong Chan +16

    cs.CLcs.AIcs.CVarXiv:2506.07044v42025
  37. MemoryWalker: Stop Training Agents on Contexts They Never Saw

    Zinco J, Xunjie Zhu, Shen Huang +3

    cs.LGcs.CLarXiv:2609.00865v12026
  38. Is Neural Machine Translation Ready for Deployment? A Case Study on 30 Translation Directions

    Marcin Junczys-Dowmunt, Tomasz Dwojak, Hieu Hoang

    cs.CLarXiv:1610.01108v32016
  39. A Fundamental Algorithm for Dependency Parsing (With Corrections)

    Michael A. Covington

    cs.CLarXiv:2510.19996v12025
  40. GUI-R1 : A Generalist R1-Style Vision-Language Action Model For GUI Agents

    Run Luo, Lu Wang, Wanwei He +3

    cs.CVcs.CLcs.HCarXiv:2504.10458v42025
  41. SWE-smith: Scaling Data for Software Engineering Agents

    John Yang, Kilian Lieret, Carlos E. Jimenez +7

    cs.SEcs.AIcs.CLarXiv:2504.21798v22025
  42. Reinforcement Learning with Verifiable Rewards Implicitly Incentivizes Correct Reasoning in Base LLMs

    Xumeng Wen, Zihan Liu, Shun Zheng +9

    cs.AIcs.CLarXiv:2506.14245v22025
  43. Findings of the BabyLM Challenge: Sample-Efficient Pretraining on Developmentally Plausible Corpora

    Alex Warstadt, Aaron Mueller, Leshem Choshen +8

    cs.CLarXiv:2504.08165v12025
  44. MoEfication: Transformer Feed-forward Layers are Mixtures of Experts

    Zhengyan Zhang, Yankai Lin, Zhiyuan Liu +3

    cs.CLarXiv:2110.01786v32021
  45. TokenSkip: Controllable Chain-of-Thought Compression in LLMs

    Heming Xia, Chak Tou Leong, Wenjie Wang +2

    cs.CLcs.AIarXiv:2502.12067v32025
  46. When Being Unseen from mBERT is just the Beginning: Handling New Languages With Multilingual Language Models

    Benjamin Muller, Antonis Anastasopoulos, Benoît Sagot +1

    cs.CLarXiv:2010.12858v22020
  47. TTRL: Test-Time Reinforcement Learning

    Yuxin Zuo, Kaiyan Zhang, Li Sheng +13

    cs.CLcs.LGarXiv:2504.16084v32025
  48. Evaluating Memory in LLM Agents via Incremental Multi-Turn Interactions

    Yuanzhe Hu, Yu Wang, Julian McAuley

    cs.CLcs.AIarXiv:2507.05257v42025
  49. Racism is a Virus: Anti-Asian Hate and Counterspeech in Social Media during the COVID-19 Crisis

    Bing He, Caleb Ziems, Sandeep Soni +3

    cs.SIcs.CLcs.CYarXiv:2005.12423v22020
  50. CODI: Compressing Chain-of-Thought into Continuous Space via Self-Distillation

    Zhenyi Shen, Hanqi Yan, Linhai Zhang +3

    cs.CLarXiv:2502.21074v32025
  51. Quit While You're Ahead: Quit for Efficient Candidate Generation in Machine Translation Reranking

    Guangyu Chen, Boxuan Lyu, Hidetaka Kamigaito +2

    cs.CLarXiv:2609.00588v22026
  52. Stochastic Answer Networks for Machine Reading Comprehension

    Xiaodong Liu, Yelong Shen, Kevin Duh +1

    cs.CLarXiv:1712.03556v22017
  53. Layer by Layer: Uncovering Hidden Representations in Language Models

    Oscar Skean, Md Rifat Arefin, Dan Zhao +4

    cs.LGcs.AIcs.CLarXiv:2502.02013v22025
  54. BrowseComp-Plus: A More Fair and Transparent Evaluation Benchmark of Deep-Research Agent

    Zijian Chen, Xueguang Ma, Shengyao Zhuang +17

    cs.CLcs.IRarXiv:2508.06600v12025
  55. Does task decomposition improve automatic NLG evaluation?

    Sebastian Steindl, Nikos Voskarides, Alberto Gasparin +1

    cs.CLarXiv:2609.01139v12026
  56. Towards Reasoning Era: A Survey of Long Chain-of-Thought for Reasoning Large Language Models

    Qiguang Chen, Libo Qin, Jinhao Liu +7

    cs.AIcs.CLarXiv:2503.09567v52025
  57. Post-hoc Alignment of LLM-judges to Human Judgment Distribution

    Sebastian Steindl, Nikos Voskarides, Alberto Gasparin +1

    cs.CLarXiv:2609.01073v12026
  58. GDPO: Group reward-Decoupled Normalization Policy Optimization for Multi-reward RL Optimization

    Shih-Yang Liu, Xin Dong, Ximing Lu +10

    cs.CLcs.AIcs.LGarXiv:2601.05242v12026
  59. CHARM: Character Hallucination for Multicultural Role Play Benchmark

    Sunkyung Han, Nahyeon Park, Gaeun Seo +2

    cs.CLcs.AIarXiv:2609.01352v12026
  60. WorldBench: Culturally Grounded Benchmark for Multilingual Agents

    Leonardo Ranaldi, Sherrie Shen, Jushi Kai +1

    cs.AIcs.CLarXiv:2609.01056v12026