Computation and Language

Papers filed under cs.CL on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

2,521 to 2,580 of 11,224

  1. Relating Word Embedding Gender Biases to Gender Gaps: A Cross-Cultural Analysis

    Scott Friedman, Sonja Schmer-Galunder, Anthony Chen +1

    cs.CLarXiv:2601.17203v12026
  2. AutoSOTA: An End-to-End Automated Research System for State-of-the-Art AI Model Discovery

    Yu Li, Chenyang Shao, Xinyang Liu +13

    cs.CLcs.CEarXiv:2604.05550v22026
  3. Knowledgeable Reader: Enhancing Cloze-Style Reading Comprehension with External Commonsense Knowledge

    Todor Mihaylov, Anette Frank

    cs.CLarXiv:1805.07858v12018
  4. Searching for Best Practices in Retrieval-Augmented Generation

    Xiaohua Wang, Zhenghua Wang, Xuan Gao +11

    cs.CLarXiv:2407.01219v12024
  5. Understanding the LLM-ification of CHI: Unpacking the Impact of LLMs at CHI through a Systematic Literature Review

    Rock Yuren Pang, Hope Schroeder, Kynnedy Simone Smith +4

    cs.HCcs.AIcs.CLarXiv:2501.12557v12025
  6. Knowledge Distillation for Large Language Models

    Alejandro Paredes La Torre, Barbara Flores, Diego Rodriguez

    cs.CLcs.AIarXiv:2603.13765v12026
  7. QuestBench: Can LLMs ask the right question to acquire information in reasoning tasks?

    Belinda Z. Li, Been Kim, Zi Wang

    cs.AIcs.CLcs.LGarXiv:2503.22674v22025
  8. RoboSpatial: Teaching Spatial Understanding to 2D and 3D Vision-Language Models for Robotics

    Chan Hee Song, Valts Blukis, Jonathan Tremblay +3

    cs.CVcs.AIcs.CLarXiv:2411.16537v52024
  9. Composable Sparse Fine-Tuning for Cross-Lingual Transfer

    Alan Ansell, Edoardo Maria Ponti, Anna Korhonen +1

    cs.CLarXiv:2110.07560v22021
  10. ChatGPT: Vision and Challenges

    Sukhpal Singh Gill, Rupinder Kaur

    cs.CYcs.CLarXiv:2305.15323v12023
  11. Scientific Agent Skills: A Library of Procedural Knowledge for Research Agents

    Timothy Kassis, Vinayak Agarwal, Yuhuan He +2

    cs.CLcs.AIarXiv:2609.00065v22026
  12. Fine-Grained Attention Mechanism for Neural Machine Translation

    Heeyoul Choi, Kyunghyun Cho, Yoshua Bengio

    cs.CLarXiv:1803.11407v22018
  13. AI Flow: Perspectives, Scenarios, and Approaches

    Hongjun An, Wenhan Hu, Sida Huang +11

    cs.AIcs.CLcs.CVarXiv:2506.12479v32025
  14. Code to Think, Think to Code: A Survey on Code-Enhanced Reasoning and Reasoning-Driven Code Intelligence in LLMs

    Dayu Yang, Tianyang Liu, Daoan Zhang +8

    cs.CLcs.AIcs.LGarXiv:2502.19411v12025
  15. ExCL: Extractive Clip Localization Using Natural Language Descriptions

    Soham Ghosh, Anuva Agarwal, Zarana Parekh +1

    cs.CLarXiv:1904.02755v12019
  16. Can Machines Learn Morality? The Delphi Experiment

    Liwei Jiang, Jena D. Hwang, Chandra Bhagavatula +12

    cs.CLarXiv:2110.07574v22021
  17. GALAXY: A Generative Pre-trained Model for Task-Oriented Dialog with Semi-Supervised Learning and Explicit Policy Injection

    Wanwei He, Yinpei Dai, Yinhe Zheng +9

    cs.CLarXiv:2111.14592v82021
  18. RouterEval: A Comprehensive Benchmark for Routing LLMs to Explore Model-level Scaling Up in LLMs

    Zhongzhan Huang, Guoming Ling, Yupei Lin +4

    cs.CLcs.AIarXiv:2503.10657v22025
  19. A Survey on Diffusion Language Models

    Tianyi Li, Mingda Chen, Bowei Guo +1

    cs.CLcs.AIcs.LGarXiv:2508.10875v32025
  20. HumanLM: Simulating Users with State Alignment Beats Response Imitation

    Shirley Wu, Evelyn Choi, Arpandeep Khatua +7

    cs.CLcs.AIarXiv:2603.03303v12026
  21. CONTaiNER: Few-Shot Named Entity Recognition via Contrastive Learning

    Sarkar Snigdha Sarathi Das, Arzoo Katiyar, Rebecca J. Passonneau +1

    cs.CLarXiv:2109.07589v22021
  22. CLadder: Assessing Causal Reasoning in Language Models

    Zhijing Jin, Yuen Chen, Felix Leeb +8

    cs.CLcs.AIcs.LGarXiv:2312.04350v32023
  23. Odysseys: Benchmarking Web Agents on Realistic Long Horizon Tasks

    Lawrence Keunho Jang, Jing Yu Koh, Daniel Fried +1

    cs.LGcs.CLarXiv:2604.24964v12026
  24. CoDEx: A Comprehensive Knowledge Graph Completion Benchmark

    Tara Safavi, Danai Koutra

    cs.CLcs.AIcs.IRarXiv:2009.07810v22020
  25. The FACTS Grounding Leaderboard: Benchmarking LLMs' Ability to Ground Responses to Long-Form Input

    Alon Jacovi, Andrew Wang, Chris Alberti +23

    cs.CLarXiv:2501.03200v12025
  26. Aligning Language Models from User Interactions

    Thomas Kleine Buening, Jonas Hübotter, Barna Pásztor +3

    cs.CLcs.AIcs.LGarXiv:2603.12273v12026
  27. LentEx: Generalizable Latent Entity Extraction via Synthetic Data and Instruction-Tuned LLMs

    Umesh Bodhwani, Yuan Ling, Cibi Chakravarthy Senthilkumar +4

    cs.CLarXiv:2609.04511v12026
  28. Training LLM-based Tutors to Improve Student Learning Outcomes in Dialogues

    Alexander Scarlatos, Naiming Liu, Jaewook Lee +2

    cs.CLcs.CYarXiv:2503.06424v22025
  29. On the generalization of language models from in-context learning and finetuning: a controlled study

    Andrew K. Lampinen, Arslan Chaudhry, Stephanie C. Y. Chan +7

    cs.CLcs.AIcs.LGarXiv:2505.00661v32025
  30. COOT: Cooperative Hierarchical Transformer for Video-Text Representation Learning

    Simon Ging, Mohammadreza Zolfaghari, Hamed Pirsiavash +1

    cs.CVcs.AIcs.CLarXiv:2011.00597v12020
  31. InteractBench: Benchmarking LLMs on Competitive Programming under Unrevealed Information

    Jiaze Li, Aocheng Shen, Bing Liu +4

    cs.SEcs.AIcs.CLarXiv:2608.29632v12026
  32. ReasonFlux: Hierarchical LLM Reasoning via Scaling Thought Templates

    Ling Yang, Zhaochen Yu, Bin Cui +1

    cs.CLcs.AIcs.LGarXiv:2502.06772v22025
  33. Few-shot Text Classification with Distributional Signatures

    Yujia Bao, Menghua Wu, Shiyu Chang +1

    cs.CLcs.LGarXiv:1908.06039v32019
  34. Reasoning-Augmented Conversation for Multi-Turn Jailbreak Attacks on Large Language Models

    Zonghao Ying, Deyue Zhang, Zonglei Jing +7

    cs.CLcs.AIcs.CRarXiv:2502.11054v42025
  35. PrivacyLens: Evaluating Privacy Norm Awareness of Language Models in Action

    Yijia Shao, Tianshi Li, Weiyan Shi +2

    cs.CLcs.AIcs.CRarXiv:2409.00138v32024
  36. Enhancing Retrieval-Augmented Generation: A Study of Best Practices

    Siran Li, Linus Stenzel, Carsten Eickhoff +1

    cs.CLcs.AIarXiv:2501.07391v12025
  37. MathCoder-VL: Bridging Vision and Code for Enhanced Multimodal Mathematical Reasoning

    Ke Wang, Junting Pan, Linda Wei +8

    cs.CVcs.AIcs.CLarXiv:2505.10557v12025
  38. The Spike, the Sparse and the Sink: Anatomy of Massive Activations and Attention Sinks

    Shangwen Sun, Alfredo Canziani, Yann LeCun +1

    cs.AIcs.CLarXiv:2603.05498v12026
  39. Parametric Retrieval Augmented Generation

    Weihang Su, Yichen Tang, Qingyao Ai +6

    cs.CLcs.IRarXiv:2501.15915v12025
  40. Towards Holistic Evaluation of Large Audio-Language Models: A Comprehensive Survey

    Chih-Kai Yang, Neo S. Ho, Hung-yi Lee

    eess.AScs.AIcs.CLarXiv:2505.15957v42025
  41. Nemotron-Research-Tool-N1: Exploring Tool-Using Language Models with Reinforced Reasoning

    Shaokun Zhang, Yi Dong, Jieyu Zhang +6

    cs.CLcs.AIarXiv:2505.00024v22025
  42. Foundation Models for Geospatial Reasoning: Assessing Capabilities of Large Language Models in Understanding Geometries and Topological Spatial Relations

    Yuhan Ji, Song Gao, Ying Nie +2

    cs.CLcs.AIarXiv:2505.17136v12025
  43. The Language of the Question Selects the Market: Query Language and Exit IP as Separable Factors in Commercial Recommendations from a Generative Search Interface

    Dmitrij Żatuchin

    cs.IRcs.CLcs.CYarXiv:2608.30052v12026
  44. Demand-Side Measurement for Generative Engine Optimization: Constructing and Validating a Million-Persona, Intent-Annotated Buyer Corpus

    Dmitrij Żatuchin, Daniil Dzemesjuk

    cs.IRcs.CLarXiv:2608.30023v12026
  45. Inner Thinking Transformer: Leveraging Dynamic Depth Scaling to Foster Adaptive Internal Thinking

    Yilong Chen, Junyuan Shang, Zhenyu Zhang +7

    cs.CLarXiv:2502.13842v22025
  46. LiveMathematicianBench: A Live Benchmark for Mathematician-Level Reasoning with Proof Sketches

    Linyang He, Qiyao Yu, Hanze Dong +5

    cs.CLcs.AIcs.LGarXiv:2604.01754v12026
  47. Internet-augmented language models through few-shot prompting for open-domain question answering

    Angeliki Lazaridou, Elena Gribovskaya, Wojciech Stokowiec +1

    cs.CLcs.LGarXiv:2203.05115v22022
  48. trajectory-judge: What Outcome-Only LLM Judges Miss on Agent Trajectories

    Hadi Mohammadi

    cs.CLcs.AIcs.SEarXiv:2609.00038v12026
  49. Towards Agentic RAG with Deep Reasoning: A Survey of RAG-Reasoning Systems in LLMs

    Yangning Li, Weizhi Zhang, Yuyao Yang +17

    cs.CLcs.AIarXiv:2507.09477v22025
  50. A Survey on LoRA of Large Language Models

    Yuren Mao, Yuhang Ge, Yijiang Fan +4

    cs.LGcs.AIcs.CLarXiv:2407.11046v42024
  51. Don't Trust ChatGPT when Your Question is not in English: A Study of Multilingual Abilities and Types of LLMs

    Xiang Zhang, Senyu Li, Bradley Hauer +2

    cs.CLcs.AIarXiv:2305.16339v22023
  52. RLHF Deciphered: A Critical Analysis of Reinforcement Learning from Human Feedback for LLMs

    Shreyas Chaudhari, Pranjal Aggarwal, Vishvak Murahari +5

    cs.LGcs.AIcs.CLarXiv:2404.08555v22024
  53. Evaluating Step-by-step Reasoning Traces: A Survey

    Jinu Lee, Julia Hockenmaier

    cs.CLarXiv:2502.12289v32025
  54. AutoAgent: A Fully-Automated and Zero-Code Framework for LLM Agents

    Jiabin Tang, Tianyu Fan, Chao Huang

    cs.AIcs.CLarXiv:2502.05957v32025
  55. A Comparative Survey of Recent Natural Language Interfaces for Databases

    Katrin Affolter, Kurt Stockinger, Abraham Bernstein

    cs.DBcs.CLcs.LGarXiv:1906.08990v12019
  56. Reduced, Reused and Recycled: The Life of a Dataset in Machine Learning Research

    Bernard Koch, Emily Denton, Alex Hanna +1

    cs.LGcs.CLcs.CVarXiv:2112.01716v12021
  57. Self-Alignment with Instruction Backtranslation

    Xian Li, Ping Yu, Chunting Zhou +5

    cs.CLarXiv:2308.06259v32023
  58. A Calibrated Reflection Approach for Enhancing Confidence Estimation in LLMs

    Umesh Bodhwani, Yuan Ling, Shujing Dong +3

    cs.CLarXiv:2609.04539v12026
  59. Deception Abilities Emerged in Large Language Models

    Thilo Hagendorff

    cs.CLcs.AIcs.LGarXiv:2307.16513v22023
  60. Accelerating LLM Inference with Staged Speculative Decoding

    Benjamin Spector, Chris Re

    cs.AIcs.CLarXiv:2308.04623v12023