Computation and Language
Papers filed under cs.CL on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
2,521 to 2,580 of 11,224
Relating Word Embedding Gender Biases to Gender Gaps: A Cross-Cultural Analysis
Scott Friedman, Sonja Schmer-Galunder, Anthony Chen +1
cs.CLarXiv:2601.17203v12026AutoSOTA: An End-to-End Automated Research System for State-of-the-Art AI Model Discovery
Yu Li, Chenyang Shao, Xinyang Liu +13
cs.CLcs.CEarXiv:2604.05550v22026Knowledgeable Reader: Enhancing Cloze-Style Reading Comprehension with External Commonsense Knowledge
Todor Mihaylov, Anette Frank
cs.CLarXiv:1805.07858v12018Searching for Best Practices in Retrieval-Augmented Generation
Xiaohua Wang, Zhenghua Wang, Xuan Gao +11
cs.CLarXiv:2407.01219v12024Understanding the LLM-ification of CHI: Unpacking the Impact of LLMs at CHI through a Systematic Literature Review
Rock Yuren Pang, Hope Schroeder, Kynnedy Simone Smith +4
cs.HCcs.AIcs.CLarXiv:2501.12557v12025Knowledge Distillation for Large Language Models
Alejandro Paredes La Torre, Barbara Flores, Diego Rodriguez
cs.CLcs.AIarXiv:2603.13765v12026QuestBench: Can LLMs ask the right question to acquire information in reasoning tasks?
Belinda Z. Li, Been Kim, Zi Wang
cs.AIcs.CLcs.LGarXiv:2503.22674v22025RoboSpatial: Teaching Spatial Understanding to 2D and 3D Vision-Language Models for Robotics
Chan Hee Song, Valts Blukis, Jonathan Tremblay +3
cs.CVcs.AIcs.CLarXiv:2411.16537v52024Composable Sparse Fine-Tuning for Cross-Lingual Transfer
Alan Ansell, Edoardo Maria Ponti, Anna Korhonen +1
cs.CLarXiv:2110.07560v22021ChatGPT: Vision and Challenges
Sukhpal Singh Gill, Rupinder Kaur
cs.CYcs.CLarXiv:2305.15323v12023Scientific Agent Skills: A Library of Procedural Knowledge for Research Agents
Timothy Kassis, Vinayak Agarwal, Yuhuan He +2
cs.CLcs.AIarXiv:2609.00065v22026Fine-Grained Attention Mechanism for Neural Machine Translation
Heeyoul Choi, Kyunghyun Cho, Yoshua Bengio
cs.CLarXiv:1803.11407v22018AI Flow: Perspectives, Scenarios, and Approaches
Hongjun An, Wenhan Hu, Sida Huang +11
cs.AIcs.CLcs.CVarXiv:2506.12479v32025Code to Think, Think to Code: A Survey on Code-Enhanced Reasoning and Reasoning-Driven Code Intelligence in LLMs
Dayu Yang, Tianyang Liu, Daoan Zhang +8
cs.CLcs.AIcs.LGarXiv:2502.19411v12025ExCL: Extractive Clip Localization Using Natural Language Descriptions
Soham Ghosh, Anuva Agarwal, Zarana Parekh +1
cs.CLarXiv:1904.02755v12019Can Machines Learn Morality? The Delphi Experiment
Liwei Jiang, Jena D. Hwang, Chandra Bhagavatula +12
cs.CLarXiv:2110.07574v22021GALAXY: A Generative Pre-trained Model for Task-Oriented Dialog with Semi-Supervised Learning and Explicit Policy Injection
Wanwei He, Yinpei Dai, Yinhe Zheng +9
cs.CLarXiv:2111.14592v82021RouterEval: A Comprehensive Benchmark for Routing LLMs to Explore Model-level Scaling Up in LLMs
Zhongzhan Huang, Guoming Ling, Yupei Lin +4
cs.CLcs.AIarXiv:2503.10657v22025A Survey on Diffusion Language Models
Tianyi Li, Mingda Chen, Bowei Guo +1
cs.CLcs.AIcs.LGarXiv:2508.10875v32025HumanLM: Simulating Users with State Alignment Beats Response Imitation
Shirley Wu, Evelyn Choi, Arpandeep Khatua +7
cs.CLcs.AIarXiv:2603.03303v12026CONTaiNER: Few-Shot Named Entity Recognition via Contrastive Learning
Sarkar Snigdha Sarathi Das, Arzoo Katiyar, Rebecca J. Passonneau +1
cs.CLarXiv:2109.07589v22021CLadder: Assessing Causal Reasoning in Language Models
Zhijing Jin, Yuen Chen, Felix Leeb +8
cs.CLcs.AIcs.LGarXiv:2312.04350v32023Odysseys: Benchmarking Web Agents on Realistic Long Horizon Tasks
Lawrence Keunho Jang, Jing Yu Koh, Daniel Fried +1
cs.LGcs.CLarXiv:2604.24964v12026CoDEx: A Comprehensive Knowledge Graph Completion Benchmark
Tara Safavi, Danai Koutra
cs.CLcs.AIcs.IRarXiv:2009.07810v22020The FACTS Grounding Leaderboard: Benchmarking LLMs' Ability to Ground Responses to Long-Form Input
Alon Jacovi, Andrew Wang, Chris Alberti +23
cs.CLarXiv:2501.03200v12025Aligning Language Models from User Interactions
Thomas Kleine Buening, Jonas Hübotter, Barna Pásztor +3
cs.CLcs.AIcs.LGarXiv:2603.12273v12026LentEx: Generalizable Latent Entity Extraction via Synthetic Data and Instruction-Tuned LLMs
Umesh Bodhwani, Yuan Ling, Cibi Chakravarthy Senthilkumar +4
cs.CLarXiv:2609.04511v12026Training LLM-based Tutors to Improve Student Learning Outcomes in Dialogues
Alexander Scarlatos, Naiming Liu, Jaewook Lee +2
cs.CLcs.CYarXiv:2503.06424v22025On the generalization of language models from in-context learning and finetuning: a controlled study
Andrew K. Lampinen, Arslan Chaudhry, Stephanie C. Y. Chan +7
cs.CLcs.AIcs.LGarXiv:2505.00661v32025COOT: Cooperative Hierarchical Transformer for Video-Text Representation Learning
Simon Ging, Mohammadreza Zolfaghari, Hamed Pirsiavash +1
cs.CVcs.AIcs.CLarXiv:2011.00597v12020InteractBench: Benchmarking LLMs on Competitive Programming under Unrevealed Information
Jiaze Li, Aocheng Shen, Bing Liu +4
cs.SEcs.AIcs.CLarXiv:2608.29632v12026ReasonFlux: Hierarchical LLM Reasoning via Scaling Thought Templates
Ling Yang, Zhaochen Yu, Bin Cui +1
cs.CLcs.AIcs.LGarXiv:2502.06772v22025Few-shot Text Classification with Distributional Signatures
Yujia Bao, Menghua Wu, Shiyu Chang +1
cs.CLcs.LGarXiv:1908.06039v32019Reasoning-Augmented Conversation for Multi-Turn Jailbreak Attacks on Large Language Models
Zonghao Ying, Deyue Zhang, Zonglei Jing +7
cs.CLcs.AIcs.CRarXiv:2502.11054v42025PrivacyLens: Evaluating Privacy Norm Awareness of Language Models in Action
Yijia Shao, Tianshi Li, Weiyan Shi +2
cs.CLcs.AIcs.CRarXiv:2409.00138v32024Enhancing Retrieval-Augmented Generation: A Study of Best Practices
Siran Li, Linus Stenzel, Carsten Eickhoff +1
cs.CLcs.AIarXiv:2501.07391v12025MathCoder-VL: Bridging Vision and Code for Enhanced Multimodal Mathematical Reasoning
Ke Wang, Junting Pan, Linda Wei +8
cs.CVcs.AIcs.CLarXiv:2505.10557v12025The Spike, the Sparse and the Sink: Anatomy of Massive Activations and Attention Sinks
Shangwen Sun, Alfredo Canziani, Yann LeCun +1
cs.AIcs.CLarXiv:2603.05498v12026Parametric Retrieval Augmented Generation
Weihang Su, Yichen Tang, Qingyao Ai +6
cs.CLcs.IRarXiv:2501.15915v12025Towards Holistic Evaluation of Large Audio-Language Models: A Comprehensive Survey
Chih-Kai Yang, Neo S. Ho, Hung-yi Lee
eess.AScs.AIcs.CLarXiv:2505.15957v42025Nemotron-Research-Tool-N1: Exploring Tool-Using Language Models with Reinforced Reasoning
Shaokun Zhang, Yi Dong, Jieyu Zhang +6
cs.CLcs.AIarXiv:2505.00024v22025Foundation Models for Geospatial Reasoning: Assessing Capabilities of Large Language Models in Understanding Geometries and Topological Spatial Relations
Yuhan Ji, Song Gao, Ying Nie +2
cs.CLcs.AIarXiv:2505.17136v12025The Language of the Question Selects the Market: Query Language and Exit IP as Separable Factors in Commercial Recommendations from a Generative Search Interface
Dmitrij Żatuchin
cs.IRcs.CLcs.CYarXiv:2608.30052v12026Demand-Side Measurement for Generative Engine Optimization: Constructing and Validating a Million-Persona, Intent-Annotated Buyer Corpus
Dmitrij Żatuchin, Daniil Dzemesjuk
cs.IRcs.CLarXiv:2608.30023v12026Inner Thinking Transformer: Leveraging Dynamic Depth Scaling to Foster Adaptive Internal Thinking
Yilong Chen, Junyuan Shang, Zhenyu Zhang +7
cs.CLarXiv:2502.13842v22025LiveMathematicianBench: A Live Benchmark for Mathematician-Level Reasoning with Proof Sketches
Linyang He, Qiyao Yu, Hanze Dong +5
cs.CLcs.AIcs.LGarXiv:2604.01754v12026Internet-augmented language models through few-shot prompting for open-domain question answering
Angeliki Lazaridou, Elena Gribovskaya, Wojciech Stokowiec +1
cs.CLcs.LGarXiv:2203.05115v22022trajectory-judge: What Outcome-Only LLM Judges Miss on Agent Trajectories
Hadi Mohammadi
cs.CLcs.AIcs.SEarXiv:2609.00038v12026Towards Agentic RAG with Deep Reasoning: A Survey of RAG-Reasoning Systems in LLMs
Yangning Li, Weizhi Zhang, Yuyao Yang +17
cs.CLcs.AIarXiv:2507.09477v22025A Survey on LoRA of Large Language Models
Yuren Mao, Yuhang Ge, Yijiang Fan +4
cs.LGcs.AIcs.CLarXiv:2407.11046v42024Don't Trust ChatGPT when Your Question is not in English: A Study of Multilingual Abilities and Types of LLMs
Xiang Zhang, Senyu Li, Bradley Hauer +2
cs.CLcs.AIarXiv:2305.16339v22023RLHF Deciphered: A Critical Analysis of Reinforcement Learning from Human Feedback for LLMs
Shreyas Chaudhari, Pranjal Aggarwal, Vishvak Murahari +5
cs.LGcs.AIcs.CLarXiv:2404.08555v22024Evaluating Step-by-step Reasoning Traces: A Survey
Jinu Lee, Julia Hockenmaier
cs.CLarXiv:2502.12289v32025AutoAgent: A Fully-Automated and Zero-Code Framework for LLM Agents
Jiabin Tang, Tianyu Fan, Chao Huang
cs.AIcs.CLarXiv:2502.05957v32025A Comparative Survey of Recent Natural Language Interfaces for Databases
Katrin Affolter, Kurt Stockinger, Abraham Bernstein
cs.DBcs.CLcs.LGarXiv:1906.08990v12019Reduced, Reused and Recycled: The Life of a Dataset in Machine Learning Research
Bernard Koch, Emily Denton, Alex Hanna +1
cs.LGcs.CLcs.CVarXiv:2112.01716v12021Self-Alignment with Instruction Backtranslation
Xian Li, Ping Yu, Chunting Zhou +5
cs.CLarXiv:2308.06259v32023A Calibrated Reflection Approach for Enhancing Confidence Estimation in LLMs
Umesh Bodhwani, Yuan Ling, Shujing Dong +3
cs.CLarXiv:2609.04539v12026Deception Abilities Emerged in Large Language Models
Thilo Hagendorff
cs.CLcs.AIcs.LGarXiv:2307.16513v22023Accelerating LLM Inference with Staged Speculative Decoding
Benjamin Spector, Chris Re
cs.AIcs.CLarXiv:2308.04623v12023