Computation and Language

Papers filed under cs.CL on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

7,861 to 7,920 of 11,278

  1. Skill Issue: Are Skills Language-Invariant in LLMs?

    Bobby Cheng, Adam Gaber, Zhengyuan Liu +4

    cs.CLcs.AIcs.GTarXiv:2608.25832v12026
  2. BioMistral: A Collection of Open-Source Pretrained Large Language Models for Medical Domains

    Yanis Labrak, Adrien Bazoge, Emmanuel Morin +3

    cs.CLcs.AIcs.LGarXiv:2402.10373v32024
  3. Unfolding Scientific Papers into Multi-Turn Generation Trajectories for Continued Pre-Training

    Qiankai Xu, Qiguang Chen, Zixin Su +4

    cs.CLcs.AIcs.LGarXiv:2608.25826v12026
  4. LongMemEval: Benchmarking Chat Assistants on Long-Term Interactive Memory

    Di Wu, Hongwei Wang, Wenhao Yu +3

    cs.CLarXiv:2410.10813v22024
  5. TOFU: A Task of Fictitious Unlearning for LLMs

    Pratyush Maini, Zhili Feng, Avi Schwarzschild +2

    cs.LGcs.CLarXiv:2401.06121v12024
  6. Learning New Facts with QLoRA: An Acquisition-Retention Frontier

    Estelle Zheng, Sébastien Warichet, Emmanuel Helbert +1

    cs.CLcs.AIcs.LGarXiv:2608.25677v12026
  7. Think-Probe-Respond: Improving Large Language Models as Judges of Research Idea Novelty

    Tim Schopf, Tobias Schreieder, Akiko Aizawa

    cs.CLcs.AIarXiv:2608.25660v12026
  8. Style Transfer in Text: Exploration and Evaluation

    Zhenxin Fu, Xiaoye Tan, Nanyun Peng +2

    cs.CLarXiv:1711.06861v22017
  9. Localize-Then-Decide Guarantees for LLM Judgments

    Xinyu Li, Yi Zhou, Guanqun Cao +3

    cs.CLarXiv:2608.25824v12026
  10. MoganBert-TR: A Turkish Encoder Foundation Model Trained from Scratch with a CLM-to-MLM Curriculum

    Furkan Yilmaz, Habibe Aleyna Tasdemir, Muhammed Faruk Gozay

    cs.CLcs.AIarXiv:2608.25768v12026
  11. Beam Search, Self-Consistency, and the Limits of Inference-Time Scaling for Grammar-Constrained Text-to-SQL in Small Language Models

    Ty Chermsirivatana, John MacCormick

    cs.CLcs.AIarXiv:2608.25761v12026
  12. When RAG Fails to Equalize: Geo-bias in Factual Question Answering over Public Companies

    Abhinav Havaldar, Enrico Santus

    cs.CLcs.AIarXiv:2608.25717v12026
  13. Overview of SHROOM-Visions 2026: A Shared Task on Hallucination Detection in Large Vision-Language Models

    Raúl Vázquez, Aman Sinha, Chuyuan Li +10

    cs.CLarXiv:2608.25662v12026
  14. Reconstructing the Right Episode: Evaluating Interleaved Conversational Memory Beyond Long Context

    Zhexi Feng, Ruiyi Zhang, Yongbo Yang +1

    cs.CLcs.AIarXiv:2608.25655v12026
  15. The Chase Is the Curriculum, the Capture Anchors the Credit: Pursuit-Evasion Self-Play for Zero-Data LLM Reasoning

    Jing Yu, Shengchao Chen, Yiyun Tan

    cs.CLcs.LGarXiv:2608.21871v12026
  16. Easily Accessible Text-to-Image Generation Amplifies Demographic Stereotypes at Large Scale

    Federico Bianchi, Pratyusha Kalluri, Esin Durmus +7

    cs.CLcs.CVarXiv:2211.03759v22022
  17. Large Language Models Are Reasoning Teachers

    Namgyu Ho, Laura Schmid, Se-Young Yun

    cs.CLcs.AIcs.LGarXiv:2212.10071v22022
  18. Causal Reasoning and Large Language Models: Opening a New Frontier for Causality

    Emre Kıcıman, Robert Ness, Amit Sharma +1

    cs.AIcs.CLcs.CYarXiv:2305.00050v32023
  19. RepoCoder: Repository-Level Code Completion Through Iterative Retrieval and Generation

    Fengji Zhang, Bei Chen, Yue Zhang +6

    cs.CLcs.AIcs.PLarXiv:2303.12570v32023
  20. RRHF: Rank Responses to Align Language Models with Human Feedback without tears

    Zheng Yuan, Hongyi Yuan, Chuanqi Tan +3

    cs.CLarXiv:2304.05302v32023
  21. LexGLUE: A Benchmark Dataset for Legal Language Understanding in English

    Ilias Chalkidis, Abhik Jana, Dirk Hartung +4

    cs.CLarXiv:2110.00976v42021
  22. DExperts: Decoding-Time Controlled Text Generation with Experts and Anti-Experts

    Alisa Liu, Maarten Sap, Ximing Lu +4

    cs.CLarXiv:2105.03023v22021
  23. DeeBERT: Dynamic Early Exiting for Accelerating BERT Inference

    Ji Xin, Raphael Tang, Jaejun Lee +2

    cs.CLcs.LGarXiv:2004.12993v12020
  24. Learning to Generate Reviews and Discovering Sentiment

    Alec Radford, Rafal Jozefowicz, Ilya Sutskever

    cs.LGcs.CLcs.NEarXiv:1704.01444v22017
  25. Learning Mixtures of Plackett-Luce Models for Multi-Objective Alignment

    Dongyue Li, Ziniu Zhang, Lu Wang +1

    cs.LGcs.AIcs.CLarXiv:2608.25200v12026
  26. Reasoning on Graphs: Faithful and Interpretable Large Language Model Reasoning

    Linhao Luo, Yuan-Fang Li, Gholamreza Haffari +1

    cs.CLcs.AIarXiv:2310.01061v22023
  27. Recurrent Neural Network Grammars

    Chris Dyer, Adhiguna Kuncoro, Miguel Ballesteros +1

    cs.CLcs.NEarXiv:1602.07776v42016
  28. Conditional Total Correlation and the Serial Depth of Adaptive Parallel Sampling

    Chuling Wen, Weijie Liang, Jian Lu

    cs.ITcs.CLarXiv:2608.25505v12026
  29. DCGC: Draft-Conditioned Global Correction for Complex Reasoning with Masked Diffusion Models

    Minhae Oh, Nakyung Lee, Jungwoo Lee

    cs.CLcs.AIarXiv:2608.25428v12026
  30. GGSS: Geodesic-Gated Spherical Steering for Inference-Time Debiasing of Generative Vision-Language Models

    Yiqun Sun, Junyu Chen, Pengfei Wei +1

    cs.CYcs.CLcs.CVarXiv:2608.25375v12026
  31. Hyena Hierarchy: Towards Larger Convolutional Language Models

    Michael Poli, Stefano Massaroli, Eric Nguyen +6

    cs.LGcs.CLarXiv:2302.10866v32023
  32. GRIP: Granular Reward-Guided Parameter Interpolation for Efficient Reasoning

    Lam So, Canhui Wu, Han Lin

    cs.CLarXiv:2608.25583v12026
  33. Adaptive Triggering for Bias Correction in LLM Reasoning

    Nayoung Kim, Mickey Mancenido, Huan Liu

    cs.CLcs.AIarXiv:2608.25379v12026
  34. Abductive Commonsense Reasoning

    Chandra Bhagavatula, Ronan Le Bras, Chaitanya Malaviya +6

    cs.CLarXiv:1908.05739v22019
  35. Demystifying Reinforcement Learning Post-Training of Language Models

    Donovan Clay, Saket Gollapudi, Sankar Harilal +4

    cs.LGcs.AIcs.CLarXiv:2608.24949v12026
  36. Retrieved But Not Reliable: A Survey on Attacks, and Defenses in Retrieval-Augmented Generation

    Minh Tran, Cuong Dang, Tuc Nguyen +9

    cs.CRcs.CLcs.LGarXiv:2608.24977v12026
  37. Modulating early visual processing by language

    Harm de Vries, Florian Strub, Jérémie Mary +3

    cs.CVcs.CLcs.LGarXiv:1707.00683v32017
  38. Chain-of-Verification Reduces Hallucination in Large Language Models

    Shehzaad Dhuliawala, Mojtaba Komeili, Jing Xu +4

    cs.CLcs.AIarXiv:2309.11495v22023
  39. FinRiskAtlas: Decision-Aligned Evaluation of Large Language Models for Financial Risk Review

    Suyang Zhong, Jingzhe Zhu, Qi Xu +5

    cs.AIcs.CLarXiv:2608.25325v12026
  40. Understanding and Improving Layer Normalization

    Jingjing Xu, Xu Sun, Zhiyuan Zhang +2

    cs.LGcs.CLstat.MLarXiv:1911.07013v12019
  41. MathAdv: What Theorem Provers Know, Reason, Formalize, and Generalize

    Jiaxin Yuan, Connor Martinez Lockhart, Xiaoyu Liu +11

    cs.CLcs.AIcs.LOarXiv:2608.25449v12026
  42. The Changing Geometry of Grammar: Dimensionality and Neighborhood Reorganization across Transformer Layers

    Samuele Vallisa, Federico Ravenda, Claudio Palominos +5

    cs.CLarXiv:2608.25166v12026
  43. Belief Cascades Drive Persuasion in LLM Agent Networks

    Haoyi Qiu, Genglin Liu, Pranav Narayanan Venkit +4

    cs.CLcs.AIarXiv:2608.25152v12026
  44. MC-CXR: A Multi-Context Chest X-ray Benchmark for Context-Induced Disruption in Vision-Language Models

    Junhyeok Lee, Songsoo Kim, Kyu Sung Choi

    cs.CLarXiv:2608.24118v12026
  45. Experts, Errors, and Context: A Large-Scale Study of Human Evaluation for Machine Translation

    Markus Freitag, George Foster, David Grangier +3

    cs.CLcs.AIcs.LGarXiv:2104.14478v12021
  46. LibriBrain100: One Hundred Hours of Broad and Deep MEG Data for Neural Speech Decoding at Scale

    Francesco Mantegna, Dulhan Jayalath, Gereon Elvers +12

    cs.LGcs.CLarXiv:2608.25204v12026
  47. Apples to Apples? Towards Comparable Crosslingual Language Model Evaluation

    Xiulin Yang, Ethan Gotlieb Wilcox, Catherine Arnett

    cs.CLarXiv:2608.25089v12026
  48. Tensor2Tensor for Neural Machine Translation

    Ashish Vaswani, Samy Bengio, Eugene Brevdo +10

    cs.LGcs.CLstat.MLarXiv:1803.07416v12018
  49. Massively Multilingual Neural Machine Translation

    Roee Aharoni, Melvin Johnson, Orhan Firat

    cs.CLarXiv:1903.00089v32019
  50. Retrieve, Match, Escalate: Accurate and Scalable Product Linking with VLM-Distilled Cross-Encoders and Agentic VLMs

    Jian Wang, Steven Xu, Sanjyot Thete +5

    cs.AIcs.CLcs.DBarXiv:2608.25037v12026
  51. RefLAM: A Reference-Grounded Line Annotation Pipeline for Historical Arabic Manuscripts

    Mohamed Guechaoui, Mohamed Diaa Zellagui, Souleyman Chaib +1

    cs.CVcs.CLarXiv:2608.25140v12026
  52. Plans You Can Check: Verifier-Grounded Learning of an Open-Weight Planner for Executable Video-Editing

    Haoyu Wang, Cheng Feng, Liuyang Bian +5

    cs.CVcs.CLarXiv:2608.25622v12026
  53. Learning What to Share and What to Personalize: Hierarchical Strategy Co-Evolution for Agent Memory

    Yupeng Han, Shuochen Liu, Kai Zhang +3

    cs.AIcs.CLarXiv:2608.25329v12026
  54. ClueWeaver: Reward-Guided Dual-Agent Evidence Reasoning for Compact LLMs on Literary Long Narratives

    Jihao Zhu, Zhiwei Yang, Wenxiao Zhang +7

    cs.CLarXiv:2608.25531v12026
  55. SelfGraphRAG: Bridging the Supervision Gap in Graph-Based RAG with Synthetic QA Generation

    Ben Lagnese, Manas Gaur

    cs.CLcs.AIarXiv:2608.25123v12026
  56. A Survey on Text Classification: From Shallow to Deep Learning

    Qian Li, Hao Peng, Jianxin Li +5

    cs.CLarXiv:2008.00364v62020
  57. EgoArgus: Benchmarking VLMs as Situational Assistants for Modality-Grounded User Supports

    Yu-Chien Tang, Yu-Hsiang Liu, An-Zi Yen

    cs.CLarXiv:2608.25561v12026
  58. OmniPhys: A Unified Multimodal Benchmark for Physics Understanding and Generation from Chinese Educational Corpora

    Hao Chen, Yumin Lin, Nadila Yushanjiang +2

    cs.CLarXiv:2608.25398v12026
  59. Leveraging Speech Acts for Low-Data and Cross-Domain Conversation Derailment Forecasting

    Angela Yifei Yuan, Christine De Kock, Christopher Leckie

    cs.CLarXiv:2608.25359v12026
  60. Can We Read the Mind of an Audio LLM? A Verbalizable, Multilingual Middle-Layer Workspace

    Jiajun Fan, Jingyuan Li, Prashanth Gurunath Shivakumar +8

    cs.SDcs.AIcs.CLarXiv:2608.24958v12026