Computation and Language
Papers filed under cs.CL on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
9,361 to 9,420 of 11,332
Good SFT Optimizes for SFT, Better SFT Prepares for Reinforcement Learning
Dylan Zhang, Yufeng Xu, Haojin Wang +2
cs.LGcs.AIcs.CLarXiv:2602.01058v22026HLE-Verified: A Systematic Verification and Structured Revision of Humanity's Last Exam
Weiqi Zhai, Zhihai Wang, Jinghang Wang +33
cs.CLarXiv:2602.13964v42026A Watermark for Large Language Models
John Kirchenbauer, Jonas Geiping, Yuxin Wen +3
cs.LGcs.CLcs.CRarXiv:2301.10226v42023DialogueRNN: An Attentive RNN for Emotion Detection in Conversations
Navonil Majumder, Soujanya Poria, Devamanyu Hazarika +3
cs.CLarXiv:1811.00405v42018Thinking to Recall: How Reasoning Unlocks Parametric Knowledge in LLMs
Zorik Gekhman, Roee Aharoni, Eran Ofek +3
cs.CLarXiv:2603.09906v12026FlashPrefill: Instantaneous Pattern Discovery and Thresholding for Ultra-Fast Long-Context Prefilling
Qihang Fan, Huaibo Huang, Zhiying Wu +3
cs.CLcs.AIarXiv:2603.06199v12026Transfer Learning in Biomedical Natural Language Processing: An Evaluation of BERT and ELMo on Ten Benchmarking Datasets
Yifan Peng, Shankai Yan, Zhiyong Lu
cs.CLarXiv:1906.05474v22019DoRA: Weight-Decomposed Low-Rank Adaptation
Shih-Yang Liu, Chien-Yi Wang, Hongxu Yin +4
cs.CLcs.CVarXiv:2402.09353v62024Memex(RL): Scaling Long-Horizon LLM Agents via Indexed Experience Memory
Zhenting Wang, Huancheng Chen, Jiayun Wang +1
cs.CLcs.LGarXiv:2603.04257v12026Joint CTC-Attention based End-to-End Speech Recognition using Multi-task Learning
Suyoun Kim, Takaaki Hori, Shinji Watanabe
cs.CLarXiv:1609.06773v22016FLAVA: A Foundational Language And Vision Alignment Model
Amanpreet Singh, Ronghang Hu, Vedanuj Goswami +4
cs.CVcs.CLarXiv:2112.04482v32021Unlocking Implicit Experience: Synthesizing Tool-Use Trajectories from Text
Zhihao Xu, Rumei Li, Jiahuan Li +4
cs.CLarXiv:2601.10355v12026Vectorizing the Trie: Efficient Constrained Decoding for LLM-based Generative Retrieval on Accelerators
Zhengyang Su, Isay Katsman, Yueqi Wang +10
cs.IRcs.CLcs.LGarXiv:2602.22647v22026$τ$-bench: A Benchmark for Tool-Agent-User Interaction in Real-World Domains
Shunyu Yao, Noah Shinn, Pedram Razavi +1
cs.AIcs.CLarXiv:2406.12045v12024Effective LSTMs for Target-Dependent Sentiment Classification
Duyu Tang, Bing Qin, Xiaocheng Feng +1
cs.CLarXiv:1512.01100v22015Multimodal Few-Shot Learning with Frozen Language Models
Maria Tsimpoukelli, Jacob Menick, Serkan Cabi +3
cs.CVcs.CLcs.LGarXiv:2106.13884v22021Jet-RL: Enabling On-Policy FP8 Reinforcement Learning with Unified Training and Rollout Precision Flow
Haocheng Xi, Charlie Ruan, Peiyuan Liao +7
cs.LGcs.CLarXiv:2601.14243v22026Does the Whole Exceed its Parts? The Effect of AI Explanations on Complementary Team Performance
Gagan Bansal, Tongshuang Wu, Joyce Zhou +5
cs.AIcs.CLcs.HCarXiv:2006.14779v32020Universal Transformers
Mostafa Dehghani, Stephan Gouws, Oriol Vinyals +2
cs.CLcs.LGstat.MLarXiv:1807.03819v32018According to Me: Long-Term Personalized Referential Memory QA
Jingbiao Mei, Jinghong Chen, Guangyu Yang +3
cs.AIcs.CLcs.CVarXiv:2603.01990v12026Assessing the Ability of LSTMs to Learn Syntax-Sensitive Dependencies
Tal Linzen, Emmanuel Dupoux, Yoav Goldberg
cs.CLarXiv:1611.01368v12016A Survey of Data Augmentation Approaches for NLP
Steven Y. Feng, Varun Gangal, Jason Wei +4
cs.CLcs.AIcs.LGarXiv:2105.03075v52021KromHC: Manifold-Constrained Hyper-Connections with Kronecker-Product Residual Matrices
Wuyang Zhou, Yuxuan Gu, Giorgos Iacovides +1
cs.CLcs.LGarXiv:2601.21579v22026DeepResearchEval: An Automated Framework for Deep Research Task Construction and Agentic Evaluation
Yibo Wang, Lei Wang, Yue Deng +7
cs.CLarXiv:2601.09688v12026Locate, Steer, and Improve: A Practical Survey of Actionable Mechanistic Interpretability in Large Language Models
Hengyuan Zhang, Zhihao Zhang, Mingyang Wang +26
cs.CLarXiv:2601.14004v42026BeaverTails: Towards Improved Safety Alignment of LLM via a Human-Preference Dataset
Jiaming Ji, Mickel Liu, Juntao Dai +7
cs.CLarXiv:2307.04657v32023Agentic Uncertainty Quantification
Jiaxin Zhang, Prafulla Kumar Choubey, Kung-Hsiang Huang +2
cs.AIcs.CLarXiv:2601.15703v12026RLAnything: Forge Environment, Policy, and Reward Model in Completely Dynamic RL System
Yinjie Wang, Tianbao Xie, Ke Shen +2
cs.LGcs.AIcs.CLarXiv:2602.02488v12026CrowS-Pairs: A Challenge Dataset for Measuring Social Biases in Masked Language Models
Nikita Nangia, Clara Vania, Rasika Bhalerao +1
cs.CLcs.AIarXiv:2010.00133v12020OpenDecoder: Open Large Language Model Decoding to Incorporate Document Quality in RAG
Fengran Mo, Zhan Su, Yuchen Hui +6
cs.CLcs.AIcs.IRarXiv:2601.09028v22026F2LLM-v2: Inclusive, Performant, and Efficient Embeddings for a Multilingual World
Ziyin Zhang, Zihan Liao, Hang Yu +2
cs.CLcs.AIarXiv:2603.19223v12026Baichuan-M3: Modeling Clinical Inquiry for Reliable Medical Decision-Making
Baichuan-M3 Team, :, Chengfeng Dou +16
cs.CLarXiv:2602.06570v12026Stable-DiffCoder: Pushing the Frontier of Code Diffusion Large Language Model
Chenghao Fan, Wen Heng, Bo Li +6
cs.CLarXiv:2601.15892v22026MTEB: Massive Text Embedding Benchmark
Niklas Muennighoff, Nouamane Tazi, Loïc Magne +1
cs.CLcs.IRcs.LGarXiv:2210.07316v32022EVA: Exploring the Limits of Masked Visual Representation Learning at Scale
Yuxin Fang, Wen Wang, Binhui Xie +6
cs.CVcs.CLcs.LGarXiv:2211.07636v22022Spark: Strategic Policy-Aware Exploration via Dynamic Branching for Long-Horizon Agentic Learning
Jinyang Wu, Shuo Yang, Changpeng Yang +4
cs.LGcs.CLarXiv:2601.20209v22026Ignore Previous Prompt: Attack Techniques For Language Models
Fábio Perez, Ian Ribeiro
cs.CLcs.AIarXiv:2211.09527v12022Scaling Embeddings Outperforms Scaling Experts in Language Models
Hong Liu, Jiaqi Zhang, Chao Wang +13
cs.CLcs.AIcs.LGarXiv:2601.21204v22026Effective Strategies for Asynchronous Software Engineering Agents
Jiayi Geng, Graham Neubig
cs.CLcs.AIarXiv:2603.21489v22026MASS: Masked Sequence to Sequence Pre-training for Language Generation
Kaitao Song, Xu Tan, Tao Qin +2
cs.CLcs.AIcs.LGarXiv:1905.02450v52019BeyondSWE: Can Current Code Agent Survive Beyond Single-Repo Bug Fixing?
Guoxin Chen, Fanzhe Meng, Jiale Zhao +12
cs.CLcs.SEarXiv:2603.03194v22026SERA: Soft-Verified Efficient Repository Agents
Ethan Shen, Daniel Tormoen, Saurabh Shah +2
cs.CLcs.LGcs.SEarXiv:2601.20789v32026Search-R2: Enhancing Search-Integrated Reasoning via Actor-Refiner Collaboration
Bowei He, Minda Hu, Zenan Xu +7
cs.AIcs.CLarXiv:2602.03647v12026WizardCoder: Empowering Code Large Language Models with Evol-Instruct
Ziyang Luo, Can Xu, Pu Zhao +7
cs.CLcs.AIarXiv:2306.08568v22023PaLI: A Jointly-Scaled Multilingual Language-Image Model
Xi Chen, Xiao Wang, Soravit Changpinyo +26
cs.CVcs.CLarXiv:2209.06794v42022Evaluating the Factual Consistency of Abstractive Text Summarization
Wojciech Kryściński, Bryan McCann, Caiming Xiong +1
cs.CLarXiv:1910.12840v12019Forest Before Trees: Latent Superposition for Efficient Visual Reasoning
Yubo Wang, Juntian Zhang, Yichen Wu +3
cs.CLcs.CVarXiv:2601.06803v22026Program Induction by Rationale Generation : Learning to Solve and Explain Algebraic Word Problems
Wang Ling, Dani Yogatama, Chris Dyer +1
cs.AIcs.CLcs.LGarXiv:1705.04146v32017Automatic Chain of Thought Prompting in Large Language Models
Zhuosheng Zhang, Aston Zhang, Mu Li +1
cs.CLcs.AIarXiv:2210.03493v12022$V_1$: Unifying Generation and Self-Verification for Parallel Reasoners
Harman Singh, Xiuyu Li, Kusha Sareen +14
cs.CLarXiv:2603.04304v12026ERNIE: Enhanced Representation through Knowledge Integration
Yu Sun, Shuohuan Wang, Yukun Li +7
cs.CLarXiv:1904.09223v12019Baichuan 2: Open Large-scale Language Models
Aiyuan Yang, Bin Xiao, Bingning Wang +52
cs.CLarXiv:2309.10305v42023SWE-Master: Unleashing the Potential of Software Engineering Agents via Post-Training
Huatong Song, Lisheng Huang, Shuang Sun +11
cs.SEcs.CLarXiv:2602.03411v22026TweetEval: Unified Benchmark and Comparative Evaluation for Tweet Classification
Francesco Barbieri, Jose Camacho-Collados, Leonardo Neves +1
cs.CLcs.SIarXiv:2010.12421v22020Accelerating Large Language Model Decoding with Speculative Sampling
Charlie Chen, Sebastian Borgeaud, Geoffrey Irving +3
cs.CLarXiv:2302.01318v12023Diachronic Word Embeddings Reveal Statistical Laws of Semantic Change
William L. Hamilton, Jure Leskovec, Dan Jurafsky
cs.CLarXiv:1605.09096v62016GPT-NeoX-20B: An Open-Source Autoregressive Language Model
Sid Black, Stella Biderman, Eric Hallahan +14
cs.CLarXiv:2204.06745v12022Toward Controlled Generation of Text
Zhiting Hu, Zichao Yang, Xiaodan Liang +2
cs.LGcs.AIcs.CLarXiv:1703.00955v42017What you can cram into a single vector: Probing sentence embeddings for linguistic properties
Alexis Conneau, German Kruszewski, Guillaume Lample +2
cs.CLarXiv:1805.01070v22018Wizard of Wikipedia: Knowledge-Powered Conversational agents
Emily Dinan, Stephen Roller, Kurt Shuster +3
cs.CLarXiv:1811.01241v22018