Computation and Language
Papers filed under cs.CL on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
8,881 to 8,940 of 11,259
DeepSeek LLM: Scaling Open-Source Language Models with Longtermism
DeepSeek-AI, :, Xiao Bi +85
cs.CLcs.AIcs.LGarXiv:2401.02954v12024GradMem: Learning to Write Context into Memory with Test-Time Gradient Descent
Yuri Kuratov, Matvey Kairov, Aydar Bulatov +2
cs.CLcs.LGarXiv:2603.13875v22026Elastic Attention: Test-time Adaptive Sparsity Ratios for Efficient Transformers
Zecheng Tang, Quantong Qiu, Yi Yang +6
cs.CLcs.AIarXiv:2601.17367v22026Beyond Surface Cues: Disentangling Sociocultural Signals in Multilingual LLMs
Yuanjun Feng, Tanzhou Liu, Stefan Feuerriegel +1
cs.CLcs.AIarXiv:2608.23026v12026Tulu 3: Pushing Frontiers in Open Language Model Post-Training
Nathan Lambert, Jacob Morrison, Valentina Pyatkin +20
cs.CLarXiv:2411.15124v52024How Auditory Knowledge in LLM Backbones Shapes Audio Language Models: A Holistic Evaluation
Ke-Han Lu, Szu-Wei Fu, Chao-Han Huck Yang +13
eess.AScs.CLcs.SDarXiv:2603.19195v12026Distilling Conversations: Abstract Compression of Conversational Audio Context for LLM-based ASR
Shashi Kumar, Esaú Villatoro-Tello, Sergio Burdisso +7
cs.CLcs.AIcs.LGarXiv:2603.26246v12026PRiSM: Benchmarking Phone Realization in Speech Models
Shikhar Bharadwaj, Chin-Jou Li, Yoonjae Kim +13
cs.CLcs.SDarXiv:2601.14046v22026TaBERT: Pretraining for Joint Understanding of Textual and Tabular Data
Pengcheng Yin, Graham Neubig, Wen-tau Yih +1
cs.CLcs.LGarXiv:2005.08314v12020Style Transfer from Non-Parallel Text by Cross-Alignment
Tianxiao Shen, Tao Lei, Regina Barzilay +1
cs.CLcs.LGarXiv:1705.09655v22017Mixture-of-Depths Attention
Lianghui Zhu, Yuxin Fang, Bencheng Liao +10
cs.CLcs.AIarXiv:2603.15619v12026Accelerating Diffusion Language Models via Structured Suffix Modeling
Zifeng Cheng, Keda Li, Zhiwei Jiang +3
cs.CLarXiv:2608.23167v12026A Novel Embedding Model for Knowledge Base Completion Based on Convolutional Neural Network
Dai Quoc Nguyen, Tu Dinh Nguyen, Dat Quoc Nguyen +1
cs.CLarXiv:1712.02121v22017A decoder-only foundation model for time-series forecasting
Abhimanyu Das, Weihao Kong, Rajat Sen +1
cs.CLcs.AIcs.LGarXiv:2310.10688v42023Beyond Scalar Rewards: Dense Feedback for LLM Policy Synthesis in Sequential Social Dilemmas
Víctor Gallego
cs.CLcs.GTarXiv:2603.19453v32026FROST: Filtering Reasoning Outliers with Attention for Efficient Reasoning
Haozheng Luo, Zhuolin Jiang, Md Zahid Hasan +2
cs.CLcs.AIcs.LGarXiv:2601.19001v22026Self-Improving Pretraining: using post-trained models to pretrain better models
Ellen Xiaoqing Tan, Jack Lanchantin, Shehzaad Dhuliawala +9
cs.CLcs.AIcs.LGarXiv:2601.21343v32026SEAR: Schema-Based Evaluation and Routing for LLM Gateways
Zecheng Zhang, Han Zheng, Yue Xu
cs.DBcs.AIcs.CLarXiv:2603.26728v12026Extract Free Dense Labels from CLIP
Chong Zhou, Chen Change Loy, Bo Dai
cs.CVcs.CLarXiv:2112.01071v22021VideoCLIP: Contrastive Pre-training for Zero-shot Video-Text Understanding
Hu Xu, Gargi Ghosh, Po-Yao Huang +5
cs.CVcs.CLarXiv:2109.14084v22021Using the Output Embedding to Improve Language Models
Ofir Press, Lior Wolf
cs.CLarXiv:1608.05859v32016Multi-task Sequence to Sequence Learning
Minh-Thang Luong, Quoc V. Le, Ilya Sutskever +2
cs.LGcs.CLstat.MLarXiv:1511.06114v42015Latent Chain-of-Thought as Planning: Decoupling Reasoning from Verbalization
Jiecong Wang, Hao Peng, Chunyang Liu
cs.AIcs.CLarXiv:2601.21358v22026UnifiedQA: Crossing Format Boundaries With a Single QA System
Daniel Khashabi, Sewon Min, Tushar Khot +4
cs.CLcs.AIarXiv:2005.00700v32020Benchmarking Large Language Models for News Summarization
Tianyi Zhang, Faisal Ladhak, Esin Durmus +3
cs.CLcs.AIcs.LGarXiv:2301.13848v12023Knowledge is Not Enough: Injecting RL Skills for Continual Adaptation
Pingzhi Tang, Yiding Wang, Muhan Zhang
cs.LGcs.AIcs.CLarXiv:2601.11258v22026MovieQA: Understanding Stories in Movies through Question-Answering
Makarand Tapaswi, Yukun Zhu, Rainer Stiefelhagen +3
cs.CVcs.CLarXiv:1512.02902v22015Aligning Agentic World Models via Knowledgeable Experience Learning
Baochang Ren, Yunzhi Yao, Rui Sun +3
cs.CLcs.AIcs.CVarXiv:2601.13247v12026S2D2: Fast Decoding for Diffusion LLMs via Training-Free Self-Speculation
Ligong Han, Hao Wang, Han Gao +2
cs.CLarXiv:2603.25702v22026A Comprehensive Survey of AI-Generated Content (AIGC): A History of Generative AI from GAN to ChatGPT
Yihan Cao, Siyu Li, Yixin Liu +4
cs.AIcs.CLcs.LGarXiv:2303.04226v12023Beyond Unimodal Shortcuts: MLLMs as Cross-Modal Reasoners for Grounded Named Entity Recognition
Jinlong Ma, Yu Zhang, Xuefeng Bai +5
cs.CLarXiv:2602.04486v12026DataFlex: A Unified Framework for Data-Centric Dynamic Training of Large Language Models
Hao Liang, Zhengyang Zhao, Meiyi Qiang +22
cs.LGcs.CLarXiv:2603.26164v12026Learning to Commit: Generating Organic Pull Requests via Online Repository Memory
Mo Li, L. H. Xu, Qitai Tan +2
cs.SEcs.CLarXiv:2603.26664v12026Learning Deep Structure-Preserving Image-Text Embeddings
Liwei Wang, Yin Li, Svetlana Lazebnik
cs.CVcs.CLcs.LGarXiv:1511.06078v22015Expectations and Practices around AI Disclosure in CS Research
Arati Mohapatra, Danish Pruthi
cs.CYcs.CLcs.HCarXiv:2608.23271v12026KG-BERT: BERT for Knowledge Graph Completion
Liang Yao, Chengsheng Mao, Yuan Luo
cs.CLcs.AIarXiv:1909.03193v22019Enrich-Retrieve-Rank: Scaling Capability Discovery Beyond In-Context Routing
Nazib Sorathiya, Daniel Zhang, Bardiya Akhbari
cs.CLcs.AIcs.IRarXiv:2608.22695v12026Adapted Large Language Models Can Outperform Medical Experts in Clinical Text Summarization
Dave Van Veen, Cara Van Uden, Louis Blankemeier +16
cs.CLarXiv:2309.07430v52023SE-Bench: Benchmarking Self-Evolution with Knowledge Internalization
Jiarui Yuan, Tailin Jin, Weize Chen +1
cs.CLcs.AIcs.LGarXiv:2602.04811v22026SAGE: Benchmarking and Improving Retrieval for Deep Research Agents
Tiansheng Hu, Yilun Zhao, Canyu Zhang +2
cs.IRcs.CLarXiv:2602.05975v22026Stop the Flip-Flop: Context-Preserving Verification for Fast Revocable Diffusion Decoding
Yanzheng Xiang, Lan Wei, Yizhen Yao +8
cs.CLcs.AIarXiv:2602.06161v22026ELI5: Long Form Question Answering
Angela Fan, Yacine Jernite, Ethan Perez +3
cs.CLarXiv:1907.09190v12019SPARC: Separating Perception And Reasoning Circuits for Test-time Scaling of VLMs
Niccolo Avogaro, Nayanika Debnath, Li Mi +6
cs.CVcs.AIcs.CLarXiv:2602.06566v32026InftyThink+: Effective and Efficient Infinite-Horizon Reasoning via Reinforcement Learning
Yuchen Yan, Liang Jiang, Jin Jiang +7
cs.CLcs.AIarXiv:2602.06960v32026Balancing Understanding and Generation in Discrete Diffusion Models
Yue Liu, Yuzhong Zhao, Zheyong Xie +5
cs.CLarXiv:2602.01362v12026Linguistic Knowledge and Transferability of Contextual Representations
Nelson F. Liu, Matt Gardner, Yonatan Belinkov +2
cs.CLarXiv:1903.08855v52019DICE: Diffusion Large Language Models Excel at Generating CUDA Kernels
Haolei Bai, Lingcheng Kong, Xueyi Chen +3
cs.LGcs.CLarXiv:2602.11715v22026HateXplain: A Benchmark Dataset for Explainable Hate Speech Detection
Binny Mathew, Punyajoy Saha, Seid Muhie Yimam +3
cs.CLcs.AIcs.SIarXiv:2012.10289v22020AutoSaddler: Automatic Harness Optimization with Durable Updates from Agent Execution Traces
Sungho Park, Wonjoong Kim, Rongyuan Tan +10
cs.AIcs.CLcs.LGarXiv:2608.23041v12026Learning Modality-Specific Representations with Self-Supervised Multi-Task Learning for Multimodal Sentiment Analysis
Wenmeng Yu, Hua Xu, Ziqi Yuan +1
cs.CLarXiv:2102.04830v12021TRACE: A Self-Evolving Skill Bank for Consistent, Limit-Aware LLM Agents
Wenhao Wu, Menghao Zhang, Xin Wang +3
cs.CLcs.AIarXiv:2608.22793v12026The Devil Behind Moltbook: Anthropic Safety is Always Vanishing in Self-Evolving AI Societies
Chenxu Wang, Chaozhuo Li, Songyang Liu +10
cs.CLarXiv:2602.09877v22026When to Memorize and When to Stop: Gated Recurrent Memory for Long-Context Reasoning
Leheng Sheng, Yongtao Zhang, Wenchang Ma +6
cs.CLcs.AIarXiv:2602.10560v12026CCNet: Extracting High Quality Monolingual Datasets from Web Crawl Data
Guillaume Wenzek, Marie-Anne Lachaux, Alexis Conneau +4
cs.CLcs.IRcs.LGarXiv:1911.00359v22019MiniCPM-SALA: Hybridizing Sparse and Linear Attention for Efficient Long-Context Modeling
MiniCPM Team, Wenhao An, Yingfa Chen +44
cs.CLcs.AIcs.LGarXiv:2602.11761v22026Privasis: Synthesizing the Largest "Public" Private Dataset from Scratch
Hyunwoo Kim, Niloofar Mireshghallah, Michael Duan +11
cs.CLcs.AIarXiv:2602.03183v12026InnoEval: On Research Idea Evaluation as a Knowledge-Grounded, Multi-Perspective Reasoning Problem
Shuofei Qiao, Yunxiang Wei, Xuehai Wang +10
cs.CLcs.AIcs.IRarXiv:2602.14367v22026Cost-Efficient RAG for Entity Matching with LLMs: A Blocking-based Exploration
Chuangtao Ma, Zeyu Zhang, Arijit Khan +2
cs.DBcs.CLarXiv:2602.05708v12026Latent Thoughts Tuning: Bridging Context and Reasoning with Fused Information in Latent Tokens
Weihao Liu, Dehai Min, Lu Cheng
cs.CLarXiv:2602.10229v22026Molecular LLM Agents: From Architectural Design to Scientific Autonomy
Jiatong Li, Wengyu Zhang, Weida Wang +8
cs.CLcs.AIarXiv:2608.23104v12026