Computation and Language
Papers filed under cs.CL on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
9,061 to 9,120 of 11,332
IQuest-Coder-V1 Technical Report
Jian Yang, Wei Zhang, Shawn Guo +35
cs.AIcs.CLcs.SEarXiv:2603.16733v12026Fast Transformer Decoding: One Write-Head is All You Need
Noam Shazeer
cs.NEcs.CLcs.LGarXiv:1911.02150v12019CRITIC: Large Language Models Can Self-Correct with Tool-Interactive Critiquing
Zhibin Gou, Zhihong Shao, Yeyun Gong +4
cs.CLcs.AIarXiv:2305.11738v42023LatentChem: From Textual CoT to Latent Thinking in Chemical Reasoning
Xinwu Ye, Yicheng Mao, Yuxuan Liao +16
physics.chem-phcs.AIcs.CLarXiv:2602.07075v62026Sentence-T5: Scalable Sentence Encoders from Pre-trained Text-to-Text Models
Jianmo Ni, Gustavo Hernández Ábrego, Noah Constant +4
cs.CLarXiv:2108.08877v32021Likelihood-Based Reward Designs for General LLM Reasoning
Ariel Kwiatkowski, Natasha Butt, Ismail Labiad +2
cs.CLarXiv:2602.03979v12026Clarify User Expertise: Towards Proactive Conversational Agents Tailoring Responses to User Proficiency
Zhihong Cao, Chen Huang
cs.AIcs.CLarXiv:2608.22266v12026Beyond What Meets the Eye: Unveiling Situational Illusions for Multimodal Large Language Models
Zhiming Yang, Zhuoxi Xiong, Donglin Zhou +3
cs.AIcs.CLcs.CVarXiv:2608.22232v12026BenchPreS: A Benchmark for Context-Aware Personalized Preference Selectivity of Persistent-Memory LLMs
Sangyeon Yoon, Sunkyoung Kim, Hyesoo Hong +5
cs.AIcs.CLarXiv:2603.16557v12026SPARKLING: Balancing Signal Preservation and Symmetry Breaking for Width-Progressive Learning
Qifan Yu, Xinyu Ma, Zhijian Zhuo +7
cs.LGcs.CLarXiv:2602.02472v22026Nanbeige4.1-3B: A Small General Model that Reasons, Aligns, and Acts
Chen Yang, Guangyue Peng, Jiaying Zhu +12
cs.AIcs.CLarXiv:2602.13367v12026Aggregation-Aware Synthetic Text Generation Against Authorship Re-Identification
Qian Ma, Anna Squicciarini, Sarah Rajtmajer
cs.AIcs.CLarXiv:2608.22161v12026Data Repetition Beats Data Scaling in Long-CoT Supervised Fine-Tuning
Dawid J. Kopiczko, Sagar Vaze, Tijmen Blankevoort +1
cs.CLarXiv:2602.11149v22026Demystifying Reinforcement Learning for Long-Horizon Tool-Using Agents: A Comprehensive Recipe
Xixi Wu, Qianguo Sun, Ruiyang Zhang +4
cs.LGcs.CLarXiv:2603.21972v12026Complementary RL: Towards Efficient Experience-Driven Agent Learning
Dilxat Muhtar, Jiashun Liu, Wei Gao +8
cs.LGcs.CLarXiv:2603.17621v22026GISA: A Benchmark for General Information-Seeking Assistant
Yutao Zhu, Xingshuo Zhang, Maosen Zhang +9
cs.CLcs.AIcs.IRarXiv:2602.08543v22026NExT-GPT: Any-to-Any Multimodal LLM
Shengqiong Wu, Hao Fei, Leigang Qu +2
cs.AIcs.CLcs.LGarXiv:2309.05519v32023The Flexibility Trap: Rethinking the Value of Arbitrary Order in Diffusion Language Models
Zanlin Ni, Shenzhi Wang, Yang Yue +8
cs.CLcs.AIcs.LGarXiv:2601.15165v42026Measuring Stability and Failure Behavior in Language Models Under Structured Perturbations
Samira Golsefid
cs.AIcs.CLarXiv:2608.22138v12026Object Hallucination in Image Captioning
Anna Rohrbach, Lisa Anne Hendricks, Kaylee Burns +2
cs.CLcs.CVarXiv:1809.02156v22018Solving math word problems with process- and outcome-based feedback
Jonathan Uesato, Nate Kushman, Ramana Kumar +6
cs.LGcs.AIcs.CLarXiv:2211.14275v12022Using DeepSpeed and Megatron to Train Megatron-Turing NLG 530B, A Large-Scale Generative Language Model
Shaden Smith, Mostofa Patwary, Brandon Norick +17
cs.CLarXiv:2201.11990v32022Zero-Shot Relation Extraction via Reading Comprehension
Omer Levy, Minjoon Seo, Eunsol Choi +1
cs.CLcs.AIcs.LGarXiv:1706.04115v12017A Safety Report on GPT-5.2, Gemini 3 Pro, Qwen3-VL, Grok 4.1 Fast, Nano Banana Pro, and Seedream 4.5
Xingjun Ma, Yixu Wang, Hengyuan Xu +18
cs.AIcs.CLcs.CVarXiv:2601.10527v22026Prompt Injection attack against LLM-integrated Applications
Yi Liu, Gelei Deng, Yuekang Li +9
cs.CRcs.AIcs.CLarXiv:2306.05499v32023TranslateGemma Technical Report
Mara Finkelstein, Isaac Caswell, Tobias Domhan +18
cs.CLcs.AIarXiv:2601.09012v32026LongWoF-Bench: Evaluating EvoMap Genes for Verifiable Long-Workflow Tasks
Xiao Zhang, Qumeng Sun, Jihao Li +4
cs.CLarXiv:2608.23200v12026Red Teaming Language Models to Reduce Harms: Methods, Scaling Behaviors, and Lessons Learned
Deep Ganguli, Liane Lovitt, Jackson Kernion +33
cs.CLcs.AIcs.CYarXiv:2209.07858v22022Perceiver-Actor: A Multi-Task Transformer for Robotic Manipulation
Mohit Shridhar, Lucas Manuelli, Dieter Fox
cs.ROcs.AIcs.CLarXiv:2209.05451v22022One Success Isn't Reliability: Thinkingbox, a Sandbox and Benchmark for Agents in Stateful Business Workflows
Zhuochun Li, Youngmin Ko, Ali Keramati +9
cs.CLcs.DBarXiv:2608.19741v12026Industrial-Instruction: An End-to-End Framework for Building Instruction-Tuning and Benchmark Datasets from Industrial Technical Reports
Parsa Bakhtiari, Hassan Bashiri, Alireza Khalilipour +2
cs.CLarXiv:2608.22817v12026MiroEval: Benchmarking Multimodal Deep Research Agents in Process and Outcome
Fangda Ye, Yuxin Hu, Pengxiang Zhu +19
cs.AIcs.CLarXiv:2603.28407v12026The Neuro-Symbolic Concept Learner: Interpreting Scenes, Words, and Sentences From Natural Supervision
Jiayuan Mao, Chuang Gan, Pushmeet Kohli +2
cs.CVcs.AIcs.CLarXiv:1904.12584v12019Beyond the Stability-Exploration Dilemma: Environmental Regularization for LLM Policy Optimization
Xianlei Zhou, Xiangdi Meng, Yu He +7
cs.CLarXiv:2608.23311v12026SciCoQA: Quality Assurance for Scientific Paper--Code Alignment
Tim Baumgärtner, Iryna Gurevych
cs.CLcs.AIarXiv:2601.12910v32026MeKi: Memory-based Expert Knowledge Injection for Efficient LLM Scaling
Ning Ding, Fangcheng Liu, Kyungrae Kim +4
cs.LGcs.AIcs.CLarXiv:2602.03359v12026Matching the Blanks: Distributional Similarity for Relation Learning
Livio Baldini Soares, Nicholas FitzGerald, Jeffrey Ling +1
cs.CLcs.AIarXiv:1906.03158v12019On the Scaling of PEFT: Towards Million Personal Models of Trillion Parameters
Mind Lab, :, Vin Bo +64
cs.LGcs.CLarXiv:2606.02437v22026Praxy Voice: Voice-Prompt Recovery + BUPS for Commercial-Class Indic TTS from a Frozen Non-Indic Base at Zero Commercial-Training-Data Cost
Venkata Pushpak Teja Menta
cs.SDcs.CLeess.ASarXiv:2604.25441v12026FAAST: Forward-Only Associative Learning via Closed-Form Fast Weights for Test-Time Supervised Adaptation
Guangsheng Bao, Hongbo Zhang, Han Cui +4
cs.LGcs.CLarXiv:2605.04651v22026Natural Language Processing (almost) from Scratch
Ronan Collobert, Jason Weston, Leon Bottou +3
cs.LGcs.CLarXiv:1103.0398v12011Summaries:한국어How Do Agents Fail on AutoResearch: End-to-End Diagnostic Evaluation on 100 Real-World Frontier Research Tasks
Yanlin Fei, Nazhou Liu, Xinmiao Yu +6
cs.CLarXiv:2608.14905v12026Learning to Retrieve from Agent Trajectories
Yuqi Zhou, Sunhao Dai, Changle Qu +3
cs.IRcs.AIcs.CLarXiv:2604.04949v12026Medusa: Simple LLM Inference Acceleration Framework with Multiple Decoding Heads
Tianle Cai, Yuhong Li, Zhengyang Geng +4
cs.LGcs.CLarXiv:2401.10774v32024Beyond the Grid: Layout-Informed Multi-Vector Retrieval with Parsed Visual Document Representations
Yibo Yan, Mingdong Ou, Yi Cao +6
cs.CLcs.IRarXiv:2603.01666v12026Post-LayerNorm Is Back: Stable, ExpressivE, and Deep
Chen Chen, Lai Wei
cs.LGcs.CLarXiv:2601.19895v22026Learning When to Act or Refuse: Guarding Agentic Reasoning Models for Safe Multi-Step Tool Use
Aradhye Agarwal, Gurdit Siyan, Yash Pandya +3
cs.CLarXiv:2603.03205v22026Table-as-Search: Formulate Long-Horizon Agentic Information Seeking as Table Completion
Tian Lan, Felix Henry, Bin Zhu +7
cs.CLarXiv:2602.06724v12026Computer Environments Elicit General Agentic Intelligence in LLMs
Daixuan Cheng, Shaohan Huang, Yuxian Gu +6
cs.CLcs.AIarXiv:2601.16206v32026MMR-Life: Piecing Together Real-life Scenes for Multimodal Multi-image Reasoning
Jiachun Li, Shaoping Huang, Zhuoran Jin +5
cs.CLcs.AIcs.CVarXiv:2603.02024v12026AdaReasoner: Dynamic Tool Orchestration for Iterative Visual Reasoning
Mingyang Song, Haoyu Sun, Jiawei Gu +4
cs.AIcs.CLcs.CVarXiv:2601.18631v22026Benchmarks Saturate When The Model Gets Smarter Than The Judge
Marthe Ballon, Andres Algaba, Brecht Verbeken +1
cs.AIcs.CLcs.LGarXiv:2601.19532v12026Precise Zero-Shot Dense Retrieval without Relevance Labels
Luyu Gao, Xueguang Ma, Jimmy Lin +1
cs.IRcs.CLarXiv:2212.10496v12022Just Ask for Calibration: Strategies for Eliciting Calibrated Confidence Scores from Language Models Fine-Tuned with Human Feedback
Katherine Tian, Eric Mitchell, Allan Zhou +5
cs.CLarXiv:2305.14975v22023OVD: On-policy Verbal Distillation
Jing Xiong, Hui Shen, Shansan Gong +7
cs.CLarXiv:2601.21968v12026AutoDAN: Generating Stealthy Jailbreak Prompts on Aligned Large Language Models
Xiaogeng Liu, Nan Xu, Muhao Chen +1
cs.CLcs.AIarXiv:2310.04451v22023Training LLMs for Divide-and-Conquer Reasoning Elevates Test-Time Scalability
Xiao Liang, Zhong-Zhi Li, Zhenghao Lin +7
cs.CLarXiv:2602.02477v12026Reading, Not Thinking: Understanding and Bridging the Modality Gap When Text Becomes Pixels in Multimodal LLMs
Kaiser Sun, Xiaochuang Yuan, Hongjun Liu +4
cs.CLcs.CVarXiv:2603.09095v32026Black-box Generation of Adversarial Text Sequences to Evade Deep Learning Classifiers
Ji Gao, Jack Lanchantin, Mary Lou Soffa +1
cs.CLcs.CRcs.IRarXiv:1801.04354v52018OpenLID-v3: Improving the Precision of Closely Related Language Identification -- An Experience Report
Mariia Fedorova, Nikolay Arefyev, Maja Buljan +4
cs.CLarXiv:2602.13139v42026