Computation and Language

Papers filed under cs.CL on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

9,061 to 9,120 of 11,332

  1. IQuest-Coder-V1 Technical Report

    Jian Yang, Wei Zhang, Shawn Guo +35

    cs.AIcs.CLcs.SEarXiv:2603.16733v12026
  2. Fast Transformer Decoding: One Write-Head is All You Need

    Noam Shazeer

    cs.NEcs.CLcs.LGarXiv:1911.02150v12019
  3. CRITIC: Large Language Models Can Self-Correct with Tool-Interactive Critiquing

    Zhibin Gou, Zhihong Shao, Yeyun Gong +4

    cs.CLcs.AIarXiv:2305.11738v42023
  4. LatentChem: From Textual CoT to Latent Thinking in Chemical Reasoning

    Xinwu Ye, Yicheng Mao, Yuxuan Liao +16

    physics.chem-phcs.AIcs.CLarXiv:2602.07075v62026
  5. Sentence-T5: Scalable Sentence Encoders from Pre-trained Text-to-Text Models

    Jianmo Ni, Gustavo Hernández Ábrego, Noah Constant +4

    cs.CLarXiv:2108.08877v32021
  6. Likelihood-Based Reward Designs for General LLM Reasoning

    Ariel Kwiatkowski, Natasha Butt, Ismail Labiad +2

    cs.CLarXiv:2602.03979v12026
  7. Clarify User Expertise: Towards Proactive Conversational Agents Tailoring Responses to User Proficiency

    Zhihong Cao, Chen Huang

    cs.AIcs.CLarXiv:2608.22266v12026
  8. Beyond What Meets the Eye: Unveiling Situational Illusions for Multimodal Large Language Models

    Zhiming Yang, Zhuoxi Xiong, Donglin Zhou +3

    cs.AIcs.CLcs.CVarXiv:2608.22232v12026
  9. BenchPreS: A Benchmark for Context-Aware Personalized Preference Selectivity of Persistent-Memory LLMs

    Sangyeon Yoon, Sunkyoung Kim, Hyesoo Hong +5

    cs.AIcs.CLarXiv:2603.16557v12026
  10. SPARKLING: Balancing Signal Preservation and Symmetry Breaking for Width-Progressive Learning

    Qifan Yu, Xinyu Ma, Zhijian Zhuo +7

    cs.LGcs.CLarXiv:2602.02472v22026
  11. Nanbeige4.1-3B: A Small General Model that Reasons, Aligns, and Acts

    Chen Yang, Guangyue Peng, Jiaying Zhu +12

    cs.AIcs.CLarXiv:2602.13367v12026
  12. Aggregation-Aware Synthetic Text Generation Against Authorship Re-Identification

    Qian Ma, Anna Squicciarini, Sarah Rajtmajer

    cs.AIcs.CLarXiv:2608.22161v12026
  13. Data Repetition Beats Data Scaling in Long-CoT Supervised Fine-Tuning

    Dawid J. Kopiczko, Sagar Vaze, Tijmen Blankevoort +1

    cs.CLarXiv:2602.11149v22026
  14. Demystifying Reinforcement Learning for Long-Horizon Tool-Using Agents: A Comprehensive Recipe

    Xixi Wu, Qianguo Sun, Ruiyang Zhang +4

    cs.LGcs.CLarXiv:2603.21972v12026
  15. Complementary RL: Towards Efficient Experience-Driven Agent Learning

    Dilxat Muhtar, Jiashun Liu, Wei Gao +8

    cs.LGcs.CLarXiv:2603.17621v22026
  16. GISA: A Benchmark for General Information-Seeking Assistant

    Yutao Zhu, Xingshuo Zhang, Maosen Zhang +9

    cs.CLcs.AIcs.IRarXiv:2602.08543v22026
  17. NExT-GPT: Any-to-Any Multimodal LLM

    Shengqiong Wu, Hao Fei, Leigang Qu +2

    cs.AIcs.CLcs.LGarXiv:2309.05519v32023
  18. The Flexibility Trap: Rethinking the Value of Arbitrary Order in Diffusion Language Models

    Zanlin Ni, Shenzhi Wang, Yang Yue +8

    cs.CLcs.AIcs.LGarXiv:2601.15165v42026
  19. Measuring Stability and Failure Behavior in Language Models Under Structured Perturbations

    Samira Golsefid

    cs.AIcs.CLarXiv:2608.22138v12026
  20. Object Hallucination in Image Captioning

    Anna Rohrbach, Lisa Anne Hendricks, Kaylee Burns +2

    cs.CLcs.CVarXiv:1809.02156v22018
  21. Solving math word problems with process- and outcome-based feedback

    Jonathan Uesato, Nate Kushman, Ramana Kumar +6

    cs.LGcs.AIcs.CLarXiv:2211.14275v12022
  22. Using DeepSpeed and Megatron to Train Megatron-Turing NLG 530B, A Large-Scale Generative Language Model

    Shaden Smith, Mostofa Patwary, Brandon Norick +17

    cs.CLarXiv:2201.11990v32022
  23. Zero-Shot Relation Extraction via Reading Comprehension

    Omer Levy, Minjoon Seo, Eunsol Choi +1

    cs.CLcs.AIcs.LGarXiv:1706.04115v12017
  24. A Safety Report on GPT-5.2, Gemini 3 Pro, Qwen3-VL, Grok 4.1 Fast, Nano Banana Pro, and Seedream 4.5

    Xingjun Ma, Yixu Wang, Hengyuan Xu +18

    cs.AIcs.CLcs.CVarXiv:2601.10527v22026
  25. Prompt Injection attack against LLM-integrated Applications

    Yi Liu, Gelei Deng, Yuekang Li +9

    cs.CRcs.AIcs.CLarXiv:2306.05499v32023
  26. TranslateGemma Technical Report

    Mara Finkelstein, Isaac Caswell, Tobias Domhan +18

    cs.CLcs.AIarXiv:2601.09012v32026
  27. LongWoF-Bench: Evaluating EvoMap Genes for Verifiable Long-Workflow Tasks

    Xiao Zhang, Qumeng Sun, Jihao Li +4

    cs.CLarXiv:2608.23200v12026
  28. Red Teaming Language Models to Reduce Harms: Methods, Scaling Behaviors, and Lessons Learned

    Deep Ganguli, Liane Lovitt, Jackson Kernion +33

    cs.CLcs.AIcs.CYarXiv:2209.07858v22022
  29. Perceiver-Actor: A Multi-Task Transformer for Robotic Manipulation

    Mohit Shridhar, Lucas Manuelli, Dieter Fox

    cs.ROcs.AIcs.CLarXiv:2209.05451v22022
  30. One Success Isn't Reliability: Thinkingbox, a Sandbox and Benchmark for Agents in Stateful Business Workflows

    Zhuochun Li, Youngmin Ko, Ali Keramati +9

    cs.CLcs.DBarXiv:2608.19741v12026
  31. Industrial-Instruction: An End-to-End Framework for Building Instruction-Tuning and Benchmark Datasets from Industrial Technical Reports

    Parsa Bakhtiari, Hassan Bashiri, Alireza Khalilipour +2

    cs.CLarXiv:2608.22817v12026
  32. MiroEval: Benchmarking Multimodal Deep Research Agents in Process and Outcome

    Fangda Ye, Yuxin Hu, Pengxiang Zhu +19

    cs.AIcs.CLarXiv:2603.28407v12026
  33. The Neuro-Symbolic Concept Learner: Interpreting Scenes, Words, and Sentences From Natural Supervision

    Jiayuan Mao, Chuang Gan, Pushmeet Kohli +2

    cs.CVcs.AIcs.CLarXiv:1904.12584v12019
  34. Beyond the Stability-Exploration Dilemma: Environmental Regularization for LLM Policy Optimization

    Xianlei Zhou, Xiangdi Meng, Yu He +7

    cs.CLarXiv:2608.23311v12026
  35. SciCoQA: Quality Assurance for Scientific Paper--Code Alignment

    Tim Baumgärtner, Iryna Gurevych

    cs.CLcs.AIarXiv:2601.12910v32026
  36. MeKi: Memory-based Expert Knowledge Injection for Efficient LLM Scaling

    Ning Ding, Fangcheng Liu, Kyungrae Kim +4

    cs.LGcs.AIcs.CLarXiv:2602.03359v12026
  37. Matching the Blanks: Distributional Similarity for Relation Learning

    Livio Baldini Soares, Nicholas FitzGerald, Jeffrey Ling +1

    cs.CLcs.AIarXiv:1906.03158v12019
  38. On the Scaling of PEFT: Towards Million Personal Models of Trillion Parameters

    Mind Lab, :, Vin Bo +64

    cs.LGcs.CLarXiv:2606.02437v22026
  39. Praxy Voice: Voice-Prompt Recovery + BUPS for Commercial-Class Indic TTS from a Frozen Non-Indic Base at Zero Commercial-Training-Data Cost

    Venkata Pushpak Teja Menta

    cs.SDcs.CLeess.ASarXiv:2604.25441v12026
  40. FAAST: Forward-Only Associative Learning via Closed-Form Fast Weights for Test-Time Supervised Adaptation

    Guangsheng Bao, Hongbo Zhang, Han Cui +4

    cs.LGcs.CLarXiv:2605.04651v22026
  41. Natural Language Processing (almost) from Scratch

    Ronan Collobert, Jason Weston, Leon Bottou +3

    cs.LGcs.CLarXiv:1103.0398v12011
    Summaries:한국어
  42. How Do Agents Fail on AutoResearch: End-to-End Diagnostic Evaluation on 100 Real-World Frontier Research Tasks

    Yanlin Fei, Nazhou Liu, Xinmiao Yu +6

    cs.CLarXiv:2608.14905v12026
  43. Learning to Retrieve from Agent Trajectories

    Yuqi Zhou, Sunhao Dai, Changle Qu +3

    cs.IRcs.AIcs.CLarXiv:2604.04949v12026
  44. Medusa: Simple LLM Inference Acceleration Framework with Multiple Decoding Heads

    Tianle Cai, Yuhong Li, Zhengyang Geng +4

    cs.LGcs.CLarXiv:2401.10774v32024
  45. Beyond the Grid: Layout-Informed Multi-Vector Retrieval with Parsed Visual Document Representations

    Yibo Yan, Mingdong Ou, Yi Cao +6

    cs.CLcs.IRarXiv:2603.01666v12026
  46. Post-LayerNorm Is Back: Stable, ExpressivE, and Deep

    Chen Chen, Lai Wei

    cs.LGcs.CLarXiv:2601.19895v22026
  47. Learning When to Act or Refuse: Guarding Agentic Reasoning Models for Safe Multi-Step Tool Use

    Aradhye Agarwal, Gurdit Siyan, Yash Pandya +3

    cs.CLarXiv:2603.03205v22026
  48. Table-as-Search: Formulate Long-Horizon Agentic Information Seeking as Table Completion

    Tian Lan, Felix Henry, Bin Zhu +7

    cs.CLarXiv:2602.06724v12026
  49. Computer Environments Elicit General Agentic Intelligence in LLMs

    Daixuan Cheng, Shaohan Huang, Yuxian Gu +6

    cs.CLcs.AIarXiv:2601.16206v32026
  50. MMR-Life: Piecing Together Real-life Scenes for Multimodal Multi-image Reasoning

    Jiachun Li, Shaoping Huang, Zhuoran Jin +5

    cs.CLcs.AIcs.CVarXiv:2603.02024v12026
  51. AdaReasoner: Dynamic Tool Orchestration for Iterative Visual Reasoning

    Mingyang Song, Haoyu Sun, Jiawei Gu +4

    cs.AIcs.CLcs.CVarXiv:2601.18631v22026
  52. Benchmarks Saturate When The Model Gets Smarter Than The Judge

    Marthe Ballon, Andres Algaba, Brecht Verbeken +1

    cs.AIcs.CLcs.LGarXiv:2601.19532v12026
  53. Precise Zero-Shot Dense Retrieval without Relevance Labels

    Luyu Gao, Xueguang Ma, Jimmy Lin +1

    cs.IRcs.CLarXiv:2212.10496v12022
  54. Just Ask for Calibration: Strategies for Eliciting Calibrated Confidence Scores from Language Models Fine-Tuned with Human Feedback

    Katherine Tian, Eric Mitchell, Allan Zhou +5

    cs.CLarXiv:2305.14975v22023
  55. OVD: On-policy Verbal Distillation

    Jing Xiong, Hui Shen, Shansan Gong +7

    cs.CLarXiv:2601.21968v12026
  56. AutoDAN: Generating Stealthy Jailbreak Prompts on Aligned Large Language Models

    Xiaogeng Liu, Nan Xu, Muhao Chen +1

    cs.CLcs.AIarXiv:2310.04451v22023
  57. Training LLMs for Divide-and-Conquer Reasoning Elevates Test-Time Scalability

    Xiao Liang, Zhong-Zhi Li, Zhenghao Lin +7

    cs.CLarXiv:2602.02477v12026
  58. Reading, Not Thinking: Understanding and Bridging the Modality Gap When Text Becomes Pixels in Multimodal LLMs

    Kaiser Sun, Xiaochuang Yuan, Hongjun Liu +4

    cs.CLcs.CVarXiv:2603.09095v32026
  59. Black-box Generation of Adversarial Text Sequences to Evade Deep Learning Classifiers

    Ji Gao, Jack Lanchantin, Mary Lou Soffa +1

    cs.CLcs.CRcs.IRarXiv:1801.04354v52018
  60. OpenLID-v3: Improving the Precision of Closely Related Language Identification -- An Experience Report

    Mariia Fedorova, Nikolay Arefyev, Maja Buljan +4

    cs.CLarXiv:2602.13139v42026