Computation and Language

Papers filed under cs.CL on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

1,681 to 1,740 of 11,307

  1. SAEScientist-Bench: Can AI Agents Conduct Autonomous SAE Interpretability Research?

    Yuqiao Tan, Shizhu He, Jun Zhao +1

    cs.AIcs.CLcs.LGarXiv:2609.09113v12026
  2. Keyformer: KV Cache Reduction through Key Tokens Selection for Efficient Generative Inference

    Muhammad Adnan, Akhil Arunkumar, Gaurav Jain +3

    cs.LGcs.AIcs.ARarXiv:2403.09054v22024
  3. Good Pretraining, Bad SFT: Checkpoint Quality Across the Training Stack

    Sohir Maskey, Philipp Scholl, Jonas Knupp +2

    cs.AIcs.CLarXiv:2609.08966v12026
  4. Answer-Distribution Trajectories: A Stochastic-Dynamics View of LLM Reasoning

    Mar Gonzàlez I Català, Haitz Sáez de Ocáriz Borde, Davide Murari +3

    cs.AIcs.CLcs.ITarXiv:2609.09030v12026
  5. Recent Progresses in Deep Learning based Acoustic Models (Updated)

    Dong Yu, Jinyu Li

    eess.AScs.CLcs.SDarXiv:1804.09298v22018
  6. Towards Safer Large Language Models through Machine Unlearning

    Zheyuan Liu, Guangyao Dou, Zhaoxuan Tan +2

    cs.CLarXiv:2402.10058v22024
  7. A Survey of Code-switched Speech and Language Processing

    Sunayana Sitaram, Khyathi Raghavi Chandu, Sai Krishna Rallabandi +1

    cs.CLcs.LGstat.MLarXiv:1904.00784v32019
  8. A Three-Tier Persona Vector for Controllable User Simulation in Agentic Evaluation

    Rahul Khedar, Eshita, Sneha Teja Sree Reddy Thondapu +6

    cs.AIcs.CLarXiv:2609.08592v12026
  9. On Feature Normalization and Data Augmentation

    Boyi Li, Felix Wu, Ser-Nam Lim +2

    cs.LGcs.CLcs.CVarXiv:2002.11102v32020
  10. Free Process Rewards without Process Labels

    Lifan Yuan, Wendi Li, Huayu Chen +6

    cs.LGcs.CLarXiv:2412.01981v12024
  11. When is multitask learning effective? Semantic sequence prediction under varying data conditions

    Héctor Martínez Alonso, Barbara Plank

    cs.CLarXiv:1612.02251v22016
  12. Real-Time Open-Domain Question Answering with Dense-Sparse Phrase Index

    Minjoon Seo, Jinhyuk Lee, Tom Kwiatkowski +3

    cs.CLarXiv:1906.05807v22019
  13. Pixtral 12B

    Pravesh Agrawal, Szymon Antoniak, Emma Bou Hanna +39

    cs.CVcs.CLarXiv:2410.07073v22024
  14. Debiasing Pre-trained Contextualised Embeddings

    Masahiro Kaneko, Danushka Bollegala

    cs.CLarXiv:2101.09523v12021
  15. Tree Transformer: Integrating Tree Structures into Self-Attention

    Yau-Shian Wang, Hung-Yi Lee, Yun-Nung Chen

    cs.CLcs.LGarXiv:1909.06639v22019
  16. SE-GoS: Self-Evolving Graph-of-Skills for Skill Library at Scale

    Dawei Fu, Cheng Jiang, Sitian Qian +2

    cs.AIcs.CLarXiv:2609.08228v12026
  17. Design2Code: Benchmarking Multimodal Code Generation for Automated Front-End Engineering

    Chenglei Si, Yanzhe Zhang, Ryan Li +3

    cs.CLcs.CVcs.CYarXiv:2403.03163v32024
  18. The Argument Reasoning Comprehension Task: Identification and Reconstruction of Implicit Warrants

    Ivan Habernal, Henning Wachsmuth, Iryna Gurevych +1

    cs.CLcs.AIarXiv:1708.01425v42017
  19. Offensive Language Identification in Greek

    Zeses Pitenis, Marcos Zampieri, Tharindu Ranasinghe

    cs.CLarXiv:2003.07459v22020
  20. Breaking Sticks and Ambiguities with Adaptive Skip-gram

    Sergey Bartunov, Dmitry Kondrashkin, Anton Osokin +1

    cs.CLarXiv:1502.07257v22015
  21. Do Dynamic Routers Need Memory? HeRo: History-Aware Routing for Efficient LLM Inference

    Hongjin Lin, Wentao Wan, Keze Wang

    cs.AIcs.CLarXiv:2609.08189v12026
  22. NewsCLIPpings: Automatic Generation of Out-of-Context Multimodal Media

    Grace Luo, Trevor Darrell, Anna Rohrbach

    cs.CVcs.CLarXiv:2104.05893v22021
  23. Does Deeper Reasoning Compromise Alignment? Revealing and Mitigating of Alignment Collapse in Large Reasoning Models

    Yu-Hang Wu, Yu-Jie Xiong, Henghua Zhang +3

    cs.AIcs.CLarXiv:2609.08186v12026
  24. Trellis Networks for Sequence Modeling

    Shaojie Bai, J. Zico Kolter, Vladlen Koltun

    cs.LGcs.AIcs.CLarXiv:1810.06682v22018
  25. Multi-View Sequence-to-Sequence Models with Conversational Structure for Abstractive Dialogue Summarization

    Jiaao Chen, Diyi Yang

    cs.CLarXiv:2010.01672v12020
  26. HydraLoRA: An Asymmetric LoRA Architecture for Efficient Fine-Tuning

    Chunlin Tian, Zhan Shi, Zhijiang Guo +2

    cs.CLcs.AIarXiv:2404.19245v22024
  27. RM-Bench: Benchmarking Reward Models of Language Models with Subtlety and Style

    Yantao Liu, Zijun Yao, Rui Min +3

    cs.CLarXiv:2410.16184v12024
  28. Fully Hyperbolic Neural Networks

    Weize Chen, Xu Han, Yankai Lin +5

    cs.CLcs.LGarXiv:2105.14686v32021
  29. Are Word Embedding-based Features Useful for Sarcasm Detection?

    Aditya Joshi, Vaibhav Tripathi, Kevin Patel +2

    cs.CLarXiv:1610.00883v12016
  30. SchemeArena: Factorized Stress Testing of Scheming in LLM Agents

    Jie Ruan, Inderjeet Nair, Amy Liu +3

    cs.AIcs.CLarXiv:2609.08126v12026
  31. SpectFormer: Frequency and Attention is what you need in a Vision Transformer

    Badri N. Patro, Vinay P. Namboodiri, Vijay Srinivas Agneeswaran

    cs.CVcs.AIcs.CLarXiv:2304.06446v22023
  32. Semi-Supervised Approach to Monitoring Clinical Depressive Symptoms in Social Media

    Amir Hossein Yazdavar, Hussein S. Al-Olimat, Monireh Ebrahimi +5

    cs.CLarXiv:1710.05429v12017
  33. Latent Diffusion for Language Generation

    Justin Lovelace, Varsha Kishore, Chao Wan +2

    cs.CLcs.LGarXiv:2212.09462v22022
  34. TransModality: An End2End Fusion Method with Transformer for Multimodal Sentiment Analysis

    Zilong Wang, Zhaohong Wan, Xiaojun Wan

    cs.CLarXiv:2009.02902v22020
  35. Eliciting Self-Verification in Multimodal Reasoning Agents with Reinforcement Learning

    Vishwas Sathish, Viresh Ranjan, Xinliang Zhu +2

    cs.AIcs.CLcs.CVarXiv:2609.08025v12026
  36. Audio ALBERT: A Lite BERT for Self-supervised Learning of Audio Representation

    Po-Han Chi, Pei-Hung Chung, Tsung-Han Wu +4

    eess.AScs.CLcs.SDarXiv:2005.08575v52020
  37. Measuring Emotions in the COVID-19 Real World Worry Dataset

    Bennett Kleinberg, Isabelle van der Vegt, Maximilian Mozes

    cs.CLcs.IRcs.SIarXiv:2004.04225v22020
  38. CERT: Continual Pre-Training on Sketches for Library-Oriented Code Generation

    Daoguang Zan, Bei Chen, Dejian Yang +6

    cs.SEcs.CLcs.PLarXiv:2206.06888v12022
  39. Reinforced Multi-Teacher Selection for Knowledge Distillation

    Fei Yuan, Linjun Shou, Jian Pei +4

    cs.CLcs.LGarXiv:2012.06048v22020
  40. BEVBert: Multimodal Map Pre-training for Language-guided Navigation

    Dong An, Yuankai Qi, Yangguang Li +4

    cs.CVcs.AIcs.CLarXiv:2212.04385v22022
  41. Evaluating the Evaluation of Diversity in Natural Language Generation

    Guy Tevet, Jonathan Berant

    cs.CLarXiv:2004.02990v32020
  42. OmniACT: A Dataset and Benchmark for Enabling Multimodal Generalist Autonomous Agents for Desktop and Web

    Raghav Kapoor, Yash Parag Butala, Melisa Russak +4

    cs.AIcs.CLcs.CVarXiv:2402.17553v32024
  43. Towards Making the Most of BERT in Neural Machine Translation

    Jiacheng Yang, Mingxuan Wang, Hao Zhou +4

    cs.CLcs.LGarXiv:1908.05672v52019
  44. CausalVerify: An Execution-Grounded Benchmark for LLM Causal Inference Workflows

    Yonghong Zhang, Ricardo Correia, Isabel M. Parra +1

    cs.AIcs.CLecon.EMarXiv:2609.07944v12026
  45. Towards Robust Neural Machine Translation

    Yong Cheng, Zhaopeng Tu, Fandong Meng +2

    cs.CLarXiv:1805.06130v12018
  46. SceneVerse: Scaling 3D Vision-Language Learning for Grounded Scene Understanding

    Baoxiong Jia, Yixin Chen, Huangyue Yu +5

    cs.CVcs.AIcs.CLarXiv:2401.09340v32024
  47. Margin Matters: Towards More Discriminative Deep Neural Network Embeddings for Speaker Recognition

    Xu Xiang, Shuai Wang, Houjun Huang +2

    eess.AScs.CLcs.SDarXiv:1906.07317v12019
  48. MM-Embed: Universal Multimodal Retrieval with Multimodal LLMs

    Sheng-Chieh Lin, Chankyu Lee, Mohammad Shoeybi +3

    cs.CLcs.AIcs.CVarXiv:2411.02571v22024
  49. A Benchmark for Systematic Generalization in Grounded Language Understanding

    Laura Ruis, Jacob Andreas, Marco Baroni +2

    cs.CLcs.AIcs.LGarXiv:2003.05161v22020
  50. Long-form factuality in large language models

    Jerry Wei, Chengrun Yang, Xinying Song +9

    cs.CLcs.AIcs.LGarXiv:2403.18802v42024
  51. Text-to-SQL Generation for Question Answering on Electronic Medical Records

    Ping Wang, Tian Shi, Chandan K. Reddy

    cs.CLcs.AIcs.IRarXiv:1908.01839v22019
  52. Intermediate Loss Regularization for CTC-based Speech Recognition

    Jaesong Lee, Shinji Watanabe

    eess.AScs.CLcs.SDarXiv:2102.03216v12021
  53. Stream of Search (SoS): Learning to Search in Language

    Kanishk Gandhi, Denise Lee, Gabriel Grand +4

    cs.LGcs.AIcs.CLarXiv:2404.03683v12024
  54. Incorporating Global Visual Features into Attention-Based Neural Machine Translation

    Iacer Calixto, Qun Liu, Nick Campbell

    cs.CLarXiv:1701.06521v12017
  55. FOIL it! Find One mismatch between Image and Language caption

    Ravi Shekhar, Sandro Pezzelle, Yauhen Klimovich +4

    cs.CVcs.CLcs.MMarXiv:1705.01359v12017
  56. Efficient Methods for Natural Language Processing: A Survey

    Marcos Treviso, Ji-Ung Lee, Tianchu Ji +19

    cs.CLarXiv:2209.00099v22022
  57. The Emerging AI Paper-Review Arms Race: Adversarial Co-Evolution in Scholarly Publishing

    Chenguang Wang, Ming Li, Adebayo Braimah +6

    cs.AIcs.CLarXiv:2609.07713v12026
  58. ORLM: A Customizable Framework in Training Large Models for Automated Optimization Modeling

    Chenyu Huang, Zhengyang Tang, Shixi Hu +5

    cs.CLcs.AIcs.CEarXiv:2405.17743v52024
  59. Evidence Aggregation for Answer Re-Ranking in Open-Domain Question Answering

    Shuohang Wang, Mo Yu, Jing Jiang +7

    cs.CLcs.AIarXiv:1711.05116v22017
  60. Relational Reflection Entity Alignment

    Xin Mao, Wenting Wang, Huimin Xu +2

    cs.IRcs.CLcs.LGarXiv:2008.07962v12020