Computation and Language
Papers filed under cs.CL on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
1,681 to 1,740 of 11,307
SAEScientist-Bench: Can AI Agents Conduct Autonomous SAE Interpretability Research?
Yuqiao Tan, Shizhu He, Jun Zhao +1
cs.AIcs.CLcs.LGarXiv:2609.09113v12026Keyformer: KV Cache Reduction through Key Tokens Selection for Efficient Generative Inference
Muhammad Adnan, Akhil Arunkumar, Gaurav Jain +3
cs.LGcs.AIcs.ARarXiv:2403.09054v22024Good Pretraining, Bad SFT: Checkpoint Quality Across the Training Stack
Sohir Maskey, Philipp Scholl, Jonas Knupp +2
cs.AIcs.CLarXiv:2609.08966v12026Answer-Distribution Trajectories: A Stochastic-Dynamics View of LLM Reasoning
Mar Gonzàlez I Català, Haitz Sáez de Ocáriz Borde, Davide Murari +3
cs.AIcs.CLcs.ITarXiv:2609.09030v12026Recent Progresses in Deep Learning based Acoustic Models (Updated)
Dong Yu, Jinyu Li
eess.AScs.CLcs.SDarXiv:1804.09298v22018Towards Safer Large Language Models through Machine Unlearning
Zheyuan Liu, Guangyao Dou, Zhaoxuan Tan +2
cs.CLarXiv:2402.10058v22024A Survey of Code-switched Speech and Language Processing
Sunayana Sitaram, Khyathi Raghavi Chandu, Sai Krishna Rallabandi +1
cs.CLcs.LGstat.MLarXiv:1904.00784v32019A Three-Tier Persona Vector for Controllable User Simulation in Agentic Evaluation
Rahul Khedar, Eshita, Sneha Teja Sree Reddy Thondapu +6
cs.AIcs.CLarXiv:2609.08592v12026On Feature Normalization and Data Augmentation
Boyi Li, Felix Wu, Ser-Nam Lim +2
cs.LGcs.CLcs.CVarXiv:2002.11102v32020Free Process Rewards without Process Labels
Lifan Yuan, Wendi Li, Huayu Chen +6
cs.LGcs.CLarXiv:2412.01981v12024When is multitask learning effective? Semantic sequence prediction under varying data conditions
Héctor Martínez Alonso, Barbara Plank
cs.CLarXiv:1612.02251v22016Real-Time Open-Domain Question Answering with Dense-Sparse Phrase Index
Minjoon Seo, Jinhyuk Lee, Tom Kwiatkowski +3
cs.CLarXiv:1906.05807v22019Pixtral 12B
Pravesh Agrawal, Szymon Antoniak, Emma Bou Hanna +39
cs.CVcs.CLarXiv:2410.07073v22024Debiasing Pre-trained Contextualised Embeddings
Masahiro Kaneko, Danushka Bollegala
cs.CLarXiv:2101.09523v12021Tree Transformer: Integrating Tree Structures into Self-Attention
Yau-Shian Wang, Hung-Yi Lee, Yun-Nung Chen
cs.CLcs.LGarXiv:1909.06639v22019SE-GoS: Self-Evolving Graph-of-Skills for Skill Library at Scale
Dawei Fu, Cheng Jiang, Sitian Qian +2
cs.AIcs.CLarXiv:2609.08228v12026Design2Code: Benchmarking Multimodal Code Generation for Automated Front-End Engineering
Chenglei Si, Yanzhe Zhang, Ryan Li +3
cs.CLcs.CVcs.CYarXiv:2403.03163v32024The Argument Reasoning Comprehension Task: Identification and Reconstruction of Implicit Warrants
Ivan Habernal, Henning Wachsmuth, Iryna Gurevych +1
cs.CLcs.AIarXiv:1708.01425v42017Offensive Language Identification in Greek
Zeses Pitenis, Marcos Zampieri, Tharindu Ranasinghe
cs.CLarXiv:2003.07459v22020Breaking Sticks and Ambiguities with Adaptive Skip-gram
Sergey Bartunov, Dmitry Kondrashkin, Anton Osokin +1
cs.CLarXiv:1502.07257v22015Do Dynamic Routers Need Memory? HeRo: History-Aware Routing for Efficient LLM Inference
Hongjin Lin, Wentao Wan, Keze Wang
cs.AIcs.CLarXiv:2609.08189v12026NewsCLIPpings: Automatic Generation of Out-of-Context Multimodal Media
Grace Luo, Trevor Darrell, Anna Rohrbach
cs.CVcs.CLarXiv:2104.05893v22021Does Deeper Reasoning Compromise Alignment? Revealing and Mitigating of Alignment Collapse in Large Reasoning Models
Yu-Hang Wu, Yu-Jie Xiong, Henghua Zhang +3
cs.AIcs.CLarXiv:2609.08186v12026Trellis Networks for Sequence Modeling
Shaojie Bai, J. Zico Kolter, Vladlen Koltun
cs.LGcs.AIcs.CLarXiv:1810.06682v22018Multi-View Sequence-to-Sequence Models with Conversational Structure for Abstractive Dialogue Summarization
Jiaao Chen, Diyi Yang
cs.CLarXiv:2010.01672v12020HydraLoRA: An Asymmetric LoRA Architecture for Efficient Fine-Tuning
Chunlin Tian, Zhan Shi, Zhijiang Guo +2
cs.CLcs.AIarXiv:2404.19245v22024RM-Bench: Benchmarking Reward Models of Language Models with Subtlety and Style
Yantao Liu, Zijun Yao, Rui Min +3
cs.CLarXiv:2410.16184v12024Fully Hyperbolic Neural Networks
Weize Chen, Xu Han, Yankai Lin +5
cs.CLcs.LGarXiv:2105.14686v32021Are Word Embedding-based Features Useful for Sarcasm Detection?
Aditya Joshi, Vaibhav Tripathi, Kevin Patel +2
cs.CLarXiv:1610.00883v12016SchemeArena: Factorized Stress Testing of Scheming in LLM Agents
Jie Ruan, Inderjeet Nair, Amy Liu +3
cs.AIcs.CLarXiv:2609.08126v12026SpectFormer: Frequency and Attention is what you need in a Vision Transformer
Badri N. Patro, Vinay P. Namboodiri, Vijay Srinivas Agneeswaran
cs.CVcs.AIcs.CLarXiv:2304.06446v22023Semi-Supervised Approach to Monitoring Clinical Depressive Symptoms in Social Media
Amir Hossein Yazdavar, Hussein S. Al-Olimat, Monireh Ebrahimi +5
cs.CLarXiv:1710.05429v12017Latent Diffusion for Language Generation
Justin Lovelace, Varsha Kishore, Chao Wan +2
cs.CLcs.LGarXiv:2212.09462v22022TransModality: An End2End Fusion Method with Transformer for Multimodal Sentiment Analysis
Zilong Wang, Zhaohong Wan, Xiaojun Wan
cs.CLarXiv:2009.02902v22020Eliciting Self-Verification in Multimodal Reasoning Agents with Reinforcement Learning
Vishwas Sathish, Viresh Ranjan, Xinliang Zhu +2
cs.AIcs.CLcs.CVarXiv:2609.08025v12026Audio ALBERT: A Lite BERT for Self-supervised Learning of Audio Representation
Po-Han Chi, Pei-Hung Chung, Tsung-Han Wu +4
eess.AScs.CLcs.SDarXiv:2005.08575v52020Measuring Emotions in the COVID-19 Real World Worry Dataset
Bennett Kleinberg, Isabelle van der Vegt, Maximilian Mozes
cs.CLcs.IRcs.SIarXiv:2004.04225v22020CERT: Continual Pre-Training on Sketches for Library-Oriented Code Generation
Daoguang Zan, Bei Chen, Dejian Yang +6
cs.SEcs.CLcs.PLarXiv:2206.06888v12022Reinforced Multi-Teacher Selection for Knowledge Distillation
Fei Yuan, Linjun Shou, Jian Pei +4
cs.CLcs.LGarXiv:2012.06048v22020BEVBert: Multimodal Map Pre-training for Language-guided Navigation
Dong An, Yuankai Qi, Yangguang Li +4
cs.CVcs.AIcs.CLarXiv:2212.04385v22022Evaluating the Evaluation of Diversity in Natural Language Generation
Guy Tevet, Jonathan Berant
cs.CLarXiv:2004.02990v32020OmniACT: A Dataset and Benchmark for Enabling Multimodal Generalist Autonomous Agents for Desktop and Web
Raghav Kapoor, Yash Parag Butala, Melisa Russak +4
cs.AIcs.CLcs.CVarXiv:2402.17553v32024Towards Making the Most of BERT in Neural Machine Translation
Jiacheng Yang, Mingxuan Wang, Hao Zhou +4
cs.CLcs.LGarXiv:1908.05672v52019CausalVerify: An Execution-Grounded Benchmark for LLM Causal Inference Workflows
Yonghong Zhang, Ricardo Correia, Isabel M. Parra +1
cs.AIcs.CLecon.EMarXiv:2609.07944v12026Towards Robust Neural Machine Translation
Yong Cheng, Zhaopeng Tu, Fandong Meng +2
cs.CLarXiv:1805.06130v12018SceneVerse: Scaling 3D Vision-Language Learning for Grounded Scene Understanding
Baoxiong Jia, Yixin Chen, Huangyue Yu +5
cs.CVcs.AIcs.CLarXiv:2401.09340v32024Margin Matters: Towards More Discriminative Deep Neural Network Embeddings for Speaker Recognition
Xu Xiang, Shuai Wang, Houjun Huang +2
eess.AScs.CLcs.SDarXiv:1906.07317v12019MM-Embed: Universal Multimodal Retrieval with Multimodal LLMs
Sheng-Chieh Lin, Chankyu Lee, Mohammad Shoeybi +3
cs.CLcs.AIcs.CVarXiv:2411.02571v22024A Benchmark for Systematic Generalization in Grounded Language Understanding
Laura Ruis, Jacob Andreas, Marco Baroni +2
cs.CLcs.AIcs.LGarXiv:2003.05161v22020Long-form factuality in large language models
Jerry Wei, Chengrun Yang, Xinying Song +9
cs.CLcs.AIcs.LGarXiv:2403.18802v42024Text-to-SQL Generation for Question Answering on Electronic Medical Records
Ping Wang, Tian Shi, Chandan K. Reddy
cs.CLcs.AIcs.IRarXiv:1908.01839v22019Intermediate Loss Regularization for CTC-based Speech Recognition
Jaesong Lee, Shinji Watanabe
eess.AScs.CLcs.SDarXiv:2102.03216v12021Stream of Search (SoS): Learning to Search in Language
Kanishk Gandhi, Denise Lee, Gabriel Grand +4
cs.LGcs.AIcs.CLarXiv:2404.03683v12024Incorporating Global Visual Features into Attention-Based Neural Machine Translation
Iacer Calixto, Qun Liu, Nick Campbell
cs.CLarXiv:1701.06521v12017FOIL it! Find One mismatch between Image and Language caption
Ravi Shekhar, Sandro Pezzelle, Yauhen Klimovich +4
cs.CVcs.CLcs.MMarXiv:1705.01359v12017Efficient Methods for Natural Language Processing: A Survey
Marcos Treviso, Ji-Ung Lee, Tianchu Ji +19
cs.CLarXiv:2209.00099v22022The Emerging AI Paper-Review Arms Race: Adversarial Co-Evolution in Scholarly Publishing
Chenguang Wang, Ming Li, Adebayo Braimah +6
cs.AIcs.CLarXiv:2609.07713v12026ORLM: A Customizable Framework in Training Large Models for Automated Optimization Modeling
Chenyu Huang, Zhengyang Tang, Shixi Hu +5
cs.CLcs.AIcs.CEarXiv:2405.17743v52024Evidence Aggregation for Answer Re-Ranking in Open-Domain Question Answering
Shuohang Wang, Mo Yu, Jing Jiang +7
cs.CLcs.AIarXiv:1711.05116v22017Relational Reflection Entity Alignment
Xin Mao, Wenting Wang, Huimin Xu +2
cs.IRcs.CLcs.LGarXiv:2008.07962v12020