Artificial Intelligence
Papers filed under cs.AI on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
8,701 to 8,760 of 15,247
Multilingual Speech Recognition With A Single End-To-End Model
Shubham Toshniwal, Tara N. Sainath, Ron J. Weiss +4
eess.AScs.AIcs.CLarXiv:1711.01694v22017Predicting the Computational Cost of Deep Learning Models
Daniel Justus, John Brennan, Stephen Bonner +1
cs.LGcs.AIstat.MLarXiv:1811.11880v12018Deep Reinforcement Learning and the Deadly Triad
Hado van Hasselt, Yotam Doron, Florian Strub +3
cs.AIcs.LGarXiv:1812.02648v12018R-Judge: Benchmarking Safety Risk Awareness for LLM Agents
Tongxin Yuan, Zhiwei He, Lingzhong Dong +9
cs.CLcs.AIarXiv:2401.10019v32024"Other-Play" for Zero-Shot Coordination
Hengyuan Hu, Adam Lerer, Alex Peysakhovich +1
cs.AIarXiv:2003.02979v32020What to talk about and how? Selective Generation using LSTMs with Coarse-to-Fine Alignment
Hongyuan Mei, Mohit Bansal, Matthew R. Walter
cs.CLcs.AIcs.LGarXiv:1509.00838v22015Learning Invariant Feature Spaces to Transfer Skills with Reinforcement Learning
Abhishek Gupta, Coline Devin, YuXuan Liu +2
cs.AIcs.ROarXiv:1703.02949v12017AugGPT: Leveraging ChatGPT for Text Data Augmentation
Haixing Dai, Zhengliang Liu, Wenxiong Liao +15
cs.CLcs.AIcs.LGarXiv:2302.13007v32023Language Models for Image Captioning: The Quirks and What Works
Jacob Devlin, Hao Cheng, Hao Fang +5
cs.CLcs.AIcs.CVarXiv:1505.01809v32015Controlling Overestimation Bias with Truncated Mixture of Continuous Distributional Quantile Critics
Arsenii Kuznetsov, Pavel Shvechikov, Alexander Grishin +1
cs.LGcs.AIstat.MLarXiv:2005.04269v12020The Hardware Lottery
Sara Hooker
cs.CYcs.AIcs.ARarXiv:2009.06489v22020What can Large Language Models do in chemistry? A comprehensive benchmark on eight tasks
Taicheng Guo, Kehan Guo, Bozhao Nan +5
cs.CLcs.AIarXiv:2305.18365v32023Supporting Qualitative Analysis with Large Language Models: Combining Codebook with GPT-3 for Deductive Coding
Ziang Xiao, Xingdi Yuan, Q. Vera Liao +2
cs.CLcs.AIcs.HCarXiv:2304.10548v12023A Symbolic Approach to Explaining Bayesian Network Classifiers
Andy Shih, Arthur Choi, Adnan Darwiche
cs.AIcs.LGarXiv:1805.03364v12018Detecting Emergent Intersectional Biases: Contextualized Word Embeddings Contain a Distribution of Human-like Biases
Wei Guo, Aylin Caliskan
cs.CYcs.AIcs.CLarXiv:2006.03955v52020A Survey of Zero-shot Generalisation in Deep Reinforcement Learning
Robert Kirk, Amy Zhang, Edward Grefenstette +1
cs.LGcs.AIarXiv:2111.09794v62021Episodic Curiosity through Reachability
Nikolay Savinov, Anton Raichuk, Raphaël Marinier +4
cs.LGcs.AIcs.CVarXiv:1810.02274v52018The KFIoU Loss for Rotated Object Detection
Xue Yang, Yue Zhou, Gefan Zhang +5
cs.CVcs.AIcs.LGarXiv:2201.12558v620223D Infomax improves GNNs for Molecular Property Prediction
Hannes Stärk, Dominique Beaini, Gabriele Corso +4
cs.LGcs.AIq-bio.BMarXiv:2110.04126v42021A Joint Speaker-Listener-Reinforcer Model for Referring Expressions
Licheng Yu, Hao Tan, Mohit Bansal +1
cs.CVcs.AIcs.CLarXiv:1612.09542v22016Delving into Out-of-Distribution Detection with Vision-Language Representations
Yifei Ming, Ziyang Cai, Jiuxiang Gu +3
cs.CVcs.AIcs.LGarXiv:2211.13445v12022Active Example Selection for In-Context Learning
Yiming Zhang, Shi Feng, Chenhao Tan
cs.CLcs.AIarXiv:2211.04486v12022NO Need to Worry about Adversarial Examples in Object Detection in Autonomous Vehicles
Jiajun Lu, Hussein Sibai, Evan Fabry +1
cs.CVcs.AIcs.CRarXiv:1707.03501v12017Automatically Correcting Large Language Models: Surveying the landscape of diverse self-correction strategies
Liangming Pan, Michael Saxon, Wenda Xu +3
cs.CLcs.AIcs.LGarXiv:2308.03188v22023Predictive Biases in Natural Language Processing Models: A Conceptual Framework and Overview
Deven Shah, H. Andrew Schwartz, Dirk Hovy
cs.CLcs.AIcs.LGarXiv:1912.11078v22019Model Merging in LLMs, MLLMs, and Beyond: Methods, Theories, Applications and Opportunities
Enneng Yang, Li Shen, Guibing Guo +4
cs.LGcs.AIcs.CLarXiv:2408.07666v52024Cross-Modality Attentive Feature Fusion for Object Detection in Multispectral Remote Sensing Imagery
Qingyun Fang, Zhaokui Wang
cs.CVcs.AIeess.IVarXiv:2112.02991v12021A Laplacian Framework for Option Discovery in Reinforcement Learning
Marlos C. Machado, Marc G. Bellemare, Michael Bowling
cs.LGcs.AIarXiv:1703.00956v22017Learning to Act by Predicting the Future
Alexey Dosovitskiy, Vladlen Koltun
cs.LGcs.AIcs.CVarXiv:1611.01779v22016Continual Unsupervised Representation Learning
Dushyant Rao, Francesco Visin, Andrei A. Rusu +3
cs.LGcs.AIcs.CVarXiv:1910.14481v12019Large Language Models Sensitivity to The Order of Options in Multiple-Choice Questions
Pouya Pezeshkpour, Estevam Hruschka
cs.CLcs.AIcs.LGarXiv:2308.11483v12023A Survey of Domain Adaptation for Neural Machine Translation
Chenhui Chu, Rui Wang
cs.CLcs.AIcs.LGarXiv:1806.00258v12018RADAR: Robust AI-Text Detection via Adversarial Learning
Xiaomeng Hu, Pin-Yu Chen, Tsung-Yi Ho
cs.CLcs.AIcs.LGarXiv:2307.03838v22023Spider 2.0: Evaluating Language Models on Real-World Enterprise Text-to-SQL Workflows
Fangyu Lei, Jixuan Chen, Yuxiao Ye +13
cs.CLcs.AIcs.DBarXiv:2411.07763v22024Precision Health Data: Requirements, Challenges and Existing Techniques for Data Security and Privacy
Chandra Thapa, Seyit Camtepe
cs.CRcs.AIarXiv:2008.10733v12020AppWorld: A Controllable World of Apps and People for Benchmarking Interactive Coding Agents
Harsh Trivedi, Tushar Khot, Mareike Hartmann +6
cs.SEcs.AIcs.CLarXiv:2407.18901v12024Agent Hospital: A Simulacrum of Hospital with Evolvable Medical Agents
Junkai Li, Yunghwei Lai, Weitao Li +8
cs.AIarXiv:2405.02957v32024Dopamine: A Research Framework for Deep Reinforcement Learning
Pablo Samuel Castro, Subhodeep Moitra, Carles Gelada +2
cs.LGcs.AIarXiv:1812.06110v12018Towards Generalization and Simplicity in Continuous Control
Aravind Rajeswaran, Kendall Lowrey, Emanuel Todorov +1
cs.LGcs.AIcs.ROarXiv:1703.02660v22017Generative Models and Model Criticism via Optimized Maximum Mean Discrepancy
Danica J. Sutherland, Hsiao-Yu Tung, Heiko Strathmann +4
stat.MLcs.AIcs.LGarXiv:1611.04488v62016CLIMATE-FEVER: A Dataset for Verification of Real-World Climate Claims
Thomas Diggelmann, Jordan Boyd-Graber, Jannis Bulian +2
cs.CLcs.AIarXiv:2012.00614v22020SGFormer: Simplifying and Empowering Transformers for Large-Graph Representations
Qitian Wu, Wentao Zhao, Chenxiao Yang +5
cs.LGcs.AIcs.SIarXiv:2306.10759v52023A Large Self-Annotated Corpus for Sarcasm
Mikhail Khodak, Nikunj Saunshi, Kiran Vodrahalli
cs.CLcs.AIcs.LGarXiv:1704.05579v42017Factuality Challenges in the Era of Large Language Models
Isabelle Augenstein, Timothy Baldwin, Meeyoung Cha +15
cs.CLcs.AIcs.LGarXiv:2310.05189v22023Combinatorial Optimization with Physics-Inspired Graph Neural Networks
Martin J. A. Schuetz, J. Kyle Brubaker, Helmut G. Katzgraber
cs.LGcond-mat.dis-nncs.AIarXiv:2107.01188v22021When LLMs Meet Cybersecurity: A Systematic Literature Review
Jie Zhang, Haoyu Bu, Hui Wen +7
cs.CRcs.AIarXiv:2405.03644v22024TorchMD-NET: Equivariant Transformers for Neural Network based Molecular Potentials
Philipp Thölke, Gianni De Fabritiis
cs.LGcs.AIphysics.chem-pharXiv:2202.02541v22022GenAttack: Practical Black-box Attacks with Gradient-Free Optimization
Moustafa Alzantot, Yash Sharma, Supriyo Chakraborty +3
cs.LGcs.AIcs.CRarXiv:1805.11090v32018tinyBenchmarks: evaluating LLMs with fewer examples
Felipe Maia Polo, Lucas Weber, Leshem Choshen +3
cs.CLcs.AIcs.LGarXiv:2402.14992v22024Navigation World Models
Amir Bar, Gaoyue Zhou, Danny Tran +2
cs.CVcs.AIcs.LGarXiv:2412.03572v22024AgentBoard: An Analytical Evaluation Board of Multi-turn LLM Agents
Chang Ma, Junlei Zhang, Zhihao Zhu +6
cs.CLcs.AIcs.LGarXiv:2401.13178v22024Net2Vec: Quantifying and Explaining how Concepts are Encoded by Filters in Deep Neural Networks
Ruth Fong, Andrea Vedaldi
cs.CVcs.AIstat.MLarXiv:1801.03454v22018DialogueCRN: Contextual Reasoning Networks for Emotion Recognition in Conversations
Dou Hu, Lingwei Wei, Xiaoyong Huai
cs.CLcs.AIarXiv:2106.01978v22021Advancements in Image Classification using Convolutional Neural Network
Farhana Sultana, A. Sufian, Paramartha Dutta
cs.CVcs.AIarXiv:1905.03288v12019Highway Long Short-Term Memory RNNs for Distant Speech Recognition
Yu Zhang, Guoguo Chen, Dong Yu +3
cs.NEcs.AIcs.CLarXiv:1510.08983v22015Attention, please! A survey of Neural Attention Models in Deep Learning
Alana de Santana Correia, Esther Luna Colombini
cs.LGcs.AIcs.CVarXiv:2103.16775v12021Domain Specialization as the Key to Make Large Language Models Disruptive: A Comprehensive Survey
Chen Ling, Xujiang Zhao, Jiaying Lu +21
cs.CLcs.AIarXiv:2305.18703v72023Unified Contrastive Learning in Image-Text-Label Space
Jianwei Yang, Chunyuan Li, Pengchuan Zhang +4
cs.CVcs.AIcs.LGarXiv:2204.03610v12022Deep learning generalizes because the parameter-function map is biased towards simple functions
Guillermo Valle-Pérez, Chico Q. Camargo, Ard A. Louis
stat.MLcs.AIcs.LGarXiv:1805.08522v52018Fractional Order Fuzzy Control of Hybrid Power System with Renewable Generation Using Chaotic PSO
Indranil Pan, Saptarshi Das
eess.SYcs.AImath.OCarXiv:1611.09809v12016