Artificial Intelligence

Papers filed under cs.AI on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

8,701 to 8,760 of 15,247

  1. Multilingual Speech Recognition With A Single End-To-End Model

    Shubham Toshniwal, Tara N. Sainath, Ron J. Weiss +4

    eess.AScs.AIcs.CLarXiv:1711.01694v22017
  2. Predicting the Computational Cost of Deep Learning Models

    Daniel Justus, John Brennan, Stephen Bonner +1

    cs.LGcs.AIstat.MLarXiv:1811.11880v12018
  3. Deep Reinforcement Learning and the Deadly Triad

    Hado van Hasselt, Yotam Doron, Florian Strub +3

    cs.AIcs.LGarXiv:1812.02648v12018
  4. R-Judge: Benchmarking Safety Risk Awareness for LLM Agents

    Tongxin Yuan, Zhiwei He, Lingzhong Dong +9

    cs.CLcs.AIarXiv:2401.10019v32024
  5. "Other-Play" for Zero-Shot Coordination

    Hengyuan Hu, Adam Lerer, Alex Peysakhovich +1

    cs.AIarXiv:2003.02979v32020
  6. What to talk about and how? Selective Generation using LSTMs with Coarse-to-Fine Alignment

    Hongyuan Mei, Mohit Bansal, Matthew R. Walter

    cs.CLcs.AIcs.LGarXiv:1509.00838v22015
  7. Learning Invariant Feature Spaces to Transfer Skills with Reinforcement Learning

    Abhishek Gupta, Coline Devin, YuXuan Liu +2

    cs.AIcs.ROarXiv:1703.02949v12017
  8. AugGPT: Leveraging ChatGPT for Text Data Augmentation

    Haixing Dai, Zhengliang Liu, Wenxiong Liao +15

    cs.CLcs.AIcs.LGarXiv:2302.13007v32023
  9. Language Models for Image Captioning: The Quirks and What Works

    Jacob Devlin, Hao Cheng, Hao Fang +5

    cs.CLcs.AIcs.CVarXiv:1505.01809v32015
  10. Controlling Overestimation Bias with Truncated Mixture of Continuous Distributional Quantile Critics

    Arsenii Kuznetsov, Pavel Shvechikov, Alexander Grishin +1

    cs.LGcs.AIstat.MLarXiv:2005.04269v12020
  11. The Hardware Lottery

    Sara Hooker

    cs.CYcs.AIcs.ARarXiv:2009.06489v22020
  12. What can Large Language Models do in chemistry? A comprehensive benchmark on eight tasks

    Taicheng Guo, Kehan Guo, Bozhao Nan +5

    cs.CLcs.AIarXiv:2305.18365v32023
  13. Supporting Qualitative Analysis with Large Language Models: Combining Codebook with GPT-3 for Deductive Coding

    Ziang Xiao, Xingdi Yuan, Q. Vera Liao +2

    cs.CLcs.AIcs.HCarXiv:2304.10548v12023
  14. A Symbolic Approach to Explaining Bayesian Network Classifiers

    Andy Shih, Arthur Choi, Adnan Darwiche

    cs.AIcs.LGarXiv:1805.03364v12018
  15. Detecting Emergent Intersectional Biases: Contextualized Word Embeddings Contain a Distribution of Human-like Biases

    Wei Guo, Aylin Caliskan

    cs.CYcs.AIcs.CLarXiv:2006.03955v52020
  16. A Survey of Zero-shot Generalisation in Deep Reinforcement Learning

    Robert Kirk, Amy Zhang, Edward Grefenstette +1

    cs.LGcs.AIarXiv:2111.09794v62021
  17. Episodic Curiosity through Reachability

    Nikolay Savinov, Anton Raichuk, Raphaël Marinier +4

    cs.LGcs.AIcs.CVarXiv:1810.02274v52018
  18. The KFIoU Loss for Rotated Object Detection

    Xue Yang, Yue Zhou, Gefan Zhang +5

    cs.CVcs.AIcs.LGarXiv:2201.12558v62022
  19. 3D Infomax improves GNNs for Molecular Property Prediction

    Hannes Stärk, Dominique Beaini, Gabriele Corso +4

    cs.LGcs.AIq-bio.BMarXiv:2110.04126v42021
  20. A Joint Speaker-Listener-Reinforcer Model for Referring Expressions

    Licheng Yu, Hao Tan, Mohit Bansal +1

    cs.CVcs.AIcs.CLarXiv:1612.09542v22016
  21. Delving into Out-of-Distribution Detection with Vision-Language Representations

    Yifei Ming, Ziyang Cai, Jiuxiang Gu +3

    cs.CVcs.AIcs.LGarXiv:2211.13445v12022
  22. Active Example Selection for In-Context Learning

    Yiming Zhang, Shi Feng, Chenhao Tan

    cs.CLcs.AIarXiv:2211.04486v12022
  23. NO Need to Worry about Adversarial Examples in Object Detection in Autonomous Vehicles

    Jiajun Lu, Hussein Sibai, Evan Fabry +1

    cs.CVcs.AIcs.CRarXiv:1707.03501v12017
  24. Automatically Correcting Large Language Models: Surveying the landscape of diverse self-correction strategies

    Liangming Pan, Michael Saxon, Wenda Xu +3

    cs.CLcs.AIcs.LGarXiv:2308.03188v22023
  25. Predictive Biases in Natural Language Processing Models: A Conceptual Framework and Overview

    Deven Shah, H. Andrew Schwartz, Dirk Hovy

    cs.CLcs.AIcs.LGarXiv:1912.11078v22019
  26. Model Merging in LLMs, MLLMs, and Beyond: Methods, Theories, Applications and Opportunities

    Enneng Yang, Li Shen, Guibing Guo +4

    cs.LGcs.AIcs.CLarXiv:2408.07666v52024
  27. Cross-Modality Attentive Feature Fusion for Object Detection in Multispectral Remote Sensing Imagery

    Qingyun Fang, Zhaokui Wang

    cs.CVcs.AIeess.IVarXiv:2112.02991v12021
  28. A Laplacian Framework for Option Discovery in Reinforcement Learning

    Marlos C. Machado, Marc G. Bellemare, Michael Bowling

    cs.LGcs.AIarXiv:1703.00956v22017
  29. Learning to Act by Predicting the Future

    Alexey Dosovitskiy, Vladlen Koltun

    cs.LGcs.AIcs.CVarXiv:1611.01779v22016
  30. Continual Unsupervised Representation Learning

    Dushyant Rao, Francesco Visin, Andrei A. Rusu +3

    cs.LGcs.AIcs.CVarXiv:1910.14481v12019
  31. Large Language Models Sensitivity to The Order of Options in Multiple-Choice Questions

    Pouya Pezeshkpour, Estevam Hruschka

    cs.CLcs.AIcs.LGarXiv:2308.11483v12023
  32. A Survey of Domain Adaptation for Neural Machine Translation

    Chenhui Chu, Rui Wang

    cs.CLcs.AIcs.LGarXiv:1806.00258v12018
  33. RADAR: Robust AI-Text Detection via Adversarial Learning

    Xiaomeng Hu, Pin-Yu Chen, Tsung-Yi Ho

    cs.CLcs.AIcs.LGarXiv:2307.03838v22023
  34. Spider 2.0: Evaluating Language Models on Real-World Enterprise Text-to-SQL Workflows

    Fangyu Lei, Jixuan Chen, Yuxiao Ye +13

    cs.CLcs.AIcs.DBarXiv:2411.07763v22024
  35. Precision Health Data: Requirements, Challenges and Existing Techniques for Data Security and Privacy

    Chandra Thapa, Seyit Camtepe

    cs.CRcs.AIarXiv:2008.10733v12020
  36. AppWorld: A Controllable World of Apps and People for Benchmarking Interactive Coding Agents

    Harsh Trivedi, Tushar Khot, Mareike Hartmann +6

    cs.SEcs.AIcs.CLarXiv:2407.18901v12024
  37. Agent Hospital: A Simulacrum of Hospital with Evolvable Medical Agents

    Junkai Li, Yunghwei Lai, Weitao Li +8

    cs.AIarXiv:2405.02957v32024
  38. Dopamine: A Research Framework for Deep Reinforcement Learning

    Pablo Samuel Castro, Subhodeep Moitra, Carles Gelada +2

    cs.LGcs.AIarXiv:1812.06110v12018
  39. Towards Generalization and Simplicity in Continuous Control

    Aravind Rajeswaran, Kendall Lowrey, Emanuel Todorov +1

    cs.LGcs.AIcs.ROarXiv:1703.02660v22017
  40. Generative Models and Model Criticism via Optimized Maximum Mean Discrepancy

    Danica J. Sutherland, Hsiao-Yu Tung, Heiko Strathmann +4

    stat.MLcs.AIcs.LGarXiv:1611.04488v62016
  41. CLIMATE-FEVER: A Dataset for Verification of Real-World Climate Claims

    Thomas Diggelmann, Jordan Boyd-Graber, Jannis Bulian +2

    cs.CLcs.AIarXiv:2012.00614v22020
  42. SGFormer: Simplifying and Empowering Transformers for Large-Graph Representations

    Qitian Wu, Wentao Zhao, Chenxiao Yang +5

    cs.LGcs.AIcs.SIarXiv:2306.10759v52023
  43. A Large Self-Annotated Corpus for Sarcasm

    Mikhail Khodak, Nikunj Saunshi, Kiran Vodrahalli

    cs.CLcs.AIcs.LGarXiv:1704.05579v42017
  44. Factuality Challenges in the Era of Large Language Models

    Isabelle Augenstein, Timothy Baldwin, Meeyoung Cha +15

    cs.CLcs.AIcs.LGarXiv:2310.05189v22023
  45. Combinatorial Optimization with Physics-Inspired Graph Neural Networks

    Martin J. A. Schuetz, J. Kyle Brubaker, Helmut G. Katzgraber

    cs.LGcond-mat.dis-nncs.AIarXiv:2107.01188v22021
  46. When LLMs Meet Cybersecurity: A Systematic Literature Review

    Jie Zhang, Haoyu Bu, Hui Wen +7

    cs.CRcs.AIarXiv:2405.03644v22024
  47. TorchMD-NET: Equivariant Transformers for Neural Network based Molecular Potentials

    Philipp Thölke, Gianni De Fabritiis

    cs.LGcs.AIphysics.chem-pharXiv:2202.02541v22022
  48. GenAttack: Practical Black-box Attacks with Gradient-Free Optimization

    Moustafa Alzantot, Yash Sharma, Supriyo Chakraborty +3

    cs.LGcs.AIcs.CRarXiv:1805.11090v32018
  49. tinyBenchmarks: evaluating LLMs with fewer examples

    Felipe Maia Polo, Lucas Weber, Leshem Choshen +3

    cs.CLcs.AIcs.LGarXiv:2402.14992v22024
  50. Navigation World Models

    Amir Bar, Gaoyue Zhou, Danny Tran +2

    cs.CVcs.AIcs.LGarXiv:2412.03572v22024
  51. AgentBoard: An Analytical Evaluation Board of Multi-turn LLM Agents

    Chang Ma, Junlei Zhang, Zhihao Zhu +6

    cs.CLcs.AIcs.LGarXiv:2401.13178v22024
  52. Net2Vec: Quantifying and Explaining how Concepts are Encoded by Filters in Deep Neural Networks

    Ruth Fong, Andrea Vedaldi

    cs.CVcs.AIstat.MLarXiv:1801.03454v22018
  53. DialogueCRN: Contextual Reasoning Networks for Emotion Recognition in Conversations

    Dou Hu, Lingwei Wei, Xiaoyong Huai

    cs.CLcs.AIarXiv:2106.01978v22021
  54. Advancements in Image Classification using Convolutional Neural Network

    Farhana Sultana, A. Sufian, Paramartha Dutta

    cs.CVcs.AIarXiv:1905.03288v12019
  55. Highway Long Short-Term Memory RNNs for Distant Speech Recognition

    Yu Zhang, Guoguo Chen, Dong Yu +3

    cs.NEcs.AIcs.CLarXiv:1510.08983v22015
  56. Attention, please! A survey of Neural Attention Models in Deep Learning

    Alana de Santana Correia, Esther Luna Colombini

    cs.LGcs.AIcs.CVarXiv:2103.16775v12021
  57. Domain Specialization as the Key to Make Large Language Models Disruptive: A Comprehensive Survey

    Chen Ling, Xujiang Zhao, Jiaying Lu +21

    cs.CLcs.AIarXiv:2305.18703v72023
  58. Unified Contrastive Learning in Image-Text-Label Space

    Jianwei Yang, Chunyuan Li, Pengchuan Zhang +4

    cs.CVcs.AIcs.LGarXiv:2204.03610v12022
  59. Deep learning generalizes because the parameter-function map is biased towards simple functions

    Guillermo Valle-Pérez, Chico Q. Camargo, Ard A. Louis

    stat.MLcs.AIcs.LGarXiv:1805.08522v52018
  60. Fractional Order Fuzzy Control of Hybrid Power System with Renewable Generation Using Chaotic PSO

    Indranil Pan, Saptarshi Das

    eess.SYcs.AImath.OCarXiv:1611.09809v12016