Artificial Intelligence

Papers filed under cs.AI on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

8,881 to 8,940 of 15,439

  1. OpenCodeInterpreter: Integrating Code Generation with Execution and Refinement

    Tianyu Zheng, Ge Zhang, Tianhao Shen +5

    cs.SEcs.AIcs.CLarXiv:2402.14658v32024
  2. GPT3Mix: Leveraging Large-scale Language Models for Text Augmentation

    Kang Min Yoo, Dongju Park, Jaewook Kang +2

    cs.CLcs.AIarXiv:2104.08826v22021
  3. Deliberative Alignment: Reasoning Enables Safer Language Models

    Melody Y. Guan, Manas Joglekar, Eric Wallace +12

    cs.CLcs.AIcs.CYarXiv:2412.16339v22024
  4. Automated Design of Agentic Systems

    Shengran Hu, Cong Lu, Jeff Clune

    cs.AIarXiv:2408.08435v22024
  5. SafeDecoding: Defending against Jailbreak Attacks via Safety-Aware Decoding

    Zhangchen Xu, Fengqing Jiang, Luyao Niu +3

    cs.CRcs.AIcs.CLarXiv:2402.08983v42024
  6. FinanceBench: A New Benchmark for Financial Question Answering

    Pranab Islam, Anand Kannappan, Douwe Kiela +3

    cs.CLcs.AIcs.CEarXiv:2311.11944v12023
  7. AutoDroid: LLM-powered Task Automation in Android

    Hao Wen, Yuanchun Li, Guohong Liu +7

    cs.AIcs.SEarXiv:2308.15272v42023
  8. RUDDER: Return Decomposition for Delayed Rewards

    Jose A. Arjona-Medina, Michael Gillhofer, Michael Widrich +3

    cs.LGcs.AImath.OCarXiv:1806.07857v32018
  9. Chasing Sparsity in Vision Transformers: An End-to-End Exploration

    Tianlong Chen, Yu Cheng, Zhe Gan +3

    cs.CVcs.AIarXiv:2106.04533v32021
  10. The Disagreement Problem in Explainable Machine Learning: A Practitioner's Perspective

    Satyapriya Krishna, Tessa Han, Alex Gu +3

    cs.LGcs.AIarXiv:2202.01602v62022
  11. Mastering the Game of Stratego with Model-Free Multiagent Reinforcement Learning

    Julien Perolat, Bart de Vylder, Daniel Hennes +31

    cs.AIcs.GTcs.MAarXiv:2206.15378v12022
  12. Transfer learning for time series classification

    Hassan Ismail Fawaz, Germain Forestier, Jonathan Weber +2

    cs.LGcs.AIstat.MLarXiv:1811.01533v12018
  13. Multilingual Speech Recognition With A Single End-To-End Model

    Shubham Toshniwal, Tara N. Sainath, Ron J. Weiss +4

    eess.AScs.AIcs.CLarXiv:1711.01694v22017
  14. Predicting the Computational Cost of Deep Learning Models

    Daniel Justus, John Brennan, Stephen Bonner +1

    cs.LGcs.AIstat.MLarXiv:1811.11880v12018
  15. Deep Reinforcement Learning and the Deadly Triad

    Hado van Hasselt, Yotam Doron, Florian Strub +3

    cs.AIcs.LGarXiv:1812.02648v12018
  16. R-Judge: Benchmarking Safety Risk Awareness for LLM Agents

    Tongxin Yuan, Zhiwei He, Lingzhong Dong +9

    cs.CLcs.AIarXiv:2401.10019v32024
  17. "Other-Play" for Zero-Shot Coordination

    Hengyuan Hu, Adam Lerer, Alex Peysakhovich +1

    cs.AIarXiv:2003.02979v32020
  18. What to talk about and how? Selective Generation using LSTMs with Coarse-to-Fine Alignment

    Hongyuan Mei, Mohit Bansal, Matthew R. Walter

    cs.CLcs.AIcs.LGarXiv:1509.00838v22015
  19. Learning Invariant Feature Spaces to Transfer Skills with Reinforcement Learning

    Abhishek Gupta, Coline Devin, YuXuan Liu +2

    cs.AIcs.ROarXiv:1703.02949v12017
  20. AugGPT: Leveraging ChatGPT for Text Data Augmentation

    Haixing Dai, Zhengliang Liu, Wenxiong Liao +15

    cs.CLcs.AIcs.LGarXiv:2302.13007v32023
  21. Language Models for Image Captioning: The Quirks and What Works

    Jacob Devlin, Hao Cheng, Hao Fang +5

    cs.CLcs.AIcs.CVarXiv:1505.01809v32015
  22. Controlling Overestimation Bias with Truncated Mixture of Continuous Distributional Quantile Critics

    Arsenii Kuznetsov, Pavel Shvechikov, Alexander Grishin +1

    cs.LGcs.AIstat.MLarXiv:2005.04269v12020
  23. The Hardware Lottery

    Sara Hooker

    cs.CYcs.AIcs.ARarXiv:2009.06489v22020
  24. What can Large Language Models do in chemistry? A comprehensive benchmark on eight tasks

    Taicheng Guo, Kehan Guo, Bozhao Nan +5

    cs.CLcs.AIarXiv:2305.18365v32023
  25. Supporting Qualitative Analysis with Large Language Models: Combining Codebook with GPT-3 for Deductive Coding

    Ziang Xiao, Xingdi Yuan, Q. Vera Liao +2

    cs.CLcs.AIcs.HCarXiv:2304.10548v12023
  26. A Symbolic Approach to Explaining Bayesian Network Classifiers

    Andy Shih, Arthur Choi, Adnan Darwiche

    cs.AIcs.LGarXiv:1805.03364v12018
  27. Detecting Emergent Intersectional Biases: Contextualized Word Embeddings Contain a Distribution of Human-like Biases

    Wei Guo, Aylin Caliskan

    cs.CYcs.AIcs.CLarXiv:2006.03955v52020
  28. A Survey of Zero-shot Generalisation in Deep Reinforcement Learning

    Robert Kirk, Amy Zhang, Edward Grefenstette +1

    cs.LGcs.AIarXiv:2111.09794v62021
  29. Episodic Curiosity through Reachability

    Nikolay Savinov, Anton Raichuk, Raphaël Marinier +4

    cs.LGcs.AIcs.CVarXiv:1810.02274v52018
  30. The KFIoU Loss for Rotated Object Detection

    Xue Yang, Yue Zhou, Gefan Zhang +5

    cs.CVcs.AIcs.LGarXiv:2201.12558v62022
  31. 3D Infomax improves GNNs for Molecular Property Prediction

    Hannes Stärk, Dominique Beaini, Gabriele Corso +4

    cs.LGcs.AIq-bio.BMarXiv:2110.04126v42021
  32. A Joint Speaker-Listener-Reinforcer Model for Referring Expressions

    Licheng Yu, Hao Tan, Mohit Bansal +1

    cs.CVcs.AIcs.CLarXiv:1612.09542v22016
  33. Delving into Out-of-Distribution Detection with Vision-Language Representations

    Yifei Ming, Ziyang Cai, Jiuxiang Gu +3

    cs.CVcs.AIcs.LGarXiv:2211.13445v12022
  34. Active Example Selection for In-Context Learning

    Yiming Zhang, Shi Feng, Chenhao Tan

    cs.CLcs.AIarXiv:2211.04486v12022
  35. NO Need to Worry about Adversarial Examples in Object Detection in Autonomous Vehicles

    Jiajun Lu, Hussein Sibai, Evan Fabry +1

    cs.CVcs.AIcs.CRarXiv:1707.03501v12017
  36. Automatically Correcting Large Language Models: Surveying the landscape of diverse self-correction strategies

    Liangming Pan, Michael Saxon, Wenda Xu +3

    cs.CLcs.AIcs.LGarXiv:2308.03188v22023
  37. Predictive Biases in Natural Language Processing Models: A Conceptual Framework and Overview

    Deven Shah, H. Andrew Schwartz, Dirk Hovy

    cs.CLcs.AIcs.LGarXiv:1912.11078v22019
  38. Model Merging in LLMs, MLLMs, and Beyond: Methods, Theories, Applications and Opportunities

    Enneng Yang, Li Shen, Guibing Guo +4

    cs.LGcs.AIcs.CLarXiv:2408.07666v52024
  39. Cross-Modality Attentive Feature Fusion for Object Detection in Multispectral Remote Sensing Imagery

    Qingyun Fang, Zhaokui Wang

    cs.CVcs.AIeess.IVarXiv:2112.02991v12021
  40. A Laplacian Framework for Option Discovery in Reinforcement Learning

    Marlos C. Machado, Marc G. Bellemare, Michael Bowling

    cs.LGcs.AIarXiv:1703.00956v22017
  41. Learning to Act by Predicting the Future

    Alexey Dosovitskiy, Vladlen Koltun

    cs.LGcs.AIcs.CVarXiv:1611.01779v22016
  42. Continual Unsupervised Representation Learning

    Dushyant Rao, Francesco Visin, Andrei A. Rusu +3

    cs.LGcs.AIcs.CVarXiv:1910.14481v12019
  43. Large Language Models Sensitivity to The Order of Options in Multiple-Choice Questions

    Pouya Pezeshkpour, Estevam Hruschka

    cs.CLcs.AIcs.LGarXiv:2308.11483v12023
  44. A Survey of Domain Adaptation for Neural Machine Translation

    Chenhui Chu, Rui Wang

    cs.CLcs.AIcs.LGarXiv:1806.00258v12018
  45. RADAR: Robust AI-Text Detection via Adversarial Learning

    Xiaomeng Hu, Pin-Yu Chen, Tsung-Yi Ho

    cs.CLcs.AIcs.LGarXiv:2307.03838v22023
  46. Spider 2.0: Evaluating Language Models on Real-World Enterprise Text-to-SQL Workflows

    Fangyu Lei, Jixuan Chen, Yuxiao Ye +13

    cs.CLcs.AIcs.DBarXiv:2411.07763v22024
  47. Precision Health Data: Requirements, Challenges and Existing Techniques for Data Security and Privacy

    Chandra Thapa, Seyit Camtepe

    cs.CRcs.AIarXiv:2008.10733v12020
  48. AppWorld: A Controllable World of Apps and People for Benchmarking Interactive Coding Agents

    Harsh Trivedi, Tushar Khot, Mareike Hartmann +6

    cs.SEcs.AIcs.CLarXiv:2407.18901v12024
  49. Agent Hospital: A Simulacrum of Hospital with Evolvable Medical Agents

    Junkai Li, Yunghwei Lai, Weitao Li +8

    cs.AIarXiv:2405.02957v32024
  50. Dopamine: A Research Framework for Deep Reinforcement Learning

    Pablo Samuel Castro, Subhodeep Moitra, Carles Gelada +2

    cs.LGcs.AIarXiv:1812.06110v12018
  51. Towards Generalization and Simplicity in Continuous Control

    Aravind Rajeswaran, Kendall Lowrey, Emanuel Todorov +1

    cs.LGcs.AIcs.ROarXiv:1703.02660v22017
  52. Generative Models and Model Criticism via Optimized Maximum Mean Discrepancy

    Danica J. Sutherland, Hsiao-Yu Tung, Heiko Strathmann +4

    stat.MLcs.AIcs.LGarXiv:1611.04488v62016
  53. CLIMATE-FEVER: A Dataset for Verification of Real-World Climate Claims

    Thomas Diggelmann, Jordan Boyd-Graber, Jannis Bulian +2

    cs.CLcs.AIarXiv:2012.00614v22020
  54. SGFormer: Simplifying and Empowering Transformers for Large-Graph Representations

    Qitian Wu, Wentao Zhao, Chenxiao Yang +5

    cs.LGcs.AIcs.SIarXiv:2306.10759v52023
  55. A Large Self-Annotated Corpus for Sarcasm

    Mikhail Khodak, Nikunj Saunshi, Kiran Vodrahalli

    cs.CLcs.AIcs.LGarXiv:1704.05579v42017
  56. Factuality Challenges in the Era of Large Language Models

    Isabelle Augenstein, Timothy Baldwin, Meeyoung Cha +15

    cs.CLcs.AIcs.LGarXiv:2310.05189v22023
  57. Combinatorial Optimization with Physics-Inspired Graph Neural Networks

    Martin J. A. Schuetz, J. Kyle Brubaker, Helmut G. Katzgraber

    cs.LGcond-mat.dis-nncs.AIarXiv:2107.01188v22021
  58. When LLMs Meet Cybersecurity: A Systematic Literature Review

    Jie Zhang, Haoyu Bu, Hui Wen +7

    cs.CRcs.AIarXiv:2405.03644v22024
  59. TorchMD-NET: Equivariant Transformers for Neural Network based Molecular Potentials

    Philipp Thölke, Gianni De Fabritiis

    cs.LGcs.AIphysics.chem-pharXiv:2202.02541v22022
  60. GenAttack: Practical Black-box Attacks with Gradient-Free Optimization

    Moustafa Alzantot, Yash Sharma, Supriyo Chakraborty +3

    cs.LGcs.AIcs.CRarXiv:1805.11090v32018