Artificial Intelligence
Papers filed under cs.AI on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
8,881 to 8,940 of 15,439
OpenCodeInterpreter: Integrating Code Generation with Execution and Refinement
Tianyu Zheng, Ge Zhang, Tianhao Shen +5
cs.SEcs.AIcs.CLarXiv:2402.14658v32024GPT3Mix: Leveraging Large-scale Language Models for Text Augmentation
Kang Min Yoo, Dongju Park, Jaewook Kang +2
cs.CLcs.AIarXiv:2104.08826v22021Deliberative Alignment: Reasoning Enables Safer Language Models
Melody Y. Guan, Manas Joglekar, Eric Wallace +12
cs.CLcs.AIcs.CYarXiv:2412.16339v22024Automated Design of Agentic Systems
Shengran Hu, Cong Lu, Jeff Clune
cs.AIarXiv:2408.08435v22024SafeDecoding: Defending against Jailbreak Attacks via Safety-Aware Decoding
Zhangchen Xu, Fengqing Jiang, Luyao Niu +3
cs.CRcs.AIcs.CLarXiv:2402.08983v42024FinanceBench: A New Benchmark for Financial Question Answering
Pranab Islam, Anand Kannappan, Douwe Kiela +3
cs.CLcs.AIcs.CEarXiv:2311.11944v12023AutoDroid: LLM-powered Task Automation in Android
Hao Wen, Yuanchun Li, Guohong Liu +7
cs.AIcs.SEarXiv:2308.15272v42023RUDDER: Return Decomposition for Delayed Rewards
Jose A. Arjona-Medina, Michael Gillhofer, Michael Widrich +3
cs.LGcs.AImath.OCarXiv:1806.07857v32018Chasing Sparsity in Vision Transformers: An End-to-End Exploration
Tianlong Chen, Yu Cheng, Zhe Gan +3
cs.CVcs.AIarXiv:2106.04533v32021The Disagreement Problem in Explainable Machine Learning: A Practitioner's Perspective
Satyapriya Krishna, Tessa Han, Alex Gu +3
cs.LGcs.AIarXiv:2202.01602v62022Mastering the Game of Stratego with Model-Free Multiagent Reinforcement Learning
Julien Perolat, Bart de Vylder, Daniel Hennes +31
cs.AIcs.GTcs.MAarXiv:2206.15378v12022Transfer learning for time series classification
Hassan Ismail Fawaz, Germain Forestier, Jonathan Weber +2
cs.LGcs.AIstat.MLarXiv:1811.01533v12018Multilingual Speech Recognition With A Single End-To-End Model
Shubham Toshniwal, Tara N. Sainath, Ron J. Weiss +4
eess.AScs.AIcs.CLarXiv:1711.01694v22017Predicting the Computational Cost of Deep Learning Models
Daniel Justus, John Brennan, Stephen Bonner +1
cs.LGcs.AIstat.MLarXiv:1811.11880v12018Deep Reinforcement Learning and the Deadly Triad
Hado van Hasselt, Yotam Doron, Florian Strub +3
cs.AIcs.LGarXiv:1812.02648v12018R-Judge: Benchmarking Safety Risk Awareness for LLM Agents
Tongxin Yuan, Zhiwei He, Lingzhong Dong +9
cs.CLcs.AIarXiv:2401.10019v32024"Other-Play" for Zero-Shot Coordination
Hengyuan Hu, Adam Lerer, Alex Peysakhovich +1
cs.AIarXiv:2003.02979v32020What to talk about and how? Selective Generation using LSTMs with Coarse-to-Fine Alignment
Hongyuan Mei, Mohit Bansal, Matthew R. Walter
cs.CLcs.AIcs.LGarXiv:1509.00838v22015Learning Invariant Feature Spaces to Transfer Skills with Reinforcement Learning
Abhishek Gupta, Coline Devin, YuXuan Liu +2
cs.AIcs.ROarXiv:1703.02949v12017AugGPT: Leveraging ChatGPT for Text Data Augmentation
Haixing Dai, Zhengliang Liu, Wenxiong Liao +15
cs.CLcs.AIcs.LGarXiv:2302.13007v32023Language Models for Image Captioning: The Quirks and What Works
Jacob Devlin, Hao Cheng, Hao Fang +5
cs.CLcs.AIcs.CVarXiv:1505.01809v32015Controlling Overestimation Bias with Truncated Mixture of Continuous Distributional Quantile Critics
Arsenii Kuznetsov, Pavel Shvechikov, Alexander Grishin +1
cs.LGcs.AIstat.MLarXiv:2005.04269v12020The Hardware Lottery
Sara Hooker
cs.CYcs.AIcs.ARarXiv:2009.06489v22020What can Large Language Models do in chemistry? A comprehensive benchmark on eight tasks
Taicheng Guo, Kehan Guo, Bozhao Nan +5
cs.CLcs.AIarXiv:2305.18365v32023Supporting Qualitative Analysis with Large Language Models: Combining Codebook with GPT-3 for Deductive Coding
Ziang Xiao, Xingdi Yuan, Q. Vera Liao +2
cs.CLcs.AIcs.HCarXiv:2304.10548v12023A Symbolic Approach to Explaining Bayesian Network Classifiers
Andy Shih, Arthur Choi, Adnan Darwiche
cs.AIcs.LGarXiv:1805.03364v12018Detecting Emergent Intersectional Biases: Contextualized Word Embeddings Contain a Distribution of Human-like Biases
Wei Guo, Aylin Caliskan
cs.CYcs.AIcs.CLarXiv:2006.03955v52020A Survey of Zero-shot Generalisation in Deep Reinforcement Learning
Robert Kirk, Amy Zhang, Edward Grefenstette +1
cs.LGcs.AIarXiv:2111.09794v62021Episodic Curiosity through Reachability
Nikolay Savinov, Anton Raichuk, Raphaël Marinier +4
cs.LGcs.AIcs.CVarXiv:1810.02274v52018The KFIoU Loss for Rotated Object Detection
Xue Yang, Yue Zhou, Gefan Zhang +5
cs.CVcs.AIcs.LGarXiv:2201.12558v620223D Infomax improves GNNs for Molecular Property Prediction
Hannes Stärk, Dominique Beaini, Gabriele Corso +4
cs.LGcs.AIq-bio.BMarXiv:2110.04126v42021A Joint Speaker-Listener-Reinforcer Model for Referring Expressions
Licheng Yu, Hao Tan, Mohit Bansal +1
cs.CVcs.AIcs.CLarXiv:1612.09542v22016Delving into Out-of-Distribution Detection with Vision-Language Representations
Yifei Ming, Ziyang Cai, Jiuxiang Gu +3
cs.CVcs.AIcs.LGarXiv:2211.13445v12022Active Example Selection for In-Context Learning
Yiming Zhang, Shi Feng, Chenhao Tan
cs.CLcs.AIarXiv:2211.04486v12022NO Need to Worry about Adversarial Examples in Object Detection in Autonomous Vehicles
Jiajun Lu, Hussein Sibai, Evan Fabry +1
cs.CVcs.AIcs.CRarXiv:1707.03501v12017Automatically Correcting Large Language Models: Surveying the landscape of diverse self-correction strategies
Liangming Pan, Michael Saxon, Wenda Xu +3
cs.CLcs.AIcs.LGarXiv:2308.03188v22023Predictive Biases in Natural Language Processing Models: A Conceptual Framework and Overview
Deven Shah, H. Andrew Schwartz, Dirk Hovy
cs.CLcs.AIcs.LGarXiv:1912.11078v22019Model Merging in LLMs, MLLMs, and Beyond: Methods, Theories, Applications and Opportunities
Enneng Yang, Li Shen, Guibing Guo +4
cs.LGcs.AIcs.CLarXiv:2408.07666v52024Cross-Modality Attentive Feature Fusion for Object Detection in Multispectral Remote Sensing Imagery
Qingyun Fang, Zhaokui Wang
cs.CVcs.AIeess.IVarXiv:2112.02991v12021A Laplacian Framework for Option Discovery in Reinforcement Learning
Marlos C. Machado, Marc G. Bellemare, Michael Bowling
cs.LGcs.AIarXiv:1703.00956v22017Learning to Act by Predicting the Future
Alexey Dosovitskiy, Vladlen Koltun
cs.LGcs.AIcs.CVarXiv:1611.01779v22016Continual Unsupervised Representation Learning
Dushyant Rao, Francesco Visin, Andrei A. Rusu +3
cs.LGcs.AIcs.CVarXiv:1910.14481v12019Large Language Models Sensitivity to The Order of Options in Multiple-Choice Questions
Pouya Pezeshkpour, Estevam Hruschka
cs.CLcs.AIcs.LGarXiv:2308.11483v12023A Survey of Domain Adaptation for Neural Machine Translation
Chenhui Chu, Rui Wang
cs.CLcs.AIcs.LGarXiv:1806.00258v12018RADAR: Robust AI-Text Detection via Adversarial Learning
Xiaomeng Hu, Pin-Yu Chen, Tsung-Yi Ho
cs.CLcs.AIcs.LGarXiv:2307.03838v22023Spider 2.0: Evaluating Language Models on Real-World Enterprise Text-to-SQL Workflows
Fangyu Lei, Jixuan Chen, Yuxiao Ye +13
cs.CLcs.AIcs.DBarXiv:2411.07763v22024Precision Health Data: Requirements, Challenges and Existing Techniques for Data Security and Privacy
Chandra Thapa, Seyit Camtepe
cs.CRcs.AIarXiv:2008.10733v12020AppWorld: A Controllable World of Apps and People for Benchmarking Interactive Coding Agents
Harsh Trivedi, Tushar Khot, Mareike Hartmann +6
cs.SEcs.AIcs.CLarXiv:2407.18901v12024Agent Hospital: A Simulacrum of Hospital with Evolvable Medical Agents
Junkai Li, Yunghwei Lai, Weitao Li +8
cs.AIarXiv:2405.02957v32024Dopamine: A Research Framework for Deep Reinforcement Learning
Pablo Samuel Castro, Subhodeep Moitra, Carles Gelada +2
cs.LGcs.AIarXiv:1812.06110v12018Towards Generalization and Simplicity in Continuous Control
Aravind Rajeswaran, Kendall Lowrey, Emanuel Todorov +1
cs.LGcs.AIcs.ROarXiv:1703.02660v22017Generative Models and Model Criticism via Optimized Maximum Mean Discrepancy
Danica J. Sutherland, Hsiao-Yu Tung, Heiko Strathmann +4
stat.MLcs.AIcs.LGarXiv:1611.04488v62016CLIMATE-FEVER: A Dataset for Verification of Real-World Climate Claims
Thomas Diggelmann, Jordan Boyd-Graber, Jannis Bulian +2
cs.CLcs.AIarXiv:2012.00614v22020SGFormer: Simplifying and Empowering Transformers for Large-Graph Representations
Qitian Wu, Wentao Zhao, Chenxiao Yang +5
cs.LGcs.AIcs.SIarXiv:2306.10759v52023A Large Self-Annotated Corpus for Sarcasm
Mikhail Khodak, Nikunj Saunshi, Kiran Vodrahalli
cs.CLcs.AIcs.LGarXiv:1704.05579v42017Factuality Challenges in the Era of Large Language Models
Isabelle Augenstein, Timothy Baldwin, Meeyoung Cha +15
cs.CLcs.AIcs.LGarXiv:2310.05189v22023Combinatorial Optimization with Physics-Inspired Graph Neural Networks
Martin J. A. Schuetz, J. Kyle Brubaker, Helmut G. Katzgraber
cs.LGcond-mat.dis-nncs.AIarXiv:2107.01188v22021When LLMs Meet Cybersecurity: A Systematic Literature Review
Jie Zhang, Haoyu Bu, Hui Wen +7
cs.CRcs.AIarXiv:2405.03644v22024TorchMD-NET: Equivariant Transformers for Neural Network based Molecular Potentials
Philipp Thölke, Gianni De Fabritiis
cs.LGcs.AIphysics.chem-pharXiv:2202.02541v22022GenAttack: Practical Black-box Attacks with Gradient-Free Optimization
Moustafa Alzantot, Yash Sharma, Supriyo Chakraborty +3
cs.LGcs.AIcs.CRarXiv:1805.11090v32018