Artificial Intelligence

Papers filed under cs.AI on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

13,981 to 14,040 of 15,368

  1. Datasheets for Datasets

    Timnit Gebru, Jamie Morgenstern, Briana Vecchione +4

    cs.DBcs.AIcs.LGarXiv:1803.09010v82018
  2. PIQA: Reasoning about Physical Commonsense in Natural Language

    Yonatan Bisk, Rowan Zellers, Ronan Le Bras +2

    cs.CLcs.AIcs.LGarXiv:1911.11641v12019
  3. Improved Semantic Representations From Tree-Structured Long Short-Term Memory Networks

    Kai Sheng Tai, Richard Socher, Christopher D. Manning

    cs.CLcs.AIcs.LGarXiv:1503.00075v32015
  4. InstructPix2Pix: Learning to Follow Image Editing Instructions

    Tim Brooks, Aleksander Holynski, Alexei A. Efros

    cs.CVcs.AIcs.CLarXiv:2211.09800v22022
  5. ImageNet-trained CNNs are biased towards texture; increasing shape bias improves accuracy and robustness

    Robert Geirhos, Patricia Rubisch, Claudio Michaelis +3

    cs.CVcs.AIcs.LGarXiv:1811.12231v32018
  6. Federated Machine Learning: Concept and Applications

    Qiang Yang, Yang Liu, Tianjian Chen +1

    cs.AIcs.CRcs.LGarXiv:1902.04885v12019
  7. SMOPD: Selective Token-Entropy Masking for Dirty-History Multi-Turn On-Policy Self-Distillation

    Chenyang Jiang, Changhan Huang

    cs.LGcs.AIarXiv:2608.14647v12026
  8. Shortcut Learning in Deep Neural Networks

    Robert Geirhos, Jörn-Henrik Jacobsen, Claudio Michaelis +4

    cs.CVcs.AIcs.LGarXiv:2004.07780v52020
  9. An Overview of Multi-Task Learning in Deep Neural Networks

    Sebastian Ruder

    cs.LGcs.AIstat.MLarXiv:1706.05098v12017
  10. The Arcade Learning Environment: An Evaluation Platform for General Agents

    Marc G. Bellemare, Yavar Naddaf, Joel Veness +1

    cs.AIarXiv:1207.4708v22012
  11. CARA: Cognitive Adaptive Recommendation Agent

    Weijun Gao, Jinyang Dong, Chuanru Ren +1

    cs.IRcs.AIarXiv:2608.16919v12026
  12. Deep learning for time series classification: a review

    Hassan Ismail Fawaz, Germain Forestier, Jonathan Weber +2

    cs.LGcs.AIstat.MLarXiv:1809.04356v42018
  13. Concrete Problems in AI Safety

    Dario Amodei, Chris Olah, Jacob Steinhardt +3

    cs.AIcs.LGarXiv:1606.06565v22016
  14. SWE-bench: Can Language Models Resolve Real-World GitHub Issues?

    Carlos E. Jimenez, John Yang, Alexander Wettig +4

    cs.CLcs.AIcs.SEarXiv:2310.06770v32023
  15. Regularized Evolution for Image Classifier Architecture Search

    Esteban Real, Alok Aggarwal, Yanping Huang +1

    cs.NEcs.AIcs.CVarXiv:1802.01548v72018
  16. Learning to summarize from human feedback

    Nisan Stiennon, Long Ouyang, Jeff Wu +6

    cs.CLcs.AIcs.LGarXiv:2009.01325v32020
  17. The Open-Strategy Dictator Game: Cooperation Under Mutual Transparency

    Michael Glass

    cs.GTcs.AIcs.MAarXiv:2608.14913v12026
  18. Automatic or Controlled? Repetition Priming Reveals Divergent Processing in Base LLMs, Instruct LLMs, and Humans

    Jinglei Ren, Yuyue Wang

    cs.CLcs.AIarXiv:2608.14681v12026
  19. A Contextual-Bandit Approach to Personalized News Article Recommendation

    Lihong Li, Wei Chu, John Langford +1

    cs.LGcs.AIcs.IRarXiv:1003.0146v22010
  20. Domain Agnostic Text Redaction from Natural Language Rules using Instruction Tuning

    Aravindhan Arunagiri, Ayaan Khan, Udayaadithya Avadhanam +1

    cs.CLcs.AIarXiv:2608.14693v12026
  21. PolyComp: A Polycube-based Benchmark for Compositional 3D Spatial Reasoning in Multimodal Models

    Siddharth Patel

    cs.CVcs.AIarXiv:2608.14741v12026
  22. Self-Instruct: Aligning Language Models with Self-Generated Instructions

    Yizhong Wang, Yeganeh Kordi, Swaroop Mishra +4

    cs.CLcs.AIarXiv:2212.10560v22022
  23. Writing Style Similarity Reflects Academic Genealogy

    Cameron Manzo

    cs.CLcs.AIarXiv:2608.14843v12026
  24. Universal and Transferable Adversarial Attacks on Aligned Language Models

    Andy Zou, Zifan Wang, Nicholas Carlini +3

    cs.CLcs.AIcs.CRarXiv:2307.15043v22023
  25. ER-KANs: Efficient and Robust Kolmogorov-Arnold Networks for Data-Scarce Scientific Machine Learning

    Harshil Lodhiya

    cs.LGcs.AIarXiv:2608.14773v12026
  26. Artificial Intelligence as a Tool for Combating Child Labour: A Real-Time Edge Vision Pipeline for Child Detection and Age Estimation

    Mark Nowak

    cs.CVcs.AIarXiv:2608.14770v12026
  27. Generated Context versus Governed State: Functional Conditions for Accountable Longitudinal Clinical Reasoning

    Augusto Bernardo Pissarra, Victor Lorena de Farias Souza

    cs.AIarXiv:2608.14804v12026
  28. Constitutional AI: Harmlessness from AI Feedback

    Yuntao Bai, Saurav Kadavath, Sandipan Kundu +48

    cs.CLcs.AIarXiv:2212.08073v12022
  29. When AI Rewrites, Classifiers Relax: Uncertainty-Aware Sentiment Analysis on Sarcastic and AI-Paraphrased Social Text

    Shresth Shroff

    cs.CLcs.AIarXiv:2608.15338v12026
  30. UC-PSRO: Utility-Conditioned Policy-Space Response Oracles with a Communication-Dropout Curriculum for Game-Theoretic Course-of-Action Generation in Adversarial Swarms

    Phillip Jiang

    cs.AIcs.MAarXiv:2608.15372v12026
  31. ReasonCast: Agentic Demand Forecasting with Selective Semantic Reasoning

    Ziyue Yang, Chaolin Xu, Yijing Wang +5

    cs.AIarXiv:2608.15291v12026
  32. A Survey on Evaluation of Large Language Models

    Yupeng Chang, Xu Wang, Jindong Wang +13

    cs.CLcs.AIarXiv:2307.03109v92023
  33. CheXpert: A Large Chest Radiograph Dataset with Uncertainty Labels and Expert Comparison

    Jeremy Irvin, Pranav Rajpurkar, Michael Ko +17

    cs.CVcs.AIcs.LGarXiv:1901.07031v12019
  34. MLP-Mixer: An all-MLP Architecture for Vision

    Ilya Tolstikhin, Neil Houlsby, Alexander Kolesnikov +9

    cs.CVcs.AIcs.LGarXiv:2105.01601v42021
  35. Obfuscated Gradients Give a False Sense of Security: Circumventing Defenses to Adversarial Examples

    Anish Athalye, Nicholas Carlini, David Wagner

    cs.LGcs.AIcs.CRarXiv:1802.00420v42018
  36. A Survey on Large Language Model based Autonomous Agents

    Lei Wang, Chen Ma, Xueyang Feng +10

    cs.AIcs.CLarXiv:2308.11432v72023
  37. Do Geometry-Aware Positional Encodings Help Transformers in Spatial Imperfect-Information Games?

    Wenji Fu

    cs.LGcs.AIstat.MLarXiv:2608.14982v12026
  38. MixMatch: A Holistic Approach to Semi-Supervised Learning

    David Berthelot, Nicholas Carlini, Ian Goodfellow +3

    cs.LGcs.AIcs.CVarXiv:1905.02249v22019
  39. Visible Reasoning and Indirect Prompt-Injection Monitorability Across English, Tamil, and Tanglish

    Madhusudhanan G

    cs.AIarXiv:2608.15392v12026
  40. Sigmoid Loss for Language Image Pre-Training

    Xiaohua Zhai, Basil Mustafa, Alexander Kolesnikov +1

    cs.CVcs.AIarXiv:2303.15343v42023
  41. Gradient Episodic Memory for Continual Learning

    David Lopez-Paz, Marc'Aurelio Ranzato

    cs.LGcs.AIarXiv:1706.08840v62017
  42. Demographic Injection in Medical Language Models under Diversity, Equity, and Inclusion Prompts

    Diego Mardian, Frank Liu

    cs.AIcs.CLarXiv:2608.15254v12026
  43. The Price of Thinking: Reasoning Effort as a Model-Specific API Contract

    Yeabin Moon

    cs.AIcs.CLcs.CYarXiv:2608.16956v12026
  44. Relational inductive biases, deep learning, and graph networks

    Peter W. Battaglia, Jessica B. Hamrick, Victor Bapst +24

    cs.LGcs.AIstat.MLarXiv:1806.01261v32018
  45. Complex Embeddings for Simple Link Prediction

    Théo Trouillon, Johannes Welbl, Sebastian Riedel +2

    cs.AIcs.LGstat.MLarXiv:1606.06357v12016
  46. Efficient Inference in Fully Connected CRFs with Gaussian Edge Potentials

    Philipp Krähenbühl, Vladlen Koltun

    cs.CVcs.AIcs.LGarXiv:1210.5644v12012
  47. Transformers in Vision: A Survey

    Salman Khan, Muzammal Naseer, Munawar Hayat +3

    cs.CVcs.AIcs.LGarXiv:2101.01169v52021
  48. Understanding Black-box Predictions via Influence Functions

    Pang Wei Koh, Percy Liang

    stat.MLcs.AIcs.LGarXiv:1703.04730v32017
  49. SAM 2: Segment Anything in Images and Videos

    Nikhila Ravi, Valentin Gabeur, Yuan-Ting Hu +15

    cs.CVcs.AIcs.LGarXiv:2408.00714v22024
  50. Mistral 7B

    Albert Q. Jiang, Alexandre Sablayrolles, Arthur Mensch +15

    cs.CLcs.AIcs.LGarXiv:2310.06825v12023
  51. Man is to Computer Programmer as Woman is to Homemaker? Debiasing Word Embeddings

    Tolga Bolukbasi, Kai-Wei Chang, James Zou +2

    cs.CLcs.AIcs.LGarXiv:1607.06520v12016
  52. A Survey on Visual Transformer

    Kai Han, Yunhe Wang, Hanting Chen +10

    cs.CVcs.AIarXiv:2012.12556v62020
  53. Retrieval-Augmented Generation for Large Language Models: A Survey

    Yunfan Gao, Yun Xiong, Xinyu Gao +7

    cs.CLcs.AIarXiv:2312.10997v52023
  54. Deep CORAL: Correlation Alignment for Deep Domain Adaptation

    Baochen Sun, Kate Saenko

    cs.CVcs.AIcs.LGarXiv:1607.01719v12016
  55. Let's Verify Step by Step

    Hunter Lightman, Vineet Kosaraju, Yura Burda +7

    cs.LGcs.AIcs.CLarXiv:2305.20050v12023
  56. Multi-column Deep Neural Networks for Image Classification

    Dan Cireşan, Ueli Meier, Juergen Schmidhuber

    cs.CVcs.AIarXiv:1202.2745v12012
  57. ADEPT: Accelerating Dexterity via Pre-Training and Post-Training using Reinforcement Learning

    Jayjun Lee, Jessica Yin, Asif Rana +7

    cs.ROcs.AIarXiv:2608.19182v12026
  58. Beyond Teacher Likelihood: Group-Calibrated On-Policy Distillation for Long-Context Reasoning

    Zhu Zhang, Jixun Wang, Xiaoang Xu +6

    cs.LGcs.AIcs.CLarXiv:2608.19181v12026
  59. Bernstein-Vazirani Networks: Quantum Machine Learning by Interference

    Natacha Kuete Meli, Tolga Birdal, Prayag Tiwari +2

    quant-phcs.AIcs.CVarXiv:2608.19043v12026
  60. The Lottery Ticket Hypothesis: Finding Sparse, Trainable Neural Networks

    Jonathan Frankle, Michael Carbin

    cs.LGcs.AIcs.NEarXiv:1803.03635v52018