Artificial Intelligence

Papers filed under cs.AI on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

13,861 to 13,920 of 15,209

  1. Sigmoid Loss for Language Image Pre-Training

    Xiaohua Zhai, Basil Mustafa, Alexander Kolesnikov +1

    cs.CVcs.AIarXiv:2303.15343v42023
  2. Gradient Episodic Memory for Continual Learning

    David Lopez-Paz, Marc'Aurelio Ranzato

    cs.LGcs.AIarXiv:1706.08840v62017
  3. Demographic Injection in Medical Language Models under Diversity, Equity, and Inclusion Prompts

    Diego Mardian, Frank Liu

    cs.AIcs.CLarXiv:2608.15254v12026
  4. The Price of Thinking: Reasoning Effort as a Model-Specific API Contract

    Yeabin Moon

    cs.AIcs.CLcs.CYarXiv:2608.16956v12026
  5. Relational inductive biases, deep learning, and graph networks

    Peter W. Battaglia, Jessica B. Hamrick, Victor Bapst +24

    cs.LGcs.AIstat.MLarXiv:1806.01261v32018
  6. Complex Embeddings for Simple Link Prediction

    Théo Trouillon, Johannes Welbl, Sebastian Riedel +2

    cs.AIcs.LGstat.MLarXiv:1606.06357v12016
  7. Efficient Inference in Fully Connected CRFs with Gaussian Edge Potentials

    Philipp Krähenbühl, Vladlen Koltun

    cs.CVcs.AIcs.LGarXiv:1210.5644v12012
  8. Transformers in Vision: A Survey

    Salman Khan, Muzammal Naseer, Munawar Hayat +3

    cs.CVcs.AIcs.LGarXiv:2101.01169v52021
  9. Understanding Black-box Predictions via Influence Functions

    Pang Wei Koh, Percy Liang

    stat.MLcs.AIcs.LGarXiv:1703.04730v32017
  10. SAM 2: Segment Anything in Images and Videos

    Nikhila Ravi, Valentin Gabeur, Yuan-Ting Hu +15

    cs.CVcs.AIcs.LGarXiv:2408.00714v22024
  11. Mistral 7B

    Albert Q. Jiang, Alexandre Sablayrolles, Arthur Mensch +15

    cs.CLcs.AIcs.LGarXiv:2310.06825v12023
  12. Man is to Computer Programmer as Woman is to Homemaker? Debiasing Word Embeddings

    Tolga Bolukbasi, Kai-Wei Chang, James Zou +2

    cs.CLcs.AIcs.LGarXiv:1607.06520v12016
  13. A Survey on Visual Transformer

    Kai Han, Yunhe Wang, Hanting Chen +10

    cs.CVcs.AIarXiv:2012.12556v62020
  14. Retrieval-Augmented Generation for Large Language Models: A Survey

    Yunfan Gao, Yun Xiong, Xinyu Gao +7

    cs.CLcs.AIarXiv:2312.10997v52023
  15. Deep CORAL: Correlation Alignment for Deep Domain Adaptation

    Baochen Sun, Kate Saenko

    cs.CVcs.AIcs.LGarXiv:1607.01719v12016
  16. Let's Verify Step by Step

    Hunter Lightman, Vineet Kosaraju, Yura Burda +7

    cs.LGcs.AIcs.CLarXiv:2305.20050v12023
  17. Multi-column Deep Neural Networks for Image Classification

    Dan Cireşan, Ueli Meier, Juergen Schmidhuber

    cs.CVcs.AIarXiv:1202.2745v12012
  18. ADEPT: Accelerating Dexterity via Pre-Training and Post-Training using Reinforcement Learning

    Jayjun Lee, Jessica Yin, Asif Rana +7

    cs.ROcs.AIarXiv:2608.19182v12026
  19. Beyond Teacher Likelihood: Group-Calibrated On-Policy Distillation for Long-Context Reasoning

    Zhu Zhang, Jixun Wang, Xiaoang Xu +6

    cs.LGcs.AIcs.CLarXiv:2608.19181v12026
  20. Bernstein-Vazirani Networks: Quantum Machine Learning by Interference

    Natacha Kuete Meli, Tolga Birdal, Prayag Tiwari +2

    quant-phcs.AIcs.CVarXiv:2608.19043v12026
  21. The Lottery Ticket Hypothesis: Finding Sparse, Trainable Neural Networks

    Jonathan Frankle, Michael Carbin

    cs.LGcs.AIcs.NEarXiv:1803.03635v52018
  22. DeepWeaver: Bridging the Evidence Synthesis Gap in Open-Ended Question Answering

    Xujia Wang, Yizhe Zhang, Bin Xu +2

    cs.CLcs.AIarXiv:2608.18988v12026
  23. Masked-attention Mask Transformer for Universal Image Segmentation

    Bowen Cheng, Ishan Misra, Alexander G. Schwing +2

    cs.CVcs.AIcs.LGarXiv:2112.01527v32021
  24. FiLM: Visual Reasoning with a General Conditioning Layer

    Ethan Perez, Florian Strub, Harm de Vries +2

    cs.CVcs.AIcs.CLarXiv:1709.07871v22017
  25. From Threat Intelligence to Detection: Knowledge-driven Enrichment and Template-based Rule Grounding for Automated Sigma Rule Generation

    Sepehr Ghaffarzadegan, Boubakr Nour, Makan Pourzandi +2

    cs.CRcs.AIarXiv:2608.19011v12026
  26. AlphaClifford: Efficient Clifford Synthesis and Transpilation with Model-based RL

    Daniele Lizzio Bosco, Jacopo Cossio, Carla Piazza +1

    quant-phcs.AIarXiv:2608.18946v12026
  27. Pre-Compiled Pipeline Shards for Distributed LLM Inference on Intel AI PC Fleets

    Tate Berenbaum, Muthaiah Venkatachalam

    cs.DCcs.AIcs.SEarXiv:2608.19147v12026
  28. Leaf Values as Coordinates: Exact Contrastive Explanation for Gradient-Boosted Ensembles

    Emanuele Luzio

    cs.LGcs.AIcs.CYarXiv:2608.19127v12026
  29. Intercepting the Kangaroo: Experimental Astrolinguistics with Constructed Lexicons, Active Probing, and Large Language Models as Informants and Hypothesis Proposers

    Francesco Cordella, Mauro Cappelli

    cs.CLcs.AIarXiv:2608.19124v12026
  30. Discretizing Continuous Time Series for Imputation with Masked Diffusion Training

    Dongbin Kim, Seungyun Lee, Geonwoo Shin +1

    cs.LGcs.AIarXiv:2608.19119v12026
  31. Learning to Prompt for Vision-Language Models

    Kaiyang Zhou, Jingkang Yang, Chen Change Loy +1

    cs.CVcs.AIcs.LGarXiv:2109.01134v62021
  32. Open-MOPD: Diagnosing and Fixing Capability Imbalance in Multi-Teacher On-Policy Distillation

    Huan-ang Gao, Haohan Chi, Yong Yan +7

    cs.LGcs.AIcs.CLarXiv:2608.19098v12026
  33. Detecting Backdoors in Object Detection via Pre-NMS Prediction Distribution Shift

    Longtian Wang, Zhengyu Zhao, Chenhao Lin +5

    cs.CVcs.AIarXiv:2608.19088v12026
  34. DA-WAM: Decision-Aligned Future Latents for Driving World Models

    Ruiguo Zhong, Benshan Ma, Xiaolong Chen +5

    cs.ROcs.AIarXiv:2608.19085v12026
  35. GS-VLA: Plug-and-Play Viewpoint Canonicalization for Frozen VLA Policies via Gaussian Splatting

    Yechan Park, HyunJin Kim

    cs.CVcs.AIarXiv:2608.19066v12026
  36. Counterfactual Contrastive Analysis

    Yunlong He, Pietro Gori

    cs.CVcs.AIarXiv:2608.19032v12026
  37. One-Stage Object Detectors in Autonomous Driving

    Jonel Roman, Ryan Sirjue, Peter Nguyen +3

    cs.CVcs.AIarXiv:2608.19014v12026
  38. A Reduction of Imitation Learning and Structured Prediction to No-Regret Online Learning

    Stephane Ross, Geoffrey J. Gordon, J. Andrew Bagnell

    cs.LGcs.AIstat.MLarXiv:1011.0686v32010
  39. Harness Continual Learning: Continual Adaptation Beyond Model Parameters

    Borui Kang, Jinrui Gu, Junhan Lv +3

    cs.LGcs.AIarXiv:2608.19013v12026
  40. GrabVG: Graph-Attentive Binding for Visual Grounding in UAV Imagery

    Chaowei Wang, Yan Di, Jingjun Sun +5

    cs.CVcs.AIarXiv:2608.18996v12026
  41. Pointer Sentinel Mixture Models

    Stephen Merity, Caiming Xiong, James Bradbury +1

    cs.CLcs.AIarXiv:1609.07843v12016
  42. Making the V in VQA Matter: Elevating the Role of Image Understanding in Visual Question Answering

    Yash Goyal, Tejas Khot, Douglas Summers-Stay +2

    cs.CVcs.AIcs.CLarXiv:1612.00837v32016
  43. rEDMRec: Distilling Large Language Model Reasoning into an Editable Experience Memory for Recommendation

    Minh Hoang Nguyen, Tung Le, Huy Tien Nguyen

    cs.IRcs.AIcs.CLarXiv:2608.18952v12026
  44. Epistemic Subordination: Generative AI and the Infrastructure of Knowledge

    Gilad Abiri, Emanuel V. Towfigh

    cs.CYcs.AIarXiv:2608.18758v12026
  45. Tree of Thoughts: Deliberate Problem Solving with Large Language Models

    Shunyu Yao, Dian Yu, Jeffrey Zhao +4

    cs.CLcs.AIcs.LGarXiv:2305.10601v22023
  46. A Critical Synthesis of Uncertainty Quantification and Foundation Models for Semantic Segmentation

    Steven Landgraf, Joceline Hinz, Markus Ulrich

    cs.CVcs.AIcs.LGarXiv:2608.18709v12026
  47. MedUAG: Unified Understanding and Generation for Medical Multimodal Models

    Zijie Meng, Yuncheng Zhang, Hualiang Wang +8

    cs.CLcs.AIarXiv:2608.18937v12026
  48. Graphical Design of Interpretable Architectures

    Pietro Barbiero

    cs.LGcs.AIcs.NEarXiv:2608.18936v12026
  49. SMTrap: Cost-Effective DoS Attacks Against Large Reasoning Models via SMT Conflict Guidance

    Jian Yang, Zhenqi Feng, Zhaoyang Yu +7

    cs.CLcs.AIarXiv:2608.18921v12026
  50. Learning-State-Aware Dynamic Generative Data Augmentation on Small-Scale Datasets

    Ting Xiang, Chenxi Deng, Jinhui Zhao +4

    cs.CVcs.AIarXiv:2608.18907v12026
  51. Density estimation using Real NVP

    Laurent Dinh, Jascha Sohl-Dickstein, Samy Bengio

    cs.LGcs.AIcs.NEarXiv:1605.08803v32016
  52. Identifying Implicit Premises for Logical Reconstruction of Argument Graphs

    Xuyao Feng, Anthony Hunter

    cs.CLcs.AIarXiv:2608.18821v12026
  53. Do Large Language Models Hallucinate Electric Fata Morganas?

    Kristina Šekrst

    cs.CLcs.AIarXiv:2608.18816v12026
  54. Forgetting, plasticity, and co-observation: a third facet of continual learning

    Timm Hess, Abhishek Jha, Gido M. van de Ven +1

    cs.LGcs.AIarXiv:2608.18803v12026
  55. The Mythos of Model Interpretability

    Zachary C. Lipton

    cs.LGcs.AIcs.CVarXiv:1606.03490v32016
  56. Decomposing Wrong-Consensus Agreement in LLM Self-Consistency: A GPT-4.1 Case Study

    Lizhuo Zhang, Mengmeng Tang, Chenfeng Long +2

    cs.CLcs.AIarXiv:2608.18795v12026
  57. Beyond Predictive Fairness: Quantifying Attribution Consistency Across Demographic Groups in Diabetic Retinopathy Screening

    Kerol Djoumessi, Philipp Berens

    cs.LGcs.AIarXiv:2608.18759v12026
  58. A Few Cases Are All You Need: An Empirical Study of Annotation-Efficient LoRA Fine-Tuning of MedSAM3

    Sachin Dudda Nagaraju, Bendik Skarre Abrahamsen, Ashkan Moradi +1

    cs.CVcs.AIarXiv:2608.18731v12026
  59. The Impact of CutMix on Reliability and Robustness in Semantic Segmentation

    Steven Landgraf, Markus Ulrich

    cs.CVcs.AIcs.LGarXiv:2608.18715v12026
  60. A Survey of Large Language Models

    Wayne Xin Zhao, Kun Zhou, Junyi Li +19

    cs.CLcs.AIarXiv:2303.18223v192023