Artificial Intelligence
Papers filed under cs.AI on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
13,861 to 13,920 of 15,209
Sigmoid Loss for Language Image Pre-Training
Xiaohua Zhai, Basil Mustafa, Alexander Kolesnikov +1
cs.CVcs.AIarXiv:2303.15343v42023Gradient Episodic Memory for Continual Learning
David Lopez-Paz, Marc'Aurelio Ranzato
cs.LGcs.AIarXiv:1706.08840v62017Demographic Injection in Medical Language Models under Diversity, Equity, and Inclusion Prompts
Diego Mardian, Frank Liu
cs.AIcs.CLarXiv:2608.15254v12026The Price of Thinking: Reasoning Effort as a Model-Specific API Contract
Yeabin Moon
cs.AIcs.CLcs.CYarXiv:2608.16956v12026Relational inductive biases, deep learning, and graph networks
Peter W. Battaglia, Jessica B. Hamrick, Victor Bapst +24
cs.LGcs.AIstat.MLarXiv:1806.01261v32018Complex Embeddings for Simple Link Prediction
Théo Trouillon, Johannes Welbl, Sebastian Riedel +2
cs.AIcs.LGstat.MLarXiv:1606.06357v12016Efficient Inference in Fully Connected CRFs with Gaussian Edge Potentials
Philipp Krähenbühl, Vladlen Koltun
cs.CVcs.AIcs.LGarXiv:1210.5644v12012Transformers in Vision: A Survey
Salman Khan, Muzammal Naseer, Munawar Hayat +3
cs.CVcs.AIcs.LGarXiv:2101.01169v52021Understanding Black-box Predictions via Influence Functions
Pang Wei Koh, Percy Liang
stat.MLcs.AIcs.LGarXiv:1703.04730v32017SAM 2: Segment Anything in Images and Videos
Nikhila Ravi, Valentin Gabeur, Yuan-Ting Hu +15
cs.CVcs.AIcs.LGarXiv:2408.00714v22024Mistral 7B
Albert Q. Jiang, Alexandre Sablayrolles, Arthur Mensch +15
cs.CLcs.AIcs.LGarXiv:2310.06825v12023Man is to Computer Programmer as Woman is to Homemaker? Debiasing Word Embeddings
Tolga Bolukbasi, Kai-Wei Chang, James Zou +2
cs.CLcs.AIcs.LGarXiv:1607.06520v12016A Survey on Visual Transformer
Kai Han, Yunhe Wang, Hanting Chen +10
cs.CVcs.AIarXiv:2012.12556v62020Retrieval-Augmented Generation for Large Language Models: A Survey
Yunfan Gao, Yun Xiong, Xinyu Gao +7
cs.CLcs.AIarXiv:2312.10997v52023Deep CORAL: Correlation Alignment for Deep Domain Adaptation
Baochen Sun, Kate Saenko
cs.CVcs.AIcs.LGarXiv:1607.01719v12016Let's Verify Step by Step
Hunter Lightman, Vineet Kosaraju, Yura Burda +7
cs.LGcs.AIcs.CLarXiv:2305.20050v12023Multi-column Deep Neural Networks for Image Classification
Dan Cireşan, Ueli Meier, Juergen Schmidhuber
cs.CVcs.AIarXiv:1202.2745v12012ADEPT: Accelerating Dexterity via Pre-Training and Post-Training using Reinforcement Learning
Jayjun Lee, Jessica Yin, Asif Rana +7
cs.ROcs.AIarXiv:2608.19182v12026Beyond Teacher Likelihood: Group-Calibrated On-Policy Distillation for Long-Context Reasoning
Zhu Zhang, Jixun Wang, Xiaoang Xu +6
cs.LGcs.AIcs.CLarXiv:2608.19181v12026Bernstein-Vazirani Networks: Quantum Machine Learning by Interference
Natacha Kuete Meli, Tolga Birdal, Prayag Tiwari +2
quant-phcs.AIcs.CVarXiv:2608.19043v12026The Lottery Ticket Hypothesis: Finding Sparse, Trainable Neural Networks
Jonathan Frankle, Michael Carbin
cs.LGcs.AIcs.NEarXiv:1803.03635v52018DeepWeaver: Bridging the Evidence Synthesis Gap in Open-Ended Question Answering
Xujia Wang, Yizhe Zhang, Bin Xu +2
cs.CLcs.AIarXiv:2608.18988v12026Masked-attention Mask Transformer for Universal Image Segmentation
Bowen Cheng, Ishan Misra, Alexander G. Schwing +2
cs.CVcs.AIcs.LGarXiv:2112.01527v32021FiLM: Visual Reasoning with a General Conditioning Layer
Ethan Perez, Florian Strub, Harm de Vries +2
cs.CVcs.AIcs.CLarXiv:1709.07871v22017From Threat Intelligence to Detection: Knowledge-driven Enrichment and Template-based Rule Grounding for Automated Sigma Rule Generation
Sepehr Ghaffarzadegan, Boubakr Nour, Makan Pourzandi +2
cs.CRcs.AIarXiv:2608.19011v12026AlphaClifford: Efficient Clifford Synthesis and Transpilation with Model-based RL
Daniele Lizzio Bosco, Jacopo Cossio, Carla Piazza +1
quant-phcs.AIarXiv:2608.18946v12026Pre-Compiled Pipeline Shards for Distributed LLM Inference on Intel AI PC Fleets
Tate Berenbaum, Muthaiah Venkatachalam
cs.DCcs.AIcs.SEarXiv:2608.19147v12026Leaf Values as Coordinates: Exact Contrastive Explanation for Gradient-Boosted Ensembles
Emanuele Luzio
cs.LGcs.AIcs.CYarXiv:2608.19127v12026Intercepting the Kangaroo: Experimental Astrolinguistics with Constructed Lexicons, Active Probing, and Large Language Models as Informants and Hypothesis Proposers
Francesco Cordella, Mauro Cappelli
cs.CLcs.AIarXiv:2608.19124v12026Discretizing Continuous Time Series for Imputation with Masked Diffusion Training
Dongbin Kim, Seungyun Lee, Geonwoo Shin +1
cs.LGcs.AIarXiv:2608.19119v12026Learning to Prompt for Vision-Language Models
Kaiyang Zhou, Jingkang Yang, Chen Change Loy +1
cs.CVcs.AIcs.LGarXiv:2109.01134v62021Open-MOPD: Diagnosing and Fixing Capability Imbalance in Multi-Teacher On-Policy Distillation
Huan-ang Gao, Haohan Chi, Yong Yan +7
cs.LGcs.AIcs.CLarXiv:2608.19098v12026Detecting Backdoors in Object Detection via Pre-NMS Prediction Distribution Shift
Longtian Wang, Zhengyu Zhao, Chenhao Lin +5
cs.CVcs.AIarXiv:2608.19088v12026DA-WAM: Decision-Aligned Future Latents for Driving World Models
Ruiguo Zhong, Benshan Ma, Xiaolong Chen +5
cs.ROcs.AIarXiv:2608.19085v12026GS-VLA: Plug-and-Play Viewpoint Canonicalization for Frozen VLA Policies via Gaussian Splatting
Yechan Park, HyunJin Kim
cs.CVcs.AIarXiv:2608.19066v12026Counterfactual Contrastive Analysis
Yunlong He, Pietro Gori
cs.CVcs.AIarXiv:2608.19032v12026One-Stage Object Detectors in Autonomous Driving
Jonel Roman, Ryan Sirjue, Peter Nguyen +3
cs.CVcs.AIarXiv:2608.19014v12026A Reduction of Imitation Learning and Structured Prediction to No-Regret Online Learning
Stephane Ross, Geoffrey J. Gordon, J. Andrew Bagnell
cs.LGcs.AIstat.MLarXiv:1011.0686v32010Harness Continual Learning: Continual Adaptation Beyond Model Parameters
Borui Kang, Jinrui Gu, Junhan Lv +3
cs.LGcs.AIarXiv:2608.19013v12026GrabVG: Graph-Attentive Binding for Visual Grounding in UAV Imagery
Chaowei Wang, Yan Di, Jingjun Sun +5
cs.CVcs.AIarXiv:2608.18996v12026Pointer Sentinel Mixture Models
Stephen Merity, Caiming Xiong, James Bradbury +1
cs.CLcs.AIarXiv:1609.07843v12016Making the V in VQA Matter: Elevating the Role of Image Understanding in Visual Question Answering
Yash Goyal, Tejas Khot, Douglas Summers-Stay +2
cs.CVcs.AIcs.CLarXiv:1612.00837v32016rEDMRec: Distilling Large Language Model Reasoning into an Editable Experience Memory for Recommendation
Minh Hoang Nguyen, Tung Le, Huy Tien Nguyen
cs.IRcs.AIcs.CLarXiv:2608.18952v12026Epistemic Subordination: Generative AI and the Infrastructure of Knowledge
Gilad Abiri, Emanuel V. Towfigh
cs.CYcs.AIarXiv:2608.18758v12026Tree of Thoughts: Deliberate Problem Solving with Large Language Models
Shunyu Yao, Dian Yu, Jeffrey Zhao +4
cs.CLcs.AIcs.LGarXiv:2305.10601v22023A Critical Synthesis of Uncertainty Quantification and Foundation Models for Semantic Segmentation
Steven Landgraf, Joceline Hinz, Markus Ulrich
cs.CVcs.AIcs.LGarXiv:2608.18709v12026MedUAG: Unified Understanding and Generation for Medical Multimodal Models
Zijie Meng, Yuncheng Zhang, Hualiang Wang +8
cs.CLcs.AIarXiv:2608.18937v12026Graphical Design of Interpretable Architectures
Pietro Barbiero
cs.LGcs.AIcs.NEarXiv:2608.18936v12026SMTrap: Cost-Effective DoS Attacks Against Large Reasoning Models via SMT Conflict Guidance
Jian Yang, Zhenqi Feng, Zhaoyang Yu +7
cs.CLcs.AIarXiv:2608.18921v12026Learning-State-Aware Dynamic Generative Data Augmentation on Small-Scale Datasets
Ting Xiang, Chenxi Deng, Jinhui Zhao +4
cs.CVcs.AIarXiv:2608.18907v12026Density estimation using Real NVP
Laurent Dinh, Jascha Sohl-Dickstein, Samy Bengio
cs.LGcs.AIcs.NEarXiv:1605.08803v32016Identifying Implicit Premises for Logical Reconstruction of Argument Graphs
Xuyao Feng, Anthony Hunter
cs.CLcs.AIarXiv:2608.18821v12026Do Large Language Models Hallucinate Electric Fata Morganas?
Kristina Šekrst
cs.CLcs.AIarXiv:2608.18816v12026Forgetting, plasticity, and co-observation: a third facet of continual learning
Timm Hess, Abhishek Jha, Gido M. van de Ven +1
cs.LGcs.AIarXiv:2608.18803v12026The Mythos of Model Interpretability
Zachary C. Lipton
cs.LGcs.AIcs.CVarXiv:1606.03490v32016Decomposing Wrong-Consensus Agreement in LLM Self-Consistency: A GPT-4.1 Case Study
Lizhuo Zhang, Mengmeng Tang, Chenfeng Long +2
cs.CLcs.AIarXiv:2608.18795v12026Beyond Predictive Fairness: Quantifying Attribution Consistency Across Demographic Groups in Diabetic Retinopathy Screening
Kerol Djoumessi, Philipp Berens
cs.LGcs.AIarXiv:2608.18759v12026A Few Cases Are All You Need: An Empirical Study of Annotation-Efficient LoRA Fine-Tuning of MedSAM3
Sachin Dudda Nagaraju, Bendik Skarre Abrahamsen, Ashkan Moradi +1
cs.CVcs.AIarXiv:2608.18731v12026The Impact of CutMix on Reliability and Robustness in Semantic Segmentation
Steven Landgraf, Markus Ulrich
cs.CVcs.AIcs.LGarXiv:2608.18715v12026A Survey of Large Language Models
Wayne Xin Zhao, Kun Zhou, Junyi Li +19
cs.CLcs.AIarXiv:2303.18223v192023