Artificial Intelligence

Papers filed under cs.AI on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

10,021 to 10,080 of 15,246

  1. ClassVision: AI-Powered Classroom Attendance System

    Ankit Kumar Aggarwal, Veerabhadra Rao Marellapudi, Ovadia Sutton +1

    cs.CYcs.AIcs.CVarXiv:2608.26173v12026
  2. GPT-Driver: Learning to Drive with GPT

    Jiageng Mao, Yuxi Qian, Junjie Ye +2

    cs.CVcs.AIcs.CLarXiv:2310.01415v32023
  3. Making Sense of Vision and Touch: Self-Supervised Learning of Multimodal Representations for Contact-Rich Tasks

    Michelle A. Lee, Yuke Zhu, Krishnan Srinivasan +5

    cs.ROcs.AIcs.LGarXiv:1810.10191v22018
  4. Evaluating AI Generated Summaries for Cancer Patients

    Muhammad Aurangzeb Ahmad, Kim Shyu, Leon Oliver +2

    cs.CLcs.AIarXiv:2608.26154v12026
  5. VFA: Empowering Multilingual MLLMs via Vision-Free Adaptation

    Yixia Li, Yaqing Shi, Zhiwen Ruan +6

    cs.CLcs.AIarXiv:2608.26155v12026
  6. Self-Generated Text Recognition: Quality Heuristics, Cross-Task Transfer, and Downstream Bias in LLM Evaluation

    Jesse St. Amand, Callum Canavan, Sohaib Imran +5

    cs.CLcs.AIarXiv:2608.26159v12026
  7. Hypertableau Reasoning for Description Logics

    Boris Motik, Rob Shearer, Ian Horrocks

    cs.LOcs.AIarXiv:1401.3485v12014
  8. Mutual Debiasing via Dual-Seed Comparison for Probabilistic Sampling in Large Language Models

    Zihao Guo, Hongtao Lv, Chaoli Zhang +4

    cs.CLcs.AIarXiv:2608.26161v12026
  9. From Sound to Symptom: Real-Time Respiratory Signal Understanding for Conversational Healthcare Agents

    Tanmay Laud, Herprit Mahal, Subhabrata Mukherjee

    cs.CLcs.AIcs.HCarXiv:2608.26163v12026
  10. End-to-End Differentiable Proving

    Tim Rocktäschel, Sebastian Riedel

    cs.NEcs.AIcs.LGarXiv:1705.11040v22017
  11. Using Poly-Encoders for Computationally Efficient Automated Creativity Assessment

    Sam Grouchnikov, Phillip Gregory, Jiho Noh

    cs.CLcs.AIarXiv:2608.26165v12026
  12. A novel time-frequency Transformer based on self-attention mechanism and its application in fault diagnosis of rolling bearings

    Yifei Ding, Minping Jia, Qiuhua Miao +1

    cs.AIcs.LGeess.SParXiv:2104.09079v32021
  13. A Comprehensive Survey on Community Detection with Deep Learning

    Xing Su, Shan Xue, Fanzhen Liu +9

    cs.SIcs.AIcs.LGarXiv:2105.12584v22021
  14. Lost in Compression: A Controlled Cross-Lingual Audit of Extractive Prompt Compressors

    Mantas Lukauskas

    cs.CLcs.AIarXiv:2608.26175v12026
  15. MedFG-VQA: Low-Frequency Memory and Graph Attention for Lightweight Medical VQA

    Haowen Gu, Gensheng Pei, Zeren Sun +4

    cs.CVcs.AIarXiv:2608.26848v12026
  16. Local Aggregation for Unsupervised Learning of Visual Embeddings

    Chengxu Zhuang, Alex Lin Zhai, Daniel Yamins

    cs.CVcs.AIarXiv:1903.12355v22019
  17. Knowledge Cards: Structured Knowledge for AI Systems

    Liliana Ferreira

    cs.AIcs.CLarXiv:2608.26176v12026
  18. Agentless: Demystifying LLM-based Software Engineering Agents

    Chunqiu Steven Xia, Yinlin Deng, Soren Dunn +1

    cs.SEcs.AIcs.CLarXiv:2407.01489v22024
  19. A Task-Centric Ontology and Deterministic Domain Rules as a Verifiable Core for AI-Assisted Chemistry Problem Solving

    Ibrokhimsho Abduchaborov

    cs.AIarXiv:2608.26164v12026
  20. A Safety-Gated Multimodal AI Backend for Mental-Health Support: Hierarchical State Representation, Conservative Risk Fusion, and Controlled Generation in Anian

    Lei Wang, Xiao Wang, Lei Li

    cs.AIarXiv:2608.26162v12026
  21. GROUND: Reducing Hallucinations in LLM-Based Enterprise Analytics Through Governed Semantic Definitions

    Aravind Sasidharan Pillai

    cs.AIarXiv:2608.26157v12026
  22. Selection Bias Correction in Retail Intelligence

    Spandan Ghose Chowdhury

    cs.AIcs.LGarXiv:2608.26156v12026
  23. Recent Advances in End-to-End Automatic Speech Recognition

    Jinyu Li

    eess.AScs.AIcs.CLarXiv:2111.01690v22021
  24. Explainable Artificial Intelligence for Customer Churn Prediction in Telecommunications: A Framework for CRM Integration

    Sandeep Gaddamwar

    cs.AIcs.LGarXiv:2608.26151v12026
  25. Cyclical Annealing Schedule: A Simple Approach to Mitigating KL Vanishing

    Hao Fu, Chunyuan Li, Xiaodong Liu +3

    cs.LGcs.AIcs.CLarXiv:1903.10145v32019
  26. Janus: Decoupling Visual Encoding for Unified Multimodal Understanding and Generation

    Chengyue Wu, Xiaokang Chen, Zhiyu Wu +8

    cs.CVcs.AIcs.CLarXiv:2410.13848v12024
  27. Refusal Is Not Robustness: Auditing Confident Fabrication in Large Language Models on a Provably Uninformative Clinical Pain Speech Transcript

    Sagnik De, Sreenija Pavuluri

    cs.AIcs.ETcs.LGarXiv:2608.26167v12026
  28. Auxiliary Deep Generative Models

    Lars Maaløe, Casper Kaae Sønderby, Søren Kaae Sønderby +1

    stat.MLcs.AIcs.LGarXiv:1602.05473v42016
  29. Transfusion: Predict the Next Token and Diffuse Images with One Multi-Modal Model

    Chunting Zhou, Lili Yu, Arun Babu +7

    cs.AIcs.CVarXiv:2408.11039v12024
  30. SAREF-based Ontology for Distributed AI Workflows across the Edge-Fog-Cloud Continuum

    Viorica Rozina Chifu, Tudor Cioara, Vasile Ofrim +4

    cs.AIarXiv:2608.26160v12026
  31. Joint Optimization of Masks and Deep Recurrent Neural Networks for Monaural Source Separation

    Po-Sen Huang, Minje Kim, Mark Hasegawa-Johnson +1

    cs.SDcs.AIcs.LGarXiv:1502.04149v42015
  32. Emotion Recognition in Conversation: Research Challenges, Datasets, and Recent Advances

    Soujanya Poria, Navonil Majumder, Rada Mihalcea +1

    cs.CLcs.AIarXiv:1905.02947v12019
  33. TSMixer: An All-MLP Architecture for Time Series Forecasting

    Si-An Chen, Chun-Liang Li, Nate Yoder +2

    cs.LGcs.AIarXiv:2303.06053v52023
  34. MedAlpaca -- An Open-Source Collection of Medical Conversational AI Models and Training Data

    Tianyu Han, Lisa C. Adams, Jens-Michalis Papaioannou +6

    cs.CLcs.AIarXiv:2304.08247v32023
  35. A Compositional Object-Based Approach to Learning Physical Dynamics

    Michael B. Chang, Tomer Ullman, Antonio Torralba +1

    cs.AIcs.LGarXiv:1612.00341v22016
  36. Methodological and Conceptual Framework for 5D Multi-Table Analysis: A Unified Approach for Complex Data Reuse

    Edouard Lansiaux, Hugo Kazzi, Aurélien Loison +2

    cs.AIcs.LGarXiv:2608.26149v12026
  37. Domain Adaptive Neural Networks for Object Recognition

    Muhammad Ghifary, W. Bastiaan Kleijn, Mengjie Zhang

    cs.CVcs.AIcs.LGarXiv:1409.6041v12014
  38. Deep Learning for Joint Source-Channel Coding of Text

    Nariman Farsad, Milind Rao, Andrea Goldsmith

    cs.ITcs.AIcs.LGarXiv:1802.06832v12018
  39. Zero-Shot Self-Orchestration with Ledger-Based Control for Improved LLM Coding Performance

    Victor Gao, Vida Khosrowshahi, Ali Khosrowshahi +4

    cs.MAcs.AIcs.CLarXiv:2608.26480v12026
  40. When the Canonical Completion Is Wrong: Formalizing and Measuring the Jump in Large Language Models

    Dai Shi, Xiaoyu Li, José Miguel Hernández-Lobato

    cs.CLcs.AIcs.LGarXiv:2608.26187v12026
  41. Separable Self-attention for Mobile Vision Transformers

    Sachin Mehta, Mohammad Rastegari

    cs.CVcs.AIcs.LGarXiv:2206.02680v12022
  42. Sequential Short-Text Classification with Recurrent and Convolutional Neural Networks

    Ji Young Lee, Franck Dernoncourt

    cs.CLcs.AIcs.LGarXiv:1603.03827v12016
  43. MMRotate: A Rotated Object Detection Benchmark using PyTorch

    Yue Zhou, Xue Yang, Gefan Zhang +9

    cs.CVcs.AIarXiv:2204.13317v42022
  44. Procedural Content Generation via Machine Learning (PCGML)

    Adam Summerville, Sam Snodgrass, Matthew Guzdial +5

    cs.AIarXiv:1702.00539v32017
  45. LanguageBind: Extending Video-Language Pretraining to N-modality by Language-based Semantic Alignment

    Bin Zhu, Bin Lin, Munan Ning +11

    cs.CVcs.AIarXiv:2310.01852v72023
  46. PMLB: A Large Benchmark Suite for Machine Learning Evaluation and Comparison

    Randal S. Olson, William La Cava, Patryk Orzechowski +2

    cs.LGcs.AIarXiv:1703.00512v12017
  47. Monkey: Image Resolution and Text Label Are Important Things for Large Multi-modal Models

    Zhang Li, Biao Yang, Qiang Liu +6

    cs.CVcs.AIcs.CLarXiv:2311.06607v42023
  48. Amnesiac Machine Learning

    Laura Graves, Vineel Nagisetty, Vijay Ganesh

    cs.LGcs.AIcs.CRarXiv:2010.10981v12020
  49. LoRA+: Efficient Low Rank Adaptation of Large Models

    Soufiane Hayou, Nikhil Ghosh, Bin Yu

    cs.LGcs.AIcs.CLarXiv:2402.12354v22024
  50. Transformers as Soft Reasoners over Language

    Peter Clark, Oyvind Tafjord, Kyle Richardson

    cs.CLcs.AIarXiv:2002.05867v22020
  51. Analysis of Explainers of Black Box Deep Neural Networks for Computer Vision: A Survey

    Vanessa Buhrmester, David Münch, Michael Arens

    cs.AIcs.CVarXiv:1911.12116v12019
  52. D3VO: Deep Depth, Deep Pose and Deep Uncertainty for Monocular Visual Odometry

    Nan Yang, Lukas von Stumberg, Rui Wang +1

    cs.CVcs.AIarXiv:2003.01060v22020
  53. A Simple Method for Commonsense Reasoning

    Trieu H. Trinh, Quoc V. Le

    cs.AIcs.CLcs.LGarXiv:1806.02847v22018
  54. Deep Reinforcement Learning from Self-Play in Imperfect-Information Games

    Johannes Heinrich, David Silver

    cs.LGcs.AIcs.GTarXiv:1603.01121v22016
  55. Quasi-Recurrent Neural Networks

    James Bradbury, Stephen Merity, Caiming Xiong +1

    cs.NEcs.AIcs.CLarXiv:1611.01576v22016
  56. MeLU: Meta-Learned User Preference Estimator for Cold-Start Recommendation

    Hoyeop Lee, Jinbae Im, Seongwon Jang +2

    cs.IRcs.AIcs.LGarXiv:1908.00413v12019
  57. HDMapNet: An Online HD Map Construction and Evaluation Framework

    Qi Li, Yue Wang, Yilun Wang +1

    cs.CVcs.AIarXiv:2107.06307v42021
  58. CorporateBench: Large-Scale Q&A Benchmarking with Temporal Knowledge Bases

    Sil Hamilton, Albert Yu Sun, Oscar J. Romero +4

    cs.AIcs.CLcs.IRarXiv:2608.27391v12026
  59. Side Adapter Network for Open-Vocabulary Semantic Segmentation

    Mengde Xu, Zheng Zhang, Fangyun Wei +2

    cs.CVcs.AIarXiv:2302.12242v22023
  60. PILOT in the Loop: Live Self-Improvement for Long-Horizon Agents

    Yang Xiao, Yusong Sun, Haoyi Wu +7

    cs.AIarXiv:2608.26530v12026