Human-Computer Interaction
Papers filed under cs.HC on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
481 to 540 of 1,136
ScreenSpot-Pro: GUI Grounding for Professional High-Resolution Computer Use
Kaixin Li, Ziyang Meng, Hongzhan Lin +5
cs.CVcs.HCcs.MMarXiv:2504.07981v12025Accelerating scientific discovery with Co-Scientist
Juraj Gottweis, Wei-Hung Weng, Alexander Daryin +48
cs.AIcs.CLcs.HCarXiv:2502.18864v22025Designing Proactive Thought Partners for Writing
Chao Zhang, Abe Davis, Chih-Wei Chen +1
cs.HCcs.AIcs.CLarXiv:2609.01588v12026"Help Me Help the AI": Understanding How Explainability Can Support Human-AI Interaction
Sunnie S. Y. Kim, Elizabeth Anne Watkins, Olga Russakovsky +2
cs.HCcs.AIcs.CVarXiv:2210.03735v22022Are We There Yet? Assessing Computer-Use Agents for Blind Users' Accessible Interaction with Desktop Applications
Satwik Ram Kodandaram, Monalika Padma Reddy, Xiaojun Bi +3
cs.HCcs.AIarXiv:2609.00524v12026VirSqueezer: Generating Realistic Deformations and Squeezing Dynamics in VR from Fine-Grained Squeezing Controls
Qian Zhang, Xiaoming Chen, Xiaorui Ma +2
cs.HCarXiv:2609.01698v12026Agent Laboratory: Using LLM Agents as Research Assistants
Samuel Schmidgall, Yusheng Su, Ze Wang +7
cs.HCcs.AIcs.CLarXiv:2501.04227v22025Who Should I Trust: AI or Myself? Leveraging Human and AI Correctness Likelihood to Promote Appropriate Trust in AI-Assisted Decision-Making
Shuai Ma, Ying Lei, Xinru Wang +4
cs.HCcs.AIcs.LGarXiv:2301.05809v12023Interactive and Visual Prompt Engineering for Ad-hoc Task Adaptation with Large Language Models
Hendrik Strobelt, Albert Webson, Victor Sanh +4
cs.CLcs.HCcs.LGarXiv:2208.07852v12022Text2Gestures: A Transformer-Based Network for Generating Emotive Body Gestures for Virtual Agents
Uttaran Bhattacharya, Nicholas Rewkowski, Abhishek Banerjee +3
cs.HCcs.GRarXiv:2101.11101v32021Design: One, but in different forms
Willemien Visser
cs.HCarXiv:0708.1725v22007End-to-end Generative Pretraining for Multimodal Video Captioning
Paul Hongsuck Seo, Arsha Nagrani, Anurag Arnab +1
cs.CVcs.AIcs.CLarXiv:2201.08264v22022"It's a Fair Game", or Is It? Examining How Users Navigate Disclosure Risks and Benefits When Using LLM-Based Conversational Agents
Zhiping Zhang, Michelle Jia, Hao-Ping Lee +5
cs.HCcs.AIcs.CRarXiv:2309.11653v22023The State of the Art in Enhancing Trust in Machine Learning Models with the Use of Visualizations
A. Chatzimparmpas, R. Martins, I. Jusufi +3
cs.LGcs.HCstat.MLarXiv:2212.11737v22022Propose to Learn, Learn to Propose: Evaluability-Aware Assistance under Bounded Rationality
Yifan Zhu, Sammie Katt, Samuel Kaski
cs.AIcs.HCcs.MAarXiv:2609.02242v12026UI-TARS: Pioneering Automated GUI Interaction with Native Agents
Yujia Qin, Yining Ye, Junjie Fang +32
cs.AIcs.CLcs.CVarXiv:2501.12326v12025Human-AI Collaboration via Conditional Delegation: A Case Study of Content Moderation
Vivian Lai, Samuel Carton, Rajat Bhatnagar +3
cs.AIcs.HCcs.LGarXiv:2204.11788v12022Occlusion-Robust Multimodal Emotion Recognition in VR via Fusion of Facial Images and EMG
Birgit Nierula, Karam Tomotaki-Dawoud, Mert Akguel +5
cs.CVcs.HCarXiv:2609.03569v12026Beyond Blur: A Semantic Tri-view Pipeline for Teledermatology Gradability via Skin Micro-relief
Robert Engel
eess.IVcs.CVcs.HCarXiv:2609.03095v12026Validation of the Virtual Reality Neuroscience Questionnaire: Maximum Duration of Immersive Virtual Reality Sessions Without the Presence of Pertinent Adverse Symptomatology
Panagiotis Kourtesis, Simona Collina, Leonidas A. A. Doumas +1
cs.HCcs.CYcs.MMarXiv:2101.08146v12021GazeRefine: Expert Gaze as a Test-Time Prompt for Training-Free Medical Image Segmentation
Mohammed Oussama Benyahia, Marouane Tliba, Mohamed Amine Kerkouri +10
eess.IVcs.AIcs.CVarXiv:2609.01310v12026Enabling Conversational Interaction with Mobile UI using Large Language Models
Bryan Wang, Gang Li, Yang Li
cs.HCcs.AIarXiv:2209.08655v22022Ask the Experts: What Should Be on an IoT Privacy and Security Label?
Pardis Emami-Naeini, Yuvraj Agarwal, Lorrie Faith Cranor +1
cs.CYcs.CRcs.HCarXiv:2002.04631v12020TEIDAN: A Multilingual Multiparty Dialogue Corpus
Taiga Mori, Koji Inoue, Mikey Elmers +2
cs.CLcs.HCarXiv:2609.00802v12026A-MEM: Agentic Memory for LLM Agents
Wujiang Xu, Zujie Liang, Kai Mei +3
cs.CLcs.HCarXiv:2502.12110v112025InSight: A Benchmark for Agentic Claim Verification in Interactive Visualizations
Maeve Hutchinson, Syed Mahbubul Huq, Mohammad Albinhassan +3
cs.CLcs.CVcs.HCarXiv:2609.01383v12026Designing Fair AI for Managing Employees in Organizations: A Review, Critique, and Design Agenda
Lionel P. Robert, Casey Pierce, Liz Morris +2
cs.HCcs.AIcs.CYarXiv:2002.09054v12020StoryBuddy: A Human-AI Collaborative Chatbot for Parent-Child Interactive Storytelling with Flexible Parental Involvement
Zheng Zhang, Ying Xu, Yanhao Wang +6
cs.HCcs.AIcs.CLarXiv:2202.06205v22022EEG-Inception: An Accurate and Robust End-to-End Neural Network for EEG-based Motor Imagery Classification
Ce Zhang, Young-Keun Kim, Azim Eskandarian
eess.SPcs.HCcs.LGarXiv:2101.10932v32021Transferring Subspaces Between Subjects in Brain-Computer Interfacing
Wojciech Samek, Frank C. Meinecke, Klaus-Robert Müller
stat.MLcs.HCcs.LGarXiv:1209.4115v22012Graphologue: Exploring Large Language Model Responses with Interactive Diagrams
Peiling Jiang, Jude Rayan, Steven P. Dow +1
cs.HCcs.AIcs.CLarXiv:2305.11473v22023Designing for Responsible Trust in AI Systems: A Communication Perspective
Q. Vera Liao, S. Shyam Sundar
cs.HCcs.AIarXiv:2204.13828v12022Beauty is in the AI of the beholder: MLLMs systematically overrate facial attractiveness
Santiago Grandas, Juan Sebastian Cely-Acosta, Mohit Mendiratta +2
cs.CVcs.HCarXiv:2609.02512v12026Real-time Driver Drowsiness Detection for Android Application Using Deep Neural Networks Techniques
Rateb Jabbar, Khalifa Al-Khalifa, Mohamed Kharbeche +3
cs.CVcs.HCarXiv:1811.01627v12018Bridging the Gulf of Envisioning: Cognitive Design Challenges in LLM Interfaces
Hariharan Subramonyam, Roy Pea, Christopher Lawrence Pondoc +2
cs.HCarXiv:2309.14459v22023Towards a Foundational Ontology for Identifying and Resolving Contradictions in Dialogue-based Human-Robot Interactions
Maitreyee Tewari, Michele Persiani
cs.HCcs.AIarXiv:2609.02364v22026Cascade and Parallel Convolutional Recurrent Neural Networks on EEG-based Intention Recognition for Brain Computer Interface
Dalin Zhang, Lina Yao, Xiang Zhang +3
cs.HCq-bio.NCarXiv:1708.06578v22017Ferret-UI: Grounded Mobile UI Understanding with Multimodal LLMs
Keen You, Haotian Zhang, Eldon Schoop +5
cs.CVcs.CLcs.HCarXiv:2404.05719v12024Sabbath Day Home Automation: "It's Like Mixing Technology and Religion"
Allison Woodruff, Sally Augustin, Brooke Foucault
cs.HCarXiv:0704.3643v12007Simulating Classroom Education with LLM-Empowered Agents
Zheyuan Zhang, Daniel Zhang-Li, Jifan Yu +9
cs.CLcs.HCarXiv:2406.19226v22024Deep Learning for Human Affect Recognition: Insights and New Developments
Philipp V. Rouast, Marc T. P. Adam, Raymond Chiong
cs.LGcs.AIcs.CVarXiv:1901.02884v12019Artificial muses: Generative Artificial Intelligence Chatbots Have Risen to Human-Level Creativity
Jennifer Haase, Paul H. P. Hanel
cs.AIcs.HCarXiv:2303.12003v12023Extension of Technology Acceptance Model by using System Usability Scale to assess behavioral intention to use e-learning
Anastasia Revythi, Nikolaos Tselios
cs.HCcs.CYarXiv:1704.06127v52017Joint Activity Recognition and Indoor Localization with WiFi Fingerprints
Fei Wang, Jianwei Feng, Yinliang Zhao +3
cs.HCarXiv:1904.04964v22019The PIONEER Project: A PrIvacy companion for mOtivatioN and knowlEdge transfER
Simon Althaus, Nina Gerber, Sara Hahn +5
cs.HCcs.CYarXiv:2609.02700v12026You Only Look at Screens: Multimodal Chain-of-Action Agents
Zhuosheng Zhang, Aston Zhang
cs.CLcs.AIcs.HCarXiv:2309.11436v42023Accessible Visualization via Natural Language Descriptions: A Four-Level Model of Semantic Content
Alan Lundgard, Arvind Satyanarayan
cs.HCcs.CLarXiv:2110.04406v12021Smart Guiding Glasses for Visually Impaired People in Indoor Environment
Jinqiang Bai, Shiguo Lian, Zhaoxiang Liu +2
cs.HCarXiv:1709.09359v12017Autonomous GIS: the next-generation AI-powered GIS
Zhenlong Li, Huan Ning
cs.AIcs.HCarXiv:2305.06453v42023Prompt Problems: A New Programming Exercise for the Generative AI Era
Paul Denny, Juho Leinonen, James Prather +4
cs.HCarXiv:2311.05943v12023Frame attention networks for facial expression recognition in videos
Debin Meng, Xiaojiang Peng, Kai Wang +1
cs.CVcs.HCcs.MMarXiv:1907.00193v22019Collective Constitutional AI: Aligning a Language Model with Public Input
Saffron Huang, Divya Siddarth, Liane Lovitt +4
cs.AIcs.CLcs.HCarXiv:2406.07814v12024Promptify: Text-to-Image Generation through Interactive Prompt Exploration with Large Language Models
Stephen Brade, Bryan Wang, Mauricio Sousa +2
cs.HCcs.AIcs.MMarXiv:2304.09337v12023Virtual World, Defined from a Technological Perspective, and Applied to Video Games, Mixed Reality and the Metaverse
Kim J. L. Nevelsteen
cs.HCcs.CYarXiv:1511.08464v22015Reward-rational (implicit) choice: A unifying formalism for reward learning
Hong Jun Jeon, Smitha Milli, Anca D. Dragan
cs.LGcs.AIcs.HCarXiv:2002.04833v42020Understanding the Role of Human Intuition on Reliance in Human-AI Decision-Making with Explanations
Valerie Chen, Q. Vera Liao, Jennifer Wortman Vaughan +1
cs.HCcs.AIarXiv:2301.07255v32023Gesticulator: A framework for semantically-aware speech-driven gesture generation
Taras Kucherenko, Patrik Jonell, Sanne van Waveren +4
cs.HCcs.LGeess.ASarXiv:2001.09326v52020Removing Speech, Keeping Activities: A Privacy Firewall for Acoustic Sensing in Assisted Living
Pavlos Nicolaou, Christos Efstratiou
cs.SDcs.CRcs.HCarXiv:2609.02376v12026Understanding Large-Language Model (LLM)-powered Human-Robot Interaction
Callie Y. Kim, Christine P. Lee, Bilge Mutlu
cs.ROcs.HCarXiv:2401.03217v12024Explainable AI is Dead, Long Live Explainable AI! Hypothesis-driven decision support
Tim Miller
cs.AIcs.HCarXiv:2302.12389v32023