Human-Computer Interaction

Papers filed under cs.HC on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

481 to 540 of 1,136

  1. ScreenSpot-Pro: GUI Grounding for Professional High-Resolution Computer Use

    Kaixin Li, Ziyang Meng, Hongzhan Lin +5

    cs.CVcs.HCcs.MMarXiv:2504.07981v12025
  2. Accelerating scientific discovery with Co-Scientist

    Juraj Gottweis, Wei-Hung Weng, Alexander Daryin +48

    cs.AIcs.CLcs.HCarXiv:2502.18864v22025
  3. Designing Proactive Thought Partners for Writing

    Chao Zhang, Abe Davis, Chih-Wei Chen +1

    cs.HCcs.AIcs.CLarXiv:2609.01588v12026
  4. "Help Me Help the AI": Understanding How Explainability Can Support Human-AI Interaction

    Sunnie S. Y. Kim, Elizabeth Anne Watkins, Olga Russakovsky +2

    cs.HCcs.AIcs.CVarXiv:2210.03735v22022
  5. Are We There Yet? Assessing Computer-Use Agents for Blind Users' Accessible Interaction with Desktop Applications

    Satwik Ram Kodandaram, Monalika Padma Reddy, Xiaojun Bi +3

    cs.HCcs.AIarXiv:2609.00524v12026
  6. VirSqueezer: Generating Realistic Deformations and Squeezing Dynamics in VR from Fine-Grained Squeezing Controls

    Qian Zhang, Xiaoming Chen, Xiaorui Ma +2

    cs.HCarXiv:2609.01698v12026
  7. Agent Laboratory: Using LLM Agents as Research Assistants

    Samuel Schmidgall, Yusheng Su, Ze Wang +7

    cs.HCcs.AIcs.CLarXiv:2501.04227v22025
  8. Who Should I Trust: AI or Myself? Leveraging Human and AI Correctness Likelihood to Promote Appropriate Trust in AI-Assisted Decision-Making

    Shuai Ma, Ying Lei, Xinru Wang +4

    cs.HCcs.AIcs.LGarXiv:2301.05809v12023
  9. Interactive and Visual Prompt Engineering for Ad-hoc Task Adaptation with Large Language Models

    Hendrik Strobelt, Albert Webson, Victor Sanh +4

    cs.CLcs.HCcs.LGarXiv:2208.07852v12022
  10. Text2Gestures: A Transformer-Based Network for Generating Emotive Body Gestures for Virtual Agents

    Uttaran Bhattacharya, Nicholas Rewkowski, Abhishek Banerjee +3

    cs.HCcs.GRarXiv:2101.11101v32021
  11. Design: One, but in different forms

    Willemien Visser

    cs.HCarXiv:0708.1725v22007
  12. End-to-end Generative Pretraining for Multimodal Video Captioning

    Paul Hongsuck Seo, Arsha Nagrani, Anurag Arnab +1

    cs.CVcs.AIcs.CLarXiv:2201.08264v22022
  13. "It's a Fair Game", or Is It? Examining How Users Navigate Disclosure Risks and Benefits When Using LLM-Based Conversational Agents

    Zhiping Zhang, Michelle Jia, Hao-Ping Lee +5

    cs.HCcs.AIcs.CRarXiv:2309.11653v22023
  14. The State of the Art in Enhancing Trust in Machine Learning Models with the Use of Visualizations

    A. Chatzimparmpas, R. Martins, I. Jusufi +3

    cs.LGcs.HCstat.MLarXiv:2212.11737v22022
  15. Propose to Learn, Learn to Propose: Evaluability-Aware Assistance under Bounded Rationality

    Yifan Zhu, Sammie Katt, Samuel Kaski

    cs.AIcs.HCcs.MAarXiv:2609.02242v12026
  16. UI-TARS: Pioneering Automated GUI Interaction with Native Agents

    Yujia Qin, Yining Ye, Junjie Fang +32

    cs.AIcs.CLcs.CVarXiv:2501.12326v12025
  17. Human-AI Collaboration via Conditional Delegation: A Case Study of Content Moderation

    Vivian Lai, Samuel Carton, Rajat Bhatnagar +3

    cs.AIcs.HCcs.LGarXiv:2204.11788v12022
  18. Occlusion-Robust Multimodal Emotion Recognition in VR via Fusion of Facial Images and EMG

    Birgit Nierula, Karam Tomotaki-Dawoud, Mert Akguel +5

    cs.CVcs.HCarXiv:2609.03569v12026
  19. Beyond Blur: A Semantic Tri-view Pipeline for Teledermatology Gradability via Skin Micro-relief

    Robert Engel

    eess.IVcs.CVcs.HCarXiv:2609.03095v12026
  20. Validation of the Virtual Reality Neuroscience Questionnaire: Maximum Duration of Immersive Virtual Reality Sessions Without the Presence of Pertinent Adverse Symptomatology

    Panagiotis Kourtesis, Simona Collina, Leonidas A. A. Doumas +1

    cs.HCcs.CYcs.MMarXiv:2101.08146v12021
  21. GazeRefine: Expert Gaze as a Test-Time Prompt for Training-Free Medical Image Segmentation

    Mohammed Oussama Benyahia, Marouane Tliba, Mohamed Amine Kerkouri +10

    eess.IVcs.AIcs.CVarXiv:2609.01310v12026
  22. Enabling Conversational Interaction with Mobile UI using Large Language Models

    Bryan Wang, Gang Li, Yang Li

    cs.HCcs.AIarXiv:2209.08655v22022
  23. Ask the Experts: What Should Be on an IoT Privacy and Security Label?

    Pardis Emami-Naeini, Yuvraj Agarwal, Lorrie Faith Cranor +1

    cs.CYcs.CRcs.HCarXiv:2002.04631v12020
  24. TEIDAN: A Multilingual Multiparty Dialogue Corpus

    Taiga Mori, Koji Inoue, Mikey Elmers +2

    cs.CLcs.HCarXiv:2609.00802v12026
  25. A-MEM: Agentic Memory for LLM Agents

    Wujiang Xu, Zujie Liang, Kai Mei +3

    cs.CLcs.HCarXiv:2502.12110v112025
  26. InSight: A Benchmark for Agentic Claim Verification in Interactive Visualizations

    Maeve Hutchinson, Syed Mahbubul Huq, Mohammad Albinhassan +3

    cs.CLcs.CVcs.HCarXiv:2609.01383v12026
  27. Designing Fair AI for Managing Employees in Organizations: A Review, Critique, and Design Agenda

    Lionel P. Robert, Casey Pierce, Liz Morris +2

    cs.HCcs.AIcs.CYarXiv:2002.09054v12020
  28. StoryBuddy: A Human-AI Collaborative Chatbot for Parent-Child Interactive Storytelling with Flexible Parental Involvement

    Zheng Zhang, Ying Xu, Yanhao Wang +6

    cs.HCcs.AIcs.CLarXiv:2202.06205v22022
  29. EEG-Inception: An Accurate and Robust End-to-End Neural Network for EEG-based Motor Imagery Classification

    Ce Zhang, Young-Keun Kim, Azim Eskandarian

    eess.SPcs.HCcs.LGarXiv:2101.10932v32021
  30. Transferring Subspaces Between Subjects in Brain-Computer Interfacing

    Wojciech Samek, Frank C. Meinecke, Klaus-Robert Müller

    stat.MLcs.HCcs.LGarXiv:1209.4115v22012
  31. Graphologue: Exploring Large Language Model Responses with Interactive Diagrams

    Peiling Jiang, Jude Rayan, Steven P. Dow +1

    cs.HCcs.AIcs.CLarXiv:2305.11473v22023
  32. Designing for Responsible Trust in AI Systems: A Communication Perspective

    Q. Vera Liao, S. Shyam Sundar

    cs.HCcs.AIarXiv:2204.13828v12022
  33. Beauty is in the AI of the beholder: MLLMs systematically overrate facial attractiveness

    Santiago Grandas, Juan Sebastian Cely-Acosta, Mohit Mendiratta +2

    cs.CVcs.HCarXiv:2609.02512v12026
  34. Real-time Driver Drowsiness Detection for Android Application Using Deep Neural Networks Techniques

    Rateb Jabbar, Khalifa Al-Khalifa, Mohamed Kharbeche +3

    cs.CVcs.HCarXiv:1811.01627v12018
  35. Bridging the Gulf of Envisioning: Cognitive Design Challenges in LLM Interfaces

    Hariharan Subramonyam, Roy Pea, Christopher Lawrence Pondoc +2

    cs.HCarXiv:2309.14459v22023
  36. Towards a Foundational Ontology for Identifying and Resolving Contradictions in Dialogue-based Human-Robot Interactions

    Maitreyee Tewari, Michele Persiani

    cs.HCcs.AIarXiv:2609.02364v22026
  37. Cascade and Parallel Convolutional Recurrent Neural Networks on EEG-based Intention Recognition for Brain Computer Interface

    Dalin Zhang, Lina Yao, Xiang Zhang +3

    cs.HCq-bio.NCarXiv:1708.06578v22017
  38. Ferret-UI: Grounded Mobile UI Understanding with Multimodal LLMs

    Keen You, Haotian Zhang, Eldon Schoop +5

    cs.CVcs.CLcs.HCarXiv:2404.05719v12024
  39. Sabbath Day Home Automation: "It's Like Mixing Technology and Religion"

    Allison Woodruff, Sally Augustin, Brooke Foucault

    cs.HCarXiv:0704.3643v12007
  40. Simulating Classroom Education with LLM-Empowered Agents

    Zheyuan Zhang, Daniel Zhang-Li, Jifan Yu +9

    cs.CLcs.HCarXiv:2406.19226v22024
  41. Deep Learning for Human Affect Recognition: Insights and New Developments

    Philipp V. Rouast, Marc T. P. Adam, Raymond Chiong

    cs.LGcs.AIcs.CVarXiv:1901.02884v12019
  42. Artificial muses: Generative Artificial Intelligence Chatbots Have Risen to Human-Level Creativity

    Jennifer Haase, Paul H. P. Hanel

    cs.AIcs.HCarXiv:2303.12003v12023
  43. Extension of Technology Acceptance Model by using System Usability Scale to assess behavioral intention to use e-learning

    Anastasia Revythi, Nikolaos Tselios

    cs.HCcs.CYarXiv:1704.06127v52017
  44. Joint Activity Recognition and Indoor Localization with WiFi Fingerprints

    Fei Wang, Jianwei Feng, Yinliang Zhao +3

    cs.HCarXiv:1904.04964v22019
  45. The PIONEER Project: A PrIvacy companion for mOtivatioN and knowlEdge transfER

    Simon Althaus, Nina Gerber, Sara Hahn +5

    cs.HCcs.CYarXiv:2609.02700v12026
  46. You Only Look at Screens: Multimodal Chain-of-Action Agents

    Zhuosheng Zhang, Aston Zhang

    cs.CLcs.AIcs.HCarXiv:2309.11436v42023
  47. Accessible Visualization via Natural Language Descriptions: A Four-Level Model of Semantic Content

    Alan Lundgard, Arvind Satyanarayan

    cs.HCcs.CLarXiv:2110.04406v12021
  48. Smart Guiding Glasses for Visually Impaired People in Indoor Environment

    Jinqiang Bai, Shiguo Lian, Zhaoxiang Liu +2

    cs.HCarXiv:1709.09359v12017
  49. Autonomous GIS: the next-generation AI-powered GIS

    Zhenlong Li, Huan Ning

    cs.AIcs.HCarXiv:2305.06453v42023
  50. Prompt Problems: A New Programming Exercise for the Generative AI Era

    Paul Denny, Juho Leinonen, James Prather +4

    cs.HCarXiv:2311.05943v12023
  51. Frame attention networks for facial expression recognition in videos

    Debin Meng, Xiaojiang Peng, Kai Wang +1

    cs.CVcs.HCcs.MMarXiv:1907.00193v22019
  52. Collective Constitutional AI: Aligning a Language Model with Public Input

    Saffron Huang, Divya Siddarth, Liane Lovitt +4

    cs.AIcs.CLcs.HCarXiv:2406.07814v12024
  53. Promptify: Text-to-Image Generation through Interactive Prompt Exploration with Large Language Models

    Stephen Brade, Bryan Wang, Mauricio Sousa +2

    cs.HCcs.AIcs.MMarXiv:2304.09337v12023
  54. Virtual World, Defined from a Technological Perspective, and Applied to Video Games, Mixed Reality and the Metaverse

    Kim J. L. Nevelsteen

    cs.HCcs.CYarXiv:1511.08464v22015
  55. Reward-rational (implicit) choice: A unifying formalism for reward learning

    Hong Jun Jeon, Smitha Milli, Anca D. Dragan

    cs.LGcs.AIcs.HCarXiv:2002.04833v42020
  56. Understanding the Role of Human Intuition on Reliance in Human-AI Decision-Making with Explanations

    Valerie Chen, Q. Vera Liao, Jennifer Wortman Vaughan +1

    cs.HCcs.AIarXiv:2301.07255v32023
  57. Gesticulator: A framework for semantically-aware speech-driven gesture generation

    Taras Kucherenko, Patrik Jonell, Sanne van Waveren +4

    cs.HCcs.LGeess.ASarXiv:2001.09326v52020
  58. Removing Speech, Keeping Activities: A Privacy Firewall for Acoustic Sensing in Assisted Living

    Pavlos Nicolaou, Christos Efstratiou

    cs.SDcs.CRcs.HCarXiv:2609.02376v12026
  59. Understanding Large-Language Model (LLM)-powered Human-Robot Interaction

    Callie Y. Kim, Christine P. Lee, Bilge Mutlu

    cs.ROcs.HCarXiv:2401.03217v12024
  60. Explainable AI is Dead, Long Live Explainable AI! Hypothesis-driven decision support

    Tim Miller

    cs.AIcs.HCarXiv:2302.12389v32023