HCI.TOPHCI, made easy
HomearXivPapersInstitutionsAuthorsGuidelinesQuestionsHandbookInnovationEvents
HomearXivPapersInstitutionsAuthorsGuidelinesQuestionsHandbookInnovationEvents
Data methodologyHCI conferencesHCI papersAbout Xue ZhirongWelcome to cooperate
search
Active Filters
search
All

Papers

Browse and search HCI research papers from All

Active Filters
Institution: UCLA
44 results

CoLyricist: Enhancing Lyric Writing with AI through Workflow-Aligned Support

We propose CoLyricist, an AI-assisted lyric writing tool designed to support the typical workflows of experienced lyricists and enhance their creative efficiency. While lyricists have unique processes, many follow common stages. Tools that fail to accommodate these stages challenge integration into creative practices.…

MY
Masahiro Yoshida et al.University of California, Los Angeles

Behavioral Indicators of Overreliance During Interaction with Conversational Language Models

LLMs are now embedded in a wide range of everyday scenarios. However, their inherent hallucinations risk hiding misinformation in fluent responses, raising concerns about overreliance on AI. Detecting overreliance is challenging, as it often arises in complex, dynamic contexts and cannot be easily captured by post-hoc…

CL
Chang Liu et al.Tsinghua University

CoSight: Exploring Viewer Contributions to Online Video Accessibility Through Descriptive Commenting

The rapid growth of online video content has outpaced efforts to make visual information accessible to blind and low vision (BLV) audiences. While professional Audio Description (AD) remains the gold standard, it is costly and difficult to scale across the vast volume of online media. In this work, we explore a comple…

RW
Ruolin Wang et al.University of California

The GenUI Study: Exploring the Design of Generative UI Tools to Support UX Practitioners and Beyond

AI can now generate high-fidelity UI mock-up screens from a high-level textual description, promising to support UX practitioners' work. However, it remains unclear how UX practitioners would adopt such Generative UI (GenUI) models in a way that is integral and beneficial to their work. To answer this question, we con…

XC
Xiang 'Anthony' Chen et al.University of California, Los Angeles
AdRecommended

Learn AI Coding at CodeNow

Structured lessons, hands-on projects, and continuous updates for people bringing AI into real development work.

Explore Nowopen_in_new

Empowering Medical Data Labeling for Non-Experts with DANNY: Enhancing Accuracy and Mitigating Over-Reliance on AI

Economic constraints on recruiting experts hinder efforts to build qualified datasets for utilizing AI in professional domains (e.g., medical diagnosis), which could provide societal benefits. To solve this issue, previous studies introduced crowdsourcing and AI to enable non-experts to perform expert-level data label…

YJ
Youngseung Jeon et al.University of California

HEPHA: A Mixed-Initiative Image Labeling Tool for Specialized Domains

Image labeling is an important task for training computer vision models. In specialized domains, such as healthcare, it is expensive and challenging to recruit specialists for image labeling. We propose HEPHA, a mixed-initiative image labeling tool that elicits human expertise via inductive logic learning to infer and…

SZ
Shiyuan Zhou et al.Computer Science

Mentigo: An Intelligent Agent for Mentoring Students in the Creative Problem Solving Process

Creative Problem-Solving (CPS) promotes creative and critical thinking while enhancing real-world problem-solving skills, making it essential for middle school education. However, providing personalized mentorship in CPS projects at scale is challenging due to resource constraints and diverse student needs. To address…

SZ
Siyu Zha et al.Tsinghua University

Proactive Conversational Agents with Inner Thoughts

One of the long-standing aspirations in conversational AI is to allow them to autonomously take initiatives in conversations, i.e. being proactive. This is especially challenging for multi-party conversations. Prior NLP research focused mainly on predicting the next speaker from contexts like preceding conversations.…

XL
Xingyu Bruce Liu et al.University of California Los Angeles

"Pinocchio had a Nose, You have a Network!": On Characterizing Fake News Spreaders on Arabic Social Media

The detection and analysis of fake news and its origins has become a main task associated with the overall objective of social media regulation in recent years. The majority of work was dedicated towards detecting misinformation with some focus on analyzing the flow of fake news over social networks. However, there is…

MF
Mahmoud Fawzi et al.University of Edinburgh
Session 2e: Echo Chambers and Fake News in Focus

Human I/O: Towards a Unified Approach to Detecting Situational Impairments

Situationally Induced Impairments and Disabilities (SIIDs) can significantly hinder user experience in contexts such as poor lighting, noise, and multi-tasking. While prior research has introduced algorithms and systems to address these impairments, they predominantly cater to specific tasks or environments and fail t…

XL
Xingyu Bruce Liu et al.University of California

WheelPose: Data Synthesis Techniques to Improve Pose Estimation Performance on Wheelchair Users

Existing pose estimation models perform poorly on wheelchair users due to a lack of representation in training data. We present a data synthesis pipeline to address this disparity in data collection and subsequently improve pose estimation performance for wheelchair users. Our configurable pipeline generates synthetic…

WH
William Huang et al.
University of California, Los Angeles

From Text to Pixels: Enhancing User Understanding through Text-to-Image Model Explanations

Recent progress in Text-to-Image (T2I) models promises transformative applications in art, design, education, medicine, and entertainment. These models, exemplified by Dall-e, Imagen, and Stable Diffusion, have the potential to revolutionize various industries. However, a primary concern is their operation as a 'black…

NE
Noyan Evirgen et al.University of California

Fingerprinting IoT Devices Using Latent Physical Side-Channels

The proliferation of low-end low-power internet-of-things (IoT) devices in "smart" environments necessitates secure identification and authentication of these devices via low-overhead fingerprinting methods. Previous work typically utilizes characteristics of the device's wireless modulation (WiFi, BLE, etc.) in the s…

JF
JUSTIN FENG et al.University of California, Los Angeles

XCreation: A Graph-Based Crossmodal Generative Creativity Support Tool

Creativity Support Tools (CSTs) aid in the efficient and effective composition of creative content, such as picture books. However, many existing CSTs only allow for mono-modal creation, whereas previous studies have become theoretically and technically mature to support multi-modal innovative creations. To overcome t…

ZY
Zihan Yan et al.Zhejiang University

NaCanva: Exploring and Enabling the Nature-Inspired Creativity for Children

Nature has been a bountiful source of materials, replenishment, inspiration, and creativity. Nature collage, as a crafting technique, offers children a fun and educational way to explore nature and express their creativity. However, the collection of raw material has been limited to static objects like leaves, ignorin…

ZY
Zihan Yan et al.Tsinghua University

Augmenting Pathologists with NaviPath: Design and Evaluation of a Human-AI Collaborative Navigation System

Artificial Intelligence (AI) brings advancements to support pathologists in navigating high-resolution tumor images to search for pathology patterns of interest. However, existing AI-assisted tools have not realized this promised potential due to a lack of insight into pathology and HCI considerations for pathologists…

HG
Hongyan Gu et al.University of California

AVscript: Accessible Video Editing with Audio-Visual Scripts

Sighted and blind and low vision (BLV) creators alike use videos to communicate with broad audiences. Yet, video editing remains inaccessible to BLV creators. Our formative study revealed that current video editing tools make it difficult to access the visual content, assess the visual quality, and efficiently navi…

MH
Mina Huh et al.University of Texas at Austin

Designing and Evaluating Interfaces that Highlight News Coverage Diversity Using Discord Questions

Modern news aggregators do the hard work of organizing a large news stream, creating collections for a given news story with tens of source options. This paper shows that navigating large source collections for a news story can be challenging without further guidance. In this work, we design three interfaces -- the A…

PL
Philippe Laban et al.Salesforce

Visual Captions: Augmenting Verbal Communication with On-the-fly Visuals

Video conferencing solutions like Zoom, Google Meet, and Microsoft Teams are becoming increasingly popular for facilitating conversations, and recent advancements such as live captioning help people better understand each other. We believe that the addition of visuals based on the context of conversations could furthe…

XL
Xingyu Bruce Liu et al.University of California Los Angeles

GANravel: User-Driven Direction Disentanglement in Generative Adversarial Networks

Generative adversarial networks (GANs) have many application areas including image editing, domain translation, missing data imputation, and support for creative work. However, GANs are considered `black boxes'. Specifically, the end-users have little control over how to improve editing directions through disentanglem…

NE
Noyan Evirgen et al.HCI Group
Paper TitleAuthorsResearch TopicsPaper DatabaseYear

CoLyricist: Enhancing Lyric Writing with AI through Workflow-Aligned Support

We propose CoLyricist, an AI-assisted lyric writing tool designed to support the typical workflows of experienced lyricists and enhance their creative efficiency. While lyricists have unique processes, many follow common stages. Tools that fail to accommodate these stages challenge integration into creative practices.…

MY
Masahiro Yoshida et al.University of California, Los Angeles

Behavioral Indicators of Overreliance During Interaction with Conversational Language Models

LLMs are now embedded in a wide range of everyday scenarios. However, their inherent hallucinations risk hiding misinformation in fluent responses, raising concerns about overreliance on AI. Detecting overreliance is challenging, as it often arises in complex, dynamic contexts and cannot be easily captured by post-hoc…

CL
Chang Liu et al.Tsinghua University

CoSight: Exploring Viewer Contributions to Online Video Accessibility Through Descriptive Commenting

The rapid growth of online video content has outpaced efforts to make visual information accessible to blind and low vision (BLV) audiences. While professional Audio Description (AD) remains the gold standard, it is costly and difficult to scale across the vast volume of online media. In this work, we explore a comple…

RW
Ruolin Wang et al.University of California

The GenUI Study: Exploring the Design of Generative UI Tools to Support UX Practitioners and Beyond

AI can now generate high-fidelity UI mock-up screens from a high-level textual description, promising to support UX practitioners' work. However, it remains unclear how UX practitioners would adopt such Generative UI (GenUI) models in a way that is integral and beneficial to their work. To answer this question, we con…

XC
Xiang 'Anthony' Chen et al.University of California, Los Angeles
AdRecommended

Learn AI Coding at CodeNow

Structured lessons, hands-on projects, and continuous updates for people bringing AI into real development work.

Explore Nowopen_in_new

Empowering Medical Data Labeling for Non-Experts with DANNY: Enhancing Accuracy and Mitigating Over-Reliance on AI

Economic constraints on recruiting experts hinder efforts to build qualified datasets for utilizing AI in professional domains (e.g., medical diagnosis), which could provide societal benefits. To solve this issue, previous studies introduced crowdsourcing and AI to enable non-experts to perform expert-level data label…

YJ
Youngseung Jeon et al.University of California

HEPHA: A Mixed-Initiative Image Labeling Tool for Specialized Domains

Image labeling is an important task for training computer vision models. In specialized domains, such as healthcare, it is expensive and challenging to recruit specialists for image labeling. We propose HEPHA, a mixed-initiative image labeling tool that elicits human expertise via inductive logic learning to infer and…

SZ
Shiyuan Zhou et al.Computer Science

Mentigo: An Intelligent Agent for Mentoring Students in the Creative Problem Solving Process

Creative Problem-Solving (CPS) promotes creative and critical thinking while enhancing real-world problem-solving skills, making it essential for middle school education. However, providing personalized mentorship in CPS projects at scale is challenging due to resource constraints and diverse student needs. To address…

SZ
Siyu Zha et al.Tsinghua University

Proactive Conversational Agents with Inner Thoughts

One of the long-standing aspirations in conversational AI is to allow them to autonomously take initiatives in conversations, i.e. being proactive. This is especially challenging for multi-party conversations. Prior NLP research focused mainly on predicting the next speaker from contexts like preceding conversations.…

XL
Xingyu Bruce Liu et al.University of California Los Angeles

"Pinocchio had a Nose, You have a Network!": On Characterizing Fake News Spreaders on Arabic Social Media

The detection and analysis of fake news and its origins has become a main task associated with the overall objective of social media regulation in recent years. The majority of work was dedicated towards detecting misinformation with some focus on analyzing the flow of fake news over social networks. However, there is…

MF
Mahmoud Fawzi et al.University of Edinburgh
Session 2e: Echo Chambers and Fake News in Focus
emoji_events

Human I/O: Towards a Unified Approach to Detecting Situational Impairments

Situationally Induced Impairments and Disabilities (SIIDs) can significantly hinder user experience in contexts such as poor lighting, noise, and multi-tasking. While prior research has introduced algorithms and systems to address these impairments, they predominantly cater to specific tasks or environments and fail t…

XL
Xingyu Bruce Liu et al.University of California

WheelPose: Data Synthesis Techniques to Improve Pose Estimation Performance on Wheelchair Users

Existing pose estimation models perform poorly on wheelchair users due to a lack of representation in training data. We present a data synthesis pipeline to address this disparity in data collection and subsequently improve pose estimation performance for wheelchair users. Our configurable pipeline generates synthetic…

WH
William Huang et al.University of California, Los Angeles

From Text to Pixels: Enhancing User Understanding through Text-to-Image Model Explanations

Recent progress in Text-to-Image (T2I) models promises transformative applications in art, design, education, medicine, and entertainment. These models, exemplified by Dall-e, Imagen, and Stable Diffusion, have the potential to revolutionize various industries. However, a primary concern is their operation as a 'black…

NE
Noyan Evirgen et al.University of California

Fingerprinting IoT Devices Using Latent Physical Side-Channels

The proliferation of low-end low-power internet-of-things (IoT) devices in "smart" environments necessitates secure identification and authentication of these devices via low-overhead fingerprinting methods. Previous work typically utilizes characteristics of the device's wireless modulation (WiFi, BLE, etc.) in the s…

JF
JUSTIN FENG et al.University of California, Los Angeles

XCreation: A Graph-Based Crossmodal Generative Creativity Support Tool

Creativity Support Tools (CSTs) aid in the efficient and effective composition of creative content, such as picture books. However, many existing CSTs only allow for mono-modal creation, whereas previous studies have become theoretically and technically mature to support multi-modal innovative creations. To overcome t…

ZY
Zihan Yan et al.Zhejiang University

NaCanva: Exploring and Enabling the Nature-Inspired Creativity for Children

Nature has been a bountiful source of materials, replenishment, inspiration, and creativity. Nature collage, as a crafting technique, offers children a fun and educational way to explore nature and express their creativity. However, the collection of raw material has been limited to static objects like leaves, ignorin…

ZY
Zihan Yan et al.Tsinghua University
emoji_events

Augmenting Pathologists with NaviPath: Design and Evaluation of a Human-AI Collaborative Navigation System

Artificial Intelligence (AI) brings advancements to support pathologists in navigating high-resolution tumor images to search for pathology patterns of interest. However, existing AI-assisted tools have not realized this promised potential due to a lack of insight into pathology and HCI considerations for pathologists…

HG
Hongyan Gu et al.University of California

AVscript: Accessible Video Editing with Audio-Visual Scripts

Sighted and blind and low vision (BLV) creators alike use videos to communicate with broad audiences. Yet, video editing remains inaccessible to BLV creators. Our formative study revealed that current video editing tools make it difficult to access the visual content, assess the visual quality, and efficiently navi…

MH
Mina Huh et al.University of Texas at Austin

Designing and Evaluating Interfaces that Highlight News Coverage Diversity Using Discord Questions

Modern news aggregators do the hard work of organizing a large news stream, creating collections for a given news story with tens of source options. This paper shows that navigating large source collections for a news story can be challenging without further guidance. In this work, we design three interfaces -- the A…

PL
Philippe Laban et al.Salesforce

Visual Captions: Augmenting Verbal Communication with On-the-fly Visuals

Video conferencing solutions like Zoom, Google Meet, and Microsoft Teams are becoming increasingly popular for facilitating conversations, and recent advancements such as live captioning help people better understand each other. We believe that the addition of visuals based on the context of conversations could furthe…

XL
Xingyu Bruce Liu et al.University of California Los Angeles

GANravel: User-Driven Direction Disentanglement in Generative Adversarial Networks

Generative adversarial networks (GANs) have many application areas including image editing, domain translation, missing data imputation, and support for creative work. However, GANs are considered `black boxes'. Specifically, the end-users have little control over how to improve editing directions through disentanglem…

NE
Noyan Evirgen et al.HCI Group