HCI.TOPHCI, made easy
HomearXivPapersInstitutionsAuthorsGuidelinesQuestionsHandbookInnovationEvents
HomearXivPapersInstitutionsAuthorsGuidelinesQuestionsHandbookInnovationEvents
Data methodologyHCI conferencesHCI papersAbout Xue ZhirongWelcome to cooperate
search
Active Filters
search
All

Papers

Browse and search HCI research papers from All

Active Filters
Author: 3803
11 results

HiSync: Spatio-Temporally Aligning Hand Motion from Wearable IMU and On-Robot Camera for Command Source Identification in Long-Range HRI

Long-range Human-Robot Interaction (HRI) remains underexplored. Within it, Command Source Identification (CSI) – determining who issued a command – is especially challenging due to multi-user and distance-induced sensor ambiguity. We introduce HiSync, an optical-inertial fusion framework that treats hand motion as bin…

CZ
Chengwen Zhang et al.Tsinghua University

TraceRing: Touchpad-like Pointing with a Single IMU Ring through Personalized Learning

Achieving touchpad-like pointing with a single IMU ring is highly desirable for portable and wearable interaction, yet challenging due to incomplete motion data and significant user variability. We present TraceRing, a finger-worn IMU system that enables precise two-dimensional cursor control. To address the limitatio…

ZH
Zhe He et al.Tsinghua University

3DRing: Enabling Low-Cost 3D Hand Position Tracking by Fusing Inertial and Low-Framerate Optical Sensing

Current mobile hand tracking systems primarily rely on high-framerate (HFR) optical sensors to capture hand positions, resulting in high computational cost and limiting the applicability in end devices. We propose 3DRing, a 3D hand position tracking method that requires only low-framerate (LFR, <10 FPS) optical data a…

ZL
Zhuojun Li et al.Tsinghua University

InterQuest: A Mixed-Initiative Framework for Dynamic User Interest Modeling in Conversational Search

In online information-seeking tasks (e.g., for products and restaurants), users seek information that aligns with their individual preferences to make informed decisions. However, existing systems often struggle to infer users' implicit interests—unstated yet essential preference factors that directly impact decision…

YM
Yu Mei et al.Tsinghua University
AdRecommended

Learn AI Coding at CodeNow

Structured lessons, hands-on projects, and continuous updates for people bringing AI into real development work.

Explore Nowopen_in_new

From Operation to Cognition: Automatic Modeling Cognitive Dependencies from User Demonstrations for GUI Task Automation

Traditional Programming by Demonstration (PBD) systems primarily automate tasks by recording and replaying operations on Graphical User Interfaces (GUIs), without fully considering the cognitive processes behind operations. This limits their ability to generalize tasks with interdependent operations to new contexts (e…

YY
Yiwen Yin et al.Tsinghua University

Investigating Context-Aware Collaborative Text Entry on Smartphones using Large Language Models

Text entry is a fundamental and ubiquitous task, but users often face challenges such as situational impairments or difficulties in sentence formulation. Motivated by this, we explore the potential of large language models (LLMs) to assist with text entry in real-world contexts. We propose a collaborative smartphone-b…

WC
Weihao Chen et al.Tsinghua University

AngleSizer: Enhancing Spatial Scale Perception for the Visually Impaired with an Interactive Smartphone Assistant

Jing 等人开发 AngleSizer 交互式智能手机助手,通过创新方式帮助视障人士感知空间尺度,提升其对物体大小和距离的理解能力。

XJ
Xiaoqing Jing et al.

ContextCam: Bridging Context Awareness with Creative Human-AI Image Co-Creation

The rapid advancement of AI-generated content (AIGC) promises to transform various aspects of human life significantly. This work particularly focuses on the potential of AIGC to revolutionize image creation, such as photography and self-expression. We introduce ContextCam, a novel human-AI image co-creation system th…

XF
Xianzhe Fan et al.Tsinghua University

From Gap to Synergy: Enhancing Contextual Understanding through Human-Machine Collaboration in Personalized Systems

This paper presents LangAware, a collaborative approach for constructing personalized context for context-aware applications. The need for personalization arises due to significant variations in context between individuals based on scenarios, devices, and preferences. However, there is often a notable gap between huma…

WC
Weihao Chen et al.Tsinghua University

VIPBoard: Improving Screen-Reader Keyboard for Visually Impaired People with Character-Level Auto Correction

Modern touchscreen keyboards are all powered by the word-level auto-correction ability to handle input errors. Unfortunately, visually impaired users are deprived of such benefit because a screen-reader keyboard offers only character-level input and provides no correction ability. In this paper, we present VIPBoard, a…

WS
Weinan Shi et al.Tsinghua University

Lip-Interact: Improving Mobile Device Interaction with Silent Speech Commands

We present Lip-Interact, an interaction technique that allows users to issue commands on their smartphone through silent speech. Lip-Interact repurposes the front camera to capture the user's mouth movements and recognize the issued commands with an end-to-end deep learning model. Our system supports 44 commands for a…

KS
Ke Sun et al.Tsinghua University
Paper TitleAuthorsResearch TopicsPaper DatabaseYear

HiSync: Spatio-Temporally Aligning Hand Motion from Wearable IMU and On-Robot Camera for Command Source Identification in Long-Range HRI

Long-range Human-Robot Interaction (HRI) remains underexplored. Within it, Command Source Identification (CSI) – determining who issued a command – is especially challenging due to multi-user and distance-induced sensor ambiguity. We introduce HiSync, an optical-inertial fusion framework that treats hand motion as bin…

CZ
Chengwen Zhang et al.Tsinghua University

TraceRing: Touchpad-like Pointing with a Single IMU Ring through Personalized Learning

Achieving touchpad-like pointing with a single IMU ring is highly desirable for portable and wearable interaction, yet challenging due to incomplete motion data and significant user variability. We present TraceRing, a finger-worn IMU system that enables precise two-dimensional cursor control. To address the limitatio…

ZH
Zhe He et al.Tsinghua University

3DRing: Enabling Low-Cost 3D Hand Position Tracking by Fusing Inertial and Low-Framerate Optical Sensing

Current mobile hand tracking systems primarily rely on high-framerate (HFR) optical sensors to capture hand positions, resulting in high computational cost and limiting the applicability in end devices. We propose 3DRing, a 3D hand position tracking method that requires only low-framerate (LFR, <10 FPS) optical data a…

ZL
Zhuojun Li et al.Tsinghua University

InterQuest: A Mixed-Initiative Framework for Dynamic User Interest Modeling in Conversational Search

In online information-seeking tasks (e.g., for products and restaurants), users seek information that aligns with their individual preferences to make informed decisions. However, existing systems often struggle to infer users' implicit interests—unstated yet essential preference factors that directly impact decision…

YM
Yu Mei et al.Tsinghua University
AdRecommended

Learn AI Coding at CodeNow

Structured lessons, hands-on projects, and continuous updates for people bringing AI into real development work.

Explore Nowopen_in_new

From Operation to Cognition: Automatic Modeling Cognitive Dependencies from User Demonstrations for GUI Task Automation

Traditional Programming by Demonstration (PBD) systems primarily automate tasks by recording and replaying operations on Graphical User Interfaces (GUIs), without fully considering the cognitive processes behind operations. This limits their ability to generalize tasks with interdependent operations to new contexts (e…

YY
Yiwen Yin et al.Tsinghua University

Investigating Context-Aware Collaborative Text Entry on Smartphones using Large Language Models

Text entry is a fundamental and ubiquitous task, but users often face challenges such as situational impairments or difficulties in sentence formulation. Motivated by this, we explore the potential of large language models (LLMs) to assist with text entry in real-world contexts. We propose a collaborative smartphone-b…

WC
Weihao Chen et al.Tsinghua University

AngleSizer: Enhancing Spatial Scale Perception for the Visually Impaired with an Interactive Smartphone Assistant

Jing 等人开发 AngleSizer 交互式智能手机助手,通过创新方式帮助视障人士感知空间尺度,提升其对物体大小和距离的理解能力。

XJ
Xiaoqing Jing et al.

ContextCam: Bridging Context Awareness with Creative Human-AI Image Co-Creation

The rapid advancement of AI-generated content (AIGC) promises to transform various aspects of human life significantly. This work particularly focuses on the potential of AIGC to revolutionize image creation, such as photography and self-expression. We introduce ContextCam, a novel human-AI image co-creation system th…

XF
Xianzhe Fan et al.Tsinghua University

From Gap to Synergy: Enhancing Contextual Understanding through Human-Machine Collaboration in Personalized Systems

This paper presents LangAware, a collaborative approach for constructing personalized context for context-aware applications. The need for personalization arises due to significant variations in context between individuals based on scenarios, devices, and preferences. However, there is often a notable gap between huma…

WC
Weihao Chen et al.Tsinghua University
emoji_events

VIPBoard: Improving Screen-Reader Keyboard for Visually Impaired People with Character-Level Auto Correction

Modern touchscreen keyboards are all powered by the word-level auto-correction ability to handle input errors. Unfortunately, visually impaired users are deprived of such benefit because a screen-reader keyboard offers only character-level input and provides no correction ability. In this paper, we present VIPBoard, a…

WS
Weinan Shi et al.Tsinghua University

Lip-Interact: Improving Mobile Device Interaction with Silent Speech Commands

We present Lip-Interact, an interaction technique that allows users to issue commands on their smartphone through silent speech. Lip-Interact repurposes the front camera to capture the user's mouth movements and recognize the issued commands with an end-to-end deep learning model. Our system supports 44 commands for a…

KS
Ke Sun et al.Tsinghua University