AMIR: 联网环境中基于视频和网络流量的主动多模态交互识别

人体姿态与行为识别环境感知与上下文计算普适计算(Ubiquitous Computing)老年护理人员家庭照顾者(Caregiver)

Activity recognition using video data is widely adopted for elder care, monitoring for safety and security, and home automation. Unfortunately, using video data as the basis for activity recognition can be brittle, since models trained on video are often not robust to certain environmental changes, such as camera angle and lighting changes. There has been a proliferation of network-connected devices in home environments. Interactions with these smart devices are associated with network activity, making network data a potential source for recognizing these device interactions. This paper advocates for the synthesis of video and network data for robust interaction recognition in connected environments. We consider machine learning-based approaches for activity recognition, where each labeled activity is associated with both a video capture and an accompanying network traffic trace. We develop a simple but effective framework AMIR (Active Multimodal Interaction Recognition)1 that trains independent models for video and network activity recognition respectively, and subsequently combines the predictions from these models using a meta-learning framework. Whether in lab or at home, this approach reduces the amount of "paired" demonstrations needed to perform accurate activity recognition, where both network and video data are collected simultaneously. Specifically, the method we have developed requires up to 70.83% fewer samples to achieve 85% F1 score than random data collection, and improves accuracy by 17.76% given the same number of samples. https://dl.acm.org/doi/10.1145/3580818

快捷操作

分享

分享当前页面

ios_share

https://hci.top/zh/papers/ubicomp/128314/2023

广告推荐

学习 AI 编程到 CodeNow

一眼看懂

论文快照

fact_check
dataset
来源
UbiComp
calendar_month
年份
2023
emoji_events
奖项
未标记奖项
group
作者
7 位作者
sell
研究子方向
人体姿态与行为识别、环境感知与上下文计算、普适计算(Ubiquitous Computing)
work
职业/产业
老年护理人员、家庭照顾者(Caregiver)
article
内容状态
仅摘要
hub
相关论文
3 篇相关论文