helpResearch questionWearable and Smart Glasses Gesture Input
How can vision-language models (VLMs) enhance object recognition and environmental interaction capabilities of wearable devices?UIST '24WatchThis: A Wearable Point-and-Ask Interface powered by Vision-Language Models for Contextual Queries