Genie in the Model: Automatic Generation of Human-in-the-Loop Deep Neural Networks for Mobile Applications
Authors
"Advances in deep neural networks (DNNs) have fostered a wide spectrum of intelligent mobile applications ranging from voice assistants on smartphones to augmented reality with smart-glasses. To deliver high-quality services, these DNNs should operate on resource-constrained mobile platforms and yield consistent performance in open environments. However, DNNs are notoriously resource-intensive, and often suffer from performance degradation in real-world deployments. Existing research strives to optimize the resource-performance trade-off of DNNs by compressing the model without notably compromising its inference accuracy. Accordingly, the accuracy of these compressed DNNs is bounded by the original ones, leading to more severe accuracy drop in challenging yet common scenarios such as low-resolution, small-size, and motion-blur. In this paper, we propose to push forward the frontiers of the DNN performance-resource trade-off by introducing human intelligence as a new design dimension. To this end, we explore human-in-the-loop DNNs (H-DNNs) and their automatic performance-resource optimization. We present H-Gen, an automatic H-DNN compression framework that incorporates human participation as a new hyperparameter for accurate and efficient DNN generation. It involves novel hyperparameter formulation, metric calculation, and search strategy in the context of automatic H-DNN generation. We also propose human participation mechanisms for three common DNN architectures to showcase the feasibility of H-Gen. Extensive experiments on twelve categories of challenging samples with three common DNN structures demonstrate the superiority of H-Gen in terms of the overall trade-off between performance (accuracy, latency), and resource (storage, energy, human labour). https://dl.acm.org/doi/10.1145/3580815"
Research Questions / Practical Problems
Question signals indexed for this paper.
Research Questions
3- How can deep neural networks (DNNs) be effectively deployed under mobile device resource constraints?Category: Model Steering, Latent Space Editing, and Knowledge InjectionSimilar questionsarrow_forward
- How can human involvement be systematically incorporated as an optimization hyperparameter in model design to improve DNN performance?Category: Model Steering, Latent Space Editing, and Knowledge InjectionSimilar questionsarrow_forward
- How does the H-Gen framework maximize the effectiveness of human involvement while balancing model accuracy, latency, and resource consumption?Category: Model Steering, Latent Space Editing, and Knowledge InjectionSimilar questionsarrow_forward
Practical Problems
1- Intelligent applications on mobile devices struggle to guarantee performance under low-resource conditions.Category: Model Steering, Latent Space Editing, and Knowledge InjectionSimilar questionsarrow_forward
- 83%
Towards Human-Guided Machine Learning
IUI '19· Human-LLM Collaboration +2
- 83%
Never-ending Learning of User Interfaces
UIST '23· Human-LLM Collaboration +2
- 80%
Comparing Sentence-Level Suggestions to Message-Level Suggestions in AI-Mediated Communication
CHI '23· Human-LLM Collaboration +1
- 80%
Model Compression in Practice: Lessons Learned from Practitioners Creating On-device Machine Learning Experiences
CHI '24· Human-LLM Collaboration +1
- 80%
VAL: Interactive Task Learning with GPT Dialog Parsing
CHI '24· Human-LLM Collaboration +1
- 80%
Need Help? Designing Proactive AI Assistants for Programming
CHI '25· Human-LLM Collaboration +1
- 80%
InstructPipe: Generating Visual Blocks Pipelines with Human Instructions and LLMs
CHI '25· Human-LLM Collaboration +1
- 80%
Interactive Hyperparameter Optimization with Paintable Timelines
DIS '21· Human-LLM Collaboration +1
- 80%
Text-to-SQL Domain Adaptation via Human-LLM Collaborative Data Annotation
IUI '25· Human-LLM Collaboration +1
- 71%
When Help Hurts: Verification Load and Fatigue with AI Coding Assistants
CHI '26· Human-LLM Collaboration +3
Based on Jaccard similarity of research subtopics & professions (≥60%)