Athena: Intermediate Representations for Iterative Scaffolded App Generation with an LLM
It is challenging to generate the code for a complete user interface using a Large Language Model (LLM). User interfaces are complex and their implementations often consist of multiple, inter-related files that together specify the contents of each screen, the navigation flows between the screens, and the data model used throughout the application. It is challenging to craft a single prompt for an LLM that contains enough detail to generate a complete user interface, and even then the result is frequently a single large and intricate file that contains all of the generated screens. In this paper, we introduce Athena, a prototype application generation environment that demonstrates how the use of shared intermediate representations, including an app storyboard, data model, and GUI skeletons, can help a developer work with an LLM in an iterative fashion to craft a complete user interface. These intermediate representations also scaffold the LLM’s code generation process, producing organized and structured code in multiple files while limiting errors. We evaluated Athena with a user study with 12 developers. Participants appreciated Athena’s support for prototyping multi-screen iOS apps, acknowledged that the intermediate representations improved their control and understanding of generated code, and discussed the limitations of the system and potential directions for improvement.
Research Questions / Practical Problems
Question signals indexed for this paper.
- 100%
The Way We Notice, That’s What Really Matters: Instantiating UI Components with Distinguishing Variations
CHI '26· Human-LLM Collaboration +2
- 83%
Rapsai: Accelerating Machine Learning Prototyping of Multimedia Applications through Visual Programming
CHI '23· Human-LLM Collaboration +1
- 83%
CoLadder: Manipulating Code Generation via Multi-Level Blocks
UIST '24· Human-LLM Collaboration +1
- 71%
Beyond Code Generation: LLM-supported Exploration of the Program Design Space
CHI '25· Generative AI (Text, Image, Music, Video) +2
- 71%
GenieWizard: Multimodal App Feature Discovery with Large Language Models
CHI '25· Full-Body Interaction & Embodied Input +2
- 71%
Prototyping with Prompts: Emerging Approaches and Challenges in Generative AI Design for Collaborative Software Teams
CHI '25· Generative AI (Text, Image, Music, Video) +2
- 71%
Assistance or Disruption? Exploring and Evaluating the Design and Trade-offs of Proactive AI Programming Support
CHI '25· Human-LLM Collaboration +2
- 71%
PointAloud: An Interaction Suite for AI-Supported Pointer-Centric Think-Aloud Computing
CHI '26· Human-LLM Collaboration +2
- 71%
Linting Style and Substance in READMEs
CHI '26· Computational Methods in HCI +2
- 71%
Where Will They Click Next? A Social Foraging Model for Collaborating Teams
CHI '26· Distributed Team Collaboration +2
Based on Jaccard similarity of research subtopics & professions (≥60%)