Tyche: Making Sense of Property-Based Testing Effectiveness
Authors
Software developers increasingly rely on automated methods to assess the correctness of their code. One such method is property-based testing (PBT), wherein a test harness generates hundreds or thousands of inputs and checks the outputs of the program on those inputs using parametric properties. Though powerful, PBT induces a sizable gulf of evaluation: developers need to put in nontrivial effort to understand how well the different test inputs exercise the software under test. To bridge this gulf, we propose Tyche, a user interface that supports sensemaking around the effectiveness of property-based tests. Guided by a formative design exploration, our design of Tyche supports developers with interactive, configurable views of test behavior with tight integrations into modern developer testing workflow. These views help developers explore global testing behavior and individual test inputs alike. To accelerate the development of powerful, interactive PBT tools, we define a standard for PBT test reporting and integrate it with a widely used PBT library. A self-guided online usability study revealed that Tyche's visualizations help developers to more accurately assess software testing effectiveness.
Research Questions / Practical Problems
Question signals indexed for this paper.
- 71%
PleaSQLarify: Visual Pragmatic Repair for Natural Language Database Querying
CHI '26· Human-LLM Collaboration +2
- 71%
XAutoML: A Visual Analytics Tool for Understanding and Validating Automated Machine Learning
IUI '24· Explainable AI (XAI) +2
- 71%
SCSimulator: An Exploratory Visual Analytics Framework for Partner Selection in Supply Chains through LLM-driven Multi-Agent Simulation
IUI '26· Human-LLM Collaboration +2
- 71%
Improving Steering and Verification in AI-Assisted Data Analysis with Interactive Task Decomposition
UIST '24· Human-LLM Collaboration +2
- 67%
Bringing AI to BI: Enabling Visual Analytics of Unstructured Data in a Modern Business Intelligence Platform
CHI '18· Explainable AI (XAI) +1
- 67%
FDHelper: Assist Unsupervised Fraud Detection Experts with Interactive Feature Selection and Evaluation
CHI '20· Explainable AI (XAI) +1
- 67%
mTSeer: Interactive Visual Exploration of Models on Multivariate Time-series Forecast
CHI '21· Interactive Data Visualization +1
- 63%
Natural Language Dataset Generation Framework for Visualizations Powered by Large Language Models
CHI '24· Human-LLM Collaboration +2
- 63%
VeriPlan: Integrating Formal Verification and LLMs into End-User Planning
CHI '25· Human-LLM Collaboration +2
- 63%
Lexara: A User-Centered Toolkit for Evaluating Large Language Models for Conversational Visual Analytics
CHI '26· Human-LLM Collaboration +2
Based on Jaccard similarity of research subtopics & professions (≥60%)