Evaluating the Impact of a Multi-Agent AI Assistant on Human Creativity in Exploratory Testing

dc.contributor.authorJiachun, Cai
dc.contributor.departmentChalmers tekniska högskola / Institutionen för data och informationstekniksv
dc.contributor.departmentChalmers University of Technology / Department of Computer Science and Engineeringen
dc.contributor.examinerStaron, Miroslaw
dc.contributor.supervisorFrancisco, Gomes
dc.date.accessioned2026-07-03T09:58:07Z
dc.date.issued2026
dc.date.submitted
dc.description.abstractExploratory Testing (ET) is an important approach for uncovering complex software defects, however its success is frequently impacted by human factors such as cognitive bias, mental fatigue, and rigid thinking patterns. This research investigates the impact of AI assistance on the creative performance of human testers. Instead of treating AI as a passive assistant, this study explores how creative stimuli can help testers break through cognitive bottlenecks and explore non-obvious failure scenarios. Using the WICKED multi-agent framework as an experimental platform, we conducted a human-centered user study involving participants with software engineering backgrounds. The research analyzes how real-time AI prompts influence the generation of test scenarios across multiple Software Requirement Specifications (SRS). We assess the interventions along two questions: their effect on the creativity of test artifacts (Quantity, Quality, Novelty), and their usability (Perceived Usefulness and Perceived Ease of Use). We find that WICKED did not reliably increase creativity. Particularly, (i) it did not raise the number of tests created, (ii) its effect on novelty was inconclusive, and (iii) quality was improved only on the more complex system, where it surfaced non-obvious faults such as a concurrency bug missed by manual testers. In terms of usability, participants found WICKED easy to learn and useful in principle, but its slow and unclear interaction limited its perceived usefulness. These results suggest that creativity chatbots are best targeted at complex, uncertain testing tasks and must protect early user trust through speed and relevance, indicating that successful adoption depends as much on the testers themselves as on the strength of the model.
dc.identifier.urihttps://hdl.handle.net/20.500.12380/311827
dc.language.isoeng
dc.setspec.uppsokTechnology
dc.subjectExploratory Testing, Human-AI Collaboration, Creativity Techniques, Large Language Models, User Study.
dc.titleEvaluating the Impact of a Multi-Agent AI Assistant on Human Creativity in Exploratory Testing
dc.type.degreeExamensarbete för masterexamensv
dc.type.degreeMaster's Thesisen
dc.type.uppsokH
local.programmeSoftware engineering and technology (MPSOF), MSc

Ladda ner

Original bundle

Visar 1 - 1 av 1
Hämtar...
Bild (thumbnail)
Namn:
CSE 26-125 JC.pdf
Size:
3.07 MB
Format:
Adobe Portable Document Format

License bundle

Visar 1 - 1 av 1
Hämtar...
Bild (thumbnail)
Namn:
license.txt
Size:
2.35 KB
Format:
Item-specific license agreed upon to submission
Description: