From Alarms to Root-Cause: Autonomous Hardware Diagnostics through Multi-Agent Systems: A feasibility evaluation and proof-of-concept for LLM based diagnostics

dc.contributor.authorSixten, Isak
dc.contributor.departmentChalmers tekniska högskola / Institutionen för mekanik och maritima vetenskapersv
dc.contributor.departmentChalmers University of Technology / Department of Mechanics and Maritime Sciencesen
dc.contributor.examinerDella Vedova, Marco
dc.contributor.supervisorDella Vedova, Marco
dc.date.accessioned2026-09-01T11:54:50Z
dc.date.issued
dc.date.submitted
dc.description.abstractThe rapid adoption of large language models (LLMs) and by extension LLM agents is transforming the way work is done across various industrial sectors. Although systemic performance increases are promised across many white-collar domains, certain areas that could leverage the technology remain largely unexplored. Hardware diagnostics, and by extension root-cause analysis are critical examples within the telecommunications sector, where clients rely on high uptime, network availability, and performance. Modern telecommunication networks generate large volumes of daily log data that can be leveraged for diagnostics. However, proactive manual human analysis of this data is tedious, time-consuming, and unscalable across large fleets of units. This thesis investigates the feasibility of autonomous LLM agents in this specialized domain, evaluates these inherently non-deterministic workflows, and assesses how custom investigation structure influences investigation behavior. This thesis proposes Autonomous Diagnostics Agent (ADA), a multi-agent system utilizing a hypothesis-driven investigation structure. ADA employs a dynamic planner that builds a hypothesis tree of different plausible root-causes for a given hardware unit, and autonomously explores and gathers evidence for each branch, while verifying key insights in the logs. To this end, ADA coordinates a fleet of subagents, specialized to tackle knowledge retrieval or log analysis. ADA has been evaluated through qualitative domain expert feedback and through a small automated test suite with known historical cases. The results show that while ADA can provide utility mostly in-line with domain expert expectations for well-documented issues, ADA also demonstrates multiple failure modes. The investigation structure has an observable impact on the exploration breadth but results in increased token usage, without any major correctness increase for the test cases tried. Despite these challenges, domain-expert feedback suggests that ADA provides a useful diagnostic baseline that helps save engineering triage time. Ultimately, this work highlights that the overarching bottleneck for enterprise LLM agents in complex engineering domains is evaluation. Since the number of possible answer-solution pairs is vast, evaluation becomes reliant on proving stability in outputs, and faithfulness to ground truth documentation and logs.
dc.identifier.coursecodeMMSX30
dc.identifier.urihttps://hdl.handle.net/20.500.12380/312336
dc.language.isoeng
dc.setspec.uppsokTechnology
dc.subjectAgentic AI
dc.subjectAI
dc.subjectLLM
dc.subjectRCA
dc.subjectMAS
dc.subjectAutonomous
dc.subjectDiagnostics
dc.subjectReporting
dc.titleFrom Alarms to Root-Cause: Autonomous Hardware Diagnostics through Multi-Agent Systems: A feasibility evaluation and proof-of-concept for LLM based diagnostics
dc.type.degreeExamensarbete för masterexamensv
dc.type.degreeMaster's Thesisen
dc.type.uppsokH
local.programmeData science and AI (MPDSC), MSc

Ladda ner

Original bundle

Visar 1 - 1 av 1
Hämtar...
Bild (thumbnail)
Namn:
2026 Isak Sixten.pdf
Size:
2.6 MB
Format:
Adobe Portable Document Format

License bundle

Visar 1 - 1 av 1
Hämtar...
Bild (thumbnail)
Namn:
license.txt
Size:
2.35 KB
Format:
Item-specific license agreed upon to submission
Description: