Air-Gapped AI SRE Agent: Run AURA on a Local LLM with llama.cpp
The model, the serving layer, and the agent asking questions all sit on local hardware. The first thing it debugs is the AI stack it is running on. No external inference provider in the path.