- Rewrite AGENTS.md: DB as source of truth, MCP knowledge tools, archive refs - Fix OIKOS.md: seeds/ paths, remove Python-era notes, update deployment status - Fix commands.md, agent-enrollment.md: archive/knowledge/ links - Fix all SKILL.md files: remove hosts/*.yaml refs, point to inventory.yaml - Fix HERMES.md, schema.md, page-templates.md, llm-wiki.md: update paths - Fix bootstrap.sh: identity check reads inventory.yaml - Fix README.md, cutover-checklist.md: stale wiki references - Move convert-wiki.py to archive/ (one-shot done)
1.3 KiB
1.3 KiB
name, risk_class, inputs, verification, docs_update_checklist
| name | risk_class | inputs | verification | docs_update_checklist | |
|---|---|---|---|---|---|
| service-health-check | read_only |
|
homelab service <name> health |
Service health check
Goal: determine whether a service is actually healthy, without ad-hoc SSH.
homelab service <name> explain— read the context card: backend, blast radius, doc pointer, risk notes.homelab service <name> health— live health probe (HTTP code against the service'surl/endpoint). Once the Week-3 scheduler ships, this reads a cached snapshot by default; pass--liveto force a fresh probe.- If unhealthy,
homelab service <name> log(or MCPtail_log) for the last 200 lines. - Cross-check blast radius:
homelab node <name> relations— is this entity's own backend host healthy? A downstream failure (e.g.strongdown) will show up here before the service's own logs explain anything. - If the fix is a restart: classify first (
seeds/policy.yaml—service-restartisreversible_lowunless the service has aservice_overridesentry, e.g.caddy/dnsareconfig_mutation). Unattended agents may act onreversible_lowwithout approval.
Docs-update checklist: none for a pure health check. If the investigation
reveals stale risk_notes or a wrong doc_page, fix inventory.yaml in
the same session.