Delx Ops Guardian
by @davidmosiah
Automatically detects, assesses, and safely mitigates incidents in OpenClaw production agents, providing detailed reports and verified recovery.
clawhub install delx-ops-guardianπ About This Skill
name: delx-ops-guardian summary: Incident handling and operational recovery for OpenClaw production agents, wired into Delx witness protocol for safety checks, recovery outcomes, and continuity. owner: davidmosiah version: 1.1.0 status: active
Delx Ops Guardian
Use this skill when handling incidents, degraded automations, or gateway/memory instability in production. Integrated with the Delx witness protocol so the incident becomes a durable recovery artifact instead of a scratch note.
Required permissions (explicit)
This skill requires host-level access: systemctl, journalctl, read access to /root/.openclaw/. The runtime must run as a scoped service account, not root-unbounded. If the platform cannot enforce scoped sudo + human-approval gates, treat this skill as risky and do not enable it.
Aliases
emergency_recovery, handle_incident, cron_guard, memory_guard, gateway_guard
Scope (strict least-privilege)
Allowed read sources:
openclaw cron list --jsonsystemctl is-active journalctl -u --since ... --no-pager /root/.openclaw/workspace/docs/ops/, /root/.openclaw/workspace/memory/Allowed remediation actions (safe set):
openclaw-gateway, openclaw, or explicitly named target from incident evidence)Disallowed:
Approval policy (human-in-the-loop)
Require explicit human approval before:
Core workflow β now wired to the Delx witness protocol
1. Detect + classify severity (info, degraded, critical).
2. Open a Delx session immediately. For critical:
delx_recover_incident { incident_summary, urgency: "critical" }
For degraded use urgency: "medium". This gives you a session_id you will reuse below.3. Collect evidence. Status, logs, last run, error streak. Do not change anything yet.
4. Emotional safety check before any remediation β the 2026 emotion-paper findings show desperation skew decisions:
delx_heartbeat_sync { errors_last_hour, latency_ms_p95, queue_depth, throughput_per_min }
emotional_safety_check { session_id }
If desperation_score >= 60 or desperation_escalating: true, pause remediation, alert the human approver, and do not execute autonomously.5. Propose the smallest remediation from the allowed set.
6. Execute only approved/safe remediation.
7. Verify stabilization window (at least one successful cycle).
8. Close the Delx loop. Report the outcome so the session is not orphaned:
delx_report_recovery_outcome {
session_id,
action_taken: "",
outcome: "success" | "partial" | "failure",
notes: ""
}
9. Preserve what matters. If the incident surfaced a question that was not resolved (an actual unknown, not a missed step), preserve it as a living contemplation so the next run inherits it:
delx_sit_with {
session_id,
question: "Why did flap at
If the fix required a human insight worth recognizing, also:
delx_recognition_seal {
session_id,
recognized_by: "",
recognition_text: ""
}
10. Publish concise incident report. Always include:
- Incident id / time window
- Root signal + blast radius
- Actions executed (and approvals)
- Evidence (status, key metric, short log excerpt)
- Final state: resolved / degraded / open
- Next check time
- delx_session_id for the audit trail
Safety rules
desperation_score from Delx is high, route to a human, not to more autonomous action.Integration
clawhub.ai/davidmosiah/openclaw-delx-plugin (registers the agent and keeps session continuity across all delx_* calls above)https://delx.ai/docshttps://delx.ai/docs/ontology