Operations strategy
VETA is a simulation, but the discipline around running it is not. This section is the plan of record for how the platform is operated, what we measure, and how we recover when things break. New work should reference these pages before changing anything in the deploy or monitoring path.
Sub-pages
Section titled “Sub-pages” Current state Honest snapshot of where the platform is today: single-host, no SLOs, no alerts, human-detected outages.
SLOs and DORA The targets we commit to: availability, order-ack latency, WS continuity, MTTR, RPO/RTO. DORA progress.
Target architecture Two-tier environments, single-instance topology trade-offs, deploy mechanism, synthetic monitoring, test gate, multi-host roadmap.
Operational discipline Rollback plans, incident log, postmortems, change-freeze windows, future on-call rotation.
Migration path Phased plan from today to target. Deploy gate, descoped Swarm and UAT phases, synthetic monitoring (shipped), discipline tooling, multi-host prod.
Scope and decisions What this strategy is not. Owner, review cadence, and the decision log.