Skip to content

Operator runbooks: document production network patterns #258

Description

@Reccetech

Summary

No documentation exists for operating a real network: node failure recovery, rolling upgrades, state backup/restore, or scaling. An operator running production infrastructure has no runbook. The state-save-and-restore example exists in the Solo repo but is not linked from any operator-facing docs today.

Proposed Scope

A new "Operating Your Network" page or section covering at minimum:

  • Handling a node failure
  • Rolling upgrade procedure
  • State backup and restore (linking the existing state-save-and-restore example)

Blocked

Requires input from someone who has operated a real network. Content cannot be written from spec alone. If you have production Solo experience, please comment.

Metadata

Metadata

Assignees

No one assigned

    Labels

    BlockedDocs issue is blocked by other needed work.ImprovementDocumentation Improvement RequestP2

    Type

    No type

    Projects

    Status
    In Progress

    Milestone

    Relationships

    None yet

    Development

    No branches or pull requests

    Issue actions