Skip to content

Startup readiness probe waits a full backoff interval after the first failure #452

Description

@JoshuaChi

Problem

A node that starts slightly earlier than its peers probes them before they are
listening, fails once, and then waits the full first backoff interval before
retrying. With the default policy this is ~3s, during which the node reports
"service not ready" although its peers become ready within milliseconds.

Observed

Three nodes started ~14ms apart: the earliest node needed ~3.0s to pass the
cluster ready check; the other two needed ~5ms.

Activity

Sign up for free to join this conversation on GitHub. Already have an account? Sign in to comment

Metadata

Metadata

Assignees

No one assigned

    Labels

    raft-clusterCluster-wide operations and coordination issuesraft-membershipDynamic membership changes (node addition/removal from cluster)

    Type

    Projects

    No projects

      Relationships

      None yet

      Development

      No branches or pull requests

      Issue actions