Windows Failover Cluster engineering

Know your cluster.
Control its risk.

ClusterTriage turns live configuration into defensible findings, priorities and verified improvement. We focus on the cluster stack where small inconsistencies become large operational risks: Windows Server, Hyper-V, storage, networking, hardware and Azure Local.

Measure · Analyse · Prioritise · Remediate · Verify · Repeat

Active incident? Preview the separate ClusterDown incident route →
CLUSTERTRIAGE · ANNUAL ASSURANCE4 MEASUREMENTS
Q1BaselineFull findings report + review
Q2VerifyDelta + residual risk
Q3DriftNew deviations exposed
Q4AssureYear-end state + priorities
Illustrative open risk12 → 3

The yearly cycle shows whether remediation actually changed the measured state.

Microsoft MVPmore than ten yearsFailover ClusteringHyper-V · S2D · Azure LocalRead-only measurementno agent or installationMeasured follow-updelta · residual risk · trend

Not a black box

The measurement path is visible.

The Method page uses the actual ClusterTriage toolchain. It shows what enters the process, which modules measure which parts of the estate, how measured state becomes a finding and which documents come out.

Every magnifier opens the underlying module, catalogue, register or report view.

CLUSTERTRIAGE TOOLCHAINMEASURE → REPORT
Measureread-only collectors at the customer
Gradecatalogue recipes against measured state
Reportfindings, delta, residual risk and trend

Real output

Evidence you can inspect, not a score you have to trust.

Examples from the reporting toolchain use fictitious environments and contain no customer data. Click to enlarge.

About us

Specialist leadership, deliberately narrow focus.

ClusterTriage concentrates on Microsoft failover-cluster infrastructure rather than presenting itself as a broad IT consultancy.

Hans Vredevoort
Server & Cluster Specialist

Hans Vredevoort

40+ years in Microsoft infrastructure, more than a decade as a Microsoft MVP and a long track record in Failover Clustering and Hyper-V.

LinkedIn profile →
Rob Scheepens
Server & Cluster Specialist

Rob Scheepens

Performance, observability and Windows internals, with a long technical background in Microsoft platforms and enterprise infrastructure.

LinkedIn profile →

Active incident

Is the cluster down now?

Assurance is preventive and continuous. An active outage needs a different first move: preserve evidence, identify the highest-priority signals and decide whether a full RCA is required.

CLUSTERDOWN.COMFree incident triage

The reactive entry point stays separate from the enterprise assurance proposition.

Preview ClusterDown →