Overview
ISO/IEC TS 22237-31:2026 provides a comprehensive framework for measuring and analyzing the resilience of data centre infrastructures. Published by ISO and IEC, this technical specification defines key performance indicators (KPIs) related to resilience, dependability, fault tolerance, and availability tolerance in data centre facilities. Its focus lies on the data centre infrastructure (DCI) elements of power distribution, power supply, and environmental control, but it can also be referenced for other infrastructures such as telecommunications cabling.
By offering standardized methods for calculating and comparing resilience KPIs, ISO/IEC TS 22237-31:2026 equips data centre planners, designers, and operators to better manage risks, set clear service level agreements (SLAs), and optimize infrastructure for maintainability, recoverability, and vulnerability.
Key Topics
-
Resilience Metrics & KPIs
The standard introduces quantifiable metrics as KPIs, allowing for objective assessment of data centre resilience. These KPIs cover:
- Dependability (reliability, availability, failure rate)
- Fault tolerance (identifying single and double points of failure)
- Availability tolerance (capacity to perform with certain failures present)
-
Data Centre Infrastructure Focus
ISO/IEC TS 22237-31:2026 addresses the core non-IT infrastructure:
- Power supply and distribution
- Environmental control (cooling)
- Extension to other infrastructure (e.g., telecommunications cabling)
-
Measurement and Calculation
The standard defines methods to measure and calculate resilience levels (RLs) and associated KPIs, enabling analytical comparison between different DCIs.
-
Life Cycle Application
Guidance is provided on incorporating resilience KPIs throughout a data centre’s life cycle:
- Design strategy and objective setting
- System specification and functional design
- Construction and operational phases
-
Practical Tools and Examples
The specification includes:
- Calculation examples
- Methods such as reliability block diagrams (RBD) and failure mode effects and criticality analysis (FMECA)
- Documentation requirements for resilience levels, dependability, fault tolerance, and availability tolerance
Applications
Implementing ISO/IEC TS 22237-31:2026 brings substantial benefits for data centre stakeholders:
-
Design and Engineering
Enables infrastructure designers to assess and optimize power and environmental systems for high resilience, reducing risk and ensuring SLA compliance.
-
Operational Management
Empowers data centre operators to monitor, document, and improve facility resilience based on standardized KPIs, addressing vulnerabilities and efficiently planning maintenance.
-
Comparative Analysis
Facilitates benchmarking of different data centre designs or sites, supporting investment decisions and continuous improvement initiatives.
-
Risk Management
Standardizes the evaluation of fault tolerance and availability tolerance, enhancing preparedness for disruptive events and supporting disaster recovery planning.
-
SLAs & Performance Reporting
Provides an objective basis for defining, tracking, and reporting SLA-related performance metrics to clients and auditors.
Related Standards
To maximize the practical value of ISO/IEC TS 22237-31:2026, consider its alignment with related standards:
- ISO/IEC 22237-1: General concepts for data centre facilities and infrastructure
- ISO/IEC 22237-3: Power distribution
- ISO/IEC 22237-4: Environmental control
- ISO/IEC 30134 series: Key performance indicators for data centres, including efficiency and sustainability
- ISO/IEC TS 22237-5: Telecommunications cabling infrastructure (for extended coverage)
By integrating ISO/IEC TS 22237-31:2026 with these standards, organizations can ensure a holistic, consistent, and future-ready approach to data centre resilience management.