Published on 24 August 2026, updated 26 August 2026
ISBN-10: 0738462748
ISBN-13: 9780738462745
IBM Form #: SG24-8446-01
Authors: Bill White, Diego Bessone, Abi Bettle-Shaffer, Nino de Carvalho, Marc Coq, Kris Dzialek, Mike Gonzales, Kevin McKenzie, David Raften, Karen Smolar and Meral Temel
IBM Z® platforms underpin mission-critical enterprise workloads where disruptions directly affect availability, data integrity, and business outcomes. Business continuity (BC) is achieved through designed resiliency and recovery that uses platform capabilities to contain failures, preserve service, and help ensure predictable outcomes.
Resiliency sustains service during disruption by isolating faults, eliminating single points of failure (SPOFs), and enabling graceful degradation. Recovery restores service after disruption through deterministic, automated processes that are aligned to defined recovery time objectives (RTOs) and recovery point objectives (RPOs). Together, they deliver consistent availability, reduced downtime, and rapid restoration across IBM Z environments.
Disruption is inevitable. Systems fail, workloads evolve, infrastructure requires maintenance, and threats continue to emerge. The real measure of success is not avoiding disruption, but helping ensure controlled behavior under stress and fast, reliable recovery.
This IBM Redbooks® publication examines resiliency on the IBM Z platform, focusing on system behavior under failure and the architectural choices that drive availability and recovery. It introduces deployment patterns ranging from restart-based recovery to multi-site continuity, providing a practical framework for aligning business requirements with design.
The publication highlights enabling capabilities across the IBM Z stack, including infrastructure, IBM z/OS®, middleware, applications, and operations, with emphasis on automation, observability, and AI-assisted predictive operations.
BC depends on predictable behavior under stress and reliable restoration after stress. This publication helps you design and operate IBM Z environments to achieve that outcome.
This publication is intended for IT leaders, enterprise architects, infrastructure architects, system programmers, system operations professionals, BC planners, DR specialists, resiliency architects, site reliability engineers (SREs), and technical decision-makers who are responsible for designing, operating, and governing HA, recoverable, and resilient IBM Z environments.
Chapter 1. Business continuity: A foundation for an intelligent digital business
Chapter 2. IBM Z deployment patterns: Behavior under disruption
Chapter 3. Infrastructure layer: Building resiliency
Chapter 4. Operating system layer: Enabling resiliency
Chapter 5. Middleware layer: Coordinating resiliency
Chapter 6. Application layer: Engineering resiliency
Chapter 7. Management layer: Orchestrating resiliency
Chapter 8. Bringing business continuity into practice
Appendix A. Data center resiliency considerations
Appendix B. Sustaining acceptable performance
Appendix C. Best practices: Reducing risk