Soru

Zorluk: OrtaDeveloping Procedures for Business Continuity and Disaster Recovery Validation

A multi-tenant IoT telemetry platform hosts its real-time processing pipeline on Google Cloud with a primary deployment in us-east1 and a secondary disaster recovery (DR) site in us-central1. The business mandates a recovery time objective (RTO) of 2 hours and a recovery point objective (RPO) of 15 minutes. During a recent DR validation drill, attempting to fail over processing nodes to the DR region caused severe service degradation because the secondary region lacked sufficient Compute Engine CPU quota to accommodate the incoming traffic volume. Which procedure should the cloud architect implement to ensure reliable disaster recovery validation?

  1. Automate pre-validation verification checks of regional resource quotas and capacity reservations in the target region prior to initiating failover drills, while conducting regular non-disruptive DR simulations.Cevap
  2. B
    Rely on initiating unannounced live failovers and submitting emergency quota increase requests to Google Cloud Support during the active recovery window whenever capacity limits are reached.
  3. C
    Transition the disaster recovery strategy to a cold standby model that backs up data snapshots to Cloud Storage and delays infrastructure provisioning until a disaster occurs.
  4. D
    Replace existing VPC Network Peering with dedicated high-bandwidth Partner Interconnect connections between regions to bypass compute provisioning limits.

Cevap

Automate pre-validation verification checks of regional resource quotas and capacity reservations in the target region prior to initiating failover drills, while conducting regular non-disruptive DR simulations.
Establishing automated pre-validation procedures that verify regional CPU/instance quotas and active capacity reservations ensures that target DR regions have sufficient capacity before failover traffic is routed, satisfying both RTO and RPO objectives.

Adım Adım Çözüm

1
Analyze the failure root cause in the DR validation drill.
Identified regional compute quota limits in the secondary region as the primary blocker preventing successful failover execution.
Compute Engine default quotas in secondary regions may not match primary region allocations unless requested in advance.
2
Evaluate validation procedures against RTO/RPO objectives.
Automated pre-checks and reserved capacity ensure target infrastructure readiness without risking SLA breaches during failover.
DR validation procedures must ensure resources exist and can scale before shifting live or simulated traffic.

Anahtar Kavram

Disaster Recovery Validation and Quota Management
Bu soruyu puanla