Skip to content
BytePatterns

SOA-C03 · Domain 2: Reliability and Business Continuity · 22% of the exam

Task 2.3: Implement backup and restore strategies.

Getting data back: backup plans and snapshots on a schedule, point-in-time restore matched to RPO, RTO and cost, versioning in S3 and FSx, and the disaster recovery strategies from backup and restore to active/active.

Study it

  • Backups: AWS Backup plans, snapshots and cross-Region copies

    Lesson coming

  • Restoring databases: point-in-time restore against RPO, RTO and cost

    Lesson coming

  • Versioning in S3 and FSx

    Lesson coming

  • Disaster recovery: backup and restore, pilot light, warm standby, active/active

    Partly covered by: Disaster Recovery: RPO & RTO

Sample questions

Try each one before opening the answer. Every option is explained, with the AWS documentation page that proves it.

Question 1 · choose 1

A company tags its Amazon EC2 instances, Amazon EBS volumes, Amazon RDS DB instances and Amazon DynamoDB tables with backup=daily. Every tagged resource must be backed up once a day, each backup kept for 35 days, and a copy of each backup stored in a second AWS Region. The operations team wants one central policy and no custom code. What should the team set up?

  1. AAmazon Data Lifecycle Manager policies that target resources with the backup=daily tag
  2. BAn AWS Backup plan with a cross-Region copy rule and resources selected by tag
  3. CRDS automated backups and DynamoDB point-in-time recovery, each set to 35 days
  4. DA scheduled AWS Lambda function that calls each service's backup API and copies the results
Show the answer and why
  • AAmazon Data Lifecycle Manager policies that target resources with the backup=daily tag

    Incorrect

    Data Lifecycle Manager automates EBS snapshots and EBS-backed AMIs. It does not back up RDS DB instances or DynamoDB tables.

  • BAn AWS Backup plan with a cross-Region copy rule and resources selected by tag

    Correct

    A backup plan schedules and retains backups across services, can copy them to another Region automatically, and can select resources by tag.

  • CRDS automated backups and DynamoDB point-in-time recovery, each set to 35 days

    Incorrect

    These are separate, service-by-service settings. They leave out EC2 and EBS and do not give one central policy.

  • DA scheduled AWS Lambda function that calls each service's backup API and copies the results

    Incorrect

    Custom code is exactly what the team wants to avoid. AWS Backup exists to replace per-service scripts with central policies.

One schedule, one retention, one cross-Region copy rule, selected by tag across several services: that is an AWS Backup plan.

Question 2 · choose 1

At 14:05 an engineer drops an important table in a production Amazon RDS for PostgreSQL Multi-AZ DB instance that also has one read replica. Automated backups are kept for 7 days, and the last automated snapshot was taken at 03:00. The team needs the data as it was at 14:04, with as little loss as possible. What should the team do?

  1. ARestore the 03:00 automated snapshot onto the existing DB instance
  2. BPromote the read replica to a standalone DB instance
  3. CUse point-in-time restore to 14:04 and repoint the application
  4. DReboot the DB instance with failover to switch to the standby
Show the answer and why
  • ARestore the 03:00 automated snapshot onto the existing DB instance

    Incorrect

    A snapshot cannot be restored onto an existing DB instance; it creates a new one. Restoring it would also lose everything written since 03:00.

  • BPromote the read replica to a standalone DB instance

    Incorrect

    RDS replays every change from the primary on the read replica, so the dropped table is gone there as well.

  • CUse point-in-time restore to 14:04 and repoint the application

    Correct

    Point-in-time restore creates a new DB instance at any time within the backup retention period, using backups and transaction logs.

  • DReboot the DB instance with failover to switch to the standby

    Incorrect

    The standby is a synchronous copy of the primary, so it holds the same dropped table.

Replicas and standbys copy mistakes as faithfully as data. Recovering from a logical error means going back in time, and point-in-time restore gets closest to the moment before it.

Question 3 · choose 1

Users of a shared Amazon S3 bucket sometimes overwrite or delete objects by mistake. The operations team must be able to bring back any earlier version of an object for 30 days after it was replaced or deleted. After that, the old versions must be removed automatically to control cost. What should the team configure?

  1. AS3 Object Lock in compliance mode with a 30-day retention period
  2. BAn S3 Lifecycle rule that moves objects to S3 Glacier Flexible Retrieval after 30 days
  3. CS3 server access logging to a separate bucket, so that every deletion is recorded
  4. DS3 Versioning, plus a Lifecycle rule that expires noncurrent versions after 30 days
Show the answer and why
  • AS3 Object Lock in compliance mode with a 30-day retention period

    Incorrect

    Compliance mode stops every user, including the root user, from deleting a protected version until retention ends. That is immutability, not recovery with automatic cleanup.

  • BAn S3 Lifecycle rule that moves objects to S3 Glacier Flexible Retrieval after 30 days

    Incorrect

    A transition changes the storage class of objects. Without versioning, an overwritten or deleted object still cannot be brought back.

  • CS3 server access logging to a separate bucket, so that every deletion is recorded

    Incorrect

    Access logs record the requests made to a bucket. They show who deleted an object but cannot restore it.

  • DS3 Versioning, plus a Lifecycle rule that expires noncurrent versions after 30 days

    Correct

    Versioning keeps every version and turns a delete into a delete marker, so earlier versions can be restored; the Lifecycle rule permanently deletes noncurrent versions after the chosen time.

Versioning gives the undo button, and a noncurrent-version expiration rule keeps the history from growing forever.

Question 4 · choose 1

A company's disaster recovery runbook says that if the primary AWS Region fails, the recovery Region must begin handling production requests immediately, at reduced capacity, and then scale up. Data must be replicated continuously. Within these requirements the company wants the lowest ongoing cost. Which disaster recovery strategy should the operations team implement?

  1. AWarm standby, with a scaled-down but working copy running in the recovery Region
  2. BPilot light, with data replicated and the application servers switched off until a failover
  3. CBackup and restore, with backups copied to the recovery Region on a schedule
  4. DMulti-site active/active, with the full workload serving users in both Regions
Show the answer and why
  • AWarm standby, with a scaled-down but working copy running in the recovery Region

    Correct

    Warm standby keeps everything deployed and running at reduced size, so it can handle traffic immediately and only needs to scale up.

  • BPilot light, with data replicated and the application servers switched off until a failover

    Incorrect

    A pilot light cannot process requests until servers are turned on and possibly more infrastructure is deployed, so it cannot serve immediately.

  • CBackup and restore, with backups copied to the recovery Region on a schedule

    Incorrect

    With backup and restore, the infrastructure must be deployed and the data restored from backups before any request can be served.

  • DMulti-site active/active, with the full workload serving users in both Regions

    Incorrect

    Active/active is the most complex and costly strategy. It goes beyond what the runbook asks for.

"Serve immediately, at reduced capacity, then scale up" is the definition AWS gives for warm standby, and it costs less than running both Regions at full size.

Question 5 · choose 1

A department keeps its documents on an Amazon FSx for Windows File Server file system. Users often need the version of a file from a few hours earlier after a bad edit. Today each request goes to an administrator, who restores a backup to a new file system to copy one file out. The users want to restore earlier versions of their own files from Windows File Explorer. What should the operations team configure?

  1. AShadow copies on the file system, taken on a regular schedule
  2. BA longer retention period for the file system's automatic daily backups
  3. CAn AWS Backup plan that backs up the file system every hour
  4. DA Multi-AZ deployment of the file system with a standby file server
Show the answer and why
  • AShadow copies on the file system, taken on a regular schedule

    Correct

    With shadow copies enabled, users can restore previous versions of files and folders themselves with Restore previous versions in File Explorer.

  • BA longer retention period for the file system's automatic daily backups

    Incorrect

    Each FSx backup is restored by creating a new file system, so an administrator would still do the work, and only once a day's versions would exist.

  • CAn AWS Backup plan that backs up the file system every hour

    Incorrect

    Restoring through AWS Backup also creates a new file system from the recovery point, which users cannot do from File Explorer.

  • DA Multi-AZ deployment of the file system with a standby file server

    Incorrect

    Multi-AZ replicates data synchronously and fails over to the standby for availability. A bad edit is replicated too.

Backups protect the whole file system and are restored by an administrator; shadow copies keep earlier versions inside the file system for users to restore themselves.

Practise domain 2 →Practise all domains →