We’re looking to implement a centralized backup solution using s3 replication to protect large s3 buckets. The solution should be designed with controls in place to reliably protect large volumes of data (backup), support restore capabilities and integrate with existing governance processes. As such the candidate should have the following experience in additional to any previous skills mentioned before:
Job Description
Solid AWS S3 expertise including versioning, storage classes (incl. Glacier), replication, lifecycle, inventory reports and experience with managing large-scale data sets
Experience with managing s3 replication, including conjuration, IAM roles, replication metrics, failure scenarios, and backfill strategies
Experience with S3 batch operations for build reprocessing (existing objects and retrying failed replications) including completion reports
Experience with managing and handling Glacier restore workflows
Experience with designing secure, isolated backup accounts (cross-account) for centralized backup architectures with secure IAM, bucket and KMS policies.
Experience in monitoring and alerting (CloudWatch, EventBridge) for monitoring replication health and failures including SQS and automation to handle and respond to failures
Experience with Athena
Experience with building operational runbacks (automation) for triage and re-drive (re-replication of failures)
Design and build a centralized solution to build repeatable onboarding patterns for current buckets into the centralized solution
J-18808-Ljbffr