Q17Design Resilient Architectures
A company stores millions of objects across several prefixes in an Amazon S3 bucket using the Amazon S3 Glacier Deep Archive storage class. The company must delete all data older than 3 years, except for a subset of data that must be retained. The company has identified the data that must be retained and wants to implement a serverless solution. Which solution meets these requirements?
← → navigate · a answer
Community votes
Discussion · 5
D 3
Takeaways:
1. S3 Inventory provides a flat file report of the objects in an S3 bucket, including metadata such as storage class, size, and last modified date. Now you can easily find files "older than 3 years".
2. The Lambda function can be used to filter objects based on the company's requirements (excluding the subset of data to retain) and to delete the objects programmatically.
3. S3 Batch Operations allow the execution of bulk operations (e.g., delete operations) on a large number of S3 objects.
A - Running a script on an EC2 instance is not serverless.
B - AWS Batch is designed for batch computing workloads, not operations like deleting S3 objects.
C - AWS Glue crawlers are designed to extract metadata and schema from data stored in S3 for querying and ETL processes, not for managing or deleting objects.
D 3
Enable S3 Inventory. Create an AWS Lambda function to filter and delete objects. Invoke the Lambda function with S3 Batch Operations to delete objects by using the inventory reports.
2
yes D is the right answer
D 2
https://docs.aws.amazon.com/AmazonS3/latest/userguide/batch-ops.html
2
D is correct