A company wants to create a cloud data repository for ML using Amazon S3. All data (about 40 TB) resides on-premises. They need a solution to transfer and continuously synchronize data between on-prem object storage and S3 that supports encryption, scheduling, monitoring, and data integrity checks. Which solution satisfies these requirements?
Choose an answer
Tap an option to check your answer.
Correct answer: Use AWS DataSync to perform an initial full copy and schedule incremental transfers of changed data until cutover..
Why this is the answer
AWS DataSync is the most suitable solution because it is designed for online data transfer between on-premises storage and S3, supporting encryption, scheduling, monitoring, and data integrity checks. It efficiently handles large datasets (40 TB) by performing an initial full copy and then incremental transfers, which is crucial for continuous synchronization. The AWS CLI s3 sync command is not designed for continuous synchronization from on-premises to S3 and lacks the robust features like scheduling, monitoring, and integrity checks needed for enterprise-grade data migration. AWS Transfer for FTPS is for transferring files using FTP, FTPS, or SFTP protocols, primarily for external users, not for continuous, managed synchronization of large datasets from on-premises storage to S3. S3 Batch Operations is for performing large-scale batch operations on S3 objects already in S3, not for pulling data from on-premises storage.
Pass your exam — without the endless answer hunt
Get every verified question and explanation for this exam in one place, and save hours of prep. 1,000+ certifications · 20+ languages · free to start.
Pass your exam faster → No card needed