Summer Sale Special Limited Time 70% Discount Offer - Ends in 0d 00h 00m 00s - Coupon code: xmasmnth

A company stores raw clickstream data in an Amazon S3 bucket.

A company stores raw clickstream data in an Amazon S3 bucket. The company needs a solution to process the data every day by using complex PySpark transformations that rely on custom internal libraries. After the data is transformed, the company must store the data in Amazon Redshift for analytics. The solution must be highly scalable to handle large data workloads.

Which solution will meet these requirements with the LEAST operational overhead?

A.

Use AWS Glue Studio to build and schedule PySpark jobs. Configure an AWS Glue data connection that includes the custom libraries.

B.

Use Amazon EC2 Auto Scaling groups with a custom AMI that contains the custom libraries to run a PySpark application.

C.

Use Amazon EMR to run PySpark jobs. Use bootstrap actions to install the custom libraries.

D.

Use Amazon SageMaker Processing jobs to run PySpark code that uses native SageMaker libraries.

Amazon Web Services Data-Engineer-Associate Summary

  • Vendor: Amazon Web Services
  • Product: Data-Engineer-Associate
  • Update on: Jul 26, 2026
  • Questions: 302
Price: $52.5  $149.99
Buy Now Data-Engineer-Associate PDF + Testing Engine Pack

Payments We Accept

Your purchase with ExamsVCE is safe and fast. Your products will be available for immediate download after your payment has been received.
The ExamsVCE website is protected by 256-bit SSL from McAfee, the leader in online security.

examsvce payment method