ETExamTower
Q7Data Ingestion and Transformation

A company created an extract, transform, and load (ETL) data pipeline in AWS Glue. A data engineer must crawl a table that is in Microsoft SQL Server. The data engineer needs to extract, transform, and load the output of the crawl to an Amazon S3 bucket. The data engineer also must orchestrate the data pipeline. Which AWS service or feature will meet these requirements MOST cost-effectively?

← → navigate · a answer
Community votes
B
100% (7)
A
0% (0)
C
0% (0)
D
0% (0)
Discussion · 7
B 12
Selected Answer: B I asked an AI. Analysis of the answers: A. AWS Step Functions: It is a solid choice for orchestrating workflows with steps across different AWS services, but it would require extra development to connect to Microsoft SQL Server. B. AWS Glue Workflows: This is the best and most cost-effective option. AWS Glue is built specifically for ETL on AWS and integrates directly with data sources such as Microsoft SQL Server through connectors. This makes configuration easier and avoids the need for additional development. C. AWS Glue Studio: It is a visual interface for AWS Glue that makes it simple to create and manage ETL jobs. However, the underlying functionality comes from AWS Glue (B) workflows. D. Amazon Managed Workflows for Apache Airflow (Amazon MWAA): It's a workable option, but it's generally more expensive than native AWS services like AWS Glue Workflows. Additionally, it requires some Airflow experience for setup and maintenance.
B 9
Selected Answer: B Glue workflows are the simplest solution here: https://aws.amazon.com/blogs/big-data/orchestrate-an-etl-pipeline-using-aws-glue-workflows-triggers-and-crawlers-with-custom-classifiers/ https://aws.amazon.com/blogs/big-data/extracting-multidimensional-data-from-microsoft-sql-server-analysis-services-using-aws-glue/
B 3
Selected Answer: B Agree with B. CRAWLING and ETL are the main functions of a Glue workflow and MS SQL is supported: https://docs.aws.amazon.com/glue/latest/dg/crawler-data-stores.html
1
Is B !
B 1
Selected Answer: B Glue is the easiest thing to pick here.
B 1
Selected Answer: B https://community.aws/content/2iBQiAGS4RvEolgSQKu4iF8InTV/choose-the-right-data-orchestration-service-for-your-data-pipeline?lang=en
B 1
Selected Answer: B AWS Glue Workflows are designed specifically to orchestrate Glue crawlers, Glue jobs, and triggers in one pipeline. Since the pipeline already uses Glue (crawler + ETL job + S3 output), Glue Workflows fit natively and do not need extra services or compute.