DEA-C01 Data Ingestion and Transformation • Set 14
DEA-C01 Data Ingestion and Transformation Practice Test 14 — 15 questions with explanations. Free, no signup.
A company uses Amazon EMR to process large datasets stored in Amazon S3. The data is in Parquet format and partitioned by date. The EMR cluster uses Spark SQL for transformations. Recently, the job has been slow and some tasks are failing due to 'java.lang.OutOfMemoryError'. The cluster has 10 core nodes of type m5.xlarge. Which configuration change would MOST improve performance and stability?
Choose an answer to begin — your selection is scored in the full session.
15 questions · instant feedback and full explanations after every question.