DP-203 Practice Question: Secure, monitor, and optimize data storage and data processing
You are monitoring an Azure Data Factory pipeline that copies data from an on-premises SQL Server to Azure Synapse Analytics using a self-hosted integration runtime. You notice that the pipeline runs are taking longer than expected, and you suspect performance bottlenecks. You need to identify the cause and optimize the copy performance. What should you do first?
⚠ Common exam trap
The trap here is jumping to performance tuning actions like increasing parallelism or scaling, but the correct first step is always to gather diagnostic data to identify the actual bottleneck.
Answer choices
Why each option matters
Answer the question above first, then reveal the full breakdown to understand why each option is right or wrong.
Correct answer & explanation
✓
Enable logging in Azure Data Factory and analyze the copy activity execution details in Azure Monitor.
To optimize copy performance, you must first diagnose the bottleneck. Azure Data Factory logging provides detailed metrics for copy activities, including duration, throughput, and errors. By analyzing these logs in Azure Monitor or Log Analytics, you can determine whether the issue is due to network, source, sink, or integration runtime configuration. Once the cause is identified, you can apply targeted optimizations such as adjusting parallelism or scaling the integration runtime.
Answer analysis
Option-by-option breakdown
For each option: why learners choose it and why it is or isn't the right answer here.
- ✗
Scale up the self-hosted integration runtime by adding more nodes to the cluster.
Why it's wrong here
Scaling up the integration runtime can increase capacity, but it is not the first step. You need to determine if the bottleneck is due to insufficient resources or other factors like network bandwidth or source limitations. Adding nodes without understanding the issue may waste resources. Log analysis should precede any scaling action.
- ✗
Change the copy activity to use a staging storage account with PolyBase for faster loading into Synapse Analytics.
Why it's wrong here
Using PolyBase with staging can improve load performance into Synapse Analytics, but it is a specific optimization that may not address the root cause. If the bottleneck is in reading from the on-premises SQL Server or network transfer, PolyBase may not help. You should first identify the bottleneck through logging. This is not the first step to take.
- ✗
Increase the degree of copy parallelism in the copy activity settings to 32.
Why it's wrong here
Increasing parallelism can improve throughput, but without diagnosing the root cause, it may not help and could overwhelm the source or sink. The self-hosted integration runtime has limited resources, and excessive parallelism can degrade performance. You should first analyze logs to determine if parallelism is the bottleneck. Making blind changes is not a recommended first step.
- ✓
Enable logging in Azure Data Factory and analyze the copy activity execution details in Azure Monitor.
Why this is correct
Azure Data Factory provides detailed logging for copy activities, including duration, data read/written, throughput, and errors. By enabling diagnostic settings and sending logs to Azure Monitor or Log Analytics, you can analyze execution details to identify bottlenecks such as network latency, parallel copy settings, or integration runtime performance. This is the first step to diagnose the issue before making changes.
Go deeper
Related to this question
Learn chapter
Monitor Data Storage and Processing
Key term
Data Transformation Pipelines
Data transformation pipelines are automated sequences of steps that take raw data from a source, clean and reshape it into a usable format, and then load it into a destination for analysis or storage.
Key term
Azure Data Factory
Azure Data Factory is a cloud-based data integration service that lets you create, schedule, and orchestrate data pipelines to move and transform data from various sources to destinations.
About these practice questions
One of 509 original DP-203 practice questions on Courseiva, each with a full explanation and wrong-answer analysis — not exam dumps or protected exam content. Learn why practice questions differ from exam dumps →
JA
Written and reviewed by Johnson Ajibi, MSc IT Security
Senior Network & Security Engineer · founder of Courseiva
Last reviewed September 2026 · checked against the official Microsoft exam blueprint
This DP-203 practice question is part of Courseiva's free Microsoft certification practice question bank. Courseiva provides original exam-style practice questions with explanations, topic-based practice, mock exams, readiness tracking, and study analytics to help learners prepare for the DP-203 exam.