Courseiva

DP-900 · topic practice

Describe an analytics workload on Azure practice questions

This domain covers designing and operating analytics solutions on Azure: ingestion with Event Hubs and Data Factory, storage in Synapse and Data Lake Storage Gen2, transformation with Databricks and Stream Analytics, and visualization with Power BI. Questions present realistic scenarios—batch ETL, real-time fraud detection, warehouse migration—and ask you to select the appropriate service, feature, or configuration.

Courseiva uses original exam-style practice questions designed for learning and revision. The goal is to understand the concepts, recognise exam patterns, and improve through explanations — not memorise copied exam dumps.

Editorial oversight:Johnson Ajibi· MSc IT Security, IEEE Senior Member
20 questionsDomain: Describe an analytics workload on Azure

What the exam tests

What to know about Describe an analytics workload on Azure

Map each scenario to the right Azure analytics service: Data Factory for orchestration, Synapse or Databricks for transformation, Event Hubs and Stream Analytics for streaming, Power BI for visualization. The single most important thing is matching batch versus real-time requirements to the correct service.

Choosing between Azure Synapse Analytics, Databricks, and Data Factory for batch versus streaming workloads

Ingesting and processing events with Azure Event Hubs and Azure Stream Analytics

Storing analytical data in Azure Data Lake Storage Gen2 and Synapse dedicated SQL pools

Building Power BI semantic models, aggregations, and reports over large datasets

Watch out for

Common Describe an analytics workload on Azure exam traps

  • ▸Confusing Azure Data Factory orchestration with Synapse or Databricks compute, and picking the wrong tool for distributed transformation.
  • ▸Assuming Stream Analytics handles model scoring; near real-time machine learning inference typically requires Azure Machine Learning endpoints.
  • ▸Treating Data Lake Storage Gen2 as a relational store, or forgetting hierarchical namespace is what enables data lake semantics.

Practice set

Describe an analytics workload on Azure questions

20 questions · select your answer, then reveal the explanation

Question 1mediummultiple choice
Review the full routing breakdown →

A logistics company receives real-time GPS tracking data from its delivery fleet via Azure Event Hubs. The data is a continuous stream of location updates (vehicle ID, latitude, longitude, timestamp). Additionally, the company has daily static route plan files in CSV format stored in Azure Data Lake Storage Gen2. The operations team needs to combine the live GPS stream with the route plans to create a near real-time dashboard showing if delivery vehicles are on schedule. They also want to run historical queries on both the stream data and route plans using T-SQL, without moving the data to another store. Which Azure service should they use as the primary analytics platform?

A retail company needs to build an analytics pipeline on Azure. They ingest sales data from multiple store systems and an online e-commerce platform. The data must be cleaned, transformed, and loaded into a data warehouse for reporting. The company wants to use a modern ELT (Extract, Load, Transform) approach where raw data is stored first and then transformed. Order the following steps in the correct sequence for this pipeline. (Drag the steps into the correct order.)

Drag or tap steps into the slots.

Steps
Order
1Step 1
2Step 2
3Step 3
4Step 4
5Step 5

A company is building a modern data warehouse on Azure using a lakehouse approach. Arrange the following steps in the correct order to implement a typical pipeline that starts with raw data ingestion and ends with business reporting.

Drag or tap steps into the slots.

Steps
Order
1Step 1
2Step 2
3Step 3
4Step 4
5Step 5

A data engineering team needs to build a batch processing pipeline that transforms large volumes of sales data stored in Azure Data Lake Storage Gen2. The transformations include aggregations and joins, and the output should be stored back in the data lake as Parquet files. The team wants a serverless compute option that automatically scales and charges per second. Which Azure service should they use?

A data engineering team wants to build a batch analytics pipeline. The raw data is stored in Azure Data Lake Storage Gen2 (ADLS Gen2). The final output will be a set of tables in Azure Synapse Analytics (dedicated SQL pool) that will be used to create reports in Power BI. Arrange the following steps in the correct order for a typical ETL process.

Drag or tap steps into the slots.

Steps
Order
1Step 1
2Step 2
3Step 3
4Step 4
5Step 5

Drag and drop the steps to perform a point-in-time restore of an Azure SQL Database in the correct order.

Drag or tap steps into the slots.

Steps
Order
1Step 1
2Step 2
3Step 3
4Step 4
5Step 5

A company uses Azure Synapse Analytics dedicated SQL pool to run large-scale analytics. The data engineering team notices that queries are slow due to excessive data movement between distributions. Which index type should be recommended to minimize data movement for fact tables that are frequently joined on a specific column?

Which THREE components are required to implement a real-time analytics solution using Azure Stream Analytics? (Choose three.)

A company uses Azure Synapse Analytics for its data warehouse. They notice that queries against a large fact table are slow. The table is partitioned by month and uses clustered columnstore index. Which action would most likely improve query performance?

An organization uses Azure Stream Analytics to process real-time IoT data from millions of devices. They need to ensure that the output is exactly once delivery semantics to a Power BI dataset. Which output configuration should they use?

You run the above Kusto query in Azure Data Explorer. What does the query return?

Exhibit

Refer to the exhibit.

```kusto
StormEvents
| where EventType == "Tornado"
| summarize TornadoCount = count() by State
| order by TornadoCount desc
| take 5
```

A company uses Azure Data Lake Storage Gen2 to store raw data files. Data engineers need to transform this data using a serverless approach without managing infrastructure. Which Azure service should they use?

A company uses Azure Data Factory to copy data from an on-premises SQL Server to Azure Data Lake Storage Gen2. The transfer must be accelerated using WAN optimization. Which Data Factory feature should the company enable?

Which TWO Azure services can be used to perform data transformation in a serverless manner? (Choose two.)

A company uses Azure Synapse Analytics to run large-scale analytics on sales data. They need to ensure that the workload can automatically scale based on demand without manual intervention. What feature should they configure?

Refer to the exhibit. An administrator runs an Azure CLI command to show the status of a Synapse SQL pool. The output shown is returned. What does this output indicate about the SQL pool?

Exhibit

Refer to the exhibit.

{
  "type": "Microsoft.Synapse/workspaces/sqlPools",
  "apiVersion": "2021-06-01",
  "properties": {
    "collation": "SQL_Latin1_General_CP1_CI_AS",
    "maxSizeBytes": 263882790666240,
    "provisioningState": "Succeeded",
    "status": "Online",
    "restorePointInTime": "2023-01-15T08:00:00Z"
  }
}

Which TWO Azure services can be used to perform interactive ad-hoc analytics on large datasets using Apache Spark?

Your company needs to build an analytics solution that can handle both batch and streaming data from IoT devices. The solution must allow complex event processing and real-time dashboards. Which Azure service should you use as the primary data ingestion and processing layer?

An organization runs a mission-critical analytics workload on Azure Synapse Analytics. They need to ensure high availability and automatic failover in case of a regional outage. Which configuration should they implement?

A data engineering team uses Azure Data Factory to orchestrate an ETL pipeline that loads data from an on-premises SQL Server to Azure Synapse Analytics. The pipeline fails intermittently with timeout errors during the copy activity. The network is stable. What should they do first to resolve the issue?

Free account

Track your progress over time

Create a free account to save your results and see which topics improve across sessions.

Focused Describe an analytics workload on Azure sessions

Start a Describe an analytics workload on Azure only practice session

Every question in these sessions is drawn from the Describe an analytics workload on Azure domain — nothing else.

Related practice questions

Related DP-900 topic practice pages

Move into related areas when this topic feels solid.

Frequently asked questions

What does the DP-900 exam test about Describe an analytics workload on Azure?
Map each scenario to the right Azure analytics service: Data Factory for orchestration, Synapse or Databricks for transformation, Event Hubs and Stream Analytics for streaming, Power BI for visualization. The single most important thing is matching batch versus real-time requirements to the correct service.
How should I use these practice questions?
Select your answer before revealing the explanation. Then read why each option is right or wrong — this active recall approach builds retention far faster than re-reading notes.
Can I practise just Describe an analytics workload on Azure questions in a focused session?
Yes — the session launcher on this page draws every question from the Describe an analytics workload on Azure domain. Use a 10-question session first to gauge your baseline, then move to 20 or 30 once the weak spots are clear.
Where can I practise other DP-900 topics?
Use the topic links above to move to related areas, or go back to the DP-900 question bank to see all topics.
Are these real exam questions or dumps?
These are original practice questions written to test the same concepts the DP-900 exam covers. They are not copied from any real exam or dump site.