Courseiva

Databricks-DE-Assoc · domain

Databricks Intelligence Platform

This domain covers the Databricks Intelligence Platform: the workspace, notebooks, clusters, Jobs, Delta Lake, and Unity Catalog. Questions test whether you can identify which component provides governance, orchestration, or compute, and how Unity Catalog governs tables, files, and ML models across workspaces.

41 questions9 easy24 medium8 hard

Focused practice

Practice Databricks Intelligence Platform questions

Scored sessions drawing only from this domain — pick a length below.

Start 20-question practice test →

What this domain covers

What to know about Databricks Intelligence Platform

Be able to pick the right platform component for governance, orchestration, or compute, and explain what Unity Catalog governs. The key point: Unity Catalog is the account-level fine-grained access control layer for tables, files, and ML models across workspaces.

Unity Catalog as the account-level governance layer for tables, files, and ML models

Jobs and workflows for task dependencies, automatic retries, and failure notifications

Cluster types and compute configuration for notebooks and job runs

Delta Lake tables and the lakehouse storage model on the platform

Watch out for

Common Databricks Intelligence Platform exam traps

  • ▸Treating Unity Catalog as workspace-scoped when it is account-level and spans multiple workspaces
  • ▸Confusing Jobs orchestration features with cluster or notebook capabilities
  • ▸Assuming legacy Hive metastore grants behave like Unity Catalog fine-grained access control

Question index

All Databricks Intelligence Platform questions (41)

Click any question to see the full explanation, or start a practice session above.

1

A data engineer is preparing a notebook that must authenticate to cloud storage using a short-lived token that is automatically rotated by Databricks and is never written into the notebook source. The engineer wants the least administrative overhead while keeping secrets out of the code. Which approach should the engineer use?

Medium
2

Refer to the exhibit. A security administrator applies this Unity Catalog policy. What is the impact on users in the 'analyst-group'?

Hard
3

Which THREE of the following are core components of the Databricks Intelligence Platform?

Medium
4

When considering the Databricks Intelligence Platform, what is the primary role of the 'Lakehouse' architecture?

Medium
5

A data engineer is configuring a Databricks job that must run on a schedule and send an email notification if the job fails. They want to minimize manual intervention. Which feature should they use to define the schedule and failure notification?

Medium
6

A data engineer needs to run a nightly transformation that reads a large Parquet dataset, writes a curated Delta table, and then immediately runs OPTIMIZE and VACUUM on that table. The engineer wants each step to be observable, retryable, and to avoid data loss if VACUUM fails. Which orchestration approach best meets these requirements?

Hard
7

A data engineer needs to store structured data in a cloud object storage location while maintaining full ACID guarantees. Which storage format is the foundation of the Databricks Lakehouse architecture that enables this functionality?

Easy
8

A data engineer is configuring a Databricks cluster to run a Spark job that processes large datasets. The job requires high memory and will run for several hours. The engineer wants to minimize costs while ensuring the job completes successfully. Which cluster configuration should the engineer choose?

Medium
9

A data engineer is designing a secure architecture using the Databricks Intelligence Platform. Which TWO of the following statements accurately describe the role and capabilities of Unity Catalog within this platform? (Choose TWO)

Hard
10

An organization is adopting the Databricks Intelligence Platform and wants to leverage Mosaic AI for building custom machine learning models. Which feature allows data engineers to track machine learning experiments, log parameters, and manage model artifacts reliably?

Medium
11

Which command is used to query the history of a Delta table to perform time travel?

Medium
12

A data engineer is using Databricks SQL to analyze data stored in a Delta table. The engineer wants to optimize query performance by leveraging Delta Lake features. Which TWO actions should the engineer take to improve query performance on the Delta table? (Choose two.)

Medium
13

A data engineer is designing a pipeline on Databricks that requires ACID transactions and schema enforcement for streaming data. Which storage abstraction should they use to ensure data reliability and support time travel?

Medium
14

Which Databricks feature should be used to securely share data with external organizations without duplicating the data?

Medium
15

A data engineer is designing a pipeline and needs to ensure that data remains consistent during concurrent read and write operations. Which Databricks feature provides the mechanism to track and validate these operations?

Medium
16

A data engineering team needs to ingest streaming data from Kafka into a Delta table while maintaining exactly-once processing guarantees and low latency. Which Databricks Intelligence Platform feature should they utilize to build this streaming pipeline declaratively?

Medium
17

Which TWO of the following statements accurately describe the functionality of Unity Catalog within the Databricks Intelligence Platform?

Medium
18

Which TWO of the following are primary benefits of using Delta Live Tables (DLT) for data pipeline development?

Medium
19

Which THREE of the following are primary components of the Databricks Lakehouse architecture?

Medium
20

A data engineer is setting up a Databricks job that runs a notebook on a schedule. The job must process data from a source that is updated daily and write results to a Delta table. The engineer wants to ensure that if the job fails, it automatically retries up to three times. Which feature should the engineer configure in the job settings to achieve this?

Medium
21

Refer to the exhibit. A Databricks administrator is using the Unity Catalog JSON policy to manage access. If the 'data_scientist_1' user attempts to execute an 'UPDATE' command on the 'orders' table, what will be the result?

Hard
22

A data engineer is working in a Databricks workspace where Unity Catalog is enabled. They need to run a SQL query that reads from the table sales in the catalog prod and schema marketing. Which fully qualified name should they use?

Easy
23

Which capability is provided by Databricks' integration with MLflow?

Easy
24

A data engineer is using Databricks Asset Bundles to deploy a data pipeline that includes a job and a notebook. The engineer wants to ensure that the deployment is idempotent and can be rolled back if needed. Which TWO of the following statements accurately describe the benefits of using Databricks Asset Bundles for this scenario? (Choose two.)

Medium
25

A data engineer is designing a pipeline on Databricks to process streaming data. Which architectural component acts as the unified storage layer, allowing both batch and streaming workloads to access the same underlying data files in a data lake?

Medium
26

Which feature of the Databricks Intelligence Platform allows users to manage fine-grained access control across workspaces for tables, files, and machine learning models?

Easy
27

Refer to the exhibit. A data engineer encounters this error while trying to list tables in a schema. Based on the error, what must the engineer do to resolve it?

Hard
28

A data engineer is designing a workflow that requires running a series of tasks with dependencies. The workflow must be able to retry failed tasks automatically and send notifications on failure. The engineer also needs to ensure that the workflow can be triggered on a schedule and via an API. Which Databricks feature should the engineer use?

Hard
29

A data engineer needs to configure a Databricks Job to orchestrate a data pipeline that includes a Python task, a SQL task, and a notebook task. The pipeline requires passing a dynamic run identifier from the Python task to the subsequent SQL and notebook tasks. Which mechanism should the data engineer use to achieve this task-to-task dependency parameter passing?

Medium
30

A data engineer working in a Databricks workspace needs to create a new notebook that will be shared with teammates in the same workspace. They want the notebook to be organized in a folder named 'team_project' and be visible to all workspace users. Which action should the data engineer take to accomplish this in the Databricks workspace?

Easy
31

A team is transitioning to the Databricks Intelligence Platform. Which TWO actions are required to successfully register a table in Unity Catalog using the three-level namespace?

Hard
32

A data engineer needs to share a Delta table with an external partner who does not have a Databricks account. The engineer wants to provide read-only access to the table for a limited time, ensuring the partner cannot access any other data. Which Databricks feature should the engineer use?

Medium
33

A data engineer needs to run a Databricks notebook that processes data stored in an external ADLS Gen2 location. The notebook must access the data securely without embedding credentials in the notebook code. The engineer has already configured a Unity Catalog external location with a storage credential. Which method should the engineer use to read the data?

Easy
34

Which TWO of the following capabilities are native to the Databricks Unity Catalog?

Medium
35

A data engineer is designing a Delta Lake table that will be used for both batch and streaming reads. The table must support upserts from a streaming source and maintain ACID transactions. The engineer wants to ensure that the table can be efficiently queried by downstream consumers using SQL while minimizing storage costs. Which TWO actions should the engineer take to meet these requirements? (Choose two.)

Hard
36

A data engineer needs to automate the ingestion and transformation of data within Databricks with minimal manual intervention. Which feature is most appropriate for orchestrating these data workflows?

Medium
37

A data engineer needs to create a Delta table in Unity Catalog that will be used by multiple teams. The engineer wants to ensure that the table is governed by Unity Catalog and that access can be controlled using SQL GRANT statements. What is the correct way to create the table?

Easy
38

Refer to the exhibit. A data engineer is deploying an automated job. Based on the provided configuration, what is the primary benefit of using the 'autoscale' attribute in this cluster definition?

Medium
39

When a data engineer needs to automate a recurring ETL job, which Databricks tool is most appropriate for orchestrating tasks and handling dependencies?

Medium
40

A data engineer needs to run a Databricks notebook on a schedule every day at 8:00 AM UTC. The notebook performs data transformations and writes results to a Delta table. The engineer wants to ensure the job runs reliably and can be monitored. Which Databricks feature should the engineer use to schedule and monitor the notebook?

Easy
41

Which Databricks compute resource is specifically optimized for running BI dashboards and SQL queries?

Easy

Frequently asked questions

What does the Databricks Intelligence Platform domain cover on the Databricks-DE-Assoc exam?
Be able to pick the right platform component for governance, orchestration, or compute, and explain what Unity Catalog governs. The key point: Unity Catalog is the account-level fine-grained access control layer for tables, files, and ML models across workspaces.
How many questions are in this domain?
This page lists all 41 Databricks Intelligence Platform questions in the Databricks-DE-Assoc question bank. The actual exam draws from this domain proportionally to its weighting in the official exam blueprint.
What is the best way to practise this domain?
Start with a short focused session (10 questions) to identify gaps, then work through explanations. Repeat with a longer session once the weak areas feel solid.
Can I practise only Databricks Intelligence Platform questions?
Yes — the session launcher on this page filters questions to this domain only. Choose any session length for inline explanations and scoring.
databricks-data-engineer-associate DATABRICKS-DATA-ENGINEER-ASSOCIATE databricks intelligence platform Practice Questions