Courseiva

Salesforce Certified Data Architecture and Management Designer (SF-Data-Arch) — Questions 1–75

222 questions total · 3pages · All types, answers revealed

Page 1 of 3

Page 2
1
MCQeasy

A Salesforce data architect is explaining the concept of a 'golden record' to business stakeholders. Which statement best describes a golden record in a master data management context?

A.A record that is stored in the system of record and never changed.
B.The most recently updated record from any source system.
C.A duplicate-free record created by merging all duplicate records in Salesforce.
D.A single, authoritative version of a master data entity that is trusted across the enterprise.
AnswerD

A golden record is the single, authoritative version of a master data entity, such as Customer or Product. It is created by applying survivorship rules to data from multiple source systems. It is trusted because it represents the best available data. This definition aligns with MDM principles and is easily understood by stakeholders.

Why this answer

A golden record is the single, authoritative version of a master data entity that is trusted across the enterprise. It is created by applying survivorship rules to data from multiple source systems, ensuring the best values are selected. It is not simply the latest, static, or a Salesforce-only merged record.

This concept is foundational to MDM.

Exam trap

The trap here is confusing a golden record with the most recent record, a static record, or a Salesforce-only merged record, when it is actually the authoritative, cross-system version.

2
MCQmedium

A Data Architect is tasked with cleaning up duplicate records in a large Salesforce organization. The duplicates have been identified across multiple objects. What is the most effective approach to maintain data integrity during this process?

A.Use Data Loader to export all records, delete them from Salesforce, and re-import unique records.
B.Delete records using a hard-delete command to ensure they are permanently removed.
C.Merge duplicate records using the native Salesforce Merge feature.
D.Update the records to a dummy state and hide them from all user profiles.
AnswerC

The native Merge feature is the safest way to consolidate duplicates. It automatically handles the reparenting of child records, such as Opportunities or Cases, to the master record. This prevents data loss and ensures that the history of interactions remains linked to the correct, surviving record in the system.

Why this answer

Using Duplicate Management rules alongside the Data Import Wizard or API-based deletion is the standard approach. By identifying duplicates through native features and merging them, the architect preserves related records (like child objects) that would otherwise be deleted during a manual record deletion. This ensures that the parent-child relationships remain intact, preventing data orphans and maintaining the overall integrity of the relational database structure.

Exam trap

Candidates frequently suggest manual deletion or custom Apex deletion scripts, forgetting that native Merge features are specifically designed to re-parent related records automatically, preserving critical relational data integrity.

3
MCQmedium

An architect is modeling a 'Subscription' service where Customers have multiple active subscriptions. Which approach best handles the 'Current' vs 'Historical' subscription data for reporting?

A.Use two separate objects for Current and Historical subscriptions.
B.Use a single object with 'Status' and 'Date' fields.
C.Delete historical records to save storage space.
D.Use a multi-select picklist to track history.
AnswerB

A single object with a status field and date range is the most efficient design. It allows for simple filtering (e.g., WHERE Status = 'Active') and makes historical reporting straightforward. This pattern is easy to maintain and scales well as the number of subscription records grows.

Why this answer

Using a combination of a status field and a date range is the standard way to model active versus historical subscriptions. By flagging records as 'Active' and using the 'End Date' field, the architect allows the system to easily filter for current records while maintaining a full history. This design supports both operational reporting and long-term analytical tracking, which is essential for subscription-based business models.

Exam trap

Candidates often try to create separate objects for historical and current records, leading to complex data synchronization challenges rather than utilizing simpler status-driven schemas.

4
MCQeasy

Universal Containers is classifying their data based on sensitivity levels (e.g., Public, Internal, Confidential, Restricted). Which Salesforce feature allows them to record this classification directly on the field definition?

A.Field-Level Security (FLS) settings.
B.Shield Platform Encryption.
C.Custom Metadata Types.
D.Data Classification Metadata fields.
AnswerD

Data Classification metadata (Data Sensitivity Level, Compliance Categorization, etc.) is a native feature in Salesforce. It allows admins to tag fields with specific values, making it easy to run reports on what data is sensitive and ensuring the Org meets various privacy and auditing standards.

Why this answer

Salesforce provides built-in Data Classification fields for every custom and standard field. This allows organizations to track the Data Owner, Field Usage, Data Sensitivity, and Compliance Categorization directly in the metadata, which is essential for data governance and regulatory compliance reporting.

Exam trap

Candidates often confuse field-level security or custom picklist fields with native governance tools, failing to recognize dedicated metadata features built specifically for data classification.

5
MCQeasy

A Salesforce data architect is migrating 100,000 Opportunity records from a legacy system. The legacy data includes a custom field 'Legacy_Region__c' that must be mapped to a new picklist field 'Region__c' in Salesforce. The picklist values in Salesforce are: 'North America', 'Europe', 'Asia', 'Latin America'. The legacy data uses values like 'NA', 'EU', 'APAC', 'LATAM'. What is the most efficient way to transform the data during migration?

A.Load the legacy values as-is and then use Data Loader to update the picklist values.
B.Use a formula field in Salesforce to convert the legacy values after loading.
C.Perform the transformation in an ETL tool before loading the data.
D.Create a workflow rule to update the picklist field based on the legacy field.
AnswerC

Using an ETL tool allows the architect to map legacy values to the correct picklist values before loading into Salesforce. This ensures data integrity and avoids load errors due to invalid picklist values. The transformation can be done via lookup tables or scripts within the ETL tool. This is the most efficient and reliable method for large-scale migrations.

Why this answer

The most efficient way to transform legacy values to match Salesforce picklist values is to perform the transformation in an ETL tool before loading. This avoids load errors and ensures data quality. Formula fields, post-load updates, and workflow rules are either not possible or inefficient for this purpose.

Exam trap

The trap here is assuming that Salesforce can automatically map legacy values to picklist values or that post-load updates are efficient.

6
MCQeasy

What is the primary architectural goal of using the Salesforce Bulk API 2.0?

A.To provide real-time updates for end-user applications.
B.To bypass all security and validation rules.
C.To efficiently manage large volume data operations.
D.To allow for direct SQL query access to the database.
AnswerC

Bulk API 2.0 is purpose-built to handle millions of records by automatically partitioning the data into optimized batches. This allows for reliable, high-throughput processing while maintaining stability in the Salesforce environment, making it the primary tool for large-scale data imports and exports.

Why this answer

The Bulk API 2.0 is designed specifically for handling large data volumes by automating the batching process. It simplifies the developer experience while optimizing for high throughput. By offloading the batching logic to the Salesforce platform, it ensures that data loads occur in a manner that respects system resources and governor limits, which is the foundational requirement for managing enterprise-scale data in the cloud.

Exam trap

Candidates often confuse the Bulk API with standard REST/SOAP APIs, failing to realize the Bulk API is specifically engineered for throughput, not for real-time or low-latency requests.

7
Multi-Selecthard

A global insurer is preparing for a data privacy audit. The architect must demonstrate that the Salesforce org can identify where personal data lives and prove who accessed it. Which two capabilities should the architect include in the governance solution? (Choose two.)

Select 2 answers
A.Data Classification metadata on fields, using the Data Classification feature to tag fields as personal or sensitive.
B.A nightly Data Loader job that exports all records for offline retention.
C.Field History Tracking enabled on every field in the org to record old and new values.
D.Event Monitoring with log retention and the ability to analyze LoginEvent and ApiEvent records.
E.Enabling Shield Platform Encryption on all text fields to mask values at rest.
AnswersA, D

Data Classification lets administrators tag fields with data sensitivity and compliance categories, producing a searchable inventory of where personal data resides. Auditors can then see which fields are classified as personal without manually inspecting every object. This directly supports the requirement to identify where personal data lives across the org.

Why this answer

The audit asks two distinct questions: where personal data resides and who accessed it. Data Classification metadata answers the first by tagging fields with sensitivity categories, and Event Monitoring answers the second by retaining access and API logs for analysis. Field history, bulk exports, and platform encryption do not provide classification inventory or access evidence.

Exam trap

The trap here is substituting a protective control such as Shield Platform Encryption or field history for the classification and access-evidence controls the audit actually requires.

8
Multi-Selectmedium

A data architect is planning to implement a Skinny Table for a custom object with 10 million records to improve query performance. Which two considerations are critical when designing the Skinny Table? (Choose two.)

Select 2 answers
A.The Skinny Table requires a synchronization mechanism to keep data current with the main object.
B.The Skinny Table can be used as the primary data source for all reporting to replace the main object.
C.The Skinny Table should be indexed on all fields to maximize query flexibility.
D.The Skinny Table must include all fields from the original object to maintain data integrity.
E.The Skinny Table should include only the fields required for the most common queries.
AnswersA, E

Because a Skinny Table is a separate custom object, it does not automatically reflect changes to the main object. A synchronization process, such as a trigger or batch job, must be implemented to update the Skinny Table when the main object changes. Without synchronization, the Skinny Table would become stale and produce incorrect query results.

Why this answer

A Skinny Table is a denormalized copy of a subset of fields from a large object, used to accelerate specific queries. The two critical considerations are selecting only the fields needed for common queries to minimize row size, and implementing a synchronization mechanism to keep the copy up to date. Without these, the Skinny Table would either be too large to be effective or become inconsistent with the source data.

Exam trap

The trap here is assuming a Skinny Table must mirror all fields or can replace the main object entirely, when it is a targeted performance optimization.

9
MCQmedium

Which governance component ensures that Salesforce Orgs are prepared for major feature updates and potential breaking changes?

A.Data Classification Policy.
B.Change Management Process.
C.Data Dictionary.
D.Master Data Management (MDM).
AnswerB

Change management provides the rigorous framework necessary for platform testing and deployment. It ensures that updates are reviewed, tested, and approved, which is essential for identifying potential breaking changes before they reach production. Without this, organizations are prone to unexpected issues and downtime during Salesforce's thrice-yearly release cycle.

Why this answer

The Change Management process is critical for handling major updates. It includes impact analysis, sandbox testing, and regression testing. By governing this process, architects ensure that features are validated in a safe environment before production deployment.

This minimizes the risk of system outages, ensures business continuity, and keeps the Salesforce environment stable during the rapid release cadence of the platform.

Exam trap

Candidates often choose 'Release Management' or 'CI/CD' as the governance component, missing that 'Change Management' is the broader framework that governs the impact and risk of those releases.

10
MCQeasy

A data architect is reviewing a custom object that has grown to 8 million records. Users report that list views and reports are slow because they sort on a text field, Priority__c, which has only three possible values. What should the architect recommend to improve performance?

A.Enable Divisions on the object and assign records to divisions based on Priority__c.
B.Request a custom index on Priority__c to speed up sorting.
C.Convert Priority__c to a picklist field with a restricted set of values to improve selectivity.
D.Remove sorting on Priority__c from list views and reports, and instead sort on an indexed field or filter on a more selective field.
AnswerD

Sorting on a low-cardinality field like Priority__c forces the query optimizer to scan many records because the field is not selective. Removing the sort or replacing it with a sort on an indexed, high-cardinality field can dramatically improve performance. Filtering on a more selective field also reduces the number of records that need to be sorted.

Why this answer

For large data volumes, query performance is heavily influenced by field selectivity. A field with only three distinct values is not selective, so sorting or filtering on it forces the platform to scan many records. The best approach is to avoid sorting on such fields and instead use an indexed, high-cardinality field for sorting or filtering.

This reduces the data set the query must process.

Exam trap

The trap here is assuming that any field used in sorting or filtering should be indexed, even when it has very low cardinality.

11
MCQeasy

The governance board at a financial services firm wants to know when critical fields on the Opportunity object were changed, who changed them, and what the previous value was, without writing custom code. Which native capability should the architect enable?

A.Enable Field History Tracking on the selected Opportunity fields
B.Build a Flow that writes a record to a custom log object on every Opportunity update
C.Schedule a nightly Data Loader export of Opportunity records for comparison
D.Create a Workflow Rule that posts to Chatter when Opportunity fields change
AnswerA

Field History Tracking is the native, no-code feature that records changes to selected fields, including the user who made the change, the timestamp, and the old and new values. Enabling it on the chosen Opportunity fields directly answers the governance board's question about when and by whom critical fields changed.

Why this answer

The board wants a native, no-code record of when critical Opportunity fields changed, who changed them, and the prior values. Field History Tracking provides exactly that for selected fields and requires only configuration. Chatter notifications, custom Flow logging, and nightly exports each capture partial or indirect information and cannot deliver a reliable, attributed field-change history.

Exam trap

The trap here is equating real-time notifications or scheduled snapshots with an auditable history, when only Field History Tracking natively stores prior values with user and timestamp attribution.

12
MCQmedium

Which strategy best supports Data Governance when scaling a Salesforce implementation across multiple business units with varying data requirements?

A.Centralized Governance.
B.Decentralized Governance.
C.Federated Governance.
D.Automated Governance.
AnswerC

Federated governance provides a scalable framework that combines central standards with localized execution. It empowers business units to manage their specific domains while staying within an enterprise-aligned policy framework. This flexibility is essential for large organizations to maintain consistency without hindering the speed of business operations and local innovation.

Why this answer

A Federated Data Governance model is best for large, complex organizations. It allows for central oversight and standardized definitions while giving local business units the autonomy to manage their domain-specific data. This balanced approach avoids the bottlenecks of a purely centralized model while preventing the chaos of a decentralized one, ensuring both compliance and business agility at scale.

Exam trap

Candidates often choose purely centralized governance models, failing to recognize that scaling across diverse business units requires balancing local autonomy with overarching central standards.

13
MCQhard

Refer to the exhibit. During a high-volume migration, the team encounters the provided error frequently. What is the most likely cause?

A.The batch size is too small, causing excessive commits to the database.
B.Multiple threads are updating child records that share the same parent account.
C.The API user does not have sufficient permissions to modify the account records.
D.The Salesforce database is experiencing a scheduled maintenance window.
AnswerB

When child records are updated, Salesforce locks the parent record to ensure data integrity. If multiple threads attempt to update children of the same parent simultaneously, they compete for the parent lock, resulting in the UNABLE_TO_LOCK_ROW error. This is a classic concurrency bottleneck in multi-threaded bulk migration processes.

Why this answer

The 'UNABLE_TO_LOCK_ROW' error indicates that multiple threads or processes are attempting to update the same record or its parent records simultaneously. In data migration, this usually happens when records sharing the same parent are processed in parallel. By reordering the data or reducing the concurrency in the bulk load, architects can prevent these deadlocks and ensure successful completion without constant retries or system instability.

Exam trap

Candidates frequently mistake locking errors for permission or data format issues, missing the core architectural problem that concurrent multi-threaded updates on shared parent records cause deadlocks.

14
MCQeasy

Universal Containers wants to ensure that when a user updates a field on an Account record, the change is captured and stored for auditing purposes. The solution must be configurable without writing code and must retain field history for up to 18 months. What should a data architect implement?

A.Configure a workflow rule to send an email notification on field change.
B.Create a trigger on the Account object to write changes to a custom object.
C.Enable Field History Tracking on the Account object for the required fields.
D.Use Salesforce Shield Event Monitoring to track field changes.
AnswerC

Field History Tracking is a declarative feature that automatically tracks changes to specified fields and retains history for up to 18 months (or 24 months in some editions). It requires no code and is ideal for auditing field changes. It stores old and new values, the user who made the change, and the timestamp, meeting the requirement for configurable auditing.

Why this answer

Field History Tracking is the standard declarative feature for auditing field changes. It automatically records changes, retains them for up to 18 months, and requires no code. Other options either involve code, track the wrong type of activity, or do not store historical data, making them unsuitable for this requirement.

Exam trap

The trap here is confusing event monitoring or triggers with field history tracking, but only Field History Tracking provides declarative, long-term field-level auditing.

15
MCQmedium

An organization wants to use 'Person Accounts' to manage B2C relationships. Which consideration is most important for a Data Architect when evaluating this change?

A.Person Accounts can be disabled by contacting Salesforce Support.
B.They replace the need for custom objects for B2C data.
C.They fundamentally change the object model and are non-reversible.
D.They require a Business Account for every Person Account.
AnswerC

Person Accounts merge the Account and Contact entities. This is a irreversible change that affects almost every part of the system, including API integrations and reporting. A data architect must perform a thorough impact analysis because this change is permanent and alters the core structure of the database.

Why this answer

Person Accounts merge the Account and Contact models into a single entity, which has long-term implications for the data schema. Once enabled, this change cannot be reversed, and it impacts existing integrations, custom code, and reporting. A Data Architect must ensure that all third-party systems are updated to handle the new object structure, as the standard Account-Contact hierarchy is fundamentally altered, potentially breaking legacy processes that expect distinct objects.

Exam trap

Candidates often treat Person Accounts as a minor configuration change, failing to realize the massive, permanent impact on the data model and existing integrations.

16
MCQhard

A company needs to migrate data into a custom object while maintaining the original CreatedDate from a legacy system. Which of the following is required?

A.A custom Apex trigger to overwrite the CreatedDate field.
B.Enable 'Set Audit Fields upon Record Creation' permission.
C.Use the Salesforce Import Wizard with mapped headers.
D.Create a formula field to store the legacy date.
AnswerB

This is the only supported mechanism for populating audit system fields during record insertion. By enabling this permission on the integration user profile, the API is granted the capability to accept values for fields like CreatedDate, allowing for accurate migration of legacy audit data into the platform.

Why this answer

Maintaining the original CreatedDate requires the 'Set Audit Fields upon Record Creation' user permission. Without this, Salesforce ignores the provided CreatedDate and assigns the current server time upon insertion. This permission is a highly sensitive setting and must be enabled specifically for the integration user through the user record or the appropriate permission set, ensuring audit traceability for historical records.

Exam trap

Candidates incorrectly assume that administrative privileges or data loader settings alone can override system fields. They fail to realize that the 'Set Audit Fields' permission is a specific, separate security requirement.

17
MCQhard

A global Salesforce org has implemented a data governance framework. The governance team needs to ensure that data is retained according to legal and regulatory requirements, and that it is securely deleted when no longer needed. They are evaluating Salesforce's native capabilities for data lifecycle management. Which statement accurately describes a limitation or behavior of Salesforce's native data retention and deletion features that the team must consider?

A.When a record is deleted in Salesforce, it is immediately and permanently removed from all backups and disaster recovery systems, ensuring compliance with right-to-erasure requests.
B.Salesforce automatically archives inactive records to a separate cold storage system after 2 years, and users can retrieve them via a standard API.
C.Salesforce Field Audit Trail can be used to retain field history for up to 10 years, but it does not automatically delete the data after the retention period; manual intervention is required.
D.Salesforce provides a built-in data retention policy engine that automatically deletes records based on custom retention rules without any development.
AnswerC

Field Audit Trail retains field history for up to 10 years, but it does not automatically purge data after the retention period. Organizations must manually delete or archive the history if they need to enforce a retention limit. This is a key limitation to consider when designing data lifecycle governance, as automated deletion is not provided.

Why this answer

Field Audit Trail retains field history for up to 10 years, but it does not automatically delete data after that period. Manual deletion or archiving is required to enforce retention limits. This is a critical limitation for data lifecycle governance, as organizations must implement their own processes to purge data when it is no longer needed, ensuring compliance with legal and regulatory requirements.

Exam trap

The trap here is assuming that Salesforce provides automated retention enforcement or immediate permanent deletion, when in fact manual intervention is often required.

18
MCQmedium

A Data Architect needs to implement a field-level security strategy where PII is visible only to the 'HR' profile. However, other profiles need to run reports on the account data without seeing the PII. What is the correct way to handle this?

A.Create two separate objects and use a lookup relationship to link them.
B.Apply field-level security to the PII field for all profiles except HR.
C.Implement a custom Lightning component to mask the field data.
D.Use an Apex trigger to clear the field value if the user is not in the HR profile.
AnswerB

Field-level security allows for granular control over who can see specific data fields. By restricting the visibility for all profiles except HR, the system ensures that sensitive information is properly protected in the UI, API, and all report types, meeting the data privacy requirement simply and effectively.

Why this answer

Field-level security (FLS) is the primary method for controlling access to specific fields. By defining FLS for the HR profile and restricting access for others, the architect ensures data privacy. When users without access run reports, Salesforce automatically hides the field, preventing unauthorized viewing while allowing them to report on other non-sensitive data.

This approach is the native, secure way to manage field visibility without complex custom components or data splitting strategies.

Exam trap

Candidates often propose complex sharing rules or record-type splitting, failing to realize that Field-Level Security is the most direct and secure way to handle visibility for sensitive fields.

19
MCQmedium

A data architect is migrating 8 million Contact records from a legacy CRM into Salesforce Sales Cloud. The legacy system stores phone numbers as free-text strings in multiple formats (e.g., '(415) 555-1212', '415.555.1212', '4155551212'). After migration, users must be able to search for Contacts by phone number and have the numbers display consistently. Which approach should the architect use to ensure phone numbers are searchable and standardized?

A.Store the original free-text phone numbers in a Text field and create a formula field that formats the number for display.
B.Store phone numbers in the standard Phone field and rely on Salesforce's automatic normalization and search indexing.
C.Create a custom Text field with a validation rule that enforces a specific phone number format, and load the raw values.
D.Pre-process the legacy data in the ETL layer to normalize all phone numbers to a single format, then load them into the standard Phone field.
AnswerD

Normalizing phone numbers in the ETL layer before loading ensures a consistent canonical format at rest, which makes the standard Phone field searchable and displayable uniformly. This approach leverages the built-in indexing of the Phone field and avoids the need for custom parsing logic in Salesforce, meeting both search and consistency requirements.

Why this answer

Normalizing phone numbers during the ETL process before loading them into the standard Phone field ensures that all values conform to one format. This makes the standard field's search indexing effective and provides consistent display without custom code. The other options either leave inconsistent data at rest or rely on features that do not automatically normalize free-text input.

Exam trap

The trap here is assuming the standard Phone field automatically normalizes or reformats inconsistent free-text values during import.

20
MCQhard

Refer to the exhibit. An architect is reviewing an error log from a high-volume data load. Which pattern should be implemented to resolve this limit violation?

A.Increase the batch size in the data loader settings.
B.Implement the Bulk DML Pattern using collections.
C.Convert the trigger to an asynchronous future method.
D.Increase the Apex CPU time limit in the Setup menu.
AnswerB

Moving DML operations outside of loops is the fundamental requirement for Salesforce Apex development. By collecting records in a List and performing a single DML call, you stay within the 150 limit per transaction. This is a critical architectural pattern for ensuring scalable code during heavy data ingestion processes.

Why this answer

The error indicates that DML operations are occurring inside a loop, exceeding the Salesforce governor limit of 150 DML statements per transaction. To resolve this, developers must use the Bulk Pattern, which involves collecting records in a collection (List or Map) and performing a single DML operation outside of the loop. This ensures efficient resource usage and prevents transactions from failing during high-volume data processing scenarios.

Exam trap

Candidates look for asynchronous processing solutions like Queueable Apex, missing that the fundamental error is simply performing individual DML statements inside a loop rather than processing collections.

21
Multi-Selectmedium

A company is defining its Data Stewardship model. Which THREE responsibilities are typically assigned to a Data Steward within a Salesforce-centric data governance framework? (Choose three.)

Select 3 answers
A.Managing daily data quality issues and resolving duplicates.
B.Configuring Salesforce Org-wide defaults and sharing rules.
C.Defining data classification and handling standards.
D.Writing APEX triggers to automate data cleanup.
E.Monitoring data compliance and policy adherence.
AnswersA, C, E

Data stewards are responsible for monitoring and remediating data quality issues. This includes identifying duplicates or incomplete records and coordinating cleanup efforts to maintain a single source of truth, ensuring the Salesforce environment remains clean and usable for all downstream business operations and analytics.

Why this answer

Data Stewards act as the bridge between technical implementation and business utility. Their roles involve maintaining data accuracy, defining standards, and acting as the subject matter experts for specific data domains. By focusing on these areas, they ensure that the data stored in Salesforce remains reliable, compliant, and useful for stakeholders.

These actions prevent the degradation of data quality over time and support ongoing compliance with organizational data policies.

Exam trap

Candidates often include 'coding new features' or 'managing server infrastructure' in the steward's responsibilities, confusing the role of a Data Steward with that of a Developer or System Admin.

22
Multi-Selecthard

Which THREE factors are essential to consider when performing a multi-org data migration consolidation?

Select 3 answers
A.Handling duplicate record IDs across source orgs.
B.Merging disparate security models and profiles.
C.Mapping custom fields with identical names but different data types.
D.Enabling the 'Set Audit Fields' for every user.
E.Using only the standard Import Wizard for migrations.
AnswersA, B, C

Since Salesforce IDs are only unique within a single org, consolidation will lead to ID collisions. Architects must map original records to new unique external keys and cross-reference them to ensure that parent-child relationships remain intact within the new, consolidated environment after the data is migrated.

Why this answer

Consolidating multiple Salesforce orgs requires meticulous planning around unique ID collisions, metadata differences (such as custom field mapping), and security model alignment. Since each org has its own unique record IDs and potentially overlapping data, an intermediary mapping strategy is vital to ensure that relationships are preserved correctly in the new, unified environment without creating duplicate records or data conflicts.

Exam trap

Candidates often focus only on record migration while neglecting the metadata and security model. They fail to realize that differing profiles and field data types will break the target org structure.

23
MCQhard

During a migration, you encounter an error stating that the record exceeds the maximum character limit for a field. What is the most robust way to handle this?

A.Automatically truncate the text in the ETL process to fit the limit.
B.Increase the Salesforce field length to the maximum possible.
C.Evaluate the data and request business approval for truncation or re-mapping.
D.Skip the records that exceed the character limit.
AnswerC

Data integrity is paramount in any migration. The architect must investigate the source of the excess data and work with stakeholders to decide the best course of action. This ensures that the business is informed of potential data loss and approves the method used, mitigating risks and ensuring compliance with business requirements.

Why this answer

Truncating data without business approval is a data integrity risk. The architect should analyze the source data to determine why it exceeds the limit and work with stakeholders to either shorten the source text, map it to a larger field, or confirm if truncation is acceptable. Simply truncating in the ETL layer may result in the loss of critical information, which can lead to compliance or business issues down the line.

Exam trap

Candidates often choose technical shortcuts like automatic ETL truncation or script-based rejection, missing the requirement to involve business stakeholders for formal impact approval before modifying data.

24
MCQmedium

A company is migrating legacy data and needs to preserve the CreatedDate and CreatedById. Which feature should they enable?

A.Data Loader batch settings
B.Set Audit Fields permission
C.Field History Tracking
D.Owner migration setting
AnswerB

The 'Set Audit Fields' permission is the specific security requirement needed to override system-managed fields. When this is enabled for the integration user, the system allows the ingestion of original creation timestamps and user IDs, ensuring that legacy data maintains its historical integrity after the migration.

Why this answer

By default, Salesforce sets the CreatedDate and CreatedById to the time and user of the import. To preserve historical audit logs, Salesforce provides the 'Set Audit Fields' permission. This is vital for data migration projects where historical accuracy is required for compliance and reporting.

Without this, the audit trail of the legacy system is lost, which can lead to significant regulatory and reporting issues.

Exam trap

Candidates frequently assume that standard administrative permissions are sufficient, forgetting that 'Set Audit Fields' is a specialized, high-privilege permission that must be explicitly enabled in the org.

25
MCQmedium

What is the main advantage of using an ETL tool rather than the standard Salesforce Data Loader for a large-scale enterprise migration?

A.It is always free to use.
B.It provides visual workflow orchestration and error handling.
C.It bypasses Salesforce security controls.
D.It allows direct database access to the Salesforce backend.
AnswerB

ETL tools offer visual design interfaces that make orchestrating complex migrations much easier than the basic, manual steps required by Data Loader. Features like automated retries and detailed logging streamline the migration process, allowing architects to handle errors gracefully and maintain a clear audit trail of all data movements.

Why this answer

Enterprise ETL tools provide advanced capabilities such as graphical orchestration, error logging, automated retry logic, and seamless connectivity to heterogeneous data sources. While Data Loader is sufficient for simple, smaller loads, ETL tools allow architects to manage complex migration workflows, parallel jobs, and sophisticated data transformations across multiple systems, which is required for large-scale enterprise integration projects.

Exam trap

Exam takers frequently choose standard Data Loader for massive enterprise migrations, forgetting that ETL tools offer critical visual workflow orchestration, error handling, and robust transformation capabilities.

26
MCQmedium

Universal Containers uses Salesforce Sales Cloud as its CRM and a legacy AS/400 system for order fulfillment. Customer records exist in both systems but are keyed differently: Salesforce uses the Account ID (15-character), while the AS/400 uses a legacy customer number. The data architect must enable ongoing synchronization of customer master data between the two systems without creating duplicate records. Which approach should be used?

A.Enable Salesforce Duplicate Management with a matching rule on the Account Name and Billing Street fields, and rely on it to prevent duplicates during integration loads.
B.Use the Salesforce Account ID as the primary key in the AS/400 system and overwrite the legacy customer number with the 15-character Salesforce ID.
C.Configure a Salesforce-to-Salesforce (S2S) connection between the two systems and let Salesforce automatically reconcile customer records.
D.Create a custom external ID field on the Account object and populate it with the legacy customer number, then use that field as the matching key in the integration.
AnswerD

A custom external ID field stores the legacy customer number and is indexed, enabling the integration to upsert records reliably. Salesforce's upsert operation matches on the external ID, so records are created or updated without duplication. This is the standard MDM pattern for cross-system key mapping when source systems use different identifiers.

Why this answer

The cross-system key mapping problem is solved by storing the legacy identifier in a custom external ID field on the Salesforce Account. This field is indexed and can be used by the integration's upsert operation to match and update the correct record, preventing duplicates. It also preserves the legacy key in the source system, avoiding disruption to existing processes.

Exam trap

The trap here is assuming that Salesforce Duplicate Management or S2S can replace a deterministic external ID mapping for cross-system synchronization.

27
MCQmedium

An enterprise Salesforce org has multiple business units that each maintain their own picklist values for the Industry field on Account. This has caused inconsistent reporting and integration failures. The Data Governance Council mandates a single, standardized Industry picklist across all business units. Which approach best enforces this standardization while allowing controlled local extensions?

A.Use a global picklist value set for the Industry field and allow each business unit to add values only through a formal governance request process that updates the global set.
B.Leave each business unit with its own custom picklist and create a formula field that maps local values to a standardized set for reporting.
C.Replace the Industry field with a text field and rely on a validation rule that checks values against a custom setting maintained by each business unit.
D.Create a single global picklist value set and assign it to the Industry field on Account for all business units, with no ability for local additions.
AnswerA

A global picklist value set ensures a single source of truth for Industry values across all business units, while a formal governance request process provides a controlled path for local extensions. This balances standardization with flexibility, prevents uncontrolled divergence, and keeps reporting and integrations consistent. It directly satisfies the mandate for a standardized field with governed local additions.

Why this answer

The governance council needs one standardized Industry picklist while still permitting controlled local additions. A global picklist value set provides the single source of truth, and a formal request process ensures any new values are reviewed and added centrally. This prevents uncontrolled divergence, keeps integrations and reporting consistent, and gives business units a governed path to request extensions.

Exam trap

The trap here is assuming that standardization requires completely locking down the picklist, when a governed extension process can preserve flexibility without sacrificing consistency.

28
MCQhard

A data architect is designing a master data management solution where customer records from Salesforce and a legacy system must be merged into a single golden record. The legacy system uses a customer ID that is a 10-digit number, while Salesforce uses a 15-character ID. The architect needs to create a unique cross-system identifier. Which approach best ensures uniqueness and traceability?

A.Use Salesforce's 15-character ID as the single identifier for all systems
B.Use a composite key consisting of the source system code and the native ID, stored in an external ID field
C.Use the customer's email address as the unique identifier
D.Generate a random UUID for each customer and store it in a custom field
AnswerB

A composite key combining the source system code and the native ID ensures global uniqueness and traceability. Storing it in an external ID field allows Salesforce to reference the legacy record and perform upserts. This approach is scalable and avoids collisions, as each system's IDs are prefixed with a unique code.

Why this answer

A composite key of source system code and native ID provides a globally unique identifier that also indicates the record's origin. This is essential for master data management, as it allows the organization to trace data back to its source and perform accurate merges. It also supports upsert operations in Salesforce via an external ID field.

Exam trap

The trap here is assuming that a single system's native ID can serve as a universal identifier, ignoring the need for cross-system uniqueness and traceability.

29
MCQeasy

An org has a custom object 'Invoice__c' with a Lookup to Account. Users frequently complain that deleting an Account leaves orphaned Invoice records, and reports show Invoices with no Account. The business wants the platform to prevent an Invoice from being saved without a valid Account. Which change should the architect make?

A.Convert the Lookup to a Master-Detail relationship from Invoice__c to Account.
B.Mark the Lookup field as required on the page layout only.
C.Create a validation rule that checks whether the Account field is blank.
D.Enable a duplicate rule that flags Invoices with a null Account.
AnswerA

Master-Detail makes the parent reference mandatory at the database level, so an Invoice cannot be saved without a valid Account regardless of whether it is created in the UI, via API, or through data loading. It also blocks Account deletion while Invoices exist, directly resolving the orphaned-record problem.

Why this answer

A Master-Detail relationship enforces the parent reference at the data layer, so records cannot be saved without a valid Account and referenced Accounts cannot be deleted while children exist. This is the native way to eliminate orphaned child records across UI, API, and data-load paths.

Exam trap

The trap here is relying on page layout requirements or validation rules, which can be bypassed or do not stop parent deletion, instead of a Master-Detail relationship that enforces integrity natively.

30
MCQmedium

Universal Containers maintains customer master data in Salesforce and a legacy ERP. The data architect needs to ensure that when a customer's address is updated in Salesforce, the change is automatically reflected in the ERP within 15 minutes. The ERP does not support outbound calls. Which Salesforce feature should be used to achieve this near real-time synchronization?

A.Platform Event published by a Process Builder on Account update
B.Scheduled Apex job that queries modified Accounts and calls the ERP API
C.Change Data Capture with a middleware subscribing to the change event stream
D.Outbound Message on the Account object
AnswerC

Change Data Capture (CDC) publishes change events for Salesforce records, including Address changes on Account. A middleware can subscribe to the CDC event stream and update the ERP. This satisfies the near real-time requirement and works with an ERP that cannot make outbound calls because the middleware initiates the update to the ERP.

Why this answer

Change Data Capture provides a near real-time stream of record changes that an external middleware can consume and use to update the ERP. It is the most direct and scalable way to synchronize address changes without requiring the ERP to call Salesforce. Scheduled jobs and outbound messages introduce latency or require specific endpoints, while Platform Events add an unnecessary layer.

Exam trap

The trap here is assuming that Scheduled Apex or Outbound Messages can achieve near real-time synchronization, when they actually introduce delays or require specific external endpoints.

31
MCQhard

A global Salesforce org maintains an Account record that is simultaneously updated by an inbound ERP integration, a nightly Data Loader batch, and a sales rep via the Lightning UI. During a data quality audit, the governance board finds that conflicting values for the 'Annual Revenue' field have overwritten each other, and no one can determine which source was authoritative at any point in time. The board wants to ensure that future conflicts are detected, attributed to a source system, and resolved according to a documented priority order before the record is committed. Which governance mechanism should the Data Architect recommend to satisfy this requirement?

A.Create a validation rule that blocks updates to Annual Revenue unless the running user has the 'Data Governance' permission set, and grant that permission set only to the ERP integration user.
B.Implement a Master Data Management (MDM) system of record with survivorship rules that encode the source priority order and write the golden record back to Salesforce.
C.Implement a Data Steward review queue using a custom object and a scheduled Flow that asks the data steward to manually approve each conflicting value.
D.Enable Field History Tracking on Annual Revenue and create a report that shows all changes, then distribute the report to the governance board weekly.
AnswerB

An MDM hub with explicit survivorship rules is the only option that both stores the precedence order as governed metadata and applies it deterministically whenever sources disagree. It can stamp the winning source and retain lineage, so the audit can show which system was authoritative at commit time. This directly satisfies detection, attribution, and documented resolution before the record is committed.

Why this answer

The requirement is a governed, deterministic conflict-resolution process that records which source system was authoritative. An MDM system of record with survivorship rules encodes the source priority order as governed metadata, applies it automatically when sources disagree, and preserves lineage for audit. Field history, manual review queues, and permission-based write exclusion either act too late, rely on human latency, or block legitimate writes without resolving the underlying conflict.

Exam trap

The trap here is assuming that field history tracking or a manual approval queue constitutes conflict resolution, when neither encodes a source priority order nor prevents conflicting overwrites from committing.

32
MCQeasy

A Salesforce data architect is asked to ensure that account records created by the sales team always include a valid 15-character or 18-character Salesforce ID in a custom external identifier field used by an ERP integration. The ERP rejects records without a valid ID. Which Salesforce feature should the architect use to enforce this at the point of data entry?

A.Field-level security settings that make the external identifier field required for all profiles.
B.Apex triggers that call the ERP synchronously and roll back the transaction if the ERP rejects the record.
C.A record-triggered flow that sends an email alert to the sales manager when the identifier is missing.
D.Validation rules on the Account object that verify the external identifier field is populated and matches the expected ID pattern.
AnswerD

Validation rules evaluate on save and can enforce required fields and format checks, including a REGEX pattern for 15- or 18-character Salesforce IDs. This prevents invalid records from being saved, so the ERP integration never receives a record without a valid identifier. This directly satisfies the requirement at the point of data entry and is the correct choice.

Why this answer

Validation rules are the native Salesforce mechanism to enforce required values and format constraints at save time. A REGEX check can confirm the external identifier is a valid 15- or 18-character ID, blocking bad records before they reach the ERP. Triggers, field-level security, and email alerts do not prevent invalid data from being saved, so they cannot satisfy the integration requirement.

Exam trap

The trap here is confusing field-level security or notification flows with actual data validation, when only validation rules can block a save based on required value and format.

33
Multi-Selectmedium

A data architect is designing a master data management strategy for customer data in Salesforce. The company wants to ensure data quality and consistency across systems. Which two practices should the architect recommend to maintain a single source of truth? (Choose two.)

Select 2 answers
A.Establish a data governance council to define policies and standards
B.Use Salesforce's duplicate management rules to prevent duplicate records
C.Schedule nightly data backups of Salesforce
D.Enable field history tracking on all customer fields
E.Implement a data stewardship program with defined roles and responsibilities
AnswersA, E

A data governance council defines data policies, standards, and decision rights. It ensures that master data is managed consistently across the organization. This is a strategic practice that supports a single source of truth by providing oversight and resolving conflicts between business units.

Why this answer

A data stewardship program and a data governance council are both strategic practices that establish accountability and policies for master data. They ensure that data is managed consistently, which is essential for a single source of truth. The other options are either tactical tools or unrelated to data quality management.

Exam trap

The trap here is confusing tactical data quality features like duplicate rules with strategic MDM practices like governance and stewardship.

34
MCQhard

A data architect at Northern Trail Outfitters is designing a master data management solution where customer records from Salesforce and a legacy ERP must be consolidated into a single golden record. The architect must ensure that the most recent address is always used, but if the most recent address is from the legacy ERP, it should only be used if it has been verified. Which MDM concept should be implemented to achieve this?

A.Deterministic matching
B.Data lineage tracking
C.Survivorship rules with trust scores
D.Data stewardship workflows
AnswerC

Survivorship rules determine which attribute value survives when merging records. By incorporating trust scores, the system can prioritize the most recent address but require verification for legacy ERP data. This ensures data quality and meets the business rule of using verified legacy addresses only.

Why this answer

Survivorship rules with trust scores allow the MDM system to apply business logic such as 'most recent, but verified for legacy ERP' when consolidating records. Trust scores can be assigned based on data source and verification status, ensuring the golden record reflects the most reliable and recent information.

Exam trap

The trap here is confusing survivorship with data stewardship; survivorship is automated conflict resolution, while stewardship is manual intervention.

35
MCQhard

A Data Architect is designing a strategy to manage 'soft-deleted' data in a Salesforce instance. The requirement is to maintain data for seven years, but only show active records in standard views. What is the most recommended approach?

A.Use the 'Archive' custom object to store all inactive records.
B.Use an 'Active' checkbox and filtered list views to hide inactive records.
C.Create a daily batch job that deletes records older than one year.
D.Change the sharing settings to 'Private' for all inactive records.
AnswerB

This approach is the standard Salesforce practice for managing record lifecycle without deleting data. By using a simple status flag, users see only what they need, while data remains in the main table for audit and reporting purposes. It maintains full relational integrity and keeps the data easily accessible for compliance.

Why this answer

Using a combination of a status field and optimized list views/reports provides a flexible, low-impact solution. By setting an 'IsActive' flag, data remains available for historical reporting and audits without cluttering the UI for operational users. This pattern is easily scalable and avoids the complexities of moving data to external systems or utilizing custom objects, while adhering to organizational data retention policies without compromising standard Salesforce object relationships.

Exam trap

Candidates often over-engineer by suggesting archiving to external databases or custom objects, ignoring that simple list view filtering on a boolean field is the most performant, native solution.

36
Multi-Selecthard

An architect is choosing between storing a new data domain in a custom object versus using Big Objects in Salesforce. The dataset is expected to reach billions of records with append-only writes and infrequent, key-based reads. Which two characteristics make Big Objects the appropriate choice for this scenario? (Choose two.)

Select 2 answers
A.Big Objects support full DML operations including update and delete through the standard user interface.
B.Big Objects replicate data across every Salesforce instance for real-time reporting in standard reports.
C.Big Objects allow roll-up summary fields and validation rules on the object.
D.Big Objects provide an indexed, query-optimized access path for retrieving records by defined key fields.
E.Big Objects are designed to store and query massive volumes of data on the Salesforce platform.
AnswersD, E

Big Objects use an index defined by the architect over selected fields, enabling efficient retrieval by those keys even at massive scale. This matches the scenario's infrequent, key-based read pattern and is a defining advantage over custom objects, which do not scale to billions of rows with predictable query performance.

Why this answer

Big Objects exist to handle data volumes beyond what custom objects can practically store, and they add an architect-defined index that makes key-based retrieval efficient at that scale. Together, scalability and indexed async query access are the two defining reasons to choose Big Objects for this append-only, high-volume domain.

Exam trap

The trap here is assuming Big Objects behave like custom objects with full DML, roll-ups, and standard reporting, when they are intentionally limited to high-volume, indexed, mostly append-only use cases.

37
MCQmedium

A data architect is migrating 50 million records into a custom object. To ensure optimal performance and avoid hitting governor limits, which strategy should be implemented during the initial data load?

A.Enable all existing triggers and validation rules to ensure data quality immediately upon insertion.
B.Use the SOAP API to perform synchronous inserts for real-time validation feedback.
C.Leverage Bulk API 2.0 and disable non-essential automation, triggers, and validation rules.
D.Increase the batch size to the maximum allowed limit for standard REST API calls.
AnswerC

Bulk API 2.0 is designed specifically for large data sets, providing efficient background processing. Disabling triggers and complex validation rules prevents unnecessary processing overhead, reduces CPU consumption, and avoids record locking. This is the industry-standard approach for large-scale data migrations to ensure stability and maximum performance throughout the load process.

Why this answer

Utilizing the Bulk API 2.0 with serial mode for initial loads minimizes contention on shared resources and prevents record locking errors during high-volume inserts. Designing around index selectivity and disabling unnecessary automation like workflow rules or triggers before loading is crucial for performance. This approach ensures data integrity while drastically reducing the time required to complete the migration compared to standard synchronous API methods.

Exam trap

Candidates often forget to disable automation. They attempt to load 50 million records while triggers and validation rules are active, causing massive performance bottlenecks and hitting governor limits almost immediately.

38
MCQmedium

Universal Containers wants to track complex relationships between Accounts where one Account can be a subsidiary of many others, and an Account can have many parent companies. Which modeling approach should the Architect recommend?

A.Add multiple self-lookup fields on the Account object.
B.Create a Text field to store a comma-separated list of parent IDs.
C.Implement a custom junction object with two lookup fields to the Account object.
D.Use a master-detail relationship between the Account and a custom object.
AnswerC

A custom junction object provides a flexible many-to-many link between two Account records. This model supports unlimited relationships, enables the inclusion of additional attributes about the connection, and integrates perfectly with standard Salesforce features like related lists, sharing rules, and report types for comprehensive hierarchy visibility.

Why this answer

Many-to-many relationships are best handled using a junction object in Salesforce. By creating a custom object, you can link two Account records while storing additional metadata like the nature or duration of the relationship. This approach overcomes the limitations of standard lookup fields, which only support one-to-many structures, ensuring data integrity and allowing for robust reporting across complex corporate hierarchies.

Exam trap

Candidates often try to use standard parent-child lookup fields on the Account object, which cannot support a many-to-many relationship structure between two Account records.

39
MCQmedium

Which security feature should an architect use to restrict access to sensitive fields based on a user's role?

A.Organization-Wide Defaults (OWD).
B.Field-Level Security (FLS).
C.Sharing Rules.
D.Role Hierarchy.
AnswerB

Field-Level Security provides the precise control needed to restrict visibility to specific fields on an object based on a user's profile or permission set. It is the standard platform feature for ensuring that sensitive data is only accessed by users authorized to view it, maintaining robust security posture.

Why this answer

Field-Level Security (FLS) is the correct mechanism for controlling field access by profile or permission set. It ensures that users only see and interact with data relevant to their role. By applying FLS, architects can enforce strict data access policies, which is essential for compliance and maintaining the 'need-to-know' principle in complex, multi-departmental Salesforce environments where data privacy is paramount.

Exam trap

Candidates often confuse Field-Level Security with Page Layouts. While layouts hide fields from the UI, they do not restrict access at the API or report level, leading to potential data exposure.

40
MCQmedium

Universal Containers has a custom object Invoice__c with 8 million records. Users report that list views and reports filtering on the Invoice_Status__c picklist field are slow. The field is not indexed. What should the data architect do to improve query performance?

A.Create a custom index on the Invoice_Status__c field.
B.Convert the picklist to a text field and enable indexing.
C.Enable the 'Allow Reports' setting on the field.
D.Create a formula field that returns the picklist value and filter on that instead.
AnswerA

Salesforce allows a custom index on a custom field via the field definition, which can significantly improve query performance for filters and list views. For an 8 million record object, filtering on an unindexed picklist forces a full table scan, so adding a custom index is the appropriate optimization. This directly addresses the slow list views and reports without changing data model.

Why this answer

Custom indexes are the standard mechanism to improve query performance on large custom objects when filtering on a custom field. Because Invoice_Status__c is not indexed, queries perform full table scans across 8 million records. Requesting a custom index on that field allows the database to quickly locate matching rows, improving list view and report responsiveness without altering the data model or field type.

Exam trap

The trap here is assuming that enabling a field for reports or converting its type will improve performance, when only an explicit custom index changes the query execution plan.

41
Multi-Selectmedium

An organization wants to improve their data quality within Salesforce. Which THREE actions should they prioritize to establish a robust Data Governance framework?

Select 3 answers
A.Assign data owners for each critical data domain.
B.Implement automated validation rules on all critical fields.
C.Mandate that every user must manually review all records created daily.
D.Document and publish clear data definitions and standard operating procedures.
E.Delete all historical data to ensure the database starts with a clean slate.
AnswersA, B, D

Data ownership establishes accountability. By assigning owners, the organization ensures there is a clear point of contact for defining data standards, resolving disputes, and maintaining the quality of specific domains, which prevents the 'tragedy of the commons' where no one takes responsibility for data maintenance.

Why this answer

A robust governance framework requires a combination of clear ownership, standardized processes, and technical automation. Without these pillars, data quality degrades over time due to inconsistent entries and lack of accountability. Prioritizing these actions ensures that all stakeholders understand their responsibilities, data is handled consistently, and the organization has a mechanism to detect and remediate issues before they impact business operations.

Exam trap

Candidates often select only technical automation options while ignoring the necessity of assigning human data owners, incorrectly assuming that software alone can govern organizational data quality and process adherence.

42
MCQhard

An architect is designing a governance framework for Master Data Management (MDM). Which approach is best for ensuring 'Golden Record' consistency in Salesforce?

A.Using Duplicate Rules to merge records.
B.Implementing an external MDM hub.
C.Creating custom formula fields.
D.Enabling Person Accounts for all users.
AnswerB

An external hub acts as the central authority for master data. It validates, cleanses, and synchronizes the 'Golden Record' across all connected systems, including Salesforce. This approach is superior for enterprise environments where data exists across multiple platforms and needs a single, unified source of truth to maintain data integrity.

Why this answer

A Master Data Management strategy requires a 'Source of Truth' identification. By defining which system holds the primary record and using integration middleware to synchronize these records, the architect creates a consistent 'Golden Record'. This prevents data fragmentation where different systems maintain conflicting versions of the same customer, which is the core challenge MDM aims to solve.

Exam trap

Many candidates choose 'standard Salesforce duplicate rules' as the solution, forgetting that MDM requires a cross-system synchronization strategy, not just internal record matching within a single Salesforce org.

43
MCQhard

A global retailer is consolidating customer master data from Salesforce, a legacy loyalty system, and an e-commerce platform. The data architect must design a matching strategy that minimizes false positives while still identifying the same customer across sources where names are spelled differently and addresses vary. Which matching approach should the architect recommend?

A.Exact matching on the combination of first name, last name, and postal code.
B.Fuzzy matching on the full name field alone using a Levenshtein distance threshold.
C.Deterministic matching on exact email address only, treating any non-match as a distinct customer.
D.Probabilistic matching with a configured match threshold and score bands, supplemented by deterministic rules for high-confidence identifiers like loyalty number.
AnswerD

Probabilistic matching compares multiple attributes with weights and tolerates variation in names and addresses, while a threshold controls false positives. Adding deterministic rules for unique identifiers such as loyalty number captures exact matches with certainty. This hybrid approach balances precision and recall, directly addressing the need to minimize false positives while still matching records across sources with inconsistent spelling and addresses.

Why this answer

Cross-source customer matching with inconsistent names and addresses requires probabilistic matching over multiple attributes with a tuned threshold, because exact keys alone miss too many true matches. Adding deterministic rules for unique identifiers like loyalty number anchors high-confidence matches. This hybrid strategy controls false positives through scoring while preserving recall, which is exactly what the scenario demands.

Exam trap

The trap here is assuming that exact matching on one or a few fields is safer, when in cross-source consolidation it creates false negatives, and that fuzzy name matching alone is sufficient, when it actually increases false positives.

44
MCQhard

A Salesforce architect is designing a data integration strategy to synchronize 5 million records daily from an external ERP system into Salesforce. The integration must handle upserts, minimize API calls, and avoid governor limits. The external system can provide data in CSV format and supports REST APIs. Which approach should the architect recommend?

A.Use the SOAP API with a batch size of 200 records per call.
B.Use the Streaming API to push records from the ERP system into Salesforce.
C.Use the REST API with composite requests to combine multiple records per call.
D.Use the Bulk API 2.0 with parallel processing and CSV data.
AnswerD

Bulk API 2.0 is designed for high-volume data loads and can process millions of records efficiently. It accepts CSV data, supports parallel processing to speed up ingestion, and automatically handles chunking and batching. It also supports upserts using external ID fields. This minimizes API calls because a single job can process large volumes, and it avoids governor limits by using asynchronous processing. This is the optimal solution for daily synchronization of 5 million records.

Why this answer

Bulk API 2.0 is specifically built for high-volume data operations, supporting CSV input and parallel processing. It minimizes API calls by handling large batches in a single job and is designed to avoid governor limits through asynchronous processing. SOAP and REST APIs are better for smaller, real-time integrations, and the Streaming API is for outbound messaging.

Therefore, Bulk API 2.0 is the correct choice for daily synchronization of millions of records.

Exam trap

The trap here is assuming that composite REST requests or SOAP batches can handle millions of records efficiently; in reality, only Bulk API is designed for such volume without hitting API limits.

45
MCQmedium

Cloud Kicks is ingesting 20 million Account records nightly from an external ERP into Salesforce using the Bulk API 2.0. The ERP also updates existing records. During testing, the data architect observes that some records are duplicated because the external system's unique identifier is not enforced in Salesforce. Which solution should the architect implement to prevent duplicates during the nightly load?

A.Create a unique External ID field on Account and use the upsert operation with that field as the external ID.
B.Schedule a nightly batch Apex job that deletes duplicates after the load completes.
C.Enable duplicate rules with a matching rule on the ERP identifier field and set the action to Block.
D.Use a before-insert trigger to query for existing records and update them if found, otherwise insert.
AnswerA

Using an External ID field marked as unique and performing an upsert allows Salesforce to match incoming records to existing ones based on the ERP's identifier. This prevents duplicates by updating existing records instead of inserting new ones when the external ID already exists. Bulk API 2.0 supports upsert with an external ID, making it the optimal solution for large-volume, recurring loads.

Why this answer

An External ID field with the unique attribute allows Salesforce to reliably match incoming records to existing ones during an upsert. This prevents duplicates by updating matched records and inserting only new ones. Bulk API 2.0 supports upsert with an external ID, making it the most efficient and scalable solution for nightly loads of millions of records while maintaining data integrity.

Exam trap

The trap here is assuming that duplicate rules alone can prevent duplicates during bulk loads, but they are not designed for upsert matching and can cause failures instead of updates.

46
MCQmedium

A company is planning to migrate legacy customer data into Salesforce. Before the load, the data architect recommends data profiling. Why is this activity mandatory for a successful MDM project?

A.It ensures that the Salesforce data model is fully compatible with the legacy SQL schema.
B.It identifies patterns, anomalies, and quality issues within the legacy data set.
C.It increases the speed of the data load by compressing the CSV files before import.
D.It automatically maps all legacy fields to the appropriate Salesforce fields without human intervention.
AnswerB

Profiling tools analyze columns to discover frequency distributions, null counts, and data format variances. This insight allows architects to create accurate data cleansing logic, ensuring that the migrated data is clean, standardized, and ready for use in Salesforce, which is vital for maintaining the integrity of the MDM strategy.

Why this answer

Data profiling uncovers hidden quality issues, such as inconsistent formatting, missing values, or unexpected dependencies. By understanding the true state of the source data, the architect can design effective cleansing and transformation rules. Failing to profile leads to 'garbage in, garbage out,' where inaccurate data undermines the value of the new Salesforce implementation and causes significant post-migration remediation costs and user frustration.

Exam trap

Candidates often assume data profiling is an optional step that can be skipped to speed up the migration timeline, failing to realize it is essential for identifying hidden data quality risks.

47
MCQhard

When dealing with high-frequency updates on a parent object, what is the best practice to prevent lock contention?

A.Always update the parent in the child's trigger.
B.Use an asynchronous approach to aggregate parent updates.
C.Set the parent record to read-only.
D.Convert the lookup relationship to a master-detail.
AnswerB

Asynchronous processing allows updates to be queued and batched, preventing multiple concurrent transactions from fighting over the same parent row. This pattern is essential for high-frequency updates, as it serializes the parent-level changes and allows the system to process them efficiently without risking transactional row locks.

Why this answer

To minimize lock contention, avoid updating the parent record every time a child record is updated. Instead, use an asynchronous pattern like a batch job or a queueable apex to aggregate updates and perform a single update to the parent. This reduces the number of locks requested on the parent object, significantly decreasing the likelihood of row-locking errors during high-volume operations.

Exam trap

Candidates often suggest using triggers for synchronous roll-ups, which causes massive row-locking issues when multiple child records are updated simultaneously in high-volume environments.

48
MCQmedium

An architect is building a data model for a subscription service. They need to calculate the total revenue across all related 'Invoice' records for an 'Account'. Which approach should they use?

A.Write a Trigger to calculate and update the total on the Account.
B.Use a Roll-up Summary field on the Account object.
C.Create a Flow that triggers on every Invoice insert.
D.Calculate the revenue dynamically using a Formula field.
AnswerB

Roll-up summary fields provide an out-of-the-box solution for aggregating data from child records. They are performance-optimized by Salesforce and maintain data consistency without the risk of logic errors associated with custom code. This is the recommended approach for any aggregation across master-detail relationships in a standard data model.

Why this answer

Roll-up summary fields are the native, declarative way to perform aggregate calculations across master-detail relationships. They are highly efficient, automatically updated by the platform, and require no code. This ensures the revenue data is always accurate and available for reporting, dashboards, and list views, which is essential for sales teams tracking customer lifetime value.

Exam trap

Candidates often propose writing custom trigger code or asynchronous batch jobs for simple aggregations, overlooking native declarative features that provide automated and efficient roll-up calculations.

49
Multi-Selectmedium

When migrating data to Salesforce, which TWO strategies help ensure successful relationship mapping between parent and child objects?

Select 2 answers
A.Load child records first, then use Apex to link parents.
B.Use External IDs to map the relationship between parent and child.
C.Perform a two-pass load, starting with parent objects followed by child objects.
D.Query the Salesforce internal IDs after the load to update the child records.
E.Only migrate parent objects to ensure no orphan records exist.
AnswersB, C

External IDs are the standard and most reliable way to maintain relationships during migration. By mapping the legacy parent ID to an External ID field on the Salesforce parent object, the data loader can resolve the lookup field on the child record automatically without needing the internal Salesforce ID.

Why this answer

Mapping relationships is the hardest part of data migration. Using external IDs for lookups ensures that child records are connected to the correct parent regardless of the internal Salesforce ID. Furthermore, performing the load in a specific sequence (parents first, then children) eliminates 'missing parent' errors.

These two strategies ensure that relational integrity is preserved from the legacy system without relying on fragile multi-step manual work or custom post-load reconciliation scripts.

Exam trap

Candidates often attempt to load child records before parent records or rely on Salesforce internal IDs. This leads to broken relationships and failed records because the parent IDs do not yet exist.

50
MCQeasy

A Salesforce data architect is asked to recommend a native capability for detecting and merging duplicate person accounts created by multiple intake channels. Which Salesforce feature should the architect recommend as the primary tool?

A.Einstein Activity Capture to reconcile duplicate activities across person accounts.
B.Validation Rules that block record creation when the last name matches an existing contact.
C.Duplicate Rules combined with Matching Rules and the built-in merge interface.
D.A scheduled Apex batch job that compares Account.Name values using SOQL LIKE queries.
AnswerC

Duplicate Rules and Matching Rules are the native Salesforce mechanism for identifying potential duplicates at the point of creation or edit, while the merge interface consolidates surviving records. Together they provide both prevention and remediation for person account duplication without external tooling, which directly addresses the intake-channel scenario and keeps the MDM workflow inside the platform.

Why this answer

Salesforce provides declarative Matching Rules and Duplicate Rules that evaluate candidate records on create and edit, plus a standard merge interface for consolidation. Together these cover both prevention and remediation natively. Validation rules, activity capture, and custom Apex each address something adjacent but not the core duplicate detection and merge requirement, so they cannot be the primary recommendation.

Exam trap

The trap here is reaching for custom Apex or validation rules when the platform already ships declarative duplicate detection and merge capabilities designed for exactly this use case.

51
MCQmedium

An organization needs to integrate Salesforce with a legacy system that does not support modern authentication. What is the most secure architectural approach?

A.Hardcode credentials in the Apex integration code.
B.Use a Middleware/Integration layer.
C.Disable TLS 1.2 in Salesforce settings.
D.Send data via unencrypted email attachments.
AnswerB

Middleware acts as a secure intermediary, handling modern auth protocols like OAuth for Salesforce while securely communicating with legacy systems. This separation ensures that Salesforce remains secure and the legacy system is protected, providing a manageable and audited path for data exchange between the two disparate environments.

Why this answer

Using an intermediate middleware or Integration Hub allows the architect to encapsulate the modern authentication requirements of Salesforce while providing a secure proxy for the legacy system. This prevents the need to downgrade security settings within Salesforce and keeps sensitive credentials secure. It is the professional standard for bridging modern cloud platforms with older on-premises systems while maintaining strict security and compliance standards.

Exam trap

Candidates frequently suggest lowering Salesforce security settings, such as enabling weak authentication protocols, to accommodate legacy systems, which is a major security violation in architectural design exams.

52
MCQhard

Refer to the exhibit. A batch job is failing consistently with the provided error message. The job updates child records that have a Master-Detail relationship with a parent Account. What is the primary cause of this error?

A.The batch job is exceeding the heap size limit.
B.Multiple concurrent threads are updating child records pointing to the same parent.
C.The sharing rules on the Account object are too complex.
D.The batch job is processing records in an incorrect order.
AnswerB

Master-Detail relationships enforce a lock on the parent record whenever a child is modified. When multiple batches process child records belonging to the same parent record at the same time, they compete for this parent-level lock, resulting in the UNABLE_TO_LOCK_ROW exception due to contention.

Why this answer

This error occurs due to row-level locking when multiple threads attempt to update records associated with the same parent simultaneously. In Salesforce, updating a child record in a Master-Detail relationship implicitly locks the parent record. When high-volume parallel batches attempt to update children of the same parent, they contend for this single lock, causing transaction failures.

Architecturally, you must serialize these updates or reduce parallel threads to avoid collision.

Exam trap

Candidates often assume that Apex batch limits or query timeouts caused the failure, missing how parent-child Master-Detail record locks trigger row contention under high-concurrency threading.

53
MCQmedium

A company is migrating legacy account data into Salesforce. They need to map legacy system primary keys to a new custom field to support future integrations. What is the recommended approach to ensure this field supports efficient data retrieval?

A.Create a standard text field and create a custom index request through Salesforce Support.
B.Use a custom field without the External ID attribute and perform lookups via SOQL.
C.Configure the custom field as an External ID with unique constraints enabled.
D.Store the legacy primary key in the standard Name field for each account.
AnswerC

Configuring the field as an External ID automatically creates an index on the database table. This allows for rapid record matching during upsert operations, which is essential for data migrations. Additionally, the unique constraint ensures data integrity by preventing duplicate source records from being inserted into the destination object.

Why this answer

Marking the custom external ID field as 'Unique' and 'External ID' ensures Salesforce automatically indexes the column. This indexing is critical for upsert operations, as it allows the platform to quickly match existing records without performing a full table scan. Proper schema design during migration prevents long-term performance degradation and simplifies the maintenance of relational integrity across interconnected systems.

Exam trap

Candidates often select standard text fields or forget to enable unique constraints, assuming that simply checking the External ID box is enough to guarantee performance and prevent duplicates during data integration upsert operations.

54
MCQhard

What is the consequence of having a 'non-selective' query running on an object with 50 million records?

A.The query will fail immediately with a compilation error.
B.The query will run faster because it ignores indexes.
C.The query will be blocked or timed out by the platform.
D.The query will automatically create an index for the fields used.
AnswerC

To protect system resources, Salesforce restricts queries that are non-selective on large datasets. If a query cannot be optimized via an index, it is likely to be blocked or time out, ensuring that the system maintains consistent performance for all users in the multi-tenant environment.

Why this answer

Non-selective queries force the system to perform a full table scan. In an environment with millions of records, this consumes massive resources, often leading to query timeouts, increased CPU usage, and potential platform instability. Preventing this is the primary goal of data architecture, as it protects the multi-tenant environment and ensures that individual system processes do not consume resources allocated for other tenants or internal tasks.

Exam trap

Candidates often assume non-selective queries will simply run slowly. They fail to realize that the Salesforce platform proactively blocks or times out these queries to protect the multi-tenant environment.

55
Multi-Selectmedium

A data architect is designing a data governance program for a Salesforce MDM implementation that spans Sales Cloud, Service Cloud, and an external data warehouse. Which two capabilities are essential to include in the governance model? (Choose two.)

Select 2 answers
A.Enabling Lightning Experience for all users to standardize the interface used for master data entry.
B.A defined data stewardship role with authority to resolve match exceptions and approve merges.
C.A nightly full refresh of all Salesforce records into the data warehouse using Bulk API 2.0.
D.Configuring field-level security so that only system administrators can edit any master data attribute.
E.A documented set of data quality dimensions and thresholds, such as completeness and uniqueness targets per attribute.
AnswersB, E

Stewardship is the human control that keeps an MDM hub accurate over time. Automated matching will always produce borderline pairs that require judgment, and without an accountable steward those exceptions accumulate as unresolved duplicates or incorrect merges. Naming the role, granting merge authority, and documenting escalation paths is therefore a foundational governance capability rather than an optional nicety.

Why this answer

A functioning MDM governance model needs both accountability and measurement. Stewardship assigns human ownership for resolving exceptions and approving merges, while defined quality dimensions and thresholds give the program objective success criteria. Integration patterns, UI standardization, and administrator-only editing are technical or operational choices that do not by themselves establish governance over master data.

Exam trap

The trap here is selecting technical integration or security configuration items as governance capabilities when governance is fundamentally about ownership, accountability, and measurable quality targets.

56
Multi-Selectmedium

Which TWO steps are critical when preparing to migrate sensitive PII (Personally Identifiable Information) into Salesforce?

Select 2 answers
A.Perform data masking on source datasets for sandboxes.
B.Use the Bulk API for all PII data transfers.
C.Enable Field-Level Security to restrict access to sensitive fields.
D.Delete all audit logs after the migration is complete.
E.Export all PII to unencrypted CSV files for mapping.
AnswersA, C

Data masking is a mandatory security practice for non-production environments to prevent the exposure of PII. By sanitizing the data before it reaches the sandbox, the risk of data leakage is minimized while still allowing developers to test against realistic, but safe, dataset structures.

Why this answer

When migrating PII, Data Architects must prioritize compliance and data protection. Data masking ensures that non-production environments do not contain real sensitive data, while Field-Level Security (FLS) ensures that only authorized users can access the data within production. These steps are essential to maintaining regulatory compliance, such as GDPR or HIPAA, throughout the lifecycle of the data migration project.

Exam trap

Test-takers often focus solely on migration speed and technical mapping, forgetting compliance requirements like PII masking in non-production environments and proper field-level security.

57
MCQmedium

A Salesforce architect is migrating 1 million Case records with related Case Comments from a legacy system. The legacy system stores comments in a separate table linked by a legacy Case ID. The architect plans to use the Bulk API to load Cases first, then load Case Comments in a second pass. During the test, the architect realizes that the legacy Case ID is not stored in Salesforce after the first load, making it impossible to link comments to the correct Cases. What should the architect have done to enable this relationship?

A.Create a custom external ID field on the Case object to store the legacy Case ID, populate it during the Case load, and then use that field to relate Case Comments during the second load.
B.Export the Salesforce Case IDs after the first load, map them back to the legacy Case IDs in the source system, and then load Comments with the new Salesforce IDs.
C.Load Cases and Case Comments simultaneously using a single Bulk API job with nested JSON to preserve relationships.
D.Use the Salesforce Data Loader's 'Insert' operation for Cases and then use 'Update' for Comments, relying on Salesforce's automatic relationship matching.
AnswerA

Storing the legacy Case ID in an external ID field on Case allows the architect to reference it when loading Case Comments. The Bulk API can use the external ID to look up the parent Case record and establish the relationship. This is a standard practice for migrating related records in multiple passes, ensuring referential integrity without manual intervention.

Why this answer

To link Case Comments to Cases in a two-pass migration, the architect must store the legacy Case ID on the Case record as an external ID. This allows the second load to use that external ID to look up the parent Case and correctly associate comments. Without it, the relationship cannot be established, leading to orphaned comments.

Exam trap

The trap here is thinking that Salesforce or Data Loader can automatically match related records without a stored external ID, when the key must be explicitly persisted for lookups.

58
MCQeasy

An administrator wants to convert an existing custom lookup relationship on the 'Project' object pointing to the 'Client' object into a master-detail relationship. Which precondition must be satisfied before this conversion can be completed successfully?

A.All existing child records must have a non-null value populated in the lookup field.
B.The parent object must have fewer than 10,000 child records associated with it.
C.The organization-wide default for the child object must be set to Public Read/Write.
D.All custom reporting types referencing the objects must be temporarily deleted prior to conversion.
AnswerA

Master-detail relationships require every child record to reference a parent, so the lookup field must be populated on all existing Project records before conversion. Null values would leave orphaned children, which Salesforce rejects during the conversion process.

Why this answer

Converting a lookup relationship to a master-detail relationship requires that every existing child record currently in the database possesses a populated, valid value for the lookup field. If any orphan records exist where the lookup field is blank, the database cannot establish the mandatory parent-child integrity required by master-detail semantics, causing the conversion operation to fail.

Exam trap

Candidates often believe Salesforce will automatically populate missing values or ignore empty lookup fields during a conversion to a master-detail relationship.

59
MCQmedium

A Salesforce architect is migrating 2 million legacy Contact records into a new Salesforce org. The legacy system does not have email addresses for 40% of the contacts, but Salesforce's standard Email field is not required. During a test load of 10,000 records using the Bulk API in serial mode, the load succeeds but the architect notices that duplicate contacts are being created because the legacy system reused a 'legacy_id__c' external ID field for different contacts across regions. What should the architect do to prevent duplicate creation during the full migration?

A.Create a new unique external ID field by concatenating region and legacy_id, populate it in the source data, and use it as the external ID for upsert operations.
B.Set the legacy_id__c field as unique in Salesforce and retry the load; Salesforce will automatically reject duplicate external IDs.
C.Enable the 'Prevent Duplicates' setting on the Contact object and use Data Loader's deduplication feature before loading.
D.Use the Bulk API in parallel mode with a batch size of 1 to ensure each record is processed individually and duplicates are avoided.
AnswerA

The duplicate creation stems from non-unique external IDs. By creating a composite external ID that combines region and legacy_id, the architect ensures each record has a unique identifier for upsert. This allows the Bulk API to correctly match existing records and prevent duplicates, while also preserving the original legacy ID for reference in a separate field.

Why this answer

The duplicate creation is due to the external ID field not being unique across regions. The correct solution is to create a new composite external ID that combines region and legacy ID, ensuring uniqueness. This allows upsert operations to correctly identify records and prevent duplicates, while the original legacy ID can be retained for audit purposes.

Exam trap

The trap here is assuming that enabling duplicate rules or setting a unique constraint on the existing non-unique field will solve the problem, when the source data itself contains duplicates that must be transformed.

60
MCQmedium

Refer to the exhibit. An organization uses the provided JSON policy to manage PII fields. As a Data Architect, what is the primary governance risk if this policy is not integrated with Salesforce's internal security controls?

A.Increased latency in database query execution.
B.Inconsistent enforcement of privacy compliance.
C.Loss of data due to automated purging.
D.Increased storage consumption in the Org.
AnswerB

When documentation deviates from actual Salesforce security settings, compliance audits will fail. The system becomes vulnerable because the policy exists in a silo, and the technical implementation fails to match the required controls. This inconsistency creates a false sense of security, exposing the organization to legal and regulatory penalties.

Why this answer

The primary risk is the discrepancy between the documented policy and the actual system implementation. If the JSON policy dictates partial masking but the Salesforce configuration allows full view access via profiles or permission sets, the policy becomes ineffective. This gap compromises compliance, as the organization cannot guarantee that data protection rules are actually enforced at the application layer where users interact.

Exam trap

Candidates often select system performance or storage issues as the primary risk, overlooking critical compliance and regulatory failures caused by policy-to-configuration alignment gaps.

61
MCQhard

A data architect at a global manufacturer is defining the match strategy for a new MDM hub. Legal names vary widely across regions, so the architect needs a technique that will correctly link records such as 'Acme Corp.' and 'Acme Corporation Ltd.' even though the strings differ. Which matching approach best satisfies this requirement?

A.Deterministic matching on the exact normalized Account Name field using a case-insensitive comparison.
B.Probabilistic (fuzzy) matching using token-based similarity scoring with a configurable match threshold.
C.Blocking on the first three characters of the account name followed by exact comparison of the remaining characters.
D.Matching on the external ERP customer number only, ignoring name fields entirely during the match.
AnswerB

Probabilistic matching compares attributes using similarity algorithms such as Jaro-Winkler or token-based edit distance, then scores candidate pairs against a threshold. This is designed precisely for cases where legal names are semantically the same but lexically different, as with 'Acme Corp.' versus 'Acme Corporation Ltd.' It trades some precision for much higher recall, which matches the stated requirement.

Why this answer

The requirement is about linking records whose names differ lexically but refer to the same real-world entity. Deterministic rules and identifier-only matching both fail when the input strings vary, while blocking is a performance layer rather than a matching algorithm. Probabilistic scoring with a tunable threshold is the standard MDM technique for this kind of fuzzy corporate-name reconciliation.

Exam trap

The trap here is confusing blocking with matching and assuming that any name-based rule will handle legal-name variants when only a similarity-based scorer can.

62
MCQhard

Universal Containers has a custom object Order__c with 10 million records. The object has a lookup relationship to Account. A data architect needs to create a report that shows the total order amount per account for the last fiscal year. The report must be accessible to 500 users and should load in under 10 seconds. Which solution should the architect recommend?

A.Create a report using a custom report type with a cross filter on Order__c and summarize by Account.
B.Create a roll-up summary field on Account that sums Order__c amounts for the last fiscal year, and use it in a report.
C.Create a Lightning Web Component that queries Order__c using SOQL with GROUP BY Account and displays the results.
D.Create a custom object to store aggregated order totals per account, populate it nightly using Batch Apex, and build a report on that object.
AnswerD

Pre-aggregating data into a custom object using Batch Apex ensures fast report performance because the report queries a small, indexed dataset. This approach scales well for many users and large data volumes. It also allows the aggregation logic to be customized for fiscal year calculations, meeting the performance and business requirements.

Why this answer

Pre-aggregating data into a custom object via Batch Apex provides a performant and scalable solution for reporting on large data volumes. The report then queries a small, indexed dataset, ensuring fast load times for many users. This pattern is commonly recommended for large-scale aggregations that cannot be efficiently handled by standard reports or real-time queries.

Exam trap

The trap here is assuming that standard reporting features like cross filters or roll-up summaries can handle millions of records efficiently, but they often hit limits or require master-detail relationships.

63
MCQmedium

When dealing with LDV, what is the primary benefit of using External Objects via Salesforce Connect instead of standard Salesforce tables?

A.External objects are automatically indexed by Salesforce.
B.It keeps the data within Salesforce for better reporting performance.
C.It avoids storing massive amounts of data in the Salesforce database.
D.External objects allow for full Apex trigger functionality.
AnswerC

By keeping high-volume, low-access data in an external system, you preserve Salesforce storage and avoid the performance overhead of managing massive tables. This architecture ensures that core business objects remain performant and responsive, while secondary data remains available through a seamless, integrated user experience.

Why this answer

External Objects allow organizations to view and interact with data stored in external systems without importing it into Salesforce. This strategy keeps the Salesforce database lean, avoiding the storage limits and performance degradation associated with managing hundreds of millions of records locally. It is the architectural standard for offloading non-critical, high-volume historical data while keeping it accessible within the platform's user interface.

Exam trap

Candidates often assume External Objects are used for performance speed, rather than storage management. They fail to realize that external calls actually add latency compared to querying local indexed data.

64
MCQmedium

Universal Containers needs to establish a formalized Data Governance framework to manage customer data across multiple Salesforce orgs and external systems. Which foundational element must the Data Governance board define first to ensure enterprise-wide alignment and accountability?

A.Implement exact-match duplicate rules across all Salesforce objects.
B.Define data owners, stewards, and organizational policies for critical data domains.
C.Configure Platform Encryption for all Personally Identifiable Information fields.
D.Establish an automated archiving job for records older than seven years.
AnswerB

Defining data owners, stewards and policies for critical data domains establishes the accountability structure the board needs before any technical controls. Without named owners and stewards per domain, cross-org alignment and enforcement cannot be assigned, so this foundational element must precede tooling or standards decisions.

Why this answer

Establishing data ownership and stewardship is the critical first step in any Data Governance framework. Without clearly defined owners responsible for data quality, definitions, and lifecycle policies across disparate systems, subsequent technical implementations like Master Data Management or deduplication rules lack institutional direction and authoritative accountability.

Exam trap

Candidates frequently jump to technical solutions like 'implementing a data warehouse' or 'data cleansing' before establishing the human governance structure of who actually owns and manages the data.

65
MCQhard

A company requires real-time reporting on millions of records. Which architecture enables this while minimizing impact on production database performance?

A.Import all data into custom Salesforce objects.
B.Use Salesforce Connect for virtualized access.
C.Copy all data into the Salesforce Big Object.
D.Use platform events to stream data records.
AnswerB

Salesforce Connect allows users to view external data in real-time without physically storing it in Salesforce. This offloads the data footprint to the external system, maintaining Salesforce performance and avoiding storage limits. It is an excellent architectural pattern for large datasets that need to be queried but not necessarily updated.

Why this answer

Salesforce Connect with External Objects is the ideal solution for large-scale data that shouldn't live directly inside Salesforce. It fetches data in real-time from external sources without consuming Salesforce storage or affecting local database performance. This pattern is critical for architects managing massive datasets, as it keeps the core Salesforce CRM instance light and performant while still providing users with seamless access to historical or external financial data.

Exam trap

Candidates often suggest custom code or batch processing to move historical data into Salesforce, ignoring the performance impact on local storage and the efficiency of virtualized external data access.

66
MCQmedium

Universal Containers requires a data model where a Child object must have exactly one Parent object that is mandatory. If the Parent is deleted, the Child must be deleted. Which relationship should the Architect select?

A.Lookup relationship
B.Self-relationship
C.Master-Detail relationship
D.External lookup relationship
AnswerC

Master-Detail relationships enforce a strict ownership model where the child record is dependent on the parent. The relationship field is mandatory, and the system automatically enforces a cascade delete when the parent is deleted, directly satisfying the requirement for mandatory association and lifecycle management.

Why this answer

A Master-Detail relationship creates a tight coupling where the child's lifecycle is strictly governed by the parent. Deleting the parent automatically triggers a cascade delete of all associated children. This is essential for maintaining referential integrity in parent-child hierarchies where the child cannot exist in isolation, ensuring the database remains clean and logically consistent without requiring custom automation for cleanup tasks.

Exam trap

Candidates might choose a Lookup relationship with required configuration, missing that only Master-Detail relationships guarantee strict lifecycle dependency and cascade deletion.

67
MCQhard

A data architect is implementing a data governance strategy for a Salesforce org with multiple business units. The requirement is to ensure that data stewards can define and enforce data quality rules centrally, and that users are alerted when they attempt to save records that violate these rules. Which feature should be used?

A.Workflow Rules
B.Lightning Data Quality Dashboard
C.Validation Rules
D.Duplicate Rules
AnswerC

Validation Rules allow administrators to define data quality criteria that are enforced when a user saves a record. They can display error messages to alert users of violations. They are centrally managed and apply to all users, making them ideal for enforcing data governance rules across business units. They directly meet the requirement of alerting users upon save.

Why this answer

Validation Rules are the standard Salesforce feature to enforce data quality at the point of record save. They can be defined centrally and apply to all users, ensuring consistent data governance. They alert users with custom error messages when rules are violated.

Duplicate Rules are for duplicates only, Workflow Rules do not block saves, and the Data Quality Dashboard is not a real-time enforcement tool.

Exam trap

The trap here is assuming that any automation tool like Workflow Rules can enforce data quality, when only Validation Rules can block saves and alert users at the point of entry.

68
MCQeasy

A nonprofit needs to record donations. Each donation is made by exactly one donor Account, and the organization wants a Donor's lifetime giving total displayed on the Account record itself, updated automatically whenever a Donation is created or edited. What should the architect configure?

A.A trigger on Donation that posts the Amount to a custom field on Account using an aggregate SOQL query.
B.A scheduled Flow that runs nightly to query all Donations and update a numeric field on each Account.
C.A cross-object formula field on Account that sums the Amount field of all related Donation records.
D.A roll-up summary field on Account that sums the Amount field of related Donation records, with Donation in a Master-Detail relationship to Account.
AnswerD

Roll-up summary fields are designed precisely for this scenario: they aggregate child records onto the parent. Because Donation is the detail and Account is the master, the Account can display a SUM of the Donation Amount field. The value recalculates automatically when donations are created, edited, or deleted, requiring no code.

Why this answer

Roll-up summary fields exist to aggregate values from detail records onto a master record. With Donation in a Master-Detail relationship to Account, the architect can create a roll-up summary field on Account that sums the Amount field of all related donations. The total updates automatically on insert, update, and delete, with no code and no scheduled job.

Exam trap

The trap here is confusing a cross-object formula field, which can only reference one related parent record, with a roll-up summary field, which aggregates many child records.

69
MCQmedium

Which of the following is an example of 'Probabilistic Matching' in the context of an MDM solution?

A.Matching two records because they share the exact same Social Security Number.
B.Comparing customer names and addresses using fuzzy logic to generate a match score.
C.Using a strict SQL join condition to link Account records with Contact records.
D.Rejecting any contact record that lacks a valid email address field.
AnswerB

Probabilistic matching evaluates multiple fields for similarity, assigning weights and scores to reach a conclusion. By using fuzzy logic to account for variations like 'John Smith' versus 'Jon Smyth' at similar addresses, the system can identify matches that would otherwise be ignored by rigid, exact-match requirements.

Why this answer

Probabilistic matching uses statistical algorithms to determine the likelihood that two records refer to the same entity based on similarity rather than exact matches. This is vital in MDM because data across systems is rarely perfectly identical due to typos, variations, or formatting differences. Using algorithms allows the system to identify matches that deterministic logic would miss, significantly improving the completeness and accuracy of the resulting golden record.

Exam trap

Candidates often confuse probabilistic matching with deterministic matching, selecting options that describe exact field-level comparisons instead of the fuzzy logic and statistical scoring that define probabilistic approaches.

70
MCQeasy

A healthcare organization is building a data governance program for its Salesforce org. The compliance officer wants a documented record of who owns each data domain, what the approved data definitions are, and how changes to those definitions are approved. Which artifact should the architect establish as the authoritative source for this information?

A.A Salesforce report listing all custom objects and their record counts.
B.A set of validation rules that enforce data entry standards on key objects.
C.An integration architecture diagram showing how data flows between Salesforce and external systems.
D.A data governance charter and data dictionary that define domain ownership, approved definitions, and the change-approval workflow.
AnswerD

A governance charter establishes the program's structure, roles, and decision rights, while a data dictionary documents approved definitions and ownership for each data domain. Together they provide the authoritative, auditable record the compliance officer requires, including how definition changes are proposed, reviewed, and approved.

Why this answer

The compliance officer is asking for documented accountability, agreed definitions, and a controlled change process. A governance charter defines the program's roles, decision rights, and approval paths, while a data dictionary records the approved definition and owner for each data domain. Together they form the authoritative governance reference that can be audited and maintained over time.

Exam trap

The trap here is selecting a technical artifact such as a report or validation rule that demonstrates data controls, when the requirement is for documented ownership, definitions, and approval authority.

71
MCQhard

Refer to the exhibit. As a Data Architect, what is the governance concern if this 'VAL-099' rule remains inactive in production?

A.Increased system storage limits.
B.Inability to run system reports.
C.Degradation of data quality over time.
D.Increased Apex trigger execution time.
AnswerC

When an active validation rule is turned off, the system stops checking for input errors. This allows bad data to accumulate, which degrades the quality of the dataset. Over time, this makes it harder to use that data for business intelligence, marketing campaigns, or sales forecasting, leading to untrustworthy results.

Why this answer

The primary concern is the uncontrolled entry of corrupt data. If a governance rule is defined but not active, the system lacks the technical enforcement to maintain quality standards. This leads to 'data rot' where the system collects invalid entries, which will eventually break downstream integrations, analytics reporting, and business processes, causing significant effort for data cleansing at a later date.

Exam trap

Candidates often focus on the immediate technical error, such as 'API failure,' rather than the long-term governance impact of 'data rot' and the resulting degradation of data quality.

72
MCQmedium

You are migrating records that contain multi-select picklist values. What is the key consideration for mapping this data from a legacy system?

A.The values must be separated by commas in the source file.
B.Each value must be imported as a separate child record.
C.The values must be formatted as a semi-colon separated string.
D.The picklist must be converted to a custom object before loading.
AnswerC

Salesforce requires multi-select picklist values to be provided as a string with each value separated by a semi-colon. This is the only format that the Salesforce API accepts for this field type. Transforming the source data to this specific format is a mandatory step in the data migration process.

Why this answer

Multi-select picklists are stored as semi-colon separated strings in Salesforce. When migrating, the source data must be transformed to match this specific delimiter format. Failure to use the exact semi-colon separator, or including values that are not currently defined in the Salesforce picklist configuration, will result in import errors.

It is essential to sanitize and format these values in the ETL layer before the load.

Exam trap

Test-takers frequently assume multi-select picklists are mapped using comma-separated lists or arrays, ignoring Salesforce's strict requirement for a semi-colon delimiter format.

73
MCQmedium

What is a 'Selective Query' in the context of Salesforce LDV?

A.A query that uses exactly five fields in the WHERE clause.
B.A query that returns fewer than 100 records total.
C.A query that utilizes indexed fields to filter within defined thresholds.
D.A query that uses only SOQL and no Apex logic.
AnswerC

A selective query is one that effectively uses indexes to limit the search space to a small percentage of total records. By staying within the platform's selectivity thresholds, the query optimizer can retrieve results quickly without the performance penalty associated with full table scans on massive data sets.

Why this answer

A selective query is one that hits an index and returns a small, efficient subset of records. Salesforce imposes selectivity thresholds; if a query exceeds these (e.g., usually 30% of the first million records), it is considered non-selective. Ensuring queries are selective is the single most important factor in maintaining performance for large datasets, as it allows the database to avoid exhaustive scanning of rows.

Exam trap

Candidates often confuse 'selective' with 'simple' queries. They fail to understand that selectivity is a mathematical threshold based on the percentage of records returned, not just query complexity.

74
MCQmedium

Northern Trail Outfitters uses Salesforce as its system of entry for accounts and has a legacy ERP that remains the system of record for billing. The data architect must ensure that when an account's billing address changes in Salesforce, the ERP is updated within near real time, and when the ERP changes a credit limit, Salesforce reflects it within the same window. No middleware is currently in place. Which approach should the architect recommend?

A.Create a duplicate of each account in the ERP as an external object and use Apex triggers to write to both systems on every account save.
B.Schedule a nightly Apex batch job that exports updated accounts to CSV and the ERP imports them, with a reciprocal nightly import.
C.Implement a bidirectional integration using Platform Events and Change Data Capture, with an external subscriber handling ERP writes and publishing ERP changes back as Platform Events.
D.Configure Salesforce Connect with an OData adapter so Salesforce reads ERP billing records live and the ERP reads Salesforce accounts live.
AnswerC

Change Data Capture publishes record changes from Salesforce in near real time, and Platform Events carry ERP-originated changes back into Salesforce. A subscriber can apply the ERP changes to accounts and publish credit limit updates as events, satisfying bidirectionality and low latency without middleware. This is the native Salesforce pattern for event-driven master data synchronization across systems.

Why this answer

Near real time bidirectional synchronization between Salesforce and an ERP without middleware is best achieved with Salesforce's native eventing: Change Data Capture streams Salesforce record changes, and Platform Events can carry ERP changes back. This decouples systems, preserves throughput, and supports the required latency. Batch or synchronous dual-write approaches either miss the latency target or create fragile coupling that undermines master data consistency.

Exam trap

The trap here is assuming that Salesforce Connect external objects or a nightly batch job can satisfy bidirectional near real time synchronization, when external objects are read-only for external data and batch jobs cannot meet the latency requirement.

75
MCQmedium

When designing a large-scale data architecture, what is the primary consideration regarding the 'Skinny Table' feature in Salesforce?

A.They automatically update whenever a record is modified by a user.
B.They are best used to improve write performance for high-volume transactions.
C.They should be requested through Salesforce Support for specific performance needs.
D.They allow for unlimited columns and data types in the table structure.
AnswerC

Skinny tables are not configurable by administrators in the UI. They must be requested via a support case. This process ensures that Salesforce engineering can evaluate the impact on the database and ensure the table is optimized correctly for the specific reporting requirements of the organization's data model.

Why this answer

Skinny tables are used to optimize read-only performance for reports and list views by denormalizing data into a single table. They are highly effective for performance, but they are maintenance-intensive because they require Salesforce Support to enable and update. They are not a general-purpose solution for transactional data; rather, they are a performance-tuning tool meant for specific, high-volume reporting needs that cannot be satisfied by standard indexes alone.

Exam trap

Candidates mistakenly believe they can create 'Skinny Tables' themselves in the setup menu, failing to realize this feature is hidden and requires a specific request to Salesforce Support.

Page 1 of 3

Page 2

All pages