Courseiva

Salesforce Certified Data Architecture and Management Designer (SF-Data-Arch) — Questions 76–150

222 questions total · 3pages · All types, answers revealed

Page 1

Page 2 of 3

Page 3
76
MCQhard

Northern Trail Outfitters has a Salesforce org where the Account object contains 40 custom fields, many of which were created by different teams over five years. No one is certain which fields are actively used, which are populated by integrations, and which are safe to remove. The CIO asks the data architect to establish ongoing visibility into field usage and data provenance. Which approach best addresses this governance gap?

A.Enable Field Audit Trail on all 40 custom fields and rely on the audit history as the primary source of field usage documentation.
B.Restrict field creation permissions to system administrators and freeze the Account object schema to prevent further sprawl.
C.Run a one-time report on field population rates and delete all fields with less than 10% fill rate.
D.Create a data dictionary and metadata inventory that documents each field's purpose, source system, and owning team, and establish a review cadence to keep it current.
AnswerD

A maintained data dictionary and metadata inventory directly address the gap: they record what each field means, where its data originates, and who owns it. Coupling the artifact with a recurring review cadence ensures the information stays accurate as teams and integrations change. This gives the CIO durable visibility and a basis for future deprecation decisions.

Why this answer

The core issue is a lack of documented knowledge about existing fields and their origins. A data dictionary paired with a metadata inventory captures purpose, source, and ownership for each field, and a review cadence keeps it trustworthy over time. This combination gives leadership the ongoing visibility needed to make informed deprecation and consolidation decisions rather than guessing from a single usage metric.

Exam trap

The trap here is treating a one-time usage metric as sufficient evidence for permanent schema decisions, when governance requires documented provenance and a repeatable review process.

77
Multi-Selecthard

A Salesforce org is implementing a data governance program to manage data quality for a large volume of customer records. The architect must recommend tools and features to monitor and improve data quality on an ongoing basis. Which two Salesforce features should be used to proactively identify and address data quality issues? (Choose two.)

Select 2 answers
A.Duplicate Management with Matching Rules and Duplicate Rules
B.Data Loader with batch size optimization
C.Salesforce Optimizer
D.Validation Rules to enforce data quality at entry
E.Reports and Dashboards with custom report types
AnswersA, D

Duplicate Management allows you to define matching rules to identify duplicate records and duplicate rules to control what happens when duplicates are detected (e.g., alert, block, or allow). This proactively identifies and prevents duplicate data entry, which is a key aspect of data quality. It can be configured for standard and custom objects, and runs in real-time on record creation and edit, as well as via batch jobs for existing data.

Why this answer

Duplicate Management and Validation Rules are proactive features that identify and prevent data quality issues. Duplicate Management detects and manages duplicate records, while Validation Rules enforce data quality at entry by blocking invalid saves. Data Loader, Salesforce Optimizer, and Reports/Dashboards are not proactive data quality tools; they serve other purposes like data loading, org optimization, or monitoring.

Exam trap

The trap here is assuming that reporting tools or data loading tools proactively enforce data quality, when they are passive or manual.

78
MCQmedium

A financial services firm stores portfolio holdings in a custom object with 4 million records. Analysts frequently filter holdings by Account and by a calculated risk score derived from multiple fields on the holding record. The architect wants to minimize query time when users filter by risk score in list views and reports. Which approach should the architect take?

A.Create a formula field for the risk score and rely on Salesforce to index it automatically for filtering.
B.Create a numeric custom field to store the risk score, populate it via a before-save record-triggered flow, and request a custom index on that field.
C.Create a roll-up summary field on a parent object to calculate the risk score and filter on the parent.
D.Enable Big Object storage for the holdings and query it with SOQL to improve filter performance.
AnswerB

A stored numeric field can be indexed, and Salesforce supports custom indexes on external IDs or fields requested through Support. Populating it with a before-save flow keeps it current without consuming additional DML, so filtering by risk score in list views and reports uses the index and performs well at scale.

Why this answer

To filter efficiently on a computed value at scale, the value must be stored in a field that can carry a custom index. A before-save record-triggered flow updates the stored field without extra DML, and a custom index lets list views and reports filter on it quickly. Formula fields, roll-up summaries, and Big Objects cannot provide indexed filtering on a per-record calculated score.

Exam trap

The trap here is believing formula fields are indexed for filtering, when in fact only stored fields with a custom index can accelerate queries at scale.

79
MCQeasy

A retail company wants to ensure that customer email addresses are consistently formatted and that invalid values are rejected at the point of entry across web, mobile, and integration channels. The data governance lead asks the architect which Salesforce feature provides a single, reusable definition of the validation that all channels can share. Which feature should the architect recommend?

A.A validation rule on the Contact object using a REGEX function to enforce a standard email pattern.
B.A duplicate rule that matches contacts with similar email addresses.
C.A formula field that concatenates the email domain with a fixed suffix.
D.A record-triggered flow that sends a notification to the data steward when the email is malformed.
AnswerA

Validation rules evaluate on insert and update regardless of channel, so the same REGEX-based rule applies to web, mobile, and integration writes. It provides one reusable definition of the format requirement and blocks invalid values at the point of entry. This directly satisfies the governance lead's request for consistency across all channels.

Why this answer

Validation rules are evaluated on every insert and update no matter which channel performs the write, so a REGEX-based rule gives one reusable definition of the email format standard. It prevents invalid values from being saved, which is what the governance policy requires. Formula fields, duplicate rules, and notification flows either do not validate or act only after the fact.

Exam trap

The trap here is confusing detection or duplication controls with prevention, when the requirement is a shared validation definition that rejects invalid values at entry.

80
MCQmedium

A multinational manufacturer runs a single Salesforce org for sales and service. Its governance team wants to enforce that all new custom objects created in production carry a business owner, a retention classification, and a data sensitivity label. Administrators frequently create objects ad hoc, and the team wants an automated check rather than a manual review. Which approach should the data architect recommend?

A.Enable Field Audit Trail and Field History Tracking on all custom objects to capture who created each object.
B.Build a custom validation rule on each custom object that checks the owner field is populated on records.
C.Use Metadata API and Apex Metadata API to scan object definitions against a governance checklist and block deployment when required attributes are missing.
D.Require every administrator to submit a change request through a ServiceNow workflow before creating a custom object.
AnswerC

Metadata API exposes object definitions, including custom fields, descriptions, and custom metadata values, so a governance service can programmatically evaluate every object against the required attributes. Apex Metadata API allows the check to run inside Salesforce and to fail a deployment or raise an alert. This gives the automated, metadata-level enforcement the governance team asked for.

Why this answer

Metadata-level governance requires inspecting object definitions, which only metadata-aware tooling can do. A programmatic check using Metadata API and Apex Metadata API lets the team compare each object against required attributes such as owner, retention classification, and sensitivity label, and it can fail the deployment automatically. Manual workflows and record-level rules cannot assert anything about the object definition itself.

Exam trap

The trap here is assuming record-level validation rules or change-request workflows can enforce metadata standards, when only metadata-aware tooling can inspect object definitions.

81
MCQmedium

Which document is essential to provide to a regulator to prove that the Data Governance program is active and effective?

A.Data Governance Charter.
B.List of user profiles.
C.Salesforce Developer Guide.
D.Individual user training materials.
AnswerA

The charter is the formal document that defines the governance program's scope, mission, and organizational authority. It is the primary artifact that regulators look for to verify that an organization has officially committed to structured data management and compliance, providing the necessary high-level evidence of a sound governance framework.

Why this answer

The Data Governance Charter is the foundational document that outlines the authority, scope, and objectives of the program. It provides the official mandate, showing that the organization has formally recognized the importance of data governance. For regulators, this is the first proof that the organization has a structured approach to managing data quality, security, and compliance in its systems.

Exam trap

Candidates often confuse strategic framework documents like the Data Governance Charter with operational artifacts like data dictionaries or technical architecture diagrams when proving program compliance to regulators.

82
Multi-Selectmedium

When planning a large data migration, which TWO tasks are essential to perform before the data load? (Choose two)

Select 2 answers
A.Deactivate all triggers, workflow rules, and flows that could interfere with the load.
B.Increase the Salesforce API limit for the specific migration user.
C.Establish a field mapping document that reconciles legacy data to Salesforce objects.
D.Delete all existing records in the environment to ensure a fresh start.
E.Enable the 'Parallelize' setting on all user profiles.
AnswersA, C

Disabling automation is essential to prevent recursive triggers or unintended updates that degrade performance. When loading large volumes, automated processes can consume CPU and DML limits, causing the load to fail. Clearing the path allows the raw data to be inserted efficiently before re-enabling business logic later.

Why this answer

Preparing the environment by deactivating automation and establishing a clean mapping strategy are foundational steps. Deactivating automation prevents unintended side effects and performance bottlenecks during the load. A robust mapping strategy, including data cleansing and validation, ensures the migration follows business requirements.

These steps minimize downtime and reduce the risk of data corruption, ensuring the system remains performant and the data remains high quality post-migration.

Exam trap

Candidates often forget to deactivate asynchronous automation like triggers and flows, leading to performance bottlenecks, governor limit exceptions, and corrupted legacy data relationships during large data loads.

83
MCQmedium

When designing a data model for a multi-region deployment, what is the best way to handle global picklist values?

A.Copy values to every object
B.Use Global Value Sets
C.Use a custom metadata type
D.Use a text field
AnswerB

Global Value Sets provide a centralized, reusable list of picklist values. This ensures that the same set of options is available across multiple objects, guaranteeing data consistency. It simplifies administration because changes are propagated automatically, making it the best architectural choice for maintaining data integrity in global deployments.

Why this answer

Global Value Sets allow for centralized management of picklist values across multiple objects and regions. This ensures consistency in reporting and data entry, as the same set of values is used consistently. If a value needs to be updated, it only needs to be changed in one place, which reduces the risk of data inconsistency and makes the maintenance of global data models much more efficient.

Exam trap

Candidates often suggest creating multiple independent picklists for different regions, which creates massive technical debt and makes global reporting impossible across the entire organization.

84
MCQmedium

A multinational company uses Salesforce across multiple regions. Each region has its own data retention requirements: the EU requires customer data to be deleted after 7 years, while the US requires retention for 10 years. The company wants to implement a data lifecycle governance policy that automatically enforces these retention periods. Which solution should the data architect recommend?

A.Create a custom object to store retention policies and use scheduled Apex to delete records based on region-specific criteria.
B.Rely on the Salesforce recycle bin and set its retention period to 7 years for EU records and 10 years for US records.
C.Use Salesforce Data Retention policies in Setup to define region-specific retention rules and automate deletion.
D.Implement a combination of Salesforce Shield Field Audit Trail for compliance and a custom batch process to archive and purge records according to regional policies.
AnswerD

Field Audit Trail provides long-term retention of field history for compliance, while a custom batch process can enforce region-specific retention by archiving and purging records. This combination addresses both the need for auditability and the requirement to delete data after the specified periods. It allows the company to tailor retention logic per region and maintain a clear audit trail.

Why this answer

The company needs region-specific retention enforcement with auditability. Field Audit Trail retains field history for compliance, while a custom batch process can archive and purge records according to each region's retention period. This approach provides the flexibility to implement different retention rules and maintains an audit trail, unlike native recycle bin or non-existent general retention settings.

Exam trap

The trap here is assuming Salesforce has a built-in data retention policy UI for arbitrary records, when in reality retention requires a combination of audit trail and custom automation.

85
MCQmedium

A Data Architect is tasked with ensuring Data Lifecycle Management is governed. Which lifecycle stage should include a requirement for automated data archiving?

A.Data Creation.
B.Data Usage.
C.Data Retention.
D.Data Acquisition.
AnswerC

Retention governance defines how long data is stored and when it should be moved or archived. Automating this phase ensures that production storage is kept clean and performant. By moving inactive data to archival storage, the organization balances regulatory requirements with the need for a lean, high-performing active Salesforce environment.

Why this answer

The 'Retention' or 'Disposition' phase is where archiving is most critical. As data ages, its value decreases, while storage costs and security risks remain. By governing this stage, the architect ensures that older data is moved to cheaper, secure storage (archiving) rather than being deleted or kept in active production tables, which optimizes system performance and maintains compliance with storage policies.

Exam trap

Candidates often confuse the 'Archiving' stage with 'Data Storage' or 'Database Optimization,' failing to recognize that retention policies specifically dictate the timing and conditions for moving data to cold storage.

86
MCQhard

Which TWO factors should be considered when choosing between a Master-Detail and a Lookup relationship for a new custom object? (Choose two)

A.Security and sharing inheritance
B.The number of records
C.Need for Rollup Summary fields
D.The color of the object
E.The API name prefix
AnswerA, C

Master-Detail relationships force child records to inherit the sharing and security settings of the parent. In contrast, Lookup relationships allow the child to have its own independent sharing rules, making this a fundamental design decision for data access control and visibility across different user profiles.

Why this answer

Choosing the correct relationship involves evaluating data ownership and security requirements. Master-Detail relationships propagate security settings from the parent to the child, which is a key architectural decision. Lookup relationships offer independent sharing, which is vital when child records should have distinct visibility regardless of the parent.

Architects must weigh these against the need for cascade deletes and rollup summaries to ensure the model aligns with business processes.

Exam trap

Candidates often confuse Lookup relationships with Master-Detail by assuming Lookups can automatically calculate rollup summary fields or enforce cascade deletes without custom automation.

87
MCQmedium

Universal Containers needs to provide a Full Sandbox for testing, but they must ensure that sensitive customer data like Social Security Numbers and Credit Card details are not visible to developers. Which solution is most appropriate?

A.Use a Partial Copy Sandbox and exclude the sensitive objects.
B.Implement Salesforce Data Mask to anonymize data during refresh.
C.Encrypt the fields in Production using Shield Platform Encryption.
D.Write a post-copy Apex script to delete sensitive field values.
AnswerB

Salesforce Data Mask is a powerful tool that automatically replaces sensitive information with random characters or mapped values during the sandbox creation or refresh. This allows developers to work with realistic data structures and volumes without being exposed to actual sensitive customer information.

Why this answer

Data Masking is a specific security process used to anonymize or pseudonymize sensitive data when it is copied from a production environment to a sandbox. Salesforce Data Mask allows administrators to mask sensitive data automatically during the sandbox refresh process, ensuring compliance with privacy regulations.

Exam trap

Candidates mistakenly suggest manual data scrubbing or writing custom Apex scripts to anonymize data after a sandbox refresh, ignoring automated platform native features.

88
MCQmedium

A Data Architect is analyzing a legacy system migration. Which THREE factors should be evaluated before deciding between a Lookup or a Master-Detail relationship?

A.Whether child records should be automatically deleted if the parent is deleted.
B.The need for roll-up summary fields on the parent object.
C.The number of child records expected to be created.
D.Whether the child record needs to have its own owner.
E.The color of the record detail page header.
AnswerA, B, D

Master-detail relationships feature cascading deletes, where child records are automatically removed when their parent is deleted. If the business requirement dictates that records should persist independently of the parent, a Lookup relationship is necessary to avoid data loss and maintain the integrity of the child records.

Why this answer

Choosing the right relationship is foundational to Salesforce architecture. Factors like data ownership, lifecycle dependency, and reporting needs drive the decision. A Master-Detail relationship implies structural dependency, while a Lookup offers flexibility.

Understanding these nuances early in the migration prevents architectural debt, ensures correct security enforcement, and aligns the data model with the organization's business process requirements for record access and summarization.

Exam trap

Candidates often overlook the security and ownership implications, forgetting that child records in a Master-Detail relationship do not have their own owner field and inherit parent security.

89
MCQeasy

A Salesforce org has a data governance policy that requires all new custom objects and fields to be reviewed and approved by a data governance council before creation. A developer needs to add a new field to the Account object to support a critical business process. What should the developer do first?

A.Ask the Salesforce administrator to create the field as a temporary workaround and document it in a technical debt log.
B.Use the Salesforce Setup menu to create the field directly in production, then notify the governance council afterward.
C.Create the field in a sandbox and then submit a change set for deployment to production.
D.Submit a request to the data governance council for review and approval, providing business justification and impact analysis.
AnswerD

The governance policy explicitly requires review and approval by the data governance council before creating new custom objects or fields. Submitting a request with justification and impact analysis is the correct first step. It ensures the council can assess the need, check for redundancy, and maintain data standards. Only after approval should the developer proceed with creation.

Why this answer

The governance policy clearly states that new custom objects and fields must be reviewed and approved before creation. The developer should first submit a request to the data governance council with business justification and impact analysis. This ensures the council can evaluate the request against data standards and avoid unnecessary or redundant fields.

Only after approval should the field be created.

Exam trap

The trap here is focusing on the technical steps of field creation and overlooking the mandatory governance approval process that must occur first.

90
MCQeasy

Which object type is most likely to cause performance issues in an LDV environment if not managed correctly?

A.A flat custom object with no relationships.
B.An object with many lookup relationships to other parent objects.
C.An object containing only standard picklist fields.
D.A setup object maintained by the system administrator.
AnswerB

Objects with many lookup relationships are problematic because every record update might involve locking multiple parent records. This leads to severe contention during large data imports and complex queries. These objects require careful planning, such as implementing indexing and optimizing batch processes to minimize the locking overhead.

Why this answer

Highly relational objects, especially those with many lookup relationships, are most prone to performance issues. Each lookup relationship can act as a locking point, and queries involving these fields often require complex joins. By understanding which objects are 'hot' in terms of frequency of access and relationship density, architects can prioritize them for indexing, archiving, and skinny table strategies to maintain system stability.

Exam trap

Candidates often assume that only 'large' objects cause issues. They fail to realize that objects with many lookup relationships create locking contention and join complexity that drastically slows down performance.

91
MCQmedium

A Salesforce data architect is migrating 500,000 Lead records from a legacy system. The legacy data contains a field 'Lead_Status' with values such as 'New', 'Working', 'Qualified', 'Unqualified'. Salesforce's Lead Status picklist has values: 'Open - Not Contacted', 'Working - Contacted', 'Closed - Converted', 'Closed - Not Converted'. The architect needs to map the legacy values to the Salesforce picklist values. Which approach should be used to ensure a successful migration?

A.Load the legacy values as-is and then run a batch Apex job to update the Lead Status field.
B.Create a custom field to store the legacy Lead Status and leave the standard Lead Status blank.
C.Modify the Salesforce Lead Status picklist to include the legacy values.
D.Use an ETL tool to map legacy values to the corresponding Salesforce picklist values before loading.
AnswerD

An ETL tool can transform the legacy Lead Status values to match the Salesforce picklist values exactly. This ensures that the standard Lead Status field is populated correctly and avoids load errors. The mapping can be defined in the ETL tool's transformation logic. This is the most reliable method for ensuring data quality.

Why this answer

Mapping legacy values to the standard Salesforce picklist values must occur before loading to avoid errors and ensure data consistency. An ETL tool provides the necessary transformation capabilities. Other options either avoid the mapping, cause load failures, or alter the Salesforce schema unnecessarily.

Exam trap

The trap here is thinking that Salesforce will automatically map legacy picklist values or that the picklist can be easily modified without impact.

92
MCQhard

Refer to the exhibit. In a 50 million record Account table, why might this query perform poorly?

A.The CreatedDate field is not indexed by default in Salesforce.
B.The custom field is likely missing an index, causing a full table scan.
C.The Id field should be the only field in the WHERE clause.
D.Apex triggers are preventing the query from completing.
AnswerB

Without an index on the custom field, the query optimizer cannot efficiently narrow down the search space. When dealing with millions of records, the lack of an index forces the system to scan every row, which is the primary cause of slow query execution and potential timeout errors.

Why this answer

This query uses a filter on CreatedDate (a system-indexed field) and a custom field (Custom_External_ID__c). If Custom_External_ID__c is not indexed, the query optimizer must perform a full table scan. Even with a system index, if the combination of filters is not selective enough, the query will exceed the platform's selectivity threshold, leading to a non-selective query error or severe performance degradation.

Exam trap

Candidates often assume that any field used in a filter is automatically indexed by the platform. They overlook the need for custom indexes on non-standard fields in large tables.

93
MCQhard

A healthcare organization uses Salesforce to manage patient cases and must comply with HIPAA. The data governance team needs to ensure that audit trails for access to Protected Health Information (PHI) are comprehensive and tamper-evident. They are evaluating Salesforce Shield Event Monitoring and Field Audit Trail. Which combination of features should the team implement to meet the requirement for tracking who accessed PHI fields and when, while also retaining the audit data for 10 years?

A.Use Salesforce Shield Platform Encryption to encrypt PHI fields, which automatically generates audit logs for all access and retains them for 10 years.
B.Enable Field Audit Trail with a retention policy of 10 years, and use Event Monitoring to track field-level access events.
C.Use Event Monitoring alone, as it captures all user activity including field-level changes, and set the retention period to 10 years.
D.Enable Field Audit Trail and configure it to also capture read access events, eliminating the need for Event Monitoring.
AnswerB

Field Audit Trail tracks changes to field values and can retain field history for up to 10 years, meeting the retention requirement. Event Monitoring captures access events, including who viewed or exported data, providing a comprehensive audit trail. Together, they address both change tracking and access tracking for PHI, ensuring HIPAA compliance.

Why this answer

Field Audit Trail provides long-term retention of field history, up to 10 years, and tracks changes to PHI fields. Event Monitoring captures access events, including reads and exports, which are not covered by Field Audit Trail. Together, they deliver a complete audit trail for HIPAA compliance, ensuring both change and access tracking with the required retention.

Exam trap

The trap here is believing that Field Audit Trail captures read access or that Event Monitoring alone can provide 10-year retention, when in fact they serve complementary roles.

94
Multi-Selecthard

An enterprise architect is evaluating data storage strategies for a telecommunications client experiencing rapid data growth of fifty million call detail records annually. Which TWO strategies should the architect implement to maintain database query performance and platform data limits? Choose 2 answers.

Select 2 answers
A.Store the call detail records in standard custom objects with indexes applied to all foreign keys.
B.Utilize Salesforce Big Objects to store historical call records asynchronously and index primary lookup keys.
C.Configure Salesforce Connect with OData adapters to federate call detail records from an external data lake.
D.Create hierarchical custom object structures linking every call record directly to the Account master object.
E.Rely on standard Salesforce report archives to automatically move historical records out of reporting tables.
AnswersB, C

Big Objects store billions of records on a separate, non-transactional index, so historical call detail records leave standard object storage and its limits. Asynchronous writes and indexed lookup keys keep queries performant while preserving the required access path.

Why this answer

When dealing with massive data volumes exceeding standard transactional limits, traditional custom objects will quickly degrade platform performance and storage capacity. Offloading historical analytical data to Salesforce Big Objects or external systems via Salesforce Connect preserves core transactional limits. Additionally, implementing strict data archiving strategies prevents data skew and maintains query optimization.

Exam trap

Candidates often suggest custom objects or external objects without considering the specific performance benefits of Big Objects for massive, non-transactional historical data sets.

95
MCQhard

A company is migrating 100 million records into Salesforce. Which TWO actions should the architect take to optimize performance and prevent row locking?

A.Enable all custom indexes before the data load.
B.Disable unnecessary triggers during the data load.
C.Use the Bulk API 2.0 with serial mode.
D.Reduce the number of indexes on the target object.
E.Increase the batch size to 10,000.
AnswerB, D

Triggers run for every record processed, consuming CPU and database resources while creating row locks on parent or related objects. Disabling them allows the system to focus exclusively on record insertion, significantly increasing throughput and avoiding potential deadlocks caused by concurrent logic execution during the migration.

Why this answer

High-volume data loads require careful management of record locking and indexing. Disabling triggers during the load prevents unnecessary processing and lock contention. Furthermore, reducing the number of indexes on the target object minimizes the overhead required for every insert operation, as Salesforce must update every index on the object for each new record processed during the migration process.

Exam trap

Candidates often overlook the impact of indexes on DML performance, assuming more indexes are always better, when in reality, every index adds overhead during mass record inserts.

96
MCQmedium

A company is migrating 50 million records into Salesforce. They need to ensure data integrity while minimizing API consumption. Which strategy is most efficient for a one-time bulk data load?

A.Use the standard REST API with single-record inserts.
B.Utilize the SOAP API with synchronous processing.
C.Use the Bulk API 2.0 with CSV files.
D.Perform a manual data import via the UI.
AnswerC

Bulk API 2.0 is purpose-built for high-volume data movement, handling large batches asynchronously to maintain system stability. It provides significant performance gains by allowing Salesforce to manage job batching internally. This minimizes the risk of hitting governor limits and ensures successful large-scale data migrations for enterprise environments.

Why this answer

Bulk API 2.0 is the recommended approach for large datasets as it optimizes throughput by automatically breaking large jobs into smaller batches. It significantly reduces the number of API calls required compared to standard REST or SOAP APIs. Properly architecting for data loads involves choosing tools that leverage Bulk API 2.0 to maintain governor limits and ensure successful processing of high-volume data without impacting real-time integrations.

Exam trap

Candidates often choose standard REST or SOAP APIs for large migrations, forgetting that they consume thousands of individual API calls and hit governor limits quickly on high-volume datasets.

97
MCQmedium

What is the primary benefit of using a staging database for data migration?

A.It provides a backup of the source data.
B.It allows for complex data transformation and cleansing.
C.It eliminates the need for field mapping documentation.
D.It speeds up the actual data insertion into Salesforce.
AnswerB

Staging databases are essential for complex transformations that are difficult to perform within Salesforce or CSV tools. They enable SQL-based joins, cleansing scripts, and deduplication logic, ensuring that the data is perfectly structured and validated before it is uploaded to the final production instance for migration.

Why this answer

A staging database provides a clean, neutral environment to perform ETL (Extract, Transform, Load) tasks such as data deduplication, value normalization, and relationship mapping. By transforming the data outside of Salesforce, architects ensure that only clean, verified data enters the target environment. This minimizes the risk of system-level errors and validation failures that occur when attempting to manipulate raw, messy data directly inside the Salesforce platform.

Exam trap

Candidates often confuse the staging database with a backup or a data warehouse. They fail to recognize its primary function as an ETL workspace for cleaning and transforming data before Salesforce ingestion.

98
MCQmedium

Which of the following is a classic symptom of poor Master Data Management in a Salesforce ecosystem?

A.Salesforce users reporting that their dashboards load slowly.
B.Sales and Marketing teams having conflicting views of the same customer.
C.Excessive use of Apex triggers that cause governor limit exceptions.
D.Users forgetting their passwords frequently, requiring support tickets.
AnswerB

When teams rely on disparate, un-synchronized data sources, they often have different versions of the truth for a single customer. This lack of a unified golden record leads to disconnected customer experiences, such as marketing sending promotional emails to a customer who has already churned in the CRM.

Why this answer

Fragmented data silos often result in the same customer existing as multiple records across different systems (e.g., Salesforce, SAP, and Marketing Cloud). This leads to poor customer service, conflicting marketing messages, and inaccurate financial reporting. Identifying these symptoms is critical for a Data Architect to justify the investment in an MDM solution, as these issues directly impact the bottom line through operational inefficiency and missed sales opportunities.

Exam trap

Candidates often look for technical symptoms like 'API errors' or 'slow page loads' rather than identifying business-level symptoms such as departmental silos and conflicting data views.

99
MCQmedium

Which design approach is best for handling a 'Data Warehouse' requirement within Salesforce when dealing with millions of records?

A.Use standard objects and archive periodically.
B.Implement Big Objects for high-volume storage.
C.Store all data in a single custom object.
D.Use a custom field to store external reference IDs.
AnswerB

Big Objects are purpose-built for massive scale, allowing storage and querying of billions of records. They do not count against the standard record count limits and offer optimized performance for analytical queries. This is the optimal architecture for data-intensive requirements, ensuring that the CRM platform remains fast and responsive.

Why this answer

For high-volume data, architects should utilize Big Objects or External Objects to prevent hitting platform storage limits and performance degradation. Big Objects are specifically designed to store and query massive amounts of data efficiently. This strategy is vital for data management, as it keeps the core CRM performance high while still maintaining accessibility to historical archives needed for operational reporting or compliance purposes.

Exam trap

Candidates mistakenly suggest standard custom objects for long-term archiving of tens of millions of records, ignoring platform storage limits and data skew.

100
MCQeasy

A Salesforce data architect is preparing to migrate 5 million Account records from a legacy CRM. The legacy system has a field 'Legacy_Owner__c' that contains the email address of the account owner. During migration, the architect must ensure that the ownership is assigned to the correct active Salesforce User. Which approach should be used to map the legacy owner email to the Salesforce User ID?

A.Load the email addresses into the OwnerId field directly; Salesforce will automatically resolve them to User IDs.
B.Use the Data Loader's 'Bulk API' with the 'Assignment Rule' option enabled to automatically assign owners based on email.
C.Create a custom field on Account to store the legacy owner email, then use a post-load process or Data Loader to update OwnerId based on a User lookup.
D.Load the legacy owner email into the Account Owner field using the 'External ID' feature of the User object.
AnswerC

This approach correctly separates the migration of Account data from the ownership assignment. Storing the legacy email in a custom field preserves the original data for audit, and a subsequent update using a User lookup (via ETL, Data Loader, or Apex) can accurately set OwnerId. This avoids load failures and ensures ownership is assigned to the correct active User. It also allows validation of email-to-user mappings before final assignment.

Why this answer

To map legacy owner emails to Salesforce User IDs, the architect should first load Accounts with the legacy email stored in a custom field. Then, using a lookup or ETL transformation, the correct User IDs can be determined and applied to the OwnerId field. This two-step process avoids errors and ensures accurate ownership assignment.

Exam trap

The trap here is assuming that Salesforce can automatically resolve email addresses to User IDs during a data load without an explicit mapping step.

101
MCQmedium

Which platform feature is best used to move data off-platform to avoid Large Data Volume issues?

A.Salesforce BigObjects.
B.Salesforce Connect.
C.Data Loader command line.
D.Platform Events.
AnswerB

Salesforce Connect integrates external data into the org via external objects. This allows companies to keep data outside of the Salesforce database, preventing native performance degradation while ensuring users still have access to the information, which is a perfect pattern for large-scale data archival and management needs.

Why this answer

Salesforce Connect allows organizations to display data from external systems as if it were stored natively in Salesforce, without actually consuming storage or impacting governor limits. By using external objects, companies can maintain massive amounts of historical data off-platform while still providing users with a seamless interface for viewing and reporting on that data within the native environment.

Exam trap

Candidates often suggest archiving data to a Big Object or a standard custom object, which still consumes storage, rather than using Salesforce Connect to keep data off-platform.

102
MCQhard

A data architect at a financial services company is designing a master data management solution for 'Product' data. The company has multiple source systems: Salesforce, a legacy ERP, and a product information management (PIM) system. The architect must ensure that the golden record for each product is always the most trusted version. Which factor is most critical when defining survivorship rules for product attributes?

A.The recency of the attribute value across all source systems.
B.The number of source systems contributing to the attribute.
C.The source system's authority for that specific attribute.
D.The data type of the attribute (e.g., text, number, date).
AnswerC

Survivorship rules should prioritize the source system that is most authoritative for a given attribute. For product data, the PIM system is typically the authority for marketing attributes, while the ERP is authoritative for pricing and inventory. This ensures the golden record reflects the most trusted value. It is the most critical factor because it directly determines trustworthiness.

Why this answer

Survivorship rules must prioritize the source system that is most authoritative for each attribute. For product data, different systems may be authoritative for different attributes: the PIM for descriptions, the ERP for pricing. This attribute-level authority ensures the golden record contains the most trusted values.

Recency, volume, or data type do not guarantee trustworthiness.

Exam trap

The trap here is assuming that the most recent value or the majority value is always the most trusted, when source authority per attribute is the key determinant.

103
MCQmedium

Which strategy should be employed when designing a data archiving solution for a high-volume object to ensure continued system performance?

A.Store all historical data in a hidden custom object within Salesforce.
B.Regularly delete records that are older than three years.
C.Offload historical records to an external system or Big Object.
D.Use Salesforce Sharing Rules to hide older records from users.
AnswerC

This is the correct approach to maintain performance. Offloading data to an external repository or Big Object reduces the record count in the transactional object, keeping indexes lean and queries fast, while still allowing access to historical data when needed for compliance or analytical purposes.

Why this answer

Archiving is critical for maintaining performance in Salesforce. By moving stale data to an external data store or a Big Object, you keep the active 'working set' of records small. This ensures that queries, reports, and DML operations remain fast.

An effective strategy involves identifying criteria for record age or status and offloading these records periodically to prevent the primary object from reaching a state that degrades system performance.

Exam trap

Many candidates incorrectly recommend creating more custom indexes or skinny tables instead of removing historical data from the active database entirely.

104
MCQmedium

A large enterprise is struggling with inconsistent data quality across multiple business units. They want to establish a Data Governance Council. Who should lead this council to ensure the initiative receives adequate executive sponsorship and alignment with corporate strategy?

A.The Lead Salesforce Administrator.
B.The Chief Data Officer or a senior business executive.
C.The Vice President of IT Infrastructure.
D.The external Salesforce Implementation Partner.
AnswerB

A Chief Data Officer or senior executive provides the necessary mandate to enforce governance policies across diverse business units. Their involvement ensures that data governance is treated as a strategic business initiative, securing the executive-level support needed to manage organizational change and resource allocation effectively.

Why this answer

Executive sponsorship is the single most important factor for the success of a Data Governance initiative. By appointing a Chief Data Officer or a senior business leader, the organization signals that data is a strategic asset rather than just an IT concern. This alignment ensures that funding, resources, and cross-departmental cooperation are prioritized, effectively breaking down silos that often impede data quality projects.

Exam trap

Candidates often suggest a 'Lead Data Architect' or 'IT Manager' to lead the council, failing to recognize that governance requires executive authority to enforce changes across business units.

105
MCQhard

Which design pattern effectively handles high-volume record updates while avoiding 'Too many SOQL queries' errors?

A.Perform a SOQL query inside the trigger loop.
B.Use a Map to cache related records.
C.Use the Future method for every record update.
D.Call the update DML statement inside the loop.
AnswerB

Caching related parent records in a Map ensures that SOQL queries are performed only once for the entire batch. This minimizes resource consumption and prevents governor limit violations, allowing developers to process thousands of records efficiently, which is vital for high-volume data operations in the Salesforce platform.

Why this answer

The use of Maps for caching parent records is essential to avoid repeated querying. By querying all necessary parent records once, storing them in a Map, and accessing them by ID in a loop, you reduce the SOQL query count to one. This pattern is the industry standard for handling large collections of records without hitting governor limits.

Exam trap

Candidates commonly write queries inside iterative loops to fetch related parent data, quickly exhausting SOQL query limits instead of utilizing collection mapping techniques.

106
Multi-Selecthard

An organization is establishing a Data Governance Council. Which THREE outcomes are primary goals of this council? (Choose Three)

Select 3 answers
A.Resolving data ownership and stewardship conflicts.
B.Setting organization-wide data standards.
C.Managing daily Apex deployment cycles.
D.Fostering data stewardship and accountability.
E.Directly configuring Salesforce security profiles.
AnswersA, B, D

Conflicts often arise regarding who 'owns' a piece of data, leading to inconsistent definitions. The council acts as the final arbiter for these disputes, ensuring that roles are clearly defined and that business units agree on the sources of truth, which is essential for uniform reporting and data usage.

Why this answer

A Data Governance Council is a cross-functional body that aligns data strategy with business objectives. By resolving data ownership conflicts, setting enterprise-wide standards, and fostering communication between business units and IT, the council ensures that data is managed as a strategic asset. This alignment prevents silos, ensures accountability, and maximizes the value derived from data across the entire organization.

Exam trap

Candidates frequently select low-level technical execution tasks like writing ETL scripts, forgetting that a Data Governance Council operates at a strategic, cross-functional policy-setting level.

107
MCQmedium

An architect is designing a schema to store product information. Each product has many versions, and each version has many components. What is the most efficient way to model this relationship?

A.Flatten everything into one giant object with 100+ custom fields.
B.Use a JSON field to store the hierarchy as a string.
C.Model as three separate objects with Master-Detail relationships.
D.Create a separate object for every possible component.
AnswerC

A Master-Detail model enforces referential integrity and allows for easy aggregation of data through roll-up summaries. This hierarchy is the most efficient way to store, query, and manage complex product data. It ensures that every component is correctly linked to its parent version, which is linked to its parent product.

Why this answer

Creating a hierarchical relationship using lookup or master-detail fields on 'Product', 'Version', and 'Component' objects is the most scalable way to represent this structure. By isolating each layer, the architect ensures that reporting, security, and maintenance are simplified. This hierarchical model is cleaner than attempting to flatten the data, as it preserves parent-child integrity at every level.

Exam trap

Candidates frequently attempt to flatten the data into a single object or use too many lookups, failing to realize that Master-Detail relationships are necessary for proper hierarchical rollups and security.

108
MCQmedium

When designing a system that requires frequent querying of very large objects, which approach provides the best performance while maintaining data integrity?

A.Always use SOQL with complex joins across many tables.
B.Implement custom indexing or Skinny tables.
C.Move all data to a custom object to simplify the schema.
D.Use the Salesforce REST API for all data retrieval.
AnswerB

Custom indexing and Skinny tables are the most effective native ways to optimize read performance. By providing the database with direct access paths or denormalized data sets, these features allow the query optimizer to return results rapidly, avoiding the performance pitfalls of full table scans.

Why this answer

Denormalization via techniques like Skinny tables or indexing is the standard way to optimize read performance for large volumes. These strategies shift the cost from query time to write time, which is usually preferable in high-volume systems where users need quick access to data. By aligning the database schema with the specific query patterns of the application, you minimize the work the database must do to return results.

Exam trap

Candidates often select standard sharing rules or caching mechanisms, missing that schema-level adjustments are required for massive data volume queries.

109
MCQhard

A bank requires a full audit trail of all record changes for compliance. What is the most effective way to track changes to sensitive fields?

A.Enable standard Field History Tracking.
B.Create a custom 'Audit' object with triggers.
C.Use Field Audit Trail (Shield).
D.Schedule daily exports of all records.
AnswerC

Field Audit Trail provides the deep, long-term history tracking required for compliance. It supports up to 10 years of data and allows for tracking a significantly higher number of fields than standard history tracking. It is the architecturally correct choice for meeting stringent regulatory data retention and audit requirements.

Why this answer

Field Audit Trail is the most reliable way to meet regulatory compliance requirements for tracking data history. Unlike standard Field History Tracking, which is limited in retention and volume, Field Audit Trail allows for long-term retention and higher storage limits. This is essential for highly regulated industries like banking, where historical data accuracy and a permanent audit trail are mandatory for legal and security compliance.

Exam trap

Candidates select standard Field History Tracking, forgetting that it has strict retention limits and maximum tracked field constraints unsuitable for strict banking compliance.

110
MCQmedium

Universal Containers maintains a single Salesforce org used by sales teams across North America, EMEA, and APAC. Each region has its own legal requirements for how long personal data may be retained. The VP of Sales wants one consistent retention policy applied everywhere to simplify administration. As the data architect, what should you recommend?

A.Retain all personal data indefinitely because Salesforce storage is inexpensive and deletion carries operational risk.
B.Implement a single global retention schedule for all personal data records and document it in the data governance policy.
C.Define region-specific retention rules in the governance policy and enforce them through record segmentation and automated deletion processes.
D.Delegate all retention decisions to regional sales managers and let each team manage its own data without a central policy.
AnswerC

Region-specific retention rules honor each jurisdiction's legal requirements while still operating under a single overarching governance framework. Enforcement requires segmenting records (for example, by region field or record type) and scheduling automated deletion or archival jobs per segment. This satisfies both the compliance mandate and the need for a consistent, documented policy.

Why this answer

Retention requirements that vary by jurisdiction demand a governance policy that encodes those differences rather than flattening them. Segmenting records by region and applying automated retention jobs per segment allows one org to comply with multiple legal regimes. This approach keeps the policy centralized and auditable while respecting local law, which is the core purpose of a data governance framework in a multi-region Salesforce deployment.

Exam trap

The trap here is assuming that a simpler, uniform retention policy is automatically better because it is easier to administer, when legal compliance actually requires differentiated handling by region.

111
Multi-Selecthard

A company is implementing a Customer Data Platform (CDP) to consolidate data from Salesforce, a web store, and an email marketing tool. Which TWO steps are critical for successful identity resolution?

Select 2 answers
A.Define deterministic and probabilistic matching rules.
B.Migrate all Salesforce CRM data into the marketing tool's database.
C.Establish survivorship rules to determine which system's data takes precedence.
D.Remove all personally identifiable information from the incoming data streams.
E.Automate the deletion of all records that do not contain a phone number.
AnswersA, C

Deterministic matching uses exact identifiers like email or ID, while probabilistic matching uses fuzzy logic to link records based on confidence scores. Combining both approaches maximizes the reach of the identity resolution process, ensuring that fragmented customer interactions are accurately linked into a single cohesive profile.

Why this answer

Identity resolution is the process of linking data points from disparate sources to create a unified profile. By defining clear matching rules and prioritizing data attributes, the CDP can accurately identify unique individuals even when they use different identifiers across channels. This is vital for personalized marketing and accurate customer analytics, ensuring that the organization does not treat the same individual as multiple distinct records.

Exam trap

Candidates frequently select only one of the two steps, or focus solely on technical matching algorithms while neglecting the business-critical aspect of deciding which system's data is the authoritative source.

112
MCQmedium

Universal Containers wants to ensure data quality by preventing the creation of duplicate Leads from various sources. The architect recommends matching rules and duplicate rules. Which action should the architect take to ensure that users are alerted to potential duplicates during manual entry while blocking automated integrations?

A.Enable 'Block' for both user-facing and integration-based duplicate rules.
B.Apply 'Report' to the record creation rule and ignore the API settings.
C.Set 'Report' for record creation and 'Block' for API insertion.
D.Use a third-party AppExchange solution exclusively for integration rules.
AnswerC

This configuration directly maps to the requirements. The 'Report' action allows users to proceed after an alert, facilitating manual data entry, while the 'Block' action on the API ensures that automated integration processes cannot bypass duplicate prevention, effectively maintaining high data quality across all system interaction types.

Why this answer

To achieve this, the architect must configure duplicate rules with distinct settings for user-based versus integration-based scenarios. By setting the 'Report' action for user-facing rules, the UI displays an alert, while selecting 'Block' for API-based rules prevents duplicates from being inserted via integration tools. This tiered approach maintains data integrity across different entry points without hindering necessary manual workflows during prospect record creation.

Exam trap

Candidates often think duplicate rules apply globally, failing to recognize that they can configure different actions for manual UI entry versus automated API integrations within the same rule.

113
MCQmedium

A multinational corporation is migrating 2 million Account records and 5 million Contact records from a legacy CRM into Salesforce. The legacy system stores the Account's legacy ID on the Contact record as a foreign key. The target Salesforce org uses an external ID field on Account called Legacy_ID__c. What is the most efficient way to associate Contacts with their Accounts during the migration?

A.Migrate Contacts first with a placeholder Account, then migrate Accounts and update Contacts via a batch Apex job.
B.Export the Salesforce Account record IDs after migration, then use those IDs to update the Contact records in the legacy system before migrating Contacts.
C.Use the Data Loader's 'Insert' operation for Contacts and rely on Salesforce's auto-association feature to link Contacts to Accounts based on matching names.
D.Migrate Accounts first, then migrate Contacts with a lookup to Account using the Legacy_ID__c field as the external ID in the Account relationship field.
AnswerD

This is correct because Salesforce allows you to populate a lookup relationship using an external ID field. After Accounts are migrated, Contacts can be loaded with the Account's legacy ID in the Account lookup field, and Salesforce will resolve it to the correct Account record. This avoids the need for a second pass or manual mapping and is efficient for large volumes.

Why this answer

Migrating Accounts first and then using the external ID field Legacy_ID__c in the Contact's Account lookup field allows Salesforce to automatically resolve the relationship. This is the most efficient method because it leverages built-in external ID resolution during data load, eliminating the need for post-migration updates or complex Apex jobs. It ensures referential integrity and scales well for large data volumes.

Exam trap

The trap here is thinking that you need to migrate Contacts first or use custom code to establish relationships, when Salesforce natively supports populating lookups via external IDs.

114
Multi-Selectmedium

A data architect is preparing to migrate 20 million Opportunity records from a legacy CRM into Salesforce. The source data includes Opportunities with related OpportunityLineItem records and historical StageName values. The architect must ensure that the migration preserves data integrity and avoids common Bulk API pitfalls. Which TWO actions should the architect take before initiating the load? (Choose two.)

Select 2 answers
A.Disable all validation rules, workflow rules, and triggers on Opportunity during the load to improve performance.
B.Load Opportunity records first, then load OpportunityLineItem records using the parent Opportunity's External ID to establish the relationship.
C.Set the Batch Size to 1 record to avoid governor limit issues and ensure maximum error isolation.
D.Enable AllOrNone on all Bulk API batches to guarantee that no partial data is committed if a single record fails.
E.Ensure that the External ID field on Opportunity is marked as Unique to support upsert and prevent duplicate Opportunities.
AnswersB, E

OpportunityLineItem records require a parent OpportunityId. By loading Opportunities first with an External ID, the architect can then load line items referencing that External ID, ensuring referential integrity. This two-pass approach is standard for parent-child migrations and avoids orphaned child records.

Why this answer

Loading parents before children using an External ID preserves referential integrity, and marking the External ID as Unique supports upsert and prevents duplicates. Together, these actions address the core data integrity concerns for a large Opportunity migration with related line items.

Exam trap

The trap here is thinking that disabling automation or using AllOrNone is a best practice for data integrity; both can introduce risk or reduce throughput without guaranteeing correctness.

115
MCQmedium

When designing a custom Data Model, when should an architect choose a 'Lookup' relationship over a 'Master-Detail' relationship?

A.When the child records must inherit the parent's security.
B.When the child record must have its own sharing settings.
C.When the child must be deleted when the parent is deleted.
D.When you need to create a Roll-up Summary field.
AnswerB

Lookup relationships are independent, allowing for granular sharing controls on both the parent and the child. If the business requirement demands that the child object has its own unique security rules separate from the parent, then a lookup relationship is the architecturally correct choice.

Why this answer

Lookup relationships offer flexibility where the parent and child records are logically distinct and can exist independently. This is ideal for scenarios where sharing is decoupled or when the child records should not be deleted upon parent deletion. By understanding these nuances, an architect avoids the 'cascading' consequences of Master-Detail relationships, which is vital for designing a loosely coupled, maintainable system that satisfies complex business requirements without unnecessary record interdependencies.

Exam trap

Candidates often choose Master-Detail relationships for convenience without considering that it forces child records to inherit parent sharing, which violates complex business security requirements.

116
MCQeasy

A data architect needs to ensure that when a user updates a field on a custom object, a related record on another object is automatically updated to reflect the change. The architect wants to avoid writing code. Which declarative feature should be used?

A.Workflow Rule with a Field Update
B.Record-Triggered Flow with an Update Records element
C.Process Builder with an Update Records action
D.Apex Trigger with a future method
AnswerB

Record-Triggered Flows are the modern declarative automation tool in Salesforce. They can be triggered when a record is created or updated, and they can update related records regardless of whether the relationship is master-detail or lookup. This meets the requirement without code. Flows are more powerful and flexible than Workflow Rules or Process Builder, and they are the recommended approach for new automation.

Why this answer

Record-Triggered Flows are the declarative tool of choice for automating updates to related records. They can be configured to run before or after a record is saved, and they can update records on the same object or related objects. They support both master-detail and lookup relationships, and they do not require code.

This makes them ideal for the scenario where a related record needs to be updated when a field changes.

Exam trap

The trap here is assuming that Workflow Rules or Process Builder are still the primary declarative automation tools, when Flow is now the strategic direction.

117
MCQeasy

A healthcare company needs to store patient consent records that must be retained for ten years and are rarely accessed after the first year. The records include sensitive data and must be queryable by compliance officers using standard reports. Which storage strategy should the architect recommend?

A.Archive consent records to an external data warehouse and delete them from Salesforce after one year.
B.Store all consent records as standard custom object records and rely on Salesforce's default data retention.
C.Store consent records in a custom object with field history tracking enabled and purge old records annually.
D.Create a Big Object to store consent records and use Async SOQL or a custom index to query them for compliance reporting.
AnswerD

Big Objects are designed for long-term retention of large data volumes and support custom indexes for efficient queries. Compliance officers can query them through Async SOQL or by exposing them via a Lightning component, and the data remains in Salesforce without consuming standard data storage.

Why this answer

Big Objects are the Salesforce-native option for retaining large volumes of rarely accessed data for long periods while keeping it queryable. They do not consume standard data storage and support indexes for compliance queries. Standard custom objects, external warehouses, and field history tracking cannot simultaneously satisfy long-term retention and native reporting requirements.

Exam trap

The trap here is assuming field history tracking or external archiving preserves full records for compliance reporting, when only Big Objects keep them queryable in Salesforce long term.

118
MCQmedium

A Salesforce org has a custom object Log__c with 5 million records. The object has a lookup to Case. Users report that when they view a Case record, the related list of Log__c records takes a long time to load. The data architect decides to create a custom index on the Case lookup field. After the index is created, performance improves. Which statement best explains why the index improved performance?

A.The index reduces the number of records that need to be queried by pre-filtering the Log__c object.
B.The index allows the related list query to use an index to quickly retrieve Log__c records for the specific Case.
C.The index enables the related list to be loaded from a cache instead of the database.
D.The index sorts the Log__c records by Case, which speeds up the display order in the related list.
AnswerB

When viewing a Case record, the related list of Log__c records is displayed by querying Log__c where Case__c equals the current Case Id. This filter is highly selective because it returns only the logs for one case. With a custom index on Case__c, the query optimizer can use the index to efficiently retrieve those records, significantly improving performance.

Why this answer

The related list on Case queries Log__c records filtered by the Case lookup. This filter is highly selective because it returns only records for one Case. Creating a custom index on the lookup field allows the query optimizer to use the index, avoiding a full table scan.

Thus, the index directly improves the performance of the related list query.

Exam trap

The trap here is assuming that indexes improve performance by sorting or caching, when in fact their primary benefit is enabling fast, selective filtering.

119
MCQhard

A data architect is migrating 10 million Case records and their related Case Comments from a legacy system to Salesforce Service Cloud. The legacy system stores Case Comments in a separate table with a foreign key to the Case. The architect plans to use Bulk API 2.0 for the migration. What is the most critical consideration for maintaining referential integrity between Cases and Case Comments during the load?

A.Load Cases and Case Comments simultaneously using parallel processing to save time.
B.Load Cases first, then load Case Comments with the correct ParentId referencing the Salesforce Case ID.
C.Disable validation rules on Case Comments to avoid errors during the load.
D.Use an External ID field on Case Comments to link to the legacy Case ID without loading Cases first.
AnswerB

To maintain referential integrity, Cases must be loaded first to generate Salesforce IDs. Then, Case Comments can be loaded with the ParentId set to the corresponding Case ID. This requires mapping legacy Case IDs to Salesforce IDs, often via an External ID field on Case. This sequential approach ensures that each comment is linked to an existing Case.

Why this answer

Referential integrity requires that parent records exist before children. Therefore, Cases must be loaded first, and their Salesforce IDs captured to populate the ParentId on Case Comments. Using an External ID on Case can facilitate mapping.

Parallel loading or relying solely on External IDs without loading Cases first would fail.

Exam trap

The trap here is thinking that External IDs can bypass the need for the parent record to exist, or that parallel loading can maintain relationships.

120
Multi-Selecthard

Which TWO of the following are consequences of having excessive indexes on a Salesforce object with large data volumes?

Select 2 answers
A.Improved insert and update performance.
B.Increased time to complete DML operations.
C.Faster SOQL query execution times.
D.Higher risk of row-level locking contention.
E.Automatic reduction in storage space usage.
AnswersB, D

Because each write requires an update to the corresponding index table, having many indexes adds cumulative overhead to every DML statement. This lengthens the time required to commit transactions, which can eventually lead to governor limit issues and performance bottlenecks in high-volume environments.

Why this answer

While indexes improve read performance, they impose a cost on write operations. Every time a record is inserted, updated, or deleted, Salesforce must update all associated index tables. With an excessive number of indexes, this overhead becomes significant, leading to slower transaction times, potential lock contention, and overall system instability during high-volume DML activities.

Balancing read performance needs with write performance costs is critical in data architecture.

Exam trap

Many candidates assume indexes only have positive impacts, forgetting that database maintenance of numerous indexes severely degrades DML performance.

121
MCQhard

Refer to the exhibit. The architect observes this error during bulk Opportunity updates. Which action resolves the issue while adhering to best practices?

A.Increase the SOQL query limit by contacting Salesforce Support.
B.Add a static variable to track if the query has been executed.
C.Move the query outside the loop and store results in a Map.
D.Change the trigger to run in a 'without sharing' context.
AnswerC

Moving the query outside the loop is the standard pattern for bulkifying Apex code. By fetching all necessary PricebookEntry records into a Map at once, the logic can access them in constant time without re-querying the database, effectively eliminating the risk of exceeding SOQL limits during bulk data processing.

Why this answer

The error indicates a SOQL query inside a loop, a common mistake in bulk operations. Moving the query outside the loop into a Map or List ensures that data is retrieved in a single call, optimizing performance and staying within governor limits. This is a crucial skill for architects to ensure that data models remain performant when processing high volumes of records in automated triggers.

Exam trap

Candidates often confuse SOQL injection prevention with bulkification, failing to recognize that querying inside a loop rapidly exhausts governor limits regardless of data syntax.

122
MCQhard

A financial services company is migrating 5 million Account records from a legacy CRM to Salesforce. The legacy system has a field 'AnnualRevenue' stored as a string with currency symbols and commas (e.g., '$1,234,567.89'). The target Salesforce field is a Currency field with 2 decimal places. During a test load of 100,000 records using the Bulk API, the architect receives errors indicating 'Invalid currency format'. What is the most efficient way to resolve this before the full migration?

A.Modify the Salesforce field type to Text to accept the legacy format, then create a formula field to display the numeric value.
B.Pre-process the source data using an ETL tool or script to remove non-numeric characters and convert the string to a decimal number before loading.
C.Use Data Loader's 'Transform' feature to strip currency symbols and commas before loading.
D.Load the data as-is and then use a scheduled Apex job to clean up the values after import.
AnswerB

The errors occur because the Bulk API expects numeric values for Currency fields, but the source data contains symbols and commas. Pre-processing the data to remove non-numeric characters and convert to decimal ensures the data matches the target format. This is the most efficient and reliable method, as it addresses the issue at the source and avoids repeated load failures.

Why this answer

The Bulk API rejects records with non-numeric characters in Currency fields. Pre-processing the source data to remove symbols and commas and convert to decimal is the most efficient solution. It ensures data conforms to Salesforce's expected format, preventing load errors and maintaining data integrity without compromising field types or requiring post-load fixes.

Exam trap

The trap here is assuming Data Loader can transform data during load or that changing the field type is a quick fix, when the correct approach is to clean the data before migration.

123
MCQhard

Refer to the exhibit. Why might 'Parallel' concurrency mode cause errors during a large migration?

A.It causes the API to exceed the daily limit for API calls.
B.It leads to record locking contention on parent objects.
C.It prevents the data from being loaded in a specific order.
D.It forces the data to be processed in a single thread.
AnswerB

Parallel processing attempts to update records as fast as possible. If multiple threads attempt to update or insert child records that share the same parent, they will contend for the parent record lock, resulting in 'UNABLE_TO_LOCK_ROW' errors that can stop the migration process.

Why this answer

Refer to the exhibit. The JSON configuration shows a Bulk API 2.0 job using 'Parallel' concurrency mode. The Data Architect must be aware that parallel processing increases the risk of record locking when parent records have many children being processed simultaneously.

This configuration requires careful monitoring of record contention to prevent failures, even though it provides the fastest throughput for independent records.

Exam trap

Candidates often confuse parallel mode performance benefits with safety, assuming it prevents errors. They overlook how simultaneous processing of child records tied to identical parents triggers record locking contentions during high-volume data migrations.

124
MCQmedium

Why is it recommended to perform large data deletes using a soft-delete approach followed by a hard-delete during off-peak hours?

A.To keep the Recycle Bin empty at all times.
B.To allow for data recovery if a mistake is made.
C.To prevent record locking and system-wide performance degradation.
D.To ensure that all triggers are fired during the deletion.
AnswerC

Large-scale deletions cause intense row-level and table-level locking. Breaking this into stages—marking records first and deleting them in batches during off-peak windows—reduces the contention for system resources, ensuring that the database remains responsive for other users while the deletion operation progresses.

Why this answer

Large-scale deletions are resource-intensive and can trigger cascading deletes if records have child relationships. By first marking records for deletion (soft-delete), you can manage the process in controlled batches. This prevents the system from locking up during a massive operation.

Performing the actual hard-delete during off-peak hours minimizes the impact on concurrent user activity and reduces the risk of reaching governor limits during peak times.

Exam trap

Candidates often try to delete records in bulk without considering the impact of cascading deletes or the performance hit of immediate record removal, causing system timeouts.

125
MCQmedium

A large enterprise discovers duplicate accounts across Salesforce and their legacy system. Before importing data, they decide to implement a 'Data Stewardship' program. What is the primary role of a data steward in this context?

A.Writing custom Apex code to automate the merging of all duplicate accounts.
B.Managing data quality, including defining rules and resolving data conflicts.
C.Performing daily backups of the Salesforce database to ensure disaster recovery.
D.Configuring Salesforce sharing rules to restrict access to sensitive account data.
AnswerB

Data stewards are responsible for establishing the quality standards and resolving disputes that automated systems cannot handle. By managing the definitions and policies governing master data, they ensure that the organization maintains a clean, reliable, and trusted data foundation that supports informed business decision-making and efficient operations.

Why this answer

Data stewards bridge the gap between IT and business by overseeing the quality and lifecycle of master data. They define policies, resolve data conflicts, and ensure compliance with governance standards. This role is crucial because technical solutions alone cannot solve issues stemming from inconsistent business processes or ambiguous data definitions, requiring human oversight to make final decisions on data accuracy and business logic interpretation.

Exam trap

Candidates confuse the Data Steward role with a Data Engineer or Developer role, focusing on the technical implementation of data cleanup scripts rather than the business-centric governance and conflict resolution.

126
MCQmedium

A Data Architect is designing a disaster recovery plan. Which data elements must be explicitly backed up outside of Salesforce to ensure business continuity?

A.Standard Account and Contact records, as they are managed by Salesforce.
B.System Metadata, Custom Settings, and complex relational data.
C.Chatter posts and historical activity logs for all users.
D.All temporary batch job execution logs.
AnswerB

Metadata and custom settings are the 'brain' of the Salesforce organization. If these are lost or corrupted, the system's custom logic and configuration cease to function. Backing them up externally is critical because they represent the effort of years of development and are not easily restorable from standard platform backups.

Why this answer

While Salesforce provides high availability, it does not provide a point-in-time restore for individual records or deleted data if the retention period has passed. Metadata, custom settings, and high-value transactional data must be exported and stored off-platform. This strategy ensures that if a catastrophic deletion occurs or if there is a need to revert to a specific state, the organization can reconstruct its configuration and data without relying solely on the platform's standard capabilities.

Exam trap

Candidates assume Salesforce's standard backup capabilities cover everything, failing to realize that metadata and specific configuration elements (like Custom Settings) require dedicated, external backup strategies for full recovery.

127
MCQhard

Refer to the exhibit. The query is part of a bulk apex process. What is the potential risk with this query in a large-scale data environment?

A.The query fails because nested SOQL is not supported.
B.The query will cause a SOQL injection vulnerability.
C.The query will hit governor limits if the result set is too large.
D.The query is invalid because it misses the 'FROM' clause.
AnswerC

Large result sets retrieved via nested queries consume significant heap space and CPU time. If the number of contacts associated with the queried accounts exceeds memory or CPU limits, the apex process will fail. Architects must use batch Apex or limit the result set to ensure reliability in large orgs.

Why this answer

The primary risk is hitting the SOQL 'Too many query rows' limit or 'CPU time' limit when joining large related datasets. While the query looks standard, processing the nested collection of contacts for 50,000 accounts can consume significant memory and CPU cycles. Architects must implement batching and pagination to ensure the process remains within governor limits and provides stable performance.

Exam trap

Candidates often assume that standard SOQL queries automatically scale infinitely in bulk contexts without considering row limits, memory footprints, or nested subquery processing overhead in large environments.

128
MCQhard

Universal Containers maintains customer master data in Salesforce and in a legacy Oracle billing system. The data architect must define a survivorship strategy that resolves field-level conflicts when the same customer is updated in both systems during the same nightly integration window. Business rules state that the most recently modified value from the system of record for that attribute should win, and that billing-related attributes must always originate from Oracle. Which approach should the architect implement to satisfy these requirements?

A.Use Salesforce Data Loader to export both data sets nightly, then run a batch Apex job that updates Salesforce records with the Oracle values for every field, ensuring Oracle always wins.
B.Configure a Salesforce Duplicate Rule with a matching rule on the customer number, and enable Alert on the rule so that integration users receive a notification when a conflict is detected.
C.Create a Salesforce Validation Rule on the customer object that prevents updates to billing fields unless the running user is an integration user, and rely on the Oracle system to overwrite Salesforce values during the nightly job.
D.Implement attribute-level survivorship rules in the MDM layer, using source priority for billing attributes and last-updated timestamp comparison for all other attributes, then publish the golden record back to both systems.
AnswerD

Attribute-level survivorship lets each field be governed independently: billing fields are hard-coded to trust Oracle as the authoritative source, while other fields are resolved by comparing last-updated timestamps. The resulting golden record is then synchronised back to Salesforce and Oracle, keeping both systems consistent. This directly satisfies the stated business rules without relying on Salesforce-native duplicate detection, which cannot perform field-level conflict resolution.

Why this answer

The requirement combines two distinct survivorship policies: source-priority for billing attributes and last-updated-wins for everything else. Only an MDM layer that resolves conflicts at the individual attribute level can apply different rules per field and then distribute the golden record back to the source systems. Record-level duplicate detection, validation rules, or blunt overwrites cannot express that per-attribute logic and would either leave conflicts unresolved or overwrite valid newer data.

Exam trap

The trap here is assuming that Salesforce Duplicate Rules or Validation Rules can perform field-level survivorship between two systems, when they only manage record-level duplicates or edit permissions.

129
MCQmedium

During a migration, you discover that the source system has a different date format than the required ISO 8601 format for Salesforce. What is the most efficient way to handle this?

A.Use an Apex trigger to format the dates after the records are inserted.
B.Transform the data using an ETL tool or script before loading it into Salesforce.
C.Update the Salesforce field type to 'Text' to accommodate any date format.
D.Manually update the records in the Salesforce UI after the migration completes.
AnswerB

Transforming data in the ETL layer is the industry standard for migration. It keeps the Salesforce org clean and prevents errors during ingestion. Using tools such as Informatica, Mulesoft, or custom scripts allows for robust validation and formatting, ensuring the incoming data meets all schema requirements before submission.

Why this answer

Data transformation should ideally occur in the ETL (Extract, Transform, Load) layer before the data reaches Salesforce. By formatting dates during the extraction or transformation process, you ensure that the data conforms to Salesforce requirements before it hits the API. This is more scalable than attempting to force Salesforce to interpret varying formats, which can lead to data loss or import errors during the ingestion process.

Exam trap

Candidates frequently choose to handle formatting inside Salesforce via complex formula fields or triggers rather than cleaning the data upstream.

130
MCQhard

Refer to the exhibit. An architect is using the Salesforce CLI to perform a high-volume data upsert. What is the significance of the '-i Legacy_ID__c' parameter in this command?

A.It specifies the internal Salesforce ID field for the upsert.
B.It indicates the field that should be used as the primary index for PK Chunking.
C.It defines the custom External ID field used to match existing records.
D.It tells the CLI to ignore all records that have a null value in that field.
AnswerC

The '-i' flag (or --externalid) tells the Bulk API which field to use for matching the incoming data against existing Salesforce records. If a match is found based on this field, the record is updated; if no match is found, a new record is created.

Why this answer

When performing an upsert operation via the Bulk API (or CLI), the system needs a way to determine if a record already exists or if a new one should be created. The '-i' parameter specifies the External ID field that the system will use as the unique key for this matching process.

Exam trap

Candidates confuse the '-i' parameter with setting a primary key or auto-number generation, forgetting its primary purpose is matching existing records during an upsert operation.

131
MCQmedium

Which object must be loaded first in a standard Salesforce migration involving Accounts, Contacts, and Opportunities?

A.Opportunities, because they drive revenue reporting.
B.Contacts, to ensure the primary contact is defined.
C.Accounts, to establish the parent records for children.
D.It does not matter if you use the Bulk API.
AnswerC

Accounts serve as the foundation of the relational data model. By loading them first, you generate the Salesforce IDs necessary to map children (Contacts and Opportunities) correctly during subsequent loads. This top-down hierarchical approach is the industry-standard sequence for successful data migration projects in Salesforce.

Why this answer

Accounts must be loaded first because they are the parent objects for both Contacts and Opportunities in the standard data model. Without the Account records present, the foreign key lookups for Contacts and Opportunities will fail, as they have no valid parent ID to link to. Establishing the hierarchy is the fundamental requirement for data referential integrity.

Exam trap

Test-takers sometimes select Opportunities first because they represent revenue, ignoring the fundamental foreign key dependency requiring parent Accounts.

132
Multi-Selecthard

Which TWO of the following are true regarding the impact of 'Formula Fields' on Large Data Volumes?

Select 2 answers
A.Formula fields are indexed by default.
B.Filtering on formula fields causes full table scans.
C.They improve query performance by pre-calculating data.
D.Values are stored in the database for fast retrieval.
E.Complex formulas can lead to 'CPU Time Exceeded' errors.
AnswersB, E

Because formula fields are not indexed and must be calculated on the fly during a query, the Salesforce database engine must perform a full table scan to evaluate every record against the filter criteria. For objects with millions of records, this is a major performance bottleneck.

Why this answer

Formula fields are calculated at runtime, which means their values are not stored in the database. When querying or filtering on a formula field, the system must compute the value for every record, which is computationally expensive. As the data volume grows, this causes significant performance issues because the database cannot utilize indexes on formula fields (unless they are deterministic and specifically indexed), leading to slow, resource-intensive queries.

Exam trap

Candidates often assume formula fields are stored in the database and indexed like standard fields. They fail to realize that calculations happen at runtime, forcing full table scans on large datasets.

133
MCQmedium

Universal Containers notices severe locking issues and transaction failures when updating a parent Account object that has over 15,000 child records in a master-detail relationship. What is the fundamental architectural cause of this behavior?

A.The platform automatically generates criteria-based sharing recalculations for every child record.
B.Parent-level database locks occur and cascade down to evaluate child records, causing row-level contention.
C.Custom triggers on the child object execute recursively because of the master-detail cascading save mechanism.
D.Roll-up summary fields exceed maximum calculation limits when evaluating more than 10,000 records.
AnswerB

Parent-level row locks cascade to child records during master-detail updates, so the 15,000-child Account triggers lock contention across all dependent rows. This directly explains the severe locking and transaction failures described, since Salesforce acquires exclusive locks on the parent and its detail records within the same transaction.

Why this answer

Master-detail relationships enforce strict parent-level locking during record saves to maintain relational integrity and roll-up summary calculations. When a parent record is updated, the platform locks the parent and cascades lock checks across associated child records. Having thousands of children spikes the probability of database contention, leading directly to row-level locking exceptions and transaction failures.

Exam trap

Candidates frequently mistake locking issues for simple network timeouts or poor indexing, failing to recognize that master-detail relationships structurally cascade parent save locks down to thousands of child records.

134
MCQmedium

Refer to the exhibit. The Data Architect is tasked with ensuring that order status options are limited based on the geographic region of the order. Which feature should be used to implement this?

A.Record Types.
B.Validation Rules.
C.Dependent Picklists.
D.Custom Metadata Types.
AnswerC

Dependent picklists provide an intuitive UI experience by dynamically filtering available options based on the controlling field. This is the most efficient and standard way to restrict picklist values, ensuring users only select valid combinations without needing custom code or complex validation logic to maintain the data model.

Why this answer

Dependent picklists are the standard Salesforce mechanism for limiting the values of one picklist based on the selection of another. By defining a controlling field (Region) and a dependent field (Status), the architect ensures data entry accuracy and prevents invalid combinations of status and geography from being saved, which is vital for regional reporting and process consistency.

Exam trap

Candidates often look for complex automation or custom code solutions, failing to recognize that standard Dependent Picklists provide an out-of-the-box solution for filtering field values.

135
MCQhard

Universal Containers maintains a custom object, Shipment__c, that has a Master-Detail relationship to Account and a Lookup to Carrier__c. A Data Architect is asked to design a new object, Shipment_Line__c, that must always belong to exactly one Shipment__c, must inherit Shipment__c's sharing rules, and must not be independently owned by a different user. Which relationship type should be used between Shipment_Line__c and Shipment__c?

A.Hierarchical relationship, using the User object to control ownership and access to Shipment_Line__c.
B.Lookup, with a required field to Shipment__c and a sharing rule that grants access based on the Shipment__c owner.
C.Self-relationship on Shipment__c, with a field pointing to the parent Shipment__c record.
D.Master-Detail, with Shipment__c as the master and Shipment_Line__c as the detail.
AnswerD

Master-Detail makes the detail record dependent on the master, so a Shipment_Line__c cannot exist without a Shipment__c. The detail inherits the master's organization-wide defaults and sharing, and the detail's owner is effectively the master's owner, so it cannot be independently owned. This matches the stated requirements for containment, sharing inheritance, and ownership.

Why this answer

The requirement is true containment with inherited sharing and no separate ownership. A Master-Detail relationship enforces that the detail cannot exist without the master, rolls up the master's sharing and organization-wide defaults, and ties the detail's ownership to the master. A required lookup can approximate containment but leaves the child independently owned and separately shared, which violates the stated constraints.

Exam trap

The trap here is assuming that a required Lookup is functionally equivalent to Master-Detail for containment and sharing inheritance.

136
MCQhard

During a large data migration, a developer notices that existing Apex triggers are causing severe performance degradation. What is the most effective way to prevent trigger execution without deleting the code?

A.Comment out the trigger logic and redeploy the code.
B.Disable the triggers in the Setup menu under 'Manage Triggers'.
C.Use a Custom Setting to control trigger execution.
D.Remove user permissions for the integration user.
AnswerC

A Custom Setting (or Custom Metadata) provides a clean, highly performant way to toggle Apex execution logic. This allows developers to disable resource-intensive triggers temporarily for the duration of the data load and reactivate them immediately afterward without needing a full code deployment cycle.

Why this answer

The most effective method to mitigate trigger-related performance issues during high-volume loads is to implement a 'Bypass' or 'Switch' pattern using Custom Settings or Custom Metadata. By checking this flag at the start of the trigger, logic can be conditionally skipped. This preserves code integrity while allowing the migration to proceed at maximum speed without the overhead of complex validation or automated process logic.

Exam trap

Candidates often select architectural solutions like permanently deleting or rewriting trigger logic, failing to recognize that configuration-based bypass mechanisms using custom settings are much safer and more efficient.

137
Multi-Selecthard

A Data Architect is designing a high-volume data model where an Account has millions of Child records. Which TWO strategies should be implemented to ensure optimal performance and avoid data skew?

Select 2 answers
A.Ensure that the owner of the parent records is not a single user or small group.
B.Convert all Lookup relationships to Master-Detail relationships to improve query speed.
C.Avoid using custom indexes on fields that have high cardinality.
D.Avoid creating parent-child relationships where the parent is a 'Person Account'.
E.Index the foreign key fields used for reporting and filtering.
AnswersA, E

When one user owns a large volume of child records, record locking can occur during updates to the parent record. Distributing ownership across multiple users helps alleviate contention, as the platform's locking mechanism is often tied to the parent record's owner and sharing rules in the underlying database.

Why this answer

Performance issues in Salesforce often arise from data skew, where a small number of parent records own a vast majority of child records. Implementing strategies like record ownership distribution and utilizing indexing on high-cardinality fields prevents row-level locking contention. These design patterns are critical for maintaining query performance and ensuring that API operations do not time out during large-scale data processing or complex analytical reporting cycles.

Exam trap

Candidates often propose increasing batch sizes or adding more hardware, failing to realize that data skew is a structural issue requiring redistribution of ownership or proper indexing.

138
MCQmedium

A Salesforce data architect at a global manufacturer is defining the master data domain for 'Customer'. The company has separate business units that each maintain their own customer records. The architect must ensure that a single, authoritative Customer record exists and is referenced by all transactional systems. Which approach best achieves this?

A.Use Salesforce Connect to virtualize customer data from each business unit system in real time.
B.Create a single Account record type and enforce a unique external ID on the Account object.
C.Implement Salesforce Duplicate Management with matching rules and duplicate rules.
D.Define a master data domain with a golden record, survivorship rules, and a data stewardship process.
AnswerD

A master data domain requires a golden record that is the single source of truth, survivorship rules to determine which source wins in conflicts, and stewardship to maintain quality. This ensures all transactional systems reference the authoritative Customer record. It directly addresses the need for a single, authoritative record.

Why this answer

A master data domain is defined by a golden record, survivorship rules, and data stewardship. The golden record is the single authoritative version of a Customer, and survivorship rules determine which source system's data prevails when attributes conflict. Data stewardship ensures ongoing quality.

This combination ensures all transactional systems reference the same authoritative Customer record, which is the goal.

Exam trap

The trap here is assuming that duplicate prevention or real-time virtualization alone creates a master data domain, when a golden record and survivorship are essential.

139
MCQmedium

An organization has 50 million records in a custom object. They need to report on historical snapshots weekly. What is the best strategy to manage storage while keeping data available?

A.Use standard Salesforce objects
B.Use Big Objects
C.Move data to a sandbox
D.Use Platform Events
AnswerB

Big Objects allow for the storage and querying of massive datasets without impacting performance or standard data storage limits. They are specifically engineered to handle high volumes, making them the optimal choice for weekly snapshots that need to be queried but not frequently updated or edited.

Why this answer

Big Objects are designed for massive scale, supporting billions of records with optimized storage costs. By moving historical snapshot data to Big Objects, the organization retains accessibility for reporting via asynchronous queries or external tools without consuming standard Salesforce storage limits. This strategy preserves database performance for the active operational records while ensuring compliance and historical availability for long-term analytical business requirements.

Exam trap

Test-takers frequently choose standard reporting snapshots or standard archival objects, ignoring the massive scale of 50 million records which will quickly breach standard data storage limits.

140
MCQmedium

A company maintains a custom object Asset__c with 5 million records. They need to archive records older than 7 years to a Big Object for compliance. After archiving, the records must be queryable via SOQL for audits, but not editable. Which approach should an architect recommend?

A.Create a Big Object with the same fields, use a batch Apex job to copy old records, and then delete them from Asset__c. Provide audit access via a Lightning component that uses Async SOQL to query the Big Object.
B.Create a Big Object with the same fields as Asset__c, use a batch Apex job to copy old records, and then delete them from Asset__c. Provide audit access via a Visualforce page that queries the Big Object using SOQL.
C.Create a custom object Asset_Archive__c to store old records, and use a batch Apex job to move them. Provide audit access via standard reports and list views.
D.Use Salesforce Connect to create an external object that points to an external database containing the archived records, and provide audit access via external object list views.
AnswerA

Big Objects are designed for large-scale archival and are queried using Async SOQL, which runs asynchronously and can handle massive datasets. Copying old records to a Big Object and deleting them from Asset__c reduces the size of the transactional object, improving performance. Async SOQL is the correct mechanism for querying Big Objects and can be invoked from a Lightning component, satisfying the audit requirement.

Why this answer

Big Objects are the Salesforce-native solution for archiving massive datasets. They support high-volume storage and are queried asynchronously using Async SOQL, which is suitable for audit scenarios where real-time access is not required. Copying old records to a Big Object and deleting them from the transactional object reduces the object's size, improving query and DML performance.

A Lightning component can invoke Async SOQL to retrieve archived records for audits.

Exam trap

The trap here is assuming that Big Objects can be queried with standard SOQL like other objects, when they actually require Async SOQL or a custom index-based query.

141
MCQmedium

Northern Trail Outfitters is preparing for a GDPR-driven data subject deletion request affecting a Contact and its related Cases, Email Messages, and custom child records. The architect must ensure the deletion is complete and evidenced. Which approach best satisfies the governance requirement?

A.Use the native Data Subject Request (DSR) tooling to locate, and then erase the individual's data across related objects, capturing the action for evidence
B.Export the individual's data to a CSV, delete the Contact, and archive the CSV in a shared drive
C.Overwrite the Contact fields with null values using Data Loader and leave related records intact
D.Delete the Contact record and rely on the recycle bin to remove related records
AnswerA

Salesforce provides Data Subject Request tooling that finds an individual's data across related records and supports erasure while producing a record of the action. It addresses the cross-object scope of the request and provides the evidence trail governance requires, unlike a manual delete.

Why this answer

A GDPR erasure request spans every object holding the individual's data, and governance requires proof the action occurred. The native Data Subject Request tooling locates data across related objects, supports erasure, and records the activity, satisfying both completeness and evidence. Manual deletes, field nulling, and off-platform exports each leave residual data or create ungoverned copies.

Exam trap

The trap here is treating a parent record delete or a field-nulling exercise as equivalent to a lawful erasure when related objects and the recycle bin still retain the individual's data.

142
MCQmedium

When designing a system for LDV, what is the role of an 'Indexed Field'?

A.To increase storage capacity in the Salesforce database.
B.To speed up the retrieval of records by narrowing search space.
C.To enforce uniqueness constraints on every custom field.
D.To automate the conversion of records to external objects.
AnswerB

Indexes function like a book index, allowing the database to jump straight to the data rows that match the filter criteria. This eliminates the need for full table scans, which are prohibitively slow on tables with millions of records, and is vital for maintaining responsive application performance.

Why this answer

Indexes are structures that the database uses to find records quickly without scanning the entire table. In an LDV environment, they are the primary defense against query timeouts. By ensuring that frequently used filters in SOQL queries are indexed, architects can significantly reduce the amount of data the database engine must process, keeping performance within acceptable levels regardless of the total record count.

Exam trap

Test-takers frequently assume that creating an index will automatically speed up every type of query, ignoring that indexes only help specific filter conditions and can actually degrade write performance.

143
MCQeasy

A data architect is migrating 500,000 records into a custom object that has a validation rule requiring the 'Status__c' field to be either 'Active' or 'Inactive'. The legacy data contains some records with a blank Status__c. The architect needs to ensure the migration succeeds without modifying the validation rule. What should be done?

A.Temporarily deactivate the validation rule during the data load, then reactivate it afterward.
B.Load the records with blank Status__c using the Data Loader's 'Bulk' mode, which bypasses validation rules.
C.Transform the legacy data by setting a default value of 'Inactive' for any records with a blank Status__c before loading.
D.Create a before insert trigger to set a default value for Status__c when it is blank.
AnswerC

This is correct because the validation rule requires a specific value. Transforming the data to populate a valid default ensures that all records pass validation. This approach is simple, does not require changing the validation rule, and maintains data quality. It is a standard data cleansing step in migration projects.

Why this answer

The validation rule requires Status__c to be 'Active' or 'Inactive'. The most straightforward solution is to transform the legacy data by setting a default value of 'Inactive' for blank entries. This ensures all records pass validation without altering the validation rule or adding custom code.

Data transformation during migration is a best practice to meet target system requirements and maintain data integrity.

Exam trap

The trap here is believing that Bulk API bypasses validation rules or that deactivating rules is an acceptable workaround, when the requirement explicitly states not to modify the validation rule.

144
MCQmedium

Which document is the most critical for Data Governance to ensure that Salesforce Orgs remain compliant with regional data residency requirements?

A.User Access Request Log.
B.Data Inventory and Mapping.
C.Service Level Agreement (SLA).
D.Salesforce Org release notes.
AnswerB

A data inventory maps the flow and storage location of data across the enterprise. It is essential for residency compliance because it identifies where data originates, moves, and resides. Without this, an organization cannot prove to regulators that they are keeping data within the required legal jurisdictions or borders.

Why this answer

The Data Inventory and Mapping document is critical because it identifies exactly where sensitive data resides. To comply with data residency laws, architects must know which Org holds data and where those Orgs are physically hosted. Without this mapping, it is impossible to audit compliance or implement the necessary architectural controls to keep data within specific geographic borders as required by law.

Exam trap

Many respondents pick technical documents like ERDs or security guides, failing to recognize that compliance tracking requires specific data location mapping.

145
MCQhard

A healthcare provider is migrating 8 million legacy Patient_Visit__c records into Salesforce. Each visit references a legacy numeric provider code, and the legacy system does not expose a stable unique identifier for each provider. The Salesforce Provider__c object already has a custom external ID field named Legacy_Provider_Code__c that is populated. The architect wants to load visits without first exporting provider Salesforce IDs. Which approach should be used?

A.Create a junction object between Provider__c and Patient_Visit__c, load the junction records with both legacy codes, and let Salesforce infer the lookups.
B.Load Patient_Visit__c records with a lookup field populated by the legacy provider code, relying on the external ID field to resolve the relationship during the load.
C.Load Patient_Visit__c records into a staging custom object, then use a scheduled Apex job to copy each legacy provider code into a text field and later reconcile manually.
D.Export all Provider__c records with their Salesforce IDs, VLOOKUP the IDs into the visit file, then load visits with the 18-character Salesforce ID in the lookup field.
AnswerB

Salesforce resolves relationship fields by matching the value supplied in the relationship field against the target object's external ID field, so loading visits with the legacy provider code in the Provider__c relationship maps each visit to the correct Provider__c record without a prior ID export. This is exactly the pattern Bulk API 2.0 and Data Loader support for external ID lookups, and it scales for multi-million record loads.

Why this answer

Relationship fields can be populated with an external ID value when the target object has that field marked as an external ID, letting Salesforce match and set the lookup during the load. Because Provider__c already has Legacy_Provider_Code__c configured as an external ID and populated, the visit load can reference the provider by that code directly, eliminating a separate ID export and join step and keeping the migration efficient at scale.

Exam trap

The trap here is assuming that a lookup must always be populated with the 15- or 18-character Salesforce record ID, when an external ID on the parent object can be used instead.

146
Multi-Selectmedium

A Salesforce architect is planning a migration of 2 million custom object records from a legacy Oracle database. The legacy system has a field 'Status' with values like 'Active', 'Inactive', 'Pending', and 'Archived'. In Salesforce, the Status field is a picklist with values 'Active', 'Inactive', and 'Pending' only. The architect needs to migrate the data while preserving as much information as possible and ensuring data quality. Which two actions should the architect take? (Choose two.)

Select 2 answers
A.Add 'Archived' as a new picklist value in Salesforce to accommodate the legacy data without transformation.
B.Create a custom text field to store the original legacy status value for audit purposes, in addition to mapping to the picklist.
C.Use a validation rule to prevent 'Archived' from being loaded, and manually update those records after migration.
D.Map the legacy 'Archived' value to 'Inactive' in Salesforce and document the transformation for future reference.
E.Exclude records with 'Archived' status from the migration to avoid picklist conflicts.
AnswersB, D

Storing the original legacy status in a separate text field preserves the exact source value for auditing and reporting, while the picklist field is used for operational purposes. This dual approach ensures data quality and traceability without compromising the picklist's integrity. It is a best practice when transforming data during migration to retain historical context.

Why this answer

The architect should map the unsupported 'Archived' value to a valid picklist value like 'Inactive' to maintain data integrity and avoid load failures. Additionally, storing the original legacy status in a separate text field preserves the exact source information for auditing. This combination ensures data quality, traceability, and compliance with Salesforce picklist constraints.

Exam trap

The trap here is assuming that adding a new picklist value or excluding records is acceptable, when the better approach is to transform the data and preserve the original value for audit.

147
MCQhard

A Salesforce org has a custom object Shipment__c with 12 million records. Users frequently run a SOQL query that filters on Status__c and Order_Date__c, and sorts by CreatedDate. The query is timing out. The org has an index on Status__c and a custom index on Order_Date__c. What is the most likely reason the query is still slow?

A.The query is slow because it sorts by CreatedDate, which is not indexed, and Salesforce cannot sort large result sets without an index on the sort field.
B.The query is non-selective because the combination of filters does not reduce the result set below the selectivity threshold, and sorting by CreatedDate requires additional processing.
C.The custom index on Order_Date__c is not used because Salesforce does not support custom indexes on date fields.
D.The query is slow because Status__c is a picklist field, and picklist fields cannot be indexed.
AnswerB

Salesforce selectivity thresholds depend on the object size; for 12 million records, a query must filter to a small percentage to use an index. Even with indexes on Status__c and Order_Date__c, if the filters are not restrictive enough, the optimizer may choose a full scan. Sorting by CreatedDate adds a sort operation that can further degrade performance when the result set is large. This combination explains the timeout.

Why this answer

For a 12-million-record object, selectivity is the critical factor. Salesforce uses a threshold based on object size to decide whether to use an index. Even with indexes on Status__c and Order_Date__c, if the filters return too many rows, the optimizer performs a full scan.

Sorting by CreatedDate adds overhead. The correct diagnosis is that the query is non-selective and the sort compounds the cost.

Exam trap

The trap here is assuming that having indexes on individual fields guarantees query performance, when the real determinant is whether the combined filters meet the selectivity threshold for that object's size.

148
MCQmedium

Universal Containers has 50 million records in a custom object. A report summarizing these records times out. Which strategy is most effective for improving report performance?

A.Increase the report timeout limit in Setup.
B.Implement a custom lightning web component for reporting.
C.Create a summary object to store pre-aggregated data.
D.Add more indexes to the custom object fields.
AnswerC

Pre-aggregation shifts the computational load from read-time to write-time. By calculating metrics asynchronously and storing them in a summary object, the report engine only processes a small number of aggregate records rather than scanning the entire base table, significantly reducing execution time and resource consumption.

Why this answer

Summarizing 50 million records directly in a standard report is inefficient due to the overhead of the reporting engine. Using a Summary Object allows Salesforce to pre-aggregate data asynchronously, moving the heavy lifting from query time to insert time. This approach ensures reports operate on a significantly smaller dataset, preventing timeout errors and improving user experience for large-scale data analysis scenarios.

Exam trap

Candidates often suggest using standard Salesforce reports or roll-up summary fields on large objects, forgetting that standard reports time out on millions of rows and roll-ups have hard limits.

149
MCQmedium

What is the primary role of a Data Dictionary in a Data Governance program?

A.Enforcing record-level security.
B.Establishing a standard metadata definition.
C.Automating data cleaning processes.
D.Managing integration authentication tokens.
AnswerB

The data dictionary serves as the single source of truth for metadata definitions. By standardizing how data is described and used across the organization, it eliminates inconsistencies. This clarity is vital for developers and business users alike, as it ensures they all interpret the data in exactly the same way.

Why this answer

A Data Dictionary acts as a central repository for technical and business metadata. It provides a common language for stakeholders, ensuring that everyone understands what a field represents, its data type, and its business rules. By centralizing this information, the dictionary prevents ambiguity and errors in reporting, which is a fundamental requirement for building a reliable, governed data environment.

Exam trap

Candidates often mistake the Data Dictionary for a 'Data Catalog' or 'Data Warehouse,' overlooking its primary purpose as a centralized repository for metadata definitions and common business language.

150
MCQeasy

A data architect needs to migrate 500,000 Account records from a legacy system into Salesforce. The legacy system uses a 15-digit alphanumeric string as the unique identifier. The architect must ensure that future updates can match existing Accounts without creating duplicates. Which field configuration should be used on the Account object to support this requirement?

A.Create a custom Text field of length 15 and use it as the Salesforce Record ID.
B.Create a custom Text field of length 15 and make it an External ID.
C.Use the standard Account Number field and set it as an External ID.
D.Create a custom Text field of length 15 and mark it as Unique.
AnswerB

An External ID field is specifically designed to store unique identifiers from external systems and is used by the upsert operation to match records. Making the custom Text field an External ID allows the architect to use upsert with this field as the match key, preventing duplicates during future updates. The length of 15 accommodates the legacy identifier.

Why this answer

To support upsert and prevent duplicates, the legacy identifier must be stored in a field marked as an External ID. A custom Text field of sufficient length, designated as External ID, allows the Data Loader or Bulk API to match existing records based on that field. This is the standard pattern for integrating external systems and maintaining data integrity.

Exam trap

The trap here is confusing the Unique attribute with the External ID attribute; only External ID enables upsert matching.

Page 1

Page 2 of 3

Page 3

All pages