Courseiva
Data Modeling →mediumMultiple Choice

C100DEV Data Modeling Practice Question

What is the primary advantage of the Subset pattern in MongoDB data modeling?

⚠ Common exam trap

Candidates often believe the Subset pattern is about security or data masking. They miss that its primary purpose is optimizing the working set to fit into the WiredTiger cache.

Answer choices

Why each option matters

Answer the question above first, then reveal the full breakdown to understand why each option is right or wrong.

Correct answer & explanation

✓

It keeps the working set small to improve cache performance.

The Subset pattern allows you to store the most frequently accessed data in the main document while offloading the rest to another collection. This ensures that the primary document remains small and fits efficiently into the WiredTiger cache, which significantly boosts read performance for common queries. By reducing the overall document size, you also avoid the 16MB BSON limit while still keeping essential information easily accessible for the application's most frequent use cases.

Answer analysis

Option-by-option breakdown

For each option: why learners choose it and why it is or isn't the right answer here.

  • ✗

    It eliminates the need for indexes on the collection.

    Why it's wrong here

    Indexes remain essential regardless of the pattern chosen. Even with a subset of data, you need indexes to ensure that queries can quickly locate the specific documents required. The Subset pattern is a storage strategy, not an indexing one, and it does not remove the performance requirements for indexing.

  • ✓

    It keeps the working set small to improve cache performance.

    Why this is correct

    By limiting the document size to only the most essential data, you maximize the number of documents that can reside in the WiredTiger cache. A smaller working set means fewer disk I/O operations, leading to faster query response times and higher overall throughput for the database cluster's performance.

  • ✗

    It automatically scales the database across sharded clusters.

    Why it's wrong here

    Scaling and sharding are infrastructure concerns handled by the cluster configuration, not by the data modeling pattern. While a well-modeled schema helps with sharding by providing good shard keys, the Subset pattern itself does not provide automatic horizontal scaling or partition data across multiple physical shards by itself.

  • ✗

    It prevents data duplication across different collections.

    Why it's wrong here

    The Subset pattern can actually lead to some degree of data duplication, as some fields may need to exist in both the primary document and the overflow collection. The primary goal is optimization of read performance, not strictly the normalization of data or the total avoidance of data duplication.

About these practice questions

This C100DEV question is part of Courseiva's 259-question bank — original exam-style content with full explanations and wrong-answer analysis, never real exam questions or exam dumps. Learn why practice questions differ from exam dumps →

How Courseiva writes practice questions · Editorial policy

JA

Written and reviewed by Johnson Ajibi, MSc IT Security

Senior Network & Security Engineer · founder of Courseiva

Last reviewed September 2026 · checked against the official MongoDB exam blueprint

This C100DEV practice question is part of Courseiva's free MongoDB certification practice question bank. Courseiva provides original exam-style practice questions with explanations, topic-based practice, mock exams, readiness tracking, and study analytics to help learners prepare for the C100DEV exam.