You must map a database symptom to the right AWS monitoring tool: CloudWatch metrics and alarms for thresholds, Performance Insights for query load, Enhanced Monitoring for OS metrics, and DMS with CloudWatch for migrations. The key skill is choosing the correct metric or service for the stated problem.
Start practicing
Monitoring and Troubleshooting — choose a session length
Free · No account required
Domain overview
The Monitoring and Troubleshooting domain covers how you observe, alert on, and diagnose Amazon RDS, Aurora, DynamoDB, and other AWS database services. Questions present a symptom — low storage, memory pressure, connection saturation, failed migration — and ask which CloudWatch metric, log, or AWS service you use to detect or resolve it.
Exam objectives
Selecting CloudWatch metrics and alarms such as FreeStorageSpace, FreeableMemory, and DatabaseConnections for RDS monitoring
Using Enhanced Monitoring, Performance Insights, and RDS log exports to diagnose query and resource bottlenecks
Applying AWS DMS and CloudWatch to track migration progress and validate data consistency
Using CloudWatch Logs, CloudTrail, and AWS Config to investigate database events and configuration changes
Confusing FreeStorageSpace (bytes remaining) with a percentage threshold; alarms must be set on the raw metric, not a percent of allocated storage.
Reaching for Performance Insights when the question asks for OS-level CPU, memory, or disk metrics, which Enhanced Monitoring provides.
Assuming CloudWatch alone validates migrated data; AWS DMS task logs and validation, plus Schema Conversion Tool, are needed for consistency checks.
Click any question to see the full explanation and answer options, or start a focused practice session above.
A company is using Amazon RDS for MySQL and notices that the Read IOPS metric is consistently high during business hours. The application is read-heavy. Which configuration change would most likely reduce Read IOPS?
2A developer reports that an application using Amazon DynamoDB is experiencing high latency during peak hours. The table has a provisioned capacity of 500 read capacity units (RCUs) and 500 write capacity units (WCUs). The application uses eventually consistent reads and the table is about 50 GB. The developer notices throttled write requests in CloudWatch. Which action would most effectively reduce write throttling?
3A company is migrating an on-premises Oracle database to Amazon RDS for Oracle. During the migration, the database administrator notices that the CPU utilization on the RDS instance is consistently above 90% during peak hours, even though the on-premises server had similar specifications. The application queries are mostly SELECT statements with occasional DML. The RDS instance is db.r5.large with 500 GB of General Purpose SSD (gp2) storage. Which change would most likely reduce CPU utilization?
4A team manages an Amazon Aurora MySQL database. They observe that the 'Deadlocks' metric in CloudWatch is spiking. The application uses a single writer instance and multiple read replicas. Which action is most effective at reducing deadlocks?
5A database engineer is troubleshooting slow query performance on an Amazon RDS for PostgreSQL instance. The instance is db.r5.large with 500 GB of General Purpose SSD (gp2) storage. CloudWatch metrics show high Read Latency and high Read IOPS, but low CPU utilization. Which TWO actions should the engineer take to improve performance?
6A company is using Amazon DynamoDB with autoscaling enabled. The table has a partition key of 'order_id' and a sort key of 'order_date'. The application performs both point queries and range queries. Recently, the 'ConsumedReadCapacityUnits' metric shows that the table is consistently using 100% of the provisioned capacity. Which THREE factors should the database engineer investigate to determine the cause?
7A company is using Amazon RDS for MySQL and notices that database connections are being rejected intermittently. The application logs show 'Too many connections' errors. The DB instance has 1000 max_connections. Which action should the DBA take to troubleshoot and resolve this issue without impacting performance?
8A company is running a production Amazon DynamoDB table with on-demand capacity. The application is experiencing increased latency and throttled requests during peak hours. Which monitoring tool should the database specialist use to identify the specific partition keys causing the throttling?
9A database engineer is monitoring an Amazon RDS for PostgreSQL instance and notices that the 'DiskQueueDepth' metric is consistently above 100. The instance uses gp2 storage with 1000 GB allocated. What is the most likely cause of the high disk queue depth?
10A company is migrating its on-premises Oracle database to Amazon RDS for Oracle. The database specialist needs to monitor the migration process and ensure data consistency. Which TWO AWS services should be used together to continuously monitor the replication lag and data integrity?
11A database specialist is troubleshooting an Amazon DynamoDB table that is experiencing high throttling on write requests. The table has on-demand capacity and uses a composite primary key (partition key and sort key). Which TWO actions should the specialist take to identify and resolve the issue?
12A database engineer is reviewing Amazon RDS for MySQL error logs and sees repeated authentication failures from the same IP address. The application team confirms the password is correct. What is the most likely cause of these errors?
13A company runs a critical e-commerce application on Amazon Aurora MySQL with a single DB instance. The database has 8 TB of data and uses the default writer endpoint. Recently, the application experienced a 10-minute outage during a primary instance failover. The failover was triggered by an underlying hardware issue. The database specialist needs to minimize downtime during future failovers. The application team is unwilling to modify the application code to handle connection retries. The company has a 99.99% SLA requirement. Which solution should the database specialist implement to meet the SLA with minimal application changes?
14A company runs a production Amazon DynamoDB table with on-demand capacity. The table stores session data for a web application. Recently, users have reported occasional slow response times. The operations team notices that the table's ConsumedWriteCapacityUnits metric shows occasional spikes that exceed the provisioned throughput (though on-demand auto-scales), and ThrottledWriteEvents metrics show occasional throttling. The application uses the AWS SDK with default retry logic. The database specialist is asked to investigate. Upon reviewing the table configuration, the specialist finds that the table has a simple primary key (partition key only) and the data access pattern is heavily skewed toward a small number of partition keys. The application writes in batches of 25 items using the BatchWriteItem API. What should the specialist recommend to reduce throttling and improve performance?
15Arrange the steps to set up cross-Region read replicas for an Amazon Aurora MySQL DB cluster in the correct order.
16A company uses Amazon ElastiCache for Redis as a caching layer for a web application. They observe a sudden increase in CPU utilization on the cache cluster, and the application experiences higher latency. Which action should be taken to diagnose the issue?
17A company runs an Amazon Aurora MySQL database cluster with one writer and two readers. The application suddenly fails with 'Too many connections' error. The writer instance's maximum connections is set to 1000. Which configuration change would best resolve the issue while maintaining high availability?
18A company uses Amazon DynamoDB with global tables. They notice that changes made in one region are not appearing in another region after several minutes. Which CloudWatch metric should be monitored to check the replication lag?
19A team is troubleshooting an Amazon RDS for SQL Server instance that is running out of storage. The instance uses General Purpose SSD (gp2) storage. The team wants to increase storage without downtime. Which action should they take?
20A developer is using AWS Database Migration Service (DMS) to migrate a database from on-premises to Amazon RDS. The migration task is failing with 'Insufficient memory' error. Which resource should be increased to resolve this?
21A developer is trying to connect to an RDS for PostgreSQL instance using the endpoint shown in the exhibit. The connection fails with a timeout. Which of the following is the most likely cause?
22A company is monitoring an Amazon Aurora MySQL DB cluster. They observe that the AuroraReplicaLagMaximum metric is consistently above 10 seconds. Which action would best reduce the replica lag?
23A database administrator notices that an Amazon RDS for MySQL instance is using 100% of its allocated storage. Which action should be taken first to prevent the instance from becoming inaccessible?
24A company uses Amazon CloudWatch to monitor an RDS for Oracle instance. They want to receive an alert when the database connection count exceeds 90% of the maximum connections. Which CloudWatch metric should be used to create the alarm?
25A developer executed a DELETE statement without a WHERE clause on an Amazon RDS for PostgreSQL instance. The transaction is still open. Which action should the developer take to undo the DELETE without affecting other operations?
26A company is using Amazon DynamoDB as the primary database for a global e-commerce application. During the holiday season, the application experiences throttling on write requests even though the read and write capacity units are well below the provisioned limits. The table uses on-demand capacity mode. What is the most likely cause of this throttling?
27A database administrator notices that an Amazon RDS for SQL Server DB instance has been in the 'storage-optimization' state for several hours after modifying the storage type from gp2 to io1. What should the administrator do to resolve this?
28A development team is using Amazon RDS for MySQL with read replicas to offload reporting queries. They notice that the read replica is consistently lagging behind the primary by several seconds. The primary handles 5000 writes per second. Which action would most likely reduce replica lag?
29A team is using Amazon RDS for Oracle with an option group that includes the Oracle Enterprise Manager (OEM) option. After modifying the option group to add a new option, the DB instance is stuck in the 'modifying' state for an extended period. What should the team do?
30An administrator is troubleshooting an Amazon RDS for PostgreSQL instance that is experiencing high CPU utilization. The administrator has enabled Performance Insights. Which metric should be examined first to identify the queries consuming the most CPU?
31Which TWO AWS services can be used to monitor the performance of an Amazon DynamoDB table and send alerts when throttling occurs? (Choose two.)
32A company is running a production Amazon RDS for MySQL DB instance. The application team reports intermittent high latency and connection timeouts. A quick check shows that the DB instance's CPU utilization is consistently above 90% during peak hours. The database size is 500 GB and the instance class is db.r5.large. Which combination of actions should a database specialist take to resolve the performance issue?
33A company is using Amazon DynamoDB for a gaming leaderboard application. Recently, users have experienced increased latency when updating scores. The DynamoDB table has on-demand capacity mode. The application performs UpdateItem calls with a condition expression. Which action is most likely to reduce the latency?
34A database specialist needs to monitor the number of deadlocks occurring in an Amazon RDS for SQL Server DB instance. Which CloudWatch metric should be used?
35A company is using Amazon Redshift for data warehousing. The data engineering team notices that queries are taking longer than expected. The cluster has two nodes of type dc2.large. The database specialist checks the system tables and finds that many queries are using the disk for temporary storage. Which action should the specialist take to improve query performance?
36A database specialist is troubleshooting an Amazon RDS for MySQL DB instance that is running out of storage. The instance has automated backups enabled. The specialist needs to free up storage space immediately without losing backup capability. Which action should the specialist take?
37Refer to the exhibit. A database specialist is investigating performance degradation on an Amazon RDS for MySQL DB instance. The BurstBalance metric shows the values above. What does this indicate, and what action should be taken?
38A company is running an Amazon RDS for MySQL database. The application team reports that the database is slow. Upon investigation, you notice that the DB instance's CPU utilization is consistently above 90%. Which initial troubleshooting step should you take?
39A developer is troubleshooting an application that uses Amazon DynamoDB. The application is experiencing throttled requests (ProvisionedThroughputExceededException). Which CloudWatch metric should be monitored to troubleshoot this issue?
40A company's Amazon RDS for PostgreSQL instance is experiencing a high number of connections, causing performance degradation. The DBA wants to identify which user and application are creating the most connections. What should the DBA do?
41A company is running an Amazon RDS for Oracle database in Multi-AZ. The primary instance fails over unexpectedly. The DBA wants to determine the cause of the failover. What should the DBA do?
42A company is using Amazon ElastiCache for Redis as a caching layer. The application performance degrades when cache misses increase. Which metric should be monitored to track the cache hit rate?
43A database administrator notices that an Amazon RDS for MySQL DB instance is using more storage than expected. Which metric should be monitored to troubleshoot storage usage?
44A company is running an Amazon DocumentDB cluster. The application is experiencing high write latency. The cluster has a single instance. What should be done to identify the cause of the latency?
45A database specialist notices that an RDS MySQL instance's FreeableMemory metric is consistently below 100 MB. Which monitoring tool should be used to identify the queries consuming the most memory?
46A company notices that its Aurora MySQL cluster has a high number of locks and deadlocks. The application uses read replicas for read scaling. What is the MOST likely cause?
47A company is using Amazon ElastiCache for Redis as a caching layer for a web application. Users report that some cached data is missing, causing slower responses. Which ElastiCache feature should be checked first to understand key evictions?
48A database administrator wants to receive an alert when an RDS instance's storage space drops below 10% of total allocated storage. Which AWS service should be used to set up this alert?
49Which TWO metrics should be monitored to detect a memory leak in an RDS for Oracle instance? (Choose 2.)
50Which THREE actions should be taken to troubleshoot an Amazon RDS for PostgreSQL instance that is unresponsive? (Choose 3.)
51An Amazon RDS for Oracle instance is experiencing high swap usage. Which metric should be monitored to determine if the instance is memory-constrained?
52A company is using Amazon Neptune and notices that some queries are slow. The DBA wants to identify which queries consume the most time. Which feature should be used?
53A company is using Amazon ElastiCache for Redis and notices that the cache hit ratio is low. The application is frequently reading data that is not in the cache. Which action would be most effective in improving the cache hit ratio?
54An Amazon RDS for MySQL instance is running out of storage. Which TWO actions can be taken to resolve this issue without downtime?
55A database administrator notices that an Amazon RDS for MySQL DB instance's CPU utilization is consistently above 90% during peak hours. Which initial troubleshooting step should the administrator take?
56A startup uses Amazon ElastiCache for Redis as a caching layer for its database. Users report that application responses are slow. The developer checks the ElastiCache metrics and sees that 'CacheHits' are low and 'CacheMisses' are high. What is the most likely cause?
57A company is running an Amazon RDS for PostgreSQL DB instance with Multi-AZ deployment. They notice that the primary DB instance is experiencing high CPU utilization. The read replica shows normal CPU. Which action should the DBA take to reduce the load on the primary instance?
58A developer is troubleshooting a slow query on Amazon RDS for MySQL. The query joins three large tables and runs frequently. What is the most effective way to identify the bottleneck?
59A company runs an Amazon Aurora MySQL DB cluster with one writer and two readers. They notice that one reader instance is consistently showing higher than expected lag. The other reader is fine. What is the most likely cause?
60A DBA is investigating a sudden increase in database connections to an Amazon RDS for SQL Server instance. The application is running on Amazon EC2 instances behind an Application Load Balancer. Which tool can provide real-time information about active connections?
61A company runs a document database using Amazon DocumentDB. They notice that some queries are taking much longer than expected. The explain plan shows a COLLSCAN. Which action would most improve query performance?
62A company uses Amazon DynamoDB global tables for a multi-region application. They notice that writes in one region are not appearing in another region after several minutes. What should they check first?
63Which TWO metrics should be monitored to detect an Amazon RDS for MySQL instance that is experiencing memory pressure? (Choose 2.)
64Refer to the exhibit. An IAM policy is attached to a user. The user reports that they cannot delete the production-db database. Which statement best explains the behavior?
65A database administrator notices that an Amazon RDS for MySQL instance's CPU utilization is consistently above 80% during peak hours. The DB instance is a db.r5.large with 16 GB memory and 500 GB gp2 storage. The application is a read-intensive web application. Which action is MOST effective to reduce CPU load without significant cost increase?
66A company is running a MongoDB-compatible Amazon DocumentDB cluster with one writer and two readers. The application writes a large amount of data during batch processing, and after a batch completes, the writer's CPU is high, and the readers have significant replica lag. The team wants to reduce replica lag without affecting the batch performance. What should they do?
67A developer is receiving timeout errors when connecting to an Amazon ElastiCache for Redis cluster from an Amazon EC2 instance. The security group for the EC2 instance allows outbound traffic to the Redis cluster's security group on port 6379. The Redis cluster's security group does not allow inbound traffic from the EC2 instance. What is the most likely cause of the timeout?
68A company uses Amazon RDS for PostgreSQL with Multi-AZ deployment. The primary instance fails, and a failover occurs. After the failover, the application is still unable to connect to the database endpoint. The database administrator checks the RDS console and sees that the new primary is in 'available' state. What should the administrator do next to diagnose the connectivity issue?
69A database administrator is troubleshooting a performance issue on an Amazon Aurora MySQL cluster. The application is experiencing high latency on write operations. Which TWO CloudWatch metrics should the administrator analyze to identify the root cause?
70A company is using Amazon DynamoDB for a gaming leaderboard. The table has a partition key of 'game_id' and a sort key of 'score'. The table is configured with on-demand capacity. During a major tournament, the application experiences high latency and some requests return 'ProvisionedThroughputExceededException' errors. The CloudWatch metric 'ThrottledRequests' spikes. The application uses a single partition key for all writes during the tournament (game_id = 'tournament_final'). What is the most likely cause of the throttling, and what is the best solution?
71A company runs an Amazon Redshift cluster with 8 dc2.large nodes for its data warehouse. The data engineering team loads data daily using COPY commands from S3. Recently, the load times have increased significantly. The cluster's CloudWatch metric 'CPUUtilization' is high during the load. The administrator runs the STL_LOAD_ERRORS table and finds no errors. The SVL_S3LOG shows that the COPY command is scanning many small files. The data in S3 is stored as 10,000 small CSV files (each ~100 KB). Which action will MOST improve the COPY performance?
72A developer is troubleshooting an application that writes to an Amazon ElastiCache for Redis cluster. The application occasionally fails with 'OOM command not allowed when used memory > maxmemory'. What is the most likely cause?
73A database specialist is troubleshooting a performance issue on a self-managed PostgreSQL database that they plan to migrate to Amazon RDS. The database has a high number of 'idle in transaction' connections. What is the impact of these connections on the database?
74A company uses Amazon ElastiCache for Redis. They want to monitor cache hit ratio. Which TWO metrics should be used to calculate the cache hit ratio?
75A retail company uses Amazon RDS for PostgreSQL as the backend for its e-commerce platform. During a flash sale, the database experienced high CPU utilization and increased the number of active connections. The application team reported that some queries timed out. The database specialist reviewed the slow query log and found that several queries were performing sequential scans on large tables due to missing indexes. The specialist created the necessary indexes, but the issue persists for some queries. Upon further investigation, the specialist notices that the query planner is still choosing sequential scans for some queries. What should the database specialist do to ensure the query planner uses the indexes?
76A startup is using Amazon ElastiCache for Redis to cache session data. They deployed a single Redis node (cache.t3.micro) in us-west-2. The application reports high latency when reading session data. CloudWatch metrics show CPUUtilization at 90% and Evictions at 100 per minute. The cache hit ratio is 80%. The database specialist suspects the node is overloaded. What should the specialist do to improve performance?
77A financial services company is using Amazon Aurora MySQL as its primary database. The database has a table 'transactions' that receives high inserts during business hours. The table is partitioned by date. Recently, the application team noticed an increase in lock wait timeouts. The database specialist reviewed the InnoDB status and found that there are frequent gap locks on the 'transaction_date' column. The isolation level is REPEATABLE READ. What should the specialist do to reduce lock waits while maintaining data consistency?
78A database administrator notices that an Amazon RDS for MySQL DB instance is experiencing high CPU utilization and increased latency during peak hours. The administrator wants to identify the queries causing the issue. Which TWO actions should be taken to diagnose the problem? (Select TWO.)
79A company is running Amazon RDS for MySQL and notices that the database CPU utilization is consistently above 80% during peak hours. The application performance is degrading. Which action should be taken first to troubleshoot the issue?
80A company has an Amazon DynamoDB table with on-demand capacity. Users report that write requests are occasionally throttled during peak hours. The application uses the AWS SDK and retries with exponential backoff. Which monitoring approach should be used to identify the cause of throttling?
81A developer notices that an Amazon ElastiCache for Redis cluster is experiencing high latency. The cluster uses a single node. Which CloudWatch metric should be reviewed first to determine if the issue is due to memory pressure?
82A database administrator notices that an Amazon RDS for Oracle instance has a high number of connections, causing performance degradation. Which tool can be used to identify the active sessions and their queries?
83A developer needs to monitor the number of throttled read requests for a DynamoDB table. Which CloudWatch metric should be used?
84Which TWO CloudWatch metrics should be monitored to detect storage performance issues for an Amazon RDS for MySQL instance? (Choose two.)
85A database administrator notices that an Amazon RDS for Oracle DB instance's CPU utilization is consistently above 90% during peak hours. The application is read-heavy. Which action can reduce CPU load?
86A company uses Amazon Redshift for data warehousing. They run a daily ETL job that loads data into the cluster. Recently, the job started failing with 'Disk Full' errors. The cluster has 5 RA3 nodes. Which step should be taken to resolve the issue?
87An application uses Amazon ElastiCache for Redis as a session store. Users report that sessions are being lost intermittently. The ElastiCache cluster has replication enabled with one replica. CloudWatch metrics show 'Evictions' spiking during peak hours. What is the MOST likely cause?
88A company uses Amazon DynamoDB and notices that some queries are taking longer than expected. The table has a partition key only. The 'ConsumedReadCapacityUnits' is below the provisioned throughput. What is the most likely cause of the slow queries?
89A database administrator is troubleshooting an Amazon RDS for SQL Server instance that is experiencing high 'ReadIOPS' and 'ReadLatency'. The instance uses General Purpose SSD (gp2) storage. The 'BurstBalance' metric is 0%. What should the administrator do to improve performance?
90A company is using Amazon RDS for PostgreSQL and needs to monitor the database for performance issues. Which TWO metrics in Amazon CloudWatch are most useful for identifying I/O bottlenecks?
91A company is experiencing slow query performance on an Amazon RDS for PostgreSQL DB instance. The DB instance is a db.r5.large with 16 GB RAM and 500 GB gp2 storage. Which metric in Amazon CloudWatch would most directly help identify if the performance issue is due to memory pressure?
92A database administrator is troubleshooting a failover event for an Amazon RDS for SQL Server Multi-AZ DB instance. The failover occurred automatically. Which AWS service or feature should the administrator use to view the failover history and the reason for the failover?
93A developer is troubleshooting an application that uses Amazon DynamoDB. The application sometimes receives ProvisionedThroughputExceededException errors. The table has on-demand capacity mode. The errors occur in short bursts. What is the most likely cause?
94A company runs an Amazon RDS for Oracle DB instance. The database administrator wants to receive an alert when the storage space is below 10% of the allocated storage. Which Amazon CloudWatch metric and alarm threshold should be used?
95A company is troubleshooting an Amazon RDS for MySQL DB instance that is experiencing high CPU utilization. The DB instance is a db.t3.medium. Which TWO actions should the database administrator take to investigate the cause?
96A database administrator is monitoring an Amazon RDS for PostgreSQL DB instance. The administrator notices that the DB instance is using more memory than expected. Which TWO metrics in Amazon CloudWatch can help diagnose memory usage?
97A company is using Amazon RDS for MySQL with Multi-AZ deployment. The application team reports increased latency during peak hours. Which AWS service should the database specialist use to identify the root cause?
98A database specialist is troubleshooting an Amazon RDS for SQL Server instance that is running out of storage. The instance has 500 GB of provisioned storage and is using General Purpose SSD (gp2). The specialist wants to set up an alarm to notify when free storage space drops below 50 GB. Which CloudWatch metric and threshold should be used?
99A company is using Amazon ElastiCache for Redis to cache frequently accessed data. Recently, the application has been experiencing increased latency. The database specialist suspects that the cache hit ratio has decreased. Which CloudWatch metric should the specialist analyze to confirm this suspicion?
100A database specialist is trying to connect to an Amazon RDS for MySQL instance from an EC2 instance but receives a 'Connection timed out' error. The security group for the RDS instance allows inbound traffic on port 3306 from the security group of the EC2 instance. What should the specialist check next?
101A company is running an Amazon Aurora MySQL database cluster. The database specialist notices that the write latency is high during peak hours. The cluster consists of one writer and two reader instances. Which action should the specialist take to reduce write latency?
102A company is using Amazon DynamoDB with provisioned capacity. The application is experiencing throttling on write requests. The database specialist needs to identify the cause. Which TWO metrics should be reviewed in CloudWatch? (Select TWO.)
103A company uses Amazon RDS for MySQL with Multi-AZ deployment. The database is experiencing increased latency and the application team reports slow queries. The DBA wants to identify the queries that consume the most resources. Which AWS service should be used to capture and analyze these queries?
104A company is using Amazon Redshift for data warehousing. The VACUUM operation is taking longer than expected, and the database administrator wants to identify the tables that require the most vacuuming effort. Which system table should be queried to find the percentage of deleted rows per table?
105A developer is troubleshooting an issue where an application using Amazon DynamoDB is receiving occasional 'ThrottlingException' errors. The application uses eventually consistent reads. What is the MOST likely cause of this error?
106A company is using Amazon ElastiCache for Redis as a caching layer for a web application. The application's response time has increased, and the operations team suspects that cache evictions are occurring frequently. Which ElastiCache metric should be monitored to confirm cache evictions?
107A database administrator is troubleshooting an issue where an Amazon RDS for PostgreSQL DB instance is not allowing connections. The administrator checks the security group and network ACLs, and they are correctly configured. What is the next step to diagnose the issue?
108A company is using Amazon DynamoDB and wants to monitor the read/write capacity utilization of a table. Which ONE AWS service can be used to set up alarms for capacity consumption?
109A company is using Amazon Redshift and has a query that is running slowly. The DBA wants to identify if the query is I/O-bound. Which TWO metrics from Amazon CloudWatch can indicate I/O-bound queries?
110Refer to the exhibit. A DBA sees the above error log entries for an Amazon RDS for MySQL DB instance. Which action should the DBA take to resolve the 'Too many connections' error?
111A company is using Amazon DynamoDB Accelerator (DAX) to improve read performance. Recently, the cache hit ratio has dropped significantly. The application uses strongly consistent reads. What is the most likely cause of the low cache hit ratio?
112A database administrator is monitoring an Amazon RDS for SQL Server DB instance and notices that the FreeableMemory metric is consistently below 200 MB. Which of the following actions is most appropriate to mitigate performance issues?
113A company is using Amazon Redshift for data warehousing. The query performance has degraded over time. The DBA suspects that the distribution style of large tables is suboptimal. Which Redshift system view should be queried to identify distribution skew?
114An application using Amazon DynamoDB is experiencing higher than expected read costs. The table uses on-demand capacity mode. The read pattern is mostly fetching small items (1 KB) using GetItem. Which of the following is the most cost-effective optimization?
115Which TWO metrics should be monitored together to detect a memory leak in an Amazon RDS for Oracle DB instance? (Choose TWO.)
116Which THREE steps should be taken to troubleshoot high replica lag in an Amazon Aurora MySQL DB cluster? (Choose THREE.)
117Which TWO CloudWatch Logs features can be used to monitor and troubleshoot Amazon RDS for SQL Server error logs? (Choose TWO.)
118A developer is troubleshooting slow queries in Amazon RDS for MySQL. The 'Threads_running' status variable is consistently above 200. The application uses connection pooling. Which metric should be monitored to identify the root cause?
119A database specialist is troubleshooting an Amazon RDS for SQL Server instance that is experiencing high CPU utilization. The instance has multiple databases. Which TWO actions should the specialist take to identify the cause?
120An Amazon DynamoDB table is experiencing throttled write requests. The table uses provisioned capacity with auto-scaling enabled. Which THREE factors could contribute to throttling despite auto-scaling?
121A company is running a production Amazon RDS for PostgreSQL database. The database has experienced a sudden spike in CPU utilization, causing application timeouts. The monitoring team needs to identify the root cause. Which AWS service or feature should be used to analyze the database load and identify the specific queries causing the high CPU?
122A company is using Amazon DynamoDB for a high-traffic application. The application is experiencing intermittent `ProvisionedThroughputExceededException` errors. The team has already increased the read and write capacity units multiple times but the errors persist. Which of the following is the MOST likely cause of the issue?
123A company is running a MongoDB-compatible Amazon DocumentDB cluster. The application team reports that write operations are failing intermittently with a `WriteConcernError` indicating that the write concern could not be satisfied. The cluster has one primary and two replicas. What is the MOST likely cause of this issue?
124A company is using Amazon ElastiCache for Redis as a caching layer. The application is experiencing higher latency than expected. The team suspects that cache evictions are occurring due to memory pressure. Which ElastiCache metric should be monitored to confirm this?
125A company is using Amazon RDS for MySQL and needs to monitor for slow queries. Which TWO AWS services can be used to capture and analyze slow query logs? (Choose TWO.)
126A company is running a MongoDB database on Amazon EC2. The database is experiencing high disk I/O latency. Which AWS service can be used to monitor the disk I/O metrics at the instance level?
127A company is experiencing slow query performance on an Amazon RDS for MySQL database. The DBA wants to identify the most time-consuming queries. Which TWO actions should the DBA take? (Choose two.)
128A company runs an e-commerce platform using Amazon Aurora MySQL with Multi-AZ deployment. The application has a read-heavy workload and uses a mix of SELECT and UPDATE queries. Recently, the company migrated from a db.r5.large to a db.r5.2xlarge instance class to handle increased traffic. However, after the migration, the CPU utilization remains high during peak hours, and the application's page load times have increased. The DBA notices that the 'Read IOPS' metric is high, but the 'Read Latency' metric is low. There is also a high number of 'Select' queries in the database. The application uses a single database endpoint. What should the DBA do to reduce CPU utilization and improve read performance?
129A company runs a production Amazon RDS for PostgreSQL Multi-AZ DB instance (db.r5.large) with 500 GB of General Purpose SSD (gp2) storage. The application experiences intermittent latency spikes every 15 minutes. Monitoring shows that during these spikes, the ReadIOPS metric on the primary instance spikes to 5,000 IOPS (the baseline is 1,500 IOPS), and the BurstBalance drops from 100% to 20% then recovers. There is no increase in CPU or connections. The application uses connection pooling with pgBouncer on an EC2 instance. The team has verified that no long-running queries or index scans are causing the spikes. Which action is MOST likely to resolve the intermittent latency?
130A company uses Amazon Aurora MySQL for its e-commerce platform. The DB cluster has one writer and two readers. Recently, the application started showing occasional deadlock errors during order processing. The error logs show: 'Transaction (Process ID 123) was deadlocked on lock resources with another process and has been chosen as the deadlock victim. Rerun the transaction.' The application retries three times before failing. The development team wants to reduce the likelihood of deadlocks. Which three actions should the team take? (Choose three.)
131A company runs a PostgreSQL database on an Amazon RDS DB instance (db.t3.medium) with 100 GB of General Purpose SSD (gp2) storage. The database is used by a web application that experiences occasional slowdowns. CloudWatch metrics show that the BurstBalance metric for the storage volume drops to 0% during peak usage and then recovers. The average IOPS during peak is 600, and the baseline IOPS for the volume is 300. The team needs a cost-effective solution to eliminate the performance issues. What should the team do?
132A database specialist is responsible for an Amazon Aurora PostgreSQL DB cluster that supports a production application. Users report that read-only queries routed to Aurora Replicas sometimes return stale data that is several seconds behind the writer. The specialist must determine whether the lag is due to Aurora Replica lag or application read-after-write expectations. Which CloudWatch metric should be monitored to directly quantify Aurora Replica lag?
133A database specialist is troubleshooting an Amazon Aurora PostgreSQL cluster. The application experiences occasional connection timeouts during peak hours. The specialist checks Amazon CloudWatch metrics and sees that 'DatabaseConnections' is high but within the cluster's maximum. They also notice that 'LoginFailures' is zero. Which of the following is the most likely cause of the timeouts?
134A database specialist is troubleshooting an Amazon RDS for MySQL DB instance that has been experiencing sudden performance degradation. The specialist suspects that a recent schema change or query pattern is causing excessive temporary table usage on disk. Which CloudWatch metric should be examined to identify the number of temporary tables created on disk?
135A database specialist is investigating intermittent connection timeouts to an Amazon Aurora MySQL cluster. The application logs show errors like 'Too many connections'. The specialist checks Amazon CloudWatch metrics and sees that the DatabaseConnections metric spikes to the maximum allowed by the DB parameter group's max_connections setting during peak traffic, but CPU and memory remain moderate. What is the MOST likely cause of the connection errors?
136A database specialist manages an Amazon Aurora PostgreSQL cluster. During a batch load, the writer instance experiences a sharp increase in commit latency. The specialist runs `SELECT * FROM aurora_stat_activity;` and observes many sessions in a wait event named `LWLock:BufferContent`, along with a high number of concurrent `INSERT` statements into a single table. The DB cluster parameter group has `max_connections` set to 2000. Which action should the specialist take to reduce the commit latency?
137A database specialist is monitoring an Amazon RDS for SQL Server DB instance. The specialist needs to identify the amount of free storage space available on the DB instance to prevent storage full issues. Which CloudWatch metric should be used?
138A database specialist manages an Amazon Aurora MySQL DB cluster with a writer and two readers. Users report that analytical queries run on the reader instances have become slower over the past week. The specialist checks Amazon CloudWatch metrics and notices that the metric 'AuroraReplicaLag' for both readers is consistently above 30 seconds. The application uses the reader endpoint for these queries. What is the MOST likely cause of the increased replica lag?
139A database specialist is monitoring an Amazon RDS for SQL Server DB instance. The specialist needs to receive an alert when the free storage space falls below 10% of the allocated storage. Which CloudWatch metric should be used to create an alarm for this condition?
140A database specialist manages an Amazon Aurora PostgreSQL cluster. Users report intermittent application timeouts. The specialist observes that the writer instance's CPU is normal, but the number of active connections is near the maximum. The application uses a connection pool that opens a new connection for each request without closing it. Which action should the specialist take to resolve the issue?
141A company uses Amazon DynamoDB with a table that has a partition key of 'CustomerID' and a sort key of 'OrderDate'. The table is configured with on-demand capacity mode. During a marketing campaign, the application experiences a sudden increase in write traffic. The operations team notices that some write requests are being throttled with 'ProvisionedThroughputExceededException' errors, even though the table is in on-demand mode. Which of the following is the MOST likely explanation for these throttling errors?
142A company uses Amazon ElastiCache for Redis to cache session data for a web application. The Redis cluster is in cluster mode disabled with one shard and two replicas. Users are reporting that they are being logged out unexpectedly. A database specialist checks the ElastiCache metrics and notices that the CurrConnections metric is very high and the Evictions metric is increasing. The cache node type is cache.m5.large with 6.38 GiB of memory. What is the MOST likely cause of the session loss?
143A database specialist is troubleshooting an Amazon Aurora PostgreSQL cluster where the application reports intermittent connection timeouts during peak hours. The specialist suspects a connection storm and wants to identify the source. Which metric should be monitored in Amazon CloudWatch to determine the number of current connections to the DB instance?
144A database specialist manages an Amazon DynamoDB table that uses on-demand capacity mode. The application experiences sudden spikes in write traffic, and the specialist observes that some write requests are being throttled with `ThrottlingException` errors. The specialist needs to identify the cause of the throttling and ensure that the application can handle the spikes without errors. Which action should the specialist take?
145A database specialist is troubleshooting an Amazon Aurora PostgreSQL DB cluster that is experiencing performance degradation. The cluster has one writer and one reader instance. The specialist observes that the writer's CPU utilization is consistently above 90%, while the reader's CPU is around 30%. The application performs a mix of read and write operations. The specialist wants to identify the cause of the high CPU on the writer. Which TWO actions should the specialist take to diagnose the issue? (Choose two.)
146A database specialist manages an Amazon Aurora PostgreSQL DB cluster. Users report that queries against a reader instance intermittently return stale data that is several seconds behind the writer. The specialist wants to quantify this lag over time and identify whether it correlates with writer activity. Which AWS mechanism should the specialist use to obtain this information?
147A database specialist is monitoring an Amazon RDS for SQL Server DB instance. The instance is experiencing intermittent connection timeouts. The specialist suspects that the issue is related to the maximum number of worker threads or the number of user connections. Which two metrics should the specialist review in Amazon CloudWatch to confirm the cause? (Choose two.)
148A database specialist manages an Amazon Aurora MySQL cluster used for an online ordering system. During peak traffic, the application reports occasional connection timeouts and slow queries. The specialist enables Performance Insights with a retention period of 7 days and observes that the wait event 'io/table/sql/handler' dominates during those periods. The cluster has one writer and one reader, both using db.r5.2xlarge instances. Which action should the specialist take to reduce the impact of this wait event?
149A database specialist is monitoring an Amazon DocumentDB cluster. The application reports slow query performance. The specialist checks the cluster's metrics and sees that 'BufferCacheHitRatio' is below 90% and 'DiskQueueDepth' is high. Which action should the specialist take to improve performance?
150A database administrator is responsible for an Amazon RDS for MySQL DB instance. The administrator needs to be alerted when the free storage space falls below 10% of the allocated storage. The administrator wants to receive an email notification. What is the MOST efficient way to achieve this?
151A database specialist is troubleshooting an Amazon RDS for SQL Server DB instance that is experiencing performance degradation. The specialist runs a query against the `sys.dm_exec_requests` dynamic management view and notices a high number of sessions with a wait type of `PAGEIOLATCH_SH`. The specialist needs to identify the cause and resolve the issue. Which action should the specialist take?
152A company uses Amazon DynamoDB with a table that has a global secondary index (GSI). The application performs frequent updates to an attribute that is part of the GSI's key schema. The database specialist observes that the GSI's provisioned write capacity is often throttled, even though the base table's write capacity is not. The GSI's write capacity is set to 100 WCU. What is the most likely cause of the throttling?
153A company runs an Amazon RDS for MySQL Multi-AZ DB instance. During a maintenance window, the DB instance fails over to the standby. After failover, the application's connections through the original DNS endpoint fail for about a minute, then recover automatically. A database specialist must explain why the application recovered without any endpoint change. What is the reason?
154A database specialist is troubleshooting an Amazon ElastiCache for Redis cluster that is experiencing increased latency. The cluster uses cluster mode disabled and has one shard with two replicas. The specialist notices that the engine CPU utilization is low, but the database is reporting many new connections per second. Which action should the specialist take to reduce latency?
155A company uses Amazon DynamoDB with a table that has a partition key of CustomerId (string) and a sort key of OrderDate (number). During a flash sale, many requests for a few popular customers cause throttling. The operations team wants to identify which partition keys are consuming the most read capacity units. Which AWS service or feature should they use to analyze the most accessed partition keys?
156A database specialist is managing an Amazon RDS for SQL Server DB instance. Users report that a critical stored procedure that previously ran in seconds now takes minutes to complete. The specialist checks Amazon CloudWatch metrics and sees that the DB instance's CPU utilization is low, but the disk queue depth is high. The specialist needs to identify the cause of the slowdown. Which action should the specialist take to diagnose the issue?
157A database specialist is monitoring an Amazon Aurora PostgreSQL DB cluster. The cluster has a writer instance and a reader instance. The application connects to the reader endpoint for read-only queries. Users report that some read queries return stale data. The specialist needs to ensure that reads are consistent with the latest writes. Which action should the specialist take?
158A company uses Amazon DynamoDB with a table that has a provisioned capacity mode. The application occasionally receives ProvisionedThroughputExceededException errors, but only for certain items. A database specialist must diagnose the root cause. Which two actions should the specialist take to identify the issue? (Choose two.)
159A database specialist is troubleshooting an Amazon Aurora PostgreSQL cluster that is experiencing replication lag on a reader instance. The application performs read-only queries on the reader endpoint. The specialist observes that the AuroraReplicaLag metric is consistently above 1000 milliseconds. The writer instance is not under heavy load. The specialist needs to reduce the replication lag. Which action should the specialist take?
160A database specialist is monitoring an Amazon RDS for MySQL DB instance. The instance is experiencing high CPU utilization. The specialist wants to identify the top SQL statements contributing to the load using Performance Insights. (Choose two.)
161A database specialist is monitoring an Amazon Aurora MySQL cluster. The specialist notices that the writer instance's CPU utilization is consistently high, and the application experiences increased latency. The specialist suspects that the issue is due to a high number of row lock waits. Which two actions should the specialist take to diagnose and mitigate the issue? (Choose two.)
162A database specialist manages an Amazon Aurora PostgreSQL cluster with a writer and a reader instance. Application logs show intermittent connection failures to the reader endpoint. The specialist suspects the reader is being restarted by Aurora. Which action will provide the most direct evidence of recent restart events on the reader instance?
163A database specialist is managing an Amazon DynamoDB table that uses on-demand capacity mode. The application experiences sudden bursts of read traffic, and the specialist notices that some read requests are being throttled with ProvisionedThroughputExceededException even though the table is in on-demand mode. The specialist needs to determine the cause of the throttling and mitigate it. Which factor is most likely causing the throttling?
164A database administrator is responsible for an Amazon RDS for SQL Server DB instance. They need to monitor the performance of a specific long-running query that is causing CPU spikes. Which AWS tool provides detailed, query-level performance data with minimal impact on the DB instance?
165A media company stores metadata in an Amazon DynamoDB table with on-demand capacity mode. Users report sporadic HTTP 500 errors from the application, and CloudWatch shows spikes in `ThrottledRequests` for the table. A database specialist must determine which two factors can cause on-demand tables to throttle requests even though the table is not provisioned with fixed capacity. (Choose two.)
166A database specialist is investigating a production Amazon Aurora PostgreSQL cluster where the application reports intermittent connection timeouts. The specialist suspects the issue is related to the database's connection handling. The cluster uses an Aurora PostgreSQL writer instance (db.r5.xlarge) and a reader instance. The application connects to the writer endpoint. CloudWatch metrics show that DatabaseConnections is consistently near the value of max_connections, and the application logs show errors like 'FATAL: remaining connection slots are reserved for non-replication superuser connections'. What is the most likely cause of the connection timeouts?
167A database specialist is troubleshooting an Amazon DynamoDB table that intermittently returns ThrottlingException to an application. The table uses on-demand capacity mode, and the application performs reads and writes with a mix of item sizes. The specialist wants to determine whether specific partitions are hot and whether request patterns exceed per-partition limits. (Choose two.)
168A database specialist is troubleshooting an Amazon Aurora MySQL DB cluster that is experiencing high write latency. The specialist suspects that the issue is related to the Aurora storage subsystem. Which two actions should the specialist take to monitor and diagnose the storage-related performance? (Choose two.)
169A company runs an Amazon ElastiCache for Redis cluster in cluster mode disabled. The operations team wants to receive an alert when the cluster's CPU utilization exceeds 80% for 5 minutes. Which AWS service should they use to create this alarm?
170A database specialist is responsible for an Amazon DynamoDB table that uses on-demand capacity mode. The application experiences sudden spikes in traffic, and the specialist notices that some read requests are being throttled. The specialist wants to monitor the number of read requests that were throttled due to exceeding the table's read capacity. Which CloudWatch metric should be used?
171A logistics company runs an Amazon RDS for SQL Server Multi-AZ instance. During a nightly batch job, the application experiences a brief outage and the RDS console shows a failover event. After the failover, the application reconnects to the new primary, but the database specialist notices that the old primary's endpoint now resolves to the standby. Which statement correctly describes what happened and how the application should be configured?
172A database specialist is investigating an Amazon Aurora MySQL DB cluster where the application reports intermittent connection timeouts. The specialist notices that the number of connections to the writer instance spikes dramatically every hour. Which metric in Amazon CloudWatch should be reviewed FIRST to determine if the DB cluster is hitting the maximum connection limit?
173A database specialist is monitoring an Amazon DynamoDB table that uses on-demand capacity mode. The table has a global secondary index (GSI) with a partition key of 'status' and a sort key of 'timestamp'. The application performs frequent queries on the GSI, and the specialist notices that the GSI's consumed read capacity is much higher than the base table's. CloudWatch metrics show that the GSI has a high number of ThrottledRequests. What is the most likely cause of the throttling on the GSI?
174A database specialist manages an Amazon Aurora MySQL DB cluster. Users report that queries against a reader instance intermittently return stale data, even though the application connects to the reader endpoint and the cluster has two Aurora Replicas. The specialist wants to quantify the replication delay and confirm whether the replicas are lagging. Which metric should the specialist monitor in Amazon CloudWatch to directly measure Aurora replica lag?
175A database specialist needs to audit all login attempts and failed authentication events on an Amazon Aurora MySQL DB cluster. The events must be retained for one year and searchable for security investigations. Which approach meets these requirements?
176A company uses Amazon ElastiCache for Redis to cache session data for a web application. The operations team reports that the cache hit ratio has dropped from 95% to 70% over the past week, and application latency has increased. The Redis cluster is running in cluster mode with 3 shards. The team needs to identify the cause of the decreased hit ratio. Which action should the team take first to diagnose the issue?
177A company uses Amazon Aurora MySQL. They have enabled the audit log and exported it to Amazon CloudWatch Logs. A security analyst needs to search for all failed login attempts in the last 24 hours. Which approach is most efficient?
178A database specialist is monitoring an Amazon DynamoDB table that uses on-demand capacity mode. The application experiences sudden traffic surges, and the specialist observes throttling errors in CloudWatch metrics. Which of the following is the MOST likely cause of the throttling?
179A database administrator is managing an Amazon RDS for MySQL DB instance. They need to identify the top SQL statements that are consuming the most CPU over the past hour to optimize performance. The DB instance has Performance Insights enabled with a retention period of 7 days. Which action should the administrator take to quickly identify the top SQL statements?
180A company runs an Amazon RDS for SQL Server DB instance in a Multi-AZ deployment. The database specialist needs to monitor the health and performance of the database and set up alerts for potential issues. The specialist wants to use Amazon CloudWatch metrics and alarms. Which two CloudWatch metrics should the specialist monitor to detect insufficient memory and excessive disk I/O on the RDS instance? (Choose two.)
181A company uses Amazon DynamoDB with on-demand capacity mode for a new application. During a load test, the application occasionally receives HTTP 400 errors with the code ProvisionedThroughputExceededException, even though the table is in on-demand mode. A database specialist must explain why throttling can still occur and which factor is most likely responsible. What is the most likely cause?
182A database specialist is investigating why an Amazon RDS for SQL Server DB instance is experiencing slow query performance. The specialist notices that the CPU utilization is high and there are many PAGEIOLATCH_SH waits. Which action should the specialist take to improve performance?
183A database specialist is monitoring an Amazon RDS for SQL Server DB instance using Amazon CloudWatch. The specialist notices that the FreeStorageSpace metric is decreasing rapidly, and the instance is approaching its storage limit. The specialist needs to identify the cause of the rapid storage consumption. Which action should the specialist take to diagnose the issue?
184A company uses an Amazon Aurora MySQL cluster with a writer and a reader. They notice that the reader's replica lag is increasing, and the application is reading stale data. The database specialist needs to identify the cause of the replica lag. Which action should the specialist take first?
185A database specialist manages an Amazon Aurora MySQL DB cluster with a single writer and two reader instances. During peak hours, the application experiences intermittent connection timeouts when connecting to the cluster writer endpoint. CloudWatch metrics show that the writer's CPU is at 45%, freeable memory is stable, and there are no long-running queries. The specialist needs to identify the cause of the connection timeouts. Which action should the specialist take to diagnose the issue?
186A database specialist is troubleshooting an Amazon RDS for SQL Server DB instance that is experiencing slow query performance. The specialist suspects that a specific query is causing high CPU usage. Which AWS tool should be used to identify the top SQL statements by CPU consumption?
187A company uses Amazon ElastiCache for Redis to cache session data for a high-traffic web application. The operations team notices that the cache hit rate has dropped from 95% to 70% over the past week, leading to increased latency. They suspect that many keys are expiring prematurely or being evicted. The Redis cluster is running in cluster mode with 3 shards. Which action should the team take to determine the cause of the reduced hit rate?
188A database specialist is managing an Amazon DocumentDB cluster. The application team reports that queries are slower than usual, and the specialist suspects that the cluster's cache hit ratio has dropped. Which Amazon CloudWatch metric should the specialist monitor to evaluate the buffer cache hit ratio for the DocumentDB cluster?
189A database specialist is troubleshooting an Amazon RDS for SQL Server DB instance that is experiencing slow query performance during peak hours. The specialist has enabled Enhanced Monitoring and Performance Insights. Which two actions should the specialist take to identify the most resource-intensive queries and their wait events? (Choose two.)
190A company runs an Amazon RDS for SQL Server DB instance with Multi-AZ. During a failover test, the application experiences a brief outage, but after the failover, the database specialist notices that the secondary AZ's DB instance is not receiving updates from the primary. The RDS console shows the DB instance status as 'available' and the Multi-AZ status as 'pending'. Which action should the specialist take to resolve this issue?
191A database specialist is monitoring an Amazon Aurora MySQL cluster with a writer and one reader. The application uses the reader endpoint for read-only queries. The specialist observes that the reader instance's ReplicaLag metric is consistently increasing and is currently at 300 seconds. The writer instance shows high write throughput. Which action should the specialist take to reduce the replica lag?
192A startup runs an Amazon Aurora MySQL cluster and wants to receive an alert when the average number of database connections exceeds a threshold for five consecutive minutes. The team already uses Amazon CloudWatch. Which combination of CloudWatch features should the database specialist configure to meet this requirement with the least effort?
193A database specialist is responsible for an Amazon RDS for MySQL DB instance. The specialist needs to receive an alert when the average number of database connections exceeds 80% of the maximum allowed for the instance's parameter group over a 5-minute period. Which combination of CloudWatch alarm configuration steps should the specialist use?
194A company uses Amazon ElastiCache for Redis to cache session data for a web application. The operations team wants to monitor the cache hit rate to ensure the cache is effective. They need to set up a CloudWatch alarm that triggers when the cache hit rate falls below 80% over a 5-minute period. Which metric should they use to calculate the cache hit rate?
195A company runs an Amazon Aurora PostgreSQL DB cluster. A database specialist notices that the writer instance's CPU utilization is consistently high, and Performance Insights shows a large number of active sessions waiting on the LWLock:BufferContent event. Which action should the specialist take to reduce the impact of this wait event?
196A database administrator is monitoring an Amazon RDS for MySQL DB instance. The administrator notices that the 'FreeableMemory' metric is consistently low, and the 'SwapUsage' metric is increasing. The DB instance is running a memory-intensive workload. What is the most likely impact on database performance, and what should the administrator do to resolve the issue?
197A database specialist is responsible for an Amazon DocumentDB cluster. The specialist needs to monitor the cluster's CPU utilization and set up an alarm if it exceeds 80% for 10 minutes. Which AWS service should be used to create this alarm?
198A company uses Amazon ElastiCache for Redis (cluster mode disabled) with a single shard and two replicas. The application experiences occasional latency spikes when reading from the primary node. CloudWatch metrics show high CPU utilization on the primary node and increased replication lag on the replicas. The application uses the primary endpoint for both reads and writes. What is the most likely cause of the latency spikes?
199A database specialist is monitoring an Amazon Aurora PostgreSQL cluster. The cluster has a writer and one reader. The application connects to the writer endpoint for writes and the reader endpoint for reads. During peak hours, the specialist notices that the reader instance's CPU is at 95% while the writer's CPU is at 40%. The application reports increased read latency. CloudWatch metrics show that the reader's 'SelectThroughput' is high and 'BufferCacheHitRatio' is low. What is the most likely cause of the high reader CPU?
200A database specialist is monitoring an Amazon RDS for SQL Server DB instance using Enhanced Monitoring. The specialist notices that the 'Disk Read Latency' metric is consistently high, while 'Disk Write Latency' is normal. The DB instance uses General Purpose SSD (gp2) storage. The application performs a high volume of read operations. What is the most likely cause of the high read latency?
201A database specialist is responsible for an Amazon Neptune cluster. The operations team reports that some Gremlin queries are timing out. The specialist needs to monitor the cluster to identify slow queries. Which AWS service or feature should the specialist use to capture and analyze the query performance?
Deep-dive questions
The most-searched questions in this domain — detailed explanations, worked examples, full answer breakdowns.
You must map a database symptom to the right AWS monitoring tool: CloudWatch metrics and alarms for thresholds, Performance Insights for query load, Enhanced Monitoring for OS metrics, and DMS with CloudWatch for migrations. The key skill is choosing the correct metric or service for the stated problem.
The Courseiva DBS-C01 question bank contains 201 questions in the Monitoring and Troubleshooting domain, covering the 18% of the exam attributed to this domain in the official Amazon Web Services blueprint. Click any question to see the full explanation and answer breakdown.
Start with a 10-question focused session to identify your baseline accuracy in this domain. Read every explanation — even for questions you answer correctly — to understand the reasoning. Once you score consistently above 80%, move to a 20–30 question session to confirm depth before moving to the next domain.
Yes — the session launcher on this page draws questions exclusively from the Monitoring and Troubleshooting domain. Choose 10, 20, 30, or 50 questions for a focused session, or click individual questions to review them one by one.
Save your results, see per-domain analytics, and get readiness scores — free, for every certification.
Sign Up FreeFree forever · Every certification included