Deal of The Day! Hurry Up, Grab the Special Discount - Save 25% - Ends In 00:00:00 Coupon code: SAVE25
Welcome to Pass4Success

- Free Preparation Discussions

Amazon AIP-C01 Exam - Topic 3 Question 15 Discussion

A financial services company is creating a Retrieval Augmented Generation (RAG) application that uses Amazon Bedrock to generate summaries of market activities. The application relies on a vector database that stores a small proprietary dataset with a low index count. The application must perform similarity searches. The Amazon Bedrock model's responses must maximize accuracy and maintain high performance. The company needs to configure the vector database and integrate it with the application. Which solution will meet these requirements?
B) Launch an Amazon MemoryDB cluster and configure the index by using the Hierarchical Navigable Small World (HNSW) algorithm. Configure a vertical scaling policy based on performance metrics.
A) Launch an Amazon MemoryDB cluster and configure the index by using the Flat algorithm. Configure a horizontal scaling policy based on performance metrics.
C) Launch an Amazon Aurora PostgreSQL cluster and configure the index by using the Inverted File with Flat Compression (IVFFlat) algorithm. Configure the instance class to scale to a larger size when the load increases.
D) Launch an Amazon DocumentDB cluster that has an IVFFlat index and a high probe value. Configure connections to the cluster as a replica set. Distribute reads to replica instances.

Amazon AIP-C01 Exam - Topic 3 Question 15 Discussion

Actual exam question for Amazon's AIP-C01 exam
Question #: 15
Topic #: 3
[All AIP-C01 Questions]

A financial services company is creating a Retrieval Augmented Generation (RAG) application that uses Amazon Bedrock to generate summaries of market activities. The application relies on a vector database that stores a small proprietary dataset with a low index count. The application must perform similarity searches. The Amazon Bedrock model's responses must maximize accuracy and maintain high performance. The company needs to configure the vector database and integrate it with the application. Which solution will meet these requirements?

Show Suggested Answer Hide Answer
Suggested Answer: B

Option B is the optimal solution because it maximizes similarity search accuracy and performance for a small, proprietary dataset while maintaining low operational complexity. Amazon MemoryDB is a fully managed, in-memory database that provides microsecond-level latency, making it ideal for real-time RAG workloads that require fast vector similarity searches.

For small datasets with low index counts, the Hierarchical Navigable Small World (HNSW) algorithm is recommended by AWS for its high recall and accuracy. Unlike approximate methods optimized for massive datasets, HNSW excels at returning the most semantically relevant vectors with minimal loss of precision, which directly improves the quality of responses generated by the Amazon Bedrock foundation model.

Vertical scaling in MemoryDB is sufficient for this use case because the dataset size is limited. Scaling up instance size provides increased memory and compute capacity without the complexity of managing distributed indexes or sharding strategies. This simplifies operations while maintaining predictable performance.

Option A's Flat algorithm is computationally expensive and inefficient at scale, even for moderate query volumes. Option C introduces higher latency and operational overhead by using a relational database not optimized for in-memory vector search. Option D is unsuitable because Amazon DocumentDB is not designed for high-performance vector similarity workloads and introduces unnecessary replica management complexity.

Therefore, Option B best meets the requirements for accuracy, performance, and efficient integration with an Amazon Bedrock--based RAG application.


Contribute your Thoughts:

0/2000 characters
Alyce
3 days ago
Not sure if MemoryDB is the right fit for this use case.
upvoted 0 times
...
Curt
8 days ago
Definitely leaning towards D for the replica set benefits.
upvoted 0 times
...
Leonard
13 days ago
Surprised that IVFFlat isn't the go-to choice here!
upvoted 0 times
...
Sarah
19 days ago
I think A is better for performance scaling.
upvoted 0 times
...
Muriel
24 days ago
Option B seems solid with HNSW for similarity searches.
upvoted 0 times
...
Mable
29 days ago
I’m leaning towards the DocumentDB option since it mentions a high probe value, which seems like it could improve accuracy. But I’m not entirely confident.
upvoted 0 times
...
Therese
1 month ago
I feel like IVFFlat was mentioned in our readings, but I can't remember if it was specifically for Amazon Aurora or another service.
upvoted 0 times
...
Charisse
1 month ago
I think we practiced a similar question about scaling policies. If I recall correctly, vertical scaling might be better for performance in this case.
upvoted 0 times
...
Mari
1 month ago
I remember we discussed vector databases in class, but I'm not sure which algorithm is best for low index counts. HNSW sounds familiar, though.
upvoted 0 times
...

Save Cancel