QuestionQ21

Operational Efficiency and Optimization for GenAI Applications

An e-commerce company is building an internal platform for developing generative AI applications with Amazon Bedrock foundation models (FMs). Developers must choose models based on evaluations aligned with e-commerce use cases. The platform must show accuracy metrics for text generation and summarization in dashboards. The company has custom e-commerce datasets to use as standardized evaluation inputs.

Which combination of steps meets these requirements with the LEAST operational overhead?

Choose two
Explanation

Amazon Bedrock managed model evaluation jobs can use custom prompt datasets stored in Amazon S3 with appropriate IAM permissions. For automatic evaluations, accuracy for general text generation is calculated as the Real World Knowledge (RWK) score, and accuracy for text summarization is calculated as BERTScore. Using the managed Amazon Bedrock evaluation capability minimizes operational work compared with running custom SageMaker jobs, open-source evaluation frameworks, or direct InvokeModel processing code.

Learn more

Community Discussion

No comments yet. Be the first to start the discussion!