Q6Operational Efficiency and Optimization for GenAI Applications
A retail company uses Amazon Bedrock to build a customer service AI assistant. Analysis indicates that 70% of customer inquiries are simple product questions that a smaller model can handle effectively. However, 30% of inquiries are complex return-policy questions requiring advanced reasoning. The company wants to implement a cost-effective model-selection framework that automatically routes customer inquiries to appropriate models according to inquiry complexity. The framework must preserve high customer satisfaction and minimize response latency. Which solution meets these requirements with the **LEAST** implementation effort?
← → navigate · a answer
Community votes
Discussion · 2
B 2
https://aws.amazon.com/blogs/machine-learning/multi-llm-routing-strategies-for-generative-ai-applications-on-aws/
B 1
Bedrock Prompt routing is the best approach for cost and LEAST operations : https://aws.amazon.com/bedrock/intelligent-prompt-routing/