Q30Fundamentals of GenAI
A company needs to choose a generative AI model to build an application. The application must return responses to users in real time. Which model characteristic should the company evaluate to meet these requirements?
← → navigate · a answer
Community votes
Discussion · 3
C 1
To deliver real-time responses in a generative AI application, the key model characteristic is:
✅ Inference speed – This is the time it takes the model to generate a response after receiving a prompt.
Faster inference speed ensures lower latency, which is essential for real-time user interactions.
Especially important in chatbots, customer support, and interactive applications.
C 1
Inference speed only which can help in lowering down the latency for real-time user interactions.
C 1
C. Inference speed
Inference speed determines how quickly a model can generate outputs after receiving input.