~/redhat-ex267 · redhat-ex267_d1_493b7bc88b2b ▊
A team wants to reduce the cost of generative AI inference while maintaining stable latency at scale as user demand grows. Which OpenShift AI component is specifically designed to manage these inference costs by using GPU resources efficiently?
Unlock this exam to join the per-question discussion and vote on questions.