An Azure service that provides access to OpenAI’s GPT-3 models with enterprise capabilities.
AI foundry performances for some OpenWeight models are unbelievably slow
Yassine
0
Reputation points
I use models such as Kimi K2, in different regions, but the performances are really slow. The time of a response can take several minutes which defeats the purpose. The RPM are set to the highest and the monitoring does not really work.
Does anyone have an idea on how to resolve these performance issues? Changing or using different regions didn't really solve it.
Thank you
Azure OpenAI in Foundry Models
Azure OpenAI in Foundry Models
Sign in to answer