An Azure real-time data ingestion service.
Hi @Yuanbo Lin ,
Thank you for your patience
The PG (Product Group) team analyzed the issue in detail and identified that the intermittent Server Busy Errors (Error Code 50011) were caused by namespace memory throttling due to uneven memory distribution across backend nodes. This occurred even though overall throughput levels were normal.
To mitigate the issue, the PG team scaled out the Event Hub cluster from 55 to 70 nodes, which improved the memory baseline across nodes and stabilized the environment. After scaling, no broker movements were observed, and the error occurrences stopped.
Currently, no Server Busy Errors have been seen in the last 24 hours, confirming that the issue is mitigated.
The issue was caused by memory pressure on certain backend nodes and has been successfully mitigated by increasing cluster capacity, ensuring stable service now.
Hope this helps. If you have any follow-up questions, please let me know. I would be happy to help.