Base Model 20250808 Internal Error when training with Structure Text file

Kaiyip Ho 40 Reputation points
2026-06-12T14:33:56.61+00:00

When I train a custom model using a structure text file and output file using 20250808 build (english) I get an failure for the model and it says internal error. The structure text and output files are error free when I uploaded them. I was successfully able to train a custom model using the same structure text and output files except with a older base model (20221013).

The docs don't say anything about newer models not accepting the structure text file. So I am seeking clarification on this matter. Thank you in advance.

User's image

Azure Speech in Foundry Tools

Answer accepted by question author
Sina Salam 31,456 Reputation points Volunteer Moderator
2026-06-20T16:32:54.6233333+00:00

Hello Kaiyip Ho,

Welcome to the Microsoft Q&A and thank you for posting your questions here.

I understand that you are experiencing an Internal Error when training an Azure AI Speech Custom Speech model with Structured Text and output-formatting files using the English base model20250808.

With information provided the issue is not caused by the Structured Text file itself, because the same validated files train successfully with base model 20221013, but fail repeatedly with base model 20250808. Microsoft documentation still lists Structured Text as a supported Custom Speech training dataset type, this points to a base-model-specific backend regression, undocumented compatibility limitation, or service-side training issue for 20250808, not a normal customer file-format problem. - https://learn.microsoft.com/en-us/azure/ai-services/speech-service/how-to-custom-speech-test-and-train, and https://learn.microsoft.com/en-us/azure/ai-services/speech-service/how-to-custom-speech-train-model

The below steps will resolve the issue:

This is the only reliable path to both continue the project safely and obtain the authoritative root cause from Microsoft engineering. Use the following official Microsoft resources for more reading and implementation steps:

I hope this is helpful! Do not hesitate to let me know if you have any other questions, steps or clarifications.


Please don't forget to close up the thread here by upvoting and accept it as an answer if it is helpful.

Was this answer helpful?

1 person found this answer helpful.
0 comments No comments

Answer accepted by question author
Anshika Varshney 15,625 Reputation points Microsoft External Staff Moderator
2026-06-12T17:30:30.2433333+00:00

Hi @Kaiyip Ho

This looks like a model/version-specific issue rather than a problem with your data.

Since the same structured text works with an older base model (20221013) but fails with 20250808, it likely indicates a limitation or backend issue with the newer model build. Similar “internal error / 500” behaviors are typically related to service-side issues or unsupported scenarios rather than file validation problems. [community.openai.com]

You can try the following:

Use the older base model temporarily (as you already confirmed it works)

Retry the training job after some time to rule out transient backend issues

Keep checking if there is any update on supported formats for the newer model

At the moment, it doesn’t look like an issue with your structured text file itself.

Hope this helps!

Thankyou!

Was this answer helpful?

1 person found this answer helpful.
0 comments No comments

0 additional answers

Sort by: Most helpful

Your answer

Answers can be marked as 'Accepted' by the question author and 'Recommended' by moderators, which helps users know the answer solved the author's problem.