Mahidol University Releases Free Thai Speech-to-Text Models for Commercial Use
Researchers from the Biomedical and Data Lab at Mahidol University have released Thonburian Whisper, a suite of Thai-language automatic speech recognition models fine-tuned from OpenAI's Whisper. The models are available in multiple sizes, ranging from small variants that run on standard hardware to a large-v3 version requiring a GPU, achieving Word Error Rates of 6.59 and 7.42 respectively on the Common Voice Thai test set. All models are licensed under Apache 2.0, making them freely usable for commercial applications without requiring additional permission. Distilled versions offering faster inference are also available for developers prioritising speed over maximum accuracy. The project was originally published in 2022 and the models are hosted publicly on Hugging Face.
This is an AI-generated summary. ShortSingh links to the original source for the complete article.


Discussion (0)
Log in to join the discussion and vote.
Log in