Developer fixes Catbot voice recognition by switching model and microphone
A developer working on a desktop AI pet called Catbot struggled with persistent voice recognition errors, including the app mishearing its own name as 'Cat bought' or 'Lawrence'. Adjusting the listening pause window from 0.5 to 1.5 seconds and switching to a close-range headset microphone improved accuracy significantly. However, Catbot continued to misrecognize its name because the primary speech model, vosk-model-en-us-0.22, does not support runtime or custom grammar graphs. Switching to the vosk-model-en-us-0.22-lgraph model, which does support custom grammar, allowed the developer to control available output choices and resolve the naming conflict. The fix highlighted a conceptual bug rooted in model architecture rather than a conventional software error.
This is an AI-generated summary. ShortSingh links to the original source for the complete article.
Discussion (0)
Log in to join the discussion and vote.
Log in