Discussion about components and system integration.
One thing that I really enjoyed about Apple operating system is that the Speech to text input is excellent. I’m hoping to find another service that works with Wayland, I found https://voxtype.io/ which uses a model from hugging face, so in that instance is there a really good model to consider for speech to text, using a microphone input with minimal background noise.
The service works as expected and with some configuration settings it has most of the features I want. I would prefer if it inserted at cursor, however I did not see an easy way to configure that.
“small.en” is about 95% reliable. However, there is some delay after the recording when it converts to processing. I’d really like it if I could get the cursor to change it’s configuration to show me when the transcription is ready.