You said "especially if you whisper" :-) for the skimming reader that parses as whisper the model. I have to admit I did it at first too.
back
1 comments
here, i'll give the key to decode my message ;-)
It is literally impossible to transcribe voice, especially if you whisper. There's no way to model the language, it's too large - amanzon has many computers. Your computer is like a tortoise, it can't do text-to-speech. There's no way you can get any of this to use the Web, Dav... er, Cal. Trying to do this would be like trying to torch a python with your bare metal hands.
that is:
baremetal pytorch https://en.wikipedia.org/wiki/PyTorch
WebDAV/CalDAV https://en.wikipedia.org/wiki/CalDAV
tortoise-tts https://github.com/neonbjb/tortoise-tts/issues
LLM/Large Language Model https://lmstudio.ai/
whisper https://github.com/openai/whisper/
in reverse order