Comment on AiOla open-sources ultra-fast ‘multi-head’ speech recognition modelparentComments−ukuina2yI haven't found a single good (working, easy to deploy cross-platform on CPU/CUDA/Apple Silicon) implementation of streaming + diarization, and I have looked at everything from WhisperX to pyannote to WhisperKit.Any suggestions would be very welcome!
Comments
I haven't found a single good (working, easy to deploy cross-platform on CPU/CUDA/Apple Silicon) implementation of streaming + diarization, and I have looked at everything from WhisperX to pyannote to WhisperKit.
Any suggestions would be very welcome!