Skip to content

Comment on Show HN: Ermine.ai – Record and transcribe speech, 100% client-side (WASM)

Comments

Can it identify speakers? For examples

SPEAKER A: blah blah

SPEAKER B: blah blah

So it can be used for transcribing phone calls?

Unfortunately not at the moment, but that feature (speaker diarization) is something I'm looking in to! AFAIK the Whisper model (which this uses currently) doesn't support that functionality, but I'm experimenting with adding an auxiliary model just for diarization.

AboutSource Built by g1lg1l

Hackerly is an independent reader for Hacker News, built on the public HN API. Not affiliated with Y Combinator.