@claude is a contributor? I find that hard to believe.
smalltorch
Stuff like this is what doctors should use instead of blasting patient data to a AI scribe who knows where. Im pretty sure that is illegal but it still takes place and I am pretty sure one of the reasons for the push to solve FHE...but its obviously way more practical and private to just do the processing locally.
show comments
dandaka
How does it compare to models from handy.computer and whisprflow?
show comments
dmwood
Question from a naive macports user. To some extent brew is orthogonal to macports. Is there an easy way to link to ffmpeg installed by macports? Thanks.
show comments
jiehong
Congrats on the launch!
1. I see it uses python, but I was wondering if it would use mlx-swift instead. Just curious if you considered it.
2. So far, qwen ASR seem usually less well supported than parakeet or whisper so far in most dictation applications (voxtral is similar in that regard). Why do you think that is?
Thanks and good luck with your roadmap!
foobarqux
This is like 2x realtime on M1 vs 60x on fluidaudio with parakeet.
@claude is a contributor? I find that hard to believe.
Stuff like this is what doctors should use instead of blasting patient data to a AI scribe who knows where. Im pretty sure that is illegal but it still takes place and I am pretty sure one of the reasons for the push to solve FHE...but its obviously way more practical and private to just do the processing locally.
How does it compare to models from handy.computer and whisprflow?
Question from a naive macports user. To some extent brew is orthogonal to macports. Is there an easy way to link to ffmpeg installed by macports? Thanks.
Congrats on the launch!
1. I see it uses python, but I was wondering if it would use mlx-swift instead. Just curious if you considered it.
2. So far, qwen ASR seem usually less well supported than parakeet or whisper so far in most dictation applications (voxtral is similar in that regard). Why do you think that is?
Thanks and good luck with your roadmap!
This is like 2x realtime on M1 vs 60x on fluidaudio with parakeet.
What languages does it support?
is it better than whisper from oai?