Thanks

#1
by Scriptease - opened

Your quant works now in audio.cpp.

I have been trying it out and it's great so far. The reason why i wanted something like r2t2 to work streaming is to get a bigger model to transcribe while i am talking and basically free. The q8 version is 1/2 real time and this like 1/3 real time so folks with a less powerful mac will definitely benefit.

I am implementing a https://www.typewhisper.com/ plugin and once my audio.cpp pr is merged will add your and https://huggingface.co/davidxifeng/Confucius4-R2T2-gguf models.

Cheers Florian

I am actually working on a fork of transcribe.cpp (usually used by Handy https://github.com/cjpais/Handy/ ) this is my fork https://github.com/NairoDorian/transcribe.cpp

and I also have a fork of Handy that is much faster and much improved would you like to be a beta tester ?

Sign up or log in to comment