ended8월 15일· 1 sources
Shaving a second off a real-time speech-to-LLM pipeline in Electron
Why it matters
Every part of a speech→LLM pipeline is fast enough on its own. Put them in a row and you get three seconds, which is far too slow when a human is waiting for you to say something. I build a desktop ov...
1
Sources
+0
24h
—
Growth
6d
Active