ended8월 15일· 1 sources

Shaving a second off a real-time speech-to-LLM pipeline in Electron

Why it matters

Every part of a speech→LLM pipeline is fast enough on its own. Put them in a row and you get three seconds, which is far too slow when a human is waiting for you to say something. I build a desktop ov...

1
Sources
+0
24h
Growth
6d
Active

Sources

Related Issues