new3시간 전· 1 sources
A weekend with TensorFold on a MacBook: the engine mattered, the quant did not
Why it matters
A 27B dense model on my MacBook decodes at about 26 tokens a second. That is fine for chat and painful for an agent that writes a few thousand tokens per turn across a hundred turns. So when a new eng...
1
Sources
+1
24h
—
Growth
1d
Active