new3시간 전· 1 sources

A weekend with TensorFold on a MacBook: the engine mattered, the quant did not

Why it matters

A 27B dense model on my MacBook decodes at about 26 tokens a second. That is fine for chat and painful for an agent that writes a few thousand tokens per turn across a hundred turns. So when a new eng...

1
Sources
+1
24h
—
Growth
1d
Active

Sources

Related Issues