ended8월 27일· 1 sources
We measured a week of inference. Routing by task difficulty cuts our cost per call roughly 48x — and flips which users are profitable.
Why it matters
We did the thing everyone building on LLMs does. We defaulted to a strong frontier model, because the demo has to be good and nobody gets fired for picking the strongest model. Then we measured a week...
1
Sources
+0
24h
—
Growth
5d
Active