#local-inference
POSTS
- Half a token per second, and nothing leaves the machine
Half a token per second sounds like a punchline until you see what it buys: a frontier-scale model that answers with no network, no…
Too many projects, too many ideas, too few hours — one learning a day anyway
Half a token per second sounds like a punchline until you see what it buys: a frontier-scale model that answers with no network, no…