I've got ~⅓ of my hardware crunching locally. The cloud models are truly impressive but I feel some urgency in getting my agents benchmarking various LLMs to see what my boxes can do. Trying to find the right balance of context and speed. It's all about vram and patience!
Post by A
The gap between remote AI and local AI is closing rapidly. We were a 1 year in terms of remote to local separation, then 6 months, now local is only lagging about 3 months behind the best frontier models.
This gap will only get smaller over time, the local models will get smaller, smarter, faster, and easily run on even the most budget hardware.
More tokens per second is great but it has an upper limit. At a certain point the human must stop, and actually think about what it wants the machine to do next.
The machines are here to serve us, not vice versa. The second you get confused and forget that, you’re dead.
DON’T GET CONFUSED!
718 850 sat
What the chain says
- Block
- 963 734
- Time
- 2026-08-24T17:17:05Z
- Signer
- 1Mg2Zb2VbidHmJh9ZEXLaYKGndVEierdWt
- App
- twetch
- Type
- reply
- Content type
- text/markdown
Fields the transaction did not carry are omitted. Open the payload to see the bytes as stored.
Signed by
1Mg2Zb2VbidHmJh9ZEXLaYKGndVEierdWt Unverified