Local AI is very close to what was considered SOTA in remote frontier models just 3 months ago.
Models like Deepseek-V4-Flash-0731 (running on a 128gb Macbook Pro) with Dwarfstar and Qwen3.8-27b running on any 32gb Mac. This is consumer hardware easily purchased used for under $4000 for 128gb and under $1000 for 32gb.
The system requirements will shrink and the models will get smarter. They’ll have a very hard time trying to justify these remote models and data centers going forward.