why I personally think the current AI “boom” is a bubble around a nascent tech that’s not really ready for lots of the situations it’s been sold for
Basically this. On top of that, we not far off from the point where you can run something like claude localy. Obviously we don’t have access to claude itself, but we do have access to local big models that come close to it in performance, like kimi-k3 ~500gb on average, deepseek ~1tb or glm5.2 (smallest quant starts from 228gb)
And it may sound like a lot, but it isn’t in reality — most amd cpus from am4 able to address 128gb of ram, while new am5 can do 256gb. Hell, my gaming PC have 128gb of DDR5. So we already technically can run some of those models on a regular office PC. Obviously, due to ram crisis, it will cost you a fortune, but it’s not like it impossible. Furthermore, demand for gpu with a lot of vram will sooner or later be satisfied, so it just matter of time when average consumer will be able to do it. And well, why would we need dedicated data centers then?
Basically this. On top of that, we not far off from the point where you can run something like claude localy. Obviously we don’t have access to claude itself, but we do have access to local big models that come close to it in performance, like kimi-k3 ~500gb on average, deepseek ~1tb or glm5.2 (smallest quant starts from 228gb)
And it may sound like a lot, but it isn’t in reality — most amd cpus from am4 able to address 128gb of ram, while new am5 can do 256gb. Hell, my gaming PC have 128gb of DDR5. So we already technically can run some of those models on a regular office PC. Obviously, due to ram crisis, it will cost you a fortune, but it’s not like it impossible. Furthermore, demand for gpu with a lot of vram will sooner or later be satisfied, so it just matter of time when average consumer will be able to do it. And well, why would we need dedicated data centers then?