hYm3nSl4yrdelirentsIMHO, the future is local models and cheaper hardware, so everyone can run AI themselves. But that would mean $0 for their magical cloud compute which is the only business they have.
i'm not sure that hardware can improve at a rate faster than the models grow. it really sucks to have to send your data over to the big companies so they can run inference (and do whatever they do with all the data)
and then there's the problem of efficiency/environmental impact, which people are making a big deal about. But running the models in datacentres is always going to be more efficient / better for the environment than people running the models locally.
really sucks
some analysts predict that local AI models can be as good as AI that ran in a datacenter half a year ago, it's not just the hardware improving but several other factors too like compression techniques and algorithmic efficiencies
a GTX 3090 can run qwen 3.8-27b which gets a better AI intelligence score (doesn't just include googling information) than Claude Opus 4.5, GPT-5.2 and Gemini 3 7 months ago