I expect so as well, and my prediction is that we’ll have LLMs that are roughly as capable as the current frontier that can be run locally within a year or two. At that point, it’s just going to be good enough for vast majority of tasks most people need to do.
Not now, but I’d expect LLMs to be much more efficient in a couple of years.
I expect so as well, and my prediction is that we’ll have LLMs that are roughly as capable as the current frontier that can be run locally within a year or two. At that point, it’s just going to be good enough for vast majority of tasks most people need to do.
I expect more efficient LLMs to come out of China, who has turned to the AI-as-software model (contrast the AI-as-service model in the US).
They already are, DeepSeek/GLM/Qwen are great models and can run on 2 DGX Sparks