It's really amazing to see how the gaps between the self hostable models and the closed models has been shrinking in the last 24 months.
And how this has been accelerating!!
I felt this very hard when I had to travel in the middle of nowhere in south america, with no network, and wanted to keep an LLM model on my macbook pro with 48GB of RAM. That was back in April 2026, a few months ago.
I downloaded Google Gemma 4 (google/gemma-4-26b-a4b) and - Oh boy - I was amazed by it's capacity!
I was able to use it to code simple things, ask it about nature, learn new stuff while traveling and make stories for the kids.
Was really amazing to observe and experiment this!
Seems to me there will be some good chance to run these great LLM locally on our hardware!
And how this has been accelerating!!
I felt this very hard when I had to travel in the middle of nowhere in south america, with no network, and wanted to keep an LLM model on my macbook pro with 48GB of RAM. That was back in April 2026, a few months ago.
I downloaded Google Gemma 4 (google/gemma-4-26b-a4b) and - Oh boy - I was amazed by it's capacity!
I was able to use it to code simple things, ask it about nature, learn new stuff while traveling and make stories for the kids.
Was really amazing to observe and experiment this!
Seems to me there will be some good chance to run these great LLM locally on our hardware!
Amazing time to be alive