3 ms·
Agreed, though I do think that LLMs are still more similar than we think. Sometime after the first release, AI labs found that coding sat in the niche space of
by fyredge 1mo ago
Agreed, though I do think that LLMs are still more similar than we think. Sometime after the first release, AI labs found that coding sat in the niche space of lots of easily digestible data and fast feedback from error messages and compiler checks etc. This allowed models to be trained with a focus on coding tasks, but the underlying technology is still the same, the infrastructure around it changed, they are still generating via probabilistic sampling.
Don't get me wrong, I'm using local coding agents myself with varying levels of success and frustration, but the models themselves behave similarly to their siblings from 2020.
The infrastructure improved, that includes the data. I have a pet theory that if they took the earlier models and retrain it with the data they used to train the latest models, we will get a similar result.