3 ms·
I am not arguing that there are perhaps other models that can run at the same quality, can coordinate between the different modalities, but are way less power h
by Phemist 1mo ago
I am not arguing that there are perhaps other models that can run at the same quality, can coordinate between the different modalities, but are way less power hungry. My point is exactly about the comparison between the token output of the LLM running on the jalapeno chip, and sneaking in the power "usage" of the brain in the "token output" of human speech.