3 ms·
But there are already voice models that do a reasonable job at a fraction of the throughput available? The real question is how expensive it is to coordinate b
by perching_aix 1mo ago
But there are already voice models that do a reasonable job at a fraction of the throughput available?
The real question is how expensive it is to coordinate between these different modalities, and I really don't see why it'd be all that much.
I half expect Boston Dynamics to show something like this off in Q4 or whatever.
- Phemist 1mo agoI am not arguing that there are perhaps other models that can run at the same quality, can coordinate between the different modalities, but are way less power hungry. My point is exactly about the comparison between the token output of the LLM running on the jalapeno chip, and sneaking in the power "usage" of the brain in the "token output" of human speech.