4 ms·
Speaking of ASICs - how likely is it that as models get better we'll see someone baking a whole model directly into the silicon? It's like having l0 cache.
by foxrider 2mo ago
Speaking of ASICs - how likely is it that as models get better we'll see someone baking a whole model directly into the silicon? It's like having l0 cache.
- SJC_Hacker 2mo agoYou could do it but there would be no point, The only advantage over would be power consumption. And it would be quite expensive. At the rate models are improving, it would be obsolete in six months.
- HPsquared 2mo agoPower consumption and latency are very important on mobile
- naasking 2mo agoThey're important everywhere of course, but especially on mobile. If AI researchers figure out how to offload knowledge and expertise from reasoning weights, then a core reasoning ASIC linked to the knowledge would totally rock.
- foxrider 2mo agoYes, right now it would be obsolete in six months, but I also must add that this never stopped crypto miners from making new ASICs. However, with how useful Kimi is right now - at some point if someone makes a dedicated hardware board with "good enough" model for daily tasks - that would be a very sought after commodity.
- kaelwd 2mo agoOnly 8B currently but it's been done: https://taalas.com/products/ https://taalas.com/products/
- apimade 2mo agoThe only _public_ example we know. This is definitely being done with private models by HFT/quant firms, data processing agencies/orgs (large intelligence agencies, _every_ data analytics org, etc).