Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
Falvyu
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
4 ms
·
1.
▲
by
Falvyu
3y ago
If I were to guess: AVX512 vector and mask register take space on the die, which is a finite resource and could be allocated to other things. Moreover, power consumption may have also been a challenge, as seen in the early Intel implementat
2.
▲
by
Falvyu
3y ago
Zen 4 lacks AVX512_FP16 (for 16-bits IEEE floating point operations), AVX512_VP2INTERSECT and also lack the Advanced Matrix eXtension (AMX) set (if you consider that part of AVX512). https://twitter.com/InstLatX64/statu
3.
▲
by
Falvyu
3y ago
I think the main difference is that the CCL would compute 'enclosed areas' on the fly. To be more specific, it would first create a mask (black & white) of pixels that are similar to the clicked pixel, and then find connection
4.
▲
by
Falvyu
3y ago
You can probably make the process extremely fast by replacing the flood-fill approach with a something based on a Connected-Component Labeling algorithm ( https://en.wikipedia.org/wiki/Connected-component_labeling ). T