3 ms·
I am curious about doing the same kind of thing for compute shaders. I'm aware of Kompute.cc (which is Vulkan based) but haven't looked at their GEMM kernels, a
by raphlinus 4y ago
I am curious about doing the same kind of thing for compute shaders. I'm aware of Kompute.cc (which is Vulkan based) but haven't looked at their GEMM kernels, and also of wonnx for WebGPU ([1] is their GEMM code).
I'm also curious whether warp shuffle operations might be useful to reduce some of the shared memory traffic.
[1]: https://github.com/webonnx/wonnx/blob/master/wonnx/templates/matrix/gemm.wgsl https://github.com/webonnx/wonnx/blob/master/wonnx/templates...
- FL33TW00D 4y agoI'll be benchmarking the WONNX GEMM in the near future - would be curious to see how far a GEMM on WebGPU can be taken. Bram wrote some interesting exploration here: https://jott.live/markdown/webgpu_safari https://jott.live/markdown/webgpu_safari