6 ms·
Here's the mempipe benchmark latency for core<->core: https://github.com/MarginResearch/cannoli/blob/main/mempipe/benchmark/graph.png https://github.com/MarginR
by gamozolabs 2y ago
Here's the mempipe benchmark latency for core<->core: https://github.com/MarginResearch/cannoli/blob/main/mempipe/benchmark/graph.png https://github.com/MarginResearch/cannoli/blob/main/mempipe/...
https://raw.githubusercontent.com/MarginResearch/cannoli/main/mempipe/benchmark/mempipe_benchmark.txt https://raw.githubusercontent.com/MarginResearch/cannoli/mai...
You can see a massive improvement for shared hyperthreads, and on-CPU-socket messages.
In this case, about ~350 cycles local core, ~700 cycles remote core, ~90 cycles same hyperthread. Divide these by your clock rate as long as you're Skylake+ for the speed in seconds. Eg. about 87.5 nanoseconds for a 4 GHz processor for local core IPC.