Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
thecompilr
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
8 ms
·
61.
▲
by
thecompilr
6y ago
What is dead may never die. Server Push never gained any traction sadly, despite having a great potential.
62.
▲
by
thecompilr
6y ago
For me a good analogy to API is driving a car. The interface (API) to the car is the steering wheel, the pedals, and some common controls like signals and the horn. You only need to learn to drive one car, and you can drive them all. But un
63.
▲
by
thecompilr
7y ago
Exactly
64.
▲
by
thecompilr
7y ago
A small nit that bugs me in the original post > The difference between AES128 and AES256 appears to be approximately 26%, rather than the 40% one might expect (AES128 does 10 rounds, AES256 does 14 rounds). There may be some limiting fac
65.
▲
by
thecompilr
7y ago
There is user mode wireguard for Linux, it is wireguard-go: https://git.zx2c4.com/wireguard-go/ . There is also BoringTun: https://github.com/cloudflare/boringtun which is faster Disclaimer: I wrot
66.
▲
by
thecompilr
7y ago
There is a setting in HTTP/2 called SETTINGS_MAX_CONCURRENT_STREAMS, if set to 1 it works like HTTP/1.1, with no multiplexing. Setting it to 4~8 would make it behave in a similar way a browser actually does with HTTP/1 (creat
67.
▲
by
thecompilr
8y ago
I have been getting the mio question a lot. Mio, in theory, is not very far off of what I ended up implementing, but it is missing a few features that I really needed, and building those features on top of Mio would be harder than just usin
68.
▲
by
thecompilr
8y ago
Currently the library part supports Windows, and it can be used from C# or C++, if you want to use something like https://docs.microsoft.com/en-us/uwp/api/Windows.Networking.... However there is no native sup
69.
▲
by
thecompilr
8y ago
It is still quite expensive, the kernel module is still faster. Performance depends a lot on your kernel version and network card. We measured anywhere from 20% to 40% advantage to kernel module. But we are working on additional performance
70.
▲
by
thecompilr
8y ago
Windows PowerShell is also smooth as butter. After using it for a bit going back to macos or linux terminals feels sluggish.
71.
▲
by
thecompilr
8y ago
I have to strongly disagree with Linus here. While he got the history up to this point right, in the present it is very easy to develop for multiple platforms at once. CI platforms make it super easy. Cloudflare made a big effort to make al
72.
▲
by
thecompilr
8y ago
From the only number I see, the specint score, it is 40% faster than Qualcomm Centriq socket vs socket, and even 6% core vs core. Which is a really great if true. Wonder what the power draw on that thing though.
73.
▲
by
thecompilr
8y ago
Symmetric MultiThreading (SMT) - that would be Simultaneous MultiThreading
74.
▲
by
thecompilr
8y ago
FWIW at Cloudflare we were running brotli for dynamic content for a while now. However the Cloudflare gzip library is much faster than brotli. You can find a benchmark in here: https://blog.cloudflare.com/arm-takes-wing/
75.
▲
by
thecompilr
8y ago
Can't say much changed over the years https://lkml.org/lkml/2003/2/26/158
76.
▲
by
thecompilr
8y ago
You could, but the premise of using a tree was to avoid unpredictable rehashing latency, if you start compacting the tree every now and then, you basically pay the same price.
77.
▲
by
thecompilr
8y ago
Thank you. Exactly. Even worse each node could live in a different page. Hash maps can span multiple pages, but the lookup should succeed in 1 or 2 tries for a good hash function.
78.
▲
by
thecompilr
8y ago
Depending on implementation and usage, radix trees can get extremely segmented, slowing down significantly over time.
79.
▲
by
thecompilr
8y ago
Yeah, but with the side effect of clearing the flag registers. So can't be used between operations that rely on flags. Alternatively it is great to break false flag dependency.
80.
▲
by
thecompilr
8y ago
You imply that there is a delay between the promise and the push, but it is not necessarily so. In fact the promise and the data may be sent in the same packet.
81.
▲
by
thecompilr
8y ago
jpegtran does not transcode. It performs lossless optimization of the Huffman coefficients.
82.
▲
by
thecompilr
8y ago
Yes and no. http://infocenter.arm.com/help/index.jsp?topic=/com.arm.doc.... > The ARM Advanced SIMD architecture, its associated implementations, and supporting software, are commonly referred to as NEON techno
83.
▲
by
thecompilr
8y ago
I don't think Intel is going to die anytime soon, but I do hope for a stronger competition between ARM based server, Intel and AMD.
84.
▲
by
thecompilr
8y ago
> which will run circles around anything arm based. Do you have a single number to back that up?
85.
▲
by
thecompilr
8y ago
I don't even claim that ARM is better than Intel on SIMD workloads. Quite the opposite. Intel has AVX2 and AVX512 that rip ARM apart on most highly parallel workloads. This very specific snippet benefits a bit more from ARM SIMD due to
86.
▲
by
thecompilr
8y ago
Our jpegtran is still faster. We use libjpegturbo for other manipulations though.
87.
▲
by
thecompilr
8y ago
I ended up writing asm for the important part: https://github.com/cloudflare/jpegtran/blob/vlad/arm/jchuff_...
88.
▲
by
thecompilr
9y ago
Intel encourages employees to file all sorts of crappy patents, but I don't see it as a bad or evil practice, but rather as a defensive strategy. Meaning that they don't really enforce patent usage on others, instead they need to
89.
▲
by
thecompilr
9y ago
Encryption is one thing Intel has a lead in, both in single thread, and as a platform. But even there QC is sufficiently fast, and crypto although extensively used actually uses very little CPU. https://blog.cloudflare.com/h
90.
▲
by
thecompilr
9y ago
Those are indeed workloads, specific for a webserver and even more specific for CF. Many other workloads that we don't care about that Intel can handily win of course. We don't have use for most Intel extensions, like AVX512 and b
More ›