Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
vient
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
7 ms
·
31.
▲
by
vient
1y ago
Oh, I see - hard refresh consistently shows HTTP/2 but after one or two soft refreshes it becomes HTTP/3 for me until next hard refresh. Edit: it is always second soft refresh for me that starts showing HTTP/3. Computers work
32.
▲
by
vient
1y ago
Limited to macOS? Does not reproduce in FF 141 and 142 on Windows.
33.
▲
by
vient
1y ago
AMX is indeed a very strong feature for AI. I've compared Ryzen 9950X with w7-2495X using single-thread inference of some fp32/bf16 neural networks, and while Zen 5 is clearly better than Zen 4, Xeon is still a lot faster even con
34.
▲
by
vient
1y ago
True, but that has nothing to do with tagged pointers.
35.
▲
by
vient
1y ago
Memory page size should be transparent for tagged pointers (any pointers, really), I don't see how they can be affected. You have an object at address 0xAB0BA, does the size of underlying page matter?
36.
▲
by
vient
1y ago
Would also be nice to remove empty categories from tree view. For example, right now you can uncheck VSX and still see "Memory Operations - VSX Unaligned ..." full of empty tags.
37.
▲
Relwarc: Mining HTTP endpoints from client-side JavaScript with static analysis
(blog.secsem.ru)
4 points
by
vient
1y ago
|
0 comments
38.
▲
Where do e-readers go from here?
(goodereader.com)
1 points
by
vient
1y ago
|
1 comments
39.
▲
by
vient
1y ago
But what's the point in acknowledging numerical issues outside of [-1,1] if polynomials do not even work there, as author explicitly notes?
40.
▲
by
vient
1y ago
> Kalimba, VE > No idea what this is, and Google won’t help me. Seems that Kalimba is a DSP, originally by CSR and now by Qualcomm. CSR8640 is using it, for example https://www.qualcomm.com/products/internet-of-th
41.
▲
by
vient
2y ago
They have some NVIDIA support in the form of external project: https://github.com/openvinotoolkit/openvino_contrib/tree/mas...
42.
▲
by
vient
2y ago
For me fails on identical setup with GET https://gnikoloff.github.io/webgpu-sponza-demo/assets/Sponza.bin net::ERR_CONNECTION_CLOSED 206 (Partial Content)
43.
▲
by
vient
2y ago
LLVM_USE_SPLIT_DWARF may help with this, some recent measurements: https://www.tweag.io/blog/2023-11-23-debug-fission/#an-examp...
44.
▲
by
vient
2y ago
Sounds right for a panic co-founder.
45.
▲
by
vient
2y ago
It is a featureful formatting library, not simply a library for slow printing of ints and strings without any modifiers. You can't create a library which is full of features, fast, and small simultaneously.
46.
▲
by
vient
2y ago
Wow, changing `count` type from uint64_t to uint32_t or int radically changes results - now gcc gets 26500 and clang gets 25000, that's just 1.7 times slower than current best solution. So you can get 25k with following code, clang -Of
47.
▲
by
vient
2y ago
Nice. Did some quick tests with your code on site, got score of ~34000 - best solution is around 14700, so this one is only 2.3 times slower. Used clang with -Ofast -march=native -static. Funnily, gcc gets only 54000 with the same options,
48.
▲
by
vient
2y ago
Note that "standard C++" solution uses std::cin while optimized one uses mmap - completely different things, a lot of speed comes just from that. Would've been nice to compare with solution having optimized input and otherwis
49.
▲
by
vient
2y ago
> I've been wanting a language with that feature Python, which realization CPython is mentioned in the article, has arbitrary precision integers.
50.
▲
by
vient
2y ago
I like speedscope.app for viewing flamegraphs. It is more interactive than traditional SVG flamegraphs, and what is relevant here is a "sandwich" view - basically a sorted list of all functions, you see what function was spent the
51.
▲
by
vient
2y ago
And, as expected, the very first reference in the post is to the Tom7's video. Of course it would be.
52.
▲
by
vient
2y ago
There is a QEMU fork used by Nyx fuzzer, may be interesting to you https://github.com/nyx-fuzz/QEMU-Nyx Basically, for the fuzzing purposes speed is paramount so they made some changes to speed up snapshot restoring. D
53.
▲
by
vient
2y ago
> don't randomly hit up bloggers Not sure if satire or you just don't know who lcamtuf is.
54.
▲
by
vient
2y ago
Looking at OpenVINO repo activity, seems that most of them got relocated and still work at Intel.
55.
▲
by
vient
2y ago
It worked only on first batches/microcode versions of hybrid CPUs, since then Intel had completely disabled AVX-512 on them.
56.
▲
by
vient
2y ago
Not the best time for testing OpenVINO at least, they plan to release NPU support in 2024.1 as I understand, current version 2024.0 has very limited NPU support if any.
57.
▲
by
vient
2y ago
Also sched-ext which seems close to be mainlined and is already a default scheduler in CachyOS: https://github.com/sched-ext/scx
58.
▲
by
vient
3y ago
Huh, seems you are right. I've read about "not really real" AVX-512 implementation in Zen 4, and also saw that on my workload turning on AVX-512 on Zen 4 indeed gives almost nothing (compared to x2 speedup on Intel). But now
59.
▲
by
vient
3y ago
Not clear how AVX-512 can provide 2x speedup on Zen 4, even more 10x (if they are comparing with AVX2 which is obvious assumption). Zen 4 does not really have proper AVX-512 units, and 10x means that there was no vectorization at all before
60.
▲
by
vient
3y ago
Huh, for me as a malware analyst previously and a reverse engineer in general, decompilation is the most important part of such tools. It's all about speed, pseudo-C of some kind lets you roughly understand what's going on in a fu
More ›