Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
WASDx
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
5 ms
·
31.
▲
by
WASDx
3mo ago
DeepSWE and FrontierCode are more realistic if you read up on what they actually measure. But the most realistic is to try it yourself. Benchmarks can only vaguely represent typical usage, and how you judge the result. Giving the same real
32.
▲
by
WASDx
3mo ago
Great explanation, thanks!
33.
▲
by
WASDx
3mo ago
> On top of that, doing research in the open amortizes the cost. Can you elaborate on this? I appreciate the open models but don't see the economics behind just giving them away like now.
34.
▲
by
WASDx
3mo ago
I see only these two possibilities: 1. If LLMs keep improving, burning models onto silicon becomes obsolete too fast and is not worth doing. Outcome: We keep getting better LLMs. 2. If LLM improvements slow down, they will be burned onto si
35.
▲
by
WASDx
4mo ago
Are you suggesting it should summarize the image in text or generate it in HTML or something else?
36.
▲
by
WASDx
4mo ago
Looking at some benchmarks, the latest ~30B Gemma/Qwen score similar as Claude or GPT versions that were released just one year earlier . That's crazy progress. I can't imagine how it will be in a few years.
37.
▲
by
WASDx
4mo ago
I think this is inevitable. Sooner or later, model-specific ASIC's will make economical sense. We're already seeing it happening with Taalas/Cerebras so I think it's sooner than 5 years. And inference is order of magnitu
38.
▲
by
WASDx
4mo ago
> distributed LLM inference This seems extremely inefficient considering data transfer between model layers if the model is distributed. I found this project called Petals that claim up to 4 tok/s for a 180B model although its repos
39.
▲
by
WASDx
4mo ago
I like this one, although its data seem to overlap with ECI. https://artificialanalysis.ai/trends
40.
▲
by
WASDx
4mo ago
https://chatjimmy.ai/ from Taalas also feels like that.
41.
▲
by
WASDx
4mo ago
I think their "code" ranking is biased towards visual aesthetics more than raw coding as the voters are just asked which generated website they prefer.
42.
▲
by
WASDx
5mo ago
I've had mostly problem-free experiences with intellij (ultimate-only feature I think). One click finds declarations both in business code and buried deep in libraries.
43.
▲
by
WASDx
5mo ago
gemma-4-31B-it-assistant is a 0.5B model. So it's performance would likely be comparable to other models of such size.
44.
▲
by
WASDx
5mo ago
I think this is the future. When models start converging at "really good" (which I think is already happening) then burning them into ASIC silicon is the natural next step. Harnesses can keep improving with a fixed model and the t
45.
▲
by
WASDx
5mo ago
I was impressed enough by AI finding vulnerabilities in source code, but doing it in binary executables is just amazing. This has so much potential, good and bad. And yet another lesson to not treat data as instructions. Sanitize all user i
46.
▲
by
WASDx
6mo ago
Creating a custom tuple class to use as key could be faster though. Nested map lookups have less efficient memory access patterns.
47.
▲
by
WASDx
6mo ago
Similar site with same features: https://xn--1-zfa.com/
48.
▲
by
WASDx
7mo ago
I think these limitations could be addressed by allowing trivial manual adjustments to the generated code before committing. And/or allowing for trivial code changes without a spec change. The judgement of "trivial" being tha
49.
▲
by
WASDx
9mo ago
I've managed a 100+ node cluster for years without seeing any corruption. Where are you getting this from?
50.
▲
by
WASDx
1y ago
You can customize it to get rid of all that. I set it to the "Robot" personality and a custom instruction to "No fluff and politeness. Be short and get straight to the point. Don't overuse bold font for emphasis."
51.
▲
by
WASDx
1y ago
Same. I recall the "stable volume" setting also eating cpu.
52.
▲
by
WASDx
1y ago
FYI here is a list of hundreds of engineering blogs: https://github.com/kilimchoi/engineering-blogs
53.
▲
by
WASDx
1y ago
The `/_cluster/reroute` endpoint lets you do that with a curl. We have aliases for common operations so I've never felt that I lack a CLI. I'm happy with Elasticsearch overall having a few years of experience.
54.
▲
by
WASDx
1y ago
I recall this article on QUIC disadvantages: https://www.reddit.com/r/programming/comments/1g7vv66/quic_i... Seems like this is a step in the right direction to resole some of those issues. I suppose not
55.
▲
by
WASDx
2y ago
https://iquilezles.org/ is a legend, see the articles and video tutorials. Aside from shadertoy I use https://glslsandbox.com/ (for some reason it has https errors now). It's the same concept and it ha
56.
▲
by
WASDx
2y ago
The closure itself is only being created once, it's essentially a singleton. Only if it would capture variables it would have to be recreated every iteration.
57.
▲
by
WASDx
2y ago
I would say no, but I wouldn't want to work there for ethical reasons either way.
58.
▲
by
WASDx
2y ago
It looks like CGI to me, the way to camera moves together with the depth of field and that things appear too shiny. They don't state anything about it so I don't know what to believe.
59.
▲
by
WASDx
3y ago
What is the gist of the algorithm? I'm not proficient at reading Go. Is it perhaps doing https://en.wikipedia.org/wiki/SWAR which is indeed SIMD?
60.
▲
by
WASDx
3y ago
Use an emulator and pick any retro variant.
More ›