Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
thethirdone
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
9 ms
·
31.
▲
by
thethirdone
2y ago
I would agree the "don't blunder" and "punish opponents blunders" are harder than endgame knowledge. However, knowing the basics of endgames is actually important to closing out games. Specifically, knowing KQvK, KR
32.
▲
by
thethirdone
2y ago
This seems like a very shallow way of thinking. "Losing all respect for the person" implies that you think this is NEVER an appropriate way to address someone. Phrasing a disagreement of opinion as a question of reasoning is often
33.
▲
by
thethirdone
2y ago
It is however important to note that allowing kings to be taken can result in both kings being traded which by the scoring function would be considered neutral. This possibility does not occur in normal games so a search with that modificat
34.
▲
by
thethirdone
2y ago
> "Completeness" is not about finishing in finite time, it also applies to completing in infinite time. Can you point to a book or article where the definition of completeness allows infinite time? Every time I have encountere
35.
▲
by
thethirdone
2y ago
What is the actual math involved? From following the link in [1] I found almost no math content.
36.
▲
by
thethirdone
2y ago
The sample answers for the horse race question are crazy. [0] Pretty much all the LLM really want to split 6 horses into two groups of three. Only LLAMA 3 makes the justification that only 2 horses can be raced at a time, but then gets its
37.
▲
by
thethirdone
2y ago
From the linked report: > The material required for construction of the cable is a carbon nanotube composite: currently under development and will be available in 2 years. Its actually crazy that the report thinks that we would have the
38.
▲
by
thethirdone
2y ago
You have not done a good job explaining/proving how they are wrong. Most of your response is only addressing a single paragraph that mentioned Netflix. > The notion that traffic ratios have anything to do with whether it makes sense
39.
▲
by
thethirdone
2y ago
The article does a good job of clearly stating that the 48% is compared to 2019, but words like "surge" and "spike" do not closely match that factual basis and imho are misleading. Attributing the 48% to AI specifically
40.
▲
by
thethirdone
2y ago
What is the surrounding text for "versus last year"? I cannot find "versus" or "year" in the article.
41.
▲
by
thethirdone
2y ago
The specific issue of using a LLM to make org decisions on how to downsize actually affects nearly all jobs equally. From what I can tell, most programmers are more ok with LLMs directly replacing them than artists are. I tend to agree that
42.
▲
by
thethirdone
2y ago
> You can say ‘the recent jumps are relatively small’ or you can notice that (1) there is an upper bound at 100 rapidly approaching for this set of benchmarks, and (2) the releases are coming quickly one after another and the slope of th
43.
▲
DeepSeek Coder V2 Released
(github.com)
1 points
by
thethirdone
2y ago
|
0 comments
44.
▲
by
thethirdone
2y ago
> They did include an 8-rollout version in the tables? I can't say as to why they didn't try using a bigger model than Llama 3 8B. That was a typo. I meant a > 8 rollout version. It doesn't seem like they have hit massi
45.
▲
by
thethirdone
2y ago
I am confused about how the MCTSr algorithm actually is. It is not clear how it is better than simply mutating potential answers (by LLM) and sorting by LLM self-eval. I have a hard time understanding how many LLM evals MCTSr actually does.
46.
▲
by
thethirdone
2y ago
I rather strongly doubt the "Norway has zero" statement. It does not directly reference any study nor does any other article stating the same. I don't doubt that it is lower or even rounds to 0 per 100,000, but actually 0 is
47.
▲
by
thethirdone
2y ago
> It has been shown that LLMs are unable to learn concepts beyond the first level of the Borel Hierarchy, which imposes severe limits on the ability of LMs, both large and small, to capture many aspects of linguistic meaning. This means
48.
▲
by
thethirdone
2y ago
It seems I read your comment as more anti-demilitarization than it was. It seemed like you were implying the US would become less safe by implementing the suggestions. If your intent was only that improving safety in one specific way does n
49.
▲
by
thethirdone
2y ago
I do not agree that it needs a qualifier. My opinion is that safety as a whole would improve with those four suggestions implemented. You can disagree, but if you think it would make safety worse you probably should point out how those chan
50.
▲
by
thethirdone
2y ago
Do note that it has 236 B parameters which makes the weights ~450 GB.
51.
▲
by
thethirdone
2y ago
> IDK about your philosophy, but any death that is 100% unavoidable invalidates the entire system of business that has been built. Assuming you meant avoidable , I don't think there is ANY system of business that avoids EVERY avoid
52.
▲
by
thethirdone
2y ago
relevant: https://stackoverflow.com/questions/74417624/how-does-clang-...
53.
▲
by
thethirdone
3y ago
Your analysis would pretty similarly apply to 5x5 Gardner's chess which has been weakly solved. Simple tree search can be quite effective as the branching factor is cut down a significant amount (2x in the starting position). Gardner&#
54.
▲
by
thethirdone
3y ago
It is not clear to me what you would define as the reason to return paradox from `halts`. It is pretty clear you can make a `halts` function that returns halts, loop or unsure. Renaming unsure to paradox would give a valid version of your 3
55.
▲
by
thethirdone
3y ago
Its not. Mostly it is an argument that turing-completeness is too mathematical to be practical. > I still had the mainstream belief that AI would be just as capable as humans.
56.
▲
by
thethirdone
3y ago
Quite a comprehensive summary of the implications of Turing-Completeness. There aren't any outright errors that pop out to me which is high bar if other articles are anything to judge by. > This was “easy” because the other major th
57.
▲
by
thethirdone
3y ago
> Tokenization counts: the impact of tokenization on arithmetic in frontier LLMs - https://arxiv.org/abs/2402.14903 Very interesting paper. It does make sense to me the R2L chunking would be better than L2R chunking
58.
▲
by
thethirdone
3y ago
The nature of numbers as `A 10^(n+1) + B 10^n` for digits `XXXABXXX` is a very important relationship for doing any arithmetic. As you tokenize strings of digits, you lose the position information within the token make more complicated rela
59.
▲
by
thethirdone
3y ago
It has seemed to me that the GPT would be considerably better at numbers if it just considered each digit as a token. Has anyone actually done an experiment to test this? I wouldn't disbelieve that the grouped version is actually bette
60.
▲
by
thethirdone
3y ago
A super basic search doesn't show any studies showing the two cup experiment (with results showing 3x reward to get 50% of people to switch) and TFA doesn't cite any sources. If anyone can cite a paper for it, I would appreciate i
More ›