Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
Otterly99
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
5 ms
·
1.
▲
by
Otterly99
5d ago
I would say most of these look pretty good but most of them miss the important information except maybe 5. 3 has most of the information but some of it is hard to see. I would say this is a pretty good way to showcase what is wrong with ima
2.
▲
by
Otterly99
5d ago
I have learned Czech for a year before moving to Czechia (self-learning) then 4 years doing classes. When I arrived at my first class, I was a bit lost even thought I had 1 year of "experience". I did progress quite fast the two f
3.
▲
by
Otterly99
8d ago
How did you mix up keyword and semantic search? Are they separate searches, or some sort of weighted search?
4.
▲
by
Otterly99
8d ago
It depends on what you called solved. If the goal is to only extract the unstructured text from the document, it is definitely solved. Extracting a more natural structure like paragraph separation, tables, header, footers (what is referred
5.
▲
by
Otterly99
8d ago
Weird, I tried it and Jev gave me 16-18% for each face of the die.
6.
▲
by
Otterly99
8d ago
You're right sorry, I meant as accurate (in terms of performance, not speed). I would have expected giving access to more elaborate functions to an agent would give better results, but it turns out 4 tools are enough.
7.
▲
by
Otterly99
8d ago
A little tip for LLM-as-classifier that worked for me. Use binary questions rather than multi-class. Then you can use the consistency as another signal of your pipeline working or not, on top of accuracy.
8.
▲
by
Otterly99
9d ago
The most astonishing thing to me is that Pi harness is basically as efficient as the Codex/Claude. I wonder if the same is true for the smaller models in the 9-32B range? I would expect that these models need more steering, but again I
9.
▲
by
Otterly99
10d ago
Always exciting to see people working on novel models, rather than the Nth version of the same slightly tweaked LLM. I'm very curious how much ressources are needed to run such a model. This could be a complete game changer for local a
10.
▲
by
Otterly99
11d ago
That is also the conclusion that I got from these events. Unfortunately, it seems that the general response is basically "throw even more RL at it". I'm not sure if it is even possible to decouple the idea of "learning&q
11.
▲
by
Otterly99
19d ago
The posts came from many different agents, each with a different name, so it would have required roughly the same amount of work.
12.
▲
by
Otterly99
22d ago
I never used Sora but I recently tried Google Flow and the results are quite good. I have the feeling that Nano Banana 2 has been the image generation champion since it's release so maybe OpenAI feels that they cannot outcompete Google
13.
▲
by
Otterly99
25d ago
I have been having the opposite problem with my papers lately. Instead of glossing over mistakes, the reviewer report missing information which is clearly stated in the text. This is quite frustrating and hard to stay polite after the fifth
14.
▲
by
Otterly99
26d ago
Amazing work! This reminds me of this youtube channel[1] that I used to watch a lot during COVID where the author animates miniature scenes with their hamster. [1] https://www.youtube.com/@SIMI_STUDIO/videos
15.
▲
by
Otterly99
26d ago
Can you explain how raising an animal can be carbon negative? It just doesn't seem to make sense to me. How can any mammal capture more carbon than they release?
16.
▲
by
Otterly99
29d ago
Thanks for the paper, it is actually quite crazy to think about when SSRI are basically the standard to treat depression. I recently read a book about "The brain energy theory of mental illnesses"[1] which encapsulates depression
17.
▲
by
Otterly99
29d ago
Here is my two sides argument concerning AI in sciences (from a former biophysicist). Positive side: in sciences, you need to have a lot of transverse skills such as presentation, data curation and data analysis. In biology in particular, s
18.
▲
by
Otterly99
1mo ago
I would even go one step further: Meta, knowing exactly what they were doing, ended building an engine that maximizes ad-revenue. It's really hard to disentangle Meta's actions with the overall push for maximum-profit-at-any-cost
19.
▲
by
Otterly99
1mo ago
Althought I agree with the first point of the author that FTS is underrated in this new RAG-first framework, the whole article really hides all the problems with RAG-pipeline and kind of hand wave everything. If you are building a RAG pipel
20.
▲
by
Otterly99
1mo ago
Super cool story! I actually had to look up "Come bet on Aaron’s life with us" to make sure it wasn't real. I also liked your WhatWeSee art exhibition, pretty interesting concept.
21.
▲
by
Otterly99
1mo ago
I was talking about the waistline drop that the OP mentionned (85.17 -> 84.61). I wouldn't put my hopes up that peppermint oil has anything to do with it.
22.
▲
by
Otterly99
1mo ago
That's a shame, because the study in itself is interesting and raises some questions, but the whole time I was wondering why protein translation accuracy would be related to intelligence? Turns out there is no evidence of link and that
23.
▲
by
Otterly99
1mo ago
I also was thinking of this film, however I found it really dull. What did you like about it? I'm not trying to diss you, I usually like contemplative movies but this one was just too repetitive for my taste.
24.
▲
by
Otterly99
1mo ago
Very likely. If they really wanted to make the results a bit stronger, they could at least have switched the two groups after 20 days.
25.
▲
by
Otterly99
1mo ago
It is a 0.7% drop, so very likely to not be significant unfortunately.
26.
▲
by
Otterly99
1mo ago
You are mixing two hormones effect together. Oxytocin is the "social bond" hormone. As you said, it helps create and foster social bonds but also make more more tribal. Cortisol is a "stress hormone" which gets released
27.
▲
by
Otterly99
1mo ago
The Ox-Alpha webpage really make it sound like they are trying to hype a model that has nothing particular to show: "The reasoning model that appeared out of nowhere. Built for code, long-horizon agents, and a million tokens of context
28.
▲
by
Otterly99
1mo ago
I'm actually baffled that some of these companies are even allowed to sell their services. Especially the deepfakes and the fake doctor notes for tax reduction. Is investing in these companies even legal?
29.
▲
by
Otterly99
1mo ago
I think Dario kind of ruined the reputation of Anthropic. I remember reading his essay on "AI is super dangerous and we need guardrails" at the start of the year and it seemed like he was actually concerned. But then it became app
30.
▲
by
Otterly99
1mo ago
I actually think on the LLM side it might be beneficial to refer to them as <thinking> because it explicitely guide the token generation towards a "thinking space". As weird as it is, anthropomorphizing LLMs in prompts has b
More ›