Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
curioussquirrel
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
6 ms
·
31.
▲
by
curioussquirrel
10mo ago
And regarding specific models - we obviously only tested a few languages, and there are thousands of them in the world. But Gemini seems to lead the pack basically regardless of the language your throw at it. YMMV.
32.
▲
by
curioussquirrel
10mo ago
In our previous tests, when it was 1.5 Pro against GPT 4o and Claude Sonnet 3.7, Gemini wasn't winning in the multilingual race, but it was definitely competitive. 2.5 and 3.0 seems to be big leaps from the 1.5 days. That said, it als
33.
▲
by
curioussquirrel
10mo ago
This is probably not a core concern for most HN readers, but at work we do multilingual testing for synthetic text data generation and natural language processing. Emphasis on multilingual. Gemini has made some serious leaps from 1.5 to 2.5
34.
▲
by
curioussquirrel
10mo ago
You could ask GPT for what it knows about you and use it to seed your personal preferences to a new model/app. Not perfect and probably quite lossy, but likely much better than starting from scratch.
35.
▲
My data should not be your cookie jar
(blog.avas.space)
3 points
by
curioussquirrel
10mo ago
|
0 comments
36.
▲
by
curioussquirrel
10mo ago
Animation generator without a single demo/example animation on the page?
37.
▲
by
curioussquirrel
10mo ago
+1 on this one! I only use LLMs once I'm done with writing, and basically using them as my editor. In case it helps anyone, here is my prompt: "You are a professional writer and editor with many years of experience. Your task is t
38.
▲
by
curioussquirrel
11mo ago
Same, even started adding new ssh keys to no avail... (I was getting some nondescript user error first, then unhealthy upstream)
39.
▲
by
curioussquirrel
11mo ago
A very similar workflow on my end, both beets as the main tagger/organizer and Picard to pick up whatever can't be processed through beets. Beets is amazing!
40.
▲
by
curioussquirrel
11mo ago
Surprised this did not get any attention whatsoever. Some really surprising findings in it: 75% of firms already have a positive return on investment from AI, less than 5% negative return. Also 46% of businesses leaders now use AI daily the
41.
▲
by
curioussquirrel
11mo ago
Although I loathe ads, I think that for new products where the presence of ads is disclosed clearly upfront, this is acceptable. Especially if this comes with a discount. We have Kindles with and without ads and people are generally fine wi
42.
▲
by
curioussquirrel
11mo ago
Azure copilot is really something. It can't see the context of the page it's embedded in, and the message you send is limited to 500 characters, so good luck pasting a log or configuration.
43.
▲
by
curioussquirrel
11mo ago
Photoshop now has a bunch of features that get used in professional environments. And in the end user space, facial recognition or magic eraser are features in apps like Google Photos that people actively use and like. People probably don&#
44.
▲
by
curioussquirrel
11mo ago
100% agree. Using it to polish your sentences or fix small grammar/syntax issues is a great use case in my opinion. I specifically ask it not to completely rewrite or change my voice. It can also double as a peer reviewer and point out
45.
▲
by
curioussquirrel
11mo ago
Interesting experiment, but I'd say aggregating the scores across models is far from ideal. Gemini 1.5 Flash got close-to-perfect scores on most languages (probably boils down to small variances in temp/top_k and statistical error
46.
▲
RSS Feeds Discovery Strategies
(blog.burkert.me)
4 points
by
curioussquirrel
11mo ago
|
0 comments
47.
▲
by
curioussquirrel
1y ago
Looking forward to Louis Rossmann's reaction. Wouldn't be surprised if this leads to a lawsuit over monopolistic behavior - this is clearly abusing their dominant position in the browser space to eliminate competitors in photos sh
48.
▲
by
curioussquirrel
1y ago
I wonder how much more telemetry and behavioral data Atlas needs to collect on its users given the rapid response system. And how well is all the session data stripped of sensitive information when transferred.
49.
▲
by
curioussquirrel
1y ago
You're right. There was also an experiment in Meta which tokenized bytes directly and it didn't hurt performance much in very small models.
50.
▲
by
curioussquirrel
1y ago
True. But does that scale to less common words? Or to other languages than English?
51.
▲
by
curioussquirrel
1y ago
This is a very good answer and I'm commenting only to bring more attention to it apart from voting up. Well put!
52.
▲
by
curioussquirrel
1y ago
Yes, but it would hurt its contextual understanding and effectively reduce the context window several times.
53.
▲
by
curioussquirrel
1y ago
Thanks for the explanation and for the tokenizer playground link!
54.
▲
by
curioussquirrel
1y ago
Why test for something? I find it fascinating if something starts being good at task it is "explicitly not designed for" (which I don't necessarily agree with - it's more of a side effect of their architecture). I also d
55.
▲
by
curioussquirrel
1y ago
Even GPT 3.5 is okay (but far from great) at Base64, especially shorter sequences of English or JSON data. Newer models might be post-trained on Base64-specific data, but I don't believe it was the case for 3.5. My guess is that as you
56.
▲
by
curioussquirrel
1y ago
Yep, there is still a room for improvement, but my point is that the LLMs are getting better at something they're "not supposed to be able to do". Quartiles sound like an especially brutal game for an LLM, though! Thanks for
57.
▲
by
curioussquirrel
1y ago
Thanks, Simon! I saw the same approach (numbering the individual characters) in GPT 4.1's answer, but not anymore in GPT 5's. It would be an interesting convergence if the models from Anthropic and OpenAI learned to do this at a s
58.
▲
LLMs are getting better at character-level text manipulation
(blog.burkert.me)
138 points
by
curioussquirrel
1y ago
|
108 comments
59.
▲
by
curioussquirrel
1y ago
Will check it out, thx!
60.
▲
by
curioussquirrel
1y ago
Could we finally get a decent opensource TTS app for Android? This project is very cool.
More ›