Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
LiamPowell
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
6 ms
·
31.
▲
by
LiamPowell
1mo ago
I always got the impression that C2PA is a way to say "this photo came from the BBC (for example) and they've only signed it because they've verified the supplied edit chain". It's always been obvious that one could
32.
▲
by
LiamPowell
2mo ago
You don't need to jailbreak it. Swipe from the top and press the big button labelled "dark mode".
33.
▲
by
LiamPowell
2mo ago
LLMs are particularly strong for this because it really doesn't matter if the code is a hacked together mess. Most user configs are already like that anyway.
34.
▲
by
LiamPowell
2mo ago
Latex rendering, inline images, inline browser showing what it's clicking on, being able to view a spreadsheet and then select a region to reference in the conversation, clickable links when it references a specific line number with mo
35.
▲
by
LiamPowell
2mo ago
Testing in that way will only tell you if the two devices work with the given cable. It won't tell you what sort of margins you have. Not all devices are made equal. A device isn't even necessarily better if it works with one cabl
36.
▲
by
LiamPowell
2mo ago
At a minimum, if plugging it in to a device and measuring waveforms is good enough for your application, you need an oscilloscope with 10 GHz bandwidth or so (don't quote me on that, I don't remember the max speed used by USB). Sc
37.
▲
by
LiamPowell
2mo ago
I also agree it's stupid when google puts UI over the webpage, but at least in that case they're not going out of their way to do it over an existing alternative. For anyone not aware of how bad google is getting about this: Did y
38.
▲
by
LiamPowell
2mo ago
> Firefox likes to draw the websites underneat the scrollbar. This was a choice that they actively made. It's not hard by default, they just chose to do something stupid because they think it looks pretty.
39.
▲
by
LiamPowell
2mo ago
Artists will at least show up further down the list usually. With songs it often just will not show the song at all when the title is an exact match, it will instead fill the entire results list with "lyric matches" that I'm
40.
▲
by
LiamPowell
2mo ago
The "gotcha" here, if you want to call it that, is that the picking the next token based on probabilities (either in training or in the sampler) is not an argument that LLMs are inherently flawed, especially when one reduces a LLM
41.
▲
by
LiamPowell
2mo ago
The natural next argument that I see a lot is "well it's still all probabilistic", which is technically true. However, don't the atoms that make up the cells that make up a human move around and interact based on probabi
42.
▲
by
LiamPowell
3mo ago
Sure, but it doesn't really fit there as a joke, it looks like it's just meant to be part of what they were trying to say.
43.
▲
by
LiamPowell
3mo ago
> Honestly? That's not just valuable—it's essential. I'm curious if you wrote this or had a LLM write it. I'm genuinely curious to be clear as I don't see why anyone would bother to go through a LLM to write such
44.
▲
by
LiamPowell
3mo ago
I suspected as much, and that brings us to the second issue where if we use a cohort of judges then the model that likes it's own code the most still wins.
45.
▲
by
LiamPowell
3mo ago
Here's the question I ask about every project that claims to make a LLMs output so much better: If it works so well then why would the model provider not just put it in the system prompt? Or in the case of interactive skills, why would
46.
▲
by
LiamPowell
3mo ago
This is not actually what the reviewer prompt says, or perhaps it is, I don't know since they don't make it public. I'm just pointing out how it seems like a bad idea to ask a LLM to make a subjective judgement on things like
47.
▲
by
LiamPowell
3mo ago
> You are a senior SWE-Bench reviewer, make no mistakes. I don't know what a better approach would look like while still remaining feasible, however this approach of telling a LLM to make a subjective judgement seems fundamentally f
48.
▲
by
LiamPowell
4mo ago
I'm not sure about Kalshi, however on most sports betting sites you actually are betting against the house. The betting sites all have in-house models (or piggyback off other sites) that are much better at predicting odds than the gene
49.
▲
by
LiamPowell
4mo ago
Most ad blockers do already use MV3, uBlock Origin is the only one still using V2 as far as I know. There are some drawbacks to V3, however none prevent creating an effective ad blocker, as demonstrated by the fact that many exist. Though s
50.
▲
by
LiamPowell
4mo ago
OP, I assume your comment[1] is getting flagged because of the obvious LLM usage. No one wants to interact with a comment that's not written by a human. [1]: https://news.ycombinator.com/item?id=48473753
51.
▲
by
LiamPowell
4mo ago
That don't fall back to Opus if their classifier thinks you might be working on anything that might be a competitor's product. It silently injects instructions into the prompt to sabotage your work. Read the policy above, it'
52.
▲
by
LiamPowell
4mo ago
The assumptions are so much worse than that: > Methodology & assumptions: No caching This is absolutely absurd. Claude code is of course using the cache (and this can be verified by looking at the traffic). It would be an incredibly
53.
▲
by
LiamPowell
4mo ago
> especially with all the stuff that SpaceX has put into orbit in recent years I've heard this repeated a lot but I've never seen anyone do the maths. StarLink satellites are all in very low orbits, so intuitively it seems like
54.
▲
by
LiamPowell
4mo ago
Maybe, but they certainly used it for marketing too. At the time they contacted a bunch of publications and gave them access but told them they could only share snippets of the output [1]. The only reason to set restrictions like that is ma
55.
▲
by
LiamPowell
4mo ago
They did it for 2 and 3, however it looks like they didn't for 4 and 5. GPT-2: https://slate.com/technology/2019/02/openai-gpt2-text-genera... GPT-3: https://www.itpro.com/technology/
56.
▲
by
LiamPowell
4mo ago
OpenAI has been pulling this marketing trick for years. Remember how GPT-3 was too dangerous to release? It's also probably bad PR if script kiddies have access to GPT model with no guardrails even if it doesn't enable any signifi
57.
▲
by
LiamPowell
4mo ago
I don't think I've ever seen LLM output as bad as this output. They sometimes write like that, but not every second sentence.
58.
▲
by
LiamPowell
4mo ago
What's this nonsensical video on the product page that allegedly shows an "all new thermal system"? https://videos.ctfassets.net/jy9s7k22hbg4/44R1LH71xb8uO4c9dD...
59.
▲
by
LiamPowell
4mo ago
LLMs are not yet capable of generating the level of marketing wankery seen here.
60.
▲
by
LiamPowell
5mo ago
TLDR: > SQLite does not (currently) accept agentic code. However the project will accept agentic bug reports that include a reproducible test case. Patches or pull requests demonstrating a possible fix, for documentation purposes, are w
More ›