Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
WhitneyLand
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
6 ms
·
61.
▲
by
WhitneyLand
4mo ago
First great work. Reminds that I wish there was a modern way to do this for the words people speak and write online with. I want to literally know when people started putting literally twice in sentences. Ngram seems is out of date a piece
62.
▲
by
WhitneyLand
4mo ago
Not sure you can go by that alone. It’s a solo project so it’s common to not use PRs in that case. People review in different ways, sometimes before the commit or at a different cadence. I get the concern, but I think it would be easy to te
63.
▲
by
WhitneyLand
4mo ago
The simplicity is really cool. To let LLMs compose music I chose json for context efficiency, but this seems like it could be better choice, simple, efficient, already a real format. https://github.com/whitneyland/riffm
64.
▲
by
WhitneyLand
4mo ago
“there aren't great corpora of training data that would connect a MusicXML representation to sheet music images or to audio” It may not be necessary…a lot of the training pairs/data for this could probably be procedurally created
65.
▲
by
WhitneyLand
4mo ago
If confirmed this is really cool and impressive work. Honestly curious how many years before it can be one shotted in a coding harness with Fable.next by someone who’s not a linguistics expert. Develop, test, and rank hypotheses about the
66.
▲
by
WhitneyLand
4mo ago
Most of the problem is that for voice chat, you usually get no reasoning at all and no tool use at all to research or ground assumptions. For example for voice ChatGPT still uses a quantized gpt40 non-reasoning model that hallucinates prett
67.
▲
by
WhitneyLand
4mo ago
It’s crucial to use for driving/walking. One problem has been ChatGpt/Claude apps don’t really do this well. They use weak and/or non-reasoning models for voice interaction and the UX is not optimized for hands free. I wrote
68.
▲
by
WhitneyLand
4mo ago
At a glance it looks like it could be just iterative texture sampling. The difference is when creating each pixel, there’s no coordinate to look up, instead it’s using only a set of rules like Conway’s game of life. But the rules come from
69.
▲
by
WhitneyLand
4mo ago
OMB = “Office of Management and Budget”. It’s a White House office run by Russell Vought, highly ideological maga institutionalist.
70.
▲
by
WhitneyLand
4mo ago
You can tell you’re slop weary when it feels enjoyable just to read text at length written purely by a human.
71.
▲
by
WhitneyLand
4mo ago
Yeah, same experience. It turned out that objectively better answers were not that easy to find plus the expense plus it’s slow.
72.
▲
by
WhitneyLand
4mo ago
I’d like to see embedding of actual video clips become practical in this type of workflow. Frame level embedding it covering a lot, but can miss out on a lot of action related searches.
73.
▲
by
WhitneyLand
4mo ago
Because having more powerful tools is a different problem from using them well. When GUI apps first made it easy for anyone to use 100 fonts and colors in a document, it did not make the quality of design go up. The tools became more power
74.
▲
by
WhitneyLand
4mo ago
The PR quality problem is legit and needs a solution. But saying his opinion hasn’t changed on this: “the main and most important reason why GenAI tools do not work for me is that they do not make me any faster.” It’s been a year and agents
75.
▲
by
WhitneyLand
4mo ago
I would guess you did not first write “CRISPR is an overhyped approach”, then after careful reflection decide, I don’t think that quite captures the intensity, better go with “extremely overhyped”. The comparison is kind of a category error
76.
▲
by
WhitneyLand
4mo ago
The frontier labs commonly trade spots at the top of the benchmarks with each new model release. The timing of these price cut discussions says to me OpenAI has no imminent release that will be edging out Mythos/Fable. If so the questi
77.
▲
by
WhitneyLand
4mo ago
This does not look good for Google. On one hand, industrial research is different from academic research. There’s no tenure and not the same level or presumption of academic freedom. Fair enough. The problem is they specifically wanted to b
78.
▲
by
WhitneyLand
4mo ago
Some things that jump out as unfortunate: - Reductionist analogies like how Microsoft Word is not conscious therefore AI is not. - Dismissive in saying LLMs are not capable of moral reasoning. Maybe he meant agency or responsibility? - Buil
79.
▲
by
WhitneyLand
4mo ago
Hate it when incompressible fluids are mentioned like it’s literally true with any qualification or explanation. Iiuc water might compress ~50% at the right place in the Earth’s mantle, maybe just not looking much like liquid.
80.
▲
by
WhitneyLand
4mo ago
Yeah and it’s pretty memory efficient with only 8 attention layers so at int8 in 16GB ram maybe you still get 64k-128k context. The part I hate though is that I’d bet none of the performance claims are based on int8. Why do we care about bf
81.
▲
by
WhitneyLand
4mo ago
I don’t think so, the HF weights are bf16 which means 24GB + cache/overhead. It sounds like marketing spin where the performance claims are based on BF16 and the “runs in 16GB” claim is on a totally different quantized version.
82.
▲
by
WhitneyLand
4mo ago
”they ported a small-to-medium service from Java to Rust. The result was such a huge performance drop that it wouldn't meet their minimum requirements” That result would say less about performance of languages than it would about com
83.
▲
by
WhitneyLand
4mo ago
The not clear comment is valid by either interpretation. To a lot of us it’s not clear that’s what’s happening. It’s speculation and one possibility. It may also be a secondary consideration and not the primary gating factor. Anthropic has
84.
▲
by
WhitneyLand
4mo ago
The article made up the claim it’s not from the paper itself. There was some improvement in cognitive scores, but no placebo group. Without a placebo group, there are a lot of explanations for the data.
85.
▲
by
WhitneyLand
4mo ago
“Maybe my own tastes are saturated now” It might be saturated for smaller scopes of work, but it’s not hard to see the cracks when you scale up what you ask of SOTA models/agents. One example, to try and single shot prompt coding a Cha
86.
▲
by
WhitneyLand
4mo ago
“Like meditation, journaling, and other contemplative practices” The big difference is that meditation and journaling do not require a belief that you are communicating with supernatural beings. “I don't think intelligence and spiri
87.
▲
by
WhitneyLand
4mo ago
So in other words if the research had tried to assign a severity to the mistakes models made the entire paper may collapse as uninteresting?
88.
▲
by
WhitneyLand
5mo ago
Agreed. This is the kind of design wisdom that’s both true and difficult to win an argument over. It reminds me of arguments related to over-engineering and complexity. The principles are super important to having a codebase that scales and
89.
▲
by
WhitneyLand
5mo ago
It’s a bizarre feeling isn’t it? Sorry you’re having to defend the act of thinking. The problem is you can’t defend it right? Someone could say your evidence came from a prompt: “Take this article and reverse engineer a hypothetical unpol
90.
▲
by
WhitneyLand
5mo ago
But you don’t really know that do you? The other day I was criticized for posting a comment people thought was AI but was actually not. I’m starting to notice that more often with others as well. Happens sometimes to those who were always
More ›