Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
strangescript
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
14 ms
·
91.
▲
by
strangescript
1y ago
there are plenty of open source LLMs trained on harry potter, is that fine?
92.
▲
by
strangescript
1y ago
I read harry potter, and you ask me about a page, and I can recite it verbatim, did I just commit copyright infringement?
93.
▲
by
strangescript
1y ago
If you look at their finances you will realize that there is nothing left in that company.
94.
▲
by
strangescript
1y ago
Yeah, but why aren't they attacking that problem? Is it just impossible, because it would be a really simple win with regards to coding. I am huge enthusiast, but I am starting to feel a peak.
95.
▲
by
strangescript
1y ago
I have used claude code a ton and I agree, I haven't noticed a single difference since updating. Its summaries I guess a little cleaner, but its has not surprised me at all in ability. I find I am correcting it and re-prompting it as m
96.
▲
by
strangescript
1y ago
it feels like openai are at a ceiling with their models, codex1 seems to be another RLHF derivative from the same base model. You can see this in their own self reported o3-high comparison where at 8 tries they converge at the same accurac
97.
▲
by
strangescript
1y ago
In open source its super useful to be able to immediately have an idea of how big the model is and what kind of hardware it could potentially run on.
98.
▲
by
strangescript
1y ago
Speed is great, but you have to set the bar a little higher than last year's tiny models
99.
▲
by
strangescript
1y ago
The smaller models have been creeping upward. They don't make headlines because they aren't leapfrogging the mainline models from the big companies, but they are all very capable. I loaded up a random 12B model on ollama the other
100.
▲
by
strangescript
1y ago
The 0.6B model is wild. I like to experiment with tiny models, and this thing is the new baseline.
101.
▲
by
strangescript
1y ago
There is more complicated math systems that computers have solved, just like Chess, and Go. Systems that seemed impossible for a machine to beat and eventually they do. Coding will be exactly the same soon.
102.
▲
by
strangescript
1y ago
Yeah, anytime I am about to do some long multiplication, I start reaching for my calculator and stop, "no, you will go to the multiplication gym, and do this by hand, need to stay sharp"
103.
▲
by
strangescript
1y ago
I would use Aider if it had an agent mode. It needs to catch up with UX, frankly just have a mode that copies what claude code does.
104.
▲
by
strangescript
1y ago
curious why you went with Phi as the default models, that seems a bit unusual compared to current trends
105.
▲
by
strangescript
1y ago
Claude Code still feels superior. o4-mini has all sorts of issues. o3 is better but at that point, you aren't saving money so who cares. I feel like people are sleeping on Claude Code for one reason or another. Its not cheap, but its b
106.
▲
by
strangescript
2y ago
"There's only one component per file" If you don't do this in React, its your own fault.
107.
▲
by
strangescript
2y ago
I think they have realized that even if OpenAI is first, it won't last long so really its just compute at scale, which is something they already do themselves.
108.
▲
by
strangescript
2y ago
I have been using Claude Code, and its a bit pricey, but has exceeded my expectations so far.
109.
▲
by
strangescript
2y ago
I think its interesting they left out Gemini 2.0 Pro in the benchmarks which I find to be markedly better than flash if you don't mind the spend.
110.
▲
by
strangescript
2y ago
This is just a bad model. I can't believe they released it. Yes it does have few interesting properties, but nothing that justifies the speed or cost when people are running R1 distillations on toasters for nothing.
111.
▲
by
strangescript
2y ago
This is sooooo dependent on where you work and your current outside of work situation (ie, you need the money bad)
112.
▲
by
strangescript
2y ago
That is what 90% of science is really. There aren't a lot of truly "aha" moments where someone discovers something fundamental with little outside influence.
113.
▲
by
strangescript
2y ago
I think its about scope and expectations. I have had some form of AI code completer in my neovim config for 3 years. It works flawlessly and saves me tons of keystrokes. Sure sometimes it suggests the incorrect completion but I just ignore
114.
▲
by
strangescript
2y ago
OpenAI is pivoting away from MS. MS also has their own internal AI interests. Need to frame this for investors that doesn't look like we are losing out. "Nadella doesn't believe in AI anymore". Done and done.
115.
▲
by
strangescript
2y ago
AI will get faster and more energy efficient over time. Deploying physical hardware will never improve in any meaningful way that fixes the biggest problem, deploying X amount of things everywhere you need it. Its a non-starter.
116.
▲
by
strangescript
2y ago
Exactly, so many first of its kind moments in HL1. The first time you saw a scripted sequence, like wtf am I playing, this is amazing. HL2 was more refined, and the art style did help elevate the story telling.
117.
▲
by
strangescript
2y ago
As context sizes get larger (and remain accurate within the size) and speeds increase, especially inference, it will start solving these large complex code bases. I think people lose sight of how much better it has gotten in just a few year
118.
▲
by
strangescript
2y ago
Shouldn't they be loaded right now with all the BTC they bought when it was low?
119.
▲
by
strangescript
2y ago
Why do articles testing ancient LLMs keep getting posted?
120.
▲
by
strangescript
2y ago
* Yes I am aware I am not running R1, and I am running a distilled version of it. If you have experience with tiny ~1B param models, its still head and shoulders above anything that has come before. IMO there have not been any other quantiz
More ›