Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
npn
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
4 ms
·
1.
▲
by
npn
5d ago
Show me a single example how is this jev thing better than a modern Bert solution? Or even llm if you claim about versatility. You can easily modify the llm inference code to make it predict a single token represent the classification choic
2.
▲
by
npn
5d ago
You got it reversed. Bert is encoder only and gpt is decoder only.
3.
▲
by
npn
9d ago
I read the readme and the guide file. There is just one thing I can comment: might as well solve the NP hard problems. I think you can do it easily, author. As you can already solved harder problems than those with your language.
4.
▲
by
npn
13d ago
People do love auditing everyone else’s donations.
5.
▲
by
npn
15d ago
As expected vibecoding bros cannot even read the manual properly. It is pretty trivial to pin a single provider for a model. Better yet, instead of calling the model directly, use presets instead. You can easily change the setting on openro
6.
▲
by
npn
16d ago
it is a way bigger model with extra 200B engram so of course the score improves. can't wait for deepseek v4.1 pro
7.
▲
by
npn
17d ago
Crazy that they still keep the price -- or actually decrease it, even -- despite it is a big improvement. I hope it retains some of the tps speed of the preview release though, 300 tps means gemini flash is no longer "the fastest optio
8.
▲
by
npn
18d ago
yeah raid or not you still get the hard limitation by the pcie lanes it is even worse with 40 macbooks. if 40 macbooks is all that take to serve a 1TB model with decent speed then you would see everyone selling the models for very cheap rig
9.
▲
by
npn
22d ago
luckily I have purchased a used amd mi50 32gb card for pretty cheap back then. while I haven't used it extensively, it feels pretty great having a backup plan that does not depend on any external 3rd parties.
10.
▲
by
npn
23d ago
openai did human crafted chain of thought dataset training. deepseek didn't have the resources so they attempted RL. doing RL correctly is hard because of the risk of model collapsing.
11.
▲
by
npn
24d ago
very important actually. just try to generate code for fresher frameworks/libraries. gemini sucks so bad in real work usage, everything it suggests are outdated and mostly useless.
12.
▲
by
npn
24d ago
Still refuse to search internet for stuff it thinks does not exist lol. And even when searching for internet, it still cannot suggest a up-to-date approach to the problem. For example I'm using crystal, it recently revamped the concurr
13.
▲
by
npn
28d ago
Because you can also use grok and a dozen other models. In fact grok is preferred choice for cursor right now, so obviously the other models get sidelined
14.
▲
by
npn
1mo ago
1. they still have revenue though. it might not enough to cover all the r&d but it is surely enough to cover the hardware cost. 2. people tend to ignore this, but the salary budget of a US frontier lab and chinese frontier lab is nowher
15.
▲
by
npn
1mo ago
I'm confused? Can you just define some presets and call them instead? With preset you can pinpoint a lot of things, especially the providers
16.
▲
by
npn
1mo ago
? Why are you assuming that I don't know about that. Actually have you really design practical webapps before? Because your post only have empty pretty words with no substance. Semantic html tags are just semantic, it is irrelevant for
17.
▲
by
npn
1mo ago
what hard is making a fully featured website, with panels no wider than 60ch. typically 60ch equal to 480px (font size 16px), so you need sidebars to the left and the right. which is fine, the holy grail was like that. but if you want to de
18.
▲
by
npn
1mo ago
I have Mi Notebook Pro. Not that bad actually. But they have stopped making premium laptop since then.
19.
▲
by
npn
1mo ago
I also did some experiment with ch many years ago. I found that 60ch is ideal width for block text for easy reading. too bad it is pretty hard to make websites with only 60ch wide.
20.
▲
by
npn
1mo ago
didn't zai already do that with their coding plan? I mean they surely had to pay users to use claude models (paying the differences). they also funded some newapi token resale websites.
21.
▲
by
npn
1mo ago
I hope it is glm air. We need more "small" models. Big models are more capable and useful, but for majority of tasks some smaller models can work just fine. It is funny that google gave up on this market, leaving the whole price r
22.
▲
by
npn
1mo ago
Ok that antise... We all know who you really want to criticize here.
23.
▲
by
npn
1mo ago
No but with 100% clean data you can easily train a model to filter ai generated content.
24.
▲
by
npn
1mo ago
I wonder how many posts in this thread are AI generated or shill posted. nobody ever reports that they installed it and ran it on production or something. personally I only use bun to replace yarn as script executor now. Used to follow it a
25.
▲
by
npn
1mo ago
There are like thousands sites with similar features all using newapi core. You can easily find them in Chinese tech forum linux.do
26.
▲
by
npn
1mo ago
> used up internet-scale data yet but it is still contain a lot of trash. you need better models to process those trash and create a curate dataset. this will happen again and again until there is no more juice to squeeze. and I'm s
27.
▲
by
npn
1mo ago
it is partly true, but like I said it is not 2025 anymore. models now get released more often, and still have notable progress so they can safely replace the old models while being faster/cheaper. and thank to chinese models the pricin
28.
▲
by
npn
1mo ago
> * For 3.6 and 3.7 Flash, introductory price expires on December 31, 2026. Starting January 1, 2027, $1.50/1M input tokens and $7.50/1M output tokens will apply. this is hilarious. it is not 2025 any more, by Jan 2027 there wi
29.
▲
by
npn
1mo ago
sorry LLM output is considered public domain in my country.
30.
▲
by
npn
1mo ago
nah, both point to the same model for deepseek. it's just openrouter weirdness.
More ›