Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
_aavaa_
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
3 ms
·
1.
▲
by
_aavaa_
5d ago
> Article says that there's no human intention or design directing the output we get I mean that’s objectively wrong for any model using RLHF.
2.
▲
by
_aavaa_
9d ago
If you are using their coding plan for coding, then yes you can easily hit such cache rates, with a good harness. I’m getting 97%.
3.
▲
by
_aavaa_
9d ago
I'm not defending their actions, but we should be clear about where the law currently stands: Anthropic was found to infringe because of the torrenting , not because of the training.
4.
▲
by
_aavaa_
9d ago
Their plans are still worth it if you use their models. You can see how many tokens you can except to get based on plan here: https://docs.z.ai/devpack/overview#estimated-token-allowance The max plan will provide ~1,10
5.
▲
by
_aavaa_
10d ago
Would that means Canadians would be protected by the GDPR and DMA?
6.
▲
by
_aavaa_
11d ago
I think cellphones are in fact PCs. Thinking of them as some other category is part of how Apple justifies it’s behaviour with regards to preventing you from installing whatever software you want on your own device; an action that would nev
7.
▲
by
_aavaa_
12d ago
Why is it stretching the definition? If compute sticks running windows are PCs why isn’t this?
8.
▲
by
_aavaa_
13d ago
The biggest one that needs changing would be the differential.
9.
▲
by
_aavaa_
16d ago
What’s better about it the say oh-my-pi?
10.
▲
by
_aavaa_
17d ago
Yes, and? This is their current off-peak pricing for their flash model [0]: $0.007 cache, $0.22 input, $0.66. 0.007 -> 0.003 0.22 -> 0.15 0.66 -> 0.60 Each one is now cheaper. [0]: https://api-docs.deepseek.com/quic
11.
▲
by
_aavaa_
17d ago
I don't see them turning people back from buying tokens through the API. Until then, I don't see why we should follow this argument.
12.
▲
by
_aavaa_
17d ago
They might need to step up their product offerings and offer cheaper.
13.
▲
by
_aavaa_
17d ago
What are you talking about? Current flash prices are 0.66 for output, this is dropping it to 0.60.
14.
▲
by
_aavaa_
17d ago
They’re trying to kidnap what we have rightfully stolen.
15.
▲
by
_aavaa_
21d ago
Whether on net they turn a profit as company overall is neither here nor there.. My point is that they are selling API tokens at a profit (or if being pedantic, then at a price higher than the cost to serve them ignoring research costs). An
16.
▲
by
_aavaa_
22d ago
Unlikely, api pricing includes a healthy profit margin (as near as we can tell from the outside) which they wouldn’t charge themselves.
17.
▲
by
_aavaa_
22d ago
Sure privacy (or legality) considerations are valid, and depending on the subscription or the API you use, you may or may not get that. But we are not talking about the same product anymore. Your $12k homelab does not provide the same produ
18.
▲
by
_aavaa_
22d ago
You are misinformed on token amounts. As one example, z.ai [0] gives you somewhere in the ballpark of 150-300 Mtok/week, so 1-2x your amounts. Plus the model has a 1M window, and will be smarter. [0]: https://docs.z.ai/
19.
▲
by
_aavaa_
22d ago
I don't think that math will work out. If you are okay with using only 7M tokens per day, and 7M from a small model, then you don't have very demanding needs. If you don't have demanding needs, I think you're unlikely
20.
▲
by
_aavaa_
22d ago
So instead of paying $20 or maybe 100$ a month for the equivalent amount of output, I can spend $10k on GPUs (plus a few more thousands for related hardware), then for ongoing electricity, and then have space to store these things. You'
21.
▲
by
_aavaa_
23d ago
I really like the 3D version, but I strongly believe you need to consider the number of tokens required to complete a task, it heavily impacts the results for certain models that rely heavily on test time compute (Glm-5.3-flash is the newes
22.
▲
by
_aavaa_
23d ago
Artificial analysis, the place where they get this data from, has a cost vs time chart. Go to https://artificialanalysis.ai/ and scroll down to the second graph under “Speed & Latency”. I think this is the most import g
23.
▲
by
_aavaa_
24d ago
Shame. Thanks for the info.
24.
▲
by
_aavaa_
24d ago
Do they officially support you using your subscription in other harnesses?
25.
▲
by
_aavaa_
24d ago
I disagree. The y-axis is some arbitrary intelligence score that we use as a proxy for performance on whatever our specific task happens to be. So it doesn't matter if a model is a 0, 1, or 20 along this axis, they are all useless for
26.
▲
by
_aavaa_
24d ago
Do they officially support you use their AI Pro subscription (or whatever the heck it's called this month, the one that gives you models in antigravity) in a 3rd party harness?
27.
▲
by
_aavaa_
24d ago
> That conveniently ignores the drilling, extracting, transporting, refining, and transporting again to make "polyethylene resin". It also ignores the energy requirements to grow the cotton. This isn't the gotcha you think
28.
▲
by
_aavaa_
29d ago
The original comment I replied to states: “A 50% reduction would change nothing, reduction to 0 would change nothing.” When talking about individual data centers. My point is that this is true of basically any since source of pollution. Y
29.
▲
by
_aavaa_
29d ago
It very well could be faster, but right now it isn’t.
30.
▲
by
_aavaa_
29d ago
It’s cheaper sure, but it’s very slow. It’s not a drop in replacement
More ›