Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
ycui7
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
5 ms
·
31.
▲
by
ycui7
4mo ago
Those resellers are simply just selling Kimi K2.5 or GLM5.1 as counterfeit Opus. We, Chinese, know how to play the counterfeit game for a long time in so many industry.
32.
▲
by
ycui7
4mo ago
in a few more months, when Chinese model gets to Mythos capacity and Fable still locked down. What Anthropic will say? Why can they just admit they are not the only people who know how to train an LLM model.
33.
▲
by
ycui7
4mo ago
At one point, I was thinking that if any of my customer send me a snail mail with an actual physical stamp on it, we will call the customer immediately and solve their problem.
34.
▲
by
ycui7
4mo ago
There are better and superior alternative of NE5532 these days. People should just move on. OPA1612 is the king in highest-end audio performance, at least on datasheet paper.
35.
▲
by
ycui7
4mo ago
does it really matter if LLM has conscious? if they produce working code, then it is useful, whoever if they have real conscious of fake intelligence. we don't know what the conscious in human brain is either.
36.
▲
by
ycui7
4mo ago
maybe it is wrong to spend 200B every year continuously to begin with.
37.
▲
by
ycui7
4mo ago
so google had spent too much money to build their own datacenter?
38.
▲
by
ycui7
5mo ago
I think what this actually means is that you can apply permanent residency in the US, but you can only get the physical green card outside of the US when the case is approved. So, the last step to get the card need to from outside the count
39.
▲
by
ycui7
5mo ago
This is not surprising at all. The biggest benefit of cloud model in terms of energy efficiency is that when running more than 1 requests, the power consumption of said GPU roughly stayed the same. The more concurrency requests the server c
40.
▲
by
ycui7
5mo ago
Every vendor defines their audio jack connector serial port differently. It is very dangerous to use 3.5mm jack. There is no pinout standard of using 3.5mm. Even as pure audio jack, the 3.5mm connector has two standards, with the difference
41.
▲
by
ycui7
5mo ago
The type of people who need spice is dead serious about accuracy. 1ppm error sometimes is not tolerable. So, an optimization in a game engine is definitely not suitable for engineering simulation.
42.
▲
by
ycui7
5mo ago
if the goal is to only get the median, you should not use sort. sort is O(nlogn). there are algo that give you medium at O(n), check quickselect.
43.
▲
by
ycui7
5mo ago
Exiting dGPU for gaming, but staying in the LLM world.
44.
▲
by
ycui7
5mo ago
B70 idles at 30W, while RTX PRO 4500 idles at 9W (measured to be 5W at wall). B70 runs at 1/3 token output rate of RTX PRO 4500 and consume 3X idle power when do nothing.
45.
▲
by
ycui7
5mo ago
When you get 4 of these, the idle power alone is 120W. That is a lot of electricity if left on 24/7. At that power consumption, you also end up being more expensive than API calls and many times slower. It starts to feel very stupid to
46.
▲
by
ycui7
5mo ago
At this speed, people end up paying more on electricity than api calls. (California electricity)
47.
▲
by
ycui7
5mo ago
You can get 120TPS (144 peak) with Qwen3.6-27B on RTX PRO 6000 with autoround when MTP enabled. It runs faster than sonnet api calls. 5090 gets maybe 100TPS with MTP
48.
▲
by
ycui7
5mo ago
Problem is the more B70 you have, the slower the inference it gets(due to terrible software atm). A single B70 is almost barely faster than CPU inference. If you have 4 B70, you might as well run interference on CPU and be faster with cheap
49.
▲
by
ycui7
5mo ago
Intel Arc B70 when released, can only produce 1/3 of the token of RTX PRO 4500. Well, it also cost 1/3 of RTX PRO 4500. It lacked software support the for the primary target application, running LLM. The officially supported vllm
50.
▲
by
ycui7
5mo ago
It is weird that the reviewer does not mention RTX PRO 6000 96GB, but mentioned RTX PRO 5000 72GB. 72GB RTX PRO 5000 is a special order, and much less people are aware of it. RTX PRO 6000 is known by mostly everyone in the LLM world. I cann