Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
WiSaGaN
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
7 ms
·
61.
▲
by
WiSaGaN
2y ago
The demon haunted world was my favorite book back in my college years more than 20 years ago.
62.
▲
by
WiSaGaN
2y ago
What if I told you most Chinese people really are more eager than the "CCP" to take over Taiwan?
63.
▲
by
WiSaGaN
2y ago
You are really underestimating how many users just forget they have such unused subscriptions, and how much of subscription based company monthly revenue is those that are not used at all.
64.
▲
by
WiSaGaN
2y ago
SearchGPT is bad because its underlying model is not a reasoning one. Deepseek one mentioned above is closer to deep research than searchgpt.
65.
▲
by
WiSaGaN
2y ago
Well you don't need to worry unless you are already on the list.
66.
▲
by
WiSaGaN
2y ago
I think the main point of local model is privacy set aside hobby and tinkering.
67.
▲
by
WiSaGaN
2y ago
I did math tests. Probably you did coding.
68.
▲
by
WiSaGaN
2y ago
My vibe question checking suggests otherwise. Even o3-mini-high is not as good as r1, even though it's faster than r1. Considering o3-mini is more expensive per token. It's not clear o3-mini-high is cheaper than r1 either even r1
69.
▲
by
WiSaGaN
2y ago
They are definitely step function improvements in regard to inferrance cost of the same level of reasoning intelligence.
70.
▲
by
WiSaGaN
2y ago
Americans need to understand that the Chinese are not obsessed with the US. They don't have a saboteur mindset. They want development not because they want US to fail and China to win. It's really sad to look at US state of affair
71.
▲
by
WiSaGaN
2y ago
Deepseek inference API has positive margin. This however does not take into account R&D like salary and training cost. I believe OpenAI is the same in these aspects, at least before now.
72.
▲
by
WiSaGaN
2y ago
This is actually good. I expect them to utilize this in code editing as well if there is some real efficiency gain under the hood.
73.
▲
by
WiSaGaN
2y ago
The one your laptop can run does not rival what OpenAI offers for money. Still, the issue is not whether third party can run it, it's just the OpenAI seems not putting API as their main product.
74.
▲
by
WiSaGaN
2y ago
Not necessarily. DeepSeek will probably only threaten the API usage of OpenAI, which could also be banned in the US if it's too sucessful. API usage is not a main revenue for OpenAI (it is for Anthropic last time I checked). The main c
75.
▲
by
WiSaGaN
2y ago
Brzezinski wrote about US not wanting an integrated Eurasia (the heartland), so that Europe can be more dependant on the US. Germany used to have more integration which make its industry thrive because it can take advantage of Russia's
76.
▲
by
WiSaGaN
2y ago
The current copyright system is clearly broken. Copyright itself is not a natural right but a construct that was created relatively recently to incentivize innovation within society. Currently, we still want to train models using copyrighte
77.
▲
by
WiSaGaN
2y ago
If what here says is true: https://x.com/teortaxesTex/status/1880768996225769738 , then R1 may as well just be the better model. You can scale up R1 with lower token count to achieve better than o1 high results.
78.
▲
by
WiSaGaN
2y ago
This is r1 not r1-lite-preview. Supposedly r1-lite is a much smaller model, where r1 here is the 600+B MoE, same size as Deepseek v3 they released earlier.
79.
▲
by
WiSaGaN
2y ago
In their benchmark, they have a tag "tuned" attached to their o3 result. I guess we need they to inform us of the exact meaning of it to gauge.
80.
▲
by
WiSaGaN
2y ago
It's interesting that nvidia takes somewhat opposite position on this "restricting adversarial nations" perspective. I think OpenAI wants the us government to restrict open source model such as deepseek or qwen from being acc
81.
▲
by
WiSaGaN
2y ago
I think the problem exacerbates as he tries so hard to appear more sincere and interesting than he truly is.
82.
▲
by
WiSaGaN
2y ago
I think it adds correct incentive to the public discourse so that you don't do bait and switch against public good will.
83.
▲
by
WiSaGaN
2y ago
Is there anything that the Chinese can do that's not seen as a threat to the American people?
84.
▲
by
WiSaGaN
2y ago
I don't think this proves that the LLM is just "pattern matcher". Human makes similar mistakes too, especially when under time pressure (similar to non-reasoning model that needs to "use system one" to generate answ
85.
▲
by
WiSaGaN
2y ago
There is also a curated benchmark just for those famous problems slightly variated: https://github.com/cpldcpu/MisguidedAttention/tree/main/eval
86.
▲
by
WiSaGaN
2y ago
He is not autistic, although perceived that way can be a competitive advantage in Silicon valley culture.
87.
▲
by
WiSaGaN
2y ago
It still fails my private physics testing question half the time, where claude 3.5 sonnet and openai o1 (both web version) most of the time passes. So I'd say close to SOTA but not quite. However given deekseek already has the r1 lite
88.
▲
by
WiSaGaN
2y ago
The withheld part is really a red flag for me. Why do you want to withhold a compute number?
89.
▲
by
WiSaGaN
2y ago
Unfortunately, the most likely outcome is they realize this is bad for both, and they should reach a private agreement under the table just like Google Apple no poaching agreement.
90.
▲
by
WiSaGaN
2y ago
Sam Altman said this around a year ago: "i expect ai to be capable of superhuman persuasion well before it is superhuman at general intelligence, which may lead to some very strange outcomes". I am wondering whether AI has any inp
More ›