Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
ozgune
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
8 ms
·
31.
▲
by
ozgune
1y ago
Question to author. Are you planning to publish CH benchmarks (TPC-C and TPC-H combined)? I'd expect Aurora to perform much worse on CH than on TPC-C/H. That's because Aurora pushes the WAL logs to replicated shared storage.
32.
▲
Trump to change controversial Biden-era restrictions on AI chip exports
(cnn.com)
1 points
by
ozgune
1y ago
|
0 comments
33.
▲
California overtakes Japan to become the world's fourth largest economy
(edition.cnn.com)
90 points
by
ozgune
1y ago
|
108 comments
34.
▲
by
ozgune
1y ago
> However, a wise man once said: “[It] ain’t about how hard you hit. It’s about how hard you can get hit and keep moving forward; how much you can take and keep moving forward.” Ujiharu may have lost Oda Castle nine times, but that means
35.
▲
by
ozgune
1y ago
I had a related, but orthogonal question about multilingual LLMs. When I ask smaller models a question in English, the model does well. When I ask the same model a question in Turkish, the answer is mediocre. When I ask the model to transla
36.
▲
by
ozgune
1y ago
In March, vLLM picked up some of the improvements in the DeepSeek paper. Through these, vLLM v0.7.3's DeepSeek performance jumped to about 3x+ of what it was before [1]. What's exciting is that there's still so much room for
37.
▲
by
ozgune
1y ago
I feel the article presents the data selectively in some places. Two examples: * The article compares Gemini 2.5 Pro Experimental to DeepSeek-R1 in accuracy benchmarks. Then, when the comparison becomes about cost, it compares Gemini 2.0 Fl
38.
▲
Economists say there's a math error in Trump's tariff formula [video]
(cnn.com)
9 points
by
ozgune
2y ago
|
2 comments
39.
▲
by
ozgune
2y ago
I agree with the blog post that using K8s + containers for GPU virtualization is a security disaster waiting to happen. Even if you configure your container right (which is extremely hard to do), you don't get seccomp-bpf. People start
40.
▲
by
ozgune
2y ago
Agreed. Here are three things that I find surreal about the s1 paper. (1) The abstract changed how I thought about this domain (advanced reasoning models). The only other paper that did that for me was the "Memory Resource Management i
41.
▲
by
ozgune
2y ago
On our side, we first saw the Cloudflare outage. Then, Docker Hub started failing, followed by GitHub API errors. It's amazing how much of the internet runs / depends on Cloudflare these days. Thank you for keeping the lights on.
42.
▲
vLLM V1: A Major Upgrade to vLLM's Core Architecture
(blog.vllm.ai)
2 points
by
ozgune
2y ago
|
0 comments
43.
▲
by
ozgune
2y ago
The official announcement from the Qwen team is also on HN right now. https://news.ycombinator.com/item?id=42831769 https://qwenlm.github.io/blog/qwen2.5-1m/
44.
▲
by
ozgune
2y ago
Yes, it is. :) Our CTO, Daniel, described why we chose Ruby for our control plane in a blog post. https://www.ubicloud.com/blog/building-infrastructure-contro... On the data plane side, we use different open source com
45.
▲
DocumentDB: Open-source MongoDB implementation based on PostgreSQL
(opensource.microsoft.com)
8 points
by
ozgune
2y ago
|
1 comments
46.
▲
by
ozgune
2y ago
I see it in the "2. Model Summary" section (for [2]). In the next section, I see links to Hugging Face to download the DeepSeek-R1 Distill Models (for [3]). https://github.com/deepseek-ai/DeepSeek-R1?tab=readm
47.
▲
by
ozgune
2y ago
Yes, o1 hid its input. Still, it also provided a summary of its reasoning steps. In the email case, o1 thought for six seconds, summarized its thinking as "summarizing the email", and then provided the answer. We saw this in other
48.
▲
by
ozgune
2y ago
The R1 GitHub repo is way more exciting than I had thought. They aren't only open sourcing R1 as an advanced reasoning model. They are also introducing a pipeline to "teach" existing models how to reason and align with human
49.
▲
by
ozgune
2y ago
> However, DeepSeek-R1-Zero encounters challenges such as endless repetition, poor readability, and language mixing. To address these issues and further enhance reasoning performance, we introduce DeepSeek-R1, which incorporates cold-sta
50.
▲
by
ozgune
2y ago
> Apple even says it will publish its software images (though unfortunately not the source code) so that security researchers can check them over for bugs. I think Apple recently changed their stance on this. Now, they say that "sou
51.
▲
by
ozgune
2y ago
Hey, author here. I took extensive notes when playing around with o1 and QwQ-32B. When reading my notes later, I realized that I used the pronoun "they" to refer to a reasoning model. Somehow, it just didn't feel right to ref
52.
▲
Show HN: QwQ-32B APIs – o1 like reasoning at 1% the cost
17 points
by
ozgune
2y ago
|
3 comments
53.
▲
by
ozgune
2y ago
I can't help but upvote this article just because of the bacteria's name.
54.
▲
by
ozgune
2y ago
+1 on your comment. I think having a description of Apple's threat model would help. I was thinking that open source would help with their verifiable privacy promise. Then again, as you've said, if Apple controls the root of trust
55.
▲
by
ozgune
2y ago
I took a guided tour of the London Zoo this summer. The zoologist said that they had to procure vegetables from special sources and that they couldn't give fruits to the animals at the zoo. When asked why, she said, "When we give
56.
▲
Microsoft Azure flaunts first custom Nvidia Blackwell racks
(tomshardware.com)
2 points
by
ozgune
2y ago
|
2 comments
57.
▲
by
ozgune
2y ago
I think the parent project, Kamal, positions itself as a simpler alternative to K8s when deploying web apps. They have a question on this on their website: https://kamal-deploy.org "Why not just run Capistrano, Kubernetes o
58.
▲
by
ozgune
2y ago
Hey there, this is a comprehensive and informative reply! I had two questions just to learn more. * What has been your experience with using local NVMes with K8s? It feels like K8s has some assumptions around volume persistence, so I'm
59.
▲
by
ozgune
2y ago
(Ozgun from Ubicloud) Thank you for the kind words! Daniel has a few gems in this blog post and we tried to italicize some of them. My favorite one is around "There is no code without a theory of testing." "If you have 100% b
60.
▲
by
ozgune
2y ago
I read this book called "How Big Things Get Done." I've seen my fair share of projects going haywire and I wanted to understand if we could do better. The book identifies uniqueness bias as an important reason for why most bi
More ›