Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
coder543
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
8 ms
·
91.
▲
by
coder543
9mo ago
Not that I've heard. I searched and I see nothing. Where has Intel said they are winding down chip fabrication?
92.
▲
by
coder543
9mo ago
All signs are that they are doing exactly that. They already have an on-device LLM which powers certain features, and I expect they will have a better-trained version of that on-device model that comes out with the "new Siri" up
93.
▲
by
coder543
9mo ago
iCloud Photos Downloader is an option, yes, but it is incorrect to say that Apple does not provide an official way to do this on Mac. Again, I direct you to the Apple Store so someone can show you in person, since you won't listen to a
94.
▲
by
coder543
9mo ago
You're arguing with a lot of people who have personally seen this work. You can listen to other people. You can also go to an Apple Store and let them show you what's going wrong here.
95.
▲
by
coder543
9mo ago
Your point has been clear the whole time. It is still not correct.
96.
▲
by
coder543
9mo ago
Please stop repeating your incorrect points that are contradicted by everyone else’s real experiences. Yes, new ones will be uploaded. That doesn’t mean old ones won’t also be downloaded.
97.
▲
by
coder543
9mo ago
No, you are not correct. How many people have to tell you this? It absolutely works the way I said it does, because I have seen it work that way. Just because you accidentally turned off iCloud Photos in your Apple Account settings on that
98.
▲
by
coder543
9mo ago
Yes, there is a button on Mac: https://support.apple.com/guide/photos/use-icloud-photos-pht... As long as you are signed into the Mac with the same iCloud account used on the iPhone, this will download them all.
99.
▲
by
coder543
9mo ago
Every app that uses tailwind builds a custom CSS bundle. Tailwind Labs does not host those; whoever is making the app has to figure out their own hosting. So I’m not seeing the significant infrastructure costs? Even if Tailwind were a share
100.
▲
by
coder543
9mo ago
To me, a closer analogy is In Context Learning. In the olden days of 2023, you didn’t just find instruct-tuned models sitting on every shelf. You could use a base model that has only undergone pretraining and can only generate text continua
101.
▲
by
coder543
9mo ago
Why wouldn’t that be one-shot voice cloning? The concept of calling it zero shot doesn’t really make sense to me.
102.
▲
by
coder543
9mo ago
> You’re talking to me like a total idiot, having assumed I know nothing about this. Sorry I tried to help? If that's the response I get for helping, good luck... > All I meant was a way to avoid storing images in git, the rest i
103.
▲
by
coder543
9mo ago
I wasn't suggesting publishing to Cloudflare, just that if you're concerned about the complexity of the workflow of getting images into the CDN, simply fronting whatever host you're using with a CDN of some kind (which coul
104.
▲
by
coder543
9mo ago
If you're okay with the images being on a CDN, why wouldn't you also be okay with the HTML and CSS also being on the CDN? Just fronting the entire static site with a pull-through CDN is an easy solution that doesn't require a
105.
▲
by
coder543
9mo ago
Is it their original launch edition keyboard, or the later refined version? The launch edition one I have is like you describe, but I hope they have improved things since then.
106.
▲
by
coder543
9mo ago
I meant large MoE models are more socially accepted now. They were not when Llama 4 launched, and I believe that worked against the Llama 4 models. The Llama 4 models are MoE models, in case you are unaware, since it feels like your comme
107.
▲
by
coder543
9mo ago
The Llama 4 models were instruct models at a time when everyone was hyped about and expecting reasoning models. As instruct models, I agree they seemed fine, and I think Meta mostly dropped the ball by taking the negative community feedback
108.
▲
by
coder543
9mo ago
Even though big, dense models aren't fashionable anymore, they are perfect for specdec, so it can be fun to see the speedup that is possible. I can get about 20 tokens per second on the DGX Spark using llama-3.3-70B with no loss in qua
109.
▲
by
coder543
9mo ago
It depends entirely on what you want to do, and how much you're willing to deal with a hardware setup that requires a lot of configuration. Buying several 3090s can be powerful. Buying one or two 5090s can be awesome, from what I'
110.
▲
by
coder543
9mo ago
They're probably referencing this article: https://blog.exolabs.net/nvidia-dgx-spark/
111.
▲
by
coder543
9mo ago
As you allude, the prompt processing speeds are a killer improvement of the Spark which even 2 Strix Halo boxes would not match. Prompt processing is literally 3x to 4x higher on GPT-OSS-120B once you are a little bit into your context wind
112.
▲
by
coder543
9mo ago
I remember another one that was popular years ago: https://news.ycombinator.com/item?id=17459204
113.
▲
by
coder543
9mo ago
Your logical fallacy is assuming two different groups of people are the same people, which never leads to productive conversation.
114.
▲
by
coder543
9mo ago
Yes, you can offload random experts to the GPU, but it will still be activating experts that are on the CPU, completely tanking performance. It won't suddenly make things fast. One of these GPUs is not enough for this model. You'
115.
▲
by
coder543
9mo ago
The Spark has more compute, so it should be faster for prefill (prompt processing). The M4 Max has double the memory bandwidth, so it should be faster for decode (token generation).
116.
▲
by
coder543
9mo ago
No… that’s not how this works. 96GB sounds impressive on paper, but this model is far, far larger than that. If you are running a REAP model (eliminating experts), then you are not running GLM-4.7 at that point — you’re running some other m
117.
▲
by
coder543
9mo ago
> I'd rather place that 10K on a RTX Pro 6000 if I was choosing between them. One RTX Pro 6000 is not going to be able to run GLM-4.7, so it's not really a choice if that is the goal.
118.
▲
by
coder543
9mo ago
$10k gets you a Mac Studio with 512GB of RAM, which definitely can run GLM-4.7 with normal, production-grade levels of quantization (in contrast to the extreme quantization that some people talk about). The point in this thread is that it w
119.
▲
by
coder543
10mo ago
A free and privacy-oriented hosted service that people have to pay to maintain? That is a confusing concept. How would the incentives be aligned?
120.
▲
by
coder543
10mo ago
If you want to be able to generate up to 128k tokens in one go successfully, then yes, that math checks out.
More ›