Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
brucethemoose2
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
5 ms
·
1.
▲
by
brucethemoose2
3y ago
Some phones have limiters to keep the battery at 60%-80%. I believe most can do this with the right software. It's not a fix, but it should extend the life considerably.
2.
▲
by
brucethemoose2
3y ago
Another benefit: modern smartphones have large GPUs, large media blocks, and fast RAM. With the right software , they can be a surprisingly powerful AI host or transcoding server.
3.
▲
by
brucethemoose2
3y ago
Real world GPU performance is hugely influenced by hand optimization of the CUDA kernels.
4.
▲
by
brucethemoose2
3y ago
Power/Weight is extremely high. A tiny wankel will do the job, and weight is everything on cars. It does prefer a narrow RPM band, which is fine. Reliability is the biggest concern TBH, but maybe that's not a huge bummer if its
5.
▲
by
brucethemoose2
3y ago
Yeah, its an unspoken but rampant thing in the llm community. Basically no one respects licenses for training data. I'd say the majority of instruct tunes, for instance, use OpenAI output (which is against their TOS). But its all just
6.
▲
by
brucethemoose2
3y ago
Yeah I know, hence its odd I found it kind of dumb for personal use. Moreso with the smaller models, which lost an objective benchmark I have to some Mistral finetunes. And I don't think I was using it wrong. I know, for instance, th
7.
▲
by
brucethemoose2
3y ago
I would note the actual leading models right now (IMO) are: - Miqu 70B (General Chat) - Deepseed 33B (Coding) - Yi 34B (for chat over 32K context) And of course, there are finetunes of all these. And there are some others in the 34B-70B ran
8.
▲
by
brucethemoose2
3y ago
The conspiracy theorist in me says thats a low priority due to perverse incentives (namely selling more storage at a huge markup). Another rationale is that the Apple ecosystems tends to not use JPEG by default anyway, right? It uses HEIC o
9.
▲
by
brucethemoose2
3y ago
Being a "hero" open source dev for a project like that can require a lot of neuroticism. Sometimes it works, but sometimes the project is just too big, I think.
10.
▲
by
brucethemoose2
3y ago
It's not either or, you can use different vendors for different tasks. tinygrad isn't in the realm of production ready though, AFAIK.
11.
▲
by
brucethemoose2
3y ago
The MI300 is the best accelerator you can buy, for many current workloads. It's technically way more advanced. Not as outrageously priced as an H100 either.
12.
▲
by
brucethemoose2
3y ago
I think you are preaching to the choir, and AMD is not listening. AMD would be selling 48GB 7900s or AI-only W7900s if they really wanted a consumer card ramp. They don't. Not because they can't (they literally prevent OEMs from d
13.
▲
by
brucethemoose2
3y ago
I never followed Hotz, so perhaps I missed something cool. But I never understood the hype myself.
14.
▲
by
brucethemoose2
3y ago
Well, personally, SDXL just blows 1.5 out of the water for me. I haven't had a reason to even touch 1.5 in months. But note that SDXL is really awful in automatic1111 or vanilla HF diffusers for me. You have to use something with prope
15.
▲
by
brucethemoose2
3y ago
Used 3090 prices are absolutely outrageous. And the 4090 MSRP was outrageous to begin with.
16.
▲
by
brucethemoose2
3y ago
I was talking about renting! There are some boutique hosts like Hot Aisle serving MI300s (who I really should reach out to), but for the immediate future our little startup is stuck with the big cloud providers. No MI300s for us mere mortal
17.
▲
by
brucethemoose2
3y ago
But is this going to blow over in a few days? Again? I can certainly appreciate frustration with the AMD stack, but be blunt, I was not impressed with Hotz's YouTube rant from before.[1] It didn't give the impression of a stable f
18.
▲
by
brucethemoose2
3y ago
SDXL is amazing. The community is entrechend in 1.5 because that's what everyone is now familiar with, IMO
19.
▲
by
brucethemoose2
3y ago
Its more like the store being a literal hedge maze, with an entrance fee, and once you get to the actual products, they are outrageously priced junk. And the store is price fixing with nearby stores. I think theres a difference between oppo
20.
▲
by
brucethemoose2
3y ago
My "oh no" moment was a vision model reading this perfectly : https://abadguide.files.wordpress.com/2012/01/jh66.jpg?w=640 Not an OCR program or anything specialized, just some generic (vision) llm that
21.
▲
by
brucethemoose2
3y ago
Tests are not out yet, but: - It's very large, yes. - It's a base model, so its not really practical to use without further finetuning. - Based on Grok-1 API performance (which itself is probably a finetune) its... not great at
22.
▲
by
brucethemoose2
3y ago
Going to leave this gem here: https://www.vttoth.com/CMS/physics-notes/311-hawking-radiati... Black holes are weird because they are essentially macroscopic particles with only one variable, mass (ignoring angul
23.
▲
by
brucethemoose2
3y ago
Groq's inference strategy appears to be "SRAM only." There is no external memory, like GGDR or HBM. Instead, large models are split between networked cards, and the inputs/outputs and pipelined. This is a great idea... I
24.
▲
by
brucethemoose2
3y ago
> Proper moderation I would point to oldschool forums (and HN!), where the communities were just large enough to moderate themselves and stop nasty off topic junk like that from appearing. One problem is modern social media has (mostly)
25.
▲
by
brucethemoose2
3y ago
Interestingly, Ollama is not popular at all in the "localllama" community (which also extends to related discords and repos). And I think thats because of capabilities... Ollama is somewhat restrictive compared to other frontends.
26.
▲
by
brucethemoose2
3y ago
I mean, they have so much potential... As an example, they have expertise designing huge silicon chips, big iron servers, and fast interconnects. They even made a ternery chip in the past, and plenty of fabrication research. They could abso
27.
▲
by
brucethemoose2
3y ago
As I understand it, the WSE-2's interconnect is actually quite good, and models are split across chips kinda like GPUs. And keep in mind that these nodes are hilariously "fat" compared to a GPU node (or even an 8x GPU node)
28.
▲
by
brucethemoose2
3y ago
And even mundane details are so difficult... thermal expansion, the sheer number of pins that have to line up. This thing is a marvel.
29.
▲
by
brucethemoose2
3y ago
Reposting the CS-2 teardown in case anyone missed it. The thermal and electrical engineering is absolutely nuts: https://vimeo.com/853557623 https://web.archive.org/web/20230812020202/https:/&
30.
▲
by
brucethemoose2
3y ago
Facebook very specifically bought and customized Intel SKUs tailored for AI workloads for some time.
More ›