Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
sawfwair
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
3 ms
·
1.
▲
by
sawfwair
2mo ago
That's a great q and point. I built it to get at a broad approach to tying multimodal capability together in one place as I'm personally building products on top of it that stitch chat, image, video together and wanted to enable o
2.
▲
by
sawfwair
2mo ago
Yes, missed that completely... Great point!
3.
▲
by
sawfwair
2mo ago
totally fair! i spent a while trying to get clever and compress what i wanted to say and finally just hit submit but prob lost too much - local inference runtime, one cli that runs image/video/music/speech/3d/etc on
4.
▲
by
sawfwair
2mo ago
some real numbers on my m4 max - image gen (zimage-nano), 1024x1024: 58s image -> textured mesh (trellis.2): 2m 49s sfx generate (5s clip): 3.6s music generate (8s, ace-step): 15s speech synth: 13s | transcribed back: 2.2s video gen w&#x
5.
▲
Show HN: Local text, image, video, music and 3D from one CLI, no Python
(github.com)
16 points
by
sawfwair
2mo ago
|
7 comments