Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
sid-the-kid
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
5 ms
·
31.
▲
by
sid-the-kid
1y ago
We are live again folks! Sorry about that. We ran out of storage space.
32.
▲
by
sid-the-kid
1y ago
The system just crashed. Sorry! Working on getting things live again as fast as we can!
33.
▲
by
sid-the-kid
1y ago
Nice! Thanks for sharing. I hadn't seen that paper before. Looks like they take in a real-world video and then re-generate the mouth to get to lip synch. In our solution, we take in an image and then generate the entire video. I am sur
34.
▲
by
sid-the-kid
1y ago
For the input, we pass the model: 1) embedded audio and 2) a single image (encoded with a causal VAE). The model outputs the final RGB video directly. The key technical unlock was getting the model to generate a video faster than real-time.
35.
▲
by
sid-the-kid
1y ago
Yes. We use Modal ( https://modal.com/ ), and are big fans of them. They are very ergonomic for development, and allow us to request GPU instances on demand. Currently, we are running our real-time model on A100s.
36.
▲
by
sid-the-kid
1y ago
We just removed email signup. You can try it out now without logging in. It was easier than expected to do technically, so we just shipped a quick update.
37.
▲
by
sid-the-kid
1y ago
That's fair. We just removed the sign-in for HN. Should be live shortly. Each person gets a dedicated GPU, so we were worried about costs before. But, let' s just go for it.
38.
▲
by
sid-the-kid
2y ago
I am with you. I want this too. Maybe somebody can make it wit their API?
39.
▲
by
sid-the-kid
2y ago
Okay. I know these guys IRL. BUT, I genuinely think they have the best music model out there. Hands down. The songs are just more unique, and have a wider range of musical variation. With Suno/Udio, the songs just sounds the same after