Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
seth_
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
7 ms
·
1.
▲
by
seth_
2y ago
Riffusion - Generative AI for Music | Research Scientist, Research Engineer | San Francisco | Full-time Riffusion is a small team training foundation models for music generation and building products that create more musicians in the world.
2.
▲
by
seth_
2y ago
Riffusion - Generative AI for Music | Research Scientist, Research Engineer | San Francisco | Full-time Riffusion is a small team training foundation models for music generation and building products that create more musicians in the world.
3.
▲
by
seth_
3y ago
love the deep dive here
4.
▲
by
seth_
3y ago
Those are awesome! Sorry if the app is down for you or anyone else right now btw. Our servers are a little overloaded but we'll be back up in no time
5.
▲
by
seth_
3y ago
We would have an iOS app if we were better mobile engineers... hopefully someday we'll make one!
6.
▲
by
seth_
3y ago
Very cool. The latent space is a wild place.
7.
▲
by
seth_
3y ago
Nice! It is best at pronouncing in English, but we've had a bunch of fun trying to get other languages too. Sometimes you can make things happen phonetically. Even for english words that it doesn't get right the first time haha
8.
▲
by
seth_
3y ago
We're happy to be building a toy! Your comment reminds me of this post: https://cdixon.org/2010/01/03/the-next-big-thing-will-start-... It's still really early innings for this technology, so we
9.
▲
by
seth_
3y ago
This one! Was a wild day for us :) https://news.ycombinator.com/item?id=33999162
10.
▲
by
seth_
3y ago
Would be neat to see the distribution of time that people stare at this... I imagine some will stare for hours
11.
▲
by
seth_
4y ago
Author here: fwiw we are running the app on a10g GPUs, which generally can turn around a 512x512 in 3.5s with 50 inference steps. This time includes converting the image into audio which should be done on the GPU as well for real-time purpo
12.
▲
by
seth_
4y ago
Author here: We were blown away too. This project started with a question in our minds about whether it was even possible for the stable diffusion model architecture to output something with the level of fidelity needed for the resulting au
13.
▲
by
seth_
4y ago
Author here: It can certainly be applied to voice, but the model would need deeper training to speak intelligibly. If you want to hear more singing, you can try a prompt like "female voice", and increase the denoising parameter in
14.
▲
by
seth_
4y ago
Author here: Indeed we are using Griffin-Lim. Would be exciting to swap it out with something faster and better though. In the real-time app we are running the conversion from spectrogram to audio on the GPU as well because it is a nontrivi
15.
▲
by
seth_
4y ago
Authors here: Fun to wake up to this surprise! We are rushing to add GPUs so you can all experience the app in real-time. Will update asap
16.
▲
by
seth_
4y ago
Looks like a much tastier version of what the times did here https://www.nytimes.com/2022/11/04/dining/ai-thanksgiving-me... , they should write a followup
17.
▲
by
seth_
4y ago
Slick app. Thanks for posting
18.
▲
by
seth_
5y ago
Hardline | iOS + Android Engineers | Remote friendly | San Francisco based company | Full-time | https://hardline.io We’re building a new way for deskless workers to communicate on the move—think Slack for construction. At the c
19.
▲
by
seth_
6y ago
Hardline | Full time | SF, Remote US Based | Lead Backend Engineer | https://www.hardline.io Help us build the fastest way for teams to communicate on the move. Our mobile app and simple hardware attachment turn any smartphone i
20.
▲
by
seth_
9y ago
Starling | Front End Engineer & Full Stack Engineer & Security Engineer | SF | Full-Time | Onsite Starling exists to make organizations better. We're an analytics platform for People data, helping companies create data-backed s