Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
zaptrem
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
9 ms
·
91.
▲
by
zaptrem
2y ago
This just sounds like gatekeeping.
92.
▲
by
zaptrem
2y ago
We aren't trying to force musicians to stop doing what they love, we're trying to give everyone else a taste of the fun of making music.
93.
▲
by
zaptrem
2y ago
Agreed! Those will be much more fun and we plan to support that. However, right now we're focused on making the base model slightly better, then we can easily add all of those controls (a-la ControlNets with Stable Diffusion).
94.
▲
by
zaptrem
2y ago
I feel like I've proven I know what I'm talking about at least from an ML perspective given we trained the model on the website. It's not strictly required, but I know you'll get much higher quality more powerful tools m
95.
▲
by
zaptrem
2y ago
Cointerintuitively we need to build the base model before we can do the advanced production stuff that’s most useful to existing pro musicians.
96.
▲
by
zaptrem
2y ago
If our sole goal was to get rich we would have pivoted to some b2bsaas thing as many suggested to us. What we’ve actually seen is so much new creativity from people who otherwise would never have made music.
97.
▲
by
zaptrem
2y ago
Just letting you know the about page has black text on a black background
98.
▲
by
zaptrem
2y ago
The creator sees/hears it! (and if they don't it really shouldn't have been generated lol, waste of compute)
99.
▲
by
zaptrem
2y ago
You can create a looped track by combining the song generation and transition generation examples from our API example repo!
100.
▲
by
zaptrem
2y ago
We're super interested in working on this (and melody conditioning) and even have some of the code written to generate the training data, but we want our base model to get a bit better before this becomes our main focus. Check back in
101.
▲
by
zaptrem
2y ago
it's because the thing we're launching today is an API for developers to use. If you want instrumental type stuff you should check out my bossa nova channel: https://sonauto.ai/radio
102.
▲
by
zaptrem
2y ago
> I have made plenty of money busking on the street That's why I specified mass market. However, given a choice between literally being on the street and working with a record label I'd probably choose the label, though I don&#
103.
▲
by
zaptrem
2y ago
The reason anything makes anyone happy is completely subjective, as evidenced by the many people who have told us our app made them and/or their friends and family happy.
104.
▲
by
zaptrem
2y ago
Over the years I've seen people get a lot of hate for things they've poured their souls into who turn around and post snarky/insulting responses that ended up getting them into even more hot water. I always wondered why they
105.
▲
by
zaptrem
2y ago
Thanks! We used SkyPilot (an open source cloud GPU worker management tool) to help out with both our small (single node) and large (many node) training runs.
106.
▲
by
zaptrem
2y ago
There was a single unhealthy worker that didn't get caught, we just killed it.
107.
▲
by
zaptrem
2y ago
I'm not sure to what extent AI music is copyrightable (I think it depends on a case-by-case amount of human influence) but our TOS assigns any rights we may have to the user.
108.
▲
by
zaptrem
2y ago
There are! Audio models are actually quite similar to image models, but there are a few key differences. First, is the autoencoder needs to be designed much more carefully as human hearing is insanely good and music requires orders of magni
109.
▲
by
zaptrem
2y ago
I think this is a good start, X high speed queries per hour then unlimited low-priority ones after. Do you know of any specific companies that do this we could take a look at?
110.
▲
by
zaptrem
2y ago
A) Agreed! B) So I guess the argument here is that this doesn't apply to AI music. I think that if someone really pours their soul into the lyrics of a song and regenerates/experiments with prompts until it's just right, and
111.
▲
by
zaptrem
2y ago
I don't see how the amount of work that went into it changes the core fact that all art is influenced by that which came before, and we don't call that stealing (unless you truly believe that "all art is theft"). My poin
112.
▲
by
zaptrem
2y ago
E.g., in the case of a future "LibreMusic" open source UI or an integration into their DAW they work with on the weekends. I'd get pretty annoyed if I had to keep putting a coin in the machine to adjust Logic Pro effects.
113.
▲
by
zaptrem
2y ago
For the consumer stuff: It's fun, and IMO that's enough. Not every song has to be peak artistic quality pushing the world forward, sometimes it's enough to bring a smile to a friend's face by making a song about them. If
114.
▲
by
zaptrem
2y ago
In my opinion training on all music is no more theft than Taylor Swift listening to the radio growing up (as long as we don't regurgitate existing songs which would be bad and useless anyway). I think an alternative legal interpretatio
115.
▲
by
zaptrem
2y ago
Suno's RVQ-token-based language model is tuned give you an acceptable song that most of their userbase would prefer every single time, but isn't very diverse. Our diffusion model is much more diverse and has higher vocal audio qua
116.
▲
by
zaptrem
2y ago
One thing I've been thinking about is how to do a better hobbyist plan system. It would be cool to do a flat rate unlimited plan, but we wouldn't want that to then be abused by larger customers/companies. Are there existing A
117.
▲
Show HN: Sonauto API – Generative music for developers
(sonauto.ai)
127 points
by
zaptrem
2y ago
|
157 comments
118.
▲
by
zaptrem
2y ago
You're thinking of "Orion" not "Omni" (GPT 4o stands for "Omni" since it's natively multimodal with image and audio input/output tokens)
119.
▲
by
zaptrem
2y ago
This seems like it should be attributed to better post training, not a bigger model.
120.
▲
by
zaptrem
2y ago
GPT 4.5 pricing is insane: Price Input: $75.00 / 1M tokens Cached input: $37.50 / 1M tokens Output: $150.00 / 1M tokens GPT 4o pricing for comparison: Price Input: $2.50 / 1M tokens Cached input: $1.25 / 1M tokens O
More ›