Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
dylanbfox
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
10 ms
·
31.
▲
AssemblyAI (YC S17) Is Hiring Python Developers and Account Execs
(apply.workable.com)
1 points
by
dylanbfox
5y ago
32.
▲
Building CLIs in Python with Click
(assemblyai.com)
1 points
by
dylanbfox
5y ago
|
0 comments
33.
▲
Fine-Tuning Transformers for NLP
(assemblyai.com)
67 points
by
dylanbfox
5y ago
|
5 comments
34.
▲
by
dylanbfox
5y ago
Love this idea. Had this on my list of side projects to build for a while - definitely see the use for this. It's one less thing for a dev team to maintain in-house. What does latency look like on delivering webhooks? From the time you
35.
▲
A Review of End-to-End Architectures for Speech Recognition
(assemblyai.com)
6 points
by
dylanbfox
6y ago
|
0 comments
36.
▲
by
dylanbfox
6y ago
That's right - most literature does show that encoder-decoder architectures outperform CTC. I think one of the main reasons for this is that CTC assumes the label outputs are conditionally independent of each other, which is a pretty b
37.
▲
by
dylanbfox
6y ago
One of the main factors for this is probably due to dataset size. Commercial STT models are trained on 10s of thousands of hours of data from real-world data. Even a decent model architecture is going to perform pretty well on that much dat
38.
▲
AssemblyAI Is Hiring a Full Stack Engineer (Remote OK)
(angel.co)
1 points
by
dylanbfox
8y ago
39.
▲
AssemblyAI Is Hiring Deep Learning Engineers
1 points
by
dylanbfox
8y ago
40.
▲
by
dylanbfox
8y ago
Those test sets are definitely more real-world than Libri Clean in our experience, but still are not as "dirty" (noise, muffling, cross talk, recording quality, etc.) as we see in production.
41.
▲
by
dylanbfox
8y ago
Thank you! Please let me know if you need any help. My email is in my profile.
42.
▲
by
dylanbfox
8y ago
Thanks for sharing this! It's awesome to see another independent company tackling STT well. If you guys ever want to try to collaborate, let me know! My email is in my profile.
43.
▲
by
dylanbfox
8y ago
There are some good open source STT solutions out there like Mozilla's DeepSpeech ( https://github.com/mozilla/DeepSpeech ) that wouldn't be too hard to put behind an API. We also plan to open source more of ou
44.
▲
by
dylanbfox
8y ago
Dylan from AssemblyAI here. Right now we only support English, but are going to be launching more languages very soon. We are also getting closer to launching a transfer learning API, so you can train your own acoustic models. This _could_
45.
▲
by
dylanbfox
8y ago
Thank you for the feedback!
46.
▲
by
dylanbfox
8y ago
Right, we noticed similar findings. We do automatic punctuation now, and do diarization when there is more than one channel in the audio file. We're launching diarization on single channel audio with multiple speakers very soon. We
47.
▲
by
dylanbfox
8y ago
Hi there! Thanks for the support. We were very surprised to see ourselves on HN tonight! We've been working towards a big launch in a few weeks, which is why we don't have the benchmarks ready just yet. But we are comparable to mo
48.
▲
by
dylanbfox
8y ago
Dylan from AssemblyAI. We're working on this! We didn't plan to launch on HN tonight which is why the benchmarks aren't ready, but we know we definitely need these for the community. We've found most public benchmarks, e
49.
▲
by
dylanbfox
8y ago
This is Dylan from AssemblyAI. We're really surprised to see ourselves on HN tonight! We had a big launch planned for 4-6 weeks from now, and have been working towards getting things ready for that. As a result, we're missing a lo
50.
▲
by
dylanbfox
8y ago
Dylan from AssemblyAI here. We're surprised to see ourselves on HN tonight -- but if you do try out the API I would love to see what you think! Thanks for your interest. DeepSpeech is an awesome open source project, and we absolutely s
51.
▲
by
dylanbfox
8y ago
Dylan from AssemblyAI here. Thanks for trying the API! Time to implementation is one thing we’re focusing a lot on. We’re really trying to make it fast to get up and running with Assembly — for example not requiring you to specify any meta
52.
▲
by
dylanbfox
8y ago
This is very important. We do not store or copy any of the audio data you send to the API. We do not train on any audio data you send to the API. We’re a small company and absolutely believe in security and privacy. We were caught off guard
53.
▲
by
dylanbfox
8y ago
Dylan from AssemblyAI here. We didn’t plan to launch on Hacker News for another few weeks — and are surprised to see ourselves on here — but really appreciate the interest. That’s why we don’t have more thorough benchmarks and technical spe
54.
▲
by
dylanbfox
8y ago
Dylan from AssemblyAI here. Sorry for the mixup. This is just a placeholder page that we don’t link to anywhere from our site yet. I’m not sure how you found this page — but we took it down for now! We didn’t realize our site got posted on
55.
▲
by
dylanbfox
8y ago
This is great -- landing page and docs are super informative and makes me want to try this.
56.
▲
by
dylanbfox
8y ago
Just FYI -- getting a: "Your connection is not secure The owner of jitx.com has configured their website improperly. To protect your information from being stolen, Firefox has not connected to this website." When hitting https:&#
57.
▲
World Models in Tensorflow
(medium.com)
2 points
by
dylanbfox
8y ago
|
0 comments
58.
▲
by
dylanbfox
9y ago
Yup! This is in the works and should be released soon!
59.
▲
by
dylanbfox
9y ago
This is something we are looking into!
60.
▲
by
dylanbfox
9y ago
The way it works right now is that you would upload one of those old speeches or radio shows to the API, and then you'd get back a transcript that includes all the "terms of art" you customized the API to be able to recognize
More ›