Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
lukax
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
7 ms
·
61.
▲
by
lukax
1y ago
It's really accurate and supports 60+ languages
62.
▲
by
lukax
1y ago
Have you tried Soniox? It's really not expensive ($0.12/h, $200 free credits when you sign up) and really accurate. https://soniox.com/ You can use it with Spokenly (free app, bring your own Soniox API key) on mac
63.
▲
by
lukax
1y ago
Maybe you have an ad-blocker that just hides the popup but does not restore scrolling (scrolling is usually prevented when popups are visible)
64.
▲
Tell HN: Opt-out of LinkedIn training content creation AI models
13 points
by
lukax
1y ago
|
7 comments
65.
▲
by
lukax
1y ago
Soniox offers real-time speech-to-text with real-time translation between 60+ languages (mostly to/from English with some additional pairs between more popular languages). It operates with minimal amount of context and produces the tra
66.
▲
by
lukax
1y ago
Is this Triton's reply to NVIDIA's tilus[1]. Tilus is suposed to be lower level (e.g. you have control over registers). NVIDIA really does not want the CUDA ecosystem to move to Triton as Triton also supports AMD and other acceler
67.
▲
by
lukax
1y ago
word
68.
▲
by
lukax
1y ago
Have you tried Soniox for speech recognition? It supports Croatian. Or are you just looking for self-hosted open-source models? Soniox is very cheap ($0.1/h for async, $0.12/h for real-time) and you get $200 free credits on signup
69.
▲
Real-time speech to text translation
(soniox.com)
2 points
by
lukax
1y ago
|
0 comments
70.
▲
by
lukax
1y ago
Have you seen Soniox? They support real-time translation (only speech to text translation for now). https://soniox.com/ (disclaimer: I worked there)
71.
▲
by
lukax
1y ago
NVIDIA does not want CUDA development (e.g. flash attention) to move to Triton because Triton also supports AMD and if ecosystem moves from pure CUDA to Triton, that's bad for NVIDIA's lock-in. That's why there is so much foc
72.
▲
by
lukax
1y ago
Japanese, not Chinese
73.
▲
by
lukax
1y ago
Inference in Python uses harmony [1] (for request and response format) which is written in Rust with Python bindings. Another OpenAI's Rust library is tiktoken [2], used for all tokenization and detokenization. OpenAI Codex [3] is also
74.
▲
by
lukax
1y ago
And 440MB tab memory usage
75.
▲
by
lukax
1y ago
Soniox also supports real-time speech-to-text translation with 60 languages. You can hook that to a TTS and you have Speech-to-Speech translation. That failed Google I/O real-time translation demo? With Soniox it just works. You can tr
76.
▲
Real-time speech translation in 60 languages
(soniox.com)
4 points
by
lukax
1y ago
|
0 comments
77.
▲
by
lukax
1y ago
Try Soniox for real-time translation (interpreting). With the limited context it has in real-time, it's actually really good. https://soniox.com Disclaimer: I work for Soniox.
78.
▲
by
lukax
1y ago
At Koofr[1] one of the most requested features was an option to prevent downloading files from public links. We didn't want to lie to our users so we added a "Hide download button" option because that's the only thing yo
79.
▲
by
lukax
1y ago
Not really. But I didn't use async (was not supported yet when I started using it). Bindings were easy, everything else (building, linking,...) was a bit pain to setup because there no good examples.
80.
▲
by
lukax
1y ago
Do not write the bindings manually. Just use the amazing uniffi-rs library from Mozilla. https://github.com/mozilla/uniffi-rs You can generate bindings for multiple languages. It supports error handling on both sides a
81.
▲
Real-time Speech-to-Text in 60 languages
(soniox.com)
1 points
by
lukax
1y ago
|
0 comments
82.
▲
by
lukax
2y ago
So people could just use a non-US VPN and a non-US Apple account to download the apps?
83.
▲
by
lukax
2y ago
Uv also bundles uvx command so you can run Python scripts without installing them manually: uvx --from 'huggingface_hub[cli]' huggingface-cli
84.
▲
by
lukax
2y ago
It's a bit funny that they use Jedi Language Server for Python as they cannot use Microsoft's own Pylance ("Pylance is licensed for use in Microsoft products and services only, so can only be used on official Microsoft builds
85.
▲
by
lukax
2y ago
Go errors in standard library does not support stack traces. errors.Wrap() only exists in github.com/pkg/errors package
86.
▲
by
lukax
2y ago
The author uses Whisper and GPT-4o to get transcriptions into a nicely formatted Markdown file. We just released Omnio, a new AI model, that can do all this in a single step, as it works with audio directly. It does not generate a transcrip
87.
▲
by
lukax
2y ago
Whoops. We've updated the website. Omnio is available to all developers and all accounts receive $5.00 in free credits.
88.
▲
by
lukax
2y ago
We built Omnio to address the limitations we kept running into with existing audio AI models (we previously built an automatic speech recognition product). Most of them rely heavily on speech-to-text, which strips out a lot of the things li
89.
▲
Omnio: First AI model that can natively reason over audio
(soniox.com)
13 points
by
lukax
2y ago
|
8 comments
90.
▲
by
lukax
2y ago
Google Docs is also migrating to Kotlin Multiplatform > The initial step in this journey is the rollout of the Google Docs app for Android, iOS, and Web, which leverages KMP for shared business logic, validating its readiness for product
More ›