4 ms·
Just added this to my benchmark site: https://multilingualsttbench.com/ https://multilingualsttbench.com/ It doesn't reach the frontier in either latency or ac
by mariano54 1mo ago
Just added this to my benchmark site:
https://multilingualsttbench.com/ https://multilingualsttbench.com/
It doesn't reach the frontier in either latency or accuracy for ai multilingual conversations.
- adamgoodapp 1mo agoThanks for this, really helpful. I would also like to see benchmark for translation. I'm looking for live translated subtitles so my Japanese wife can enjoy any show with out waiting months for official VOD streams to release them.
- Kokouane 1mo agoI'm confused, doesn't your leaderboard clearly show it is the most accurate model? It's number one in the leaderboard. Am I missing something?
- Kokouane 1mo agoFigured it out. 3.5 Flash and 3.5 Transcribe are different models
- alxndr13 1mo agomissing aqua voice's avalon 1.5 model there.