4 ms·
No, they didn't start with dictionaries, or any other parallel corpora; they learned the word by word translations as well from monolingual corpora, by finding
by apendleton 8y ago
No, they didn't start with dictionaries, or any other parallel corpora; they learned the word by word translations as well from monolingual corpora, by finding alignments between monolingual word embeddings in the target languages.
- deleted 8y ago[deleted]
- pc2g4d 8y agoNext step: unsupervised word segmentation. That way they could maybe apply this unsupervised translation system to undeciphered texts, e.g. Linear A, Rongorong, etc. I doubt it will work since most of the undeciphered scripts have a very small corpus, but maybe worth a try.