3 ms·
Yeah, a random sample won't be (numerically) ideal, but given the size of the set, probably safe. A really quick and easy way to optimise this would be to just
by Mumps 5y ago
Yeah, a random sample won't be (numerically) ideal, but given the size of the set, probably safe.
A really quick and easy way to optimise this would be to just grab a well-trained word embedding space and pick _k_ furthest words by your favourite distance metric.