3 ms·
This is an implementation of a transformer and in README it's presented as text->text. Tokens are just integers going in and out. Is it possible to use it to t
by milansuk 2y ago
This is an implementation of a transformer and in README it's presented as text->text. Tokens are just integers going in and out.
Is it possible to use it to train other types of LLMs(text->image, image->text, speech->text, etc.)?
- bootsmann 2y agoThe transformer itself just takes arrays of numbers and turns them into arrays of numbers. What you are interested in is the process that happens before and after the transformer.
- _giorgio_ 2y agoYes, anything can be an input token. Patch of pixels ---> token Fragment of input Audio ---> token etc