4 ms·
I used “all” instead of “most”, which is indeed inaccurate. A better response would have been “a collection of humanity’s creativity”. However, you fail to rec
by maxdoop 4y ago
I used “all” instead of “most”, which is indeed inaccurate. A better response would have been “a collection of humanity’s creativity”.
However, you fail to recognize that OpenAI also created Whisper, which is a quote capable speech-to-text transcriber; and this tool easily converts audio and video (those verbal bits you mentioned) into text.
So the pool of creativity which OpenAI can train its models on is far larger than just original text; further, they demoed image modality a couple weeks back which would allow for VISUAL creative works to be parsed as well.
- NhanH 4y agoI understand your points, but what I wanted to highlight was something slightly different: verbal communication and dialogs should still consists of magnitude more human conversation than text, or presentation, or basically anything digital.
- maxdoop 4y agoApologies, I’m maybe low on coffee this morning. Do you mean things like tone, facial expression, the general “energy” around a conversation, etc?
- NhanH 4y agoIt's simpler even: there are conversations that are straight up never happen in writings. So by training only on written text, you will never got those conversations. For example: teacher - student conversation where one side is confused and need to be explained, or many kind of debates and discussions where there is no end conclusion. Basically, pick a random human from the street and they would be talking and listening a lot more than they would be writing. Those talking and listening is never going to be captured by any digital system. It might end up that written text has enough similarity with verbal communication that it doesn't matter anyway. But that's hard to guess as a priori that it would be the case.