4 ms·
I don't buy this argument. Specifically, Modalities. The claim is that "[b]y using modality-specific tokenizers or processing raw data streams, frontier models
by pointlessone 3y ago
I don't buy this argument. Specifically, Modalities. The claim is that "[b]y using modality-specific tokenizers or processing raw data streams, frontier models can, in principle, handle any known sensory or motor modality". Sure, you may get some sort of output but I'm quite confident you won't get anything useful without training the model on that kind of input. We barely started sticking narrow AI together. E.g. ChatGPT (LLM) can invoke Dall-e (text to image). Maybe AGI will turn out to be a bunch of narrow AIs in a trench coat. So far I haven't seen such a system.
Tasks claim is a little stretched, too. For example, simple arithmetic. Does model actually do arithmetic or does it do text generation and the right answer just happens to be the most likely next token? Can we reliably tell one from the other to claim that model actually perform tasks other than just text generation?