3 ms·
I doublt it's fine tuning (actually changing the model weights). It's more like "im going to paste a text blob then in the following chats i will ask questions
by Racing0461 3y ago
I doublt it's fine tuning (actually changing the model weights). It's more like "im going to paste a text blob then in the following chats i will ask questions about it) type inner prompt.
- singularity2001 3y agoFor longer documents it uses vector embeddings
- Racing0461 3y agoHow's that different from pasting the text in the first chat and running the vector embedding step on the text on the server (maybe at least bypassing the chat text limit)? Does this fix the amnesia issue where the info from chats longer than the context length is forgotton because the document isn't baked directly into the weights like fine tuning?