4 ms·
The paper's implementation is to essentially append memory text as part of the prompt. Why don't they use a storage/retrieval system that doesn't consume conte
by samanator 3y ago
The paper's implementation is to essentially append memory text as part of the prompt.
Why don't they use a storage/retrieval system that doesn't consume context window tokens? E.g. Storage could be to automatically categorize data with tags at insertion time (i.e. upon user prompt) and retrieval could be a query that filters using a tag guessed by the LLM (before responding to the user).
With a few initial rules like hard coded tag names/styles, my intuition is that this would produce great results.