3 ms·
Yeah, 100% agree. That's one thing I just thought about yesterday also - Every session should be summarised and written to the data store also, that way session
by chsitter 7mo ago
Yeah, 100% agree. That's one thing I just thought about yesterday also - Every session should be summarised and written to the data store also, that way sessions become portable contexts. There's potentially a case to be made to have full replay, i.e. the plaintext sessions stored but I'm not entirely sure on how much more valuable that is over a summary.
Do you have thoughts, or a take on that?
- jlongo78 7mo ago[flagged]
- chsitter 7mo agoRight - that makes sense. I see it similarly. My assumption though is that as models get better, the likelyhood of the model missing context that matters the most will get lower and lower. The hybrid approach though is nice, I'll have a think about that and see if that's something I can incoporate into it. Thanks for the feedback, very much appreciated
- jlongo78 7mo ago[flagged]
- chsitter 7mo ago100% - I'll implement and add that now :)
- jlongo78 7mo ago[flagged]
- chsitter 7mo agoI'm not gonna lie - at the moment it's pretty basic as a chunked semantic store where relevant chunks are retrieved in conjunction with some criteria the Agent can pass to the MCP server. Context usage is definitely a problem that's on my mind also and semantic anchors are one area I'm exploring but don't have a clear architecture for it jotted down yet. The real problem I'm facing right now is how to inject this into say claude or chatgpt and have those agents default use it as a memory layer
- jlongo78 7mo ago[flagged]
- chsitter 7mo agoThat is what I do at the moment - I gotta update the instructions on the website to prod users to set things up this way. The issue I perceive though is that adding an MCP server alone is not enough to modify the system prompt of the AI Agent. I tried to have the mcp server description be an injected prompt to add these instructions to the system prompt but that doesn't seem to work, I tried adding sampling to the MCP server which supposedly should be able to plug into messages without luck, tried to optimise for chatGPT with an OpenAPI spec etc. The only way I found that I can get those clients to use my memory layer is by doing what you describe - which is not necessarily the most user friendly/one-click setup I desire