6 ms·
That's a good question! We tried using Anthropic 100k before (Claude 1.3 was a lot worse), and I think that it's really important to figure out how to be contex
by williamzeng0 3y ago
That's a good question! We tried using Anthropic 100k before (Claude 1.3 was a lot worse), and I think that it's really important to figure out how to be context efficient, at least for GPT4.
My stance is with models ignoring long contexts(https://arxiv.org/pdf/2307.03172.pdf https://arxiv.org/pdf/2307.03172.pdf), we'll have this problem for a long time. I could be wrong though.
Also we did try function calling, but it doesn't allow for a chain of thought step. This made the plan/code way worse. Cool to see you found the same!