3 ms·
Honestly, it's hard! We try our best through a few strategies: - At this time Marvin is only tested with GPT-3.5 and GPT-4 to reduce the surface area of LLM di
by jlowin 4y ago
Honestly, it's hard! We try our best through a few strategies:
- At this time Marvin is only tested with GPT-3.5 and GPT-4 to reduce the surface area of LLM differences. You're welcome to try others, but the prompts are optimized for performance of those two models
- I expect that prompts optimized for one family of models will not automatically work with others, so as time allows we may end up with branching prompts based on your model choice. Marvin does a lot of work to manipulate the user prompt before sending it to the LLM so I think there's an excellent chance that we could deliver similar results for the same user input.
- Even with GPT-3.5 and GPT-4 we see a remarkable difference between them, with 4 needing far less complex prompts than 3.5 to get the same result. However, given both the cost and availability of GPT-4, we decided to make 3.5 our default to make sure everyone can use the library. Therefore we do our best to make sure prompts work with 3.5 and expect that that is a good bet of compatibility with 4.
- akiselev 4y agoCan you offer the ability to commit the generated code? Only generate code if the function body is empty and write it back out to the Python source file which would allow users to check in functions and quickly regenerate them just by rerunning the executable.
- jlowin 4y agoWe have a related issue open (https://github.com/PrefectHQ/marvin/issues/64 https://github.com/PrefectHQ/marvin/issues/64) but haven't designed anything yet.