5 ms·
I recently quit my job to build specialized tooling in this space. We’re broadly focusing on eval in general, but are starting with high quality question and an
by maxrmk 3y ago
I recently quit my job to build specialized tooling in this space. We’re broadly focusing on eval in general, but are starting with high quality question and answer generation for testing these kinds of RAG pipelines. It’s surprisingly hard!
- resiros 3y agoSounds very interesting. I am building an open-source LLM building platform (agenta.ai) and looking for eval approaches to integrate for our users. Do you have already a product/api that we could use?
- maxrmk 3y agoWe're in closed beta right now, but shoot me an email (max@talc.ai) and I can get you API access