4 ms·
Here is a queueing api server for self hosted inference backends: https://github.com/aime-team/aime-api-server https://github.com/aime-team/aime-api-server from
by mitjam 2y ago
Here is a queueing api server for self hosted inference backends: https://github.com/aime-team/aime-api-server https://github.com/aime-team/aime-api-server from a friend of mine. Very light weight and easy to use. You can even serve models from Jupyter Notebooks with it without needing to worry about overwhelming the server. It just gets slower the more load you send to it.
- _akhe 2y agoReally cool! I like that they have live demos to prove it out. Thanks for sharing