3 ms·
Self-hosted inference orchestrators compared: LocalAI, exo, GPUStack, vLLM
- nextime 12d ago[flagged]
- hypfer 12d agoThis feels agentically generated. The blog, the post here, the (auto?)killed LLM comment.
- nextime 11d agonot the whole blog ( but yes that post ), and yes it WAS autogenerated, it was supposed to be only a test but i did a gating wrong, apologise for it.
- SahAssar 12d agoSeems generated. Also why no llamafile?
- nextime 11d agosorry it was generated and was supposed to be a test and not to be posted. thanks for the llamafile suggestion!
- polotics 12d agowhat exactly did you mean when xou wrote this paragraph title: "LiteLLM — a router, not a runtime' ?
- nextime 11d agoI wouldn't wrote as it, sorry, the post was an unintended generated post, wasn't supposed to post for real, until i was going to rewrite it by hand at least. The meaning for that sentence is that litellm just route requests, it doesn't execute a model.
- nextime 11d agofor whoever is reading and rightly say "it seems llm generated", you are right, sorry, wasn't supposed to really post it and it did by error, my bad for not supervision my test while i was doing it.
- anotherCodder 11d ago[flagged]