2 ms·
Scaleway have two separate (one fully managed one a bit less) services for that: https://www.scaleway.com/en/generative-apis/ https://www.scaleway.com/en/gener
by sofixa 3mo ago
Scaleway have two separate (one fully managed one a bit less) services for that:
https://www.scaleway.com/en/generative-apis/ https://www.scaleway.com/en/generative-apis/
https://www.scaleway.com/en/inference/ https://www.scaleway.com/en/inference/
- jonas_scholz 3mo agoyea, but no prompt caching right? This makes it unusable for my usecase at least, the cost would be insane
- PaoloBarbolini 3mo agoThey confirmed on LinkedIn that they are working on it. Also, they have a feature request that has been getting many votes recently: https://feature-request.scaleway.com/posts/1251/prompt-caching-to-reduce-input-tokens-cost https://feature-request.scaleway.com/posts/1251/prompt-cachi...