3 ms·
I think it depends on usage pattern. You trade speed for lower memory usage. Maybe engine specialisation and faster SSDs is the future for local inference, who
by gitpusher42 2mo ago
I think it depends on usage pattern. You trade speed for lower memory usage.
Maybe engine specialisation and faster SSDs is the future for local inference, who knows