4 ms·
also no updates since then. does any library follow this reference implementation? llama.cpp surely is more optimized?
by singularity2001 4y ago
also no updates since then.
does any library follow this reference implementation?
llama.cpp surely is more optimized?
- JJJollyjim 4y agollama.cpp runs on the CPU, not the ANE or GPU.