4 ms·
You can use llama.cpp server's tokenize endpoint to tokenize and count the tokens: https://github.com/ggerganov/llama.cpp/blob/master/examples/server/README.md
by xyc 2y ago
You can use llama.cpp server's tokenize endpoint to tokenize and count the tokens:
https://github.com/ggerganov/llama.cpp/blob/master/examples/server/README.md#post-tokenize-tokenize-a-given-text https://github.com/ggerganov/llama.cpp/blob/master/examples/...