3 ms·
Check it out, you might be able to speed it up using this https://github.com/Anemll/anemll-flash-mlx https://github.com/Anemll/anemll-flash-mlx https://x.com/an
by anemll 6mo ago
Check it out,
you might be able to speed it up using this
https://github.com/Anemll/anemll-flash-mlx https://github.com/Anemll/anemll-flash-mlx
https://x.com/anemll/status/2038684375425200360 https://x.com/anemll/status/2038684375425200360
- aegis_camera 6mo agoThanks, pure Swift was the design idea and since I found nothing could be used for my project https://www.sharpai.org https://www.sharpai.org then I created Swift version. Python is too heavy to be delivered with application, user mentioned they want to use MLX, that's why I've been working on it for 1-2 weeks for bug fixing and testing , then suddenly TurboQuant proposed, I had a quick integration. My 64GB M5 Pro is already good for my local security task, now it's able to use M1/M2 Mini w/ 8GB memory.