3 ms·
How does it compare to some of the newer mlx inference engines like optiq that support turboquantization - https://mlx-optiq.pages.dev/ https://mlx-optiq.pages.
by codelion 6mo ago
How does it compare to some of the newer mlx inference engines like optiq that support turboquantization - https://mlx-optiq.pages.dev/ https://mlx-optiq.pages.dev/