Qwen3.5-TurboQuant-MLX-LM
4d old3★C++
Accelerate Qwen3.5 KVCache with TurboQuant MLX-LM tools for prompt caching, quantized attention, and research-ready evals
Get it → calaogreskk.github.ioAccelerate Qwen3.5 KVCache with TurboQuant MLX-LM tools for prompt caching, quantized attention, and research-ready evals
Get it → calaogreskk.github.io