Devadex

Qwen3.5-TurboQuant-MLX-LM

github   Free   by CalaoGreskk
4d old3★C++

Accelerate Qwen3.5 KVCache with TurboQuant MLX-LM tools for prompt caching, quantized attention, and research-ready evals

Get it → calaogreskk.github.io
Related:
aiai-agentai-agentsai-codingai-governanceai-modelai-safetyai-securityai-toolsairflow

Found on Devadex — the discovery index for independent software the big search engines bury. More from github.

Report this listing