ThinkingCap-Qwen3.8-27B oQe MTP
Collection
8 items • Updated
How to use scottlowry/ThinkingCap-Qwen3.8-27B-oQ4e-fp16-mtp with MLX:
# Download the model from the Hub pip install huggingface_hub[hf_xet] hf download scottlowry/ThinkingCap-Qwen3.8-27B-oQ4e-fp16-mtp --local-dir ThinkingCap-Qwen3.8-27B-oQ4e-fp16-mtp
This model was quantized using oQ (oMLX v0.7.0) mixed-precision quantization.
4-bit