dicksondickson/gpt-oss-20b-oQ5e-bf16-MLX

This model is a quantization of:
https://huggingface.co/openai/gpt-oss-20b

This checkpoint was quantized using oMLX 0.7.0 with imatrix enabled.

Important tensors are left at bf16 which is for Apple M3 chips and later.

Run the model using oMLX:
https://github.com/jundot/omlx

Downloads last month
57
Safetensors
Model size
21B params
Tensor type
U32
·
BF16
·
MLX
Hardware compatibility
Log In to add your hardware

5-bit

Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support

Model tree for dicksondickson/gpt-oss-20b-oQ5e-bf16-MLX

Quantized
(250)
this model

Collection including dicksondickson/gpt-oss-20b-oQ5e-bf16-MLX