This model is an MLX conversion quantized to q8 for the prompt enhancer included with Sulphur-2-base. It's tested in LM Studio, and roughly doubles the tokenization rate vs. running the original GGUF model on Apple M-series Macs.

To use, download the files and place them in your LM Studio models directory using the "owner/model" path structure other models use.

Or use HF Hub and just symlink it

hf download MLXBits/sulphur-promptenhancer-mlx-q8
mkdir -p $HOME/.lmstudio/models/sulphur/prompt-enhancer-mlx-q8
ln -s $HOME/.cache/huggingface/hub/models--MLXBits--sulphur-prompt-enhancers-mlx-q8/snapshots/<snapshot hash> $HOME/.lmstudio/models/sulphur/prompt-enhancer-mlx-q8
Downloads last month
174
Safetensors
Model size
3B params
Tensor type
BF16
·
U32
·
F32
·
MLX
Hardware compatibility
Log In to add your hardware

8-bit

Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support

Model tree for MLXBits/sulphur-promptenhancer-mlx-q8

Quantized
(15)
this model

Collection including MLXBits/sulphur-promptenhancer-mlx-q8