Ling-3.0-tiny-oQ6e

This is an oQ6e MLX conversion produced with oMLX.

Source and credit

Please follow the original model's license, acceptable-use policy, and attribution requirements.

Conversion

  • Quantization: oQ6e (enhanced=True)
  • Importance calibration: imatrix, 128 samples 脳 512 tokens
  • oMLX version: 0.6.0.dev1

oQe uses an imatrix sensitivity calibration to assign precision selectively, rather than applying a uniform bit-width to every tensor. The generated oq_imatrix_report.json records the resulting quantization allocation.

python run_oqe_conversion.py Ling-3.0-tiny-bf16 --models-dir /path/to/models --oq-level 6.0 --imatrix-samples 128 --imatrix-seq-length 512 --source-repo https://huggingface.co/inclusionAI/Ling-3.0-tiny

Files

  • model*.safetensors: quantized MLX weights
  • model.safetensors.index.json: shard index
  • config.json: model configuration
  • oq_imatrix_report.json: imatrix calibration and quantization report
Downloads last month
292
Safetensors
Model size
2B params
Tensor type
BF16
U32
F32
MLX
Hardware compatibility
Log In to add your hardware

6-bit

Inference Providers NEW
This model isn't deployed by any Inference Provider. 馃檵 Ask for provider support

Model tree for djrsystemservices/Ling-3.0-tiny-oQ6e

Quantized
(12)
this model