YAML Metadata Warning:empty or missing yaml metadata in repo card

Check out the documentation for more information.

Qwen3.5-4B-MLX-E3-3bit-g128

Mixed-precision MLX checkpoint for on-device VLM inference on Apple Silicon.

  • Base model: Qwen3.5-4B
  • Vision tower: BF16 (unquantized)
  • LM backbone: 3-bit, group-size 128
  • Framework: MLX (mlx-vlm 0.6.3)
  • Conversion source: mlx-community/Qwen3.5-4B-MLX-bf16
  • Mac inference peak: 4.499 GB
  • Disk size: 2.2 GB
  • Semantic validation: Pass on 3 benchmark images (simple scene, text-heavy dashboard, complex lab)
  • Part of: Z Lab / UCSD on-device VLM deployment project
Downloads last month
9
Safetensors
Model size
0.8B params
Tensor type
BF16
U32
F32
Inference Providers NEW
This model isn't deployed by any Inference Provider. 馃檵 Ask for provider support