YAML Metadata Warning:empty or missing yaml metadata in repo card
Check out the documentation for more information.
Qwen3.5-4B-MLX-E3-3bit-g128
Mixed-precision MLX checkpoint for on-device VLM inference on Apple Silicon.
- Base model: Qwen3.5-4B
- Vision tower: BF16 (unquantized)
- LM backbone: 3-bit, group-size 128
- Framework: MLX (mlx-vlm 0.6.3)
- Conversion source: mlx-community/Qwen3.5-4B-MLX-bf16
- Mac inference peak: 4.499 GB
- Disk size: 2.2 GB
- Semantic validation: Pass on 3 benchmark images (simple scene, text-heavy dashboard, complex lab)
- Part of: Z Lab / UCSD on-device VLM deployment project
- Downloads last month
- 9
Inference Providers NEW
This model isn't deployed by any Inference Provider. 馃檵 Ask for provider support