SenseNova U1.5 8B MoT - GGUF Quantizations
Optimized GGUF quantization formats for SenseNova-U1.5-8B-MoT, tailored for lower-VRAM setups (such as 8GB NVIDIA GPUs) via ComfyUI.
The tracking tags above link directly back to the main repository model page to aggregate download statistics.
Available Quants
Files are engineered to balance performance and memory footprints using high-precision F16 tensor swaps for sensitive layers (patch_embedding, dense_embedding, and fm_head) to preserve generation quality:
- Q8_0 (~18.6 GB) β Near-lossless quality.
- Q6_K (~14.5 GB) β High fidelity.
- Q5_K_M (~12.5 GB) β Balanced tier.
- Q4_K_M (~10.5 GB) β Standard compressed tier.
- Q3_K_M (~8.5 GB) β Recommended tight VRAM profile for 8GB cards.
- Q2_K (~6.4 GB) β Maximum compression.
Required Custom Node Fork
To run these GGUF files natively without structural AttributeError crashes or shape mismatches, use the dedicated node fork:
π GitHub: ComfyUI_SenseNova_U1_REBEL
Installation & Usage
- Place your chosen
.gguffile into your ComfyUI models directory (e.g.,ComfyUI/models/gguf/). - Ensure you have installed the ComfyUI_SenseNova_U1_REBEL custom node suite.
- Load the model using the
SenseNova_SM_Modelloader node and hook it up to theSenseNova_SM_Sampler.
- Downloads last month
- -
Hardware compatibility
Log In to add your hardware
2-bit
3-bit
4-bit
5-bit
6-bit
8-bit
Model tree for realrebelai/SenseNova-U1.5-8B_GGUFs
Base model
sensenova/SenseNova-U1.5-8B-MoT