deepseek-ai/DeepSeek-V4-Flash-0731 Text Generation β’ 304B β’ Updated 2 days ago β’ 236k β’ β’ 1.97k
thinkingmachines/Inkling-Small Image-Text-to-Text β’ 266B β’ Updated 4 days ago β’ 8.5k β’ β’ 254
thinkingmachines/Inkling-Small-NVFP4 Image-Text-to-Text β’ 156B β’ Updated 4 days ago β’ 67.1k β’ 61
nvidia/NVIDIA-Nemotron-Labs-3-Puzzle-75B-A9B-NVFP4 Text Generation β’ 45B β’ Updated 27 days ago β’ 89.8k β’ 128
view post Post 4266 Weβre releasing new Qwen3.6 quants that run 2.5Γ faster on your GPU. β‘Qwen3.6-27B NVFP4 runs on 24GB VRAM.35B-A3B can hit 17,561 tok/s (B200).We also improved accuracy, tool calling, agent use, and looping.Qwen3.6 NVFP4: https://huggingface.co/collections/unsloth/nvfp4Guide: https://unsloth.ai/docs/models/qwen3.6#nvfp4 See translation 1 reply Β· π 15 15 π₯ 11 11 π€ 1 1 + Reply
unsloth/Qwen3.6-35B-A3B-NVFP4-Fast Image-Text-to-Text β’ 22B β’ Updated 22 days ago β’ 412k β’ 95