Quantized via the updated crispasr-quantize tool. This CLI provides superior performance over the older cohere-quantize tool due to better tensor handling and model-agnostic alignment logic. Empirical tests confirm higher quality compared to versions using the legacy tool, specifics are in the manifest.json

quantized versons of nvidia/parakeet-tdt-0.6b-v3 in gguf format via CrispASR (ggml backend)

  • made for use with parakit, a daemon-style local dictation tool I made
  • Empirical tests comparing the Q8, the F16, and the original Nemo file (all in this repo) on a ~60s dictation show essentially no degradation in both Q8 and F16
Downloads last month
93
GGUF
Model size
0.6B params
Architecture
parakeet
Hardware compatibility
Log In to add your hardware

8-bit

16-bit

Inference Providers NEW
This model isn't deployed by any Inference Provider. ๐Ÿ™‹ Ask for provider support

Model tree for pszemraj/parakeet-tdt-0.6b-v3-gguf

Quantized
(62)
this model