X-ASR zh-en 960ms streaming zipformer2 transducer โ€” GGUF

GGUF conversion of GilgameshWind/X-ASR-zh-en (k2-fsa / sherpa-onnx streaming zipformer2 transducer, 960 ms chunk) for the RapidSpeech.cpp ggml runtime (branch zipformer), targeting CUDA on the Jetson Nano gen1 (sm_53).

File Size Notes
x-asr-zh-en-960ms-f16.gguf 292 MB f16 weights, 922 tensors
x-asr-zh-en-960ms-q4_k.gguf 85 MB Q4_K (block-32 fallback where ne0 % 256 != 0)

arch XAsrZipformer2. 16 kHz, 80-dim kaldi fbank (povey window). Vocab 5000 (zh+en, punctuation). Converted with scripts/xasr/convert_xasr_to_gguf.py. encoder_embed forward validated vs sherpa-onnx (corr 0.999998); full encoder/transducer port in progress.

Downloads last month
37
GGUF
Model size
0.2B params
Architecture
XAsrZipformer2
Hardware compatibility
Log In to add your hardware

16-bit

Inference Providers NEW
This model isn't deployed by any Inference Provider. ๐Ÿ™‹ Ask for provider support