Automatic Speech Recognition
Safetensors
English
speech-to-text
canary-qwen
nvidia
qwen
optimum-quanto
qint8
int8
open-webui
openai-compatible
rtx-5090
quanto
quantized
openwebui
1.7gb
audio-encoder-unquantized
llm-only-quant
asr-preserved
systemd
fastapi
webm-opus
browser-mic
quanto-qint8
qint8-safetensors
8-bit precision
8bit
File size: 477 Bytes
39530e1 | 1 2 3 4 5 6 7 8 9 10 11 | {
"base_model": "nvidia/canary-qwen-2.5b",
"artifact_type": "optimum-quanto-qint8-llm-state-dict-plus-quantization-map",
"created_at": "2026-06-04T17:28:19-0700",
"scope": "SALM LLM component only",
"notes": [
"The SALM perception/audio encoder, tokenizer, and config are loaded from the upstream base model.",
"This artifact is intended for this service wrapper and reload path.",
"It is not advertised as a generic Transformers AutoModel drop-in."
]
} |