VibeVoice ASR full fine-tune on EN_Emilia_Yodas_616h

Training setup:

  • Base model: microsoft/VibeVoice-ASR
  • Hardware: 1x H200
  • Precision: bf16
  • Optimizer: Adafactor
  • Effective weighting: rows with non-empty events_scribe emitted 3x
  • Target content field: text_scribe

Caveat:

  • events_scribe != empty is a noisy proxy for nonverbals and also includes some lexical or annotation artifacts.
Downloads last month
4
Safetensors
Model size
9B params
Tensor type
BF16
·
Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support

Model tree for mrfakename/vibevoice-asr-en-emilia-yodas-616h-fft-events3x-20260322

Finetuned
(18)
this model

Dataset used to train mrfakename/vibevoice-asr-en-emilia-yodas-616h-fft-events3x-20260322