Instructions to use ibm-granite/granite-4.0-1b-speech with libraries, inference providers, notebooks, and local apps. Follow these links to get started.
- Libraries
- Transformers
How to use ibm-granite/granite-4.0-1b-speech with Transformers:
# Use a pipeline as a high-level helper from transformers import pipeline pipe = pipeline("automatic-speech-recognition", model="ibm-granite/granite-4.0-1b-speech")# Load model directly from transformers import AutoProcessor, AutoModelForMultimodalLM processor = AutoProcessor.from_pretrained("ibm-granite/granite-4.0-1b-speech") model = AutoModelForMultimodalLM.from_pretrained("ibm-granite/granite-4.0-1b-speech", device_map="auto") - Notebooks
- Google Colab
- Kaggle
Update README.md
Browse files
README.md
CHANGED
|
@@ -89,7 +89,7 @@ assert wav.shape[0] == 1 and sr == 16000 # mono, 16kHz
|
|
| 89 |
|
| 90 |
# Create text prompt
|
| 91 |
user_prompt = "<|audio|>can you transcribe the speech into a written format?"
|
| 92 |
-
# Add "Keywords: <kw1>, <kw2>
|
| 93 |
chat = [
|
| 94 |
{"role": "user", "content": user_prompt},
|
| 95 |
]
|
|
|
|
| 89 |
|
| 90 |
# Create text prompt
|
| 91 |
user_prompt = "<|audio|>can you transcribe the speech into a written format?"
|
| 92 |
+
# Add "Keywords: <kw1>, <kw2> ..." at the end for keyword biasing
|
| 93 |
chat = [
|
| 94 |
{"role": "user", "content": user_prompt},
|
| 95 |
]
|